We provide IT Staff Augmentation Services!

Jr. Hadoop Developer Resume

0/5 (Submit Your Rating)

PROFESSIONAL SUMMARY:

  • Around 3 years of experience in Big Data and Hadoop Adminstration
  • Expertise with tools in Hadoop Ecosystem including HDFS, MapReduce, Hive, Sqoop, Pig, Spark, Kafka, Yarn, Oozie, and Zookeeper.
  • Worked on Hadoop distributions like Cloudera, MapR and Horton Works.
  • Worked on importing and exporting data using stream processing platforms like Flume and Kafka.
  • Experience in Enterprise Service Bus(ESB) such as WebSphere Message Brokering.
  • Comprehensive knowledge of Software Development Life Cycle (SDLC), having thorough understanding of various phases like Requirements Analysis, Design, Development and Testing.
  • Ability to quickly master new concepts and applications.

TECHNICAL SKILLS:

Hadoop/Big Data Technologies/Programming Languages/Databases/Operating Systems: Hadoop 2.x, HDFS, Map Reduce, Hive, Pig, Sqoop, HBase, Flume, Oozie, Zookeeper, Spark, ScalaJava, C, Python, SQL/HiveQL, ESQLSQL Server, NoSQL, MongoDB, CassandraLinux, Unix, Windows, Mac

PROFESSIONAL EXPERIENCE:

Jr. Hadoop Developer

Confidential

Responsibilities:

  • Importing data into HDFS from Relational Database Systems and vice - versa by making use of Sqoop.
  • CreatingHiveTables, loading with data and writingHivequeries which will invoke and run Map Reduce jobs in the backend.
  • Analyzing large datasets to find patterns and insights within structured and unstructured data to help business intelligence with the help of Tableau.
  • Integrating MapReduce with HBase to import huge clusters of data using MapReduce programs.
  • Implementing several workflows using Apache Oozie framework to automate tasks.
  • Used Zookeeper to co-ordinate and run different cluster services.
  • Designing and developing applications in Spark using Scala to compare the performance of Spark with Hive and SQL/Oracle.
  • Developed and Configured Kafka brokers to pipeline server logs data into spark streaming.

Environment: Hadoop, Cloudera, Pig, Hive, Sqoop, Kafka, Spark, Storm, Tableau, HBase, Scala, Kerberos, Agile, Zookeeper, AWS, MySQL.

Hadoop Consultant

Confidential

Responsibilities:

  • Helping in Design, implementation and deployment of Hadoop cluster and providing solutions based on issues using big data analytics.
  • Part of the team that built scalable distributed data solutions using Hadoop cluster environment using Horton Works distribution.
  • Loading data into the Hadoop distributed file system (HDFS) with the help of Kafka and REST API
  • Worked on Sqoop to load data into HDFS from Relational Database Management Systems.
  • Carried out transforming huge data of Structured, Semi-Structured and Unstructured types and analyzing them using Hive queries and Pig scripts.
  • Worked with NoSQL databases like Cassandra and MongoDB for developing and implementing programs in Hadoop Environment.
  • Using Apache Oozie to execute the workflows.

Environment: Hadoop, Pig, Hive, HBase, Sqoop, Spark, Scala, Oozie, Zookeeper, RHEL, Java, Eclipse, SQL, NoSQL, Talend, Tableau.

Confidential

Responsibilities:

  • Facilitated communication between faculty and students on studend-related issues.
  • Represented over 300 students in various cultural, academic, professional and orientation meetings.
  • Served to enhance the academic experience of fellow classmates as well as the teaching experience of faculty at NEC
  • Provided feedbacks, commends and served a great hand in solving many practical day-to-day problems that students faced at NEC.
  • Guiding freshers through better paths with the past experiences and experiences of fellow students.
  • Filled in for teaching assistants when the professor needed help with classwork or Lab work.

Hadoop Intern

Confidential

Responsibilities:

  • Used Hadoop Cloudera Distribution. Involved in all phases of the Big Data Implementation including requirement analysis, design and development of Hadoop cluster.
  • Using Spark RDD and Spark SQL to convert MapReduce jobs into Spark transformations by using data sets and Spark Data frames.
  • Coding Scala for various Spark jobs to analyse customer data and sales history among other data.
  • Design, build and support pipelines of data ingestion, transformation, conversion and validation.
  • Performing different types joins on hive tables along with experience in partitioning, bucketing and collection concepts in Hive for efficient data access.
  • Worked on NoSQL databases including HBase and Cassandra.
  • Developed the data model to manage the summarized data.

Environment: Hadoop, Cloudera Hive, Java, Oozie, Cassandra, Zookeeper, HiveQl/SQL, MongoDB, Tableau.

We'd love your feedback!