Jr. Hadoop Developer Resume
PROFESSIONAL SUMMARY:
- Around 3 years of experience in Big Data and Hadoop Adminstration
- Expertise with tools in Hadoop Ecosystem including HDFS, MapReduce, Hive, Sqoop, Pig, Spark, Kafka, Yarn, Oozie, and Zookeeper.
- Worked on Hadoop distributions like Cloudera, MapR and Horton Works.
- Worked on importing and exporting data using stream processing platforms like Flume and Kafka.
- Experience in Enterprise Service Bus(ESB) such as WebSphere Message Brokering.
- Comprehensive knowledge of Software Development Life Cycle (SDLC), having thorough understanding of various phases like Requirements Analysis, Design, Development and Testing.
- Ability to quickly master new concepts and applications.
TECHNICAL SKILLS:
Hadoop/Big Data Technologies/Programming Languages/Databases/Operating Systems: Hadoop 2.x, HDFS, Map Reduce, Hive, Pig, Sqoop, HBase, Flume, Oozie, Zookeeper, Spark, ScalaJava, C, Python, SQL/HiveQL, ESQLSQL Server, NoSQL, MongoDB, CassandraLinux, Unix, Windows, Mac
PROFESSIONAL EXPERIENCE:
Jr. Hadoop Developer
Confidential
Responsibilities:
- Importing data into HDFS from Relational Database Systems and vice - versa by making use of Sqoop.
- CreatingHiveTables, loading with data and writingHivequeries which will invoke and run Map Reduce jobs in the backend.
- Analyzing large datasets to find patterns and insights within structured and unstructured data to help business intelligence with the help of Tableau.
- Integrating MapReduce with HBase to import huge clusters of data using MapReduce programs.
- Implementing several workflows using Apache Oozie framework to automate tasks.
- Used Zookeeper to co-ordinate and run different cluster services.
- Designing and developing applications in Spark using Scala to compare the performance of Spark with Hive and SQL/Oracle.
- Developed and Configured Kafka brokers to pipeline server logs data into spark streaming.
Environment: Hadoop, Cloudera, Pig, Hive, Sqoop, Kafka, Spark, Storm, Tableau, HBase, Scala, Kerberos, Agile, Zookeeper, AWS, MySQL.
Hadoop Consultant
Confidential
Responsibilities:
- Helping in Design, implementation and deployment of Hadoop cluster and providing solutions based on issues using big data analytics.
- Part of the team that built scalable distributed data solutions using Hadoop cluster environment using Horton Works distribution.
- Loading data into the Hadoop distributed file system (HDFS) with the help of Kafka and REST API
- Worked on Sqoop to load data into HDFS from Relational Database Management Systems.
- Carried out transforming huge data of Structured, Semi-Structured and Unstructured types and analyzing them using Hive queries and Pig scripts.
- Worked with NoSQL databases like Cassandra and MongoDB for developing and implementing programs in Hadoop Environment.
- Using Apache Oozie to execute the workflows.
Environment: Hadoop, Pig, Hive, HBase, Sqoop, Spark, Scala, Oozie, Zookeeper, RHEL, Java, Eclipse, SQL, NoSQL, Talend, Tableau.
Confidential
Responsibilities:
- Facilitated communication between faculty and students on studend-related issues.
- Represented over 300 students in various cultural, academic, professional and orientation meetings.
- Served to enhance the academic experience of fellow classmates as well as the teaching experience of faculty at NEC
- Provided feedbacks, commends and served a great hand in solving many practical day-to-day problems that students faced at NEC.
- Guiding freshers through better paths with the past experiences and experiences of fellow students.
- Filled in for teaching assistants when the professor needed help with classwork or Lab work.
Hadoop Intern
Confidential
Responsibilities:
- Used Hadoop Cloudera Distribution. Involved in all phases of the Big Data Implementation including requirement analysis, design and development of Hadoop cluster.
- Using Spark RDD and Spark SQL to convert MapReduce jobs into Spark transformations by using data sets and Spark Data frames.
- Coding Scala for various Spark jobs to analyse customer data and sales history among other data.
- Design, build and support pipelines of data ingestion, transformation, conversion and validation.
- Performing different types joins on hive tables along with experience in partitioning, bucketing and collection concepts in Hive for efficient data access.
- Worked on NoSQL databases including HBase and Cassandra.
- Developed the data model to manage the summarized data.
Environment: Hadoop, Cloudera Hive, Java, Oozie, Cassandra, Zookeeper, HiveQl/SQL, MongoDB, Tableau.
