We provide IT Staff Augmentation Services!

Hadoop Developer Resume

0/5 (Submit Your Rating)

Parsippany, NJ

SUMMARY

  • Over all 8+ years of experience in IT which includes 3+ years' experience using Apache Hadoop and experience for analyzing the Big Data as per the requirement and 1+ years’ experience in Data warehouse implementation and 4+ years’ experience as a java developer.
  • Good Exposure on Apache Hadoop Map Reduce programming, Hive, PIG scripting and HDFS.
  • Hands on experience in installing, configuring, monitoring, and using Hadoop ecosystem components like Hadoop MapReduce, HDFS, HBase, Hive, Sqoop, Pig, Zookeeper, Horton works, Flume, spark.
  • In depth knowledge of understanding the Hadoop architecture and its components such as HDFS, Job tracker, Task tracker, Name Node, Data Node, Resource Manager, Node Manager, Map Reduce programs and YARN paradigm.
  • Strong experience in writing Map Reduce programs for Data Analysis. Hands on experience in writing custom partitions for Map Reduce.
  • Experience with distributed systems, large - scale non-relational data stores, RDBMS, data modeling, database performance, and multi-terabyte data warehouses.
  • Efficient in writing MapReduce Programs and using Apache Hadoop API for analyzing the structured and unstructured data.
  • Worked on developing ETL processes to load data from multiple data sources to HDFS using FLUME and SQOOP, perform structural modifications using Map-Reduce, HIVE and analyze data using visualization/ reporting tools.
  • Hands on experience and Good Knowledge on real time data feeding platform-KAFKA, integration software like Talend,
  • Good Knowledge on NOSQL databases like MongoDB, HBase and Cassandra.
  • Experience in using Sqoop to import data and export data into HDFS from RDBMS.
  • Solid understanding of source monitoring tool: Cloudera Manager.
  • Experience in using and understanding of Pig, Hive and HBase and Hive Built-in functions and Hive partitioning, bucketing and perform different types of joins on Hive tables and HDFS Designs, Daemons, HDFS high availability (HA).
  • Familiar with data warehousing and ETL tools like Informatica.
  • Familiar in Core Java with strong understanding and working knowledge in Object Oriented Concepts like Collections, Multithreading, Data Structures, Algorithms, JSP, Servlets, Multi-Threading, JDBC, HTML.
  • Excellent interpersonal and communication skills, creative, research-minded, technically competent, result oriented with problem solving as well maintaining the leadership skills and ability to work well with people and to maintain a good relation with the organization.

TECHNICAL SKILLS

Programming: C, C++, Java.

Frameworks: JDBC, Struts.

Data Base: SQL, MySQL, HBase, MongoDB, Cassandra.

Script: JavaScript, Shell Scripting.

Web Technology: HTML, CSS, JSP, Web Services, XML, JavaScript.

IDEs: Eclipse, Net Beans, MS Office, Microsoft Visual Studio

Web/Application: servers Apache Tomcat, Web logic.

Cluster Monitoring Tools: Apache Tomcat, Web logic.

Big Data: Hive, Map Reduce, HDFS, Sqoop, R, Flume, Spark, Scala, Apache Kafka, HBase, Pig, Oozie, Zookeeper, YARN, Talend.

PROFESSIONAL EXPERIENCE

Confidential, Parsippany, NJ

Hadoop Developer

Responsibilities:

  • Developing Hive scripts to select the Delta (CDC) and load into HBase tables using pig script
  • Transforming data using pig scripts
  • Developing MapReduce scripts to count large number of records in hbase tables
  • Working on different hive optimization and performance tuning techniques.
  • Working on Ingestion of logs into Hadoop using Flume and Kafka
  • Processing logs using spark streaming and loaded into hive tables. Using Spark jobs written in Scala used to transformed data and loaded into hive tables.
  • Using Hive SerDe to read and write data in different formats.
  • Involved in loading data from UNIX file system to HDFS.
  • Responsible for building scalable distributed data solutions using Hadoop.
  • Involved in loading data from edge node to HDFS.
  • Involved in Design, Architecture and Installation of Big Data and Hadoop ecosystem components.
  • Involved in creating Hive tables, loading with data and writing hive queries which will run internally in map reduce way.
  • Creating workflows using oozie and Automating Hadoop jobs using Oozie scheduler.
  • Gained very good business knowledge on health insurance, claim processing, fraud suspect identification, appeals process etc.

Environment: Hadoop, MapReduce, HDFS, Hive, HBase, Oozie, Kafka, Scala, Spark 1.6

Confidential, Minneapolis, MN

Hadoop Developer

Responsibilities:

  • Installed and configured Hadoop MapReduce, Flume, Avro and HBase
  • Involved in configuring Flume and Avro and HBase
  • Written HBase queries for finding different metrics.
  • Gained very good business knowledge on Transactions processing, fraud suspect identification, appeals process etc.
  • Developed pig scripts to transform data and loaded into Hbase tables
  • Created Hive snapshot tables and Hive ORC tables from hive tables.
  • Worked on evaluation and analysis of Hadoop cluster and different big data analytic tools like Hbase. Developed MapReduce programs to perform data filtering for unstructured data.
  • Optimized hive joins for large tables and developed map reduce code for the full outer join of two large tables.
  • Experience in using HDFS and My SQL and deployed HBase integration to perform OLAP operations on HBase data.
  • Creating Hbase tables for random read/writes by the map reduce programs.
  • Used Talend Big Data Open Studio 5.6.2 to create framework for executing extract framework
  • Used spark to parse XML files and extract values from tags and load it into multiple hive tables using map classes.
  • Used different bigdata components in Talend like thiverow, thiveInput, tHDFSCopy, tHDFSput, tHDFSGet, tMap, tdenormalize, tFlowtoIterate etc.,
  • Scheduled different talend jobs using TAC (Talend Admin Console)

Environment: Hadoop, MapReduce, HDFS, Flume, Hbase, Hive, Pig, TAC, Azure

Confidential, Detroit, MI

Informatica Consultant

Responsibilities:

  • Involved in the design and development of Data Warehousing project for the improvement of Account Management System.
  • Used Transformations like look up, Router, Filter, Joiner, Stored procedure, Source Qualifier, Aggregator, and Update strategy extensively.
  • Created Mapp let and used them in different Mappings.
  • Performed incremental aggregation to load incremental data into Aggregate tables.
  • Done extensive bulk loading into the target using Oracle SQL Loader.
  • Created standard and reusable Informatica mappings/moppets.
  • Creating and Run Sessions using Workflow Manager and Monitoring using Workflow Monitor.
  • Involved in Unit Testing and Resolution of various Bottlenecks came across.

Environment: Informatica Power Center 8.6, Flat files, Oracle 10G, TOAD, SQL server 2008, SQL, PL/SQL, T-SQL Windows 7

Confidential

JAVA/J2EE Developer

Responsibilities:

  • Developed activity, sequence and class diagrams using Unified Modeling Language and Rational Rose.
  • Developed HTML and JSP pages using Struts.
  • Developed Controller Servlets, Action and Form objects for process of interacting with Oracle database and retrieving dynamic data.
  • Responsible for coding SQL Statements and Stored procedures for back end communication using JDBC.
  • Used Apache Log4J logging API to log errors and messages.
  • Developed XML parser for File parsing.
  • Involved in writing Detail Design Documents with UML Specifications.
  • Involved in unit testing and system testing and also responsible for preparing test scripts for the system testing.
  • Responsible for packaging and deploying components in to the WebSphere.
  • Developed backend components, DBScripts for the backend communication.
  • Responsible for performance tuning of the product and eliminating memory leakages in the product.

Environment: Java, Java Beans, JSP, Servlets, JDBC, LOG4J, IBM DB2, XML, HTML, Struts 1.2, Web Sphere, CVS

Confidential

Java Developer

Responsibilities:

  • Designed and developed Struts like MVC 2 Web framework using the front-controller design pattern, which is used successfully in many production systems.
  • Spearheadedthe “Quick Wins” project by working very closely with the business and end users to improve the current website’s ranking from being 23rdto 6thin just 3 months.
  • Normalized Oracle database, conforming to design concepts and best practices.
  • Maintained records in Excel Spread Sheet and exported data into SQL Server Database using SQL Server Integration Services (SSIS).
  • Experience in providing Logging, Error handling by using Event Handler, and Custom Logging for SSIS Packages.
  • Resolved product complications at customer sites and funneled the insights to the development and deployment teams to adopt long term product development strategy with minimal roadblocks.
  • Convinced business users and analysts with alternative solutions that are more robust and simpler to implement from technical perspective while satisfying the functional requirements from the business perspective.
  • Applied design patterns and OO design conceptsto improve the existing Java/JEE based code base.
  • Identified and fixed transactional issues due to incorrect exception handling and concurrency issues due to unsynchronized block of code.

Environment: Java 1.2/1.3, Swing, Applet, Servlets, JSP, custom tags, JNDI, JDBC, XML, XSL, DTD, HTML, CSS, Java Script, Oracle, DB2, PL/SQL, WebLogic, JUnit, Log4J and CVS.

We'd love your feedback!