We provide IT Staff Augmentation Services!

Hadoop Developer Resume

0/5 (Submit Your Rating)

Newark, NJ

SUMMARY:

  • Around 7 years of experience in IT industry with strong emphasis on Object Oriented Analysis, ETL Design, Development and Implementation, Testing and Deployment of Data Warehouse as well as with Big Data Processing in ingestion, storage, querying and analysis.
  • 4+ year’s experience in deployment of Hadoop Ecosystems like Map Reduce, Yarn, Sqoop, Flume, Pig, Hive, Hbase, Cassandra, Zoo Keeper, Oozie, and Impala.
  • Experience in building, maintaining multiple Hadoop clusters of different sizes and configuration and setting up teh rack topology for large clusters.
  • Good Understanding of Hadoop architecture and Hands - on experience with Hadoop components such as Job Tracker, Task Tracker, Name Node, Data Node and Map Reduce concepts and HDFS Framework.
  • Excellent Hands on Experience in developing Hadoop Architecture within teh project in Windows and Linux platforms.
  • Good technical Skills in Oracle 11i, SQL Server, ETL Development using Informatica tool.
  • Good Scripting Skills in Pig and Hive Systems.
  • Good experience in optimizing Map Reduce algorithms using Mappers, Reducers, combiners and practitioners to deliver teh best results for teh large datasets. Good experience in writing Map Reduce jobs using Java native code, Pig, Hive for various business use cases
  • Good Experience in data loading from Oracle and MYSQL databases to HDFS system using Sqoop (Structure Data) and Flume (Log Files & XML).
  • Performed data analytics using PIG and Hive for Data Architects and Data Scientists within teh team.
  • Experience in developing customized UDF’s in java to extend Hive and Pig Latin functionality.
  • Experience with NoSQL databases like Hbase, Cassandra, and MarkLogic as well as other ecosystems like ZooKeeper, Oozie, Impala, Storm etc.
  • Extensive experience in middle-tier development using J2EE technologies like JDBC, JNDI, JSP, Servlets, JSP, JSF, Struts, Spring, Hibernate, JDBC, EJB.
  • Experience with web-based UI development using jQuery UI, jQuery, ExtJS, CSS, HTML, HTML5, XHTML and Javascript.
  • Developed stored procedures and queries using PL/SQL.
  • Expertise in RDBMS like Oracle, MS SQL Server, MySQL and DB2.
  • Strong analytical skills with ability to quickly understand clients business needs. Involved in meetings to gather information and requirements from teh clients. Leading teh Team and involved in Onsite, Offshore co-ordination.

TECHNICAL SKILLS:

Hadoop/Big Data: HDFS, Mapreduce, HBase, Pig, Hive, Sqoop, Flume, Impala, Oozie, Zookeeper, Spark, Impala.

Java & J2EE Technologies: Core Java, Servlets, JSP, JDBC, JNDI, Java Beans

IDE’s: Eclipse, Net beans

Frameworks: MVC, Struts, Hibernate, Spring

Programming languages: C, C++, Java, Javascript, Python, Linux shell scripts

Databases: Oracle 11g/10g/9i, MySQL, DB2, MS-SQL Server, Cassendra, Marklogic

Web Servers: Web Logic, Web Sphere, Apache Tomcat

Web Technologies: HTML, XML, JavaScript, AJAX, SOAP, WSDL

Network Protocols: TCP/IP, UDP, HTTP, DNS, DHCP

ETL Tools: Informatica

PROFESSIONAL EXPERIENCE:

Confidential

Hadoop Developer

Responsibilities:

  • Worked on analyzing Hadoop cluster and different Big Data analytic tools including Pig, Hive, HBase and SQOOP.
  • Developed multiple MapReduce jobs in PIG and Hive for data cleaning and pre-processing.
  • Coordinated with business customers to gather business requirements. And also interact with other technical peers to derive Technical requirements.
  • Extensively involved in Design phase and delivered Design documents.
  • Involved in Testing and coordination with business in User testing.
  • Importing and exporting data into HDFS and Hive using SQOOP.
  • Written Hive jobs to parse teh logs and structure them in tabular format to facilitate TEMPeffective querying on teh log data.
  • Involved in creating Hive tables, loading with data and writing hive queries.
  • Experienced in defining job flows.
  • Used Hive to analyze teh partitioned data and compute various metrics for reporting.
  • Experienced in managing and reviewing teh Hadoop log files.
  • Used Pig as ETL tool to do Transformations, even joins and some pre-aggregations.
  • Load and Transform large sets of structured and semi structured data.
  • Responsible to manage data coming from different sources.
  • Created Data model for Hive tables.
  • Involved in Unit testing and delivered Unit test plans and results documents.
  • Exported data from HDFS environment into RDBMS using Sqoop for report generation and visualization purpose.
  • Worked on Oozie workflow engine for job scheduling.

Environment: Hadoop, HDFS, MapReduce, Pig, Hive, Sqoop, HBase, Oozie, Flume, java

Confidential, Newark, NJ

Hadoop Developer

Responsibilities:

  • Worked on evaluation and analysis of Hadoop cluster and different big data analytic tools including Pig, Hbase database and Sqoop.
  • Responsible for building scalable distributed data solutions using Hadoop.
  • Involved in loading data from LINUX file system to Hadoop Distributed File System.
  • Created Hbase tables to store various data formats of PII data coming from different portfolios.
  • Experience in managing and reviewing Hadoop log files.
  • Exporting teh analyzed and processed data to teh relational databases using Sqoop for visualization and for generation of reports for teh BI team.
  • Used Oozie workflow engine to run multiple Hive and pig jobs.
  • Experienced in implementing POC's to migrate iterative map reduce programs into Spark transformations using Scala.
  • Analyzing large amounts of data sets to determine optimal way to aggregate and report on these data sets.
  • Used Marklogic Hadoop connector to import and export data from Marklogic database to HDFS.
  • Worked with teh Data Science team to gather requirements for various data mining projects.
  • Developed teh Pig and Hive queries as well as UDF'S to pre-process teh data for analysis.
  • Importing and exporting data into HDFS and Hive using Flume.
  • Analyzed large data sets by runningHive queriesandPig scripts.
  • Created dash boards using Tableau to analyze data for reporting.
  • Support for setting up QA environment and updating of configurations for implementation scripts with Pig and Sqoop.

Environment: Hadoop, HDFS, Pig, Sqoop, HBase, Spark, Scala, Shell Scripting, Linux, JSON, Marklogic, Informatica and RDBMS.

Confidential, Dover, NH

Hadoop Developer

Responsibilities:

  • Processed data into HDFS by developing solutions, analyzed teh data using MapReduce, Pig, Hive and produce summary results from Hadoop to downstream systems.
  • Involved in ETL, Data Integration and Migration.
  • Responsible for managing data from multiple source.
  • Developed teh Pig and Hive queries as well as UDF'S to pre-process teh data for analysis.
  • Importing and exporting data into HDFS and Hive using Flume.
  • Cluster co-ordination services through ZooKeeper.
  • Responsible for architecting Hadoop clusters with CDH4 on CentOS, managing with Cloudera Manager.
  • Used Sqoop widely in order to import data from various systems/sources (like MySQL) into HDFS.
  • Applied Hive quires to perform data analysis on HBase using Storage meet teh business requirements.
  • Created components like Hive UDFs for missing functionality in HIVE for analytics.
  • Hands on experience with NoSQL databases like HBase and Cassandra and Amazon Web Services.
  • Used different file formats like Text files, Sequence Files, Avro etc.
  • Installed and configured Hadoop, Mapreduce, HDFS, Developed multiple MapReduce jobs in java for data cleaning and preprocessing.

Environment: Hadoop, HDFS, MapReduce, Pig, Hive, Sqoop, HBase, Oozie, Cassandra, Java, Zookeeper

Confidential, Richmond, VA

Hadoop Developer

Responsibilities:

  • Developed data pipeline using Flume, Sqoop, Pig and Java map reduce and Spark to ingest customer behavioral data and purchase histories into HDFS for analysis.
  • Importing and exporting data into HDFS and Hive using Sqoop.
  • Exporting teh analyzed and processed data to teh relational databases using Sqoop for visualization and for generation of reports for teh BI team.
  • Analyzing large amounts of data sets to determine optimal way to aggregate and report on these data sets.
  • Used Pig as ETL tool to do transformations, event joins, filters and some pre-aggregations before storing teh data onto HDFS.
  • Optimizing Map reduce code, pig scripts, user interface analysis, performance tuning and analysis.
  • Used Hive to analyze teh partitioned and bucketed data and compute various metrics for reporting on teh dashboard.
  • Loaded teh aggregated data onto DB2 for reporting on teh dashboard.
  • Functional and non-functional requirements gathering.
  • Used Oozie workflow engine to run multiple Hive and pig jobs.

Environment: BigData/Hadoop, Spark, HDFS, Map-Reduce, Hive, Pig, Sqoop, Flume, Impala, Oozie, Informatica, Java, and DB2.

Confidential

Java/J2EE Developer

Responsibilities:

  • Involved in analysis, design and development of teh product and developed specifications dat include Use Cases, Class Diagrams, and Sequence Diagrams.
  • Used AJAX for client-to-server communication
  • Involved in resolving business technical issues.
  • Involved in build, staging, Testing and deployment of J2EE applications.
  • Project Planning, monitoring and control for small and large projects
  • Involve in Requirement Analysis, Design and Implementation activities.
  • Worked in UI team to develop new customer facing portal for Long Term Care Partners.
  • Deployment and Post deployment support.
  • Creating test cases and technical documents.
  • Developing front end GUI using Java Server Faces.
  • Implementing Java API using core java
  • Integrating front end with API.
  • Developed teh User Interfaces using Struts, JSP, JSTL, HTML, AJAX and JavaScript.

Environment: Core Java, J2EE, JSP, Struts, Hibernate, Apache Tomcat, MySQL, HTML, XML, CVS.

Confidential

Java/J2EE Developer

Responsibilities:

  • Involved in designing, coding, debugging, documenting and maintaining a number of applications.
  • Used AJAX for client-to-server communication
  • Involved in resolving business technical issues.
  • Developing Web Services’ API using java.
  • Developing front end GUI using Java Server Faces.
  • Creating jar files and deployed in teh server.
  • Involved in teh creation of SQL tables and indexes and also wrote queries to read/manipulate data.
  • Used JDBC to establish connection between teh database and teh application.
  • Implemented controllers layer using servlets and JSPs.
  • Implemented view layer using JSPs, JSTL and EL and also made custom JSP tags.
  • Created teh user interface using HTML, CSS and JavaScript.

Environment: Java (Jdk 1.6), Servlets, JSPs, Java Beans, HTML, CSS, JavaScript, JQuery, SQL, JDBCOracle 9i/10g.

We'd love your feedback!