We provide IT Staff Augmentation Services!

Sr. Hadoop Developer Resume

2.00/5 (Submit Your Rating)

Woonsocket, RI

PROFESSIONAL SUMMARY:

  • 7+ years of professional experience in IT, including 4 years of work experience in Big Data Analytics, Distributed Computing using Hadoop Ecosystem tools.
  • String hands on experience using major components in Hadoop Ecosystem like Spark, Map Reduce, HIVE, PIG, HBase, Sqoop, Oozie, Flume and Kafka.
  • Excellent knowledge and understanding of Distributed Computing and Parallel processing frameworks.
  • Strong experience with developing end - to-end spark applications in scala.
  • Worked extensively on troubleshooting issues related to memory management, resource management, with in spark applications.
  • Strong knowledge on fine-tuning spark applications and hive scripts.
  • Written complex mapreduce jobs to perform various data transformations on large scale datasets.
  • Experience in installation, configuration, and monitoring Hadoop clusters both in house and on the cloud (AWS).
  • Extending Hive and Pig core functionality by writing custom UDF’s for Data Analysis.
  • Handling importing of data from various data source, performed transformation, and hands on developing and debugging MR2 jobs to process large data sets.
  • Experience in using Apache Flume for collecting, aggregation, moving large amount of data from application server.
  • Hands on experience using tableau to perform operations like report generation.
  • Experience in creating tables on top of data on AWS S3 obtained from different data sources and providing them to analytics team building reports using Tableau.
  • Used sqoop extensively for ingesting data from relational data sources.
  • Strong knowledge of entire SDLCE - Requirement Gathering & Analysis, Planning, Design, Development, Testing and Implementation.
  • Used NoSql technologies like Hbase, Mongo dB for data extraction and storing huge volume of data. Knowledge in job/workflow scheduling and monitoring tools like Oozie, Zookeeper.
  • Expertise in writing Map Reduce jobs using Java native code, Pig, Hive for data Processing.
  • Used SVN repository for version control of the developed code.
  • Experience working on NoSQL databases including Cassandra and Hbase.
  • Major strengths are familiarity with multiple software systems, ability to learn quickly new technologies, adapt to new environments, self-motivated, team player, focused adaptive and quick learner with excellent interpersonal, technical and communication skills.
  • Strong oral and written communication, initiation, interpersonal learning and organizing skills matched with the ability to manage time and people effectively.

TECHNICAL SKILLS:

Hadoop Stack: HDFS, Map Reduce, Apache Pig, Hbase, Hive, Impala, Oozie, ZooKeeper, Flume, Sqoop, MRUnit.

Database: Oracle 10g/11g, Sql Server 2005/2008 R2, My SQL, DB2, HBase, MongoDB, Cassandra.

Framework: Structs2, Spring 2.5/2.0, Hibernate 3.0.

Operating Systems: Windows 2008, 2003, 2000 Server, Windows 95/98/XP/Vista/7, DOS, Red Hat Linux, Macintosh OSX.

Database Tools: SQL Enterprise Manager, SQL Profiler, Query Analyser, SQL Server Setup, Security Manager, Service manager, DTS, Import Export Data, Bulk Insert, SQL Server Reporting Services(SSRS)

Programming Languages: Java, J2EE, PL/SQL.

Script Languages: JavaScript, JQuery, Python, Shell Script(BASH)

Methodologies: Waterfall, Iterative, Agile/Scrum

WORK EXPERIENCE:

Confidential, Woonsocket, RI

Sr. Hadoop developer

Responsibilities:

  • Uploaded and processed more than 40 terabytes of data from various structured and unstructured sources into HDFS using Sqoop and Kafka and Custom Input Adapters.
  • Developed spark applications for performing large scale transformations and denormalization of relational datasets.
  • Utilized Spark-SQL and Dataframe API in spark for writing custom transformations and data aggregations.
  • Generated various reports using tableau.
  • Created data visualization to view reports with tableau.
  • Writing the queries using Impala query engine to get the faster results.
  • Developed shell scripts for running Hive scripts in Hive and Impala.
  • Modeled various hive tables and optimized the access by designing partitions and bucketing.
  • Proactively monitored systems and services, architecture design and implementation of Hadoop deployment, configuration management, backup, and disaster recovery systems and procedures.
  • Designed HBase tables for time series data. Designed row key to avoid region server hot spotting.
  • Involved in Analyzing system failures, identifying root causes, and recommended course of actions.
  • Documented the systems processes and procedures for future references.
  • Worked with systems engineering team to plan and deploy new Hadoop environments and expand existing Hadoop clusters.
  • Involved in setting up a 80 node Hadoop cluster with the help of devops team.
  • Designed and configured Kafka cluster to accommodate heavy throughput of 1 million messages per second. Used kafka producer Java API’s to write messages to kafka topics.
  • Used HBase API’s to get and scan events data stored in HBase.
  • Used Java extensively to write custom input adapters and rest services utilized by down stream applications teams.

Environment: HDP 2.2 - Hadoop, Kafka, Impala, Spark, HBase, HDFS, ZooKeeper, Java, Spark, Shell Scripting, JSON, JUnit MySQL, JIRA, Confluence, Putty.

Sr. Hadoop developer

Confidential, Kansas City, MO

Responsibilities:

  • Responsible for building scalable distributed data solutions using Hadoop.
  • Responsible for fine tuning long running hive queries.
  • Responsible for Cluster maintenance, adding and removing cluster nodes, Cluster Monitoring and Troubleshooting, Manage and review data backups and log files.
  • Analyzed data using Hive and Pig and written custom UDFs in Hive.
  • Worked hands on with ETL process.
  • Involved in coding the modules and Testing
  • Involved in Code walkthrough, Test Plan and Test Script Reviews
  • Involved in Component Integration testing and System testing.
  • Worked with application teams to install operating system, Hadoop updates, patches, version upgrades as needed.
  • Responsible for running Hadoop MR jobs to process terabytes of xml's data.
  • Load and transform large sets of structured, semi structured and unstructured data using Hadoop/Big Data concepts.
  • Responsible for creating Hive tables, loading data and writing Hive queries.
  • Handled importing data from various data sources, performed transformations using Hive, Map Reduce, and loaded data into HDFS.
  • Extracted the data from Teradata into HDFS using the Sqoop.
  • Exported the patterns analyzed back to Teradata using Sqoop.
  • Installed Oozie workflow engine to run multiple Hive and MR jobs which run independently with time and data availability.
  • Involved in running Hadoop jobs for processing millions of records of text data.

Environment: CDH 5.4, Apache Solr, MapReduce,Hive,Pig,Sqoop, Java, Eclipse Kepler, SVN repository, Linux, Putty, WinSCP.

Hadoop Developer

Confidential, Chicago, IL

Responsibilities:

  • Responsible for building scalable distributed data solutions using Hadoop.
  • Responsible for fine tuning long running hive queries.
  • Responsible for Cluster maintenance, adding and removing cluster nodes, Cluster Monitoring and Troubleshooting, Manage and review data backups and log files.
  • Analyzed data using Hive and Pig and written custom UDFs in Hive.
  • Worked hands on with ETL process.
  • Involved in coding the modules and Testing
  • Involved in Code walkthrough, Test Plan and Test Script Reviews
  • Involved in Component Integration testing and System testing.
  • Worked with application teams to install operating system, Hadoop updates, patches, version upgrades as needed.
  • Responsible for running Hadoop MR jobs to process terabytes of xml's data.
  • Load and transform large sets of structured, semi structured and unstructured data using Hadoop/Big Data concepts.
  • Responsible for creating Hive tables, loading data and writing Hive queries.
  • Handled importing data from various data sources, performed transformations using Hive, Map Reduce, and loaded data into HDFS.
  • Extracted the data from Teradata into HDFS using the Sqoop.
  • Exported the patterns analyzed back to Teradata using Sqoop.
  • Installed Oozie workflow engine to run multiple Hive and MR jobs which run independently with time and data availability.
  • Involved in running Hadoop jobs for processing millions of records of text data.

Environment: CDH 5.4, Apache Solr, MapReduce,Hive,Pig,Sqoop, Java, Eclipse Kepler, SVN repository, Linux, Putty, WinSCP.

Java Developer

Confidential, New York, NY

Responsibilities:

  • Involved in requirements analysis and prepared Requirements Specifications document.
  • Designed implementation logic for core functionalities
  • Developed service layer logic for core modules using JSPs and Servlets and involved in integration with presentation layer
  • Involved in implementation of presentation layer logic using HTML, CSS, JavaScript and XHTML
  • Design of MySQL database to store customer's general and billing details
  • Used JDBC connections to store and retrieve data from the database.
  • Development of complex SQL queries and stored procedures to process and store the data
  • Used SAX parsers for parsing, XSL/XSLT transformation in customizing the statements reports that are retrieved from the database.
  • Used ANT, a build tool to configure application
  • Developed test cases using JUnit
  • Involved in unit testing and bug fixing.
  • Prepared design documents for code developed and defect tracker maintenance.

Environment: Java, J2EE (JSPs & Servlets), JUnit, HTML, CSS, JavaScript, Apache Tomcat, MySQL.

Java Developer

Confidential, Houston, TX

Responsibilities:

  • Involved in almost all the phases of SDLC
  • Complete involvement in Requirement Analysis and documentation on Requirement Specification.
  • Developed prototype based on the requirements using Struts2 framework as part of POC (Proof of Concept)
  • Prepared use-case diagrams, class diagrams and sequence diagrams as part of requirement specification documentation
  • Involved in design of the core implementation logic using MVC architecture
  • Used Apache Maven to build and configure the application.
  • Configured struts.xml file with required action-mappings for all the required services.
  • Developed implementation logic using struts2 framework
  • Developed JAX-WS web services to provide services to the other systems.
  • Developed JAX-WS client to utilize few of the services provided by the other systems.
  • Involved in developing EJB 3.0 Stateless Session beans for business tier to expose business to services component as well as web tier.
  • Implemented Hibernate at DAO layer by configuring hibernate configuration file for different databases.
  • Developed business services to utilize hibernate service classes that connect to the database and to perform the required action.
  • Developed JSP pages using struts JSP-tags and inhouse tags to meet business requirements
  • Developed JavaScript validations to validate form fields
  • Developed design documents for the code developed.
  • Used SVN repository for version control of the developed code

Environment: Java, J2EE, MVC, Struts2, Hibernate, Apache Tomcat Server, XML.

Jr. Java Developer

Confidential

Responsibilities:

  • Involved in requirements analysis and prepared Requirements Specifications document.
  • Designed implementation logic for core functionalities
  • Developed service layer logic for core modules using JSPs and Servlets and involved in integration with presentation layer
  • Involved in implementation of presentation layer logic using HTML, CSS, JavaScript and XHTML
  • Design of MySQL database to store customer's general and billing details
  • Used JDBC connections to store and retrieve data from the database.
  • Development of complex SQL queries and stored procedures to process and store the data
  • Used database modeling, administration and development using SQL and PL/SQL in Oracle (8i, 9i and 10g), DB2 and SQL Server environments.
  • Used ANT, a build tool to configure application
  • Developed test cases using JUnit
  • Involved in unit testing and bug fixing.
  • Prepared design documents for code developed and defect tracker maintenance.
  • Involved in Technical Discussions, Design, and Workflow.
  • Participate in the Requirement Gathering and Analysis.
  • Employed JAXB to unmarshall XML into Java Objects.

Environment: Java, J2EE 1.4, JSF, IBM WAS, EJB 2.1, UML, Rational Rose, XML, XSLT, SOAP, SAX, JSP 2.0, JMS, HTML, JDBC, JavaScript, OOAD, Servlets 2.4, Eclipse, Confidential, PL/SQL. Oracle 9i.

We'd love your feedback!