Sr. Hadoop Developer Resume
Woonsocket, RI
PROFESSIONAL SUMMARY:
- 7+ years of professional experience in IT, including 4 years of work experience in Big Data Analytics, Distributed Computing using Hadoop Ecosystem tools.
- String hands on experience using major components in Hadoop Ecosystem like Spark, Map Reduce, HIVE, PIG, HBase, Sqoop, Oozie, Flume and Kafka.
- Excellent knowledge and understanding of Distributed Computing and Parallel processing frameworks.
- Strong experience with developing end - to-end spark applications in scala.
- Worked extensively on troubleshooting issues related to memory management, resource management, with in spark applications.
- Strong knowledge on fine-tuning spark applications and hive scripts.
- Written complex mapreduce jobs to perform various data transformations on large scale datasets.
- Experience in installation, configuration, and monitoring Hadoop clusters both in house and on the cloud (AWS).
- Extending Hive and Pig core functionality by writing custom UDF’s for Data Analysis.
- Handling importing of data from various data source, performed transformation, and hands on developing and debugging MR2 jobs to process large data sets.
- Experience in using Apache Flume for collecting, aggregation, moving large amount of data from application server.
- Hands on experience using tableau to perform operations like report generation.
- Experience in creating tables on top of data on AWS S3 obtained from different data sources and providing them to analytics team building reports using Tableau.
- Used sqoop extensively for ingesting data from relational data sources.
- Strong knowledge of entire SDLCE - Requirement Gathering & Analysis, Planning, Design, Development, Testing and Implementation.
- Used NoSql technologies like Hbase, Mongo dB for data extraction and storing huge volume of data. Knowledge in job/workflow scheduling and monitoring tools like Oozie, Zookeeper.
- Expertise in writing Map Reduce jobs using Java native code, Pig, Hive for data Processing.
- Used SVN repository for version control of the developed code.
- Experience working on NoSQL databases including Cassandra and Hbase.
- Major strengths are familiarity with multiple software systems, ability to learn quickly new technologies, adapt to new environments, self-motivated, team player, focused adaptive and quick learner with excellent interpersonal, technical and communication skills.
- Strong oral and written communication, initiation, interpersonal learning and organizing skills matched with the ability to manage time and people effectively.
TECHNICAL SKILLS:
Hadoop Stack: HDFS, Map Reduce, Apache Pig, Hbase, Hive, Impala, Oozie, ZooKeeper, Flume, Sqoop, MRUnit.
Database: Oracle 10g/11g, Sql Server 2005/2008 R2, My SQL, DB2, HBase, MongoDB, Cassandra.
Framework: Structs2, Spring 2.5/2.0, Hibernate 3.0.
Operating Systems: Windows 2008, 2003, 2000 Server, Windows 95/98/XP/Vista/7, DOS, Red Hat Linux, Macintosh OSX.
Database Tools: SQL Enterprise Manager, SQL Profiler, Query Analyser, SQL Server Setup, Security Manager, Service manager, DTS, Import Export Data, Bulk Insert, SQL Server Reporting Services(SSRS)
Programming Languages: Java, J2EE, PL/SQL.
Script Languages: JavaScript, JQuery, Python, Shell Script(BASH)
Methodologies: Waterfall, Iterative, Agile/Scrum
WORK EXPERIENCE:
Confidential, Woonsocket, RI
Sr. Hadoop developerResponsibilities:
- Uploaded and processed more than 40 terabytes of data from various structured and unstructured sources into HDFS using Sqoop and Kafka and Custom Input Adapters.
- Developed spark applications for performing large scale transformations and denormalization of relational datasets.
- Utilized Spark-SQL and Dataframe API in spark for writing custom transformations and data aggregations.
- Generated various reports using tableau.
- Created data visualization to view reports with tableau.
- Writing the queries using Impala query engine to get the faster results.
- Developed shell scripts for running Hive scripts in Hive and Impala.
- Modeled various hive tables and optimized the access by designing partitions and bucketing.
- Proactively monitored systems and services, architecture design and implementation of Hadoop deployment, configuration management, backup, and disaster recovery systems and procedures.
- Designed HBase tables for time series data. Designed row key to avoid region server hot spotting.
- Involved in Analyzing system failures, identifying root causes, and recommended course of actions.
- Documented the systems processes and procedures for future references.
- Worked with systems engineering team to plan and deploy new Hadoop environments and expand existing Hadoop clusters.
- Involved in setting up a 80 node Hadoop cluster with the help of devops team.
- Designed and configured Kafka cluster to accommodate heavy throughput of 1 million messages per second. Used kafka producer Java API’s to write messages to kafka topics.
- Used HBase API’s to get and scan events data stored in HBase.
- Used Java extensively to write custom input adapters and rest services utilized by down stream applications teams.
Environment: HDP 2.2 - Hadoop, Kafka, Impala, Spark, HBase, HDFS, ZooKeeper, Java, Spark, Shell Scripting, JSON, JUnit MySQL, JIRA, Confluence, Putty.
Sr. Hadoop developer
Confidential, Kansas City, MO
Responsibilities:
- Responsible for building scalable distributed data solutions using Hadoop.
- Responsible for fine tuning long running hive queries.
- Responsible for Cluster maintenance, adding and removing cluster nodes, Cluster Monitoring and Troubleshooting, Manage and review data backups and log files.
- Analyzed data using Hive and Pig and written custom UDFs in Hive.
- Worked hands on with ETL process.
- Involved in coding the modules and Testing
- Involved in Code walkthrough, Test Plan and Test Script Reviews
- Involved in Component Integration testing and System testing.
- Worked with application teams to install operating system, Hadoop updates, patches, version upgrades as needed.
- Responsible for running Hadoop MR jobs to process terabytes of xml's data.
- Load and transform large sets of structured, semi structured and unstructured data using Hadoop/Big Data concepts.
- Responsible for creating Hive tables, loading data and writing Hive queries.
- Handled importing data from various data sources, performed transformations using Hive, Map Reduce, and loaded data into HDFS.
- Extracted the data from Teradata into HDFS using the Sqoop.
- Exported the patterns analyzed back to Teradata using Sqoop.
- Installed Oozie workflow engine to run multiple Hive and MR jobs which run independently with time and data availability.
- Involved in running Hadoop jobs for processing millions of records of text data.
Environment: CDH 5.4, Apache Solr, MapReduce,Hive,Pig,Sqoop, Java, Eclipse Kepler, SVN repository, Linux, Putty, WinSCP.
Hadoop Developer
Confidential, Chicago, IL
Responsibilities:
- Responsible for building scalable distributed data solutions using Hadoop.
- Responsible for fine tuning long running hive queries.
- Responsible for Cluster maintenance, adding and removing cluster nodes, Cluster Monitoring and Troubleshooting, Manage and review data backups and log files.
- Analyzed data using Hive and Pig and written custom UDFs in Hive.
- Worked hands on with ETL process.
- Involved in coding the modules and Testing
- Involved in Code walkthrough, Test Plan and Test Script Reviews
- Involved in Component Integration testing and System testing.
- Worked with application teams to install operating system, Hadoop updates, patches, version upgrades as needed.
- Responsible for running Hadoop MR jobs to process terabytes of xml's data.
- Load and transform large sets of structured, semi structured and unstructured data using Hadoop/Big Data concepts.
- Responsible for creating Hive tables, loading data and writing Hive queries.
- Handled importing data from various data sources, performed transformations using Hive, Map Reduce, and loaded data into HDFS.
- Extracted the data from Teradata into HDFS using the Sqoop.
- Exported the patterns analyzed back to Teradata using Sqoop.
- Installed Oozie workflow engine to run multiple Hive and MR jobs which run independently with time and data availability.
- Involved in running Hadoop jobs for processing millions of records of text data.
Environment: CDH 5.4, Apache Solr, MapReduce,Hive,Pig,Sqoop, Java, Eclipse Kepler, SVN repository, Linux, Putty, WinSCP.
Java Developer
Confidential, New York, NY
Responsibilities:
- Involved in requirements analysis and prepared Requirements Specifications document.
- Designed implementation logic for core functionalities
- Developed service layer logic for core modules using JSPs and Servlets and involved in integration with presentation layer
- Involved in implementation of presentation layer logic using HTML, CSS, JavaScript and XHTML
- Design of MySQL database to store customer's general and billing details
- Used JDBC connections to store and retrieve data from the database.
- Development of complex SQL queries and stored procedures to process and store the data
- Used SAX parsers for parsing, XSL/XSLT transformation in customizing the statements reports that are retrieved from the database.
- Used ANT, a build tool to configure application
- Developed test cases using JUnit
- Involved in unit testing and bug fixing.
- Prepared design documents for code developed and defect tracker maintenance.
Environment: Java, J2EE (JSPs & Servlets), JUnit, HTML, CSS, JavaScript, Apache Tomcat, MySQL.
Java Developer
Confidential, Houston, TX
Responsibilities:
- Involved in almost all the phases of SDLC
- Complete involvement in Requirement Analysis and documentation on Requirement Specification.
- Developed prototype based on the requirements using Struts2 framework as part of POC (Proof of Concept)
- Prepared use-case diagrams, class diagrams and sequence diagrams as part of requirement specification documentation
- Involved in design of the core implementation logic using MVC architecture
- Used Apache Maven to build and configure the application.
- Configured struts.xml file with required action-mappings for all the required services.
- Developed implementation logic using struts2 framework
- Developed JAX-WS web services to provide services to the other systems.
- Developed JAX-WS client to utilize few of the services provided by the other systems.
- Involved in developing EJB 3.0 Stateless Session beans for business tier to expose business to services component as well as web tier.
- Implemented Hibernate at DAO layer by configuring hibernate configuration file for different databases.
- Developed business services to utilize hibernate service classes that connect to the database and to perform the required action.
- Developed JSP pages using struts JSP-tags and inhouse tags to meet business requirements
- Developed JavaScript validations to validate form fields
- Developed design documents for the code developed.
- Used SVN repository for version control of the developed code
Environment: Java, J2EE, MVC, Struts2, Hibernate, Apache Tomcat Server, XML.
Jr. Java Developer
Confidential
Responsibilities:
- Involved in requirements analysis and prepared Requirements Specifications document.
- Designed implementation logic for core functionalities
- Developed service layer logic for core modules using JSPs and Servlets and involved in integration with presentation layer
- Involved in implementation of presentation layer logic using HTML, CSS, JavaScript and XHTML
- Design of MySQL database to store customer's general and billing details
- Used JDBC connections to store and retrieve data from the database.
- Development of complex SQL queries and stored procedures to process and store the data
- Used database modeling, administration and development using SQL and PL/SQL in Oracle (8i, 9i and 10g), DB2 and SQL Server environments.
- Used ANT, a build tool to configure application
- Developed test cases using JUnit
- Involved in unit testing and bug fixing.
- Prepared design documents for code developed and defect tracker maintenance.
- Involved in Technical Discussions, Design, and Workflow.
- Participate in the Requirement Gathering and Analysis.
- Employed JAXB to unmarshall XML into Java Objects.
Environment: Java, J2EE 1.4, JSF, IBM WAS, EJB 2.1, UML, Rational Rose, XML, XSLT, SOAP, SAX, JSP 2.0, JMS, HTML, JDBC, JavaScript, OOAD, Servlets 2.4, Eclipse, Confidential, PL/SQL. Oracle 9i.
