We provide IT Staff Augmentation Services!

Senior Hadoop Developer Resume

2.00/5 (Submit Your Rating)

Atlanta, GA

PROFESSIONAL SUMMARY:

  • Having 6 years of experience in the Information Technology field, with comprehensive hands - on expertise in the Application development, Production support, and Enhancement of Big Data Hadoop and mainframe platform applications. Having 4+ years of experience as a Hadoop Developer in developing and supporting Hadoop framework applications using Java, Pig, Hive, MapReduce, Talend ETL. Domain expertise in Banking, Engineering & Manufacturing and Health Insurance experience.
  • Hadoop Eco System: Excellent understanding/knowledge of Hadoop architecture and various components such as HDFS, JobTracker, TaskTracker, NameNode, DataNode and MapReduce programming paradigm.
  • Hadoop Administration: Coordinated with administrators in installing, configuring and using Hadoop ecosystem components like Hadoop MapReduce, HDFS, HBase, Oozie, Sqoop, Flume, Pig, Hive and Spark.
  • Hadoop Enterprise Distribution: Experience in installing, maintainingand upgrading Hadoop distributions like Cloudera CDH 5.x/4.x, Horton works 2.x, MapR 1.x and DataStax 4.x.
  • Hadoop Tools: Experience in analyzing data using HiveQL, HBase, Pig and custom MapReduce programs in Java. Hands-on experience in creating scripts for data-analysis with Pig, Hive and Impala. Experience in migrating applications from MapReduceto Spark using Scala.
  • Cassandra Developing and Modeling: Configuring and setting up Cassandra Cluster. Expertise in data modeling and analysis of Cassandra and Cassandra Query Language.
  • HBase: Ingested Data from RDBMS and HivetoHBase. Based on the requirements implemented custom coded MapReduce for HBase and used Java client API.
  • Data Ingestion: Using Flume designed the flow and configured the individual component. Efficiently transferred the data from and to traditional databases with Sqoop
  • Data Storage: Experience in maintaining distributed storage like HDFS, HBase and Cassandra.
  • SQL to NoSQL: Experience in migrating data from traditional database to NoSQL. As well included migration tools such Sqoop and Flume whileingesting data.
  • Management and Monitoring: Maintained and coordinated centralized services using Zookeeper.
  • Messaging System: Good knowledge of Kafka for fast messaging transfer across systems.
  • Cloud Platforms: Experience in using cloud based Hadoop clusters like Open Stackand Amazon Web Services (AWS).
  • Job Schedulers: Experience in scheduling the jobs using Apache Oozie
  • Scripting: Experienced in Hive, Pig and Shell Scripting
  • Very good working knowledge on Performance Tuning, Debugging, Testing on various platforms.
  • Hands-on experience in Linux, UNIX Shell Scripting.
  • Design and strong programming experience as a java developer in internet applications, client/server technologies using Java, J2EE, JSP, JDBC, XML and web based development tools.
  • UI Design: Comprehensive knowledge in HTML5, CSS3, JavaScript and Bootstrap
  • Project Management: Experience in Agile and Scrum project management.

TECHNICAL SKILLS:

Big Data Technologies: HDFS, MapReduce, Hive, Pig, Sqoop, Flume, Hbase, Storm, Impala, Kafka, Oozie, Spark, Zookeeper, Yarn, Mahout.

Programming Languages: Java, Objective C, Shell scripting, Pig Latin, Scala

Web Technologies: HTML 5, CSS 3, Java Script, Servlets, JSP, XML, JSON

Frameworks: Spring and Hibernate.

Web/App Servers: Apache Tomcat server, Apache HTTP webserver

Databases: Hive, MySql, SQL Server 2003/2008, DB Oracle.

NoSQL Databases: Hbase and Cassandra

Operating Systems: Linux (Centos, Ubuntu), Windows (XP/7/8)

IDE Tools& Utilities: Eclipse, NetBeans, Git, Maven.

Reporting/ETL Tools: Tableau, Informatica, SSIS, SSRS,Pentaho

PROFESSIONAL EXPERIENCE:

Confidential, Atlanta, GA

Senior Hadoop Developer

Responsibilities:

  • Coordinated with Administrators in setting up, configuring, initializing and troubleshooting DS Enterprise 4.7
  • Involved in loading data from UNIX file system to HDFS
  • Involved in defining job flows and running data streaming jobs to process terabytes of text data
  • Developed multiple MapReduce jobs in Java for data cleaning and preprocessing
  • Wrote MapReduce jobs to discover trends in data usage by users
  • Involved in managing and reviewing Hadoop log files
  • Installed and configured Hive and written HiveQL scripts.
  • Involved in loading and transforming large sets of structured, semi structured and unstructured data
  • Involved in creating Hive tables, loading data and writing Hive queries as per business requirements
  • Implemented static partitioning, dynamic partitioning and bucketing of data in Hive for improving the performance
  • Supported Map Reduce programs those are running on the cluster.
  • Involved in writing both DML and DDL operations in NoSQL database Cassandra
  • Responsible to manage data coming from different sources
  • Implemented POC on writing programs in Scala using Spark
  • Worked on migrating MapReduce programs into Spark using Scala
  • Assisted the team in their development &deployment activities
  • Used Web services concepts like SOAP to interact with other project within organization and for sharing information.
  • Involved in developing database access components using Spring DAO integrated with Hibernate for accessing the data.
  • Followed Hybrid (Waterfall - Scrum) principles in developing the project

ENVIRONMENT: DataStax Enterprise 4.7, Linux, Hadoop 2.4.0, MapReduce, Hive 0.12.0, Pig 0.10.1, Impala, HBase 0.96.1, Sqoop 1.4.5, Flume 1.4.0, ZooKeeper 3.4.5, Cassandra 2.1.5, Spark 1.2.1, Spark Cassandra Connector 1.2.1, SparkQL, Soap, Spring 4.1, Hibernate 4.3, Oracle 12c

Confidential, New York city,NY

Senior Hadoop Developer

Responsibilities:

  • Developed data pipeline using Flume, SQOOP, Pig and Java map reduce to ingest customer behavioral data and financial histories into HDFS for analysis
  • Involved in Sqoop, HDFS Put or Copy from Local to ingest data and Map Reduce jobs.
  • Used Pig to do transformations, event joins, filter boot traffic and some pre-aggregations before storing the data onto HDFS.
  • Involved in developing Pig UDFs for the needed functionality that is not out of the box available from Apache Pig.
  • Expertise with the tools in Hadoop Ecosystem including Pig, Hive, HDFS, Map Reduce, SQOOP, Kafka, Yarn, Oozie, and Zookeeper. Hadoop architecture and its components.
  • Extensive experience in using the MOM with Active MQ, Apache storm, Apache Spark & Kafka and Zookeeper.
  • Used Hive to analyze the partitioned and bucketed data and compute various metrics for reporting.
  • Involved in developing Hive DDLs to create, alter and drop Hive tables and storm.
  • Managed works including indexing data, tuning relevance, developing custom tokenizers and filters, adding functionality includes playlist, custom sorting and regionalization with Solr Search Engine.
  • Involved in loading data from UNIX file system to HDFS. Installed and configured Hive and also written Hive UDFs and Cluster coordination services through Zoo Keeper.
  • Involved in creating Hive tables, loading with data and writing hive queries which will run internally in map reduce way.
  • Experienced in managing Hadoop Cluster using Cloudera Manager Tool.
  • Involved in developing Hive UDFs for the needed functionality that is not out of the box available from Apache Hive.
  • Involved in using HCATALOG to access Hive table metadata from Map Reduce or Pig code.
  • Computed various metrics using Java Map Reduce to calculate metrics that define user experience, Revenue etc.
  • Responsible for developing data pipeline using flume, Sqoop and pig to extract the data from weblogs and store in HDFS.
  • Extracted and updated the data into Monod using Mongo import and export command line utility interface.
  • Extracted and updated the data into Monod using Mongo import and export command line utility interface. Involved in using SQOOP for importing and exporting data into HDFS.
  • Used Eclipse and ant to build the application. Proficient work experience with NOSQL, Monod databases also the HDFS data from Rows to Columns and Columns to Rows.
  • Involved in developing Shell scripts to orchestrate execution of all other scripts (Pig, Hive, and Map Reduce) and move the data files within and outside of HDFS.

Environment: Hadoop, Map Reduce, Mongo, Yarn, Hive, Solr, Pig, HBase, Oozie, Sqoop, Flume, Oracle 11g, Core Java, Cloudera, HDFS, Eclipse.

Confidential, Richardson, TX

Hadoop Consultant

Responsibilities:

  • Interacting with the Business Requirements and the design team and preparing the Low Level Design and high level design documents.
  • Provide in-depth technical and business knowledge to ensure efficient design, programming, implementation and on-going support for the application.
  • Involved in identifying possible ways to improve the efficiency of the system.
  • Developed multiple MapReduce jobs in java for log data cleaning and preprocessing and scheduled the job to collect aggregate the log on an hourly basis.
  • Implemented MapReduce programs using Java.
  • Logical implementation and interaction with HBase.
  • Efficiently put and fetched data to/from HBase by writing Map/Reduce job.
  • Developed Map Reduce jobs to automate transfer of data from/to HBase.
  • Used flume to collect all the web log from the online ad-servers and push into HDFS.
  • Wrote efficient map reduce code to aggregate the log data from the Ad-server.
  • Used Hive to analyze the partitioned and bucketed data and compute various metrics for reporting.
  • Prepared multi-cluster test harness to exercise the system for better performance.
  • Developed high-performance cache, making the site stable and improving its performance.

Environment: HDFS, Map Reduce, Hive, PIG, UNIX Shell, Cloudera, OOZIE, HUE Editor, Spark, Qlick Stream, MYSQL, DB2, Netezza.

Confidential

Java Developer

Responsibilities:

  • Used Spring Framework for dependency injection using spring configuration files
  • Developed the presentation layer using JSP, HTML, CSS and client validations using JavaScript
  • Involved in the installation and configuration of Tomcat Server
  • Involved in dynamic form generation, auto completion of forms and user validation functionalities using AJAX
  • Involved in the design part of the User Stories in Scrum process.
  • Worked in Scrum Agile process with two weeks iteration delivering new features at iteration.
  • Develop the User Stories within the timeline and discuss the status in Daily Scrum meetings.
  • Used Collections interface to create a composite pattern and provide a dynamic user interface.
  • Implemented the Singleton pattern for the data source.
  • Developed the web application using MVC design pattern.
  • Implemented the front-end screens using Spring MVC.
  • Created validations utility classes for front-end validations.
  • Integration of Springand Hibernate frameworks.
  • Designed, developed and maintained the data layer using Hibernate and performed configuration of Spring application framework.
  • Created stored procedures using PL/SQL for data access layer
  • Worked on tuning of back-end Oracle stored procedures using TOAD
  • Developed test cases for unit testing using JUnit
  • Developed Stored Procedures to extract data from Oracle database

ENVIRONMENT: Java/J2EE, JSP, Servlets, Spring Framework, Hibernate, SQL/PLSQL, Web Services, WSDL, JUnit, Tomcat, Oracle 9i

We'd love your feedback!