We provide IT Staff Augmentation Services!

Hadoop Developer Resume

5.00/5 (Submit Your Rating)

Los Angeles, CA

SUMMARY

  • Over 8 years of professional IT experience which includes 2 years of experience in Big data ecosystem related technologies.
  • Excellent understanding / knowledge of Hadoop architecture and various components such as HDFS, Job Tracker, Task Tracker, NameNode, Data Node and MapReduce programming paradigm.
  • Proficient in Installation, Configuration and migrating and upgrading of data from Hadoop MapReduce, HIVE, HDFS, HBase, Sqoop, Oozie, Pig, Cloudera, Zookeeper, Flume and Cassandra.
  • Experience in installation, configuration, supporting and managing - CloudEra's Hadoop platformalong with CDH3&4 clusters.
  • Experience with leveraging Hadoop ecosystem components including Pig and Hive for data analysis, Sqoop for data migration, Oozie for scheduling and HBase as a NoSQL data store.
  • Good Exposure on Apache Hadoop Map Reduce programming, PIG Scripting and Distribute Application and HDFS.
  • Experience in NoSQL database MongoDB and Cassandra.
  • Experience in importing and exporting data using Sqoop from HDFS to Relational Database Systems and vice-versa.
  • Experience in deployment of Hadoop Cluster using Puppet tool.
  • Experience in Hadoop Shell commands, writing MapReduce Programs, verifying managing and reviewing Hadoop Log files.
  • Proficient in configuring Zookeeper, Cassandra & Flume to the existing Hadoop cluster.
  • In depth knowledge of JobTracker, TaskTracker, NameNode, DataNodes and MapReduce concepts.
  • Experience in understanding the security requirements for Hadoop and integrate with Kerberos authentication and authorization infrastructure.
  • Experience in Big Data analysis using PIG and HIVE and understanding of SQOOP and Puppet.
  • Good understanding of HDFS Designs, Daemons, federation and HDFS high availability (HA).
  • Experience in developing MapReduce programs using Apache Hadoop for working with Big Data.
  • Experience in developing customized UDF’s in java to extend Hive and Pig Latin functionality.
  • Good experience in design and development of various web and enterprise applicationsusing J2EE technologies like JSP, Servlets, EJB, JavaScript, DOJO, JDBC, JMS, JNDI, IBM RMI, XML, and Web Services.
  • Experience in developing MVC architecture using Servlets, JSP, Struts Framework, Hibernate Framework and Spring Framework.
  • Experience in using Java IDE tools like IBM WebSphere Studio Application Developer (WSAD), Rational Application Developer, Eclipse and familiar with other IDE's like Net Beans, JBuilder, and JDeveloper.
  • Good experience in creating enterprise applications on web/application servers such as JBOSS, Tomcat, IBM WebSphere, WebLogic.
  • Good understanding of Oracle architecture and Oracle internals (data dictionary, SQL execution, system waits, resource allocation, latches).
  • Experience in configuration and maintaining Oracle Standby Database and Oracle Data Guard. Experience on Logical & Physical Designing of databases, Data Modeling, Conceptual Design & Data Architecture using MS Visio, Erwin, Power Designer and TOAD Data modeler as a modeling tool.
  • Experience in using Oracle 11g/10g, DB2, SQL Server and MySQL databases and writing complex SQL queries.
  • Experience in Data Extraction, Transformation and Loading (ETL) using SQL Server Integration Services.
  • Strong team player, ability to work independently and in a team as well, ability to adapt to a rapidly changing environment, commitment towards learning.
  • Ability to blend technical expertise with strong Conceptual, Business and Analytical skills to provide quality solutions and result-oriented problem solving technique and leadership skills.

TECHNICAL SKILLS

Big Data: Hadoop, MapReduce, HDFS, Hive, Pig, Sqoop, Zookeeper and HBase

Languages: Java, SQL, HTML, DHTML, JavaScript, XML, C/C++.

Java/J2EE Technologies: JSP, Servlets, JavaBeans, JDBC, JNDI, JTA, JPA, EJB 3.0

Design Patterns: MVC, Session Facade, Service Locator, Data Access Object, Data Transfer Object / Value Object, Business Delegate

Web Design Tools: HTML, DHTML, AJAX, JavaScript, jQuery and CSS

Version Control Tools: CVS, Rational Clear Case

Frameworks: Struts 1.1/2.0, Spring 2.5, Hibernate 3.0

IDEs: NetBeans, Eclipse, IntelliJ, SQL Developer, Rational Rose for Java, JBuilder 4.0, Aptana

Databases: Oracle 9i / 10g/11g, SQL Server 2008, MS-SQL Server, SQL*Plus.

Operating systems: Windows XP, Linux, UNIX, DOS

PROFESSIONAL EXPERIENCE

Confidential, Los Angeles, CA

Hadoop Developer

Responsibilities:

  • Evaluated suitability of Hadoop and its ecosystem to the project and implementing / validating with various proof of concept (POC) applications to eventually adopt them to benefit from the Big Data Hadoop initiative.
  • Extracted the needed data from the server into HDFS and Bulk Loaded the cleaned data into HBase using MapReduce.
  • Used MRUnit for unit testing.
  • Developed HIVE queries for the analysts.
  • Performed ETL using Pig, Hive and MapReduce to transform transactional data to de-normalized form.
  • Configured periodic incremental imports of data from DB2 into HDFS using Sqoop.
  • Worked extensively with importing metadata into Hive using Sqoop and migrated existing tables and applications to work on Hive.
  • Wrote Pig and Hive User Defined Functions to analyze the complex data to find specific user behavior.
  • Created an e-mail notification service upon completion of job for the particular team which requested for the data.
  • Defined job work flows as per their dependencies in Oozie.
  • Played a key role in productionizing the application after testing by BI analysts.
  • Involved in code reviews and peer reviews.
  • Maintain System integrity of all sub-components related to Hadoop.
  • Worked with the Data Science team to gather requirements for various data mining projects.
  • Involved in running Hadoop jobs for processing data coming from different sources.

Environment: Apache Hadoop, HDFS, Hive,Pig, Sqoop, HBase, Map Reduce, Java, Cloudera CDH4, Oozie, Flume,DB2,,Maven, shell, Kafka, Eclipse, LINUX.

Confidential, Boston, MA

Hadoop Consultant

Responsibilities:

  • Installed and configured Hadoop, MapReduce, HDFS (Hadoop Distributed File System), developed multiple MapReduce jobs in java for data cleaning and cessing.
  • Developed data pipeline using Flume, Sqoop, Pig and Java MapReduce to ingest customer behavioral data and financial histories into HDFS for analysis.
  • Used Pig as ETL tool to do transformations, event joins and some pre-aggregations before storing the data onto HDFS.
  • Involved in launching and Setup of HADOOP/ HBASE Cluster which includes configuring different components of HADOOP and HBASE Cluster.
  • Installed and configured Cloudera Hadoop on a 100 node cluster.
  • Implemented the workflows using Apache Oozie framework to automate tasks.
  • Developed Pig Latin scripts to extract the data from the web server output files to load into HDFS.
  • Applied MapReduce frameworkjobs in java for data processing by installing and configuring Hadoop, HDFS.
  • Load log data into HDFS using Flume. Worked extensively in creating MapReduce jobs to power data for search and aggregation.
  • Wrote the shell scripts to monitor the health check of Hadoop daemon services and respond accordingly to any warning or failure conditions.
  • Created Hive External tables and loaded the data in to tables and query data using HQL.
  • Developed Pig Latin scripts to extract the data from the web server output files to load into HDFS.
  • Developed PIG Latin scripts to extract the data from the web server output files to load into HDFS.
  • Implemented Fair schedulers on the Job tracker to share the resources of the Cluster for the Map Reduce jobs given by the users.
  • Developed workflow in Oozie to automate the tasks of loading the data into HDFS and pre-processing with Pig.
  • Responsible for architecting Hadoop clusters with CDH3.
  • Involved in writing Flume and Hive scripts to extract, transform and load the data into Database.
  • Importing and exporting data into HDFS and Hive using Sqoop.
  • Created HBase tables to store various data formats of PII data coming from different portfolios.
  • Performed cluster co-ordination through Zookeeper.
  • Involved in creating Hive tables, loading with data and writing hive queries which will run internally in map reduce way.
  • Installed and configured Hive and also written Hive UDFs.
  • Involved in HDFS maintenance and WEBUI it through Hadoop-Java API.
  • Performed data analysis in Hive by creating tables, loading it with data and writing hive queries which will run internally in a MapReduce way.
  • Worked on analyzing Hadoop cluster and different big data analytic tools including Pig, HBase NoSQL database and Sqoop.
  • Extracted files from MongoDB through Sqoop and placed in HDFS and processed.
  • Developed shell script to pull the data from third party system’s into Hadoop file system.
  • Supported in setting up QA environment and updating configurations for implementing scripts with Pig.

Environment: Hadoop, MapReduce, HDFS, Flume, Sqoop, Pig, HBase, Hive, ZooKeeper, Cloudera, Oozie, MongoDB, Sqoop, Kafka, NoSQL, UNIX/LINUX.

Confidential, Bellevue WA

Hadoop Consultant

Responsibilities:

  • Installed/Configured/Maintained Apache Hadoop clusters for application development and Hadoop tools like Hive, Pig, HBase, Flume, Oozie Zookeeper and Sqoop.
  • Extensively involved in Installation and configuration of Cloudera distribution Hadoop 2, 3, NameNode, Secondary NameNode, JobTracker, TaskTrackers and DataNodes.
  • Created POC to store Server Log data in MongoDB to identify System Alert Metrics.
  • Implemented Hadoop framework to capture user navigation across the application to validate the user interface and provide analytic feedback/result to the UI team.
  • Loaded data into the cluster from dynamically generated files using Flume and from relational database management systems using Sqoop.
  • Performed analysis on the unused user navigation data by loading into HDFS and writing MapReduce jobs. The analysis provided inputs to the new APM front end developers and lucent team.
  • Wrote MapReduce jobs using Java API and Pig Latin.
  • Loaded the data from Teradata to HDFS using Teradata Hadoop connectors.
  • Used Flume to collect, aggregate and store the web log data onto HDFS.
  • Wrote Pig scripts to run ETL jobs on the data in HDFS.
  • Used Hive to do analysis on the data and identify different correlations.
  • Involved in HDFS maintenance and administering it through Hadoop-Java API.
  • Worked on importing and exporting data from Oracle and DB2 into HDFS and HIVE using Sqoop.
  • Imported data using Sqoop to load data from MySQL to HDFS on regular basis.
  • Written Hive queries for data analysis to meet the business requirements.
  • Automated all the jobs, for pulling data from FTP server to load data into Hive tables, using Oozie workflows.
  • Involved in creating Hive tables and working on them using Hive QL.
  • Supported Map Reduce Programs those are running on the cluster.
  • Maintaining and monitoring clusters. Loaded data into the cluster from dynamically generated files using Flume and from relational database management systems using Sqoop.
  • Weekly meetings with technical collaborators and active participation in code review sessions with senior and junior developers.

Environment: Hadoop, MapReduce, HDFS, Pig, Hive, HBase, Flume, ZooKeeper, Cloudera, Oozie, Java (jdk1.6), Oracle, PL/SQL, SQL*PLUS, Windows NT, UNIX Shell Scripting.

Confidential, South Farmingdale, NY

PL/SQL Developer

Responsibilities:

  • Involved in creating the requirement analysis and design the system as per the requirement.
  • Created the database objects like Tables, Indexes and Sequences as per requirements.
  • Worked on Database design with the team.
  • Measured the performance of the shell scripts by getting the expected results.
  • Worked on Logical, Physical and conceptual design of Database Data Modeling, and Data Architecture with the use of Erwin as a modeling tool.
  • Wrote PL/SQL code using the technical and functional specifications.
  • Generated SQL and PL/SQL scripts to install create and drop database objects including Tables, Views, Primary Keys, Indexes, Constraints, Sequences and Synonyms.
  • Involved in writing Packages, Functions, Stored Procedures and Database Triggers.
  • Created stored procedures, functions, packages, data base triggers and cursors based on requirement.
  • Developed complex queries to retrieve the data.
  • Created indexes for faster access of data.
  • Documented business rules for writing triggers and constraints, functional and technical design, test cases and user guides.
  • Involved in testing all forms, PL/SQL code for logic correction. Performed Unit testing on queries and reports.
  • Used IN and OUT parameters with TYPE, ROWTYPE, PL/SQL tables and PL/SQL records.
  • Involved in creating Index’s, passing hints, analyzing the table.
  • Built custom forms using Oracle Forms Builder to fulfill the business requirements of the client.
  • Involved in optimizing database performance by analyzing database objects, creating indexes, creating materialized views etc.
  • Wrote UNIX Shell scripts.
  • Worked on Database tuning and performance monitoring.
  • Involved in debugging and error handling.

Environment: Oracle, SQL, PL/SQL, PL/SQL Developer, SQL*Plus, Reports6i, Discoverer 4i, Toad, Ultra edit, UNIX shell script.

Confidential, Louisville, KY

PL/SQL Developer

Responsibilities:

  • Performed analysis, design, development, and testing of UI modules for Unemployment Benefits.
  • Collected user requirement, analyzed and defined business information.
  • Developed several PL/SQL procedures for Unemployment Benefits Claim/Continued Claim/Check processing.
  • Utilized PL/SQL developer tools and TOAD in developing all back end database interfaces.
  • Utilized SQL loader to load data from legacy system to database tables.
  • Performed testing of the database applications.
  • Performed trouble shooting of database related production problems.
  • Coordinated with technical project managers, PL/SQL developers, and DBA team to ensure database performance and reliability.
  • Developed Oracle package to calculate Unemployment Benefits based on the UI law.
  • Wrote procedures for EUC check registers, reprint and void processes.
  • Used Oracle PL/SQL to develop the Overpayment detection and Collection subsystem.
  • Wrote PL/SQL code to generate internal and federal reports.
  • Responsible for all technical aspects of Oracle and application development.

Environment: Oracle 10g/11i, Toad, SQL loader, SQL, PL/SQL, Windows.

Confidential

Java Developer

Responsibilities:

  • Involved in the design and implementation of the architecture for the project using OOAD, UML design patterns.
  • Involved in design and development of server side layer using XML, JSP, JDBC, JNDI, EJB and DAO patterns using eclipse IDE.
  • Work involved extensive usage of HTML, CSS, JavaScript and Ajax for client side development and validations.
  • Used parsers for the conversion of XML files to java objects and vice versa.
  • Developed screens using XML documents and XSL
  • Developed Client programs for consuming the Web services published by the Country Defaults Department which keeps in track of the information regarding life span, inflation rates, retirement age, etc. using Apache Axis.
  • Developed java beans and jsp's by using spring and JSTL tag libs for supplements.
  • Development of EJB’s, Servlets and JSP files for implementing Business rules and Security options using IBM Web Sphere.
  • Involved in creating tables, stored procedures in SQL for data manipulation and retrieval using SQL Server, Oracle and DB2.
  • Trained end users on developed application.

Environment: Java, JSF Framework, Eclipse IDE, Ajax, Apache Axis, OOAD, Web Logic, Java script, HTML, XML, CSS, SQL Server, Oracle, Web services, Ajax, Spring, OOAD and UML, Windows.

Confidential

Java Developer

Responsibilities:

  • Generated everyday reports for data services using SQL commands.
  • Involved in validating and approving new wireless data service products such as Java-On-Mobile application platform and Qualcomm’s Binary Runtime Environment for Wireless (BREW) application server.
  • Led team in proposing and implementing automated content billing system for prepaid subscribers.
  • Successful in proposing and working with the development team to build an automated data billing system for prepaid subscribers.
  • Single point of contact in providing Revenue Assurance clearances.
  • Worked with design team to implement wireless data services reports into Revenue Assurance System.
  • Conducted audit of all the billing packages across the company for postpaid and prepaid subscribers.
  • Ensured that no revenue leakages are prevalent in launch of the new products under wireless data services.

Environment: Java, HTML, XML, SQL Server, MS-Excel, Windows.

We'd love your feedback!