We provide IT Staff Augmentation Services!

Sr. Hadoop/big Data Developer Resume

0/5 (Submit Your Rating)

Columbus, OH

SUMMARY

  • 10+ years of extensive IT experience with multinational clients which includes 4+ years of Hadoop related experience developing Big Data / Hadoop applications.
  • Experience with installation, configuration, supporting and managing of Big Data and underlying infrastructure of Hadoop Cluster.
  • Hands on experience in installing configuring and using Hadoop ecosystem components like HadoopMapReduce, HDFS, HBase, Hive, Sqoop, Pig, Oozie, Zookeeper and Flume with CDH4&5 distributions and EC2 cloud computing.
  • MapR Certified Hadoop Developer.
  • Having 4+ years of experience on Big data Analytics and eco systems.
  • Knowledge in Hive Query Language and PIG scripting for Data Analysis and creating Business reports.
  • Experience in exporting/importing the data using Sqoop and Active Archive from HDFS to Relational Database systems / Mainframe and vice - versa.
  • Worked on writing Unix shell scripts for execution of parallel jobs. Also generated Hive HQL scrips to execute from the shell.
  • Extensive experiance as Data Analyst using DMT tool in Data Warehousing.
  • Experience in SQL and good knowledge in SQL programming and developed Stored Procedures and Triggers. Worked on creating PL/SQL both external and internal procedures and cursors in Oracle (10g/11i).
  • Having experience on Performance tuning of PL/SQL queries.Expertise in Data mapping & data analysis using SQL queries and DMT tool.
  • Used Hadoop to analyse outcome data for generating recommendations, performing predictive analytics, and detecting fraudulent activities.
  • Experience in Data transformation and Data mapping from source to target database schemas and also data cleansing.
  • Excellent knowledge in Data Analysis, Data Validation, Data Cleansing, Data Verification and involved in designing and Preparing Test Scenarios, Test Plans, Test Cases and Test Data.
  • Provided various Re-engineering ideas for optimization of existing Assessment.
  • Experience in preparing Testing Strategies documents and documentation of Test Plans and Test Cases.
  • Experienced in all facets of Software Development Life Cycle (Analysis, Design, Development, Testing and Maintenance) using Waterfall and Agile methodologies.

TECHNICAL SKILLS

Hadoop Ecosystem: Hadoop, MapReduce, HDFS, HIVE, Hbase, Zookeeper, Pig, Sqoop, Flume, Ambari, Hadoop Active Archive, Oozie, Ambari, DBeaver, Squirrel

Programming Languages: Core Java, SQL, PL/SQL, Cobol, JCL, Easytrieve, CICS and Unix

Databases: Oracle 10g/11i, DB2, IMS DB, SQL Server 2000/2005

Tools: SQL Loader, Tivoli, Toad, PVCS, Isql Plus, DMT, Changeman, Expeditor, SPUFI, FileAid, QMF, CA7, Panvalet, Control-M.

Hardware: Hortonworks Hadoop Cluster, IBM Big Insights, IBM System Z and HP Server

Methodologies: Agile and Waterfall

PROFESSIONAL EXPERIENCE

Sr. Hadoop/Big Data Developer

Confidential, Columbus, OH

Responsibilities:

  • Data Analysis of the Claims Legacy Systems and understanding the Confidential ’s business required for Data Mapping
  • Data Lineage of Risk Data elements to its source system and capturing all the calculation/derivation logic involved in underlying Teradata /SQL codes.
  • Developed Map Reduce programs that filter bad and un-necessary records and find out unique records based on different criteria.
  • Developed Secondary sorting implementation to get sorted values at reduce side to improve map reduce performance.
  • Worked on documentation of all Extract, Transform and Load, designed, developed, validated and deploy the Talend ETL processes for Data ware house team using PIG, HIVE on Hortonworks Hadoop.
  • Experience with distributed systems, map reduce systems, data modeling and Big Data systems
  • Responsible for performing extensive data validation using Hive.
  • Implemented Map Reduce programs to classified data organizations into different classifieds based on different type of records.
  • Worked on Sequence files, RC files, Map side joins, bucketing, partitioning for Hive performance enhancement and storage improvement.
  • Implemented Daily Oozi jobs that automate parallel tasks of loading the data into HDFS and pre-processing with Pig using Oozie co-coordinator jobs.
  • Importing and exporting data into HDFS and Hive using Sqoop and Active Archive.
  • Using Active Archive tool scheduling the jobs for Data Ingestion.
  • Perform data analysis using Hive and Pig.
  • Worked intuning Hive and Pig scriptsto improve performance.
  • Knowledge on handling Hive queries using Spark SQL that integrate Spark environment.
  • Involved in submitting and tracking Map Reduce jobs using JobTracker.
  • Involved in loading the created HFiles into HBase for faster access of large customer base without taking Performance hit.
  • Configured build scripts for multi module projects with Maven.
  • Conduct Unit and Integration Testing. Performing code-walkthroughs and reviews.

Environment: HDFS, Map Reduce, Hive, Sqoop, Flume, Zookeeper, Active Archive, HBase, Hadoop, IBM Big Insights, Java, Linux, Oracle 11g/10g, JDK 1.7, Agile, ETL, XML, DBeaver, UNIX Shell Scripting, Ambari, Squirrel, Beeline.

Sr. Hadoop/Big Data Developer

Confidential, Columbus, OH

Responsibilities:

  • Data Analysis of the Donor Bank and understanding the Huntington’s business required for Data Mapping.
  • Developing the transformations programs based on Data Mapping.
  • Written SQL triggers and Stored Procedures.
  • Analyzed the data, which is using the maximum number of resources and made changes in the back-end code using SQL stored procedures and triggers.
  • Generated comprehensive analytical reports by running SQL queries against current databases to conduct data analysis pertaining to Loan products
  • Data Lineage of Risk Data elements to its source system and capturing all the calculation/derivation logic involved in underlying Teradata /SQL codes.
  • Used Cognos reports for the data analysis on Business demand
  • Maintained security and data integrity of the database.
  • Worked with the ETL team to document the transformation rules for data migration from OLTP to Warehouse environment for reporting purposes.
  • Performing code-walkthroughs and reviews.
  • Finding the Business/Process improvements.
  • Involved in writing Unix shell scripts for faster resolution.
  • Support for QA team and UAT.
  • Involved in data ingestion process from SAS to Hive tables.
  • Experience in HDFS, MapReduce and Hadoop Framework.
  • Developed Pig Latin scripts to extract the data from the web server output files to load into HDFS.
  • Moving all the log information into HDFS.
  • Retrieved data using HQL from Hive.
  • Written Map Reduce code to convert semi Structured Data to Structured data.
  • Developed a Framework that will create external and manageable tables in a batch processing based on the metadata files.
  • Scheduling the jobs in Oozie.
  • Involved in migrating the data from development cluster to QA cluster and from there to production cluster.
  • Created the developer Unit test plans and executed unit testing in the development cluster.

Environment: HDFS, Map Reduce, Hive, Sqoop, Flume, Zookeeper, Active Archive, HBase, Hadoop, Hortonworks, Linux, SQL Assistant, SQL Server 2012, Oracle 11g/10g, DMT, JDK 1.7, Agile, ETL, XML, DBeaver, UNIX Shell Scripting, Ambari, Squirrel, Beeline.

Hadoop/Big Data Developer

Confidential, Plantation, FL

Responsibilities:

  • Involved in POC for migration of tape data from Mainframe to Hadoop.
  • Involved in creating the Java API’s to connect to IMS and DB2 databases.
  • Written use case requirements based on business requirements.
  • Developed Map Reduce applications for Risk-Assessment Process.
  • Designed HIVE tables and developed both one time and Delta load of Mainframe database to HIVE tables.
  • Implemented MapReduce classes for processing Bureau file and storing in HBase tables.
  • Implemented process for one-time data storing in HBase tables.
  • Implemented SQL Scripts to modify the data to resolve assigned defects.
  • Performance tuning of the database and SQL queries.
  • Tuned SQL queries and indexed database tables wherever necessary and also partitioned databases
  • Developing the transformations programs based on Data Mapping.
  • Involved in migrating the data from development cluster to QA cluster and from there to production cluster.
  • Created the developer Unit test plans and executed unit testing in the development cluster.

Environment: HDFS, Map Reduce, Hive, Sqoop, Zookeeper, HBase, Hadoop, MapR, Linux, SQL Server 2012, Oracle 11g/10g, Agile, ETL, XML, UNIX Shell Scripting, Ambari, Corner Stone, Drools, IBM Mainframe, IMS DB, COBOL, JCL, DB2, Changeman, Expeditor, SPUFI, FileAid, QMF, Control-M, Easytrieve, CICS.

Data Architect

Confidential, Parsippany, NJ

Responsibilities:

  • Analysis and design of the Enhancements.
  • Involved in requirement gathering along with the business analysts group.
  • Gathered all the Sales Analysis report prototypes from the business analysts belonging to different Business units
  • Prepare Analysis and High Level Design for the Change Requests and Enhancements. Design and Code Documentation.
  • Documenting the Enhancements and new Business Functionalities.
  • Adhoc analysis based on business requirements and handling client queries.
  • Conduct Unit and Integration Testing.
  • Daily interactions with Client for Reporting/resolutions of Tickets.
  • Support for QA team and UAT.
  • Coordinate and Communicate between customers and offshore team.
  • Performing code-walkthroughs and reviews. Finding the business/process improvements. Preparing Review Checklist.
  • Preparing the code fixes to reduce the Production Job abends
  • Implemented SQL Scripts to modify the data to resolve assigned defects.
  • Performance tuning of the database and SQL queries.
  • Process improvement of batch jobs in Production.
  • Tuned DB2 queries and indexed database tables wherever necessary and also partitioned databases
  • Developing the transformations programs based on Data Mapping.
  • Created the developer Unit test plans and executed unit testing in the development cluster.

Environment: IBM Mainframe, IMS DB, Cobol, JCL, DB2, Changeman, Expeditor, SPUFI, FileAid, QMF, Control-M, Easytrieve, CICS, VSA, TSO/ISPFM, CA7, PANVALET.

Mainframe Developer

Confidential

Responsibilities:

  • Analysis and design of the Enhancements.
  • Involved in requirement gathering along with the business analysts group.
  • Gathered all the Sales Analysis report prototypes from the business analysts belonging to different Business units
  • Prepare Analysis and High Level Design for the Change Requests and Enhancements. Design and Code Documentation.
  • Understanding the business problems and documenting them.
  • Converting Business problems into Functional requirements.
  • Converting the Functional requirements into Technical requirements.
  • Analysis and design of the new requirements.
  • Design and Code Documentation.
  • Development of programs in order to meet the Business requirements.
  • Conduct Unit and Integration Testing.
  • Performing code-walkthroughs and reviews.
  • Finding the Business/Process improvements.
  • Support for QA team and UAT.
  • Production Support with handling P1 issues.
  • Coordinate and Communicate between customers and team.

Environment: MF Cobol, Pro-Cobol Oracle 10g/11i, Unix Shell Scripting, PVCS, ISQL Plus, TOAD, HP - UX11i, Windows - 2000/XP.

We'd love your feedback!