Sr. Hadoop/big Data Developer Resume
Columbus, OH
SUMMARY
- 10+ years of extensive IT experience with multinational clients which includes 4+ years of Hadoop related experience developing Big Data / Hadoop applications.
- Experience with installation, configuration, supporting and managing of Big Data and underlying infrastructure of Hadoop Cluster.
- Hands on experience in installing configuring and using Hadoop ecosystem components like HadoopMapReduce, HDFS, HBase, Hive, Sqoop, Pig, Oozie, Zookeeper and Flume with CDH4&5 distributions and EC2 cloud computing.
- MapR Certified Hadoop Developer.
- Having 4+ years of experience on Big data Analytics and eco systems.
- Knowledge in Hive Query Language and PIG scripting for Data Analysis and creating Business reports.
- Experience in exporting/importing the data using Sqoop and Active Archive from HDFS to Relational Database systems / Mainframe and vice - versa.
- Worked on writing Unix shell scripts for execution of parallel jobs. Also generated Hive HQL scrips to execute from the shell.
- Extensive experiance as Data Analyst using DMT tool in Data Warehousing.
- Experience in SQL and good knowledge in SQL programming and developed Stored Procedures and Triggers. Worked on creating PL/SQL both external and internal procedures and cursors in Oracle (10g/11i).
- Having experience on Performance tuning of PL/SQL queries.Expertise in Data mapping & data analysis using SQL queries and DMT tool.
- Used Hadoop to analyse outcome data for generating recommendations, performing predictive analytics, and detecting fraudulent activities.
- Experience in Data transformation and Data mapping from source to target database schemas and also data cleansing.
- Excellent knowledge in Data Analysis, Data Validation, Data Cleansing, Data Verification and involved in designing and Preparing Test Scenarios, Test Plans, Test Cases and Test Data.
- Provided various Re-engineering ideas for optimization of existing Assessment.
- Experience in preparing Testing Strategies documents and documentation of Test Plans and Test Cases.
- Experienced in all facets of Software Development Life Cycle (Analysis, Design, Development, Testing and Maintenance) using Waterfall and Agile methodologies.
TECHNICAL SKILLS
Hadoop Ecosystem: Hadoop, MapReduce, HDFS, HIVE, Hbase, Zookeeper, Pig, Sqoop, Flume, Ambari, Hadoop Active Archive, Oozie, Ambari, DBeaver, Squirrel
Programming Languages: Core Java, SQL, PL/SQL, Cobol, JCL, Easytrieve, CICS and Unix
Databases: Oracle 10g/11i, DB2, IMS DB, SQL Server 2000/2005
Tools: SQL Loader, Tivoli, Toad, PVCS, Isql Plus, DMT, Changeman, Expeditor, SPUFI, FileAid, QMF, CA7, Panvalet, Control-M.
Hardware: Hortonworks Hadoop Cluster, IBM Big Insights, IBM System Z and HP Server
Methodologies: Agile and Waterfall
PROFESSIONAL EXPERIENCE
Sr. Hadoop/Big Data Developer
Confidential, Columbus, OH
Responsibilities:
- Data Analysis of the Claims Legacy Systems and understanding the Confidential ’s business required for Data Mapping
- Data Lineage of Risk Data elements to its source system and capturing all the calculation/derivation logic involved in underlying Teradata /SQL codes.
- Developed Map Reduce programs that filter bad and un-necessary records and find out unique records based on different criteria.
- Developed Secondary sorting implementation to get sorted values at reduce side to improve map reduce performance.
- Worked on documentation of all Extract, Transform and Load, designed, developed, validated and deploy the Talend ETL processes for Data ware house team using PIG, HIVE on Hortonworks Hadoop.
- Experience with distributed systems, map reduce systems, data modeling and Big Data systems
- Responsible for performing extensive data validation using Hive.
- Implemented Map Reduce programs to classified data organizations into different classifieds based on different type of records.
- Worked on Sequence files, RC files, Map side joins, bucketing, partitioning for Hive performance enhancement and storage improvement.
- Implemented Daily Oozi jobs that automate parallel tasks of loading the data into HDFS and pre-processing with Pig using Oozie co-coordinator jobs.
- Importing and exporting data into HDFS and Hive using Sqoop and Active Archive.
- Using Active Archive tool scheduling the jobs for Data Ingestion.
- Perform data analysis using Hive and Pig.
- Worked intuning Hive and Pig scriptsto improve performance.
- Knowledge on handling Hive queries using Spark SQL that integrate Spark environment.
- Involved in submitting and tracking Map Reduce jobs using JobTracker.
- Involved in loading the created HFiles into HBase for faster access of large customer base without taking Performance hit.
- Configured build scripts for multi module projects with Maven.
- Conduct Unit and Integration Testing. Performing code-walkthroughs and reviews.
Environment: HDFS, Map Reduce, Hive, Sqoop, Flume, Zookeeper, Active Archive, HBase, Hadoop, IBM Big Insights, Java, Linux, Oracle 11g/10g, JDK 1.7, Agile, ETL, XML, DBeaver, UNIX Shell Scripting, Ambari, Squirrel, Beeline.
Sr. Hadoop/Big Data Developer
Confidential, Columbus, OH
Responsibilities:
- Data Analysis of the Donor Bank and understanding the Huntington’s business required for Data Mapping.
- Developing the transformations programs based on Data Mapping.
- Written SQL triggers and Stored Procedures.
- Analyzed the data, which is using the maximum number of resources and made changes in the back-end code using SQL stored procedures and triggers.
- Generated comprehensive analytical reports by running SQL queries against current databases to conduct data analysis pertaining to Loan products
- Data Lineage of Risk Data elements to its source system and capturing all the calculation/derivation logic involved in underlying Teradata /SQL codes.
- Used Cognos reports for the data analysis on Business demand
- Maintained security and data integrity of the database.
- Worked with the ETL team to document the transformation rules for data migration from OLTP to Warehouse environment for reporting purposes.
- Performing code-walkthroughs and reviews.
- Finding the Business/Process improvements.
- Involved in writing Unix shell scripts for faster resolution.
- Support for QA team and UAT.
- Involved in data ingestion process from SAS to Hive tables.
- Experience in HDFS, MapReduce and Hadoop Framework.
- Developed Pig Latin scripts to extract the data from the web server output files to load into HDFS.
- Moving all the log information into HDFS.
- Retrieved data using HQL from Hive.
- Written Map Reduce code to convert semi Structured Data to Structured data.
- Developed a Framework that will create external and manageable tables in a batch processing based on the metadata files.
- Scheduling the jobs in Oozie.
- Involved in migrating the data from development cluster to QA cluster and from there to production cluster.
- Created the developer Unit test plans and executed unit testing in the development cluster.
Environment: HDFS, Map Reduce, Hive, Sqoop, Flume, Zookeeper, Active Archive, HBase, Hadoop, Hortonworks, Linux, SQL Assistant, SQL Server 2012, Oracle 11g/10g, DMT, JDK 1.7, Agile, ETL, XML, DBeaver, UNIX Shell Scripting, Ambari, Squirrel, Beeline.
Hadoop/Big Data Developer
Confidential, Plantation, FL
Responsibilities:
- Involved in POC for migration of tape data from Mainframe to Hadoop.
- Involved in creating the Java API’s to connect to IMS and DB2 databases.
- Written use case requirements based on business requirements.
- Developed Map Reduce applications for Risk-Assessment Process.
- Designed HIVE tables and developed both one time and Delta load of Mainframe database to HIVE tables.
- Implemented MapReduce classes for processing Bureau file and storing in HBase tables.
- Implemented process for one-time data storing in HBase tables.
- Implemented SQL Scripts to modify the data to resolve assigned defects.
- Performance tuning of the database and SQL queries.
- Tuned SQL queries and indexed database tables wherever necessary and also partitioned databases
- Developing the transformations programs based on Data Mapping.
- Involved in migrating the data from development cluster to QA cluster and from there to production cluster.
- Created the developer Unit test plans and executed unit testing in the development cluster.
Environment: HDFS, Map Reduce, Hive, Sqoop, Zookeeper, HBase, Hadoop, MapR, Linux, SQL Server 2012, Oracle 11g/10g, Agile, ETL, XML, UNIX Shell Scripting, Ambari, Corner Stone, Drools, IBM Mainframe, IMS DB, COBOL, JCL, DB2, Changeman, Expeditor, SPUFI, FileAid, QMF, Control-M, Easytrieve, CICS.
Data Architect
Confidential, Parsippany, NJ
Responsibilities:
- Analysis and design of the Enhancements.
- Involved in requirement gathering along with the business analysts group.
- Gathered all the Sales Analysis report prototypes from the business analysts belonging to different Business units
- Prepare Analysis and High Level Design for the Change Requests and Enhancements. Design and Code Documentation.
- Documenting the Enhancements and new Business Functionalities.
- Adhoc analysis based on business requirements and handling client queries.
- Conduct Unit and Integration Testing.
- Daily interactions with Client for Reporting/resolutions of Tickets.
- Support for QA team and UAT.
- Coordinate and Communicate between customers and offshore team.
- Performing code-walkthroughs and reviews. Finding the business/process improvements. Preparing Review Checklist.
- Preparing the code fixes to reduce the Production Job abends
- Implemented SQL Scripts to modify the data to resolve assigned defects.
- Performance tuning of the database and SQL queries.
- Process improvement of batch jobs in Production.
- Tuned DB2 queries and indexed database tables wherever necessary and also partitioned databases
- Developing the transformations programs based on Data Mapping.
- Created the developer Unit test plans and executed unit testing in the development cluster.
Environment: IBM Mainframe, IMS DB, Cobol, JCL, DB2, Changeman, Expeditor, SPUFI, FileAid, QMF, Control-M, Easytrieve, CICS, VSA, TSO/ISPFM, CA7, PANVALET.
Mainframe Developer
Confidential
Responsibilities:
- Analysis and design of the Enhancements.
- Involved in requirement gathering along with the business analysts group.
- Gathered all the Sales Analysis report prototypes from the business analysts belonging to different Business units
- Prepare Analysis and High Level Design for the Change Requests and Enhancements. Design and Code Documentation.
- Understanding the business problems and documenting them.
- Converting Business problems into Functional requirements.
- Converting the Functional requirements into Technical requirements.
- Analysis and design of the new requirements.
- Design and Code Documentation.
- Development of programs in order to meet the Business requirements.
- Conduct Unit and Integration Testing.
- Performing code-walkthroughs and reviews.
- Finding the Business/Process improvements.
- Support for QA team and UAT.
- Production Support with handling P1 issues.
- Coordinate and Communicate between customers and team.
Environment: MF Cobol, Pro-Cobol Oracle 10g/11i, Unix Shell Scripting, PVCS, ISQL Plus, TOAD, HP - UX11i, Windows - 2000/XP.
