Big Data Developer Resume
CharlottE
SUMMARY
- Around 8.2 years of experience in all phases of SDLC including application design, development, support, maintenance, testing, change request management and Enhancement support to the client.
- Having 4 years of experience on BigData Hadoop major echo systems Hdfs, Map Reduce, Sqoop, Apache Hive, Apache Pig, HBase, Apache Spark and Cassandra with all distributed systems.
- Hands on experience in programming Scala and Python language for data analytics.
- Expertise in Cloudera CDH3, CDH4, Hortonworks, MapR and IBM BIG Insights Info Sphere Tools.
- Experience in Setup cluster installing and configuring ecosystem like Hdfs, Pig, Hbase, Sqoop, Hive, Ambari and Spark single and multinode cluster with Apache, Clodera(CDH), IBM BIGInsights and Hortonworks(HDP).
- Developed data pipelines to import/export datasets using Sqoop to move data in and out of the Hadoop.
- Extensive experience in creating and schedule big data workflows in Spark, MapReduce, Pig scripts.
- Developed big data Oozie workflows to process very large datasets with using hadoop eco systems.
- Experience on Oracle11g and Informatica for data integration in developing ETL process implementation
- Expert and Certified on IBM BIG Insights InfoSphere Tools for distributed file system.
- Created Spark jobs for migrating vertica data base to oracle data base with using spark sql and Vertica.
- Set up Oozie workflows to run data validation check jobs and placed into hdfs for next curation process.
- Performed file control checks using UNIX script from LFS location using Autosys jobs and the files will be placed in HDFS landing zone once file control check is successful.
- Developed proof of concepts for enterprise adoption of Hadoop and related tools for business analytics.
- Expert and Certified on IBM Netteza and good exposure on Netteza distributed file system.
- Good knowledge and expertise in Banking, Financial, Health Insurance and Telecom Domains.
- Experience on Mainframe large scale applications and their Maintenance and Enhancements using DB2, COBOL, JCL and CICS screens are developed for financial domain in Transunion implementation.
- Trained on Virtual networks components Open Stack, Glance, Cinder, Keystone, Neutron, Devstack.
- Worked in Agile process, Participated in Project Requirement and planning meetings with the customers.
- Experience in performing Peer & Code reviews, unit test cases preparation and documentation.
- Exceptional ability to learn new technologies and to deliver outputs in short deadlines based on business.
- Experience in schedule jobs and created for automated process in Autosys, $Universe schedule tools.
- Excellent verbal, written communication and interpersonal skills and manage different environments.
- Experienced in trained and mentoring teams with functional knowledge and business processes.
- Excellent Customer interaction skills and project coordination skills with onsite/offshore model.
- A Strong team player having Good analytical skills to identify key issues and provide solution, design of Technical Specification document as per schedule.
TECHNICAL SKILLS
Big Data Ecosystem: HDFS, MapReduce, Hive, Pig, Sqoop, HBase, Oozie, Spark, Knox, Storm, Kafka, HCatalog, Ambari, Apache Kylin, AtScale and Cassandra.
Operating Systems: Windows, RedHat Linux, CentOS, Ubuntu
Programming/ Scripting Languages: Scala(2.10), Java, Python, Shell Scripting, Cobol
Databases/Database Languages: MySQL, Oracle 9i/11g, NoSQL (HBase, Cassandra), SQL, DB2
Web Technologies: HTML, XML, JSON
Tools: Autosys, $Universe, Zeppelin Notebook, Traffodian, EsgynDB Manager, IBM Big Info Sphere, HDP
IDEs: Eclipse, intellij, NetBeans, Toad, Rapid SQL
Version Tools: SVN, GitHub, Win CVS
PROFESSIONAL EXPERIENCE
Confidential, Charlotte
Big Data Developer
Responsibilities:
- Created a design mapping for transactional data from different sources.
- Pulling required data from 8 major sources in boa like Cardhub, Bntol, Impacs, Apte, Teller OSI, Gasper, V3 and Safebox using Sqoop and NDM.
- Written Sqoop scripts to pull the data from teradata tables using oozie work flow.
- Performed file control checks using UNIX script from LFS location using Autosys jobs and the files will be placed in HDFS landing zone once file control check is successful.
- Created and implemented spark jobs for data control checks using Scala.
- Set up Oozie workflows to run the data validation check jobs and placed into hdfs for next curation process.
- Identified lineage from sources and joined the data for curation process.
- Implemented the process for Curation and loaded into Parquet files.
- Performed the process for aggregation on top of curated data.
- Created autosys jobs for data control, Curation and aggregation process on daily basis.
- Involved and written in Unite test cases in Scala.
Environment: Spark1.6.1, Scala 2.10, HDP2, Oozie, Python, Sqoop
Confidential
Data Analyst
Responsibilities:
- Prepared the design process for spark jobs in Intellij environment and setup the environment.
- Conducted daily meetings and gathered requirements for the preparation of spark template.
- Created 125 Spark jobs in java for migrating vertica data base to oracle for each and every table loading.
- Prepared for SPARKRDD to JDBCRDD for getting the data from oracle data and process for partition data.
- Running spark jobs and monitoring data loads manually and tested with UC4 schedule tool.
- Tested and compare the data in both Vertica table data and Oracle tables.
- Schedule the spark jobs using UC4 job scheduler
Environment: HDP 2.2, Spark 1.4.1, UC4, HP Vertica, Oracle
Confidential
Hadoop Developer
Responsibilities:
- Prepared the design documents for ETL process of sql jobs
- Implemented the SQL Scripts and loaders to retrieve the data from files.
- Prepared the informatica workflows for loading data.
- Test the tables in different environments Dev and Testing phases.
- Created the Unit test cases and documents.
- Prepared the Schedule trigger jobs.
Environment: Oracle 11g SQL, Informatica
Confidential
Hadoop Developer
Responsibilities:
- Prepared the design documents for sql jobs
- Implemented the SQL Scripts and loaders to retrieve the data from files.
- Created Unit test cases and Unit test documents.
- Mentored the jobs in advanced SQL concepts
Environment: Oracle 11g SQL, InformaticaWebLog
Confidential
Developer
Responsibilities:
- Built a Text Extractor using AQL(Annotation Query Language), published and deployed it as an application on IBM Big Insights.
- Developed an application in JAQL to process the output from the text extractor.
- Developed Reports Using Bigsheets.
Environment: IBMBiginsights2.1, MapReduce, hive, Eclipse Helios 3.6
Confidential
Hadoop Developer
Responsibilities:
- Analyze the requirements and making design.
- Capturing data from DB2 and store into hdfs.
- Created heap analysis of watchtower data using MapReduce.
- Built in Pig Scripts on top of Hadoop data and making time interval analysis for alerts generating system.
- Exporting data to visualization purpose.
Environment: Hadoop - Hdfs, Apache Pig, Sqoop, Hive, and Map Reduce, Cassandra and DB2.
Confidential
Associate consultant
Responsibilities:
- Analyzing business requirements and preparation of program specifications accordingly.
- Writing new batch, Jobs and Procs for imparting new business requirements.
- Preparation of Pre Design, DTD and UTP process
- Coding and Unit Testing documents preparation for each module.
- Worked on Production Assurance testing support.
- Pre and post production Implementation support.
- Involved in production support activities, resolving production IM tickets.
- Working on client queries, and ado requests.
Environment: Mainframe, COBOL, JCL and DB2
Confidential
Software Developer
Responsibilities:
- Requirement Analysis for CICS screens preparation for the module.
- Prepared Functional and Design documents for flow of Billing Module.
- Created each and flow of coding, code review and documentation.
- Prepared Test cases and Unit testing documents for the CICS billing screens.
- Handling change requests. Completion of service requests with in time.
- Online and batch change requests as per the client requirements.
