We provide IT Staff Augmentation Services!

Big Data Developer Resume

2.00/5 (Submit Your Rating)

CharlottE

SUMMARY

  • Around 8.2 years of experience in all phases of SDLC including application design, development, support, maintenance, testing, change request management and Enhancement support to the client.
  • Having 4 years of experience on BigData Hadoop major echo systems Hdfs, Map Reduce, Sqoop, Apache Hive, Apache Pig, HBase, Apache Spark and Cassandra with all distributed systems.
  • Hands on experience in programming Scala and Python language for data analytics.
  • Expertise in Cloudera CDH3, CDH4, Hortonworks, MapR and IBM BIG Insights Info Sphere Tools.
  • Experience in Setup cluster installing and configuring ecosystem like Hdfs, Pig, Hbase, Sqoop, Hive, Ambari and Spark single and multinode cluster with Apache, Clodera(CDH), IBM BIGInsights and Hortonworks(HDP).
  • Developed data pipelines to import/export datasets using Sqoop to move data in and out of the Hadoop.
  • Extensive experience in creating and schedule big data workflows in Spark, MapReduce, Pig scripts.
  • Developed big data Oozie workflows to process very large datasets with using hadoop eco systems.
  • Experience on Oracle11g and Informatica for data integration in developing ETL process implementation
  • Expert and Certified on IBM BIG Insights InfoSphere Tools for distributed file system.
  • Created Spark jobs for migrating vertica data base to oracle data base with using spark sql and Vertica.
  • Set up Oozie workflows to run data validation check jobs and placed into hdfs for next curation process.
  • Performed file control checks using UNIX script from LFS location using Autosys jobs and the files will be placed in HDFS landing zone once file control check is successful.
  • Developed proof of concepts for enterprise adoption of Hadoop and related tools for business analytics.
  • Expert and Certified on IBM Netteza and good exposure on Netteza distributed file system.
  • Good knowledge and expertise in Banking, Financial, Health Insurance and Telecom Domains.
  • Experience on Mainframe large scale applications and their Maintenance and Enhancements using DB2, COBOL, JCL and CICS screens are developed for financial domain in Transunion implementation.
  • Trained on Virtual networks components Open Stack, Glance, Cinder, Keystone, Neutron, Devstack.
  • Worked in Agile process, Participated in Project Requirement and planning meetings with the customers.
  • Experience in performing Peer & Code reviews, unit test cases preparation and documentation.
  • Exceptional ability to learn new technologies and to deliver outputs in short deadlines based on business.
  • Experience in schedule jobs and created for automated process in Autosys, $Universe schedule tools.
  • Excellent verbal, written communication and interpersonal skills and manage different environments.
  • Experienced in trained and mentoring teams with functional knowledge and business processes.
  • Excellent Customer interaction skills and project coordination skills with onsite/offshore model.
  • A Strong team player having Good analytical skills to identify key issues and provide solution, design of Technical Specification document as per schedule.

TECHNICAL SKILLS

Big Data Ecosystem: HDFS, MapReduce, Hive, Pig, Sqoop, HBase, Oozie, Spark, Knox, Storm, Kafka, HCatalog, Ambari, Apache Kylin, AtScale and Cassandra.

Operating Systems: Windows, RedHat Linux, CentOS, Ubuntu

Programming/ Scripting Languages: Scala(2.10), Java, Python, Shell Scripting, Cobol

Databases/Database Languages: MySQL, Oracle 9i/11g, NoSQL (HBase, Cassandra), SQL, DB2

Web Technologies: HTML, XML, JSON

Tools: Autosys, $Universe, Zeppelin Notebook, Traffodian, EsgynDB Manager, IBM Big Info Sphere, HDP

IDEs: Eclipse, intellij, NetBeans, Toad, Rapid SQL

Version Tools: SVN, GitHub, Win CVS

PROFESSIONAL EXPERIENCE

Confidential, Charlotte

Big Data Developer

Responsibilities:

  • Created a design mapping for transactional data from different sources.
  • Pulling required data from 8 major sources in boa like Cardhub, Bntol, Impacs, Apte, Teller OSI, Gasper, V3 and Safebox using Sqoop and NDM.
  • Written Sqoop scripts to pull the data from teradata tables using oozie work flow.
  • Performed file control checks using UNIX script from LFS location using Autosys jobs and the files will be placed in HDFS landing zone once file control check is successful.
  • Created and implemented spark jobs for data control checks using Scala.
  • Set up Oozie workflows to run the data validation check jobs and placed into hdfs for next curation process.
  • Identified lineage from sources and joined the data for curation process.
  • Implemented the process for Curation and loaded into Parquet files.
  • Performed the process for aggregation on top of curated data.
  • Created autosys jobs for data control, Curation and aggregation process on daily basis.
  • Involved and written in Unite test cases in Scala.

Environment: Spark1.6.1, Scala 2.10, HDP2, Oozie, Python, Sqoop

Confidential

Data Analyst

Responsibilities:

  • Prepared the design process for spark jobs in Intellij environment and setup the environment.
  • Conducted daily meetings and gathered requirements for the preparation of spark template.
  • Created 125 Spark jobs in java for migrating vertica data base to oracle for each and every table loading.
  • Prepared for SPARKRDD to JDBCRDD for getting the data from oracle data and process for partition data.
  • Running spark jobs and monitoring data loads manually and tested with UC4 schedule tool.
  • Tested and compare the data in both Vertica table data and Oracle tables.
  • Schedule the spark jobs using UC4 job scheduler

Environment: HDP 2.2, Spark 1.4.1, UC4, HP Vertica, Oracle

Confidential

Hadoop Developer

Responsibilities:

  • Prepared the design documents for ETL process of sql jobs
  • Implemented the SQL Scripts and loaders to retrieve the data from files.
  • Prepared the informatica workflows for loading data.
  • Test the tables in different environments Dev and Testing phases.
  • Created the Unit test cases and documents.
  • Prepared the Schedule trigger jobs.

Environment: Oracle 11g SQL, Informatica

Confidential

Hadoop Developer

Responsibilities:

  • Prepared the design documents for sql jobs
  • Implemented the SQL Scripts and loaders to retrieve the data from files.
  • Created Unit test cases and Unit test documents.
  • Mentored the jobs in advanced SQL concepts

Environment: Oracle 11g SQL, InformaticaWebLog

Confidential

Developer

Responsibilities:

  • Built a Text Extractor using AQL(Annotation Query Language), published and deployed it as an application on IBM Big Insights.
  • Developed an application in JAQL to process the output from the text extractor.
  • Developed Reports Using Bigsheets.

Environment: IBMBiginsights2.1, MapReduce, hive, Eclipse Helios 3.6

Confidential

Hadoop Developer

Responsibilities:

  • Analyze the requirements and making design.
  • Capturing data from DB2 and store into hdfs.
  • Created heap analysis of watchtower data using MapReduce.
  • Built in Pig Scripts on top of Hadoop data and making time interval analysis for alerts generating system.
  • Exporting data to visualization purpose.

Environment: Hadoop - Hdfs, Apache Pig, Sqoop, Hive, and Map Reduce, Cassandra and DB2.

Confidential

Associate consultant

Responsibilities:

  • Analyzing business requirements and preparation of program specifications accordingly.
  • Writing new batch, Jobs and Procs for imparting new business requirements.
  • Preparation of Pre Design, DTD and UTP process
  • Coding and Unit Testing documents preparation for each module.
  • Worked on Production Assurance testing support.
  • Pre and post production Implementation support.
  • Involved in production support activities, resolving production IM tickets.
  • Working on client queries, and ado requests.

Environment: Mainframe, COBOL, JCL and DB2

Confidential

Software Developer

Responsibilities:

  • Requirement Analysis for CICS screens preparation for the module.
  • Prepared Functional and Design documents for flow of Billing Module.
  • Created each and flow of coding, code review and documentation.
  • Prepared Test cases and Unit testing documents for the CICS billing screens.
  • Handling change requests. Completion of service requests with in time.
  • Online and batch change requests as per the client requirements.

We'd love your feedback!