Lead Hadoop Developer Resume
Atlanta, GA
SUMMARY:
- Result oriented, proactive and industrious professional with eleven (11) years of hands on experience in the field of data warehousing in big data, Spark, Spark SQL, Spark Streaming, Kafka, Scala, Python, AWS (EMR, S3, EC2, Lambda), Cloudera Hadoop, HDFS, Hive, Pig, Impala, MySQL, UNIX, Teradata, Informatica. Strong knowledge in data warehouse methodology and have end - to-end experience in data warehouse projects, core experience in developing Full Life Cycle data warehousing projects, which includes requirement gathering, analysis, design, development, testing, implementation and maintenance of data warehouse using ETL. Strong in writing UNIX shell scripts, good knowledge in RDBMS concepts and writing SQL queries
- Core working experience in data warehousing projects with big data Hadoop in Banking and payment domains
- Strong knowledge in agile, V models and Waterfall models
- Extensive experience in interaction with clients and functional people for gathering the business requirements and functional specifications
- Enthusiastic and adaptive to new technologies, working well under pressure and communicating ideas clearly and effectively. Good communication skills, self-motivated and organized in delivering high quality software solutions and Adroit at learning new concepts quickly
- Dedicated and highly ambitious to achieve personal goals as well as the organizational goals
- Experience in banking and payment, insurance and retail domains
- Implemented ETL in Hadoop using Spark, Hive, Kafka and Scoop
- Developed Teradata BTEQ scripts and TPT utility to Extract, Transform and Load the data
- Accountable for ETL design, performance truing on Teradata platform, hands-on development in big data platform and well experienced in agile and Waterfall methodology
- Excellent communication and interpreted the business needs to technical terms
- Add-on hands-on on multi-platform and technologies, proficient in Scala, Python and Spark
- Expert in DB tools and SQL concepts
- Managed the teams along with individual project delivery as per the needs of the project
TECHNICAL SKILLS:
Operating Systems: Windows2000/NT/7, UNIX
Languages: SQL, Scala, Python, HQL, Spark, Unix Korn Shell, C/C++
Databases: AWS - S3, HDFS, Informatica, MySQL, Teradata
Tools: & Utilities: AWS, Spark, Hadoop, HIVE, Sqoop, Kafka, flume, Pig
Domain Knowledge: Banking and payments, insurance and retails
PROFESSIONAL EXPERIENCE:
Lead Hadoop Developer
Confidential, Atlanta, GA
Responsibilities:
- Implemented ETL using Spark, Kafka, Scala, Python, AWS (EMR, S3, EC2, Lambda), Hadoop, HDFS, Hive, Pig, Impala, MySQL, JIRA.
- Developed Auto file trigger for AWS and Oozie workflow scheduling for Hadoop HDFS batch processing.
- Developed HIVE SQL queries to generate ad-hoc reports for the business analysts to implement complex business rules.
- Work as Development Lead and is responsible for ETL Design, helping developers in developing efficient code, ensure coding standards are followed, perform code reviews, delivery of quality code, meeting the timelines for development, perform end-to-end testing and provide weekly status on projects to management
- Interacting with clients, business analysts, end user to get business requirements from them, providing estimates, Translating business requirements into equivalent technical requirements, Designing and Implementing a solution for the same, coordinating with offshore to ensure the defect free deliverables
- Resolve major Performance Issues in production, and prepared performance tuning guidelines to meet SLA and coding standards
- Exchanged data between S3 and Hadoop HDFS.
- Adapted agile scrum methodology
Environment: Spark, Spark SQL, Spark Streaming, Kafka, Scala, Python, AWS (EMR, S3, EC2, Lambda), Hadoop, HDFS, Hive, Pig, Impala, MySQL, JIRA, UNIX, Windows
Sr. Informatica Developer
Confidential
Responsibilities:
- Helped solution architects to design the process, performed the EDW impact analysis, Gap fit report, maintained the data lineage doc, ETL spec creation, mappings, BTEQ script and UNIX shell as per the requirements
- Tested creations bases on the requirements and maintained the RTM tracker. Coordinated with other module leads, aligned the code versioning, code base-lining and retrofit
- Helped UAT on test planning, test scope estimations, test execution and defect management
- Coordinated between the offshore and the onshore teams for the timely deliverables
Environment: UNIX, Informatica, Teradata, Windows
Module Lead
Confidential
Responsibilities:
- Gathered the information and analyzed the business and technical requirements
- ETL spec creation and mappings as per the requirements
- Developed BTEQ scripts and TPT utility to Extract, Transform and Load data
- Lead Scrum teams and had responsibilities to delivery different sprint goals for “Domain Page Stats”
- Developed shell scripts
Environment: UNIX, Informatica, Oracle, Windows
Software Engineer
Confidential
Responsibilities:
- Good understanding of the requirements and functionalities
- Developed ETL mappings, BTEQ scripts and UNIX shell scripts as per the requirements
- Test creation based on the requirements and maintained the RTM tracker
- Handled defect management and prepared migration docs and back-out plans
Environment: UNIX, Informatica, Netezza, DB2, Windows
