We provide IT Staff Augmentation Services!

Talend Etl Developer / Administrator Resume

2.00/5 (Submit Your Rating)

NJ

SUMMARY:

  • 7+ years of experience in full life cycle of software project development in various areas like design, Applications development of Enterprise Data Warehouse on large scale development efforts using best practices with Talend(7.x/6.x/5.x), Ab Initio and SSIS
  • 4+ years of IT experience with using Talend Data Integration/Big Data Integration tools
  • Extensive knowledge of various business domains Health Care, Mortgage, Financial and Retail sectors
  • Extensive experience working in AGILE team
  • Expertise in Talend integration components like tMap, tJoin, tReplicate, tParallelize, tConvertType,, tflowtoIterate, tAggregate, tSortRow, tFlowMeter, tLogCatcher, tRowGenerator, tNormalize, tDenormalize, tSetGlobalVar, tHashInput, tHashOutput, tJava, tJavarow, tAggregateRow, tWarn, tLogCatcher, tMysqlScd, tFilter, tGlobalmap etc
  • Well versed with Talend Big data components like tHDFSInput, tHDFSOutput, tPigLoad, tPigFilterRow, tPigFilterColumn, tPigStoreResult, tHiveLoad, tHiveInput, tHbaseInput, tHbaseOutput, tSqoopImport and tSqoopExport
  • Expertise in using context variables, Routines and metedata.
  • Integrated Talend jobs with AWS SNS Jason topics
  • Spinning up AWS EC2 instance with auto scheduling with Lambda
  • CICD integration and code promotion with Jinkins
  • Experience with On - Call Support to Production Systems
  • Good Knowledge in UNIX shell scripting
  • Hands on experience working on RESTful API
  • Talend code migration and environment setup
  • Talend Project creation, Branch creation to support multiple versions in GIT and enabling access controls
  • Talend jobs deployment in TAC and user support
  • Installing talend on on remote engine.
  • Build and Deploy the Talend jobs via Jenkins CI/CD process
  • ETL server maintenance and oncall support
  • Public cloud enablement - Talend on AWS
  • Extensive experience in utilizing ETL Process for designing and building very large-scale data OLAP and OLTP instance using ETL tools Ab-Initio and custom scripts
  • Experienced in Ab Initio parallelism techniques and implemented Ab Initio graphs using Data parallelism and Multi File System (MFS) techniques.
  • Designed and Developed the graphs using GDE, with components partition by round robin, partition by key, rollup, sort, scan, dedup sort, reformat, join, merge, gather and concatenate components, filter by expression, partition by expression, replicate, partition by key and sort.
  • Experience in developing and deploying Ab Initio graphs, testing, trouble shooting and performance enhancement.
  • Tuned AB INITIO graph using max core, phasing and lookup file component.
  • Debugged Ab Initio graphs by enabling Watchers at flow level, Breakpoints at transformation level and Isolating individual components.
  • Worked with EME for Ab Initio jobs migration, version control, and Dependency analysis operation
  • Expertise in deploying from DEV to QA, UAT and PROD with both Deployment group and Import/Exports method
  • Experience in gathering and writing detailed business requirement and translating them into technical specifications and design
  • Good Knowledge on Big data Hadoop architecture.
  • Expertise in Data modeling techniques like Data Modeling- Dimensional/ Star Schema and Snowflake modeling, Slowly Changing Dimensions (SCD Type 1, Type 2, and Type 3).
  • Extensive experience in gathering requirements and documenting the same as per the industry best practices
  • Designed the data conversion strategy, development of data mappings, source data profiling and the design of Extraction, Transformation and Load (ETL) routines for migrating data from non-relational or source relational to target relational.
  • Strong analytical, logical, Communication and problem solving skills and ability to quickly adapt to new technologies by self-learning.

TECHNICAL SKILLS

ETL Tools: Talend Data Integration/ Big Data Integration(7.x/6/x/5.x), Talend Administrator Console, AbInitio GDE 3.1.7/3.2.6/3.2.7 and Co>operating System 3.1.7/3.2.5/3.2.7

Programming Languages: Core Java, SQL, PL/SQL

Methodologies: Agile, Waterfall

Testing Tools: Test Director 8.0, Quality Center 9.2, HP ALM

Other Tools: MS Office, MS Excel, MS Project, MS PowerPoint and MS Visio

Operating Systems: Windows 10/8/7/XP, UNIX

Scheduling Tools: Autosys

PROFESSIONAL EXPERIENCE

Talend ETL Developer / Administrator

Confidential, NJ

Responsibilities:

  • Participated in Requirement gathering, Business Analysis, User meetings and translating user inputs into ETL mapping documents.
  • Designed and customized data models for Data warehouse supporting data from multiple sources on real time
  • Involved in building the Data Ingestion architecture and Source to Target mapping to load data into Data warehouse
  • Extensively leveraged the Talend Big Data components (tSpark) for Hadoop on AWS cloud platform
  • Developing Talend DI and spark jobs and onboarding them to AWS cloud infrastructure
  • Analyzed the requirements and framed the business logic and implemented it using Talend.
  • Involved in ETL design and documentation.
  • Developed Talend jobs from the mapping documents and loaded the data into the warehouse.
  • Involved in end-to-end Testing of Talend jobs.
  • Worked on Talend components like tReplace, tmap, tsort and tFilterColumn, tFilterRow etc.
  • Worked with various File components like tFileCopy, tFileCompare, tFileExist.
  • Created complex mappings by using different transformations like Filter, Router, lookups, Stored procedure, Joiner, Update Strategy, Expressions and Aggregator transformations to pipeline data to Data Mart.
  • Scheduling and Automation of ETL processes with AWS Lambda and Jason integration.
  • Scheduled the workflows using Shell script.
  • Troubleshoot database, Joblets, mappings, source, and target to find out the bottlenecks and improved the performance.
  • Migrated Talend mappings/Job /Joblets from Development to Test and to production environment.
  • Performing transformations, cleaning and filtering on imported data using Hive, Map Reduce, and loaded final data into HDFS.
  • Performing transformations using Hive, MapReduce and loaded data into HDFS.
  • Developed talend generic jobs to load the bulk load from oracle to HDFS.
  • Executed queries using sparkSQL for complex joins and data validations.
  • Writing complex queries on sparkSQl for implementing the business functionality.
  • Worked on parquet files in multiple scenarios to load the data onto Impala.
  • Worked on Talend Administration Console (TAC) for scheduling jobs and adding users.
  • Developed stored procedure to automate the testing process to ease QA efforts and also reduced the test timelines for data comparison on tables.
  • Migrated code from Talend 6.5 version to 7.1.
  • Involved in production n deployment activities, creation of the deployment guide for migration of the code to production, also prepared production run books.

Environment: Talend Bigdata Data Integration 6.5/7.1, Talend Administrator Console, Oracle 11g, SQL Navigator, Cloudera Distribution, Cloudera Hue, AWS EC2, Parquet files.

Talend ETL Developer/Administrator

Confidential

Responsibilities:

  • Talend Project creation, Branch creation to support multiple versions in GIT and enabling access controls
  • Code Migration to Higher Environments
  • Talend jobs deployment in TAC and user support
  • Gathering the requirement from the source team and conducting design reviews.
  • Installing the talend cloud on remote engine.
  • Build and Deploy the Talend jobs via Jenkins CI/CD process
  • ETL server maintenance and oncall support
  • Developed mappings on data mapper to parse the EBCDIC files.
  • Transformed data using Redshift SQL and loading of stage data into fact and dimensional structure.
  • Used tStatsCatcher, tDie, tLogRow to create a generic job let to store processing stats.
  • Participated in JAD sessions with business users and SME's for better understanding of the requirements.
  • Design and developed end-to-end ETL process from various source systems to Staging area, from staging to Data Marts.
  • Control Center configuring and application setup
  • Coordinate Server checkouts after weekly reboots
  • Converting EBCDIC files into flat files using Talend data mapper.
  • Tested the data and data integrity among various sources and targets. Tested to verify that all data were synchronized after the data is troubleshoot, and also used SQL to verify/validate test cases.
  • Written Test Cases for ETL to compare Source and Target database systems and check all the transformation rules.
  • Defects identified in testing environment where communicated to the developers using defect tracking tool HP Quality Center
  • Performed Verification, Validation, and Transformations on the Input data

Environment: Talend open studio 5.x/6.x, Talend data mapper, Informatica, Bitbucket, Jira, EBCDIC files, flat files, AWS EC2, Service now, Autosys.

Abinitio ETL Developer

Confidential

Responsibilities:

  • Had worked on Terabytes of data in multi files system and partitioned components to fine-tune the graphs
  • Involved in all the stages of SDLC during the projects. Analyzed, designed and tested the new system for performance, efficiency and maintainability using ETL tool AB INITIO.
  • Worked to support daily Development, UAT and Production cycles execution from data prospects, by using a powerful scheduling tool Autosys, I had created, updated and maintained the Autosys jobs to automate the job sequence
  • Responsible for requirement gathering and development of Expiration Date Fictionalization New Enhancements and adding of New Source Systems in the projects.
  • Extensively used the Ab Initio components like Reformat, Join, Partition by Key, Partition by Expression, Merge, Gather, Sort, Dedup Sort, Rollup, Scan, Lookup, Normalize and De normalize.
  • Build various Sandboxes in order to create Regular applications in AbInitio.
  • Check in/Check Out existing applications Using EME in order to perform the necessary modifications.
  • FTP the New & Existing Abinitio applications to the Local Host.
  • Creating several Complex Graphs in AbInitio to Automate many process.
  • Wide usage of Lookup Files while getting data from multiple sources and size of the data is limited.
  • Worked with EME / sandbox for version control and did impact analysis for various Ab Initio projects across the organization.
  • Extensively used EME for Version Control System and for Code Promotion.
  • Involved heavily in writing complex SQL queries based on the given requirements.
  • Developed UNIX Korn shell wrapper scripts to accept parameters and scheduled the processes using Autosys.
  • Involved in building a complete new environment LAB where we had upgraded the environment from Abinitio version 1.15 to 3.0 by Co-ordinating with different offshore-teams globally.
  • Co-ordinate with development for future changes in the file or table structures to accommodate future testing requirements.
  • Managed a combination of project, technical and managerial skills with proven abilities in solving complex problems, while exceeding performance expectations. Also cooperated and communicated among architects and business groups to achieve common business goals.

Environment: Ab-Initio 3.0/3.2 GDE with co-op 3.0/3.2.,IBM AIX, HP-UX PL/SQL, Oracle 10g/11g, Todd, Aqua Data Studio, Windows XP.

We'd love your feedback!