Sr. Etl/talend Developer Resume
Norfolk, VA
PROFESSIONAL SUMMARY:
- Over 9+ years of IT Experience in analysis, design, development, implementation and troubleshooting of Data warehouse/Data Services applications and Talend BDE/DI.
- Expertise in designing/building Enterprise Data Warehouses (EDW), Operational Data Store (ODS), Data Marts, and Decision Support Systems (DSS) using Multidimensional and Ralph Kimball Dimensional modeling (Star and Snowflake schema) Concepts.
- Experience with Big data, Hadoop, HDFS, Map Reduce, Spark and Hadoop Ecosystem (Pig & Hive) Technologies.
- Experience in CDC and daily load strategies of Data warehouse and Data marts, slowly changing Dimensions (Type1, Type2, and Type3) and Surrogate Keys and Data warehouse concepts.
- Extensive experience in ETL methodology for performing Data Profiling, Data Migration, Extraction, Transformation and Loading using Talend and designed data conversions from wide variety of source systems including Oracle, DB2, SQL server, Teradata, Hive and non - relational sources like flat files, XML and Mainframe files.
- Experience of Hadoop Architecture and various components such as HDFS, Name Node, Data Node, Job Tracker, Task Tracker, Spark, Zookeeper, YARN and Map Reduce.
- Experience in using cloud components and connectors to make API calls for accessing data from cloud storage (Jira, Amazon S3) in Talend Open Studio.
- Experience with Salesforce connectors and implemented SFDC CDC for tables Account, Opportunity, Opportunity History, Case, Users.
- Strong Experience in developing Sessions/tasks, Worklets, Workflows using Workflow Manager Tools - Task Developer, Workflow & Worklet Designer.
- Excellent knowledge in identifying performance bottlenecks and tuning the ETL Loads for better performance and efficiency.
- Expertized in creating specifications documents from source to Target systems.
- Experience in UNIX shell scripting, FTP and file management in various UNIX environments.
- Strong understanding of Data warehouse project development life cycle.
- Expertise in documenting all the phases of DWH projects.
TECHNICAL SKILLS:
Operating Systems: Windows 2008/2007/, UNIX, LINUX
ETL Tools: Talend Open Studio, Talend Data Fabric, Informatica, Mainframe system
Databases: Teradata, Oracle 11g, My SQL, MS SQL Server, DB2, HBase.
Methodologies: Agile, Waterfall.
Languages: SQL, PL/SQL, UNIX, Shell scripts, C++, JSP, Java Script, HTML, Eclipse
Scheduling Tools: Autosys, TIC, TAC.
Version Control: Subversion (SVN), TFS, Git Hub, Nexus, Jenkins
PROFESSIONAL EXPERIENCE:
Confidential, Norfolk, VA
Sr. ETL/Talend Developer
Responsibilities:
- Worked with Business Analyst to design Business Requirement Documents.
- Created a data dictionary to map the business requirements to attributes to design logical data model implementing star schema.
- Responsible in producing report analysis by column level for all oracle tables, primary key analysis, foreign key analysis, cross domain analysis.
- Provided tips for the DW team to improve the performance of the ETL jobs.
- Designed and Developed ETL process using Talend Open Studio (Data Integration) & Worked on Enterprise Latest Version in pararallel to Development and acted as an Talend Admin creating Projects/ Scheduling Jobs / Migration to Higher Environments & Version Upgrades
- Developed complex ETL mappings for Stage, Dimensions, Facts and Data marts load
- Involved in Data Extraction for various Databases & Files using Talend
- Used Talend most used components (tMap, tDie, tConvertType, tLogCatcher, tRowGenerator, tSetGlobalVar, tHashInput & tHashOutput, tFilterRow, tAggregateRow, tFileExist, tFileCopy, tFileList, tDie and many more).
- Created many complex ETL jobs for data exchange from and to Database Server and various other systems Including RDBMS, XML, CSV, and Flat file structures.
- Responsible for developing, support and maintenance for the ETL (Extract, Transform and Load) processes using Talend Integration Suite.
- Performing code reviews with client teams and the peers. Clarifying and helping the team resolving technical issues
- Hands of Experience on many components which are there in the palette to design Jobs & used Context Variables/Groups to Parameterize Talend Jobs.
- Implemented Error Logging, Error Recovery, and Performance Enhancement’s & created Audit Process (generic) for various Application teams.
- Experience in using Repository Manager for Migration of Source code from Lower to higher environments.
- Performing unit testing, regression testing and supporting QA testing.
- Taking care of check-in, checkouts of the code and deploying across the various project environments, such as development, test, stage and then to production.
- Prepared ETL mapping Documents for every mapping and Data Migration document for smooth transfer of project from development to testing environment and then to production environment.
- Involved in Unit testing, User Acceptance Testing to check whether the data loads into target are accurate, which was extracted from different source systems according to the user requirements.
- Responsible for prioritizing the issues and assign them to the production support team and planning the deployment of fixes for the same.
Environment: Talend 7.1, Sqlworkbench, Git, Jenkins, Spark Sql, Putty, JIRA, Teradata, Oracle 11g, Cloudera, Hive, HDFS, Sqoop, TOAD, UNIX.
Confidential, Bloomington, IL
Sr. ETL/ Talend Developer
Responsibilities:
- Participated in all phases of development life-cycle with extensive involvement in the definition and design meetings, functional and technical walkthroughs.
- Created Talend jobs to copy the files from one server to another and utilized Talend FTP components
- Created and managed Source to Target mapping documents for all Facts and Dimension tables
- Used ETL methodologies and best practices to create Talend ETL jobs. Followed and enhanced programming and naming standards.
- Created and deployed physical objects including custom tables, custom views, stored procedures, and Indexes to SQL Server for Staging and Data-Mart environment.
- Design and Implemented ETL for data load from heterogeneous Sources to SQL Server and Oracle as target databases and for Fact and Slowly Changing Dimensions SCD-Type1 and SCD-Type2.
- Setup ETL Framework around Talend for STC Big data implementation.
- Excellent experience working on tHDFSInput, tHDFSOutput, tPigLoad, tPigFilterRow, tPigFilterColumn, tPigStoreResult, tHiveLoad, tHiveInput, tHbaseInput, tHbaseOutput, tSqoopImport and tSqoopExport.
- Developed jobs to expose HDFS files to Hive tables and Views depending up on the schema versions.
- Created Hive tables, partitions and implemented incremental imports to perform ad-hoc queries on structured data.
- Implementing best practices and taking delivery standards to the next level.
- Worked on Talend Administrator Console (TAC) for scheduling jobs and adding users.
- Developed jobs to move inbound files to HDFS file location based on monthly, weekly, daily and hourly partitioning.
- Worked extensively on design, development and deployment of talend jobs to extract data, filter the data and load them into datalake.
- Used Apache Spark for streaming data from large datasets.
- Worked on Standard, Big Data Spark and Big Data Streaming jobs based on the user requirements.
- Manage and Review Hadoop log files and hands on with executing Linux and HDFS Commands
- Designed the hive jobs and scheduled them using the framework (an in-house scheduling framework in the organization).
- Experience in using cloud components and connectors to make API calls for accessing data from cloud storage (Jira, Salesforce, Amazon S3) in Talend Open Studio.
- Responsible for writing Talend Routines in Java.
- Developed ODS/OLAP data model in Erwin and also created source to target mapping documents.
- Experience working with web services using tSOAP components for sending XML requests and receiving response XML files. Expertized in reading XMLs files on a loop and sending to webservice end point for generating output XML files.
- Responsible for digging into PL/SQL code for investigating data issues.
- Involved in the development of Talend Jobs and preparation of design documents, technical specification documents.
- Implemented job parallelism in Talend BDE 6.0.1.
- Experience working with Big data components for extracting and loading data into HDFS file system.
- Production Support activities like application checkout, batch cycle monitoring and resolving User Queries.
- Responsible for deploying code to different environments using GIT.
Environment: Talend Big Data 6.0.1/6.3, Hortonworks HDP 2.3, Hive, HDFS, Sqoop, Spark, Jaspersoft Professional 6 Oracle 11g, Web services (SOAP), GIT, Jira, Jenkins, and AWS
ETL Developer
Confidential
Responsibilities:
- Meeting with business analysts and the application architects to understand the project requirements in terms of functional and technical.
- Preparing technical designs and getting reviewed with the client teams.
- Developed templates for Dimensional tables, Slowly Changing dimension tables, Fact tables.
- Designed and Developed ETL logic for implementing CDC by tracking the changes in critical fields required by the user.
- Developed standard and re-usable mappings and mapplets using various transformations like expression, aggregator, joiner, source qualifier, router, lookup Connected/Unconnected, and filter.
- Extensive use of Persistent cache to reduce session processing time. Identified performance issues in existing sources, targets and mappings by analyzing the data flow, evaluating transformations and tuned accordingly for better performance.
- Maintained warehouse metadata, naming standards and warehouse standards for future application development.
- Used Workflow Manager for creating, validating, testing and running the sequential and concurrent sessions and scheduling them to run at specified time and as well to read data from different sources and write it to target databases.
- Implemented type II slowly changing dimension techniques on changing attributes.
- Implemented Retry logic for reprocessing records in next run if the foreign key data missed in first run due to timing issues.
- Automated coding process to create ETL jobs from Enterprise to Stage for faster delivery.
- Developed UNIX shell scripts for automating process.
Environment: Informatica 8.5, Oracle 9i/10g, Teradata, SQL plus, Teradata SQL Assistant, Control M, PVCS.
ETL Developer
Confidential, Michigan, USA.
Responsibilities:
- Involved in Analysis of the requirements.
- Responsible for writing scripts in the data members for ETL process using Fexport, Mload, FastLoad, and BTEQ utilities.
- Tuning the Mappings for Optimum Performance, Dependencies and Batch Design.
- Port Test data using file manager and Teradata utilities.
- Assisting the team for production move.
- Schedule and Run Extraction and Load process and monitor sessions using CA7 Scheduler.
- Implementing code changes and support production issues.
- Involved in entire life cycle of replenishment work bench i.e. owing the responsibilities until the component moved to production.
- Involved in report generation for the forecasting purpose using BI Query tool.
Environment: Teradata SQL Assistant, Oracle 8.i, PL/SQL, Teradata, DB2, Z/OS, JCL, IBM Data Studio, BI Query, IBM Data studio.
