We provide IT Staff Augmentation Services!

Sr. Etl/talend Developer Resume

4.00/5 (Submit Your Rating)

Chicago, IL

SUMMARY

  • 8+ years of IT experience in Analysis, Design, Developing and Testing and Implementation of business application systems.
  • Highly skilled ETL Engineer with 6+ years of software development in tools like Talend and DataStage.
  • Strong experience in the Analysis, design, development, testing and Implementation of Business Intelligence solutions using Data Warehouse/Data Mart Design, ETL, OLAP, BI, Client/Server applications.
  • 6+ years’ Experience on Talend ETL Enterprise Edition for Big data/Data integration/Data Quality.
  • Experience in Big Data technologies like Hadoop/Map Reduce, Pig, HBASE, Hive, Sqoop, HBase, DynamoDB, Elastic Search and Spark SQL.
  • Experienced in ETL methodology for performing Data Migration, Data Profiling, Extraction, Transformation and Loading using Talend and designed data conversions from large variety of source systems including Oracle 10g/9i/8i/7.x, DB2, Netezza, SQL server, Teradata, Hive, Hana and non - relational sources like flat files, XML and Mainframe files.
  • Involved in code migrations from Devto QA and production and providing operational instructions for deployments.
  • Utilized AWS Services (S3, EC2, EMR, RDS, Amazon RedShift) in project scope.
  • Experienced in using TBD and Talend Data Fabric tools (Talend DI, Talend MDM, Talend DQ, Talend Data Preparation, ESB, TAC).
  • Experienced in using All Talend DI, MDM, DQ, DP, ESB components.
  • Worked hands on DataStage 8.7 ETL migration to Talend Studio ETL process.
  • Expertise in Data Warehouse/Data mart, ODS, OLTP and OLAP implementations teamed with project scope, Analysis, requirements gathering, data modeling, Effort Estimation, ETL Design, development, System testing, Implementation and production support.
  • Experienced in Talend Service Oriented Web Services using SOAP, REST and XML/HTTP technologies using Talend ESB components.
  • AWS Pipeline knowledge to develop ETL for data movement to Redshift
  • Experience in EDI formats like X12, EDIFACT etc.
  • Hands on experience in Pentaho Business Intelligence Server Studio.
  • Expertise in Using transformations like Joiner, Expression, Connected and Unconnected lookups, Filter, Aggregator, Store Procedure, Rank, Update Strategy, Java Transformation, Router and Sequence generator.
  • Experienced on writing Hive queries to load the data into HDFS.
  • Excellent experience on designing and developing of multi-layer Web Based information systems using Web Services including Java and JSP.
  • Strong experience in Dimensional Modeling using Star and Snowflake Schema, Identifying Facts and Dimensions, Physical and logical data modeling using ERwin and ER-Studio.
  • Expertise in working with relational databases such as Oracle 12c/11g/10g/9i/8x, SQL Server 2012/2008/2005 , DB2 8.0/7.0, UDB, MS Access and Teradata, Netezza.

TECHNICAL SKILLS

BI Tools: Micro Strategy, IBM Cognos, Pentaho.

ETL Tools: Talend Bigdata 6.3/5.6/5.1 Informatica Power Center, IBM Infosphere DataStage.

Databases: Hive, Impala, Mongo DB, Teradata, Oracle, Netezza, Oracle Exadata, SQL Server, DB2, Access, AWS Services (S3, EC2, EMR, RDS, Amazon RedShift) etc.

Document management: Visual Source Safe 6.0, Share point, Ultra Edit, Documentum

Defect management tool: HP Quality center 10.0

Programming Languages: HTML, SQL, PL/SQL, Core Java, Unix

Data modeling: Power Designer, ERwin, ER Studio, MS Visio

Code repository: SVN, GIT, Bitbucket.

Scheduling tools: AutoSys, control, tidal.

PROFESSIONAL EXPERIENCE

Confidential, Chicago, IL

Sr. ETL/Talend Developer

Responsibilities:

  • Participated in Requirement gathering, Business Analysis, User meetings and translating user inputs into ETL mapping documents.
  • Designed and customized data models for Data warehouse supporting data from multiple sources on real time
  • Involved in building the Data Ingestion architecture and Source to Target mapping to load data into Data warehouse
  • Worked with Data mapping team to understand the source to target mapping rules. Analyzed the requirements and framed the business logic and implemented it using Talend.
  • Involved in ETL design and documentation.
  • Developed Talend jobs from the mapping documents and loaded the data into the various target Data warehouses.
  • Perform unit testing and capturing the run metrics against Redshift cluster and Talend jobs working closely with SIT and UAT.
  • Involved in end-to-end Testing of Talend jobs.
  • Created several Java routines for complex transformation and utilities to use across the integrations.
  • Experience in migrating the data to cloud data warehouses, and Salesforce CRM from legacy systems, like Oracle, Sugar CRM, Netezza, Salesforce.
  • Developed and Designed SOAP/REST Services using SOA/OSB to support the data on the User Interface like Mobile and Web applications
  • Worked with business to collect requirements, design the integration, document the specifications and communicate and coordinate the development and testing effort. As part of the change management team, revisit and readdress P2P policy and procedures impacted by these initiatives
  • Worked with multiple ICS adapters (File, FTP, Salesforce, ERP Cloud, HCM, Twillo, DB, SOAP & REST)
  • Developed a migration utility to automatically migrate/deploy and activate interfaces from one Oracle ICS Cloud Environment to another using CI / CD tools such as Bit Bucket Pipelines and Oracle ICS REST API’s.
  • EDI conversion XML to EDI and EDI to XML using b2b adaptor.
  • Continuous Integration and Continuous Delivery (CI-CD), Dependency Injection (DI), Inversion of Control (IoA), Github, Source Tree, Bitbucket, Maven, Jenkins
  • Worked on the cloud data storages like Amazon S3, Azure data lakes,
  • Worked with the Amazon S3 and EC2 components to migrate the Data from Different source systems to Amazon S3 buckets.
  • Gained a knowledge on Azure services also achieved the Azure certification.
  • Wrote complex SQL queries to take data from various sources and integrated it with Talend.
  • Worked on Context variables and defined contexts for database connections, file paths for easily migrating to different environments in a project.
  • Involved in loading the data into Netezza from legacy and flat files using Unix scripts. Worked on Performance Tuning of Netezza queries with proper understanding of joins and Distribution
  • Created ETL job infrastructure using Talend Open Studio.
  • Developed the business rules for cleansing/validating/standardization of data using Informatica Data Quality.
  • Developed standards for ETL framework for the ease of reusing similar logic across the board.
  • Created complex mappings by using different transformations like Filter, Router, lookups, Stored procedure, Joiner, Update Strategy, Expressions and Aggregator transformations to pipeline data to Data Mart.
  • Utilized Talend components like tS3Put, tS3Get, tS3File List, tRedshift Row, tRedshiftUnload, tRedshift BulkExec.
  • Involved in the Development of copying Data from S3 to RedShift using the Talend Process.
  • Responsible for developing the jobs using ESB components like tESBConsumer, tESBProviderFault, tESBProviderRequest, tESBProviderResponse, tRESTClient, tRESTRequest, tRESTResponse to get the service calls for customers DUNS numbers.
  • Creating Talend Development Standards. This document describes the general guidelines for Talend developers, the naming conventions to be used in the Transformations and development and production environment structures.
  • Troubleshoot database, Joblets, mappings, source, and target to find out the bottlenecks and improved the performance.
  • Involved rigorously in Data Cleansing and Data Validation to validate the corrupted data.

Environment: Talend 7.1.1, XML files, DB2, Oracle 11g, Netezza 4.2, SQL server 2008, SQL, MS Excel, MS Access, UNIX Shell Scripts, Talend Administrator Console, Cassandra, Redshift, Oracle, Jira, SVN, Teradata, Quality Center, and Agile Methodology, TOAD, Autosys.

Confidential, Deerfield, IL

Sr. Talend Bigdata Developer

Responsibilities:

  • Interacted with business team to understand business needs and to gather requirements.
  • Designed target tables as per the requirement from the reporting team and designed Extraction, Transformation and Loading (ETL) using Talend.
  • Hadoop Cluster using TAC in the following nodes, Managing and scheduling jobs on Hadoop cluster using Autosys tool.
  • Work in SQL and shell scripting to automate processes
  • Handle importing of data in Talend from various data sources, performed transformations using Hive, Pig and Spark and loaded data into HDFS.
  • Design and development experience with Talend Integration Suite and knowledge in Performance Tuning of mappings
  • Good proficiency in writing SQL queries.
  • Developed the data ingestion i.e. Transformation on data source raw data file into Redshift cluster using Talend.
  • Expertise in creating mappings in TALEND using tMap tJoin tReplicate tParallelize tConvertType tflowtoIterate tAggregate tSortRow tFlowMeter tLogCatcher tRowGenerator tNormalize tDenormalize tSetGlobalVar tHashInput tHashOutput tJava tJavarow tAggregateRow tWarn tLogCatcher tMysqlScd tFilter tGlobalmap tDie etc.
  • Created Talend ETL jobs using ETL methodologies and have best practices by following naming standards.
  • Integrated Redshift SSO cluster with Talend.
  • Data ingestion with different data sources and load into redshift.
  • Defining parameter files context variables and assign them e Creating and using parallel Talend jobs and reusable Talend jobs
  • Developed multiple SPARK jobs in java for data cleaning and accessing.
  • Work on importing and exporting data from Oracle and DB2 into HDFS using Sqoop.
  • Extract source data from relational databases (SQL/ Flat File/Mainframes) are processed using Extract, Transform and Load into another database.
  • Developed jobs to move inbound files to vendor server location based on monthly, weekly and daily frequency in Talend.
  • Test data flows within Talend to match functionality and performance of Ab Initio. This includes data matching and run time comparisons.
  • Troubleshoot and debug the code to normalizing the data flows, redesigning a job, optimizing code and properly utilizing resources such as memory.

Environment: Talend Data Integration 7.1.1, Talend Enterprise Big Data Edition 7.7.1, Talend Administrator Console 7.1.1, Abinitio, Redshift, Oracle Data Integrator, HDFS, Sqoop, SSIS, SSRS, SSAS, Pig, SQL Navigator, XML files, Flat files, HL7 files, JSON,Teradata, TWS, Hadoop 2.4.1, HBase.

We'd love your feedback!