Senior Etl Datastage Developer Resume
MN
8SUMMARY
- A Qualified IT Professional with 14+ years of experience in Software Development with significant experience on ETL Datastage and as Snowflake Developer
- 6+ Years of experience leading onsite and offshore teams
- Experience in developing ETL processes to extract, transform and load data from various source and archive systems into Data Warehouses and Data Marts using Datastage Manager, Datastage Designer and Datastage Director.
- Ability to solve complex technical problems and have unique ability to understand long - term project development issues at all levels and interpersonal skills with strong desire to achieve specified goals within the given time frame
- Converted PL/SQL type of code to both bigquery - python architecture as well as azure databricks and pyspark in dataproc.
TECHNICAL SKILLS
ETL Tools/ Hadoop: Infromatica powercenter 10.4; Data Stage 8.5, 9.1, 11.5 & 11.7; HDFS, Mapreduce, Hive, Impala, pig, Sqoop, Flume, Hue, Spark
Languages/Technologies: Python, Shell scripting, SQL, R, PL/SQL, SnowSQL, Unix, Linux, COBOL, T-SQL.
Data Warehouses: Snowflake, Netezza, Teradata, Oracle,Sybase, SAP HANA
Data modeling Tools: Star & Snow-Flake schema Modeling, Fact and Dimensions, Physical and Logical Data Modeling,Erwin, Micro Soft Visio
Cloud Computing: AWS (S3, EMR, Lambda, Glue, Step Functions, Orchestrator)
PROFESSIONAL EXPERIENCE
Confidential, MN
Senior ETL Datastage Developer
Responsibilities:
- Migrated existing on perm Datastage to AWS environment
- Identified potential improvements to the current ETL design/process post migration
- Worked on Performance tuning and SAP integration project primarily focused on finding tuning options for the data load methods through Datastage including BAPI, IDOC, and LSMW.
- Worked on tuning of existing jobs, tuned jobs and reduced the execution time by 40% on an average
- Tuning tasks performed include gathering functional requirements, exporting DataStage projects into new environment (DEV, STG)
- Prepare technical solution of the applications, present to Architecture leadership team for approvals and feedback.
- Convert legacy reports from SAS, Looker, Access, Excel and SSRS into Azure Power BI and Tableau
- Responsible for maintaining the application code in Git Lab repository and perform code reviews of pull requests, ensure accurate code in release branches for production deployment.
Environment: ORACLE, SAPHANA, SQL, Looker, DB2, T-SQL, BigQuery, Teradata, PostgreSQL, Linux, Unix DB2, Snowflake, AWS services, GIT, Autosys, Tableau.
Confidential, Dallas, TX
SnowFlake / Datastage Developer
Responsibilities:
- Work with Snowflake utilities, Snowflake SQL, Snow Pipe, etc.
- Migrated existing on perm Datastage to AWS environment
- Identified potential improvements to the current ETL design/process post migration
- Integrated data from multiple source systems which includes JASON/PARQUET into snowflake using AWS S3
- Assisted to map existing data ingestion platform to AWS cloud stack for enterprise cloud intiative
- Create Snow pipe for continuous load data and used copy to bulk load the data.
- Work in Snowflake advanced concepts like setting up Resource Monitors, Role Based Access Controls, Data Sharing, Virtual Warehouse Sizing, Query Performance Tuning, Snow Pipe, Tasks, Streams, Zero- copy cloning etc.
- Developed UNIX shell scripts for scheduling sessions in Informatica.
- Migrate Matillion pipelines and Looker reports from Amazon Redshift to Snowflake data warehouse.
- Using snowflake as SaaS and migrated DB2 data to snowflake using SnowSQL and data movement servers.
- Create Internal, External stage and transformed data during load.
- Use COPY to bulk load the data.
- Updating the Python Unit tests regularly to ensure its accuracy and usefulness
- Work related to downloading BigQuery data into pandas or Spark data frames for advanced ETL capabilities.
- Created T-SQL stored procedures, functions, triggers, cursors and tables.
- Create several types of data visualizations using Python and Tableau.
- Migrate existing Netezza Data Warehouse to Snowflake.
- Create data sharing between two snowflake accounts.
- Develop common data Ingestion framework to load various source data formats to Data Lake using Control-M and UNIX scripting.
- Develop and modify LookML code in Looker business intelligence solution.
- Worked and developed python utilities for Modelling Data before loading to Staging Area.
- Implement batch processing using Control M tool.
- Configure DBT and developed models for Data Transformations on snowflake.
- Strong understanding of the principles of Data Warehousing using fact tables, dimension tables and star/snowflake schema modeling.
- Use different processing stages like Transformer, Aggregator, Lookup, Join, Sort, Copy, Merge, Funnel, CDC and Filter in IBM DataStage.
- Develop the mappings using needed Transformations in Informatica tool according to technical specifications
- Highly proficient in the use of T-SQL for developing complex stored procedures, tally tables, triggers, functions, and merges.
- Manage large datasets using Python Panda data frames.
- Use Python in data migration from DB2 to Snowflake during history loads and quality checks.
- Troubleshoot jobs and address production issues like data issues, performance tuning and enhancements.
- Design, develop, test and maintain Tableau functional reports based on user requirements and converted existing BO reports to tableau dashboards.
- Created BigQuery authorized views for row level security or exposing the data to other teams.
- Prepare technical solution of the applications, present to Architecture leadership team for approvals and feedback.
- Responsible for maintaining the application code in Git Lab repository and perform code reviews of pull requests, ensure accurate code in release branches for production deployment.
Environment: Python, SQL, Linux, Unix DB2, Snowflake, T-SQL,Looker, BigQuery, AWS services, GIT, Control M, DBT and Tableau.
Confidential, Charlotte, NC
Snowflake Developer
Responsibilities:
- Worked on SnowSQL and Snowpipe
- Created Snowpipe for continuous data load.
- Used COPY to bulk load the data.
- Migrated existing SQL Server Data Warehouse to Snowflake.
- Created data sharing between two snowflake accounts.
- Created internal and external stage and transformed data during load.
- After the processing the Project records, used Reports to generates the customize the reports.
- Created the tableau group on SQL Server and granted permission to publish on the tableau server.
Environment: Snowflake, SQL server, AWS, Looker, GitHub, Tableau, Oracle, SSIS, Informatica
Confidential, Charlotte, NC
Hadoop Developer
Responsibilities:
- Responsible for building scalable distributed data solutions using Hadoop.
- Responsible for creating Hive tables, loading the structured data resulted from MapReduce jobs into the tables and writing Hive queries to further analyze the logs to identify issues and behavioral patterns.
- Involved in submitting and tracking MapReduce jobs using Job Tracker.
- Used Pig as ETL tool to do transformations, event joins, filter and some pre-aggregations.
- Exported data to Tableau and excel with Power view for presentation and refining.
- Implemented Hive Generic UDF's to implement business logic.
- Implemented test scripts to support test driven development and continuous integration.
Confidential, Boston MA
Senior Datastage/ ETL Developer
Responsibilities:
- Worked on NDW project creating a new datawarehouse for maintaining exception data.
- Worked on Credit MIS project extracting and loading data into Oracle database.
- Extensively worked with Datastage utilities.
- Migrating existing processes from Informatica powercenter to Datastage
- Used Info sphere Information Analyzer for generating profiling reports helping data cleansing.
- Worked on Quality stage for data standardization.
- Used Webservices, XML stages to extract and load data.
- Used ISD stages to expose datastage jobs as Webservices.
- Used Autosys for scheduling.
- Developed complex SQL to efficiently extract data from multiple data sources databases.
- Used the Data Stage Director for scheduling, validating, running and monitoring the jobs.
- Used Datastage Administrator for defining various environmental settings and variables
- Automated Qualitystage reports for user convenience enabling execution of the reports throughSharePoint.
- Good knowledge on Agile working model.
Environment: DataStage 11.5, Informatica Powercenter, IA, Quality Stage, Unix, Autosys, Clearcase
Confidential, Dallas, Tx
Senior Datastage/ETL Developer
Responsibilities:
- Data stage 8.5 was used to transform a variety of financial transaction files from different product platforms into standardized data.
- Designing ETL jobs incorporating complex transform methodologies using Data Stage tool resulting in development of efficient interfaces between source and target systems.
- Developed ETL jobs to load data from VSAM, GDG, IMS, DB2 databases, Flat files, CSV files to Target and experience with high volume databases on Mainframes.
- Worked with stages like Complex Flat File, Transformer, Aggregator, Sort, Join, Lookup, and Data masking pack.
- Performed the Unit testing for jobs developed to ensure that it meets the requirements.
- Involved in documenting the Frame Work Templates and the process of developing jobs using Templates.
Environment: IBM Data stage 8.5 (Director, Designer, Administrator), IBM DB2, IBM Optim, Zlinux, Oracle, DB2, SQL server, Mainframes.
Confidential, Chicago, Illinois
Datastage Developer
Responsibilities:
- Involved in meetings to gather information and requirements from the clients.
- Involved in Designing the ETL process to Extract translates and load data from OLTP Oracle database system to Teradata data warehouse.
- Gathered information from different data warehouse systems and loaded into Database using Fast Load, Fast Export, Multi Load, Bteq and UNIX shell scripts.
- Used the Ascential Datastage Designer to develop processes for extracting, cleansing, transforming, integrating, and loading data into data warehouse database.
- Worked in the areas of relational database logical design, physical design, and performance tuning of the RDBMS.
- Used the Datastage Director and its run-time engine to schedule running the solution, testing and debugging its components, and monitoring the resulting executable versions (on an ad hoc or scheduled basis).
- Scheduled jobs dependencies using Control-M Scheduler.
- Implemented Unit, Functionality, Performance and Stress testing on Mappings and created Testing Documents.
- Involved in unit testing, systems testing, integrated testing and user acceptance testing.
Environment: Ascential Datastage 7.5.1 (Manager, Designer, Director), Teradata V2R6, Tools & Utilities (BTEQ, Fast Export, Multi Load, Fast load, TPUMP), Oracle 10g, Windows 2000/NT, Unix, Control-M Scheduling.
