Sr. Data Warehouse Engineer Resume
Hayward, CA
SUMMARY
- Experienced with all phases of data warehouse development lifecycle, including requirements gathering, Analysis, design, development, implementation, testing and support/maintenance.
- Exceptional background in analysis, design, development, customization, and implementation and testing of software applications and products.
- Expertize with the tools in Hadoop Ecosystem including Hive, HDFS, MapReduce and Spark.
- Excellent knowledge on Hadoop ecosystems such as HDFS, Job Tracker, Task Tracker, Name Node, Data Node and Map Reduce programming paradigm
- Experience in designing and developing applications in Spark using Scala to compare the performance of Spark with Hive and SQL/Oracle.
- Experience in manipulating/analyzing large datasets and finding patterns and insights within structured and unstructured data.
- Demonstrated expertise utilizing Data Integration tools, including Informatica, SSIS, Talend and various RDBMS including SQL Servers, Oracle, and Teradata.
- Extensive experience in Data Modeling, Data Quality, Data mart / Data Warehouse and Enterprise data architecture development and enhancement.
- Expertise in developing BI reports using various BI tools including Tableau, QlikView and SSRS.
- Experience in scheduling ofETLjobs using Tivoli work scheduler andAutosys.
- Extensive knowledge of data transformation, understanding of data modeling concepts, ability to understand and work with complex data models.
- Strong leadership with experience training and mentoring developers and advising technical groups on ETL best practices.
- Excellent technical and analytical skills with clear understanding of design goals of ER modeling for OLTP and dimension modeling for OLAP.
- Extensive experience in designing and implementing star schemas and snowflake schemas.
- Worked closely with data scientists and analysts to design and maintain scalable data models and pipelines.
- Experience with biotechnology lab functions and bioinformatics data pipeline with reporting platforms.
- Proven expertise in Production support Tier 2 for Data Warehouse.
- Used an analytical and detail oriented approach for problem solving.
- A highly motivated self - starter and a goodteam-player with excellent verbal and written communication skills.
- Managed and mentoredteamof developers distributed globally in all facets of the SDLC forETL Production Operations & System Support.
- Used TFS and Jira to manage Production Support issues and resolutions.
- Used ScrumBan boards and agile methodologies to manage and develop projects through various phases of the SDLC.
- Volunteered for Informatica User Group leader for Iowa chapter.
TECHNICAL SKILLS
Databases: Teradata, ORACLE, Sybase, Amazon Redshift, Hive, DB2, MS SQL Server
Database Tools: Aqua Data Studio, SQL Assistant, Toad, SQL Developer, MS SQL Server, SQL Server Integration Services, SQL Server Analysis Services, Business Intelligence & Development Studio, SSMS, Attunity, Aginity for Redshift, SQL Workbench, Hadoop, AWS.
Development skills: T-SQL, PL/SQL, SSIS/SSRS/SSAS
Languages: VB.net, VB Script, T-SQL, PL/SQL, HTML, CSS, XML, ASP.net, JavaScript, Python, C, C++ and MDX, Shell scripting, Hive, Spark
ETL Tools: Informatica 9.6.1, SSIS (SQL Server Integration Services), Pentaho, Talend
Reporting Packages: SQL Server Reporting Services, Tableau, QlikView, SAP Business Objects, Spotfire, Microstrategy
Tools: /Methodologies: MS Visio, Erwin, TFS, SVN, Scrumban, Agile, SQL Profiler, Jira, Hive, Spark
PROFESSIONAL EXPERIENCE
Confidential, Hayward, CA
Sr. Data Warehouse Engineer
Environment: Informatica 8.x, 9.x, Oracle 11g, PL-SQL, ETL, Erwin, MS Visio, Windows 7, SVN, Tortoise SVN, Toad, Dimensional/Relational Database, TFS, Scrumban, Spotfire, Microsoft-Office Suite, Sybase, SQL Server, MS Excel, Teradata, Aqua Data Studio, SQL Assistant, Jira, Attunity, Hadoop, HDFS, Hive, Spark, AWS, salesforce, ERP
Responsibilities:
- Involved in the all phases of the project such as Requirement gathering and analysis, Data analysis, ETL design, development and maintenance.
- Responsible for building scalable distributed data solutions using Hadoop.
- Developed Spark scripts by using Scala shell commands as per the requirement.
- Responsible for developing data pipeline with Amazon AWS to extract the data from weblogs and store in HDFS.
- Involved in creating Hive tables, and loading and analyzing data using hive queries.
- Developed Hive queries to process the data and generate the data cubes for visualizing.
- Developed the CDC pipeline for Data Warehouse. Using Attunity tool for CDC.
- Leading the production support team and responsible for scheduling (Maestro) and monitoring the jobs.
- Developed ETL programs using Informatica to implement the business requirements.
- Used Informatica file watch events to pole the FTP sites for the external files.
- Communicated with business customers to discuss the issues and requirements.
- Created shell scripts to fine tune the ETL flow of the Informatica workflows.
- Production Support has been done to resolve the ongoing issues and troubleshoot the problems.
- Performance tuning was done at the functional level and map level. Used relational SQL wherever possible to minimize the data transfer over the network.
- Effectively used Informatica parameter files for defining mapping variables, workflow variables, FTP connections and relational connections.
- Involved in enhancements and maintenance activities of the data warehouse including tuning, modifying of stored procedures for code enhancements.
- Effectively worked in Informatica version based environment and used deployment groups to migrate the objects.
- Used debugger in identifying bugs in existing mappings by analyzing data flow, evaluating transformations.
- Effectively worked on Onsite and Offshore work model.
- Pre and post session assignment variables were used to pass the variable values from one session to other.
- Designed workflows with many sessions with decision, assignment task, event wait, and event raise tasks, used informatica scheduler to schedule jobs.
- Reviewed and analyzed functional requirements, mapping documents, problem solving and trouble shooting.
- Performed unit testing at various levels of the ETL and actively involved in team code reviews.
- Designed and implemented historic, incremental (type 1 and type 2), delete mappings Using Pushdown optimization, TPT loader for Teradata.
- Used various transformations in Informatica Designer including Expression, Aggregator, Joiner, Transaction control, Filter, Lookup, Dynamic Lookup, Update Strategy, Sorter, Sequence Generator, Router, Stored procedure etc.
- Using Parameter files, workflow and mapping variables in mappings to make them easy to maintain in future.
- Used various tasks in Workflow Manager including Session, Timer, Event wait, Event Raise, Decision, Link condition, Email, Command task etc.
- Integrated data to Data warehouse from different LIMS systems, Weather data and experimental data. Created Tables, Views, DB Links, Indexes, Partitioned Tables, Procedures, Triggers, Sequence and Functions.
- Developed database models using star and snowflake schemas and implemented them on database level.
- Worked to tune the queries to improve the performance of code. Analyzed tables, used explain plan and used hints.
- Refactor all the existing views and procedures to improve efficiency and performance.
- Created best practices and standards document for the team to follow.
- Designed and developed dimensional schema for green house data.
- Profile, Integrate and manage data duality on data coming from various databases like SQL Server, Sybase, Excel and Oracle databases.
- Load all the required data in reporting database and schedule the loads as per refresh rate requirements.
- Performed POC for various tools and appliances.
- Working on Implementing Amazon Web Services for creating subscriptions to different applications.
- Lead the Production support tier 2 and participate in on-call rotation, job monitoring and resolving data issues.
Confidential, Minneapolis, MN
SQL/SSRS/BI Developer
Environment: SQL Server 2008/2005, Oracle-10g, 11g, T-SQL/PL-SQL, Windows XP, Linux, Visual Studio 2008/2005, SVN, Ankh SVN, Report Builder 2.0, Toad, Reporting Tools, Business Objects, Crystal reports, Report Manager, SharePoint Site, Quality Center, SSIS, SSRS, Dimensional Database, Microsoft-Office Suite, Hadoop, Hive.
Responsibilities:
- Built various SSIS packages having different tasks and transformations for various business areas and Scheduled SSIS packages
- Created different reports along with shared Data Sources using SSRS.
- Created various functions, views, procedures using PL/SQL and Complex Datasets with SQL.
- Created Parameterized reports and build Single-valued, Multi-valued, Drop- down, Textbox parameters.
- Created reports including KPIs, metrics and measures for required business objectives.
- Developed Cascaded Parameters for various reports.
- Designed drill-down and drill-through reports in SSRS, drill-through reports that can navigate to other reports or to other URL.
- Made subscriptions for reports through Report Manager.
- Designed SSIS work-flows and standards for various ETL jobs and Data transformations.
- Involved in database modification process and data re-modeling process for Sales, Orders
- Created SSIS packages to move data from Oracle to SQL Server and also from other sources including CSV, Text, Access and Excel files
- Involved in Resolving advanced and complex application bugs and configuration issues.
- Migrated reports from SQL Server 2005 Version to SQL Server 2008 Version
- Applied Version control to the Project. Used Ankh SVN on client machines.
- Augmented the system Triggers and defined the Triggers to perform the required action.
- Configured and setup Autosys for scheduling Jobs with specific logic.
- Used Quality Center to assign the bugs, defects, issues and the update the status of them. Also to retrieve the requirement documents.
- Load and transform large sets of structured, semi structured and unstructured data.
- Involved in loading data from LINUX file system to HDFS
- Importing and exporting data into HDFS and Hive.
- Implemented Partitioning, Dynamic Partitions, Buckets in Hive.
- Worked in creating HBase tables to load large sets of semi structured data coming from various sources.
- Experienced in running Hadoop streaming jobs to process terabytes of xml format data.
Confidential, NJ
SQL/BI Developer
Environment: SQL Server 2008/2005, Oracle-10g, T-SQL/PL-SQL, Windows XP/2003, Visual Studio 2008/2005, TFS2008/2005, Report Builder 2.0, Crystal Reports, Erwin, SSIS, SSRS, SSAS, MS-Visio, Share-Point Designer, MOSS-2007, Microsoft-Office 2003/2007, Rational Clear Quest and Rational Clear case
Responsibilities:
- Involved in gathering business requirements from business and clients to develop various reports and cubes
- Created functional requirement specifications and supporting documents for business systems
- Designed SSIS work-flows and standards for various ETL jobs and Data transformations.
- Used Agile Methodology in various projects.
- Developed complex SQL and Database objects in SQL Server 2008, Oracle 10g, 11g
- Involved in database modification process and data re-modeling process for Sales, Orders
- Created SSIS packages to move data from Oracle to SQL Server and also from other sources including CSV, Text, Access and Excel files
- Built various SSIS packages having different tasks and transformations for various business areas and Scheduled SSIS packages
- Migrated reports from SQL Server 2005 Version to SQL Server 2008 Version
- Created Partitions in Cubes on time dimension for optimizing the performance
- Designed Hierarchies by setting name column and key column as there was no strong hierarchies defined
- Designed drill-down and drill-through reports in SSRS, drill-through reports that can navigate to other reports or to other URL
- Migrated reports from SQL Server 2005 Version to SQL Server 2008 Version
- Published reports on to Share-Point and Report Server and set the data-driven subscriptions for these reports
- Used MS-Office Web Components to show data directly from the cube on a web-browser as a OWC Report
Confidential, Atlanta, GA
SQL Server / BI Consultant
Environment: SQL Server 2008/2005, T-SQL, Windows XP/2003, Visual Studio 2008/2005, Report Builder 2.0, SQL Profiler, Erwin, SQL Server Reporting services, SQL Server Analysis Services, Visio, Share-Point.
Responsibilities:
- Worked as part of a team for gathering and analyzing Business Requirements
- Participated actively in designing the system database structure
- Designed Data-Mart for Prescription using Relational Model and wrote SSIS Packages to extract data from the old PDX system and load it into the new data mart of SQL Server 2008
- Designed SSIS packages to migrate the data from PDX files to the Staging Area in SQL server 2008
- Developed SSIS Packages for migrating data from Staging Area of SQL Server 2005 to SQL Server 2008.
- Installed and administered Microsoft SQL Server 2008/2005, SQL Server Integration Services 2008/2005, SQL Server Reporting Services2008/2005 and Team Foundation Server 2005/2008.
- Built Reports in SSRS for errors generated in SSIS Packages.
- Generated Reports using Data Source Views in Report Builder 2.0
- Involved in the development of custom stored procedures, functions, triggers, SQL, T-SQL
- Involved in the optimization of SQL queries which resulted in substantial performance improvement for the conversion processes
- Performed performance tuning of stored procedures using SQL Server Profiler
- Used Query Analyzer, Profiler, Index Wizard and Performance Monitor for performance tuning on SQL Server
- Involved in creating and maintaining SQL Server Analysis Services.
Confidential
SQL Server / BI Developer
Environment: MS SQL Server 2005/2000, T-SQL,UNIX, Excel, Access, Reporting Services, Analysis Services, DTS, Data Analyzer, Visual Studio 2005, Crystal Reports 8.0, Windows XP/2000
Responsibilities:
- Extracted large volumes of data from different data sources and loaded the data into target data sources by Performing different kinds of transformations using SQL Server Integration Services (SSIS).
- Experience in SSIS script task, look up transformations and data flow tasks using T- SQL and Visual Basic (VB) scripts. .
- Involved in SQL joins, sub queries, tracing and performance tuning for better running of queries.
- Involved in Error Handling using try and catch blocks and performance tuning using counters in SSIS.
- Participated in developing Logical design of database incorporating business logic and user requirement
- Created stored procedures, functions, triggers (database objects) and called them in the SSIS packages.
- Involved in designing, building and deploying multidimensional cubes using SQL Server Analysis Services
- Used best practices method to build the cube with respect to the performance
- Worked with advance properties of the cubes like calculations, partitions and aggregations
- Configured database mail, created operators, jobs, alerts for automating databases
- Ad-hoc report design, grouping and sorting using Visual Studio 2005 and also importing sub-reports into it.
