We provide IT Staff Augmentation Services!

Sr. Data Warehouse Engineer Resume

0/5 (Submit Your Rating)

Hayward, CA

SUMMARY

  • Experienced with all phases of data warehouse development lifecycle, including requirements gathering, Analysis, design, development, implementation, testing and support/maintenance.
  • Exceptional background in analysis, design, development, customization, and implementation and testing of software applications and products.
  • Expertize with the tools in Hadoop Ecosystem including Hive, HDFS, MapReduce and Spark.
  • Excellent knowledge on Hadoop ecosystems such as HDFS, Job Tracker, Task Tracker, Name Node, Data Node and Map Reduce programming paradigm
  • Experience in designing and developing applications in Spark using Scala to compare the performance of Spark with Hive and SQL/Oracle.
  • Experience in manipulating/analyzing large datasets and finding patterns and insights within structured and unstructured data.
  • Demonstrated expertise utilizing Data Integration tools, including Informatica, SSIS, Talend and various RDBMS including SQL Servers, Oracle, and Teradata.
  • Extensive experience in Data Modeling, Data Quality, Data mart / Data Warehouse and Enterprise data architecture development and enhancement.
  • Expertise in developing BI reports using various BI tools including Tableau, QlikView and SSRS.
  • Experience in scheduling ofETLjobs using Tivoli work scheduler andAutosys.
  • Extensive knowledge of data transformation, understanding of data modeling concepts, ability to understand and work with complex data models.
  • Strong leadership with experience training and mentoring developers and advising technical groups on ETL best practices.
  • Excellent technical and analytical skills with clear understanding of design goals of ER modeling for OLTP and dimension modeling for OLAP.
  • Extensive experience in designing and implementing star schemas and snowflake schemas.
  • Worked closely with data scientists and analysts to design and maintain scalable data models and pipelines.
  • Experience with biotechnology lab functions and bioinformatics data pipeline with reporting platforms.
  • Proven expertise in Production support Tier 2 for Data Warehouse.
  • Used an analytical and detail oriented approach for problem solving.
  • A highly motivated self - starter and a goodteam-player with excellent verbal and written communication skills.
  • Managed and mentoredteamof developers distributed globally in all facets of the SDLC forETL Production Operations & System Support.
  • Used TFS and Jira to manage Production Support issues and resolutions.
  • Used ScrumBan boards and agile methodologies to manage and develop projects through various phases of the SDLC.
  • Volunteered for Informatica User Group leader for Iowa chapter.

TECHNICAL SKILLS

Databases: Teradata, ORACLE, Sybase, Amazon Redshift, Hive, DB2, MS SQL Server

Database Tools: Aqua Data Studio, SQL Assistant, Toad, SQL Developer, MS SQL Server, SQL Server Integration Services, SQL Server Analysis Services, Business Intelligence & Development Studio, SSMS, Attunity, Aginity for Redshift, SQL Workbench, Hadoop, AWS.

Development skills: T-SQL, PL/SQL, SSIS/SSRS/SSAS

Languages: VB.net, VB Script, T-SQL, PL/SQL, HTML, CSS, XML, ASP.net, JavaScript, Python, C, C++ and MDX, Shell scripting, Hive, Spark

ETL Tools: Informatica 9.6.1, SSIS (SQL Server Integration Services), Pentaho, Talend

Reporting Packages: SQL Server Reporting Services, Tableau, QlikView, SAP Business Objects, Spotfire, Microstrategy

Tools: /Methodologies: MS Visio, Erwin, TFS, SVN, Scrumban, Agile, SQL Profiler, Jira, Hive, Spark

PROFESSIONAL EXPERIENCE

Confidential, Hayward, CA

Sr. Data Warehouse Engineer

Environment: Informatica 8.x, 9.x, Oracle 11g, PL-SQL, ETL, Erwin, MS Visio, Windows 7, SVN, Tortoise SVN, Toad, Dimensional/Relational Database, TFS, Scrumban, Spotfire, Microsoft-Office Suite, Sybase, SQL Server, MS Excel, Teradata, Aqua Data Studio, SQL Assistant, Jira, Attunity, Hadoop, HDFS, Hive, Spark, AWS, salesforce, ERP

Responsibilities:

  • Involved in the all phases of the project such as Requirement gathering and analysis, Data analysis, ETL design, development and maintenance.
  • Responsible for building scalable distributed data solutions using Hadoop.
  • Developed Spark scripts by using Scala shell commands as per the requirement.
  • Responsible for developing data pipeline with Amazon AWS to extract the data from weblogs and store in HDFS.
  • Involved in creating Hive tables, and loading and analyzing data using hive queries.
  • Developed Hive queries to process the data and generate the data cubes for visualizing.
  • Developed the CDC pipeline for Data Warehouse. Using Attunity tool for CDC.
  • Leading the production support team and responsible for scheduling (Maestro) and monitoring the jobs.
  • Developed ETL programs using Informatica to implement the business requirements.
  • Used Informatica file watch events to pole the FTP sites for the external files.
  • Communicated with business customers to discuss the issues and requirements.
  • Created shell scripts to fine tune the ETL flow of the Informatica workflows.
  • Production Support has been done to resolve the ongoing issues and troubleshoot the problems.
  • Performance tuning was done at the functional level and map level. Used relational SQL wherever possible to minimize the data transfer over the network.
  • Effectively used Informatica parameter files for defining mapping variables, workflow variables, FTP connections and relational connections.
  • Involved in enhancements and maintenance activities of the data warehouse including tuning, modifying of stored procedures for code enhancements.
  • Effectively worked in Informatica version based environment and used deployment groups to migrate the objects.
  • Used debugger in identifying bugs in existing mappings by analyzing data flow, evaluating transformations.
  • Effectively worked on Onsite and Offshore work model.
  • Pre and post session assignment variables were used to pass the variable values from one session to other.
  • Designed workflows with many sessions with decision, assignment task, event wait, and event raise tasks, used informatica scheduler to schedule jobs.
  • Reviewed and analyzed functional requirements, mapping documents, problem solving and trouble shooting.
  • Performed unit testing at various levels of the ETL and actively involved in team code reviews.
  • Designed and implemented historic, incremental (type 1 and type 2), delete mappings Using Pushdown optimization, TPT loader for Teradata.
  • Used various transformations in Informatica Designer including Expression, Aggregator, Joiner, Transaction control, Filter, Lookup, Dynamic Lookup, Update Strategy, Sorter, Sequence Generator, Router, Stored procedure etc.
  • Using Parameter files, workflow and mapping variables in mappings to make them easy to maintain in future.
  • Used various tasks in Workflow Manager including Session, Timer, Event wait, Event Raise, Decision, Link condition, Email, Command task etc.
  • Integrated data to Data warehouse from different LIMS systems, Weather data and experimental data. Created Tables, Views, DB Links, Indexes, Partitioned Tables, Procedures, Triggers, Sequence and Functions.
  • Developed database models using star and snowflake schemas and implemented them on database level.
  • Worked to tune the queries to improve the performance of code. Analyzed tables, used explain plan and used hints.
  • Refactor all the existing views and procedures to improve efficiency and performance.
  • Created best practices and standards document for the team to follow.
  • Designed and developed dimensional schema for green house data.
  • Profile, Integrate and manage data duality on data coming from various databases like SQL Server, Sybase, Excel and Oracle databases.
  • Load all the required data in reporting database and schedule the loads as per refresh rate requirements.
  • Performed POC for various tools and appliances.
  • Working on Implementing Amazon Web Services for creating subscriptions to different applications.
  • Lead the Production support tier 2 and participate in on-call rotation, job monitoring and resolving data issues.

Confidential, Minneapolis, MN

SQL/SSRS/BI Developer

Environment: SQL Server 2008/2005, Oracle-10g, 11g, T-SQL/PL-SQL, Windows XP, Linux, Visual Studio 2008/2005, SVN, Ankh SVN, Report Builder 2.0, Toad, Reporting Tools, Business Objects, Crystal reports, Report Manager, SharePoint Site, Quality Center, SSIS, SSRS, Dimensional Database, Microsoft-Office Suite, Hadoop, Hive.

Responsibilities:

  • Built various SSIS packages having different tasks and transformations for various business areas and Scheduled SSIS packages
  • Created different reports along with shared Data Sources using SSRS.
  • Created various functions, views, procedures using PL/SQL and Complex Datasets with SQL.
  • Created Parameterized reports and build Single-valued, Multi-valued, Drop- down, Textbox parameters.
  • Created reports including KPIs, metrics and measures for required business objectives.
  • Developed Cascaded Parameters for various reports.
  • Designed drill-down and drill-through reports in SSRS, drill-through reports that can navigate to other reports or to other URL.
  • Made subscriptions for reports through Report Manager.
  • Designed SSIS work-flows and standards for various ETL jobs and Data transformations.
  • Involved in database modification process and data re-modeling process for Sales, Orders
  • Created SSIS packages to move data from Oracle to SQL Server and also from other sources including CSV, Text, Access and Excel files
  • Involved in Resolving advanced and complex application bugs and configuration issues.
  • Migrated reports from SQL Server 2005 Version to SQL Server 2008 Version
  • Applied Version control to the Project. Used Ankh SVN on client machines.
  • Augmented the system Triggers and defined the Triggers to perform the required action.
  • Configured and setup Autosys for scheduling Jobs with specific logic.
  • Used Quality Center to assign the bugs, defects, issues and the update the status of them. Also to retrieve the requirement documents.
  • Load and transform large sets of structured, semi structured and unstructured data.
  • Involved in loading data from LINUX file system to HDFS
  • Importing and exporting data into HDFS and Hive.
  • Implemented Partitioning, Dynamic Partitions, Buckets in Hive.
  • Worked in creating HBase tables to load large sets of semi structured data coming from various sources.
  • Experienced in running Hadoop streaming jobs to process terabytes of xml format data.

Confidential, NJ

SQL/BI Developer

Environment: SQL Server 2008/2005, Oracle-10g, T-SQL/PL-SQL, Windows XP/2003, Visual Studio 2008/2005, TFS2008/2005, Report Builder 2.0, Crystal Reports, Erwin, SSIS, SSRS, SSAS, MS-Visio, Share-Point Designer, MOSS-2007, Microsoft-Office 2003/2007, Rational Clear Quest and Rational Clear case

Responsibilities:

  • Involved in gathering business requirements from business and clients to develop various reports and cubes
  • Created functional requirement specifications and supporting documents for business systems
  • Designed SSIS work-flows and standards for various ETL jobs and Data transformations.
  • Used Agile Methodology in various projects.
  • Developed complex SQL and Database objects in SQL Server 2008, Oracle 10g, 11g
  • Involved in database modification process and data re-modeling process for Sales, Orders
  • Created SSIS packages to move data from Oracle to SQL Server and also from other sources including CSV, Text, Access and Excel files
  • Built various SSIS packages having different tasks and transformations for various business areas and Scheduled SSIS packages
  • Migrated reports from SQL Server 2005 Version to SQL Server 2008 Version
  • Created Partitions in Cubes on time dimension for optimizing the performance
  • Designed Hierarchies by setting name column and key column as there was no strong hierarchies defined
  • Designed drill-down and drill-through reports in SSRS, drill-through reports that can navigate to other reports or to other URL
  • Migrated reports from SQL Server 2005 Version to SQL Server 2008 Version
  • Published reports on to Share-Point and Report Server and set the data-driven subscriptions for these reports
  • Used MS-Office Web Components to show data directly from the cube on a web-browser as a OWC Report

Confidential, Atlanta, GA

SQL Server / BI Consultant

Environment: SQL Server 2008/2005, T-SQL, Windows XP/2003, Visual Studio 2008/2005, Report Builder 2.0, SQL Profiler, Erwin, SQL Server Reporting services, SQL Server Analysis Services, Visio, Share-Point.

Responsibilities:

  • Worked as part of a team for gathering and analyzing Business Requirements
  • Participated actively in designing the system database structure
  • Designed Data-Mart for Prescription using Relational Model and wrote SSIS Packages to extract data from the old PDX system and load it into the new data mart of SQL Server 2008
  • Designed SSIS packages to migrate the data from PDX files to the Staging Area in SQL server 2008
  • Developed SSIS Packages for migrating data from Staging Area of SQL Server 2005 to SQL Server 2008.
  • Installed and administered Microsoft SQL Server 2008/2005, SQL Server Integration Services 2008/2005, SQL Server Reporting Services2008/2005 and Team Foundation Server 2005/2008.
  • Built Reports in SSRS for errors generated in SSIS Packages.
  • Generated Reports using Data Source Views in Report Builder 2.0
  • Involved in the development of custom stored procedures, functions, triggers, SQL, T-SQL
  • Involved in the optimization of SQL queries which resulted in substantial performance improvement for the conversion processes
  • Performed performance tuning of stored procedures using SQL Server Profiler
  • Used Query Analyzer, Profiler, Index Wizard and Performance Monitor for performance tuning on SQL Server
  • Involved in creating and maintaining SQL Server Analysis Services.

Confidential

SQL Server / BI Developer

Environment: MS SQL Server 2005/2000, T-SQL,UNIX, Excel, Access, Reporting Services, Analysis Services, DTS, Data Analyzer, Visual Studio 2005, Crystal Reports 8.0, Windows XP/2000

Responsibilities:

  • Extracted large volumes of data from different data sources and loaded the data into target data sources by Performing different kinds of transformations using SQL Server Integration Services (SSIS).
  • Experience in SSIS script task, look up transformations and data flow tasks using T- SQL and Visual Basic (VB) scripts. .
  • Involved in SQL joins, sub queries, tracing and performance tuning for better running of queries.
  • Involved in Error Handling using try and catch blocks and performance tuning using counters in SSIS.
  • Participated in developing Logical design of database incorporating business logic and user requirement
  • Created stored procedures, functions, triggers (database objects) and called them in the SSIS packages.
  • Involved in designing, building and deploying multidimensional cubes using SQL Server Analysis Services
  • Used best practices method to build the cube with respect to the performance
  • Worked with advance properties of the cubes like calculations, partitions and aggregations
  • Configured database mail, created operators, jobs, alerts for automating databases
  • Ad-hoc report design, grouping and sorting using Visual Studio 2005 and also importing sub-reports into it.

We'd love your feedback!