We provide IT Staff Augmentation Services!

Sr. Data Engineer Resume

2.00/5 (Submit Your Rating)

CaliforniA

SUMMARY

  • 15+ years of overall experience of designing and implementing data information systems empowering analytical decision meeting industry standard performance, scalability and reliability.
  • 15+ years of progressive experience with architecting, designing and developing all parts of Data Warehouses (ETL, data modeling, database design, metric design and reporting /data visualization) using RDBMS as well as Big Data for large - scale complex business information collaborating with product managers, data scientists, business users
  • Data warehousing concepts, including dimensional data models, STAR/SNOW-FLAKE Schema, and ETL
  • Big Data strategy, architecture, design, implementation and integration.
  • Hands on Experience working on the analytics, reporting and visualization of the data using Microstrategy.
  • Identify data quality issues and their root causes. Propose fixes and design data audits
  • Proven track record of SQL tuning/optimization using Oracle Tools (Explain Plan, SQL Trace, TKPROF, Extended trace events).
  • Performance tuning of Hive queries/ Big Data processes.
  • Performance tuning for high volume Terabytes of Data loads and aggregation in most optimized time
  • Expert knowledge in writing highly optimized SQL/Analytical queries for Data Warehousing and aggregation/Data Mining.
  • Ensure validity and integrity of data in the Data Warehouse and BI layers.
  • Understand data, data quality and best practices to derive value of data empowering stake owners to make more informed decisions.

TECHNICAL SKILLS

Big Data: AWS, Hadoop, Map Reduce, Hive, Pig and Sqoop with Oraoop, Amazon S3, Amazon EMR, Spark

Languages: Languages SQL, Oracle PL/SQL, Shell Scripting(bash), Perl, Python

Database: Teradata, Oracle, MySQL, Amazon Redshift

Reporting: Microstrategy, OBIEE, Tableau

ETL Tools: Abinitio, SQL*Loader, Sync Sort, SQL*Plus

Other Tools: Toad, UC4, SQLTrace (with extended trace events), Power Designer

PROFESSIONAL EXPERIENCE

Confidential, California

Sr. Data Engineer

Responsibilities:

  • Involved in converting Hive SQL queries into Spark transformations using Spark RDD's, PYSPARK
  • Used Spark API over Amazon EMR, YARN to perform analytics on data in Hive.
  • Python scripts to load data from RDBMS to S3 using Sqoop.

Confidential, Confidential, California

Lead Data Warehouse Engineer

Responsibilities:

  • Technical Lead for enterprise data warehousing initiatives by driving requirements gathering, analysis, design, functional/technical specification, development, deployment and testing.
  • Support, maintenance and enhancement, bug fixing to existing data warehouse using Amazon Redshift, Syncsort, Amazon S3 and Python/Shell scripts.
  • Designed a new ETL pipe line for ingestion of transaction data from Amazon S3 to EMR using hive and Spark for data Science team.
  • Exploring Spark for improving performance and optimization of the existing algorithms in Hadoop using Spark Context, Spark-SQL, Data Frame and RDD.
  • Partnered with analysts and stakeholders to gather, understand and develop technical requirements and plan projects from concept to completion.
  • Developed a new data mart for Inventory management system using Oracle, Amazon Redshift and bash scripts.
  • Developed a python script to load data in an automated way from Google sheets to Amazon Redshift.
  • Developed data model for landed cost for the retail items and integrated with existing finance data mart using Amazon Redshift
  • Developed a data mart for fraud orders and new dynamic reporting on existing current data to help fraud team to catch the fraudulent orders.
  • Developed a new report for Fit and sell shop using SQL, Redshift.
  • Performance improvement of tableau data extracts/Live queries.

Confidential, California

Sr. Data Warehouse/BI Architect

Responsibilities:

  • Responsible for the overall BI & DW solution architecture, design, development and implementation leveraging Enterprise BI & DW platform and architecture.
  • Designed a new data mart from scratch using Oracle PL/SQL, Hadoop/Hive and shell scripts.
  • Tuned existing Materialized Views queries to run in 30 minutes from 7 hours.

Confidential, California

Sr. Data Warehouse/BI Architect

Responsibilities:

  • Technical lead, designer, developer for numerous data warehousing, financial processing, data visualization and reporting infrastructure back end applications.
  • Tuning of the ETL process to scale the increased Data capacity and met the defined SLA for reporting
  • Collaborated with engineering teams, business teams, product managers, data scientists and BI analysts to understand data requirements and design end to end solution right from data acquisitions from multiple sources to data visualization using Oracle, Teradata, Hadoop Hive, Sqoop and Perl/shell scripts and Microstrategy.
  • Worked with data scientist team to develop data load processes to acquire data from multiple sources and transformation and push to Hadoop/Hive to meet their ongoing need for algorithm development.
  • Designed and developed the process for the data integration with Confidential using Oracle, Teradata, bash scripts and Hive.
  • Identified and addressed root cause for data quality/anomalies issues in data marts.
  • Identified and addressed performance/bottlenecks issues in data marts ETL process for Microstrategy based data marts.
  • Designed and developed and modified existing DWH ETL process to cater for a new attribute device type (Smart Phone/Tablet) in all data marts.
  • Develop a framework for financial process and dependent DWH processes to handle backlog gracefully in automated way using Oracle PL/SQL.
  • Designed and developed 3 new DataMart for sku Analytics reporting using Microstarategy, Hadoop/Hive, shell scripts and Oracle PL/SQL
  • Designed and developed a process for uploading finance data to Oracle DB (input Excel sheet) using Oracle Express.
  • Tech lead for DWH Oracle 11g upgrade. Implemented successfully and addressed query performance issues.

Confidential, San Jose California

Data Warehouse Architect

Responsibilities:

  • Architected and build prototypes A Data Warehouse model for Cisco's devices to track hardware failures and overall performance of Cisco Supplied hardware using SQL-LOADER (DIRECT PATH), PL/SQL and Oracle processing upward of 10,000 transactions per second. Designed of ETL architecture, Dimension and facts tables.
  • Installed Hadoop/Hive on 5 node clusters for POC. Developed process for extracting data from Oracle to HDFS.

Confidential, California

Sr. Database Architect

Responsibilities:

  • Responsibilities include Database Architecture, Design, Code Review and Performance Testing & Maintenance of Production Systems.
  • Maintained, bug fixing, enhancement, performance tuning and production support for Confidential ’s ad .system using back end PL/SQL, external tables.
  • Designed and implemented a framework for publishing daily average 10 TB production data to GRID/HADOOP data pipe lines using PL/SQL, Oracle scheduler and perl scripts.
  • Supported /Enhanced and maintained home-grown ETL system including Mat views, Partitions, Parallelism, Data Replication, IOT, Advanced SQL Analytical Functions, External Tables.
  • Architecture, design, performance tuning and DB related support to DOTS pipeline, which feeds serving from ad system.

Confidential, California

Sr. Database Engineer

Responsibilities:

  • Maintain/bug fixing/enhancement/performance tuning and production support for homegrown ETL processes using PL/SQL, SQL-Loader and Perl/Shell scripts data warehouse ETL processes.
  • Lead ETL architecture, design, development and implementation for migrating existing Microsoft cube server data marts to Microstrategy reporting environment. Designed new Snow-flake schema using dimensional modeling for each DataMart. Developed ETL Framework using ab initio graphical development environment tool and support. Tuned Microstrategy queries for optimal performance. Developed Perl/Unix wrapper scripts to run ab initio graph for all data marts for monitoring. Use Appworx to schedule ETL processes.
  • Owned and supported the Business Intelligence Reports PL/SQL code, bug fixing and enhancement. Extensively tuned the PL/SQL packages using hints/SQLTrace/tkprof and rewriting long running queries.
  • Mentor database engineers and front end developers for performance tuning in oracle database.

Confidential, California

Data Warehouse Lead

Responsibilities:

  • Designed and developed Marketing Performance Dashboard data warehouse using Oracle database Oracle Warehouse Builder. Implemented multi-dimensional database using logical and physical data modeling techniques (Star schema) . ETL mapping to source system and data warehouse
  • Developed applications(ETL processes) that extract/transform/load/cleanse data from various source systems to the Data Warehouse using PL/SQL, SQL-LOADER, Perl, Shell scripts . Tuned /optimizeed above packages using SQLTrace and tkprof .Worked on Oracle express to maintain Analytical Workspace AW/OLAP data structures and AW loading processes.

We'd love your feedback!