Data Engineer Resume
Palo Alto, Ca
SUMMARY:
- Dedicated, result - oriented professional with seven years of progressive hands-on experience in Data Analytics, Data Modeling, ETL processes and Lead Generation/Reporting Tools.
- Extensive experience in acquiring, analyzing and integrating competitive data from multiple sources.
- Develop and support programs that scale analytics-driven lead gen across all account sizes & regions.
- Design and develop automation processes to download and analyze hundreds of files into Hadoop systems.
- Generate reports utilizing third party data such as Double Click to track and analyze digital campaigns.
- Develop Python based APIs to ingest user attributes into Oracle Bluekai Platform for score based media targeting.
- Build, manage and support data platform needs for marketing, sales and data science team.
- Expert understanding of the Marketing and Finance domain - Campaign, Performance, Response, Sales, Revenue, Acquisition, Services and Segments.
- Principled team member known for integrity, enthusiasm, business acumen and technical expertise.
TECHNICAL SKILLS:
Databases: Vertica, Teradata, Oracle 10g/9i/8i, MySQL
Languages: Python, SQL, HTML
Reporting Tools: Tableau, Hyperion Essbase, Business Object webI
Tools: Aqua data studio, Toad, SQL Developer, SQL Assistant, Visio, Erwin, TeamForge
ETL Tools: Informatica Power Center 9.1/8.x
Environment: Unix, Windows 7/XP/Vista, Linux, Mac OS X
IT Concepts: OLAP, OLTP, SDLC, Data Structures, Agile Methodology
Big Data: Hadoop, HDFS, Hive, Sqoop
PROFESSIONAL EXPERIENCE:
Confidential, Palo Alto,CA
Data Engineer
Responsibilities:
- Develop and automate end to end process for downloading ~200 files of Double Click raw data from google storage into Hadoop on a daily basis. The framework was developed in Python.
- Implement Partitioning in HIVE for efficient data access.
- Create Hive aggregated tables to help build tableau dashboard that provide high level view on digital campaign performance with respect to clicks, impressions and media spend.
- Written Hive queries for data analysis and extract reports to meet the business requirement.
- Export rollup tables using Sqoop from HDFS to Vertica on daily basis.
- Generate reports with calculated metrics like media cost, click rate, cost per click to analyze digital marketing campaign performance.
- Develop Python based API (RESTful Web Service) to generate the authentication signature, construct the user data API request URL, and make the HTTP call to upload user data into Oracle Bluekai DMP for score based media targeting.
Confidential, Palo Alto,CA
Data Engineer
Responsibilities:
- Assess and acquire competitive data from multiple vendors based on coverage by location, account, industry segment and industry vertical.
- Create ER diagrams, data models and data objects (DDL) to load and integrate competitive data from multiple sources.
- Define the workflow and implement the Rules Engine in InfoR to automate and enhance the lead generation program across Marketing and Sales
- Automate end-to-end process of data collection, cleansing and loading; built using Python framework.
- Support Tableau dashboard implementation to track and analyze campaign performance.
Environment: Vertica, MySQL, Aqua Data Studio, InfoR, Python2.7, PyCharm, Hadoop, HDFS, Hive, Sqoop, Tableau, MS Excel, TeamForge, Visio
Confidential,San Jose,CA
ETL Developer
Responsibilities:
- Actively participate in all phases of SDLC: Requirement Analysis, Design, Coding, Testing and Documentation.
- Coordinate with various cross functional teams such as Enterprise Data Warehouse (EDW), System Admins and Platform Services to ensure smooth end-to-end process.
- Identify and resolve bottlenecks in source, target transformations, mappings and sessions to improve performance.
- Create and document ETL Test Plans, Test Cases, Expected Results, Assumptions and Validations.
- Extensive involvement with the Quality Assurance team for building test plans and exhaustive set of test cases.
- Participate in Integration Testing, User Acceptance Testing (UAT) and Functionality Testing.
Environment: Hyperion Essbase, BO WebI, Informatica Power Center 9.1, Teradata, Oracle10g, SQL Assistant, Visio, Toad, MS Excel, Power Point
Confidential, Bellevue, WA
ETL Developer
Responsibilities:
- Worked on creating staging Tables, Constraints, Indexes, and Views.
- Interpret and convert business processes into Informatica mappings.
- Used various Transformations like Joiner, Aggregator, Expression, Lookup, Filter, Update Strategy, Stored Procedures, and Router while developing the ETL mappings.
- Responsible for retrofitting the code to QA environment, and extending the support for the QA and UAT for fixing the bugs.
Environment: Informatica Power Center 8.6.1, Oracle 10g, Toad, SQL Developer, Shell Scripts
