Data Architect Resume
SUMMARY
- He is an expert level Data Architect/ETL Architect/Data Modeler with 14 years of system analysis, ETL design, and implementation experience of numerous ETL Projects.
- Data Steward with hands on experience in design of the data movement from source system to Data Lake, Data warehouse to enable reporting and analytics capabilities using Big Data platform
- Strong expertise using ETL tools such as SSIS, Informatica 9.x, Ab Initio, PL - SQL & Teradata
- As a Data Architect my core strengths include but not limited to:
- Requirements gathering and business rules analysis Logical/Physical database modeling.
- ETL Strategy & ETL Architecture Design.
- Data Modeling and Data Architecture
- Data Profiling using SQL and tools.
- Big Data Hive Data warehouse design
- Big Data ETL design for ingestion, Hive Data warehouse load and export to DWH.
- Master Data Management and Reference Data Management
- Extraction/Transformation/Loading (ETL) design and implementation
- Error and Exception management
- ETL & SSIS Package Performance Tuning.
- Design of ETL mapping & Tech Specifications
- Design of ETL Auditing and Balancing mappings
- Design of complex SQL queries for ETL.
- Performance tuning of SQL queries for ETL
- Design ETL Architecture for ETL Systems
- Design of PowerShell scripts to process flat files
- Unit testing, Integration testing of Data to assure Data Quality
- Product Evaluations
- Work with vendors to establish the required infrastructure and frameworks for ETL and other tools
- Team leadership/mentoring/training
- Coordination of offshore development efforts
- Liaison between business users and development staff
- Shell scripting in Unix and ETL process Automation
- Iterative development processes
- “Agile” development processes
TECHNICAL SKILLS
Primary Skills: Business Problem Analysis, ETL Architecture and Design, ETL Development, Leading and Mentoring ETL Developers, ETL Development
Techniques: OOA, D & Programming, Entity-Relationship (ER) Modeling, Agile techniques, Hive ETL, Hive Warehouse Design, BIG Data
Analysis/Design Tools: ER/Win
Database Platforms: Oracle, DB2, MS SQL Server
Languages/Tools: SQL/PLSQL, INFORMATICA, SSIS, Business Objects, UNIX Shell, PowerShell, Java.
PROFESSIONAL EXPERIENCE
Confidential
Data Architect
Responsibilities:
- Understand new source system FIS
- Work with Subject Matter Experts to clarify open and scenario questions.
- Based on the source system design updates to the existing data models to be able to house data from FIS.
- Design change to keys/granularity of tables to accommodate new data.
- Design new tables and relationships to accommodate new business process.
- Add/Drop columns to the data model to achieve reporting requirement.
- Analyze data files received from Fiserv and Design data model for them.
- Get Data model reviewed and approved by the Modeling lead.
- Attend Design meeting to discuss open issues.
- Attend Code review to review code created by Developers.
- Write SQL scripts to dedup data and re-load them back into tables after undergoing data transformation.
- Tune SQL queries that will be embedded into applications and Business Objects
- Deploy DDL in DEV and QA
- Check in DDL in Visual Studio and version and label them as required.
Tools: and Technologies: Informatica 9.x, Oracle 11g, PL-SQL, UNIX, Erwin, SQL Server
Confidential
Data Architect
Responsibilities:
- Design of the data movement from source system to Data Lake, Data warehouse to enable reporting and analytics capabilities using Big Data platform
- Observed the Set up and monitoring of a scalable distributed system based on HDFS
- Code reviewed and understood MapReduce jobs in java and use of different Hive UDF's for data cleaning and processing.
- Understood business requirements from the BSA and Business representative.
- Extracted the data from Oracle database into HDFS using SQOOP. more efficient data use
- Used HIVE queries for aggregating the data and mining information sorted by volume and grouped by vendor and product.
- Profiled source data to better understand the source data.
- Work with Source System SME to understand the source system tables
- Design the presentation data marts for reporting on user activity
- Extracted the data from Hive to Oracle for reporting.
- Schedule jobs using Oozie.
Tools: and Technologies: Hadoop (HDFS/MapReduce), HIVE, SQOOP, Oracle, Linux, Oozie
