We provide IT Staff Augmentation Services!

Sr. Data Architect/modeler/engineer Resume

0/5 (Submit Your Rating)

Atlanta, GA

SUMMARY:

  • 9+ years of IT industry experience in Application Design, Development, and Data Management - Data Governance, Data Architecture, Data Modeling, Data Warehousing and BI, Data Integration, Meta-data, Reference Data and MDM.
  • Experience as Architect UML models and leverage the advanced executable code generators to target different domains.
  • 3+ YearsIBM DataStageCoding Experience
  • 2+ Years Unit and Functional Testing and Debugging.
  • Strong Experience in Big Data Hadoop Ecosystem in ingestion, storage, querying, processing and analysis of big data.
  • Experience in Dimensional Data Modeling, Star/Snowflake schema, FACT & Dimension tables.
  • Experience in Mater Data Management (MDM) Multi Domain Edition HUB Console Configurations, Informatica Data Director (IDD) Application creations.
  • Extensive experience in Master Data Management (MDM) HUB Console Configurations such as Staging Process Configurations, Landing Process Configurations, Match and Merge Process, Cleansing Functions, User Exits.
  • Experience in Designing, Developing, Documenting, Testing of ETL jobs and mappings in Server and Parallel jobs using Data Stage to populate tables in Data Warehouse and Data marts.
  • Expertise in UNIX shell scripts using K-shell for the automation of processes and scheduling the Data Stage.
  • Knowledge of programming languages like Java & Python.
  • Experience wif emerging technologies such Big Data, Hadoop, and NoSQL.
  • Strong experience in analyzing/ Data Transformation of large amounts of data sets writing Pig scripts and Hive, AWS EMR, AWS RDS. Extensive noledge in Hadoop stack components viz. Apache Hive, Pig Scripting, etc.
  • Experience in analyzing data using Hadoop Ecosystem including HDFS, Hive, Spark, Spark Streaming, Elastic Search, Kibana, Kafka, HBase, Zookeeper, PIG, Sqoop, and Flume.
  • Hands on experience in Normalization (1NF, 2NF, 3NF and BCNF) Denormalization techniques for TEMPeffective and optimum performance in OLTP and OLAP environments.
  • Experience in cloud development architecture on Amazon AWS, EC2, EC3, Elastic Search, Redshift and Basic on Azure.
  • Experience in BI/DW solution (ETL, OLAP, Data mart), Informatica, BI Reporting tool like Tableau and QlikView and also experienced leading the team of application, ETL, BI developers, Testing team.
  • Experience wif Agile Extreme Programming (XP) development and Scrum lifecycle practices, or a strong desire to learn including: pair programming, test driven development, continuous integration, iterative delivery, retrospection
  • Good experience in working wif different ETL tool environments like SSIS, Informatica and reporting tool environments like SQL Server Reporting Services (SSRS), Cognos and Business Objects.
  • Knowledge of transformdata, usingdatamapping anddataprocessing inApacheBeam.
  • Proficient in UML Modeling like Use Case Diagrams, Activity Diagrams, and Sequence Diagrams wif Rational Rose and MS Visio.
  • Experienced in Technical consulting and end-to-end delivery wif architecture, data modeling, data governance and design - development - implementation of solutions.
  • Solid noledge of Data Marts, Operational Data Store (ODS), OLAP, Dimensional Data Modeling wif Ralph Kimball Methodology (Star Schema Modeling, Snow-Flake Modeling for FACT and Dimensions Tables) using Analysis Services.
  • Solid hands-on experience in creating and implementation of Conceptual, Logical and Physical Models for Online Transaction Processing and Online Analytical Processing. Efficient in developing Logical and Physical Data model and organizing data as per the business requirements using Erwin, ER Studio in both OLTP and OLAP applications.
  • Worked on data modeling using ERWIN tool to build logical and physical models.
  • Skillful in Data Analysis using SQL on Oracle, MS SQL Server, DB2 & Teradata.
  • Extensive experience in development of T-SQL, Oracle PL/SQL Scripts, Stored Procedures and Triggers for business logic implementation.
  • Decode the Teradata and SQL queries to find all the data attributes involved and document for the purpose of development.
  • Excellent understanding of Hub Architecture Style for MDM hubs the registry, repository and hybrid approach.
  • Mapping the Risk Data elements to the Authoritative Data Source and documenting the Schema, Database, Table details for data modelling purpose.
  • Good exposure on usage of NoSQL database.
  • Experienced in understanding the ETL framework metadata to understand the current state ETL implementation.

TECHNICAL SKILLS:

Data Modeling Tools: Erwin R6/R9, Rational System Architect, Confidential Info sphere Data Architect, ER Studio and Oracle Designer.

Database Tools: Microsoft SQL Server12.0, Teradata 15.0, Oracle 12c/11G/9i, and MS Access

BI Tools: Tableau 7.0/8.2, Tableau server 8.2, Tableau Reader 8.1, SAP Business Objects, Crystal Reports

Big Data: PIG, Hive, HBase, Spark, Sqoop, Flume.

Cloud Platforms: AWS EMR, AWS RDS, EC2, S3, Azure.

Packages: Microsoft Office 2010, Microsoft Project 2010, SAP and Microsoft Visio, Share point.

Operating Systems: Windows, Centos, Sun Solaris, UNIX, Ubuntu Linux

Version Tool: VSS, SVN, CVS, SAP BO 4.1

Tools: & Utilities: TOAD 9.6, Microsoft Visio 2010.

Methodologies: RAD, JAD, RUP, UML, System Development Life Cycle (SDLC), Waterfall Model.

PROFESSIONAL EXPERIENCE:

Confidential, Atlanta, GA

Sr. Data Architect/Modeler/Engineer

Responsibilities:

  • Owned and managed all changes to the data models. Created data models, solution designs and data architecture documentation for complex information systems.
  • Architected, researched, evaluated and deployed new tools, frameworks, and patterns to build sustainable Big Data platforms for our clients.
  • Designed, Installed, Configured core Informatica Master Data Management (MDM) Hub components such as Informatica Master Data Management (MDM) Hub Console, Hub Store, Hub Server, Cleanse Match Server, Cleanse Adapter & Data Modeling.
  • Created the Mater Data Management (MDM) Hub Console configurations like Stage Process configuration, Load Process Configuration
  • SIF API was created to access Master Data Management (MDM) Hub from External Application.
  • Designed the Logical Data Model using ERWIN 9.64 wif the entities and attributes for each subject areas.
  • Used Tableau for BI Reporting and Data Analysis.
  • Working as a Sr. Data Architect/Modeler to generate Data Models using Erwin r9.64 and developed relational database system.
  • Design and developed architecture for data services ecosystem spanning Relational, NoSQL, and Big Data technologies.
  • Architected solutions using MS Azure PaaS services such as SQL Server, HDInsight, service bus, etc
  • Implemented CI/CD based application development methodology using tools like Jenkins/TFS/powershell etc.
  • Provided technical oversight and guidance during clients engagement execution
  • Provided Cloud / Azure thought leadership through regular publications and speaking engagements
  • Develop and maintain data architecture, including master data and data quality, using Toad Data Modeler and Microsoft Master Data Manager (MDS) as well as Oracle Data Integrator.
  • Configured Hunk to read customer transaction data from Hadoop Ecosystems such as HDFS and Hive.
  • Used DataStage as an ETL tool to extract data from sources systems, loaded the data into the ORACLE database.
  • Designed and Developed Data stage Jobs to Extract data from heterogeneous sources, Applied transform logics to extracted data and Loaded into Data Warehouse Databases.
  • Created Datastage jobs using different stages like Transformer, Aggregator, Sort, Join, Merge, Lookup, Data Set, Funnel, Remove Duplicates, Copy, Modify, Filter, Change Data Capture, Change Apply, Sample, Surrogate Key, Column Generator, Row Generator, Etc.
  • Designed facts and dimension tables and defined relationship between facts and dimensions wif Star Schema andSnowflakeSchema in SSAS.
  • Used Flume extensively in gathering and moving log data files from Application Servers to a central location in Hadoop Distributed File System (HDFS) for data science.
  • Involved in Normalization / Denormalization techniques for optimum performance in relational and dimensional database environments.
  • Developed Data Mapping, Data Governance, and Transformation and cleansing rules for the Master Data Management Architecture involving OLTP, ODS.
  • Working wif project management, business teams and departments to assess and refine requirements to design/develop BI solutions using Azure.
  • Created user-friendly and a dynamically rendered custom dashboard to visualize the output data using the Pentaho CDE and CDF.
  • Created various types of chart reports in Pentaho Business Analytics having Pie Charts, 3D Pie Charts, and Line Charts, Bar Charts, Stacked Bar Charts and Percentage Bar charts.
  • Designed and developed architecture for data services ecosystem spanning Relational, NoSQL, and Big Data technologies.
  • Guide Teams (onsite and offshore) in creating Unit and Integration tests. Setup deployment plans and dependency management using Maven. SetupJenkinsContinuous Integration Jobs for automated deployments to integration servers.
  • Collected large amounts of log data using Apache Flume and aggregating using PIG/HIVE in HDFS for further analysis.
  • Created Logical and Physical Data Model using Confidential Data Architect tool.
  • Specifies overall Data Architecture for all areas and domains of the enterprise, including Data Acquisition, ODS, MDM, Data Warehouse, Data Provisioning, ETL and BI.
  • Developed PL/SQL scripts to validate and load data into interface tables
  • Participated in maintaining data integrity between Oracle and SQL databases.
  • Participated in OLAP model based on Dimension and FACTS for efficient loads of data based on Star Schema structure on levels of reports using multi-dimensional models such as Star Schemas and Snowflake Schema.

Environment: Oracle 12c, MS-Office, SQL Architect, Spark, TOAD Benchmark Factory, Teradatav15, Hadoop, SQL Loader, SharePoint, ERwin r 9.64, DB2, MS-Office, SQL Server 2008/2012, Azure, HBase, Hive.

Confidential, Hunt Valley MD

Sr. Data Architect/Modeler/Engineer

Responsibilities:

  • Involve in Data Architect role to review business requirement and compose source to target data mapping documents.
  • Designed and build relational database models and defines data requirements to meet the business requirements.
  • Worked wif Data Steward Team for designing, documenting and configuring Informatica Data Director for supporting management of MDM data.
  • Actively involved in the Design and development of the Star schema data model.
  • Implemented slowly changing and rapidly changing dimension methodologies; created aggregate fact tables for the creation of ad-hoc reports.
  • Created and maintained surrogate keys on the master tables to handle SCD type 2 changes TEMPeffectively.
  • Developed Star andSnowflakeschemas based dimensional model to develop the data warehouse.
  • Setup automatic code review, testing and deployment pipelines usingJenkinsand bitBucket for continuous integration and deployment (CICD).
  • Worked wif the developers in deciding the application architecture.
  • Designed and implemented a Data Lake to consolidate data from multiple sources, using Hadoop stack technologies like SQOOP, HIVE/HQL.
  • Written complex SQL queries for validating the data against different kinds of reports generated by Business Objects XIR2.
  • Designing Logical data models and Physical Data Models using ER Studio.
  • Designed semantic layer data model. Conducted performance optimization for BI infrastructure.
  • Used the DataStage Designer to develop processes for extracting, cleansing, transforming, integrating and loading data into staging tables.
  • Worked wif Metadata Definitions, Import and Export of Datastage jobs using Data stage Manager.
  • Created source table definitions in the DataStage Repository.
  • Involved in the creation, maintenance of Data Warehouse and repositories containing Metadata.
  • Performed Hive programming for applications that were migrated to big data using Hadoop.
  • As an Architect implement MDM hub to provide clean, consistent data for a SOA implementation.
  • Installing and configuring the a 3-node Cluster in AWS EC2 Linux Servers.
  • Designed different type of STAR schemas like detailed data marts and Plan data marts, Monthly Summary data marts using ER studio wif various Dimensions Like Time, Services, Customers and various FACT Tables.
  • Developed and maintained data dictionary to create metadata reports for technical and business purpose.
  • ETL processing using Pig & Hive in AWS EMR, S3
  • Implemented Data Vault Modeling Concept solved the problem of dealing wif change in the environment by separating the business keys and the associations between those business keys, from the descriptive attributes of those keys using HUB, LINKS tables and Satellites.
  • Extensive Data validation by writing several complex SQL queries and Involved in back-end testing and worked wif data quality issues.
  • Data Profiling, Mapping and Integration from multiple sources to AWS S3.
  • Created single value as well as multi-value drop down and list type of parameters wif cascading prompt in the reports.
  • Integrating Kettle (ETL) wif Hadoop, Pig, Hive, Spark, Storm, HBase, Kafka and other Big Data component for various functionalities and other various NoSQL data stores can be found in the Pentaho Big Data Plugin.
  • Design and development of ETL routines to extract data from heterogeneous sources and loading to Actuarial Data Warehouse.
  • Participated in preparing Logical Data Models/Physical Data Models.
  • Identify source systems, their connectivity, related tables and fields and ensure data suitably for mapping.
  • Designed aDataVault for Deal transactions for POC usingSnowflake.
  • Worked wif BTEQ to submit SQL statements, import and export data, and generate reports in Teradata.
  • Worked on HL7 2.x file format (ADT and clinical messages) on MEDIFAX and a thorough understanding of how interface development projects work.
  • Developed company-wide data standards, data policies and data warehouse/business intelligence architectures.
  • Designed and documented Use Cases, Activity Diagrams, Sequence Diagrams, OOD (Object Oriented Design) using UML and Visio.
  • Performed data cleaning and data manipulation activities using NOSQL utility.
  • Designed and Developed Oracle PL/SQL Procedures and UNIX Shell Scripts for Data Import/Export and Data Conversions.
  • Developing the Conceptual Data Models, Logical data models and transformed them to creating schema using ER Studio.

Environment: DB2, ER Studio, Oracle 11g, MS-Office, SQL Architect, Hadoop, Hive, Pig, TOAD Benchmark Factory, Sqoop, SQL Loader, AWS S3, PL/SQL, SharePoint, MS-Office, SQL Server 2014

Confidential, Atlanta, GA

Data Modeler/Data Engineer

Responsibilities:

  • Analyzed the business requirements by dividing them into subject areas and understood the data flow wifin the organization
  • Database Design (Conceptual, Logical and Physical) for OLTP and OLAP systems.
  • Created and developed Slowly Changing Dimensions tables SCD2, SCD3 to facilitate maintenance of history.
  • Created documents for technical & business user requirements during requirements gathering sessions.
  • Turned SQL queries to make use of data base indexes, and analyzed the data base objects.
  • Created Logical and Physical EDW models and data marts.
  • Experienced in data migration and cleansing rules for the integrated architecture (OLTP, ODS, DW).
  • Managed all indexing, debugging and query optimization techniques for performance tuning using T-SQL.
  • Developed the logical and physical model from the conceptual model developed using a tool Erwin by understanding and analyzing business requirements.
  • Handled data loading operations from flat files to tables using NZLOAD utility.
  • Experienced in data cleansing for accurate reporting. Thoroughly analyzed the data and integrated different data sources to process matching functions.
  • Provided Azure technical expertise including strategic design and architectural mentorship, assessments, POCs, etc., in support of the overall sales lifecycle or consulting engagement process
  • Implement solutions using Azure PaaS features like web jobs, cloud services, Azure SQL Server, service bus, notification hubs etc
  • Configure large database solutions in Azure using SQL Server or Oracle database solutions
  • Applied data naming standards, created the data dictionary and documented data model translation decisions and also maintained DW metadata.
  • Created DDL scripts for implementing Data Modeling changes. Created ERWIN reports in HTML, RTF format depending upon the requirement, Published Data model in model mart, created naming convention files, co-coordinated wif DBAs' to apply the data model changes.
  • Extensively used Normalization techniques (up to 3NF).
  • Writing complex queries using Teradata SQL.
  • Worked wif the ETL team to document the transformation rules for data migration from source to target systems.
  • Developed source to target mapping documents to support ETL design.

Environment: ERWIN r7.2, PL/SQL, MS SQL, MS Visio, Business Objects, Windows NT, Linux, Sybase Power Designer, Oracle 8i, SQL Server, Windows, MS Excel, Informatica.

Confidential, Phoenix, AR

Data Modeler/Engineer

Responsibilities:

  • Designed and implemented business intelligence to support sales and operations functions to increase customer satisfaction
  • Developed Data Mapping, Data Governance, Transformation and Cleansing rules for the Master Data Management Architecture involving OLTP, ODS and OLAP.
  • Created logical data model from the conceptual model and it's conversion into the physical database design using Erwin.
  • Analyzed the data and provide resolution by writing analytical/complex SQL in case of data discrepancies.
  • Tuning and code optimization using different techniques like dynamic SQL, dynamic cursors, and tuning SQL queries, writing generic procedures, functions and packages.
  • Responsible for Relational data modeling (OLTP) using MS Visio (Logical, Physical and Conceptual).
  • Designed, developed and implemented solutions wif data warehouse, ETL, data analysis, and BI reporting technologies.
  • Design and development of ETL processes using Informatica ETL tool for dimension and fact file creation
  • Create and execute test scripts, cases, and scenarios that will determine optimal system performance according to specifications.
  • Reverse Engineered DB2 databases and tan forward engineered them to Teradata using Erwin.
  • Tested the database to check field size validation, check constraints, stored procedures and cross verifying the field size defined wifin the application wif metadata.
  • Extensively worked on development of mappings wif BODS Transformations like Map Operation, Table Comparison, History Preserving, Key Generation, Pivot, Reverse Pivot etc.
  • Responsible for design of logical and physical Data model for client's investment management ODS using dimensional modeling.

Environment: Erwin, Oracle 8i, Developer 2000 wif Forms 5.0 and Reports 3.0, O, Windows XP, PL/SQL, MS-Access, Sql Server, MS Office, MS Visio, Informatica Power center 5.1, Teradata SQL Assistant

We'd love your feedback!