We provide IT Staff Augmentation Services!

Data Analyst Resume

5.00/5 (Submit Your Rating)

Dallas, TX

SUMMARY

  • Over 8+ years of IT experience and technical proficiency as Data Analyst.
  • My business users often complement me by saying - “Compelling story, yet easy to understand”, “Thank you, for the super-fast analysis”, “I use your dashboards every day, it saves me lot of time”, “Your code is efficient and bug free”.
  • As a seasoned data analyst, I supported data teams in building - ETL systems, SQL queries, data modeling, data validation, data governance, enterprise reporting, business intelligence capabilities, and machine learning systems.
  • Extensive experience in ETL and business intelligence tools such as - Tableau, Informatica, and Alteryx
  • Hands - on experience on performing data analysis on the data stored in variety of systems such as - Oracle, MS Access, SQL server, MongoDB, Redshift, and Snowflake.
  • Well versed with big data on AWS cloud services i.e., EC2, S3, Glue, Athena, EMR and RedShift
  • Expert at scoping the business problem, gathering the appropriate data, blending various data sources, and telling compelling stories using stellar visuals
  • Extensive knowledge in Business Intelligence and Data Warehousing Concepts with emphasis on ETL and System Development Life Cycle (SDLC).
  • Experience using Python to perform - advance data blending and developing data pipelines for data science teams.
  • Developed several enterprise scale business intelligence systems in tools such as Tableau and Power BI.
  • Optimized the BI user experience by designing - aggregate tables, Tableau extracts, scheduled reports, etc.
  • Experience in Performance Tuning and Debugging of mappings and sessions. Strong in optimizing the Mappings by creating/using Re-usable transformations, Mapplets and PL/SQL stored procedures
  • Proficient with Data Warehouse models like Star Schema and Snowflake Schema.
  • Involved in all aspects of ETL-requirement gathering, coming up with standard interfaces to be used by operational sources, data cleaning, coming up with data load strategies, designing various mappings, developing mappings, unit testing, integration testing, regression testing and UAT in development
  • Experience working on a variety of enterprise systems such as - Salesforce, Oracle ERP, Workday HRMS, and Peoplesoft HRMS.
  • Data Processing Experience in Designing and Implementing Data Mart applications, mainly Transformation Process using Informatica.

TECHNICAL SKILLS

ETL Tools: Informatica, Alteryx, Python

Languages: SQL, PL SQL, MS SQL, Python, R

Databases: Teradata, Oracle, DB2/UDB, SQL Server, MS Access, Snowflake, MongoDB, Redshift

Operating Systems: Windows 95/98/NT/2000/XP, UNIX, Linux, NCR MP-RAS UNIX

Data Modeling: Erwin, ER Studio

Scheduling tools: Control M, Autosys

Business Intelligence: Tableau, Power BI, Alteryx, Excel

SERVERS: Windows 2003, 2008, 2012R, Microsoft SQL Server 2012/2008-R 2/2008/2005.

OTHER SKILLS: Unix Scripting/Linux scripting, HTML, XML

PROFESSIONAL EXPERIENCE

Confidential, Dallas, TX

Data Analyst

Responsibilities:

  • Coordinate with Business partners, end users and MDs to gather requirements for their data needs.
  • Perform data analysis and gather data from multiple sources to assist the end user’s business need.
  • Perform extensive data analysis and data validations to ensure highest levels of data quality and integrity.
  • Responsible for redesigning the existing process and infrastructure to pull data from Azure data lake.
  • Rewrite all the existing SQL queries in Snowflake.
  • Design and Implement data strategies in TDV (Tibco Data Virtualization tool).
  • Write SQL scripts to pull data from different sources to assist business functions
  • Perform data mining in various data warehouses and data marts to integrate the data into TDV.
  • Maintain data environments and set up new infrastructure as needed.
  • Responsible for quality assurance by configuring test plans in SOAP UI, executing test cases and first level triage of defects found.
  • Create and maintain ETL process in SSIS.
  • Develop dashboards in Tableau.
  • Conduct peer testing, code review and impact analysis of the scripts developed by team members.
  • Act as a release engineer to deploy to lower environments (INT and ACP) and coordinate with DBA team for Prod Deployment.

Environment: TDV (Tibco Data Virtualization), DB Visualizer, SQL Server, SSIS, Tableau, Snowflake, Microsoft Azure, Salesforce, TFS, SOAP UI, Putty.

Confidential, Philadelphia, PA

Data Analyst

Responsibilities:

  • Implemented Spark using Python and utilized Data frames and Spark SQL API for processing and querying data.
  • Developed Spark Applications by using python Driver and Implemented Apache Spark data processing project to handle data from various RDBMS and Streaming sources.
  • Worked with data team in designing and building a multi-terabyte, full end-to-end MPP Data Warehouse infrastructure to handle several records.
  • Experienced in working real time streaming with Kafka as a data pipeline using python, spark streaming module.
  • Consumed Kafka messages and loaded data into Cassandra cluster deployed in containers.
  • Developed python scripts to parse and transform Jason files. Proficient in deploying Kubernetes cluster using Helm Charts.
  • Managed Kubernetes cluster using Helm charts, contributed to deployment and service files, managed release of Helm Charts.
  • Assisted in Key space creation, Table creation, Secondary Index creation in Cassandra database.
  • Performed Query optimization of the tables through load testing using Cassandra stress tool.
  • Worked with various teams in deploying containers on site using Kubernetes to run Cassandra clusters in Linux Environment.
  • Familiar in developing Materialized view, denormalization and aggregation for analysis needs.
  • Good Knowledge on DevOps, Kubernetes architecture like scheduler, pods, nodes, kubectl api and etcd database.
  • Experience in creating tables in Cassandra clusters, building images using docker and deploying the images.
  • Assisted in developing and creating schemas, tables for Cassandra cluster to ensure good query performance for the front-end Application.
  • Experience in managing MongoDB environments from availability, performance, and scalability perspectives.
  • Worked on the Ad hoc queries, Indexing, Replication, Load balancing, Aggregation in MongoDB.
  • Used GitHub as a version control.
  • Worked on the UNIX environment, developing bash shell scripts running CRON jobs.

Environment: Python, Spark, Kafka, JSON, GitHub, LINUX, Flask, Nginx, REST, CI CD, Kubernetes, Helm, MongoDB, Cassandra.

Confidential, San Jose, CA

Data Analyst

Responsibilities:

  • Created and analyzed business requirements to compose functional and implementable technical data solutions.
  • Identified integration impact, data flows and data stewardship.
  • Created new data constraints and or leveraged existing constraints for reuse.
  • Created data dictionary, Data mapping for ETL and application support, DFD, ERD, mapping documents, metadata, DDL and DML as required.
  • Anticipated JAD sessions as primary modeler in expanding existing DB and developing new ones.
  • Evaluated and enhanced current data models to reflect business requirements.
  • Generated, wrote, and run SQL script to implement the DB changes including table update, addition or update of indexes, creation of views and store procedures.
  • Consolidated and updated various data models through reverse and forward engineering.
  • Compared different WFHM DB environments and determined, resolved, and documented discrepancies.
  • Analyzed DB discrepancies and synchronized the Staging, Development, UAT and Production DB environments with data models.
  • Reviewed and revised data models for soundness of data structures and adherence to client standards.
  • Restructured Logical and physical data models to respond to changing business needs and to assured data integrity using Power Designer.
  • Created naming convention files and co-coordinated with DBAs to apply the data model change.

Environment: Tableau Server, Oracle, Tab Jolt, MySQL, Hadoop, Cloud database, Phyton, MongoDB.

Confidential

Data Analyst

Responsibilities:

  • Consulted with application development business analysts to translate business requirements into data design requirements used for driving innovative data designs that meet business objectives.
  • Involved in information-gathering meetings and JAD sessions to gather business requirements, deliver business requirements document and preliminary logical data model.
  • Performed requirements gathering and design analysis for customer contact history for customer relationship management.
  • Worked as business requirements analysts with subject matter experts to identify and understand requirements.
  • Supported data warehouse projects with logical and physical data modeling in Oracle environment and assured accurate delivery of DDL to the development teams.
  • Documented the data requirements and system changes into detailed functional specifications.
  • Took ownership of Source to Target Mapping and tracked and maintained changes.
  • Performed Extraction, Transformation and Loading (ETL) using Informatica power center.
  • Designed and implemented data profiling and data quality improvement solution to analyze, match, cleanse, and consolidate data before loading into data warehouse.
  • Extensively used Erwin to design Logical, Physical Data Models, and Relational database and to perform forward/reverse engineering.
  • Responsible in designing new FACT or Dimension Tables to existing Models.
  • Implemented stored procedures, functions, views, triggers, packages in PL/SQL.

Environment: Tableau Server, Oracle, Tab Jolt, MySQL, Hadoop, Cloud database, Salesforce, Microsoft PowerPivot,Pivotal HD, HDFS,AWS,Microsoft SQL server 2012, MS Excel/Access.

Confidential

Data Analyst

Responsibilities:

  • Identify business, functional, and technical requirements through meetings and interviews and JAD sessions.
  • Define the ETL mapping specification and Design the ETL process to source the data from sources and load it into DWH tables.
  • Designed the logical and physical schema for data marts and integrated the legacy system data into data marts.
  • Integrate Data stage Metadata to Informatica Metadata and created ETL mappings and workflows.
  • Designed mapping and identified and resolved performance bottlenecks in Source to Target, Mappings.
  • Developed Mappings using Source Qualifier, Expression, Filter, Look up, Update Strategy, Sorter, Joiner, Normalizer and Router transformations.
  • Involved in writing, testing, and implementing triggers, stored procedures and functions at Database level using PL/SQL.
  • Developed Stored Procedures to test ETL Load per batch and provided performance optimized solution to eliminate duplicate records.
  • Provide the team with technical leadership on ETL design and development best practices.

Environment: TDV (Tibco Data Virtualization), DB Visualizer, SQL Server, SSIS, Tableau, Snowflake, Microsoft Azure, Salesforce, TFS, SOAP UI, Putty.

We'd love your feedback!