Sr Python Data Developer Resume
Plano, TX
SUMMARY
- Experienced Data Analyst, with 9+ years of expertise and hands - on multiple languages and experience in Design, Development, Implementation of data flows, datasets and databases using SQL, Python, Dash, Plotly, Tableau, PowerBI, Cloud Services using Azure data services.
- Experienced in design and extraction of encounter data from multiple source systems into the data warehouse relational database (Oracle) while ensuring data integrity.
- Experienced in Design, development, and implementation of performant ETL pipelines using python API (pyspark) of Apache Spark on Azure and AWS.
- Experienced with event-driven and scheduledAWSLambda functions to trigger variousAWS resources.
- Experience using Azure app services for code deployment and Azure Devops for version control.
- IT experience in Data Warehousing, Data Analysis, ETL, BI and Business Analytics. Experience in Building AWS code pipelines using code commit and code build, and code deploy.
- Hands on Data modeling for Data Warehouse/Data Mart development, Data Analysis for Online Transaction Processing (OLTP) and Data Warehousing (OLAP)/Business Intelligence (BI) applications created.Experience in writing complexSQLQueries, created materialized views.
- Excellent understanding of the software development lifecycle (SDLC).
- Highly skilled in deployment, data security and troubleshooting of the applications usingAWS services.
- Experience in setting up CI/CD pipelines using AWS DevOps tools
- Experience in all phases of Software Development Life Cycle (SDLC)-Waterfall, agile Process across various workflows (Requirement study, Analysis, Design, Coding, Testing, Deployment and Maintenance) in Web & Client/Server application development.
- Experience with AWS infrastructure and concepts like EC2 instance, S3 storage (cross region replications), RDS, IPM, route 53, AWS lambda functions
- Experience in working with several python libraries including beautiful soup, NumPy, matplotlib, SciPy, PyQt, Scapy, SQLAlchemy.
- Good experience in developing web applications implementing MVT architecture using Django, Flask, Webapp2 web application frameworks, with good understanding of Django ORM and SQLAlchemy.
- Hands-on experience with industry-standard IDEs like PyCharm, Sublime, NetBeans.
- Good experience in Shell Scripting, Oracle RDBMS, SQL Server, UNIX, and Linux.
- Good knowledge of TCP/IP, UDP, HTTP, HTTPS, SSH and SSL protocols.
- Expert at version control systems like Git, SVN.
- Proficient in writing SQL Queries, Stored procedures, functions, packages, tables, views, triggers using relational databases like Oracle, MSSQL Server.
- Experience with Arc Objects SDK and ArcGIS JS API
- Good experience in working with Amazon Web Services (AWS) like AWS EC2, S3, VPC, SES, ELB, EBS, RDS, Glacier, DynamoDB…etc.
- Experience in using XML, SOAP, REST web Service for inter operable software applications.
- Hands on experience with bug tracking tools JIRA and Bugzilla.
- Experience in Agile development processes ensuring rapid and high-quality software delivery
- Highly motivated, quality minded developer, with proven ability to deliver applications against tight deadlines.
TECHNICAL SKILLS
Software Development: Python 3.6/2.7, R, Pyspark, Hadoop/Spark
Web Technologies: HTML, CSS
Project management: SCRUM
Python Libraries: NumPy, Pandas, Matplotlib, Seaborn, SciPy, NLTK, Dash, Plotly
Database Management: SQL server, MS-Access, PostgreSQL, Oracle
Reporting Tools: Tableau, PowerBI.
Version Control: GIT (GitHub), SVN.
Cloud Technologies: AWS, AZURE
PROFESSIONAL EXPERIENCE
Confidential, Plano, TX
Sr Python Data Developer
Responsibilities:
- Built data dashoboards for machine learning and data scientists teams using Python Dash, Plotly and Flask framework.
- Developed resubale OOPS code in Python for data analysis and UI dashboards use cases to support business and research needs.
- Developed Python Dash code in Local, tested and shared UI for daily reporting and analysis with internal stakeholders.
- Deployed Python Dash/Plotly UI to Azure Fucntions using custom built Docker images and Files.
- Analyzed huge volume of Pipelines data using Jupyter Notebooks using Pandas, Numpy and devloped weekly and monthly business review reports using Matplotlib and Seaborn Python data visualization packages and frameworks.
- Created insightful visualizations and custom reports for multiple stakeholder teams, which helped them to make critical business impact decisions.
- Developed Tableau reports fetching data from files and Databases using Sql queris.
- Created ETL pipelines in Azure for file based data loads and Tranformations using Python and Azure SDK.
- Created custom UI using HTML, CSS and basic Javascript functionalites with Python Dash Framework to support business and application needs.
- Participated in daily Scrum calls and weekly code reviews.
- Made use of Azure Devops to Push and make changes to code for version control.
Environment: Python, SQL, Azure, Tableau, NumPy, Pandas, PySpark, Dash, Ploytly, Flask, HTML, CSS.
Confidential, Chicago, IL
Senior Data Analyst
Responsibilities:
- Worked with PySpark and Spark SQL in building the performance and optimized spark frame work for applications running on AWS EMR cluster data stored on Hadoop, HDFS, HIVE tables.
- Understand existing system business logic, perform enhancement and impact analysis of the applications.
- Used python Libraries like PySpark, Pytest, cxOracle and PyMongo based on the modules and business requirement.
- Developed Spark applications using Spark tools like RDD transformations, Spark core, Spark streaming and SparkSQL.
- Used GO lang scripts for uploading a file to S3 and deploying them and creating GO serverless application and deploying it to AWS lambda.
- Developed end to end Spark applications using PySpark to perform various data cleansing, transformation and summarization activities according to the business requirements.
- Python Script reviewing for the data collection, analytical reports development scripts, altercation. Created data frames for data type testing, did normalization analysis on data supporting digital data science team.
- Wrote pre-processing queries in PySpark for internal spark jobs and validated using HQL
- Loaded data from many different sources to local analytical area & data lakes with the environment creating CI/CD pipelines.
- Create reports for the BI team using Imply-Druid exporting data from HDFS and Hive
- Enhanced current Automation Regressing Scripts for validation of ETL process between multiple databases like Oracle, SQL Server, Spark, Mongo DB using Python.
Environment: Python, SQL, Databricks, Tableau, NumPy, Pandas, AWS, PySpark
Confidential, Newport Beach, CA
Senior Python Data Analyst
Responsibilities:
- Responsibilities include analyzing, trouble shooting, resolving and documenting reports. Understand existing system business logic, perform enhancement and impact analysis of the applications.
- Document all the requirements received from business team, optimized for effective solutions.
- Coordinate with business analyst, business client and various functional teams across the application to gather requirements and provide best solution approach.
- Developed ETL programs in python to move data from source systems to analytics area.
- Used python to retrieve and manipulate data from AWS Redshift, Oracle DB, T-SQL, MS SQL Server, Excel and Flat files.
- Demonstrated to move data between production systems and across multiple platforms.
- Once the data was dumped in to the analytics area, identified the business requirements to transform the data for analytics purposes based on the business requirements.
- Developed Automation Regressing Scripts for validation of ETL process between multiple databases like AWS Redshift, Oracle, MongoDB, T-SQL, SQL Server using Python.
- Used python Libraries like Pytest, Pymongo, cxOracle, PyExcel, PyExcel, Boto3, Psycopg, SOAP, embedPy NumPy and BeautifulSoup based on the modules and business requirement.
- Configured AWS environment to extract data from various sources and loaded the data in Redshift using distribution and sorting.
- Create complex data models and process flow diagrams. Create and maintain high level design (HLD), detailed design (DD), Unit test (UTP) documentations & business-process documentation based on the requirements gathered from business analysts & business user, using industry-standard methodology.
- Development/enhancement of Oracle PL/SQL programs for creating Tables, Views, Sequences, Database triggers, Cursors, Stored Procedures and Functions, Exception Handling and Indexing, Optimization and Tuning of Procedures, SQL Queries to improve performance.
- Developed python programs and excel functions using VB Script to move data and transform data.
- Development/enhancement of UNIX shell scripts. Troubleshoot the Production issues.
- Developed data analysis tools using SQL and Python code.
- Work closely with upper management and consultants onshore and offshore of various teams in development, maintenance, QA & Testing and production support of compensation system.
- Design, Development and Enhancements of various types of reports.
- Working on reporting converting and making reports using python (PyExcel Module).
- Expertise in data quality, data organization, Meta data and data profiling.
Environment: Python, AWS, Oracle 11g/12C,MS SQL Server, Teradata, Vertica, PL/SQL, Linux, Microsoft Access, Power BI, VBA, PyCharm, Sql Developer, DBeaver, Jupyter Notebook.
Confidential
Senior Data and Reporting Engineer
Responsibilities:
- Involved in the analysis, design, and development and testing phases of Software Development Life Cycle (SDLC).Managed and resolved design issues during deployment.
- Maintains the Information Systems Development Methodology (ISDM) document. Coordinates with system project administrators to promote integration, consistency, and quality. Reviews applications and deliverables to ensure compliance with the ISDM.
- Gathered the requirements from the data analysts and architects and performed technical analysis to reach a viable solution.Involved in tuning and optimization ofSQLstatements.
- Converted, loaded and integrated data fromMicrosoft Access, Excel, text.csv files and internal database applications including the mainframe with data captured for security applications.
- Created and managed primary data base objects such as Tables, Views, Indexes Sequences and Synonyms.Involved in writing test scripts, unit testing, system testing and documentation.
- Built complex queries usingSQLand wrote stored procedures &packages usingPL/SQL.
- Using OracleSQLDeveloperData Modeler create and edit logical, relational, physical, multi-dimensional and data type models. GeneratedDDLfrom data models using ODI.
- Worked withMicrosoftAccess, andSQLServer … to write functions, procedures and views to meet the user's requirements.
- Functionally discussed the requirement with Business and worked on all Major Oracle Apps ERP functional modules like CRM, SCM and Finance.
- Advanced execution in MS EXCEL, EXCEL macros, MS access,VBAscripting, MySQL, SSIS, SSRS, SSAS, OLAP, Reporting and Analytics, Oracle 9i, Oracle 8i,PL/SQL(Stored Procedures), T-SQL.
- WrotePL/SQLStored Procedures, Functions, and Packages to implement Business Rules.
- Designed and developedPL/SQLfunctions/ stored procedures/ cursors/ triggers/ packages.
- Involved in day-to-day duties of DBA work like taking backup and recovery of database environment.Worked with production support team in resolving production issues.
- Experience in AccessVBAto extract data from a variety of operational data sources on multiple platforms and build a data warehouse and data marts that integrate the extracted data.
- Designed database structures for effective data extraction, validation, run detail notification and error logging.Supported code deployment from lower to higher environment.
- DevelopedPL/SQLETL scripts to transfer data across the schemas and databases.
- Involved in development and scheduled jobs using Autosys for automation.
