Data Engineer Resume
5.00/5 (Submit Your Rating)
FL
SUMMARY
- Data Engineer with 3+ years of IT experience in various technologies, tools and databases likeBig Data, AWS, S3, Hadoop, Hive, Spark, Python, etc.
- Knowledge of all phases of Software Development (SDLC) like Requirement Analysis, Implementation, and Maintenance and good experience with Agile and Waterfall.
- Proficiency in Data Science libraries in R and PythonlikeNumPy, Pandas, Matplotlib, SciPy, Scikit - learn, Seaborn, ggplot2 and TensorFlow.
- Understanding of big data processing using Hadoop technologiesMap Reduce, Apache Spark, Hive, HDFS and Pig.
- Well-Versed in SQL Server Business intelligence tools like SQL Server Integration Services (SSIS) and SQL Server Reporting Services (SSRS).
- Capable of AWS EC2, configuring the servers for Auto scaling and Elastic load balancing.
- Proficient in Extraction, Transformation and Loading (ETL) data from various sources into Data Warehouses, as well as data processing like collecting, aggregating and moving data from various sources using PowerBI and Microsoft SSIS.
- Good Technical, Analytical, Problem-Solving skills, strict attention to detail, and ability to work independently, work within a team environment.
TECHNICAL SKILLS
Methodologies: SDLC, Agile, Waterfall
Programming Language: R, Python, SQL
IDE’s: PyCharm, Jupyter Notebook
Big Data Eco system: Hadoop, MapReduce, Hive, Apache Spark, Pig
ETL Tools: SSIS
Cloud Technologies: AWS, Azure
Packages: NumPy, Pandas, Matplotlib, SciPy, Scikit-learn, Seaborn, TensorFlow
Reporting Tools: Tableau, Power BI, SSRS
Database: MongoDB, MySQL
Other Tools: Git, MS Office
Operating System: Windows, Linux
PROFESSIONAL EXPERIENCE
Confidential, FL
Data Engineer
Responsibilities:
- Working with an Agile environment, with an ability to accommodate and test the newly propose changes at any point of time during the release.
- Processing third-party spending data into maneuverable deliverables within a specific format with Excel macros and python and R libraries like NumPy, ggplot2 and Matplotlib.
- Performing data transformations by writing MapReduce and Pig scripts as per business requirements.
- Using AWS for the Tableau server scaling and secure Tableau server on AWS to protect the Tableau environment using Amazon VPC, security group, AWS IAM and AWS Direct Connect.
- Developing SSRS reports, SSIS integration packages, SSAS analysis cubes using Microsoft BIDS and Tableau.
- Writing complex views, functions and store procedures using SQL to be using Report’s development.
- Creating and facilitates presentations and demonstrations for Business Intelligence tools, Business Objects and ETL tools-SSIS.
Confidential
Data Engineer
Responsibilities:
- Involved throughout the Software Development Life Cycle (SDLC) using Waterfall.
- Processed third-party spending data into maneuverable deliverables within a specific format with Excel macros and python libraries like NumPy, SciPy, Pandas and Matplotlib.
- Designed data architecture for one project with cloud computing environment using Amazon Web Services AWS for hosting the databases.
- Wrote MapReduce jobs using Pig, Optimized the existing Hive and Pig Scripts.
- Used Microsoft Power BI Power Query to extract data from external sources and modify data to certain format as required in Excel and created SSIS packages to load excel sheets from PC to database.
- Analyzed new dimensions into spark application upon the business requirements.
- Involved in migration of data from existing RDBMS (MySQL and SQL server) to Hadoop.
- Worked with OLAP cubes to generate drill through reports in SSRS.
