Data Analytics Intern Resume
5.00/5 (Submit Your Rating)
PROFESSIONAL PROFILE:
- 3 years of professional experience in big data analytics with strong interest in Data Science & Machine learning
- Implemented machine learning pipeline from exploratory data analysis, feature engineering, model building, performance evaluation, and testing with large data set.
- Data extracting/preprocessing, features selection, optimization of classification algorithms by tuning parameters
- Built a near real time discovery platform to analyze sentiment from tweets using big data analytics framework.
- It can continuously measure and compare sentiments of given input terms.
- Developed algorithm to analyze the structure of a drug trafficking organization by calculating different Centrality Measures using Visualization of plotly/R shiny
- RFM (regency, frequency, monetary) analysis to find out best customer group for targeted marketing
- Built storylines and dashboards with the KPI’s to identify strong customer relations using Tableau
- Market segmentation (hierarchical clustering) based upon customer data (needs/demographic)
- Built sample big data warehouse architecture using 80% Hadoop system.
- Detailed comparison over of data warehouse system expansion using traditional methods Vs data warehouse system expansion using Hadoop (hosted on cloud using Microsoft Azure and AWS)
TECHNICAL SKILLS:
Programming Languages: Python, R, Java, SQL/PLSQL, SAS, STATA, Matlab, Linux/Unix scripting, C, C++,HTML,CSS, JSON, JavaScript
Big data: Spark,PySpark,Hadoop, NoSQL, MapReduce, Sqoop, Hive, Flume, Solr, Nifi
Library: Scikit Learn, Pandas, Numpy, Keras, Tensorflow, MLlib, Pandas, NLTK, flask
Cloud technologies: Amazon web services (AWS)
Database: Oracle 11g/12g, MySQL, Postgres, Mongo DB
Data Analytics Tools: Tableau, SAS, Google Analytics, IBM SPSS, SSIS
WORK EXPERIENCE:
Data Analytics Intern
Confidential
Responsibilities:
- Developed a web based analytics/visualization tool using R shiny/ggplot2/plotly for architectural acoustic data
- Modified existing application to add statistical analytics/visualization of business data using python flask
Project Engineer
Confidential
Responsibilities:
- Built automated data pipelines to import Asset/Inventory data by developing Java/SQL/PLSQL,ETL jobs to collect, centralized large - scale data from sources like third party vendors and database application to facilitate data integration for IBM Maximo asset management system(ERP) for clients like SFR(France) and Cisco(USA)
- Built automated reporting solutions using BIRT and IBM Cognos Reporting tools
