Python Developer / Data Scientist Resume
Pennington, NJ
SUMMARY:
- 7+ years of experience in Data mining,Data Acquisition, Data Validation, Data Science, Predictive modeling,Machine Learning, Forecasting fromproblem definition to deployment. I valuetechnical competency and enjoy deep diving into problems.
- Python (5 years), R (4 years), MATLAB (2 years), Tableau (3 Years), ML (2 years), AutoCAD (2 Years),GitHub (2 Years), Excel (5 Years), SQL (4 years), C++ (2years), Django (3 years)
- Proficient in managing entire data science project life cycle and actively involved in all the phases of project life cycle including data acquisition, data cleaning, data engineering, features scaling, features engineering, statistical modeling(decision trees, regression models, neural networks, SVM, clustering),dimensionality reduction usingPrincipal Component Analysis and Factor Analysis,testing and validation usingROC plot, K - fold cross validationanddata visualization.
- Adept and deep understanding of Statistical modeling, Multivariate Analysis, model testing, problem analysis, model comparison and validation.
- Expertise in transforming business requirements into analytical models, designing algorithms, building models, developing data mining and reporting solutions that scales across massive volume of structured and unstructured data.
- Skilled in performing data parsing, data manipulation and data preparation with methods including describe data contents, compute descriptive statistics ofdata, regex, split and combine, Remap, merge, subset, reindex, meltandreshape.
- Experience in using various packages inRandpythonlikeggplot2, caret, dplyr, Rweka, gmodels, RCurl, tm, C50, twitteR, NLP, Reshape2, rjson, plyr, pandas, numpy, seaborn, scipy, matplotlib,scikit-learn, Beautiful Soup, Rpy2.
- Extensive experience in Text Analytics, generating data visualizations usingR,Pythonand creating dashboards using tools likeTableau.
- Strong SQL Server programming skills, with experience in working with functions, packages andtriggers.
- Hands on experience in implementingLDA, Naive Bayesand skilled inRandom Forests, Decision Trees, Linear and Logistic Regression, SVM, Clustering, neural networks, Principle Component Analysis.
- Good Knowledge inProof of Concepts(PoC s), gap analysis and gathered necessary data for analysis from different sources, prepared data for data exploration using data munging.
- Good industry noledge, analytical &problem solving skills and ability to work well with in a team as well as an individual.
- Highly creative, innovative, committed, intellectually curious, business savvy with good communication and interpersonal skills.
- Extensive experience inData Visualizationincluding producing tables, graphs, listings using various procedures and tools such asTableau.
PROFESSIONAL EXPERIENCE:
Confidential, Pennington, NJ
Python Developer / Data Scientist
Responsibilities:
- Usedpythonlibraries like Beautiful Soap, NumPy and SQLAlchemy.
- Created various types of data visualizations usingPythonandTableau.
- Monitoring and tracking process performance using analytics tools likeTableaudashboard, Rand TSQL.
- Utilized standardPythonmodules such as csv, robot parser, itertools and pickle for development.
- Created views inTableauDesktop that were published to internal team for review and further data analysis and customization using filters and actions.
- Worked onPythonOpenStack APIs and used NumPy for Numerical analysis.
- UsedPythonscripts to update content in thedatabaseand manipulate files.
- GeneratedPythonDjangoForms to record data of online users.
- UsedPythonandDjangocreating graphics, XML processing, data exchange and business logic implementation.
- Performed troubleshooting, fixed and deployed manyPythonbug fixes of the applications and involved in fine tuning of existing processes followed advance patterns and methodologies.
- Skilled in using collections inPythonfor manipulating and looping through different user defined objects.
- Installed numerouspythonpackages using pip and easy install.
- Created independent libraries inPythonwhich can be used by multiple projects which have common functionalities.
- Developed test plan, test scripts and test procedures from the specification document inPythonand automating them to run in the real time HIL environment.
- Used machine learning algorithms such as, decision trees, regression models, SVM, clustering to make prediction based on the given financial data
- Utilized RandpythonforExploratory Data Analysis, A/B testing, ANOVA testand Hypothesis testto compare and identify the correlation of other financial factors
Environment: Python2.7/3.4,Django1.4,Tableau, R - console, T - SQL, MySQL, HTML5, CSS, XML, Linux, Shell Scripting
Confidential, Mahwah, NJ
Python Developer / Data Scientist
Responsibilities:
- Performed Data Profiling to learn about behavior with various features such as traffic pattern, location, time, Date and Time etc.
- Application of various machine learning algorithms and statistical modeling likedecision trees, regression models, neural networks, SVM, clusteringto identify Volume usingscikit-learnpackage inpython
- Used clustering techniqueK-Meansto identify outliers and to classify unlabeled data.
- Evaluated models usingCross Validation, Log loss function, ROC curvesand usedAUCfor feature selection.
- Analyze traffic patterns by calculating autocorrelation with different time lags.
- Ensured that the model haslow False Positive Rate.
- Addressed overfitting by implementing of the algorithm regularization methods likeL2andL1.
- Used Principal Component Analysis in feature engineering to analyze high dimensional data.
- Created and designed reports that will use gathered metrics to infer and draw logical conclusions of past and future behavior.
- PerformedMultinomial Logistic Regression, Random forest, Decision Tree, SVMto classify package is going to deliver on time for the new route.
- UsedMLlib, Spark’sMachine learning library to build and evaluate different models.
- Implemented rule based expertise system from the results of exploratory analysis and information gathered from the people from different departments.
- Performed Data Cleaning, features scaling, features engineering usingpandasandnumpypackages inpython.
- CreatedTableaustories, dashboards using stack bars, bar graphs, scattered plots, geographical maps, Gantt charts using show me functionality.
- Created various types of data visualizations usingPythonandTableau.
- UsedRandpythonforExploratory Data Analysis, A/B testing, ANOVA testand Hypothesis testto compare and identify the TEMPeffectiveness of Creative Campaigns.
- Communicated the results with operations team for taking best decisions.
- Collected data needs and requirements by Interacting with the other departments.
Environment: Python 2.x, Linux, R, TableauDesktop, SQLServer 2012, Microsoft Excel,Matlab
Confidential
Python Developer
Responsibilities:
- Developed entire frontend and backend modules usingPythononDjangoWeb Framework.
- Implemented the presentation layer withHTML,CSSandJavaScript.
- Involved in writing stored procedures usingMySQL.
- Optimized thedatabasequeries to improve the performance.
- Designed and developed data management system usingMySQL.
Environment: MySQL 5.x,HTML5, CSS3, JavaScript, Shell, Linux & Windows,Django1.x,Python2.x, Matlab, R Console
Confidential
Embedded System Engineer/ Programmer
Responsibilities:
- Designed the specific system structure on hardware layers and fulfilled the technology selection of main modules in the alpha version.
- Designed the communication protocols.
- Programmed the management interface. (C++ on QT4.6)
- Tested and fixed fixed hardware and software bugs and improved the algorithm and data structure for exception handling.
- Used Excel and Matlab to analyze by using finite element method
Environment: C, C++, QT 4.6, embedded programming (Keil uVision 3), MySQL, Linux, Windows, Matlab, AutoCAD, Excel
