Big Data Cloud Analytics Developer Resume
St Louis, MO
SUMMARY:
- Comprehensive experience of 6 years, with over 4 years in Hadoop and Scala(spark) development and Administration experience along with 2+ of experience in Java/J2EE enterprise application design, development and maintenance.
- Extensive experience implementing Big Data solutions using various distributions of Hadoop and its ecosystem tools.
- Hands - on experience in installing, configuring and monitoring HDFS clusters (on premise & cloud AWS).
- In depth understanding of MapReduce programs to scrub, sort, filter, join and query data
- Planning, deployment, and tuning of SQL (SQL Server, MySQL) and NoSQL (elasticsearch, Redis, memcached) databases.
- Implemented innovative solutions using various Hadoop ecosystem tools like Pig, Hive, Impala, Sqoop, Flume, Kafka, HBase, ElasticSearch, Zookeeper, Couchbase, Storm, Solr, Cassandra and Spark.
- Experience developing PigLatin and HiveQL scripts for Data Analysis and ETL purposes and also extended the default functionality by writing User Defined Functions (UDF s) for data specific processing.
- Experience with migrating data to and from RDBMS and unstructured sources into HDFS using Sqoop & Flume.
- Hands-on experience developing workflows that execute MapReduce, Sqoop, Flume, Hive and Pig scripts using oozie.
- Hands-on experience on R language using Shiny web application framework.
- Well-versed database development knowledge using SQL data types, Joins, Views, Transactions, Large Objects and Performance tuning.
- Good knowledge of Data warehousing concepts and ETL and Teradata.
- Experience writing Shell scripts in Linux OS and integrating them with other solutions.
- Intensive work experience in developing enterprise solutions using Java, J2EE, Servlets, JSP, JDBC, Struts, spring, Hibernate, JavaBeans, JSF, MVC.
- Fluent with the core Java concepts like I/O. Multi-threading, Exceptions, RegEx. Collections, Data-structures and Serialization.
- Excellent problem-solving analytical, communication, presentation and interpersonal skills that help me be a core member of any team.
- Experience mentoring and working with offshore and distributed teams.
TECHNICAL SKILLS:
Big Data Ecosystems: Hadoop, MapReduce, HDFS, HBase, Zookeeper, Hive, Parquet, Pig, Sqoop, Cassandra, Oozie, Flume, Chukwa, Pentaho Kettle and Talend
Programming Languages: Java, Scala
Scripting Languages: JSP & Servlets, PHP, JavaScript, XML, HTML, Python and Bash
Databases: NoSQL, Oracle, SQLtalend
UNIX Tools: Apache, Yum, RPM
Tools: IntelliJ IDEA, Eclipse, R Studio, JDeveloper, JProbe, CVS, Ant, MS Visual Studio, MATLAB
Platforms: Mac OSX, Unix, Linux, Solaris, Windows
Application Servers: Apache Tomcat 5.x 6.0, Jboss 4.0
Automation tools: Chef, Puppet
Testing Tools: NetBeans, Eclipse, WSAD, RAD
Methodologies: Agile, UML, Design Patterns
WORK EXPERIENCE:
Big Data Cloud Analytics Developer
Confidential, St Louis, MO
Responsibilities:
- Partnered with other engineering teams to help architect and build the data pipeline that ingest hundreds of billions of data points for Field Analytics Platform utilizing AWS.
- Expanded capability using various open source data processing technologies like Hadoop, Kafka, Spark.
- Integrated big data infrastructure in the AWS cloud.
- Built services, deployed models, algorithms, performed model training and provided tools to make our infrastructure more accessible to all our data scientists.
- Integrated R models into a Scala Project.
- Helped in generating Log files.
Environment: Scala, Kafka, OpenCPU, ASReml, IntelliJ IDEA, R Studio, Git, AWS
Hadoop Developer
Confidential, Atlanta, GA
Responsibilities:
- Responsible for developing efficient MapReduce on AWS cloud programs for more than 20 years' worth of claim data to detect and separate fraudulent claims.
- Uploaded and processed more than 30 terabytes of data from various structured and unstructured sources into HDFS (AWS cloud) using Sqoop and Flume.
- Played a key-role is setting up a 40 node Hadoop cluster utilizing Apache Spark by working closely with the Hadoop Administration team.
- Worked with the advanced analytics team to design fraud detection algorithms and then developed
- MapReduce programs to efficiently run the algorithm on the huge datasets.
- Developed Scala programs to perform data scrubbing for unstructured data.
- Responsible for designing and managing the Sqoop jobs that uploaded the data from Oracle to HDFS and Hive.
- Worked with the development team to create appropriate cloud solutions for client needs.
- Helped in troubleshooting Scala problems while working with Micro Strategy to produce illustrative reports and dashboards along with ad-hoc analysis.
- Used Flume to collect the logs data with error messages across the cluster.
- Designed and Maintained Oozie workflows to manage the flow of jobs in the cluster.
- Experience in DevOps using Unix, Java, Chef and Puppet
- Played a key role in installation and configuration of the various Hadoop ecosystem tools such as Solr, Kafka, Pig, HBase and Cassandra.
- Tibco Jasper Soft studio was used for the ireport analysis using AWScloud
- Teradata concepts were used for the early instance creation with the DBMS concepts.
- Actively updated the upper management with daily updates on the progress of project that include the classification levels that were achieved on the data.
Environment: Java, Hadoop, Hive, Pig, Sqoop, Flume, HBase, Oracle 10g, Teradata, Cassandra
Hadoop Developer/ Hadoop Admin
Confidential, Phoenix, AZ
Responsibilities:
- Responsible for architecting Hadoop clusters with CDH3
- Involved in the installation of CDH3 and up-gradation from CDH3 to CDH4
- Developer in Big Data team, worked with Hadoop AWS cloud, and its ecosystem.
- Installed and configured Hadoop, Map Reduce, and HDFS.
- Used Hive QL to do analysis on the data and identify different correlations.
- Developed multiple Map Reduce jobs in Java for data cleaning and preprocessing.
- Installed and configured Pig and also written Pig Latin scripts.
- Wrote Map Reduce job using Scala.
- Great understanding of REST architecture style and its application to well performing web sites for global usage.
- Developed and maintained Hive QL, Pig Latin Scripts, Scala and Map Reduce.
- Worked on the RDBMS system using PL/SQL to create packages, procedures, functions, triggers as per the business requirements.
- Involved in ETL, Data Integration and Migration.
- Worked on Talend to run ETL jobs on the data in HDFS.
- Imported data using Sqoop to load data from Oracle to HDFS on a regular basis.
- Developing scripts and batch jobs to schedule various Hadoop Programs.
- Have written Hive Queries for data analysis to meet the business requirements.
- Creating Hive Tables and working on them using Hive QL.
- Importing and exporting data into HDFS from Oracle Database, and vice versa using Sqoop.
- Experienced in defining job flows.
- Experience with NoSQL database HBase.
- Experience in automation and configuration management of Hadoop 2.0 using Chef and Puppet
- Wrote and modified stored procedures to load and modifying of data according to business rule changes.
- Involved in creating Hive Tables, loading the data and writing Hive Queries that will run internally in a map reduce way.
- Developed a custom file system plugin for Hadoop to access files on data platform.
- The custom file system plugin allows Hadoop Map Reduce programs, HBase, Pig, and Hive to access files directly.
- Extracted feeds from social media sites such as Facebook, Twitter using Python scripts.
- Organized and benchmarked Hadoop/HBase Clusters for internal use.
Environment: Hadoop, HDFS, HBase, Pig, Hive, MapReduce, Sqoop, Chef, Puppet, Flume, ETL, REST, Java, Python, PL/SQL, Oracle 11g, Unix/Linux, CDH3, CDH4.
Hadoop Admin/ Data Mining Analyst
Confidential, Portland OR
Responsibilities:
- Define scope and architecture requirements
- Performed Strategic data analysis and research to support business needs.
- Maintained, designed, modified and constructed tools like UNIX and Oracle database
- Specified and analyzed requirements for scalable database and result reporting
- Created world class web data-mining system
- Handled the tasks of code development and data analysis
- Designed and developed transformations required for generating machine learning data sets
- Installation of required apache and related packages
Environment: Hive, PIG, MapReduce, Java, Shell scripting, SQL, Python, HDFS, CDH4.3, Sqoop, Flume, Oozie, Web Crawler, Cent OS, Fedora, Unix.
Core Java/ J2EE Developer
Confidential, Phoenix, AZ
Responsibilities:
- Responsible for the implementation of application system with core Java and Spring framework.
- Used Spring framework for dependence injection and integrated it with Hibernate.
- Developed Restful APIs using Spring Cloud.
- Developed the Spring Xml file for database configuration using different spring beans.
- Implemented and Used Web Services with the help of WSDL and SOAP to get updates from the third parties.
- Involved in implementation of MVC pattern using Angular JS, JSF and Spring Controller.
- Developed specialized search system (using Angular JS, Java Servlets, JUnit).
- Involved in Java and EJB Design Patterns.
- Used Multi-threading to overcome the errors in the process of transactions.
- Used JAXB parser for parsing the valid XML files.
- Used Spring DAO concept in order to interact with database (DB2) using JDBC template.
- Used MVC Framework and integrated Struts Web Module with Java Server Faces (JSF).
- Involved with Master Data Management (MDM) for Customer Data Integration.
- Involved in module testing using JUnit.
- Implemented Hibernate to map all the tables from different data sources to make database updating.
- Maven is used to build and deploy the application.
- Used Spring Framework for Dependency Injection and spring bean wiring
- Used Hibernate 3.0 object relational data mapping framework to persist and retrieve the data from database.
- Implemented Hibernate to map all the tables from different data sources to make database updating.
- Applications are designed using J2EE, JSP, Struts, WSDL, Web Services, JMS
Environment: Spring, JUnit, JDBC, MDM, Eclipse, CSS3, HTML5, Multithreading, JQuery, Oracle 11g, DB2, JSF, Hibernate, Angular JS, RESTFUL APIs, SOAP, LINUX SCRIPTS Shell Scripting, Jboss, SQL Server 2012, GIT, Maven. .
Junior Java Developer
Confidential, Chicago, IL
Responsibilities:
- Developed JavaScript behavior code for user interaction.
- Used HTML, JavaScript, and JSP and developedUI
- Used JDBC and managed connectivity, for inserting/querying& data management including stored procedures and triggers.
- Involved in the design and coding of the data capture templates, presentation and component templates.
- Developed an API to write XML documents from database.
- Used JavaScript and designed user-interface and checking validations.
- Part of a team which is responsible for metadata maintenance and synchronization of data from database.
Environment: java script, JSP, JDBC, HTML, XML.
