Big Data Developer Resume
Union, NJ
SUMMARY:
- Over 6 years of professional experience in IT industry, having 2+ years of experience in Big Data and 4 years in Java/J2EE in various domains like Retail, Banking, and Tele Communication.
- Experience in Big Data processing using Apache Hadoop and its ecosystems like HDFS, MapReduce, Pig, Hive, Flume, Kafka, Oozie, Sqoop, HBase, ZooKeeper, Impala, and Zepplin.
- Experience in working with NoSQL databases like HBase 1.3.0 and Cassandra 3.10
- Proficient in implementing and optimizing Spark program to perform data collection, data cleaning, and data integration.
- Experienced in manipulating the streaming data to clusters through Kafka and Spark - Streaming.
- Proficient in Spark RDD, DataFrame API with Scala to support ETL procedure.
- Experienced in developing Spark code using Scala and Spark-SQL/Streaming for faster processing of data.
- Experienced in using, configuring and optimizing Kafka, Flume to do efficient data ingestion from the different data source.
- Experienced in transferring the bulk data between RDBMS and HDFS and vice-versa using Sqoop.
- Experienced in all phase of Data Warehouse life Cycle involving requirement analysis, design, coding, testing, and deployment.
- Good in writing HiveQL queries and Pig Latin script to do batch processing and storage.
- Expertise in writing Hive queries for data analysis to meet the requirements, created Hive tables to store data in HDFS and process data using Hive QL.
- Experienced in RDBMS including Oracle, MySQL.
- Expertize in utilizing Sqoop to import and export bulk data to HDFS and Hive Metastore from MySQL.
- Well experienced in implementing MapReduce jobs using Java/Scala to process and perform various analytics on large datasets
- Knowledge in writing custom UDF’s for extending Hive and Pig core functionality
- Proficiency in JSON, XML, CSV, Avro, Apache Parquet and other formats of data.
- Experience with Cloudera and Hortonworks as distributions of Hadoop.
- Experience in Web application development using J2EE 7, HTML 5, CSS 3, Bootstrap, JavaScript, JSON, JQuery, AJAX and Spring MVC 4.x
- Very good understanding and Working Knowledge of Object Oriented Programming (OOPS), Multithreading in Core Java, J2EE 7, JDBC, JavaScript, and JQuery.
- Good knowledge of ETL Scripts for Data Acquisition and Transformation
- Experience in Unit Testing with JUnit and Scala Test.
- Experienced to make the data available for BI team and generated reports based on data using Tableau.
- Involved in Agile Scrum and Test-Driven Development methodology that leverages the Client big data platform and used version control tool Git.
- Strong problem-solving skills, a Self-motivated learner with passion about new technologies, good communication skills and an excellent team player.
TECHNICAL SKILLS:
Big Data Ecosystem \ Scripting Language: Hadoop 2.7.x, MapReduce, Spark 2.1.1, \ HTML5, CSS3, XML, Scala 2.12.0, UNIX \ Zookeeper 3.4.6, Hive 2.1.1, Sqoop 1.99.7, \ Shell, Git Shell 2.12.0, JSP 3.1, Java\ Oozie 4.2.0, Flume 1.7.0, Yarn 0.21.3, Pig 0.14, \ Kafka 2.1.0, Cloudera 5.8.0, Hortonworks 2.5\
Web Development \ Development Tools/IDE: J2EE 7(JavaScript, jQuery, Spring MVC 4.0), \ Eclipse 4.6, Visual Studio code, Documentum\ HTML5, CSS3, SOAP, REST \ Composer, Sublime, Zeppelin.\
Operating Systems \ Databases: Mac OS, Ubuntu, CentOS, Windows XP/\ SQL Server 2008, Oracle 11g, Mysql 5.0, \ Vista/7/8/10\ Cassandra 3.10 HBase 1.3.0. PL/SQL 11g.\
Data Analysis & Visualization\ Collaboration & Environment: Python, Tableau\ Git 2.12.0, Scala Test 3.0.1 & Agile, Waterfall\
PROFESSIONAL EXPERIENCE:
Confidential - UNION, NJ
Big Data Developer
Roles & Responsibilities:
- Loaded and transformed large sets of structured, semi-structured and unstructured data from relational databases into HDFS using Sqoop imports.
- Processed the real-time data using Spark Streaming for faster processing of data
- Imported bulk data from various data sources into Hadoop and transform data in flexible ways by using Kafka.
- Developed Sqoop scripts to import-export data from relational sources and handled incremental loading on the customer transaction data by date.
- Developed Spark code using Scala and Scala Test for testing.
- Experienced with batch processing of data sources using Apache Spark.
- Implemented Spark RDD transformations, actions to implement business analysis.
- Worked on partitioning Hive tables and running scripts in parallel to reduce run-time of the scripts
- Worked on Data serialization formats for converting complex objects into sequence bits by using AVRO, JSON formats.
- Implemented business logic by writing Hive UDFs in Scala.
- Aggregated and stored the data result into HDFS and Cassandra.
- Used Oozie operational services for batch processing and scheduling workflows dynamically.
- Processed the data which extracted from Cassandra utilizes scala programming in Spark Framework.
- Wrote XML scripts to build Oozie functionality.
- Implemented all components following test-driven development(TDD) methodology and used Scala Test 3.0.1 for unit testing
- Used to monitor and manage the Hadoop cluster using Ambari.
Environment: Spark 2.1.1, Scala 2.11, Hive 2.2.0, Sqoop 1.4.6, Kafka 2.1, Hadoop 2.0, HDFS, Oozie, NoSQL, ScalaTest 2.2.0, Putty, Cassandra 3.10 and Hortonworks 2.5
Confidential - Jersey City NJ
Big data developer
Roles & Responsibilities:
- Collaborated with Internal/Client BA's to understand the requirement and to architect a data flow system.
- Imported and exported data from different RDBMS into HDFS using Sqoop.
- Responsible for utilizing data pipeline using Kafka, Hive, and Sqoop to ingest, transform and analyzing customer behavioral data and data of transactional environments.
- Loading a large amount of application server logs from different web servers using Kafka and performed various analytics using Hive queries.
- Involved in running Hive Jobs for processing millions of records and compression techniques.
- Wrote HiveQL queries and Pig Latin script to do batch processing and store
- Experience in querying data using Data Warehouse like Hive.
- Scheduled and executed workflows in Oozie to run Hive jobs.
- Created hive schemas using performance techniques like partitioning and bucketing.
- Developed multiple MapReduce jobs in java for data cleansing and pre-processing.
- Used HBase as storage for a large volume of data, integrate with Apache Phoenix.
- Monitored Hadoop cluster job performance and performed capacity planning and managed nodes on Hadoop cluster.
- Exported the analyzed data from HDFS to the relational database using Sqoop to further visualize and generate reports using Tableau 9.0.
Environment: Hadoop 2.0, HDFS, Hive 2.0.0, Sqoop 1.99.6, Flume 1.7.0, MapReduce, Kafka 2.0, Spark, Scala 2.11, Oozie 3.3.0, MySQL, Java 1.7/1.8, Pig 0.14, HBase 0.94.6, Tableau 9.0, Cloudera 5.8.0
Confidential
Java Developer
Roles & Responsibilities:
- Created the Mock-ups using JSP, JavaScript to understand the flow of the web application and created class diagrams using MS Visio 2005
- Involved in the process of analysis, design, and development of the application.
- Developed the user interfaces using JSP and Servlets for different User Interfaces using RSA tool.
- Created dynamic HTML pages, used JavaScript to create interactive front-end GUI.
- Used Spring IoC and created the Dependency Injection for the Action classes by ApplicationContext.xml.
- Configured the deployment descriptors in Hibernate to achieve object-relational mapping.
- Developed Hibernate persistence layer modules.
- Involved in writing procedures, SQL queries to process the data.
- Performed regression testing, unit testing using JUnit.
- Used Apache Maven 2.2.1 as a build tool.
- Used CVS as version control tool for maintaining source code and project documents.
- Perform deployment of Application on IBM Web Sphere Application Server.
Environment: Java, J2EE, HTML, Servlets, JSP, MS SQL Server 2005, JUnit, Web Sphere Application Server
Confidential
Jr. Java Developer
Roles & Responsibilities:
- Developed application using Java Spring Framework and used Eclipse Integrated Development Environment
- Developed the UI Layer using Struts, CSS, JSP, JavaScript, JSTL, XML
- Developed the presentation layer using MVC.
- Developed various SOAP-based Web services using Apache Axis2 implementation.
- Developed various service codes to provision the lines and configured them with Rest Web services.
- Used Spring framework for wiring and managing business objects.
- Developed PL/SQL programming on Oracle database using Oracle SQL Developer and Java JDBC technologies.
- Involved in creating various Data Access Objects for Addition, modification, and deletion of records using various specification files.
- Managed Service dependencies using Spring Dependency Injection.
- Written Junit test cases for unit testing and load testing for various service codes.
- Wrote Maven scripts for building project modules.
- Monitored the error logs using Log4j.
Environment: Java/J2EE, Spring 3.0, SOAP web services, Restful web services, JMS, Oracle 11g, WebLogic, Angular JS, Junit 3.0, Maven and Log4j
Confidential
Jr. Software developer
Roles & Responsibilities:
- Writing the document for the application.
- Involved in developing integration and system test cases based on the business requirements.
- Analyzed and fixed the defects for various modules in all the QA stages
- Implemented functionality based on the business requirements for all the major releases
- Participated in the design review of the application to come up with UI and provide best possible recommendations for the application from UI standpoint.
- Experienced in web design and development for UI interface design, graphic design, illustration, and logo design.
- Used HTML, DHTML, CSS, Flash, Photoshop, JavaScript.
- Worked on Java Servlet Pages (JSP).
- Developed screen functionality using HTML, CSS, JavaScript.
- Developed UI using HTML, CSS, and JavaScript validations.
- Implemented interaction between Frontend and backend using JSON object.
- Wrote Cross Browser code of CSS and JavaScript for Internet Explorer and Firefox
- Debugging and troubleshooting the errors.
Environment: Eclipse IDE, HTML, DHTML, CSS, Flash, Photoshop, JavaScript, JSP
