We provide IT Staff Augmentation Services!

Big-data And Hadoop Developer Resume

3.00/5 (Submit Your Rating)

MD

SUMMARY

  • Over 7+ years of experience in the field of IT including four years of experience in Hadoop environment and good object oriented programming skills.
  • Good noledge of Hadoop Architecture and various components such as HDFS, Job Tracker, Task Tracker, Name Node, Data Node, MapReduce concepts and setting up standards and processes for Hadoop based application design and implementation.
  • Expertise with the tools in Hadoop Ecosystem including Pig, Hive, HDFS, MapReduce, Sqoop, Spark, Kafka, Yarn, Oozie, and Zookeeper.
  • Experience in installation, configuration and deployment of Big Data solutions.
  • Hands on experience using Kafka, Spark, Cassandra.
  • Expertise with the tools in Hadoop Ecosystem including Pig, Hive, HDFS, MapReduce, Sqoop, Spark, Storm, Scala, Impala, Kafka, Yarn, Oozie, and Zookeeper.
  • Expertise in MapReduce programs in HIVE and PIG to validate and cleanse the data in HDFS, obtained from heterogeneous data sources, to make it suitable for analysis.
  • Analyzed or transformed stored data by writing MapReduce jobs based on business requirements.
  • Experience in developing Pig scripts and Hive Query Language.
  • Experience with Hortonworks Hadoop distribution components and custom packages.
  • Hands on experience in Cloudera.
  • Expertise in Talend.
  • Hands on experience in Python, Scala and Shell.
  • Written Hive queries for data analysis and to process the data for visualization.
  • Hands on experience in Web Platform Development.
  • Managing and scheduling batch Jobs on a Hadoop Cluster using Oozie and managing Metadata with Big Data.
  • Experience in using Apache Storm for real time processing.
  • Hands on experience with Teradata.
  • Hands on experience in JSON, Avro.
  • Exported the analyzed data to the relational databases using Sqoop for visualization and to generate reports for the BI team.
  • Hands on experience working with NoSQL database including Monod and HBase.
  • Experience in optimizing MapReduce jobs to use HDFS efficiently by using various compression mechanisms.
  • Experience in managing and reviewing Hadoop Log files.
  • Expertize in data centric application development.
  • Experience in developing the Pig UDFs to pre - process the data for analysis.
  • Participated in multiple big data POCs to evaluate different architectures, tools and vendor products.
  • Used Zookeeper to provide coordination services to the cluster.
  • Experience in developing Pig Latin scripts to extract the data from the web server output files to load into HDFS.
  • Hands on experience in using Sqoop to import data into HDFS from RDBMS and vice-versa.
  • Hands on experience in application development using Big Data, Java, RDBMS and Linux/bash Shell Scripting and using application servers like Tomcat, WebLogic and Glassfish.
  • Hands on experience in using version control system Git.
  • Expertise in Web technologies using Core Java, J2EE, Servlets, EJB, JSP, JDBC, Java Beans, and Design Patterns.
  • Expertise in MVC Technologies Struts MVC, Spring MVC, Hibernate and JSF.
  • Experience in developing J2EE applications using various other Open Source tools, Persistence frame work (Hibernate) and implementing JPA (Java Persistence API).
  • Hands on experience with SOAP, REST Web Services.
  • Detailed understanding of Software Development Life Cycle (SDLC) and sound noledge of project implementation methodologies including Waterfall and Agile.
  • Experience in Test Driven Development and Test Automation.

TECHNICAL SKILLS

Operating Systems: Windows 8/7/XP, Unix, Ubuntu 13.X, Mac OSX

Hadoop Eco System: Hadoop 1.x/2.x(Yarn), Hortonworks, HDFS, Map Reduce, Mongo, HBase, Hive, Pig, Zookeeper, Sqoop, Oozie, Spark, Storm, Scala, Impala, Flume, Avro, Talend, Eclipse, Cloudera-desktop and SVN

Java Tools: Java, MapReduce, J2EE (JSP, Servlets, EJB, JDBC, JMS, JNDI, RMI), Struts, Springs, Hibernate, AJAX, XSLT, HTML, JavaScript, CSS, Junit, Java, JSP, JSON, AVRO, J2EE, Web Services, DHTML, Javascript, DOM, SAX, JQuery, XML,XSLT

API’s: Servlets, EJB, Java Naming and Directory Interface(JNDI), MapReduce

Development Tools: Eclipse, RAD/RSA (Rational Software Architect), IBM DB2 Command Editor, SQL Developer, Microsoft Suite (Word, Excel, PowerPoint, Access), Open Office Suite (Editor, Calc etc..),VM Ware

Databases: MySQL, SQL, IBM DB2 9.x, Oracle 11g/10g

No SQL Databases: HBase, Cassandra, Monod

Servers: Web sphere (WAS) 6.x/7.0, Web Logic 10-12c, Tomcat, Glassfish

Version Control: Git Bash, Bitbucket, Trello

Programming Languages: C, C++, Java, Python, Scala

PROFESSIONAL EXPERIENCE

Confidential, MD

Big-Data and Hadoop Developer

Responsibilities:

  • Worked on analyzing Hadoop cluster using different big data analytic tools including Hive, MapReduce, Pig and Kafka.
  • Involved in loading data from LINUX file system to HDFS.
  • Implemented Partitioning, Dynamic Partitions, Buckets in Hive.
  • Provided L3 engineering support, diagnosed and provided solutions for support requests for Hadoop in the project.
  • Involved in scheduling Oozie workflow engine to run multiple Hive and pig jobs.
  • Used Talend for Data Integration.
  • Optimized BigData performance in the cloud using Talend.
  • Used Hortonworks Hadoop distribution components and custom packages.
  • Written MapReduce code in Python.
  • Used Teradata to implement Hadoop in the project.
  • Exported the result set from Hive to MySQL using Shell scripts.
  • Monitored job executions regularly.
  • Involved in developing Pig UDFs for the needed functionality that is not out of the box available from Apache Pig.
  • Used Big Data for aggregation and extraction with ETL for Big Data Analytics.
  • Involved in processing ingested raw data using MapReduce, Apache Pig and Hive.
  • Importing and Exporting of data from RDBMS to HDFS and vice versa using Sqoop.
  • Analyzed the data using Pig and written Pig scripts by grouping, joining and sorting the data.
  • Used Spark for Bid Data Processing and Apache Storm for real time processing.
  • Used Impala for parallel processing the data in SQL in the cluster.
  • Load and transform large sets of structured, semi structured and unstructured data.
  • Developed Pig Latin Scripts to extract data from the web server output files to load into HDFS.
  • Made Data Analysis on the data which comes from Logs, graphs, ERPs.
  • Configured Spark streaming to receive real time data from the Kafka and store the stream data to HDFS using Scale.
  • Worked on debugging, performance tuning of Hive & Pig Jobs.
  • Involved in developing Hive DDLs to create, alter and drop Hive tables.
  • Actively involved in code review and bug fixing for improving the performance.
  • Developed screens using JSP, DHTML, CSS, AJAX, JavaScript, Java and XML.
  • Supported MapReduce programs those are running on the cluster.
  • Extensively used Pig for data cleansing.
  • Used Kafka to publish messages.
  • Used NoSQL database with Cassandra and Monod.
  • Computed various metrics using Java, MapReduce to calculate metrics that define user experience, revenue etc.
  • Involved in using Sqoop for importing and exporting data into HDFS.
  • Actively participated in weekly meetings with the technical teams to review the code.

Environment: Hadoop, BigData, HDFS, Pig, Hive, MapReduce, Sqoop, Kafka, Linux, Cloudera, Talend, Big Data, Java APIs, Java collection, SQL, NoSQL, Cassandra, Monod, AJAX.

Confidential, Atlanta, GA

Big-Data and Hadoop Developer

Responsibilities:

  • Involved in Installing, Configuring Hadoopecosystem, and Cloudera Manager using CDH3 Distribution.
  • Involved in creating Hive tables, loading the data and writing hive queries that will run internally in MapReduce.
  • Managed Metadata with Big Data.
  • Involved in writing MapReduce jobs.
  • Used Talend for BigData Integration to generate the native code to work with Hadoop and Spark.
  • Used Impala for parallel processing the data.
  • Written MapReduce programs using Python.
  • Performed Data Serialization using Talend.
  • Real time streaming the data using Spark with Kafka.
  • Responsible for developing data pipeline using flume, Sqoop and pig to extract the data from weblogs and store in HDFS.
  • Used Scala to write MapReduce programs.
  • Installed and configured Hive and also written Hive UDFs.
  • Involved in emitting processed data from Hadoop to relational databases or external file systems using Sqoop, HDFS GET or CopyToLocal.
  • Developed data pipeline using Flume, Sqoop, Pig and Java MapReduce to ingest customer behavioral data and financial histories into HDFS for analysis.
  • Experienced in managing and reviewing Hadoop log files.
  • Used Pig to do transformations, event joins, filter boot traffic and some pre-aggregations before storing the data onto HDFS.
  • Experience as Data Engineer SME.
  • Extracted and updated the data intoMonodusing Mongo import and export command line utility interface.
  • Written Hive queries for data to meet the business requirements.
  • Importing and exporting data into HDFS and Hive using Sqoop and Kafka.
  • Worked on tuning the performance of Pig queries.
  • Supported MapReduce programs those are running on the cluster.
  • Involved in developing Pig Scripts for data change capture and delta record processing between newly arrived data and already existing data in HDFS.
  • Designed and Developed Dashboards using Tableau.
  • Gained experience in managing and reviewing Hadoop log files.
  • Involved in pivoting the HDFS data from Rows to Columns and Columns to Rows.

Environment: Hadoop, BigData, MapReduce, Mongo, Yarn, Hive, Pig, HBase, Oozie, Sqoop, Flume, Talend, Oracle 11g, Core Java, Cloudera, HDFS, Eclipse.

Confidential, Phoenix, AZ

Hadoop Developer

Responsibilities:

  • Collecting and aggregating large amounts of log data using Apache Flume and staging data in HDFS for further analysis.
  • Experienced in running Hadoop streaming jobs to process terabytes of xml format data.
  • Participated in requirement gathering and analysis phase of the project in documenting the business requirements by conducting workshops/meetings with various business users.
  • Involved in Sqoop, HDFS Put or CopyFromLocal to ingest data.
  • Used Hive to analyze the partitioned and bucketed data and compute various metrics for reporting.
  • Created and maintained technical documentation for launching Hadoop clusters and for executing Hive queries and Pig Scripts.
  • Implemented test scripts to support test driven development and continuous integration.
  • Implemented SQL, PL/SQL Stored Procedures.
  • Involved in developing Shell scripts to orchestrate execution of all other scripts (Pig, Hive, and MapReduce) and move the data files within and outside of HDFS.
  • Involved in developing Hive UDFs for the needed functionality that is not out of the box available from Apache Hive.
  • Experience working on processing unstructured data using Pig and Hive.
  • Proficient work experience with NoSQL,Monoddatabases.
  • Strong experience on Apache server configuration.

Environment: Hadoop, MapReduce, Monod, Hive, Pig, Sqoop, Core Java, Cloudera, HDFS, Eclipse.

Confidential - San Roman, CA

Hadoop Developer

Responsibilities:

  • Developed simple to complex MapReduce Jobs.
  • Installed and configured MapReduce, Hive and the HDFS; implemented CDH3 Hadoop cluster on CentOS. Assisted with performance tuning and monitoring.
  • Monitoring the running MapReduce programs on the cluster.
  • Pre-processing data using Hive and Pig.
  • Participated in translating complex functional and technical requirements into detailed design.
  • Had been a part of a POC effort to help build new Hadoop clusters.
  • Written Pig Latin Scripts to analyze and process the data.
  • Importing and exporting data into the HDFS and Hive using Sqoop.
  • Created partitioned tables in Hive.
  • Imported the data from various data sources, performed transformations by using MapReduce, Hive.
  • Developed and implemented the workflows using Apache Oozie for tasks automation.
  • Continuous monitoring and managing the Hadoop cluster using Cloudera Manager.
  • Worked on Oozie workflow to run multiple jobs.

Environment: Hadoop, HDFS, MapReduce, Hive, Pig, Sqoop, Oozie, Cloudera, Eclipse

Confidential, Plano, TX

Java Developer

Responsibilities:

  • Responsible for requirement gathering and analysis through interaction with end users.
  • Implemented the DAO pattern.
  • Designed and developed front end using HTML, JSP and Servlets.
  • Developed the application using Struts Framework to implement a MVC design approach.
  • Validated all forms using Struts validation framework.
  • Involved in developing JSP pages using Struts custom tags, JQuery and Tiles Framework.
  • Used SOAP messaging service to implement information exchange.
  • Worked in writing commands using UNIX Shell scripting.
  • Implemented action classes, form beans and JSP pages interaction with these components.
  • Actively involved in backend tuning SQL queries/DB script.
  • Involved in developing other subsystems’ server-side components.
  • Developed a Web service to communicate with the database using SOAP.
  • Created UML class diagrams that depict the code’s design and its compliance with the functional requirements.
  • Wrote a controller Servlet that dispatched requests to appropriate classes.
  • Debugged and developed applications using Rational Application Developer (RAD).

Environment: Java EE 6, IBM WebSphere Application Server 7, Apache-Struts 2.0, EJB 3, JSP 2.0, Web services, JQuery 1.7, Servlet 3.0, Struts-Validator, Struts-Tiles, Tag Libraries, JDBC, Oracle 11g/SQL, JUNIT 3.8, Rational clear case, Eclipse 4.2, JSTL, DHTML.

Confidential

Java Developer

Responsibilities:

  • Develop GUI related changes using JSP, HTML and client validations using Javascript.
  • Implemented client side validation using JavaScript.
  • Developed user interface using JSP, Struts Tag Libraries to simplify the complexities of the application.
  • Developed business logic using Stateless session beans for calculating asset depreciation on Straight line and written down value approaches.
  • Involved coding SQL Queries, Stored Procedures and Triggers.
  • Created REST based web service in JSON, RSS and CSV format.
  • Created java classes to communicate with database using JDBC.
  • Responsible for design and implementation of various modules of the application using Struts-Spring-Hibernate architecture.
  • Developed the Web Interface using Servlets, Java Server Pages, HTML and CSS.
  • Extensively used the JDBC Prepared Statement to embed the SQL queries into the java code.
  • Developed DAO (Data Access Objects) using Spring Framework 3.
  • Developed Web applications with Rich Internet applications using Java applets, Silverlight, Java.
  • Used JavaScript to perform client side validations and Struts-Validator Framework for server-side validation.
  • Designed and developed the application using various design patterns, such as session facade, business delegate and service locator.
  • Involved in designing use-case diagrams, class diagrams, interaction using UML model with Rational Rose.

Environment: Java, Servlets, JSP, EJB, J2EE, STRUTS, XML, XSLT, Javascript, HTML, CSS, Spring 3.2, SQL, PL/SQL, MS Visio, Eclipse, JDBC, Windows XP.

We'd love your feedback!