Hadoop Developer Resume
Chicago, IL
SUMMARY:
- Over 6 years of overall IT experience in a variety of industries, which includes hands on experience of 3+ years in Big Data technologies and designing and implementing Map
- Reduce.
- Well versed in installation, configuration, supporting and managing of Big Data and underlying infrastructure of Hadoop Cluster.
- Hands on experience on major components in Hadoop Ecosystem like Hadoop Map
- Reduce, HDFS, HIVE, PIG, HBase
- Experience and responsible for functional as well as technical track of a project.
- Experience in developing Pig scripts and Hive Query Language.
- Experience writing custom UDFs in pig and hive based on the user requirement.
- Written Hive queries for data analysis and to process the data for visualization.
- Experience in importing and exporting the different formats of data into HDFS, HBASE from different RDBMS databases and vice versa.
- Very good experience in complete project life cycle (design, development, testing and implementation) of Client Server and Web applications.
- Excellent Java development skills using J2EE, J2SE, Servlets, JSP, JDBC.
- Hands on experience on IDE tools like Eclipse, NetBeans, Visual Studio
- Good interpersonal and communication skills. Team player with strong problem solving skills.
TECHNICAL SKILLS:
- Sqoop
- Flume
- Spark
- YARN
- Java & J2EE Technologies
- Core Java
- Hibernate
- Spring
- JSP
- Servlets
- Java Beans
- JDBC
- Oracle
- MySQL
- DB2
- Windows … UNIX
- Mac OS
- Putty
- WinScp
- Stream weaver.
PROFESSIONAL EXPERIENCE:
Hadoop Developer
Confidential, Chicago, IL
Responsibilities:
- Worked on analyzing Hadoop cluster using different big data analytic tools including Pig, Hive and Map Reduce.
- Exported the result set from Hive to MySQL using Shell scripts.
- Implemented SQL, PL/SQL Stored Procedures.
- Actively involved in code review and bug fixing for improving the performance.
- Write sql scripts as needed to query/update data from Oracle data base.
- Collecting and aggregating large amounts of log data using Apache Flume and staging data in HDFS for further analysis.
- Real time streaming the data using Spark.
- Used Java Collections for iterating over collection of data
- Implemented test scripts to support test driven development and continuous integration.
- Used Java Threads for concurrent processing of data.
- Involved in loading data from LINUX file system to HDFS.
- Importing and exporting data into HDFS using Sqoop.
- Experience working on processing unstructured data using Pig.
- Implemented Partitioning, Dynamic Partitions, Buckets in Hive.
- Supported Map Reduce Programs those are running on the cluster.
- Involved in scheduling Oozie workflow engine to run multiple pig jobs.
- Involved in using HCATALOG to access Hive table metadata from Map Reduce or Pig code.
- Computed various metrics using Java Map Reduce to calculate metrics that define user experience, revenue etc.
- Installed and configured Hive.
Environment: Hadoop, HDFS, Pig, Hive, Map Reduce, Cloudera, Big Data, Java APIs, Java collection, SQL, NoSQL
Hadoop Developer
Confidential - Dallas, TX
Responsibilities:
- Responsible for building scalable distributed data solutions using Hadoop.
- Developed job processing scripts using Oozie workflow.
- Developed Simple to complex Map/reduce Jobs using Hive and Pig.
- Involved in Hadoop cluster task like commissioning & decommissioning Nodes without any effect to running jobs and data.
- Wrote Map Reduce jobs to discover trends in data usage by users.
- Involved in running Hadoop streaming jobs to process terabytes of text data.
- Worked extensively with Sqoop for importing metadata from Oracle.
- Involved in creating Hive tables, and loading and analyzing data using hive queries.
- Designed, developed and did maintenance of data integration programs in a Hadoop and RDBMS environment with both traditional and non-traditional source systems as RDBMS and NoSQL data stores for data access and analysis.
- Experience in running Hadoop streaming jobs to process terabytes of xml format data.
- Load and transform large sets of structured, semi structured and unstructured data.
- Responsible to manage data coming from different sources.
- Assisted in exporting analyzed data to relational databases using Sqoop.
- Wrote Hive Queries and UDF's.
- Created Pig Latin scripts to sort, group, join and filter the enterprise wise data.
- Implemented Partitioning, Dynamic Partitions, Buckets in HIVE.
Environment: Hadoop, MapReduce, Sqoop, HDFS, Hive, Pig, Spark, Java, Oracle 10g, MySQL.
Hadoop Developer
Confidential - Plano, TX
Responsibilities:
- Worked on analyzing Hadoop cluster using different big data analytic tools including Pig, Hive and Map Reduce.
- Real time streaming the data using Spark.
- Implemented test scripts to support test driven development and continuous integration.
- Involved in loading data from LINUX file system to HDFS.
- Importing and exporting data into HDFS using Sqoop.
- Experience working on processing unstructured data using Pig.
- Implemented Partitioning, Dynamic Partitions, Buckets in Hive.
- Supported Map Reduce Programs those are running on the cluster.
- Involved in scheduling Oozie workflow engine to run multiple pig jobs.
- Computed various metrics using Java Map Reduce to calculate metrics that define user experience, revenue etc.
- Installed and configured Hive.
- Exported the result set from Hive to MySQL using Shell scripts.
- Implemented SQL, PL/SQL Stored Procedures.
- Actively involved in code review and bug fixing for improving the performance.
Environment: Hadoop, HDFS, Pig, Hive, Map Reduce, Cloudera, Big Data, Java APIs, Java collection, SQL, NoSQL.
Java Developer
Confidential - Dallas, TX
Responsibilities:
- Played a very vital role starting from Requirements analysis, designing the Presentation templates and CTDs based on the business expectations.
- Actively participated in meetings with Business Analysts and Architects to identify the scope, requirements and architecture of the project.
- Followed MVC model and used spring frameworks for developing the Web layer of the application.
- Developed application using Spring MVC, JSP, and AJAX on the presentation layer, the business layer is built using spring and the persistent layer using Hibernate.
- Developed User Interface and web page screens for various modules using JSF, JavaScript, and AJAX using RAD.
- Developed interfaces and their implementation classes to communicate with the mid-tier (services) using JMS.
- Extensively used JavaScript to provide dynamic User Interface and for the client side validations.
- Used AJAX framework for asynchronous data transfer between the browser and the server.
- Extensively used Java Multi-Threading concept for downloading files from a URL.
- Written Java classes to test UI and Web services through JUnit.
- Performed functional and integration testing, extensively involved in release/deployment related critical activities. Responsible for designing Rich user Interface Applications using JSP, JavaScript, CSS, HTML.
- Used spring, Hibernate module as an Object Relational mapping tool for back end operations over SQL database.
- Involved in Database design for new modules and developed the persistence layer based on Hibernate.
- Implemented the J2EE design patterns Data Access Object (DAO), Session Façade and Business Delegate.
Environment: Java, J2EE, JSP, Spring, Hibernate, Oracle, Eclipse, Web services, HTML
Java/J2EE developer
Confidential
Responsibilities:
- Analyzed Business Requirements and identified mapping documents required for system.
- Performed requirement gathering, analyzing and negotiating customer requirements.
- Hands on experience on IDE tools like RAD (Rational Application Developer) and Eclipse.
- Responsible for creating, sending and receiving messages by using SOAP protocols.
- Configured Hibernate and spring through required configuration and XML files.
- Extensively used Spring Inversion-of-Control for Dependency Injection.
- Experience with WebSphere, Apache and IBM HTTP.
- Involved with writing SQL queries using Joins and Stored Procedures.
- Experience in Waterfall Software Development Lifecycle Model.
Environment: Servlet, Struts, Hibernate and DB2
