Hadoop Developer Resume
Confidential Plano, TX
PROFFESIONAL SUMMARY:
- Over 6 years of overall IT experience in a variety of industries, which includes hands on experience of 3+ years in Big Data technologies and designing and implementing Map Reduce.
- Well versed in installation, configuration, supporting and managing of Big Data and underlying infrastructure of Hadoop Cluster.
- Hands on experience on major components in Hadoop Ecosystem like Hadoop Map Reduce, HDFS, HIVE, PIG, HBase, Zookeeper, Sqoop, Oozie and Flume.
- Experience and responsible for functional as well as technical track of a project.
- Experience in developing Pig scripts and Hive Query Language.
- Experience writing custom UDFs in pig and hive based on the user requirement.
- Experience in storing, processing unstructured data using NOSQL databases like HBase, Cassandra.
- Written Hive queries for data analysis and to process the data for visualization.
- Experience in managing and reviewing Hadoop Log files.
- Experience in importing and exporting the different formats of data into HDFS, HBASE from different RDBMS databases and vice versa.
- Very good experience in complete project life cycle (design, development, testing and implementation) of Client Server and Web applications.
- Excellent Java development skills using J2EE, J2SE, Servlets, JSP, EJB, JDBC.
- Hands on experience on IDE tools like Eclipse, NetBeans, Visual Studio
- Good interpersonal and communication skills. Team player with strong problem solving skills.
SKILLS:
Hadoop/Big Data HDFS, MapReduce, Pig, Hive, HBase, Sqoop, Oozie, Flume, Spark, YARN, Java & J2EE Technologies Core Java. Hibernate, Spring, JSP, Servlets, Java Beans, JDBC, Oracle, MySQL, DB2, Windows … UNIX, Mac OS, Putty, WinScp, Stream weaver.: PROFESSIONAL EXPERIENCE:Hadoop DeveloperConfidential, Dallas, TX
Responsibilities:
- Responsible for building scalable distributed data solutions using Hadoop.
- Developed job processing scripts using Oozie workflow.
- Installed and configured Hive, Pig, Sqoop, Flume and Oozie on the Hadoop cluster.
- Developed Simple to complex Map/reduce Jobs using Hive and Pig.
- Involved in Hadoop cluster task like commissioning & decommissioning Nodes without any effect to running jobs and data.
- Wrote Map Reduce jobs to discover trends in data usage by users.
- Involved in running Hadoop streaming jobs to process terabytes of text data.
- Job management using Fair scheduler.
- Worked extensively with Sqoop for importing metadata from Oracle.
- Involved in creating Hive tables, and loading and analyzing data using hive queries.
- Designed, developed and did maintenance of data integration programs in a Hadoop and RDBMS environment with both traditional and non - traditional source systems as we as RDBMS and NoSQL data stores for data access and analysis.
- Experienced in running Hadoop streaming jobs to process terabytes of xml format data.
- Load and transform large sets of structured, semi structured and unstructured data.
- Responsible to manage data coming from different sources.
- Assisted in exporting analyzed data to relational databases using Sqoop.
- Wrote Hive Queries and UDF's.
- Developed Hive queries to process the data and generate the data cubes for visualizing.
- Created Pig Latin scripts to sort, group, join and filter the enterprise wise data.
- Implemented Partitioning, Dynamic Partitions, Buckets in HIVE.
- Gained experience in managing and reviewing Hadoop log files.
Environment: Hadoop, MapReduce, Sqoop, HDFS, Hive, Pig, Oozie, Spark, Java, Oracle 10g, MySQL.
Hadoop Developer
Confidential, Plano,TX
Responsibilities
- Worked on analyzing Hadoop cluster using different big data analytic tools including Pig, Hive and Map Reduce.
- Collecting and aggregating large amounts of log data using Apache Flume and staging data in HDFS for further analysis.
- Real time streaming the data using Spark.
- Configured Spark streaming to receive real time data from the Kafka and store the stream data to HDFS using Scala.
- Worked on debugging, performance tuning of Hive & Pig Jobs.
- Implemented test scripts to support test driven development and continuous integration.
- Worked on tuning the performance of Pig queries.
- Involved in loading data from LINUX file system to HDFS.
- Importing and exporting data into HDFS using Sqoop.
- Experience working on processing unstructured data using Pig.
- Implemented Partitioning, Dynamic Partitions, Buckets in Hive.
- Supported Map Reduce Programs those are running on the cluster.
- Gained experience in managing and reviewing Hadoop log files.
- Involved in scheduling Oozie workflow engine to run multiple pig jobs.
- Involved in using HCATALOG to access Hive table metadata from Map Reduce or Pig code.
- Computed various metrics using Java Map Reduce to calculate metrics that define user experience, revenue etc.
- Installed and configured Hive.
- Exported the result set from Hive to MySQL using Shell scripts.
- Implemented SQL, PL/SQL Stored Procedures.
- Actively involved in code review and bug fixing for improving the performance.
Environment: Hadoop, HDFS, Pig, Hive, Map Reduce, Sqoop, LINUX, Cloudera, Big Data, Java APIs, Java collection, SQL, NoSQL, HBase.
Java Developer
Confidential, Dallas, TX
Responsibilities:
- Played a very vital role starting from Requirements analysis, designing the Presentation templates and CTDs based on the business expectations.
- Actively participated in meetings with Business Analysts and Architects to identify the scope, requirements and architecture of the project.
- Followed MVC model and used spring frameworks for developing the Web layer of the application.
- Developed application using Spring MVC, JSP, JSTL and AJAX on the presentation layer, the business layer is built using spring and the persistent layer uses Hibernate.
- Developed User Interface and web page screens for various modules using JSF, JavaScript, and AJAX using RAD.
- Developed interfaces and their implementation classes to communicate with the mid-tier (services) using JMS.
- Extensively used JavaScript to provide dynamic User Interface and for the client side validations.
- Used AJAX framework for asynchronous data transfer between the browser and the server.
- Extensively used Java Multi-Threading concept for downloading files from a URL. Involved in developing XML compilers using XQuery.
- Written Java classes to test UI and Web services through JUnit.
- Performed functional and integration testing, extensively involved in release/deployment related critical activities. Responsible for designing Rich user Interface Applications using JSP, JSP Tag libraries, Spring Tag libraries, JavaScript, CSS, HTML.
- Used spring, Hibernate module as an Object Relational mapping tool for back end operations over SQL database.
- Developed the business components using EJB Session Beans.
- Involved in Database design for new modules and developed the persistence layer based on Hibernate.
- Implemented the J2EE design patterns Data Access Object (DAO), Session Façade and Business Delegate.
Environment: Java, J2EE, JSP, Spring, Hibernate, CSS, JavaScript, Oracle, Eclipse, JUnit, Web services, JNDI, JMS, HTML, XML, XSD, XML Schema.
Java/J2EE developer
Confidential
Responsibilities:
- Analyzed Business Requirements and identified mapping documents required for system.
- Performed requirement gathering, analyzing and negotiating customer requirements.
- Hands on experience on IDE tools like RAD (Rational Application Developer) and Eclipse.
- Responsible for creating, sending and receiving messages by using SOAP protocols.
- Configured Hibernate and spring through required configuration and XML files.
- Extensively used Spring Inversion-of-Control for Dependency Injection.
- Experience with WebSphere, Apache and IBM HTTP.
- Used Log4j for logging and JUnit for Unit Testing.
- Involved with writing SQL queries using Joins and Stored Procedures.
- Experience in Waterfall Software Development Lifecycle Model.
Environment: Servlet, Struts, Hibernate, Junit and DB2
