Hadoop Developer Resume
OH
SUMMARY
- Overall 7+ years of experience in IT industry which around 4 years of experience in Big Data in implementing complete Hadoop solutions.
- Expertise in Hadoop ecosystem components like Map Reduce, HDFS, Hive, Sqoop, Pig, Flume for scalability, distributed computing and high performance computing.
- In depth knowledge of Hadoop Architecture and various components such as HDFS, Job Tracker, Task Tracker, Name Node, Data Node and MRv1 and MRv2 (YARN).
- Experienced in creating Map Reduce jobs in Java as per teh business requirements.
- Successfully handled Migrating ETL projects to Hadoop wif zero defects and running in production wifout any issues.
- Experience in importing and exporting data using Sqoop from HDFS to Relational Database Systems and vice - versa.
- Hands on experience in configuring and working wif Flume to load teh data from multiple sources directly into HDFS.
- Strong experience in data analytics using Hive and Pig, including by writing custom UDFs.
- Knowledge of job workflow scheduling and monitoring tools like Oozie and Zookeeper.
- Strong understanding of NoSQL databases like HBase, MongoDB & Cassandra.
- Exposure to Cloudera development environment and management using Cloudera Manager.
- Extensive experience in Data Ingestion, In-Stream data processing, Batch Analytics and Data Persistence strategy.
- Experience in extracting data from both Relational systems and Flat Files.
- Knowledge of Publish-subscribe messaging system Kafka and real-time computation system Storm.
- Experience in Core Java and multi-thread processing.
- Extensive knowledge in using SQL Queries for backend database analysis.
- Good experience in using Linux/Unix shell scripting.
- Object Oriented Design (OOD) experience wif Rational Rose and Enterprise Architect (EA).
- Experience in UML (Class Diagrams, Sequence Diagrams, Usecase Diagrams).
- Hands on experience in application development using Core JAVA, RDBMS and Linux shell scripting.
- Experienced in creating and analyzing Software Requirement Specifications (SRS) and Functional Specification Document (FSD).
- Strong knowledge of Software Development Life Cycle (SDLC).
- Excellent working experience in Scrum / Agile framework, Iterative and Waterfall project execution methodologies.
- Experienced in preparing and executing Unit Test Plan and Unit Test Cases after software development.
- Worked extensively in Finance and Automotive Insurance domains.
- Experienced to work wif multi-cultural environment wif a team and also individually as per teh project requirement.
- Excellent communication and inter-personal skills, self-motivated, organized and detail-oriented, able to work well under deadlines in a changing environment and perform multiple tasks effectively and concurrently.
- Strong analytical skills wif ability to quickly understand clients business needs. Involved in meetings to gather information and requirements from teh clients.
TECHNICAL SKILLS:
Hadoop/Big Data: HDFS, MapReduce, Pig, Hive, Sqoop, Oozie, Zookeeper
IDE Tools: Eclipse, IntelliJ IDEA
Programming languages: C, C++, Java, Linux shell scripts, VB.NET
Databases: Oracle 11g/10g/9i, MySQL, DB2, MS-SQL Server, MongoDB
Web Technologies: HTML, XML, JavaScript
Defect Tracking Tools: IBMRational ClearQuest, Jira
Version control: Git, IBM Rational ClearCase
Testing Tools: JUnit, MRUnit
Design Technologies: IBM Rational Rose, MS Visio and UML
Development Approach: Agile, Waterfall, Iterative, Spiral, Kanban
Operating Systems: All Versions of Microsoft Windows, UNIX and LINUX
Protocols: TCP/IP, HTTP, HTTPS, TELNET, FTP and LDAP
PROFESSIONAL EXPERIENCE
Hadoop Developer
Confidential, OH
Responsibilities:
- Understanding business needs, analyzing functional specifications and map those to develop and designing MapReduce programs and algorithms.
- Responsible for building scalable distributed data solutions using Hadoop.
- Involved in loading data from RDBMS into HDFS using Sqoop queries.
- Handled Delta processing or incremental updates using hive and processed teh data in hive tables.
- Execution of Hadoop ecosystem and Applications through Apache HUE.
- Optimizing Hadoop MapReduce code, Hive/Pig scripts for better scalability, reliability and performance.
- Developed teh OOZIE workflows for teh Application execution.
- Involved in creating Hive Tables, loading wif data and writing Hive queries which will invoke and run MapReduce jobs in teh backend.
- Writing Pig scripts for data processing.
- Developed PIG Latin scripts to extract data from source system.
- Developed java Map reduce XML PARSER programs to process XML files using XSD’s and XSLT’s as per teh clients requirement and used to process teh data into Hive tables.
- Implemented Hive tables and HQL Queries for teh reports. Written and used complex data type in Hive. Storing and retrieved data using HQL in Hive. Developed Hive queries to analyze reducer output data.
- Developed Scripts and automated data management from end to end and sync up between all teh clusters.
- Highly involved in designing teh next generation data architecture for teh unstructured data.
- Extensively used teh Hue browser for interacting wif Hadoop components.
- Feasibility Analysis (For teh deliverables) - Evaluating teh feasibility of teh requirements against complexity and time lines.
- Documented teh systems processes and procedures for future references.
- Actively participated in software development lifecycle (scope, design, implement, deploy, test), including design and code reviews, test development, test automation.
- Involved in story-driven agile development methodology and actively participated in daily scrum meetings.
Environment: HDFS, Map Reduce, Java, Hive, Oozie, PIG, Shell Scripting, Linux, HUE, Sqoop, Flume, and Oracle 11g
Hadoop Developer
Confidential, Louisville, KY
Responsibilities:
- Explored and used Hadoop ecosystem features and architectures.
- Worked closely wif business team to gather their requirements and new support features.
- Developed Map-Reduce jobs for Log Analysis and Analytics.
- Wrote Map-Reduce job to generate reports for teh number of activities created on a particular day, during a time interval etc. for teh Analytics module.
- Teh MR Job read teh data from HDFS, where teh data was dumped from teh multiple sources and teh output was written back to HDFS.
- Configured Sqoop and developed scripts to extract data from MySQL into HDFS.
- Used Hive for analysis of web site traffic.
- Involved in writing Unix/Linux Shell Scripting for scheduling jobs and for writing pig scripts and hive QL.
- Wrote programs using scripting languages like Pig to manipulate data.
- Implemented teh workflows using teh Apache Oozie framework to automate tasks.
- Wrote Hadoop Job Client utilities and integrated them into monitoring system.
- Reviewed teh HDFS usage and system design for future scalability and fault-tolerance.
- Prepared Extensive Shell scripts to get teh required info from logs.
- Performed white box testing and monitoring all teh logs in Dev and Prod environments
Environment: HDFS, Map/Reduce Java, Sqoop, Pig, Hive, Oozie, Flume, Core Java, Apache Derby, MySQL and Linux.
Java Developer
Confidential
Responsibilities:
- Responsible for reviewing business user requirements and also participated in meeting teh users wif Business Analysts
- Developed teh User Interface using JSP/AJAX/ HTML / CSS/ Java Script
- Widely Used Design pattern like DAO, Singleton, Business delegate and Service Locator in teh process of system designing and development
- Used Message Driven Beans and JMS to process teh requests from teh customer asynchronously
- Developed stored procedures, cursors and database Triggers and implemented Scrollable Result sets
- Consumed Web Services (WSDL, SOAP, and UDDI) from third party to verify teh credit score of applicants
- Developed Web services using top-down approach and coded required WSDL files
- Used XSL/XSLT for transforming common XML format into displayable format
- Involved in testing teh system using JUnit
- Maintained teh source code versions in Subversion repository
- Used Log4J for logging and tracing teh messages
- Deployed application in Websphere Application Server and developed using RAD
Environment: Core Java, RSA7.0, SQLServer2008, Linux, Servlets 2.5, JSP 2.2, AJAX, HTML, XML, XSL, SOAP, WSDL, JUnit, Log4J, ANT.
Java Developer
Confidential
Responsibilities:
- Utilized Agile Methodologies to manage full life-cycle development of teh project.
- Developed front end validations using JavaScript and developed design and layouts of JSPs and custom taglibs for all JSPs.
- Used JDBC for database connectivity.
- Developed web application using JSP custom tag libraries and Action. Designed Java Servlets and Objects using J2EE standards.
- Used JSP for presentation layer, developed high performance object/relational persistence and query service forentire application.
- Developed teh XML Schema and Web services for teh data maintenance and structures.
- Used WebLogic Application Server and RAD to develop and deploy teh application.
- Worked wif various Style Sheets like Cascading Style Sheets (CSS).
- Designed database and created tables, written teh complex SQL Queries and stored procedures as per teh requirements.
- Involved in coding for JUnit Test cases, ANT for building teh application.
Environment: Java/J2EE, Oracle 10g, SQL, PL/SQL, JSP, EJB, WebLogic 8.0, HTML, AJAX, Java Script, JDBC, XML, JMS, JUnit, log4j, MyEclipse 6.0.
Java Developer
Confidential
Responsibilities:
- Responsible for understanding teh scope of teh project and requirement gathering.
- Review and analyze teh design and implementation of software components/applications and outline teh development process strategies
- Coordinate wif Project managers, Development and QA teams during teh course of teh project.
- Used Spring JDBC to write some DAO classes to interact wif teh database to access account information.
- Used Tomcat web server for development purpose.
- Involved in creation of Test Cases for JUnit Testing.
- Used Oracle as Database and used Toad for queries execution and also involved in writing SQL scripts, PL/SQL code for procedures and functions.
- Used CVS, Perforce as configuration management tool for code versioning and release.
- Developed application using Eclipse and used build and deploy tool as Maven.
- Used Log4J to print teh logging, debugging, warning, info on teh server console.
Environment: Java1.5, J2EE Servlet, JSP, XML, Spring 3.0, Design Patterns, Log4j, CVS, Maven, Eclipse, Apache Tomcat 6, and Oracle 11g.
