Sr. Hadoop Developer Resume
Chicago
SUMMARY:
- Over 10 years of IT experience in Software development in various domains with Hadoop Ecosystems and Java J2EE technologies.
- 3 years of hands on experience Technical Designing and Developing BIG DATA solutions using Hadoop Ecosystems (HDFS, MapReduce, Pig, Hive, Sqoop, Hbase, Cassandra and Oozie scheduler).
- 1 Year of hands on experience as Spark with scala developer.
- Have experience in Apache Spark for Batch processing and Spark Streaming.
- Used Apache Spark API over Hortonworks Hadoop YARN cluster to perform analytics on data in Hive.
- Exploring with Spark improving the performance and optimization of the existing algorithms in Hadoop using Spark context, Spark - SQL, Data Frame, pair RDD's, Spark YARN.
- Load the data into Spark RDD and performed in-memory data computation to generate the output response.
- 6 years of hands on experience with proficiency in technical designing, development, maintenance and support of applications using Java J2EE technologies.
- Hands on experience with Hadoop applications (such as development, configuration management, monitoring, debugging, and performance tuning). Experience on Hadoop clusters using major Hadoop distributions like Hortonworks, Cloudera (CDH) and AWS Servers.
- Worked on Hadoop admin activities as basic level and good understanding on Kerberos.
- Good Knowledge and understanding of Hadoop Architecture and various components in Hadoop ecosystems - HDFS, Map Reduce, Pig, Sqoop and Hive.
- Good Understanding in using Flume and Oozie Scheduler.
- Experience in writing Map Reduce jobs using Java and scheduling workflows with Oozie scheduler.
- Hands-on experience in using Sqoopto import and export data from RDBMS to Hadoop and vice-versa.
- Developed Pig Latin scripts using operators such as LOAD, STORE, DUMP, FILTER, DISTINCT, FOREACH, GENERATE, GROUP, COGROUP, ORDER, LIMIT, UNION, SPLIT to extract data from data files to load into HDFS.
- Hands-on experience in writing Pig Latin scripts, working with grunt shells and scheduling workflows with Oozie scheduler.
- Have Good Knowledge on Talend for Integration and Hadoop.
- Collecting and aggregating large amount of Log data using Apache Flume and storing data in HDFS for further analysis.
- Basic knowledge on Apache Mahout and Python.
- Experience using different file formats like Binary, XML, JSON and CSV files.Experience in developing solutions to analyze large data sets efficiently.
- Have experience on Apache Kafka.
- Apache Kafka is a distributed commit log service.
- Kafka integrates this unique abstraction with traditional publish/subscribe messaging concepts such as producers, consumers, and brokers.
- Kafka supports High throughput, supporting hundreds of thousands of messages per second.
- Kafka Supports for parallel data load into Hadoop.
- Have experience on Cassandra.
- Have experience on Impala.
- Worked on cassandra for maintaing larger data and versions.
- Have good knowledge on Java 8 features.
- Extensively worked in entire system development life cycle (SDLC) phases namely Analysis, Design, Development, Testing, Deployment and Maintenance.
- Have experience on SQLR and Kafka.
- Proficient in Test Driven Development (TDD) process and extensive experience with Agile and SCRUM programming methodology.
- Experience in performing data validation using HIVE dynamic partitioning and bucketing.
- Extensive experience working in Oracle and Microsoft SQL Server database.
- Experience in application development using J2EE technologies like Servlets, JSP, JDBC, JNDI, and RESTfull Web Services and HTML, CSS, Java Script, JSTL, Custom tags in designing web pages.
- Experience in Java API design and Debugging.
- Have experience on JVM.
- Have experience on Business Data Analys.
- Have experience on Requirement Analysis, FIT and GAP Analysis.
- Have experience on Machine Learning and data analytics on Big data set.
- Have experience on Healthcare domain.
- Have knowlege on Socket programing.
- Ability to work independently with minimum supervision in a team.
- Drive the team and collaborate to meet project timelines.
- Support the implementation and drive it to stable state in production.
- Interacted with the end users, Business Analysts for understanding the Business requirements.
- Involved in the walkthroughs and worked closely with business partners to understand their changing needs and made appropriate recommendations to project scope and approach.
- Worked to give End to end support during User acceptance testing and regression testing.
- Good experience in working with team members from vendors and internal departments to coordinate activities across multiple applications.
- Have good knowledge on ETL's - Informatica for pulling the data from different databases.
- Involved in training sessions conducted by the organization to make work users feel comfortable with the new application.
- Highly organized with ability to manage multiple projects and meet deadlines.
- Good knowledge in Enterprise Logistics - Warehouse Management System and Retail Store domains.
- Proactive and well organized with effective time management skills and problem solving skills.
- Good Inter personnel skills and ability to work as part of a team.Exceptional ability to learn and master new technologies and to deliver outputs in short deadlines.
TECHNICAL SKILLS:
Programming Languages: Scala and Java
Open Source Frameworks: Hadoop ecosystems, Struts, Hibernate
Web Services: Web Service (RESTful and SOAP).
Web Technologies: JavaScript,HTML, jQuery, Angular JS, CSS3, XML, AJAX.
Databases: Oracle 11i,Sql Server 2008 and My SQL5, HSQL (Hyper SQL),No-SQL Database (Hbase, Mongodb)
Database Tools: Microsoft SQL Server Management Studio, Toad and PLSQL Developer.
Application/ Web Servers: JBoss 5.1, Websphere Application Server, Weblogic 10.3 and Oc4J full suit.
IDE (Integrated Development Environment): Scala IDE, Eclipse and Jdeveloper.
Methodologies: Agile, Water Fall.
Configuration Management Tools: Ambari, WinCVS, GIT, TFS, SVN, Araxis Merge and Exam Diff.
Operating Systems: Linux Cent OS, Windows 98/2000/NT /XP/Windows 7/Windows 8.1.
Development Methodologies: Agile, Scrum and Waterfall
Domain Experience: Logistics and Health and Finance
Build Tool: Apache Ant and Maven
BigData Technologies: HDFS, MapReduce, Hive, Pig, Sqoop, Oozie, Hadoop Streaming, Zookeeper, Apache Spark, Kafka
Hadoop Distributions: Hortonworks, Cloudera
Other Tool: JUnit, Soap UI, Putty, WINSCP, Notepad++ and Edit Plus
PROFESSIONAL EXPERIENCE:
Confidential, Chicago
Sr. Hadoop Developer
Responsibilities:
- Interactions with Mc-Donald’s team for status updates, discussions for business requirements.
- Creating technical design specification’s to be used for coding by development team.
- Actively involved in writing and coding the business logic for the application.
- Ensuring timely monitoring and completion of TLD jobs.
- Handling Migration requests for PROD and TEST environments.
- Performing Root Cause Analysis (RCA) for TLD job failures and data mismatches.
- Transitioning knowledge from meetings, discussions, brainstorming sessions and Task assignment and tracking status, query tracking & resolution and review of deliverables.
- Worked on improving the performance for MapReduce job's.
- Analyzed large amounts of data sets to determine optimal way to aggregate and report on it.
- Setup Oozie workflow/Sub workflow Jobs for MapReduce/Hive/Scoop/HDFS actions.
- Involved in loading data from UNIX file system to HDFS.
- Extracted the data from Oracle into HDFS using Sqoop.
- Handled importing of data from various data sources, performed transformations using MapReduce, Spark and loaded data into HDFS.
- Manage and review Hadoop log files.
- Troubleshooting production support issues post-deployment and come up with solutions as required.
- Worked on data analysis and giving reports on daily basis.
- Check the registered logs in database whether the file status is properly updated or not.
- Handling the backup for input files in HDFS.
- Unit tested and supported QA testing of applications.
- Participated in code reviews and technical design meetings.
- Created product files for deployment to higher environments.
- Participated in code reviews and worked closely with external teams within the organization.
- Involved in Bug Fixing being the support team during the regression testing.
- Used JIRA to fix the Defects that are raised by the QA team.
- Have worked on Hive SQL queries.
- Have done performance challenges on Hive.
- Have worked with different formats on Hive.
- Have done Batch process by using apache spark with scala programing.
- Have done Real Time data process by using Apache spark Streaming.
- Have integrated Apache spark frame work with Hive componet.
- Have worked with parttions on Hive to get the query results.
- Have done SCALA POC's on this project to implement in the other project.
- Have done OOP's and functional programing on SCALA.
Environment: and Technologies: Hadoop Ecosystems - (HDFS, Spark with Scala, MapReduce with java, Sqoop and PIG, Hive, Impala and HBase), Java, SQL Server, Oracle11g, Eclipse, WinCVS, Putty, CentOS, JIRA and Hortonworks distribution.
ConfidentialHadoop Developer
Responsibilities:
- Participating in meetings and interactions with RIOTINTO team for status updates, discussions for business requirements.
- Creating technical/design specification’s to be used for development/coding by development team.
- Ensuring timely monitoring and completion of PAHS jobs.
- Handling Migration requests for PROD and TEST environments.
- Performing Root Cause Analysis (RCA) for PAHS job failures and data mismatches.
- Transitioning knowledge from meetings, discussions, brainstorming sessions. Task assignment and tracking status, query tracking & resolution and review of deliverables.
- Worked on Performance tuning on Mapreduce jobs.
- Worked on AWS technical support for Migration.
- Worked on data analysis and send the reports to clients on daily basis.
- Worked with team and collaborate to meet project timelines.
- Reviewing code and providing feedback relative to best practices, improving performance etc.
- Worked on Performance tuning on Mapreduce jobs.
Environment: and Technologies: Hadoop Ecosystems - (HDFS and MapReduce), Java, SQL Server, Oracle11g, Eclipse, WinCVS, Putty, Linux CentOS, JIRA and Hortonworks distribution.
ConfidentialHadoop Developer
Responsibilities:
- Support all business areas of DELL with critical data analysis that helps team members make profitable decisions as a forecast expert and business analyst and utilize tools for business optimization and analytics.
- Experience and talents to be a part of ground breaking thinking and visionary goals. As an Executive Analytics, we take the lead to Delivery analyses/ ad-hoc reports including data extraction and summarization using big data tool set.
- Practical work experience with Hadoop Ecosystem (i.e. Hadoop, Hive, Pig, Sqoop etc.)
- Experience with Unix and/or Linux.
- Conduct Trainings on Hadoop MapReduce, Pig and Hive. Demonstrates up-to-date expertise in Hadoop and applies this to the development, execution, and improvement.
- Ensures technology roadmaps are incorporated into data and database designs.
- Experience in extracting large data sets is a HUGE plus.
- Experience in data management and analysis technologies like Hadoop, HDFS.
- Create list and summary view reports.
- Handling and communicating with business and understanding the problems from business perspective rather than as a developer perspective.
- Preparing the Unit Test Plan and System Test Plan documents.
- Preparation & Execution of unit test cases and Troubleshooting and debugging.
Environment: and Technologies: Cloudera Hadoop, Linux, HDFS, Maprduce, Hive, Pig, Sqoop, Oracle, SQL Server, Eclise, Java and Oozie scheduler.
ConfidentialJava Technical Designer
Responsibilities:
- Participated in CR discussions and design activities for new requirements.
- Preparing DTD against the CR/Design Document.
- Prepared Unit Test Specification and Integration test Specification.
- Part of coding and development.
- Project Documentation and maintenance.
- Provided knowledge sharing sessions to the team members
- Reviewing the technical design documents for Flows and Flow Actions.
- Participated in creating validation rules
- Involved in Bug fixes and Unit testing.
- Involved in SQL tuning and database tuning on Oracle Server.
- Involved in configuring, deploying the application and worked with Trouble Shootings.
Environment: and Technologies: Java J2EE, Restful Services, Oracle 10g, Eclipse, WAS 6.0 and Weblogic.
ConfidentialJava Technocal Designer
Responsibilities:
- Involved in fixing bugs and enhancements for the modules.
- Involved in Requirements gathering and participating in design discussions
- Prepared DTD's for CR's.
- Developed critical implementations.
- Involved in license key implementation.
- Reviewed UTS, ITS.
- Prepared Code review documents.
- Reviewed CR Test cases.
- Involved in Integration Testing and Bug fixing.
- Involved in UAT Deployments.
- Involved in maintenance and support.
- Prepared Project Approach Document based on designed documents
- Involved in training efforts for junior developers, resolving issues.
Environment: and Technologies: Eclipse IDE, Oracle 10g database, jboss-4.2.2.GA container and J2EE Technologies.
ConfidentialJava Technical Designer
Responsibilities:
- Involved in jboss-4.2.2.GA migration from oc4j 10.1.3.
- Prepared DTDs for CRs.
- Reviewed UTS, ITS.
- Involved in SAAS model implementation.
- Involved in CR development.
- Involved in Integration Testing and Bug fixing.
- Involved in maintenance and support.
- Updating the client with weekly calls about the project status.
Environment: and Technologies: Java J2EE, Restful web services, Oracle10g, Eclipse3.6, PL/SQL, 4S PMP Tool, Jprofiler and Connection leakage Tool.
ConfidentialJava Developer
Responsibilities:
- Developed the required GUIs.
- Developed Product Requirements.
- Reviewed UTS, ITS.
- Prepared and reviewed PR Test cases.
- Involved in Unit Testing and Bug fixing.
- Involved in Load Testing.
- Involved in maintenance and support.
Environment: and Technologies: Java J2EE, Oracle10g, PL/SQL, Oc4j10.1.3 Container and Rational ClearCase.
ConfidentialJava Developer
Responsibilities:
- Developed Product Requirements.
- Prepared UTS and ITS.
- Involved in Unit Testing and Bug fixing.
- Involved in creating help documentation for the end user.
- Involved in maintenance and support.
Environment: and Technologies: Java J2EE, Oracle10g, Eclipse, PL/SQL, Oc4j10.1.3 Container.
