Senior Hadoop Developer Resume
Beaverton, OregoN
SUMMARY
- 10 years of overall experience in IT.
- 9 +years of hands - on experience in development and design of Java and related frameworks.
- 5 years of experience in Hadoop and it's Eco-system components.
- Experience in working with MapReduce programs using Apache Hadoop and CASCADING framework.
- Experience in setting up Hadoop clusters, both in-house as well as on the cloud (Amazon EC2).
- Experience in leading a team of 10-12 members at Confidential .
- Developed state-of-the-art Dependency manager (in place of Oozie) in Hadoop and also an application parallel to EMR. dis saved ~$20/hour/EC2 instance at Confidential .
- Developed applications with a combination of Pig, Hive and Sqoop/Oraoop.
- Developed Python and shell scripts.
- Extended Hive and Pig core functionality by writing custom UDFs.
- Experience in analyzing data using Hive and Map Reduce programs in Java.
- Familiar with Java virtual machine (JVM) and multi-threaded processing.
- Worked on NoSQL database - MongoDB.
- Set-up Asgard to spin-up clusters in AWS.
- Implemented Chaos Monkey on the AWS cluster at Confidential .
- Conducted POC’s on Spark.
- Experience on working with YARN.
- Set-up SENTRY after moving to CDH 5.1.2 in the Hadoop cluster.
- Set-up code reviews with the developers to maintain best coding practices for JAVA.
- Knowledge in job work flow scheduling and monitoring tools like Oozie, Autosys and Zookeeper.
- Experience as a Java Developer in Web/intranet, client/server technologies using Java, JDBC and SQL.
- Developed ETL pipe-lines to export/import the data from/into Salesforce.
- Contributor to the Apache Sqoop project. Submitted multiple patches to the open source community.
TECHNICAL SKILLS
Languages: JAVA, J2EE, Python
Frameworks: Struts 1.3, Spring 3.0, Play, Cascading 2.5
IDE: Eclipse (Helios)
Servers: Heroku, RAD
Database: MySql,Oracle and MongoDB (on MongoHQ)
Scripting languages: JavaScript, JQuery 1.6.4, JQuery UI
Big Data Technologies: Hadoop, Hbase, Zookeeper, Sentry, Hive, Pig, Oozie
Hadoop Distribution: Cloudera (CDH4.3.1 and CDH 5.1.2)
Others: JIRA, Bit-Bucket, Jenkins, Autosys, Teamcity, Kraken, SVN, AWS
PROFESSIONAL EXPERIENCE
Confidential, Beaverton, Oregon
Senior Hadoop Developer
Responsibilities:
- Helped to architect the production and pre-production (development) Hadoop clusters.
- Leading a team of 10 Hadoop developers, 2 admin's and 2 tester's.
- Set-up Cloudera's CDH4.3.1 and upgraded to CDH 5.1.2 after multiple interactions with the Cloudera team.
- Developed Map-Reduce programs to clean and aggregate the data. We use cascading framework for the latest map-reduce jobs.
- Developed Hive queries to aggregate the click-stream data, that was imported into HDFS using Sqoop.
- Developed PIG scripts which can perform multiple aggregations on a single data set.
- Configured the above jobs in Autosys.
- Meeting with the business on a weekly basis to collect their requirements, which will need the Big Data architecture to step-in.
- Working on a POC on Spark in AWS.
- Developed a migration plan to move the Big Data platform onto the Cloud with Cloudera CDH 5.1.2.
Environment: JAVA, HADOOP, CASCADING, HIVE, OOZIE, CHAOS MONKEY, KERBEROS, PIG, AUTOSYS, TEAMCITY, KRAKEN, CDH 4.3.1 (migrated to CDH 5.1.2), ASGARD
Confidential
Senior Hadoop Developer
Responsibilities:
- Set-up a Hadoop cluster in Amazon EC2.
- Developed Map-Reduce programs and Hive queries to aggregate the raw data (granularity - daily), that was pushed to HDFS.
- Developed Map-reduce programs to make logical changes to the incoming granular data. dis data was then sent to PIG scripts for processing.
- dis aggregated data is then pushed to MongoDB, which is hosted on MongoHQ, to serve the BI/ad-hoc reporting functionality that we offer in Salesforce.
- Set-up Hadoop cluster locally as well as on Amazon EC2.
- Set-up MongoDB on MongoHQ, to place the aggregated data into it.
- Utilized Oozie to manage the flow of jobs in the cluster.
- Developed a Bulk API, to load/extract the data into/from Salesforce.
- Used Flume to pull the data from Amazon S3 to Amazon EC2 Hadoop cluster.
- Implemented the Kerberos security in the Amazon EC2 Hadoop cluster.
- Also used HIVE to aggregate the raw data that we get from clients.
- Contact point for the infrastructure and development team in the organization.
- Collecting on-the-fly changes from the client. These changes went into the primary version of our component.
- Presiding over the scrum team on a weekly basis and making sure that the tasks are completed in JIRA.
Environment: JAVA HADOOP, HIVE, OOZIE, FLUME, PLAY, HEROKU, MONGODB, LZO, KERBEROS, AMAZON EC2, AMAZON S3, SALESFORCE
Confidential
Hadoop Developer
Responsibilities:
- We have written Map/Reduce programs, Pig scripts to specify the conditions to separate the fraudulent claims.
- These conditions were derived after processing the past 20 years of health care data that we possesses.
- dis data was exported into HDFS using Sqoop.
- Hive was used to integrate with Micro-Strategy - to produce results quickly based on the report that was requested.
- Flume was used to collect the logs with error messages across the cluster.
- Oozie and Zoo-keeper were used to manage the flow of jobs and coordination in the cluster respectively.
- LZO compression technique was used to make optimum utilization of the network bandwidth.
- Kerberos security was implemented to safeguard the cluster.
- Played a key-role is setting up a Hadoop cluster (in-house).
- Lead a team which consisted of 3 resources.
- Worked on a stand-alone as well as a distributed Hadoop.
- Worked on a Hortonworks & Cloudera Hadoop distribution.
- Updated the classification levels that were achieved on the data every day to the on-site team.
- Integrated RHadoop using connectors, with the cluster that we have set-up.
Environment: JAVA, HADOOP, HIVE, PIG, SQOOP, FLUME, HBASE, CDH 4.0.1
Confidential, Chicago, IL
Programmer Analyst
Responsibilities:
- Worked on STRUTS and ORACLE. Gained a very good hands-on experience.
- Developed the UI screens using web technologies like Javascript, JQuery, CSS.
- Worked on JAVA code to write map-reduce programs and SQOOP the data from ORACLE database.
- Made enhancements to the applications which presented me with an opportunity to go through the entire SDLC.
- Providing daily updates to the on-site team over a call and making enhancements.
- Mentored in a few enhancements which have gone through the PERF and UAT environments and are now live in production.
Environment: JSP, STRUTS, JQUERY, CSS, ORACLE, DOS, JAVA, HIBERNATE, XML, HADOOP, MAP-REDUCE, SQOOP.
Confidential, New York City, NY
Programmer Analyst
Responsibilities:
- Implemented reward points pages for AMEX using Java and JSP's at server and client end respectively.
- Designed unit-test cases using J-Unit.
- Collected the enhancement requirements from the clients through weekly calls.
- Updated the client manager on the tasks accomplished for the day.
Environment: JSP, STRUTS, ORACLE, JAVA.
Confidential,
Responsibilities:
- Formulated a plan by discussing with the team member and jotted down the requirements.
- Zeroed on the technologies that we can use for dis project by researching on multiple alternatives.
- Designed 3 pages using HTML, CSS, JavaScript and also have written JAVA code for the same.
- Wrote SQL queries to fetch the data from the database - MySQL.
Environment: JAVA, HTML, CSS, JavaScript, Tomcat, MySQL, SQL
Confidential
Responsibilities:
- Collected the requirements from the IT and CSE departments.
- Created data flow diagrams in UML.
- Participated in discussions which revolved around the enhancements that are needed in the website.
- Wrote SQL queries to fetch the data from the database.
Environment: JAVA, HTML, CSS, JavaScript, Tomcat, Oracle, SQL
