We provide IT Staff Augmentation Services!

Hadoop Developer Resume

5.00/5 (Submit Your Rating)

Los Angeles, CA

SUMMARY

  • Over 8 years of experience in the field of IT including three years of experience in Hadoop ecosystem and4+ years of experience as a java/J2EE developer with different client environments.
  • Good knowledge of Hadoop Architecture and various components such as HDFS, Job Tracker, Task Tracker,Data Node, Name Node and Map - Reduce conceptsResponsible for writing Map Reduce programs.
  • Experience in creating custom Lucene/Solr Query components.
  • Excellent knowledge on MySQL, Oracle, DB2 and RDBMS, Couchbase, Cassandra and MangoDB
  • Experience in developing Shell scripts and Python Scripts for system management.
  • Knowledge of administrative tasks such as installing Hadoop and its ecosystem components.
  • Expertise with the tools in Hadoop Ecosystem including Pig, Hive, HDFS, Map Reduce, Sqoop, Spark, Kafka,Yarn, Oozie, and Zookeeper. Hadoop architecture and its components.
  • Experience in tuning and troubleshooting performance issues in Hadoop cluster.
  • Worked on Multi Clustered environment and setting up Hotonworks Hadoop echo -System.
  • Worked on Ignite and Data Greed for the Apache, Ignite file system and distributed system.
  • Experience in the concept of natural language processing (NLP).
  • Experience in Datameer as well as big data hadoop. Experienced in NoSQL databases such as HBase, and MongoDB
  • Experience in using Cloudera Manager for installation and management of single-node and multi-nodeHadoop cluster.
  • Strong knowledge of Drools an JBoss Enterprise BRMS.
  • Worked on real-time, in-memory processing engines such as Spark, Impala and integration with BI Toolssuch as Tableau, OBIEE.
  • Worked on Service Oriented Architecture (SOA) such as Apache Axis web services which use SOAP.
  • Assisted in testing of AutoSys R11.3.5 instances for JIL and other scripts
  • Proficient in programming with Java and strong experience in technologies such as JSP Servlets, Struts,Spring, Hibernate,JDBC, Solr.
  • Strong data warehousing and OLTP system knowledge from database/ETL development perspective.
  • Efficient in packaging & deploying applications using ANT, Maven & Cruise Control on WebLogic, WebSphere& JBoss. Worked on the performance & load test related tools like JProfiler and JMeter.
  • Extensive experience in developing the SOA middleware based out of Fuse ESB and Mule ESB.AndConfigured, Elastic Search logstash, kibana to monitor spring batch jobs.
  • Worked with Cassandra for non-relational data storage and retrieval on enterprise use cases.
  • Written Apache Spark streaming API on Big Data distributions in the active cluster environment.
  • Expertise in n-tier and three-tier Client/Server development architecture and Distributed ComputingArchitecture.
  • Analytical thinker that consistently resolves on-going issues or defects, often called upon to consult onproblems as well a fast learner.
  • Detailed understanding of Software Development Life Cycle (SDLC) and sound knowledge of projectimplementation methodologies including Waterfall and Agile
  • Design and development of web-based applications using different Web and application servers such asApacheTomcat, Web Sphere, JBoss and WebLogic.
  • Developing applications using all Java technologies like Servlets, JSP, EJB, JDBC, JNDI, JMS etc.

TECHNICAL SKILLS:

Hadoop/Big Data: HDFS, Mapreduce, Solr, HBase, Pig, Hive, Sqoop, Flume, MongoDB, HBase, Oozie,Python, Perl,Map Reduce Zookeeper, YARN, NLP spark, storm, & Kafka, Cloudera, Hortonworks

Java Technologies: Java, J2EE, JSTL, JDBC3.0/2.1, JSP1.2/1.1, Java Servlets, JMS, JUnit, Log4j

IDE Development Tools: Eclipse 3.5, Net Beans, My Eclipse, Oracle JDeveloper 10.1.3, SOAP UI, Ant

Big data Analytics: Datameer 2.0.5

Frameworks: MVC Struts, Hibernate, spring, Apache Struts2.0

Programming languages: C, Java, Python, Ruby, Ant scripts, Linux shell scripts.

Databases: Oracle 11g, MySQL, MS SQL Server, IBM DB2 NoSQL Databases HBaseMongoDB, Cassandra.

RDBMS: Oracle 10g, MS Access, MS SQL Server, IBM DB2, PL/SQL.

Web Technologies: HTML, XHTML, CSS, XML, XML, JavaScript, AJAX, SOAP, WSDL.

Network Protocols: TCP/IP, UDP, HTTP, FTP.

Operating Systems: Windows XP/Vista/Windows 7 Windows 8, Windows 10, UNIX, LINUX.

WORK EXPERIENCE:

Hadoop Developer

Confidential, Los Angeles, CA

Responsibilities:

  • Created Hive Tables, loaded retail transactional data from Teradata using Sqoop.
  • Loaded home mortgage data from the existing DWH tables (SQL Server) to HDFS using Sqoop.
  • Orchestrated hundreds of sqoop scripts, pig scripts, hive queries using oozie workflows and sub - workflows.Managing Hadoop Clusters using Apache, Horton works, Cloudera and MapReduce
  • Loaded the load ready files from mainframes to Hadoop and files were converted to ASCII format.
  • Research Authorization re design using No-SQL DB like MongoDB. Used of Datameer for big data Analyticsand big data integration
  • Worked on major and minor upgrades of Couchbase and Cassandra cluster.
  • Generating Scala and java classes from the respective APIs so that they can be incorporated in the overallapplication.
  • Solved performance issues in Hive and Pig scripts with understanding of Joins, Group and Aggregation andhow does it translate to MapReduce job
  • Responsible to manage data coming from different sources and loaded into HDFS. Written Java Programsto load data into HDFS.
  • Writing Map Reduce programs to convert JSON, XML data to CSV data and loaded in HDFS.
  • Written Java program to move data to archive automatically while deleting data from HDFS.
  • Solid understanding of application program interfaces (APIs), messaging software and interoperabilitytechniques and standards
  • Developed Spark code and Spark-SQL/Streaming for faster testing and processing of data
  • Used of Data agreed we can improve performance and scalability of the application.
  • Installing and configuring Apache and supporting them on Linux production servers. Build Linux serversUpgrade and patch existing servers. Compile, built and upgrade Linux kernel.
  • Hands-on experience of system management, system setup and managing Linux or Solaris based serversas well as configuring them.
  • Configured nine nodes CDH5 Hadoop cluster on Red hat LINUX. Also involved in loading data from LINUX/UNIX file system to HDFS.
  • Extensively used Pig for data cleansing. Proficient work experience with NOSQL, Monod databases. Involvedin developing Hive DDLs to create, alter and drop Hive tables.
  • Data is loaded back to the Teradata for the BASEL reporting and for the business users to analyze andvisualize the data using Datameer. Real time streaming data using Spark with Kafka.
  • Installed and configured Hadoop MapReduce, HDFS, developed multiple MapReduce jobs in java for datacleaning and pre-processing.
  • Using of Elasticsearch in real time search which is scalable in search. Used Scala for coding the componentsin Play and Akka.
  • Configured Spark streaming to receive real time data from the Kafka and store the stream data to HDFSusing Scale.And databases such as HBase, and MongoDB
  • Installed cloudera, IBM biginsight across nodes and monitores nodes using cloudera and Ambari
  • Developed MapReduce programs to write data with headers and footers and Shell scripts to convert the datato fixed-length format suitable for Mainframes CICS consumption.
  • Elasticsearch handle the all long term persistence of the index
  • Involved in processing ingested raw data using Map Reduce, Apache Pig and Hive. Migrated databases fromDB2 to Hadoop eco system as part of Atlas data lake pro
  • Experience in collecting metrics for Hadoop clusters using Ganglia and Ambari.
  • Populated HDFS and Cassandra with huge amounts of data using Apache Kafka. Worked with NoSQLdatabases like Cassandra and Mongo DB for POC purpose.
  • Recommendation engine for Portfolio and Research articles using Apache Spark and MongoDB.
  • Agile methodology was used for development using XP Practices (TDD, Continuous Integration).

Environment: Hadoop, Kerberos Map Reduce, Ambari, MongoDB, NLP, Spark, Hive, Pig, Sqoop, Avro,Teradata, Scala, cloudera, SQL Server, Java 7.0, Solr, Log4J, Junit, MRUnit, SVN, HDFS.

Hadoop Developer

Confidential, Atlanta, GA

Responsibilities:

  • Involved in Sqoop, HDFS Put or Copy from Local to ingest data and Map Reduce jobs.
  • Used Kibana to delivering the products with data
  • Used Pig to do transformations, event joins, filter boot traffic and some pre - aggregations before storing thedata onto HDFS.
  • Written Hive queries for data analysis to meet the business requirements. Creating Hive tables and workingon them using Hive QL.
  • Kibana designed to take data from any source, search and analyze in it real time
  • Collected the log data from web servers and integrated into HDFS using Flume.
  • Developed workflow in Oozie to automate the tasks of loading the data into HDFS and pre-processing withPig.
  • Kerberos/LDAP skills Falcon, Atlas, Ranger, Knox, Ambari - Deep knowledge w/practical exp.
  • Elasticsearch is used to searchto all kind of documents
  • Developed data pipeline using Flume, Sqoop, Pig and Java map reduce to ingest customer behavioral dataand financial histories into HDFS for analysis.
  • Data ingestion experience with profiling exp is a strong plus. Experience on Apache Knox Gateway security for Hadoop Clusters
  • Created and written queries in SQL to update the changes in MySql when we upload or delete file in HDFS.
  • Extended support for application to work with Hive, Pig, Oozie and Sqoop. Experienced in Elasticsearch toprovide scalable search
  • Involved in developing Pig UDFs for the needed functionality that is not out of the box available from Apache
  • Pig. Collected, analyzed, monitored structured and unstructured data using ELK (Elastic Search.
  • Maintain the Autosys documentation and performance. Involved in handling level1 Support for Autosys WebPortal.
  • Involved in generating analytics data using Map/Reduce programs written in Python.
  • Used Pig as ETL tool to do Transformations, even joins and some pre-aggregations before storing the dataon to HDFS.
  • Used Hive to analyze the partitioned and bucketed data and compute various metrics for reporting. POCwork is going on using Spark and Kafka for real time processing.
  • Involved in developing Hive DDLs to create, alter and drop Hive tables and storm, & Kafka.
  • Managed works including indexing data, tuning relevance, developing custom tokenizers and filters, addingfunctionality includes playlist, custom sorting and regionalization with Solr Search Engine.
  • Involved in loading data from UNIX file system to HDFS. Installed and configured Hive and also written HiveUDFs and Cluster coordination services through Zoo Keeper.
  • Involved in creating Hive tables, loading with data and writing hive queries which will run internally in mapreduce way.
  • Involved in developing Hive UDFs for the needed functionality that is not out of the box available from ApacheHive.
  • Involved in using HCATALOG to access Hive table metadata from Map Reduce or Pig code.
  • Computed various metrics using Java Map Reduce to calculate metrics that define user experience, revenueetc.
  • Only Developed Simple to Quebec and Python Map/reduce streaming jobs using Python language that are implemented using Hive and Pig.
  • Responsible for developing data pipeline using flume, Sqoop and pig to extract the data from weblogs andstore in HDFS.
  • Extracted and updated the data into Monod using Mongo import and export command line utility interface.
  • Manage data in different databases in different datacenter using Couchbase, Elastic Search
  • Used in developingtechniques a model deriving insights from Social Media postings.
  • IncludingNLP techniques Kafka, zookeeper, and Elastic Search, rabbitMQ, Cassandra.
  • Extracted and updated the data into Monod using Mongo import and export command line utility interface.Involved in using Sqoop for importing and exporting data into HDFS.
  • Used Eclipse and ant to build the application. Proficient work experience with NOSQL, Mongodbdatabases.Also the HDFS data from Rows to Columns and Columns to Rows.
  • Involved in developing Shell scripts to orchestrate execution of all other scripts (Pig, Hive, and Map Reduce)and move the data files within and outside of HDFS.

Environment: Hadoop, Map Reduce, Knox, Mongo, Yarn, NLP, Spark, Hive, Solr, Pig, HBase, Oozie, Sqoop,Flume, Oracle 11g, Core Java, Cloudera, HDFS, Eclipse, zookeeper.

Java Developer

Confidential, New York, NY

Responsibilities:

  • Involved in the analysis, design, and development phases of Software Development Lifecycle (SDLC) usingagile development methodology.
  • Involved in business requirement gathering and technical specifications.
  • Implemented J2EE standards, MVC2 architecture using Struts Framework. And Implementing Servlets, JSPand Ajax to design the user interface.
  • Used JSP, Java Script, HTML5, and CSS for manipulating, validating, customizing, error messages to theUser Interface. Used JBoss for EJB and JTA, for caching and clustering purpose.
  • Presentation components in JSP pages are built using ICE faces tag libraries.
  • ICE Faces libraries are used in all presentation pages like Search/Inquiry and data collection pages.
  • Used EJBs (Session beans) to implement the business logic, JMS for communication for sending updatesto various other applications and MDB for routing priority requests.
  • All the Business logic in all the modules is written in core Java.
  • GUI was developed using JSP, AJAX and JavaScript, spring framework. Involved in the Development ofSpring Framework Controllers.
  • Configured the URL mappings and bean classes using Springapp - servlet.xml. Sybase was the databaseand Mybatis was used.
  • Used of Drools for Business Rule Management System Solution and it provide core business rule engine
  • Developed Mybatis in Data Access Layer to access and update information in the database.
  • Worked with Flied level engineers and teams to make the product more user-friendly. Performed testing forGUI and back end.
  • Wrote Web Services using SOAP for sending and getting data from the external interface.
  • Used XSL/XSLT for transforming and displaying reports Developed Schemas for XML.
  • Involved in writing the ANT scripts to build and deploy the application. Developed a web-based reporting formonitoring system with HTML and Tiles using Struts framework.
  • Used Design patterns such as Business delegate, Service locator, Model View Controller, Session, DAO.

Environment: JAVA multithreading, collections, SQL, PHP, Sybase, Eclipse, JavaScript, WebSphere, Drools,JBOSS, HTML5, DHTML, CSS, XML, Log4j, ANT, STRUTS 1.3.8, JUNIT, JSP.

Java Developer

Confidential

Responsibilities:

  • Worked on both WebLogic Portal 9.2 for Portal development and WebLogic 8.1 for Data ServicesProgramming. Developed the presentation layer using JSP, HTML, CSS and client validations usingJavaScript.
  • Used GWT to send AJAX requests to the server and updating data in the UI dynamically.
  • Developed Hibernate 3.0 in Data Access Layer to access and update information in the database.
  • Used JDBC, SQL and PL/SQL programming for storing, retrieving, manipulating the data.
  • Involved in designing and development of the ecommerce site using JSP, Servlet, EJBs, JavaScript andJDBC
  • Used Eclipse 6.0 as IDE for application development Configured Struts framework to implement MVC designpatterns. also Validated all forms using Struts validation framework and implemented Tiles framework in thepresentation layer and Designed and developed GUI using JSP, HTML, DHTML and CSS. Worked with JMSfor messaging interface
  • Used Hibernate for handling database transactions and persisting objects Deployed the entire project onWebLogic application server
  • Used AJAX for interactive user operations and client side validations Used XSL transforms on certain XMLdata.
  • Used XML for ORM mapping relations with the java classes and the database
  • Developed ANT script for compiling and deployment. Performed unit testing using JUnit
  • Used Subversion as the version control system. Extensively used Log4j for logging the log files.

Environment: Java/J2EE, Oracle 10g, SQL, PL/SQL, JSP, HTML, AJAX, Java Script, JDBC, XML, JMS, UML,JUnit, Log4j, Eclipse 6.0.

Java/J2EE Developer

Confidential

Responsibilities:

  • Written ANT Scripts to deploy the application into Tomcat application server for dev.
  • Monitored UNIX Switches utilizing Exceed, IBM Tivoli Netcool, Dotcom Monitor, WebStats and Nagios.Monitored network status via HP Open View.
  • Developed integrated systems by Implementing dozer mapping, Java, Spring, JAXBDiagnose and solve Application performance and stability issues.
  • Facilitated Daily Scrum, Sprint Planning, Sprint Review, & Sprint Retrospectives
  • Developed the application using J2EE Design Patterns like Delegate, Singleton, and DAO
  • Developed Visualization to use of kibana
  • Consumed web services from different applications within the network.
  • Developed Custom Tags to simplify the JSP2.0 code. Designed UI screens using JSP 2.0, CSS, XML1.1and HTML. Used JavaScript for client side validation.
  • Perform advance data analysis and visualize data in form of charts, maps and tables to use of akibana
  • Used Spring 2.5 Framework for Dependency injection and integrated with Hibernate and Struts frameworks.
  • Developed and implemented UI controls and APIs with ExtJS.
  • Created shell and perl scripts required in the project maintenance and software migration.
  • Built a framework for Agile Project and Program management office and aligned processes and tools.Implemented scalable server code and conducted unit testing.
  • Kibana plugins are helpful to dashaboard and discover the data of databases
  • Configured Hibernate's second level cache using EHCache to reduce the number of hits to the configurationtable data.
  • Designed and developed Utility Class that consumed the messages from the Java messageQueue andgenerated emails to be sent to the customers. Used Java Mail API for sending emails.
  • Used JUnit framework for unit testing of application and Log4j 1.2 to capture the log that includes runtimeexceptions.
  • Used Confidential for version control and used IBM RAD 6.0 as the IDE for implementing the application.

Environment: Weblogic Portal server 10.2, Java/J2EE, Spring, EJB 2.1, Struts 1.2, JMS, Windows XP, Unix,Oracle 10i, JQuery1.7.1, Ext-JS 3.1, BIRT Chart Library 3.0, Weblogic Workspace studio 10.2 and Eclipse

We'd love your feedback!