We provide IT Staff Augmentation Services!

Sr. Big Data Developer Resume

4.00/5 (Submit Your Rating)

Burlington, NJ

SUMMARY:

  • Over 10 years of experience in Information Technology in various business domains
  • Working experience with large scale Hadoop environments build and support including design, configuration, installation, performance tuning and monitoring
  • Hands - on experience in Big Data and HADOOP Ecosystem components. (HDFS, Map Reduce, PIG, HIVE, HBASE, SQOOP, Flume, Oozie, Zookeeper, Ambari, Ranger,Hue, Kafka, Storm, MongoDB, R language)
  • Experience in Adding, Configuring and Deploying New services in Hadoop Eco-System.
  • Deep understanding of schedulers, workload management, availability, scalability and distributed data platforms.
  • Expertise in Datalake process on Acquiring, preprocessing & ingesting different data formats from Clients.
  • Experience interacting with business users, understanding requirements and provide quick solution.
  • Experience building a commercially viable scalable solution using Big Insights Horton works and Cloudera distributions of Hadoop, HIVE and HBase, Hue, YARN, SPARK, sentry
  • Proficient in Map-Reduce Java and streaming APIs, Ruby, Python
  • Experience with Hadoop & Big Data architecture
  • Configured and monitored Hadoop clusters with BigInsights Enterprise distribution(3.0/4.0)
  • Configured and monitored Hadoop clusters with Cloudera Enterprise distribution(CDH4)
  • Experience with AWS (Amazon Web Services) and MapR
  • Experience working with Big Data Computing Platform (Spark, Map Reduce) and Big Data Manipulation Tools (Hive, and Pig).
  • Experience with Spark, Spark Streaming and Spark SQL
  • Deep understanding of schedulers, workload management, availability, scalability and distributed data platforms.
  • Expert understanding of ETL principles and how to apply them within Hadoop
  • Experience with Messaging and collection frameworks like Flume, Kafka
  • Experience with Oozie Workflow Engine to automate and parallelize Hadoop Map/Reduce, Hive and Pig jobs.
  • Experience with IBM BigSQL (Massively Parallel Processing) IBM STREAM
  • Experienced in developing application using the R, Big R programming language.
  • Experience with R and statistical modeling with RStudio.
  • Experience with Source control and Configuration Management tools and technologies (Git/GitHub)
  • Experience with developing and driving solution architectures employing portal,SOA, data management, data conversion and BI strategies, architecture and design p Confidential erns
  • Experience building business applications using RDBMS such as Oracle, DB2, MS SQL Server and MySQL
  • Importing/Exporting data from MYSQL/Teradata/Netezza/DB2 /Oracle to HDFS (Hadoop) using sqoop
  • Good understanding of performance tuning with both NoSQL and SQL technologies.
  • Good understanding of file formats including JSON, Parquet, Avro, and others.
  • Extensive experience building and designing large-scale distributed applications
  • Experience with agile/scrum methodologies to iterate quickly on product changes, developing user stories and working through backlogs
  • Experience with Talend RTX ETL tool, devalop jobs and scheduled jobs in Talend integration suite
  • Experience with ETL development with Sync sort/NiFi/Talend/TAC

PROFESSIONAL EXPERIENCE:

Confidential, Burlington, NJ

Sr. Big Data Developer

Responsibilities:

  • Developed Sqoop Framework to Source Historical Data from Oracle, DB2, Sql Server, Oracle
  • Ingest Flat files received via ECG FTP tool and files received from Sqoop into UHG Data Lake Hive and HBase using Data Fabric functionalities.
  • Validate Transactional Data Files coming from IBM Confidential (Change Data Capture) tool
  • Used Splunk Dashboard to record and monitor incoming file frequency from Confidential tool
  • Developed and automated jobs in Talend open studio to validate the Ingested Data.
  • Performed address standardization per business requirement and captured Golden Records.
  • Ingested XML files capture from RabbitMQ and Stored in HBase tables.
  • Developed HBase table for Monitor Ingestion Logs and snapshot logs.
  • Loaded the data to HBASE using bulk load and HBASE API.
  • Performing analytics using Talend Spark components on Insurance Claims Data.
  • Developed Data transformation module employing hive, Map reduce.
  • Involved in new development as well as bug fixing, performance tuning in the existing application.
  • Wrote Data transformation script using hive, Map reduce (Python)
  • Scheduled backups jobs by implementing Talend Tac Scheduler
  • Data Integrated, Extracted and transformed from mainframe and loaded to Hadoop using Talend
  • Configured BI tools to access hadoop cluster from windows.
  • Designed/Developed framework/api to leverage platform capabilities using map reduce /HDFS

Environment:: MapR, Hive, Hbase, Sqoop, Pig, Talend TAC, Map Reduce, Splunk, Eclipse, Maven,, Shell Script, JSP, Json, JDBC, XML, MYBATIS,UNIX, LINUX, DB2,Oracle,TERADATA, SVN,JBOSS, MYECLIPSE, JIRA, Subversion, Jenkins, Codehub, Autosys, Talend/TAC.

Confidential, NY

Sr. Big Data Developer

Responsibilities:

  • Developed Data transformation module employing hive, Map reduce.
  • Geospatial modeling, scripting and geostatistical application development for cloud based solution utilizing Map Reduce Hadoop
  • Setting Up ESRI Spatial Frame Work for Hadoop
  • Adding configuring ESRI jar to execute geospatial quires in hive and beeline.
  • Involved in new development as well as bug fixing, performance tuning in the existing application
  • Importing/Exporting data from MYSQL /Oracle to Azure cluster and AWS cluster using sqoop.
  • Worked on Hive Jason Serde to parse spatial / Enclosed JSON, Unenclosed JSON Data format.
  • Analyzing data using Hive & Pig Latin Scripting
  • Involved in new development as well as bug fixing, performance tuning in the existing application.
  • Wrote Data transformation script using hive, Map reduce (Python)
  • Scheduled backups jobs by implementing cron job
  • Data Integrated, Extracted and transformed from mainframe and loaded to Hadoop using (HDF) Nifi
  • Configured BI tools to access hadoop cluster from windows.
  • Designed/Developed framework/api to leverage platform capabilities using map reduce /HDFS

Environment:: AWS (Amazon Web Services), MICROSOFT AZURE, Hive, Sqoop, crontab, Map Reduce, Eclipse, Maven,, Shell Script, JSP, Json, JDBC, XML, MYBATIS,UNIX, LINUX, DB2,TERADATA, SVN,JBOSS, MYECLIPSE, JIRA, Subversion, Control M/Autosys. NIFI (HDF).

Confidential, Chicago, IL

Sr. Big Data Developer

Responsibilities:

  • Configured and monitored Hadoop clusters with Big Insights Enterprise distribution (3.0/4.0)
  • Configured and monitored Hadoop clusters with Cloudera Enterprise distribution(CDH4/CDH5)
  • Review Big Data Architect/Design/Release to maintain platform integrity
  • Installing and Configuring SysnSort
  • Developed Data transformation module employing hive, Map reduce
  • Developed User Defined Function (UDF) for hive.
  • Loaded the data to HBASE using bulk load and HBASE API.
  • Developing APIs to remotely access the hadoop cluster
  • Configuring BI tools to access hadoop cluster from windows desktops
  • Architect and Implement Hadoop eco system component such as HDFS, Map Reduce, HBase, Zookeeper, Pig, Hadoop streaming, Sqoop, Oozie, hive, hive server 2.
  • Designed applications for storing data to HDFS by using Kafka to get more performance.
  • Involved in new development as well as bug fixing, performance tuning in the existing application
  • Mentor other team members with the development of hive/udf/Map Reduce scripts
  • Designed and developed the presentation and web layers based on Java, J2EE
  • Developed Restful Web service using Jersey and JBOSS
  • Demonstrated and strong working knowledge of Global Delivery Models comprising onsite and offshore staffing model.
  • Data Integrated, Extract, transform from mainframe and loaded to Hadoop using Syncsort,
  • Talend open studio.
  • Assisted with data capacity planning and node forecasting.
  • Importing/exporting heterogeneous data(Ebcdic,Ascii..) using sqoop (native/fast connector ), custom api (map reduce sftp push & pull) and flume
  • Involved in new development as well as bug fixing, performance tuning in the existing application
  • Proactively recognizing potential performance improvements and mitigating potential issues.
  • Working Closely with IBM (Big Insights) team for future updates and recommendations.

Environment: Java/J2EE, Hadoop, Hbase, Hive, Map Reduce, Eclipse, Maven,, JavaScript, JSP, Python, HTML, JDBC, XML, MYBATIS,UNIX, LINUX, DB2,TERADATA, SVN,JBOSS, MYECLIPSE, JIRA, Subversion, Control M/Autosys. Sync sort. Talend . R Big R.

Confidential, NC

Hadoop Developer

Responsibilities:

  • Developed Data transformation module employing hive, Map reduce
  • Developed User Defined Function(UDF) for hive (Java)
  • Loaded the data to HBASE using bulk load and HBASE API.
  • Accessed the data from HBASE using Hbase Rest interface
  • Mentored other team members with the development of hive/udf/Map Reduce scripts
  • Designed and developed the presentation and web layers based on Java, J2EE
  • Developed Restful Web service using Jersey and JBOSS
  • Demonstrated and strong working knowledge of Global Delivery Models comprising onsite and offshore staffing model.
  • Configure and monitor Hadoop clusters with Cloudera Enterprise distribution(CDH4)

Environment: Java/J2EE, HAdoop, HBAase, Hive, Map Reduce, Eclipse, Maven,, JavaScript, JSP, HTML, JDBC, XML, MYBATIS, MySQL, DB2,TERADATA, SVN,JBOSS, MYECLIPSE, JIRA, Subversion, Control M/Autosys

Confidential, Pleasanton, CA

Hadoop Developer

Responsibilities:

  • Analyzing and reviewing business, functional and high-level technical requirements.
  • Analytical and problem solving skills, applied to a “Big Data” environment (map-reduce,hive,hive-server2)
  • Designing detailed technical components for complex applications utilizing high-level architecture, design p Confidential erns and reusable code.
  • Supporting detailed project estimates and project work plans.
  • Programming code independently for intermediate to complex modules following development standards.
  • Planning and conducting code reviews for changes and enhancements that ensure standards compliance and systems interoperability.
  • Creating test plans for module-level integration and tracking/resolving defects.
  • Preparing and packaging production-ready code modules for staging.
  • Ability to problem solve and reach resolution with a broader perspective in mind.
  • Able to effectively communicate with peers, business analysts, project managers, quality control, and across other technology team boundaries

Environment: Java/J2EE, Hadoop, HBase, Hive, Map Reduce, Impala, Eclipsw, JUnit,Linux,Windows, JIRA, Subversion, Oracle, MS-SQL server.

Confidential, Detroit, MI

Software Developer

Responsibilities:

  • Designed and developed the application architecture, use cases, and flowcharts using Microsoft Visio
  • Designed and developed the presentation and web layers (Transactional application) based on Java, J2EE (STRUTS, spring, Hibernate, Velocity), web services (Axis2) using Eclipse 3.0, Visual Basic, ASP using Visual Studio.
  • Worked with technologies such as JavaScript, XML, AJAX, HTML, and CSS.
  • Developed Web services Architecture for supporting common business functions with direct access to web services from PL/SQL using Oracle9i/ AXIS 2
  • Developed internal Tracking system to track marketing mediums, including organic and paid search, banner ads, referral links and email newsletters
  • Developed and deployed applications in UNIX and Windows environments using ANT tool and shell script.

Environment: J2EE, AXIS2,AJAX, JavaScript, Struts, TILES, ANT, JSP, HTML, JDBC, XML, Hibernate 3.0 Spring, Oracle 9i, Web Sphere 5, JBOSS 3/4 X

Confidential, St. Louis, MO

IT Analyst /Programmer

Responsibilities:

  • Worked both independently and in a team-oriented collaborative environment.
  • Worked with Microsoft SQL server.
  • Documented and provided status of project and technical information related to the application/software supported (Web2Py & .NET).
  • Supported remote users at their home office, hotel, or customer site utilizing remote tools and troubleshooting over phone using VPN.
  • Knowledge of DSL/Cable Modem Routers, Windows Server 2012 and 2008, CISCO Switches and Routers and VOIP (i.e. Cisco Call Manager).

Environment:: Environment: Java, JDBC, Tomcat, JSP, Servlets, Oracle7, PL/SQL, HTML, DHTML, JavaScript, ASP, UML and XML, Rational Robot, Oracle, Java Servlets, XML, XPATH,JBOSS.

We'd love your feedback!