Sr. Big Data Developer Resume
Burlington, NJ
SUMMARY:
- Over 10 years of experience in Information Technology in various business domains
- Working experience with large scale Hadoop environments build and support including design, configuration, installation, performance tuning and monitoring
- Hands - on experience in Big Data and HADOOP Ecosystem components. (HDFS, Map Reduce, PIG, HIVE, HBASE, SQOOP, Flume, Oozie, Zookeeper, Ambari, Ranger,Hue, Kafka, Storm, MongoDB, R language)
- Experience in Adding, Configuring and Deploying New services in Hadoop Eco-System.
- Deep understanding of schedulers, workload management, availability, scalability and distributed data platforms.
- Expertise in Datalake process on Acquiring, preprocessing & ingesting different data formats from Clients.
- Experience interacting with business users, understanding requirements and provide quick solution.
- Experience building a commercially viable scalable solution using Big Insights Horton works and Cloudera distributions of Hadoop, HIVE and HBase, Hue, YARN, SPARK, sentry
- Proficient in Map-Reduce Java and streaming APIs, Ruby, Python
- Experience with Hadoop & Big Data architecture
- Configured and monitored Hadoop clusters with BigInsights Enterprise distribution(3.0/4.0)
- Configured and monitored Hadoop clusters with Cloudera Enterprise distribution(CDH4)
- Experience with AWS (Amazon Web Services) and MapR
- Experience working with Big Data Computing Platform (Spark, Map Reduce) and Big Data Manipulation Tools (Hive, and Pig).
- Experience with Spark, Spark Streaming and Spark SQL
- Deep understanding of schedulers, workload management, availability, scalability and distributed data platforms.
- Expert understanding of ETL principles and how to apply them within Hadoop
- Experience with Messaging and collection frameworks like Flume, Kafka
- Experience with Oozie Workflow Engine to automate and parallelize Hadoop Map/Reduce, Hive and Pig jobs.
- Experience with IBM BigSQL (Massively Parallel Processing) IBM STREAM
- Experienced in developing application using the R, Big R programming language.
- Experience with R and statistical modeling with RStudio.
- Experience with Source control and Configuration Management tools and technologies (Git/GitHub)
- Experience with developing and driving solution architectures employing portal,SOA, data management, data conversion and BI strategies, architecture and design p Confidential erns
- Experience building business applications using RDBMS such as Oracle, DB2, MS SQL Server and MySQL
- Importing/Exporting data from MYSQL/Teradata/Netezza/DB2 /Oracle to HDFS (Hadoop) using sqoop
- Good understanding of performance tuning with both NoSQL and SQL technologies.
- Good understanding of file formats including JSON, Parquet, Avro, and others.
- Extensive experience building and designing large-scale distributed applications
- Experience with agile/scrum methodologies to iterate quickly on product changes, developing user stories and working through backlogs
- Experience with Talend RTX ETL tool, devalop jobs and scheduled jobs in Talend integration suite
- Experience with ETL development with Sync sort/NiFi/Talend/TAC
PROFESSIONAL EXPERIENCE:
Confidential, Burlington, NJ
Sr. Big Data Developer
Responsibilities:
- Developed Sqoop Framework to Source Historical Data from Oracle, DB2, Sql Server, Oracle
- Ingest Flat files received via ECG FTP tool and files received from Sqoop into UHG Data Lake Hive and HBase using Data Fabric functionalities.
- Validate Transactional Data Files coming from IBM Confidential (Change Data Capture) tool
- Used Splunk Dashboard to record and monitor incoming file frequency from Confidential tool
- Developed and automated jobs in Talend open studio to validate the Ingested Data.
- Performed address standardization per business requirement and captured Golden Records.
- Ingested XML files capture from RabbitMQ and Stored in HBase tables.
- Developed HBase table for Monitor Ingestion Logs and snapshot logs.
- Loaded the data to HBASE using bulk load and HBASE API.
- Performing analytics using Talend Spark components on Insurance Claims Data.
- Developed Data transformation module employing hive, Map reduce.
- Involved in new development as well as bug fixing, performance tuning in the existing application.
- Wrote Data transformation script using hive, Map reduce (Python)
- Scheduled backups jobs by implementing Talend Tac Scheduler
- Data Integrated, Extracted and transformed from mainframe and loaded to Hadoop using Talend
- Configured BI tools to access hadoop cluster from windows.
- Designed/Developed framework/api to leverage platform capabilities using map reduce /HDFS
Environment:: MapR, Hive, Hbase, Sqoop, Pig, Talend TAC, Map Reduce, Splunk, Eclipse, Maven,, Shell Script, JSP, Json, JDBC, XML, MYBATIS,UNIX, LINUX, DB2,Oracle,TERADATA, SVN,JBOSS, MYECLIPSE, JIRA, Subversion, Jenkins, Codehub, Autosys, Talend/TAC.
Confidential, NY
Sr. Big Data Developer
Responsibilities:
- Developed Data transformation module employing hive, Map reduce.
- Geospatial modeling, scripting and geostatistical application development for cloud based solution utilizing Map Reduce Hadoop
- Setting Up ESRI Spatial Frame Work for Hadoop
- Adding configuring ESRI jar to execute geospatial quires in hive and beeline.
- Involved in new development as well as bug fixing, performance tuning in the existing application
- Importing/Exporting data from MYSQL /Oracle to Azure cluster and AWS cluster using sqoop.
- Worked on Hive Jason Serde to parse spatial / Enclosed JSON, Unenclosed JSON Data format.
- Analyzing data using Hive & Pig Latin Scripting
- Involved in new development as well as bug fixing, performance tuning in the existing application.
- Wrote Data transformation script using hive, Map reduce (Python)
- Scheduled backups jobs by implementing cron job
- Data Integrated, Extracted and transformed from mainframe and loaded to Hadoop using (HDF) Nifi
- Configured BI tools to access hadoop cluster from windows.
- Designed/Developed framework/api to leverage platform capabilities using map reduce /HDFS
Environment:: AWS (Amazon Web Services), MICROSOFT AZURE, Hive, Sqoop, crontab, Map Reduce, Eclipse, Maven,, Shell Script, JSP, Json, JDBC, XML, MYBATIS,UNIX, LINUX, DB2,TERADATA, SVN,JBOSS, MYECLIPSE, JIRA, Subversion, Control M/Autosys. NIFI (HDF).
Confidential, Chicago, IL
Sr. Big Data Developer
Responsibilities:
- Configured and monitored Hadoop clusters with Big Insights Enterprise distribution (3.0/4.0)
- Configured and monitored Hadoop clusters with Cloudera Enterprise distribution(CDH4/CDH5)
- Review Big Data Architect/Design/Release to maintain platform integrity
- Installing and Configuring SysnSort
- Developed Data transformation module employing hive, Map reduce
- Developed User Defined Function (UDF) for hive.
- Loaded the data to HBASE using bulk load and HBASE API.
- Developing APIs to remotely access the hadoop cluster
- Configuring BI tools to access hadoop cluster from windows desktops
- Architect and Implement Hadoop eco system component such as HDFS, Map Reduce, HBase, Zookeeper, Pig, Hadoop streaming, Sqoop, Oozie, hive, hive server 2.
- Designed applications for storing data to HDFS by using Kafka to get more performance.
- Involved in new development as well as bug fixing, performance tuning in the existing application
- Mentor other team members with the development of hive/udf/Map Reduce scripts
- Designed and developed the presentation and web layers based on Java, J2EE
- Developed Restful Web service using Jersey and JBOSS
- Demonstrated and strong working knowledge of Global Delivery Models comprising onsite and offshore staffing model.
- Data Integrated, Extract, transform from mainframe and loaded to Hadoop using Syncsort,
- Talend open studio.
- Assisted with data capacity planning and node forecasting.
- Importing/exporting heterogeneous data(Ebcdic,Ascii..) using sqoop (native/fast connector ), custom api (map reduce sftp push & pull) and flume
- Involved in new development as well as bug fixing, performance tuning in the existing application
- Proactively recognizing potential performance improvements and mitigating potential issues.
- Working Closely with IBM (Big Insights) team for future updates and recommendations.
Environment: Java/J2EE, Hadoop, Hbase, Hive, Map Reduce, Eclipse, Maven,, JavaScript, JSP, Python, HTML, JDBC, XML, MYBATIS,UNIX, LINUX, DB2,TERADATA, SVN,JBOSS, MYECLIPSE, JIRA, Subversion, Control M/Autosys. Sync sort. Talend . R Big R.
Confidential, NC
Hadoop Developer
Responsibilities:
- Developed Data transformation module employing hive, Map reduce
- Developed User Defined Function(UDF) for hive (Java)
- Loaded the data to HBASE using bulk load and HBASE API.
- Accessed the data from HBASE using Hbase Rest interface
- Mentored other team members with the development of hive/udf/Map Reduce scripts
- Designed and developed the presentation and web layers based on Java, J2EE
- Developed Restful Web service using Jersey and JBOSS
- Demonstrated and strong working knowledge of Global Delivery Models comprising onsite and offshore staffing model.
- Configure and monitor Hadoop clusters with Cloudera Enterprise distribution(CDH4)
Environment: Java/J2EE, HAdoop, HBAase, Hive, Map Reduce, Eclipse, Maven,, JavaScript, JSP, HTML, JDBC, XML, MYBATIS, MySQL, DB2,TERADATA, SVN,JBOSS, MYECLIPSE, JIRA, Subversion, Control M/Autosys
Confidential, Pleasanton, CA
Hadoop Developer
Responsibilities:
- Analyzing and reviewing business, functional and high-level technical requirements.
- Analytical and problem solving skills, applied to a “Big Data” environment (map-reduce,hive,hive-server2)
- Designing detailed technical components for complex applications utilizing high-level architecture, design p Confidential erns and reusable code.
- Supporting detailed project estimates and project work plans.
- Programming code independently for intermediate to complex modules following development standards.
- Planning and conducting code reviews for changes and enhancements that ensure standards compliance and systems interoperability.
- Creating test plans for module-level integration and tracking/resolving defects.
- Preparing and packaging production-ready code modules for staging.
- Ability to problem solve and reach resolution with a broader perspective in mind.
- Able to effectively communicate with peers, business analysts, project managers, quality control, and across other technology team boundaries
Environment: Java/J2EE, Hadoop, HBase, Hive, Map Reduce, Impala, Eclipsw, JUnit,Linux,Windows, JIRA, Subversion, Oracle, MS-SQL server.
Confidential, Detroit, MI
Software Developer
Responsibilities:
- Designed and developed the application architecture, use cases, and flowcharts using Microsoft Visio
- Designed and developed the presentation and web layers (Transactional application) based on Java, J2EE (STRUTS, spring, Hibernate, Velocity), web services (Axis2) using Eclipse 3.0, Visual Basic, ASP using Visual Studio.
- Worked with technologies such as JavaScript, XML, AJAX, HTML, and CSS.
- Developed Web services Architecture for supporting common business functions with direct access to web services from PL/SQL using Oracle9i/ AXIS 2
- Developed internal Tracking system to track marketing mediums, including organic and paid search, banner ads, referral links and email newsletters
- Developed and deployed applications in UNIX and Windows environments using ANT tool and shell script.
Environment: J2EE, AXIS2,AJAX, JavaScript, Struts, TILES, ANT, JSP, HTML, JDBC, XML, Hibernate 3.0 Spring, Oracle 9i, Web Sphere 5, JBOSS 3/4 X
Confidential, St. Louis, MO
IT Analyst /Programmer
Responsibilities:
- Worked both independently and in a team-oriented collaborative environment.
- Worked with Microsoft SQL server.
- Documented and provided status of project and technical information related to the application/software supported (Web2Py & .NET).
- Supported remote users at their home office, hotel, or customer site utilizing remote tools and troubleshooting over phone using VPN.
- Knowledge of DSL/Cable Modem Routers, Windows Server 2012 and 2008, CISCO Switches and Routers and VOIP (i.e. Cisco Call Manager).
Environment:: Environment: Java, JDBC, Tomcat, JSP, Servlets, Oracle7, PL/SQL, HTML, DHTML, JavaScript, ASP, UML and XML, Rational Robot, Oracle, Java Servlets, XML, XPATH,JBOSS.
