Big Data Developer Resume
Richmond, VA
SUMMARY
- 9+ years of IT experience in Software Development, with 2+ years of experience in Big DataHadoop in domains like Finance, Transportation and Logistics.
- Extensive experience in analyzing data using Hadoop Ecosystem including Hive, PIG, Sqoop and Map Reduce.
- Experience in developing custom Map Reduce Programs in Java using Apache Hadoopfor analyzing Big Data as per the requirement.
- In depth understanding and knowledge of Hadoop Architecture and various components such as HDFS, Job Tracker, Task Tracker, Name Node, Data Node and MapReduce concepts.
- Experience in extending HIVE and PIG core functionality by using custom UDF’s.
- Experience in importing and exporting the data using Sqoop from HDFS to Relational Database systems and vice - versa.
- Involved in business requirements gathering for successful implementation and POC (proof-of-concept) of Hadoop and its ecosystem
- Experience in managing and reviewing Hadoop Log files.
- Experience in installation, configuration, management and deployment of Big Data solutions and the underlying infrastructure of Hadoop Cluster.
- Experience in the integration of various data sources like Java, RDBMS, Shell Scripting, Spreadsheets and Text files.
- Extensive experience with SQL, PL/SQL and database concepts
- Ability to blend technical expertise with strong Conceptual and Analytical skills to provide quality solutions and result-oriented problem solving technique and leadership skills.
- Knowledge of Agile Software Development using Scrum process.
- Excellent leadership, interpersonal, problem solving and time management skills.
TECHNICAL SKILLS
Big data/Hadoop Ecosystem: HDFS, Map Reduce, HIVE, PIG, HBase, Sqoop, Oozie
Programming Languages: C, C++, Java, SQL, PL/SQL, Windows scripting.
NoSQL Databases: MongoDB, HBase
Database: Oracle 11g/10g, DB2, MS-SQL Server, MySQL.
Web Technologies: HTML, XML, JDBC, JavaScript.
Tools: Used: Eclipse, Putty.
Operating System: Ubuntu (Linux), Win 95/98/2000/XP, Mac OS, RedHat
ETL Tools: Job Runner.
Testing: Hadoop Testing, Hive Testing, Quality Center (QC), Unified Functional Test (UFT).
PROFESSIONAL EXPERIENCE
Confidential -Richmond, VA
Big Data Developer
RESPONSIBILITIES:
- Develop Map Reduce jobs for metrics.
- Develop PIG script / MR Jobs to structure the data.
- Connect the Reducer output for the dashboard to consume.
- Implement wrapper scripts if required.
- Designed and developed MapReduce programs.
- Involved in moving all log files generated from various sources to HDFS for further processing through Flume.
- Involved in loading data from UNIX file system to HDFS.
- Worked on Hue interface for querying the data.
- Created Hive tables to store the processed results in a tabular format.
- Involved in designing the Hive table partitioning (Yearly partitioning, Monthly partitioning).
- Designed and developed UDF for Hive Scripts to handle the data and business logic.
- Implemented test scripts to support test driven development and continuous integration.
- Responsible to manage data coming from different sources.
- Experienced on loading and transforming of large sets of structured and semi structured data.
- Exported the analyzed data to the relational databases using Sqoop for visualization and to generate reports for the BI team.
- Analyzed large amounts of data sets to determine optimal way to aggregate and report on it.
- Participate in requirement gathering and analysis phase of the project in documenting the business requirements by conducting workshops/meetings with various business users.
Environment: Cloudera CDH4, HDFS, Map Reduce, Hive 0.10, Pig 0.11, Sqoop, Java API.
Confidential - Jacksonville, FL
Hadoop Developer
RESPONSIBILITIES:
- Interact with business partners to understand their analytical needs.
- Conduct data analysis and perform feasibility study.
- Develop Map Reduce.
- Develop PIG scripts / MR Jobs to structure the unstructured data.
- Developed solutions to process data into HDFS (Hadoop Distributed File System), process within Hadoop and emit the summary results from Hadoop to downstream systems.
- Used Sqoop extensively to ingest data from various source systems into HDFS.
- Worked on different file formats like Text files, Sequence Files, Record columnar files (RC).
- Worked on a stand-alone as well as a distributed Hadoop application.
- Understood complex data structures of different type (structured, semi structured) and de-normalizing for storage in Hadoop.
Environment: Apache Hortonworks distribution 1.x, HDFS, Pig 0.10, Hive, MapReduce, Sqoop, Java Eclipse.
Confidential - Jacksonville, FL
Sr Developer
RESPONSIBILITIES:
- Automation framework implementation for QA team.
- Responsible for gathering business and functional requirements for the development and support of in-house and vendor developed applications
- Gathered and analyzed information for developing, supporting, and modifying existing web applications based on prioritized business needs
- Maintenance and framework support.
Environment: Java 1.6, Eclipse, Selenium.
Confidential
Sr Technical Associate
RESPONSIBILITIES:
- Develop Automation Scripts and scheduling.
- Generate Test Data as requested.
- Bug Logging / Tracking.
- Bug Regression.
- Enhance Automation scripts.
- Analysis and development of test data profile for each application.
- Identify and implement Data Management processes.
- Develop and maintain comprehensive data request process.
- Develop QTP automation scripts to generate Siebel orders and also COX legacy applications.
- Allocate data between the testing work streams in order to prevent conflict.
- Generate volumes of data as and when needed for performance and load testing.
Confidential
Software Engineer
RESPONSIBILITIES:
- Develop Automation Scripts and scheduling.
- Sanity Testing / Regression Testing.
- Bug Logging / Tracking.
- Bug Regression.
