Talend Developer Resume
Dallas, Tx
SUMMARY:
- Hortonworks certified hadoop developer with overall 5 years of IT experience in design, development of web applications using Java EE technologies including 3 years of experience working in hadoop ecosystem.
- Rich exposure to Java EE and Hadoop stack - MapReduce, HDFS, Pig, Hive, Sqoop, HBase, Flume .
- Good understanding of Hadoop architecture and have experience working on different hadoop distributions like MapR, HortonWorks .
- Experience using Sqoop to import data into HDFS from RDBMS and vice-versa.
- Experience in writing custom MapReduce programs & UDF's in Java to extend Hive and Pig core functionality.
- Good experience in developing projects using TALEND studio for Big data.
- Basic knowledge in Marklogic tools like mlcp for big data and SPARK.
- Experience in dealing with structured and semi-structured data in HDFS.
- Comprehensive knowledge and working experience with relational databases MySQL, Oracle and DB2.
- Knowledge in UNIX shell scripting .
- Experience in working/leading onsite/offshore model projects independently.
- Proficient in development methodologies such as Agile, Waterfall .
- Passion to excel in any assignment and have good debugging and problem solving skills.
TECHNICAL SKILLS:
Big data technologies: Hadoop, MapReduce, HDFS, Flume, Sqoop, Yarn, Pig, Hive
Programming Languages: Core Java, JSP, Servlets, JDBC, HTML 4.0, JavaScript, SQL, JSON, Unix shell scripting
Frameworks: Spring JDBC, Jquery 1.10, JUnit, Maven, Android
Tools: / IDE: Eclipse,Talend, IBM iRAD (Rational Application Developer), HP Performance Center, Rally(agile tool),Bed Rock 3.0,Sonar, AntHillPro, SVN
Servers: Tomcat, Oracle Web logic 10.3, IBM Web Sphere 7.0
Database: Oracle 10g, DB2 9.0, MYSQL
WORK EXPERIENCE:
Confidential, Dallas,TX
Talend Developer
Responsibilities:
- Designed partial restartability of workflows and logging mechanism.
- Ingested the data from various Databases into HDFS by SQOOP tool.
- Responsible for Build and deployment of all the jobs from TAC(Talend Admin Console).
- Experienced in coordinating and working with multiple teams for application development
- Loaded data from various sources, pre/post-processing using Hive and created tables in our cluster.
- Responsible for scheduling monitoring and troubleshooting workflows.
Hadoop Developer
Responsibilities:
- Designed jobs for ingesting and provisioning data to various customers using Talend Studio for BigData in second version of project.
- Ingested the data from various systems like Claims, Mail Service, Pharmacies, Providers into HDFS.
- Developed workflows using Talend for data pre-processing.
- Designed the work flows using Bed Rock tool in first version of project.
- Used Avro serialization technique to serialize data for handling schema evolution .
- Created hive external tables on the avro data and data is loaded in hive partitions .
- Transformed existing PL/SQL procedures to hive queries.
- Developed UDF’s to handle common functionalities across various procedures.
- Worked with different file format and compression techniques to ensure optimal performance of hive queries.
- Good knowledge on DataLake implementation including source registration, ingestion and snapshot process with familiarity in Talend .
- Published hive views above the hive base tables as per the business logic for the downstream systems to pull data from Hadoop.
- Developed clean up shell scripts for purging old partitions as per the business requirement.
- Developed automated workflows for monitoring the landing zone for the files and ingestion into HDFS in Bedrock Tool and Talend .
- Used HBase for storing the Meta data of files and maintaining the file patterns .
- One year exposure on Data Lake implementation, documented, Logged and resolved defects in the roll out phase.
- Participate in code review, test case review, scrum calls, sprint retro.
Technology/Tools used : MapR Hadoop, Map Reduce, Hive, Pig, BedRock, HBase, Maven, Sonar, Eclipse
Confidential, Detroit, MIHadoop Developer
Responsibilities:
- Understanding of the BRD and preparing high level design document.
- Intake of the fixed width files from landing zone to HDFS.
- Schema Validation of the input files with the Meta file received.
- Rejection of any duplicate files based on the MD5 check sum.
- Maintaining of metadata in the HBase table.
- Develop mapreduce job for avro conversion and load the avro data to hive table using the SerDe’s.
- Used BeanIO library to map the data from hive managed table according to the input XML and provision the file to the downstream for further analysis.
- Scheduling the hadoop jobs using Oozie .
- Collected metrics of various mapreduce jobs like elapsed time, number of mappers, reducers, I/O, Shuffle time, spilled records etc and ensured optimal performance of the queries.
- Mavenized the code and ensured proper builds management, release management are practiced.
- Done POC’s on loading semi structured data like JSON, XML to hive tables directly.
- Explored on using Talend Open Studio for big data as development platform for hadoop jobs.
- Technology/Tools used: Hadoop, Hive, Map Reduce, Oozie, Java, Eclipse, Oracle 11g, SQL, SVN, Maven
Java Developer
Responsibilities:
- Involved in various phases of Software Development Life Cycle (SDLC) of the application like Requirement gathering, validating SRS, DDS, Design, Analysis and Code development, bug fixing, testing and 24/7 production support.
- Worked on bug fixing and enhancements on change requests.
- Analyzed and responded well on all the critical issues.
- Provided on call support for both level 2 and level 3.
- Daily performed server and application health check up.
- Written Unit Test cases and performed unit testing, regression testing.
- Worked in parallel with clients in few enhancements and proposed optimized code solutions.
Technology/Tools used: JAVA/J2EE, Weblogic, Oracle 10g, JAVA Script, $Universe
