We provide IT Staff Augmentation Services!

Talend Developer Resume

5.00/5 (Submit Your Rating)

Dallas, Tx

SUMMARY:

  • Hortonworks certified hadoop developer with overall 5 years of IT experience in design, development of web applications using Java EE technologies including 3 years of experience working in hadoop ecosystem.
  • Rich exposure to Java EE and Hadoop stack - MapReduce, HDFS, Pig, Hive, Sqoop, HBase, Flume .
  • Good understanding of Hadoop architecture and have experience working on different hadoop distributions like MapR, HortonWorks .
  • Experience using Sqoop to import data into HDFS from RDBMS and vice-versa.
  • Experience in writing custom MapReduce programs & UDF's in Java to extend Hive and Pig core functionality.
  • Good experience in developing projects using TALEND studio for Big data.
  • Basic knowledge in Marklogic tools like mlcp for big data and SPARK.
  • Experience in dealing with structured and semi-structured data in HDFS.
  • Comprehensive knowledge and working experience with relational databases MySQL, Oracle and DB2.
  • Knowledge in UNIX shell scripting .
  • Experience in working/leading onsite/offshore model projects independently.
  • Proficient in development methodologies such as Agile, Waterfall .
  • Passion to excel in any assignment and have good debugging and problem solving skills.

TECHNICAL SKILLS:

Big data technologies: Hadoop, MapReduce, HDFS, Flume, Sqoop, Yarn, Pig, Hive

Programming Languages: Core Java, JSP, Servlets, JDBC, HTML 4.0, JavaScript, SQL, JSON, Unix shell scripting

Frameworks: Spring JDBC, Jquery 1.10, JUnit, Maven, Android

Tools: / IDE: Eclipse,Talend, IBM iRAD (Rational Application Developer), HP Performance Center, Rally(agile tool),Bed Rock 3.0,Sonar, AntHillPro, SVN

Servers: Tomcat, Oracle Web logic 10.3, IBM Web Sphere 7.0

Database: Oracle 10g, DB2 9.0, MYSQL

WORK EXPERIENCE:

Confidential, Dallas,TX

Talend Developer

Responsibilities:

  • Designed partial restartability of workflows and logging mechanism.
  • Ingested the data from various Databases into HDFS by SQOOP tool.
  • Responsible for Build and deployment of all the jobs from TAC(Talend Admin Console).
  • Experienced in coordinating and working with multiple teams for application development
  • Loaded data from various sources, pre/post-processing using Hive and created tables in our cluster.
  • Responsible for scheduling monitoring and troubleshooting workflows.
Confidential, Minneapolis, MN

Hadoop Developer

Responsibilities:

  • Designed jobs for ingesting and provisioning data to various customers using Talend Studio for BigData in second version of project.
  • Ingested the data from various systems like Claims, Mail Service, Pharmacies, Providers into HDFS.
  • Developed workflows using Talend for data pre-processing.
  • Designed the work flows using Bed Rock tool in first version of project.
  • Used Avro serialization technique to serialize data for handling schema evolution .
  • Created hive external tables on the avro data and data is loaded in hive partitions .
  • Transformed existing PL/SQL procedures to hive queries.
  • Developed UDF’s to handle common functionalities across various procedures.
  • Worked with different file format and compression techniques to ensure optimal performance of hive queries.
  • Good knowledge on DataLake implementation including source registration, ingestion and snapshot process with familiarity in Talend .
  • Published hive views above the hive base tables as per the business logic for the downstream systems to pull data from Hadoop.
  • Developed clean up shell scripts for purging old partitions as per the business requirement.
  • Developed automated workflows for monitoring the landing zone for the files and ingestion into HDFS in Bedrock Tool and Talend .
  • Used HBase for storing the Meta data of files and maintaining the file patterns .
  • One year exposure on Data Lake implementation, documented, Logged and resolved defects in the roll out phase.
  • Participate in code review, test case review, scrum calls, sprint retro.

Technology/Tools used : MapR Hadoop, Map Reduce, Hive, Pig, BedRock, HBase, Maven, Sonar, Eclipse

Confidential, Detroit, MI

Hadoop Developer

Responsibilities:

  • Understanding of the BRD and preparing high level design document.
  • Intake of the fixed width files from landing zone to HDFS.
  • Schema Validation of the input files with the Meta file received.
  • Rejection of any duplicate files based on the MD5 check sum.
  • Maintaining of metadata in the HBase table.
  • Develop mapreduce job for avro conversion and load the avro data to hive table using the SerDe’s.
  • Used BeanIO library to map the data from hive managed table according to the input XML and provision the file to the downstream for further analysis.
  • Scheduling the hadoop jobs using Oozie .
  • Collected metrics of various mapreduce jobs like elapsed time, number of mappers, reducers, I/O, Shuffle time, spilled records etc and ensured optimal performance of the queries.
  • Mavenized the code and ensured proper builds management, release management are practiced.
  • Done POC’s on loading semi structured data like JSON, XML to hive tables directly.
  • Explored on using Talend Open Studio for big data as development platform for hadoop jobs.
  • Technology/Tools used: Hadoop, Hive, Map Reduce, Oozie, Java, Eclipse, Oracle 11g, SQL, SVN, Maven
Confidential, Memphis

Java Developer

Responsibilities:

  • Involved in various phases of Software Development Life Cycle (SDLC) of the application like Requirement gathering, validating SRS, DDS, Design, Analysis and Code development, bug fixing, testing and 24/7 production support.
  • Worked on bug fixing and enhancements on change requests.
  • Analyzed and responded well on all the critical issues.
  • Provided on call support for both level 2 and level 3.
  • Daily performed server and application health check up.
  • Written Unit Test cases and performed unit testing, regression testing.
  • Worked in parallel with clients in few enhancements and proposed optimized code solutions.

Technology/Tools used: JAVA/J2EE, Weblogic, Oracle 10g, JAVA Script, $Universe

We'd love your feedback!