Application Development Resume
0/5 (Submit Your Rating)
SUMMARY:
- 8 years of strong experience in Application Development using Pyspark, Java, Python, Scala, and R & in depth understanding of Distributed Systems Architecture and Parallel Processing Frameworks.
- Strong understanding of Java Virtual Machines and multi - threading process.
- Experience in writing complex SQL queries, creating reports and dashboards.
- Proficient in using Unix based Command Line Interface, Expertise in handling ETL tools like Informatica.
- Strong experience using pyspark, HDFS, MapReduce, Hive, Pig, Spark, Sqoop, Oozie, and HBase.
- Deep knowledge of troubleshooting and tuning Spark applications and Hive scripts to achieve optimal performance.
- Experienced working with various Hadoop Distributions (Cloudera, Hortonworks, MapR, Amazon EMR) to fully implement and leverage new features.
- Experience in developing Spark Applications using Spark RDD, Spark-SQL and Data frame APIs.
- Worked with real-time data processing and streaming techniques using Spark streaming and Kafka.
- Experience in moving data into and out of the HDFS and Relational Database Systems (RDBMS) using
- Apache Sqoop.
- Expertise in working with HIVE data warehouse infrastructure-creating tables, data distribution by implementing Partitioning and Bucketing, developing, and tuning the HQL queries.
- Significant experience writing custom UDFs in Hive and custom Input Formats in MapReduce.
- Involved in creating Hive tables, loading with data, and writing Hive ad-hoc queries that will run internally in MapReduce and TEZ, replaced existing MR jobs and Hive scripts with Spark SQL & Spark data transformations for efficient data processing, Experience developing Kafka producers and Kafka
- Consumers for streaming millions of events per second on streaming data.
- Good knowledge in Database Creation and maintenance of physical data models with Oracle, Teradata,
- Netezza, DB2, MongoDB, HBase and SQL Server databases.
- Deep understanding of Ma
