We provide IT Staff Augmentation Services!

Big Data Solution Architect Resume

3.00/5 (Submit Your Rating)

PittsburgH

SUMMARY:

  • Highly accomplished Big Data Solution Architect with over 13 years of IT experience and over 3 years of experience in Big Data Architecture, Design and implementation. Adept at leading cross - functional teams to develop solutions to complex business problems through innovation and technology.
  • Hands-on experience in designing and building scalable, distributed enterprise applications using Big Data technologies like Hadoop, MapReduce, HDFS, MAPR-FS, MAPR-DB,YARN, Pig, Hive, Drill, Sqoop, HBase, Elastic Search, Flume, Oozie, Zookeeper, Apache Spark, Kafka, Spark SQL, Control-M.
  • Successfully Designed and implemented three key Big Data projects for different Clients using Big Data technologies like Hadoop, HDFS, MAPR-FS, Map Reduce, YARN, Flume, HBase, Hive, Elastic Search, Flume, Sqoop, Oozie, Spark, Kafka.
  • Led several POCs to develop Analytics workflows for retail clients using Apache Spark, Scala, SparkSQL, MAPR-FS, MAPR-DB, Drill, Elastic Search, Hive, Kafka, Spring Framework, PostgreSQL, Web Focus etc.
  • Proven track record of end-to-end implementation of top notch data management solutions leveraging both Big Data (Hadoop Ecosystem) and traditional data warehousing technologies.

TECHNICAL SKILLS:

Big Data: Hadoop, HDFS, MapReduce, Java, Scala, MAPR-FS, MAPR-DB, Drill, Elastic search, Hive, Pig, HBase, Flume, Sqoop, Oozie, Control-M, Zookeeper, Impala, Cloudera, Hue, YARN, Apache Spark, RDD, Data Frames, Spark Streaming, Spark SQL, MLlib, Apache Kafka, NoSQL, Cassandra,MDM,SBT, Maven, Eclipse, Spring Framework.

Data Warehouse: Teradata, Web Focus, Microsoft SSIS, MS SQL Server, Informatica, ETL, MDM,Cognos, QlikView

PROFESSIONAL EXPERIENCE:

Confidential, Pittsburgh

Big Data Solution Architect

Responsibilities:

  • Designed and developed components of the product for Data ingestion, transformation, MDM(Master Data Management), analysis and visualization.
  • Designed and developed algorithm to perform ETL workloads to pull data from MAPR-DB and push on to Hive for warehouse aggregation and reporting.
  • Developed ETL jobs using Scala and Sqoop and used Control-M for scheduling the work flow.
  • Developed Hive warehouse to store the data after the aggregation are done on daily, weekly, monthly yearly basis. Web focus used for reporting and Dashboards. Drill and Hue used for querying data.
  • Developing and launching services via Rancher Catalog.
  • Built a Cleansing-UI application with Spring Java frame work, where Business users can cleanse the data from MAPR-DB as part of MDM process.
  • Created Elastic Search Indices to facilitate easy data retrieval from MAPR-DB. This was useful for business users and Data Scientist Team.
  • Designed and developed real-time data synchronization jobs using Spark streaming.
  • Built automation scripts to build and deploy code to Hadoop through Jenkins using Shell and Java as part of CICD.
  • Worked in Agile/Scrum methodology and used Atlassian tools like Jira, bitbucket, bamboo.

Tools: MAPR-FS, MAPR-DB, Hive, Drill, Scala, Spark 2.0, SparkSQL, Elastic Search, Zookeeper, HUE, Sqoop, Control-M, PostgreSQL, Java, Linux, Web Focus, JIRA, Bitbucket, Bamboo

Confidential

Big Data Architect

Responsibilities:

  • Led the initial Proof of concept for the system and played key role for getting stakeholder’s.
  • Worked with the customer and end users to define application and technical requirements.
  • Solution Architecture and end to end platform implementation as per Enterprise standards.
  • Formulated and optimized Data Ingestion, Storage, Analysis and Visualization.
  • Performance tuning, troubleshooting, Team building and Training.
  • Technically contributed to several project activities like creation of Design Documents, MapReduce Jobs, Hive tables, Sqoop injection, Oozie work flows, Hive scripts etc.

Tools: HDFS, Hadoop, MapReduce, Hbase, Hive, Pig, Sqoop, Oozie, Zookeeper, Impala, Cloudera, HUE, SQL Server, Oracle, DB2, Java, Linux, Cognos, QlikView.

Big Data Delivery Lead

Confidential

Responsibilities:

  • Collaborated to provide technical expertise and lead team for proof of concept development which played vital role in obtaining stakeholder’s buy-in for the project.
  • Technically contributed to several project activities and helped team in resolving Technical issues.
  • Led full cycle big data project management from inception including defining problem statement, scope, planning schedule, risk mitigation and all deliverables.

Tools: HDFS, MapReduce, Pig, Hive, Hbase, Java, Linux, XML, Flume, Oozie, Sqoop, Teradata, HUE, Impala, Cloudera CDH, Eclipse, Mainframes, DB2, Teradata, Cognos and QlikView

IT Delivery Lead

Confidential

Responsibilities:

  • Successfully Designed, managed and implemented key Data warehouse projects like LRRS Reporting, Premium Warehouse, CSR Projects for Travelers through Onsite-Offshore Model.
  • Successfully transitioned engagement from staff augmentation model to project-collaboration model with the final transition effectively moving to managed-service model. Received accolades from clients.
  • Adept in SPRINT planning, product backlog management, agile estimation, and daily Scrum metrics such as Burn-down and Velocity charts, Sprint summary report, and Sprint review and retro meetings.

Tools: CSR Loss, CSR Premium, CSPT, Wind pool, WARP, WCOMP, LRRS, PICS POCS, DMV

Onsite Tech Lead

Confidential

Team Lead

Confidential

Team Member

Confidential

We'd love your feedback!