Software Engineer Resume
SUMMARY:
- I am Technical professional with 6.2 years of varied and extensive experience in Enterprise Data Warehousing, CIDR repository, cloud oracle ERP HRMS and currently, working on Big data analytics technology in that, I has to create Big data lake from scratch.
- Key Asset is working on Rack space cloud servers in clustered mode using centos6.5 flavour. And SDLC(Agile Scrum)
- Worked on Big data analytics, A Large distributed system to manage data.one of keen project is CIDR (central identities data repository) repository of Government for Unique identification authority of India.
- Currently, working on Hadoop - based solutions, or relational databases, or Greenplum systems
- Database administration of any RDBMS is my strong asset.
- Intermediate programming and computation skills, Java, BASH/Shell Scripting.
- Currently, Working on Linux System flavor centos6.5 as a Rackspace commodity server. Recently working on java spring-xd distributed runtime architecture.
- Worked on traditional DW/BI components (ETL, Staging, DW, and BI Tools) Using PHD3.0 Hadoop distribution.
- Knowledge of parallel processing approaches and concepts. Also, Improved Understanding on teh hardware and networking architecture behind large scale clustered systems
- When I learn something new, my immediate goal is to document it and share it. I enjoy mentoring others and empowering people around me to become experts.
- Acadmic’s:AIEEE (TOP -200 -SR)Rank Holder, Advanced GERMAN LANGUAGE done
- Courses completed during Pivotal training:
TECHNICAL SKILLS
Deployment and scheduling Environment: Windows server 2012, Centos6.5.
Data Integration tool: Talend Big Data suite, Talend Data prepration.
Operating Systems: RHEL, CentOS6.5, Ubuntu,Windows server 2012 R2,Windows server 2008 R2.
Monitoring: Cron,Nagios,Ganglia,Basic knowledge of Newrelic.
Database: MySQL,Oracle, Pivotal HDB(HAWQ),Greenplum (GPDB),sql server 2012.
IN-memory: Redis
Data Ingestion tool: Spring XD,Spring XD Distributed mode(DIRT)
File system: Hadoop (Pivotal distribution PHD), FTP Streaming for EDI system.
Security: iptables.
Hadoop Cluster Monitoring and Administration: Apache Ambari.
Data Management tool: Workbench developer force for Salesforce, Aginity Pivotal Greenplum Workbench,SSMS for sql server 2012,RealForce Explorer,Salesforce Sandbox Environment.
Cloud Storage: Rackspace
Data Api: Syspro e.NET
BI Tool: Excel2013, Tableau.
PROFESSIONAL EXPERIENCE
Confidential
Software engineer
Responsibilities:
- Currently, I am working on job related to Salesforce Sandbox environment. In that, source system is SFDC AND Target system is SYSPRO (ERP).
- Worked on talend job Using SQL server 2012 and Pivotal Hadoop distribution(that is PHD3.0)
- Worked on Talend Job Using source sfdc and SQL server 2012(vice versa).
- Worked on Talend Job Using Sql server 2012 and Greenplum(vice versa).
- Worked on Talend Job Using Sql server 2012 and Hawq (vice versa).
- Following Talend job created using source-target system.
- In all task common part is Data Mapping using talend, we are frequently doing Data mapping between source and target system using tMap component along with Data Transformation.
- Expertise with processing structured (CSV) and Processing XML files using XSLT stylesheet and transforming them to before loading to SFDC databases.
- Retrieving and Loading data from Greenplum and Hawq Database.
- Working knowledge of UNIX, SQL, and Windows Environment.
- Working Knowledge of Salesforce Data Management using Workbench sandbox environment.
- Working Knowledge of Data Transformations & Mappings.
- Working Knowledge of Data Migration - Retrieval from Source and Load to target.
- Working Knowledge of Data Profiling & Quality and Data Cleansing.
- Retrieving and Integrating data from following web services:
Salesforce CRM
Greenplum cluster Role/skills flavor
Responsibilities:
- Expertise in Greenplum Administration using Aginity Workbench and, PGadminIII.
- Expertise in Implementing high availability Greenplum using master/segment mirroring
- Good Experience in Installing, Configuring & Administering teh Greenplum Clusters.
- Ability to provide root cause analysis to teh customers
- Ability to recover Greenplum DB from Hardware & DB issues
- Experience of Greenplum SQL. Creating External table and core table for data pipeline AT staging area is my daily basis task.
Hadoop
Confidential
Responsibilities:
- Set up production Hadoop clusters with optimum configurations
- Drive automation of Hadoop deployments, cluster maintenance operations
- Manage Hadoop cluster, monitoring alerts and notification
- Deployment of upgrades, updates and patches
- Provide 24x7 tier-3 troubleshooting and break-fix support for production services
- Diagnosis of installation & configuration issues
- Diagnosis of cluster management issues
- Job scheduling, monitoring, debugging and troubleshooting
- Monitoring and management of teh cluster in all respects, notably availability, performance and security
- File system management and monitoring
- Data transfer between Hadoop and other data stores ( RDBMS,Salesforce cloud service)
- Ingesting additional data sources into Hadoop in either streaming or batch mode
- Configuring teh cluster to be rack-awareness
- Set up High Availability/Disaster Recovery environment
- Debug/Troubleshoot environment failures/downtime
- Manage and review Hadoop log files
