Sr. Gcpdevops Engineer Resume
3.00/5 (Submit Your Rating)
Austin, TX
SUMMARY
- Having 8+ years of experience asGCP/DevOps/Java/J2EE technologies and frameworks.
- Experience and expertise in Google Pub/sub, Cloud Functions, Compute Engine, Storage, Sinks.
- Full understanding and experience of SDLC, Agile Methodologies and process.
- Experience of working in CI/CD environment to merge the developed code with Master branch (CI) and code is ready to ship to next level (CD).
- Hands on experience on Google Cloud Platform (GCP) in all the bigdata products BigQuery, CloudDataProc, Google Cloud Storage, Composer (Air Flow as a service).
- Hands of experience inGCP, Big Query, GCS bucket, G - cloud function, cloud dataflow, Pub/suB cloud shell, GSUTIL, BQ command line utilities,DataProc, Stack driver
- Well noledge and experience in Cloudera ecosystem (hdfs, yarn, hive, sqoop, flume, hbase, oozie, kafka, pig),datapipeline,dataanalysis and processing with hive, impala, spark, spark sql.
- Experience in building theDevOpsPlatform with high end scripting languages of Python.
- Experience in creating new Team Projects with security in AzureDevops, TFS and configured Areas, Iterations, builds and releases.
- Effectively implementsDevOpsculture to improve time to market, minimize production failures, improve efficiency, and gain a competitive advantage over the competition.
- Experience in building integration process for migratingdatafrom Oracle platforms, SQL Server platforms and even excel spreadsheets via DataStage.
- Experience on DevOps tools such as Chef, Docker, Puppet, Jenkins, Maven, ANT, SVN, Vagrant, and Virtual Box.
- Experience in using Flume, Kafka and Spark streaming to ingest real time or near real timedatain HDFS.
- Experience in analyzingdataand provided insights with R Programming and Python Pandas.
- Experience in playing an important role in automating the deployments onGCPusing GitHub, Terraform, Jenkins and Chef.
- Extensive experience in setting up the CI/CD pipelines using Jenkins, Terraform andGCP.
- Experience in migratingGCPprojects from one organization to other organization
- Hands on experience on architecting the ETL transformation layers and writing spark jobs to do the processing.
- Experience in developing real time/batchdatastreaming systems, using the latest technologies.
- Expert in monitoring servers using Nagios,Datadog, Cloud watch and using EFK Stack Elasticsearch Fluentd Kibana.
- Has good Programming experience with Python and Scala.
- Hands in experience on No SQL database like Hbase, Cassandra.
- Analyzing the way to migrate oracle database to redshift.
- Experience with scripting languages like PowerShell, Perl, Shell, etc.
- Expert noledge and experience in fact dimensional modeling (Star schema, Snow flake schema), transactional modeling and SCD (Slowly changing dimension)
- Extensive experience in writing MS SQL, T-SQL procedures, ORACLE TOAD functions and queries.
- Expert in achieving oracle SQL plan stability, maintaining baselines with SQL plans, ASH, AWR, ADDM, Sql Advisor for pro-active follow up and SQL rewrites.
- Experience on Shell scripting to automate various activities.
- Experience in Application development with oracle forms and report with OBIEE, discoverer, report builder and ETL development.
- Effective team member, collaborative and comfortable working independently.
PROFESSIONAL EXPERIENCE
Sr. GCPDevops Engineer
Confidential - Austin, TX
RESPONSIBILITIES:
- Knowledge in building and architecting multipleDatapipelines, end to end ETL and ELT process forDataingestion and transformation inGCPand coordinate task among the team.
- Automated Cloud Operations onGCP(Project Creation and Deletion, Provisioning Compute Engines, Locking Services inGCPproject) using Python Client Libraries, REST API, GIT and Jenkins.
- Implemented a CI/CD pipeline using AzureDevOps(VSTS, TFS) in both cloud and on-premises.
- DevOpsrole converting existing AWS infrastructure to Server-less architecture (AWS Lambda, Kinesis) deployed via CloudFormation
- Drive the process in whichDevOpsteam operates and iterates/releases.
- Involved in DevOps migration/automation processes for build and deploy systems.
- Developed multi cloud strategies in better usingGCP(for its PAAS) and Azure (for its SAAS).
- Has Knowledge in designing and deployment of Hadoop cluster and different BigDataanalytic tools including Hive, Sqoop, Apache Spark with Cloudera Distribution.
- Involved in loading and transforming large sets of the structured, semi-structured dataset and analyzed them by running Hive queries.
- Used Stackdriver logging for monitoringGCPcloud resources and the applications dat deployed onGCPby creating new alarm, enable notification service.
- Developed custom python program including CI/CD rules for google clouddatacatalog for metadata management.
- Builddatapipelines in airflow inGCPfor ETL related jobs using different airflow operators both old and newer operators.
- UsedGCPCloud Deployment Manager to automate configuration, deployment, scaling and monitoring of systems.
- Design and architect various layer ofDatalake, Design star schema in Big Query.
- Loading salesforceDataevery 15 min on incremental basis to BIGQUERY raw and UDM layer using SQL, Google DataProc, GCS bucket, HIVE, Spark, Scala, Python, Gsutil And Shell Script.
- Using rest API with Python to ingestDatafrom and some other site to BIGQUERY.
- Build a program with Python and apache beam and execute it in cloud Dataflow to runDatavalidation between raw source file and Bigquery tables.
- Building a Scala and spark based configurable framework to connect commonDatasources like MYSQL, Oracle, Postgres, SQL Server, Salesforce, Bigquery and load it in Bigquery.
- Monitoring Bigquery, Dataproc and cloudDataflow jobs via Stackdriver for all the environment.
- Open SSH tunnel to Google DataProc to access to yarn manager to monitor spark jobs.
- Design, build and manage the ELK (Elasticsearch, Logstash, and Kibana) cluster for centralized logging and search functionalities.
- Submit spark jobs using gsutil and spark submission get it executed in Dataproc cluster
- Write a Python program to maintain raw file archival in GCS bucket.
- Analyze various type of raw file like Json, Csv, Xml with Python using Pandas, Numpy etc.
- Write Scala program for spark transformation in Dataproc.
Environment: Cloud Dataflow, Cloud Shell, Vm Instances, Cloud SQL, MySQL, Posgres, SQL Server, Salesforce SQL, Python, Scala, Spark, Hive.
GCPDevopsEngineer
Confidential - Highland Heights, KY
RESPONSIBILITIES:
- Using g-cloud function with Python to loadDatain to Bigquery for on arrival csv files in GCS bucket.
- Write a program to download a SQL Dump from there equipment maintenance site and then load it in GCS bucket. On the other side load this SQL dump from GCS bucket to MYSQL (hosted in Google cloud SQL) and load theDatafrom MYSQL to Bigquery using Python, Scala, spark and Dataproc.
- Responsible for implementingDevOpsprocesses usingAzureCloudandAzureDevOpsfor Digital Marketing Team.
- Created WebApps (PaaS solutions) on Azure portal for deploying web applications using AzureDevOpspipelines.
- As a member of onboarding team, my responsibility is to implementDevOpstransformation by working with Agile teams to migrate applications to Azure platform.
- Worked on google cloud platform (GCP) services like compute engine, cloud load balancing, cloud storage, cloud SQL, stack driver monitoring, and cloud deployment manager.
- Mentored to Dev team for setting up infrastructure using PostgreSQL for troubleshooting and Orchestrated using Mesos on Google Cloud Platform (GCP).
- Deliver OpenText EIM products as an EMS on a secure, globally scaled platform from Google.
- Optimized OpenText performance with Kubernetes container management.
- Coordinated with the development team as a backup, and resolved issues based on configuration management tool like Chef.
- Worked with Docker containers and at least one cluster management software - Mesos, Kubernetes, OpenShift
- Hands on noledge in using Google cloud platform for bigquery, cloud dataproc and apache airflow services.
- Got involved in migrating on prem Hadoop system to usingGCP(Google Cloud Platform).
- Responsible for designing developing application frameworks using Java Programming language, Java Script & related approved software technologies
- Process and load bound and unboundDatafrom Google pub/sub topic to Bigquery using cloud Dataflow with Python.
- Wrote scripts in Hive SQL/Presto SQL, using python plugin for both spark and presto for creating complex tables with high performance metrics like partitioning, clustering and skewing.
- Migrated previously written cron jobs to airflow/composer inGCP.
- Has noledge in using stackdriver service/ dataproc clusters inGCPfor accessing logs for debugging.
- Has written python DAGs in airflow which orchestrate end to enddatapipelines for multiple applications.
- Knowledge in using python API for spark and python pandas package for ETL jobs
- Create firewall rules to access GoogleDataproc from other machines.
- Write Scala program for spark transformation in Dataproc.
Environment: GCP, GCS Bucket, G-Cloud Function, Apache Beam, Gsutil, Vm Instances, Cloud SQL, MySQL, SQL Server, Python, Scala, Spark, Spark-SQL.
Software Engineer
Confidential - Los Angeles, CA
RESPONSIBILITIES:
- Involved in business requirements gathering and preparing architecture design documents.
- Involved in preparing Technical Design Documents for the ETL development.
- Extensively used DataStage Designer to develop various Parallel jobs to extract, cleanse, transform, integrate and loaddatainto EDW and then toDataMart.
- Worked with DataStage Designer to import/export jobs, metadata, DataStage components between the projects.
- Worked briefly in writing pyspark core methods for speeding up Hive SQL queries like non-equi joins.
- Was involved in setting up of apache airflow service inGCP.
- Builddatapipelines in airflow inGCPfor ETL related jobs using different airflow operators.
- Used apache Sqoop import and export and handled datatypes after moving.
- Extracted, compiled, and trackeddataand analyzeddatato generate reports.
- Participated with troubleshooting, and problem resolution efforts.
- Performed dailydatamanipulation using SQL and prepared reports on weekly, monthly, and quarterly basis.
- Used Excel functions to generate spreadsheets and pivot tables.
- Identified, analyze, and interpret trends or patterns in complexdatasets.
- Cleansing and Blending multipledatasources to allow for different views on applicationdatain a single dashboard.
- Performdatamanipulation operations like import/exportdatafrom various external file formats using SQL Generated variety of business reports from SQL server using excel with pivot table and pivot chart.Checked for invalid/out of rangedata.
- Develop and implement databases,datacollection systems,dataanalytics and other strategies dat optimize statistical efficiency and quality.
- Developed Job Sequences to run Staging load, EDW load,DataMart Load and to run store procedures.
- Extensively worked on Parallel Jobs using Various Stages like Sequential File, CFF Stage,Dataset, Lookup, Join, Aggregator, Remove Duplicate, Modify, Filter, Funnel, Copy, Surrogate Key Generator, Sort, Store
- Develop and implement databases,datacollection systems,dataanalytics and other strategies dat optimize statistical efficiency and quality.
- Acquiredatafrom primary or secondarydatasources and maintain databases/datasystems. Procedure Stage, Transformer, Change Capture, Oracle Connector/Oracle Enterprise, and ODBC Stage.
- Used Parameter Sets, Environment Variables, Stage Variables and Routines for developing Parameter Driven Jobs.
Environment: AWS, Python, Oracle, SQL, Shell Script, PySpark, MySQL, Cloudera, Agile.
