Devops Engineer Resume
4.00/5 (Submit Your Rating)
SUMMARY
- Experienced IT professional working as a Systems/DevOps Engineer for the last 5 years.
- Proficient and knowledgeable in DevOps tools and architecture designs.
- Equipped with a combination of exceptional interpersonal soft skills in agile methodologies enhancing collaboration and transparency.
- Advanced technical knowledge in the realm of IT, automation, and cloud computing.
TECHNICAL SKILLS
- RHEL 7
- RHEL 6
- CentOS
- Ubuntu
- Dockers
- Kuberentes
- VMware
- Azure
- AWS
- GCP
- Git
- Bash
- Python
- Openshift
- Bamboo
- Jenkins
- CircleCI
- Microsoft
- Agile
- Ansible
- Puppet
- Terraform
- Cloudformation
- JSON
- YML
- XML
- Deployments and Releases
- Support
- Engineering
PROFESSIONAL EXPERIENCE
Confidential
DevOps Engineer
Responsibilities:
- Configuring, provisioning, and managing AWS services such as EC2, IAM, S3, EBS, EFS, CloudWatch, Auto - scaling groups, VPC, lambda, and cloudformation.
- Creating and managing IAM permission policies, users, groups, and roles to manage resource access between multiple teams and environments with best security practices.
- Setting up ELB for high availability, making sure all resources are utilized evenly throughout and appointing resources as needed to the EC2 instances.
- Creating and provisioning S3 buckets and assigning storage classes for various objects
- Managing S3 buckets storage class manually and through object life cycle policies for optimizing storage utilization and costs for the organization.
- Creating EC2 instances with EBS and EFS for specific tasks, in addition to EBS volumes, IAM roles, custom AMI’s, security groups for multiple applications.
- Configuring Route53 for delegated domains by mapping domains to endpoints through configuration of resource records for multiple applications and web servers.
- Monitored servers and application health through Cloudwatch and Cloudtrail
- Configuring security for the infrastructure by enabling guard duty, cognito, multi-factor authentication and inspector.
- Built infrastructure for applications using Terraform, including adjusting variables based on the requirements of the new machines and environments.
- Deployed and troubleshoot Terraform scripts to provision infrastructure on cloud platforms.
- Managed state buckets and configured release pipelines for Terraform using cloud native services.
- Deployed Ansible playbooks and ran ad-hoc for automation and configuration management.
- Creating Ansible playbooks for automating older processes, service configurations, patching, and applications installs and upgrades.
- Running Ansible playbooks for upgrading and deploying in-house applications with various customers, including performing pre and post checks.
- Creating Ansible playbooks for specific services and operations such as NFS file shares, web server applications such as HTTP, HTTPS etc.
- Managing GIT repositories for source code by pushing the code to various branches, creating tags for revisions, adding changes and committing new updates.
- GIT commands to clone the central repository to the local server, merging the local changes with the feature, dev or test branch, and creating pull requests.
- Managing, provisioning and configuring web services (Apache,HTTP,NGINX)
- Created custom image in Dockerfile for Docker containers, and uploaded custom images provided by lead architects to JFrog artifactory, as well as migrated onto cloud
- Managed Docker containers in production and deployed new containers using Docker Files, writing a Docker compose YAML file, and also helped in containerizing applications.
- Configuring and attaching volumes to docker containers to provide persistent data volume, in addition to troubleshooting containers through logs.
- Deployed and troubleshooted CICD pipelines through Jenkins and Bamboo for upgrading existing infrastructure code and applications.
- Experience in continuous integration and continuous deployment using Jenkins for managing infrastructure and applications in dev, qa, stage and production environments
- Deployed and troubleshooted Kubernetes clusters regarding pod failures,resource utilization and corrupt images.
- Building out Kubernetes clusters by configuring Kubernetes services, managing volume resources, updating namespaces, managing deployment and replicas.
- Involved in setting up JIRA as a defect tracking system and configured various workflows, customizations and plug-ins for the JIRA bug/issue tracker.
- Handing and solving tickets for different troubleshooting issues on JIRA, in addition to being involved in release planning and executing the release build request through JIRA.
- Working in Agile environments, communicating with the team members for planning and resolving an issue and for bug development tasks.
- Performing bi-weekly sprints with my teams, communicating through JIRA to resolve issues and debugging the task in a faster and less time consuming way.
Confidential
Linux Systems Analyst
Responsibilities:
- Experienced with configuring different flavors of linux (RHEL 6/7, CentOS, Ubuntu).
- Troubleshooting slow servers using system utility tools such as top, ps, iostat, reprioritizing a process, postponing and killing processes on production systems,
- Managed swap space storage on systems with memory overutilization, to avoid server hung issues and facing any out of memory errors.
- Maintaining the system’s load by prioritizing the processes values through cli or top.
- Maintaining linux services on an expert level ( DNS, HTTP, HTTPS, FTP, NFS, FIREWALL, SSH, DHCP, TFTP, NGINX, LVM, PXE, CORN)
- Creating and managing storage through the use of partitions (Primary, Extended, Logical) including using fdisk command line tool to configure and mount partitions.
- Managing dynamic storage on Linux systems using Logical Volume manager (physical volume, volume group, logical volume), including extending and resizing logical volumes.
- Creating, managing and configuring file share services such as NFS, FTP, TFTP. Making file systems available for a certain amount of time to prevent security threats.
- Managing system load and processes using tools such as top, sar, and isotat.
- Troubleshooting network issues using traceroute, tcpdump, netstat.Using nmap to perform vulnerability tests on the system.
- Working with different networking protocol ( DNS DHCP, PING, TCP,UDP,HTTPS)
- Managing and troubleshooting web servers (Tomcat, apache, nginx), in addition to using nginx as a load balancer and proxy.
- Configuring DNS, managing resource records, and creating forward and reverse zones
- Deploying and configuring machines using PXE servers, and deploying customized kickstart configuration files based on application and environment needs.
- Modifying grub configuration file and tuning kernel parameters based on server needs.
- Configuring DHCP by assigning network configuration to servers upon creation.
- Worked with different streamline editing tools to find specific lines or words in large files and directories ( sed,awk, grep, cut, sort).
- Configuring and provisioning bare metal server machines. (HP Proliant, Dell R710).
- Configuring RAIDS level 0,1,5,6 on a bare metal server based on server usage.
- Accessing IDRAC and ILO through and troubleshooting the hung servers.
- Deploying multiple VMWare hypervisor virtual machines based on team and applications needs with custom resources and OVF templates.
- Deploying clusters for teams and applications. Managing the cluster on VMWare for redundancy using HA and DRS, including migrating machines using vmotion.
- Creating and scheduling automated weekly and daily jobs using crontab and bash scripts, such as backing up sensitive data on mission critical systems to disaster recovery servers.
- Install and configure mySQL and oracle database on linux servers.
- Creating and configuring ssh keys for users and systems with password-less access, and troubleshooting any user access issues revolving with keys.
- Creating and managing users and groups, including managing sudo access for users.
- Managing permissions for files and directories as well as umask permission.
- Modifying the security of a linux system using iptables, firewall, selinux, and control access.
- Monitoring and maintaining linux servers using nagios, including creating and resolving any alerts to resource utilization, server down, or service error on production systems.
- Managing Nagios hosts configuration through NRPE plugins for new servers in environments.
- Compressing and archiving large files and directories and transferring them over to a different environment using scp or rsync protocol.
- Using package management tools such as YUM, RPM, APT, and setting up a local repository for dependency applications for applications and security.
- Creating and managing nic-bonding using NIC cards for redundancy, load balancing, and managing load on a network.
- Resolving a high number of incident tickets in a fast paced environment in accordance to SLA’s and escalating severity one issue with senior managers.
- Collaborating with developers, QA, Database admins, and other engineers to carry out large projects and improve existing processes.
