Post Job Free
Sign in

Senior DevOps Engineer - AWS, Kubernetes, GitOps

Location:
Denton, TX
Posted:
October 02, 2026

Contact this candidate

Resume:

NAGARAJU NAGAM

Senior DevOps Engineer

Texas, USA *********@*****.*** 321-***-**** LinkedIn

Professional Summary

DevOps Engineer with nearly 7 years of experience designing, automating, and operating cloud-native infrastructure across banking, healthcare, and e-commerce environments. Expertise in AWS, Kubernetes, Docker, Terraform, Ansible, GitOps, Infrastructure as Code, CI/CD, Linux administration, and cloud automation. Proven success building scalable cloud platforms, automating infrastructure provisioning, implementing enterprise CI/CD pipelines, modernizing application deployments, and improving developer productivity. Experienced in designing secure, resilient, and highly available cloud infrastructure supporting mission-critical enterprise applications while reducing operational overhead through automation and platform standardization. Technical Skills & Core Competencies

Core Competencies: Site Reliability Engineering, Platform Engineering, Kubernetes, AWS Cloud, Infrastructure as Code, Observability, GitOps, Production Operations, Reliability Automation, Incident Response, FinOps, Disaster Recovery Cloud Platforms: AWS (EC2, EKS, RDS, S3, Lambda, CloudFront, VPC), GCP, Azure, Hybrid Operating Systems & Version Control: Linux, Windows, MacOS, Git Containers/Platform: Kubernetes, OpenShift, Docker, Helm, Istio, Envoy, Argo Rollouts, KEDA, Karpenter, Cluster Autoscaler, Cilium, Linkerd

Platform Engineering: Internal Developer Platform, Golden Paths, Self-Service Infrastructure, Platform Automation, GitOps, Developer Experience

Infrastructure as Code (IaC): Terraform, Terraform Cloud, Ansible, CloudFormation, Pulumi CI/CD: Jenkins, GitHub Actions, GitLab CI/CD, TeamCity, Argo CD, Flux, JFrog Artifactory, Nexus Observability: Prometheus, Grafana, Loki, Tempo, Mimir, OpenTelemetry, Datadog, CloudWatch, PagerDuty, Splunk, ELK, AppDynamics, eBPF

Security: Vault, IAM, Kyverno, Gatekeeper, Falco

Networking: TCP/IP, DNS, Load Balancing

Languages: Python, Bash, PowerShell, Go

Reliability: SLI/SLO, Error Budgets, Availability Engineering, Capacity Engineering, Performance Engineering, Scalability Engineering, RCA, Chaos Engineering, Incident Management, MTTR/MTTD Reduction, High Availability, Operational Efficiency Professional Experience

Citi, Senior DevOps Engineer (Aug 2024 – Present)

• Designed reusable Terraform modules and Ansible playbooks to automate AWS infrastructure provisioning, significantly reducing manual deployment effort and improving infrastructure consistency across multiple production environments.

• Built and managed enterprise Kubernetes platforms on Amazon EKS and OpenShift supporting more than 40 production services running highly available containerized applications.

• Designed GitOps deployment workflows using Helm, Argo CD, and Argo Rollouts, reducing deployment lead time by 35% while standardizing application releases across engineering teams.

• Automated cloud operations by integrating Terraform, CloudWatch, Datadog, PagerDuty, and Python automation scripts, improving deployment reliability and reducing operational overhead.

• Developed reusable Infrastructure-as-Code templates enabling self-service infrastructure provisioning and standardized cloud deployments across AWS environments.

• Engineered reusable CI/CD pipeline templates supporting multiple engineering teams, improving deployment consistency and reducing manual release effort.

• Implemented Kubernetes autoscaling using KEDA, Cluster Autoscaler, and Karpenter to optimize resource utilization and improve workload scalability.

• Strengthened Kubernetes platform security by implementing Kyverno, Gatekeeper, Falco, and Cilium security policies across production clusters.

• Managed AWS networking infrastructure including VPC, IAM, Security Groups, Elastic Load Balancers, CloudWatch, Auto Scaling, and Route 53 to support secure enterprise cloud environments.

• Developed Python, Bash, and PowerShell automation scripts for infrastructure provisioning, application deployments, health checks, log management, and routine operational activities.

• Optimized infrastructure supporting PostgreSQL, Oracle, MongoDB, Redis, Kafka, and Elasticsearch workloads, improving platform performance and scalability.

• Led disaster recovery automation initiatives and platform resilience improvements, reducing recovery time by over 40%.

• Collaborated closely with development teams to standardize infrastructure, improve deployment automation, and enhance developer productivity through platform engineering initiatives. IBM, DevOps Engineer (May 2021 – Jul 2023)

• Led migration of more than 20 enterprise healthcare applications to Amazon EKS using Terraform, Docker, Helm, and Kubernetes, improving scalability, resilience, and deployment consistency.

• Designed and maintained enterprise CI/CD pipelines using Jenkins, GitHub Actions, GitLab CI/CD, TeamCity, Argo CD, and Flux, enabling automated build, testing, deployment, and rollback workflows.

• Automated AWS infrastructure provisioning using Terraform, Ansible, and CloudFormation, developing reusable Infrastructure-as-Code modules that significantly reduced environment provisioning time.

• Containerized enterprise applications using Docker and deployed workloads to Amazon EKS, improving application portability and accelerating software delivery.

• Implemented GitOps deployment strategies using Argo CD and Flux, enabling version-controlled infrastructure management and standardized release processes.

• Built centralized monitoring and observability platforms using Prometheus, Grafana, CloudWatch, Splunk, ELK Stack, and PagerDuty, contributing to a 55% reduction in MTTR.

• Automated operational processes using Python, Bash, and PowerShell, eliminating repetitive manual activities and improving engineering productivity.

• Configured Kubernetes Horizontal Pod Autoscaler, KEDA, Cluster Autoscaler, and Karpenter to improve application scalability while optimizing cloud infrastructure costs.

• Collaborated with development teams to streamline application onboarding, standardize deployment processes, and improve developer experience through platform automation.

• Participated in production support and on-call rotations supporting healthcare platforms serving over 1.2 million users while maintaining 99.99% platform availability.

Deloitte, Server Administrator (DevOps / SRE) (Jan 2019 – May 2021)

• Automated AWS infrastructure provisioning using Terraform, creating reusable Infrastructure-as-Code templates that accelerated environment provisioning and improved deployment consistency.

• Built and maintained Jenkins CI/CD pipelines supporting automated application build, testing, deployment, and release processes across development, QA, and production environments.

• Containerized enterprise applications using Docker and assisted in Kubernetes adoption, improving deployment portability and reducing application delivery time.

• Administered Linux servers supporting business-critical e-commerce applications, performing patch management, operating system upgrades, performance tuning, capacity planning, and production troubleshooting.

• Configured AWS services including EC2, IAM, VPC, Auto Scaling, Elastic Load Balancers, S3, CloudWatch, and Route 53 to support secure, scalable, and highly available cloud infrastructure.

• Developed automation scripts using Python and Bash for server provisioning, application deployments, log management, health monitoring, and routine operational tasks, significantly reducing manual engineering effort.

• Implemented centralized monitoring and alerting using Prometheus, Grafana, ELK Stack, and CloudWatch, reducing incident detection time by 40% while improving operational visibility.

• Collaborated with software development teams to automate deployments, troubleshoot production issues, improve infrastructure reliability, and enhance CI/CD processes.

• Supported disaster recovery activities, production incident resolution, infrastructure maintenance, and operational support for high-volume e-commerce platforms operating in a 24 7 production environment. Certifications

• AWS Certified DevOps Engineer – Professional

• AWS Certified Solutions Architect – Associate

• Certified Cloud Security Professional (CCSP)

Education

• Master of Information Technology & Cloud Management — Indiana Wesleyan University

• Bachelor of Computer Applications — Mahatma Gandhi University



Contact this candidate