Pawan Yadav
DevOps & Platform Engineer
***************@*****.*** +91-817******* Faridabad India
Summary
DevOps & Platform Engineer with 4+ years of experience designing, automating, and managing cloud-native and on-premises infrastructure across AWS, GCP, and Kubernetes environments. Strong expertise in Kubernetes, Docker, Terraform, GitLab CI/CD, Argo CD, Helm, AWS
(EKS/ECS), and platform engineering, delivering scalable, highly available infrastructure for production workloads. Proven experience migrating 100+ microservices from cloud to on-premises Kubernetes, building GitOps-driven deployment pipelines, implementing observability with Prometheus, Grafana, and Loki, and improving platform reliability through automation and infrastructure best practices. Actively expanding into MLOps and AIOps, applying platform engineering and observability expertise to ML pipeline automation, model deployment infrastructure, and AI workload operations. AWS Certified Solutions Architect – Associate, HashiCorp Certified: Terraform Associate, and Google Cloud Certified Associate Cloud Engineer. Skills
AWS Cloud
EC2, ECS, EKS, VPC,
IAM, S3, RDS, Lambda,
DynamoDB, API Gateway,
WAF, MediaLive, MediaTailor,
MediaConvert
GCP Cloud & On-Prem
Compute Engine, GKE, VPC,
IAM, Cloud Run, App
Engine, Cloud Monitoring,
On-Premises Infrastructure,
Bare-Metal Kubernetes
IaC
Terraform, Pulumi, AWS
CloudFormation, Ansible
Containers &
Orchestration
Docker, Kubernetes
(On-Premises, EKS), Amazon
ECS, Helm, Argo CD, Docker
Compose
CI/CD & GitOps
GitLab CI/CD, Argo CD,
Jenkins, GitOps
Networking & API
Management
HAProxy, MetalLB, Kong API
Gateway, MikroTik Firewall,
Ingress, DNS, Load Balancing,
SSL/TLS, Nginx
Monitoring &
Observability
Prometheus, Grafana, Loki,
Dynatrace, Datadog, New
Relic, ELK, Alerting, Log
Aggregation
Databases & Messaging
PostgreSQL, Redis, RabbitMQ,
Qdrant, RDS
Operating Systems &
Scripting
Linux (Ubuntu), Bash, Python
Version Control &
Collaboration
Git, GitLab, GitHub, Bitbucket,
JIRA
Experience
Dizzaract FZ LLC Remote
DevOps Engineer June 2026 - Present
• Led migration of 80+ microservices and supporting infrastructure from Google Cloud Platform (GCP) to on-premises Kubernetes, reducing cloud dependency and increasing infrastructure control.
•Architected a highly available Kubernetes cluster (3 control-plane, 4 worker nodes) to support production workloads with zero single points of failure.
• Engineered bare-metal load balancing using HAProxy and MetalLB, ensuring consistent, highly available external traffic routing for Kubernetes services.
• Participated in architecture and design reviews for production platform components, focusing on scalability, reliability, maintainability, and high availability.
• Configured Kong API Gateway (public and private) for optimized API routing, authentication, and controlled service exposure.
• Built a centralized observability stack (Prometheus, Grafana, Loki) delivering real-time monitoring, log aggregation, and operational dashboards.
•Automated CI/CD and GitOps workflows using GitLab CI/CD and Argo CD, enabling continuous, reliable deployments.
•Owned production service health and incident troubleshooting across Kubernetes workloads, using Prometheus, Grafana, and Loki telemetry to identify issues, perform root cause analysis, and drive corrective actions.
•Managed containerized databases/messaging (PostgreSQL, Redis, RabbitMQ, Qdrant) and owned TLS/SSL, DNS, and ingress configuration for high-availability, secure application delivery.
• Created and maintained technical documentation, deployment procedures, troubleshooting guides, and operational runbooks for Kubernetes and application infrastructure.
To The New Pvt Ltd Noida India
Senior DevOps Engineer October 2025 - May 2026
• Led the design and scaling of AWS infrastructure using Terraform and Pulumi, developing 15+ reusable IaC modules adopted across multiple microservices projects and standardized into shared repositories for org-wide use.
•Managed Kubernetes (EKS) and ECS workloads supporting high-traffic platforms across a large multi-cluster environment, maintaining 99.9% uptime.
•Owned and optimized CI/CD pipelines (Jenkins, GitLab CI, Buildkite), cutting deployment time by ~40% through pipeline automation and release standardization.
• Implemented centralized observability (Dynatrace, Prometheus, Grafana), improving incident detection speed by ~30% while reducing operational monitoring costs by ~15%.
• Built a Python-based AWS security auditing tool scanning 30+ resource types per account with automated remediation recommendations.
•Mentored 3–5 engineers and led IaC/CI-CD code reviews, embedding security best practices (IAM, WAF) and deployment standards across the team.
•Developed reusable Python-based automation and internal tooling for AWS infrastructure and security operations, following modular, maintainable, and scalable software development practices.
•Designed and delivered an end-to-end live-streaming solution using AWS MediaLive, MediaTailor, Lambda, and DynamoDB, supporting scalable concurrent streams.
• Performed root cause analysis for production incidents using monitoring and service telemetry, implementing corrective and preventive actions to improve system reliability and reduce recurring issues. To The New Pvt Ltd Noida India
DevOps Engineer Feb 2022 - Oct 2025
• Built cloud infrastructure from scratch using Terraform, Pulumi (Python), and CloudFormation across 5+ production projects.
•Developed and maintained CI/CD pipelines using Jenkins, enabling automated deployment of 20+ microservices and serverless workloads on AWS (ECS, EKS, Lambda).
•Automated IAM role provisioning and operational workflows using Python, reducing manual effort by ~50% and accelerating service onboarding.
• Implemented monitoring and observability using Dynatrace, Prometheus, and Grafana, enabling real-time insights and proactive incident management.
• Collaborated with cross-functional teams to support high-volume data platforms and API workloads using MongoDB, Cassandra, RDS, and Redshift.
•Managed EKS clusters and optimized CI/CD pipelines for an OTT streaming platform, reducing infrastructure costs by ~20%.
• Built serverless data pipelines using AWS Lambda and ECS for a gaming and betting platform, and automated IAM role management across environments.
Certifications
AWS Certified Solutions Architect – Associate
HashiCorp Certified: Terraform Associate
Google Cloud Certified Associate Cloud Engineer
Professional Development
MLOps Fundamentals — ML lifecycle, model deployment patterns, monitoring for ML systems, and running AI/ML workloads on Kubernetes (studying via KodeKloud)
Education
Poornima University Bachelor of Technology • 8.2
Computer Science Jaipur India • Aug 2018 - July 2022 Languages
English
Fluent
Hindi
Native