Post Job Free
Sign in

Kubernetes SRE, DevOps & Cloud Platform Engineer

Location:
New Delhi, Delhi, India
Posted:
September 16, 2026

Contact this candidate

Resume:

PREETU SHARMA

DEVOPS SRE KUBERNETES CLOUD & PLATFORM ENGINEERING

New Delhi, India ******.******@*******.*** +91-959**-***** linkedin.com/in/preetu-sharma PROFESSIONAL SUMMARY

Site Reliability Engineer with 5.5+ years of experience building, operating and troubleshooting production cloud-native environments across Kubernetes, Azure, Linux and Docker. Strong hands-on experience with Rancher/RKE2, AKS, Helm, Kustomize, Terraform, GitHub Actions, Argo CD/Argo Rollouts, FluxCD, Prometheus/Grafana and incident response, with growing hands-on exposure to Azure DevOps Pipelines. Additional hands-on engineering in Golang REST APIs, backend services, unit testing and Kubernetes CRDs/Operators using Kubebuilder and controller-runtime, with Python and Bash automation. CORE TECHNICAL SKILLS

Kubernetes / DevOps: Kubernetes, Rancher, RKE2, AKS, Helm, Kustomize, RBAC, cluster upgrades, workload troubleshooting, Docker, container lifecycle

CI/CD / GitOps: GitHub Actions, Azure DevOps Pipelines, Argo CD, Argo Rollouts, FluxCD, Helm, Kustomize, GitOps, progressive delivery, release automation, Git

Cloud / IaC: Azure, AWS, OpenStack, Terraform, Azure VMs, networking, NSGs, load balancers, storage, IAM Golang / Backend: Golang, REST APIs, HTTP/JSON, backend services, Go interfaces, unit testing, Kubernetes API, CRDs, Kubebuilder, controller-runtime

SRE / Automation: Linux, Prometheus, Grafana, Zabbix, Python, Bash/Shell, SLIs/SLOs, error budgets, incident response, RCA, SQL, PostgreSQL, MySQL

PROFESSIONAL EXPERIENCE

Site Reliability Engineer Genpact India Private Limited Jun 2024 - Present

● Own reliability and day-to-day operations for 10+ production Kubernetes clusters on Rancher covering provisioning, version upgrades, RBAC, certificates management, node health, workload scheduling and production troubleshooting.

● Deploy, operate and troubleshoot 30+ containerized microservices across on-premises Kubernetes and Azure Kubernetes Service (AKS) using Docker and Helm; support application configuration, workload health and platform issues while sustaining 99.9% uptime.

● Engineer repeatable CI/CD and GitOps delivery with GitHub Actions, Argo CD/Argo Rollouts and FluxCD, automating application releases, environment promotion and progressive production rollouts with consistent deployment workflows.

● Manage templated, environment-specific Kubernetes manifests in production using Helm charts and Kustomize overlays alongside FluxCD, improving release consistency and reducing configuration drift across clusters.

● Build Terraform automation for Azure compute, networking, storage and IAM, improving repeatability and reducing manual release and infrastructure effort by 30% across secure, highly available environments.

● Lead production incident response, troubleshooting and root-cause analysis using Kubernetes diagnostics, Prometheus and Grafana; contribute to a 25% MTTR reduction and strengthen SLI/SLO, alerting and error-budget practices.

● Develop engineering automation in Golang and Python/Bash, including REST API/backend development and unit testing; build Kubernetes CRDs and Operators/Controllers with Kubebuilder and controller-runtime for reconciliation, status management and workload automation.

● Own Docker image lifecycle, optimization and vulnerability scanning, reducing image build time by 20%; support healthcare and medical-imaging infrastructure across 100+ hospital servers using PACS, DICOM and HL7.

● Harden cluster and workload security by enforcing RBAC least-privilege policies, network policies and centralized secrets management, reducing unauthorized-access risk across shared multi-tenant clusters.

● Right-size CPU/memory requests-limits and autoscaling configuration for 30+ microservices, improving cluster resource utilization and cutting infrastructure spend without impacting SLAs.

● Participate in the on-call rotation as primary escalation point for platform-level incidents, driving faster acknowledgement and structured post-incident reviews to prevent repeat outages.

● Author and maintain runbooks, architecture docs and onboarding guides for the platform team, and evaluate Azure DevOps Pipelines as an alternate CI/CD path alongside GitHub Actions for select workloads. Cloud Support Associate UST Nov 2021 - Jun 2024

● Supported production Kubernetes environments across on-premises and Azure platforms, managing workloads, Helm releases, Rancher operations, service availability and day-to-day cluster troubleshooting.

● Administered Azure infrastructure and cloud services including virtual machines, AKS, networking, load balancers and storage, supporting enterprise application environments and resolving infrastructure-level issues.

● Performed Linux system administration across production servers running Nginx, MySQL and PostgreSQL, handling service management, resource utilization, connectivity issues, configuration changes and operational troubleshooting.

● Designed and maintained infrastructure monitoring and alerting using Zabbix, Prometheus and Grafana, building dashboards and alerts for Kubernetes, Linux and application infrastructure to improve visibility and accelerate incident detection.

● Managed Docker-based application environments, including container deployment, lifecycle operations, configuration, networking and troubleshooting across cloud and on-premises platforms.

● Handled production support tickets against SLA/ITIL processes, triaging incidents by severity and coordinating with L2/L3 teams to restore service and document resolutions for the knowledge base.

● Applied OS patching, security updates and configuration hardening across Linux and Azure VM fleets on a scheduled cadence, reducing exposure to known vulnerabilities.

● Participated in backup and disaster-recovery testing for critical workloads, validating restore procedures and RTO/RPO targets across on-premises and Azure environments. Engineer, Monitoring NOC Progressive Infotech Pvt. Ltd. Jan 2021 - Oct 2021

● Administered Linux (CentOS/Ubuntu) and OpenStack-based hybrid cloud infrastructure across on-premises and Azure environments covering compute, networking, storage, backup and site recovery.

● Used Bash scripting to automate recurring operational tasks, simplify system administration and reduce manual troubleshooting effort.

● Administered Azure Active Directory including user/group management, SSO integration, and conditional access policies.

● Coordinated with application, infrastructure and network teams during incidents, providing monitoring data, initial root- cause analysis and timely escalation for production issues. EDUCATION

B.Tech in Computer Science & Engineering - DCRUST, Murthal, India 2022 - 2025 CGPA 6.98/10 Diploma Engineering - Aryabhat Institute of Technology, Delhi, India 2014 - 2017 CGPA 8.80/10 Secondary School Certificate (SSC) - RBRR Vidya Mandir, Delhi, India 2013 - 2014 CGPA: 8.40 / 10 CERTIFICATIONS

• Certified Kubernetes Administrator (CKA) 2025 - 2027

• Certified Kubernetes Application Developer (CKAD) 2025 - 2027

• Certified Kubernetes Security Specialist (CKS) 2025 - 2027

• Kubernetes and Cloud Native Associate (KCNA) 2025 - 2027

• Kubernetes and Cloud Native Security Associate (KCSA) 2025 - 2027

• CNCF Kubestronaut 2026

• Linux Foundation Certified System Administrator (LFCS) 2024 - 2026

• Microsoft Certified: Azure Administrator Associate (AZ-104) 2022 - 2027

• Microsoft Certified: Azure Fundamentals (AZ-900) 2021 LANGUAGES

Hindi — Native English — Fluent (professional working proficiency)



Contact this candidate