PREETU SHARMA
DEVOPS SRE KUBERNETES CLOUD & PLATFORM ENGINEERING
New Delhi, India ******.******@*******.*** +91-959**-***** linkedin.com/in/preetu-sharma PROFESSIONAL SUMMARY
Site Reliability Engineer with 5.5+ years of experience building, operating and troubleshooting production cloud-native environments across Kubernetes, Azure, Linux and Docker. Strong hands-on experience with Rancher/RKE2, AKS, Helm, Kustomize, Terraform, GitHub Actions, Argo CD/Argo Rollouts, FluxCD, Prometheus/Grafana and incident response, with growing hands-on exposure to Azure DevOps Pipelines. Additional hands-on engineering in Golang REST APIs, backend services, unit testing and Kubernetes CRDs/Operators using Kubebuilder and controller-runtime, with Python and Bash automation. CORE TECHNICAL SKILLS
Kubernetes / DevOps: Kubernetes, Rancher, RKE2, AKS, Helm, Kustomize, RBAC, cluster upgrades, workload troubleshooting, Docker, container lifecycle
CI/CD / GitOps: GitHub Actions, Azure DevOps Pipelines, Argo CD, Argo Rollouts, FluxCD, Helm, Kustomize, GitOps, progressive delivery, release automation, Git
Cloud / IaC: Azure, AWS, OpenStack, Terraform, Azure VMs, networking, NSGs, load balancers, storage, IAM Golang / Backend: Golang, REST APIs, HTTP/JSON, backend services, Go interfaces, unit testing, Kubernetes API, CRDs, Kubebuilder, controller-runtime
SRE / Automation: Linux, Prometheus, Grafana, Zabbix, Python, Bash/Shell, SLIs/SLOs, error budgets, incident response, RCA, SQL, PostgreSQL, MySQL
PROFESSIONAL EXPERIENCE
Site Reliability Engineer Genpact India Private Limited Jun 2024 - Present
● Own reliability and day-to-day operations for 10+ production Kubernetes clusters on Rancher covering provisioning, version upgrades, RBAC, certificates management, node health, workload scheduling and production troubleshooting.
● Deploy, operate and troubleshoot 30+ containerized microservices across on-premises Kubernetes and Azure Kubernetes Service (AKS) using Docker and Helm; support application configuration, workload health and platform issues while sustaining 99.9% uptime.
● Engineer repeatable CI/CD and GitOps delivery with GitHub Actions, Argo CD/Argo Rollouts and FluxCD, automating application releases, environment promotion and progressive production rollouts with consistent deployment workflows.
● Manage templated, environment-specific Kubernetes manifests in production using Helm charts and Kustomize overlays alongside FluxCD, improving release consistency and reducing configuration drift across clusters.
● Build Terraform automation for Azure compute, networking, storage and IAM, improving repeatability and reducing manual release and infrastructure effort by 30% across secure, highly available environments.
● Lead production incident response, troubleshooting and root-cause analysis using Kubernetes diagnostics, Prometheus and Grafana; contribute to a 25% MTTR reduction and strengthen SLI/SLO, alerting and error-budget practices.
● Develop engineering automation in Golang and Python/Bash, including REST API/backend development and unit testing; build Kubernetes CRDs and Operators/Controllers with Kubebuilder and controller-runtime for reconciliation, status management and workload automation.
● Own Docker image lifecycle, optimization and vulnerability scanning, reducing image build time by 20%; support healthcare and medical-imaging infrastructure across 100+ hospital servers using PACS, DICOM and HL7.
● Harden cluster and workload security by enforcing RBAC least-privilege policies, network policies and centralized secrets management, reducing unauthorized-access risk across shared multi-tenant clusters.
● Right-size CPU/memory requests-limits and autoscaling configuration for 30+ microservices, improving cluster resource utilization and cutting infrastructure spend without impacting SLAs.
● Participate in the on-call rotation as primary escalation point for platform-level incidents, driving faster acknowledgement and structured post-incident reviews to prevent repeat outages.
● Author and maintain runbooks, architecture docs and onboarding guides for the platform team, and evaluate Azure DevOps Pipelines as an alternate CI/CD path alongside GitHub Actions for select workloads. Cloud Support Associate UST Nov 2021 - Jun 2024
● Supported production Kubernetes environments across on-premises and Azure platforms, managing workloads, Helm releases, Rancher operations, service availability and day-to-day cluster troubleshooting.
● Administered Azure infrastructure and cloud services including virtual machines, AKS, networking, load balancers and storage, supporting enterprise application environments and resolving infrastructure-level issues.
● Performed Linux system administration across production servers running Nginx, MySQL and PostgreSQL, handling service management, resource utilization, connectivity issues, configuration changes and operational troubleshooting.
● Designed and maintained infrastructure monitoring and alerting using Zabbix, Prometheus and Grafana, building dashboards and alerts for Kubernetes, Linux and application infrastructure to improve visibility and accelerate incident detection.
● Managed Docker-based application environments, including container deployment, lifecycle operations, configuration, networking and troubleshooting across cloud and on-premises platforms.
● Handled production support tickets against SLA/ITIL processes, triaging incidents by severity and coordinating with L2/L3 teams to restore service and document resolutions for the knowledge base.
● Applied OS patching, security updates and configuration hardening across Linux and Azure VM fleets on a scheduled cadence, reducing exposure to known vulnerabilities.
● Participated in backup and disaster-recovery testing for critical workloads, validating restore procedures and RTO/RPO targets across on-premises and Azure environments. Engineer, Monitoring NOC Progressive Infotech Pvt. Ltd. Jan 2021 - Oct 2021
● Administered Linux (CentOS/Ubuntu) and OpenStack-based hybrid cloud infrastructure across on-premises and Azure environments covering compute, networking, storage, backup and site recovery.
● Used Bash scripting to automate recurring operational tasks, simplify system administration and reduce manual troubleshooting effort.
● Administered Azure Active Directory including user/group management, SSO integration, and conditional access policies.
● Coordinated with application, infrastructure and network teams during incidents, providing monitoring data, initial root- cause analysis and timely escalation for production issues. EDUCATION
B.Tech in Computer Science & Engineering - DCRUST, Murthal, India 2022 - 2025 CGPA 6.98/10 Diploma Engineering - Aryabhat Institute of Technology, Delhi, India 2014 - 2017 CGPA 8.80/10 Secondary School Certificate (SSC) - RBRR Vidya Mandir, Delhi, India 2013 - 2014 CGPA: 8.40 / 10 CERTIFICATIONS
• Certified Kubernetes Administrator (CKA) 2025 - 2027
• Certified Kubernetes Application Developer (CKAD) 2025 - 2027
• Certified Kubernetes Security Specialist (CKS) 2025 - 2027
• Kubernetes and Cloud Native Associate (KCNA) 2025 - 2027
• Kubernetes and Cloud Native Security Associate (KCSA) 2025 - 2027
• CNCF Kubestronaut 2026
• Linux Foundation Certified System Administrator (LFCS) 2024 - 2026
• Microsoft Certified: Azure Administrator Associate (AZ-104) 2022 - 2027
• Microsoft Certified: Azure Fundamentals (AZ-900) 2021 LANGUAGES
Hindi — Native English — Fluent (professional working proficiency)