Post Job Free
Sign in

Site Reliability Engineer - DevOps Kubernetes

Location:
Community Center, CA, 94087
Posted:
July 27, 2026

Contact this candidate

Resume:

Meghana Puvvadi

Email: ***************@*****.***

Mobile: +1-832-***-****

LinkedIn: www.linkedin.com/in/meghana-puvvadi-a24b25292 Site Reliability Engineer

PROFESSIONAL SUMMARY

DevOps Engineer with 4+ years of experience supporting CI/CD, cloud infrastructure, automation, monitoring, security, and scalable deployments across production environments for distributed enterprise systems.

Skilled in Terraform, Kubernetes, Docker, Jenkins, GitHub Actions, Azure DevOps, AWS, Azure, GCP, and observability practices improving reliability and release consistency across cloud platforms.

Experienced in Infrastructure as Code, GitOps, monitoring, alerting, incident response, IAM, RBAC, and change control, aligning secure operations with documentation standards across enterprise teams.

Collaborative engineer delivering runbooks, deployment pipelines, environment lifecycle practices, configuration management, and troubleshooting support for cloud-native, SaaS, and platform engineering teams across production environments.

Facilitated strong communication skills by leading team meetings, ensuring clarity in project goals, and enhancing collaboration, which resulted in improved teamwork and a more cohesive work environment.

Applied a customer-centric mindset by actively listening to client feedback and tailoring solutions accordingly, ultimately boosting customer satisfaction and fostering long-term relationships with key clients. TECHNICAL SKILLS

Containerization & Orchestration - Kubernetes, Docker, Helm, ArgoCD, Flux, GitOps, Container Registry

SRE & Observability - Datadog, Prometheus, Grafana, ELK Stack, PagerDuty, CloudWatch, Azure Monitor, New Relic, SLO/SLI Management, Incident Response, MTTR Optimization, SLIs, SLOs, monitoring gaps, alerting strategies

Programming & Scripting - Python (Boto3), Bash, PowerShell, Go, YAML, JSON

Collaboration Tools - Git, GitHub, GitLab, Jira, Confluence, Slack, Agile/Scrum

Cloud Platforms - AWS (EKS, EC2, Lambda, S3, RDS, CloudFormation, CloudWatch, VPC), Azure (AKS, Azure DevOps, Functions, App Services, ARM Templates, Azure Monitor)

CI/CD & Automation - Azure DevOps, Jenkins, GitLab CI, GitHub Actions, Azure Pipelines, AWS CodePipeline, CI/CD practices

Security & Compliance - HashiCorp Vault, Prisma Cloud, SonarQube, Azure Security Center, IAM (AWS/Azure), DevSecOps, PCI-DSS, SOC 2

Databases - PostgreSQL, MySQL, MongoDB, DynamoDB, Redis, Azure SQL Database

Operating Systems - Linux (Ubuntu, CentOS, RHEL), Windows Server, Amazon Linux, Linux kernel internals

Service Mesh & Infrastructure as Code - Terraform, Ansible, CloudFormation, ARM Templates, Bicep

System Performance & Debugging - system performance metrics, analyzing system logs, system-level debugging, kernel panic analysis, kdump, memory allocation, scheduler, bare-metal cloud infrastructure

Storage Solutions - distributed storage systems, object storage, block storage, file storage

Resilience & Scaling - resilient code, scaling systems PROFESSIONAL EXPERIENCE

JPMorgan Chase January 2025 – Present

Senior DevOps Engineer New York, NY, USA

Orchestrated CI/CD practices using AWX to automate deployment pipelines, achieving a 40% reduction in release cycle time while enhancing system reliability and increasing overall deployment frequency across microservices architecture.

Modernized object storage solutions alongside block storage implementations, enhancing data retrieval speeds by 40% and optimizing storage costs by 30% across the organization.

Architected design patterns for scaling systems, resulting in a 200% increase in user capacity without performance degradation while maintaining system stability during peak loads.

Orchestrated system-level debugging initiatives focusing on Linux kernel internals, which improved fault tolerance and reduced system downtime by 50% in high-availability environments.

Engineered a bare-metal cloud infrastructure utilizing TCP/IP network programming and an OVN/OVS-based networking stack, resulting in a 40% increase in deployment speed and a 99.99% reduction in latency.

Designed resilient CI/CD pipelines with Jenkins, GitHub Actions, and Terraform, improving deployment consistency, rollback readiness, and release governance for regulated applications across production environments.

Engineered Kubernetes and Docker deployment workflows with Helm and ArgoCD, strengthening container orchestration, environment parity, and repeatable application delivery across teams in production operations.

Automated Infrastructure as Code provisioning with Terraform, Ansible, and AWS, reducing configuration drift while improving cloud scalability, reliability, and operational transparency for critical workloads.

Integrated Prometheus, Grafana, CloudWatch, and runbooks for observability, accelerating incident response, root cause analysis, and proactive system health management across distributed service platforms securely. Toyota Financial Services August 2023 – December 2024 DevOps Engineer Texas, USA

Revamped file storage architectures to align with SLIs, leading to a 25% increase in data access speeds while maintaining compliance with stringent SLOs and performance standards.

Streamlined monitoring gaps by implementing advanced alerting strategies, which enabled proactive issue resolution and reduced average downtime by 60% across critical services.

Executed kernel panic analysis within distributed storage systems, identifying root causes that led to a 99.99% platform availability and enhanced data integrity across services.

Automated routine processes by integrating system performance metrics and Hardware GPU troubleshooting, which eliminated 40+ weekly operational hours and improved overall system responsiveness by 30%.

Refined release workflows with Azure Pipelines, GitLab CI, and Maven, increasing build traceability, approval consistency, and predictable deployments for financial services applications.

Secured cloud resources with IAM, RBAC, SAML, and Key Vault, improving access governance, credential protection, and audit readiness across regulated application environments.

Administered ServiceNow change records, Jira work items, and Confluence documentation, improving release coordination, operational transparency, and handoff quality across distributed DevOps teams.

Monitored application performance with Splunk, Datadog, and New Relic, improving alert investigation, service stability, and production support efficiency across cloud-hosted platforms. Tata Consultancy Services May 2021 – December 2022 Cloud Engineer Hyderabad, India

Optimized architecture to meet evolving SLOs, facilitating a 50% improvement in deployment efficiency and enabling rapid scaling to accommodate a 300% increase in user demand.

Consolidated memory allocation strategies within the driver subsystem, achieving a 50% reduction in memory leaks and improving system performance metrics significantly during high-load scenarios.

Pioneered post-mortems for production incidents, fostering a culture of resilient code development that decreased future incident rates by 35% and improved overall software reliability.

Analyzed system logs to enhance monitoring capabilities, leading to the identification of critical bottlenecks and a 25% improvement in incident response times across the infrastructure.

Streamlined cloud migrations with Azure, GCP, and Terraform, improving environment lifecycle management, resource optimization, and reliable modernization delivery across client platforms for enterprise programs.

Orchestrated Docker, Kubernetes, ECS, and EKS workloads for scalable containerized environments, improving deployment repeatability, availability, and platform engineering outcomes across staging and production systems.

Hardened Linux and Windows Server operations with patch management, vulnerability scanning, and firewalls, strengthening system administration, security posture, and operational resilience across hybrid infrastructure.

Documented standards, runbooks, and configuration management practices with Jira and Confluence, improving collaboration, onboarding, and repeatable support across cloud operations for distributed engineering teams. CERTIFICATIONS

AWS Certified Cloud Practitioner

EDUCATION

Master's in Computer Science - Lamar University



Contact this candidate