Post Job Free
Sign in

Senior Platform Engineer (AI-Powered SRE)

Location:
San Diego, CA
Posted:
August 18, 2026

Contact this candidate

Resume:

GURUSHANKAR SEETHARAMAN

San Diego, CA 804-***-**** ***********@*****.*** linkedin.com/in/gurushankar-seetharaman SENIOR PLATFORM ENGINEER SRE AI-POWERED CLOUD INFRASTRUCTURE Platform engineer and SRE with 18+ years building cloud-native infrastructure, Kubernetes platforms, and CI/CD systems across fintech, semiconductor, healthcare, and enterprise banking. Currently designing AI-powered operational tooling at a high-growth fintech, including LLM-integrated platforms for incident intelligence, automated alert triage, and cluster-wide diagnostics across 20+ EKS environments. Track record of driving $1M+ in infrastructure cost savings, reducing deployment times by 70%, and leading engineering teams through large-scale cloud transformations. AREAS OF EXPERTISE

Cloud Platforms & Infrastructure CI/CD Automation Infrastructure as Code (IaC) DevSecOps Platform Engineering & SRE Kubernetes & Container Orchestration AI-Powered Operational Tooling Security & Compliance Observability & Monitoring Cloud Cost Optimization System Architecture Cross-Functional Engineering Leadership Automation & Scripting High-Concurrency Systems Incident Intelligence & Response Developer Productivity Tools Cloud Migrations & Strategy Programming: Python FastAPI Bash PowerShell C# ASP.NET Java JavaScript Node.js Cloud Platforms: AWS (EKS, EC2, IAM, VPC, S3, ECR, Secrets Manager, CloudWatch) Azure DevOps GCP (GKE) Kubernetes & Platform: EKS Karpenter Helm Kustomize Docker Linkerd CoreDNS Kyverno GitOps & IaC: Flux CD ArgoCD Argo Workflows Terraform Terraform Cloud Ansible AI & LLM Integration: OpenAI GPT-4o API Claude API Prompt Engineering Agentic Workflows LLM-Assisted Analysis CI/CD Tools: Jenkins TeamCity GitLab CI GitHub Actions Octopus Deploy Azure DevOps Observability: Datadog Prometheus Grafana ELK Stack (Logstash, Kibana) Splunk Cortex CloudWatch Networking & Security: DNS VPC VPN IAM Security Groups Load Balancers SAST/DAST Secrets Management Messaging & Queues: Kafka AWS SNS SQS Redis Memcached Certifications: PMP Certified Scrum Master (CSM) ITIL Foundation MCSD KEY ACHIEVEMENTS

AI-Powered Platform Intelligence

• Built an Incident Intelligence Platform that aggregates real-time data from incident.io, Datadog, Slack, and JIRA, using OpenAI GPT-4o to automate root cause analysis, recurrence risk scoring, and executive reporting.

• Created an LLM-powered alert triage system that classifies incoming alerts as Noise, Warranted, or Unclear, cutting manual review time and improving signal-to-noise ratio for on-call engineers.

• Developed a Platform Health Analyzer that scans 20+ EKS clusters across multiple AWS accounts and generates context-aware remediation steps aligned to the team's GitOps workflows. CI/CD & Automation Efficiency

• Reduced deployment time by 70% across 100+ microservices by standardizing CI/CD pipelines with Jenkins, Terraform, and Ansible, cutting manual overhead and improving release frequency.

• Eliminated $500K+ in annual costs tied to manual release processes through end-to-end pipeline automation.

• Improved developer productivity 30% with auto-triggered, environment-specific pipelines and GitHub Actions. AI/ML Infrastructure & Observability

• Cut model deployment cycles in half by building autoscaling Kubernetes clusters paired with a Cortex and Prometheus observability layer for real-time monitoring.

• Increased platform uptime by 40% through self-healing infrastructure, optimized scheduling, and proactive alerting.

• Introduced AI/ML observability workflows that integrate model health checks, inference latency tracking, and performance telemetry into existing production monitoring systems. DevSecOps & Security Transformation

• Reduced vulnerability exposure time by 80% by embedding automated SAST/DAST scanning, policy enforcement, and compliance gates directly into CI/CD pipelines across all environments.

• Improved audit readiness by 90% through automated encryption, role-based access controls, and IAM governance.

• Minimized production risk with container image scanning, Vault integration, and centralized secrets management. Cloud Cost Optimization

• Delivered $1M+ in annual infrastructure savings through consolidation, rightsizing, and strategic use of spot instances.

• Led cost attribution analysis across 9 AWS resource categories totaling $440K+ in monthly spend, mapping costs to teams and identifying key optimization targets.

Scalability & Reliability

• Designed high-concurrency systems with transactional locking and optimized request handling to support latency- sensitive workloads across production financial services infrastructure.

• Implemented zero-downtime Kubernetes deployments with automated rollback and progressive delivery via ArgoCD.

• Drove site reliability improvements tracked through OKRs for infra health, observability, and availability. Leadership & Strategic Impact

• Led a team of 10+ engineers to build unified DevSecOps pipelines serving 150+ production services.

• Accelerated release velocity by 3x through GitOps adoption, Helm chart standardization, and pipeline consolidation.

• Recognized by CTO for platform reliability improvements and building automated incident response capabilities. PROFESSIONAL EXPERIENCE

EarnIn Remote 02/2025 – Present

SENIOR PLATFORM ENGINEER

• Designed and built an AI-powered Incident Intelligence Platform that centralizes data from incident.io, Datadog, Slack, and JIRA, giving engineering leadership a single view into incident patterns and response quality.

• Integrated OpenAI GPT-4o APIs to automate incident analysis, generating recurrence risk scores, identifying response gaps, and assessing real-time business impact across production systems.

• Built a Platform Health Analyzer that scans 20+ EKS clusters across multiple AWS accounts, surfacing context-aware remediation steps aligned to the team's GitOps workflows.

• Developed an AI-powered alert triage system that classifies alerts as Noise, Warranted, or Unclear using LLM-assisted analysis, cutting on-call fatigue and sharpening incident response.

• Own and operate platform infrastructure spanning production, staging, sandbox, management, and PCI-compliant environments across 20+ EKS clusters with full lifecycle ownership.

• Manage core platform components including Karpenter, Linkerd, Kyverno, CoreDNS, cert-manager, ExternalSecrets, ECR replication, and the GitOps delivery layer.

• Led AWS cost attribution analysis across 9 resource categories and 15+ accounts, mapping $440K+ in monthly spend to specific teams, services, and infrastructure components. Qualcomm San Diego, CA 11/2022 – 11/2024

SENIOR STAFF DEVOPS ENGINEER (AI SOFTWARE TEAM)

• Built CI/CD pipelines for TensorFlow-based AI/ML inference models targeting edge devices, automating the full path from training artifacts to production.

• Managed EKS clusters and containerized environments for distributed AI teams, ensuring secure, reproducible deployments across AWS regions.

• Automated infrastructure provisioning with Terraform and Jenkins, reducing environment setup time and eliminating configuration drift across accounts.

• Deployed a Prometheus and Grafana observability stack that improved pipeline reliability and gave teams early visibility into cluster health issues.

• Embedded vulnerability scanning, IAM policy enforcement, and container image validation directly into every stage of the deployment pipeline.

• Supported model validation automation and cross-geography infrastructure replication, enabling consistent deployment lifecycles across global teams.

Truist Bank San Diego, CA 01/2018 – 11/2022

SENIOR DEVOPS (CI/CD) ENGINEER

• Designed CI/CD pipelines for enterprise banking applications using TeamCity, Octopus Deploy, and Jenkins, automating a previously manual release workflow.

• Built infrastructure-as-code automation with Ansible and Terraform, standardizing provisioning workflows across regulated, SOC-audited AWS environments.

• Managed Kubernetes clusters and Docker-based container platforms in a SOC-audited banking environment under strict change control and documented release procedures.

• Implemented CloudWatch and Grafana dashboards giving operations teams real-time visibility into system health, performance metrics, and deployment status.

• Optimized source control workflows and automated testing pipelines, significantly reducing deployment failures and strengthening team-wide release confidence.

Mitchell International San Diego, CA 06/2015 – 12/2017 SENIOR DEVOPS (CI/CD) ENGINEER

• Built Jenkins-based CI/CD pipelines for .NET and Java applications, automating previously manual build and deployment processes with repeatable, version-controlled workflows.

• Decomposed monolithic builds into testable components, cutting deployment cycle times and broadening unit test coverage across the full application codebase.

• Integrated AWS services into the CI/CD pipeline and managed infrastructure automation with Terraform and Ansible.

• Deployed a Prometheus and Grafana monitoring stack that improved pipeline reliability and gave teams faster visibility into cluster and node health degradation.

• Managed Kubernetes clusters and Docker environments, improving infrastructure scalability and streamlining day-to- day platform operations and incident response.

Humana Louisville, KY 12/2014 – 06/2015

LEAD DEVELOPER

• Led the rebuild of legacy applications into responsive web platforms using .NET MVC, JavaScript, and Kendo UI.

• Provided technical direction to onshore and offshore development teams, ensuring consistent delivery standards and engineering quality across all project streams.

• Implemented SHA-256 encryption and led new engineer onboarding through structured training, mentoring, and hands-on coaching across all active project streams. SWBC San Antonio, TX 07/2014 – 11/2014

LEAD PROGRAMMER ANALYST

• Modernized legacy applications with ASP.NET MVC and WCF, improving page load performance and user experience.

• Redesigned application data flow using MVVM architecture and RESTful APIs, simplifying service communication and reducing tight coupling between system components. Enterprise Products Houston, TX 02/2014 – 06/2014 SENIOR PROGRAMMER ANALYST

• Built frontend automation for oil and gas operations using Kendo UI, Knockout JS, C#, and Entity Framework. Intel India Technology Pvt Ltd Bangalore, India 01/2005 – 02/2014 APPLICATION DEVELOPMENT LEAD / PROJECT MANAGER

• Led global project delivery across planning, estimation, resource allocation, and cross-functional coordination.

• Managed end-to-end delivery timelines, balancing resource allocation and schedule constraints across regions.

• Streamlined global support operations, reducing costs through targeted process improvements and automation. ADDITIONAL EXPERIENCE

Accenture Bangalore, India SENIOR SOFTWARE ENGINEER

• Delivered Microsoft and Java-based solutions across full project life cycles from requirements and design through production deployment, testing, and ongoing support. EDUCATION

MASTER OF COMPUTER APPLICATIONS (MCA) Bharathidasan University Tiruchirappalli, India



Contact this candidate