Resume

A focused, one-page view of my experience.

Read the latest version below or download the ATS-ready PDF.

Sandeep Vangala

Staff Site Reliability Engineer | Platform Engineering | Cloud Infrastructure & Security

Charlotte, NC (901) 605-8536 sandeepreddyvv@gmail.com LinkedIn sandeepvangala.dev

Professional Summary

Staff Site Reliability Engineer with 10+ years of experience building secure, reliable cloud platforms across GCP, AWS, Kubernetes/OpenShift and Terraform. Known for applying automation and AI to transform complex infrastructure and operational challenges into scalable platforms, intelligent self-service and secure engineering workflows.

Experience

Staff Site Reliability Engineer | Intuit Credit Karma

Charlotte, NC | Aug 2022 - Present

  • Architected an event-driven self-service platform using Next.js, Django/Python, Celery and GCP Pub/Sub to automate Terraform workspaces, Kubernetes resources and cloud projects; reduced fulfillment turnaround by ~60% for pre-approved requests.
  • Built a VPC Service Controls remediation workflow that correlates GCP logs, guides troubleshooting in Slack, creates Jira requests and generates remediation pull requests to reduce manual handoffs.
  • Engineered a threat landscape platform combining trust-model data and external threat intelligence, with AI-assisted attack-path analysis and Google ADK-based project/resource investigation.
  • Established Kubernetes and GCP policy-as-code guardrails with OPA Gatekeeper (Rego), Terraform Sentinel and time-bound, auditable exception workflows.
  • Developed an IaC pull-request review agent that evaluates internal standards, infrastructure context, risk signals and vulnerability data earlier in the secure SDLC.
  • Built scheduled Python ETL pipelines to normalize trust-model and threat-intelligence data from MITRE ATT&CK, Anomali and other sources into SQL with daily refreshes.
  • Designed a reusable Google ADK + MCP agent framework integrating New Relic and Splunk for platform/security troubleshooting and self-service support.

Principal Infrastructure Engineer | Discover Financial Services

Remote / Memphis, TN | Oct 2020 - Aug 2022

  • Led design and operations of AWS/OpenShift infrastructure for internal web and mobile services; automated cluster provisioning across AWS and vSphere with Terraform, Ansible and Jenkins.
  • Built Go-based OpenShift operators with Operator SDK for service onboarding and on-demand cloud resource provisioning through Kubernetes-native APIs.
  • Implemented Red Hat Advanced Cluster Management and GitOps governance across the OpenShift fleet for centralized lifecycle and configuration consistency.
  • Engineered readiness/liveness validation in Go with GoBDD and Terratest before production handoff to reduce configuration-related risk.

Senior DevOps Engineer | Idea Evolver

Remote / Memphis, TN | Jul 2020 - Oct 2020

  • Stabilized Docker and GKE delivery environments by applying repeatable CI/CD standards and platform controls.
  • Built Terraform + Helm automation for GCP infrastructure and Kubernetes deployments, replacing ad hoc console changes with version-controlled provisioning.
  • Integrated Google Secret Manager into CI/CD workflows to centralize credentials and reduce secret exposure.

Lead Release Engineer | Hilton Worldwide

Memphis, TN | Feb 2016 - Jun 2020

  • Led Bamboo-to-GitLab CI/CD migration across application portfolios, consolidating delivery tooling and standardizing release workflows.
  • Built GitLab pipelines and deployment automation for OpenShift microservices and static assets on Akamai NetStorage and Amazon S3.
  • Supported production reliability and performance using Dynatrace, Datadog and Splunk; implemented synthetic monitoring for critical application journeys.
  • Built performance-testing automation for microservices with Python/Grinder and helped drive monolith-to-microservices delivery practices using 12-factor principles.