Sandeep Vangala
Staff Site Reliability Engineer | Platform Engineering | Cloud Infrastructure & Security
Charlotte, NC (901) 605-8536 sandeepreddyvv@gmail.com LinkedIn sandeepvangala.devProfessional Summary
Staff Site Reliability Engineer with 10+ years of experience building secure, reliable cloud platforms across GCP, AWS, Kubernetes/OpenShift and Terraform. Known for applying automation and AI to transform complex infrastructure and operational challenges into scalable platforms, intelligent self-service and secure engineering workflows.
Experience
Staff Site Reliability Engineer | Intuit Credit Karma
- Architected an event-driven self-service platform using Next.js, Django/Python, Celery and GCP Pub/Sub to automate Terraform workspaces, Kubernetes resources and cloud projects; reduced fulfillment turnaround by ~60% for pre-approved requests.
- Built a VPC Service Controls remediation workflow that correlates GCP logs, guides troubleshooting in Slack, creates Jira requests and generates remediation pull requests to reduce manual handoffs.
- Engineered a threat landscape platform combining trust-model data and external threat intelligence, with AI-assisted attack-path analysis and Google ADK-based project/resource investigation.
- Established Kubernetes and GCP policy-as-code guardrails with OPA Gatekeeper (Rego), Terraform Sentinel and time-bound, auditable exception workflows.
- Developed an IaC pull-request review agent that evaluates internal standards, infrastructure context, risk signals and vulnerability data earlier in the secure SDLC.
- Built scheduled Python ETL pipelines to normalize trust-model and threat-intelligence data from MITRE ATT&CK, Anomali and other sources into SQL with daily refreshes.
- Designed a reusable Google ADK + MCP agent framework integrating New Relic and Splunk for platform/security troubleshooting and self-service support.
Principal Infrastructure Engineer | Discover Financial Services
- Led design and operations of AWS/OpenShift infrastructure for internal web and mobile services; automated cluster provisioning across AWS and vSphere with Terraform, Ansible and Jenkins.
- Built Go-based OpenShift operators with Operator SDK for service onboarding and on-demand cloud resource provisioning through Kubernetes-native APIs.
- Implemented Red Hat Advanced Cluster Management and GitOps governance across the OpenShift fleet for centralized lifecycle and configuration consistency.
- Engineered readiness/liveness validation in Go with GoBDD and Terratest before production handoff to reduce configuration-related risk.
Senior DevOps Engineer | Idea Evolver
- Stabilized Docker and GKE delivery environments by applying repeatable CI/CD standards and platform controls.
- Built Terraform + Helm automation for GCP infrastructure and Kubernetes deployments, replacing ad hoc console changes with version-controlled provisioning.
- Integrated Google Secret Manager into CI/CD workflows to centralize credentials and reduce secret exposure.
Lead Release Engineer | Hilton Worldwide
- Led Bamboo-to-GitLab CI/CD migration across application portfolios, consolidating delivery tooling and standardizing release workflows.
- Built GitLab pipelines and deployment automation for OpenShift microservices and static assets on Akamai NetStorage and Amazon S3.
- Supported production reliability and performance using Dynatrace, Datadog and Splunk; implemented synthetic monitoring for critical application journeys.
- Built performance-testing automation for microservices with Python/Grinder and helped drive monolith-to-microservices delivery practices using 12-factor principles.