Sharavan Kumar Puppala
Site Reliability Engineer II • Maryland Heights, MO, 63043 • s****************@gmail.com • +18******536 • drivetube.ai/•••••
Professional Summary
Site Reliability Engineer with 5+ years of experience designing, automating, and operating highly available cloud-native platforms for financial services and healthcare. Expertise in Kubernetes, AWS (EKS), Terraform, observability (Prometheus, Grafana, Splunk), SLO/SLI engineering, incident response, and infrastructure automation to reduce toil, improve availability, and accelerate safe delivery.
Technical Skills
Programming Languages: Python,Bash
Cloud and DevOps: AWS,EKS,EC2,Lambda,S3,RDS,SNS,CloudWatch,Secrets Manager,Cost Explorer,Kubernetes,Docker,Helm,ECS,Terraform,Ansible,Jenkins,GitHub Actions,XL Release,SonarQube,Linux
Tools and Methodologies: Bitbucket,Git,GitHub
Infrastructure as Code: CloudFormation
Observability & Monitoring: Prometheus,Grafana,Dynatrace,Splunk,OpenTelemetry,PagerDuty
Reliability Engineering: SLO,SLI,Error Budgets,Incident Response,Root Cause Analysis,Postmortems,Capacity Planning,Disaster Recovery,High Availability
Networking & Systems: TCP,IP,DNS,Load Balancers,VPC,Route 53
Security & Governance: IAM,Least-Privilege Access,Secrets Management,Compliance Support
Work Experience
Mastercard
Saint Louis, MO
Site Reliability Engineer II
Jun 2025 – Present
Worked on SRE for global payment processing platforms, supporting production reliability and operational excellence for transaction services.
Tech Stack: AWS, EKS, Kubernetes, Terraform, CloudFormation, Splunk, Dynatrace, Prometheus, Grafana, Python, Bash, Helm, PagerDuty, Git
- Owned reliability and availability for 20+ production payment services; drove reliability initiatives and runbook standardization to align platform behaviour with business SLAs.
- Led incident response, RCA, and postmortems across a 6-member SRE group; standardized troubleshooting playbooks and reduced MTTR by 35%.
- Redesigned Splunk and Dynatrace alerting across 40+ microservices; implemented noise reduction and actionable thresholds, cutting alert fatigue by 25% and improving signal quality.
- Partnered with engineering and product teams to define SLOs, SLIs, and error budgets for 12 business-critical services, enabling measurable reliability targets tied to customer experience.
- Implemented Blue-Green deployment patterns on EKS using Helm and CI pipelines to eliminate deployment-related production incidents and increase deployment confidence.
- Standardized AWS infrastructure provisioning across environments with Terraform and CloudFormation and delivered Python/Bash automation that reduced operational toil and accelerated diagnostics.
Humana
Chicago, IL
Cloud Engineer
Jan 2024 – Jun 2025
Supported cloud platforms for healthcare applications, focusing on AWS infrastructure, observability, CI/CD, cost optimization and compliance.
Tech Stack: AWS, EKS, Terraform, CloudFormation, Jenkins, GitHub Actions, SonarQube, Prometheus, Grafana, CloudWatch, Splunk, Cost Explorer, IAM, Secrets Manager, Python, Bash
- Automated provisioning and environment management using Terraform and CloudFormation across dev/QA/prod, reducing environment deployment times by 65% and improving consistency.
- Administered and tuned Amazon EKS clusters hosting 50+ containerized microservices; implemented scaling and resiliency patterns to achieve 99.95% platform availability.
- Designed and maintained CI/CD pipelines with Jenkins, GitHub Actions, and SonarQube to reduce release cycles by 40% and increase deployment success rates.
- Implemented observability across services using Prometheus, Grafana, CloudWatch, and Splunk; reduced incident detection time by 45% and improved operational visibility.
- Led cloud cost optimization using AWS Cost Explorer and rightsizing strategies to cut infrastructure spend by 22% while maintaining performance SLAs.
- Hardened cloud security posture by implementing IAM least-privilege policies and Secrets Manager integration, lowering security audit findings by 35%.
Infosys
Hyderabad, India
Cloud Engineer
Sep 2019 – Jul 2022
Delivered cloud engineering and platform support for enterprise AWS environments hosting customer-facing applications with large user bases.
Tech Stack: AWS, Terraform, CloudFormation, Jenkins, Bitbucket, CloudWatch, Splunk, Python, Bash, Linux, Docker, Kubernetes
- Supported and operated AWS production environments serving enterprise applications used by 100,000+ end users; maintained 99.9% service availability.
- Implemented Infrastructure as Code using Terraform and CloudFormation, cutting provisioning effort by 60% and standardizing deployments.
- Built and enhanced CI/CD pipelines with Jenkins and Bitbucket to shorten deployment windows by 50% and increase release frequency.
- Developed Python and Bash automation to eliminate repetitive operational tasks, saving approximately 30 hours of manual effort per month.
- Configured monitoring and alerting with CloudWatch and Splunk to enable proactive detection, reducing critical production incidents by 25%.
- Performed Linux system administration and environment maintenance across 100+ servers and collaborated with dev/QA to reduce recurring production defects by 35%.
Education
California Lutheran University
MS in Management Science • Thousand Oaks, CA • Aug 2022 – Nov 2023
Osmania University
BE in Computer Science • Hyderabad, India • Jun 2016 – May 2019
Powered by Drivetube · Create your own profile at drivetube.ai