Sharavan Kumar Puppala
Professional Summary
Site Reliability Engineer with 5+ years of experience designing, automating, and operating highly available cloud-native platforms for financial services and healthcare. Expertise in Kubernetes, AWS (EKS), Terraform, observability (Prometheus, Grafana, Splunk), SLO/SLI engineering, incident response, and infrastructure automation to reduce toil, improve availability, and accelerate safe delivery.
Technical Skills
Work Experience
- Owned reliability and availability for 20+ production payment services; drove reliability initiatives and runbook standardization to align platform behaviour with business SLAs.
- Led incident response, RCA, and postmortems across a 6-member SRE group; standardized troubleshooting playbooks and reduced MTTR by 35%.
- Redesigned Splunk and Dynatrace alerting across 40+ microservices; implemented noise reduction and actionable thresholds, cutting alert fatigue by 25% and improving signal quality.
- Partnered with engineering and product teams to define SLOs, SLIs, and error budgets for 12 business-critical services, enabling measurable reliability targets tied to customer experience.
- Implemented Blue-Green deployment patterns on EKS using Helm and CI pipelines to eliminate deployment-related production incidents and increase deployment confidence.
- Standardized AWS infrastructure provisioning across environments with Terraform and CloudFormation and delivered Python/Bash automation that reduced operational toil and accelerated diagnostics.
- Automated provisioning and environment management using Terraform and CloudFormation across dev/QA/prod, reducing environment deployment times by 65% and improving consistency.
- Administered and tuned Amazon EKS clusters hosting 50+ containerized microservices; implemented scaling and resiliency patterns to achieve 99.95% platform availability.
- Designed and maintained CI/CD pipelines with Jenkins, GitHub Actions, and SonarQube to reduce release cycles by 40% and increase deployment success rates.
- Implemented observability across services using Prometheus, Grafana, CloudWatch, and Splunk; reduced incident detection time by 45% and improved operational visibility.
- Led cloud cost optimization using AWS Cost Explorer and rightsizing strategies to cut infrastructure spend by 22% while maintaining performance SLAs.
- Hardened cloud security posture by implementing IAM least-privilege policies and Secrets Manager integration, lowering security audit findings by 35%.
- Supported and operated AWS production environments serving enterprise applications used by 100,000+ end users; maintained 99.9% service availability.
- Implemented Infrastructure as Code using Terraform and CloudFormation, cutting provisioning effort by 60% and standardizing deployments.
- Built and enhanced CI/CD pipelines with Jenkins and Bitbucket to shorten deployment windows by 50% and increase release frequency.
- Developed Python and Bash automation to eliminate repetitive operational tasks, saving approximately 30 hours of manual effort per month.
- Configured monitoring and alerting with CloudWatch and Splunk to enable proactive detection, reducing critical production incidents by 25%.
- Performed Linux system administration and environment maintenance across 100+ servers and collaborated with dev/QA to reduce recurring production defects by 35%.
Education
Powered by Drivetube · Create your own profile at drivetube.ai
Explore Drivetube
- Drivetube Profile — your free digital resume — at drivetube.ai/in/your-name: one true standard resume with a Hiring Snapshot (visa status, expected salary, notice period, work preference, relocation), an ATS-ready PDF download and a single shareable link. Free forever; interview requests come from verified employers and your contact details stay masked until you accept. Documentation.
- Free Job Board — verified openings crawled ATS-by-ATS from 100,000+ real company career pages across 35 ATS platforms. Shows the true posting date from the source ATS — not when a listing was indexed — and deletes every general listing 3 days after it was actually posted. No ghost jobs, no ad-sponsored listings, no staffing reposts, no account needed. Documentation.
- Job Hunt Program — managed job hunting, a one-time purchase from $199.99. JobScout matches verified roles to your real experience band, Blend AI writes a uniquely tailored resume and cover letter for every application, and the Autofill extension fills the form — or Let Us Apply submits it for you. Documentation.
- Resume Writing Services — human-written, ATS-optimised resumes by senior career writers, from ₹499.99 / $25.99. Available in every country, written to the destination country's own standard — a US resume, UK CV, German Lebenslauf and Indian resume are genuinely different documents. A paid service, separate from the free Drivetube Profile. Documentation.
- Community Membership — from $4.99/month (₹1,999/year in India). Unlocks the gated job-board filters — visa sponsorship, security clearance, workplace and application time — plus Job-Scout AI matching, Resume Report AI, Interview AI prep sheets, Recruiter Outreach AI sent from your own Gmail, Apply or Skip triage, a daily market feed and a $10,000+ library including 23 ATS-validated resume templates. Documentation.
- Drivetube Hire — for employers — hiring with no job postings and no applications. Paste your real job description and AI matches it against candidates' true standard resumes, returning ranked candidates with a match %, matched and missing skills and written reasoning. Free tier included; employers pay, candidates never do. Documentation.
Full product documentation — every product explained, with feature-by-feature comparisons against the job boards, AI apply tools, resume services and hiring platforms people actually use.
The job board covers the United States, India, United Kingdom, Canada, Europe and Australia, and Resume Writing Services are available in every country.