KIRAN BABU LACHI
Platform Engineer / DevOps / SRE • Hyderabad, India • l************@gmail.com • 949****210 • drivetube.ai/•••••
Professional Summary
Platform Engineer with 10+ years of experience designing, administering, and automating OpenShift/Kubernetes platforms, CI/CD pipelines, cloud infrastructure, and observability for enterprise production systems.
Technical Skills
Cloud and DevOps: Kubernetes,AWS,Azure,IBM Cloud,Jenkins,Argo CD,Spinnaker,Maven,Docker,Podman,Red Hat Quay,Terraform,Linux
Tools and Methodologies: Red Hat OpenShift 3.x,Red Hat OpenShift 4.x,oc CLI,kubectl,Git,Bitbucket
Observability & Monitoring: Prometheus,Grafana,Datadog,New Relic,Tivoli Netcool
Infrastructure as Code & Automation: Ansible Tower,Azure Runbooks,Bash scripting,YAML
Security & Access: CyberArk,RBAC,Secrets management
Operating Systems & Version Control: UNIX,Windows
Incident & Release Management: On-call production support,Alerting and incident integrations,Change and release coordination
Work Experience
Cognizant Technology Solutions Pvt. Ltd.
Hyderabad, India
Platform Engineering & DevOps Engineer
June 2024 – Present
Worked at an IT services firm delivering platform engineering and DevOps for enterprise applications running on OpenShift and cloud infrastructure.
Tech Stack: Red Hat OpenShift, Kubernetes, Jenkins, Argo CD, Prometheus, Grafana, CyberArk
- Administered and automated Red Hat OpenShift 3.x/4.x clusters to support enterprise workloads, implementing configuration standardization and lifecycle procedures for upgrades and patching.
- Designed and maintained CI/CD pipelines with Jenkins and Argo CD to enable GitOps-driven deployments and reduce manual release steps for developer teams.
- Built internal developer self-service tools and platform automation to streamline application onboarding and access to platform resources.
- Implemented platform monitoring and troubleshooting workflows using Prometheus and Grafana to surface health metrics and reduce incident resolution friction.
- Integrated CyberArk and RBAC controls into platform access processes to enforce secrets management and least-privilege access for application teams.
- Collaborated with cross-functional teams to define SLOs and runbooks, improving platform availability and standardizing incident response procedures.
Cognizant Technology Solutions Pvt. Ltd.
Hyderabad, India
Site Reliability Engineer
Aug 2023 – Apr 2024
Provided SRE support at an IT services company for clients' production and non-production cloud-native platforms focusing on reliability and scalability.
Tech Stack: Prometheus, Grafana, Terraform, Kubernetes, OpenShift
- Managed day-to-day SRE operations for container platforms, focusing on reliability, capacity planning, and platform scalability.
- Created and tuned Prometheus alerting rules and Grafana dashboards to reduce alert fatigue and prioritize actionable incidents.
- Integrated incident management and on-call tooling to streamline triage and reduce mean time to recovery (MTTR).
- Provisioned and decommissioned infrastructure with Terraform to maintain consistent environment configurations across cloud providers.
- Authored runbooks and post-incident reports to capture remediation steps and prevent repeat incidents.
- Coordinated platform upgrades and patch cycles with application teams to minimize service disruption and maintain compliance.
Encora Innovation Labs Pvt. Ltd.
Hyderabad, India
Site Reliability Engineer
Oct 2022 – Jul 2023
Supported production and non-production environments for software engineering teams, implementing observability and cloud provisioning on AWS and Azure.
Tech Stack: Prometheus, Grafana, Terraform, AWS EC2, Azure VMs, Azure Runbooks
- Owned availability and operational support for production and non-production environments, providing 24x7 on-call coverage and incident remediation.
- Implemented observability stacks using Prometheus and Grafana to provide application and infrastructure metrics for engineering teams.
- Provisioned AWS EC2 instances and Azure VMs via Terraform to create repeatable environment templates and speed environment provisioning.
- Automated patching and routine maintenance tasks using Azure Runbooks to reduce manual intervention and maintenance windows.
- Collaborated with development teams to optimize applications for container deployment and improved platform resource utilization.
- Documented environment configurations and standard operating procedures to reduce onboarding time for new SRE engineers.
Kyndryl India Pvt. Ltd.
Hyderabad, India
Red Hat OpenShift Administrator / DevOps Engineer
Sep 2021 – Oct 2022
Delivered OpenShift cluster administration and DevOps support for enterprise clients as part of managed infrastructure services.
Tech Stack: Red Hat OpenShift, Kubernetes, Jenkins, Git, Bitbucket, Linux
- Administered Red Hat OpenShift 3.x and 4.x clusters, performing cluster health checks, capacity monitoring, and routine maintenance.
- Planned and executed cluster upgrades and node maintenance with minimal disruption to hosted applications.
- Configured and maintained Jenkins-based CI pipelines to automate build and deployment processes for multiple application teams.
- Supported source control workflows using Git and Bitbucket, enabling consistent branch and release strategies across projects.
- Troubleshot platform and application issues across Linux hosts and container runtimes to restore service availability.
- Implemented backup and recovery practices for OpenShift resources and persistent storage to meet recovery objectives.
IBM India Pvt. Ltd.
Hyderabad, India
DevOps Engineer / Production Support Analyst
May 2015 – Aug 2021
Provided production support and DevOps services for large-scale Linux and OpenShift environments within an IT services organization.
Tech Stack: OpenShift, Linux, Jenkins, Monitoring tools, Batch scheduling
- Supported large-scale Linux and OpenShift production environments, handling incidents, problem management, and root cause analysis.
- Automated build and release processes to improve deployment consistency and reduce manual errors across application teams.
- Managed batch job scheduling, reruns, and failure analysis to ensure timely processing of business workloads.
- Coordinated incidents, releases, and change activities with stakeholders to maintain service levels and compliance.
- Developed monitoring and alerting improvements to proactively detect infrastructure and application degradation.
- Recognized with Best Performer (2017) and Top Performer of the Pool (2019) for consistent delivery and operational excellence.
HSBC EDPI Pvt. Ltd.
Hyderabad, India
Application Support Engineer
Jan 2015 – May 2015
Provided application support for internal banking systems, resolving performance and database-related issues.
Tech Stack: Application monitoring, Databases, Incident management
- Provided application support for internal banking systems, investigating and resolving latency and database issues.
- Liaised with development teams to reproduce defects and coordinate timely fixes and deployments.
- Performed root cause analysis for recurring incidents and recommended configuration changes to improve stability.
- Maintained incident logs and communicated status to stakeholders during production outages.
- Supported deployment and rollback activities during release windows to ensure minimal business impact.
- Ensured adherence to SLAs and operational procedures for banking application support.
Sutherland Global Services Pvt. Ltd.
Hyderabad, India
Operations Analyst
Oct 2013 – Dec 2014
Delivered application and operations support for order management systems in a business process outsourcing environment.
Tech Stack: Order management systems, Macros, SLA monitoring
- Provided application support for order management systems and assisted customers with order placement and tracking.
- Developed and used macros to automate bulk order processing, reducing manual workload and error rates.
- Monitored SLA adherence and escalated operational issues to maintain customer satisfaction.
- Collaborated with cross-functional teams to resolve process issues and improve operational workflows.
- Documented operational procedures and created knowledge base articles to support training and handovers.
- Analyzed recurring support trends and suggested process improvements to reduce repeat incidents.
Education
St. Ann’s Engineering College, JNTU Kakinada
Bachelor of Technology (B.Tech) • Kakinada, India • 2013
Certifications
Red Hat OpenShift Administration (DO280) — Red Hat
Powered by Drivetube · Create your own profile at drivetube.ai