Cloud-Native Reliability
Production Kubernetes, Karpenter, Helm, cluster security, rollout safety, and performance tuning for high-availability workloads.
Site Reliability / DevOps Engineer
I build cloud-native infrastructure, GitOps delivery, observability systems, and agentic automation that reduce production risk, accelerate RCA, and make engineering workflows easier to operate.
Platform Focus
Production Kubernetes, Karpenter, Helm, cluster security, rollout safety, and performance tuning for high-availability workloads.
Agentic workflows that connect infrastructure tools, CI systems, and developer operations with traceable LLM observability.
New Relic, Dynatrace, Prometheus, Grafana, Langfuse, incident review loops, and production dashboards built for fast diagnosis.
Work Experience
Feb 2023 - Present / Intern to Site Reliability Engineer I to Site Reliability Engineer II
Selected Work
Led a zero-downtime migration strategy with phased traffic validation, rollback safety, and full data consistency across critical messaging streams.
Directed workload compatibility checks, performance validation, and cluster right-sizing to improve efficiency and reduce infrastructure spend by 30%.
Built a hackathon-winning orchestration layer that helps debug CI/CD and infrastructure workflows through connected platform tools.
Technical Skills
Technical Writing
A production migration case study covering blue-green architecture, canary rollout phases, rollback strategy, and observability validation.
Read articleA technical guide to zero-downtime EKS migration, compatibility validation, performance improvements, and cloud cost optimization.
Read articleA CI/CD automation writeup for mobile certificate management that removes manual bottlenecks and reduces expired credential risk.
Read articleEducation & Recognition
Open to SRE / Platform Engineering roles
Bangalore, Hyderabad, or cloud-native teams building serious production systems.