Open to Staff / Senior Backend, Platform Engineering & SRE roles · Bangalore · Hyderabad · Pune · Remote
I'm a Senior Software Engineer with 10+ years of experience building backend services, distributed systems, and the Kubernetes platforms other teams deploy on. My recent work is multi-cluster platform engineering at fleet scale — control-plane design, GitOps architecture, and deployment workflows operating across 500+ Kubernetes clusters.
I enjoy solving complex engineering problems around scalability, reliability, concurrency, infrastructure automation, and distributed systems, in Go and Kubernetes.
- 500+ Kubernetes clusters · 60+ applications migrated with zero downtime — designed the GitOps architecture and platform workflows operating across hundreds of workload clusters, migrating 60+ applications onto it without a single service interruption.
- Multi-cluster control plane — designed systems for securely onboarding and managing remote Kubernetes clusters at scale.
- 90% reduction in deployment wait time — designed parallel, non-blocking deployment strategies in place of serialized rollouts.
Ordered Upgrade Operator — A Kubernetes operator that upgrades a callee service before its caller, so in-flight requests don't fail mid-rollout; Go and a v1alpha1 CRD, proven by Kind end-to-end tests that terminate pods mid-rollout.
OTel Trace Propagation — Distributed tracing across the boundary where it usually breaks: a custom OpenTelemetry TextMapCarrier carries W3C trace context on RabbitMQ message metadata, so an HTTP request and a worker in a separate process land in one five-span trace — with a CI job that fails unless the trace spans both services.
Train Ticket Booking — A gRPC booking service in Go built around one invariant, never sell the same seat twice: seat allocation is a read-then-write, so a single write lock spans the whole decision, proven by 200 buyers contending for 20 seats under the race detector in CI.
I contribute to open-source infrastructure and cloud-native projects, with a focus on backend systems, Kubernetes, platform engineering and infrastructure automation.
Airship View Commits
- Contributed to the Airship Deckhand configuration management platform by enhancing the secret substitution engine to support one-to-many configuration propagation and designing a development-mode authentication framework that removed Keystone dependencies from local development workflows.
Openstack-helm-infra View Commits
- Extended the OpenStack-Helm ecosystem by contributing new infrastructure components, Kubernetes controller integrations, and automated integration tests, strengthening deployment automation and release quality.
Neutron-tempest-plugin View Commits
- Expanded OpenStack Neutron's integration test coverage, improving release quality.
Languages
Go Python Bash SQL
Cloud & Infrastructure
AWS Terraform Ansible Docker Kubernetes Linux
Kubernetes Ecosystem
Kubebuilder Operator SDK Cluster API Helm Argo CD Rancher CRDs & Controllers
Reliability Engineering
SLIs & SLOs Error Budgets Incident Response Postmortems Capacity Planning Progressive Delivery
Observability
Prometheus Grafana OpenTelemetry Loki Alertmanager Thanos
CI/CD & Delivery
GitHub Actions Argo Workflows Zuul GitOps
Data & Messaging
Kafka RabbitMQ PostgreSQL MongoDB Redis gRPC
- AI/LLM infrastructure — GPU scheduling, vLLM and inference platforms on Kubernetes
- MLOps & Kubernetes-native ML infrastructure
- Multi-cluster fleet management and internal platform APIs
- eBPF and Cilium for network-level observability
- SLO-driven reliability at scale — Thanos, long-retention metrics, alert quality
- Infrastructure cost efficiency and capacity modelling
I'm interested in conversations around:
Backend Engineering · Platform Engineering · Kubernetes · Distributed Systems · Cloud Infrastructure · Observability · AI Infrastructure


