Bengaluru, India · open to onsite / remote
Eknatha Reddy Puli
Senior Staff Engineer · Platform Engineering & SRE
Fourteen years keeping enterprise infrastructure alive in production — Linux, virtualization, Kubernetes and OpenStack, most of it on call. I build the paved road: automation that removes the repetitive work, and observability that surfaces failure before a customer does.
About
I design, operate and troubleshoot mission-critical platforms — the infrastructure other teams depend on and rarely think about until it breaks.
Fourteen years across enterprise infrastructure, virtualization, Linux and production support, from carrier networks at Ericsson to ad-tech at scale to hybrid cloud at IBM. The through-line is reliability: making systems legible, making failure recoverable, and making the fix repeatable so nobody has to remember it at 3am.
Day to day that means Kubernetes, Terraform, OpenStack and cloud platforms across AWS, Azure and GCP — automating repetitive operations, simplifying how infrastructure gets managed, and giving engineering teams a path to production that doesn't route through a ticket queue. Currently going deeper on Go and Google Cloud.
Outside of work I make practical visual guides on Kubernetes, Linux, Docker, Git, Terraform and cloud infrastructure, for engineers trying to get a grip on concepts that are usually explained badly.
At a glance
Operating numbers
- years.infrastructure.engineering
- 14+
- primary.clouds
- AWS · GCP · Azure · hybrid
- orchestration
- Kubernetes · OpenStack
- base
- Bengaluru, India · onsite / remote
Service history
Seven teams since 2010
Newest first, all expanded. Collapse any role you want out of the way.
Toolchain
What I run
Listed honestly — what I've operated at scale, and what I'm still growing into.
Orchestration
- Kubernetes
- OpenShift
- Helm
- Docker
- OpenStack
Infrastructure as code
- Terraform
- OpenTofu
- Ansible
- Packer
Cloud
- AWS
- Azure
- Hybrid & private cloud
- GCP — working knowledge
Databases
- MySQL
- PostgreSQL
- Oracle
- Redis
Systems
- Linux internals
- Networking
- Storage
- Capacity & performance
Reliability
- SLI / SLO / error budgets
- Incident command
- Prometheus
- Grafana
- Splunk
Automation
- Python
- Bash
- Git
- CI/CD pipelines
- Go — learning
Shell access
Ask me directly
Everything here is generated from the same data as the rest of the page. Type help, or tap a command.
Credentials
Certifications
Notes from production
Dogfooding
This site's own health
Same instincts, smaller blast radius. Build metadata comes from the deploy; load time is measured in your browser right now.
Get in touch
Tell me what's breaking.
LinkedIn is the fastest way to reach me. I reply to anything specific about platform engineering, reliability or Kubernetes work.