
Staff DevOps Engineer/SRE
- ₹40L – ₹70L
- |
- |6 years of exp
- |Full Time
In office
Not Available

About the job
Roles and Responsibility
Own Reliability & Architecture:
Design and evolve the infrastructure backbone for our AI and PaaS platform
Build highly available, fault-tolerant, and scalable systems
Define and drive SRE practices (SLIs, SLOs, error budgets)
Build Infrastructure at Scale:
Lead Infrastructure as Code using Pulumi
Own and scale Kubernetes clusters and containerized workloads
Standardize and automate infrastructure for global deployments
CI/CD & Automation:
Design and scale CI/CD pipelines for fast, reliable releases
Build self-healing systems and automated remediation workflows
Drive GitOps and platform engineering practices
Observability & Performance:
Implement end-to-end observability using VictoriaMetrics and Grafana (metrics, logs, traces)
Identify and resolve performance bottlenecks (latency, throughput, cost)
Lead incident response, root cause analysis, and postmortems
Leadership & Collaboration:
Partner with backend, AI, runtime, and security teams
Guide infrastructure decisions and scaling strategy
Mentor engineers and raise the bar on reliability and engineering standards
Security & Resilience:
Embed security into infrastructure and deployment workflows
Design for resilience (disaster recovery, chaos testing, capacity planning)
About the company
Similar Jobs








