Senior Site Reliability Engineer
Xebia · Gurgaon
What the role covers
Xebia is looking for experienced Site Reliability Engineers (SREs) to help drive reliability, scalability, performance, and resilience across technology platforms. Key Responsibilities include design and implement highly available, scalable, and fault-tolerant systems, define and drive SLOs, SLAs, and reliability engineering practices, build automation-first, self-healing platforms and operational processes, lead observability, monitoring, incident management, and continuous improvement, work across AWS, Azure, GCP, Kubernetes, and Infrastructure as Code (IaC), and drive engineering best practices across reliability, performance, and operational excellence. What we are looking for: 10+ years of experience across SRE, DevOps, Platform Engineering, or Software Engineering, strong expertise in distributed systems and cloud architecture, hands-on experience with Kubernetes, Terraform, Prometheus, Grafana, and modern observability platforms, and strong problem-solving, technical leadership, and stakeholder management capabilities.
At a glance
- Employment
- Full-time
- Experience
- 10+ years
- Notice period
- Not stated
- Work mode
- Not Specified
- Indicative pay
- Not specified (market estimate, not an offer)
- Posted
- 1 wk ago
- Hiring contact
getjob