Valorant
Website:
valorant.com
Job details:
Site Reliability Engineer (SRE)
Experience: 3–5 years
Location: Bengaluru — Hybrid
Type: Full-time
Reports to: Tech Lead / EM
About the role
You'll keep our SaaS platform fast, available, and observable as we scale. This is an engineering-first SRE role — you'll spend more time automating reliability than firefighting, and you'll set the standards for how we run services in production.
What you'll do
• Define and track SLOs/SLIs and error budgets with product and engineering teams.
• Build and own observability: metrics, logging, tracing, dashboards, and alerting (e.g., Prometheus,
Grafana, Datadog, OpenTelemetry).
• Drive incident management — on-call rotation, runbooks, blameless postmortems, and
follow-through on action items.
• Own infrastructure as code and automation (Terraform), container orchestration on AWS EKS and
ECS with Helm, and CI/CD reliability.
• Improve performance, capacity planning, and cost efficiency of cloud infrastructure.
• Partner with security to maintain a hardened, compliant production environment.
What we're looking for
• 3–5 years in SRE, DevOps, or production-focused software engineering.
• Hands-on with AWS EKS (Kubernetes) and ECS for container orchestration.
• Strong with Helm for packaging and deploying services.
• Strong IaC (Terraform) and CI/CD pipeline experience on AWS.
• Solid observability tooling experience and a track record of reducing MTTR.
• Comfortable with automation/scripting in Go, Python, or Bash.
• Has carried production on-call and run real incident response.
Bonus
• Reliability for multi-tenant SaaS at scale.
• Service mesh, progressive delivery (Argo, Flagger), or chaos engineering.
• Exposure to security/compliance frameworks (SOC 2, ISO 27001).
Click on Apply to know more.