Live opening · Posted 7 days ago

Site Reliability Engineer

Recro · Bengaluru, Karnataka, India (On-site)
Linkedin No
You are 7 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 7 days ago
CompanyRecro
LocationBengaluru, Karnataka, India (On-site)
Salary600K INR/yr - 700K INR/yr
Work modeNo
SourceLinkedin
Listed7 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
10 min from Linkedin publishing this role to us finding it
15 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
62,526 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Site Reliability Engineer
Experience: 3+ years
Location: Bangalore
Shift: Rotational shifts
Role Overview
We are looking for an SRE to ensure the reliability, scalability, availability, and observability of high-traffic production systems. The role involves infrastructure automation, Kubernetes operations, monitoring, incident management, and continuous improvement of production reliability.
Key Responsibilities
Manage and troubleshoot AWS/GCP infrastructure and Kubernetes environments in production.
Work with Kubernetes, Redis, Kafka, Solr/Elasticsearch and related infrastructure components.
Build and maintain CI/CD pipelines, Terraform/Helm-based infrastructure, and automation.
Implement monitoring and alerting using Prometheus, Grafana, ELK/Loki and distributed tracing tools.
Participate in rotational on-call and shifts, handling P1/P2 incidents, troubleshooting, RCA, and post-incident reviews.
Develop SLIs/SLOs, alerts, dashboards, runbooks, and reliability improvements.
Automate repetitive operational tasks using Python, Shell/Bash, or Go to reduce manual toil.
Work on capacity planning, autoscaling, performance optimization, security patching, and cost optimization.
Troubleshoot Linux, networking, DNS, TCP/IP, load balancing, TLS/HTTPS, and application/infrastructure issues.
Drive preventive actions through RCA, automation, self-healing, and improved deployment/recovery processes.
Must-Have Skills
3+ years of experience in SRE / DevOps / Infrastructure Engineering
Experience with high-traffic or large-scale production environments
Hands-on Kubernetes in production
Strong experience with AWS or GCP
Terraform and Infrastructure as Code (IaC)
Docker and CI/CD
Prometheus & Grafana
Linux administration and troubleshooting
Python / Bash / Shell scripting
Production incident management, RCA and on-call experience
Good understanding of networking fundamentals
Good to Have
Helm, ArgoCD/GitOps
ELK/Loki
Redis, Kafka, Solr/Elasticsearch
SLI/SLO, SLA and error-budget concepts
Ansible
Distributed tracing / OpenTelemetry
Note: This is a rotational-shift/on-call role, so candidates should be comfortable supporting production systems across different shifts.

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App