Live opening · Posted 6 days ago

Site Reliability Engineer

Quess Corp Limited · Hyderabad, Telangana, India (Hybrid)
Linkedin No
You are 6 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 6 days ago
CompanyQuess Corp Limited
LocationHyderabad, Telangana, India (Hybrid)
Work modeNo
SourceLinkedin
Listed6 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
8 min from Linkedin publishing this role to us finding it
8 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
16,120 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Key Responsibilities
Set up and tune alerts and dashboards (ELK, Grafana, Dynatrace) to catch issues before they become customer-facing incidents
Define and monitor SLAs/SLOs for the POS platform, including telemetry setup for key services
Reduce alert noise and false positives through systematic tuning and threshold refinement
Partner with development and platform teams to instrument new services for observability from day one
Perform root cause analysis on monitoring/alerting gaps that contributed to missed or delayed incident detection
Document alerting logic, dashboards, and SLO definitions; contribute to the team's observability runbook
Participate in the team's weekend and holiday roster as part of a rotating shift schedule
Required Skills & Experience
Hands-on experience with observability tooling – ELK, Grafana, Dynatrace, or equivalent – including alert configuration and tuning
Working knowledge of SLA/SLO definition and reliability monitoring practices
Scripting ability (Python, Bash, or similar) for automation and telemetry setup
Comfortable working with ticketing/ITSM tools (ServiceNow, Jira) and log/monitoring platforms (Kibana, Splunk, or Grafana/Dynatrace)
Strong technical root cause analysis (RCA) skills, with a structured, evidence-based approach to problem-solving
Willingness and availability to work weekends as part of a team roster/rotation
Ability to work under pressure during live incidents while maintaining a calm, methodical approach
Preferred / Added Advantage
Prior experience building observability practice from scratch on a growing platform
Exposure to cloud platforms and containerized/microservices architectures
Prior subject-matter knowledge of Point of Sale (POS) systems or retail infrastructure
ITIL Foundation certification or equivalent exposure to incident/problem management practices
Technical Stack
Languages
Java, Python, React.js
Messaging
RabbitMQ
Database
SQL, Mongo
Monitoring
Dynatrace, ELK, Splunk, Grafana
ITSM
ServiceNow, JIRA
Cloud
GCP and Azure
Domain
POS transaction flows, payments, loyalty, receipting, store hardware

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App