Live opening · Posted 27 days ago

Senior Site Reliability Engineer

Zeta · Hyderabad
Instahyre 5-9 yrs
You are 27 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 27 days ago
CompanyZeta
LocationHyderabad
Experience5-9 yrs
SourceInstahyre
Listed27 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
0 min from Instahyre publishing this role to us finding it
8 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
16,808 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Responsibilities:
Work to understand any arising issues and overall application performance by enacting monitoring solutions.
Conduct consistent and thorough analysis of current systems and work to reduce the quantity of existing problems, suggesting new solutions to help upgrade and refine such systems.
Provide support across a broad range of areas, including monitoring, processes and tools, architecture, and root cause analysis.
Develop and maintain monitoring and alerting systems to proactively detect and resolve issues.
Automate routine tasks to improve system efficiency and reduce downtime.
Troubleshoot and resolve incidents and outages.
Develop and implement automation scripts and tools to improve the efficiency and effectiveness of system management tasks.
Identifying areas for improvement and designing solutions that are scalable, reliable, and easy to maintain.
Monitoring and acting on alerts to avoid production outages & incidents.
Upkeeping of Run books for the Alerts
Requirements:
4-6 years of sysadmin experience in handling large-scale distributed system software deployments in the cloud or in an on-premises environment.
Strong cloud management foundation.
Unix shells, Python and Go programming proficiency.
Experience in MySQL or PostgreSQL databases.
Outstanding teammate who can collaborate and influence in a multifaceted environment.
Excellent interpersonal and written communication skills.
Excellent debugging and troubleshooting skills.
Ability to define standard operating procedures for supported platform features.
Experience working with observability tools and practices (Prometheus, Grafana).
Experience in troubleshooting and resolving incidents.
Cloud experience in AWS (preferred), including hands-on experience with AWS-CLI.
Hands-on experience in orchestration and containerisation, like Kubernetes and containers.
Experience with CI/CD (i. e., Jenkins and ArgoCD).
Solid Understanding of Networking (firewall, connectivity, routing, iptables, subnet configuration, etc. ).
Experience with Linux OS and shell/Python scripting.
Experience in programming with Python and Go.
Experience with API gateways like Kong and Nginx-based systems.
Experience with security best practices and technologies.
BS degree in computer science or a related technical field involving coding, or equivalent practical experience.
4-6 years of sysadmin experience in handling large-scale distributed system software deployments in the cloud or in an on-premises environment.

Experience
5-9 yrs

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App