Live opening · Posted 9 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
The Senior DevOps Engineer position drives reliability strategy across multiple platforms and services. Defines cross‑team SLO governance, error‑budget policy, and release/change risk controls; architects resilient, scalable topologies (active‑active, partition tolerance) and leads modernization of platform, observability, and incident capabilities.
What You’ll Do:
Set enterprise SRE standards (SLO/SLI taxonomy, alerting contracts, escalation/run‑of‑show) and audit adherence.
Lead design of active‑active/multi‑region architectures and coordinate fault‑injection and region evacuation exercises.
Establish AIOps/event correlation to reduce MTTD/MTTR and drive automated remediation adoption.
Govern IaC/platform modules, golden pipelines, artifact provenance, and deployment verification patterns.
Own capacity/efficiency strategy (FinOps guardrails, autoscaling policies, caching/CDN economics).
Mentor engineers and contribute to the development of engineering talent through technical reviews and knowledge sharing.
Support enterprise-level incident response and post-mortem analysis for critical issues.
Represent SRE in strategic discussions and contribute to the development of technology roadmaps.
Oversee technical planning, estimation, and execution of high-impact projects, ensuring alignment with business priorities.
Drive continuous improvement in development processes and methodologies to enhance quality.
What You’ll Bring to Zelis:
Typically has Bach. +12 yrs; Masters +8 yrs Attends company training and regional training in area of expertise. Presents own work to members of department and external consultants and external professional conferences, Collaborates externally, Hosts external guests, visitors and consultants
Requires proficiency using AI tools skillfully with an understanding of how to develop/create intelligent prompts and/or agents.
Mastery of cloud platforms, Kubernetes at scale, global networking, DR tiering, and observability architectures.
Proven leadership across incidents, architecture councils, and cross‑functional reliability initiatives.
Exceptional ability to influence senior stakeholders and drive consensus.
Strong strategic thinking and decision-making capabilities.
Manages enterprise-level stakeholder relationships and communicates reliability strategy.
Chairs architecture reviews and ensures clear decision-making and follow-through.
Mentors staff and principal engineers, modeling calm and decisive leadership.
Drives organizational change through effective storytelling and consensus-building.
Work arrangement
No
More openings worth a look
Recently tracked roles with full details and direct application links.