Live opening · Posted 6 hours ago

Lead Site Reliability Engineer, Athena Core

JPMorgan Chase · LONDON, United Kingdom
Oracle
You are 6 hours behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 6 hours ago
CompanyJPMorgan Chase
LocationLONDON, United Kingdom
SkillsPython, Java, Spring Boot, .NET
SourceOracle
Listed6 hours ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
2 min from Oracle publishing this role to us finding it
12 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
71,381 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability.
As a Lead Site Reliability Engineer at JPMorgan Chase within the Commercial and Investment Banking, Markets Technology – Athena Core team, you hold a leadership role in your team, demonstrate strong knowledge across multiple technical domains, and advise others on the technical and business issues facing them. Take lead and conduct resiliency design reviews, break up complex problems into digestible work for other engineers, act as a technical lead for medium to large-sized products, and provide advice and mentoring to other engineers.
Job responsibilities
Consistently models and champions site reliability culture and practices, documents and shares knowledge within your organization via internal forums and communities of practice
Leads initiatives to improve the reliability and stability of your team’s applications and platforms using data-driven analytics to improve service levels, proactively identifying and solving technology-related bottlenecks in areas of expertise
Drives collaboration with your team to identify comprehensive service level indicators and the stakeholder partners to establish reasonable service level objectives and error budgets with your customers
Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
Serves as the main point of contact during major incidents for your application and has the skills to identify and solve the issue quickly to avoid financial loss to the business
Offers a high level of technical expertise within one or more technical domains and provides advice and mentorship to other engineers
Leads reuse-first adoption of AI-assisted reliability workflows across SDLC/toolchain practices (e.g., CI/CD quality checks, test/validation automation, and operational readiness), ensuring traceability/auditability, resiliency, and security controls.
Be part of the first line of support for the wider Athena Core team.
Required qualifications, capabilities, and skills
Formal training or certification on site reliability engineering concepts and applied experience
Bachelor’s Degree in Computer Science or equivalent
Demonstrated proficiency in reliability, scalability, performance, security, enterprise system architecture, toil reduction, and other site reliability best practices
An understanding of programming languages such as: Python, Java/Spring Boot, .Net
Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge capture) with strong validation habits and awareness of data sensitivity.
Ability to evaluate AI-assisted operational recommendations for correctness and risk, define appropriate guardrails for team usage, and ensure outcomes align to resiliency and security expectations.
A strong understanding of Linux, Proficient knowledge and experience in observability such as white and black box monitoring, service level objective alerting, and telemetry collection
Proficient with continuous integration and continuous delivery practices and tooling
Proficient with container and container orchestration
Experience with troubleshooting common networking technologies and issues
Advanced knowledge of software applications and technical processes with emerging depth in one or more technical disciplines, and actively self-educates to evaluate and recommend suitable new technologies

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App