Live opening · Posted 26 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
Requirements:
5+ years of experience in site reliability engineering, infrastructure engineering, or platform engineering.
Experience operating distributed production systems.
Experience with asynchronous processing systems and event-driven architectures.
Strong experience with cloud infrastructure such as AWS, GCP, or similar platforms.
Familiarity with containerised services and distributed worker systems.
Experience operating production databases and storage systems (PostgreSQL, MySQL, DynamoDB, Redis, or similar).
Experience building monitoring, alerting, and observability systems.
Experience working with infrastructure as code tools such as Terraform or Pulumi.
Experience defining service reliability targets such as SLIs, SLOs, and error budgets.
Experience handling incidents, running postmortems, and improving systems after production failures.
Ability to read and instrument application code (Python, Node.js, or similar).
Strong focus on reliability, operational clarity, and cost efficiency.
Nice to Have:
Experience operating large-scale workflow orchestration or distributed job systems.
Experience with AI or ML inference infrastructure.
Familiarity with mobile device cloud platforms.
Experience running high-concurrency systems.
Experience working in fast-moving engineering teams or startup environments.
Experience
4-8 yrs
More openings worth a look
Recently tracked roles with full details and direct application links.