Live opening · Posted 13 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
We are looking for an experienced DevOps / SRE Engineer with strong hands-on expertise in Google Cloud Platform (GCP), Kubernetes, Terraform, Docker, CI/CD, and API Gateway administration. The ideal candidate will be responsible for building, managing, monitoring, and improving highly available cloud platforms and API infrastructure. The role requires strong experience in Site Reliability Engineering (SRE), production support, incident management, automation, observability, and cloud-native technologies.
Responsibilities:
Administer and manage API Gateway / API Management platforms, with hands-on experience in Apigee.
Design, deploy, and maintain cloud infrastructure primarily on Google Cloud Platform (GCP).
Manage Kubernetes / GKE clusters and containerised workloads.
Develop and maintain infrastructure using Terraform / Infrastructure as Code (IaC).
Build and manage Docker-based application environments.
Design, implement, and maintain CI/CD pipelines for automated application and infrastructure deployments.
Implement monitoring, alerting, logging, and observability across cloud and application environments.
Work with Prometheus, Grafana, Elasticsearch, and Kibana for monitoring, logging, and troubleshooting.
Implement and manage Helm charts for Kubernetes deployments.
Support GitOps-based deployment and infrastructure management practices.
Work with Istio / Service Mesh for traffic management, security, and observability.
Troubleshoot production issues and participate in incident management, root-cause analysis, and service restoration.
Implement and maintain secure authentication and authorisation mechanisms using OAuth2 / OIDC.
Collaborate with development, security, infrastructure, and product teams in an Agile / SRE environment.
Continuously improve platform reliability, scalability, performance, security, and automation.
The core requirements for the job include the following:
Mandatory Skills:
Google Cloud Platform (GCP) - Mandatory
Kubernetes / GKE
Terraform
Docker
Linux
CI/CD
API Gateway Administration
Apigee
Cloud Monitoring, Alerting and Logging
SRE / Site Reliability Engineering experience
Incident Management and Production Support
OAuth2 / OIDC operational knowledge
Good to Have:
Helm
GitOps
Prometheus
Grafana
Elasticsearch / Kibana
Istio / Service Mesh
Distributed tracing/observability
Experience with cloud-native architecture
Automation and scripting
Ideal Candidate:
The ideal candidate should have strong hands-on experience managing GCP-based cloud platforms and Kubernetes environments, along with exposure to API Gateway / Apigee administration, infrastructure automation, CI/CD, observability, and SRE practices.
Strong troubleshooting skills, production ownership, automation mindset, and experience handling critical incidents in a cloud-native environment will be highly valued.
Core Technology Stack: GCP, Apigee, GKE/Kubernetes, Terraform, Docker, Linux, CI/CD, Helm, GitOps, Prometheus, Grafana, Elasticsearch/Kibana, Istio, OAuth2/OIDC, SRE.
Experience
5-8 yrs
More openings worth a look
Recently tracked roles with full details and direct application links.