Live opening · Posted 27 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
We are looking for a highly skilled Senior DevOps Engineer with 5-9 years of experience in designing, implementing, and managing cloud-native infrastructure. The ideal candidate should have strong expertise in Kubernetes, cloud platforms (preferably GCP), CI/CD, infrastructure automation, and monitoring. This role requires someone who can build scalable, secure, and highly available systems while collaborating closely with engineering teams.
Responsibilities:
Design, build, and manage scalable cloud infrastructure on GCP (preferred) or AWS.
Develop, maintain, and optimise CI/CD pipelines for faster and more reliable software delivery.
Deploy, manage, and troubleshoot Kubernetes-based applications in production environments.
Automate infrastructure provisioning and operational tasks using Python and Infrastructure as Code (IaC) tools.
Implement autoscaling strategies using HPA, Cluster Autoscaler, and KEDA.
Monitor application and infrastructure health using Prometheus, Grafana, New Relic, and logging/observability tools.
Ensure platform security by implementing IAM, RBAC, secrets management, and container security best practices.
Collaborate with engineering teams to improve deployment processes, reliability, and platform performance.
Troubleshoot production issues and drive continuous improvements in infrastructure and operational efficiency.
Contribute to infrastructure architecture, scalability planning, and cloud optimisation initiatives.
Requirements:
5-9 years of experience in DevOps, Platform Engineering, or Site Reliability Engineering (SRE).
Strong hands-on experience with Kubernetes in production environments.
Experience with GCP (preferred) or AWS cloud platforms.
Strong understanding of CI/CD architecture and deployment automation.
Experience designing and managing cloud-native infrastructure.
Proficiency in Python for automation and scripting.
Strong knowledge of monitoring, logging, metrics, dashboards, and distributed tracing.
Experience with Prometheus, Grafana, and/or New Relic.
Good understanding of IAM, RBAC, Secrets Management, and container security.
Experience implementing autoscaling using HPA, Cluster Autoscaler, and KEDA.
Good to Have:
Experience with Event-Driven Architecture (EDA).
Exposure to cloud and infrastructure architecture design.
Knowledge of SOC 2 compliance and security practices.
Understanding of FinOps and cloud cost optimisation.
Experience working with AI/LLM infrastructure or GPU-based workloads.
Experience
5-9 yrs
More openings worth a look
Recently tracked roles with full details and direct application links.