Live opening · Posted 1 day ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
We are looking for an experienced SDE III - DevOps Engineer to join our DevOps team and help build, scale, and maintain reliable cloud-native infrastructure and services. The ideal candidate will have strong hands-on expertise in Kubernetes, cloud platforms, CI/CD, Python automation, cloud architecture, and observability. You will work on building scalable infrastructure, improving system reliability, and driving engineering best practices across the organization.
Responsibilities:
Design, build, and maintain highly scalable and reliable Kubernetes-based infrastructure.
Develop and improve CI/CD architectures and deployment pipelines.
Contribute to cloud architecture and infrastructure design across AWS/GCP, with preference for GCP.
Build automation and tooling using Python.
Implement and improve system monitoring and observability, including logs, metrics, dashboards, and distributed tracing.
Work on infrastructure scalability, reliability, availability, and performance.
Implement and manage Kubernetes autoscaling solutions such as HPA, Cluster Autoscaler, and KEDA.
Apply security best practices, including IAM, secrets management, RBAC, and container security.
Collaborate with engineering teams to improve deployment, operational, and reliability practices.
Participate in architecture discussions, troubleshooting, and production problem-solving.
Requirements:
5-9 years of relevant software engineering / DevOps / SRE experience.
Strong hands-on experience with Kubernetes.
Experience with AWS or GCP, preferably GCP.
Strong understanding of CI/CD architecture and pipelines.
Experience with cloud architecture and cloud-native systems.
Strong Python scripting/automation experience.
Experience with monitoring and observability, including logs, metrics, dashboards, and tracing.
Strong understanding of production systems, reliability, and troubleshooting.
Strong problem-solving and debugging skills.
Ability to work effectively across engineering teams.
Strong ownership of production systems and infrastructure.
Ability to contribute to architectural decisions and technical strategy.
Good communication and collaboration skills.
Good to Have:
Experience with HPA, Cluster Autoscaler, or KEDA.
Knowledge of IAM, RBAC, secrets management, and container security.
Experience with Prometheus, Grafana, or New Relic.
Experience with Event-Driven Architecture (EDA).
Strong architecture and system-design experience.
Exposure to SOC2 environments.
Understanding of FinOps principles.
Experience with AI/LLM infrastructure or GPU workloads.
Experience
5-9 yrs
More openings worth a look
Recently tracked roles with full details and direct application links.