Live opening · Posted 5 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
Experience: 5–8 Years
Domain: Enterprise Banking Platforms & Financial Systems
Employment Type: Full-time
About the Role
We are looking for an experienced Senior / Lead Site Reliability Engineer to drive reliability engineering for mission-critical banking infrastructure and distributed transaction-processing platforms.
The ideal candidate will have strong expertise in SRE, DevOps, Kubernetes, cloud platforms, Infrastructure as Code, observability, automation, and high-availability systems, with experience supporting high-volume transactional environments.
Key Responsibilities
Design fault-tolerant, highly available, and multi-region distributed systems.
Define and manage SLOs, SLIs, SLAs, and Error Budgets in collaboration with product and engineering teams.
Implement Infrastructure as Code (IaC) using Terraform and Ansible.
Develop Kubernetes-based automation and self-healing mechanisms.
Lead response to P1/P2 incidents, perform root-cause analysis, and drive long-term remediation.
Conduct capacity planning, load testing, and Chaos Engineering exercises.
Build automation to improve system reliability and operational efficiency.
Support infrastructure security, compliance, and automated auditing for financial platforms.
Drive zero-downtime and reliability-focused architecture practices.
Required Technical Skills
5–8 years of experience in SRE, DevOps, or Software Engineering.
2+ years of experience in banking, fintech, or high-volume transactional environments.
Strong expertise in Kubernetes; CKA certification is preferred.
Experience with AWS, Azure, or GCP and public/hybrid cloud architectures.
Strong programming skills in Go, Python, or Java.
Hands-on experience with Terraform, Ansible, and Kubernetes operators.
Experience with PostgreSQL, Oracle, Cassandra, Redis, and Kafka.
Strong observability experience with ELK/EFK, OpenTelemetry, Prometheus, Cortex/Thanos, and distributed tracing tools such as Jaeger/Zipkin.
Exposure to Chaos Engineering tools such as Gremlin or Chaos Mesh.
Understanding of PCI-DSS, SOC 2, and banking regulatory requirements.
Interested candidates can send their updated CV to rekha@icorepioneer.com
Work arrangement
No
More openings worth a look
Recently tracked roles with full details and direct application links.