Live opening · Posted 7 hours ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
Job Title:DevOps / Site Reliability Engineer (SRE)
Experience: 6 to 10 Years
Location: PAN India
Notice Period: 0 to 30 Days Preferred
Job Summary
We are seeking an experienced DevOps / Site Reliability Engineer (SRE) to join a high-performing engineering team. The ideal candidate will be responsible for ensuring the reliability, scalability, performance, and availability of critical applications and infrastructure. This role requires strong expertise in DevOps practices, cloud technologies, automation, monitoring, and incident management.
Key Responsibilities
Design, implement, and maintain scalable, secure, and highly available infrastructure.
Automate deployment, monitoring, and operational processes using modern DevOps tools and practices.
Improve system reliability, performance, and operational efficiency through continuous improvement initiatives.
Manage CI/CD pipelines and drive automation across the software development lifecycle.
Monitor application and infrastructure health using observability and monitoring tools.
Conduct root cause analysis for production incidents and implement preventive measures.
Collaborate closely with development, security, and infrastructure teams to ensure platform stability.
Support cloud infrastructure management and optimization.
Establish and maintain SRE best practices, including SLIs, SLOs, and error budgets.
Required Skills
Strong experience in DevOps and Site Reliability Engineering (SRE) concepts.
Hands-on expertise with CI/CD tools such as Jenkins, GitHub Actions, GitLab CI, or similar.
Experience with Infrastructure as Code (IaC) tools such as Terraform, CloudFormation, or Ansible.
Strong knowledge of Linux/Unix administration.
Experience with Containerization and Orchestration using Docker and Kubernetes.
Hands-on experience with Cloud Platforms (AWS, Azure, or GCP).
Expertise in monitoring and observability tools such as Prometheus, Grafana, ELK, Splunk, Dynatrace, or Datadog.
Strong scripting skills in Python, Shell, or similar automation languages.
Experience in incident management, production support, and troubleshooting complex systems.
Preferred Skills
Knowledge of security best practices within DevOps environments.
Experience implementing high-availability and disaster recovery solutions.
Understanding of microservices architecture.
Exposure to Agile and DevSecOps practices.
Educational Qualification
Bachelor's or Master's degree in Computer Science, Information Technology, Engineering, or a related field.
Work arrangement
No
More openings worth a look
Recently tracked roles with full details and direct application links.