Live opening · Posted 27 days ago

Site Reliability Engineer

Flipkart · Bangalore
Instahyre 5-9 yrs
You are 27 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 27 days ago
CompanyFlipkart
LocationBangalore
Experience5-9 yrs
SourceInstahyre
Listed27 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
0 min from Instahyre publishing this role to us finding it
11 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
16,937 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

We are seeking an experienced Site Reliability Engineer (SRE) to join our dynamic Cloud Infrastructure team. This role demands a deep understanding of cloud-native technologies, particularly containers and Kubernetes, along with strong programming skills in languages such as Go and Python. The ideal candidate will have a proven track record of at least 3 years in the field, focusing on enhancing the reliability, scalability, design, development, deployment, and operation of self-service platforms that facilitate the lifecycle management of applications supporting products and services.
Responsibilities:
Collaborate with internal customers and partners to deliver key business outcomes.
Ensure that cloud products are reliable, scalable, efficient, and compliant with security and operational standards.
Enhance observability practices to ensure comprehensive monitoring and alerting across cloud services.
Respond to cloud incidents, perform root cause analysis, and implement corrective actions to prevent future occurrences.
Develop and maintain incident response plans.
Analyse system performance metrics and make recommendations for improvements.
Implement changes to optimise resource utilisation and improve application performance.
Drive improvements in CI/CD processes to increase deployment velocity and reliability.
Develop and maintain automation to streamline operations, reduce manual work, and enhance system reliability.
Requirements:
Minimum of 5+ years of programming experience with Go or Python.
5+ years of experience in implementing large-scale, distributed, high-availability, fault-tolerant systems and infrastructure in a production environment.
Proficiency in delivering products within a multi-functional team environment.
Demonstrated expertise in observability tools and practices, ensuring system reliability and performance.
Extensive experience with Kubernetes as an SRE or related cloud infrastructure and cloud-native technologies.
Experience in developing with Kubernetes and/or building Kubernetes controllers is highly desirable.
Deep understanding of API design and RESTful principles, with experience in building web services at scale.
Preferred Skills:
Certifications in Kubernetes, lifecycle management, or related fields.
Understanding application lifecycle management and CI/CD is a plus.
Experience in a high-traffic, large-scale environment.
Familiarity with additional programming languages or frameworks.
Proficiency in Agile development methodologies.
Experience in participating in open-source standards and contributing to open-source projects is a plus.

Experience
5-9 yrs

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App