Live opening · Posted 10 hours ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
Senior Site Reliability Engineer (SRE) – AWS Certified
Location: LATAM – Remote
Experience: 8+ years
Schedule: Monday–Friday, 5:00 PM–1:00 AM EST
AWS Certification: Mandatory
We are looking for an experienced Senior Site Reliability Engineer (SRE) to support highly available, enterprise-scale applications and cloud infrastructure across production environments.
Key Responsibilities
Monitor, troubleshoot, and resolve production application and infrastructure issues.
Act as a first responder for production alerts and major incidents.
Participate in P1/P2 incident bridge calls and communicate troubleshooting progress, impact, and resolution to technical and leadership teams.
Perform root cause analysis and implement corrective/preventive actions.
Support highly available applications across AWS, Kubernetes, and hybrid environments.
Improve reliability, monitoring, alerting, automation, and operational processes.
Troubleshoot Kubernetes workloads, networking, application connectivity, and cloud infrastructure.
Support CI/CD pipelines and collaborate with development, infrastructure, networking, and security teams.
Required Skills
8+ years of experience in SRE, DevOps, Cloud Engineering, or Production Support.
Active AWS Certification is mandatory.
Strong hands-on experience with AWS, including:
EC2
ECS
S3
Lambda
Load Balancers
Strong Kubernetes/EKS experience.
Hands-on Terraform/IaC experience.
Strong production incident response, P1/P2 support, RCA, and on-call experience.
Experience deploying and troubleshooting Java, Node.js, and React-based applications in cloud environments.
Strong understanding of Linux, TCP/IP, DNS, HTTP/HTTPS, and load balancing.
Experience with CI/CD tools such as GitHub, GitLab, Jenkins, or Harness.
Experience with monitoring/observability tools such as Prometheus, Grafana, Datadog, Splunk, or CloudWatch.
Experience troubleshooting database connectivity, including Oracle, MariaDB, or Microsoft SQL Server.
Strong communication skills, particularly during major production incidents.
Preferred Qualifications
Terraform certification.
Experience with Akamai or other CDN technologies.
VMware or hybrid/on-premises infrastructure experience.
Multi-cloud experience across AWS, Azure, or GCP.
Experience supporting large-scale, high-availability enterprise environments.
Work arrangement
Yes
More openings worth a look
Recently tracked roles with full details and direct application links.