Live opening · Posted 4 days ago

Site Reliability Engineer

THUNDERYARD SOLUTIONS · United States (Remote)
Linkedin Yes
You are 4 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 4 days ago
CompanyTHUNDERYARD SOLUTIONS
LocationUnited States (Remote)
Salary$120K/yr - $140K/yr · 401(k), +4 benefits
Work modeYes
SourceLinkedin
Listed4 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
12 min from Linkedin publishing this role to us finding it
5 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
70,076 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Site Reliability Engineer – AWS Cloud Tenant Support
Position Summary
The Site Reliability Engineer will support tenants operating within a multi-tenant AWS Cloud environment by ensuring platform reliability, availability, scalability, security, and operational efficiency. This role partners with application teams, cloud engineers, security teams, and tenant stakeholders to maintain resilient infrastructure, automate operational processes, improve deployment practices, and reduce service disruption through proactive monitoring and incident response.
Platform operations: Managing Kubernetes environments (EKS or similar), ensuring availability, scaling, and lifecycle management.
CI/CD: Building and optimizing Jenkins pipelines for secure, automated, repeatable deployments; improving rollback and release strategies with dev/platform teams.
Automation: Creating infrastructure-as-code, scripts, runbooks, and self-service tools to cut manual work.
Observability & reliability: Setting up monitoring, alerting, SLIs/SLOs, and dashboards; leading incident response, RCAs, and post-incident fixes.
Tenant support: Handling onboarding, provisioning, troubleshooting, and service health communication.
Security: Applying AWS best practices — IAM, encryption, network segmentation, least privilege, compliance controls.
Environment:
Requires on-call rotation, maintenance windows, incident response, and cross-team collaboration in a production environment where reliability and security are top priorities.
Required Qualifications
5+ years of experience supporting production workloads in AWS, including services such as EC2, VPC, IAM, CloudWatch, S3, Route 53, Elastic Load Balancing, RDS, Lambda, or related AWS services.
Hands-on experience administering, troubleshooting, and supporting Kubernetes clusters, containers, Helm charts, deployments, services, ingress, and workload scaling.
Experience creating, maintaining, and troubleshooting Jenkins jobs, shared libraries, build agents, credentials, plugins, and automated deployment workflows.
Strong understanding of CI/CD concepts, including build automation, artifact management, automated testing, environment promotion, release orchestration, and rollback practices.
Working knowledge of infrastructure as code and automation tools such as Terraform, AWS CloudFormation, Ansible, Python, Bash, or PowerShell.
Experience with observability tools and practices, including log aggregation, metrics collection, distributed tracing, alert design, and dashboarding.
Strong Linux administration and troubleshooting skills, including networking, system performance, process management, storage, and permissions.
Understanding of cloud networking concepts such as VPCs, subnets, routing, security groups, network ACLs, DNS, load balancing, and VPN or Direct Connect patterns.
Ability to troubleshoot complex application, infrastructure, deployment, and network issues across multiple teams and tenant environments.
Strong written and verbal communication skills, including the ability to document procedures, explain technical issues, and communicate incident status to stakeholders.
Preferred Qualifications
AWS certification such as AWS Certified Solutions Architect, AWS Certified SysOps Administrator, AWS Certified DevOps Engineer, or equivalent cloud experience.
Kubernetes certification such as Certified Kubernetes Administrator or Certified Kubernetes Application Developer.
Experience supporting multi-tenant cloud platforms, shared services, managed service environments, or enterprise cloud landing zones.
Experience with Git, GitHub, GitLab, Bitbucket, Maven, Gradle, Docker, Argo CD, AWS CodePipeline, or other DevOps tooling.
Knowledge of SRE practices, including error budgets, toil reduction, blameless postmortems, capacity planning, and reliability engineering principles.
Experience implementing security and compliance controls in regulated or enterprise environments.
Compensation:
$120,000.00-$140,000.00 annually, plus benefits, including medical/dental/vision coverage, PTO and a partial 401k match.
Vetting:
Applicants selected will be subject to a government investigation and may need to meet eligibility requirements of the U.S. government client.
ThunderYard Solutions is proud to be an Equal Opportunity Employer. We don’t just accept difference – we celebrate it, we support it, and we thrive on it for the benefit of our employees, our community, and our customers. All applicants will be considered for employment without discrimination of race, color, religion, or belief, national, social, or ethnic origin, sex, age, physical, mental, or sensory disability, HIV status, sexual orientation, gender identity and/or expression, marital, civil union, or domestic partnership status, protected veteran status, family medical history or genetic information.

Work arrangement
Yes

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App