Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Fri, Sep 18, 2026
This job expires in: 30 days
Job Summary
Supporting and maintaining Kubernetes infrastructure across AWS and on-premises environments, the full-time Site Reliability Engineer will drive reliability, automation, and operational excellence within highly available production systems in a remote setting.
Key responsibilities
- Manage, build, and maintain Kubernetes clusters across AWS and on-premises environments
- Monitor, troubleshoot, and resolve issues impacting Kubernetes infrastructure and production services
- Implement and maintain Infrastructure as Code (IaC) solutions using tools like Terraform or CloudFormation
Required qualifications
- 4+ years of experience in Site Reliability Engineering, DevOps, or a related role
- Strong hands-on experience managing Kubernetes clusters in production environments
- Experience with Linux administration and troubleshooting
- Familiarity with Infrastructure as Code (IaC) tools such as Terraform or CloudFormation
- Knowledge of monitoring and observability tools like Prometheus or Grafana
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 41,590 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee