Site Reliability Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Thu, Oct 01, 2026
This job expires in: 30 days
Job Summary
To build reliable, observable systems, the remote contract Site Reliability Engineer will manage system reliability across the backend stack, design observability systems, and participate in 24/7 on-call rotations while mentoring new team members.
Key Responsibilities
- Design observability systems and define SLOs, error budgets, and monitoring strategies
- Own 24/7 P0 on-call rotation and validate AI-generated incident reports
- Implement reliability improvements and collaborate with teams to integrate reliability into delivery
Required Qualifications
- 6+ years of hands-on experience in SRE, DevOps, or Platform Engineering
- Strong proficiency in AWS services such as ALB, ECS/Fargate, and RDS Aurora
- Production experience in Python backend engineering and debugging
- Experience with incident response, on-call rotations, and SLOs
- Hands-on experience with infrastructure-as-code tools like Terraform or CloudFormation
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 41,336 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee