Site Reliability Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Fri, Aug 14, 2026
This job expires in: 29 days
Job Summary
To enhance platform reliability and drive automation, the contract Site Reliability Engineer will work remotely, focusing on improving operational health and collaborating with cross-functional teams to manage critical production incidents and optimize monitoring solutions.
Key responsibilities:
- Act as the primary technical escalation point for critical production incidents, providing hands-on support during high-severity outages
- Improve platform reliability by reviewing new product launches and production readiness before release
- Design, implement, and optimize monitoring and observability solutions across cloud infrastructure and applications
Required qualifications:
- 4+ years of experience as a Site Reliability Engineer, DevOps Engineer, or similar infrastructure-focused role
- Strong experience supporting production systems running on AWS
- Hands-on experience with monitoring and observability platforms such as Datadog or AWS CloudWatch
- Experience with incident management platforms such as PagerDuty
- Strong understanding of production incident management and root cause analysis processes
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 47,198 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee