Senior Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Fri, Sep 18, 2026
This job expires in: 30 days
Job Summary
Seeking a hands-on Senior Site Reliability Engineer to join the Infrastructure team in a fully remote, full-time role, where the engineer will own the reliability of core production systems, define SLIs and SLOs, and lead incident response efforts.
Key responsibilities
- Own the reliability of core production systems, including instrumentation, target setting, and operational accountability
- Define and maintain SLIs and SLOs, integrating them into dashboards and alerts to drive improvements based on error budgets
- Lead incident response efforts, conducting thorough investigations, restoring services, and writing actionable postmortems
Required qualifications
- 6-10 years of experience in SRE, production engineering, or backend engineering within cloud-based environments (AWS preferred)
- Proven track record of owning systems end to end, including design, operation, and accountability for performance
- Hands-on experience with SLIs, SLOs, and error budgets, along with strong incident management skills
- Depth in distributed systems and cloud infrastructure fundamentals, including containerization and networking
- Strong programming skills in Go, Python, or similar, with experience managing infrastructure through code (Terraform or equivalent)
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 41,590 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee