Senior Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Wed, Oct 07, 2026
This job expires in: 30 days
Job Summary
To enhance system reliability, the full-time Senior Reliability Engineer will build observability, tooling, and automation to proactively identify issues, ensuring optimal performance during high-traffic events while working remotely.
Key responsibilities
- Own reliability for Tier 1 customer journeys by defining and instrumenting SLOs and monitors to detect degradation
- Lead capacity and resilience work for high-stakes events, including load-testing and backend investigation
- Automate incident response processes and maintain tooling for operational excellence reviews
Required qualifications
- 5+ years of experience as a Software, SRE, Platform, or Infrastructure Engineer with a focus on reliability outcomes
- Strong software engineering fundamentals with hands-on depth in observability and SLO engineering
- Production experience with AWS, Kubernetes/EKS, Terraform, and PostgreSQL
- Experience managing end-to-end incident management processes and tools like FireHydrant or PagerDuty
- Practical use of AI coding and analysis tools to enhance investigation and reporting
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 40,207 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee