Senior Site Reliability Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Mon, Oct 05, 2026
This job expires in: 30 days
Job Summary
To operationalize reliability in production, the full-time remote Senior Site Reliability Engineer will define service level objectives, build observability and incident response practices, and lead performance engineering efforts for a platform used by MEP contractors.
Key Responsibilities:
- Define and own service level indicators, objectives, and error budgets for customer-facing services while building a trustworthy measurement pipeline
- Establish and run incident response practices, including on-call rotations, escalation paths, and blameless post-mortems
- Lead performance and capacity engineering efforts, including query tuning, capacity modeling, and ensuring production readiness for new services
Required Qualifications:
- 6+ years of professional engineering experience, with at least 3+ years in a dedicated SRE or production engineering role at a B2B SaaS company
- Demonstrated ownership of SLOs and error budgets in production, with experience in defining and measuring them
- Deep observability skills with tools like Prometheus, Loki, Tempo, and Grafana
- Hands-on incident command experience and the ability to build on-call and escalation practices
- Strong database performance skills, particularly with MongoDB, and production Kubernetes experience
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 38,620 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee