Senior Site Reliability Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Mon, Oct 05, 2026
This job expires in: 30 days
Job Summary
Owning the reliability, availability, and performance of the GCP platform, the full-time Senior Site Reliability Engineer will define and operationalize service level objectives, build observability and incident response systems, and enhance infrastructure as code practices, all while working remotely.
Key responsibilities
- Own reliability, availability, and performance of the GCP platform, driving initiatives related to service level objectives and disaster recovery
- Build observability and incident response capabilities, extending GCP Cloud Operations and participating in on-call rotations
- Grow infrastructure as code and delivery pipeline, expanding Terraform usage and improving CI/CD processes for safer deployments
Required qualifications
- 5+ years of engineering experience, with at least 3 years in a cloud infrastructure, DevOps, or SRE-focused role
- Expert-level GCP fluency, including serverless compute and Cloud Operations, with strong Terraform skills preferred
- Experience in production operations, including incident response and defining service level objectives
- Quantitative fluency applied to production systems, with a strong understanding of SLO and error-budget math
- Hands-on experience using AI tools in engineering work and building automation with them
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 38,620 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee