Principal Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Fri, Aug 14, 2026
This job expires in: 30 days
Job Summary
Working remotely within the United States, the full-time Principal Site Reliability Engineer will lead project work to enhance platform reliability, mentor junior engineers, and engage in incident response while collaborating closely with product stakeholders and architects.
Key responsibilities
- Lead the development and maintenance of platform features that enhance reliability and cloud infrastructure
- Mentor service owners on deploying and operating their services effectively at scale
- Participate in incident response, triage, and root cause analysis to uphold service reliability standards
Required qualifications
- Significant experience operating Kubernetes in distributed environments
- Experience with cloud platforms such as GCP or AWS
- Familiarity with monitoring and observability infrastructure practices
- Understanding of infrastructure-as-code tools and methodologies
- Six years of experience in systems operations or development
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 48,944 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee