Staff Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Tue, Sep 08, 2026
This job expires in: 30 days
Job Summary
Seeking an experienced Staff Site Reliability Engineer for a full-time role within the Emerging Products Group, who will design and operate large-scale cloud infrastructure, lead incident response efforts, and drive systemic improvements while ensuring compliance with FedRAMP standards.
Key responsibilities
- Design, build, and operate large-scale cloud infrastructure and production services
- Lead incident response efforts and drive post-incident reviews focused on systemic improvements
- Develop software and automation to eliminate operational toil and improve deployment safety
Required qualifications
- Extensive experience architecting large-scale production services in AWS and/or GCP
- Deep expertise in Kubernetes patterns and Linux-based system standards
- Strong software engineering skills in Golang and/or Python
- Experience with Infrastructure as Code technologies such as Terraform and Helm
- Understanding of cloud security fundamentals and compliance standards like FedRAMP
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 40,027 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee