Staff Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Thu, Aug 27, 2026
This job expires in: 30 days
Job Summary
Owning the reliability, security, and automation of a platform that processes billions of dollars annually, the full-time Staff Site Reliability Engineer will manage uptime and incident response, drive containerization efforts, and engineer AWS infrastructure while working remotely.
Key responsibilities
- Ensure production availability and responsiveness, managing incident response for payment processing systems
- Lead the containerization of internal DevOps services, transitioning legacy systems to a modern container platform
- Design and maintain AWS infrastructure using Terraform, focusing on reliability, security, and cost efficiency
Required qualifications
- 10+ years of experience in development, DevOps, or site reliability engineering
- 5+ years of experience administering Linux servers in a production environment
- 3+ years of experience building and operating production AWS infrastructure
- 3+ years of experience running production container workloads
- 3+ years of experience with configuration management and infrastructure as code (Terraform, Puppet, Ansible, etc.)
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 47,980 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee