Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Thu, Aug 20, 2026
This job expires in: 30 days
Job Summary
Leading Site Reliability Engineering initiatives, the full-time Sr. Site Reliability Engineer will define and improve SLIs, SLOs, and reliability standards for clinical-grade applications while managing production incidents and designing observability solutions in a remote environment.
Key responsibilities
- Define and improve SLIs, SLOs, and reliability standards for platform services
- Manage production incident response, including root cause analysis and corrective action planning
- Design observability solutions using modern monitoring platforms and optimize Kubernetes-based environments
Required qualifications
- 5+ years of experience in Site Reliability Engineering or related fields supporting high-availability systems
- Bachelor's degree in Information Technology, Computer Science, Engineering, or a related field; Master's preferred
- Deep expertise in Kubernetes, Linux, and cloud platforms such as AWS, Azure, or GCP
- Strong hands-on experience with Python, Go, or Bash, and Infrastructure as Code tools like Terraform
- Experience supporting healthcare IT environments and knowledge of HIPAA and HITRUST compliance
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 48,221 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee