Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Thu, Sep 17, 2026
This job expires in: 30 days
Job Summary
To improve the reliability and operability of production services in a growing 24/7 SaaS environment, the full-time Site Reliability Engineer will apply software and systems engineering practices while working remotely.
Key responsibilities
- Own and enhance the use of Datadog for visibility and actionability of reliability metrics, logs, and alerts
- Design, implement, and maintain high-availability systems supporting the SaaS platform
- Participate in incident management, serving as an incident commander and coordinating service restoration efforts
Required qualifications
- BS degree in Information Systems, Engineering, or equivalent experience
- 5+ years of experience in Systems Engineering, Cloud Engineering, DevOps, Software Engineering, or SRE
- Experience with cloud-based technologies, particularly Azure and Kubernetes
- Strong Linux systems knowledge and experience troubleshooting complex distributed systems
- Proficiency in scripting and operational automation using tools like Bash, PowerShell, or Python
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 41,606 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee