Site Reliability Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Wed, Sep 02, 2026
This job expires in: 30 days
Job Summary
Focused on enhancing site reliability and scalability, the full-time Site Reliability Engineer will lead post-incident investigations, develop preventive strategies, and collaborate with cross-functional teams to improve performance metrics, working remotely.
Key responsibilities
- Leads post-incident investigations and conducts in-depth analyses to identify root causes
- Collaborates with teams to implement system enhancements and develops client-focused dashboards for performance tracking
- Maintains observability tools and provides actionable feedback to improve incident response and performance metrics
Required qualifications
- Bachelor's degree in Computer Science or a related field (or equivalent experience)
- 5+ years of proven experience in a Site Reliability Engineering role
- Strong knowledge of SRE best practices and incident management protocols
- Deep experience with observability tools such as New Relic or Data Dog
- Proficiency in coding (e.g., JavaScript, .NET, SQL) and familiarity with cloud platforms (e.g., AWS, Azure)
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 43,981 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee