Senior Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Mon, Oct 05, 2026
This job expires in: 30 days
Job Summary
Focused on improving the reliability and operational maturity of cloud-based data platforms, the full-time Senior Site Reliability Engineer will work remotely to enhance production health, incident response, and automation while collaborating with various engineering teams.
Key responsibilities
- Own and improve the reliability, availability, and performance of production data platforms and pipelines
- Troubleshoot complex production issues across AWS and data platforms, leading incident response and root-cause analysis
- Design observability solutions using platforms like Datadog and Splunk, and automate operational tasks with Python
Required qualifications
- 5+ years of hands-on AWS experience supporting production environments
- Demonstrated experience in Site Reliability Engineering or similar roles
- Strong practical understanding and implementation of SRE principles
- Experience supporting production data pipelines or data-intensive services
- Proficiency in Python for automation and SQL for production troubleshooting
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 38,620 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee