Senior Site Reliability Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Thu, Aug 20, 2026
This job expires in: 30 days
Job Summary
Working remotely, the full-time Senior Site Reliability Engineer will enhance system reliability, performance, and scalability by collaborating with application engineers and DevOps teams to build automation tools, improve observability, and streamline incident response.
Key responsibilities
- Assess and enhance system visibility by improving dashboards, metrics, and logs
- Tighten monitoring and alerting for critical services to enable faster issue detection and response
- Automate routine checks and monitoring tasks to reduce manual effort and improve overall reliability
Required qualifications
- Solid experience in Python for automation and operational tasks
- Proficiency in at least one programming language such as Java, C++, or Go
- Strong understanding of Linux systems and cloud infrastructure (AWS, GCP, or Azure)
- Experience with CI/CD pipelines and automated testing frameworks
- Familiarity with observability tools like Prometheus, Grafana, or Datadog
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 47,805 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee