Senior Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Mon, Sep 21, 2026
This job expires in: 30 days
Job Summary
Passionate about building reliable infrastructure, the full-time Senior Site Reliability Engineer will design and maintain a Kubernetes-based multi-cloud platform, improve monitoring tools, and collaborate with product and engineering teams in a fully remote environment.
Key responsibilities
- Design and maintain the Kubernetes-based multi-cloud platform architecture to ensure availability, scalability, and fault tolerance
- Implement and enhance monitoring and alerting tools for system health and performance visibility
- Participate in on-call rotations and create runbooks and automation for efficient incident response and management
Required qualifications
- Deep hands-on experience with Kubernetes in production environments
- Expertise in infrastructure as code tools like Terraform
- Demonstrated experience with monitoring and observability tools such as Prometheus or Grafana
- Strong incident response skills with experience in root cause analysis
- A passion for automation and improving system reliability and maintainability
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 38,952 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee