Senior Site Reliability Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Thu, Sep 24, 2026
This job expires in: 30 days
Job Summary
Leading the design and operation of a Kubernetes-based platform, the remote Senior Site Reliability Engineer will ensure high availability and security in highly regulated environments while driving reliability engineering initiatives and developing secure CI/CD pipelines.
Key responsibilities
- Design, build, and operate large-scale Kubernetes platforms supporting FedRAMP High and DoD IL5 environments
- Lead reliability engineering initiatives by defining SLOs, SLIs, and error budgets while improving platform performance through automation
- Drive incident response and root cause analysis, participating in on-call support and reducing operational toil through automation
Required qualifications
- 10+ years of experience in Site Reliability Engineering, DevOps, or related fields with a focus on large-scale platform initiatives
- Extensive hands-on experience with Kubernetes platforms in production environments, including EKS, AKS, or GKE
- Strong background in supporting FedRAMP High, DoD IL5, or similar regulated environments, with knowledge of compliance-driven practices
- Expertise in cloud infrastructure, Linux administration, and Infrastructure as Code (Terraform)
- Proficiency in Python, Go, or similar programming languages, with experience in building observability solutions
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 41,975 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee