Senior Site Reliability Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Thu, Jul 23, 2026
This job expires in: 21 days
Job Summary
Owning the reliability and operational excellence of the platform, the full-time Senior Site Reliability Engineer will manage Kubernetes-based infrastructure, improve AWS services, and lead incident response efforts while working remotely.
Key responsibilities
- Own and enhance platform reliability, capacity planning, and production readiness
- Design, build, and maintain Kubernetes infrastructure and deployment workflows
- Lead incident response processes including triage, mitigation, and postmortems
Required qualifications
- 5+ years of experience in site reliability engineering
- Proven experience managing reliability for production systems and defining SLOs
- Deep hands-on experience with Kubernetes/Helm on EKS in production
- Experience with AWS core services beyond EKS, including networking and IAM
- Familiarity with CI/CD or GitOps deployment patterns
COMPLETE JOB DESCRIPTION
The job description is available to subscribers. Subscribe today to get the full benefits of a premium membership with Virtual Vocations. We offer the largest remote database online...