Site Reliability Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Wed, Aug 05, 2026
This job expires in: 22 days
Job Summary
To ensure the reliability and performance of software infrastructure, the full-time remote Site Reliability Engineer will monitor Kubernetes clusters, troubleshoot system issues, and support observability tools while participating in on-call rotations.
Key responsibilities
- Monitoring and maintaining Kubernetes clusters and containerized workloads
- Troubleshooting system issues and improving operational documentation
- Supporting and maintaining the observability stack, including metrics and alerting rules
Required qualifications
- Expertise in Docker and Kubernetes concepts
- Fluency with Linux and Bash scripting
- Familiarity with core AWS concepts and services, particularly IAM, VPCs, EC2, and EKS
- Experience using Terraform to deploy cloud infrastructure
- Ability to leverage Prometheus and Grafana for metrics and logs collection
COMPLETE JOB DESCRIPTION
The job description is available to subscribers. Subscribe today to get the full benefits of a premium membership with Virtual Vocations. We offer the largest remote database online...