Site Reliability Engineering Manager
Location: Remote
Compensation: To Be Discussed
Reviewed: Mon, Jul 27, 2026
This job expires in: 29 days
Job Summary
Leading a US-based team, the full-time Site Reliability Engineering Manager will mentor engineers, drive operational excellence, and manage incident responses while collaborating closely with global teams in a remote environment.
Key Responsibilities
- Lead, mentor, and develop a team of Site Reliability Engineers, conducting performance reviews and fostering a culture of continuous improvement
- Oversee day-to-day production operations, serving as an incident commander for major incidents and driving root cause analysis processes
- Enhance system reliability and automation, partnering with various teams to integrate reliability into the software development lifecycle
Required Qualifications
- 5-8 years of experience in Site Reliability Engineering, DevOps, Production Support, or Platform Engineering
- 1-2+ years of experience in a leadership role, mentoring or managing engineers
- Strong hands-on experience with production incident management and escalation processes
- Proficiency with Datadog or similar observability platforms and experience with Kubernetes and Docker
- Strong scripting or programming skills in PowerShell, Bash, Python, Java, or C#
COMPLETE JOB DESCRIPTION
The job description is available to subscribers. Subscribe today to get the full benefits of a premium membership with Virtual Vocations. We offer the largest remote database online...