Senior Manager, Site Reliability Engineering
Location: Remote
Compensation: Hourly
Reviewed: Wed, Oct 07, 2026
This job expires in: 30 days
Job Summary
Leading a high-performing team, the full-time Senior Manager, Site Reliability Engineering will oversee the reliability, performance, and availability of Linux-based systems supporting digital commerce applications, while driving modernization of SRE and DevOps practices in a remote role.
Key responsibilities:
- Lead and develop a team of Site Reliability Engineers to ensure the stability and performance of digital commerce systems
- Set technical direction for SRE initiatives, promoting the adoption of Kubernetes and CI/CD standards
- Champion incident management and observability practices to enhance system reliability and engineering improvements
Required qualifications:
- 8+ years of experience administering Linux systems, with 3+ years in a leadership role within engineering or SRE teams
- Strong hands-on experience with Kubernetes, Docker, and CI/CD tools like GitHub Actions
- Extensive knowledge of configuration management tools such as Puppet and automation using Python
- Proven expertise in load balancing, application delivery, and tuning web servers like Nginx and Apache Tomcat
- Experience with observability tools such as Datadog and incident response process development
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 40,207 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee