Senior Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Wed, Aug 05, 2026
This job expires in: 22 days
Job Summary
Seeking a full-time Senior Site Reliability Engineer to own the reliability, scalability, and observability of critical financial SaaS applications and infrastructure, working remotely to ensure seamless, secure, and performant services for customers.
Key Responsibilities
- Design, implement, and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs) across critical systems
- Lead observability strategy by designing monitoring, logging, and tracing architectures and deploying observability tools
- Build and own runbooks, incident response procedures, and mentor the team on incident management and postmortems
Required Qualifications
- 7+ years in Site Reliability Engineering, DevOps, or related roles with significant responsibility for production systems
- Expert-level experience with Azure or AWS and knowledge of managed services at scale
- Demonstrated expertise in observability and hands-on experience with platforms like Prometheus or Grafana
- Strong background in SLOs, SLIs, and SLAs, with experience defining objectives and building systems to meet them
- Proficiency in Python, PowerShell, or bash for automation and systems programming
COMPLETE JOB DESCRIPTION
The job description is available to subscribers. Subscribe today to get the full benefits of a premium membership with Virtual Vocations. We offer the largest remote database online...