Senior Infrastructure Engineer
Location: Remote
Compensation: Salary
Reviewed: Fri, Sep 04, 2026
This job expires in: 30 days
Job Summary
To lead the reliability and operational evolution of a growing platform, the full-time Senior Infrastructure Engineer will manage the improvement of system reliability, establish service-level indicators, and evolve disaster recovery strategies while working remotely or from various U.S. locations.
Key responsibilities
- Build and enhance the reliability and resiliency of systems and services
- Establish and review SLIs, SLOs, and error budgets for critical services
- Own and evolve the disaster recovery strategy, including recovery objectives and regular exercises
Required qualifications
- 5+ years of hands-on cloud or infrastructure engineering experience focused on reliability and production operations
- Experience defining SLIs and SLOs for production services
- Proficiency with an observability platform in production, preferably Datadog
- Strong coding skills in Python, Go, TypeScript, or similar for internal tooling and automation
- Experience with Terraform and AWS, including managing disaster recovery plans
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 44,401 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee