Site Reliability Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Thu, Oct 01, 2026
This job expires in: 30 days
Job Summary
Owning the uptime, performance, and observability of the platform, the midlevel or senior Site Reliability Engineer will work remotely to establish operational best practices and enhance the reliability of API and platform services.
Key responsibilities
- Manage uptime, SLAs, and SLOs across API and platform services
- Build and maintain observability through logging, metrics, and tracing
- Design and implement high-availability deployment and containerization strategies
Required qualifications
- Full-time experience in SRE, DevOps, or backend engineering, preferably in a trading firm or high-growth startup
- Hands-on experience with observability tools (e.g., Prometheus, OpenTelemetry)
- Strong proficiency in Python, including application development and performance optimization
- Experience with containerization and deployment strategies (e.g., Docker, Kubernetes)
- Familiarity with Linux debugging and profiling tools (e.g., strace, gdb)
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 41,336 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee