Senior Platform Monitoring Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Wed, Aug 26, 2026
This job expires in: 13 days
Job Summary
Seeking a Senior Platform Monitoring Engineer, this full-time position will focus on developing monitoring solutions, leading incident investigations, and enhancing platform reliability while working remotely.
Key responsibilities:
- Lead platform incident investigations and coordinate cross-functional teams for rapid detection and resolution
- Conduct post-incident root cause analyses to identify patterns and prevent future occurrences
- Design and implement customer-focused alerting pipelines and observability workflows to improve detection coverage
Required qualifications:
- Minimum of 6 years of experience in SRE, DevOps, or similar roles
- Production-level experience with major cloud providers (AWS, Azure, GCP) and container technologies (Docker, Kubernetes)
- Hands-on experience with monitoring and alerting tools such as ELK, Prometheus, and Grafana
- Strong proficiency in Python or similar languages for building automation tools
- BS, Master's, or PhD in Computer Science, Computer Engineering, or a related field
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 41,015 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee