AI DevOps Site Reliability Engineer
Job Summary
Driving innovation in a remote full-time capacity, the Enterprise AI DevOps Site Reliability Engineer will architect and deploy scalable Infrastructure as Code (IaC) and multi-cloud CI/CD pipelines, lead the design of intelligent AI-driven workflows, and implement observability frameworks to enhance operational efficiency across complex enterprise ecosystems.
Key Responsibilities
- Architect and deploy scalable Infrastructure as Code (IaC) and multi-cloud CI/CD pipelines across AWS, Azure, and GCP
- Lead the design of intelligent AI-driven workflows for non-engineering systems using Agentic AI and the Model Context Protocol (MCP)
- Implement robust observability frameworks with Grafana, Prometheus, and New Relic to monitor performance and optimize costs
Required Qualifications
- 5+ years of hands-on experience managing enterprise-grade Kubernetes and Docker environments
- Proficiency in integrating complex enterprise APIs and managing secure identity frameworks like OIDC and SAML
- Experience with modern CI/CD tooling and observability platforms
- Ability to act as a consultative advisor, translating technical requirements into strategic guidance
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 48,944 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee