Lead Site Reliability Engineer
Location: Remote
Compensation: Salary
Reviewed: Thu, Oct 08, 2026
This job expires in: 30 days
Job Summary
Driving the reliability and operational excellence of cloud-native platforms, the full-time Lead Site Reliability Engineer (AI & Cloud Operations) will implement modern DevOps practices, manage AWS-based infrastructure, and optimize AI-driven operations in a remote environment.
Key responsibilities
- Lead the design and continuous improvement of Site Reliability Engineering (SRE) practices for scalable cloud platforms
- Architect and support AWS infrastructure, including containerized and serverless environments
- Develop automation solutions and maintain CI/CD pipelines to standardize deployments and enhance operational efficiency
Required qualifications
- 5+ years of experience in Site Reliability Engineering (SRE), DevOps, or Cloud Operations
- Strong hands-on experience with AWS services such as EC2, EKS, and Lambda
- Expertise in CI/CD pipeline management using tools like Jenkins or GitHub Actions
- Proficient in scripting and automation with Python or Shell
- Experience with containerization and orchestration technologies, particularly Kubernetes
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 41,387 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee