Senior GPU Cloud K8S Expert
Location: Remote
Compensation: To Be Discussed
Reviewed: Fri, Sep 11, 2026
This job expires in: 30 days
Job Summary
To support the development of an AI-operated GPU cloud, the full-time Senior GPU Cloud K8S Expert will design, deploy, and manage production Kubernetes clusters optimized for GPU workloads while working remotely from San Jose, CA or Austin, TX.
Key responsibilities
- Manage production Kubernetes clusters for GPU workloads, ensuring topology-aware scheduling and multi-tenant isolation
- Implement and oversee automated provisioning and lifecycle management for Bare-Metal as a Service (BMaaS) environments
- Develop and maintain monitoring and incident management systems to ensure cluster availability and performance
Required qualifications
- 5+ years of experience in Kubernetes operations, with at least 2 years focused on GPU workloads
- Deep understanding of Nvidia GPU operator and GPU scheduling within Kubernetes
- Experience building multi-tenant Kubernetes platforms with strong isolation guarantees
- Proficiency in Terraform, Helm, and GitOps workflows
- Strong programming skills in Go or Python for operator/CRD development
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 42,189 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee