GPU Compute and DPU Engineer
Location: Remote
Compensation: To Be Discussed
Reviewed: Thu, Aug 27, 2026
This job expires in: 29 days
Job Summary
Owning the full lifecycle of bare-metal GPU nodes, the full-time GPU Compute and DPU Engineer will manage provisioning, delivery, and operations while working remotely within San Jose, CA or Austin, TX, to support the scaling of GPU capacity toward a 10,000+ node fleet.
Key Responsibilities
- Oversee the lifecycle of bare-metal GPU nodes including provisioning, delivery, operation, break-fix, and decommissioning across multiple regions
- Develop and maintain automated node-delivery pipelines to address delivery backlogs and support fleet growth
- Manage DPU/SmartNIC and server firmware, ensuring version baselines, upgrades, and validation are effectively executed
Required Qualifications
- 3+ years of experience in large-scale bare-metal/server-fleet operations, HPC, or cloud infrastructure (6+ years for senior roles)
- Hands-on experience with GPU servers at scale, including driver/CUDA and firmware management
- Strong Linux systems skills with experience in PXE/IPMI/Redfish, OS imaging, and automated provisioning
- Familiarity with DPU/SmartNIC technologies and bare-metal networking
- Proficiency in infrastructure automation tools such as Ansible, Terraform, and programming languages like Python or Go
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 47,528 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee