Remote Jobs Sign In

AI Inference Engineer

Location: Remote
Compensation: Salary
Reviewed: Mon, Jul 20, 2026
This job expires in: 28 days

Job Summary

To define and build the inference serving strategy at scale, the founding AI Inference Engineer will work remotely, focusing on designing the serving stack and optimizing model-level strategies while collaborating closely with CUDA and GPU engineering teams.

Key responsibilities
  • Define the inference serving strategy and architecture from first principles
  • Design and build the serving stack for high-throughput, latency-sensitive inference workloads
  • Own the model-level optimisation strategy, partnering with CUDA/GPU engineers for integration
Required qualifications
  • 4+ years of experience building or operating large-scale inference serving systems
  • Deep, hands-on experience with inference serving frameworks and optimisation techniques
  • Strong systems thinking with the ability to reason about the full request-response path
  • Proven track record of making high-stakes architecture calls
  • Comfort operating without a playbook in an early-stage function

COMPLETE JOB DESCRIPTION

The job description is available to subscribers. Subscribe today to get the full benefits of a premium membership with Virtual Vocations. We offer the largest remote database online...