Remote Jobs Sign In

Tech Lead - AI Inference

Location: Remote
Compensation: To Be Discussed
Reviewed: Thu, Aug 20, 2026
This job expires in: 30 days

Job Summary

Leading a high-performing AI Inference team, the full-time Tech Lead - AI Inference will manage a squad of developers, balancing hands-on technical contributions with leadership to optimize Large Language Model (LLM) serving and drive execution on high-performance systems.

Key responsibilities
  • Take end-to-end ownership of core inference infrastructure, driving technical decisions and delivery outcomes
  • Guide engineers through the design, implementation, and delivery of high-throughput, low-latency LLM inference systems
  • Mentor and coach engineers, fostering a culture of ownership, collaboration, and technical excellence within the team
Required qualifications
  • 5+ years of professional software engineering experience, with a focus on AI/ML infrastructure or high-performance computing
  • Hands-on expertise with LLM serving systems and familiarity with frameworks like vLLM and LMCache
  • Strong Python and C++ skills, with knowledge of CUDA and high-performance I/O systems
  • Experience deploying and scaling GPU workloads on Kubernetes and bare-metal GPU clusters
  • Proven ability to mentor and develop engineers while balancing technical execution with team health

Complete Job Description

The complete job description is available to members. Premium membership includes:

Full access to 47,805 remote jobs from human-vetted companies, updated daily

Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job

Twice-monthly live group coaching and the full Remote Career Center

20% member discount on Career Services

Backed by a 30-day money-back guarantee