Inference Engineer

San Francisco
Hybridmid
Confirmed open34 minutes ago
First seentoday
Remoet re-checks every listing roughly once a day.
Apply on AdaptionYou will apply on Adaption's own site.Star Adaption to hear about new roles →

Mid-level Inference Engineer role in San Francisco focusing on model serving performance using vLLM, SGLang, and TensorRT-LLM.

Summary, tech stack, seniority and salary here are extracted or inferred by Remoet, not the employer's own words.

Tech stack

PythonC++RustvLLMSGLangTensorRT-LLMCUDANCCL

Benefits

Flexible workAnnual travel stipendWeekly meal allowanceComprehensive medical benefitsGenerous paid time off
Confirmed open34 minutes ago
First seentoday
Remoet re-checks every listing roughly once a day.
Apply on AdaptionYou will apply on Adaption's own site.Star Adaption to hear about new roles →