LLM Inference Frameworks and Optimization Engineer

San Francisco, Singapore, Amsterdam
Hybridmid
Confirmed open63 days ago
First seen2 months ago
Remoet re-checks every listing roughly once a day.
Apply on Together AIYou will apply on Together AI's own site.Star Together AI to hear about new roles →

Mid-level engineer for LLM inference optimization using Python, C++, and CUDA in a hybrid work environment.

Summary, tech stack, seniority and salary here are extracted or inferred by Remoet, not the employer's own words.

Salary

$160,000 - $230,000/year

Tech stack

PythonC++CUDATritonTensorRTTensorRT-LLMvLLMSGLangTGIPyTorchKubernetes

Benefits

competitive compensationstartup equityhealth insurance
Confirmed open63 days ago
First seen2 months ago
Remoet re-checks every listing roughly once a day.
Apply on Together AIYou will apply on Together AI's own site.Star Together AI to hear about new roles →