Senior Applied Scientist, Efficient LLM Inference & Model Optimization

Palo Alto
Hybridsenior
Confirmed open26 days ago
First seen4 weeks ago
Remoet re-checks every listing roughly once a day.
Apply on NebiusYou will apply on Nebius's own site.Star Nebius to hear about new roles →

Senior Applied Scientist role focusing on LLM inference optimization using PyTorch and CUDA, offering $195k-$262k annually.

Summary, tech stack, seniority and salary here are extracted or inferred by Remoet, not the employer's own words.

Salary

$195,200 - $262,200/year

Tech stack

PythonPyTorchTritonCUDAvLLMSGLangTensorRT-LLMFlashAttentionFlashInfer

Benefits

100% company-paid medical, dental, and vision coverage401(k) plan with 4% company match20 weeks paid parental leave for primary caregivers12 weeks paid parental leave for secondary caregiversRemote work reimbursementCompany-paid short-term and long-term disability insuranceCompany-paid life insurance
Confirmed open26 days ago
First seen4 weeks ago
Remoet re-checks every listing roughly once a day.
Apply on NebiusYou will apply on Nebius's own site.Star Nebius to hear about new roles →