AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1296 open roles · 85 companies · last verified today
85 roles
Baseteninference provider
Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown
cudagpu-kernelscpp-langgpu-genericinference
posted 14mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · $180K–$255K · senior
inferenceperformance-engineercpp-langdistributed-inferenceinference-engines
posted 9mo ago · verified today
Anyscaleai startup
Palo Alto · San Francisco · hybrid · $170K–$245K/year · unknown
distributed-inferenceinferenceinference-enginessoftware-engineervllm-engine
posted 3mo ago · verified today
Fireworks AIinference provider
San Mateo · hybrid · $175K–$220K/year · unknown
gpu-kernelsperformance-engineercudagpu-genericsoftware-engineer
posted 16mo ago · verified today
xAIfrontier lab
Palo Alto, CA · $180K–$440K/year est. · unknown
cpp-langgpu-genericgpu-kernelsinferenceinference-engines
posted 23mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA · $188K–$275K/year est. · staff plus
inferenceinference-enginesdistributed-inferencegpu-generickubernetes-ops
posted 4mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · Bellevue, WA - US +5 more · $188K–$303K/year est. · manager
eng-managerinferenceinference-engineskubernetes-opsreliability-sre
posted 9mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA · $206K–$333K/year est. · staff plus
cudaperformance-engineergpu-genericinferenceinference-engines
posted 9mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $293K–$385K/year · unknown
evaluationgpu-kernelsinferencesoftware-engineercollectives
posted 4mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $342K–$555K/year · unknown
cudagpu-kernelsperformance-engineertriton-langcluster-datacenter
posted 5mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $293K–$445K/year · senior
inference-enginesml-platformcpp-langgo-langinference
posted 8w ago · verified today
Anthropicfrontier lab
New York City, NY · Remote-Friendly (Travel-Required) +2 more · remote · $405K–$485K · staff plus
tputrainiumgpu-genericsoftware-engineercuda
posted 3mo ago · verified today
Anthropicfrontier lab
San Francisco, CA · $350K–$850K · unknown
cudareinforcement-learningjax-pallasresearch-engineerrocm-hip
posted 5mo ago · verified today
Anthropicfrontier lab
New York City, NY · San Francisco, CA +1 more · $280K–$850K · unknown
cudagpu-genericgpu-kernelsperformance-engineercollectives
posted 11mo ago · verified today
Nebiusneocloud
Amsterdam, Netherlands · Berlin, Germany +6 more · remote · senior
gpu-genericinferenceinference-enginesml-platformperformance-engineer
posted 6mo ago · verified today
Nebiusneocloud
Amsterdam, Netherlands · Israel +3 more · senior
cudagpu-genericgpu-kernelsinference-enginesinference
posted 7mo ago · verified today
Nebiusneocloud
Palo Alto · Palo Alto, California, United States · $195K–$262K · senior
model-parallelismpost-trainingpytorch-distreinforcement-learningtraining-frameworks
posted 8w ago · verified today
Nebiusneocloud
Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior
inferenceinference-enginespython-langquantizationvllm-engine
posted 8w ago · verified today
Nebiusneocloud
Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior
research-engineercudainferencepython-langtriton-lang
posted 8w ago · verified today
Crusoeneocloud
San Francisco, CA - US · Sunnyvale, CA - US · onsite · manager
eng-managerml-platformkubernetes-opsscheduling-orchestrationgo-lang
posted 4w ago · verified today
Together AIneocloud
San Francisco · $200K–$300K/year est. · unknown
cudagpu-kernelsperformance-engineerresearch-engineertriton-lang
posted 32mo ago · verified today
Together AIneocloud
San Francisco · $200K–$290K/year est. · senior
inferenceinference-enginesnvidiasoftware-engineerdistributed-inference
posted 13mo ago · verified today
Together AIneocloud
San Francisco · $200K–$290K/year est. · unknown
research-engineerfine-tuninggo-langinferenceinference-engines
posted 10w ago · verified today
Together AIneocloud
San Francisco · $200K–$290K/year est. · unknown
cudafine-tuningml-platformnccl-libnvidia
posted 6w ago · verified today
Together AIneocloud
San Francisco · $200K–$300K/year est. · mid
inferenceinference-enginespython-langpytorch-distsoftware-engineer
posted 27mo ago · verified today