AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

85 roles

triton-lang

Baseteninference provider

Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown

cudagpu-kernelscpp-langgpu-genericinference

posted 14mo ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $180K–$255K · senior

inferenceperformance-engineercpp-langdistributed-inferenceinference-engines

posted 9mo ago · verified today

Anyscaleai startup

Palo Alto · San Francisco · hybrid · $170K–$245K/year · unknown

distributed-inferenceinferenceinference-enginessoftware-engineervllm-engine

posted 3mo ago · verified today

Fireworks AIinference provider

San Mateo · hybrid · $175K–$220K/year · unknown

gpu-kernelsperformance-engineercudagpu-genericsoftware-engineer

posted 16mo ago · verified today

xAIfrontier lab

Palo Alto, CA · $180K–$440K/year est. · unknown

cpp-langgpu-genericgpu-kernelsinferenceinference-engines

posted 23mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $188K–$275K/year est. · staff plus

inferenceinference-enginesdistributed-inferencegpu-generickubernetes-ops

posted 4mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Bellevue, WA - US +5 more · $188K–$303K/year est. · manager

eng-managerinferenceinference-engineskubernetes-opsreliability-sre

posted 9mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $206K–$333K/year est. · staff plus

cudaperformance-engineergpu-genericinferenceinference-engines

posted 9mo ago · verified today

OpenAIfrontier lab

San Francisco · Seattle · hybrid · $293K–$385K/year · unknown

evaluationgpu-kernelsinferencesoftware-engineercollectives

posted 4mo ago · verified today

OpenAIfrontier lab

San Francisco · Seattle · hybrid · $342K–$555K/year · unknown

cudagpu-kernelsperformance-engineertriton-langcluster-datacenter

posted 5mo ago · verified today

OpenAIfrontier lab

San Francisco · Seattle · hybrid · $293K–$445K/year · senior

inference-enginesml-platformcpp-langgo-langinference

posted 8w ago · verified today

Anthropicfrontier lab

New York City, NY · Remote-Friendly (Travel-Required) +2 more · remote · $405K–$485K · staff plus

tputrainiumgpu-genericsoftware-engineercuda

posted 3mo ago · verified today

Anthropicfrontier lab

San Francisco, CA · $350K–$850K · unknown

cudareinforcement-learningjax-pallasresearch-engineerrocm-hip

posted 5mo ago · verified today

Anthropicfrontier lab

New York City, NY · San Francisco, CA +1 more · $280K–$850K · unknown

cudagpu-genericgpu-kernelsperformance-engineercollectives

posted 11mo ago · verified today

Nebiusneocloud

Amsterdam, Netherlands · Berlin, Germany +6 more · remote · senior

gpu-genericinferenceinference-enginesml-platformperformance-engineer

posted 6mo ago · verified today

Nebiusneocloud

Amsterdam, Netherlands · Israel +3 more · senior

cudagpu-genericgpu-kernelsinference-enginesinference

posted 7mo ago · verified today

Nebiusneocloud

Palo Alto · Palo Alto, California, United States · $195K–$262K · senior

model-parallelismpost-trainingpytorch-distreinforcement-learningtraining-frameworks

posted 8w ago · verified today

Nebiusneocloud

Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior

inferenceinference-enginespython-langquantizationvllm-engine

posted 8w ago · verified today

Nebiusneocloud

Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior

research-engineercudainferencepython-langtriton-lang

posted 8w ago · verified today

Crusoeneocloud

San Francisco, CA - US · Sunnyvale, CA - US · onsite · manager

eng-managerml-platformkubernetes-opsscheduling-orchestrationgo-lang

posted 4w ago · verified today

Together AIneocloud

San Francisco · $200K–$300K/year est. · unknown

cudagpu-kernelsperformance-engineerresearch-engineertriton-lang

posted 32mo ago · verified today

Together AIneocloud

San Francisco · $200K–$290K/year est. · senior

inferenceinference-enginesnvidiasoftware-engineerdistributed-inference

posted 13mo ago · verified today

Together AIneocloud

San Francisco · $200K–$290K/year est. · unknown

research-engineerfine-tuninggo-langinferenceinference-engines

posted 10w ago · verified today

Together AIneocloud

San Francisco · $200K–$290K/year est. · unknown

cudafine-tuningml-platformnccl-libnvidia

posted 6w ago · verified today

Together AIneocloud

San Francisco · $200K–$300K/year est. · mid

inferenceinference-enginespython-langpytorch-distsoftware-engineer

posted 27mo ago · verified today