AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

39 roles

deepspeed-lib

Baseteninference provider

New York · San Francisco · hybrid · $165K–$330K/year · unknown

ml-platformsoftware-engineerfine-tuningkubernetes-opspost-training

posted 7mo ago · verified today

Baseteninference provider

New York · San Francisco · hybrid · $165K–$330K/year · unknown

go-langkubernetes-opsscheduling-orchestrationsoftware-engineerml-platform

posted 12mo ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $180K–$255K · senior

inferenceperformance-engineercpp-langdistributed-inferenceinference-engines

posted 9mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $206K–$333K/year est. · staff plus

cudaperformance-engineergpu-genericinferenceinference-engines

posted 9mo ago · verified today

Mistral AIfrontier lab

Palo Alto · San Francisco · hybrid · mid

research-engineerdeepspeed-libfsdpgpu-genericml-platform

posted 9w ago · verified today

Mistral AIfrontier lab

London · Paris +2 more · unknown

research-engineerdeepspeed-libfsdppython-langslurm-admin

posted 10w ago · verified today

Nebiusneocloud

Palo Alto · Palo Alto, California, United States · $195K–$262K · senior

model-parallelismpost-trainingpytorch-distreinforcement-learningtraining-frameworks

posted 7w ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Fremont St) · remote · $240K–$356K · senior

reliability-sresreansiblecluster-datacenterinfiniband-ops

posted 15d ago · verified today

Together AIneocloud

San Francisco · $200K–$290K/year est. · unknown

cudafine-tuningml-platformnccl-libnvidia

posted 6w ago · verified today