AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1297 open roles · 85 companies · last verified today
153 roles
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA · $152K–$204K/year est. · senior
cudadistributed-inferencego-langgpu-genericgpu-kernels
posted 11mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA · $139K–$204K/year est. · senior
software-engineerinferenceinference-engineskubernetes-opsgpu-generic
posted 7mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA · $206K–$333K/year est. · staff plus
cudaperformance-engineergpu-genericinferenceinference-engines
posted 9mo ago · verified today
CoreWeaveneocloud
Atlanta, GA · San Francisco, CA · $182K–$242K/year est. · unknown
solutions-architectpython-langnvidiafine-tuninginfiniband-ops
posted 6w ago · verified today
CoreWeaveneocloud
Bellevue, WA · Seattle, WA · $182K–$242K/year est. · unknown
solutions-architectnvidiapython-langvllm-enginegpu-generic
posted 6w ago · verified today
CoreWeaveneocloud
Atlanta, GA · Austin, TX +1 more · $182K–$242K/year est. · unknown
solutions-architectpython-langevaluationfine-tuninginference
posted 4w ago · verified today
CoreWeaveneocloud
San Francisco, CA · $143K–$210K/year est. · mid
python-langsolutions-architectgpu-generickubernetes-opsslurm-admin
posted 3mo ago · verified today
CoreWeaveneocloud
Manhattan, NY · New York, NY · $143K–$210K/year est. · unknown
python-langsolutions-architectgpu-genericnvidiafine-tuning
posted 3mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $293K–$385K/year · unknown
evaluationgpu-kernelsinferencesoftware-engineercollectives
posted 4mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $293K–$445K/year · senior
inference-enginesml-platformcpp-langgo-langinference
posted 8w ago · verified today
OpenAIfrontier lab
San Francisco · $295K–$555K/year · unknown
gpu-genericinferenceinference-enginessoftware-engineerdistributed-inference
posted 16mo ago · verified today
Nebiusneocloud
Amsterdam · Amsterdam, Netherlands +7 more · senior
go-langgpu-generickubernetes-opsscheduling-orchestrationsoftware-engineer
posted 7w ago · verified today
Nebiusneocloud
Amsterdam · Amsterdam, Netherlands +14 more · remote · senior
go-langinferenceinference-engineskv-cache-systemspython-lang
posted 11w ago · verified today
Nebiusneocloud
Amsterdam · Amsterdam, Netherlands +6 more · remote · senior
sregpu-genericinferenceinference-engineskubernetes-ops
posted 15mo ago · verified today
Nebiusneocloud
Remote - Europe · remote · senior
inferencepython-langsolutions-architectfine-tuningvllm-engine
posted 6mo ago · verified today
Nebiusneocloud
Singapore · senior
solutions-architectinferencevllm-enginefine-tuningpython-lang
posted 4mo ago · verified today
Nebiusneocloud
Amsterdam, Netherlands · Berlin, Germany +6 more · remote · senior
gpu-genericinferenceinference-enginesml-platformperformance-engineer
posted 6mo ago · verified today
Nebiusneocloud
Amsterdam, Netherlands · Israel +3 more · senior
cudagpu-genericgpu-kernelsinference-enginesinference
posted 7mo ago · verified today
Nebiusneocloud
Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior
inferenceinference-enginespython-langquantizationvllm-engine
posted 8w ago · verified today
Nebiusneocloud
Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior
research-engineercudainferencepython-langtriton-lang
posted 8w ago · verified today
Nebiusneocloud
United States · $208K–$261K · staff plus
solutions-architectfine-tuninginferenceinference-engineskv-cache-systems
posted 11w ago · verified today
Nebiusneocloud
Amsterdam, Netherlands · Czech Republic +5 more · remote · unknown
cudagpu-genericgpu-kernelsnccl-libperformance-engineer
posted 4mo ago · verified today
Nebiusneocloud
Remote - United States · remote · $180K–$224K · senior
python-langsglang-enginesolutions-architecttensorrt-stackvllm-engine
posted 4mo ago · verified today
Crusoeneocloud
San Francisco, CA - US · Sunnyvale, CA - US · onsite · manager
eng-managerml-platformkubernetes-opsscheduling-orchestrationgo-lang
posted 4w ago · verified today
Crusoeneocloud
San Francisco, CA - US · Sunnyvale, CA - US · onsite · staff plus
fine-tuningsoftware-engineerml-platformpost-trainingreinforcement-learning
posted 6mo ago · verified today
Crusoeneocloud
San Francisco, CA - US · Sunnyvale, CA - US · onsite · staff plus
cpp-langcudadistributed-inferencegpu-genericgpu-kernels
posted 8w ago · verified today
Lambdaneocloud
Bellevue Office · San Francisco Office (Second St) +1 more · remote · $226K–$355K · senior
gpu-genericnvidiasolutions-architectansiblecpp-lang
posted 6w ago · verified today
Together AIneocloud
San Francisco · $220K–$280K/year est. · staff plus
inferenceinference-enginespython-langcudaevaluation
posted 4mo ago · verified today
Together AIneocloud
San Francisco · $200K–$260K/year est. · senior
gpu-kernelsperformance-engineertensorrt-stackvllm-enginegpu-generic
posted 5mo ago · verified today
Together AIneocloud
San Francisco · $200K–$290K/year est. · senior
inferenceinference-enginesnvidiasoftware-engineerdistributed-inference
posted 13mo ago · verified today