AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1295 open roles · 85 companies · last verified today

295 roles

inference-engines

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands +7 more · senior

go-langgpu-generickubernetes-opsscheduling-orchestrationsoftware-engineer

posted 7w ago · verified today

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands +14 more · remote · senior

go-langinferenceinference-engineskv-cache-systemspython-lang

posted 11w ago · verified today

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands +6 more · remote · senior

sregpu-genericinferenceinference-engineskubernetes-ops

posted 15mo ago · verified today

Nebiusneocloud

Amsterdam, Netherlands · Berlin, Germany +6 more · remote · senior

gpu-genericinferenceinference-enginesml-platformperformance-engineer

posted 6mo ago · verified today

Nebiusneocloud

Germany · Israel +5 more · remote · senior

fine-tuninginferencedistributed-inferenceinference-enginesperformance-engineer

posted 7mo ago · verified today

Nebiusneocloud

Amsterdam, Netherlands · Israel +3 more · senior

cudagpu-genericgpu-kernelsinference-enginesinference

posted 7mo ago · verified today

Nebiusneocloud

Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior

inferenceinference-enginespython-langquantizationvllm-engine

posted 8w ago · verified today

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands +3 more · remote · $200–$350K · senior

solutions-architectinference-enginesml-platformgpu-genericinference

posted 6mo ago · verified today

Nebiusneocloud

United States · $208K–$261K · staff plus

solutions-architectfine-tuninginferenceinference-engineskv-cache-systems

posted 11w ago · verified today

Nebiusneocloud

Amsterdam, Netherlands · Czech Republic +5 more · remote · unknown

cudagpu-genericgpu-kernelsnccl-libperformance-engineer

posted 4mo ago · verified today

Nebiusneocloud

Amsterdam, Netherlands · Finland +7 more · remote · mid

python-langsolutions-architectgpu-genericdistributed-inferencekubernetes-ops

posted 19mo ago · verified today

Crusoeneocloud

San Francisco, CA - US · Sunnyvale, CA - US · onsite · manager

eng-managerml-platformkubernetes-opsscheduling-orchestrationgo-lang

posted 4w ago · verified today

Crusoeneocloud

San Francisco, CA - US · Sunnyvale, CA - US · onsite · senior

reliability-sresredistributed-inferenceinferenceinference-engines

posted 3mo ago · verified today

Crusoeneocloud

San Francisco, CA - US · Sunnyvale, CA - US · onsite · staff plus

cpp-langcudadistributed-inferencegpu-genericgpu-kernels

posted 8w ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Second St) +1 more · remote · $226K–$355K · senior

gpu-genericnvidiasolutions-architectansiblecpp-lang

posted 6w ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Fremont St) +1 more · remote · $314K–$465K · staff plus

go-langgpu-generickubernetes-opsml-platformnvidia

posted 5w ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Fremont St) +1 more · remote · $266K–$395K · senior

go-langgpu-generickubernetes-opsml-platformnvidia

posted 4w ago · verified today

Together AIneocloud

Pune or Bangalore, India · Remote · senior

ansibleinferenceinference-engineskubernetes-opsobservability

posted 6mo ago · verified today

Together AIneocloud

San Francisco · $220K–$280K/year est. · staff plus

inferenceinference-enginespython-langcudaevaluation

posted 4mo ago · verified today

Together AIneocloud

San Francisco · $200K–$260K/year est. · senior

gpu-kernelsperformance-engineertensorrt-stackvllm-enginegpu-generic

posted 5mo ago · verified today

Together AIneocloud

San Francisco · $200K–$290K/year est. · senior

inferenceinference-enginesnvidiasoftware-engineerdistributed-inference

posted 13mo ago · verified today

Together AIneocloud

San Francisco · $200K–$290K/year est. · unknown

research-engineerfine-tuninggo-langinferenceinference-engines

posted 10w ago · verified today

Together AIneocloud

San Francisco · $200K–$300K/year est. · mid

inferenceinference-enginespython-langpytorch-distsoftware-engineer

posted 27mo ago · verified today