AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
985 open roles · 19 companies · last verified today
82 roles
Nebiusneocloud
Amsterdam · Amsterdam, Netherlands; London, United Kingdom +5 more · senior
go-langkubernetes-opsscheduling-orchestrationsoftware-engineergpu-generic
posted 3d ago · verified today
Nebiusneocloud
Remote - United States · remote · mid
inferenceinference-enginestensorrt-stackvllm-engineml-platform
posted 3d ago · verified today
Nebiusneocloud
Canada · Canada; Remote - United States +1 more · remote · manager
solutions-architectcluster-datacenterdistributed-inferencefine-tuninginference
posted 3d ago · verified today
Nebiusneocloud
Amsterdam, Netherlands; Remote - Europe; Remote - United States · Czech Republic +3 more · remote · mid
performance-engineercudagpu-genericgpu-kernelsnccl-lib
posted 3d ago · verified today
Nebiusneocloud
Remote - Europe · remote · senior
inferenceinference-enginespython-langsolutions-architectvllm-engine
posted 3d ago · verified today
Nebiusneocloud
Amsterdam, Netherlands; Berlin, Germany; Israel; London, United Kingdom; Prague, Czech Republic; Remote - Europe · Israel +3 more · remote · senior
gpu-genericinferenceinference-enginesperformance-engineerpython-lang
posted 3d ago · verified today
Nebiusneocloud
Amsterdam, Netherlands · Israel +3 more · senior
cpp-langcudagpu-genericgpu-kernelsinference
posted 3d ago · verified today
Nebiusneocloud
Singapore · senior
inferencesolutions-architectfine-tuninginference-enginespython-lang
posted 3d ago · verified today
Nebiusneocloud
Netherlands · Remote +2 more · remote · mid
cpp-langgpu-genericgpu-kernelssoftware-engineerinference
posted 3d ago · verified today
Nebiusneocloud
Amsterdam · Amsterdam, Netherlands +13 more · senior
distributed-inferenceinferenceinference-enginesvllm-enginego-lang
posted 3d ago · verified today
Nebiusneocloud
Palo Alto, California, United States · San Francisco Bay Area · senior
inferenceinference-enginespython-langresearch-engineergpu-generic
posted 3d ago · verified today
Nebiusneocloud
Palo Alto, California, United States · San Francisco Bay Area · senior
distributed-inferencegpu-genericinferenceinference-engineskv-cache-systems
posted 3d ago · verified today
Nebiusneocloud
United States · staff plus
solutions-architectfine-tuninginferenceinference-enginespython-lang
posted 3d ago · verified today
Lambdaneocloud
Bellevue Office · San Francisco Office (Second St) +1 more · remote · $226K–$355K · senior
nvidiasolutions-architectdistributed-inferencegpu-genericinference
posted 16d ago · verified today
Together AIneocloud
Remote · Singapore · senior
fine-tuninginferenceinference-engineskv-cache-systemspost-training
posted 3d ago · verified today
Together AIneocloud
Remote · San Francisco, Singapore, Amsterdam · mid
cpp-langcudadistributed-inferencegpu-genericinference
posted 3d ago · verified today
Together AIneocloud
San Francisco · staff plus
inferenceinference-enginesperformance-engineerpython-langgpu-generic
posted 3d ago · verified today
Together AIneocloud
San Francisco · mid
fine-tuninginference-enginespython-langreinforcement-learningsglang-engine
posted 3d ago · verified today
Together AIneocloud
San Francisco · senior
inferenceinference-enginescudaperformance-engineerpython-lang
posted 3d ago · verified today
Together AIneocloud
San Francisco · senior
distributed-inferenceinferenceinference-enginespost-trainingpython-lang
posted 3d ago · verified today
Together AIneocloud
San Francisco · senior
inferenceinference-enginespost-trainingpython-langreinforcement-learning
posted 3d ago · verified today
Together AIneocloud
San Francisco · San Francisco · mid
inferenceinference-enginespython-langsoftware-engineergpu-generic
posted 3d ago · verified today