ai-infra-jobs

AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

987 open roles · 19 companies · last verified today

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands +13 more · senior

distributed-inferenceinferenceinference-enginesvllm-enginego-lang

posted 3d ago · verified today

Nebiusneocloud

Amsterdam, Netherlands; Remote - Europe; Remote - United States · Czech Republic +3 more · remote · mid

performance-engineercudagpu-genericgpu-kernelsnccl-lib

posted 3d ago · verified today

Nebiusneocloud

Remote - United States · United States · senior

inferenceinference-enginessolutions-architectfine-tuningpython-lang

posted 3d ago · verified today

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands; Berlin, Germany; London, United Kingdom; Prague, Czech Republic; Remote - Europe +3 more · remote · senior

distributed-inferencegpu-genericinferenceinference-engineskubernetes-ops

posted 3d ago · verified today

Crusoeneocloud

San Francisco, CA - US · Sunnyvale, CA - US · senior

reliability-sreinference-enginesgpu-genericobservabilitypython-lang

posted 11w ago · verified today

Crusoeneocloud

San Francisco, CA - US · Sunnyvale, CA - US · staff plus

fine-tuningreinforcement-learningml-platformpost-trainingtraining-frameworks

posted 5mo ago · verified today

Crusoeneocloud

San Francisco, CA - US · Sunnyvale, CA - US · staff plus

cudadistributed-inferencegpu-genericgpu-kernelsinference

posted 4w ago · verified today

Crusoeneocloud

San Francisco, CA - US · Sunnyvale, CA - US · manager

eng-managerkubernetes-opsml-platformsoftware-engineergo-lang

posted 7d ago · verified today

Crusoeneocloud

San Francisco, CA - US · Sunnyvale, CA - US · senior

fine-tuningtraining-frameworksgpu-genericml-platformpytorch-dist

posted 5mo ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Second St) +1 more · remote · $226K–$355K · senior

nvidiasolutions-architectdistributed-inferencegpu-genericinference

posted 17d ago · verified today

Together AIneocloud

Remote · San Francisco, Singapore, Amsterdam · mid

cpp-langcudadistributed-inferencegpu-genericinference

posted 3d ago · verified today

Together AIneocloud

San Francisco · San Francisco · mid

inferenceinference-enginespython-langsoftware-engineergpu-generic

posted 3d ago · verified today