AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

198 roles

cuda

Nebiusneocloud

Remote - United States · remote · $180K–$224K · senior

python-langsglang-enginesolutions-architecttensorrt-stackvllm-engine

posted 4mo ago · verified today

Nebiusneocloud

Amsterdam · Finland +6 more · remote · unknown

solutions-architectansiblegpu-generickubernetes-opspython-lang

posted 4mo ago · verified today

Crusoeneocloud

San Francisco, CA - US · Sunnyvale, CA - US · onsite · staff plus

cpp-langcudadistributed-inferencegpu-genericgpu-kernels

posted 8w ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Second St) +1 more · remote · $226K–$355K · senior

gpu-genericnvidiasolutions-architectansiblecpp-lang

posted 6w ago · verified today

Lambdaneocloud

Remote, USA · remote · $122K–$162K · senior

collectivescudagpu-genericinfiniband-opskubernetes-ops

posted 3w ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Fremont St) +1 more · remote · $314K–$465K · staff plus

cluster-datacentersoftware-engineercpp-langgpu-genericpython-lang

posted 4w ago · verified today

Together AIneocloud

San Francisco · $200K–$300K/year est. · unknown

cudagpu-kernelsperformance-engineerresearch-engineertriton-lang

posted 32mo ago · verified today

Together AIneocloud

San Francisco · $220K–$280K/year est. · staff plus

inferenceinference-enginespython-langcudaevaluation

posted 4mo ago · verified today

Together AIneocloud

San Francisco · $220K–$290K/year est. · senior

cluster-datacentergpu-genericnvidiasoftware-engineeransible

posted 15mo ago · verified today

Together AIneocloud

San Francisco · $200K–$290K/year est. · senior

inferenceinference-enginesnvidiasoftware-engineerdistributed-inference

posted 13mo ago · verified today

Together AIneocloud

San Francisco · $200K–$290K/year est. · unknown

research-engineerfine-tuninggo-langinferenceinference-engines

posted 10w ago · verified today

Together AIneocloud

San Francisco · $200K–$290K/year est. · unknown

cudafine-tuningml-platformnccl-libnvidia

posted 6w ago · verified today

Together AIneocloud

San Francisco · $200K–$300K/year est. · mid

inferenceinference-enginespython-langpytorch-distsoftware-engineer

posted 27mo ago · verified today

Together AIneocloud

San Francisco · $190K–$270K/year est. · mid

cluster-datacentergpu-genericpython-langreliability-sresoftware-engineer

posted 4mo ago · verified today

Together AIneocloud

Bangalore, India · Remote · unknown

ansiblecluster-datacentergo-langgpu-generickubernetes-ops

posted 9w ago · verified today