AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1294 open roles · 85 companies · last verified today

130 roles

slurm-admin

Lambdaneocloud

Remote, USA · remote · $122K–$162K · senior

collectivescudagpu-genericinfiniband-opskubernetes-ops

posted 3w ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Fremont St) +1 more · remote · $314K–$465K · staff plus

go-langgpu-generickubernetes-opsml-platformnvidia

posted 5w ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Fremont St) +1 more · remote · $266K–$395K · senior

go-langgpu-generickubernetes-opsml-platformnvidia

posted 4w ago · verified today

Together AIneocloud

Remote · San Francisco · remote · $160K–$230K/year est. · senior

gpu-genericinferencekubernetes-opsreliability-sresre

posted 6w ago · verified today

Together AIneocloud

Pune or Bangalore, India · Remote · senior

ansibleinferenceinference-engineskubernetes-opsobservability

posted 6mo ago · verified today

Together AIneocloud

Pune or Bangalore, India · Remote · mid

ansiblegpu-generickubernetes-opscluster-datacenterslurm-admin

posted 12mo ago · verified today

Together AIneocloud

San Francisco · $220K–$290K/year est. · senior

cluster-datacentergpu-genericnvidiasoftware-engineeransible

posted 15mo ago · verified today