AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

10 roles

gpu-virtualization

Lila Sciencesai startup

Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus

inferenceinference-engineskubernetes-opsml-platformnvidia

posted 7w ago · verified today

falinference provider

Remote - USA · remote · $180K–$250K/year · senior

gpu-generickubernetes-opssoftware-engineercluster-datacentergpu-virtualization

posted 4w ago · verified today

CoreWeaveneocloud

Bellevue, WA · San Francisco, CA +1 more · $198K–$264K/year est. · staff plus

amdcluster-datacenterdeepspeed-libdistributed-inferencefsdp

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior

solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks

posted 3mo ago · verified today

FluidStackneocloud

Austin, TX · New York, NY +2 more · onsite · $224K–$344K/year · manager

cluster-datacentereng-managergpu-virtualizationkubernetes-opsreliability-sre

posted 8w ago · verified today

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands +7 more · senior

go-langgpu-generickubernetes-opsscheduling-orchestrationsoftware-engineer

posted 7w ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Fremont St) +1 more · remote · $314K–$465K · staff plus

go-langgpu-generickubernetes-opsml-platformnvidia

posted 5w ago · verified today

Lambdaneocloud

Bellevue Office · San Francisco Office (Fremont St) +1 more · remote · $266K–$395K · senior

go-langgpu-generickubernetes-opsml-platformnvidia

posted 4w ago · verified today