AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1297 open roles · 85 companies · last verified today
195 roles
Radiantneocloud
UK · staff plus
kubernetes-opsansiblecluster-datacentergo-langgpu-virtualization
posted 1d ago · verified today
NexGen Cloudneocloud
Finland · Finland - Remote · mid
cluster-datacenterkubernetes-opsgpu-genericreliability-srescheduling-orchestration
posted 1d ago · verified today
Nscaleneocloud
Houston · New York +2 more · manager
network-engineernetwork-fabriccluster-datacenterreliability-sregpu-generic
posted 2d ago · verified today
xAIfrontier lab
Palo Alto, CA · $150K–$250K/year est. · unknown
network-engineernetwork-fabriccluster-datacentergpu-genericroce-net
posted 2d ago · verified today
xAIfrontier lab
Dublin · Dublin, Ireland · $100K–$150K/year est. · unknown
network-engineeransiblecluster-datacenternetwork-fabricpython-lang
posted 2d ago · verified today
Amazon (AWS)hyperscaler
Austin, Texas, USA · Dallas, Texas, USA +2 more · onsite · staff plus
cluster-datacentercollectivesdeepspeed-libdistributed-inferenceefa-fabric
posted 7w ago · verified today
CoreWeaveneocloud
Bellevue, WA · San Francisco, CA +2 more · $182K–$242K/year est. · senior
kubernetes-opssolutions-architectnvidiainfiniband-opsnccl-lib
posted 5d ago · verified today
NexGen Cloudneocloud
London · mid
network-engineernetwork-fabriccluster-datacentergpu-genericnvidia
posted 4w ago · verified today
Nscaleneocloud
Singapore · senior
cluster-datacenterdatacenter-engineernetwork-fabricobservabilitygpu-generic
posted 5d ago · verified today
Reactorinference provider
San Francisco · onsite · unknown
inferenceinference-enginescudagpu-kernelsperformance-engineer
added 7d ago · verified today
Boson AIfrontier lab
Toronto · onsite · CA$125K–CA$250K/year · unknown
nvidiareliability-sresreansiblecluster-datacenter
posted 9w ago · verified today
Boson AIfrontier lab
Barrie · onsite · CA$50K–CA$100K/year · unknown
cluster-datacenterdatacenter-engineernvidiagpu-genericnetwork-fabric
posted 19d ago · verified today
DigitalOceanneocloud
Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus
amdcudainferencekubernetes-opsnvidia
posted 4mo ago · verified today
DigitalOceanneocloud
Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus
amdcudadistributed-inferencego-langinference
posted 4mo ago · verified today
DigitalOceanneocloud
Boston · Seattle Metro · $191K–$239K/year est. · staff plus
cudagpu-kernelsinferenceinference-enginesperformance-engineer
posted 7w ago · verified today
DigitalOceanneocloud
Denver · Seattle Metro · $191K–$239K/year est. · staff plus
gpu-kernelscudainferenceinference-enginesperformance-engineer
posted 7w ago · verified today
DigitalOceanneocloud
San Francisco · Seattle Metro · $191K–$239K/year est. · staff plus
amdcudadistributed-inferenceflash-attentiongpu-kernels
posted 7w ago · verified today
DigitalOceanneocloud
Austin · Seattle Metro · $191K–$239K/year est. · staff plus
amdcudadistributed-inferenceflash-attentiongpu-kernels
posted 7w ago · verified today
DigitalOceanneocloud
Seattle · Seattle Metro · $191K–$239K/year est. · staff plus
gpu-kernelsinferenceinference-enginescudadistributed-inference
posted 11w ago · verified today
DigitalOceanneocloud
Bay Area Metro · New York · $150K–$186K/year est. · senior
solutions-architectcudagpu-genericinferencekubernetes-ops
posted 5mo ago · verified today
DigitalOceanneocloud
Bay Area Metro · San Francisco · $150K–$215K/year est. · senior
cudakubernetes-opssolutions-architectfine-tuninggpu-generic
posted 5mo ago · verified today
DigitalOceanneocloud
Bangalore Metro · Bengaluru · senior
cudakubernetes-opsnvidiaamdrocm-hip
posted 11w ago · verified today
Scalewayneocloud
Paris · hybrid · manager
eng-managercluster-datacenterkubernetes-opsnvidiareliability-sre
posted 3w ago · verified today
DeepInfrainference provider
Palo Alto, United States · onsite · unknown
inferenceinference-enginesnvidiapython-langsolutions-architect
posted 4w ago · verified today
Sakana AIfrontier lab
Tokyo · unknown
reliability-srenvidiainferencesregpu-generic
added 7d ago · verified today
Runwareinference provider
United Kingdom · remote · senior
gpu-genericinferencenvidiareliability-sresre
posted 4mo ago · verified today
Black Forest Labsfrontier lab
Freiburg (Germany) · onsite · unknown
cluster-datacentergo-langkubernetes-opsnvidiaobservability
posted 12mo ago · verified today
Baseteninference provider
San Francisco · hybrid · $170K–$230K/year · manager
cluster-datacentergo-langgpu-generickubernetes-opsnvidia
posted 8d ago · verified today
Nebiusneocloud
Abu Dhabi, UAE · Dubai +1 more · senior
solutions-architectansiblecudagpu-generickubernetes-ops
posted 9d ago · verified today
Lila Sciencesai startup
Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus
inferenceinference-engineskubernetes-opsml-platformnvidia
posted 7w ago · verified today