AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1297 open roles · 85 companies · last verified today
380 roles
Modalinference provider
San Francisco · onsite · $300K–$350K/year · manager
eng-managergpu-genericinferenceinference-enginesml-platform
posted today · verified today
Nscaleneocloud
Houston · New York +2 more · $220K–$293K · staff plus
inferenceinference-engineskv-cache-systemspost-trainingpython-lang
posted 1d ago · verified today
Amazon (AWS)hyperscaler
Austin, Texas, USA · New York, New York, USA +2 more · onsite · senior
solutions-architectkubernetes-opsmpinccl-libtraining-frameworks
posted 2d ago · verified today
OpenAIfrontier lab
San Francisco · hybrid · $266K–$500K/year · unknown
software-engineerml-platformscheduling-orchestrationevaluationgpu-generic
posted 2d ago · verified today
SambaNovachip vendor
Tokyo, Japan · Tokyo Prefecture, Japan · 105K–130K JPY · senior
inferencepython-langsolutions-architectperformance-engineerevaluation
posted 2d ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
inference-enginessoftware-engineersglang-enginetrainiumvllm-engine
posted 4d ago · verified today
DigitalOceanneocloud
DigitalOcean Seattle Office · Seattle · $167K–$209K/year est. · senior
go-langinferencekubernetes-opsgpu-genericinference-engines
posted 5d ago · verified today
Ai2frontier lab
Seattle · Seattle, WA · $147K–$220K/year est. · senior
python-langresearch-engineertraining-frameworksml-platformpre-training
posted 5mo ago · verified today
DoorDashenterprise
San Francisco · San Francisco, CA +2 more · senior
fine-tuninggpu-genericinferenceinference-enginesml-platform
posted 10w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
trainiuminferenceinference-enginesmoe-systemsperformance-engineer
posted 6d ago · verified today
Inferactinference provider
San Francisco · onsite · junior
amdcudagpu-kernelsinferenceinference-engines
posted 1d ago · verified today
Mistral AIfrontier lab
Montréal · New York +1 more · remote · senior
solutions-architectgpu-genericscheduling-orchestrationcluster-datacenterinference
posted 6d ago · verified today
Anthropicfrontier lab
San Francisco, CA · $320K–$485K · staff plus
inferenceinference-enginessoftware-engineerreliability-sre
posted 7d ago · verified today
Anthropicfrontier lab
San Francisco, CA · $320K–$485K · staff plus
inferenceinference-enginesobservabilitypython-langsoftware-engineer
posted 7d ago · verified today
Anthropicfrontier lab
New York City, NY · San Francisco, CA · $350K–$850K · unknown
inferenceinference-enginesperformance-engineergpu-genericrust-lang
posted 6d ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
trainiuminferenceinference-enginesmoe-systemsperformance-engineer
posted 7d ago · verified today
Amazon (AWS)hyperscaler
Austin, Texas, USA · Chicago, Illinois, USA +5 more · onsite · staff plus
solutions-architectcluster-datacentercollectivesefa-fabricfsdp
posted 7d ago · verified today
Scale AIai startup
Washington, DC · $214K–$267K · manager
eng-managerinferenceinference-engineskubernetes-opsreliability-sre
posted 7d ago · verified today
Runwareinference provider
United Kingdom · remote · senior
inferenceinference-enginesgpu-genericgpu-kernelsperformance-engineer
posted 7d ago · verified today
Reactorinference provider
San Francisco · onsite · unknown
gpu-generickubernetes-opsml-platformobservabilityterraform-iac
added 7d ago · verified today
Reactorinference provider
San Francisco · onsite · mid
inferencepython-langsolutions-architectinference-enginesperformance-engineer
added 7d ago · verified today
Reactorinference provider
San Francisco · onsite · unknown
inferenceinference-enginescudagpu-kernelsperformance-engineer
added 7d ago · verified today
Boson AIfrontier lab
Toronto · onsite · CA$125K–CA$250K/year · unknown
nvidiareliability-sresreansiblecluster-datacenter
posted 9w ago · verified today
MongoDBai startup
Sydney · mid
inferenceinference-enginessoftware-engineerml-platformgo-lang
posted 5w ago · verified today
DigitalOceanneocloud
Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus
amdcudainferencekubernetes-opsnvidia
posted 4mo ago · verified today
DigitalOceanneocloud
Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus
amdcudadistributed-inferencego-langinference
posted 4mo ago · verified today
DigitalOceanneocloud
Boston · Seattle Metro · $191K–$239K/year est. · staff plus
cudagpu-kernelsinferenceinference-enginesperformance-engineer
posted 7w ago · verified today
DigitalOceanneocloud
Denver · Seattle Metro · $191K–$239K/year est. · staff plus
gpu-kernelscudainferenceinference-enginesperformance-engineer
posted 7w ago · verified today
DigitalOceanneocloud
San Francisco · Seattle Metro · $191K–$239K/year est. · staff plus
amdcudadistributed-inferenceflash-attentiongpu-kernels
posted 7w ago · verified today
DigitalOceanneocloud
Austin · Seattle Metro · $191K–$239K/year est. · staff plus
amdcudadistributed-inferenceflash-attentiongpu-kernels
posted 7w ago · verified today