AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

380 roles

inference

Modalinference provider

San Francisco · onsite · $300K–$350K/year · manager

eng-managergpu-genericinferenceinference-enginesml-platform

posted today · verified today

Nscaleneocloud

Houston · New York +2 more · $220K–$293K · staff plus

inferenceinference-engineskv-cache-systemspost-trainingpython-lang

posted 1d ago · verified today

Amazon (AWS)hyperscaler

Austin, Texas, USA · New York, New York, USA +2 more · onsite · senior

solutions-architectkubernetes-opsmpinccl-libtraining-frameworks

posted 2d ago · verified today

OpenAIfrontier lab

San Francisco · hybrid · $266K–$500K/year · unknown

software-engineerml-platformscheduling-orchestrationevaluationgpu-generic

posted 2d ago · verified today

SambaNovachip vendor

Tokyo, Japan · Tokyo Prefecture, Japan · 105K–130K JPY · senior

inferencepython-langsolutions-architectperformance-engineerevaluation

posted 2d ago · verified today

DigitalOceanneocloud

DigitalOcean Seattle Office · Seattle · $167K–$209K/year est. · senior

go-langinferencekubernetes-opsgpu-genericinference-engines

posted 5d ago · verified today

Ai2frontier lab

Seattle · Seattle, WA · $147K–$220K/year est. · senior

python-langresearch-engineertraining-frameworksml-platformpre-training

posted 5mo ago · verified today

DoorDashenterprise

San Francisco · San Francisco, CA +2 more · senior

fine-tuninggpu-genericinferenceinference-enginesml-platform

posted 10w ago · verified today

Inferactinference provider

San Francisco · onsite · junior

amdcudagpu-kernelsinferenceinference-engines

posted 1d ago · verified today

Mistral AIfrontier lab

Montréal · New York +1 more · remote · senior

solutions-architectgpu-genericscheduling-orchestrationcluster-datacenterinference

posted 6d ago · verified today

Anthropicfrontier lab

San Francisco, CA · $320K–$485K · staff plus

inferenceinference-enginessoftware-engineerreliability-sre

posted 7d ago · verified today

Anthropicfrontier lab

San Francisco, CA · $320K–$485K · staff plus

inferenceinference-enginesobservabilitypython-langsoftware-engineer

posted 7d ago · verified today

Anthropicfrontier lab

New York City, NY · San Francisco, CA · $350K–$850K · unknown

inferenceinference-enginesperformance-engineergpu-genericrust-lang

posted 6d ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · senior

trainiuminferenceinference-enginesmoe-systemsperformance-engineer

posted 7d ago · verified today

Amazon (AWS)hyperscaler

Austin, Texas, USA · Chicago, Illinois, USA +5 more · onsite · staff plus

solutions-architectcluster-datacentercollectivesefa-fabricfsdp

posted 7d ago · verified today

Runwareinference provider

United Kingdom · remote · senior

inferenceinference-enginesgpu-genericgpu-kernelsperformance-engineer

posted 7d ago · verified today

Reactorinference provider

San Francisco · onsite · unknown

gpu-generickubernetes-opsml-platformobservabilityterraform-iac

added 7d ago · verified today

Reactorinference provider

San Francisco · onsite · mid

inferencepython-langsolutions-architectinference-enginesperformance-engineer

added 7d ago · verified today

Reactorinference provider

San Francisco · onsite · unknown

inferenceinference-enginescudagpu-kernelsperformance-engineer

added 7d ago · verified today

Boson AIfrontier lab

Toronto · onsite · CA$125K–CA$250K/year · unknown

nvidiareliability-sresreansiblecluster-datacenter

posted 9w ago · verified today

DigitalOceanneocloud

Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus

amdcudainferencekubernetes-opsnvidia

posted 4mo ago · verified today

DigitalOceanneocloud

Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus

amdcudadistributed-inferencego-langinference

posted 4mo ago · verified today

DigitalOceanneocloud

Boston · Seattle Metro · $191K–$239K/year est. · staff plus

cudagpu-kernelsinferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

Denver · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelscudainferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

San Francisco · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Austin · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today