AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1297 open roles · 85 companies · last verified today
83 roles
Nscaleneocloud
Houston · New York +2 more · $220K–$293K · staff plus
inferenceinference-engineskv-cache-systemspost-trainingpython-lang
posted 1d ago · verified today
SambaNovachip vendor
Tokyo, Japan · Tokyo Prefecture, Japan · 105K–130K JPY · senior
inferencepython-langsolutions-architectperformance-engineerevaluation
posted 2d ago · verified today
DoorDashenterprise
San Francisco · San Francisco, CA +2 more · senior
fine-tuninggpu-genericinferenceinference-enginesml-platform
posted 10w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
trainiuminferenceinference-enginesmoe-systemsperformance-engineer
posted 6d ago · verified today
Inferactinference provider
San Francisco · onsite · junior
amdcudagpu-kernelsinferenceinference-engines
posted 1d ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
trainiuminferenceinference-enginesmoe-systemsperformance-engineer
posted 7d ago · verified today
Reactorinference provider
San Francisco · onsite · unknown
inferenceinference-enginescudagpu-kernelsperformance-engineer
added 7d ago · verified today
DigitalOceanneocloud
Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus
amdcudainferencekubernetes-opsnvidia
posted 4mo ago · verified today
DigitalOceanneocloud
Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus
amdcudadistributed-inferencego-langinference
posted 4mo ago · verified today
DigitalOceanneocloud
Boston · Seattle Metro · $191K–$239K/year est. · staff plus
cudagpu-kernelsinferenceinference-enginesperformance-engineer
posted 7w ago · verified today
DigitalOceanneocloud
Denver · Seattle Metro · $191K–$239K/year est. · staff plus
gpu-kernelscudainferenceinference-enginesperformance-engineer
posted 7w ago · verified today
DigitalOceanneocloud
San Francisco · Seattle Metro · $191K–$239K/year est. · staff plus
amdcudadistributed-inferenceflash-attentiongpu-kernels
posted 7w ago · verified today
DigitalOceanneocloud
Austin · Seattle Metro · $191K–$239K/year est. · staff plus
amdcudadistributed-inferenceflash-attentiongpu-kernels
posted 7w ago · verified today
DigitalOceanneocloud
Seattle · Seattle Metro · $191K–$239K/year est. · staff plus
gpu-kernelsinferenceinference-enginescudadistributed-inference
posted 11w ago · verified today
DigitalOceanneocloud
DigitalOcean Seattle Office · Seattle · junior
go-langgpu-genericinferencepython-langsoftware-engineer
posted 3w ago · verified today
DigitalOceanneocloud
Bay Area Metro · New York · $150K–$186K/year est. · senior
solutions-architectcudagpu-genericinferencekubernetes-ops
posted 5mo ago · verified today
DigitalOceanneocloud
Bay Area Metro · San Francisco · $150K–$215K/year est. · senior
cudakubernetes-opssolutions-architectfine-tuninggpu-generic
posted 5mo ago · verified today
DigitalOceanneocloud
Bangalore Metro · Bengaluru · senior
cudakubernetes-opsnvidiaamdrocm-hip
posted 11w ago · verified today
DigitalOceanneocloud
Bangalore Metro · Bengaluru · senior
distributed-inferenceinferenceinference-engineskv-cache-systemsperformance-engineer
posted 6w ago · verified today
DeepInfrainference provider
Palo Alto, United States · onsite · unknown
inferenceinference-enginesnvidiapython-langsolutions-architect
posted 4w ago · verified today
Sakana AIfrontier lab
Tokyo · unknown
reliability-srenvidiainferencesregpu-generic
added 7d ago · verified today
Interfazeinference provider
San Francisco, CA · senior
fine-tuninginferenceinference-enginespython-langquantization
added 7d ago · verified today
Inceptionfrontier lab
San Mateo, United States · onsite · unknown
cudagpu-genericinferenceinference-engineskubernetes-ops
posted 6mo ago · verified today
Black Forest Labsfrontier lab
San Francisco (United States) · onsite · unknown
cudagpu-genericinferenceinference-enginesperformance-engineer
posted 24mo ago · verified today
Together AIneocloud
San Francisco · $270K–$300K/year est. · senior
distributed-inferencefine-tuninginferenceinference-engineskv-cache-systems
posted 4mo ago · verified today
Lila Sciencesai startup
Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus
inferenceinference-engineskubernetes-opsml-platformnvidia
posted 7w ago · verified today
Genesis Molecular AIai startup
New York, NY · San Mateo, CA · hybrid · senior
research-engineergpu-genericpre-trainingpython-langcuda
posted 13mo ago · verified today
ElevenLabsai startup
Bulgaria · Poland +2 more · remote · unknown
cudagpu-genericgpu-kernelsinferenceinference-engines
posted 19d ago · verified today
Figureai startup
HQ · San Jose, CA · $180K–$275K/year est. · staff plus
cpp-langpython-langquantizationgpu-genericperformance-engineer
posted 11w ago · verified today
Tenstorrentchip vendor
Cyprus · unknown
software-engineergpu-kernelsperformance-engineerml-platformpre-training
posted 11mo ago · verified today