AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1297 open roles · 85 companies · last verified today
190 roles
Nscaleneocloud
Houston · New York +2 more · $220K–$293K · staff plus
inferenceinference-engineskv-cache-systemspost-trainingpython-lang
posted 1d ago · verified today
Cerebraschip vendor
Canada · United States · unknown
cpp-langpython-langreliability-sresoftware-engineercustom-asic
posted 1d ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
inference-enginessoftware-engineersglang-enginetrainiumvllm-engine
posted 4d ago · verified today
Amazon (AWS)hyperscaler
Toronto, Ontario, CAN · onsite · junior
trainiumpython-langsoftware-engineergpu-kernelsperformance-engineer
posted 5d ago · verified today
Amazon (AWS)hyperscaler
Austin, Texas, USA · Cupertino, California, USA +1 more · onsite · senior
cluster-datacentergpu-genericreliability-srelinux-kernelsoftware-engineer
posted 5d ago · verified today
Inferactinference provider
San Francisco · onsite · junior
amdcudagpu-kernelsinferenceinference-engines
posted 1d ago · verified today
Runwareinference provider
United Kingdom · remote · senior
inferenceinference-enginesgpu-genericgpu-kernelsperformance-engineer
posted 7d ago · verified today
Reactorinference provider
San Francisco · onsite · unknown
inferenceinference-enginescudagpu-kernelsperformance-engineer
added 7d ago · verified today
DigitalOceanneocloud
Boston · Seattle Metro · $191K–$239K/year est. · staff plus
cudagpu-kernelsinferenceinference-enginesperformance-engineer
posted 7w ago · verified today
DigitalOceanneocloud
Denver · Seattle Metro · $191K–$239K/year est. · staff plus
gpu-kernelscudainferenceinference-enginesperformance-engineer
posted 7w ago · verified today
DigitalOceanneocloud
San Francisco · Seattle Metro · $191K–$239K/year est. · staff plus
amdcudadistributed-inferenceflash-attentiongpu-kernels
posted 7w ago · verified today
DigitalOceanneocloud
Austin · Seattle Metro · $191K–$239K/year est. · staff plus
amdcudadistributed-inferenceflash-attentiongpu-kernels
posted 7w ago · verified today
DigitalOceanneocloud
Seattle · Seattle Metro · $191K–$239K/year est. · staff plus
gpu-kernelsinferenceinference-enginescudadistributed-inference
posted 11w ago · verified today
DigitalOceanneocloud
Bay Area Metro · San Francisco · $150K–$215K/year est. · senior
cudakubernetes-opssolutions-architectfine-tuninggpu-generic
posted 5mo ago · verified today
DeepInfrainference provider
Sofia, Bulgaria · onsite · junior
inferenceinference-enginespython-langsoftware-engineercpp-lang
posted 11w ago · verified today
DeepInfrainference provider
Bulgaria - Remote · remote · mid
cpp-langcudagpu-genericinferenceinference-engines
posted 5mo ago · verified today
DeepInfrainference provider
Palo Alto, United States · onsite · mid
cudainferencepython-langsoftware-engineercpp-lang
posted 5mo ago · verified today
DeepInfrainference provider
Palo Alto, United States · onsite · $140K–$150K/year est. · junior
inference-enginessoftware-engineercpp-langcudaml-platform
posted 8mo ago · verified today
DeepInfrainference provider
Bulgaria - Remote · remote · junior
cpp-langcudainferenceinference-enginespython-lang
posted 11mo ago · verified today
Waferinference provider
San Francisco · onsite · $200K–$300K/year · unknown
gpu-kernelsinferenceinference-enginescluster-datacenterperformance-engineer
posted 8w ago · verified today
Relaceinference provider
San Francisco · onsite · mid
cudagpu-kernelsperformance-engineersoftware-engineercpp-lang
posted 10mo ago · verified today
Black Forest Labsfrontier lab
Freiburg (Germany) · onsite · staff plus
pre-trainingfsdppython-langpytorch-distresearch-engineer
posted 4mo ago · verified today
Crusoeneocloud
San Francisco, CA - US · onsite · staff plus
inferenceinference-enginesperformance-engineersglang-enginevllm-engine
posted 7d ago · verified today
Xaira Therapeuticsai startup
Seattle, WA · Seattle, Washington, United States +1 more · $185K–$308K/year est. · senior
gpu-genericgpu-kernelsinferenceresearch-engineerinference-engines
posted 10mo ago · verified today
Fireworks AIinference provider
New York · San Mateo · hybrid · $200K–$230K/year · unknown
scheduling-orchestrationcpp-langnetwork-fabricpython-langstorage-checkpointing
posted 8d ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
distributed-inferenceinferenceinference-enginespython-langpytorch-dist
posted 9d ago · verified today
Lila Sciencesai startup
Cambridge, MA USA · One Charles Park, Cambridge, MA +1 more · $189K–$289K · unknown
post-trainingpython-langresearch-engineerreinforcement-learningtraining-frameworks
posted 11mo ago · verified today
Prior Labsai startup
Berlin · Freiburg +1 more · onsite · unknown
cluster-datacentergpu-genericpython-langpytorch-distscheduling-orchestration
posted 6w ago · verified today
Abridgeai startup
SF Office · hybrid · $221K–$260K/year · senior
inferenceinference-engineskubernetes-opssoftware-engineercuda
posted 12mo ago · verified today
Genesis Molecular AIai startup
New York, NY · San Mateo, CA · hybrid · senior
research-engineergpu-genericpre-trainingpython-langcuda
posted 13mo ago · verified today