AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

190 roles

gpu-kernels

Nscaleneocloud

Houston · New York +2 more · $220K–$293K · staff plus

inferenceinference-engineskv-cache-systemspost-trainingpython-lang

posted 1d ago · verified today

Cerebraschip vendor

Canada · United States · unknown

cpp-langpython-langreliability-sresoftware-engineercustom-asic

posted 1d ago · verified today

Amazon (AWS)hyperscaler

Austin, Texas, USA · Cupertino, California, USA +1 more · onsite · senior

cluster-datacentergpu-genericreliability-srelinux-kernelsoftware-engineer

posted 5d ago · verified today

Inferactinference provider

San Francisco · onsite · junior

amdcudagpu-kernelsinferenceinference-engines

posted 1d ago · verified today

Runwareinference provider

United Kingdom · remote · senior

inferenceinference-enginesgpu-genericgpu-kernelsperformance-engineer

posted 7d ago · verified today

Reactorinference provider

San Francisco · onsite · unknown

inferenceinference-enginescudagpu-kernelsperformance-engineer

added 7d ago · verified today

DigitalOceanneocloud

Boston · Seattle Metro · $191K–$239K/year est. · staff plus

cudagpu-kernelsinferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

Denver · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelscudainferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

San Francisco · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Austin · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Seattle · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelsinferenceinference-enginescudadistributed-inference

posted 11w ago · verified today

DigitalOceanneocloud

Bay Area Metro · San Francisco · $150K–$215K/year est. · senior

cudakubernetes-opssolutions-architectfine-tuninggpu-generic

posted 5mo ago · verified today

DeepInfrainference provider

Sofia, Bulgaria · onsite · junior

inferenceinference-enginespython-langsoftware-engineercpp-lang

posted 11w ago · verified today

DeepInfrainference provider

Bulgaria - Remote · remote · mid

cpp-langcudagpu-genericinferenceinference-engines

posted 5mo ago · verified today

DeepInfrainference provider

Palo Alto, United States · onsite · mid

cudainferencepython-langsoftware-engineercpp-lang

posted 5mo ago · verified today

DeepInfrainference provider

Palo Alto, United States · onsite · $140K–$150K/year est. · junior

inference-enginessoftware-engineercpp-langcudaml-platform

posted 8mo ago · verified today

DeepInfrainference provider

Bulgaria - Remote · remote · junior

cpp-langcudainferenceinference-enginespython-lang

posted 11mo ago · verified today

Waferinference provider

San Francisco · onsite · $200K–$300K/year · unknown

gpu-kernelsinferenceinference-enginescluster-datacenterperformance-engineer

posted 8w ago · verified today

Relaceinference provider

San Francisco · onsite · mid

cudagpu-kernelsperformance-engineersoftware-engineercpp-lang

posted 10mo ago · verified today

Black Forest Labsfrontier lab

Freiburg (Germany) · onsite · staff plus

pre-trainingfsdppython-langpytorch-distresearch-engineer

posted 4mo ago · verified today

Crusoeneocloud

San Francisco, CA - US · onsite · staff plus

inferenceinference-enginesperformance-engineersglang-enginevllm-engine

posted 7d ago · verified today

Xaira Therapeuticsai startup

Seattle, WA · Seattle, Washington, United States +1 more · $185K–$308K/year est. · senior

gpu-genericgpu-kernelsinferenceresearch-engineerinference-engines

posted 10mo ago · verified today

Fireworks AIinference provider

New York · San Mateo · hybrid · $200K–$230K/year · unknown

scheduling-orchestrationcpp-langnetwork-fabricpython-langstorage-checkpointing

posted 8d ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · senior

distributed-inferenceinferenceinference-enginespython-langpytorch-dist

posted 9d ago · verified today

Lila Sciencesai startup

Cambridge, MA USA · One Charles Park, Cambridge, MA +1 more · $189K–$289K · unknown

post-trainingpython-langresearch-engineerreinforcement-learningtraining-frameworks

posted 11mo ago · verified today

Prior Labsai startup

Berlin · Freiburg +1 more · onsite · unknown

cluster-datacentergpu-genericpython-langpytorch-distscheduling-orchestration

posted 6w ago · verified today

Abridgeai startup

SF Office · hybrid · $221K–$260K/year · senior

inferenceinference-engineskubernetes-opssoftware-engineercuda

posted 12mo ago · verified today