AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

85 roles

triton-lang

Nscaleneocloud

Houston · New York +2 more · $220K–$293K · staff plus

inferenceinference-engineskv-cache-systemspost-trainingpython-lang

posted 1d ago · verified today

DigitalOceanneocloud

DigitalOcean Seattle Office · Seattle · $167K–$209K/year est. · senior

go-langinferencekubernetes-opsgpu-genericinference-engines

posted 5d ago · verified today

Inferactinference provider

San Francisco · onsite · junior

amdcudagpu-kernelsinferenceinference-engines

posted 1d ago · verified today

Runwareinference provider

United Kingdom · remote · senior

inferenceinference-enginesgpu-genericgpu-kernelsperformance-engineer

posted 7d ago · verified today

DigitalOceanneocloud

Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus

amdcudainferencekubernetes-opsnvidia

posted 4mo ago · verified today

DigitalOceanneocloud

Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus

amdcudadistributed-inferencego-langinference

posted 4mo ago · verified today

DigitalOceanneocloud

Boston · Seattle Metro · $191K–$239K/year est. · staff plus

cudagpu-kernelsinferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

Denver · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelscudainferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

San Francisco · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Austin · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Seattle · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelsinferenceinference-enginescudadistributed-inference

posted 11w ago · verified today

Fish Audio (39 AI)ai startup

Location not specified · unknown

cpp-langinference-enginespython-langinferencemegatron-lm

added 7d ago · verified today

Runwareinference provider

United Kingdom · remote · senior

gpu-genericinferencenvidiareliability-sresre

posted 4mo ago · verified today

Fireworks AIinference provider

New York · San Mateo · hybrid · $200K–$230K/year · unknown

scheduling-orchestrationcpp-langnetwork-fabricpython-langstorage-checkpointing

posted 8d ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · senior

distributed-inferenceinferenceinference-enginespython-langpytorch-dist

posted 9d ago · verified today

Prior Labsai startup

Berlin · Freiburg +1 more · onsite · unknown

cluster-datacentergpu-genericpython-langpytorch-distscheduling-orchestration

posted 6w ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · unknown

cudagpu-kernelstriton-langgpu-genericnvidia

posted 7w ago · verified today

Cartesiaai startup

*HQ - San Francisco, CA · onsite · $180K–$250K/year · unknown

inferenceinference-enginessoftware-engineerml-platformdistributed-inference

posted 21mo ago · verified today

ElevenLabsai startup

Bulgaria · Poland +2 more · remote · unknown

cudagpu-genericgpu-kernelsinferenceinference-engines

posted 19d ago · verified today

Figureai startup

HQ · San Jose, CA · $200K–$400K/year est. · senior

performance-engineercollectivescpp-langcudafsdp

posted 4w ago · verified today

1Xai startup

San Carlos, CA · onsite · $250K–$350K/year · unknown

gpu-genericpre-trainingpython-langresearch-engineertraining-data-infra

posted 3mo ago · verified today

Lightning AIai startup

London, UK · New York, New York +6 more · remote · $165K–$310K · senior

python-langresearch-engineertraining-frameworksevaluationfine-tuning

posted 5w ago · verified today

Lightning AIai startup

London, England, United Kingdom · London, UK +6 more · remote · $165K–$310K · unknown

post-trainingresearch-engineercudatraining-frameworksdeepspeed-lib

posted 26mo ago · verified today

Coherefrontier lab

London · Montreal +3 more · hybrid · unknown

cudagpu-kernelsperformance-engineertriton-langpre-training

posted 19mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

gpu-genericpost-traininggpu-kernelsmodel-parallelismtraining-frameworks

posted 7mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

distributed-inferenceinferenceinference-enginesperformance-engineernvidia

posted 4mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

amdnvidiagpu-kernelscollectivescpp-lang

posted 6w ago · verified today