AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

100 roles

sglang-engine

Nscaleneocloud

Houston · New York +2 more · $220K–$293K · staff plus

inferenceinference-engineskv-cache-systemspost-trainingpython-lang

posted 1d ago · verified today

SambaNovachip vendor

Tokyo, Japan · Tokyo Prefecture, Japan · 105K–130K JPY · senior

inferencepython-langsolutions-architectperformance-engineerevaluation

posted 2d ago · verified today

DoorDashenterprise

San Francisco · San Francisco, CA +2 more · senior

fine-tuninggpu-genericinferenceinference-enginesml-platform

posted 10w ago · verified today

Inferactinference provider

San Francisco · onsite · junior

amdcudagpu-kernelsinferenceinference-engines

posted 1d ago · verified today

Mistral AIfrontier lab

Montréal · New York +1 more · remote · senior

solutions-architectgpu-genericscheduling-orchestrationcluster-datacenterinference

posted 6d ago · verified today

DigitalOceanneocloud

Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus

amdcudainferencekubernetes-opsnvidia

posted 4mo ago · verified today

DigitalOceanneocloud

Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus

amdcudadistributed-inferencego-langinference

posted 4mo ago · verified today

DigitalOceanneocloud

Bangalore Metro · Bengaluru · senior

cudakubernetes-opsnvidiaamdrocm-hip

posted 11w ago · verified today

DeepInfrainference provider

Palo Alto, United States · onsite · unknown

inferenceinference-enginesnvidiapython-langsolutions-architect

posted 4w ago · verified today

Fish Audio (39 AI)ai startup

Location not specified · unknown

cpp-langinference-enginespython-langinferencemegatron-lm

added 7d ago · verified today

Inceptionfrontier lab

San Mateo, United States · onsite · unknown

cudagpu-genericinferenceinference-engineskubernetes-ops

posted 6mo ago · verified today

Novita AIinference provider

San Mateo · onsite · unknown

kubernetes-opspython-langsolutions-architectinferenceinference-engines

posted 11w ago · verified today

Crusoeneocloud

San Francisco, CA - US · onsite · staff plus

inferenceinference-enginesperformance-engineersglang-enginevllm-engine

posted 7d ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · unknown

inferenceinference-enginessoftware-engineerkubernetes-opspython-lang

posted 7w ago · verified today

Sesameai startup

Bellevue · New York +1 more · onsite · $175K–$280K/year · senior

inferenceinference-enginesperformance-engineersglang-enginevllm-engine

posted 18mo ago · verified today

Cartesiaai startup

*HQ - San Francisco, CA · onsite · $180K–$250K/year · unknown

inferenceinference-enginessoftware-engineerml-platformdistributed-inference

posted 21mo ago · verified today

ElevenLabsai startup

Bulgaria · Poland +2 more · remote · unknown

cudagpu-genericgpu-kernelsinferenceinference-engines

posted 19d ago · verified today

Tenstorrentchip vendor

Austin, Texas, United States · Santa Clara +2 more · $100K–$500K/year est. · staff plus

inferencekubernetes-opsinference-enginesobservabilitysglang-engine

posted 5w ago · verified today

Lightning AIai startup

London, UK · New York, New York +6 more · remote · $165K–$310K · senior

python-langresearch-engineertraining-frameworksevaluationfine-tuning

posted 5w ago · verified today

Lightning AIai startup

London, England, United Kingdom · London, UK +6 more · remote · $165K–$310K · unknown

post-trainingresearch-engineercudatraining-frameworksdeepspeed-lib

posted 26mo ago · verified today

Etchedchip vendor

San Jose · onsite · junior

cpp-langinferencepython-langdistributed-inferencecollectives

posted 9mo ago · verified today

Coherefrontier lab

Montreal · New York +2 more · remote · unknown

cpp-langgpu-genericinferenceinference-enginesperformance-engineer

posted 10mo ago · verified today

Coherefrontier lab

Montreal · New York +2 more · remote · senior

cpp-langcudagpu-genericgpu-kernelsinference

posted 10mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

gpu-genericpost-traininggpu-kernelsmodel-parallelismtraining-frameworks

posted 7mo ago · verified today