AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

57 roles

kv-cache-systems

Modalinference provider

San Francisco · onsite · $300K–$350K/year · manager

eng-managergpu-genericinferenceinference-enginesml-platform

posted today · verified today

Nscaleneocloud

Houston · New York +2 more · $220K–$293K · staff plus

inferenceinference-engineskv-cache-systemspost-trainingpython-lang

posted 1d ago · verified today

DoorDashenterprise

San Francisco · San Francisco, CA +2 more · senior

fine-tuninggpu-genericinferenceinference-enginesml-platform

posted 10w ago · verified today

Inferactinference provider

San Francisco · onsite · junior

amdcudagpu-kernelsinferenceinference-engines

posted 1d ago · verified today

DigitalOceanneocloud

Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus

amdcudainferencekubernetes-opsnvidia

posted 4mo ago · verified today

DigitalOceanneocloud

Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus

amdcudadistributed-inferencego-langinference

posted 4mo ago · verified today

DigitalOceanneocloud

DigitalOcean Seattle Office · Seattle · junior

go-langgpu-genericinferencepython-langsoftware-engineer

posted 3w ago · verified today

DeepInfrainference provider

Palo Alto, United States · onsite · unknown

inferenceinference-enginesnvidiapython-langsolutions-architect

posted 4w ago · verified today

Fireworks AIinference provider

New York · San Mateo · hybrid · $200K–$230K/year · unknown

scheduling-orchestrationcpp-langnetwork-fabricpython-langstorage-checkpointing

posted 7d ago · verified today

Lila Sciencesai startup

Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus

inferenceinference-engineskubernetes-opsml-platformnvidia

posted 7w ago · verified today

ElevenLabsai startup

Bulgaria · Poland +2 more · remote · unknown

cudagpu-genericgpu-kernelsinferenceinference-engines

posted 19d ago · verified today

Coherefrontier lab

Montreal · New York +2 more · remote · senior

cpp-langcudagpu-genericgpu-kernelsinference

posted 10mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid

gpu-genericinferenceinference-enginesreliability-sreamd

posted 4mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

distributed-inferenceinferenceinference-enginesperformance-engineernvidia

posted 4mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior

inference-enginesgpu-genericinferenceperformance-engineerdistributed-inference

posted 7mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid

cudagpu-kernelsinference-enginescpp-langrocm-hip

posted 7w ago · verified today

Inferactinference provider

Remote · remote · unknown

inferenceinference-enginesvllm-enginepython-langsoftware-engineer

posted 12d ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

amdgpu-kernelsinferenceperformance-engineerrocm-hip

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

tpuinferenceinference-enginesjax-pallasperformance-engineer

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

amdgpu-kernelsinferenceinference-enginesperformance-engineer

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

vllm-engineinferenceinference-engineskv-cache-systemspython-lang

posted 12w ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

inferenceinference-enginespython-langvllm-enginesoftware-engineer

posted 3mo ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

tpuinference-enginesjax-pallasperformance-engineerxla-compiler

posted 3mo ago · verified today

Nscaleneocloud

London · UK · senior

evaluationfine-tuninggpu-genericinferenceinference-engines

posted 5mo ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer

posted 20d ago · verified today

OpenAIfrontier lab

San Francisco · hybrid · $266K–$445K/year · unknown

software-engineerinference-enginesinferencekv-cache-systemsscheduling-orchestration

posted 3w ago · verified today