AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

100 roles

sglang-engine

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid

distributed-inferenceinferencejax-pallastpuxla-compiler

posted 8mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

amdnvidiagpu-kernelscollectivescpp-lang

posted 6w ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior

inference-enginesgpu-genericinferenceperformance-engineerdistributed-inference

posted 7mo ago · verified today

RadixArkai startup

Palo Alto, CA · Radixark · $200K–$400K/year est. · mid

ml-platformgo-langkubernetes-opsobservabilitypython-lang

posted 8mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · junior

cpp-langinferenceinference-enginespython-langsglang-engine

posted 8mo ago · verified today

Inferactinference provider

Remote · remote · unknown

inferenceinference-enginesvllm-enginepython-langsoftware-engineer

posted 12d ago · verified today

Inferactinference provider

San Francisco · onsite · manager

eng-managerinferenceinference-enginesvllm-enginedistributed-inference

posted 6w ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

amdgpu-kernelsinferenceperformance-engineerrocm-hip

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

tpuinferenceinference-enginesjax-pallasperformance-engineer

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

amdgpu-kernelsinferenceinference-enginesperformance-engineer

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

vllm-engineinferenceinference-engineskv-cache-systemspython-lang

posted 12w ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

inferenceinference-enginespython-langvllm-enginesoftware-engineer

posted 3mo ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

tpuinference-enginesjax-pallasperformance-engineerxla-compiler

posted 3mo ago · verified today

Periodic Labsfrontier lab

Menlo Park, CA · onsite · unknown

collectivescudacutlass-cutefsdpgpu-generic

posted 4mo ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · senior

ml-platformsoftware-engineerinferencejax-pallaspytorch-dist

posted 19d ago · verified today

Crusoeneocloud

Denver, CO - US · onsite · staff plus

inferenceinference-enginesperformance-engineersglang-enginevllm-engine

posted 3w ago · verified today

SambaNovachip vendor

Austin, Texas, United States · Austin, TX +2 more · $210K–$280K · staff plus

inferenceml-platformobservabilityreliability-sresre

posted 9mo ago · verified today

Nebiusneocloud

Remote - United States · United States · remote · $228K–$285K · manager

eng-managerfine-tuninginferenceinference-engineskubernetes-ops

posted 4w ago · verified today

Baseteninference provider

New York · San Francisco · hybrid · $200K–$400K/year · mid

inferenceinference-enginessolutions-architectevaluationperformance-engineer

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · onsite · mid

gpu-kernelsinferenceinference-enginesdistributed-inferenceml-platform

posted 6w ago · verified today