AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1297 open roles · 85 companies · last verified today
100 roles
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid
distributed-inferenceinferencejax-pallastpuxla-compiler
posted 8mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown
amdnvidiagpu-kernelscollectivescpp-lang
posted 6w ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior
inference-enginesgpu-genericinferenceperformance-engineerdistributed-inference
posted 7mo ago · verified today
RadixArkai startup
Palo Alto, CA · Radixark · $200K–$400K/year est. · mid
ml-platformgo-langkubernetes-opsobservabilitypython-lang
posted 8mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · junior
cpp-langinferenceinference-enginespython-langsglang-engine
posted 8mo ago · verified today
Inferactinference provider
Remote · remote · unknown
inferenceinference-enginesvllm-enginepython-langsoftware-engineer
posted 12d ago · verified today
Inferactinference provider
San Francisco · onsite · manager
eng-managerinferenceinference-enginesvllm-enginedistributed-inference
posted 6w ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
amdgpu-kernelsinferenceperformance-engineerrocm-hip
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
tpuinferenceinference-enginesjax-pallasperformance-engineer
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
amdgpu-kernelsinferenceinference-enginesperformance-engineer
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
vllm-engineinferenceinference-engineskv-cache-systemspython-lang
posted 12w ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
inferenceinference-enginespython-langvllm-enginesoftware-engineer
posted 3mo ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
tpuinference-enginesjax-pallasperformance-engineerxla-compiler
posted 3mo ago · verified today
Prime Intellectneocloud
San Francisco · onsite · unknown
kubernetes-opsml-platformfine-tuninggpu-genericnvidia
posted 10w ago · verified today
Prime Intellectneocloud
San Francisco · unknown
post-trainingreinforcement-learningresearch-engineerevaluationml-platform
posted 10w ago · verified today
Prime Intellectneocloud
New York City, USA · hybrid · unknown
evaluationpost-trainingdistributed-inferenceray-distributedreinforcement-learning
posted 10w ago · verified today
Prime Intellectneocloud
Remote · San Francisco · unknown
inference-enginesreinforcement-learningresearch-engineersglang-enginevllm-engine
posted 10w ago · verified today
Periodic Labsfrontier lab
Menlo Park, CA · onsite · unknown
collectivescudacutlass-cutefsdpgpu-generic
posted 4mo ago · verified today
Thinking Machines Labfrontier lab
San Francisco · hybrid · unknown
reinforcement-learningresearch-engineerdistributed-inferenceinferenceinference-engines
posted 3w ago · verified today
Thinking Machines Labfrontier lab
San Francisco · hybrid · unknown
distributed-inferenceinference-enginesresearch-engineergpu-genericinference
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · mid
distributed-inferencegpu-kernelsinferenceinference-enginessoftware-engineer
posted 16d ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
ml-platformsoftware-engineerinferencejax-pallaspytorch-dist
posted 19d ago · verified today
Crusoeneocloud
Denver, CO - US · onsite · staff plus
inferenceinference-enginesperformance-engineersglang-enginevllm-engine
posted 3w ago · verified today
SambaNovachip vendor
Austin, Texas, United States · Austin, TX +2 more · $210K–$280K · staff plus
inferenceml-platformobservabilityreliability-sresre
posted 9mo ago · verified today
Nebiusneocloud
Remote - United States · United States · remote · $228K–$285K · manager
eng-managerfine-tuninginferenceinference-engineskubernetes-ops
posted 4w ago · verified today
Baseteninference provider
New York · San Francisco · hybrid · $200K–$400K/year · mid
inferenceinference-enginessolutions-architectevaluationperformance-engineer
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
trainiumdistributed-inferenceinferencesoftware-engineercpp-lang
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
trainiumcpp-langdistributed-inferenceinferenceinference-engines
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · mid
software-engineerml-platformpost-trainingreinforcement-learningtraining-frameworks
posted 3mo ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · mid
gpu-kernelsinferenceinference-enginesdistributed-inferenceml-platform
posted 6w ago · verified today