AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

105 roles

tensorrt-stack

ElevenLabsai startup

Bulgaria · Poland +2 more · remote · unknown

cudagpu-genericgpu-kernelsinferenceinference-engines

posted 19d ago · verified today

Figureai startup

HQ · San Jose, CA · $180K–$275K/year est. · staff plus

cpp-langpython-langquantizationgpu-genericperformance-engineer

posted 11w ago · verified today

1Xai startup

San Carlos, CA · onsite · $250K–$350K/year · unknown

gpu-genericpre-trainingpython-langresearch-engineertraining-data-infra

posted 3mo ago · verified today

Lightning AIai startup

London, UK · New York, New York +6 more · remote · $165K–$310K · senior

python-langresearch-engineertraining-frameworksevaluationfine-tuning

posted 5w ago · verified today

Lightning AIai startup

London, England, United Kingdom · New York, New York +5 more · $180K–$250K · unknown

python-langsolutions-architectdistributed-inferenceinferenceinference-engines

posted 3mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

gpu-genericpost-traininggpu-kernelsmodel-parallelismtraining-frameworks

posted 7mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior

inference-enginesgpu-genericinferenceperformance-engineerdistributed-inference

posted 7mo ago · verified today

Inferactinference provider

Remote · remote · unknown

inferenceinference-enginesvllm-enginepython-langsoftware-engineer

posted 12d ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

amdgpu-kernelsinferenceperformance-engineerrocm-hip

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

tpuinferenceinference-enginesjax-pallasperformance-engineer

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

amdgpu-kernelsinferenceinference-enginesperformance-engineer

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

vllm-engineinferenceinference-engineskv-cache-systemspython-lang

posted 12w ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

inferenceinference-enginespython-langvllm-enginesoftware-engineer

posted 3mo ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

tpuinference-enginesjax-pallasperformance-engineerxla-compiler

posted 3mo ago · verified today

Periodic Labsfrontier lab

Menlo Park, CA · onsite · unknown

collectivescudacutlass-cutefsdpgpu-generic

posted 4mo ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · senior

ml-platformsoftware-engineerinferencejax-pallaspytorch-dist

posted 19d ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer

posted 20d ago · verified today

Nebiusneocloud

Remote - United States · United States · remote · $228K–$285K · manager

eng-managerfine-tuninginferenceinference-engineskubernetes-ops

posted 4w ago · verified today

Baseteninference provider

New York · San Francisco · hybrid · $200K–$400K/year · mid

inferenceinference-enginessolutions-architectevaluationperformance-engineer

posted 4w ago · verified today

Amazon (AWS)hyperscaler

San Francisco, California, USA · onsite · senior

cudagpu-kernelsnvidiapython-langresearch-engineer

posted 3mo ago · verified today

Amazon (AWS)hyperscaler

New York, New York, USA · Seattle, Washington, USA · onsite · senior

software-engineerinferenceinference-enginesdistributed-inferencetensorrt-stack

posted 6w ago · verified today

Amazon (AWS)hyperscaler

Tel Aviv-Yafo, Tel Aviv, ISR · onsite · mid

cpp-langsoftware-engineercustom-asicdistributed-inferencegpu-kernels

posted 5w ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

cudacutlass-cutedistributed-inferenceevaluationflash-attention

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior

solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks

posted 3mo ago · verified today