AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

39 roles

amd

Inferactinference provider

San Francisco · onsite · junior

amdcudagpu-kernelsinferenceinference-engines

posted 1d ago · verified today

DigitalOceanneocloud

Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus

amdcudainferencekubernetes-opsnvidia

posted 4mo ago · verified today

DigitalOceanneocloud

Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus

amdcudadistributed-inferencego-langinference

posted 4mo ago · verified today

DigitalOceanneocloud

Boston · Seattle Metro · $191K–$239K/year est. · staff plus

cudagpu-kernelsinferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

Denver · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelscudainferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

San Francisco · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Austin · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Seattle · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelsinferenceinference-enginescudadistributed-inference

posted 11w ago · verified today

DigitalOceanneocloud

Bangalore Metro · Bengaluru · senior

cudakubernetes-opsnvidiaamdrocm-hip

posted 11w ago · verified today

Lila Sciencesai startup

Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus

inferenceinference-engineskubernetes-opsml-platformnvidia

posted 7w ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · $168K–$252K/year · senior

reliability-sresregpu-genericnvidiaamd

posted 7w ago · verified today

Figureai startup

HQ · San Jose, CA · $200K–$400K/year est. · senior

performance-engineercollectivescpp-langcudafsdp

posted 4w ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid

gpu-genericinferenceinference-enginesreliability-sreamd

posted 4mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

amdnvidiagpu-kernelscollectivescpp-lang

posted 6w ago · verified today

falinference provider

San Francisco · $180K–$250K/year · mid

ansiblecudanvidiaobservabilitypython-lang

posted 6mo ago · verified today

falinference provider

Remote - Global · unknown

python-langgpu-genericcudanvidiareliability-sre

posted 6mo ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

amdgpu-kernelsinferenceperformance-engineerrocm-hip

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

amdgpu-kernelsinferenceinference-enginesperformance-engineer

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · $200K–$400K/year · unknown

gpu-kernelscudaperformance-engineercpp-langinference

posted 12w ago · verified today

FriendliAIinference provider

Seoul · onsite · mid

amdcpp-langcudacutlass-cutegpu-kernels

posted 4mo ago · verified today

FriendliAIinference provider

San Francisco · hybrid · mid

amdcpp-langcudagpu-genericgpu-kernels

posted 6mo ago · verified today

TensorWaveneocloud

Las Vegas, Nevada · onsite · mid

amdgpu-generickubernetes-opsreliability-srescheduling-orchestration

posted 3w ago · verified today

TensorWaveneocloud

Las Vegas, Nevada · Remote · remote · senior

ansiblecudakubernetes-opsnetwork-fabricpython-lang

posted 7w ago · verified today

TensorWaveneocloud

Remote · remote · senior

amdterraform-iacobservabilityreliability-sresre

posted 3mo ago · verified today

Nscaleneocloud

Houston · San Francisco +1 more · $120K–$170K · senior

nvidiasrecluster-datacenterinfiniband-opsnccl-lib

posted 6mo ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · onsite · senior

gpu-kernelsperformance-engineertrainiumcudanvidia

posted 15d ago · verified today

CoreWeaveneocloud

Bellevue, WA · San Francisco, CA +1 more · $198K–$264K/year est. · staff plus

amdcluster-datacenterdeepspeed-libdistributed-inferencefsdp

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Toronto, Ontario, CAN · onsite · senior

gpu-kernelsperformance-engineertrainiumcudatriton-lang

posted 14d ago · verified today

Cerebraschip vendor

Sunnyvale, CA · Toronto, CAN · hybrid · staff plus

amdinference-enginespython-langsoftware-engineervllm-engine

posted 7w ago · verified today