AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

83 roles

quantization

Nscaleneocloud

Houston · New York +2 more · $220K–$293K · staff plus

inferenceinference-engineskv-cache-systemspost-trainingpython-lang

posted 1d ago · verified today

SambaNovachip vendor

Tokyo, Japan · Tokyo Prefecture, Japan · 105K–130K JPY · senior

inferencepython-langsolutions-architectperformance-engineerevaluation

posted 2d ago · verified today

DoorDashenterprise

San Francisco · San Francisco, CA +2 more · senior

fine-tuninggpu-genericinferenceinference-enginesml-platform

posted 10w ago · verified today

Inferactinference provider

San Francisco · onsite · junior

amdcudagpu-kernelsinferenceinference-engines

posted 1d ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · senior

trainiuminferenceinference-enginesmoe-systemsperformance-engineer

posted 7d ago · verified today

Reactorinference provider

San Francisco · onsite · unknown

inferenceinference-enginescudagpu-kernelsperformance-engineer

added 7d ago · verified today

DigitalOceanneocloud

Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus

amdcudainferencekubernetes-opsnvidia

posted 4mo ago · verified today

DigitalOceanneocloud

Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus

amdcudadistributed-inferencego-langinference

posted 4mo ago · verified today

DigitalOceanneocloud

Boston · Seattle Metro · $191K–$239K/year est. · staff plus

cudagpu-kernelsinferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

Denver · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelscudainferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

San Francisco · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Austin · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Seattle · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelsinferenceinference-enginescudadistributed-inference

posted 11w ago · verified today

DigitalOceanneocloud

DigitalOcean Seattle Office · Seattle · junior

go-langgpu-genericinferencepython-langsoftware-engineer

posted 3w ago · verified today

DigitalOceanneocloud

Bay Area Metro · New York · $150K–$186K/year est. · senior

solutions-architectcudagpu-genericinferencekubernetes-ops

posted 5mo ago · verified today

DigitalOceanneocloud

Bay Area Metro · San Francisco · $150K–$215K/year est. · senior

cudakubernetes-opssolutions-architectfine-tuninggpu-generic

posted 5mo ago · verified today

DigitalOceanneocloud

Bangalore Metro · Bengaluru · senior

cudakubernetes-opsnvidiaamdrocm-hip

posted 11w ago · verified today

DeepInfrainference provider

Palo Alto, United States · onsite · unknown

inferenceinference-enginesnvidiapython-langsolutions-architect

posted 4w ago · verified today

Interfazeinference provider

San Francisco, CA · senior

fine-tuninginferenceinference-enginespython-langquantization

added 7d ago · verified today

Inceptionfrontier lab

San Mateo, United States · onsite · unknown

cudagpu-genericinferenceinference-engineskubernetes-ops

posted 6mo ago · verified today

Black Forest Labsfrontier lab

San Francisco (United States) · onsite · unknown

cudagpu-genericinferenceinference-enginesperformance-engineer

posted 24mo ago · verified today

Lila Sciencesai startup

Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus

inferenceinference-engineskubernetes-opsml-platformnvidia

posted 7w ago · verified today

ElevenLabsai startup

Bulgaria · Poland +2 more · remote · unknown

cudagpu-genericgpu-kernelsinferenceinference-engines

posted 19d ago · verified today

Figureai startup

HQ · San Jose, CA · $180K–$275K/year est. · staff plus

cpp-langpython-langquantizationgpu-genericperformance-engineer

posted 11w ago · verified today