AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

57 roles

kv-cache-systems

Nebiusneocloud

Remote - United States · United States · remote · $228K–$285K · manager

eng-managerfine-tuninginferenceinference-engineskubernetes-ops

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Tel Aviv-Yafo, Tel Aviv, ISR · onsite · mid

cpp-langsoftware-engineercustom-asicdistributed-inferencegpu-kernels

posted 5w ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

cudacutlass-cutedistributed-inferenceevaluationflash-attention

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior

solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks

posted 3mo ago · verified today

Perplexityai startup

New York City · Palo Alto +1 more · $220K–$485K/year · mid

cudagpu-kernelsinferenceinference-enginescutlass-cute

posted 5mo ago · verified today

Cerebraschip vendor

Sunnyvale, CA · Toronto, CAN · hybrid · staff plus

amdinference-enginespython-langsoftware-engineervllm-engine

posted 7w ago · verified today

Cerebraschip vendor

Canada · United States · senior

amdcpp-langinferenceinference-enginespython-lang

posted 9mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $165K–$330K/year · unknown

cpp-langdistributed-inferenceinfiniband-opsnetwork-fabricroce-net

posted 6mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $165K–$330K/year · senior

performance-engineerevaluationinferenceinference-enginesnvidia

posted 8mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown

cpp-langcudainferenceinference-engineskv-cache-systems

posted 30mo ago · verified today

Modalinference provider

New York · San Francisco · $150K–$350K/year · unknown

inferenceinference-engineskv-cache-systemsquantizationresearch-engineer

posted 10w ago · verified today

xAIfrontier lab

Palo Alto, CA · $180K–$440K/year est. · unknown

cpp-langgpu-genericgpu-kernelsinferenceinference-engines

posted 23mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $92K–$135K/year est. · junior

inferencepython-langsoftware-engineergo-langinference-engines

posted 10mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $152K–$204K/year est. · senior

cudadistributed-inferencego-langgpu-genericgpu-kernels

posted 11mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $139K–$204K/year est. · senior

software-engineerinferenceinference-engineskubernetes-opsgpu-generic

posted 7mo ago · verified today

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands +14 more · remote · senior

go-langinferenceinference-engineskv-cache-systemspython-lang

posted 11w ago · verified today

Nebiusneocloud

Amsterdam, Netherlands · Berlin, Germany +6 more · remote · senior

gpu-genericinferenceinference-enginesml-platformperformance-engineer

posted 6mo ago · verified today

Nebiusneocloud

Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior

inferenceinference-enginespython-langquantizationvllm-engine

posted 7w ago · verified today

Nebiusneocloud

Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior

research-engineercudainferencepython-langtriton-lang

posted 7w ago · verified today

Nebiusneocloud

United States · $208K–$261K · staff plus

solutions-architectfine-tuninginferenceinference-engineskv-cache-systems

posted 11w ago · verified today

Nebiusneocloud

Remote - United States · remote · $180K–$224K · senior

python-langsglang-enginesolutions-architecttensorrt-stackvllm-engine

posted 4mo ago · verified today

Together AIneocloud

San Francisco · $200K–$290K/year est. · senior

inferenceinference-enginesnvidiasoftware-engineerdistributed-inference

posted 13mo ago · verified today