AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

100 roles

sglang-engine

Amazon (AWS)hyperscaler

Tel Aviv-Yafo, Tel Aviv, ISR · onsite · mid

cpp-langsoftware-engineercustom-asicdistributed-inferencegpu-kernels

posted 5w ago · verified today

Amazon (AWS)hyperscaler

Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior

solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks

posted 3mo ago · verified today

Cerebraschip vendor

Sunnyvale, CA · Toronto, CAN · hybrid · staff plus

amdinference-enginespython-langsoftware-engineervllm-engine

posted 7w ago · verified today

Cerebraschip vendor

Canada · United States · senior

amdcpp-langinferenceinference-enginespython-lang

posted 9mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown

software-engineerdistributed-inferenceinferenceinference-enginesml-platform

posted 3mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $165K–$330K/year · unknown

cpp-langnetwork-fabricnvidiasoftware-engineercollectives

posted 6mo ago · verified today

Baseteninference provider

Montreal · New York +3 more · hybrid · $165K–$330K/year · unknown

solutions-architectinferenceinference-enginessglang-enginetensorrt-stack

posted 6mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $180K–$360K/year · mid

inferenceinference-enginessoftware-engineercudadistributed-inference

posted 11mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown

inferenceinference-engineskv-cache-systemsperformance-engineerquantization

posted 30mo ago · verified today

SambaNovachip vendor

Austin, Texas, United States · Austin, TX +2 more · $200K–$275K · senior

inferenceinference-enginessoftware-engineerspeculative-decodingvllm-engine

posted 3mo ago · verified today

SambaNovachip vendor

Bengaluru, India · Bengaluru, Karnataka, India · 3324K–4062K INR · unknown

python-langsolutions-architectinferencefine-tuningkubernetes-ops

posted 9w ago · verified today

Modalinference provider

New York · San Francisco · $150K–$350K/year · unknown

inferenceinference-engineskv-cache-systemsquantizationresearch-engineer

posted 10w ago · verified today

Modalinference provider

New York · San Francisco · $180K–$250K/year · unknown

solutions-architectsglang-enginevllm-engineperformance-engineerinference

posted 6mo ago · verified today

Fireworks AIinference provider

New York · San Mateo +1 more · hybrid · $200K–$260K/year · senior

inference-enginesquantizationsglang-enginevllm-enginefine-tuning

posted 3mo ago · verified today

Fireworks AIinference provider

New York · San Mateo · hybrid · $200K–$260K/year · unknown

fine-tuningpython-langsglang-enginesolutions-architectvllm-engine

posted 3mo ago · verified today

Fireworks AIinference provider

London · senior

fine-tuninginferenceinference-enginespython-langsglang-engine

posted 8w ago · verified today

Fireworks AIinference provider

Singapore · 200K–350K SGD/year · senior

fine-tuninginferenceinference-enginespython-langsglang-engine

posted 7w ago · verified today

Fireworks AIinference provider

San Mateo · hybrid · $175K–$220K/year · unknown

ml-platformpython-langsoftware-engineerinferenceinference-engines

posted 10mo ago · verified today

xAIfrontier lab

Palo Alto, CA · $180K–$440K/year est. · unknown

cpp-langgpu-genericgpu-kernelsinferenceinference-engines

posted 23mo ago · verified today

xAIfrontier lab

Palo Alto, CA · $180K–$440K/year est. · unknown

cpp-langcudadistributed-inferenceinferenceinference-engines

posted 10w ago · verified today

xAIfrontier lab

London · London, England, United Kingdom · £107K–£262K/year est. · unknown

cpp-langinferenceinference-enginesml-platformobservability

posted 4mo ago · verified today

OpenAIfrontier lab

San Francisco · Seattle · hybrid · $293K–$385K/year · unknown

evaluationgpu-kernelsinferencesoftware-engineercollectives

posted 4mo ago · verified today

OpenAIfrontier lab

San Francisco · Seattle · hybrid · $293K–$445K/year · senior

inference-enginesml-platformcpp-langgo-langinference

posted 8w ago · verified today

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands +7 more · senior

go-langgpu-generickubernetes-opsscheduling-orchestrationsoftware-engineer

posted 7w ago · verified today

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands +14 more · remote · senior

go-langinferenceinference-engineskv-cache-systemspython-lang

posted 11w ago · verified today

Nebiusneocloud

Remote - Europe · remote · senior

inferencepython-langsolutions-architectfine-tuningvllm-engine

posted 6mo ago · verified today

Nebiusneocloud

Amsterdam, Netherlands · Berlin, Germany +6 more · remote · senior

gpu-genericinferenceinference-enginesml-platformperformance-engineer

posted 6mo ago · verified today