AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

Modalinference provider

San Francisco · onsite · $300K–$350K/year · manager

eng-managergpu-genericinferenceinference-enginesml-platform

posted today · verified today

Nscaleneocloud

Houston · New York +2 more · $220K–$293K · staff plus

inferenceinference-engineskv-cache-systemspost-trainingpython-lang

posted 1d ago · verified today

Amazon (AWS)hyperscaler

Austin, Texas, USA · New York, New York, USA +2 more · onsite · senior

solutions-architectkubernetes-opsmpinccl-libtraining-frameworks

posted 2d ago · verified today

DoorDashenterprise

San Francisco · San Francisco, CA +2 more · senior

fine-tuninggpu-genericinferenceinference-enginesml-platform

posted 10w ago · verified today

Inferactinference provider

San Francisco · onsite · junior

amdcudagpu-kernelsinferenceinference-engines

posted 1d ago · verified today

Anthropicfrontier lab

New York City, NY · San Francisco, CA · $350K–$850K · unknown

inferenceinference-enginesperformance-engineergpu-genericrust-lang

posted 6d ago · verified today

Boson AIfrontier lab

Toronto · onsite · CA$125K–CA$250K/year · unknown

nvidiareliability-sresreansiblecluster-datacenter

posted 9w ago · verified today

DigitalOceanneocloud

Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus

amdcudainferencekubernetes-opsnvidia

posted 4mo ago · verified today

DigitalOceanneocloud

Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus

amdcudadistributed-inferencego-langinference

posted 4mo ago · verified today

DigitalOceanneocloud

Boston · Seattle Metro · $191K–$239K/year est. · staff plus

cudagpu-kernelsinferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

Denver · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelscudainferenceinference-enginesperformance-engineer

posted 7w ago · verified today

DigitalOceanneocloud

San Francisco · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Austin · Seattle Metro · $191K–$239K/year est. · staff plus

amdcudadistributed-inferenceflash-attentiongpu-kernels

posted 7w ago · verified today

DigitalOceanneocloud

Seattle · Seattle Metro · $191K–$239K/year est. · staff plus

gpu-kernelsinferenceinference-enginescudadistributed-inference

posted 11w ago · verified today

DigitalOceanneocloud

Bangalore Metro · Bengaluru · senior

cudakubernetes-opsnvidiaamdrocm-hip

posted 11w ago · verified today

DigitalOceanneocloud

DigitalOcean Seattle Office · Seattle · $106K–$132K/year est. · senior

cluster-datacentergpu-genericinferencedistributed-inferenceml-platform

posted 19d ago · verified today

DeepInfrainference provider

Bulgaria - Remote · remote · mid

cpp-langcudagpu-genericinferenceinference-engines

posted 5mo ago · verified today

DeepInfrainference provider

Palo Alto, United States · onsite · mid

cudainferencepython-langsoftware-engineercpp-lang

posted 5mo ago · verified today

DeepInfrainference provider

Bulgaria - Remote · remote · junior

cpp-langcudainferencepython-langsoftware-engineer

posted 7mo ago · verified today

DeepInfrainference provider

Bulgaria - Remote · remote · junior

cpp-langcudainferenceinference-enginespython-lang

posted 11mo ago · verified today

Fish Audio (39 AI)ai startup

Location not specified · unknown

cpp-langinference-enginespython-langinferencemegatron-lm

added 7d ago · verified today

Interfazeinference provider

San Francisco, CA · senior

fine-tuninginferenceinference-enginespython-langquantization

added 7d ago · verified today

Waferinference provider

San Francisco · onsite · $200K–$300K/year · unknown

gpu-kernelsinferenceinference-enginescluster-datacenterperformance-engineer

posted 8w ago · verified today

Exaai startup

Singapore · onsite · 90K–300K SGD/year · unknown

cluster-datacenterkubernetes-opsscheduling-orchestrationsoftware-engineerdistributed-inference

posted 6mo ago · verified today

Exaai startup

San Francisco, California · onsite · $180K–$350K/year · unknown

software-engineergpu-generickubernetes-opsml-platformray-distributed

posted 12mo ago · verified today

Inceptionfrontier lab

San Mateo, United States · onsite · senior

inferenceinference-enginespython-langsoftware-engineergpu-generic

posted 6mo ago · verified today

Inceptionfrontier lab

San Mateo, United States · onsite · unknown

cudagpu-genericinferenceinference-engineskubernetes-ops

posted 6mo ago · verified today