AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

1296 roles

DigitalOceanneocloud

DigitalOcean Seattle Office · Seattle · $234K–$292K/year est. · manager

cluster-datacenterdatacenter-engineereng-managergpu-generic

posted 12w ago · verified today

DigitalOceanneocloud

All Other Locations · Seattle · $128K–$161K/year est. · senior

datacenter-engineercluster-datacentergpu-generic

posted 5mo ago · verified today

DigitalOceanneocloud

DigitalOcean Seattle Office · Seattle · $106K–$132K/year est. · senior

cluster-datacentergpu-genericinferencedistributed-inferenceml-platform

posted 19d ago · verified today

DigitalOceanneocloud

All Other Locations · Seattle · $184K–$231K/year est. · staff plus

datacenter-engineercluster-datacentergpu-generic

posted 3mo ago · verified today

DigitalOceanneocloud

All Other Locations · San Francisco · $184K–$231K/year est. · staff plus

datacenter-engineercluster-datacentergpu-generic

posted 3mo ago · verified today

DigitalOceanneocloud

Atlanta · *United States · manager

cluster-datacenterdatacenter-engineergpu-genericreliability-sre

posted 20d ago · verified today

DigitalOceanneocloud

Seattle Metro · Wenatchee · manager

cluster-datacenterdatacenter-engineergpu-genericreliability-sre

posted 20d ago · verified today

DigitalOceanneocloud

Austin Metro · Seattle · $107K–$134K/year est. · mid

reliability-sresregpu-generickubernetes-opsgo-lang

posted 7mo ago · verified today

Scalewayneocloud

Paris · hybrid · manager

eng-managercluster-datacenterkubernetes-opsnvidiareliability-sre

posted 3w ago · verified today

Scalewayneocloud

Paris · hybrid · senior

datacenter-engineercluster-datacenter

posted 11mo ago · verified today

DeepInfrainference provider

Palo Alto, United States · onsite · unknown

inferenceinference-enginesnvidiapython-langsolutions-architect

posted 4w ago · verified today

DeepInfrainference provider

Sofia, Bulgaria · onsite · junior

inferenceinference-enginespython-langsoftware-engineercpp-lang

posted 11w ago · verified today

DeepInfrainference provider

Bulgaria - Remote · remote · mid

cpp-langcudagpu-genericinferenceinference-engines

posted 5mo ago · verified today

DeepInfrainference provider

Palo Alto, United States · onsite · mid

cudainferencepython-langsoftware-engineercpp-lang

posted 5mo ago · verified today

DeepInfrainference provider

Bulgaria - Remote · remote · junior

cpp-langcudainferencepython-langsoftware-engineer

posted 7mo ago · verified today

DeepInfrainference provider

Palo Alto, United States · onsite · $140K–$150K/year est. · junior

inference-enginessoftware-engineercpp-langcudaml-platform

posted 8mo ago · verified today

DeepInfrainference provider

Bulgaria - Remote · remote · junior

cpp-langcudainferenceinference-enginespython-lang

posted 11mo ago · verified today

DeepInfrainference provider

Palo Alto, United States · onsite · junior

cpp-langcudanccl-libpython-langsoftware-engineer

posted 11mo ago · verified today

Fish Audio (39 AI)ai startup

Location not specified · senior

cluster-datacentergpu-genericreliability-sresrekubernetes-ops

added 7d ago · verified today

Fish Audio (39 AI)ai startup

Location not specified · unknown

cpp-langinference-enginespython-langinferencemegatron-lm

added 7d ago · verified today

Interfazeinference provider

San Francisco, CA · senior

fine-tuninginferenceinference-enginespython-langquantization

added 7d ago · verified today

Waferinference provider

San Francisco · onsite · $200K–$300K/year · unknown

gpu-kernelsinferenceinference-enginescluster-datacenterperformance-engineer

posted 8w ago · verified today

Runwareinference provider

France · Germany +3 more · remote · senior

software-engineergo-langscheduling-orchestrationinferenceml-platform

posted 20d ago · verified today

Runwareinference provider

Remote · remote · senior

reliability-sresrego-langkubernetes-opsobservability

posted 8w ago · verified today

Runwareinference provider

United Kingdom · remote · senior

gpu-genericinferencenvidiareliability-sresre

posted 4mo ago · verified today

Relaceinference provider

San Francisco · onsite · mid

cudagpu-kernelsperformance-engineersoftware-engineercpp-lang

posted 10mo ago · verified today

Relaceinference provider

San Francisco · onsite · mid

software-engineercluster-datacenterscheduling-orchestrationinferenceml-platform

posted 10mo ago · verified today

Exaai startup

Singapore · onsite · 90K–300K SGD/year · unknown

cluster-datacenterkubernetes-opsscheduling-orchestrationsoftware-engineerdistributed-inference

posted 6mo ago · verified today