AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

13 roles

triton-server

DigitalOceanneocloud

Bangalore Metro · Bengaluru · senior

cudakubernetes-opsnvidiaamdrocm-hip

posted 11w ago · verified today

Inceptionfrontier lab

San Mateo, United States · onsite · senior

inferenceinference-enginespython-langsoftware-engineergpu-generic

posted 6mo ago · verified today

Inceptionfrontier lab

San Mateo, United States · onsite · unknown

cudagpu-genericinferenceinference-engineskubernetes-ops

posted 6mo ago · verified today

Lila Sciencesai startup

Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus

inferenceinference-engineskubernetes-opsml-platformnvidia

posted 7w ago · verified today

Abridgeai startup

SF Office · hybrid · $221K–$260K/year · senior

inferenceinference-engineskubernetes-opssoftware-engineercuda

posted 12mo ago · verified today

Amazon (AWS)hyperscaler

New York, New York, USA · Seattle, Washington, USA · onsite · senior

software-engineerinferenceinference-enginesdistributed-inferencetensorrt-stack

posted 6w ago · verified today

Amazon (AWS)hyperscaler

Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior

solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks

posted 3mo ago · verified today

Cerebraschip vendor

Sunnyvale, CA · Toronto, CAN · hybrid · staff plus

amdinference-enginespython-langsoftware-engineervllm-engine

posted 7w ago · verified today

Cerebraschip vendor

Canada · United States · senior

amdcpp-langinferenceinference-enginespython-lang

posted 9mo ago · verified today

OpenAIfrontier lab

San Francisco · Seattle · hybrid · $293K–$385K/year · unknown

evaluationgpu-kernelsinferencesoftware-engineercollectives

posted 4mo ago · verified today

OpenAIfrontier lab

San Francisco · Seattle · hybrid · $293K–$445K/year · senior

inference-enginesml-platformcpp-langgo-langinference

posted 8w ago · verified today

Nebiusneocloud

Amsterdam · Amsterdam, Netherlands +7 more · senior

go-langgpu-generickubernetes-opsscheduling-orchestrationsoftware-engineer

posted 7w ago · verified today

Nebiusneocloud

Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior

inferenceinference-enginespython-langquantizationvllm-engine

posted 7w ago · verified today