AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1296 open roles · 85 companies · last verified today
13 roles
DigitalOceanneocloud
Bangalore Metro · Bengaluru · senior
cudakubernetes-opsnvidiaamdrocm-hip
posted 11w ago · verified today
Inceptionfrontier lab
San Mateo, United States · onsite · senior
inferenceinference-enginespython-langsoftware-engineergpu-generic
posted 6mo ago · verified today
Inceptionfrontier lab
San Mateo, United States · onsite · unknown
cudagpu-genericinferenceinference-engineskubernetes-ops
posted 6mo ago · verified today
Lila Sciencesai startup
Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus
inferenceinference-engineskubernetes-opsml-platformnvidia
posted 7w ago · verified today
Abridgeai startup
SF Office · hybrid · $221K–$260K/year · senior
inferenceinference-engineskubernetes-opssoftware-engineercuda
posted 12mo ago · verified today
Amazon (AWS)hyperscaler
New York, New York, USA · Seattle, Washington, USA · onsite · senior
software-engineerinferenceinference-enginesdistributed-inferencetensorrt-stack
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior
solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks
posted 3mo ago · verified today
Cerebraschip vendor
Sunnyvale, CA · Toronto, CAN · hybrid · staff plus
amdinference-enginespython-langsoftware-engineervllm-engine
posted 7w ago · verified today
Cerebraschip vendor
Canada · United States · senior
amdcpp-langinferenceinference-enginespython-lang
posted 9mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $293K–$385K/year · unknown
evaluationgpu-kernelsinferencesoftware-engineercollectives
posted 4mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $293K–$445K/year · senior
inference-enginesml-platformcpp-langgo-langinference
posted 8w ago · verified today
Nebiusneocloud
Amsterdam · Amsterdam, Netherlands +7 more · senior
go-langgpu-generickubernetes-opsscheduling-orchestrationsoftware-engineer
posted 7w ago · verified today
Nebiusneocloud
Palo Alto, California, United States · San Francisco Bay Area · $195K–$262K · senior
inferenceinference-enginespython-langquantizationvllm-engine
posted 7w ago · verified today