AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

380 roles

inference

Fireworks AIinference provider

New York · San Mateo · hybrid · $200K–$230K/year · unknown

scheduling-orchestrationcpp-langnetwork-fabricpython-langstorage-checkpointing

posted 8d ago · verified today

Nebiusneocloud

Abu Dhabi, UAE · Dubai +1 more · senior

solutions-architectansiblecudagpu-generickubernetes-ops

posted 9d ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · senior

distributed-inferenceinferenceinference-enginespython-langpytorch-dist

posted 9d ago · verified today

Harveyai startup

San Francisco · hybrid · $231K–$340K/year · staff plus

inferenceinference-enginessoftware-engineerml-platformobservability

posted 10d ago · verified today

Lila Sciencesai startup

Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus

inferenceinference-engineskubernetes-opsml-platformnvidia

posted 7w ago · verified today

Xaira Therapeuticsai startup

Seattle, WA · Seattle, Washington, United States · $185K–$226K/year est. · senior

evaluationml-platformpython-langscheduling-orchestrationsoftware-engineer

posted 3mo ago · verified today

Abridgeai startup

SF Office · hybrid · $221K–$260K/year · senior

inferenceinference-engineskubernetes-opssoftware-engineercuda

posted 12mo ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · manager

inferenceinference-enginesml-platformscheduling-orchestrationcluster-datacenter

posted 7w ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · $235K–$353K/year · staff plus

gpu-genericreliability-sresrecluster-datacenterkubernetes-ops

posted 7w ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · unknown

inferenceinference-enginessoftware-engineerkubernetes-opspython-lang

posted 7w ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · $168K–$252K/year · senior

reliability-sresregpu-genericnvidiaamd

posted 7w ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · unknown

cudagpu-kernelstriton-langgpu-genericnvidia

posted 7w ago · verified today

Hedraai startup

San Francisco · $175K–$275K/year · senior

kubernetes-opsml-platformpython-langsoftware-engineerinference

posted 3mo ago · verified today

Sesameai startup

Bellevue · New York +1 more · onsite · $175K–$280K/year · senior

inferenceinference-enginesperformance-engineersglang-enginevllm-engine

posted 18mo ago · verified today

Cartesiaai startup

Bangalore · remote · 7000K–9000K INR/year · mid

go-langinferenceinference-enginesml-platformpython-lang

posted 9w ago · verified today

Cartesiaai startup

*HQ - San Francisco, CA · onsite · $180K–$250K/year · unknown

inferenceinference-enginessoftware-engineerml-platformdistributed-inference

posted 21mo ago · verified today

Cartesiaai startup

*HQ - San Francisco, CA · onsite · $180K–$250K/year · unknown

software-engineergo-langinferenceinference-enginespython-lang

posted 9w ago · verified today

ElevenLabsai startup

Bulgaria · Poland +2 more · remote · unknown

cudagpu-genericgpu-kernelsinferenceinference-engines

posted 19d ago · verified today

Harveyai startup

San Francisco · $236K–$290K/year · staff plus

inferenceinference-enginesml-platformobservabilitysoftware-engineer

posted 8w ago · verified today

Harveyai startup

San Francisco · hybrid · $260K–$340K/year · manager

eng-managerml-platformobservabilityreliability-sreinference

posted 8w ago · verified today

Cursorai startup

New York · San Francisco · onsite · unknown

inferencesoftware-engineerinference-enginesml-platformreliability-sre

posted 5mo ago · verified today

Cursorai startup

New York · San Francisco · onsite · unknown

research-engineerinferenceinference-enginesreinforcement-learningsoftware-engineer

posted 7mo ago · verified today

Figureai startup

HQ · San Jose, CA · $180K–$275K/year est. · staff plus

cpp-langpython-langquantizationgpu-genericperformance-engineer

posted 11w ago · verified today

Tenstorrentchip vendor

Austin, Texas, United States · Santa Clara +2 more · $100K–$500K/year est. · staff plus

inferencekubernetes-opsinference-enginesobservabilitysglang-engine

posted 5w ago · verified today

Tenstorrentchip vendor

Tokyo · Tokyo, Japan · unknown

performance-engineerpython-langfine-tuninginferencepre-training

posted 9mo ago · verified today