AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

Novita AIinference provider

San Mateo · onsite · unknown

kubernetes-opspython-langsolutions-architectinferenceinference-engines

posted 11w ago · verified today

Crusoeneocloud

San Francisco, CA - US · onsite · staff plus

inferenceinference-enginesperformance-engineersglang-enginevllm-engine

posted 8d ago · verified today

Fireworks AIinference provider

New York · San Mateo · hybrid · $200K–$230K/year · unknown

scheduling-orchestrationcpp-langnetwork-fabricpython-langstorage-checkpointing

posted 8d ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · senior

distributed-inferenceinferenceinference-enginespython-langpytorch-dist

posted 9d ago · verified today

Harveyai startup

San Francisco · hybrid · $231K–$340K/year · staff plus

inferenceinference-enginessoftware-engineerml-platformobservability

posted 10d ago · verified today

Lila Sciencesai startup

Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus

inferenceinference-engineskubernetes-opsml-platformnvidia

posted 7w ago · verified today

Lila Sciencesai startup

San Francisco, CA · San Francisco, CA USA · $268K–$384K · staff plus

ml-platformgpu-genericsoftware-engineertraining-frameworksinference-engines

posted 4mo ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · unknown

inferenceinference-enginessoftware-engineerkubernetes-opspython-lang

posted 7w ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · $168K–$252K/year · senior

reliability-sresregpu-genericnvidiaamd

posted 7w ago · verified today

Cartesiaai startup

*HQ - San Francisco, CA · onsite · $180K–$250K/year · unknown

inferenceinference-enginessoftware-engineerml-platformdistributed-inference

posted 21mo ago · verified today

Tenstorrentchip vendor

Austin, Texas, United States · Santa Clara +2 more · $100K–$500K/year est. · staff plus

inferencekubernetes-opsinference-enginesobservabilitysglang-engine

posted 5w ago · verified today

Tenstorrentchip vendor

Austin, Texas, United States · Santa Clara +2 more · $100K–$500K/year est. · unknown

network-fabricsoftware-engineercollectivescpp-langdistributed-inference

posted 18mo ago · verified today

Tenstorrentchip vendor

Austin · Austin, Texas, United States +4 more · $100K–$500K/year est. · unknown

software-engineercpp-langmpicollectivescluster-datacenter

posted 17mo ago · verified today

Generalistai startup

San Francisco Bay Area (San Mateo) or Boston (Somerville) · onsite · $260K–$350K/year · unknown

gpu-genericnvidiasoftware-engineerinferencekubernetes-ops

posted 7mo ago · verified today

1Xai startup

San Carlos, CA · onsite · $250K–$350K/year · unknown

gpu-genericpre-trainingpython-langresearch-engineertraining-data-infra

posted 3mo ago · verified today

Lightning AIai startup

London, England, United Kingdom · New York, New York +5 more · $180K–$250K · unknown

python-langsolutions-architectdistributed-inferenceinferenceinference-engines

posted 3mo ago · verified today

Lightning AIai startup

New York, New York, United States · San Francisco, California +3 more · $115K–$140K · unknown

cudagpu-generickubernetes-opsnccl-libobservability

posted 3mo ago · verified today

Lightning AIai startup

London, England, United Kingdom · London, UK · £75K–£95K · unknown

cudagpu-generickubernetes-opsml-platformnccl-lib

posted 3mo ago · verified today

Lightning AIai startup

Philippines · Remote +1 more · unknown

reliability-sresrecluster-datacentergpu-generickubernetes-ops

posted 4mo ago · verified today

Etchedchip vendor

San Jose · onsite · $200K–$300K/year · manager

observabilitydistributed-inferenceperformance-engineercpp-langeng-manager

posted 6mo ago · verified today

Etchedchip vendor

San Jose · onsite · $175K–$275K/year · unknown

custom-asicperformance-engineerinferencedistributed-inferencegpu-generic

posted 4mo ago · verified today

Etchedchip vendor

San Jose · onsite · junior

cpp-langinferencepython-langdistributed-inferencecollectives

posted 9mo ago · verified today

Coherefrontier lab

Montreal · New York +2 more · hybrid · staff plus

inferenceinference-enginessoftware-engineergpu-generickubernetes-ops

posted 8mo ago · verified today

Coherefrontier lab

London · Montreal +4 more · remote · senior

training-frameworksdistributed-inferencepre-trainingcudakubernetes-ops

posted 9mo ago · verified today

Coherefrontier lab

Montreal · New York +2 more · remote · unknown

cpp-langgpu-genericinferenceinference-enginesperformance-engineer

posted 10mo ago · verified today