AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1295 open roles · 85 companies · last verified today
410 roles
Baseteninference provider
Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown
inferenceinference-engineskv-cache-systemsperformance-engineerquantization
posted 30mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · $245K–$325K · manager
eng-managerinferenceinference-enginesml-platformgo-lang
posted 3mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · $245K–$325K · staff plus
inferenceinference-enginespython-langsoftware-engineerkubernetes-ops
posted 3mo ago · verified today
SambaNovachip vendor
Austin, Texas, United States · Austin, TX +2 more · $200K–$275K · senior
inferenceinference-enginessoftware-engineerspeculative-decodingvllm-engine
posted 3mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · staff plus
custom-asicgpu-kernelsnetwork-fabricrdma-verbsroce-net
posted 5mo ago · verified today
Scale AIai startup
New York, NY · San Francisco, CA · $290K–$363K · manager
cudaeng-managerpost-trainingtraining-frameworksflash-attention
posted 11mo ago · verified today
Scale AIai startup
New York, NY · San Francisco, CA · $180K–$225K · mid
software-engineertraining-data-inframl-platformscheduling-orchestrationevaluation
posted 13mo ago · verified today
Scale AIai startup
New York, NY · San Francisco, CA +2 more · unknown
inferenceinference-enginessoftware-engineerkubernetes-opsnetwork-fabric
posted 31mo ago · verified today
Scale AIai startup
New York, NY · San Francisco, CA +1 more · $190K–$237K · unknown
cudaflash-attentioninference-enginesml-platformresearch-engineer
posted 18mo ago · verified today
Scale AIai startup
London, UK · unknown
inferenceinference-enginessoftware-engineercpp-langgo-lang
posted 7w ago · verified today
Anyscaleai startup
San Francisco · hybrid · $200K–$240K/year · mid
reliability-sresrego-langkubernetes-opspython-lang
posted 3w ago · verified today
Anyscaleai startup
Bengaluru, Karnataka · unknown
go-langkubernetes-opspython-langsoftware-engineerml-platform
posted 6w ago · verified today
Anyscaleai startup
San Francisco · hybrid · $200K–$240K/year · mid
kubernetes-opsray-distributedsoftware-engineergo-langml-platform
posted 4w ago · verified today
Anyscaleai startup
Palo Alto · San Francisco · hybrid · $170K–$245K/year · unknown
distributed-inferenceinferenceinference-enginessoftware-engineervllm-engine
posted 3mo ago · verified today
Anyscaleai startup
Bengaluru, Karnataka · unknown
ray-distributedsoftware-engineercpp-langscheduling-orchestrationml-platform
posted 7w ago · verified today
Modalinference provider
Stockholm · mid
inferencesolutions-architectdistributed-inferencefine-tuninggpu-generic
posted 6mo ago · verified today
Modalinference provider
London · New York +2 more · onsite · $200K–$280K/year · senior
solutions-architectkubernetes-opsml-platformpython-langterraform-iac
posted 6mo ago · verified today
Modalinference provider
New York · $275K–$330K/year · manager
eng-managercpp-langml-platformrust-langlinux-kernel
posted 7mo ago · verified today
Modalinference provider
Stockholm · $140K–$200K/year · unknown
linux-kernelsoftware-engineerml-platformscheduling-orchestrationrust-lang
posted 8mo ago · verified today
Modalinference provider
London · New York +2 more · $150K–$220K/year · unknown
solutions-architectgpu-genericsoftware-engineerml-platforminference
posted 11mo ago · verified today
Modalinference provider
New York · San Francisco · $200K–$350K/year · senior
performance-engineercudagpu-kernelsinference-enginesnvidia
posted 4mo ago · verified today
Modalinference provider
New York · San Francisco · $180K–$250K/year · unknown
solutions-architectsglang-enginevllm-engineperformance-engineerinference
posted 6mo ago · verified today
Modalinference provider
New York · San Francisco · onsite · $220K–$300K/year · senior
ml-platformsoftware-engineerlinux-kernelscheduling-orchestrationreliability-sre
posted 23mo ago · verified today
Fireworks AIinference provider
San Mateo · hybrid · $175K–$220K/year · unknown
ml-platformpython-langsoftware-engineerinferenceinference-engines
posted 10mo ago · verified today
Fireworks AIinference provider
San Mateo · hybrid · $210K–$320K/year · mid
evaluationpost-trainingfine-tuningml-platforminference
posted 10mo ago · verified today
Fireworks AIinference provider
New York · San Mateo · hybrid · $175K–$220K/year · senior
software-engineerkubernetes-opsml-platformpython-langobservability
posted 15mo ago · verified today
Fireworks AIinference provider
New York · San Mateo · hybrid · $210K–$320K/year · unknown
software-engineertraining-frameworksfsdpkubernetes-opsml-platform
posted 6w ago · verified today
xAIfrontier lab
Palo Alto, CA · $180K–$440K/year est. · unknown
cpp-langkubernetes-opsrust-langsoftware-engineercluster-datacenter
posted 8w ago · verified today
xAIfrontier lab
Palo Alto, CA · $180K–$440K/year est. · mid
gpu-genericml-platformpython-langsoftware-engineercpp-lang
posted 8w ago · verified today
xAIfrontier lab
Palo Alto, CA · $180K–$440K/year est. · unknown
gpu-genericpost-trainingpre-trainingpython-langresearch-engineer
posted 5mo ago · verified today