AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1293 open roles · 85 companies · last verified today
212 roles
Baseteninference provider
Montreal · New York +3 more · hybrid · $165K–$330K/year · unknown
inferencekubernetes-opsml-platformpython-langsoftware-engineer
posted 18mo ago · verified today
Baseteninference provider
Montreal · New York +3 more · hybrid · $165K–$330K/year · unknown
inferencekubernetes-opsml-platformsoftware-engineerinference-engines
posted 11mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · $245K–$325K · staff plus
inferenceinference-enginespython-langsoftware-engineerkubernetes-ops
posted 3mo ago · verified today
SambaNovachip vendor
Austin, Texas, United States · Austin, TX +2 more · $200K–$275K · senior
inferenceinference-enginessoftware-engineerspeculative-decodingvllm-engine
posted 3mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · staff plus
custom-asicgpu-kernelsnetwork-fabricrdma-verbsroce-net
posted 5mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · $180K–$255K · senior
inferenceperformance-engineercpp-langdistributed-inferenceinference-engines
posted 9mo ago · verified today
Scale AIai startup
New York, NY · San Francisco, CA · $290K–$363K · manager
cudaeng-managerpost-trainingtraining-frameworksflash-attention
posted 11mo ago · verified today
Scale AIai startup
New York, NY · San Francisco, CA +1 more · $190K–$237K · unknown
cudaflash-attentioninference-enginesml-platformresearch-engineer
posted 18mo ago · verified today
Scale AIai startup
London, UK · unknown
inferenceinference-enginessoftware-engineercpp-langgo-lang
posted 7w ago · verified today
Anyscaleai startup
Palo Alto · San Francisco · hybrid · $170K–$245K/year · unknown
distributed-inferenceinferenceinference-enginessoftware-engineervllm-engine
posted 3mo ago · verified today
Modalinference provider
New York · San Francisco · $150K–$350K/year · unknown
inferenceinference-engineskv-cache-systemsquantizationresearch-engineer
posted 10w ago · verified today
Modalinference provider
Stockholm · mid
inferencesolutions-architectdistributed-inferencefine-tuninggpu-generic
posted 6mo ago · verified today
Modalinference provider
New York · San Francisco · $200K–$350K/year · senior
performance-engineercudagpu-kernelsinference-enginesnvidia
posted 4mo ago · verified today
Modalinference provider
New York · San Francisco · $180K–$250K/year · unknown
solutions-architectsglang-enginevllm-engineperformance-engineerinference
posted 6mo ago · verified today
Fireworks AIinference provider
London · senior
fine-tuninginferenceinference-enginespost-trainingpython-lang
posted 14d ago · verified today
Fireworks AIinference provider
San Mateo · hybrid · $175K–$220K/year · unknown
ml-platformpython-langsoftware-engineerinferenceinference-engines
posted 10mo ago · verified today
Fireworks AIinference provider
San Mateo · hybrid · $175K–$220K/year · unknown
gpu-kernelsperformance-engineercudagpu-genericsoftware-engineer
posted 16mo ago · verified today
xAIfrontier lab
Palo Alto, CA · $180K–$440K/year est. · unknown
cpp-langgpu-genericgpu-kernelsinferenceinference-engines
posted 23mo ago · verified today
xAIfrontier lab
Palo Alto, CA · $180K–$440K/year est. · unknown
cpp-langcudadistributed-inferenceinferenceinference-engines
posted 10w ago · verified today
xAIfrontier lab
London · London, England, United Kingdom · £107K–£262K/year est. · unknown
cpp-langinferenceinference-enginesml-platformobservability
posted 4mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA · $188K–$275K/year est. · staff plus
inferenceinference-enginesdistributed-inferencegpu-generickubernetes-ops
posted 4mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA · $152K–$204K/year est. · senior
cudadistributed-inferencego-langgpu-genericgpu-kernels
posted 11mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA · $139K–$204K/year est. · senior
software-engineerinferenceinference-engineskubernetes-opsgpu-generic
posted 7mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA · $206K–$333K/year est. · staff plus
cudaperformance-engineergpu-genericinferenceinference-engines
posted 9mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $293K–$385K/year · unknown
evaluationgpu-kernelsinferencesoftware-engineercollectives
posted 4mo ago · verified today
OpenAIfrontier lab
San Francisco · $266K–$500K/year · unknown
inferenceinference-enginesperformance-engineersoftware-engineergpu-generic
posted 4mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $293K–$385K/year · senior
collectivescpp-langcudadistributed-inferencegpu-generic
posted 5mo ago · verified today
OpenAIfrontier lab
San Francisco · $295K–$555K/year · unknown
gpu-genericinferenceinference-enginessoftware-engineerdistributed-inference
posted 16mo ago · verified today
OpenAIfrontier lab
San Francisco · $266K–$500K/year · unknown
cudainferenceinference-enginesnvidiasoftware-engineer
posted 19mo ago · verified today
Anthropicfrontier lab
San Francisco, CA · $320K–$485K · staff plus
inferenceinference-enginessoftware-engineerobservabilitykubernetes-ops
posted 3mo ago · verified today