AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1297 open roles · 85 companies · last verified today
295 roles
DeepInfrainference provider
Palo Alto, United States · onsite · unknown
inferenceinference-enginesnvidiapython-langsolutions-architect
posted 4w ago · verified today
DeepInfrainference provider
Sofia, Bulgaria · onsite · junior
inferenceinference-enginespython-langsoftware-engineercpp-lang
posted 11w ago · verified today
DeepInfrainference provider
Bulgaria - Remote · remote · mid
cpp-langcudagpu-genericinferenceinference-engines
posted 5mo ago · verified today
DeepInfrainference provider
Palo Alto, United States · onsite · mid
cudainferencepython-langsoftware-engineercpp-lang
posted 5mo ago · verified today
DeepInfrainference provider
Bulgaria - Remote · remote · junior
cpp-langcudainferencepython-langsoftware-engineer
posted 7mo ago · verified today
DeepInfrainference provider
Palo Alto, United States · onsite · $140K–$150K/year est. · junior
inference-enginessoftware-engineercpp-langcudaml-platform
posted 8mo ago · verified today
DeepInfrainference provider
Bulgaria - Remote · remote · junior
cpp-langcudainferenceinference-enginespython-lang
posted 11mo ago · verified today
DeepInfrainference provider
Palo Alto, United States · onsite · junior
cpp-langcudanccl-libpython-langsoftware-engineer
posted 11mo ago · verified today
Sakana AIfrontier lab
Tokyo · unknown
reliability-srenvidiainferencesregpu-generic
added 7d ago · verified today
Fish Audio (39 AI)ai startup
Location not specified · senior
cluster-datacentergpu-genericreliability-sresrekubernetes-ops
added 7d ago · verified today
Fish Audio (39 AI)ai startup
Location not specified · unknown
cpp-langinference-enginespython-langinferencemegatron-lm
added 7d ago · verified today
Interfazeinference provider
San Francisco, CA · senior
fine-tuninginferenceinference-enginespython-langquantization
added 7d ago · verified today
Waferinference provider
San Francisco · onsite · $200K–$300K/year · unknown
gpu-kernelsinferenceinference-enginescluster-datacenterperformance-engineer
posted 8w ago · verified today
Runwareinference provider
United Kingdom · remote · senior
gpu-genericinferencenvidiareliability-sresre
posted 4mo ago · verified today
Relaceinference provider
San Francisco · onsite · mid
software-engineercluster-datacenterscheduling-orchestrationinferenceml-platform
posted 10mo ago · verified today
Exaai startup
Singapore · onsite · 90K–300K SGD/year · unknown
cluster-datacenterkubernetes-opsscheduling-orchestrationsoftware-engineerdistributed-inference
posted 6mo ago · verified today
Inceptionfrontier lab
San Mateo, United States · onsite · senior
inferenceinference-enginespython-langsoftware-engineergpu-generic
posted 6mo ago · verified today
Inceptionfrontier lab
San Mateo, United States · onsite · unknown
cudagpu-genericinferenceinference-engineskubernetes-ops
posted 6mo ago · verified today
Novita AIinference provider
San Mateo · onsite · unknown
kubernetes-opspython-langsolutions-architectinferenceinference-engines
posted 11w ago · verified today
Black Forest Labsfrontier lab
San Francisco (United States) · onsite · unknown
cudagpu-genericinferenceinference-enginesperformance-engineer
posted 24mo ago · verified today
Crusoeneocloud
San Francisco, CA - US · onsite · staff plus
inferenceinference-enginesperformance-engineersglang-enginevllm-engine
posted 7d ago · verified today
Xaira Therapeuticsai startup
Seattle, WA · Seattle, Washington, United States +1 more · $185K–$308K/year est. · senior
gpu-genericgpu-kernelsinferenceresearch-engineerinference-engines
posted 10mo ago · verified today
Fireworks AIinference provider
New York · San Mateo · hybrid · $200K–$230K/year · unknown
scheduling-orchestrationcpp-langnetwork-fabricpython-langstorage-checkpointing
posted 8d ago · verified today
Together AIneocloud
San Francisco · $270K–$300K/year est. · senior
distributed-inferencefine-tuninginferenceinference-engineskv-cache-systems
posted 4mo ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
distributed-inferenceinferenceinference-enginespython-langpytorch-dist
posted 9d ago · verified today
Harveyai startup
San Francisco · hybrid · $231K–$340K/year · staff plus
inferenceinference-enginessoftware-engineerml-platformobservability
posted 9d ago · verified today
Lila Sciencesai startup
Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus
inferenceinference-engineskubernetes-opsml-platformnvidia
posted 7w ago · verified today
Lila Sciencesai startup
San Francisco, CA · San Francisco, CA USA · $268K–$384K · staff plus
ml-platformgpu-genericsoftware-engineertraining-frameworksinference-engines
posted 4mo ago · verified today
Abridgeai startup
SF Office · hybrid · $221K–$260K/year · senior
inferenceinference-engineskubernetes-opssoftware-engineercuda
posted 12mo ago · verified today
Genesis Molecular AIai startup
New York, NY · San Mateo, CA · hybrid · senior
research-engineergpu-genericpre-trainingpython-langcuda
posted 13mo ago · verified today