AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
984 open roles · 19 companies · last verified today
110 roles
Baseteninference provider
New York · San Francisco · remote · $165K–$330K · senior
software-engineerfine-tuninggpu-genericml-platformpost-training
posted 6mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $165K–$330K · mid
inferenceinference-enginessoftware-engineergpu-genericdistributed-inference
posted 3mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $180K–$360K · mid
inferenceinference-enginesml-platformsoftware-engineertensorrt-stack
posted 10mo ago · verified today
Baseteninference provider
Montreal · New York +3 more · remote · $165K–$330K · mid
inferenceinference-enginessolutions-architectgpu-genericvllm-engine
posted 5mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $260K–$380K · manager
eng-managerinferenceinference-enginespython-langdistributed-inference
posted 3mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $165K–$330K · senior
collectivescpp-langdistributed-inferencegpu-genericinference
posted 5mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $180K–$360K · mid
software-engineerinference-enginesinferencekubernetes-opsml-platform
posted 11w ago · verified today
SambaNovachip vendor
Remote - US · remote · senior
inferenceinference-enginespython-langsolutions-architectfine-tuning
posted 3d ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · senior
performance-engineergpu-genericgpu-kernelsinferenceinference-engines
posted 3d ago · verified today
SambaNovachip vendor
Austin, Texas, United States; San Jose, California, United States · Austin, TX +1 more · senior
inferenceinference-enginespython-langquantizationcustom-asic
posted 3d ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · staff plus
inferenceinference-enginesperformance-engineercustom-asicquantization
posted 3d ago · verified today
SambaNovachip vendor
Remote - US · remote · senior
inferenceinference-enginespython-langvllm-enginecustom-asic
posted 3d ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · staff plus
observabilityreliability-sresreinferencekubernetes-ops
posted 3d ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · senior
inferenceinference-engineskubernetes-opspython-langsglang-engine
posted 3d ago · verified today
Scale AIai startup
New York, NY · San Francisco, CA +1 more · senior
inferenceinference-enginessoftware-engineerkubernetes-opsml-platform
posted 3d ago · verified today
Anyscaleai startup
Palo Alto · San Francisco · remote · $170K–$245K · mid
distributed-inferencegpu-genericinferenceinference-enginessoftware-engineer
posted 12w ago · verified today
Modalinference provider
New York · San Francisco · $200K–$350K · senior
inferenceinference-enginesperformance-engineercudagpu-generic
posted 4mo ago · verified today
Modalinference provider
New York · San Francisco · $180K–$250K · mid
sglang-enginesolutions-architectvllm-enginedistributed-inferenceinference-engines
posted 5mo ago · verified today
Modalinference provider
Stockholm · mid
gpu-genericsolutions-architectinferenceinference-enginespython-lang
posted 5mo ago · verified today
Fireworks AIinference provider
San Mateo · $175K–$220K · mid
inferencesoftware-engineerdistributed-inferenceinference-engineskubernetes-ops
posted 9mo ago · verified today
Fireworks AIinference provider
London · senior
inferenceinference-enginessolutions-architectfine-tuninggpu-generic
posted 4w ago · verified today
Fireworks AIinference provider
Singapore · 200K–350K SGD · senior
solutions-architectfine-tuninggpu-genericinferenceinference-engines
posted 4w ago · verified today
Fireworks AIinference provider
New York · San Mateo · $200K–$260K · senior
inference-enginespython-langsolutions-architectfine-tuninginference
posted 10w ago · verified today
Fireworks AIinference provider
New York · San Mateo · $210K–$320K · mid
gpu-genericpython-langpytorch-distcpp-langtraining-frameworks
posted 6w ago · verified today
Fireworks AIinference provider
New York · San Mateo +1 more · $200K–$260K · senior
python-langsolutions-architectinferenceinference-enginesfine-tuning
posted 10w ago · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · mid
cpp-langcudadistributed-inferenceinferenceinference-engines
posted 3d ago · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · senior
cpp-langgpu-genericinferenceinference-enginessoftware-engineer
posted 3d ago · verified today
xAIfrontier lab
London · London, England, United Kingdom · mid
software-engineerinferenceinference-enginesrust-langcpp-lang
posted 3d ago · verified today
CoreWeaveneocloud
Seattle, WA · mid
python-langsolutions-architectinferencepytorch-distfine-tuning
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · senior
distributed-inferenceinferenceinference-engineskubernetes-opsml-platform
posted 3d ago · verified today