AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
984 open roles · 19 companies · last verified today
82 roles
Baseteninference provider
New York · San Francisco · remote · $165K–$330K · manager
eng-managerinferenceinference-enginessglang-enginesolutions-architect
posted 3mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $260K–$380K · manager
eng-managerinferenceinference-enginespython-langdistributed-inference
posted 3mo ago · verified today
Baseteninference provider
New York · San Francisco · remote · $165K–$330K · manager
inferenceinference-enginesml-platformdistributed-inferencescheduling-orchestration
posted 4mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $165K–$330K · mid
inferenceinference-enginessoftware-engineergpu-genericdistributed-inference
posted 3mo ago · verified today
SambaNovachip vendor
Austin, Texas, United States; San Jose, California, United States · Austin, TX +1 more · senior
inferenceinference-enginespython-langquantizationcustom-asic
posted 3d ago · verified today
SambaNovachip vendor
Remote - US · remote · senior
inferenceinference-enginespython-langvllm-enginecustom-asic
posted 3d ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · senior
performance-engineergpu-genericgpu-kernelsinferenceinference-engines
posted 3d ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · staff plus
inferenceinference-enginesperformance-engineercustom-asicquantization
posted 3d ago · verified today
Scale AIai startup
New York, NY · San Francisco, CA +1 more · senior
inferenceinference-enginessoftware-engineerkubernetes-opsml-platform
posted 3d ago · verified today
Anyscaleai startup
Palo Alto · San Francisco · remote · $170K–$245K · mid
distributed-inferencegpu-genericinferenceinference-enginessoftware-engineer
posted 12w ago · verified today
Modalinference provider
New York · San Francisco · $200K–$350K · senior
inferenceinference-enginesperformance-engineercudagpu-generic
posted 4mo ago · verified today
Fireworks AIinference provider
London · senior
inferenceinference-enginessolutions-architectfine-tuninggpu-generic
posted 4w ago · verified today
Fireworks AIinference provider
New York · San Mateo · $200K–$260K · senior
inference-enginespython-langsolutions-architectfine-tuninginference
posted 10w ago · verified today
Fireworks AIinference provider
New York · San Mateo +1 more · $200K–$260K · senior
python-langsolutions-architectinferenceinference-enginesfine-tuning
posted 10w ago · verified today
Fireworks AIinference provider
San Mateo · $175K–$220K · mid
inferencesoftware-engineerdistributed-inferenceinference-engineskubernetes-ops
posted 9mo ago · verified today
Fireworks AIinference provider
Singapore · 200K–350K SGD · senior
solutions-architectfine-tuninggpu-genericinferenceinference-engines
posted 4w ago · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · senior
cpp-langgpu-genericinferenceinference-enginessoftware-engineer
posted 3d ago · verified today
xAIfrontier lab
London · London, England, United Kingdom · mid
software-engineerinferenceinference-enginesrust-langcpp-lang
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · senior
gpu-generickubernetes-opsml-platformpython-langnvidia
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · senior
cudagpu-kernelsperformance-engineercpp-langinference
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · senior
distributed-inferenceinferenceinference-engineskubernetes-opsml-platform
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Livingston, NJ +4 more · senior
inference-enginesinferencesolutions-architectvllm-enginedistributed-inference
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Bellevue, WA/ San Francisco, CA/ Sunnyvale, CA +3 more · mid
inferenceinference-enginespython-langperformance-engineervllm-engine
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · staff plus
gpu-genericinferenceinference-engineskubernetes-opssoftware-engineer
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · junior
cpp-langgo-langgpu-genericinferenceinference-engines
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · senior
inferenceinference-engineskubernetes-opspython-langcuda
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · staff plus
inferenceinference-enginesnvidiaperformance-engineertraining-frameworks
posted 3d ago · verified today
OpenAIfrontier lab
San Francisco · $295K–$555K · mid
inferenceinference-enginessoftware-engineerdistributed-inferencegpu-generic
posted 15mo ago · verified today
Nebiusneocloud
Remote - United States · United States · senior
inferenceinference-enginessolutions-architectfine-tuningpython-lang
posted 3d ago · verified today
Nebiusneocloud
Remote - Europe · remote · manager
kubernetes-opspython-langgo-langgpu-genericinfiniband-ops
posted 3d ago · verified today