AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1296 open roles · 85 companies · last verified today
100 roles
Amazon (AWS)hyperscaler
Tel Aviv-Yafo, Tel Aviv, ISR · onsite · mid
cpp-langsoftware-engineercustom-asicdistributed-inferencegpu-kernels
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior
solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks
posted 3mo ago · verified today
Cerebraschip vendor
Sunnyvale, CA · Toronto, CAN · hybrid · staff plus
amdinference-enginespython-langsoftware-engineervllm-engine
posted 7w ago · verified today
Cerebraschip vendor
Canada · United States · senior
amdcpp-langinferenceinference-enginespython-lang
posted 9mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown
software-engineerdistributed-inferenceinferenceinference-enginesml-platform
posted 3mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $165K–$330K/year · unknown
cpp-langnetwork-fabricnvidiasoftware-engineercollectives
posted 6mo ago · verified today
Baseteninference provider
Montreal · New York +3 more · hybrid · $165K–$330K/year · unknown
solutions-architectinferenceinference-enginessglang-enginetensorrt-stack
posted 6mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $180K–$360K/year · mid
inferenceinference-enginessoftware-engineercudadistributed-inference
posted 11mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown
inferenceinference-engineskv-cache-systemsperformance-engineerquantization
posted 30mo ago · verified today
SambaNovachip vendor
Austin, Texas, United States · Austin, TX +2 more · $200K–$275K · senior
inferenceinference-enginessoftware-engineerspeculative-decodingvllm-engine
posted 3mo ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · 3324K–4062K INR · unknown
python-langsolutions-architectinferencefine-tuningkubernetes-ops
posted 9w ago · verified today
Scale AIai startup
London, UK · unknown
inferenceinference-enginessoftware-engineercpp-langgo-lang
posted 7w ago · verified today
Modalinference provider
New York · San Francisco · $150K–$350K/year · unknown
inferenceinference-engineskv-cache-systemsquantizationresearch-engineer
posted 10w ago · verified today
Modalinference provider
Stockholm · mid
inferencesolutions-architectdistributed-inferencefine-tuninggpu-generic
posted 6mo ago · verified today
Modalinference provider
New York · San Francisco · $180K–$250K/year · unknown
solutions-architectsglang-enginevllm-engineperformance-engineerinference
posted 6mo ago · verified today
Fireworks AIinference provider
New York · San Mateo +1 more · hybrid · $200K–$260K/year · senior
inference-enginesquantizationsglang-enginevllm-enginefine-tuning
posted 3mo ago · verified today
Fireworks AIinference provider
New York · San Mateo · hybrid · $200K–$260K/year · unknown
fine-tuningpython-langsglang-enginesolutions-architectvllm-engine
posted 3mo ago · verified today
Fireworks AIinference provider
London · senior
fine-tuninginferenceinference-enginespython-langsglang-engine
posted 8w ago · verified today
Fireworks AIinference provider
Singapore · 200K–350K SGD/year · senior
fine-tuninginferenceinference-enginespython-langsglang-engine
posted 7w ago · verified today
Fireworks AIinference provider
San Mateo · hybrid · $175K–$220K/year · unknown
ml-platformpython-langsoftware-engineerinferenceinference-engines
posted 10mo ago · verified today
xAIfrontier lab
Palo Alto, CA · $180K–$440K/year est. · unknown
cpp-langgpu-genericgpu-kernelsinferenceinference-engines
posted 23mo ago · verified today
xAIfrontier lab
Palo Alto, CA · $180K–$440K/year est. · unknown
cpp-langcudadistributed-inferenceinferenceinference-engines
posted 10w ago · verified today
xAIfrontier lab
London · London, England, United Kingdom · £107K–£262K/year est. · unknown
cpp-langinferenceinference-enginesml-platformobservability
posted 4mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $293K–$385K/year · unknown
evaluationgpu-kernelsinferencesoftware-engineercollectives
posted 4mo ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · hybrid · $293K–$445K/year · senior
inference-enginesml-platformcpp-langgo-langinference
posted 8w ago · verified today
Nebiusneocloud
Amsterdam · Amsterdam, Netherlands +7 more · senior
go-langgpu-generickubernetes-opsscheduling-orchestrationsoftware-engineer
posted 7w ago · verified today
Nebiusneocloud
Amsterdam · Amsterdam, Netherlands +14 more · remote · senior
go-langinferenceinference-engineskv-cache-systemspython-lang
posted 11w ago · verified today
Nebiusneocloud
Remote - Europe · remote · senior
inferencepython-langsolutions-architectfine-tuningvllm-engine
posted 6mo ago · verified today
Nebiusneocloud
Singapore · senior
solutions-architectinferencevllm-enginefine-tuningpython-lang
posted 4mo ago · verified today
Nebiusneocloud
Amsterdam, Netherlands · Berlin, Germany +6 more · remote · senior
gpu-genericinferenceinference-enginesml-platformperformance-engineer
posted 6mo ago · verified today