AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1294 open roles · 85 companies · last verified today
210 roles
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
trainiumcpp-langdistributed-inferenceinferenceinference-engines
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · manager
eng-managertrainiuminferenceinference-enginesdistributed-inference
posted 8w ago · verified today
Amazon (AWS)hyperscaler
New York, New York, USA · onsite · mid
software-engineertrainiumdistributed-inferencegpu-kernelstraining-frameworks
posted 7w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · Seattle, Washington, USA · onsite · mid
trainiumperformance-engineersoftware-engineercollectivesgpu-generic
posted 7mo ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
software-engineertrainiuminference-enginesvllm-enginedistributed-inference
posted 7w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · Seattle, Washington, USA · onsite · mid
trainiummodel-parallelismtraining-frameworkscollectivesperformance-engineer
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Tel Aviv-Yafo, Tel Aviv, ISR · onsite · senior
inferenceinference-enginesvllm-enginecudagpu-generic
posted 10w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
gpu-kernelsvllm-enginemodel-parallelismdistributed-inferenceml-platform
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · mid
gpu-kernelsinferenceinference-enginesdistributed-inferenceml-platform
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Tel Aviv-Yafo, Tel Aviv, ISR · onsite · mid
cpp-langsoftware-engineercustom-asicdistributed-inferencegpu-kernels
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Houston, Texas, USA · onsite · senior
cudagpu-kernelsmodel-parallelismpytorch-disttraining-frameworks
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
cudacutlass-cutedistributed-inferenceevaluationflash-attention
posted 4w ago · verified today
Anthropicfrontier lab
Ontario, CA · Ontario, CAN · staff plus
distributed-inferenceinferenceinference-enginessoftware-engineerscheduling-orchestration
posted 4w ago · verified today
Perplexityai startup
London · unknown
cudagpu-kernelsinferenceinference-enginesgpu-generic
posted 5mo ago · verified today
Perplexityai startup
New York City · Palo Alto +1 more · $220K–$485K/year · mid
cudagpu-kernelsinferenceinference-enginescutlass-cute
posted 5mo ago · verified today
Cerebraschip vendor
Sunnyvale, CA · Toronto, CAN · hybrid · staff plus
amdinference-enginespython-langsoftware-engineervllm-engine
posted 7w ago · verified today
Cerebraschip vendor
Sunnyvale, CA · hybrid · staff plus
inferencereliability-sresreobservabilitycluster-datacenter
posted 10w ago · verified today
Cerebraschip vendor
Sunnyvale, CA · onsite · senior
ansibleinferenceinference-engineskubernetes-opsml-platform
posted 4mo ago · verified today
Cerebraschip vendor
Canada · United States · unknown
custom-asicperformance-engineercpp-langinferencepython-lang
posted 7mo ago · verified today
Cerebraschip vendor
Canada · United States · senior
amdcpp-langinferenceinference-enginespython-lang
posted 9mo ago · verified today
Cerebraschip vendor
Canada · United States · senior
cpp-langsoftware-engineerml-platforminferencepython-lang
posted 10mo ago · verified today
Cerebraschip vendor
Sunnyvale, CA · onsite · staff plus
inferenceinference-enginessoftware-engineercpp-langdistributed-inference
posted 26mo ago · verified today
Cerebraschip vendor
Sunnyvale, CA · Toronto, CAN · onsite · staff plus
inferencekubernetes-opssoftware-engineergo-langml-platform
posted 12w ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown
inference-enginessoftware-engineerdistributed-inferenceinferencekubernetes-ops
posted 3mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $165K–$330K/year · unknown
inferenceinference-enginessoftware-engineerdistributed-inferenceperformance-engineer
posted 4mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $165K–$330K/year · unknown
cpp-langnetwork-fabricnvidiasoftware-engineercollectives
posted 6mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $260K–$380K/year · manager
eng-managerinferenceinference-enginespython-langdistributed-inference
posted 4mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $180K–$360K/year · mid
cudainferenceinference-enginestensorrt-stackgpu-generic
posted 11mo ago · verified today
Baseteninference provider
Montreal · New York +3 more · hybrid · $165K–$330K/year · unknown
inferencekubernetes-opsml-platformpython-langsoftware-engineer
posted 18mo ago · verified today
Baseteninference provider
Montreal · New York +3 more · hybrid · $165K–$330K/year · unknown
inferencekubernetes-opsml-platformsoftware-engineerinference-engines
posted 11mo ago · verified today