AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1294 open roles · 85 companies · last verified today
186 roles
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
gpu-kernelsvllm-enginemodel-parallelismdistributed-inferenceml-platform
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · mid
gpu-kernelsinferenceinference-enginesdistributed-inferenceml-platform
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Tel Aviv-Yafo, Tel Aviv, ISR · onsite · mid
cpp-langsoftware-engineercustom-asicdistributed-inferencegpu-kernels
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Santa Clara, California, USA · onsite · unknown
cudagpu-kernelstriton-langgpu-genericmodel-parallelism
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Houston, Texas, USA · onsite · senior
cudagpu-kernelsmodel-parallelismpytorch-disttraining-frameworks
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
cudacutlass-cutedistributed-inferenceevaluationflash-attention
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Toronto, Ontario, CAN · onsite · senior
gpu-kernelsperformance-engineertrainiumcudatriton-lang
posted 15d ago · verified today
Perplexityai startup
London · unknown
cudagpu-kernelsinferenceinference-enginesgpu-generic
posted 5mo ago · verified today
Perplexityai startup
New York City · Palo Alto +1 more · $220K–$485K/year · mid
cudagpu-kernelsinferenceinference-enginescutlass-cute
posted 5mo ago · verified today
Cerebraschip vendor
Sunnyvale, CA · Toronto, CAN · hybrid · staff plus
amdinference-enginespython-langsoftware-engineervllm-engine
posted 7w ago · verified today
Cerebraschip vendor
Sunnyvale, CA · hybrid · junior
cpp-langgpu-kernelssoftware-engineerperformance-engineerpython-lang
posted 7w ago · verified today
Cerebraschip vendor
Sunnyvale, CA · hybrid · manager
cpp-langeng-managerinferencemlir-llvmpython-lang
posted 7w ago · verified today
Cerebraschip vendor
Sunnyvale, CA · Toronto, CAN · hybrid · unknown
custom-asicperformance-engineercpp-langinferencepython-lang
posted 9w ago · verified today
Cerebraschip vendor
Bengaluru, IND · mid
custom-asicperformance-engineercpp-langgpu-kernelsinference
posted 3mo ago · verified today
Cerebraschip vendor
Remote · staff plus
cpp-langpython-langgpu-kernelsperformance-engineersoftware-engineer
posted 7mo ago · verified today
Cerebraschip vendor
Sunnyvale, CA · onsite · junior
custom-asiccpp-langgpu-kernelsperformance-engineerpython-lang
posted 8mo ago · verified today
Cerebraschip vendor
Canada · United States · senior
amdcpp-langinferenceinference-enginespython-lang
posted 9mo ago · verified today
Cerebraschip vendor
Bengaluru, IND · hybrid · manager
cpp-langeng-managerpython-langsoftware-engineerperformance-engineer
posted 11mo ago · verified today
Cerebraschip vendor
Toronto, CAN · hybrid · mid
cpp-langinference-enginesmlir-llvmpython-langsoftware-engineer
posted 14mo ago · verified today
Baseteninference provider
San Francisco · hybrid · $200K–$275K/year · unknown
post-traininggpu-genericmodel-parallelismpytorch-distresearch-engineer
posted 5mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $165K–$330K/year · unknown
cpp-langnetwork-fabricnvidiasoftware-engineercollectives
posted 6mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $180K–$360K/year · mid
cudainferenceinference-enginestensorrt-stackgpu-generic
posted 11mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown
cudagpu-kernelscpp-langgpu-genericinference
posted 14mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · staff plus
custom-asicgpu-kernelsnetwork-fabricrdma-verbsroce-net
posted 5mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · $180K–$255K · senior
inferenceperformance-engineercpp-langdistributed-inferenceinference-engines
posted 9mo ago · verified today
Modalinference provider
New York · San Francisco · $150K–$350K/year · unknown
inferenceinference-engineskv-cache-systemsquantizationresearch-engineer
posted 10w ago · verified today
Modalinference provider
Stockholm · mid
inferencesolutions-architectdistributed-inferencefine-tuninggpu-generic
posted 6mo ago · verified today
Modalinference provider
New York · San Francisco · $200K–$350K/year · senior
performance-engineercudagpu-kernelsinference-enginesnvidia
posted 4mo ago · verified today
Fireworks AIinference provider
London · senior
fine-tuninginferenceinference-enginespost-trainingpython-lang
posted 14d ago · verified today
Fireworks AIinference provider
San Mateo · hybrid · $175K–$220K/year · unknown
gpu-kernelsperformance-engineercudagpu-genericsoftware-engineer
posted 16mo ago · verified today