AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1294 open roles · 85 companies · last verified today

186 roles

gpu-kernels

Amazon (AWS)hyperscaler

Cupertino, California, USA · onsite · senior

gpu-kernelsvllm-enginemodel-parallelismdistributed-inferenceml-platform

posted 6w ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · onsite · mid

gpu-kernelsinferenceinference-enginesdistributed-inferenceml-platform

posted 6w ago · verified today

Amazon (AWS)hyperscaler

Tel Aviv-Yafo, Tel Aviv, ISR · onsite · mid

cpp-langsoftware-engineercustom-asicdistributed-inferencegpu-kernels

posted 5w ago · verified today

Amazon (AWS)hyperscaler

Santa Clara, California, USA · onsite · unknown

cudagpu-kernelstriton-langgpu-genericmodel-parallelism

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Houston, Texas, USA · onsite · senior

cudagpu-kernelsmodel-parallelismpytorch-disttraining-frameworks

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

cudacutlass-cutedistributed-inferenceevaluationflash-attention

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Toronto, Ontario, CAN · onsite · senior

gpu-kernelsperformance-engineertrainiumcudatriton-lang

posted 15d ago · verified today

Perplexityai startup

New York City · Palo Alto +1 more · $220K–$485K/year · mid

cudagpu-kernelsinferenceinference-enginescutlass-cute

posted 5mo ago · verified today

Cerebraschip vendor

Sunnyvale, CA · Toronto, CAN · hybrid · staff plus

amdinference-enginespython-langsoftware-engineervllm-engine

posted 7w ago · verified today

Cerebraschip vendor

Sunnyvale, CA · hybrid · junior

cpp-langgpu-kernelssoftware-engineerperformance-engineerpython-lang

posted 7w ago · verified today

Cerebraschip vendor

Sunnyvale, CA · hybrid · manager

cpp-langeng-managerinferencemlir-llvmpython-lang

posted 7w ago · verified today

Cerebraschip vendor

Sunnyvale, CA · Toronto, CAN · hybrid · unknown

custom-asicperformance-engineercpp-langinferencepython-lang

posted 9w ago · verified today

Cerebraschip vendor

Sunnyvale, CA · onsite · junior

custom-asiccpp-langgpu-kernelsperformance-engineerpython-lang

posted 8mo ago · verified today

Cerebraschip vendor

Canada · United States · senior

amdcpp-langinferenceinference-enginespython-lang

posted 9mo ago · verified today

Cerebraschip vendor

Bengaluru, IND · hybrid · manager

cpp-langeng-managerpython-langsoftware-engineerperformance-engineer

posted 11mo ago · verified today

Cerebraschip vendor

Toronto, CAN · hybrid · mid

cpp-langinference-enginesmlir-llvmpython-langsoftware-engineer

posted 14mo ago · verified today

Baseteninference provider

San Francisco · hybrid · $200K–$275K/year · unknown

post-traininggpu-genericmodel-parallelismpytorch-distresearch-engineer

posted 5mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $165K–$330K/year · unknown

cpp-langnetwork-fabricnvidiasoftware-engineercollectives

posted 6mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $180K–$360K/year · mid

cudainferenceinference-enginestensorrt-stackgpu-generic

posted 11mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown

cudagpu-kernelscpp-langgpu-genericinference

posted 14mo ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · staff plus

custom-asicgpu-kernelsnetwork-fabricrdma-verbsroce-net

posted 5mo ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $180K–$255K · senior

inferenceperformance-engineercpp-langdistributed-inferenceinference-engines

posted 9mo ago · verified today

Modalinference provider

New York · San Francisco · $150K–$350K/year · unknown

inferenceinference-engineskv-cache-systemsquantizationresearch-engineer

posted 10w ago · verified today

Modalinference provider

New York · San Francisco · $200K–$350K/year · senior

performance-engineercudagpu-kernelsinference-enginesnvidia

posted 4mo ago · verified today

Fireworks AIinference provider

San Mateo · hybrid · $175K–$220K/year · unknown

gpu-kernelsperformance-engineercudagpu-genericsoftware-engineer

posted 16mo ago · verified today