AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

114 roles

collectives

SambaNovachip vendor

Stockholm, Sweden · staff plus

cpp-langsoftware-engineerinferencelinux-kernelnetwork-fabric

posted 19d ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer

posted 20d ago · verified today

OpenAIfrontier lab

San Francisco · hybrid · $266K–$445K/year · unknown

software-engineerinference-enginesinferencekv-cache-systemsscheduling-orchestration

posted 3w ago · verified today

Amazon (AWS)hyperscaler

Austin, Texas, USA · Cupertino, California, USA · onsite · mid

cpp-langsoftware-engineercollectivespython-langrust-lang

posted 5mo ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · Seattle, Washington, USA · onsite · manager

collectivescpp-langcudagpu-kernelsnccl-lib

posted 6w ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · Seattle, Washington, USA · onsite · mid

trainiumperformance-engineersoftware-engineercollectivesgpu-generic

posted 7mo ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · Seattle, Washington, USA · onsite · mid

trainiummodel-parallelismtraining-frameworkscollectivesperformance-engineer

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · Seattle, Washington, USA · onsite · senior

software-engineertrainiummodel-parallelismpre-trainingtraining-frameworks

posted 5w ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · junior

network-fabricefa-fabricmpinetwork-engineersoftware-engineer

posted 8w ago · verified today

Amazon (AWS)hyperscaler

Santa Clara, California, USA · onsite · unknown

cudagpu-kernelstriton-langgpu-genericmodel-parallelism

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

cudacutlass-cutedistributed-inferenceevaluationflash-attention

posted 4w ago · verified today

Perplexityai startup

New York City · Palo Alto +1 more · $220K–$485K/year · mid

cudagpu-kernelsinferenceinference-enginescutlass-cute

posted 5mo ago · verified today

Cerebraschip vendor

Sunnyvale, CA · Toronto, CAN · hybrid · staff plus

amdinference-enginespython-langsoftware-engineervllm-engine

posted 7w ago · verified today

Cerebraschip vendor

Canada · United States · senior

amdcpp-langinferenceinference-enginespython-lang

posted 9mo ago · verified today

Baseteninference provider

San Francisco · hybrid · $200K–$275K/year · unknown

post-traininggpu-genericmodel-parallelismpytorch-distresearch-engineer

posted 5mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $165K–$330K/year · unknown

cpp-langnetwork-fabricnvidiasoftware-engineercollectives

posted 6mo ago · verified today