AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

37 roles

xla-compiler

Inferactinference provider

San Francisco · onsite · junior

amdcudagpu-kernelsinferenceinference-engines

posted 1d ago · verified today

Inceptionfrontier lab

San Mateo, United States · onsite · unknown

python-langsoftware-engineertraining-frameworksgpu-genericpytorch-dist

posted 6mo ago · verified today

Luma AIai startup

Redwood City, CA · hybrid · unknown

cudagpu-kernelstriton-langgpu-genericnvidia

posted 7w ago · verified today

Coherefrontier lab

Canada · Europe +1 more · remote · senior

python-langsoftware-engineertraining-frameworksjax-pallaskubernetes-ops

posted 4w ago · verified today

Coherefrontier lab

London · Montreal +3 more · remote · unknown

post-trainingpython-langresearch-engineerjax-pallaspytorch-dist

posted 13mo ago · verified today

Coherefrontier lab

London · Montreal +4 more · hybrid · unknown

post-trainingfine-tuningjax-pallaskubernetes-opsml-platform

posted 15mo ago · verified today

Coherefrontier lab

London · Montreal +4 more · remote · mid

jax-pallasml-platformpython-langpytorch-distsoftware-engineer

posted 19mo ago · verified today

Coherefrontier lab

London · Montreal +3 more · hybrid · unknown

cudagpu-kernelsperformance-engineertriton-langpre-training

posted 19mo ago · verified today

Coherefrontier lab

London · Montreal +4 more · remote · unknown

cudapython-langgpu-genericgpu-kernelsmlir-llvm

posted 22mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

distributed-inferenceinferenceinference-enginesperformance-engineernvidia

posted 4mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid

distributed-inferenceinferencejax-pallastpuxla-compiler

posted 8mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

amdnvidiagpu-kernelscollectivescpp-lang

posted 6w ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior

cpp-langcudagpu-kernelsperformance-engineercollectives

posted 7mo ago · verified today

Inferactinference provider

San Francisco · onsite · manager

eng-managerinferenceinference-enginesvllm-enginedistributed-inference

posted 6w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

tpuinferenceinference-enginesjax-pallasperformance-engineer

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · $200K–$400K/year · unknown

gpu-kernelscudaperformance-engineercpp-langinference

posted 12w ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

tpuinference-enginesjax-pallasperformance-engineerxla-compiler

posted 3mo ago · verified today

OpenAIfrontier lab

San Francisco · hybrid · $295K–$380K/year · unknown

software-engineertrainiumgpu-kernelsinferenceinference-engines

posted 3w ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · Seattle, Washington, USA · onsite · senior

software-engineercpp-langmlir-llvmtrainiumxla-compiler

posted 3mo ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · Seattle, Washington, USA · onsite · unknown

trainiumsoftware-engineerml-platformtraining-frameworksperformance-engineer

posted 9w ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · unknown

software-engineertrainiumcpp-langlinux-kernelml-platform

posted 5mo ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · onsite · manager

eng-managermlir-llvmtrainiumxla-compilertpu

posted 7mo ago · verified today