AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

148 roles

pytorch-dist

Amazon (AWS)hyperscaler

Houston, Texas, USA · onsite · senior

cudagpu-kernelsmodel-parallelismpytorch-disttraining-frameworks

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Bellevue, Washington, USA · onsite · mid

software-engineerml-platformtrainiumscheduling-orchestrationfsdp

posted 6w ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

cudacutlass-cutedistributed-inferenceevaluationflash-attention

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior

solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks

posted 3mo ago · verified today

Perplexityai startup

Berlin · London +2 more · $220K–$485K/year · unknown

post-trainingreinforcement-learningpython-langtraining-frameworksresearch-engineer

posted 3w ago · verified today

Perplexityai startup

Palo Alto · San Francisco · $220K–$405K/year · unknown

cluster-datacentercpp-langgpu-genericinferencekubernetes-ops

posted 5mo ago · verified today

Perplexityai startup

New York City · Palo Alto +1 more · $220K–$485K/year · mid

cudagpu-kernelsinferenceinference-enginescutlass-cute

posted 5mo ago · verified today

Cerebraschip vendor

Sunnyvale, CA · Toronto, CAN · hybrid · staff plus

amdinference-enginespython-langsoftware-engineervllm-engine

posted 7w ago · verified today

Cerebraschip vendor

Canada · United States · senior

amdcpp-langinferenceinference-enginespython-lang

posted 9mo ago · verified today

Baseteninference provider

San Francisco · hybrid · $200K–$275K/year · unknown

post-traininggpu-genericmodel-parallelismpytorch-distresearch-engineer

posted 5mo ago · verified today

Baseteninference provider

New York · San Francisco · hybrid · $165K–$330K/year · unknown

go-langkubernetes-opsscheduling-orchestrationsoftware-engineerml-platform

posted 12mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown

inferenceinference-engineskv-cache-systemsperformance-engineerquantization

posted 30mo ago · verified today

Scale AIai startup

New York, NY · San Francisco, CA · $290K–$363K · manager

cudaeng-managerpost-trainingtraining-frameworksflash-attention

posted 11mo ago · verified today

Scale AIai startup

New York, NY · San Francisco, CA +1 more · $190K–$237K · unknown

cudaflash-attentioninference-enginesml-platformresearch-engineer

posted 18mo ago · verified today

Fireworks AIinference provider

New York · San Mateo +1 more · hybrid · $200K–$260K/year · senior

inference-enginesquantizationsglang-enginevllm-enginefine-tuning

posted 3mo ago · verified today

Fireworks AIinference provider

San Mateo · hybrid · $175K–$220K/year · unknown

ml-platformpython-langsoftware-engineerinferenceinference-engines

posted 10mo ago · verified today

Fireworks AIinference provider

San Mateo · hybrid · $175K–$220K/year · unknown

gpu-kernelsperformance-engineercudagpu-genericsoftware-engineer

posted 16mo ago · verified today

Fireworks AIinference provider

New York · San Mateo · hybrid · $210K–$320K/year · unknown

software-engineertraining-frameworksfsdpkubernetes-opsml-platform

posted 6w ago · verified today

xAIfrontier lab

Palo Alto, CA · $180K–$440K/year est. · unknown

cpp-langcudadistributed-inferenceinferenceinference-engines

posted 10w ago · verified today

xAIfrontier lab

Palo Alto, CA · $180K–$440K/year est. · unknown

gpu-genericpost-trainingpre-trainingpython-langresearch-engineer

posted 5mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $182K–$242K/year est. · senior

kubernetes-opsml-platformreinforcement-learningsoftware-engineerevaluation

posted 11w ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $182K–$242K/year est. · senior

gpu-genericpost-trainingpython-langreinforcement-learningresearch-engineer

posted 5mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $206K–$333K/year est. · staff plus

cudaperformance-engineergpu-genericinferenceinference-engines

posted 9mo ago · verified today

CoreWeaveneocloud

Atlanta, GA · San Francisco, CA · $182K–$242K/year est. · unknown

solutions-architectpython-langnvidiafine-tuninginfiniband-ops

posted 6w ago · verified today

OpenAIfrontier lab

San Francisco · Seattle · hybrid · $293K–$385K/year · senior

collectivescpp-langcudadistributed-inferencegpu-generic

posted 5mo ago · verified today

OpenAIfrontier lab

San Francisco · hybrid · $295K–$500K/year · unknown

performance-engineercollectivescpp-langcudagpu-generic

posted 11mo ago · verified today