ai-infra-jobs

AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

984 open roles · 19 companies · last verified today

136 roles

cuda

Amazon (AWS)hyperscaler

Cupertino, California, USA · manager

collectivescpp-langnccl-libnetwork-fabricnvshmem-lib

posted 20d ago · verified today

OpenAIfrontier lab

San Francisco · remote · $302K–$445K · manager

cluster-datacentercustom-asicdatacenter-engineereng-managergpu-generic

posted 3d ago · verified today

Perplexityai startup

Palo Alto · San Francisco · $220K–$485K · mid

post-trainingpython-langmegatron-lmpytorch-distreinforcement-learning

posted 4mo ago · verified today

Perplexityai startup

New York City · Palo Alto +1 more · $220K–$485K · senior

cudacutlass-cutegpu-kernelsinferenceinference-engines

posted 4mo ago · verified today

Cerebraschip vendor

Headquarters/Sunnyvale Office · Toronto Office · remote · staff plus

cpp-langinferenceinference-enginesrocm-hipvllm-engine

posted 3w ago · verified today

Cerebraschip vendor

Headquarters/Sunnyvale Office · senior

inferenceinference-enginesperformance-engineercudavllm-engine

posted 4mo ago · verified today

Cerebraschip vendor

US and Canada Offices · senior

inference-enginesamdcpp-langdistributed-inferenceinference

posted 8mo ago · verified today

Cerebraschip vendor

US and Canada Offices · remote · junior

cpp-langcustom-asicgpu-kernelsperformance-engineerpython-lang

posted 4w ago · verified today

Cerebraschip vendor

Headquarters/Sunnyvale Office · staff plus

post-trainingcustom-asicreinforcement-learningpre-trainingpython-lang

posted 8mo ago · verified today

Cerebraschip vendor

US and Canada Offices · mid

software-engineercpp-langcustom-asicgpu-kernelspython-lang

posted 5mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · remote · $180K–$360K · mid

cpp-langcudagpu-genericgpu-kernelsinference

posted 13mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · remote · $180K–$360K · mid

gpu-genericinferenceinference-enginescpp-langnvidia

posted 29mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · remote · $165K–$330K · senior

gpu-genericinferenceinference-enginesnvidiaperformance-engineer

posted 7mo ago · verified today

Baseteninference provider

Montreal · New York +2 more · remote · $180K–$360K · mid

inferenceinference-enginesml-platformsoftware-engineertensorrt-stack

posted 10mo ago · verified today

SambaNovachip vendor

Remote - US · remote · senior

inferenceinference-enginespython-langsolutions-architectfine-tuning

posted 3d ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · staff plus

inferenceinference-enginesperformance-engineercustom-asicquantization

posted 3d ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · senior

performance-engineergpu-genericgpu-kernelsinferenceinference-engines

posted 3d ago · verified today

SambaNovachip vendor

Bengaluru, India · Bengaluru, Karnataka, India · senior

inferenceinference-engineskubernetes-opspython-langsglang-engine

posted 3d ago · verified today

Scale AIai startup

San Francisco, CA · San Francisco, CA; New York, NY · manager

ml-platformpost-trainingtraining-frameworkseng-managergpu-generic

posted 3d ago · verified today

Scale AIai startup

San Francisco, CA · San Francisco, CA; Seattle, WA; New York, NY · mid

cudaml-platformgpu-genericpytorch-distsoftware-engineer

posted 3d ago · verified today

Anyscaleai startup

Palo Alto · San Francisco · remote · $170K–$245K · mid

distributed-inferencegpu-genericinferenceinference-enginessoftware-engineer

posted 12w ago · verified today

Modalinference provider

New York · San Francisco · $200K–$350K · senior

inferenceinference-enginesperformance-engineercudagpu-generic

posted 4mo ago · verified today

Fireworks AIinference provider

Singapore · 200K–350K SGD · senior

solutions-architectfine-tuninggpu-genericinferenceinference-engines

posted 4w ago · verified today

Fireworks AIinference provider

New York · San Mateo · $210K–$320K · mid

gpu-genericpython-langpytorch-distcpp-langtraining-frameworks

posted 6w ago · verified today

Fireworks AIinference provider

San Mateo · $175K–$220K · senior

cudagpu-genericgpu-kernelsinferenceperformance-engineer

posted 15mo ago · verified today

xAIfrontier lab

Palo Alto, CA · Palo Alto, California · senior

cpp-langgpu-genericinferenceinference-enginessoftware-engineer

posted 3d ago · verified today

xAIfrontier lab

Palo Alto, CA · Palo Alto, California · mid

cpp-langcudadistributed-inferenceinferenceinference-engines

posted 3d ago · verified today

xAIfrontier lab

Palo Alto, CA · Palo Alto, California · senior

cpp-langcudacutlass-cutegpu-kernelssoftware-engineer

posted 3d ago · verified today