AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1293 open roles · 85 companies · last verified today

Baseteninference provider

Montreal · New York +3 more · hybrid · $165K–$330K/year · unknown

inferencekubernetes-opsml-platformpython-langsoftware-engineer

posted 18mo ago · verified today

Baseteninference provider

Montreal · New York +3 more · hybrid · $165K–$330K/year · unknown

inferencekubernetes-opsml-platformsoftware-engineerinference-engines

posted 11mo ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $245K–$325K · staff plus

inferenceinference-enginespython-langsoftware-engineerkubernetes-ops

posted 3mo ago · verified today

SambaNovachip vendor

Austin, Texas, United States · Austin, TX +2 more · $200K–$275K · senior

inferenceinference-enginessoftware-engineerspeculative-decodingvllm-engine

posted 3mo ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · staff plus

custom-asicgpu-kernelsnetwork-fabricrdma-verbsroce-net

posted 5mo ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $180K–$255K · senior

inferenceperformance-engineercpp-langdistributed-inferenceinference-engines

posted 9mo ago · verified today

Scale AIai startup

New York, NY · San Francisco, CA · $290K–$363K · manager

cudaeng-managerpost-trainingtraining-frameworksflash-attention

posted 11mo ago · verified today

Scale AIai startup

New York, NY · San Francisco, CA +1 more · $190K–$237K · unknown

cudaflash-attentioninference-enginesml-platformresearch-engineer

posted 18mo ago · verified today

Anyscaleai startup

Palo Alto · San Francisco · hybrid · $170K–$245K/year · unknown

distributed-inferenceinferenceinference-enginessoftware-engineervllm-engine

posted 3mo ago · verified today

Modalinference provider

New York · San Francisco · $150K–$350K/year · unknown

inferenceinference-engineskv-cache-systemsquantizationresearch-engineer

posted 10w ago · verified today

Modalinference provider

New York · San Francisco · $200K–$350K/year · senior

performance-engineercudagpu-kernelsinference-enginesnvidia

posted 4mo ago · verified today

Modalinference provider

New York · San Francisco · $180K–$250K/year · unknown

solutions-architectsglang-enginevllm-engineperformance-engineerinference

posted 6mo ago · verified today

Fireworks AIinference provider

San Mateo · hybrid · $175K–$220K/year · unknown

ml-platformpython-langsoftware-engineerinferenceinference-engines

posted 10mo ago · verified today

Fireworks AIinference provider

San Mateo · hybrid · $175K–$220K/year · unknown

gpu-kernelsperformance-engineercudagpu-genericsoftware-engineer

posted 16mo ago · verified today

xAIfrontier lab

Palo Alto, CA · $180K–$440K/year est. · unknown

cpp-langgpu-genericgpu-kernelsinferenceinference-engines

posted 23mo ago · verified today

xAIfrontier lab

Palo Alto, CA · $180K–$440K/year est. · unknown

cpp-langcudadistributed-inferenceinferenceinference-engines

posted 10w ago · verified today

xAIfrontier lab

London · London, England, United Kingdom · £107K–£262K/year est. · unknown

cpp-langinferenceinference-enginesml-platformobservability

posted 4mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $188K–$275K/year est. · staff plus

inferenceinference-enginesdistributed-inferencegpu-generickubernetes-ops

posted 4mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $152K–$204K/year est. · senior

cudadistributed-inferencego-langgpu-genericgpu-kernels

posted 11mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $139K–$204K/year est. · senior

software-engineerinferenceinference-engineskubernetes-opsgpu-generic

posted 7mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $206K–$333K/year est. · staff plus

cudaperformance-engineergpu-genericinferenceinference-engines

posted 9mo ago · verified today

OpenAIfrontier lab

San Francisco · Seattle · hybrid · $293K–$385K/year · unknown

evaluationgpu-kernelsinferencesoftware-engineercollectives

posted 4mo ago · verified today

OpenAIfrontier lab

San Francisco · $266K–$500K/year · unknown

inferenceinference-enginesperformance-engineersoftware-engineergpu-generic

posted 4mo ago · verified today

OpenAIfrontier lab

San Francisco · Seattle · hybrid · $293K–$385K/year · senior

collectivescpp-langcudadistributed-inferencegpu-generic

posted 5mo ago · verified today

OpenAIfrontier lab

San Francisco · $295K–$555K/year · unknown

gpu-genericinferenceinference-enginessoftware-engineerdistributed-inference

posted 16mo ago · verified today

OpenAIfrontier lab

San Francisco · $266K–$500K/year · unknown

cudainferenceinference-enginesnvidiasoftware-engineer

posted 19mo ago · verified today