AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1295 open roles · 85 companies · last verified today

1295 roles

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $150K–$190K · unknown

datacenter-engineercluster-datacentergo-langpython-langnetwork-fabric

posted 6w ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $180K–$255K · staff plus

cpp-langlinux-kernelnetwork-fabricrdma-verbssoftware-engineer

posted 5mo ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $250K–$330K · senior

network-engineernetwork-fabricinfiniband-opsroce-netnvlink-topology

posted 12w ago · verified today

SambaNovachip vendor

Bengaluru, India · Bengaluru, Karnataka, India · 3324K–4062K INR · unknown

python-langsolutions-architectinferencefine-tuningkubernetes-ops

posted 9w ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $210K–$280K · senior

software-engineermlir-llvmperformance-engineer

posted 9mo ago · verified today

SambaNovachip vendor

Austin, Texas, United States · Austin, TX +3 more · remote · $250K–$350K · staff plus

network-fabriccustom-asicnetwork-engineercpp-langpython-lang

posted 10w ago · verified today

Scale AIai startup

New York, NY · San Francisco, CA · $290K–$363K · manager

cudaeng-managerpost-trainingtraining-frameworksflash-attention

posted 11mo ago · verified today

Scale AIai startup

New York, NY · San Francisco, CA · $180K–$225K · mid

software-engineertraining-data-inframl-platformscheduling-orchestrationevaluation

posted 13mo ago · verified today

Scale AIai startup

New York, NY · San Francisco, CA +2 more · unknown

inferenceinference-enginessoftware-engineerkubernetes-opsnetwork-fabric

posted 31mo ago · verified today

Scale AIai startup

New York, NY · San Francisco, CA +1 more · $190K–$237K · unknown

cudaflash-attentioninference-enginesml-platformresearch-engineer

posted 18mo ago · verified today

Anyscaleai startup

San Francisco · hybrid · $200K–$240K/year · mid

reliability-sresrego-langkubernetes-opspython-lang

posted 3w ago · verified today

Anyscaleai startup

San Francisco · hybrid · $200K–$240K/year · mid

kubernetes-opsray-distributedsoftware-engineergo-langml-platform

posted 4w ago · verified today

Anyscaleai startup

Palo Alto · San Francisco · hybrid · $170K–$245K/year · unknown

distributed-inferenceinferenceinference-enginessoftware-engineervllm-engine

posted 3mo ago · verified today

Modalinference provider

New York · San Francisco · $150K–$350K/year · unknown

inferenceinference-engineskv-cache-systemsquantizationresearch-engineer

posted 10w ago · verified today

Modalinference provider

Stockholm · $175K–$250K/year · manager

cpp-langeng-managerlinux-kernelrust-langcluster-datacenter

posted 4mo ago · verified today

Modalinference provider

London · New York +2 more · onsite · $200K–$280K/year · senior

solutions-architectkubernetes-opsml-platformpython-langterraform-iac

posted 6mo ago · verified today

Modalinference provider

New York · $275K–$330K/year · manager

eng-managercpp-langml-platformrust-langlinux-kernel

posted 7mo ago · verified today

Modalinference provider

Stockholm · senior

solutions-architectkubernetes-opsterraform-iacgpu-generic

posted 8mo ago · verified today

Modalinference provider

Stockholm · $140K–$200K/year · unknown

linux-kernelsoftware-engineerml-platformscheduling-orchestrationrust-lang

posted 8mo ago · verified today

Modalinference provider

New York · San Francisco · onsite · $180K–$240K/year · mid

solutions-architectkubernetes-opsterraform-iac

posted 11mo ago · verified today

Modalinference provider

London · New York +2 more · $150K–$220K/year · unknown

solutions-architectgpu-genericsoftware-engineerml-platforminference

posted 11mo ago · verified today

Modalinference provider

New York · San Francisco · $200K–$350K/year · senior

performance-engineercudagpu-kernelsinference-enginesnvidia

posted 4mo ago · verified today

Modalinference provider

New York · San Francisco · $180K–$250K/year · unknown

solutions-architectsglang-enginevllm-engineperformance-engineerinference

posted 6mo ago · verified today

Modalinference provider

New York · San Francisco · onsite · $220K–$300K/year · senior

ml-platformsoftware-engineerlinux-kernelscheduling-orchestrationreliability-sre

posted 23mo ago · verified today

Fireworks AIinference provider

New York · San Mateo +1 more · hybrid · $200K–$260K/year · senior

inference-enginesquantizationsglang-enginevllm-enginefine-tuning

posted 3mo ago · verified today

Fireworks AIinference provider

New York · San Mateo · hybrid · $200K–$260K/year · unknown

fine-tuningpython-langsglang-enginesolutions-architectvllm-engine

posted 3mo ago · verified today