AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1293 open roles · 85 companies · last verified today

382 roles

inference

SambaNovachip vendor

Stockholm, Sweden · staff plus

cpp-langsoftware-engineerinferencelinux-kernelnetwork-fabric

posted 19d ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · senior

ml-platformsoftware-engineerinferencejax-pallaspytorch-dist

posted 20d ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer

posted 3w ago · verified today

Nebiusneocloud

Abu Dhabi, UAE · Middle East +1 more · remote · senior

solutions-architectgpu-generickubernetes-opsml-platforminference

posted 9w ago · verified today

Perplexityai startup

Palo Alto · San Francisco · $220K–$405K/year · senior

software-engineerml-platformpython-langrust-langinference

posted 3w ago · verified today

Amazon (AWS)hyperscaler

New York, New York, USA · Sunnyvale, California, USA · onsite · senior

kubernetes-opsml-platforminferenceinference-enginesquantization

posted 3w ago · verified today

Crusoeneocloud

Denver, CO - US · onsite · staff plus

inferenceinference-enginesperformance-engineersglang-enginevllm-engine

posted 3w ago · verified today

Anthropicfrontier lab

New York City, NY · San Francisco, CA +1 more · $320K–$485K · staff plus

inferenceinference-enginesscheduling-orchestrationsoftware-engineerdistributed-inference

posted 3w ago · verified today

OpenAIfrontier lab

San Francisco · hybrid · $266K–$445K/year · unknown

software-engineerinference-enginesinferencekv-cache-systemsscheduling-orchestration

posted 3w ago · verified today

Nebiusneocloud

New York City, New York, United States · Remote - United States +1 more · remote · $200K–$245K · senior

gpu-genericnvidiasolutions-architectinferenceinference-engines

posted 7mo ago · verified today

SambaNovachip vendor

Bengaluru, India · Bengaluru, Karnataka, India · 10300K–12500K INR · staff plus

inference-enginessoftware-engineergo-langinferencekubernetes-ops

posted 4w ago · verified today

SambaNovachip vendor

Bengaluru, India · Bengaluru, Karnataka, India · 10300K–12500K INR · manager

cluster-datacenterobservabilitysoftware-engineerml-platformreliability-sre

posted 3w ago · verified today

SambaNovachip vendor

Bengaluru, India · Bengaluru, Karnataka, India · 8900K–10800K INR · staff plus

go-langinferenceinference-engineskubernetes-opspython-lang

posted 3w ago · verified today

OpenAIfrontier lab

San Francisco · hybrid · $295K–$380K/year · unknown

software-engineertrainiumgpu-kernelsinferenceinference-engines

posted 4w ago · verified today

SambaNovachip vendor

Austin, Texas, United States · Austin, TX +2 more · $210K–$280K · staff plus

inferenceml-platformobservabilityreliability-sresre

posted 9mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · San Francisco, CA +1 more · $198K–$264K/year est. · staff plus

amdcluster-datacenterdeepspeed-libdistributed-inferencefsdp

posted 4w ago · verified today

Nebiusneocloud

Remote - United States · United States · remote · $228K–$285K · manager

eng-managerfine-tuninginferenceinference-engineskubernetes-ops

posted 4w ago · verified today

Baseteninference provider

New York · San Francisco · hybrid · $200K–$400K/year · mid

inferencesolutions-architectevaluationinference-enginespost-training

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · manager

eng-managernvidiacluster-datacentercpp-langml-platform

posted 5w ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · manager

cudaeng-managerml-platformperformance-engineercluster-datacenter

posted 3mo ago · verified today

Amazon (AWS)hyperscaler

New York, New York, USA · Seattle, Washington, USA · onsite · senior

software-engineerinferenceinference-enginesdistributed-inferencetensorrt-stack

posted 6w ago · verified today

Amazon (AWS)hyperscaler

Austin, Texas, USA · Cupertino, California, USA · onsite · mid

custom-asicrtl-designdatacenter-engineerinference

posted 4mo ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · senior

distributed-inferencegpu-kernelsinferencepython-langsoftware-engineer

posted 10mo ago · verified today