AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1293 open roles · 85 companies · last verified today
382 roles
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · mid
distributed-inferencegpu-kernelsinferenceinference-enginessoftware-engineer
posted 17d ago · verified today
SambaNovachip vendor
Stockholm, Sweden · staff plus
cpp-langsoftware-engineerinferencelinux-kernelnetwork-fabric
posted 19d ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
ml-platformsoftware-engineerinferencejax-pallaspytorch-dist
posted 20d ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer
posted 3w ago · verified today
Nebiusneocloud
Abu Dhabi, UAE · Middle East +1 more · remote · senior
solutions-architectgpu-generickubernetes-opsml-platforminference
posted 9w ago · verified today
Perplexityai startup
Palo Alto · San Francisco · $220K–$405K/year · senior
software-engineerml-platformpython-langrust-langinference
posted 3w ago · verified today
Amazon (AWS)hyperscaler
New York, New York, USA · Sunnyvale, California, USA · onsite · senior
kubernetes-opsml-platforminferenceinference-enginesquantization
posted 3w ago · verified today
Nebiusneocloud
South Korea · staff plus
solutions-architectcudagpu-genericansiblekubernetes-ops
posted 3w ago · verified today
Nebiusneocloud
Japan · staff plus
solutions-architectcudagpu-generickubernetes-opsml-platform
posted 3w ago · verified today
Crusoeneocloud
Denver, CO - US · onsite · staff plus
inferenceinference-enginesperformance-engineersglang-enginevllm-engine
posted 3w ago · verified today
Anthropicfrontier lab
New York City, NY · San Francisco, CA +1 more · $320K–$485K · staff plus
inferenceinference-enginesscheduling-orchestrationsoftware-engineerdistributed-inference
posted 3w ago · verified today
OpenAIfrontier lab
San Francisco · hybrid · $266K–$445K/year · unknown
software-engineerinference-enginesinferencekv-cache-systemsscheduling-orchestration
posted 3w ago · verified today
Nebiusneocloud
New York City, New York, United States · Remote - United States +1 more · remote · $200K–$245K · senior
gpu-genericnvidiasolutions-architectinferenceinference-engines
posted 7mo ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · 10300K–12500K INR · staff plus
inference-enginessoftware-engineergo-langinferencekubernetes-ops
posted 4w ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · 10300K–12500K INR · manager
cluster-datacenterobservabilitysoftware-engineerml-platformreliability-sre
posted 3w ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · 8900K–10800K INR · staff plus
go-langinferenceinference-engineskubernetes-opspython-lang
posted 3w ago · verified today
Together AIneocloud
Amsterdam · London · staff plus
gpu-genericinferencekubernetes-opsscheduling-orchestrationsoftware-engineer
posted 3w ago · verified today
OpenAIfrontier lab
San Francisco · hybrid · $295K–$380K/year · unknown
software-engineertrainiumgpu-kernelsinferenceinference-engines
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Tel Aviv-Yafo, Tel Aviv, ISR · onsite · junior
python-langsoftware-engineerml-platformreliability-sreobservability
posted 4w ago · verified today
Together AIneocloud
India · Remote · unknown
go-langgpu-generickubernetes-opspython-langrust-lang
posted 4w ago · verified today
SambaNovachip vendor
Austin, Texas, United States · Austin, TX +2 more · $210K–$280K · staff plus
inferenceml-platformobservabilityreliability-sresre
posted 9mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · San Francisco, CA +1 more · $198K–$264K/year est. · staff plus
amdcluster-datacenterdeepspeed-libdistributed-inferencefsdp
posted 4w ago · verified today
Nebiusneocloud
Remote - United States · United States · remote · $228K–$285K · manager
eng-managerfine-tuninginferenceinference-engineskubernetes-ops
posted 4w ago · verified today
Baseteninference provider
New York · San Francisco · hybrid · $200K–$400K/year · mid
inferencesolutions-architectevaluationinference-enginespost-training
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · manager
eng-managernvidiacluster-datacentercpp-langml-platform
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · Seattle, Washington, USA · onsite · senior
cpp-langdistributed-inferencenetwork-fabricsoftware-engineercollectives
posted 8w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · manager
cudaeng-managerml-platformperformance-engineercluster-datacenter
posted 3mo ago · verified today
Amazon (AWS)hyperscaler
New York, New York, USA · Seattle, Washington, USA · onsite · senior
software-engineerinferenceinference-enginesdistributed-inferencetensorrt-stack
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Austin, Texas, USA · Cupertino, California, USA · onsite · mid
custom-asicrtl-designdatacenter-engineerinference
posted 4mo ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
distributed-inferencegpu-kernelsinferencepython-langsoftware-engineer
posted 10mo ago · verified today