AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1294 open roles · 85 companies · last verified today
294 roles
SambaNovachip vendor
Stockholm, Sweden · staff plus
cpp-langsoftware-engineerinferencelinux-kernelnetwork-fabric
posted 19d ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer
posted 3w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
training-frameworksdistributed-inferencecluster-datacentercollectivesdatacenter-engineer
posted 3w ago · verified today
Amazon (AWS)hyperscaler
New York, New York, USA · Sunnyvale, California, USA · onsite · senior
kubernetes-opsml-platforminferenceinference-enginesquantization
posted 3w ago · verified today
Crusoeneocloud
Denver, CO - US · onsite · staff plus
inferenceinference-enginesperformance-engineersglang-enginevllm-engine
posted 3w ago · verified today
Anthropicfrontier lab
New York City, NY · San Francisco, CA +1 more · $320K–$485K · staff plus
inferenceinference-enginesscheduling-orchestrationsoftware-engineerdistributed-inference
posted 3w ago · verified today
OpenAIfrontier lab
San Francisco · hybrid · $266K–$445K/year · unknown
software-engineerinference-enginesinferencekv-cache-systemsscheduling-orchestration
posted 3w ago · verified today
Nebiusneocloud
New York City, New York, United States · Remote - United States +1 more · remote · $200K–$245K · senior
gpu-genericnvidiasolutions-architectinferenceinference-engines
posted 7mo ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · 10300K–12500K INR · staff plus
inference-enginessoftware-engineergo-langinferencekubernetes-ops
posted 4w ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · 10300K–12500K INR · manager
cluster-datacenterobservabilitysoftware-engineerml-platformreliability-sre
posted 3w ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · 8900K–10800K INR · staff plus
go-langinferenceinference-engineskubernetes-opspython-lang
posted 3w ago · verified today
OpenAIfrontier lab
San Francisco · hybrid · $295K–$380K/year · unknown
software-engineertrainiumgpu-kernelsinferenceinference-engines
posted 4w ago · verified today
SambaNovachip vendor
Austin, Texas, United States · Austin, TX +2 more · $210K–$280K · staff plus
inferenceml-platformobservabilityreliability-sresre
posted 9mo ago · verified today
Nebiusneocloud
Remote - United States · United States · remote · $228K–$285K · manager
eng-managerfine-tuninginferenceinference-engineskubernetes-ops
posted 4w ago · verified today
Baseteninference provider
New York · San Francisco · hybrid · $200K–$400K/year · mid
inferencesolutions-architectevaluationinference-enginespost-training
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · mid
cpp-langdistributed-inferencenetwork-fabricsoftware-engineercollectives
posted 7w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · Seattle, Washington, USA · onsite · senior
cpp-langdistributed-inferencenetwork-fabricsoftware-engineercollectives
posted 8w ago · verified today
Amazon (AWS)hyperscaler
New York, New York, USA · Seattle, Washington, USA · onsite · senior
software-engineerinferenceinference-enginesdistributed-inferencetensorrt-stack
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
custom-asicdistributed-inferencesoftware-engineertrainiuminference
posted 10mo ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
trainiumdistributed-inferenceinferencesoftware-engineercpp-lang
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
trainiumcpp-langdistributed-inferenceinferenceinference-engines
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · manager
eng-managertrainiuminferenceinference-enginesdistributed-inference
posted 8w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
software-engineertrainiuminference-enginesvllm-enginedistributed-inference
posted 7w ago · verified today
Amazon (AWS)hyperscaler
Tel Aviv-Yafo, Tel Aviv, ISR · onsite · senior
inferenceinference-enginesvllm-enginecudagpu-generic
posted 10w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · mid
gpu-kernelsinferenceinference-enginesdistributed-inferenceml-platform
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Tel Aviv-Yafo, Tel Aviv, ISR · onsite · mid
cpp-langsoftware-engineercustom-asicdistributed-inferencegpu-kernels
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Palo Alto, California, USA · onsite · senior
software-engineergpu-genericinferenceinference-engines
posted 4mo ago · verified today
Amazon (AWS)hyperscaler
Santa Clara, California, USA · onsite · unknown
cudagpu-kernelstriton-langgpu-genericmodel-parallelism
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Austin, Texas, USA · Cupertino, California, USA · onsite · senior
custom-asicgpu-genericcluster-datacentercollectivessoftware-engineer
posted 5mo ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
cudacutlass-cutedistributed-inferenceevaluationflash-attention
posted 4w ago · verified today