AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1296 open roles · 85 companies · last verified today
94 roles
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
training-frameworksdistributed-inferencecluster-datacentercollectivesdatacenter-engineer
posted 3w ago · verified today
Amazon (AWS)hyperscaler
New York, New York, USA · Sunnyvale, California, USA · onsite · senior
kubernetes-opsml-platforminferenceinference-enginesquantization
posted 3w ago · verified today
Nebiusneocloud
Remote - United States · United States · remote · $228K–$285K · manager
eng-managerfine-tuninginferenceinference-engineskubernetes-ops
posted 4w ago · verified today
Baseteninference provider
New York · San Francisco · hybrid · $200K–$400K/year · mid
inferenceinference-enginessolutions-architectevaluationperformance-engineer
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
software-engineertraining-frameworkstrainiumpytorch-distfsdp
posted 11mo ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · Seattle, Washington, USA · onsite · mid
trainiummodel-parallelismtraining-frameworkscollectivesperformance-engineer
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · mid
software-engineerml-platformpost-trainingreinforcement-learningtraining-frameworks
posted 3mo ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · staff plus
solutions-architectcudagpu-generickubernetes-opscluster-datacenter
posted 10w ago · verified today
Amazon (AWS)hyperscaler
Santa Clara, California, USA · onsite · unknown
cudagpu-kernelstriton-langgpu-genericmodel-parallelism
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Houston, Texas, USA · onsite · senior
cudagpu-kernelsmodel-parallelismpytorch-disttraining-frameworks
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Bellevue, Washington, USA · onsite · mid
software-engineerml-platformtrainiumscheduling-orchestrationfsdp
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
cudacutlass-cutedistributed-inferenceevaluationflash-attention
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior
solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks
posted 3mo ago · verified today
Cerebraschip vendor
Canada · United States · unknown
python-langevaluationpost-trainingreinforcement-learningfine-tuning
posted 6mo ago · verified today
Cerebraschip vendor
Canada · United States · senior
amdcpp-langinferenceinference-enginespython-lang
posted 9mo ago · verified today
Baseteninference provider
San Francisco · hybrid · $200K–$275K/year · unknown
post-traininggpu-genericmodel-parallelismpytorch-distresearch-engineer
posted 5mo ago · verified today
Baseteninference provider
New York · San Francisco · hybrid · $165K–$330K/year · unknown
ml-platformsoftware-engineerfine-tuningkubernetes-opspost-training
posted 7mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · $220K–$300K · staff plus
fine-tuninginferencesoftware-engineerinference-enginespre-training
posted 4w ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · 3324K–4062K INR · unknown
python-langsolutions-architectinferencefine-tuningkubernetes-ops
posted 9w ago · verified today
Modalinference provider
Stockholm · mid
inferencesolutions-architectdistributed-inferencefine-tuninggpu-generic
posted 6mo ago · verified today
Modalinference provider
New York · San Francisco · $180K–$250K/year · unknown
solutions-architectsglang-enginevllm-engineperformance-engineerinference
posted 6mo ago · verified today
Modalinference provider
New York · San Francisco · onsite · $220K–$300K/year · senior
ml-platformsoftware-engineerlinux-kernelscheduling-orchestrationreliability-sre
posted 23mo ago · verified today
Fireworks AIinference provider
New York · San Mateo +1 more · hybrid · $200K–$260K/year · senior
inference-enginesquantizationsglang-enginevllm-enginefine-tuning
posted 3mo ago · verified today
Fireworks AIinference provider
New York · San Mateo · hybrid · $200K–$260K/year · unknown
fine-tuningpython-langsglang-enginesolutions-architectvllm-engine
posted 3mo ago · verified today
Fireworks AIinference provider
London · senior
fine-tuninginferenceinference-enginespost-trainingpython-lang
posted 14d ago · verified today
Fireworks AIinference provider
London · senior
fine-tuninginferenceinference-enginespython-langsglang-engine
posted 8w ago · verified today
Fireworks AIinference provider
Singapore · 200K–350K SGD/year · senior
fine-tuninginferenceinference-enginespython-langsglang-engine
posted 7w ago · verified today
Fireworks AIinference provider
San Mateo · hybrid · $210K–$320K/year · mid
evaluationpost-trainingfine-tuningml-platforminference
posted 10mo ago · verified today
xAIfrontier lab
Palo Alto, CA · $180K–$440K/year est. · unknown
gpu-genericpost-trainingpre-trainingpython-langresearch-engineer
posted 5mo ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA · $185K–$275K/year est. · staff plus
reinforcement-learningpost-trainingresearch-engineerfine-tuninggpu-generic
posted 5mo ago · verified today