AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1296 open roles · 85 companies · last verified today
20 roles
Inferactinference provider
San Francisco · onsite · junior
amdcudagpu-kernelsinferenceinference-engines
posted 1d ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown
gpu-genericpost-traininggpu-kernelsmodel-parallelismtraining-frameworks
posted 7mo ago · verified today
FriendliAIinference provider
Seoul · onsite · mid
amdcpp-langcudacutlass-cutegpu-kernels
posted 4mo ago · verified today
FriendliAIinference provider
San Francisco · hybrid · mid
amdcpp-langcudagpu-genericgpu-kernels
posted 6mo ago · verified today
Liquid AIfrontier lab
Boston · Remote +1 more · hybrid · unknown
cudagpu-genericgpu-kernelsperformance-engineercpp-lang
posted 13mo ago · verified today
Periodic Labsfrontier lab
Menlo Park, CA · onsite · unknown
collectivescudacutlass-cutefsdpgpu-generic
posted 4mo ago · verified today
Thinking Machines Labfrontier lab
San Francisco · hybrid · unknown
cudagpu-kernelstriton-langresearch-engineercutlass-cute
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · mid
distributed-inferencegpu-kernelsinferenceinference-enginessoftware-engineer
posted 16d ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer
posted 20d ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
trainiumdistributed-inferenceinferencesoftware-engineercpp-lang
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
trainiumcpp-langdistributed-inferenceinferenceinference-engines
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
cudacutlass-cutedistributed-inferenceevaluationflash-attention
posted 4w ago · verified today
Perplexityai startup
London · unknown
cudagpu-kernelsinferenceinference-enginesgpu-generic
posted 5mo ago · verified today
Perplexityai startup
New York City · Palo Alto +1 more · $220K–$485K/year · mid
cudagpu-kernelsinferenceinference-enginescutlass-cute
posted 5mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · hybrid · $180K–$360K/year · unknown
cudagpu-kernelscpp-langgpu-genericinference
posted 14mo ago · verified today
xAIfrontier lab
Palo Alto, CA · $180K–$440K/year est. · unknown
cudacutlass-cutegpu-kernelsperformance-engineersoftware-engineer
posted 6mo ago · verified today
Anthropicfrontier lab
New York City, NY · San Francisco, CA +1 more · $280K–$850K · unknown
cudagpu-genericgpu-kernelsperformance-engineercollectives
posted 11mo ago · verified today
Nebiusneocloud
Amsterdam, Netherlands · Berlin, Germany +6 more · remote · senior
gpu-genericinferenceinference-enginesml-platformperformance-engineer
posted 6mo ago · verified today
Nebiusneocloud
Amsterdam, Netherlands · Israel +3 more · senior
cudagpu-genericgpu-kernelsinference-enginesinference
posted 7mo ago · verified today
Together AIneocloud
San Francisco · $200K–$290K/year est. · unknown
research-engineerfine-tuninggo-langinferenceinference-engines
posted 10w ago · verified today