AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
984 open roles · 19 companies · last verified today
10 roles
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · senior
cudadistributed-inferencegpu-kernelsinferenceinference-engines
posted 3d ago · verified today
Perplexityai startup
New York City · Palo Alto +1 more · $220K–$485K · senior
cudacutlass-cutegpu-kernelsinferenceinference-engines
posted 4mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $180K–$360K · mid
cpp-langcudagpu-genericgpu-kernelsinference
posted 13mo ago · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · senior
cpp-langcudacutlass-cutegpu-kernelssoftware-engineer
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · senior
cudagpu-kernelsperformance-engineercpp-langinference
posted 3d ago · verified today
Anthropicfrontier lab
San Francisco, CA · San Francisco, CA | New York City, NY | Seattle, WA · senior
cudaperformance-engineergpu-genericgpu-kernelscollectives
posted 3d ago · verified today
Nebiusneocloud
Amsterdam, Netherlands; Berlin, Germany; Israel; London, United Kingdom; Prague, Czech Republic; Remote - Europe · Israel +3 more · remote · senior
gpu-genericinferenceinference-enginesperformance-engineerpython-lang
posted 3d ago · verified today
Nebiusneocloud
Amsterdam, Netherlands · Israel +3 more · senior
cpp-langcudagpu-genericgpu-kernelsinference
posted 3d ago · verified today
Nebiusneocloud
Netherlands · Remote +2 more · remote · mid
cpp-langgpu-genericgpu-kernelssoftware-engineerinference
posted 3d ago · verified today
Together AIneocloud
San Francisco · mid
fine-tuninginference-enginespython-langreinforcement-learningsglang-engine
posted 3d ago · verified today