AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
984 open roles · 19 companies · last verified today
23 roles
Amazon (AWS)hyperscaler
Cupertino, California, USA · senior
inferenceinference-enginessoftware-engineertrainiumdistributed-inference
posted 3w ago · verified today
Amazon (AWS)hyperscaler
Tel Aviv-Yafo, Tel Aviv, ISR · senior
cpp-langcustom-asicinferenceinference-enginessoftware-engineer
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · mid
ml-platformperformance-engineersoftware-engineercluster-datacentercuda
posted 5mo ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · manager
deepspeed-libeng-managerfsdpjax-pallasmegatron-lm
posted 9w ago · verified today
Amazon (AWS)hyperscaler
Tel Aviv-Yafo, Tel Aviv, ISR · senior
cpp-langcustom-asicdistributed-inferenceinferenceinference-engines
posted 3mo ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · manager
distributed-inferenceeng-managerinferenceinference-enginestrainium
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · senior
custom-asicinferenceinference-enginessoftware-engineergpu-kernels
posted 3mo ago · verified today
Cerebraschip vendor
US and Canada Offices · senior
inference-enginesamdcpp-langdistributed-inferenceinference
posted 8mo ago · verified today
Cerebraschip vendor
Headquarters/Sunnyvale Office · Toronto Office · remote · staff plus
cpp-langinferenceinference-enginesrocm-hipvllm-engine
posted 3w ago · verified today
Cerebraschip vendor
Toronto Office · remote · mid
cpp-langcustom-asicpython-langinference-enginesmlir-llvm
posted 13mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $180K–$360K · mid
cpp-langcudagpu-genericgpu-kernelsinference
posted 13mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $165K–$330K · senior
collectivescpp-langdistributed-inferencegpu-genericinference
posted 5mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · staff plus
custom-asicmoe-systemsreinforcement-learningspeculative-decodingfine-tuning
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · senior
cudagpu-kernelsperformance-engineercpp-langinference
posted 3d ago · verified today
OpenAIfrontier lab
San Francisco · Seattle · remote · $293K–$385K · mid
inferenceinference-enginessoftware-engineerdistributed-inferenceml-platform
posted 3mo ago · verified today
Nebiusneocloud
Amsterdam, Netherlands; Berlin, Germany; Israel; London, United Kingdom; Prague, Czech Republic; Remote - Europe · Israel +3 more · remote · senior
gpu-genericinferenceinference-enginesperformance-engineerpython-lang
posted 3d ago · verified today
Nebiusneocloud
Palo Alto, California, United States · San Francisco Bay Area · senior
inferenceinference-enginespython-langresearch-engineergpu-generic
posted 3d ago · verified today
Nebiusneocloud
Germany; Israel; Netherlands; Prague, Czech Republic; Remote - Europe; United Kingdom · Israel +3 more · remote · senior
fine-tuninggpu-genericinferencejax-pallaspython-lang
posted 3d ago · verified today
Nebiusneocloud
Canada · Canada; Remote - United States +1 more · remote · manager
solutions-architectcluster-datacenterdistributed-inferencefine-tuninginference
posted 3d ago · verified today
Crusoeneocloud
Sunnyvale, CA - US · senior
gpu-genericperformance-engineercluster-datacenterpython-langinfiniband-ops
posted 3w ago · verified today
Crusoeneocloud
San Francisco, CA - US · Sunnyvale, CA - US · staff plus
cluster-datacenterdatacenter-engineernvidiaperformance-engineerinference
posted 6w ago · verified today
Together AIneocloud
San Francisco · mid
fine-tuninginference-enginespython-langreinforcement-learningsglang-engine
posted 3d ago · verified today
Together AIneocloud
Remote · San Francisco, Singapore, Amsterdam · mid
cpp-langcudadistributed-inferencegpu-genericinference
posted 3d ago · verified today