AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
984 open roles · 19 companies · last verified today
136 roles
Amazon (AWS)hyperscaler
Cupertino, California, USA · manager
collectivescpp-langnccl-libnetwork-fabricnvshmem-lib
posted 20d ago · verified today
OpenAIfrontier lab
San Francisco · remote · $302K–$445K · manager
cluster-datacentercustom-asicdatacenter-engineereng-managergpu-generic
posted 3d ago · verified today
Perplexityai startup
Palo Alto · San Francisco · $220K–$485K · mid
post-trainingpython-langmegatron-lmpytorch-distreinforcement-learning
posted 4mo ago · verified today
Perplexityai startup
New York City · Palo Alto +1 more · $220K–$485K · senior
cudacutlass-cutegpu-kernelsinferenceinference-engines
posted 4mo ago · verified today
Cerebraschip vendor
Headquarters/Sunnyvale Office · Toronto Office · remote · staff plus
cpp-langinferenceinference-enginesrocm-hipvllm-engine
posted 3w ago · verified today
Cerebraschip vendor
Headquarters/Sunnyvale Office · senior
inferenceinference-enginesperformance-engineercudavllm-engine
posted 4mo ago · verified today
Cerebraschip vendor
US and Canada Offices · senior
inference-enginesamdcpp-langdistributed-inferenceinference
posted 8mo ago · verified today
Cerebraschip vendor
US and Canada Offices · remote · junior
cpp-langcustom-asicgpu-kernelsperformance-engineerpython-lang
posted 4w ago · verified today
Cerebraschip vendor
Headquarters/Sunnyvale Office · staff plus
post-trainingcustom-asicreinforcement-learningpre-trainingpython-lang
posted 8mo ago · verified today
Cerebraschip vendor
US and Canada Offices · mid
software-engineercpp-langcustom-asicgpu-kernelspython-lang
posted 5mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $180K–$360K · mid
cpp-langcudagpu-genericgpu-kernelsinference
posted 13mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $180K–$360K · mid
gpu-genericinferenceinference-enginescpp-langnvidia
posted 29mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $165K–$330K · senior
gpu-genericinferenceinference-enginesnvidiaperformance-engineer
posted 7mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $180K–$360K · mid
inferenceinference-enginesml-platformsoftware-engineertensorrt-stack
posted 10mo ago · verified today
SambaNovachip vendor
Remote - US · remote · senior
inferenceinference-enginespython-langsolutions-architectfine-tuning
posted 3d ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · staff plus
inferenceinference-enginesperformance-engineercustom-asicquantization
posted 3d ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · senior
performance-engineergpu-genericgpu-kernelsinferenceinference-engines
posted 3d ago · verified today
SambaNovachip vendor
Bengaluru, India · Bengaluru, Karnataka, India · senior
inferenceinference-engineskubernetes-opspython-langsglang-engine
posted 3d ago · verified today
Scale AIai startup
San Francisco, CA · San Francisco, CA; New York, NY · manager
ml-platformpost-trainingtraining-frameworkseng-managergpu-generic
posted 3d ago · verified today
Scale AIai startup
San Francisco, CA · San Francisco, CA; New York, NY · mid
post-trainingresearch-engineertraining-frameworksgpu-genericinference-engines
posted 3d ago · verified today
Scale AIai startup
San Francisco, CA · San Francisco, CA; Seattle, WA; New York, NY · mid
cudaml-platformgpu-genericpytorch-distsoftware-engineer
posted 3d ago · verified today
Anyscaleai startup
Palo Alto · San Francisco · remote · $170K–$245K · mid
distributed-inferencegpu-genericinferenceinference-enginessoftware-engineer
posted 12w ago · verified today
Modalinference provider
Stockholm · mid
gpu-genericsolutions-architectinferenceinference-enginespython-lang
posted 5mo ago · verified today
Modalinference provider
New York · San Francisco · $200K–$350K · senior
inferenceinference-enginesperformance-engineercudagpu-generic
posted 4mo ago · verified today
Fireworks AIinference provider
Singapore · 200K–350K SGD · senior
solutions-architectfine-tuninggpu-genericinferenceinference-engines
posted 4w ago · verified today
Fireworks AIinference provider
New York · San Mateo · $210K–$320K · mid
gpu-genericpython-langpytorch-distcpp-langtraining-frameworks
posted 6w ago · verified today
Fireworks AIinference provider
San Mateo · $175K–$220K · senior
cudagpu-genericgpu-kernelsinferenceperformance-engineer
posted 15mo ago · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · senior
cpp-langgpu-genericinferenceinference-enginessoftware-engineer
posted 3d ago · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · mid
cpp-langcudadistributed-inferenceinferenceinference-engines
posted 3d ago · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · senior
cpp-langcudacutlass-cutegpu-kernelssoftware-engineer
posted 3d ago · verified today