AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
985 open roles · 19 companies · last verified today
171 roles
Cerebraschip vendor
US and Canada Offices · senior
cpp-langsoftware-engineercustom-asicinference-enginestraining-frameworks
posted 9mo ago · verified today
Cerebraschip vendor
Headquarters/Sunnyvale Office · Toronto Office · remote · staff plus
cpp-langinferenceinference-enginesrocm-hipvllm-engine
posted 3w ago · verified today
Cerebraschip vendor
India Office · manager
cpp-langcustom-asicinference-enginespython-langsoftware-engineer
posted 10w ago · verified today
Cerebraschip vendor
India Office · staff plus
custom-asicdatacenter-engineerpython-langcluster-datacentercpp-lang
posted 13d ago · verified today
Baseteninference provider
New York · San Francisco · remote · $165K–$330K · senior
software-engineerfine-tuninggpu-genericml-platformpost-training
posted 7mo ago · verified today
Baseteninference provider
San Francisco · remote · $200K–$275K · senior
post-trainingresearch-engineermodel-parallelismfine-tuninggpu-generic
posted 5mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $180K–$360K · mid
gpu-genericinferenceinference-enginescpp-langnvidia
posted 29mo ago · verified today
Baseteninference provider
San Francisco · remote · $210K–$285K · staff plus
inferencepost-trainingresearch-engineerinference-engineskv-cache-systems
posted 5mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $165K–$330K · senior
gpu-genericinferenceinference-enginesnvidiaperformance-engineer
posted 7mo ago · verified today
Baseteninference provider
New York · San Francisco · remote · $165K–$330K · senior
go-langkubernetes-opsml-platformnetwork-fabricobservability
posted 11mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · senior
performance-engineergpu-genericgpu-kernelsinferenceinference-engines
posted 3d ago · verified today
SambaNovachip vendor
Austin, Texas, United States; San Jose, California, United States · Austin, TX +1 more · senior
inferenceinference-enginespython-langquantizationcustom-asic
posted 3d ago · verified today
SambaNovachip vendor
Austin, Texas, United States; San Jose, California, United States · San Jose, CA · senior
linux-kernelcpp-langcustom-asicsoftware-engineergpu-kernels
posted 3d ago · verified today
Scale AIai startup
San Francisco, CA · San Francisco, CA; New York, NY · mid
post-trainingresearch-engineertraining-frameworksgpu-genericinference-engines
posted 3d ago · verified today
Scale AIai startup
San Francisco, CA · San Francisco, CA; New York, NY · manager
ml-platformpost-trainingtraining-frameworkseng-managergpu-generic
posted 3d ago · verified today
Scale AIai startup
San Francisco, CA · San Francisco, CA; Seattle, WA; New York, NY · mid
cudaml-platformgpu-genericpytorch-distsoftware-engineer
posted 3d ago · verified today
Anyscaleai startup
Palo Alto · San Francisco · remote · $170K–$245K · mid
distributed-inferencegpu-genericinferenceinference-enginessoftware-engineer
posted 12w ago · verified today
Modalinference provider
New York · San Francisco · $150K–$350K · senior
research-engineercluster-datacenterpost-trainingreinforcement-learninggpu-generic
posted 6w ago · verified today
Modalinference provider
New York · San Francisco · $180K–$250K · mid
sglang-enginesolutions-architectvllm-enginedistributed-inferenceinference-engines
posted 5mo ago · verified today
Modalinference provider
New York · San Francisco · $200K–$350K · senior
inferenceinference-enginesperformance-engineercudagpu-generic
posted 4mo ago · verified today
Fireworks AIinference provider
San Mateo · $175K–$220K · senior
cudagpu-genericgpu-kernelsinferenceperformance-engineer
posted 15mo ago · verified today
Fireworks AIinference provider
Singapore · 180K–250K SGD · mid
software-engineerml-platformgpu-genericinferenceinference-engines
posted 8w ago · verified today
Fireworks AIinference provider
New York · San Mateo · remote · $210K–$320K · mid
gpu-generickubernetes-opsml-platformmodel-parallelismpython-lang
posted 20d ago · verified today
Fireworks AIinference provider
New York · San Mateo · $210K–$320K · mid
gpu-genericpython-langpytorch-distcpp-langtraining-frameworks
posted 6w ago · verified today
Fireworks AIinference provider
San Mateo · $175K–$220K · mid
inferencesoftware-engineerdistributed-inferenceinference-engineskubernetes-ops
posted 9mo ago · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · mid
gpu-genericml-platformpython-langcpp-langnvidia
posted 3d ago · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · mid
cpp-langcudadistributed-inferenceinferenceinference-engines
posted 3d ago · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · staff plus
fine-tuningpost-trainingpre-trainingpython-langtraining-data-infra
posted 3d ago · verified today
CoreWeaveneocloud
Bellevue, WA · Sunnyvale, CA +1 more · staff plus
inferenceinference-enginesnvidiaperformance-engineertraining-frameworks
posted 3d ago · verified today
CoreWeaveneocloud
San Francisco, CA · San Francisco, CA · mid
solutions-architectpython-langgpu-genericinferenceml-platform
posted 3d ago · verified today