AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1297 open roles · 85 companies · last verified today
57 roles
Modalinference provider
San Francisco · onsite · $300K–$350K/year · manager
eng-managergpu-genericinferenceinference-enginesml-platform
posted today · verified today
Nscaleneocloud
Houston · New York +2 more · $220K–$293K · staff plus
inferenceinference-engineskv-cache-systemspost-trainingpython-lang
posted 1d ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
inference-enginessoftware-engineersglang-enginetrainiumvllm-engine
posted 4d ago · verified today
DoorDashenterprise
San Francisco · San Francisco, CA +2 more · senior
fine-tuninggpu-genericinferenceinference-enginesml-platform
posted 10w ago · verified today
Inferactinference provider
San Francisco · onsite · junior
amdcudagpu-kernelsinferenceinference-engines
posted 1d ago · verified today
DigitalOceanneocloud
Bay Area Metro · Seattle · $220K–$239K/year est. · staff plus
amdcudainferencekubernetes-opsnvidia
posted 4mo ago · verified today
DigitalOceanneocloud
Bay Area Metro · San Francisco · $220K–$239K/year est. · staff plus
amdcudadistributed-inferencego-langinference
posted 4mo ago · verified today
DigitalOceanneocloud
DigitalOcean Seattle Office · Seattle · junior
go-langgpu-genericinferencepython-langsoftware-engineer
posted 3w ago · verified today
DigitalOceanneocloud
Bangalore Metro · Bengaluru · senior
distributed-inferenceinferenceinference-engineskv-cache-systemsperformance-engineer
posted 6w ago · verified today
DeepInfrainference provider
Palo Alto, United States · onsite · unknown
inferenceinference-enginesnvidiapython-langsolutions-architect
posted 4w ago · verified today
Fireworks AIinference provider
New York · San Mateo · hybrid · $200K–$230K/year · unknown
scheduling-orchestrationcpp-langnetwork-fabricpython-langstorage-checkpointing
posted 7d ago · verified today
Together AIneocloud
San Francisco · $270K–$300K/year est. · senior
distributed-inferencefine-tuninginferenceinference-engineskv-cache-systems
posted 4mo ago · verified today
Lila Sciencesai startup
Alewife, Cambridge, MA · Cambridge, MA USA · $192K–$272K · staff plus
inferenceinference-engineskubernetes-opsml-platformnvidia
posted 7w ago · verified today
ElevenLabsai startup
Bulgaria · Poland +2 more · remote · unknown
cudagpu-genericgpu-kernelsinferenceinference-engines
posted 19d ago · verified today
Coherefrontier lab
Montreal · New York +2 more · remote · senior
cpp-langcudagpu-genericgpu-kernelsinference
posted 10mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid
gpu-genericinferenceinference-enginesreliability-sreamd
posted 4mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown
distributed-inferenceinferenceinference-enginesperformance-engineernvidia
posted 4mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior
inference-enginesgpu-genericinferenceperformance-engineerdistributed-inference
posted 7mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid
cudagpu-kernelsinference-enginescpp-langrocm-hip
posted 7w ago · verified today
Inferactinference provider
Remote · remote · unknown
inferenceinference-enginesvllm-enginepython-langsoftware-engineer
posted 12d ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
amdgpu-kernelsinferenceperformance-engineerrocm-hip
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
tpuinferenceinference-enginesjax-pallasperformance-engineer
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
amdgpu-kernelsinferenceinference-enginesperformance-engineer
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
vllm-engineinferenceinference-engineskv-cache-systemspython-lang
posted 12w ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
inferenceinference-enginespython-langvllm-enginesoftware-engineer
posted 3mo ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
tpuinference-enginesjax-pallasperformance-engineerxla-compiler
posted 3mo ago · verified today
Thinking Machines Labfrontier lab
San Francisco · hybrid · unknown
reinforcement-learningresearch-engineerdistributed-inferenceinferenceinference-engines
posted 3w ago · verified today
Nscaleneocloud
London · UK · senior
evaluationfine-tuninggpu-genericinferenceinference-engines
posted 5mo ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer
posted 20d ago · verified today
OpenAIfrontier lab
San Francisco · hybrid · $266K–$445K/year · unknown
software-engineerinference-enginesinferencekv-cache-systemsscheduling-orchestration
posted 3w ago · verified today