AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1296 open roles · 85 companies · last verified today
187 roles
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid
distributed-inferenceinferencejax-pallastpuxla-compiler
posted 8mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown
amdnvidiagpu-kernelscollectivescpp-lang
posted 6w ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior
cpp-langcudagpu-kernelsperformance-engineercollectives
posted 7mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior
inference-enginesgpu-genericinferenceperformance-engineerdistributed-inference
posted 7mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid
cudagpu-kernelsinference-enginescpp-langrocm-hip
posted 7w ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · junior
cpp-langinferenceinference-enginespython-langsglang-engine
posted 8mo ago · verified today
Inferactinference provider
San Francisco · onsite · manager
eng-managerinferenceinference-enginesvllm-enginedistributed-inference
posted 6w ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
amdgpu-kernelsinferenceperformance-engineerrocm-hip
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
tpuinferenceinference-enginesjax-pallasperformance-engineer
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
amdgpu-kernelsinferenceinference-enginesperformance-engineer
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
software-engineerdistributed-inferenceinference-enginesinfiniband-opsnvlink-topology
posted 12w ago · verified today
Inferactinference provider
Singapore · onsite · $200K–$400K/year · unknown
gpu-kernelscudaperformance-engineercpp-langinference
posted 12w ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
tpuinference-enginesjax-pallasperformance-engineerxla-compiler
posted 3mo ago · verified today
Inferactinference provider
Remote · remote · unknown
gpu-kernelsinferenceinference-enginesvllm-enginedistributed-inference
posted 7mo ago · verified today
FriendliAIinference provider
Seoul · onsite · mid
amdcpp-langcudacutlass-cutegpu-kernels
posted 4mo ago · verified today
FriendliAIinference provider
Seoul · onsite · senior
gpu-kernelsinferencecpp-langinference-enginespython-lang
posted 4mo ago · verified today
FriendliAIinference provider
San Francisco · hybrid · mid
amdcpp-langcudagpu-genericgpu-kernels
posted 6mo ago · verified today
FriendliAIinference provider
San Francisco · hybrid · mid
solutions-architectgpu-kernelsinference-engineskubernetes-opsobservability
posted 6mo ago · verified today
FriendliAIinference provider
San Francisco · hybrid · senior
cpp-langgpu-kernelsinferenceinference-enginespython-lang
posted 6mo ago · verified today
Prime Intellectneocloud
Remote · San Francisco · unknown
cudapytorch-distreinforcement-learningresearch-engineertriton-lang
posted 10w ago · verified today
Prime Intellectneocloud
Remote · San Francisco · unknown
cudadeepspeed-libfsdpgpu-kernelsmodel-parallelism
posted 10w ago · verified today
Liquid AIfrontier lab
Boston · Remote +1 more · hybrid · unknown
cudagpu-genericgpu-kernelsperformance-engineercpp-lang
posted 13mo ago · verified today
Rekafrontier lab
US, UK, Singapore, Remote · remote · unknown
cpp-langcudafine-tuninggpu-genericgpu-kernels
posted 8mo ago · verified today
Periodic Labsfrontier lab
Menlo Park, CA · onsite · unknown
collectivescudacutlass-cutefsdpgpu-generic
posted 4mo ago · verified today
Thinking Machines Labfrontier lab
San Francisco · hybrid · unknown
cudagpu-kernelstriton-langresearch-engineercutlass-cute
posted 6w ago · verified today
Thinking Machines Labfrontier lab
San Francisco · hybrid · unknown
distributed-inferenceinference-enginesresearch-engineergpu-genericinference
posted 6w ago · verified today
Thinking Machines Labfrontier lab
San Francisco · hybrid · unknown
gpu-genericresearch-engineergpu-kernelsmodel-parallelismquantization
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Vancouver, British Columbia, CAN · onsite · senior
ml-platformsoftware-engineerevaluationinferencekubernetes-ops
posted 13d ago · verified today
Amazon (AWS)hyperscaler
Toronto, Ontario, CAN · onsite · unknown
gpu-kernelsperformance-engineertrainium
posted 14d ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
gpu-kernelsperformance-engineertrainiumcudanvidia
posted 15d ago · verified today