AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

311 roles

performance-engineer

Coherefrontier lab

Montreal · New York +2 more · remote · senior

cpp-langcudagpu-genericgpu-kernelsinference

posted 10mo ago · verified today

Coherefrontier lab

London · Montreal +3 more · hybrid · unknown

cudagpu-kernelsperformance-engineertriton-langpre-training

posted 19mo ago · verified today

Reflection AIfrontier lab

London · New York, NY +1 more · onsite · unknown

pre-trainingmodel-parallelismsoftware-engineertraining-frameworkscollectives

posted 5mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

gpu-genericpost-traininggpu-kernelsmodel-parallelismtraining-frameworks

posted 7mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

distributed-inferenceinferenceinference-enginesperformance-engineernvidia

posted 4mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid

distributed-inferenceinferencejax-pallastpuxla-compiler

posted 8mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown

amdnvidiagpu-kernelscollectivescpp-lang

posted 6w ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior

cpp-langcudagpu-kernelsperformance-engineercollectives

posted 7mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior

inference-enginesgpu-genericinferenceperformance-engineerdistributed-inference

posted 7mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · mid

cudagpu-kernelsinference-enginescpp-langrocm-hip

posted 7w ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior

cluster-datacenterscheduling-orchestrationgpu-generickubernetes-opsnetwork-fabric

posted 7mo ago · verified today

RadixArkai startup

Palo Alto, CA · Palo Alto Office · junior

cpp-langinferenceinference-enginespython-langsglang-engine

posted 8mo ago · verified today

falinference provider

Remote - Global · unknown

python-langgpu-genericcudanvidiareliability-sre

posted 6mo ago · verified today

Parasailinference provider

San Mateo · senior

inferencereliability-sresreobservabilitykubernetes-ops

posted 7w ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

amdgpu-kernelsinferenceperformance-engineerrocm-hip

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

tpuinferenceinference-enginesjax-pallasperformance-engineer

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

amdgpu-kernelsinferenceinference-enginesperformance-engineer

posted 11w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

software-engineerdistributed-inferenceinference-enginesinfiniband-opsnvlink-topology

posted 12w ago · verified today

Inferactinference provider

Singapore · onsite · $200K–$400K/year · unknown

gpu-kernelscudaperformance-engineercpp-langinference

posted 12w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

vllm-engineinferenceinference-engineskv-cache-systemspython-lang

posted 12w ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

inferenceinference-enginespython-langvllm-enginesoftware-engineer

posted 3mo ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

tpuinference-enginesjax-pallasperformance-engineerxla-compiler

posted 3mo ago · verified today

FriendliAIinference provider

Seoul · onsite · mid

amdcpp-langcudacutlass-cutegpu-kernels

posted 4mo ago · verified today

FriendliAIinference provider

Seoul · onsite · senior

gpu-kernelsinferencecpp-langinference-enginespython-lang

posted 4mo ago · verified today

FriendliAIinference provider

San Francisco · hybrid · mid

amdcpp-langcudagpu-genericgpu-kernels

posted 6mo ago · verified today

FriendliAIinference provider

San Francisco · hybrid · senior

cpp-langgpu-kernelsinferenceinference-enginespython-lang

posted 6mo ago · verified today