AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

294 roles

inference-engines

Inferactinference provider

Singapore · onsite · $200K–$400K/year · unknown

gpu-kernelscudaperformance-engineercpp-langinference

posted 12w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

vllm-engineinferenceinference-engineskv-cache-systemspython-lang

posted 12w ago · verified today

Inferactinference provider

Singapore · onsite · 200K–400K SGD/year · unknown

cluster-datacentergo-langgpu-genericinferencekubernetes-ops

posted 12w ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

inferenceinference-enginespython-langvllm-enginesoftware-engineer

posted 3mo ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

tpuinference-enginesjax-pallasperformance-engineerxla-compiler

posted 3mo ago · verified today

Inferactinference provider

San Francisco · onsite · $200K–$400K/year · unknown

distributed-inferenceinference-enginesvllm-enginecpp-langgo-lang

posted 7mo ago · verified today

FriendliAIinference provider

Seoul · onsite · mid

amdcpp-langcudacutlass-cutegpu-kernels

posted 4mo ago · verified today

FriendliAIinference provider

Seoul · onsite · senior

gpu-kernelsinferencecpp-langinference-enginespython-lang

posted 4mo ago · verified today

FriendliAIinference provider

San Francisco · hybrid · mid

amdcpp-langcudagpu-genericgpu-kernels

posted 6mo ago · verified today

FriendliAIinference provider

San Francisco · hybrid · senior

cpp-langgpu-kernelsinferenceinference-enginespython-lang

posted 6mo ago · verified today

NexGen Cloudneocloud

London · London, England, United Kingdom, UK - Remote +1 more · remote · unknown

solutions-architectcudacudnn-libdeepspeed-libgpu-generic

posted 8w ago · verified today

Liquid AIfrontier lab

Boston · Remote +1 more · hybrid · unknown

cudagpu-genericgpu-kernelsperformance-engineercpp-lang

posted 13mo ago · verified today

Rekafrontier lab

US, UK, Singapore, Remote · remote · unknown

cpp-langcudafine-tuninggpu-genericgpu-kernels

posted 8mo ago · verified today

Periodic Labsfrontier lab

Menlo Park, CA · onsite · unknown

collectivescudacutlass-cutefsdpgpu-generic

posted 4mo ago · verified today

Amazon (AWS)hyperscaler

Palo Alto, California, USA · Seattle, Washington, USA · onsite · mid

ml-platformsoftware-engineerinference-enginesscheduling-orchestrationinference

posted 14d ago · verified today

Nscaleneocloud

London · UK · senior

evaluationfine-tuninggpu-genericinferenceinference-engines

posted 5mo ago · verified today

Anthropicfrontier lab

New York City, NY · San Francisco, CA +1 more · $405K–$625K · manager

eng-managerinferencescheduling-orchestrationinference-enginescluster-datacenter

posted 15d ago · verified today