AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

311 roles

performance-engineer

NexGen Cloudneocloud

UK - Remote · remote · senior

cudanvidiacluster-datacenternetwork-fabriccudnn-lib

posted 4mo ago · verified today

San Francisco Compute Companyneocloud

Remote · San Francisco, CA · hybrid · $220K–$300K/year · senior

cluster-datacenterdatacenter-engineergpu-genericnetwork-fabricinfiniband-ops

posted 2d ago · verified today

Prime Intellectneocloud

Remote · San Francisco · unknown

cudadeepspeed-libfsdpgpu-kernelsmodel-parallelism

posted 10w ago · verified today

Radiantneocloud

Gloucestershire · London · hybrid · senior

reliability-sresrecluster-datacentergpu-genericnvidia

posted 3mo ago · verified today

Radiantneocloud

Gloucestershire · hybrid · senior

ansibleobservabilitypython-langreliability-sresre

posted 5mo ago · verified today

Liquid AIfrontier lab

Boston · Remote +1 more · hybrid · unknown

cudagpu-genericgpu-kernelsperformance-engineercpp-lang

posted 13mo ago · verified today

Rekafrontier lab

US, UK, Singapore, Remote · remote · unknown

cpp-langcudafine-tuninggpu-genericgpu-kernels

posted 8mo ago · verified today

Periodic Labsfrontier lab

Menlo Park, CA · onsite · unknown

collectivescudacutlass-cutefsdpgpu-generic

posted 4mo ago · verified today

Amazon (AWS)hyperscaler

Vancouver, British Columbia, CAN · onsite · senior

ml-platformsoftware-engineerevaluationinferencekubernetes-ops

posted 14d ago · verified today

Amazon (AWS)hyperscaler

Toronto, Ontario, CAN · onsite · unknown

gpu-kernelsperformance-engineertrainium

posted 15d ago · verified today

Anthropicfrontier lab

London, UK · £325K–£390K · staff plus

ebpfobservabilitygpu-genericsoftware-engineertpu

posted 14d ago · verified today

Nscaleneocloud

London · UK · senior

evaluationfine-tuninggpu-genericinferenceinference-engines

posted 5mo ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · onsite · senior

gpu-kernelsperformance-engineertrainiumcudanvidia

posted 16d ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · manager

eng-managerml-platformperformance-engineercluster-datacentercuda

posted 17d ago · verified today

CoreWeaveneocloud

Bellevue, WA · Livingston, NJ +4 more · $182K–$242K/year est. · senior

gpu-generickubernetes-opssoftware-engineergo-langlinux-kernel

posted 19d ago · verified today

CoreWeaveneocloud

Manhattan, NY · New York, NY · $153K–$204K/year est. · senior

object-storagesoftware-engineerstorage-checkpointinggo-langobservability

posted 19d ago · verified today

SambaNovachip vendor

Stockholm, Sweden · staff plus

cpp-langsoftware-engineerinferencelinux-kernelnetwork-fabric

posted 19d ago · verified today

Amazon (AWS)hyperscaler

Austin, Texas, USA · Cupertino, California, USA +2 more · onsite · junior

trainiumcpp-langpython-langsoftware-engineerdistributed-inference

posted 3w ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer

posted 3w ago · verified today

Amazon (AWS)hyperscaler

New York, New York, USA · Sunnyvale, California, USA · onsite · senior

kubernetes-opsml-platforminferenceinference-enginesquantization

posted 3w ago · verified today