AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1297 open roles · 85 companies · last verified today

83 roles

post-training

Prime Intellectneocloud

Remote · San Francisco · unknown

cudadeepspeed-libfsdpgpu-kernelsmodel-parallelism

posted 10w ago · verified today

TensorWaveneocloud

Las Vegas, Nevada · onsite · mid

amdgpu-generickubernetes-opsreliability-srescheduling-orchestration

posted 3w ago · verified today

Liquid AIfrontier lab

Boston · Remote +1 more · hybrid · unknown

cudagpu-genericgpu-kernelsperformance-engineercpp-lang

posted 13mo ago · verified today

Rekafrontier lab

US, UK, Singapore, Remote · remote · unknown

cpp-langcudafine-tuninggpu-genericgpu-kernels

posted 8mo ago · verified today

Mistral AIfrontier lab

Amsterdam · Lausanne +3 more · hybrid · unknown

post-trainingreinforcement-learningresearch-engineerevaluationfine-tuning

posted 12d ago · verified today

Thinking Machines Labfrontier lab

New York · San Francisco · onsite · unknown

post-trainingreliability-sresrepython-langreinforcement-learning

posted 15d ago · verified today

Amazon (AWS)hyperscaler

Vancouver, British Columbia, CAN · onsite · senior

ml-platformsoftware-engineerevaluationinferencekubernetes-ops

posted 13d ago · verified today

Nscaleneocloud

London · UK · senior

evaluationfine-tuninggpu-genericinferenceinference-engines

posted 5mo ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · manager

eng-managerml-platformperformance-engineercluster-datacentercuda

posted 16d ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer

posted 20d ago · verified today

Baseteninference provider

New York · San Francisco · hybrid · $200K–$400K/year · mid

inferenceinference-enginessolutions-architectevaluationperformance-engineer

posted 4w ago · verified today

Amazon (AWS)hyperscaler

San Francisco, California, USA · onsite · senior

cudagpu-kernelsnvidiapython-langresearch-engineer

posted 3mo ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · Seattle, Washington, USA · onsite · mid

trainiumperformance-engineersoftware-engineercollectivesgpu-generic

posted 7mo ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · Seattle, Washington, USA · onsite · mid

trainiummodel-parallelismtraining-frameworkscollectivesperformance-engineer

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

cudacutlass-cutedistributed-inferenceevaluationflash-attention

posted 4w ago · verified today

Cerebraschip vendor

Canada · United States · senior

amdcpp-langinferenceinference-enginespython-lang

posted 9mo ago · verified today

Baseteninference provider

San Francisco · hybrid · $200K–$275K/year · unknown

model-parallelismpost-trainingpytorch-distresearch-engineertraining-frameworks

posted 5mo ago · verified today

Baseteninference provider

New York · San Francisco · hybrid · $165K–$330K/year · unknown

ml-platformsoftware-engineerfine-tuningkubernetes-opspost-training

posted 7mo ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $220K–$300K · staff plus

fine-tuninginferencesoftware-engineerinference-enginespre-training

posted 4w ago · verified today

Scale AIai startup

New York, NY · San Francisco, CA · $290K–$363K · manager

cudaeng-managerpost-trainingtraining-frameworksflash-attention

posted 11mo ago · verified today

Fireworks AIinference provider

New York · San Mateo +1 more · hybrid · $200K–$260K/year · senior

inference-enginesquantizationsglang-enginevllm-enginefine-tuning

posted 3mo ago · verified today

Fireworks AIinference provider

New York · San Mateo · hybrid · $200K–$260K/year · unknown

fine-tuningpython-langsglang-enginesolutions-architectvllm-engine

posted 3mo ago · verified today

Fireworks AIinference provider

London · senior

fine-tuninginferenceinference-enginespython-langsglang-engine

posted 7w ago · verified today