AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

1296 open roles · 85 companies · last verified today

94 roles

fine-tuning

Amazon (AWS)hyperscaler

New York, New York, USA · Sunnyvale, California, USA · onsite · senior

kubernetes-opsml-platforminferenceinference-enginesquantization

posted 3w ago · verified today

Nebiusneocloud

Remote - United States · United States · remote · $228K–$285K · manager

eng-managerfine-tuninginferenceinference-engineskubernetes-ops

posted 4w ago · verified today

Baseteninference provider

New York · San Francisco · hybrid · $200K–$400K/year · mid

inferenceinference-enginessolutions-architectevaluationperformance-engineer

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Cupertino, California, USA · Seattle, Washington, USA · onsite · mid

trainiummodel-parallelismtraining-frameworkscollectivesperformance-engineer

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Seattle, Washington, USA · onsite · staff plus

solutions-architectcudagpu-generickubernetes-opscluster-datacenter

posted 10w ago · verified today

Amazon (AWS)hyperscaler

Santa Clara, California, USA · onsite · unknown

cudagpu-kernelstriton-langgpu-genericmodel-parallelism

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Houston, Texas, USA · onsite · senior

cudagpu-kernelsmodel-parallelismpytorch-disttraining-frameworks

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Bellevue, Washington, USA · onsite · mid

software-engineerml-platformtrainiumscheduling-orchestrationfsdp

posted 6w ago · verified today

Amazon (AWS)hyperscaler

Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior

cudacutlass-cutedistributed-inferenceevaluationflash-attention

posted 4w ago · verified today

Amazon (AWS)hyperscaler

Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior

solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks

posted 3mo ago · verified today

Cerebraschip vendor

Canada · United States · senior

amdcpp-langinferenceinference-enginespython-lang

posted 9mo ago · verified today

Baseteninference provider

San Francisco · hybrid · $200K–$275K/year · unknown

post-traininggpu-genericmodel-parallelismpytorch-distresearch-engineer

posted 5mo ago · verified today

Baseteninference provider

New York · San Francisco · hybrid · $165K–$330K/year · unknown

ml-platformsoftware-engineerfine-tuningkubernetes-opspost-training

posted 7mo ago · verified today

SambaNovachip vendor

San Jose, CA · San Jose, California, United States · $220K–$300K · staff plus

fine-tuninginferencesoftware-engineerinference-enginespre-training

posted 4w ago · verified today

SambaNovachip vendor

Bengaluru, India · Bengaluru, Karnataka, India · 3324K–4062K INR · unknown

python-langsolutions-architectinferencefine-tuningkubernetes-ops

posted 9w ago · verified today

Modalinference provider

New York · San Francisco · $180K–$250K/year · unknown

solutions-architectsglang-enginevllm-engineperformance-engineerinference

posted 6mo ago · verified today

Modalinference provider

New York · San Francisco · onsite · $220K–$300K/year · senior

ml-platformsoftware-engineerlinux-kernelscheduling-orchestrationreliability-sre

posted 23mo ago · verified today

Fireworks AIinference provider

New York · San Mateo +1 more · hybrid · $200K–$260K/year · senior

inference-enginesquantizationsglang-enginevllm-enginefine-tuning

posted 3mo ago · verified today

Fireworks AIinference provider

New York · San Mateo · hybrid · $200K–$260K/year · unknown

fine-tuningpython-langsglang-enginesolutions-architectvllm-engine

posted 3mo ago · verified today

Fireworks AIinference provider

London · senior

fine-tuninginferenceinference-enginespython-langsglang-engine

posted 8w ago · verified today

Fireworks AIinference provider

Singapore · 200K–350K SGD/year · senior

fine-tuninginferenceinference-enginespython-langsglang-engine

posted 7w ago · verified today

Fireworks AIinference provider

San Mateo · hybrid · $210K–$320K/year · mid

evaluationpost-trainingfine-tuningml-platforminference

posted 10mo ago · verified today

xAIfrontier lab

Palo Alto, CA · $180K–$440K/year est. · unknown

gpu-genericpost-trainingpre-trainingpython-langresearch-engineer

posted 5mo ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA · $185K–$275K/year est. · staff plus

reinforcement-learningpost-trainingresearch-engineerfine-tuninggpu-generic

posted 5mo ago · verified today