AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
1296 open roles · 85 companies · last verified today
105 roles
ElevenLabsai startup
Bulgaria · Poland +2 more · remote · unknown
cudagpu-genericgpu-kernelsinferenceinference-engines
posted 19d ago · verified today
Figureai startup
HQ · San Jose, CA · $180K–$275K/year est. · staff plus
cpp-langpython-langquantizationgpu-genericperformance-engineer
posted 11w ago · verified today
1Xai startup
San Carlos, CA · onsite · $250K–$350K/year · unknown
gpu-genericpre-trainingpython-langresearch-engineertraining-data-infra
posted 3mo ago · verified today
Lightning AIai startup
London, UK · New York, New York +6 more · remote · $165K–$310K · senior
python-langresearch-engineertraining-frameworksevaluationfine-tuning
posted 5w ago · verified today
Lightning AIai startup
London, England, United Kingdom · New York, New York +5 more · $180K–$250K · unknown
python-langsolutions-architectdistributed-inferenceinferenceinference-engines
posted 3mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · unknown
gpu-genericpost-traininggpu-kernelsmodel-parallelismtraining-frameworks
posted 7mo ago · verified today
RadixArkai startup
Palo Alto, CA · Palo Alto Office · $200K–$400K/year est. · senior
inference-enginesgpu-genericinferenceperformance-engineerdistributed-inference
posted 7mo ago · verified today
Inferactinference provider
Remote · remote · unknown
inferenceinference-enginesvllm-enginepython-langsoftware-engineer
posted 12d ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
amdgpu-kernelsinferenceperformance-engineerrocm-hip
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
tpuinferenceinference-enginesjax-pallasperformance-engineer
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
amdgpu-kernelsinferenceinference-enginesperformance-engineer
posted 11w ago · verified today
Inferactinference provider
Singapore · onsite · 200K–400K SGD/year · unknown
vllm-engineinferenceinference-engineskv-cache-systemspython-lang
posted 12w ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
inferenceinference-enginespython-langvllm-enginesoftware-engineer
posted 3mo ago · verified today
Inferactinference provider
San Francisco · onsite · $200K–$400K/year · unknown
tpuinference-enginesjax-pallasperformance-engineerxla-compiler
posted 3mo ago · verified today
FriendliAIinference provider
San Francisco · hybrid · mid
solutions-architectgpu-kernelsinference-engineskubernetes-opsobservability
posted 6mo ago · verified today
Prime Intellectneocloud
San Francisco · onsite · unknown
kubernetes-opsml-platformfine-tuninggpu-genericnvidia
posted 10w ago · verified today
Periodic Labsfrontier lab
Menlo Park, CA · onsite · unknown
collectivescudacutlass-cutefsdpgpu-generic
posted 4mo ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · mid
distributed-inferencegpu-kernelsinferenceinference-enginessoftware-engineer
posted 16d ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · senior
ml-platformsoftware-engineerinferencejax-pallaspytorch-dist
posted 19d ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
gpu-kernelsinferenceinference-engineskv-cache-systemsperformance-engineer
posted 20d ago · verified today
Nebiusneocloud
Remote - United States · United States · remote · $228K–$285K · manager
eng-managerfine-tuninginferenceinference-engineskubernetes-ops
posted 4w ago · verified today
Baseteninference provider
New York · San Francisco · hybrid · $200K–$400K/year · mid
inferenceinference-enginessolutions-architectevaluationperformance-engineer
posted 4w ago · verified today
Amazon (AWS)hyperscaler
San Francisco, California, USA · onsite · senior
cudagpu-kernelsnvidiapython-langresearch-engineer
posted 3mo ago · verified today
Amazon (AWS)hyperscaler
New York, New York, USA · Seattle, Washington, USA · onsite · senior
software-engineerinferenceinference-enginesdistributed-inferencetensorrt-stack
posted 6w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
trainiumdistributed-inferenceinferencesoftware-engineercpp-lang
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Cupertino, California, USA · onsite · senior
trainiumcpp-langdistributed-inferenceinferenceinference-engines
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Seattle, Washington, USA · onsite · mid
software-engineerml-platformpost-trainingreinforcement-learningtraining-frameworks
posted 3mo ago · verified today
Amazon (AWS)hyperscaler
Tel Aviv-Yafo, Tel Aviv, ISR · onsite · mid
cpp-langsoftware-engineercustom-asicdistributed-inferencegpu-kernels
posted 5w ago · verified today
Amazon (AWS)hyperscaler
Boston, Massachusetts, USA · Seattle, Washington, USA +1 more · onsite · senior
cudacutlass-cutedistributed-inferenceevaluationflash-attention
posted 4w ago · verified today
Amazon (AWS)hyperscaler
Herndon, Virginia, USA · New York, New York, USA +1 more · onsite · senior
solutions-architectgpu-genericinference-enginesmodel-parallelismtraining-frameworks
posted 3mo ago · verified today