ai-infra-jobs

AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

883 open roles · 17 companies · last verified today

306 roles

inference

Anyscaleai startup

Palo Alto · San Francisco · remote · $170K–$245K · mid

distributed-inferenceinference-enginesinferencegpu-genericsoftware-engineer

posted 11w ago · verified today

Modalinference provider

London or Stockholm · New York +1 more · $150K–$220K · senior

software-engineerml-platformgpu-genericdistributed-inferencefine-tuning

posted 10mo ago · verified today

Modalinference provider

New York · San Francisco · $200K–$350K · senior

gpu-kernelsperformance-engineerinference-enginesnvidiainference

posted 3mo ago · verified today

Modalinference provider

Stockholm · senior

solutions-architectml-platformfine-tuninggpu-genericinference

posted 7mo ago · verified today

Modalinference provider

New York · San Francisco · $150K–$350K · senior

inference-enginesinferencedistributed-inferencegpu-kernelsresearch-engineer

posted 6w ago · verified today

Fireworks AIinference provider

Singapore · 200K–350K SGD · senior

solutions-architectinferenceinference-enginesfine-tuninggpu-generic

posted 3w ago · verified today

Fireworks AIinference provider

New York · Remote +1 more · $200K–$260K · senior

fine-tuninginferenceinference-enginesml-platformsolutions-architect

posted 9w ago · verified today

Fireworks AIinference provider

San Mateo · $175K–$220K · mid

inference-enginesml-platformsoftware-engineerinferencedistributed-inference

posted 9mo ago · verified today

Fireworks AIinference provider

New York · San Mateo · $175K–$220K · senior

software-engineerml-platformscheduling-orchestrationdistributed-inferenceinference

posted 14mo ago · verified today

Fireworks AIinference provider

San Mateo · $175K–$220K · senior

gpu-kernelsperformance-engineerdistributed-inferencegpu-genericinference

posted 15mo ago · verified today

Fireworks AIinference provider

New York · San Francisco Bay Area +1 more · $140K–$210K · mid

inferenceinference-enginessolutions-architectml-platformfine-tuning

posted 3w ago · verified today

Fireworks AIinference provider

New York · San Mateo · $200K–$260K · senior

inferenceinference-enginessolutions-architectfine-tuningdistributed-inference

posted 9w ago · verified today

xAIfrontier lab

Dublin · Dublin, Ireland · mid

network-fabricsoftware-engineergpu-genericinferencepre-training

posted 1d ago · verified today

xAIfrontier lab

Palo Alto, CA · Palo Alto, California

inferenceinference-enginesgpu-kernelssoftware-engineerdistributed-inference

posted 1d ago · verified today

xAIfrontier lab

London · London, England, United Kingdom · senior

inferenceinference-enginesdistributed-inferencesoftware-engineergpu-generic

posted 1d ago · verified today

xAIfrontier lab

Palo Alto, CA · Palo Alto, California · mid

gpu-kernelsnvidiasoftware-engineercluster-datacenterinference

posted 1d ago · verified today

xAIfrontier lab

Palo Alto, CA · Palo Alto, California · senior

cluster-datacentergpu-genericml-platformreliability-sresre

posted 1d ago · verified today

xAIfrontier lab

Memphis, Tennessee; Remote International · Memphis, TN · remote · manager

cluster-datacenterdatacenter-engineerreliability-sreeng-managerpre-training

posted 1d ago · verified today

xAIfrontier lab

Memphis, Tennessee; Southaven, Mississippi · Memphis, TN +1 more · manager

cluster-datacentereng-managernetwork-fabricdatacenter-engineerinference

posted 1d ago · verified today

xAIfrontier lab

Memphis, Tennessee · Memphis, TN · junior

cluster-datacenterreliability-sresreinferencepre-training

posted 1d ago · verified today

xAIfrontier lab

Memphis, Tennessee; Southaven, Mississippi · Memphis, TN +1 more · mid

reliability-sresrecluster-datacentergpu-genericinference

posted 1d ago · verified today

xAIfrontier lab

Palo Alto, CA · Palo Alto, California; Seattle, Washington +1 more · mid

network-fabricsoftware-engineergpu-genericnetwork-engineercollectives

posted 1d ago · verified today

CoreWeaveneocloud

Bellevue, WA · Livingston, NJ +5 more · remote · manager

eng-managerinferenceinference-enginesdistributed-inferenceml-platform

posted 1d ago · verified today

CoreWeaveneocloud

Bellevue, WA · Livingston, NJ +4 more · senior

inference-enginessolutions-architectdistributed-inferenceinference

posted 1d ago · verified today

CoreWeaveneocloud

Bellevue, WA · Livingston, NJ +4 more · senior

gpu-genericperformance-engineercluster-datacenterml-platforminference

posted 1d ago · verified today

CoreWeaveneocloud

Bellevue, WA · New York, NY +5 more · mid

cluster-datacentergpu-genericreliability-sresreinference

posted 1d ago · verified today

CoreWeaveneocloud

Bellevue, WA · New York, NY +1 more · senior

cluster-datacenterdatacenter-engineersoftware-engineergpu-genericreliability-sre

posted 1d ago · verified today

CoreWeaveneocloud

Bellevue, WA · Sunnyvale, CA +1 more · senior

performance-engineergpu-genericml-platforminferencedistributed-inference

posted 1d ago · verified today