ai-infra-jobs

AI infrastructure roles, filterable by the stack you actually work on.

Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.

883 open roles · 17 companies · last verified today

Anyscaleai startup

San Francisco · remote · $215K–$265K · senior

scheduling-orchestrationsoftware-engineercluster-datacenterml-platformtraining-frameworks

posted 3mo ago · verified today

Anyscaleai startup

Palo Alto · San Francisco · remote · $170K–$245K · mid

distributed-inferenceinference-enginesinferencegpu-genericsoftware-engineer

posted 11w ago · verified today

Anyscaleai startup

San Francisco · remote · $215K–$265K · staff plus

ml-platformscheduling-orchestrationsoftware-engineercluster-datacenter

posted 3mo ago · verified today

Anyscaleai startup

San Francisco · remote · $200K–$240K · mid

scheduling-orchestrationsoftware-engineerml-platformcluster-datacenterreliability-sre

posted 5w ago · verified today

Anyscaleai startup

San Francisco · $215K–$230K · mid

software-engineerml-platformtraining-data-infradistributed-inferenceinference-engines

posted 5w ago · verified today

Anyscaleai startup

Palo Alto · San Francisco · remote · $215K–$275K · senior

reliability-sresreml-platformscheduling-orchestrationsoftware-engineer

posted 8w ago · verified today

Modalinference provider

New York · San Francisco · $220K–$300K · senior

ml-platformsoftware-engineercluster-datacenterperformance-engineergpu-generic

posted 22mo ago · verified today

Modalinference provider

New York · San Francisco · $180K–$250K · mid

gpu-genericinference-enginesml-platformsoftware-engineersolutions-architect

posted 5mo ago · verified today

Modalinference provider

New York · $150K–$270K · mid

cluster-datacentersoftware-engineerml-platformreliability-sre

posted 3mo ago · verified today

Modalinference provider

Stockholm · $140K–$200K · senior

ml-platformsoftware-engineerinference-enginesgpu-genericscheduling-orchestration

posted 7mo ago · verified today

Modalinference provider

London or Stockholm · New York +1 more · $150K–$220K · senior

software-engineerml-platformgpu-genericdistributed-inferencefine-tuning

posted 10mo ago · verified today

Fireworks AIinference provider

New York · San Mateo · $250K–$400K · mid

research-engineergpu-generictraining-frameworksnvidiapre-training

posted 5w ago · verified today

Fireworks AIinference provider

San Mateo · $200K–$300K · mid

software-engineerinference-enginesml-platform

posted 17d ago · verified today

Fireworks AIinference provider

New York · San Mateo · $175K–$220K · senior

software-engineerml-platformscheduling-orchestrationdistributed-inferenceinference

posted 14mo ago · verified today

Fireworks AIinference provider

New York · Remote +1 more · $200K–$260K · senior

fine-tuninginferenceinference-enginesml-platformsolutions-architect

posted 9w ago · verified today

Fireworks AIinference provider

New York · San Mateo · remote · $200K–$350K · mid

software-engineertraining-frameworksml-platformscheduling-orchestrationstorage-checkpointing

posted 17d ago · verified today

Fireworks AIinference provider

New York · San Mateo · $200K–$260K · senior

inferenceinference-enginessolutions-architectfine-tuningdistributed-inference

posted 9w ago · verified today

Fireworks AIinference provider

San Mateo · $175K–$220K · senior

gpu-kernelsperformance-engineerdistributed-inferencegpu-genericinference

posted 15mo ago · verified today

Fireworks AIinference provider

New York · $175K–$220K · mid

ml-platformsoftware-engineertraining-frameworksscheduling-orchestrationgpu-generic

posted 9w ago · verified today

Fireworks AIinference provider

San Mateo · $175K–$220K · mid

inference-enginesml-platformsoftware-engineerinferencedistributed-inference

posted 9mo ago · verified today

xAIfrontier lab

Palo Alto, CA · Palo Alto, California; Seattle, Washington +1 more · mid

network-fabricsoftware-engineergpu-genericnetwork-engineercollectives

posted 1d ago · verified today

xAIfrontier lab

Palo Alto, CA · Palo Alto, California · senior

cluster-datacentergpu-genericml-platformreliability-sresre

posted 1d ago · verified today

xAIfrontier lab

Palo Alto, CA · Palo Alto, California

inferenceinference-enginesgpu-kernelssoftware-engineerdistributed-inference

posted 1d ago · verified today