AI infrastructure roles, filterable by the stack you actually work on.
Distributed training · inference serving · GPU fleets · network fabric — aggregated straight from company boards, never a copy of a copy.
884 open roles · 17 companies · last verified today
222 roles
Baseteninference provider
Montreal · New York +3 more · remote · $165K–$330K · mid
inferenceinference-enginesml-platformsoftware-engineerdistributed-inference
posted 17mo ago · verified today
Baseteninference provider
Montreal · New York +2 more · remote · $180K–$360K · mid
distributed-inferenceinferenceinference-enginessoftware-engineergpu-generic
posted 11w ago · verified today
Baseteninference provider
San Francisco · remote · $200K–$275K · senior
post-trainingresearch-engineertraining-frameworksfine-tuninggpu-generic
posted 4mo ago · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · manager
eng-managercustom-asicinferenceinference-enginesml-platform
posted today · verified today
SambaNovachip vendor
Remote - US · remote · senior
inferenceinference-enginessoftware-engineercustom-asicdistributed-inference
posted today · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · staff plus
inference-enginescustom-asicinferenceml-platformsoftware-engineer
posted today · verified today
SambaNovachip vendor
San Jose, CA · San Jose, California, United States · staff plus
custom-asicsoftware-engineercollectivesgpu-kernelsml-platform
posted today · verified today
Scale AIai startup
New York, NY · San Francisco, CA +1 more · manager
ml-platforminference-enginessoftware-engineerscheduling-orchestrationgpu-generic
posted today · verified today
Scale AIai startup
London, UK · manager
ml-platformgpu-genericinference-enginesscheduling-orchestrationsoftware-engineer
posted today · verified today
Anyscaleai startup
San Francisco · remote · $200K–$240K · mid
scheduling-orchestrationsoftware-engineerml-platformcluster-datacenterreliability-sre
posted 5w ago · verified today
Anyscaleai startup
San Francisco · remote · $215K–$265K · staff plus
ml-platformscheduling-orchestrationsoftware-engineercluster-datacenter
posted 3mo ago · verified today
Anyscaleai startup
Palo Alto · San Francisco · remote · $170K–$245K · mid
distributed-inferenceinference-enginesinferencegpu-genericsoftware-engineer
posted 11w ago · verified today
Anyscaleai startup
Bengaluru, Karnataka · mid
software-engineerml-platformscheduling-orchestrationcluster-datacenternetwork-fabric
posted 17d ago · verified today
Anyscaleai startup
Palo Alto · San Francisco · remote · $215K–$275K · senior
reliability-sresreml-platformscheduling-orchestrationsoftware-engineer
posted 8w ago · verified today
Anyscaleai startup
San Francisco · remote · $215K–$265K · senior
scheduling-orchestrationsoftware-engineercluster-datacenterml-platformtraining-frameworks
posted 3mo ago · verified today
Modalinference provider
Stockholm · $140K–$200K · senior
ml-platformsoftware-engineerinference-enginesgpu-genericscheduling-orchestration
posted 7mo ago · verified today
Modalinference provider
New York · San Francisco · $220K–$300K · senior
ml-platformsoftware-engineercluster-datacenterperformance-engineergpu-generic
posted 22mo ago · verified today
Modalinference provider
Stockholm · $175K–$250K · manager
eng-managercluster-datacenterscheduling-orchestration
posted 3mo ago · verified today
Fireworks AIinference provider
New York · San Mateo · remote · $200K–$350K · mid
software-engineertraining-frameworksml-platformscheduling-orchestrationstorage-checkpointing
posted 17d ago · verified today
Fireworks AIinference provider
New York · San Mateo · $250K–$400K · mid
research-engineergpu-generictraining-frameworksnvidiapre-training
posted 5w ago · verified today
Fireworks AIinference provider
New York · San Mateo · $175K–$220K · senior
software-engineerml-platformscheduling-orchestrationdistributed-inferenceinference
posted 14mo ago · verified today
Fireworks AIinference provider
New York · $175K–$220K · mid
ml-platformsoftware-engineertraining-frameworksscheduling-orchestrationgpu-generic
posted 9w ago · verified today
Fireworks AIinference provider
San Mateo · $175K–$220K · mid
inference-enginesml-platformsoftware-engineerinferencedistributed-inference
posted 9mo ago · verified today
xAIfrontier lab
Dublin · Dublin, Ireland · mid
network-fabricsoftware-engineergpu-genericinferencepre-training
posted today · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · mid
gpu-genericml-platformsoftware-engineernvidiatraining-frameworks
posted today · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · mid
gpu-kernelsnvidiasoftware-engineercluster-datacenterinference
posted today · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California · senior
cluster-datacentergpu-genericml-platformreliability-sresre
posted today · verified today
xAIfrontier lab
London · London, England, United Kingdom · senior
reinforcement-learningscheduling-orchestrationsoftware-engineercluster-datacenterml-platform
posted today · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California
inferenceinference-enginesgpu-kernelssoftware-engineerdistributed-inference
posted today · verified today
xAIfrontier lab
Palo Alto, CA · Palo Alto, California; Seattle, Washington +1 more · mid
network-fabricsoftware-engineergpu-genericnetwork-engineercollectives
posted today · verified today