Semantic-Aware Scheduling for GPU Clusters with Large Language Models

Abstract

SchedMate brings semantic context to GPU-cluster scheduling by extracting information from source code, runtime logs, and historical jobs. Its LLM-based scheduling advisor, metric tracker, and failure handler enhance existing schedulers through a non-intrusive interface.

Publication
arXiv preprint arXiv:2510.03334
Zerui Wang
Zerui Wang
Ph.D. Student · Research Intern