Scheduler-Step GPU Energy Attribution for Continuously Batched LLM Serving
Probir Roy · Zenodo (CERN European Organization for Nuclear Research) · 2026
GPU energy attribution for continuously batched LLM serving. SEAM aligns serving-engine scheduler steps with measured GPU-board energy, conserves each measured sample, reports bounded attribution where ownership is ambiguous, and evaluates attribution-guided energy optimizations.