Scheduler-Step GPU Energy Attribution for Continuously Batched LLM Serving

Probir Roy · Zenodo (CERN European Organization for Nuclear Research) · 2026

GPU energy attribution for continuously batched LLM serving. SEAM aligns serving-engine scheduler steps with measured GPU-board energy, conserves each measured sample, reports bounded attribution where ownership is ambiguous, and evaluates attribution-guided energy optimizations.

Read the paper · More papers on PaperTik