JouleShare Framework for Request-Level Energy Attribution in Batched Serving
August 4, 2026
JouleShare provides a method for assigning specific energy costs to individual requests in batched LLM serving environments like vLLM. It uses an offline Shapley value ground truth and a lightweight calibration model, JCalib, to predict energy consumption based on request features at serving time.
HOW THIS AFFECTS YOU
●
builderYou can implement precise energy-based chargebacks and sustainability reporting for your inference API.
●
founderThis allows for more accurate unit economic modeling and workload-based cost analysis.