1 repository on SrcLog
Measured energy cost of on-prem LLM inference on AMD ROCm: 8.7× tokens-per-joule from batching, with a cost-governance gateway and agent built on the measurements.