1 paper
Bohua Zou, Nian Liu, Binqi Sun +6
On-device LLM inference is increasingly attractive for privacy-preserving, reliable, and cost-effective deployment, yet its energy and thermal costs remain a critical bottleneck. E…