arXiv cs.AI
· Papers
Keyword Matters: Unveiling the Energy Sensitivity of On-Device LLM Prompting
arXiv:2607.22568v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed on mobile and embedded devices to improve privacy and reduce network latency. Yet on-device inference faces a fundamental constraint: high energy consumption on battery-powered, resource-limited hardware. While model