Energy- and Memory-Efficient PEFT Methods for Personalized On-Device SLMs on Consumer GPUs
Results show that compact SLMs paired with PEFT provide a practical, energy-aware path to personalized on-device deployment, with the optimal method set by the dominant constraint: LoRA+ for energy and QLoRA for memory.