Energy- and Memory-Efficient PEFT Methods for Personalized On-Device SLMs on Consumer GPUs
arXiv:2608.04488v1 Announce Type: new Abstract: Despite rapid advances in large language models (LLMs), deploying and personalizing them on resource-constrained devices remains impractical due to high