Research
Resource Consumption Threats in Large Language Models
arXiv:2603.16068v3 Announce Type: replace-cross Abstract: Given limited and costly computational infrastructure, resource efficiency is a key requirement for large language models (LLMs). Efficient LL
arXiv:2603.16068v3 Announce Type: replace-cross Abstract: Given limited and costly computational infrastructure, resource efficiency is a key requirement for large language models (LLMs). Efficient LLMs increase service capacity for providers and reduce latency and API costs for users. Recent resource consumption threats induce excessive generation, degrading model efficiency and harming both service availability and economic sustainability. This survey presents a systematic review of threats to resource consumption in LLMs. We further establish a unified view of this emerging area by clarifying its scope and examining the problem along the full pipeline from threat induction to mechanism understanding and mitigation. Our goal is to clarify the problem landscape for this emerging area, thereby providing a clearer foundation for characterization and mitigation.
Related
- LongSpec: Long-Context Lossless Speculative Decoding with Efficient Drafting and Verification
- On the Limits of Layer Pruning for Generative Reasoning in Large Language Models
- Resource-constrained Amazons chess decision framework integrating large language models and graph attention
- GRACE: A Dynamic Coreset Selection Framework for Large Language Model Optimization
- FBS: Modeling Native Parallel Reading inside a Transformer
- SpecBound: Adaptive Bounded Self-Speculation with Layer-wise Confidence Calibration
Source: arXiv cs.AI | 2026-04-14