Intelligence per Watt: Measuring Intelligence Efficiency of Local AI
DGX agentarXiv:2511.07885v4 Announce Type: replace-cross Abstract: Large language model (LLM) queries are predominantly processed by frontier models in centralized cloud infrastructure. Demand growth strains t