AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

Simulating Tenant Responses to Energy Policy Interventions with Transaction-Cost-Aware LLM Age

DGX agent

arXiv:2607.24341v1 Announce Type: new Abstract: Recent studies use Large language models (LLMs) to simulate human opinions and decisions by prompting models with demographic, attitudinal, or persona-b

model-releasesarxiv-cs-ai
28 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SIREN: Towards End-to-End Extreme-Weather Early Warning with Experience-Grounded LLM Agents

DGX agent

arXiv:2607.24588v1 Announce Type: new Abstract: Early warning of extreme weather is essential for mitigating the societal, economic, and environmental risks posed by hazardous weather events. However,

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Skillware: A Software Ontology and Engineering Lifecycle for Persistent Behavioral Artifacts

DGX agent

arXiv:2607.18970v2 Announce Type: replace-cross Abstract: Agent Skills have become persistent behavioral artifacts across independent AI agent systems. They combine natural-language task specification

agentsarxiv-cs-ai
28 Jul 2026
Hardware

Source-Aware Reranking for Retrieval-Augmented Generation: A Reliability Prior Approach

DGX agent

arXiv:2607.22584v1 Announce Type: new Abstract: Standard Retrieval-Augmented Generation pipelines rank retrieved documents by semantic similarity alone, without accounting for source provenance or cre

hardwarearxiv-cs-ai
28 Jul 2026
Research

Sparse Autoencoders Encode Both Concepts and Functions: The Downstream Geometry of Feature Effects

DGX agent

arXiv:2607.24645v1 Announce Type: cross Abstract: The wide-scale use of sparse autoencoders (SAEs) as interpretability tools is limited by inconsistent links between SAE features and model behavior. F

researcharxiv-cs-ai
28 Jul 2026
Agents

Sparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation Detection

DGX agent

arXiv:2607.18080v2 Announce Type: replace-cross Abstract: Multimodal video misinformation detection is commonly formulated as a holistic video-understanding task, where the entire video and its associ

agentsarxiv-cs-ai
28 Jul 2026
Hardware

Spatial-IQ: Deconstructing Spatial Intelligence via Hierarchical Capability Tests

DGX agent

arXiv:2607.22864v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) excel at visual interpretation but fail on spatial reasoning tasks that humans solve reliably. Existing bench

hardwarearxiv-cs-ai
28 Jul 2026
Local Ai

Spatial Prediction of Soil Microplastics and Organic Matter Using Graph Attention Networks

DGX agent

arXiv:2607.22875v1 Announce Type: cross Abstract: Accurate estimation of soil microplastics and organic matter is essential to assess ecosystem health and support sustainable land use. This study pres

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Spatial Reasoning in LLM Game Agents: Impact of Causal Context and Multi-Step Planning

DGX agent

arXiv:2607.22732v1 Announce Type: new Abstract: LLM-based game agents often perform poorly on more complex tasks. This work examines whether these failures are linked to limited spatial reasoning and

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Spatula: Exploring On-Demand In-Situ Interfaces and Interaction for Attribute Control

DGX agent

arXiv:2607.10405v2 Announce Type: replace-cross Abstract: Controlling attributes is a critical step toward achieving the final creative outcome, yet current approaches fall short in supporting users i

model-releasesarxiv-cs-ai
28 Jul 2026
Local Ai

SpecAHD: Localize to Specialize for Automated Heuristic Design in Large-Scale Routing Problems

DGX agent

arXiv:2607.23676v1 Announce Type: new Abstract: LLM-based automated heuristic design (AHD) typically scores executable programs on complete instances or within fixed solver components. In large-scale

local-aiarxiv-cs-ai
28 Jul 2026
Agents

SpecBox: Speculative Sandbox Scheduling for Efficient LLM Agent Serving

DGX agent

arXiv:2607.23933v1 Announce Type: cross Abstract: As LLM agents increasingly rely on the Model Context Protocol (MCP) to invoke isolated external sandboxes, disaggregated sandbox deployment introduces

agentsarxiv-cs-ai
28 Jul 2026
Safety

Spectral Dynamics of Semantic Drift in Clinical Multi-Agent Language Model Networks

DGX agent

arXiv:2607.22758v1 Announce Type: cross Abstract: The integration of iterative LLMs within multi-agent diagnostic frameworks requires a rigorous quantitative reevaluation of underlying communication t

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Speed Reading Tool Powered by Artificial Intelligence for Students with ADHD, Dyslexia, and Short Attention Span

DGX agent

arXiv:2307.14544v2 Announce Type: replace-cross Abstract: This paper presents an artificial intelligence tool designed to assist students with dyslexia, ADHD, and short attention spans in processing t

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

SQBench: A Benchmark for Evaluating Task Delivery by Language-Model Agents in Production-Oriented Workflows

DGX agent

arXiv:2607.23123v1 Announce Type: new Abstract: Existing evaluations of large language models cover knowledge, reasoning, coding, and tool use, but they rarely treat a verifiable deliverable produced

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Stability of AI Governance Systems: A Coupled Dynamics Model of Public Trust and Social Disruptions

DGX agent

arXiv:2603.20248v2 Announce Type: replace-cross Abstract: AI systems are increasingly entrenched in public governance, yet scholarship lacks formal tools to determine when deviations of public trust i

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

STAIF: A Stage-wise Optimization for Complex Instruction Following

DGX agent

arXiv:2607.22649v1 Announce Type: new Abstract: Following complex instructions with multiple explicit constraints remains a fundamental challenge for large language models (LLMs). Existing alignment m

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

StanceBench: A Benchmark for Audio LLM-Based Interpersonal Stance Evaluation from Speech

DGX agent

arXiv:2607.22658v1 Announce Type: new Abstract: Speech-to-speech dialogue models increasingly depend on prosody and interactional nuance to convey social intent, yet benchmarks for these cues remain l

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

StanceFlip: A Comprehensive Multi-Dimensional Benchmark for Multimodal Conversational Stance Flipping Forecasting

DGX agent

arXiv:2607.24191v1 Announce Type: cross Abstract: Conversational stance detection has shifted from static text analysis to dynamic multimodal modeling. However, existing benchmarks exhibit three key l

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Statistically Supported LLM Ingredient and Recipe Data Collection in Computational Nutrition

DGX agent

arXiv:2607.23273v1 Announce Type: cross Abstract: Computational nutrition needs precise ingredient data, but current databases are incomplete, inconsistent, and built for human reference rather than a

researcharxiv-cs-ai
28 Jul 2026
Research

Steerable Chatbots: Exploring Personalization Control Interfaces via LLM Activation Steering

DGX agent

arXiv:2505.04260v3 Announce Type: replace-cross Abstract: Personalizing LLM responses typically requires users to articulate their preferences through prompting, which can be burdensome at cold start

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Stress-Testing EEG Foundation Models for Clinical Decoding: Dataset Identity and Targeted Negative Controls

DGX agent

arXiv:2607.24519v1 Announce Type: cross Abstract: Pretrained EEG foundation models are increasingly proposed for clinical decoding, but their transfer across populations and robustness to negative con

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Stress-testing large language model agents in a robotic chemistry laboratory

DGX agent

arXiv:2607.23045v1 Announce Type: new Abstract: AI is evaluated through knowledge, reasoning and plan generation, yet scientific agency requires reliable physical action and adaptation to evidence. He

agentsarxiv-cs-ai
28 Jul 2026
Research

Structural Preservation Governs Data Augmentation in Deep Learning-Based Laser Speckle Material Classification

DGX agent

arXiv:2607.22725v1 Announce Type: cross Abstract: Data augmentation is routinely used to improve generalization in image classification, but the assumptions underlying standard policies are poorly mat

researcharxiv-cs-ai
28 Jul 2026
Tutorials

Structure over Depth: A Single-Block Spatio-Temporal Transformer for Multi-Entity Reasoning

DGX agent

arXiv:2607.23077v1 Announce Type: new Abstract: Modeling multi-entity temporal data requires capturing dependencies across entities, time, and their interactions. Transformer-based approaches perform

tutorialsarxiv-cs-ai
28 Jul 2026
Research

Structure Over Scale: Schema-Constrained Causal Graphs for RAG

DGX agent

arXiv:2607.22592v1 Announce Type: new Abstract: Graph-based retrieval-augmented generation (GraphRAG) grounds answers in structured knowledge, but current systems extract entities and relationships ex

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Success Is Not Self-Explanatory: Auditing Success Provenance in Agent Evaluation

DGX agent

arXiv:2607.24054v1 Announce Type: new Abstract: A correct answer can conceal why an agent succeeded. Once agents change their information state during evaluation, correctness no longer distinguishes i

model-releasesarxiv-cs-ai
28 Jul 2026
Research

SwitchBraidNet: Quantisation-Aware Lightweight Architecture for Hybrid Brain-Computer Interface

DGX agent

arXiv:2606.18816v2 Announce Type: replace-cross Abstract: Hybrid brain-computer interfaces (BCIs) that integrate motor imagery (MI) and steady-state visual evoked potentials (SSVEP) provide high-dimen

researcharxiv-cs-ai
28 Jul 2026
Model Releases

SymStep: Symbolic Step Verification for Logical Reasoning

DGX agent

arXiv:2607.23055v1 Announce Type: new Abstract: Chain-of-thought (CoT) prompting can fail severely on constraint-dense logical reasoning tasks, where unverified errors accumulate silently across steps

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Synthetic Scenario Generation for Evaluation of Industry 4.0 Agents

DGX agent

arXiv:2607.22563v1 Announce Type: new Abstract: Industrial agent benchmarks require realistic evaluation scenarios that integrate telemetry, failure modes, maintenance records, and domain standards. H

safetyarxiv-cs-ai
28 Jul 2026
Safety

SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding

DGX agent

arXiv:2607.23991v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly controlled through system prompts that specify roles, styles, formats, and safety requirements. However,

safetyarxiv-cs-ai
28 Jul 2026
Agents

TableMind: An Autonomous Programmatic Agent for Tool-Augmented Table Reasoning

DGX agent

arXiv:2509.06278v4 Announce Type: replace Abstract: Table reasoning requires models to jointly perform comprehensive semantic understanding and precise numerical operations. Although recent large lang

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Tag Questions and the Generational Reversal of Sycophancy Across 45 Language Models

DGX agent

arXiv:2607.23976v1 Announce Type: cross Abstract: Appending a two-word confirmation tag to a decision question -- 'Is X the better choice?' versus 'X is the better choice, right?' -- changes whether a

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Task-Conditional Faithfulness Auditing of Multimodal LLMs for Grid Diagnosis

DGX agent

arXiv:2607.24539v1 Announce Type: new Abstract: Multimodal large language models (LLMs) can combine topology, measurements, and incident text for grid diagnosis, yet answer accuracy does not establish

researcharxiv-cs-ai
28 Jul 2026
Research

Teacher Knows It Best: Spontaneous Symmetry Breaking and Tipping Points in Networked Langevin Dynamics AI Sycophancy

DGX agent

arXiv:2607.24304v1 Announce Type: cross Abstract: We formulate a statistical physics framework to model a networked stochastic dynamical system exhibiting bistability, driven by additive noise and soc

researcharxiv-cs-ai
28 Jul 2026
Research

Temporal Context Reinstatement Drives Episodic-Like Order Memory in Long-Context Language Models

DGX agent

arXiv:2607.22575v1 Announce Type: new Abstract: Human episodic memory supports the retrieval of experiences that unfold over extended timescales, yet the computational mechanisms underlying this abili

researcharxiv-cs-ai
28 Jul 2026
Agents

Test-Time Coverage: Test-Conditioned Data Curation for Deployment-Aware Learning

DGX agent

arXiv:2607.22697v1 Announce Type: new Abstract: Deployed AI systems are often trained from broad candidate data pools, necessitating data curation towards the deployment test distribution. However, st

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

TextRich: A Multi-Domain Benchmark for Detecting AI-Generated Text-Rich Images from GPT-Image-2

DGX agent

arXiv:2606.19259v2 Announce Type: replace-cross Abstract: Text-rich images often contain privacy-sensitive, transactional, or decision-relevant information. As recent multimodal image generation model

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

The Cost of Knowing: A Resource-Aware Protocol for Benchmarking Hallucination Beyond Static Leaderboards

DGX agent

arXiv:2607.24063v1 Announce Type: new Abstract: On standard factuality tasks, frontier models now cluster near the top of the scale. The question is therefore shifting from how factual a system is tow

model-releasesarxiv-cs-ai
28 Jul 2026
Research

The Equalizer: Introducing Shape-Gain Decomposition in Neural Audio Codecs

DGX agent

arXiv:2602.15491v2 Announce Type: replace-cross Abstract: Neural audio codecs (NACs) typically encode the short-term energy (gain) and normalized structure (shape) of speech/audio signals jointly with

researcharxiv-cs-ai
28 Jul 2026
Model Releases

The Half-Lives of Generative-AI Evidence: A 40-Record Audit, a Claim-Currency Framework, and a Reflexive Case of Frontier-Model-Assisted Research

DGX agent

arXiv:2607.24032v1 Announce Type: new Abstract: Generative-AI evaluations can become historical before publication, yet calendar age does not affect every conclusion equally. This paper has two linked

model-releasesarxiv-cs-ai
28 Jul 2026
Applications

The Illusion of Secure LLM Code: Closing the Security Gap via Iterative Reprompting

DGX agent

arXiv:2607.23710v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly integrated into software development workflows, yet their ability to autonomously generate secure authen

applicationsarxiv-cs-ai
28 Jul 2026
Safety

The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation

DGX agent

arXiv:2607.24720v1 Announce Type: cross Abstract: Multi-turn long-horizon planning is critical for foundation model agents, yet how to fundamentally improve it remains unclear. Existing models are tra

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

The Scaffold Effect in Coding Agents: Harness Choice as a Hidden Variable in Coding-Agent Evaluation

DGX agent

arXiv:2607.22585v1 Announce Type: new Abstract: Public leaderboards for coding agents typically rank systems by model name and pass rate, while the surrounding harness (the scaffold that issues tools,

model-releasesarxiv-cs-ai
28 Jul 2026
Applications

The SpiNNaker2 chip: a many-core platform for flexible and scalable brain-inspired computing

DGX agent

arXiv:2607.24396v1 Announce Type: cross Abstract: In deep learning, efficiency gets more and more important to compensate for the ongoing growth in model sizes and applications. Neuromorphic hardware

applicationsarxiv-cs-ai
28 Jul 2026
Model Releases

The Tokenizer Tax: Quantifying and Explaining the Cross-Lingual Cost of Subword Tokenization for Indian Languages

DGX agent

arXiv:2607.24276v1 Announce Type: cross Abstract: Large language models (LLMs) process text through subword tokenizers rather than directly reading characters or words. Because these tokenizers are tr

model-releasesarxiv-cs-ai
28 Jul 2026
Local Ai

The Visual Bottleneck: Sparse-Frame Adaptation of MLLMs for Joint Spatial-Temporal Video Grounding

DGX agent

arXiv:2607.24570v1 Announce Type: cross Abstract: Large-scale video platforms process millions of uploads hourly, requiring moderation systems that can localize when and where policy violations occur

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Through the Bottleneck: How Multi-head Latent Attention Separates Content from Position in Language Models

DGX agent

arXiv:2607.23054v1 Announce Type: cross Abstract: Multi-head Latent Attention (MLA), introduced in DeepSeek-V2, compresses key-value pairs through a shared low-rank bottleneck (cKV), achieving 81% KV-

model-releasesarxiv-cs-ai
28 Jul 2026
← Previous
1…7071727374…448
Next →