AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Agents

Autoreflection: How Agentic Strange Loops Turn Human Culture into AI Infrastructure

DGX agent

arXiv:2608.03800v1 Announce Type: cross Abstract: An LLM-based agent is a loop that reads itself. Agentic frameworks externalize identity, memory, and disposition into editable files. The agent loads

agentsarxiv-cs-ai
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

AutoSND: From Execution Evidence to Structural Policies for Automated Network Dismantling Heuristic Discovery

DGX agent

arXiv:2608.03653v1 Announce Type: new Abstract: Network dismantling is fundamental to analyzing the robustness and vulnerability of complex systems, yet practical heuristics must balance effectiveness

applicationsarxiv-cs-ai
5 Aug 2026
Model Releases

Balancing Efficiency and Efficacy: Training-Free Attention-Guided Switching Between Explicit and Latent Thoughts for MLLMs

DGX agent

arXiv:2608.03450v1 Announce Type: cross Abstract: Reasoning in Multimodal Large Language Models (MLLMs) requires both fine-grained visual perception and rigorous logical deduction. Explicit text-based

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

BAP-SQL: Budget-Aware Observation Planning for Agentic Text-to-SQL

DGX agent

arXiv:2608.02876v1 Announce Type: new Abstract: Tool-using agents do not merely consume observations: their actions determine what arrives next. In agentic text-to-SQL, a broad query can spend context

safetyarxiv-cs-ai
5 Aug 2026
Agents

Before Reasoning Can Fail: Pre-Evidence Procedural Failures in Agentic RAG

DGX agent

arXiv:2608.02011v2 Announce Type: replace Abstract: Agentic retrieval-augmented generation (RAG) systems can fail before evidence-conditioned reasoning is tested: an agent may retrieve candidate snipp

agentsarxiv-cs-ai
5 Aug 2026
Research

Behaviorally Adaptive Visual Diversion for Inclusive and Resilient Digital Assessment Delivery

DGX agent

arXiv:2608.03531v1 Announce Type: new Abstract: Institutions increasingly rely on browser lockdown, webcam monitoring, and behavioral analytics to secure high-stakes digital assessments, yet these mec

researcharxiv-cs-ai
5 Aug 2026
Applications

Beyond Average Performance: Dynamic Instance Clustering and Specialized Algorithm Design in LLM-Assisted Evolutionary Search

DGX agent

arXiv:2608.03129v1 Announce Type: new Abstract: Large Language Model-assisted Evolutionary Search (LES) has emerged as a powerful paradigm for automated algorithm design. However, existing LES methods

applicationsarxiv-cs-ai
5 Aug 2026
Model Releases

Beyond Either-Or Reasoning: Transduction and Induction as Cooperative Problem-Solving Paradigms

DGX agent

arXiv:2505.14744v3 Announce Type: replace-cross Abstract: Traditionally, in Programming-by-example (PBE) the goal is to synthesize a program from a small set of input-output examples. Lately, PBE has

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Beyond Representational Similarity: Source-Conditioned Description-Length Gain for Generative Plagiarism Detection and Candidate Source Reranking

DGX agent

arXiv:2608.03859v1 Announce Type: cross Abstract: Large language models (LLMs) pose challenges to academic integrity and peer review. Yet generative plagiarism detection remains an underexplored and l

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Beyond the Hivemind: Escaping LLM Homogeneity via Meta-Persona Anchoring and Sequential Temperature Scaling

DGX agent

arXiv:2608.02618v1 Announce Type: new Abstract: Recent studies have identified an ``Artificial Hivemind'' effect in Large Language Models (LLMs) causing models to converge on a narrow, homogenized con

researcharxiv-cs-ai
5 Aug 2026
Safety

BODHI: Do LLMs Branch Out and Discover Heterogeneous Inferences?

DGX agent

arXiv:2608.02867v1 Announce Type: cross Abstract: Although reinforcement learning with verifiable rewards (RLVR) has improved the performance of large language models (LLMs) across a variety of reason

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

BulkPR-Bench: Benchmarking Queue-Level Governance of Interacting Pull Requests

DGX agent

arXiv:2608.02685v1 Announce Type: cross Abstract: Coding-agent benchmarks increasingly cover long-horizon, end-to-end, and interactive development, but typically retain one requested outcome or a fixe

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

CADET: Physics-Grounded Causal Auditing and Training-Free Deconfounding of End-to-End Driving Planners

DGX agent

arXiv:2606.14438v3 Announce Type: replace-cross Abstract: End-to-end (E2E) autonomous-driving planners trained by imitation are prone to statistical shortcuts: they associate scene elements that merel

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Can Large Language Models Recover Semantic Optimization Opportunities That Compilers Miss?

DGX agent

arXiv:2608.03983v1 Announce Type: cross Abstract: Optimizing compilers miss profitable transformations when their enabling semantics are absent from the analyzed program representation. We ask whether

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Can LLM design high-quality experiments? A Comprehensive and Systematic Benchmark on Autonomous Experimental Design

DGX agent

arXiv:2608.03501v1 Announce Type: new Abstract: AI for Research (AI4Research) leverages AI to automate and improve scientific workflows. While experimental design is a critical stage of the research p

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Can LLMs Test Terminal User Interfaces?

DGX agent

arXiv:2608.03743v1 Announce Type: cross Abstract: Terminal User Interfaces (TUIs) combine the stateful, screen-oriented behaviour of GUIs with terminal deployment and are now common in developer tools

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Can Training Logs Make Model Comparisons More Precise?

DGX agent

arXiv:2608.02705v1 Announce Type: cross Abstract: Comparing stochastically trained models requires estimating both a performance difference and its uncertainty from repeated runs. We study whether tra

researcharxiv-cs-ai
5 Aug 2026
Model Releases

CARE-Bench: Benchmarking Patient-Facing LLM Triage

DGX agent

arXiv:2608.03731v1 Announce Type: new Abstract: Patient-facing medical LLMs and agents increasingly answer symptom questions before clinician contact, where the key safety question is what action the

model-releasesarxiv-cs-ai
5 Aug 2026
Local Ai

CARE-X: Towards Clinically Useful Radiology VLMs with Auxiliary Supervision, Reward-Aligned Learning, and Tool-Augmented Measurement

DGX agent

arXiv:2608.03890v1 Announce Type: cross Abstract: A clinically useful chest X-ray system must go beyond fluent report generation: it should classify findings with tunable decision thresholds, localize

local-aiarxiv-cs-ai
5 Aug 2026
Agents

CastFSR: A Fast--Slow--Reflect Agentic Reasoning Framework for Context-Aware Time Series Forecasting

DGX agent

arXiv:2608.03031v1 Announce Type: new Abstract: Time series forecasting is fundamental to decision-making in complex systems, where future dynamics are influenced not only by historical observations b

agentsarxiv-cs-ai
5 Aug 2026
Model Releases

ChartAnno: Evaluating MLLMs for Chart Annotation Generation

DGX agent

arXiv:2608.03464v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have made significant progress in chart understanding, generation, and editing, but their ability to annotate e

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Chat Debugging: An Exploratory Study of Human-AI Collaboration to Debug Analog Circuits

DGX agent

arXiv:2608.02955v1 Announce Type: cross Abstract: This research paper describes an exploratory study on the effectiveness of Chat Debugging: troubleshooting malfunctioning analog circuits on breadboar

researcharxiv-cs-ai
5 Aug 2026
Model Releases

ChiEngMixBench: Evaluating Large Language Models on Expert-Style Chinese-English Terminology Mixing

DGX agent

arXiv:2601.16217v2 Announce Type: replace-cross Abstract: Large language models increasingly mediate multilingual professional communication, where useful generation requires adapting to community con

model-releasesarxiv-cs-ai
5 Aug 2026
Research

ChronoLens: Measuring Language Change Across Time, Languages, and Linguistic Levels

DGX agent

arXiv:2608.03507v1 Announce Type: cross Abstract: Historical language change affects morphology, syntax, semantics, and pragmatics, yet computational studies typically examine these levels with incomp

researcharxiv-cs-ai
5 Aug 2026
Research

CLIP-EBC: CLIP Can Count Accurately through Enhanced Blockwise Classification

DGX agent

arXiv:2403.09281v3 Announce Type: cross Abstract: We propose CLIP-EBC, the first fully CLIP-based model for accurate crowd density estimation. While the CLIP model has demonstrated remarkable success

researcharxiv-cs-ai
5 Aug 2026
Hardware

Compound and Parallel Modes of Tropical Convolutional Neural Networks

DGX agent

arXiv:2504.06881v2 Announce Type: replace-cross Abstract: Convolutional neural networks (CNNs) are foundational to many state-of-the-art computer vision systems, yet their reliance on multiplication-i

hardwarearxiv-cs-ai
5 Aug 2026
Applications

Computing Actual Causes for Neural Network Predictions under Structured Causal Inputs

DGX agent

arXiv:2608.03772v1 Announce Type: new Abstract: Explaining the predictions of neural networks is a central challenge in trustworthy AI. Existing explanation methods, such as those based on feature att

applicationsarxiv-cs-ai
5 Aug 2026
Agents

ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities?

DGX agent

arXiv:2608.03874v1 Announce Type: new Abstract: Modern agent frameworks equip large language models with external skill libraries to solve complex tasks. However, it remains unclear whether these syst

agentsarxiv-cs-ai
5 Aug 2026
Safety

Continue or Replan? Bernoulli-Continuation Policy Learning for Adaptive Horizon Execution

DGX agent

arXiv:2608.03483v1 Announce Type: cross Abstract: Existing chunk-based Vision-Language-Action (VLA) models execute a fixed number of actions (i.e., execution horizon) before replanning, turning replan

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

CorePath: A Breast-Specialized Pathology Foundation Model for Core Needle Biopsy Diagnosis and Risk-Controlled Report Generation

DGX agent

arXiv:2608.03079v1 Announce Type: cross Abstract: Breast core needle biopsy (CNB) is central to breast cancer diagnosis yet remains challenging because limited tissue sampling, lesion heterogeneity, a

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Cross-Anesthetic ECoG State Decoding Fails at the Decision Threshold, Not the Representation

DGX agent

arXiv:2608.02646v1 Announce Type: cross Abstract: Decoders of anesthetic state from cortical activity fail across drug classes, most notoriously ketamine, but reported accuracy cannot say whether the

researcharxiv-cs-ai
5 Aug 2026
Research

Cross-Layer Interaction under Weight-Space Ablation: A Closed-Form Attention Jacobian Bound and a Test on a Real Pretrained Model

DGX agent

arXiv:2608.03629v1 Announce Type: new Abstract: A companion paper studies when activation patching and weight-space ablation agree, inside an idealized model where a conditional computation is carried

researcharxiv-cs-ai
5 Aug 2026
Safety

CT-HEG: A Bidirectional, Timestamp-Attributed Event Graph for ICU In-Hospital Mortality Prediction - An Architectural Ablation Study

DGX agent

arXiv:2608.02663v1 Announce Type: cross Abstract: Accurate ICU mortality prediction requires modeling irregular clinical observations across heterogeneous entity types. Existing sequence models handle

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

CUADebug: Diagnosing and Repairing Computer-Use Agent Failures

DGX agent

arXiv:2608.02643v1 Announce Type: cross Abstract: Computer-use agents (CUAs) operate real desktop and web interfaces through screenshots, mouse and keyboard actions, and stateful UI feedback, yet thei

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Cura 1T: Specialized Model for Agentic Healthcare

DGX agent

arXiv:2607.15314v2 Announce Type: replace Abstract: Healthcare AI agents handle patient consultation, clinical reasoning over text and images, interactive diagnosis, and electronic health record (EHR)

model-releasesarxiv-cs-ai
5 Aug 2026
Applications

CURV: Enhancing Chart Understanding Through Curriculum Visual Grounded Reasoning

DGX agent

arXiv:2608.02833v1 Announce Type: cross Abstract: Chart question answering (CQA) requires multimodal large language models (MLLMs) to integrate visual comprehension with logical reasoning, yet current

applicationsarxiv-cs-ai
5 Aug 2026
Safety

CVPO: Enhancing LLM Reinforcement Learning Reasoning via Value-Variance Adaptation and Dynamic Curriculum Learning

DGX agent

arXiv:2608.03068v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as an effective method for enhancing the reasoning capabilities of large language models (LLMs). However, exis

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces

DGX agent

arXiv:2608.03451v1 Announce Type: new Abstract: Data agents enable natural-language analytics over organizational workspaces, where relevant evidence may be scattered across databases, structured file

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Decoupling Generation and Selection for Budget-Constrained Faithful Summarization

DGX agent

arXiv:2608.03655v1 Announce Type: cross Abstract: Abstractive summarization models remain vulnerable to factual inconsistency, redundancy, and weak length control. We propose a modular generation-and-

researcharxiv-cs-ai
5 Aug 2026
Research

Deep Divide-and-Reduce in Symbolic Regression

DGX agent

arXiv:2608.02628v1 Announce Type: cross Abstract: Symbolic regression (SR) is the task of discovering underlying patterns from data and representing them using mathematical expressions. Current machin

researcharxiv-cs-ai
5 Aug 2026
Safety

Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs

DGX agent

arXiv:2608.01755v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving (AD) increasingly utilize chain-of-thought (CoT) supervision to enhance the reason

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

DenialRAG: Single-Document RAG Poisoning via Embedded Parametric Denial

DGX agent

arXiv:2608.02678v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are vulnerable to corpus poisoning: an attacker who inserts a crafted document into the retrieval corpus

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Designing a Good Virtual Node: Addressable and Cardinality-Preserving Global Memory for Message Passing Architectures

DGX agent

arXiv:2608.02709v1 Announce Type: cross Abstract: Virtual nodes give message-passing neural networks a simple global communication route, but the standard node--VN--node pipeline compresses the graph

researcharxiv-cs-ai
5 Aug 2026
Model Releases

DiagChain: A Diagnostic Benchmark for Evaluating LLM Agents on Evidence-Grounded Attack Chain Reconstruction

DGX agent

arXiv:2608.03591v1 Announce Type: cross Abstract: Large Language Model (LLM) agents offer a promising approach to attack chain reconstruction by retrieving and interpreting heterogeneous telemetry to

model-releasesarxiv-cs-ai
5 Aug 2026
Tutorials

DiffImaginE: Imagine to Verify Entity Types with Diffusio

DGX agent

arXiv:2608.03025v1 Announce Type: new Abstract: Multimodal named entity recognition (MNER) determines whether each candidate span and entity-type hypothesis is supported by joint textual and visual ev

tutorialsarxiv-cs-ai
5 Aug 2026
Research

DigitCode: Symbolic Tokenization of Hand Motion by Anatomical Units

DGX agent

arXiv:2608.03127v1 Announce Type: cross Abstract: Hand motion carries the finest-grained information in human activity, yet the representations behind hand generation, understanding, and robot learnin

researcharxiv-cs-ai
5 Aug 2026
Research

Distilled Roads: Generalisable Road Network Extraction Across Sensors, Resolutions, and Region

DGX agent

arXiv:2608.03407v1 Announce Type: cross Abstract: Road network segmentation from satellite imagery remains challenging due to large geographic variation in road appearance, occlusions, and domain shif

researcharxiv-cs-ai
5 Aug 2026
Model Releases

Distractor-Aware Truncation: Disentangling Context-Length Effects from Signal Loss in Long-Context LLM Benchmarks

DGX agent

arXiv:2608.03297v1 Announce Type: new Abstract: A standard claim in the literature on retrieval-augmented and memory-augmented language models is that shorter context is better when the relevant infor

model-releasesarxiv-cs-ai
5 Aug 2026
← Previous
1…3839404142…443
Next →