AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
12 May 2026

The Safety-Aware Denoiser for Text Diffusion Models

SafetyDGX agent

arXiv:2605.08116v1 Announce Type: cross Abstract: Recent work on text diffusion models offers a promising alternative to autoregressive generation, but controlling their safety remains underexplored.

The Silent Vote: Improving Zero-Shot LLM Reliability by Aggregating Semantic Neighborhoods

Model ReleasesDGX agent

arXiv:2605.09739v1 Announce Type: cross Abstract: Large Language Models are increasingly used as zero-shot classifiers in complex reasoning tasks. However, standard constrained decoding suffers from a

The Trap of Trajectory: Towards Understanding and Mitigating Spurious Correlations in Agentic Memory

Model ReleasesDGX agent

arXiv:2605.09330v1 Announce Type: cross Abstract: Agentic memory enables LLMs to persist information beyond a single context window and reuse it in later decisions, but it also introduces a new vulner


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The two clocks and the innovation window: When and how generative models learn rules

TutorialsDGX agent

arXiv:2605.10019v1 Announce Type: cross Abstract: Generative models trained on finite data face a fundamental tension: their score-matching or next-token objective converges to the empirical training

The Wittgensteinian Representation Hypothesis: Is Language the Attractor of Multimodal Convergence?

SafetyDGX agent

arXiv:2605.09352v1 Announce Type: new Abstract: Understanding why independently trained neural networks from different modalities converge toward shared representations, and where this convergence lea

The World is Not Mono: Enabling Spatial Understanding in Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2601.02954v3 Announce Type: replace-cross Abstract: Large audio-language models have made rapid progress in recognizing what is present in an audio clip, but spatial audio-language understanding

Think as Needed: Geometry-Driven Adaptive Perception for Autonomous Driving

AgentsDGX agent

arXiv:2605.10117v1 Announce Type: cross Abstract: Autonomous driving scenes range from empty highways to dense intersections with dozens of interacting road users, yet current 3D detection models appl

Threat Modelling using Domain-Adapted Language Models: Empirical Evaluation and Insights

ResearchDGX agent

arXiv:2605.10808v1 Announce Type: cross Abstract: Large Language Models(LLMs) are increasingly explored for cybersecurity applications such as vulnerability detection. In the domain of threat modellin

ThreatCore: A Benchmark for Explicit and Implicit Threat Detection

Model ReleasesDGX agent

arXiv:2605.10563v1 Announce Type: cross Abstract: Threat detection in Natural Language Processing lacks consistent definitions and standardized benchmarks, and is often conflated with broader phenomen

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning

Model ReleasesDGX agent

arXiv:2605.09544v1 Announce Type: new Abstract: Tool-integrated reasoning has emerged as a promising paradigm for enhancing large language models with external computation, retrieval, and execution ca

TIDES: Implicit Time-Awareness in Selective State Space Models

Model ReleasesDGX agent

arXiv:2605.09742v1 Announce Type: cross Abstract: Selective state space models (SSMs), such as Mamba, achieve strong per-token expressivity by making the time discretization step Tilde{Delta} a learne

TiledAttention: a CUDA Tile SDPA Kernel for PyTorch

Model ReleasesDGX agent

arXiv:2603.01960v2 Announce Type: replace-cross Abstract: TiledAttention is a scaled dot-product attention (SDPA) forward operator for SDPA research on NVIDIA GPUs. Implemented in cuTile Python (TileI

TimeClaw: A Time-Series AI Agent with Exploratory Execution Learning

AgentsDGX agent

arXiv:2605.10038v1 Announce Type: new Abstract: Time series analysis underpins forecasting, monitoring, and decision making in domains such as finance and weather, where solving a task often requires

TinySSL: Distilled Self-Supervised Pretraining for Sub-Megabyte MCU Models

ResearchDGX agent

arXiv:2605.08241v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has transformed representation learning for large models, yet remains unexplored for microcontroller (MCU)-class models

TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit

AgentsDGX agent

arXiv:2507.09788v2 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLM) have led to a new class of autonomous agents, renewing and expanding interest in the area. LLM-

TMAS: Scaling Test-Time Compute via Multi-Agent Synergy

AgentsDGX agent

arXiv:2605.10344v1 Announce Type: new Abstract: Test-time scaling has become an effective paradigm for improving the reasoning ability of large language models by allocating additional computation dur

To Redact, or not to Redact? A Local LLM Approach to Deliberative Process Privilege Classification

Model ReleasesDGX agent

arXiv:2605.10211v1 Announce Type: cross Abstract: Government transparency laws, like the Freedom of Information (FOIA) acts in the United States and United Kingdom, and the Woo (Open Government Act) i

TodyComm: Task-Oriented Dynamic Communication for Multi-Round LLM-based Multi-Agent System

SafetyDGX agent

arXiv:2602.03688v2 Announce Type: replace Abstract: Multi-round LLM-based multi-agent systems rely on effective communication structures to support collaboration across rounds. However, most existing

Token Economics for LLM Agents: A Dual-View Study from Computing and Economics

AgentsDGX agent

arXiv:2605.09104v1 Announce Type: new Abstract: As LLM agents evolve, tokens have emerged as the core economic primitives of Agentic AI. However, their exponential consumption introduces severe comput

Top-H Decoding: Adapting the Creativity and Coherence with Bounded Entropy in Text Generation

ResearchDGX agent

arXiv:2509.02510v2 Announce Type: replace-cross Abstract: Large language models (LLMs), despite their impressive performance across a wide range of tasks, often struggle to balance two competing objec

Toward an Engineering of Science: Rebalancing Generation and Verification in the Age of AI

ResearchDGX agent

arXiv:2605.10425v1 Announce Type: cross Abstract: AI systems can now cheaply generate plausible scientific artifacts such as papers, reviews, and surveys. This creates a risk of epistemic pollution in

Toward Optimal Regret in Robust Pricing: Decoupling Corruption and Time

ResearchDGX agent

arXiv:2605.08290v1 Announce Type: cross Abstract: We design the first regret guarantees for robust dynamic pricing that decouple the dependence on the corruption C and the time horizon T. In dynamic p

Towards a Large Language-Vision Question Answering Model for MSTAR Automatic Target Recognition

Model ReleasesDGX agent

arXiv:2605.10772v1 Announce Type: cross Abstract: Large language-vision models (LLVM), such as OpenAI's ChatGPT and GPT-4, have gained prominence as powerful tools for analyzing text and imagery. The

Towards a Virtual Neuroscientist: Autonomous Neuroimaging Analysis via Multi-Agent Collaboration

AgentsDGX agent

arXiv:2605.09366v1 Announce Type: new Abstract: Transforming neuroimaging data into clinically actionable biomarkers is a knowledge-intensive and labor-intensive process. Standardized workflows such a

Towards Autonomous Railway Operations: A Semi-Hierarchical Deep Reinforcement Learning Approach to the Vehicle Rescheduling Problem

AgentsDGX agent

arXiv:2605.10257v1 Announce Type: new Abstract: Managing disruptions in railway traffic management is a major challenge. Rising traffic density and infrastructure limits increase complexity, making th

Towards Backdoor-Based Ownership Verification for Vision-Language-Action Models

ResearchDGX agent

arXiv:2605.09005v1 Announce Type: cross Abstract: Vision-Language-Action models (VLAs) support generalist robotic control by enabling end-to-end decision policies directly from multi-modal inputs. As

Towards Batch-to-Streaming Deep Reinforcement Learning for Continuous Control

Local AiDGX agent

arXiv:2603.08588v2 Announce Type: replace-cross Abstract: State-of-the-art deep reinforcement learning (RL) methods have achieved remarkable performance in continuous control tasks, yet their computat

Towards Conversational Medical AI with Eyes, Ears and a Voice

Model ReleasesDGX agent

arXiv:2605.09272v1 Announce Type: new Abstract: The practice of medicine relies not only upon skillful dialogue but also on the nuanced exchange and interpretation of rich auditory and visual cues bet

Towards Cross-lingual Values Judgment: A Consensus-Pluralism Perspective

Model ReleasesDGX agent

arXiv:2602.17283v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are employed worldwide, existing evaluation paradigms for their multilingual capabilities primarily focus on f

Towards Effective Theory of LLMs: A Representation Learning Approach

ResearchDGX agent

arXiv:2605.09294v1 Announce Type: cross Abstract: We propose Representational Effective Theory (RET), a framework for describing large language model computation in terms of learned macrostates rather

Towards Robust Sequential Decomposition for Complex Image Editing

ApplicationsDGX agent

arXiv:2605.09233v1 Announce Type: cross Abstract: Recent advances in visual generative models have enabled high-fidelity image editing guided by human instructions. However, these models often struggl

Towards Shutdownable Agents: Generalizing Stochastic Choice in RL Agents and LLMs

Model ReleasesDGX agent

arXiv:2604.17502v2 Announce Type: replace Abstract: Misaligned artificial agents might resist shutdown. One proposed solution is to train agents to lack preferences between different-length trajectori

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm

TutorialsDGX agent

arXiv:2605.10640v1 Announce Type: cross Abstract: Continual Pre-Training (CPT) is essential for enabling Language Models (LMs) to integrate new knowledge without erasing old. While classical CPT techn

Towards Universal Gene Regulatory Network Inference: Unlocking Generalizable Regulatory Knowledge in Single-cell Foundation Models

Model ReleasesDGX agent

arXiv:2605.08128v1 Announce Type: cross Abstract: Gene Regulatory Network (GRN) inference is essential for understanding complex cellular mechanisms, rendered tractable through single-cell transcripto

TRACE: Distilling Where It Matters via Token-Routed Self On-Policy Alignment

SafetyDGX agent

arXiv:2605.10194v1 Announce Type: new Abstract: On-policy self-distillation (self-OPD) densifies reinforcement learning with verifiable rewards (RLVR) by letting a policy teach itself under privileged

Tracing Moral Foundations in Large Language Models

Model ReleasesDGX agent

arXiv:2601.05437v2 Announce Type: replace-cross Abstract: Large language models often produce human-like moral judgments, but it is unclear whether this reflects an internal conceptual structure or su

Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models

Model ReleasesDGX agent

arXiv:2605.08974v1 Announce Type: cross Abstract: While multimodal large language models (MLLMs) have advanced video understanding, they remain highly prone to hallucinations in dynamic scenes. We arg

Training-Free Cultural Alignment of Large Language Models via Persona Disagreement

SafetyDGX agent

arXiv:2605.10843v1 Announce Type: cross Abstract: Large language models increasingly mediate decisions that turn on moral judgement, yet a growing body of evidence shows that their implicit preference

Training Reasoning Models on Saturated Problems via Failure-Prefix Conditioning

ResearchDGX agent

arXiv:2601.20829v2 Announce Type: replace-cross Abstract: As Reinforcement Learning with Verifiable Rewards (RLVR) substantially improves the reasoning abilities of large language models (LLMs), a new

Trajectory Supervision for Continual Tool-Use Learning in LLMs

Model ReleasesDGX agent

arXiv:2605.09734v1 Announce Type: cross Abstract: Most language-model training data shows final artifacts, not the process that produced them. We study a tractable version of this question in tool use

TrajPrism: A Multi-Task Benchmark for Language-Grounded Urban Trajectory Understanding

Model ReleasesDGX agent

arXiv:2605.10782v1 Announce Type: new Abstract: Urban mobility is naturally expressed both as trajectories in space and as natural-language descriptions of travel intent, constraints, and preferences.

TRAM: Training Approximate Multiplier Structures for Low-Power AI Accelerators

ResearchDGX agent

arXiv:2605.08231v1 Announce Type: cross Abstract: Reducing power consumption in AI accelerators is increasingly important. Approximate computing can reduce power consumption while keeping the accuracy

Transformer autoencoder with local attention for sparse and irregular time series with application on risk estimation

ApplicationsDGX agent

arXiv:2605.08914v1 Announce Type: cross Abstract: This paper introduces a framework specifically designed for sparse and irregular time series {risk estimation}. It is based on a Transformer Autoencod

Transformers Can Implement Preconditioned Richardson Iteration for In-Context Gaussian Kernel Regression

ResearchDGX agent

arXiv:2605.08475v1 Announce Type: cross Abstract: Mechanistic accounts of in-context learning (ICL) have identified iterative algorithms for linear regression and related linear prediction tasks, ofte

Trapping Attacker in Dilemma: Examining Internal Correlations and External Influences of Trigger for Defending GNN Backdoors

ResearchDGX agent

arXiv:2605.08278v1 Announce Type: cross Abstract: GNNs have become a standard tool for learning on relational data, yet they remain highly vulnerable to backdoor attacks. Prior defenses often depend o

TTCD:Transformer Integrated Temporal Causal Discovery from Non-Stationary Time Series Data

Model ReleasesDGX agent

arXiv:2605.08111v1 Announce Type: cross Abstract: The widespread availability of complex time series data in various domains such as environmental science, epidemiology, and economics demands robust c

UFO: A Unified Flow-Oriented Framework for Robust Continual Graph Learning

Model ReleasesDGX agent

arXiv:2605.09862v1 Announce Type: cross Abstract: Graph learning research has increasingly shifted toward continual graph learning (CGL), which better reflects real-world scenarios where graphs evolve

UMEDA: Unified Multi-modal Efficient Data Fusion for Privacy-Preserving Graph Federated Learning via Spectral-Gated Attention and Diffusion-Based Operator Alignment

Model ReleasesDGX agent

arXiv:2605.08288v1 Announce Type: cross Abstract: Device-free localization trains models from heterogeneous wireless and visual sensors (e.g., Wi-Fi, LiDAR) distributed across edge devices. Federated

Uncovering Intra-expert Activation Sparsity for Efficient Mixture-of-Expert Model Execution

ResearchDGX agent

arXiv:2605.08575v1 Announce Type: cross Abstract: Mixture of Experts (MoE) architecture has become the standard for state-of-the-art large language models, owing to its computational efficiency throug

Understand and Accelerate Memory Processing Pipeline for Disaggregated LLM Inference

HardwareDGX agent

arXiv:2603.29002v2 Announce Type: replace-cross Abstract: Modern large language models (LLMs) increasingly depends on efficient long-context processing and generation mechanisms, including sparse atte

Understanding Asynchronous Inference Methods for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.08168v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models offer a promising path to generalist robot control, but their inference latency causes observation staleness when

Uniform Inductive Spatio-Temporal Kriging

SafetyDGX agent

arXiv:2603.05301v2 Announce Type: replace Abstract: Inductive spatio-temporal kriging infers signals at unobserved locations from observed sensors, but real-world observations are often incomplete and

Unifying Perspectives: Plausible Counterfactual Explanations on Global, Group-wise, and Local Levels

ResearchDGX agent

arXiv:2405.17642v3 Announce Type: replace-cross Abstract: The growing complexity of AI systems has intensified the need for transparency through Explainable AI (XAI). Counterfactual explanations (CFs)

Unlearners Can Lie: Evaluating and Improving Honesty in LLM Unlearning

SafetyDGX agent

arXiv:2605.08765v1 Announce Type: cross Abstract: Unlearning in large language models (LLMs) aims to remove harmful training data while preserving overall utility. However, we find that existing metho

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why

Model ReleasesDGX agent

arXiv:2605.10889v1 Announce Type: cross Abstract: On-policy distillation offers dense, per-token supervision for training reasoning models; however, it remains unclear under which conditions this sign

Unpredictability dissociates from structured control in language agents

Model ReleasesDGX agent

arXiv:2605.09692v1 Announce Type: new Abstract: Unpredictable behavior is often taken as evidence of control, yet stochastic dispersion and structured action control need not coincide. This paper test

Upholding Epistemic Agency: A Brouwerian Assertibility Constraint for Responsible AI

SafetyDGX agent

arXiv:2603.03971v2 Announce Type: replace-cross Abstract: Generative AI can convert uncertainty into hypersuasive, authoritative-seeming verdicts, displacing the justificatory work on which democratic

Useful for Exploration, Risky for Precision: Evaluating AI Tools in Academic Research

ResearchDGX agent

arXiv:2605.10125v1 Announce Type: new Abstract: Artificial intelligence (AI) tools are being incorporated into scientific research workflows with the potential to enhance efficiency in tasks such as d

Users as Annotators: LLM Preference Learning from Comparison Mode

SafetyDGX agent

arXiv:2510.13830v2 Announce Type: replace-cross Abstract: Pairwise preference data have played an important role in the alignment of large language models (LLMs). Each sample of such data consists of

UTS at PsyDefDetect: Multi-Agent Councils and Absence-Based Reasoning for Defense Mechanism Classification

Model ReleasesDGX agent

arXiv:2605.09769v1 Announce Type: new Abstract: This paper describes our system for classifying psychological defense mechanisms in emotional support dialogues using the Defense Mechanism Rating Scale

← Previous
1…275276277278279…358
Next →