AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
25 May 2026

HTMuon: Improving Muon via Heavy-Tailed Spectral Correction

Model ReleasesDGX agent

arXiv:2603.10067v2 Announce Type: replace-cross Abstract: Muon has recently shown promising results in LLM training. In this work, we study how to further improve Muon. We argue that Muon's orthogonal

Human-Centered Learning Mechanics: A Dynamical Framework for Entropy-Regulated Representation Learning

Model ReleasesDGX agent

arXiv:2605.22940v1 Announce Type: cross Abstract: Deep learning is increasingly viewed as a dynamical process in parameter space, yet many existing theories still treat training as a closed optimizati

Human Decision-Making with Persuasive and Narrative LLM Explanations

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.23867v1 Announce Type: cross Abstract: Large language models (LLMs) have the potential to aid and improve human decision-making in classification tasks, not only by providing fairly accurat

Human-in-the-Loop Multi-Agent Ventilator Decision Support with Contextual Bandit Preference Learning

SafetyDGX agent

arXiv:2605.23320v1 Announce Type: new Abstract: Ventilator decision support requires sequential decisions that track evolving physiology and disease trajectories while respecting safety boundaries and

ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization

Model ReleasesDGX agent

arXiv:2605.22885v1 Announce Type: new Abstract: Formal mathematics libraries are rapidly expanding, creating a growing need to refactor verified proofs for maintainability and to improve training data

Inductive Deductive Synthesis: Enabling AI to Generate Formally Verified Systems

Model ReleasesDGX agent

arXiv:2605.23109v1 Announce Type: new Abstract: AI agents increasingly excel at generating, testing, and refining code. However, they fall short on tasks requiring formal guarantees of full coverage t

Information Access of the Oppressed: Freirean Design for Emancipatory Information Access

SafetyDGX agent

arXiv:2601.09600v3 Announce Type: replace-cross Abstract: Online information access (IA) platforms are targets of authoritarian capture. We explore the question of how to safeguard our platforms and e

Infra-Bayesian Reinforcement Learning Agents Outperform Classical RL For Worst-Case Robustness

SafetyDGX agent

arXiv:2605.23146v1 Announce Type: cross Abstract: Classical reinforcement learning assumes the agent interacts with a fixed environment whose behavior does not depend on the agent's policy. This assum

IntentScore: Intent-Conditioned Action Evaluation for Computer-Use Agents

SafetyDGX agent

arXiv:2604.05157v2 Announce Type: replace Abstract: Computer-Use Agents (CUAs) leverage large language models to execute GUI operations on desktop environments, yet they generate actions without evalu

Interactive Query Answering on Knowledge Graphs with Soft Entity Constraints

ApplicationsDGX agent

arXiv:2508.13663v5 Announce Type: replace Abstract: Methods for query answering over incomplete knowledge graphs retrieve entities that are likely to be answers, which is particularly useful when such

Investigating Robot Control Policy Learning for Autonomous X-ray-guided Spine Procedures

Model ReleasesDGX agent

arXiv:2511.03882v2 Announce Type: replace-cross Abstract: Imitation learning-based robot control policies are enjoying renewed interest in video-based robotics. However, it remains unclear whether thi

Is Capability a Liability? More Capable Language Models Make Worse Forecasts When It Matters Most

Model ReleasesDGX agent

arXiv:2605.22672v2 Announce Type: replace Abstract: We document inverse scaling in LLMs on forecasting problems whose underlying time series exhibit superlinear growth and tail risk of regime change,

It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amplified by the language of the prompt

Model ReleasesDGX agent

arXiv:2605.23825v1 Announce Type: cross Abstract: It has generally been assumed that geopolitical bias in language models originates from the training data used during the pre-training phase. We teste

KAPLAN: Kolmogorov-Arnold Prognostic Learnable Activation Networks for Survival Analysis

Model ReleasesDGX agent

arXiv:2605.23082v1 Announce Type: cross Abstract: Survival analysis aims to model how covariates and time jointly shape the time-to-event distribution under right censoring. Classical methods such as

KPI2KVI: A Multi Agent Workflow for Calculating Key Value Indicators from Service Descriptions

AgentsDGX agent

arXiv:2605.22825v1 Announce Type: cross Abstract: Key Value Indicators (KVIs) provide a decision oriented view of a service by summarizing how operational performance translates into stakeholder value

LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation

AgentsDGX agent

arXiv:2511.02239v2 Announce Type: replace-cross Abstract: Learning generalizable policies for robotic manipulation increasingly relies on large-scale models that map language instructions to actions (

Learning Individual Dynamics from Sparse Cross-Sectional Snapshots

ApplicationsDGX agent

arXiv:2605.23470v1 Announce Type: cross Abstract: Predicting how a dynamical unit evolves over time - how an individual ages, an epidemic spreads, or a physical system degrades - typically requires de

Learning Through Noise: Why Subliminal Learning Works and When It Fails

ResearchDGX agent

arXiv:2605.23645v1 Announce Type: cross Abstract: In the context of artificial neural networks, subliminal learning refers to the transfer of task-relevant knowledge or unintended biases from teacher

Leveraging Foundation Models for Causal Generative Modeling

ResearchDGX agent

arXiv:2605.23861v1 Announce Type: cross Abstract: Causal generative modeling is essential for developing reliable and transparent AI systems capable of counterfactual reasoning. While existing approac

LFRAG: Layout-oriented Fine-grained Retrieval-Augmented Generation on Multimodal Document Understanding

Model ReleasesDGX agent

arXiv:2605.22829v1 Announce Type: cross Abstract: Multimodal Retrieval-Augmented Generation (RAG) has emerged as an effective paradigm for enhancing Large Language Models (LLMs) with external knowledg

Lipschitz Optimization for Formal Verification of Homographies

Model ReleasesDGX agent

arXiv:2605.23203v1 Announce Type: cross Abstract: The adoption of vision neural networks in regulated industries requires formal robustness guarantees, especially in safety-critical domains such as he

LLM Code Smells: A Taxonomy and Detection Approach

ResearchDGX agent

arXiv:2605.22976v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly integrated into software systems for diverse purposes, due to their versatility, flexibility, and abilit

LLMs as Noisy Channels: A Shannon Perspective on Model Capacity and Scaling Laws

ResearchDGX agent

arXiv:2605.23901v1 Announce Type: cross Abstract: Existing scaling laws for Large Language Models (LLMs), predominantly monotonic power laws, fail to explain emerging non-monotonic phenomena such as c

MadEvolve: Evolutionary Optimization of Trading Systems with Large Language Models

Model ReleasesDGX agent

arXiv:2605.23007v1 Announce Type: cross Abstract: We explore the application of LLM-driven algorithm optimization to several common tasks in quantitative finance. MadEvolve, a general-purpose algorith

MAS-Orchestra: Understanding and Improving Multi-Agent Reasoning Through Holistic Orchestration and Controlled Benchmarks

Model ReleasesDGX agent

arXiv:2601.14652v5 Announce Type: replace Abstract: While multi-agent systems (MAS) promise elevated intelligence through coordination of agents, current approaches to automatic MAS design under-deliv

MedExpMem: Adapting Experience Memory for Differential Diagnosis

Model ReleasesDGX agent

arXiv:2605.22872v1 Announce Type: cross Abstract: Experienced physicians develop diagnostic expertise through clinical practice, acquiring not only disease knowledge but also the ability to differenti

Mediative Fuzzy Logic: From Type-1 Foundations to Type-2, Type-3 and Quantum Extensions

SafetyDGX agent

arXiv:2605.22900v1 Announce Type: new Abstract: Mediative Fuzzy Logic was conceived as a practical scheme for reconciling hesitant or conflicting assessments in fuzzy control and decision-making. Howe

MedSAE: Dissecting MedCLIP Representations with Sparse Autoencoders

ApplicationsDGX agent

arXiv:2510.26411v2 Announce Type: replace Abstract: Artificial intelligence in healthcare requires models that are accurate and interpretable. We advance mechanistic interpretability in medical vision

MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structural Anomaly Detection

AgentsDGX agent

arXiv:2605.23723v1 Announce Type: new Abstract: Large language model agents increasingly rely on persistent memory to store past interactions, retrieve relevant demonstrations, and improve long-horizo

Memorization Dynamics of Fill-in-the-Middle Pretraining

Model ReleasesDGX agent

arXiv:2605.22981v1 Announce Type: cross Abstract: Fill-in-the-middle (FIM) is a pretraining objective widely used to equip causal language models with infilling ability, yet its effect on verbatim mem

Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals

ResearchDGX agent

arXiv:2605.23384v1 Announce Type: cross Abstract: Recent RL methods have substantially improved the reasoning abilities of LLMs. Existing reward designs mainly follow two paradigms: (1) Reinforcement

MirrorCheck: Efficient Adversarial Defense for Vision-Language Models

ResearchDGX agent

arXiv:2406.09250v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly susceptible to sophisticated adversarial attacks, including adaptive strategies specifically de

Model Collapse as Cultural Evolution

Model ReleasesDGX agent

arXiv:2605.23054v1 Announce Type: cross Abstract: Model collapse, the progressive degradation of LLMs trained on their own outputs, has been characterized statistically but lacks a linguistic explanat

Moonwalk: Inverse-Forward Differentiation

Model ReleasesDGX agent

arXiv:2402.14212v4 Announce Type: replace-cross Abstract: Backpropagation's main limitation is its need to store intermediate activations (residuals) during the forward pass, which restricts the depth

Multi-Gate Residuals

ResearchDGX agent

arXiv:2605.23259v1 Announce Type: cross Abstract: While Attention Residuals has shown some effectiveness in addressing the widespread issue of unbounded activation growth across deep residual layers,

Multimodal Crystal Flow: Any-to-Any Modality Generation for Unified Crystal Modeling

ResearchDGX agent

arXiv:2602.20210v2 Announce Type: replace-cross Abstract: Crystal modeling spans a family of conditional and unconditional generation tasks, including crystal structure prediction (CSP) and de novo ge

Multimodal Distribution Matching for Vision-Language Dataset Distillation

SafetyDGX agent

arXiv:2605.23482v1 Announce Type: cross Abstract: Dataset distillation compresses large training sets into compact synthetic datasets while preserving downstream performance. As modern systems increas

MUSEKG: A Knowledge Graph Over Museum Collections

ResearchDGX agent

arXiv:2511.16014v2 Announce Type: replace Abstract: Digitisation in the cultural heritage sector has produced large but fragmented repositories of museum collection data, spanning structured catalogue

NeuroNL2LTL: A Neurosymbolic Framework for Natural Language Translation of Linear Temporal Logic

SafetyDGX agent

arXiv:2605.22874v1 Announce Type: new Abstract: Effectively translating between natural language (NL) and formal logics like Linear Temporal Logic (LTL) requires expertise that limits formal verificat

NeuroWeaver: An Autonomous Evolutionary Agent for Exploring the Programmatic Space of EEG Analysis Pipelines

AgentsDGX agent

arXiv:2602.13473v2 Announce Type: replace Abstract: Although foundation models have demonstrated remarkable success in general domains, the application of these models to electroencephalography (EEG)

Not Too Generative, Not Too Discriminative: The Human Alignment Sweet Spot

SafetyDGX agent

arXiv:2605.23819v1 Announce Type: cross Abstract: A central question in computational vision is whether human-like visual representations are better explained by discriminative or generative learning.

ObjectCache: Layerwise Object-Storage Retrieval for KV Cache Reuse

Local AiDGX agent

arXiv:2605.22850v1 Announce Type: cross Abstract: Prefix KV caching has become a key mechanism in LLM serving: it reduces time to first token (TTFT) by avoiding redundant computation across requests t

On the Infinite Width and Depth Limits of Predictive Coding Networks

ResearchDGX agent

arXiv:2602.07697v2 Announce Type: replace-cross Abstract: Predictive coding (PC) is a biologically plausible alternative to standard backpropagation (BP) that minimises an energy function with respect

On the Koopman-Based Generalization Bounds for Multi-Task Deep Learning

ResearchDGX agent

arXiv:2512.19199v2 Announce Type: replace-cross Abstract: The paper establishes generalization bounds for multitask deep neural networks using operator-theoretic techniques. The authors propose a tigh

One-Forcing: Towards Stable One-Step Autoregressive Video Generation

ResearchDGX agent

arXiv:2605.23458v1 Announce Type: cross Abstract: Recent advances have substantially improved real-time interactive video generation in the autoregressive regime. However, most existing few-step autor

One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents

Model ReleasesDGX agent

arXiv:2605.23652v1 Announce Type: new Abstract: On a 300-persona life-simulation benchmark, pcsp achieves compositional zero-shot persona identification up to 17x above chance, Spearman rho approx 0.7

OnePred: Next-Query Prediction via Recursive Intent Memory in Multi-Turn Conversations

ResearchDGX agent

arXiv:2605.23668v1 Announce Type: cross Abstract: Although large language model (LLM) conversational systems process millions of multi-turn dialogues daily, they remain fundamentally reactive: they re

Online Hand Gesture Recognition Using 3D Convolutional Neural Networks

ResearchDGX agent

arXiv:2605.23409v1 Announce Type: cross Abstract: In human computer interaction, real-time detection and classification of dynamic hand gestures is challenging as: 1) the system must run in a real-tim

Online Learning with Multiple Fairness Regularizers via Graph-Structured Feedback

SafetyDGX agent

arXiv:2508.14311v2 Announce Type: replace-cross Abstract: There is an increasing need to enforce multiple, often competing, measures of fairness within automated decision systems. The appropriate weig

Ontological Knowledge Blocks: Executable Compliance and Profile-Based Validation for Trustworthy AI Systems

Model ReleasesDGX agent

arXiv:2605.23297v1 Announce Type: new Abstract: AI-enabled services deployed in critical digital infrastructure are subject to governance obligations spanning transparency, accountability, fairness, a

Operator-Based Generalization Bound for Deep Learning: Insights on Multi-Task Learning

ResearchDGX agent

arXiv:2512.19184v2 Announce Type: replace-cross Abstract: This paper presents novel generalization bounds for vector-valued neural networks and deep kernel methods, focusing on multi-task learning thr

PaP-NF: Probabilistic Long-Term Time Series Forecasting via Prefix-as-Prompt Reprogramming and Normalizing Flows

ApplicationsDGX agent

arXiv:2605.23219v1 Announce Type: cross Abstract: Time series forecasting plays a central role in many real-world applications and has been extensively studied. Most existing approaches rely on determ

Parallel Context Compaction for Long-Horizon LLM Agent Serving

Model ReleasesDGX agent

arXiv:2605.23296v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate growing conversation histories that eventually exceed the model's context window. Context compaction via LLM-based su

Parametric Prior Mapping Framework for Non-stationary Probabilistic Time Series Forecasting

ResearchDGX agent

arXiv:2605.23402v1 Announce Type: cross Abstract: Effectively modeling non-stationary dynamics in probabilistic multivariate time series(MTS) forecasting requires balancing expressiveness with robustn

PathCal: State-Aware Reflection-Marker Calibration for Efficient Reasoning

ResearchDGX agent

arXiv:2605.23074v1 Announce Type: new Abstract: The emergence of Large Reasoning Language Models (LRMs) has paved the way for tackling complex reasoning tasks through test-time scaling by generating l

PathNavigate: A Training-Free Pathology Agent with Surprise-Guided Scan and Shared Slide Memory for Whole-Slide Image VQA

Local AiDGX agent

arXiv:2605.23559v1 Announce Type: cross Abstract: Whole-slide image visual question answering (WSI-VQA) frames pathology as an extreme-context search problem: to answer a free-form clinical query, a s

PGT: Procedurally Generated Tasks for improving visual grounding in MLLMs

Model ReleasesDGX agent

arXiv:2605.23883v1 Announce Type: cross Abstract: Despite remarkable progress in Multimodal Large Language Models (MLLMs), these models still struggle with fine-grained understanding tasks. In this wo

PhenoYieldNet: Learning Crop-Aware Phenological Responses for Multi-Crop Yield Prediction

TutorialsDGX agent

arXiv:2605.23478v1 Announce Type: cross Abstract: Accurate crop yield prediction is crucial for sustainable agriculture and global food security. While existing methods are predominantly developed for

Philosophical Dispositions as Behavioral Constraints for AI-Assisted Code Review: An Empirical Study

Model ReleasesDGX agent

arXiv:2605.23108v1 Announce Type: cross Abstract: AI-assisted code review tools typically operate as generic 'expert reviewer' agents, producing homogeneous findings regardless of the analysis type ne

PhotoFlow: Agentic 3D Virtual Photography Missions

Model ReleasesDGX agent

arXiv:2605.23771v1 Announce Type: cross Abstract: Virtual photography asks an agent to enter a prepared 3D scene with no preselected camera pose or reference image, infer a suitable shot from scene in

← Previous
1…222223224225226…358
Next →