AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
1 Jun 2026

The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability

Local AiDGX agent

arXiv:2605.30628v1 Announce Type: cross Abstract: Universal LLM reliability is not a finite-library problem: across all possible tasks, tools, schemas, knowledge sources, and evaluator expectations, n

The Gaussian-Head OFL Family: One-Shot Federated Learning from Client Global Statistics

ResearchDGX agent

arXiv:2602.01186v2 Announce Type: replace-cross Abstract: Classical Federated Learning relies on a multi-round iterative process of model exchange and aggregation between server and clients, with high

The Global Landscape of Environmental AI Regulation: From the Cost of Reasoning to a Right to Green AI

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2603.00068v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI) systems impose substantial and growing environmental costs, yet transparency about these impacts has declined eve

The Information Geometry of Softmax: Probing and Steering

ResearchDGX agent

arXiv:2602.15293v2 Announce Type: replace-cross Abstract: This paper concerns the question of how AI systems encode semantic structure into the geometric structure of their representation spaces. The

The Refutability Gap: Challenges in Validating Reasoning by Large Language Models

SafetyDGX agent

arXiv:2601.02380v4 Announce Type: replace-cross Abstract: Recent reports claim that Large Language Models (LLMs) have achieved the ability to derive new science and exhibit human-level general intelli

The Surface You Test Is Not the Surface That Breaks

Model ReleasesDGX agent

arXiv:2605.30454v1 Announce Type: cross Abstract: Tool-augmented LLM agents are vulnerable to prompt injection: a third party who controls part of the agent's context can plant instructions that the a

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning

SafetyDGX agent

arXiv:2605.31404v1 Announce Type: cross Abstract: Large Language Model (LLM)-based navigation systems commonly construct explicit spatial representations (e.g., topological graphs, semantic raster map

The Terminal Representation in Reinforcement Learning

TutorialsDGX agent

arXiv:2605.31289v1 Announce Type: cross Abstract: Representation learning is a powerful tool for spatio-temporal abstraction within reinforcement learning (RL). Two well established approaches are thr

Towards Atoms of Large Language Models

ResearchDGX agent

arXiv:2509.20784v3 Announce Type: replace-cross Abstract: The fundamental representational units (FRUs) of large language models (LLMs) remain undefined, limiting further understanding of their underl

Toxic HallucinAItions: Perturbing Prompts and Tracing LLM Circuits

ResearchDGX agent

arXiv:2605.30913v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in conversational settings where user tone ranges from polite to adversarial or toxic, yet less

TraceGraph: Shared Decision Landscapes for Diagnosing and Improving Agent Trajectories

Model ReleasesDGX agent

arXiv:2605.31308v1 Announce Type: new Abstract: Agent benchmarks increasingly record rich interaction trajectories, yet evaluation often reduces each rollout to a pass rate or reward score. We introdu

Transforming and Encoding FTS for SAT Solving: What Helps, What Hurts (Extended Version)

TutorialsDGX agent

arXiv:2605.30563v1 Announce Type: new Abstract: Factored tasks are a classical planning representation that extends SAS+ with limited forms of disjunctive preconditions, conditional effects, and angel

TRINE: A Token-Aware, Runtime-Adaptive FPGA Inference Engine for Multimodal AI

ResearchDGX agent

arXiv:2603.22867v1 Announce Type: cross Abstract: Multimodal stacks that mix ViTs, CNNs, GNNs, and transformer NLP strain embedded platforms because their compute/memory patterns diverge and hard real

Trust-Region Behavior Blending for On-Policy Distillation

SafetyDGX agent

arXiv:2605.31159v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student on prefixes sampled from its own policy while matching a stronger teacher. This addresses the prefix mis

TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation

SafetyDGX agent

arXiv:2605.31590v1 Announce Type: cross Abstract: Text-to-video (T2V) generation faces challenging questions when generating videos with long horizons containing multiple events. Inspired by the intri

TUX: Measuring Human--AI Tacit Understanding

SafetyDGX agent

arXiv:2605.30930v1 Announce Type: cross Abstract: As large language models (LLMs) increasingly act as collaborative partners, human--AI alignment is often evaluated through explicit task success, accu

Uncertainty-Aware and Temporally Regulated Expert Advice in Reinforcement Learning for Autonomous Driving

SafetyDGX agent

arXiv:2605.30576v1 Announce Type: new Abstract: Exploration in reinforcement learning for autonomous driving is inherently unsafe: agents must experience novel behaviors to learn, yet exploration can

Understanding the Fundamental Design Decisions of Retrieval-Augmented Generation Systems

TutorialsDGX agent

arXiv:2411.19463v3 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has emerged as a critical technique for enhancing large language model (LLM) capabilities. However, pract

Unicorn: Scaling High-Dimensional Time Series Forecasting via Universal Correlation Modeling

ResearchDGX agent

arXiv:2605.30376v1 Announce Type: cross Abstract: Modern time series architectures face a fundamental trade-off: channel-independent models scale well with increasing data volume but ignore critical i

Unifying and Optimizing Data Values for Selection via Sequential Decision-Making

ResearchDGX agent

arXiv:2502.04554v2 Announce Type: replace Abstract: Data selection has emerged as a crucial downstream application of data valuation, yet the theoretical foundations for using data values in selection

UniScale: Adaptive Unified Inference Scaling via Online Joint Optimization of Model Routing and Test-Time Scaling

ApplicationsDGX agent

arXiv:2605.30898v1 Announce Type: new Abstract: In real-world deployments of large language models (LLMs), balancing inference quality and computational cost has become a central challenge. Existing a

Unlearning in Diffusion Models: A Unified Framework with KL Divergence and Likelihood Constraints

ResearchDGX agent

arXiv:2605.30825v1 Announce Type: cross Abstract: Unlearning in diffusion models aims to remove undesirable data or concepts while preserving the utility of pretrained models -- two fundamentally conf

Unlearning's Blind Spots: Over-Unlearning and Prototypical Relearning Attack

ResearchDGX agent

arXiv:2506.01318v4 Announce Type: replace-cross Abstract: Machine unlearning (MU) aims to expunge a designated forget set from a trained model without costly retraining, yet the existing techniques ov

Unraveling LoRA Interference: Orthogonal Subspaces for Robust Model Merging

Model ReleasesDGX agent

arXiv:2505.22934v2 Announce Type: replace-cross Abstract: Fine-tuning large language models (LMs) for individual tasks yields strong performance but is expensive for deployment and storage. Recent wor

Updating the standard neuron model in artificial neural networks

ResearchDGX agent

arXiv:2605.30370v1 Announce Type: cross Abstract: From their inception in the 1950s, artificial neural networks (ANNs) started using the so-called point neuron model then prevalent in neuroscience, ho

Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Information

SafetyDGX agent

arXiv:2605.31445v1 Announce Type: cross Abstract: In this work we study agents in simulated bargaining scenarios, where a buyer and a seller communicate through a text channel and attempt to negotiate

UXR PoV for Neuroinclusive Emotion Regulation

SafetyDGX agent

arXiv:2605.31131v1 Announce Type: cross Abstract: Attention-deficit/hyperactivity disorder (ADHD) is a psychiatric disorder which presents itself in individuals through patterns of developmentally ina

Variational Adapter for Cross-modal Similarity Representation

ResearchDGX agent

arXiv:2605.30968v1 Announce Type: cross Abstract: The core of vision-language models lies in measuring cross-modal similarity within a unified representation space. However, most image-text matching o

Variational Routing: A Scalable Bayesian Framework for Calibrated Mixture-of-Experts Transformers

Model ReleasesDGX agent

arXiv:2603.09453v3 Announce Type: replace-cross Abstract: Foundation models are increasingly being deployed in contexts where understanding the uncertainty of their outputs is critical to ensuring res

Vector Linking via Cross-Model Local Isometric Consistency

Local AiDGX agent

arXiv:2605.31100v1 Announce Type: new Abstract: We study Vector Linking: given two embedding clouds produced by different black-box encoders over partially overlapping datasets, recover cross-model ob

Vision-Language Models Suppress Female Representations Under Ambiguous Input

SafetyDGX agent

arXiv:2605.31556v1 Announce Type: cross Abstract: Alignment teaches vision-language models (VLMs) to avoid expressing demographic biases, and when gender is clearly visible they largely succeed. Far l

VLM3: Vision Language Models Are Native 3D Learners

ResearchDGX agent

arXiv:2605.30561v1 Announce Type: cross Abstract: Vision Language Models (VLMs) enable a unified model to solve various vision tasks through prompting. They have shown promising performance in semanti

Weight Decay Improves Language Model Plasticity

Model ReleasesDGX agent

arXiv:2602.11137v2 Announce Type: replace-cross Abstract: Large language models are typically trained in two broad phases: pretraining to produce a base model, followed by further training to improve

What changes after deployment? A survey on On-device Learning in TinyML

Local AiDGX agent

arXiv:2605.31226v1 Announce Type: cross Abstract: Machine learning models on microcontroller-class devices (TinyML) face a fundamental challenge: post-deployment distribution change undermines static

What Gets Unmasked First? Trajectory Analysis of Diffusion Models for Graph-to-Text Generation

ResearchDGX agent

arXiv:2605.31564v1 Announce Type: cross Abstract: We present the first systematic study of masked diffusion language models (MDLMs) for graph-to-text generation. We analyze MDLM generation trajectorie

What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucination Robustness

Model ReleasesDGX agent

arXiv:2605.30911v1 Announce Type: cross Abstract: Hallucination remains one of the key challenges undermining the reliability of Large Vision-Language Models (LVLMs). But what makes an LVLM hallucinat

When are LLMs Sufficient Policy Optimizers for Sequential RL Tasks?

SafetyDGX agent

arXiv:2605.30719v1 Announce Type: cross Abstract: We study when large language models (LLMs) can serve as effective black-box policy optimizers for reinforcement learning (RL) tasks, i.e., when can we

When LLMs Learn to Be Consistently Wrong: A Multi-Model Study of Linear Representations of Synthetic Deception

Model ReleasesDGX agent

arXiv:2605.30381v1 Announce Type: cross Abstract: Deceptive alignment, in which models maintain accurate internal representations while deliberately producing false outputs, remains a central challeng

Who Gets Credit or Blame? Attributing Accountability in Modern AI Systems

SafetyDGX agent

arXiv:2506.00175v5 Announce Type: replace-cross Abstract: Modern AI systems are typically developed through multiple stages-pretraining, fine-tuning rounds, and subsequent adaptation or alignment, whe

Why Linear Recurrent Memory Works in Partially Observable Reinforcement Learning

SafetyDGX agent

arXiv:2605.31261v1 Announce Type: cross Abstract: The family of linear recurrent neural networks has shown strong performance as recurrent memory units in partially observable reinforcement learning.

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry

SafetyDGX agent

arXiv:2604.01985v2 Announce Type: replace-cross Abstract: General-purpose world models promise scalable policy evaluation, optimization, and planning, yet achieving the required level of robustness re

XLGoBench: Detecting cross-lingual skill gaps with algorithmic tasks

Model ReleasesDGX agent

arXiv:2605.30788v1 Announce Type: cross Abstract: We introduce a set of synthetic algorithmic tasks to detect cross-lingual gaps in the abilities of large language models. Our benchmark is commensurat

XOResNet: Exclusive-OR Meta-Residuals Facilitate Deep Spiking Neural Networks Learning

ResearchDGX agent

arXiv:2605.30362v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) hold promise for demonstrating superior learning and representation capabilities in deep models. Given the tremendous s

Your Teacher Can't Help You Here: Combating Supervision Fidelity Decay in On-Policy Distillation

SafetyDGX agent

arXiv:2605.30833v1 Announce Type: cross Abstract: On-policy distillation transfers reasoning capabilities by training a student model on its own generated trajectories using token-level feedback from

29 May 2026

A comparative study of transformer-based embeddings for topic coherence

Model ReleasesDGX agent

arXiv:2605.28832v1 Announce Type: cross Abstract: Topic modeling is a branch of Natural Language Processing (NLP) that aims to organize large collections of texts into coherent groups according to wor

A Composable Multimodal Framework for cine CMR-Text-Driven Prediction of Heart Failure Outcomes

ResearchDGX agent

arXiv:2502.16548v3 Announce Type: replace-cross Abstract: Objective. Heart failure is one of the leading causes of death worldwide, with millions of deaths each year, according to data from the World

A Language-Guided Bayesian Optimization for Efficient LoRA Hyperparameter Search

ResearchDGX agent

arXiv:2602.11171v2 Announce Type: replace-cross Abstract: Fine-tuning Large Language Models (LLMs) with Low-Rank Adaptation (LoRA) offers a resource-efficient way to personalize or specialize. However

A Matter of Interest: Understanding Interestingness of Math Problems in Humans and Language Models

ApplicationsDGX agent

arXiv:2511.08548v2 Announce Type: replace Abstract: The evolution of mathematics is shaped importantly by interestingness: researchers choose which problems to pursue, and students choose which proble

A Minimal Bifurcation Model of Load Imbalance in a Softmax Mixture-of-Experts Router

Model ReleasesDGX agent

arXiv:2605.29121v1 Announce Type: cross Abstract: We propose a minimal dynamical model of adaptive softmax routing for a two-expert Mixture-of-Experts (MoE) layer. The model is obtained as a mean-fiel

A Predictive Law for On-Policy Self-Distillation From World Feedback

SafetyDGX agent

arXiv:2605.30070v1 Announce Type: cross Abstract: Moving beyond simple scalar rewards toward richer world feedback is a natural path to more scalable RL post-training. On-policy self-distillation (OPS

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach

SafetyDGX agent

arXiv:2512.11944v2 Announce Type: replace-cross Abstract: Motion planning for autonomous driving (AD) faces a critical trade-off. While traditional rule-based pipelines offer verifiable safety and int

A Survey on Recent Advances in Conversational Data Generation

ResearchDGX agent

arXiv:2405.13003v2 Announce Type: replace-cross Abstract: Recent advancements in conversational systems have significantly enhanced human-machine interactions across various domains. However, training

A unified deeplearning framework for contrast-phase-specific virtual monochromatic imaging

ResearchDGX agent

arXiv:2605.29753v1 Announce Type: cross Abstract: Dual-energy CT (DECT) enables virtual monochromatic imaging (VMI) and improved contrast resolution, but its clinical adoption is limited by hardware c

Accelerating Constrained Decoding with Token Space Compression

ResearchDGX agent

arXiv:2605.29986v1 Announce Type: new Abstract: To guarantee that an LLM's outputs conform to a specified structure, context-free grammar (CFG) decoding engines force the selection of next tokens that

Adaptive Interviewing for Persona Simulation in LLMs: Evidence-Grounded Reasoning Improves Decision Alignment

SafetyDGX agent

arXiv:2605.29458v1 Announce Type: cross Abstract: Accurately simulating the decisions of a specific individual remains challenging for large language models (LLMs), partly because persona information

Adopt neq Adapt: Longitudinal Analyses of LLM Conversations in the Wild

ResearchDGX agent

arXiv:2605.29018v1 Announce Type: new Abstract: Although a growing body of research has begun to describe user--LLM interactions, the picture it paints is largely static; little is known about how ind

AG-REPA: Causal Layer Selection for Representation Alignment in Audio Flow Matching

SafetyDGX agent

arXiv:2603.01006v2 Announce Type: replace-cross Abstract: REPresentation Alignment (REPA) improves the training of generative flow models by aligning intermediate hidden states with pretrained teacher

Agent4Edu: Generating Learner Response Data by Generative Agents for Intelligent Education Systems

AgentsDGX agent

arXiv:2501.10332v2 Announce Type: replace-cross Abstract: Personalized learning represents a promising educational strategy within intelligent educational systems, aiming to enhance learners' practice

AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security

Model ReleasesDGX agent

arXiv:2605.29801v1 Announce Type: new Abstract: Modern open-world agents such as OpenClaw exhibit powerful cross-environment execution capabilities yet introduce broad new safety risk sources. Meanwhi

AgentDropoutV2: Optimizing Information Flow in Multi-Agent Systems via Test-Time Rectify-or-Reject Pruning

Model ReleasesDGX agent

arXiv:2602.23258v2 Announce Type: replace Abstract: While Multi-Agent Systems (MAS) excel in complex reasoning, they suffer from the cascading impact of erroneous information from individual agents. C

← Previous
1…187188189190191…358
Next →