AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
Human
88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
29 May 2026

Uncertainty Estimation via Hyperspherical Confidence Mapping

SafetyDGX agent

arXiv:2605.05964v2 Announce Type: replace Abstract: Quantifying uncertainty in neural network predictions is essential for high-stakes domains such as autonomous driving, healthcare, and manufacturing

Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge

ResearchDGX agent

arXiv:2505.16178v2 Announce Type: replace Abstract: While fine-tuning is the standard for injecting factual knowledge into large language models (LLMs), the mechanisms enabling reliable fact recall vi

Understanding Safety-Sensitive Expert Behavior in Mixture-of-Experts LLMs

Model ReleasesDGX agent

arXiv:2605.29708v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) LLMs rely on sparse, router-driven expert activation, yet how safety alignment interacts with routed expert specialization rema

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Understanding the Ability of LLMs to Handle Character-Level Perturbation

ResearchDGX agent

arXiv:2510.14365v4 Announce Type: replace Abstract: This work investigates the resilience of contemporary large language models (LLMs) against frequent character-level perturbations. We examine three

Uni-RCM: Unified Reference-guided Cross-modal Mapping for Multi-Class Anomaly Detection

TutorialsDGX agent

arXiv:2605.29455v1 Announce Type: new Abstract: Multi-modal industrial anomaly detection typically relies on separate models for each product category, fundamentally limiting practical scalability. Wh

Unifying Temporal and Structural Credit Assignment in LLM-Based Multi-Agent Prompt Optimization

Local AiDGX agent

arXiv:2605.30227v1 Announce Type: cross Abstract: While Multi-Agent Systems (MAS) empower Large Language Models to tackle complex reasoning tasks through collaborative interaction, optimizing their dy

UniNote: A Unified Embedding Model for Multimodal Representation and Ranking

Local AiDGX agent

arXiv:2605.29287v1 Announce Type: cross Abstract: Item-to-Item (I2I) retrieval is a fundamental part of modern content platforms, supporting critical industrial workflows from recommendation engines t

UniSteer: Text-Guided Flow Matching in Activation Space for Versatile LLM Steering

ResearchDGX agent

arXiv:2605.30076v1 Announce Type: new Abstract: Activation-based control steers large language models (LLMs) by intervening on their internal representations during inference, and has emerged as an ef

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning

AgentsDGX agent

arXiv:2605.29115v1 Announce Type: cross Abstract: Unix competence is the ability to use shell and operating-system primitives as first-class tools, not merely to write programs through a terminal. Cur

Unlocking the Working Memory of Large Language Models for Latent Reasoning

ResearchDGX agent

arXiv:2605.30343v1 Announce Type: cross Abstract: To improve the reasoning capabilities of large language models, test-time compute is typically scaled by generating intermediate tokens before the fin

Unsupervised Hierarchical Skill Discovery

ResearchDGX agent

arXiv:2601.23156v2 Announce Type: replace Abstract: We consider the problem of unsupervised skill segmentation and hierarchical structure discovery in reinforcement learning. While recent approaches h

Unsupervised Semantic Segmentation Facilitates Model Understanding

Local AiDGX agent

arXiv:2605.29691v1 Announce Type: new Abstract: Self-supervised learning (SSL) has produced a diverse landscape of vision transformers (ViTs) whose pretrained representations support a wide range of d

Unveiling Multi-regime Patterns in SciML: Distinct Failure Modes and Regime-specific Optimization

ResearchDGX agent

arXiv:2605.29153v1 Announce Type: cross Abstract: Neural networks trained under different hyperparameter settings can fall into distinct training 'regimes,' with consistent behavior within regimes and

Unveiling the Visual Counting Bottleneck in Vision-Language Models

ResearchDGX agent

arXiv:2605.30170v1 Announce Type: cross Abstract: While Large Vision-Language Models (VLMs) excel at interpolation, they suffer catastrophic failures in systematic generalization, most notably in visu

User-Aware Active Knowledge Acquisition for Emotional Support Dialogue

SafetyDGX agent

arXiv:2605.29715v1 Announce Type: new Abstract: Emotional support plays an important role in dialogue systems, and its success depends on adapting to a user's evolving and implicit needs across multi-

V2XCrafter: Learning to Generate Driving Scene Across Agents

SafetyDGX agent

arXiv:2605.29471v1 Announce Type: new Abstract: Collaborative driving systems leverage vehicle-to-everything (V2X) communication for multi-agent collaborative perception to enhance driving safety, yet

Valency Classification of Mapudungun Verbal Roots. Established by the language's own morphotactics

ResearchDGX agent

arXiv:2604.00789v3 Announce Type: replace Abstract: In the previous work, a lexical (re)categorisation -- or confirmation of the given category -- of roots identified as verbal was undertaken to deter

ValueFlow: Measuring the Propagation of Value Perturbations in Multi-Agent LLM Systems

SafetyDGX agent

arXiv:2602.08567v2 Announce Type: replace-cross Abstract: Multi-agent large language model (LLM) systems increasingly consist of agents that observe and respond to one another's outputs. While value a

VE2VF: Vision-Enabled to Vision-Free Distillation via Real-world Reinforcement Learning for Robust Contact-Rich Manipulation

Model ReleasesDGX agent

arXiv:2605.29564v1 Announce Type: new Abstract: When using reinforcement learning (RL) for contact-rich robotic manipulation, vision can provide task-relevant information that accelerates learning bey

Veda: Scalable Video Diffusion via Distilled Sparse Attention

ResearchDGX agent

arXiv:2605.30325v1 Announce Type: new Abstract: Scaling Diffusion Transformers to generate high-resolution, long videos is constrained by the quadratic cost of self-attention, and existing sparse atte

Verifiable Rewards Beyond Math and Code: Lightweight Corpus-Grounded Process Supervision for Factual Question Answering

Model ReleasesDGX agent

arXiv:2605.29648v1 Announce Type: new Abstract: Applying reinforcement learning to improve factual accuracy in knowledge-intensive question answering faces a reward design dilemma. Response-level rewa

VFEAgent: A Multimodal Agent Framework for End-to-End Automated Finite Element Analysis

AgentsDGX agent

arXiv:2605.28978v1 Announce Type: new Abstract: Finite Element Analysis (FEA) serves as the cornerstone of modern engineering design. However, its workflow is inherently complex and relies heavily on

ViASNet: A Video Ad Saliency Network for Predicting Dynamic Saliency and Viewer Engagement

ResearchDGX agent

arXiv:2605.29302v1 Announce Type: new Abstract: The digital media landscape has seen a pervasive shift toward short-form video advertising on TV, social media and e-commerce platforms. The present stu

Video Individual Counting and Tracking from Moving Drones: A Benchmark and Methods

Model ReleasesDGX agent

arXiv:2601.12500v2 Announce Type: replace Abstract: Counting and tracking dense crowds in large-scale scenes is a highly practical yet challenging problem. Existing methods mostly rely on fixed-camera

VideoFDB: Evaluating Full-Duplex Vision-Speech Capabilities in Conversational Agents

Model ReleasesDGX agent

arXiv:2605.30256v1 Announce Type: cross Abstract: Natural human conversation is full-duplex and audio-visual: people simultaneously speak and listen while continuously interpreting and producing nonve

VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion

HardwareDGX agent

arXiv:2605.30351v1 Announce Type: cross Abstract: Long-rollout causal video diffusion has converged on a fixed-size sliding-window KV cache, with recent progress innovating within this layout by chang

VikingMem: A Memory Base Management System for Stateful LLM-based Applications

AgentsDGX agent

arXiv:2605.29640v1 Announce Type: new Abstract: Large Language Models have revolutionized interactive applications; however, their finite context windows pose a critical data management challenge for

Visual Spatial Learning: Single-Field Spatial Interpolation Using Convolutional Neural Networks

ResearchDGX agent

arXiv:2605.30167v1 Announce Type: cross Abstract: Predicting a complete spatially correlated field from sparse observations is a fundamental challenge in spatial statistics and environmental modelling

VisualThink-VLA: Visual Intermediate Reasoning for Effective and Low-Latency Vision-Language-Action Policies

AgentsDGX agent

arXiv:2605.30011v1 Announce Type: cross Abstract: Recent work has begun to equip vision-language-action (VLA) policies with explicit intermediate reasoning. In embodied control, however, textual chain

VitalAgent: A Tool-Augmented Agent for Reactive and Proactive Physiological Monitoring over Wearable Health Data

Model ReleasesDGX agent

arXiv:2605.29483v1 Announce Type: new Abstract: Wearable devices enable continuous monitoring of physiological signals such as ECG and PPG, but existing mHealth systems are largely limited to task-spe

VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2605.29562v1 Announce Type: cross Abstract: Vision-Language-Action~(VLA) models have shown strong potential for general-purpose robotic manipulation, yet they still struggle to generalize to uns

VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior Tracing

SafetyDGX agent

arXiv:2605.30117v1 Announce Type: new Abstract: Understanding how Vision-Language-Action (VLA) models transform multimodal knowledge into embodied control remains an open challenge. We present VLA-Tra

VLAConf: Calibrated Task-Success Confidence for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.29605v1 Announce Type: new Abstract: Confidence estimation for Vision-Language-Action (VLA) models is essential for robots to perform manipulation tasks in the open world, providing crucial

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation

Model ReleasesDGX agent

arXiv:2605.30317v1 Announce Type: new Abstract: Autoregressive image and video generators are trained with teacher-forced histories but must sample from their own generated prefixes at inference time,

Wait! There's a Way Out: A Decision Mechanism for Forecasting Conversational Derailment

ResearchDGX agent

arXiv:2605.29243v1 Announce Type: cross Abstract: Forecasting conversational derailment is the task of predicting, as the conversation unfolds, whether it will eventually derail into personal attacks.

WASHH: An Anchor-Aware Whale-Guided Selection Hyper-Heuristic for Continuous Optimization and SVC Configuration

Model ReleasesDGX agent

arXiv:2605.28844v1 Announce Type: cross Abstract: Learning-assisted algorithm design often has to make reliable search decisions under small evaluation budgets, where committing to a single metaheuris

Wasserstein Contraction of Coordinate Ascent Variational Inference

ResearchDGX agent

arXiv:2605.30253v1 Announce Type: cross Abstract: We study the contraction in Wasserstein distance of the coordinate ascent variational inference algorithm. This is shown to hold under a transport-inf

WaterSearch: A Quality-Aware Search-based Watermarking Framework for Large Language Models

ResearchDGX agent

arXiv:2512.00837v2 Announce Type: replace Abstract: Watermarking acts as a critical safeguard in text generated by Large Language Models (LLMs). By embedding identifiable signals into model outputs, w

Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bioacoustic Data

ResearchDGX agent

arXiv:2502.20838v3 Announce Type: replace-cross Abstract: Passive acoustic monitoring (PAM) systems generate continuous recordings spanning months, yet automated bioacoustic analysis of whale calls re

What are They Thinking? Delineation, Probing and Tracking of Concepts in LLMs

ResearchDGX agent

arXiv:2605.28823v1 Announce Type: new Abstract: As the influence of LLMs expands, it is imperative to gain insight into their decisions. One way to do that is to develop probes that detect the presenc

What drives performance in molecular MPNNs? An operator-level factorial benchmark

Model ReleasesDGX agent

arXiv:2605.30195v1 Announce Type: cross Abstract: Message-passing neural networks (MPNNs) are widely used for molecular property prediction, but their deployment as monolithic architectures makes it d

What Exactly do Children Receive in Language Acquisition? A Case Study on CHILDES with Automated Detection of Filler-Gap Dependencies

TutorialsDGX agent

arXiv:2603.02082v2 Announce Type: replace Abstract: Children's acquisition of filler-gap dependencies has been argued by some to depend on innate grammatical knowledge, while others suggest that the d

When and How Human Curation Backfires: Preference Alignment under Multi-Model Self-Consuming Loop

SafetyDGX agent

arXiv:2605.29267v1 Announce Type: new Abstract: Foundation models are increasingly trained on synthetic data generated by prior model iterations rather than exclusively on real data. This self-consumi

When and How Long? The Readout-Mediator Angle in Temporal Reasoning

SafetyDGX agent

arXiv:2605.29126v1 Announce Type: cross Abstract: A linear probe can decode a representation almost perfectly and yet be completely irrelevant to how the model uses it. On calendar-date duration reaso

When Cloud Agents Meet Device Agents: Lessons from Hybrid Multi-Agent Systems

Local AiDGX agent

arXiv:2605.30102v1 Announce Type: cross Abstract: The design space of agentic AI inference spans two extremes: frontier large language models (LLMs), typically hosted in the cloud and offering strong

When Do Graph Foundation Models Transfer? A Data-Centric Theory

ResearchDGX agent

arXiv:2605.29828v1 Announce Type: new Abstract: Graph foundation models (GFMs) aim to reuse a single backbone across diverse graph domains, yet their transfer is often uneven and can exhibit negative

When Does Persona Prompting Actually Help? A Retrieval and Metric Analysis of Expert Role Injection in LLMs

ApplicationsDGX agent

arXiv:2605.29420v1 Announce Type: new Abstract: Persona prompting is widely used to steer large language models, yet its practical value remains unclear. Prior work often evaluates persona prompting u

When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL

Model ReleasesDGX agent

arXiv:2605.28918v1 Announce Type: new Abstract: For sparse, structured reinforcement-learning tasks with semantic reward-function interfaces, LLM-generated reward shaping is better framed as debugging

When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis

ResearchDGX agent

arXiv:2605.29025v1 Announce Type: new Abstract: Federal agencies are deploying large language models (LLMs) to categorize public comment corpora, where the model's organization of the record shapes wh

When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-Language Models

TutorialsDGX agent

arXiv:2603.23085v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have enabled interpretable medical diagnosis by integrating visual perception with linguistic reasoning. Yet, existing

When RL Suppresses Its Own Vocabulary: Recovering Reasoning Diversity in Puzzle-to-Math Transfer

ResearchDGX agent

arXiv:2605.29190v1 Announce Type: cross Abstract: Reinforcement learning using verifiable rewards (RLVR) improves LLM reasoning, but the conditions under which it transfers across domains -- and why i

When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making

Model ReleasesDGX agent

arXiv:2603.16673v4 Announce Type: replace-cross Abstract: Embodied robotic systems increasingly rely on large language model (LLM)-based agents to support high-level reasoning, planning, and decision-

When Should Models Change Their Minds? Contextual Belief Management in Large Language Models

Model ReleasesDGX agent

arXiv:2605.30219v1 Announce Type: new Abstract: Long-horizon interactions require language models to manage accumulating information: when to update their state, when to preserve their state, and what

When the Same Coefficients Reach Different Places: Asymmetric Realizability in Transplanting Tokenizers across Large Language Models

Model ReleasesDGX agent

arXiv:2601.00065v3 Announce Type: replace-cross Abstract: Tokenizer transplant in cross-vocabulary model composition reconstructs donor-only embedding rows as weighted combinations over shared lexical

When, why, and how do diffusion posterior samplers fail? A finite-sample lens

ResearchDGX agent

arXiv:2605.30330v1 Announce Type: new Abstract: Diffusion models have excellent capacity to model complex distributions of natural data, which has made them a popular and effective choice for posterio

Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialogues

ApplicationsDGX agent

arXiv:2605.30051v1 Announce Type: new Abstract: A key part of developing large language model (LLM)-powered, automated tutoring tools is student simulation, i.e., using LLMs to role-play as students,

Who can we trust? LLM-as-a-jury for Comparative Assessment

Model ReleasesDGX agent

arXiv:2602.16610v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied as automatic evaluators for natural language generation assessment often using pairwise

Why Far Looks Up: Probing Spatial Representation in Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.30161v1 Announce Type: new Abstract: Vision-language models (VLMs) achieve strong performance on spatial reasoning benchmarks, yet it remains unclear whether this reflects structured 3D und

Why Larger Models Learn More: Effects of Capacity, Interference, and Rare-Task Retention

TutorialsDGX agent

arXiv:2605.29548v1 Announce Type: new Abstract: Larger models learn tasks smaller models do not. What drives this phenomenon? We develop a simple phenomenological argument that power-law scaling alrea

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

Model ReleasesDGX agent

arXiv:2605.29744v1 Announce Type: new Abstract: The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-speci

← Previous
1…565566567568569…1049
Next →