AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
Model Releases

VideoFDB: Evaluating Full-Duplex Vision-Speech Capabilities in Conversational Agents

DGX agent

arXiv:2605.30256v1 Announce Type: cross Abstract: Natural human conversation is full-duplex and audio-visual: people simultaneously speak and listen while continuously interpreting and producing nonve

model-releasesarxiv-cs-cl
29 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion

DGX agent

arXiv:2605.30351v1 Announce Type: cross Abstract: Long-rollout causal video diffusion has converged on a fixed-size sliding-window KV cache, with recent progress innovating within this layout by chang

hardwarearxiv-cs-ai
29 May 2026
Agents

VikingMem: A Memory Base Management System for Stateful LLM-based Applications

DGX agent

arXiv:2605.29640v1 Announce Type: new Abstract: Large Language Models have revolutionized interactive applications; however, their finite context windows pose a critical data management challenge for

agentsarxiv-cs-ai
29 May 2026
Research

Visual Spatial Learning: Single-Field Spatial Interpolation Using Convolutional Neural Networks

DGX agent

arXiv:2605.30167v1 Announce Type: cross Abstract: Predicting a complete spatially correlated field from sparse observations is a fundamental challenge in spatial statistics and environmental modelling

researcharxiv-cs-cv
29 May 2026
Agents

VisualThink-VLA: Visual Intermediate Reasoning for Effective and Low-Latency Vision-Language-Action Policies

DGX agent

arXiv:2605.30011v1 Announce Type: cross Abstract: Recent work has begun to equip vision-language-action (VLA) policies with explicit intermediate reasoning. In embodied control, however, textual chain

agentsarxiv-cs-ai
29 May 2026
Model Releases

VitalAgent: A Tool-Augmented Agent for Reactive and Proactive Physiological Monitoring over Wearable Health Data

DGX agent

arXiv:2605.29483v1 Announce Type: new Abstract: Wearable devices enable continuous monitoring of physiological signals such as ECG and PPG, but existing mHealth systems are largely limited to task-spe

model-releasesarxiv-cs-ai
29 May 2026
Applications

VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models

DGX agent

arXiv:2605.29562v1 Announce Type: cross Abstract: Vision-Language-Action~(VLA) models have shown strong potential for general-purpose robotic manipulation, yet they still struggle to generalize to uns

applicationsarxiv-cs-ai
29 May 2026
Safety

VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior Tracing

DGX agent

arXiv:2605.30117v1 Announce Type: new Abstract: Understanding how Vision-Language-Action (VLA) models transform multimodal knowledge into embodied control remains an open challenge. We present VLA-Tra

safetyarxiv-cs-ai
29 May 2026
Model Releases

VLAConf: Calibrated Task-Success Confidence for Vision-Language-Action Models

DGX agent

arXiv:2605.29605v1 Announce Type: new Abstract: Confidence estimation for Vision-Language-Action (VLA) models is essential for robots to perform manipulation tasks in the open world, providing crucial

model-releasesarxiv-cs-ro
29 May 2026
Model Releases

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation

DGX agent

arXiv:2605.30317v1 Announce Type: new Abstract: Autoregressive image and video generators are trained with teacher-forced histories but must sample from their own generated prefixes at inference time,

model-releasesarxiv-cs-cv
29 May 2026
Research

Wait! There's a Way Out: A Decision Mechanism for Forecasting Conversational Derailment

DGX agent

arXiv:2605.29243v1 Announce Type: cross Abstract: Forecasting conversational derailment is the task of predicting, as the conversation unfolds, whether it will eventually derail into personal attacks.

researcharxiv-cs-ai
29 May 2026
Model Releases

WASHH: An Anchor-Aware Whale-Guided Selection Hyper-Heuristic for Continuous Optimization and SVC Configuration

DGX agent

arXiv:2605.28844v1 Announce Type: cross Abstract: Learning-assisted algorithm design often has to make reliable search decisions under small evaluation budgets, where committing to a single metaheuris

model-releasesarxiv-cs-lg
29 May 2026
Research

Wasserstein Contraction of Coordinate Ascent Variational Inference

DGX agent

arXiv:2605.30253v1 Announce Type: cross Abstract: We study the contraction in Wasserstein distance of the coordinate ascent variational inference algorithm. This is shown to hold under a transport-inf

researcharxiv-cs-lg
29 May 2026
Research

WaterSearch: A Quality-Aware Search-based Watermarking Framework for Large Language Models

DGX agent

arXiv:2512.00837v2 Announce Type: replace Abstract: Watermarking acts as a critical safeguard in text generated by Large Language Models (LLMs). By embedding identifiable signals into model outputs, w

researcharxiv-cs-cl
29 May 2026
Research

Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bioacoustic Data

DGX agent

arXiv:2502.20838v3 Announce Type: replace-cross Abstract: Passive acoustic monitoring (PAM) systems generate continuous recordings spanning months, yet automated bioacoustic analysis of whale calls re

researcharxiv-cs-ai
29 May 2026
Research

What are They Thinking? Delineation, Probing and Tracking of Concepts in LLMs

DGX agent

arXiv:2605.28823v1 Announce Type: new Abstract: As the influence of LLMs expands, it is imperative to gain insight into their decisions. One way to do that is to develop probes that detect the presenc

researcharxiv-cs-cl
29 May 2026
Model Releases

What drives performance in molecular MPNNs? An operator-level factorial benchmark

DGX agent

arXiv:2605.30195v1 Announce Type: cross Abstract: Message-passing neural networks (MPNNs) are widely used for molecular property prediction, but their deployment as monolithic architectures makes it d

model-releasesarxiv-cs-ai
29 May 2026
Tutorials

What Exactly do Children Receive in Language Acquisition? A Case Study on CHILDES with Automated Detection of Filler-Gap Dependencies

DGX agent

arXiv:2603.02082v2 Announce Type: replace Abstract: Children's acquisition of filler-gap dependencies has been argued by some to depend on innate grammatical knowledge, while others suggest that the d

tutorialsarxiv-cs-cl
29 May 2026
Safety

When and How Human Curation Backfires: Preference Alignment under Multi-Model Self-Consuming Loop

DGX agent

arXiv:2605.29267v1 Announce Type: new Abstract: Foundation models are increasingly trained on synthetic data generated by prior model iterations rather than exclusively on real data. This self-consumi

safetyarxiv-cs-ai
29 May 2026
Safety

When and How Long? The Readout-Mediator Angle in Temporal Reasoning

DGX agent

arXiv:2605.29126v1 Announce Type: cross Abstract: A linear probe can decode a representation almost perfectly and yet be completely irrelevant to how the model uses it. On calendar-date duration reaso

safetyarxiv-cs-ai
29 May 2026
Local Ai

When Cloud Agents Meet Device Agents: Lessons from Hybrid Multi-Agent Systems

DGX agent

arXiv:2605.30102v1 Announce Type: cross Abstract: The design space of agentic AI inference spans two extremes: frontier large language models (LLMs), typically hosted in the cloud and offering strong

local-aiarxiv-cs-ai
29 May 2026
Research

When Do Graph Foundation Models Transfer? A Data-Centric Theory

DGX agent

arXiv:2605.29828v1 Announce Type: new Abstract: Graph foundation models (GFMs) aim to reuse a single backbone across diverse graph domains, yet their transfer is often uneven and can exhibit negative

researcharxiv-cs-lg
29 May 2026
Applications

When Does Persona Prompting Actually Help? A Retrieval and Metric Analysis of Expert Role Injection in LLMs

DGX agent

arXiv:2605.29420v1 Announce Type: new Abstract: Persona prompting is widely used to steer large language models, yet its practical value remains unclear. Prior work often evaluates persona prompting u

applicationsarxiv-cs-ai
29 May 2026
Model Releases

When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL

DGX agent

arXiv:2605.28918v1 Announce Type: new Abstract: For sparse, structured reinforcement-learning tasks with semantic reward-function interfaces, LLM-generated reward shaping is better framed as debugging

model-releasesarxiv-cs-lg
29 May 2026
Research

When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis

DGX agent

arXiv:2605.29025v1 Announce Type: new Abstract: Federal agencies are deploying large language models (LLMs) to categorize public comment corpora, where the model's organization of the record shapes wh

researcharxiv-cs-ai
29 May 2026
Tutorials

When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-Language Models

DGX agent

arXiv:2603.23085v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have enabled interpretable medical diagnosis by integrating visual perception with linguistic reasoning. Yet, existing

tutorialsarxiv-cs-ai
29 May 2026
Research

When RL Suppresses Its Own Vocabulary: Recovering Reasoning Diversity in Puzzle-to-Math Transfer

DGX agent

arXiv:2605.29190v1 Announce Type: cross Abstract: Reinforcement learning using verifiable rewards (RLVR) improves LLM reasoning, but the conditions under which it transfers across domains -- and why i

researcharxiv-cs-cl
29 May 2026
Model Releases

When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making

DGX agent

arXiv:2603.16673v4 Announce Type: replace-cross Abstract: Embodied robotic systems increasingly rely on large language model (LLM)-based agents to support high-level reasoning, planning, and decision-

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

When Should Models Change Their Minds? Contextual Belief Management in Large Language Models

DGX agent

arXiv:2605.30219v1 Announce Type: new Abstract: Long-horizon interactions require language models to manage accumulating information: when to update their state, when to preserve their state, and what

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

When the Same Coefficients Reach Different Places: Asymmetric Realizability in Transplanting Tokenizers across Large Language Models

DGX agent

arXiv:2601.00065v3 Announce Type: replace-cross Abstract: Tokenizer transplant in cross-vocabulary model composition reconstructs donor-only embedding rows as weighted combinations over shared lexical

model-releasesarxiv-cs-cl
29 May 2026
Research

When, why, and how do diffusion posterior samplers fail? A finite-sample lens

DGX agent

arXiv:2605.30330v1 Announce Type: new Abstract: Diffusion models have excellent capacity to model complex distributions of natural data, which has made them a popular and effective choice for posterio

researcharxiv-cs-lg
29 May 2026
Applications

Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialogues

DGX agent

arXiv:2605.30051v1 Announce Type: new Abstract: A key part of developing large language model (LLM)-powered, automated tutoring tools is student simulation, i.e., using LLMs to role-play as students,

applicationsarxiv-cs-cl
29 May 2026
Model Releases

Who can we trust? LLM-as-a-jury for Comparative Assessment

DGX agent

arXiv:2602.16610v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied as automatic evaluators for natural language generation assessment often using pairwise

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Why Far Looks Up: Probing Spatial Representation in Vision-Language Models

DGX agent

arXiv:2605.30161v1 Announce Type: new Abstract: Vision-language models (VLMs) achieve strong performance on spatial reasoning benchmarks, yet it remains unclear whether this reflects structured 3D und

model-releasesarxiv-cs-cv
29 May 2026
Tutorials

Why Larger Models Learn More: Effects of Capacity, Interference, and Rare-Task Retention

DGX agent

arXiv:2605.29548v1 Announce Type: new Abstract: Larger models learn tasks smaller models do not. What drives this phenomenon? We develop a simple phenomenological argument that power-law scaling alrea

tutorialsarxiv-cs-lg
29 May 2026
Model Releases

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

DGX agent

arXiv:2605.29744v1 Announce Type: new Abstract: The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-speci

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

World Models in Words: Auditing Physical State-Transition Commitments in Vision-Language Models

DGX agent

arXiv:2605.29585v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used to answer questions about physical scenes, yet most evaluations reduce performance to a final answer

model-releasesarxiv-cs-cl
29 May 2026
Agents

WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction

DGX agent

arXiv:2605.29341v1 Announce Type: cross Abstract: Multimodal large language models are increasingly deployed as long-horizon agents, where memory must do more than recall: it must track an evolving wo

agentsarxiv-cs-cl
29 May 2026
Research

X-GS: An Extensible Framework for Perceiving and Thinking via 3D Gaussian Splatting

DGX agent

arXiv:2603.09632v3 Announce Type: replace-cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a powerful technique for novel view synthesis, subsequently extending into numerous spatial AI app

researcharxiv-cs-cl
29 May 2026
Research

Xetrieval: Mechanistically Explaining Dense Retrieval

DGX agent

arXiv:2605.29507v1 Announce Type: new Abstract: Explaining why dense retrievers assign high relevance scores remains challenging because retrieval decisions are made through opaque high-dimensional em

researcharxiv-cs-ai
29 May 2026
Safety

xModel-KD: Cross-modal Knowledge Distillation for 3D Scene Perception using LiDAR

DGX agent

arXiv:2605.30111v1 Announce Type: cross Abstract: Point cloud segmentation is a fundamental task in 3D scene understanding. Its progress is constrained by the high cost and time required for dense 3D

safetyarxiv-cs-ai
29 May 2026
Model Releases

YoCausal: How Far is Video Generation from World Model? A Causality Perspective

DGX agent

arXiv:2605.30346v1 Announce Type: new Abstract: As video diffusion models (VDMs) advance toward world models, a key question arises: do they truly understand causality, or merely overfit to statistica

model-releasesarxiv-cs-cv
29 May 2026
Tutorials

Zero-shot CT Super-Resolution using Diffusion-based 2D Projection Priors and Signed 3D Gaussians

DGX agent

arXiv:2508.15151v3 Announce Type: replace-cross Abstract: Computed tomography (CT) is important in clinical diagnosis, but acquiring high-resolution (HR) CT is constrained by radiation exposure risks.

tutorialsarxiv-cs-cv
29 May 2026
Model Releases

A Bayesian Nonparametric Perspective on Mahalanobis Distance for Out of Distribution Detection

DGX agent

arXiv:2502.08695v2 Announce Type: replace-cross Abstract: Bayesian nonparametric methods are naturally suited to the problem of out-of-distribution (OOD) detection. However, these techniques have larg

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

A Broader View of Thompson Sampling

DGX agent

arXiv:2510.07208v2 Announce Type: replace Abstract: Thompson Sampling is one of the most widely used and studied bandit algorithms, known for its simple structure, low regret performance, and solid th

model-releasesarxiv-cs-lg
28 May 2026
Safety

A Comparative Study of Rule-Based and Data-Driven Approaches in Industrial Monitoring

DGX agent

arXiv:2509.15848v2 Announce Type: replace Abstract: Industrial monitoring systems, especially when deployed in Industry 4.0 environments, are experiencing a shift in paradigm from traditional rule-bas

safetyarxiv-cs-ai
28 May 2026
Research

A Conflict-Aware Penalty and Statistical Loss Framework for Balancing Modalities and Enhancing Stability in Multimodal Sentiment Analysis

DGX agent

arXiv:2605.28575v1 Announce Type: new Abstract: Multimodal Sentiment Analysis (MSA) fuses text, acoustic, and visual streams to infer sentiment. Because pre-trained text encoders are far more expressi

researcharxiv-cs-ai
28 May 2026
Research

A Digital Twin Framework for Virtual Visuo-Haptic Teleoperation of Complex-Shaped Optical Microrobots

DGX agent

arXiv:2605.28448v1 Announce Type: new Abstract: Optical tweezers (OT) provide piconewton-scale manipulation for delicate biomedical tasks, where visuo-haptic feedback can improve operator awareness by

researcharxiv-cs-ro
28 May 2026
← Previous
1…707708709710711…1311
Next →