AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
Human
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
14 May 2026

Towards Unified Surgical Scene Understanding:Bridging Reasoning and Grounding via MLLMs

ApplicationsDGX agent

arXiv:2605.13530v1 Announce Type: cross Abstract: Surgical scene understanding is a cornerstone of computer-assisted intervention. While recent advances, particularly in surgical image segmentation, h

Tracing Persona Vectors Through LLM Pretraining

SafetyDGX agent

arXiv:2605.13329v1 Announce Type: cross Abstract: How large language models internally represent high-level behaviors is a core interpretability question with direct relevance to AI safety: it determi

TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking

SafetyDGX agent

arXiv:2605.12587v1 Announce Type: new Abstract: Dense 3D tracking from monocular video is fundamental to dynamic scene understanding. While recent 3D foundation models provide reliable per-frame geome

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Training Large Language Models to Predict Clinical Events

Model ReleasesDGX agent

arXiv:2605.12817v1 Announce Type: cross Abstract: Longitudinal clinical notes contain rich evidence of how patients evolve over time, but converting this signal into training supervision for clinical

Training LLMs with Reinforcement Learning for Intent-Aware Personalized Question Answering

Model ReleasesDGX agent

arXiv:2605.12645v1 Announce Type: cross Abstract: Effective personalized question answering (PQA) in language models requires grounding responses in the user's underlying intent, where intent refers t

Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context

AgentsDGX agent

arXiv:2605.13831v1 Announce Type: new Abstract: Long-context modeling is becoming a core capability of modern large vision-language models (LVLMs), enabling sustained context management across long-do

Trajectory-Level Data Augmentation for Offline Reinforcement Learning

SafetyDGX agent

arXiv:2605.13401v1 Announce Type: new Abstract: We propose a data augmentation method for offline reinforcement learning, motivated by active positioning problems. Particularly, our approach enables t

TRIAGE: Evaluating Prospective Metacognitive Control in LLMs under Resource Constraints

AgentsDGX agent

arXiv:2605.13414v1 Announce Type: new Abstract: Deploying language models as autonomous agents requires more than per-task accuracy: when an agent faces a queue of problems under a finite token budget

TurboGR: An Accelerated Training System for Large-Scale Generative Recommendation

ResearchDGX agent

arXiv:2605.13433v1 Announce Type: cross Abstract: Generative recommendation (GR) has emerged as a promising paradigm that replaces fragmented, scenario-specific architectures with unified Transformer-

Twincher: Bijective Representation Learning for Robust Inversion of Continuous Systems

TutorialsDGX agent

arXiv:2605.13470v1 Announce Type: new Abstract: Recent advances in AI have been primarily driven by large-scale neural architectures that excel at function approximation, rather than by tailored induc

U-HNO: A U-shaped Hybrid Neural Operator with Sparse-Point Adaptive Routing for Non-stationary PDE Dynamics

ResearchDGX agent

arXiv:2605.12965v1 Announce Type: new Abstract: Solutions to many partial differential equations (PDEs) display coexisting smooth global transport and localized sharp features within a single trajecto

UFO: A Domain-Unification-Free Operator Framework for Generalized Operator Learning

ResearchDGX agent

arXiv:2605.12700v1 Announce Type: new Abstract: Neural operators have become an effective framework for learning mappings between function spaces, yet most existing architectures realize operators wit

Uncertainty-Aware 3D Position Refinement for Multi-UAV Systems

AgentsDGX agent

arXiv:2605.13500v1 Announce Type: new Abstract: Reliable real-time 3D localization is essential for multi-UAV navigation, collision avoidance, and coordinated flight, yet onboard estimates can degrade

Uncertainty-Aware Prediction of Lung Tumor Growth from Sparse Longitudinal CT Data via Bayesian Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2605.13560v1 Announce Type: new Abstract: This work studies lung tumor growth prediction from sparse and irregular longitudinal computed tomography (CT) observations with measurement variability

Uncertainty-aware Spatial-Frequency Registration and Fusion for Infrared and Visible Images

SafetyDGX agent

arXiv:2605.13049v1 Announce Type: new Abstract: Infrared and Visible Image Fusion (IVIF) has shown promise in visual tasks under challenging environments, but fusion under unregistered conditions face

Uncertainty-Driven Anomaly Detection for Psychotic Relapse Using Smartwatches: Forecasting and Multi-Task Learning Fusion

Model ReleasesDGX agent

arXiv:2605.13816v1 Announce Type: new Abstract: Digital phenotyping enables continuous passive monitoring of behavior and physiology, offering a promising paradigm for early detection of psychotic rel

Uncovering Latent Pathological Signatures in Pulmonary CT via Cross-Window Knowledge Distillation

TutorialsDGX agent

arXiv:2605.12562v1 Announce Type: cross Abstract: Multi-window CT imaging captures complementary pathological information across anatomical structures of differing densities, yet existing deep learnin

Uncovering Symmetry Transfer in Large Language Models via Layer-Peeled Optimization

ResearchDGX agent

arXiv:2605.12756v1 Announce Type: cross Abstract: Large language models (LLMs) are pretrained by minimizing the cross-entropy loss for next-token prediction. In this paper, we study whether this optim

Understanding and Accelerating the Training of Masked Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.13026v1 Announce Type: cross Abstract: Masked diffusion models (MDMs) have emerged as a promising alternative to autoregressive models (ARMs) for language modeling. However, MDMs are known

Understanding Catastrophic Forgetting In LoRA via Mean-Field Attention Dynamics

Model ReleasesDGX agent

arXiv:2402.15415v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) is the dominant parameter-efficient fine-tuning method due to its favorable compute-performance trade-off, yet it suffers

Understanding Generalization through Decision Pattern Shift

ResearchDGX agent

arXiv:2605.13148v1 Announce Type: cross Abstract: Understanding why deep neural networks (DNNs) fail to generalize to unseen samples remains a long-standing challenge. Existing studies mainly examine

Unified generalization analysis for physics informed neural networks

ResearchDGX agent

arXiv:2605.13260v1 Announce Type: new Abstract: Physics-Informed Neural Networks (PINNs) and their variational counterparts (VPINNs) are neural networks that incorporate physical laws, making them use

Unify Robot Actions in Camera Frame

TutorialsDGX agent

arXiv:2511.17001v2 Announce Type: replace Abstract: Cross-embodiment robot learning requires a unified action representation with consistent semantics across robot platforms. Existing representations

Unifying Entropy Regularization in Optimal Control: From and Back to Classical Objectives via Iterated Soft Policies and Path Integral Solutions

SafetyDGX agent

arXiv:2512.06109v3 Announce Type: replace-cross Abstract: This paper develops a unified perspective on several optimal control formulations through the lens of Kullback-Leibler (KL) regularization. We

Unifying Physically-Informed Weather Priors in A Single Model for Image Restoration Across Multiple Adverse Weather Conditions

ResearchDGX agent

arXiv:2605.13158v1 Announce Type: new Abstract: Image restoration under multiple adverse weather conditions aims to develop a single model to recover the underlying scene with high visibility. Weather

UniJEPA: Enhancing Robot Policy via Unified Continuous and Discrete Representation Learning

SafetyDGX agent

arXiv:2510.10642v3 Announce Type: replace-cross Abstract: Building generalist robot policies that can handle diverse tasks in open-ended environments is a central challenge in robotics. To leverage kn

UNIV: Unified Foundation Model for Infrared and Visible Modalities

Model ReleasesDGX agent

arXiv:2509.15642v3 Announce Type: replace Abstract: Joint RGB-infrared perception is essential for achieving robustness under diverse weather and illumination conditions. Although foundation models ex

Universal Representation of Generalized Convex Functions and their Gradients

Model ReleasesDGX agent

arXiv:2509.04477v3 Announce Type: replace-cross Abstract: A wide range of optimization problems can often be written in terms of generalized convex functions (GCFs). When this structure is present, it

Unlocking Patch-Level Features for CLIP-Based Class-Incremental Learning

Model ReleasesDGX agent

arXiv:2605.13835v1 Announce Type: new Abstract: Class-Incremental Learning (CIL) enables models to continuously integrate new knowledge while mitigating catastrophic forgetting. Driven by the remarkab

Unweighted ranking for value-based decision making with uncertainty

SafetyDGX agent

arXiv:2605.13601v1 Announce Type: new Abstract: As intelligent systems are increasingly implemented in our society to make autonomous decisions, their commitment to human values raises serious concern

Useful Memories Become Faulty When Continuously Updated by LLMs

Model ReleasesDGX agent

arXiv:2605.12978v1 Announce Type: new Abstract: Learning from past experience benefits from two complementary forms of memory: episodic traces -- raw trajectories of what happened -- and consolidated

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2605.13277v1 Announce Type: cross Abstract: Visual evidence selection is a critical component of multimodal retrieval-augmented generation (RAG), yet existing methods typically rely on semantic

Vector-Quantized Discrete Latent Factors Meet Financial Priors: Dynamic Cross-Sectional Stock Ranking Prediction for Portfolio Construction

ResearchDGX agent

arXiv:2605.13407v1 Announce Type: new Abstract: Predicting cross-sectional stock returns is challenging due to low signal-to-noise ratios and evolving market regimes. Classical factor models offer int

VectorSmuggle: Steganographic Exfiltration in Embedding Stores and a Cryptographic Provenance Defense

Model ReleasesDGX agent

arXiv:2605.13764v1 Announce Type: cross Abstract: Modern retrieval-augmented generation (RAG) systems convert sensitive content into high-dimensional embeddings and store them in vector databases that

VERA-MH Concept Paper

Model ReleasesDGX agent

arXiv:2510.15297v4 Announce Type: replace-cross Abstract: We introduce VERA-MH (Validation of Ethical and Responsible AI in Mental Health), an automated evaluation of the safety of AI chatbots used in

VERA-MH: Validation of Ethical and Responsible AI in Mental Health

SafetyDGX agent

arXiv:2605.13318v1 Announce Type: new Abstract: Chatbot usage has increased, including in fields for which they were never developed for--notably mental health support. To that end, we introduce Valid

VideoSEAL: Mitigating Evidence Misalignment in Agentic Long Video Understanding by Decoupling Answer Authority

SafetyDGX agent

arXiv:2605.12571v1 Announce Type: cross Abstract: Long video question answering requires locating sparse, time-scattered visual evidence within highly redundant content. Although current MLLMs perform

ViDR: Grounding Multimodal Deep Research Reports in Source Visual Evidence

Model ReleasesDGX agent

arXiv:2605.13034v1 Announce Type: new Abstract: Recent deep research systems have improved the ability of large language models to produce long, grounded reports through iterative retrieval and reason

VIP-COP: Context Optimization for Tabular Foundation Models

ResearchDGX agent

arXiv:2605.12904v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have emerged as a powerful paradigm for in-context learning on structured data, enabling direct prediction on new tabul

VIPO: Value Function Inconsistency Penalized Offline Reinforcement Learning

TutorialsDGX agent

arXiv:2504.11944v3 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL) learns effective policies from pre-collected datasets, offering a practical solution for applications wher

Visual Accommodation: Rethinking Image Scale as a Learnable Variable for Object Detection

TutorialsDGX agent

arXiv:2412.06341v2 Announce Type: replace-cross Abstract: We propose Ciliary-DETR (previous name: Elastic-DETR), a framework for test-time resolution adjustment analogous to biological accommodation.

Visual Aesthetic Benchmark: Can Frontier Models Judge Beauty?

Model ReleasesDGX agent

arXiv:2605.12684v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are now routinely deployed for visual understanding, generation, and curation. A substantial fraction of thes

ViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation

ApplicationsDGX agent

arXiv:2506.15953v2 Announce Type: replace Abstract: Dexterous manipulation is a cornerstone capability for robotic systems aiming to interact with the physical world in a human-like manner. Although v

Vividh-ASR: A Complexity-Tiered Benchmark and Optimization Dynamics for Robust Indic Speech Recognition

Model ReleasesDGX agent

arXiv:2605.13087v1 Announce Type: cross Abstract: Fine-tuning multilingual ASR models like Whisper for low-resource languages often improves read speech but degrades spontaneous audio performance, a p

VoxCor: Training-Free Volumetric Features for Multimodal Voxel Correspondence

ResearchDGX agent

arXiv:2605.13798v1 Announce Type: new Abstract: Cross-modal 3D medical image analysis requires voxelwise representations that remain anatomically consistent across imaging contrasts, scanners, and acq

WARDEN: Endangered Indigenous Language Transcription and Translation with 6 Hours of Training Data

ResearchDGX agent

arXiv:2605.13846v1 Announce Type: cross Abstract: This paper introduces WARDEN, an early language model system capable of transcribing and translating Wardaman, an endangered Australian indigenous lan

Watermarking Should Be Treated as a Monitoring Primitive

SafetyDGX agent

arXiv:2605.13095v1 Announce Type: cross Abstract: Watermarking is widely proposed for provenance, attribution, and safety monitoring in generative models, yet is typically evaluated only under adversa

WD-FQDet: Multispectral Detection Transformer via Wavelet Decomposition and Frequency-aware Query Learning

SafetyDGX agent

arXiv:2605.13621v1 Announce Type: new Abstract: Infrared-visible object detection improves detection performance by combining complementary features from multispectral images. Existing backbone-specif

Weakly Supervised Segmentation as Semantic-Based Regularization

ResearchDGX agent

arXiv:2605.13674v1 Announce Type: cross Abstract: Weakly supervised semantic segmentation (WSSS) trains dense pixel-level segmentation models from partial or coarse annotations such as bounding boxes,

Weakly-Supervised Spatiotemporal Anomaly Detection

ResearchDGX agent

arXiv:2605.13746v1 Announce Type: cross Abstract: In this paper, we explore a weakly supervised method for anomaly detection. Since annotating videos is time-consuming, we only look at weak video-leve

What Do You Think I Think? Accounting for Human Beliefs Using Second-Order Theory of Mind

AgentsDGX agent

arXiv:2605.12745v1 Announce Type: cross Abstract: Discrepancies between an agent's actual knowledge and what a person thinks the agent knows can hinder interactions. If an agent could detect such disc

What Happens Before Decoding? Prefill Determines GUI Grounding in VLMs

ResearchDGX agent

arXiv:2605.12549v1 Announce Type: new Abstract: Existing training-free approaches for GUI grounding often rely on multiple inference runs, such as iterative cropping or candidate aggregation, to ident

What Information Matters? Graph Out-of-Distribution Detection via Tri-Component Information Decomposition

ResearchDGX agent

arXiv:2605.13032v1 Announce Type: new Abstract: Graph neural networks are widely used for node classification, but they remain vulnerable to out-of-distribution (OOD) shifts in node features and graph

What is Learnable in Valiant's Theory of the Learnable?

ResearchDGX agent

arXiv:2605.13840v1 Announce Type: cross Abstract: Valiant's 1984 paper is widely credited with introducing the PAC learning model, but it, in fact, introduced a different model: unlike PAC learning, t

What Limits Vision-and-Language Navigation ?

AgentsDGX agent

arXiv:2605.13328v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) is a cornerstone of embodied intelligence. However, current agents often suffer from significant performance degr

What properties of reasoning supervision are associated with improved downstream model quality?

SafetyDGX agent

arXiv:2605.13290v1 Announce Type: new Abstract: Validating training data for reasoning models typically requires expensive trial-and-error fine-tuning cycles. In this work, we investigate whether the

What to Ignore, What to React: Visually Robust RL Fine-Tuning of VLA Models

SafetyDGX agent

arXiv:2605.13105v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning has shown promise for Vision-Language-Action (VLA) models in robotic manipulation, but deployment-time visual sh

When Absolute State Fails: Evaluating Proprioceptive Encodings for Robust Manipulation

ApplicationsDGX agent

arXiv:2605.13067v1 Announce Type: cross Abstract: As end-to-end robotic policies are progressively deployed in the real world to solve real tasks, they face a gap between the training and inference co

When and Why is Optimistic Multiplicative Weights Slow? The Geometry of Energy Dissipation

ResearchDGX agent

arXiv:2605.13242v1 Announce Type: cross Abstract: This paper studies the convergence of the Optimistic Multiplicative Weights Update algorithm (OMWU) in two player zero-sum games. Recent works have id

When Attention Closes: How LLMs Lose the Thread in Multi-Turn Interaction

Model ReleasesDGX agent

arXiv:2605.12922v1 Announce Type: new Abstract: Large language models can follow complex instructions in a single turn, yet over long multi-turn interactions they often lose the thread of instructions

← Previous
1…708709710711712…1025
Next →