AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
Safety

Trust the Batch, On- or Off-Policy: Adaptive Policy Optimization for RL Post-Training

DGX agent

arXiv:2605.12380v1 Announce Type: new Abstract: Reinforcement learning is structurally harder than supervised learning because the policy changes the data distribution it learns from. The resulting fr

safetyarxiv-cs-lg
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

U-STS-LLM A Unified Spatio-Temporal Steered Large Language Model for Traffic Prediction and Imputation

DGX agent

arXiv:2605.11735v1 Announce Type: new Abstract: The efficient operation of modern cellular networks hinges on the accurate analysis of spatio-temporal traffic data. Mastering these patterns is essenti

model-releasesarxiv-cs-lg
13 May 2026
Safety

UGround: Towards Unified Visual Grounding with Unrolled Transformers

DGX agent

arXiv:2510.03853v4 Announce Type: replace Abstract: We present UGround, a extbf{U}nified visual extbf{Ground}ing paradigm that dynamically selects intermediate layers across extbf{U}nrolled transforme

safetyarxiv-cs-cv
13 May 2026
Model Releases

UHR-Micro: Diagnosing and Mitigating the Resolution Illusion in Earth Observation VLMs

DGX agent

arXiv:2605.12237v1 Announce Type: new Abstract: Vision-Language Models (VLMs) increasingly operate on ultra-high-resolution (UHR) Earth observation imagery, yet they remain vulnerable to a severe scal

model-releasesarxiv-cs-cv
13 May 2026
Safety

Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization

DGX agent

arXiv:2605.11491v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning ability of large language models. How

safetyarxiv-cs-lg
13 May 2026
Safety

Understanding Sample Efficiency in Predictive Coding

DGX agent

arXiv:2605.11911v1 Announce Type: new Abstract: Predictive Coding (PC) is an influential account of cortical learning. Much of recent work has focused on comparing PC to Backpropagation (BP) to find w

safetyarxiv-cs-lg
13 May 2026
Safety

Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPO

DGX agent

arXiv:2505.19770v5 Announce Type: replace-cross Abstract: We present a fine-grained theoretical analysis of the performance gap between two-stage reinforcement learning from human feedback~(RLHF) and

safetyarxiv-cs-cl
13 May 2026
Model Releases

UnfoldLDM: Degradation-Aware Unfolding with Iterative Latent Diffusion Priors for Blind Image Restoration

DGX agent

arXiv:2511.18152v3 Announce Type: replace Abstract: Deep unfolding networks (DUNs) combine the interpretability of model-based methods with the learning ability of deep networks, yet remain limited fo

model-releasesarxiv-cs-cv
13 May 2026
Tutorials

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation

DGX agent

arXiv:2605.12088v1 Announce Type: new Abstract: Multi-reference image generation aims to synthesize images from textual instructions while faithfully preserving subject identities from multiple refere

tutorialsarxiv-cs-cv
13 May 2026
Safety

UniFixer: A Universal Reference-Guided Fixer for Diffusion-Based View Synthesis

DGX agent

arXiv:2605.12169v1 Announce Type: new Abstract: With the recent surge of generative models, diffusion-based approaches have become mainstream for view synthesis tasks, either in an explicit depth-warp

safetyarxiv-cs-cv
13 May 2026
Research

Uniform Scaling Limits in AdamW-Trained Transformers

DGX agent

arXiv:2605.11059v1 Announce Type: cross Abstract: We study the large-depth limit of transformers trained with AdamW, by modelling the hidden-state dynamics as an interacting particle system (IPS) coup

researcharxiv-cs-lg
13 May 2026
Applications

UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs

DGX agent

arXiv:2605.11856v1 Announce Type: cross Abstract: Multimodal large language models are increasingly expected to perform thinking with images, yet existing visual latent reasoning methods still rely on

applicationsarxiv-cs-cl
13 May 2026
Research

Unlearning with Asymmetric Sources: Improved Unlearning-Utility Trade-off with Public Data

DGX agent

arXiv:2605.11170v1 Announce Type: new Abstract: Noise-based certified machine unlearning currently faces a hard ceiling: the noise magnitude required to certify unlearning typically destroys model uti

researcharxiv-cs-lg
13 May 2026
Research

Unlocking Compositional Generalization in Continual Few-Shot Learning

DGX agent

arXiv:2605.11710v1 Announce Type: cross Abstract: Object-centric representations promise a key property for few-shot learning: Rather than treating a scene as a single unit, a model can decompose it i

researcharxiv-cs-cv
13 May 2026
Agents

Unlocking LLM Creativity in Science through Analogical Reasoning

DGX agent

arXiv:2605.11258v1 Announce Type: cross Abstract: Autonomous science promises to augment scientific discovery, particularly in complex fields like biomedicine. However, this requires AI systems that c

agentsarxiv-cs-cl
13 May 2026
Model Releases

Unlocking UML Class Diagram Understanding in Vision Language Models

DGX agent

arXiv:2605.11634v1 Announce Type: new Abstract: Although Vision Language Models (VLMs) have seen tremendous progress across all kinds of use cases, they still fall behind in answering questions regard

model-releasesarxiv-cs-cv
13 May 2026
Research

Unpacking the Eye of the Beholder: Social Location, Identity, and the Moving Target of Political Perspectives

DGX agent

arXiv:2605.11166v1 Announce Type: new Abstract: Political and social identities structure how people evaluate political information, a finding decades deep in political science and routinely discarded

researcharxiv-cs-cv
13 May 2026
Model Releases

Urban Risk-Aware Navigation via VQA-Based Event Maps for People with Low Vision

DGX agent

arXiv:2605.11782v1 Announce Type: new Abstract: Visual impairment affects hundreds of millions of people worldwide, severely limiting their ability to navigate urban environments safely and independen

model-releasesarxiv-cs-cv
13 May 2026
Research

USEMA: a Scalable Efficient Mamba Like Attention for Medical Image Segmentation

DGX agent

arXiv:2605.11131v1 Announce Type: new Abstract: Accurate medical image segmentation is an integral part of the medical image analysis pipeline that requires the ability to merge local and global infor

researcharxiv-cs-cv
13 May 2026
Applications

Variance-aware Reward Modeling with Anchor Guidance

DGX agent

arXiv:2605.11865v1 Announce Type: cross Abstract: Standard Bradley--Terry (BT) reward models are limited when human preferences are pluralistic. Although soft preference labels preserve disagreement i

applicationsarxiv-cs-lg
13 May 2026
Research

Variational Linear Attention: Stable Associative Memory for Long-Context Transformers

DGX agent

arXiv:2605.11196v1 Announce Type: new Abstract: Linear attention reduces the quadratic cost of softmax attention to O(T), but its memory state grows as O(T) in Frobenius norm, causing progressive inte

researcharxiv-cs-lg
13 May 2026
Research

Vector Scaffolding: Inter-Scale Orchestration for Differentiable Image Vectorization

DGX agent

arXiv:2605.11913v1 Announce Type: new Abstract: Differentiable vector graphics have enabled powerful gradient-based optimization of vector primitives directly from raster images. However, existing fra

researcharxiv-cs-cv
13 May 2026
Model Releases

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference

DGX agent

arXiv:2605.11334v1 Announce Type: cross Abstract: LLM-as-Judge systems are widely deployed for automated evaluation, yet practitioners lack reliable methods to know when a judge's verdict should be tr

model-releasesarxiv-cs-cl
13 May 2026
Research

Vertex-Softmax: Tight Transformer Verification via Exact Softmax Optimization

DGX agent

arXiv:2605.10974v1 Announce Type: new Abstract: Certified verification of transformer attention requires bounding the softmax function over interval constraints on the pre-softmax scores. Existing ver

researcharxiv-cs-lg
13 May 2026
Model Releases

Very Efficient Listwise Multimodal Reranking for Long Documents

DGX agent

arXiv:2605.11864v1 Announce Type: cross Abstract: Listwise reranking is a key yet computationally expensive component in vision-centric retrieval and multimodal retrieval-augmented generation (M-RAG)

model-releasesarxiv-cs-cv
13 May 2026
Research

Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives

DGX agent

arXiv:2406.05615v4 Announce Type: replace Abstract: Humans use multiple senses to comprehend the environment. Vision and language are two of the most vital senses since they allow us to easily communi

researcharxiv-cs-cl
13 May 2026
Research

VidSplat: Gaussian Splatting Reconstruction with Geometry-Guided Video Diffusion Priors

DGX agent

arXiv:2605.11424v1 Announce Type: new Abstract: Gaussian Splatting has achieved remarkable progress in multi-view surface reconstruction, yet it exhibits notable degradation when only few views are av

researcharxiv-cs-cv
13 May 2026
Safety

VIP: Visual-guided Prompt Evolution for Efficient Dense Vision-Language Inference

DGX agent

arXiv:2605.12325v1 Announce Type: new Abstract: Pursuing training-free open-vocabulary semantic segmentation in an efficient and generalizable manner remains challenging due to the deep-seated spatial

safetyarxiv-cs-cv
13 May 2026
Research

Vision-aligned Latent Reasoning for Multi-modal Large Language Model

DGX agent

arXiv:2602.04476v2 Announce Type: replace Abstract: Despite recent advancements in Multi-modal Large Language Models (MLLMs) on diverse understanding tasks, these models struggle to solve problems whi

researcharxiv-cs-cv
13 May 2026
Model Releases

Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics

DGX agent

arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation

DGX agent

arXiv:2605.11307v1 Announce Type: new Abstract: Image-to-code generation tests whether a vision-language model (VLM) can recover the structure of an image enough to express it as executable code. Exis

model-releasesarxiv-cs-cv
13 May 2026
Safety

VNDUQE: Information-Theoretic Novelty Detection using Deep Variational Information Bottleneck

DGX agent

arXiv:2605.11551v1 Announce Type: cross Abstract: Detecting out-of-distribution (OOD) samples is critical for safe deployment of neural networks in safety-critical applications. While maximum softmax

safetyarxiv-cs-cv
13 May 2026
Model Releases

Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

DGX agent

arXiv:2605.11654v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL), which matches an oblique drone view to a geo-referenced satellite tile, has emerged as a key alternative for autonom

model-releasesarxiv-cs-cv
13 May 2026
Research

Welfare as a Guiding Principle for Machine Learning -- From Compass, to Lens, to Roadmap

DGX agent

arXiv:2502.11981v3 Announce Type: replace Abstract: Decades of research in machine learning have given us powerful tools for making accurate predictions. But when used in social settings and on human

researcharxiv-cs-lg
13 May 2026
Model Releases

What Does It Mean for a Medical AI System to Be Right?

DGX agent

arXiv:2605.11963v1 Announce Type: new Abstract: This paper examines what it means for a medical AI system to be right by grounding the question in a specific clinical context: the automatic classifica

model-releasesarxiv-cs-cv
13 May 2026
Tutorials

What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty

DGX agent

arXiv:2605.12281v1 Announce Type: new Abstract: What makes a word difficult to learn, and how does the difficulty depend on the learner's native language? We computationally model vocabulary difficult

tutorialsarxiv-cs-cl
13 May 2026
Safety

What-Where Transformer: A Slot-Centric Visual Backbone for Concurrent Representation and Localization

DGX agent

arXiv:2605.12021v1 Announce Type: new Abstract: Many image understanding tasks involve identifying what is present and where it appears. However, tasks that address where, such as object discovery, de

safetyarxiv-cs-cv
13 May 2026
Tutorials

When and How to Canonize: A Generalization Perspective

DGX agent

arXiv:2605.11008v1 Announce Type: new Abstract: While invariant architectures are standard for processing symmetric data, there is growing interest in achieving invariance by applying group averaging

tutorialsarxiv-cs-lg
13 May 2026
Research

When Brains Disagree: Biological Ambiguity Underlies the Challenge of Amyloid PET Synthesis from Structural MRI

DGX agent

arXiv:2605.11867v1 Announce Type: new Abstract: Structural MRI-to-amyloid PET synthesis has been proposed as a non-invasive alternative for amyloid assessment in Alzheimer's disease (AD). However, rep

researcharxiv-cs-cv
13 May 2026
Safety

When Does ell_2-Boosting Overfit Benignly? High-Dimensional Risk Asymptotics and the ell_1 Implicit Bias

DGX agent

arXiv:2605.06314v2 Announce Type: replace Abstract: Benign overfitting is well-characterized in ell_2 geometries, but its behavior under the ell_1 implicit bias of greedy ensembles remains challenging

safetyarxiv-cs-lg
13 May 2026
Research

When Emotion Becomes Trigger: Emotion-style dynamic Backdoor Attack Parasitising Large Language Models

DGX agent

arXiv:2605.11612v1 Announce Type: new Abstract: Backdoor vulnerabilities widely exist in the fine-tuning of large language models(LLMs). Most backdoor poisoning methods operate mainly at the token lev

researcharxiv-cs-cl
13 May 2026
Research

When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs

DGX agent

arXiv:2605.11559v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have become a key interface for visual reasoning and grounded question answering, yet they remain vulnerable to

researcharxiv-cs-cv
13 May 2026
Safety

When Policy Entropy Constraint Fails: Preserving Diversity in Flow-based RLHF via Perceptual Entropy

DGX agent

arXiv:2605.12112v1 Announce Type: new Abstract: RLHF is widely used to align flow-matching text-to-image models with human preferences, but often leads to severe diversity collapse after fine-tuning.

safetyarxiv-cs-cv
13 May 2026
Research

When the Gold Standard Isn't Necessarily Standard: Challenges of Evaluating the Translation of User-Generated Content

DGX agent

arXiv:2512.17738v2 Announce Type: replace Abstract: User-generated content (UGC) is characterised by frequent use of non-standard language, from spelling errors to expressive choices such as slang, ch

researcharxiv-cs-cl
13 May 2026
Safety

When to Ask a Question: Understanding Communication Strategies in Generative AI Tools

DGX agent

arXiv:2605.11240v1 Announce Type: cross Abstract: Generative AI models differ from traditional machine learning tools in that they allow users to provide as much or as little information as they choos

safetyarxiv-cs-lg
13 May 2026
Applications

Why Conclusions Diverge from the Same Observations: Formalizing World-Model Non-Identifiability via an Inference

DGX agent

arXiv:2605.12255v1 Announce Type: cross Abstract: When people share the same documents and observations yet reach different conclusions, the disagreement often shifts into a judgment that the other pa

applicationsarxiv-cs-lg
13 May 2026
Model Releases

WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting

DGX agent

arXiv:2605.11696v1 Announce Type: new Abstract: Recent single-image relighting methods, powered by advanced generative models, have achieved impressive photorealism on synthetic benchmarks. However, t

model-releasesarxiv-cs-cv
13 May 2026
Safety

World Action Models: The Next Frontier in Embodied AI

DGX agent

arXiv:2605.12090v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved strong semantic generalization for embodied policy learning, yet they learn reactive observation-to-

safetyarxiv-cs-cl
13 May 2026
← Previous
1…903904905906907…1282
Next →