AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
Human
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
13 May 2026

Towards Uncertainty-Aware Federated Granger Causal Learning

ApplicationsDGX agent

arXiv:2602.13004v2 Announce Type: replace Abstract: Granger causality recovers directed interactions from time-series data, but in many distributed systems, the data are vertically partitioned across

Towards Visually-Guided Movie Subtitle Translation for Indic Languages

ApplicationsDGX agent

arXiv:2605.11993v1 Announce Type: new Abstract: Movie subtitle translation is inherently multimodal, yet text-only systems often miss visual cues needed to convey emotion, action, and social nuance, e

Toxicity Detection Should Measure Contextual Harm, Not Text-Intrinsic Badness

SafetyDGX agent

arXiv:2503.16072v4 Announce Type: replace-cross Abstract: Toxicity detection has become core safety infrastructure for online moderation, dataset filtering, and deployed language-model systems. Yet mo

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TRACE: Temporal Routing with Autoregressive Cross-channel Experts for EEG Representation Learning

ResearchDGX agent

arXiv:2605.11380v1 Announce Type: new Abstract: Learning transferable representations for electroencephalography (EEG) remains challenging because EEG signals are inherently multi-channel and non-stat

Training-Inference Consistent Segmented Execution for Long-Context LLMs

ResearchDGX agent

arXiv:2605.11744v1 Announce Type: new Abstract: Transformer-based large language models face severe scalability challenges in long-context generation due to the computational and memory costs of full-

Training Transformers for KV Cache Compressibility

SafetyDGX agent

arXiv:2605.05971v2 Announce Type: replace Abstract: Long-context language modeling is increasingly constrained by the Key-Value (KV) cache, whose memory and decode-time access costs scale linearly wit

Trajectory-Agnostic Asteroid Detection in TESS with Deep Learning

Model ReleasesDGX agent

arXiv:2605.12391v1 Announce Type: cross Abstract: We present a novel method for extracting moving objects from TESS data using machine learning. Our approach uses two stacked 3D U-Nets with skip conne

Trajectory First: A Curriculum for Discovering Diverse Policies

SafetyDGX agent

arXiv:2506.01568v3 Announce Type: replace Abstract: Being able to solve a task in diverse ways makes agents more robust to task variations and less prone to local optima. In this context, constrained

Transferable Delay-Aware Reinforcement Learning via Implicit Causal Graph Modeling

SafetyDGX agent

arXiv:2605.12312v1 Announce Type: new Abstract: Random delays weaken the temporal correspondence between actions and subsequent state feedback, making it difficult for agents to identify the true prop

Transformer-Based Autonomous Driving Models and Deployment-Oriented Compression: A Survey

SafetyDGX agent

arXiv:2304.10891v2 Announce Type: replace-cross Abstract: Transformer-based models are becoming a central paradigm in autonomous driving because they can capture long-range spatial dependencies, multi

TriBand-BEV: Real-Time LiDAR-Only 3D Pedestrian Detection via Height-Aware BEV and High-Resolution Feature Fusion

HardwareDGX agent

arXiv:2605.12220v1 Announce Type: new Abstract: Safe autonomous agents and mobile robots need fast real time 3D perception, especially for vulnerable road users (VRUs) such as pedestrians. We introduc

Trust Region Inverse Reinforcement Learning: Explicit Dual Ascent using Local Policy Updates

SafetyDGX agent

arXiv:2605.11020v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) is typically formulated as maximizing entropy subject to matching the distribution of expert trajectories. Classica

Trust the Batch, On- or Off-Policy: Adaptive Policy Optimization for RL Post-Training

SafetyDGX agent

arXiv:2605.12380v1 Announce Type: new Abstract: Reinforcement learning is structurally harder than supervised learning because the policy changes the data distribution it learns from. The resulting fr

U-STS-LLM A Unified Spatio-Temporal Steered Large Language Model for Traffic Prediction and Imputation

Model ReleasesDGX agent

arXiv:2605.11735v1 Announce Type: new Abstract: The efficient operation of modern cellular networks hinges on the accurate analysis of spatio-temporal traffic data. Mastering these patterns is essenti

UGround: Towards Unified Visual Grounding with Unrolled Transformers

SafetyDGX agent

arXiv:2510.03853v4 Announce Type: replace Abstract: We present UGround, a extbf{U}nified visual extbf{Ground}ing paradigm that dynamically selects intermediate layers across extbf{U}nrolled transforme

UHR-Micro: Diagnosing and Mitigating the Resolution Illusion in Earth Observation VLMs

Model ReleasesDGX agent

arXiv:2605.12237v1 Announce Type: new Abstract: Vision-Language Models (VLMs) increasingly operate on ultra-high-resolution (UHR) Earth observation imagery, yet they remain vulnerable to a severe scal

Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization

SafetyDGX agent

arXiv:2605.11491v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning ability of large language models. How

Understanding Sample Efficiency in Predictive Coding

SafetyDGX agent

arXiv:2605.11911v1 Announce Type: new Abstract: Predictive Coding (PC) is an influential account of cortical learning. Much of recent work has focused on comparing PC to Backpropagation (BP) to find w

Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPO

SafetyDGX agent

arXiv:2505.19770v5 Announce Type: replace-cross Abstract: We present a fine-grained theoretical analysis of the performance gap between two-stage reinforcement learning from human feedback~(RLHF) and

UnfoldLDM: Degradation-Aware Unfolding with Iterative Latent Diffusion Priors for Blind Image Restoration

Model ReleasesDGX agent

arXiv:2511.18152v3 Announce Type: replace Abstract: Deep unfolding networks (DUNs) combine the interpretability of model-based methods with the learning ability of deep networks, yet remain limited fo

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation

TutorialsDGX agent

arXiv:2605.12088v1 Announce Type: new Abstract: Multi-reference image generation aims to synthesize images from textual instructions while faithfully preserving subject identities from multiple refere

UniFixer: A Universal Reference-Guided Fixer for Diffusion-Based View Synthesis

SafetyDGX agent

arXiv:2605.12169v1 Announce Type: new Abstract: With the recent surge of generative models, diffusion-based approaches have become mainstream for view synthesis tasks, either in an explicit depth-warp

Uniform Scaling Limits in AdamW-Trained Transformers

ResearchDGX agent

arXiv:2605.11059v1 Announce Type: cross Abstract: We study the large-depth limit of transformers trained with AdamW, by modelling the hidden-state dynamics as an interacting particle system (IPS) coup

UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs

ApplicationsDGX agent

arXiv:2605.11856v1 Announce Type: cross Abstract: Multimodal large language models are increasingly expected to perform thinking with images, yet existing visual latent reasoning methods still rely on

Unlearning with Asymmetric Sources: Improved Unlearning-Utility Trade-off with Public Data

ResearchDGX agent

arXiv:2605.11170v1 Announce Type: new Abstract: Noise-based certified machine unlearning currently faces a hard ceiling: the noise magnitude required to certify unlearning typically destroys model uti

Unlocking Compositional Generalization in Continual Few-Shot Learning

ResearchDGX agent

arXiv:2605.11710v1 Announce Type: cross Abstract: Object-centric representations promise a key property for few-shot learning: Rather than treating a scene as a single unit, a model can decompose it i

Unlocking LLM Creativity in Science through Analogical Reasoning

AgentsDGX agent

arXiv:2605.11258v1 Announce Type: cross Abstract: Autonomous science promises to augment scientific discovery, particularly in complex fields like biomedicine. However, this requires AI systems that c

Unlocking UML Class Diagram Understanding in Vision Language Models

Model ReleasesDGX agent

arXiv:2605.11634v1 Announce Type: new Abstract: Although Vision Language Models (VLMs) have seen tremendous progress across all kinds of use cases, they still fall behind in answering questions regard

Unpacking the Eye of the Beholder: Social Location, Identity, and the Moving Target of Political Perspectives

ResearchDGX agent

arXiv:2605.11166v1 Announce Type: new Abstract: Political and social identities structure how people evaluate political information, a finding decades deep in political science and routinely discarded

Urban Risk-Aware Navigation via VQA-Based Event Maps for People with Low Vision

Model ReleasesDGX agent

arXiv:2605.11782v1 Announce Type: new Abstract: Visual impairment affects hundreds of millions of people worldwide, severely limiting their ability to navigate urban environments safely and independen

USEMA: a Scalable Efficient Mamba Like Attention for Medical Image Segmentation

ResearchDGX agent

arXiv:2605.11131v1 Announce Type: new Abstract: Accurate medical image segmentation is an integral part of the medical image analysis pipeline that requires the ability to merge local and global infor

Variance-aware Reward Modeling with Anchor Guidance

ApplicationsDGX agent

arXiv:2605.11865v1 Announce Type: cross Abstract: Standard Bradley--Terry (BT) reward models are limited when human preferences are pluralistic. Although soft preference labels preserve disagreement i

Variational Linear Attention: Stable Associative Memory for Long-Context Transformers

ResearchDGX agent

arXiv:2605.11196v1 Announce Type: new Abstract: Linear attention reduces the quadratic cost of softmax attention to O(T), but its memory state grows as O(T) in Frobenius norm, causing progressive inte

Vector Scaffolding: Inter-Scale Orchestration for Differentiable Image Vectorization

ResearchDGX agent

arXiv:2605.11913v1 Announce Type: new Abstract: Differentiable vector graphics have enabled powerful gradient-based optimization of vector primitives directly from raster images. However, existing fra

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference

Model ReleasesDGX agent

arXiv:2605.11334v1 Announce Type: cross Abstract: LLM-as-Judge systems are widely deployed for automated evaluation, yet practitioners lack reliable methods to know when a judge's verdict should be tr

Vertex-Softmax: Tight Transformer Verification via Exact Softmax Optimization

ResearchDGX agent

arXiv:2605.10974v1 Announce Type: new Abstract: Certified verification of transformer attention requires bounding the softmax function over interval constraints on the pre-softmax scores. Existing ver

Very Efficient Listwise Multimodal Reranking for Long Documents

Model ReleasesDGX agent

arXiv:2605.11864v1 Announce Type: cross Abstract: Listwise reranking is a key yet computationally expensive component in vision-centric retrieval and multimodal retrieval-augmented generation (M-RAG)

Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives

ResearchDGX agent

arXiv:2406.05615v4 Announce Type: replace Abstract: Humans use multiple senses to comprehend the environment. Vision and language are two of the most vital senses since they allow us to easily communi

VidSplat: Gaussian Splatting Reconstruction with Geometry-Guided Video Diffusion Priors

ResearchDGX agent

arXiv:2605.11424v1 Announce Type: new Abstract: Gaussian Splatting has achieved remarkable progress in multi-view surface reconstruction, yet it exhibits notable degradation when only few views are av

VIP: Visual-guided Prompt Evolution for Efficient Dense Vision-Language Inference

SafetyDGX agent

arXiv:2605.12325v1 Announce Type: new Abstract: Pursuing training-free open-vocabulary semantic segmentation in an efficient and generalizable manner remains challenging due to the deep-seated spatial

Vision-aligned Latent Reasoning for Multi-modal Large Language Model

ResearchDGX agent

arXiv:2602.04476v2 Announce Type: replace Abstract: Despite recent advancements in Multi-modal Large Language Models (MLLMs) on diverse understanding tasks, these models struggle to solve problems whi

Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics

Model ReleasesDGX agent

arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W

Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation

Model ReleasesDGX agent

arXiv:2605.11307v1 Announce Type: new Abstract: Image-to-code generation tests whether a vision-language model (VLM) can recover the structure of an image enough to express it as executable code. Exis

VNDUQE: Information-Theoretic Novelty Detection using Deep Variational Information Bottleneck

SafetyDGX agent

arXiv:2605.11551v1 Announce Type: cross Abstract: Detecting out-of-distribution (OOD) samples is critical for safe deployment of neural networks in safety-critical applications. While maximum softmax

Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

Model ReleasesDGX agent

arXiv:2605.11654v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL), which matches an oblique drone view to a geo-referenced satellite tile, has emerged as a key alternative for autonom

Welfare as a Guiding Principle for Machine Learning -- From Compass, to Lens, to Roadmap

ResearchDGX agent

arXiv:2502.11981v3 Announce Type: replace Abstract: Decades of research in machine learning have given us powerful tools for making accurate predictions. But when used in social settings and on human

What Does It Mean for a Medical AI System to Be Right?

Model ReleasesDGX agent

arXiv:2605.11963v1 Announce Type: new Abstract: This paper examines what it means for a medical AI system to be right by grounding the question in a specific clinical context: the automatic classifica

What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty

TutorialsDGX agent

arXiv:2605.12281v1 Announce Type: new Abstract: What makes a word difficult to learn, and how does the difficulty depend on the learner's native language? We computationally model vocabulary difficult

What-Where Transformer: A Slot-Centric Visual Backbone for Concurrent Representation and Localization

SafetyDGX agent

arXiv:2605.12021v1 Announce Type: new Abstract: Many image understanding tasks involve identifying what is present and where it appears. However, tasks that address where, such as object discovery, de

When and How to Canonize: A Generalization Perspective

TutorialsDGX agent

arXiv:2605.11008v1 Announce Type: new Abstract: While invariant architectures are standard for processing symmetric data, there is growing interest in achieving invariance by applying group averaging

When Brains Disagree: Biological Ambiguity Underlies the Challenge of Amyloid PET Synthesis from Structural MRI

ResearchDGX agent

arXiv:2605.11867v1 Announce Type: new Abstract: Structural MRI-to-amyloid PET synthesis has been proposed as a non-invasive alternative for amyloid assessment in Alzheimer's disease (AD). However, rep

When Does ell_2-Boosting Overfit Benignly? High-Dimensional Risk Asymptotics and the ell_1 Implicit Bias

SafetyDGX agent

arXiv:2605.06314v2 Announce Type: replace Abstract: Benign overfitting is well-characterized in ell_2 geometries, but its behavior under the ell_1 implicit bias of greedy ensembles remains challenging

When Emotion Becomes Trigger: Emotion-style dynamic Backdoor Attack Parasitising Large Language Models

ResearchDGX agent

arXiv:2605.11612v1 Announce Type: new Abstract: Backdoor vulnerabilities widely exist in the fine-tuning of large language models(LLMs). Most backdoor poisoning methods operate mainly at the token lev

When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs

ResearchDGX agent

arXiv:2605.11559v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have become a key interface for visual reasoning and grounded question answering, yet they remain vulnerable to

When Policy Entropy Constraint Fails: Preserving Diversity in Flow-based RLHF via Perceptual Entropy

SafetyDGX agent

arXiv:2605.12112v1 Announce Type: new Abstract: RLHF is widely used to align flow-matching text-to-image models with human preferences, but often leads to severe diversity collapse after fine-tuning.

When the Gold Standard Isn't Necessarily Standard: Challenges of Evaluating the Translation of User-Generated Content

ResearchDGX agent

arXiv:2512.17738v2 Announce Type: replace Abstract: User-generated content (UGC) is characterised by frequent use of non-standard language, from spelling errors to expressive choices such as slang, ch

When to Ask a Question: Understanding Communication Strategies in Generative AI Tools

SafetyDGX agent

arXiv:2605.11240v1 Announce Type: cross Abstract: Generative AI models differ from traditional machine learning tools in that they allow users to provide as much or as little information as they choos

Why Conclusions Diverge from the Same Observations: Formalizing World-Model Non-Identifiability via an Inference

ApplicationsDGX agent

arXiv:2605.12255v1 Announce Type: cross Abstract: When people share the same documents and observations yet reach different conclusions, the disagreement often shifts into a judgment that the other pa

WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting

Model ReleasesDGX agent

arXiv:2605.11696v1 Announce Type: new Abstract: Recent single-image relighting methods, powered by advanced generative models, have achieved impressive photorealism on synthetic benchmarks. However, t

World Action Models: The Next Frontier in Embodied AI

SafetyDGX agent

arXiv:2605.12090v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved strong semantic generalization for embodied policy learning, yet they learn reactive observation-to-

← Previous
1…722723724725726…1025
Next →