AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
Human
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,904 results
Research

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling

DGX agent

arXiv:2605.05922v2 Announce Type: replace Abstract: Recent advances in generative video models are increasingly driven by post-training and test-time scaling, both of which critically depend on the qu

researcharxiv-cs-cv
13 May 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Three Regimes of Context-Parametric Conflict: A Predictive Framework and Empirical Validation

DGX agent

arXiv:2605.11574v1 Announce Type: new Abstract: The literature on how large language models handle conflict between their training knowledge and a contradicting document presents a persistent empirica

model-releasesarxiv-cs-cl
13 May 2026
Safety

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment

DGX agent

arXiv:2605.10983v1 Announce Type: cross Abstract: Reinforcement learning (RL) has shown extraordinary potential in aligning diffusion models to downstream tasks, yet most of them still suffer from sig

safetyarxiv-cs-cv
13 May 2026
Safety

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning

DGX agent

arXiv:2605.12236v1 Announce Type: cross Abstract: Fine-tuning pre-trained robot policies with reinforcement learning (RL) often inherits the bottlenecks introduced by pre-training with behavioral clon

safetyarxiv-cs-lg
13 May 2026
Hardware

To Err Is Human; To Annotate, SILICON? Toward Robust Reproducibility in LLM Annotation

DGX agent

arXiv:2412.14461v4 Announce Type: replace Abstract: Unstructured text data annotation is foundational to management research. LLMs offer a cost-effective and scalable alternative to human annotation,

hardwarearxiv-cs-cl
13 May 2026
Safety

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching

DGX agent

arXiv:2605.12288v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is a widely used RL-free method for aligning language models from pairwise preferences, but it models preferences o

safetyarxiv-cs-cl
13 May 2026
Model Releases

TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing

DGX agent

arXiv:2605.11473v1 Announce Type: cross Abstract: Soft Actor-Critic (SAC) and its variants dominate Multi-Task Reinforcement Learning (MTRL) due to their off-policy sample efficiency, while on-policy

model-releasesarxiv-cs-lg
13 May 2026
Applications

Towards Affordable Energy: A Gymnasium Environment for Electric Utility Demand-Response Programs

DGX agent

arXiv:2605.12462v1 Announce Type: cross Abstract: Extreme weather and volatile wholesale electricity markets expose residential consumers to catastrophic financial risks, yet demand response at the di

applicationsarxiv-cs-lg
13 May 2026
Safety

Towards Fine-Grained Code-Switch Speech Translation with Semantic Space Alignment

DGX agent

arXiv:2511.10670v2 Announce Type: replace Abstract: Code-switching (CS) speech translation (ST) aims to translate speech that alternates between multiple languages into a target language text, posing

safetyarxiv-cs-cl
13 May 2026
Safety

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization

DGX agent

arXiv:2605.11974v1 Announce Type: new Abstract: Large Language Models (LLMs) suffer from order bias, where their performance is affected by the arrangement order of input elements. This unfairness lim

safetyarxiv-cs-lg
13 May 2026
Applications

Towards Uncertainty-Aware Federated Granger Causal Learning

DGX agent

arXiv:2602.13004v2 Announce Type: replace Abstract: Granger causality recovers directed interactions from time-series data, but in many distributed systems, the data are vertically partitioned across

applicationsarxiv-cs-lg
13 May 2026
Applications

Towards Visually-Guided Movie Subtitle Translation for Indic Languages

DGX agent

arXiv:2605.11993v1 Announce Type: new Abstract: Movie subtitle translation is inherently multimodal, yet text-only systems often miss visual cues needed to convey emotion, action, and social nuance, e

applicationsarxiv-cs-cl
13 May 2026
Safety

Toxicity Detection Should Measure Contextual Harm, Not Text-Intrinsic Badness

DGX agent

arXiv:2503.16072v4 Announce Type: replace-cross Abstract: Toxicity detection has become core safety infrastructure for online moderation, dataset filtering, and deployed language-model systems. Yet mo

safetyarxiv-cs-cl
13 May 2026
Research

TRACE: Temporal Routing with Autoregressive Cross-channel Experts for EEG Representation Learning

DGX agent

arXiv:2605.11380v1 Announce Type: new Abstract: Learning transferable representations for electroencephalography (EEG) remains challenging because EEG signals are inherently multi-channel and non-stat

researcharxiv-cs-lg
13 May 2026
Research

Training-Inference Consistent Segmented Execution for Long-Context LLMs

DGX agent

arXiv:2605.11744v1 Announce Type: new Abstract: Transformer-based large language models face severe scalability challenges in long-context generation due to the computational and memory costs of full-

researcharxiv-cs-cl
13 May 2026
Safety

Training Transformers for KV Cache Compressibility

DGX agent

arXiv:2605.05971v2 Announce Type: replace Abstract: Long-context language modeling is increasingly constrained by the Key-Value (KV) cache, whose memory and decode-time access costs scale linearly wit

safetyarxiv-cs-lg
13 May 2026
Model Releases

Trajectory-Agnostic Asteroid Detection in TESS with Deep Learning

DGX agent

arXiv:2605.12391v1 Announce Type: cross Abstract: We present a novel method for extracting moving objects from TESS data using machine learning. Our approach uses two stacked 3D U-Nets with skip conne

model-releasesarxiv-cs-lg
13 May 2026
Safety

Trajectory First: A Curriculum for Discovering Diverse Policies

DGX agent

arXiv:2506.01568v3 Announce Type: replace Abstract: Being able to solve a task in diverse ways makes agents more robust to task variations and less prone to local optima. In this context, constrained

safetyarxiv-cs-lg
13 May 2026
Safety

Transferable Delay-Aware Reinforcement Learning via Implicit Causal Graph Modeling

DGX agent

arXiv:2605.12312v1 Announce Type: new Abstract: Random delays weaken the temporal correspondence between actions and subsequent state feedback, making it difficult for agents to identify the true prop

safetyarxiv-cs-lg
13 May 2026
Safety

Transformer-Based Autonomous Driving Models and Deployment-Oriented Compression: A Survey

DGX agent

arXiv:2304.10891v2 Announce Type: replace-cross Abstract: Transformer-based models are becoming a central paradigm in autonomous driving because they can capture long-range spatial dependencies, multi

safetyarxiv-cs-cv
13 May 2026
Hardware

TriBand-BEV: Real-Time LiDAR-Only 3D Pedestrian Detection via Height-Aware BEV and High-Resolution Feature Fusion

DGX agent

arXiv:2605.12220v1 Announce Type: new Abstract: Safe autonomous agents and mobile robots need fast real time 3D perception, especially for vulnerable road users (VRUs) such as pedestrians. We introduc

hardwarearxiv-cs-cv
13 May 2026
Safety

Trust Region Inverse Reinforcement Learning: Explicit Dual Ascent using Local Policy Updates

DGX agent

arXiv:2605.11020v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) is typically formulated as maximizing entropy subject to matching the distribution of expert trajectories. Classica

safetyarxiv-cs-lg
13 May 2026
Safety

Trust the Batch, On- or Off-Policy: Adaptive Policy Optimization for RL Post-Training

DGX agent

arXiv:2605.12380v1 Announce Type: new Abstract: Reinforcement learning is structurally harder than supervised learning because the policy changes the data distribution it learns from. The resulting fr

safetyarxiv-cs-lg
13 May 2026
Model Releases

U-STS-LLM A Unified Spatio-Temporal Steered Large Language Model for Traffic Prediction and Imputation

DGX agent

arXiv:2605.11735v1 Announce Type: new Abstract: The efficient operation of modern cellular networks hinges on the accurate analysis of spatio-temporal traffic data. Mastering these patterns is essenti

model-releasesarxiv-cs-lg
13 May 2026
Safety

UGround: Towards Unified Visual Grounding with Unrolled Transformers

DGX agent

arXiv:2510.03853v4 Announce Type: replace Abstract: We present UGround, a extbf{U}nified visual extbf{Ground}ing paradigm that dynamically selects intermediate layers across extbf{U}nrolled transforme

safetyarxiv-cs-cv
13 May 2026
Model Releases

UHR-Micro: Diagnosing and Mitigating the Resolution Illusion in Earth Observation VLMs

DGX agent

arXiv:2605.12237v1 Announce Type: new Abstract: Vision-Language Models (VLMs) increasingly operate on ultra-high-resolution (UHR) Earth observation imagery, yet they remain vulnerable to a severe scal

model-releasesarxiv-cs-cv
13 May 2026
Safety

Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization

DGX agent

arXiv:2605.11491v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning ability of large language models. How

safetyarxiv-cs-lg
13 May 2026
Safety

Understanding Sample Efficiency in Predictive Coding

DGX agent

arXiv:2605.11911v1 Announce Type: new Abstract: Predictive Coding (PC) is an influential account of cortical learning. Much of recent work has focused on comparing PC to Backpropagation (BP) to find w

safetyarxiv-cs-lg
13 May 2026
Safety

Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPO

DGX agent

arXiv:2505.19770v5 Announce Type: replace-cross Abstract: We present a fine-grained theoretical analysis of the performance gap between two-stage reinforcement learning from human feedback~(RLHF) and

safetyarxiv-cs-cl
13 May 2026
Model Releases

UnfoldLDM: Degradation-Aware Unfolding with Iterative Latent Diffusion Priors for Blind Image Restoration

DGX agent

arXiv:2511.18152v3 Announce Type: replace Abstract: Deep unfolding networks (DUNs) combine the interpretability of model-based methods with the learning ability of deep networks, yet remain limited fo

model-releasesarxiv-cs-cv
13 May 2026
Tutorials

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation

DGX agent

arXiv:2605.12088v1 Announce Type: new Abstract: Multi-reference image generation aims to synthesize images from textual instructions while faithfully preserving subject identities from multiple refere

tutorialsarxiv-cs-cv
13 May 2026
Safety

UniFixer: A Universal Reference-Guided Fixer for Diffusion-Based View Synthesis

DGX agent

arXiv:2605.12169v1 Announce Type: new Abstract: With the recent surge of generative models, diffusion-based approaches have become mainstream for view synthesis tasks, either in an explicit depth-warp

safetyarxiv-cs-cv
13 May 2026
Research

Uniform Scaling Limits in AdamW-Trained Transformers

DGX agent

arXiv:2605.11059v1 Announce Type: cross Abstract: We study the large-depth limit of transformers trained with AdamW, by modelling the hidden-state dynamics as an interacting particle system (IPS) coup

researcharxiv-cs-lg
13 May 2026
Applications

UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs

DGX agent

arXiv:2605.11856v1 Announce Type: cross Abstract: Multimodal large language models are increasingly expected to perform thinking with images, yet existing visual latent reasoning methods still rely on

applicationsarxiv-cs-cl
13 May 2026
Research

Unlearning with Asymmetric Sources: Improved Unlearning-Utility Trade-off with Public Data

DGX agent

arXiv:2605.11170v1 Announce Type: new Abstract: Noise-based certified machine unlearning currently faces a hard ceiling: the noise magnitude required to certify unlearning typically destroys model uti

researcharxiv-cs-lg
13 May 2026
Research

Unlocking Compositional Generalization in Continual Few-Shot Learning

DGX agent

arXiv:2605.11710v1 Announce Type: cross Abstract: Object-centric representations promise a key property for few-shot learning: Rather than treating a scene as a single unit, a model can decompose it i

researcharxiv-cs-cv
13 May 2026
Agents

Unlocking LLM Creativity in Science through Analogical Reasoning

DGX agent

arXiv:2605.11258v1 Announce Type: cross Abstract: Autonomous science promises to augment scientific discovery, particularly in complex fields like biomedicine. However, this requires AI systems that c

agentsarxiv-cs-cl
13 May 2026
Model Releases

Unlocking UML Class Diagram Understanding in Vision Language Models

DGX agent

arXiv:2605.11634v1 Announce Type: new Abstract: Although Vision Language Models (VLMs) have seen tremendous progress across all kinds of use cases, they still fall behind in answering questions regard

model-releasesarxiv-cs-cv
13 May 2026
Research

Unpacking the Eye of the Beholder: Social Location, Identity, and the Moving Target of Political Perspectives

DGX agent

arXiv:2605.11166v1 Announce Type: new Abstract: Political and social identities structure how people evaluate political information, a finding decades deep in political science and routinely discarded

researcharxiv-cs-cv
13 May 2026
Model Releases

Urban Risk-Aware Navigation via VQA-Based Event Maps for People with Low Vision

DGX agent

arXiv:2605.11782v1 Announce Type: new Abstract: Visual impairment affects hundreds of millions of people worldwide, severely limiting their ability to navigate urban environments safely and independen

model-releasesarxiv-cs-cv
13 May 2026
Research

USEMA: a Scalable Efficient Mamba Like Attention for Medical Image Segmentation

DGX agent

arXiv:2605.11131v1 Announce Type: new Abstract: Accurate medical image segmentation is an integral part of the medical image analysis pipeline that requires the ability to merge local and global infor

researcharxiv-cs-cv
13 May 2026
Applications

Variance-aware Reward Modeling with Anchor Guidance

DGX agent

arXiv:2605.11865v1 Announce Type: cross Abstract: Standard Bradley--Terry (BT) reward models are limited when human preferences are pluralistic. Although soft preference labels preserve disagreement i

applicationsarxiv-cs-lg
13 May 2026
Research

Variational Linear Attention: Stable Associative Memory for Long-Context Transformers

DGX agent

arXiv:2605.11196v1 Announce Type: new Abstract: Linear attention reduces the quadratic cost of softmax attention to O(T), but its memory state grows as O(T) in Frobenius norm, causing progressive inte

researcharxiv-cs-lg
13 May 2026
Research

Vector Scaffolding: Inter-Scale Orchestration for Differentiable Image Vectorization

DGX agent

arXiv:2605.11913v1 Announce Type: new Abstract: Differentiable vector graphics have enabled powerful gradient-based optimization of vector primitives directly from raster images. However, existing fra

researcharxiv-cs-cv
13 May 2026
Model Releases

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference

DGX agent

arXiv:2605.11334v1 Announce Type: cross Abstract: LLM-as-Judge systems are widely deployed for automated evaluation, yet practitioners lack reliable methods to know when a judge's verdict should be tr

model-releasesarxiv-cs-cl
13 May 2026
Research

Vertex-Softmax: Tight Transformer Verification via Exact Softmax Optimization

DGX agent

arXiv:2605.10974v1 Announce Type: new Abstract: Certified verification of transformer attention requires bounding the softmax function over interval constraints on the pre-softmax scores. Existing ver

researcharxiv-cs-lg
13 May 2026
Model Releases

Very Efficient Listwise Multimodal Reranking for Long Documents

DGX agent

arXiv:2605.11864v1 Announce Type: cross Abstract: Listwise reranking is a key yet computationally expensive component in vision-centric retrieval and multimodal retrieval-augmented generation (M-RAG)

model-releasesarxiv-cs-cv
13 May 2026
Research

Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives

DGX agent

arXiv:2406.05615v4 Announce Type: replace Abstract: Humans use multiple senses to comprehend the environment. Vision and language are two of the most vital senses since they allow us to easily communi

researcharxiv-cs-cl
13 May 2026
← Previous
1…911912913914915…1290
Next →