AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
Human
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
60,292 results
11 May 2026

TTF: Temporal Token Fusion for Efficient Video-Language Model

ResearchDGX agent

arXiv:2605.07355v1 Announce Type: cross Abstract: Video-language models (VLMs) face rapid inference costs as visual token counts scale with video length. For example, 32 frames at 448{imes}448 resolut

TUANDROMD-X: Advanced Entropy and Visual Analytics Dataset for Enhanced Malware Detection and Classification

ResearchDGX agent

arXiv:2605.06718v1 Announce Type: cross Abstract: Malware and malware-based attacks are becoming more prevalent and complex. Attackers regularly come up with new techniques that have the ability to ev

Tyche: One Step Flow for Efficient Probabilistic Weather Forecasting

ResearchDGX agent

arXiv:2605.06916v1 Announce Type: new Abstract: Probabilistic weather forecasting requires not only accurate trajectories, but calibrated distributions over plausible atmospheric futures. Recent data-

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

UFT: Unifying Fine-Tuning of SFT and RLHF/DPO/UNA through a Generalized Implicit Reward Function

SafetyDGX agent

arXiv:2410.21438v3 Announce Type: replace Abstract: By pretraining on trillions of tokens, an LLM gains the capability of text generation. However, to enhance its utility and reduce potential harm, SF

UNA: A Unified Supervised Framework for Efficient LLM Alignment Across Feedback Types

SafetyDGX agent

arXiv:2408.15339v4 Announce Type: replace-cross Abstract: RL alignment methods, including RLHF and DPO, are primarily based on pairwise preference data. Although scalar or score-based feedback has bee

Uncertainty-Aware Structured Data Extraction from Full CMR Reports via Distilled LLMs

ResearchDGX agent

arXiv:2605.08045v1 Announce Type: new Abstract: Converting free-text cardiac magnetic resonance (CMR) reports into auditable structured data remains a bottleneck for cohort assembly, longitudinal cura

Uncertainty Quantification for Cardiac Shape Reconstruction with Deep Signed Distance Functions via MCMC methods

ResearchDGX agent

arXiv:2605.07987v1 Announce Type: cross Abstract: Atlas-based approaches allow high-quality, patient-specific shape reconstructions of cardiac anatomy from sparse and/or noisy data such as point cloud

Uncertainty Quantification for Prior-Data Fitted Networks using Martingale Posteriors

ApplicationsDGX agent

arXiv:2505.11325v4 Announce Type: replace-cross Abstract: Prior-data fitted networks (PFNs) have emerged as promising foundation models for prediction from tabular datasets, achieving state-of-the-art

UNCOM: Zero-shot Context-Aware Command Understanding for Tabletop Scenarios

Model ReleasesDGX agent

arXiv:2410.06355v3 Announce Type: replace-cross Abstract: This paper presents UNCOM, a novel hybrid framework for interpreting natural human commands in tabletop scenarios. The system integrates multi

Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models

ApplicationsDGX agent

arXiv:2605.07148v1 Announce Type: new Abstract: Decades of cognitive science establish that humans navigate environments by forming cognitive maps, defined as allocentric and topology-preserving repre

Uncovering Hidden Systematics in Neural Network Models for High Energy Physics

TutorialsDGX agent

arXiv:2605.07470v1 Announce Type: new Abstract: Neural networks (NNs) are inherently multidimensional classifiers that learn complex, non-linear relationships among input observables. While their flex

Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions

ResearchDGX agent

arXiv:2605.07271v1 Announce Type: cross Abstract: Layer pruning efficiently reduces Large Language Model (LLM) computational costs but often triggers sudden performance collapse. Existing representati

Understanding Robustness of Model Editing in Code LLMs

Model ReleasesDGX agent

arXiv:2511.03182v2 Announce Type: replace-cross Abstract: Large language models (LLMs) for code are increasingly used in software development, but they remain static after pretraining while APIs and s

Uneven Evolution of Cognition Across Generations of Generative AI Models

Model ReleasesDGX agent

arXiv:2605.06815v1 Announce Type: new Abstract: The pursuit of artificial general intelligence necessitates robust methods for evaluating the cognitive capabilities of models beyond narrow task perfor

UniD-Shift: Towards Unified Semantic Segmentation via Interpretable Share-Private Multimodal Decomposition

SafetyDGX agent

arXiv:2605.07356v1 Announce Type: new Abstract: Semantic segmentation of large-scale 3D point clouds is crucial for applications such as autonomous driving and urban digital twins. However, the sparse

UniISP: A Unified ISP Framework for Both Human and Machine Vision

ResearchDGX agent

arXiv:2605.07359v1 Announce Type: new Abstract: Compared to RGB images, raw sensor data provides a richer representation of information, which is crucial for accurate recognition, particularly under c

UniV2D: Bridging Visual Restoration and Semantic Perception for Underwater Salient Object Detection

TutorialsDGX agent

arXiv:2605.07146v1 Announce Type: new Abstract: Underwater salient object detection (USOD) plays a vital role in marine vision tasks but remains fundamentally challenging due to severe visual degradat

Unlocking High-Fidelity Molecular Generation from Mass Spectra via Dual-Stream Line Graph Diffusion

Local AiDGX agent

arXiv:2605.07048v1 Announce Type: cross Abstract: De novo molecular generation from tandem mass spectra is a challenging inverse problem whose core difficulty lies in the circular dependency between a

Unsolvability Ceiling in Multi-LLM Routing: An Empirical Study of Evaluation Artifacts

Model ReleasesDGX agent

arXiv:2605.07395v1 Announce Type: cross Abstract: Efficient routing across multiple LLMs enables cost-quality tradeoffs by directing queries to the cheapest capable model. Prior work attributes routin

Upper Generalization Bounds for Neural Oscillators

ResearchDGX agent

arXiv:2603.09742v2 Announce Type: replace Abstract: Neural oscillators that originate from second-order ordinary differential equations (ODEs) have shown competitive performance in learning mappings b

User eXperience Perception Insights Dataset (UXPID): Synthetic User Feedback from Public Industrial Forums

ApplicationsDGX agent

arXiv:2509.11777v2 Announce Type: replace Abstract: Customer feedback in industrial forums offers rich but underexplored insights into real-world product experience. Yet systematic analysis remains ch

Utility-Preserving De-Identification for Math Tutoring: Investigating Numeric Ambiguity in the MathEd-PII Benchmark Dataset

Model ReleasesDGX agent

arXiv:2602.16571v2 Announce Type: replace Abstract: Large-scale sharing of dialogue data is key to advancing the science of teaching and learning, yet rigorous de-identification remains a major barrie

Vaporizer: Breaking Watermarking Schemes for Large Language Model Outputs

ApplicationsDGX agent

arXiv:2605.07481v1 Announce Type: cross Abstract: In this paper, we investigate the recent state-of-the-art schemes for watermarking large language models (LLMs) outputs. These techniques are claimed

Variable Aerodynamic Damping via Co-Contraction: A Dynamic Isomorphism with Variable Stiffness Actuators

Local AiDGX agent

arXiv:2605.07292v1 Announce Type: new Abstract: We prove that aerodynamic co-contraction in a redundant dual-rotor actuator can tune a passive, trim-defined aero-mechanical damping while keeping the c

VDCook:DIY video data cook your MLLMs

AgentsDGX agent

arXiv:2603.05539v2 Announce Type: replace-cross Abstract: We introduce VDCook: a self-evolving video data operating system, a configurable video data construction platform for researchers and vertical

VDEGaussian: Video Diffusion Enhanced 4D Gaussian Splatting for Dynamic Urban Scenes Modeling

SafetyDGX agent

arXiv:2508.02129v2 Announce Type: replace Abstract: Dynamic urban scene modeling is a rapidly evolving area with broad applications. While current approaches leveraging neural radiance fields or Gauss

VecCISC: Improving Confidence-Informed Self-Consistency with Reasoning Trace Clustering and Candidate Answer Selection

ResearchDGX agent

arXiv:2605.08070v1 Announce Type: new Abstract: A standard technique for scaling inference-time reasoning is Self-Consistency, whereby multiple candidate answers are sampled from an LLM and the most c

Velocity-Space 3D Asset Editing

ResearchDGX agent

arXiv:2605.07385v1 Announce Type: cross Abstract: Editing a 3D asset locally, modifying a target region while preserving the rest, is a fundamental requirement of native 3D editing. Existing methods e

Versatile yet Efficient Network Traffic Analysis: Offloading Network Foundation Model to SmartNIC

HardwareDGX agent

arXiv:2508.02001v2 Announce Type: replace-cross Abstract: Pervasive encryption makes large-scale labeling infeasible for traffic analysis, while security operations demand edge analysis to avert servi

VESPO: Variational Sequence-Level Soft Policy Optimization for Stable Off-Policy LLM Training

SafetyDGX agent

arXiv:2602.10693v3 Announce Type: replace-cross Abstract: Off-policy updates are inevitable in reinforcement learning (RL) for large language models (LLMs) due to rollout staleness from asynchronous t

VesselRW: Weakly Supervised Subcutaneous Vessel Segmentation via Learned Random Walk Propagation

TutorialsDGX agent

arXiv:2508.06819v3 Announce Type: replace Abstract: The task of parsing subcutaneous vessels in clinical images is often hindered by the high cost and limited availability of ground truth data, as wel

Vibe coding before the trend

ApplicationsDGX agent

arXiv:2605.07751v1 Announce Type: cross Abstract: Early 2025 we ran a series of vibe coding challenges across four different student cohorts. The cohorts included 54 ICT students, 24 digital marketing

VIDEE: Visual and Interactive Decomposition, Execution, and Evaluation of Text Analytics with Intelligent Agents

AgentsDGX agent

arXiv:2506.21582v5 Announce Type: replace-cross Abstract: Text analytics has traditionally required specialized knowledge in Natural Language Processing (NLP) or text analysis, which presents a barrie

Video Understanding Reward Modeling: A Robust Benchmark and Performant Reward Models

Model ReleasesDGX agent

arXiv:2605.07872v1 Announce Type: cross Abstract: Multimodal reward models have advanced substantially in text and image domains, yet progress in video understanding reward modeling remains severely l

VideoRouter: Query-Adaptive Dual Routing for Efficient Long-Video Understanding

SafetyDGX agent

arXiv:2605.05848v2 Announce Type: replace-cross Abstract: Video large multimodal models increasingly face a scalability bottleneck: long videos produce excessively long visual-token sequences, which s

VIMCAN: Visual-Inertial 3D Human Pose Estimation with Hybrid Mamba-Cross-Attention Network

ResearchDGX agent

arXiv:2605.07552v1 Announce Type: new Abstract: The rapid advances in deep learning have significantly enhanced the accuracy of multimodal 3D human pose estimation (HPE). However, the state-of-the-art

VISD: Enhancing Video Reasoning via Structured Self-Distillation

SafetyDGX agent

arXiv:2605.06094v2 Announce Type: replace-cross Abstract: Training VideoLLMs for complex reasoning remains challenging due to sparse sequence level rewards and the lack of fine grained credit assignme

Visual Text Compression as Measure Transport

Model ReleasesDGX agent

arXiv:2605.06708v1 Announce Type: cross Abstract: Visual text compression (VTC) promises efficient long-context processing by rendering text into an image and re-encoding it with a vision-language mod

VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing

ResearchDGX agent

arXiv:2605.06765v1 Announce Type: cross Abstract: Human speech conveys expressiveness beyond linguistic content, including personality, mood, or performance elements, such as a comforting tone or humm

VLA-GSE: Boosting Parameter-Efficient Fine-Tuning in VLA with Generalized and Specialized Experts

Model ReleasesDGX agent

arXiv:2605.06175v2 Announce Type: replace Abstract: Vision-language-action (VLA) models inherit rich visual-semantic priors from pre-trained vision-language backbones, but adapting them to robotic con

VNN-LIB 2.0: Rigorous Foundations for Neural Network Verification

ResearchDGX agent

arXiv:2605.07451v1 Announce Type: new Abstract: Neural network verification is an active and rapidly maturing research area, with a growing ecosystem of solvers and tools. The VNN-LIB standard was int

Weather-Robust Scene Semantics with Vision-Aligned 4D Radar

ResearchDGX agent

arXiv:2605.07367v1 Announce Type: cross Abstract: Cameras and LiDAR degrade in rain, fog, and snow, while millimeter-wave radar remains largely unaffected. We align a radar encoder to frozen SigLIP vi

WeatherSyn: An Instruction Tuning MLLM For Weather Forecasting Report Generation

ResearchDGX agent

arXiv:2605.07522v1 Announce Type: new Abstract: Accurate weather forecast reporting enables individuals and communities to better plan daily activities and agricultural operations. However, the curren

WebClipper: Efficient Evolution of Web Agents with Graph-based Trajectory Pruning

AgentsDGX agent

arXiv:2602.12852v2 Announce Type: replace Abstract: Deep Research systems based on web agents have shown strong potential in solving complex information-seeking tasks, yet their search efficiency rema

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents

ApplicationsDGX agent

arXiv:2605.06761v1 Announce Type: new Abstract: The web is complex, open-ended, and constantly changing, making it challenging to scale training data for visual web agents. Existing data collection at

What if AI systems weren't chatbots?

ApplicationsDGX agent

arXiv:2605.07896v1 Announce Type: cross Abstract: The rapid convergence of artificial intelligence (AI) toward conversational chatbot interfaces marks a critical moment for the industry. This paper ar

What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion

ResearchDGX agent

arXiv:2605.07915v1 Announce Type: new Abstract: Tokenizers are a crucial component of latent diffusion models, as they define the latent space in which diffusion models operate. However, existing toke

When Are Experts Misrouted? Counterfactual Routing Analysis in Mixture-of-Experts Language Models

Model ReleasesDGX agent

arXiv:2605.07260v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) language models route each token to a small subset of experts, but whether the routes selected by a trained top-k router are

When Descent Is Too Stable: Event-Triggered Hamiltonian Learning to Optimize

SafetyDGX agent

arXiv:2605.06868v1 Announce Type: new Abstract: Fixed-budget nonconvex optimization can fail not because local descent is unstable, but because it is too stable: after reaching a nearby stationary poi

When Diffusion Model Can Ignore Dimension: An Entropy-Based Theory

ResearchDGX agent

arXiv:2605.07969v1 Announce Type: new Abstract: Diffusion models perform remarkably well on high-dimensional data such as images, often using only a modest number of reverse-time steps. Despite this p

When Does a Language Model Commit? A Finite-Answer Theory of Pre-Verbalization Commitment

ResearchDGX agent

arXiv:2605.06723v1 Announce Type: new Abstract: Language models often generate reasoning before giving a final answer, but the visible answer does not reveal when the model's answer preference became

When Does Critique Improve AI-Assisted Theoretical Physics? SCALAR: Structured Critic--Actor Loop for Agentic Reasoning

Model ReleasesDGX agent

arXiv:2605.06772v1 Announce Type: new Abstract: As large language models (LLMs) show increasing promise on research-level physics reasoning tasks and agentic AI becomes more common, a practical questi

When Does Embedding Magnitude Matter? A Cross-Task Functional-Symmetry Framework

ResearchDGX agent

arXiv:2602.09229v3 Announce Type: replace Abstract: Cosine similarity normalizes both sides; dot product normalizes neither. We propose a 2x2 framework that independently controls query-side and docum

When Losses Align: Gradient-Based Composite Loss Weighting for Efficient Pretraining

ResearchDGX agent

arXiv:2605.07756v1 Announce Type: cross Abstract: Modern deep models are often pretrained on large-scale data with missing labels using composite objectives, where the relative weights of multiple los

When Routine Chats Turn Toxic: Unintended Long-Term State Poisoning in Personalized Agents

Model ReleasesDGX agent

arXiv:2605.06731v1 Announce Type: cross Abstract: Personalized LLM agents maintain persistent cross-session state to support long-horizon collaboration. Yet, this persistence introduces a subtle but c

When Stored Evidence Stops Being Usable: Scale-Conditioned Evaluation of Agent Memory

AgentsDGX agent

arXiv:2605.07313v1 Announce Type: new Abstract: Memory-agent evaluations report fixed-snapshot accuracy or retrieval quality, but these scores do not show whether evidence remains usable as irrelevant

When Symbol Names Should Not Matter: A Logistic Theory of Fresh-Symbol Classification

TutorialsDGX agent

arXiv:2605.07120v1 Announce Type: new Abstract: Template tasks have emerged as a clean testbed for asking whether transformers reason with abstract symbols rather than concrete token names. We study t

Where to Spend Rollouts: Hit-Utility Optimal Rollout Allocation for Group-Based RLVR

Model ReleasesDGX agent

arXiv:2605.07114v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a central paradigm for improving the reasoning capabilities of large language model

Where's the Plan? Locating Latent Planning in Language Models with Lightweight Mechanistic Interventions

Model ReleasesDGX agent

arXiv:2605.07984v1 Announce Type: cross Abstract: We study planning site formation in language models -- where internal representations of structurally-constrained future tokens form during the forwar

Who Prices Cognitive Labor in the Age of Agents? Compute-Anchored Wages

SafetyDGX agent

arXiv:2605.05558v2 Announce Type: replace Abstract: A natural intuition about the economics of AI agents is that, because agents can be replicated at very low marginal cost, agent labor may be supplie

← Previous
1…754755756757758…1005
Next →