AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
Human
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
21 Apr 2026

Triples and Knowledge-Infused Embeddings for Clustering and Classification of Scientific Documents

Model ReleasesDGX agent

arXiv:2601.08841v2 Announce Type: replace Abstract: The increasing volume and complexity of scientific literature demand robust methods for organizing and understanding research documents. In this stu

TriTS: Time Series Forecasting from a Multimodal Perspective

Model ReleasesDGX agent

arXiv:2604.16748v1 Announce Type: new Abstract: Time series forecasting plays a pivotal role in critical sectors such as finance, energy, transportation, and meteorology. However, Long-term Time Serie

Trustworthy deep domain adaptation for wearable photoplethysmography signal analysis with decision-theoretic uncertainty quantification

ApplicationsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.17480v1 Announce Type: new Abstract: In principle, deep generative models can be used to perform domain adaptation; i.e. align the input feature representations of test data with that of a

Trustworthy Endoscopic Super-Resolution

Local AiDGX agent

arXiv:2604.18001v1 Announce Type: new Abstract: Super-resolution (SR) models are attracting growing interest for enhancing minimally invasive surgery and diagnostic videos under hardware constraints.

TSegAgent: Zero-Shot Tooth Segmentation via Geometry-Aware Vision-Language Agents

ResearchDGX agent

arXiv:2603.19684v2 Announce Type: replace Abstract: Automatic tooth segmentation and identification from intra-oral scanned 3D models are fundamental problems in digital dentistry, yet most existing a

TSM-Pose: Topology-Aware Learning with Semantic Mamba for Category-Level Object Pose Estimation

Model ReleasesDGX agent

arXiv:2604.16954v1 Announce Type: new Abstract: Category-level object pose estimation is fundamental for embodied intelligence, yet achieving robust generalization to unseen instances remains challeng

TSVer: A Benchmark for Fact Verification Against Time-Series Evidence

Model ReleasesDGX agent

arXiv:2511.01101v2 Announce Type: replace Abstract: Reasoning over temporal and numerical data, such as time series, is a crucial aspect of fact-checking. While many systems have recently been develop

TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts

Local AiDGX agent

arXiv:2604.16542v1 Announce Type: cross Abstract: Safety guardrails have become an active area of research in AI safety, aimed at ensuring the appropriate behavior of large language models (LLMs). How

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models

Model ReleasesDGX agent

arXiv:2604.18518v1 Announce Type: new Abstract: Uniform Discrete Diffusion Model (UDM) has recently emerged as a promising paradigm for discrete generative modeling; however, its integration with rein

UGD: An Unsupervised Geometric Distance for Evaluating Real-world Noisy Point Cloud Denoising

TutorialsDGX agent

arXiv:2604.16976v1 Announce Type: new Abstract: Point cloud denoising is a fundamental and crucial challenge in real-world point cloud applications. Existing quantitative evaluation metrics for point

Uncertainty Quantification in PINNs for Turbulent Flows: Bayesian Inference and Repulsive Ensembles

ResearchDGX agent

arXiv:2604.17156v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have emerged as a promising framework for solving inverse problems governed by partial differential equations (

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment

SafetyDGX agent

arXiv:2405.13068v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have revolutionized various applications, making robust safety alignment essential to prevent harmful outputs. Cu

Understanding Counting Mechanisms in Large Language and Vision-Language Models

ResearchDGX agent

arXiv:2511.17699v2 Announce Type: replace Abstract: Counting is one of the fundamental abilities of large language models (LLMs) and large vision-language models (LVLMs). This paper examines how these

Understanding the Prompt Sensitivity

ResearchDGX agent

arXiv:2604.18389v1 Announce Type: new Abstract: Prompt sensitivity, which refers to how strongly the output of a large language model (LLM) depends on the exact wording of its input prompt, raises con

Understanding Tool-Augmented Agents for Lean Formalization: A Factorial Analysis

Model ReleasesDGX agent

arXiv:2604.16538v1 Announce Type: cross Abstract: Automatic translation of natural language mathematics into faithful Lean 4 code is hindered by the fundamental dissonance between informal set-theoret

Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark

Model ReleasesDGX agent

arXiv:2510.13759v3 Announce Type: replace Abstract: Unified multimodal models aim to jointly enable visual understanding and generation, yet current benchmarks rarely examine their true integration. E

UniComp: A Unified Evaluation of Large Language Model Compression via Pruning, Quantization and Distillation

SafetyDGX agent

arXiv:2602.09130v3 Announce Type: replace Abstract: Model compression is increasingly essential for deploying large language models (LLMs), yet existing comparative studies largely focus on pruning an

UniCon: Unified Framework for Efficient Contrastive Alignment via Kernels

SafetyDGX agent

arXiv:2604.16678v1 Announce Type: new Abstract: Contrastive objectives power state-of-the-art multimodal models, but their training remains slow, relying on long stochastic optimization. We propose a

UniCSG: Unified High-Fidelity Content-Constrained Style-Driven Generation via Staged Semantic and Frequency Disentanglement

SafetyDGX agent

arXiv:2604.17850v1 Announce Type: new Abstract: Style transfer must match a target style while preserving content semantics. DiT-based diffusion models often suffer from content-style entanglement, le

UniDomain: Pretraining a Unified PDDL Domain from Real-World Demonstrations for Generalizable Robot Task Planning

ApplicationsDGX agent

arXiv:2507.21545v3 Announce Type: replace Abstract: Robotic task planning in real-world environments requires reasoning over implicit constraints from language and vision. While LLMs and VLMs offer st

Unified Multimodal Brain Decoding via Cross-Subject Soft-ROI Fusion

Model ReleasesDGX agent

arXiv:2512.20249v3 Announce Type: replace-cross Abstract: Multimodal brain decoding aims to reconstruct semantic information that is consistent with visual stimuli from brain activity signals such as

Unified Ultrasound Intelligence Toward an End-to-End Agentic System

AgentsDGX agent

arXiv:2604.16914v1 Announce Type: new Abstract: Clinical ultrasound analysis demands models that generalize across heterogeneous organs, views, and devices, while supporting interpretable workflow-lev

UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models

ResearchDGX agent

arXiv:2604.17565v1 Announce Type: new Abstract: Camera-controllable image editing aims to synthesize novel views of a given scene under varying camera poses while strictly preserving cross-view geomet

UniMamba: A Unified Spatial-Temporal Modeling Framework with State-Space and Attention Integration

Model ReleasesDGX agent

arXiv:2604.16325v1 Announce Type: new Abstract: Multivariate time series forecasting is fundamental to numerous domains such as energy, finance, and environmental monitoring, where complex temporal de

UniMesh: Unifying 3D Mesh Understanding and Generation

ResearchDGX agent

arXiv:2604.17472v1 Announce Type: new Abstract: Recent advances in 3D vision have led to specialized models for either 3D understanding (e.g., shape classification, segmentation, reconstruction) or 3D

UniSim: A Unified Simulator for Time-Coarsened Dynamics of Biomolecules

TutorialsDGX agent

arXiv:2506.03157v4 Announce Type: replace-cross Abstract: Molecular Dynamics (MD) simulations are essential for understanding the atomic-level behavior of molecular systems, giving insights into their

Universal Diffusion-Based Probabilistic Downscaling

ResearchDGX agent

arXiv:2602.11893v3 Announce Type: replace Abstract: We introduce a universal diffusion-based downscaling framework that lifts deterministic low-resolution weather forecasts into probabilistic high-res

Universally Empowering Zeroth-Order Optimization via Adaptive Layer-wise Sampling

Model ReleasesDGX agent

arXiv:2604.18264v1 Announce Type: new Abstract: Zeroth-Order optimization presents a promising memory-efficient paradigm for fine-tuning Large Language Models by relying solely on forward passes. Howe

Unleashing Spatial Reasoning in Multimodal Large Language Models via Textual Representation Guided Reasoning

Model ReleasesDGX agent

arXiv:2603.23404v2 Announce Type: replace-cross Abstract: Existing Multimodal Large Language Models (MLLMs) struggle with 3D spatial reasoning, as they fail to construct structured abstractions of the

Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2604.18000v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) models report impressive success rates on standard robotic benchmarks, fueling optimism about general-purpose physic

Unraveling the Key of Machine Learning-based Android Malware Detection

TutorialsDGX agent

arXiv:2402.02953v2 Announce Type: replace-cross Abstract: With the rapid advancement of machine learning (ML), ML-based Android malware detection has gained significant popularity due to its ability t

Unsupervised Discovery of Intermediate Phase Order in the Frustrated J_1-J_2 Heisenberg Model via Prometheus Framework

Model ReleasesDGX agent

arXiv:2602.21468v4 Announce Type: replace-cross Abstract: The spin-1/2 J_1-J_2 Heisenberg model on the square lattice exhibits a debated intermediate phase between Neel antiferromagnetic and stripe or

Untrained CNNs Match Backpropagation at V1: A Systematic RSA Comparison of Four Learning Rules Against Human fMRI

SafetyDGX agent

arXiv:2604.16875v1 Announce Type: new Abstract: A central question in computational neuroscience is whether the learning rule used to train a neural network determines how well its internal representa

Unveiling Deepfakes: A Frequency-Aware Triple Branch Network for Deepfake Detection

Model ReleasesDGX agent

arXiv:2604.17477v1 Announce Type: new Abstract: Advanced deepfake technologies are blurring the lines between real and fake, presenting both revolutionary opportunities and alarming threats. While it

Upper Approximation Bounds for Neural Oscillators

ResearchDGX agent

arXiv:2512.01015v2 Announce Type: replace Abstract: Neural oscillators, originating from second-order ordinary differential equations (ODEs), have demonstrated strong performance in stably learning ca

User-Assistant Bias in LLMs

Model ReleasesDGX agent

arXiv:2508.15815v3 Announce Type: replace Abstract: Modern large language models (LLMs) are typically trained and deployed using structured role tags (e.g. system, user, assistant, tool) that explicit

Using large language models for embodied planning introduces systematic safety risks

Model ReleasesDGX agent

arXiv:2604.18463v1 Announce Type: cross Abstract: Large language models are increasingly used as planners for robotic systems, yet how safely they plan remains an open question. To evaluate safe plann

Using Perspectival Words Is Harder Than Vocabulary Words for Humans and Even More So for Multimodal Language Models

ResearchDGX agent

arXiv:2506.00065v2 Announce Type: replace Abstract: Multimodal language models (MLMs) increasingly demonstrate human-like communication, yet their use of everyday perspectival words remains poorly und

VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning

Model ReleasesDGX agent

arXiv:2402.13243v2 Announce Type: replace Abstract: Learning a human-like driving policy from large-scale driving demonstrations is promising, but the uncertainty and non-deterministic nature of plann

Variational Autoencoder Domain Adaptation for Cross-System Generalization in ML-Based SOP Monitoring

ResearchDGX agent

arXiv:2604.18035v1 Announce Type: new Abstract: Machine learning (ML) models trained to detect physical-layer threats on one optical fiber system often fail catastrophically when applied to a differen

VC-Inspector: Advancing Reference-free Evaluation of Video Captions with Factual Analysis

ResearchDGX agent

arXiv:2509.16538v3 Announce Type: replace-cross Abstract: We propose VC-Inspector, a lightweight, open-source large multimodal model (LMM) for reference-free evaluation of video captions, with a focus

VCORE: Variance-Controlled Optimization-based Reweighting for Chain-of-Thought Supervision

Model ReleasesDGX agent

arXiv:2510.27462v2 Announce Type: replace Abstract: Supervised fine-tuning (SFT) on long chain-of-thought (CoT) trajectories has emerged as a crucial technique for enhancing the reasoning abilities of

VIBE: Voice-Induced open-ended Bias Evaluation for Large Audio-Language Models via Real-World Speech

SafetyDGX agent

arXiv:2604.17248v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) are increasingly integrated into daily applications, yet their generative biases remain underexplored. Existing sp

Video Panels for Long Video Understanding

Model ReleasesDGX agent

arXiv:2509.23724v2 Announce Type: replace Abstract: Recent Video-Language Models (VLMs) achieve promising results on long-video understanding, but their performance still lags behind that achieved on

Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation

Local AiDGX agent

arXiv:2604.17656v1 Announce Type: cross Abstract: Video-to-music (V2M) is the fundamental task of creating background music for an input video. Recent V2M models achieve audiovisual alignment by typic

VIDEOP2R: Video Understanding from Perception to Reasoning

SafetyDGX agent

arXiv:2511.11113v2 Announce Type: replace Abstract: Reinforcement fine-tuning (RFT), a two-stage framework consisting of supervised fine-tuning (SFT) and reinforcement learning (RL) has shown promisin

VideoThinker: Building Agentic VideoLLMs with LLM-Guided Tool Reasoning

Local AiDGX agent

arXiv:2601.15724v2 Announce Type: replace Abstract: Long-form video understanding remains a fundamental challenge for current Video Large Language Models. Most existing models rely on static reasoning

VIDS: A Verified Imaging Dataset Standard for Medical AI

Model ReleasesDGX agent

arXiv:2604.17525v1 Announce Type: cross Abstract: Medical imaging AI development is fundamentally dependent on annotated datasets, yet no existing standard provides machine-enforceable validation acro

View-Consistent 3D Scene Editing via Dual-Path Structural Correspondense and Semantic Continuity

ResearchDGX agent

arXiv:2604.17801v1 Announce Type: new Abstract: Text-driven 3D scene editing has recently attracted increasing attention. Most existing methods follow a render-edit-optimize pipeline, where multi-view

ViPS: Video-informed Pose Spaces for Auto-Rigged Meshes

ResearchDGX agent

arXiv:2604.17623v1 Announce Type: new Abstract: Kinematic rigs provide a structured interface for articulating 3D meshes, but they lack an inherent representation of the plausible manifold of joint co

Vision-Braille: A Curriculum Learning Toolkit and Braille-Chinese Corpus for Braille Translation

ApplicationsDGX agent

arXiv:2407.06048v2 Announce Type: replace Abstract: We present Vision-Braille, the first publicly available end-to-end system for translating Chinese Braille extracted from images into written Chinese

Vision Language Models are Biased

ResearchDGX agent

arXiv:2505.23941v4 Announce Type: replace-cross Abstract: Large language models (LLMs) memorize a vast amount of prior knowledge from the Internet that helps them on downstream tasks but also may noto

Visual-RRT: Finding Paths toward Visual-Goals via Differentiable Rendering

ApplicationsDGX agent

arXiv:2604.16388v1 Announce Type: cross Abstract: Rapidly-exploring random trees (RRTs) have been widely adopted for robot motion planning due to their robustness and theoretical guarantees. However,

ViT^3: Unlocking Test-Time Training in Vision

ResearchDGX agent

arXiv:2512.01643v2 Announce Type: replace Abstract: Test-Time Training (TTT) has recently emerged as a promising direction for efficient sequence modeling. TTT reformulates attention operation as an o

VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction

Model ReleasesDGX agent

arXiv:2505.20279v4 Announce Type: replace-cross Abstract: The rapid advancement of Large Multimodal Models (LMMs) for 2D images and videos has motivated extending these models to understand 3D scenes,

Vocab Diet: Reshaping the Vocabulary of LLMs via Vector Arithmetic

ResearchDGX agent

arXiv:2510.17001v2 Announce Type: replace Abstract: Large language models (LLMs) often encode word-form variation (e.g., walk vs. walked) as linear directions in the embedding space. However, standard

VocabTailor: Dynamic Vocabulary Selection for Downstream Tasks in Small Language Models

Local AiDGX agent

arXiv:2508.15229v3 Announce Type: replace Abstract: Small Language Models (SLMs) provide computational advantages in resource-constrained environments, yet memory limitations remain a critical bottlen

Voronoi-guided Bilateral 2D Gaussian Splatting for Arbitrary-Scale Hyperspectral Image Super-Resolution

Model ReleasesDGX agent

arXiv:2604.17727v1 Announce Type: new Abstract: Most existing hyperspectral image super-resolution methods require modifications for different scales, limiting their flexibility in arbitrary-scale rec

Waking Up Blind: Cold-Start Optimization of Supervision-Free Agentic Trajectories for Grounded Visual Perception

SafetyDGX agent

arXiv:2604.17475v1 Announce Type: cross Abstract: Small Vision-Language Models (SVLMs) are efficient task controllers but often suffer from visual brittleness and poor tool orchestration. They typical

Wasserstein Distributionally Robust Risk-Sensitive Estimation via Conditional Value-at-Risk

ResearchDGX agent

arXiv:2604.18546v1 Announce Type: new Abstract: We propose a distributionally robust approach to risk-sensitive estimation of an unknown signal x from an observed signal y. The unknown signal and obse

← Previous
1…902903904905906…998
Next →