AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
Human
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
16 Apr 2026

TIP: Token Importance in On-Policy Distillation

Model ReleasesDGX agent

arXiv:2604.14084v1 Announce Type: new Abstract: On-policy knowledge distillation (OPD) trains a student on its own rollouts under token-level supervision from a teacher. Not all token positions matter

TLoRA+: A Low-Rank Parameter-Efficient Fine-Tuning Method for Large Language Models

Model ReleasesDGX agent

arXiv:2604.13368v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) aims to adapt pre-trained models to specific tasks using relatively small and domain-specific datasets. Among P

Tokenizing Semantic Segmentation with Run Length Encoding

TutorialsDGX agent

arXiv:2602.21627v3 Announce Type: replace Abstract: This paper presents a new unified approach to semantic segmentation in both images and videos by using language modeling to output the masks as sequ

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution

Model ReleasesDGX agent

arXiv:2604.13787v1 Announce Type: new Abstract: Large Language Models (LLMs) enhance their problem-solving capability by utilizing external tools. However, in open-world scenarios with massive and evo

ToolSpec: Accelerating Tool Calling via Schema-Aware and Retrieval-Augmented Speculative Decoding

AgentsDGX agent

arXiv:2604.13519v1 Announce Type: new Abstract: Tool calling has greatly expanded the practical utility of large language models (LLMs) by enabling them to interact with external applications. As LLM

Towards Generalizable Robotic Manipulation in Dynamic Environments

Model ReleasesDGX agent

arXiv:2603.15620v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models excel in static manipulation but struggle in dynamic environments with moving targets. This performance gap prim

Towards Multi-Object-Tracking with Radar on a Fast Moving Vehicle: On the Potential of Processing Radar in the Frequency Domain

AgentsDGX agent

arXiv:2604.14013v1 Announce Type: cross Abstract: We promote in this paper the processing of radar data in the frequency domain to achieve higher robustness against noise and structural errors, especi

Towards Patient-Specific Deformable Registration in Laparoscopic Surgery

ResearchDGX agent

arXiv:2604.13186v1 Announce Type: new Abstract: Unsafe surgical care is a critical health concern, often linked to limitations in surgeon experience, skills, and situational awareness. Integrating pat

Towards Successful Implementation of Automated Raveling Detection: Effects of Training Data Size, Illumination Difference, and Spatial Shift

Model ReleasesDGX agent

arXiv:2604.13322v1 Announce Type: new Abstract: Raveling, the loss of aggregates, is a major form of asphalt pavement surface distress, especially on highways. While research has shown that machine le

Towards Unconstrained Human-Object Interaction

ResearchDGX agent

arXiv:2604.14069v1 Announce Type: new Abstract: Human-Object Interaction (HOI) detection is a longstanding computer vision problem concerned with predicting the interaction between humans and objects.

Training-Free Semantic Multi-Object Tracking with Vision-Language Models

ResearchDGX agent

arXiv:2604.14074v1 Announce Type: new Abstract: Semantic Multi-Object Tracking (SMOT) extends multi-object tracking with semantic outputs such as video summaries, instance-level captions, and interact

Training-Free Test-Time Contrastive Learning for Large Language Models

AgentsDGX agent

arXiv:2604.13552v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong reasoning capabilities, but their performance often degrades under distribution shift. Existing test-tim

Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability

Model ReleasesDGX agent

arXiv:2604.05478v2 Announce Type: replace-cross Abstract: Immune checkpoint inhibitors (ICIs) have transformed cancer therapy; yet substantial proportion of patients exhibit intrinsic or acquired resi

TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration

Model ReleasesDGX agent

arXiv:2604.14116v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have empowered AI research agents to perform isolated scientific tasks, automating complex, real-world workflows, s

TRIM: Hybrid Inference via Targeted Stepwise Routing in Multi-Step Reasoning Tasks

SafetyDGX agent

arXiv:2601.10245v2 Announce Type: replace-cross Abstract: Multi-step reasoning tasks like mathematical problem solving are vulnerable to cascading failures, where a single incorrect step leads to comp

Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations

ResearchDGX agent

arXiv:2601.07422v2 Announce Type: replace Abstract: Despite their impressive capabilities, large language models (LLMs) frequently generate hallucinations. Previous work shows that their internal stat

Two-Stage Regularization-Based Structured Pruning for LLMs

Model ReleasesDGX agent

arXiv:2505.18232v3 Announce Type: replace-cross Abstract: The deployment of large language models (LLMs) is largely hindered by their large number of parameters. Structural pruning has emerged as a pr

UHR-BAT: Budget-Aware Token Compression Vision-Language model for Ultra-High-Resolution Remote Sensing

ResearchDGX agent

arXiv:2604.13565v1 Announce Type: new Abstract: Ultra-high-resolution (UHR) remote sensing imagery couples kilometer-scale context with query-critical evidence that may occupy only a few pixels. Such

UI-Copilot: Advancing Long-Horizon GUI Automation via Tool-Integrated Policy Optimization

Model ReleasesDGX agent

arXiv:2604.13822v1 Announce Type: new Abstract: MLLM-based GUI agents have demonstrated strong capabilities in complex user interface interaction tasks. However, long-horizon scenarios remain challeng

UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Grounding

Local AiDGX agent

arXiv:2604.14113v1 Announce Type: cross Abstract: GUI grounding, which localizes interface elements from screenshots given natural language queries, remains challenging for small icons and dense layou

UMI-3D: Extending Universal Manipulation Interface from Vision-Limited to 3D Spatial Perception

SafetyDGX agent

arXiv:2604.14089v1 Announce Type: new Abstract: We present UMI-3D, a multimodal extension of the Universal Manipulation Interface (UMI) for robust and scalable data collection in embodied manipulation

UNBOX: Unveiling Black-box visual models with Natural-language

SafetyDGX agent

arXiv:2603.08639v2 Announce Type: replace Abstract: Ensuring trustworthiness in open-world visual recognition requires models that are interpretable, fair, and robust to distribution shifts. Yet moder

UniBlendNet: Unified Global, Multi-Scale, and Region-Adaptive Modeling for Ambient Lighting Normalization

Model ReleasesDGX agent

arXiv:2604.13383v1 Announce Type: new Abstract: Ambient Lighting Normalization (ALN) aims to restore images degraded by complex, spatially varying illumination conditions. Existing methods, such as IF

UniGeoSeg: Towards Unified Open-World Segmentation for Geospatial Scenes

Model ReleasesDGX agent

arXiv:2511.23332v2 Announce Type: replace Abstract: Instruction-driven segmentation in remote sensing generates masks from guidance, offering great potential for accessible and generalizable applicati

Universality of Gaussian-Mixture Reverse Kernels in Conditional Diffusion

ResearchDGX agent

arXiv:2604.13470v1 Announce Type: new Abstract: We prove that conditional diffusion models whose reverse kernels are finite Gaussian mixtures with ReLU-network logits can approximate suitably regular

Unleashing Implicit Rewards: Prefix-Value Learning for Distribution-Level Optimization

Local AiDGX agent

arXiv:2604.13197v1 Announce Type: new Abstract: Process reward models (PRMs) provide fine-grained reward signals along the reasoning process, but training reliable PRMs often requires step annotations

UNRIO: Uncertainty-Aware Velocity Learning for Radar-Inertial Odometry

Model ReleasesDGX agent

arXiv:2604.13584v1 Announce Type: new Abstract: We present UNRIO, an uncertainty-aware radar-inertial odometry system that estimates ego-velocity directly from raw mmWave radar IQ signals rather than

Unsupervised Anomaly Detection in Process-Complex Industrial Time Series: A Real-World Case Study

Model ReleasesDGX agent

arXiv:2604.13928v1 Announce Type: new Abstract: Industrial time-series data from real production environments exhibits substantially higher complexity than commonly used benchmark datasets, primarily

Unsupervised domain transfer: Overcoming signal degradation in sleep monitoring by increasing scoring realism

TutorialsDGX agent

arXiv:2604.13988v1 Announce Type: new Abstract: Objective: Investigate whether hypnogram 'realism' can be used to guide an unsupervised method for handling arbitrary types of signal degradation in mob

Using reasoning LLMs to extract SDOH events from clinical notes

ResearchDGX agent

arXiv:2604.13502v1 Announce Type: new Abstract: Social Determinants of Health (SDOH) refer to environmental, behavioral, and social conditions that influence how individuals live, work, and age. SDOH

Utilizing Inpainting for Keypoint Detection for Vision-Based Control of Robotic Manipulators

ResearchDGX agent

arXiv:2604.13309v1 Announce Type: new Abstract: In this paper we present a novel visual servoing framework to control a robotic manipulator in the configuration space by using purely natural visual fe

ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs

Model ReleasesDGX agent

arXiv:2604.06484v2 Announce Type: replace Abstract: Cultural values are expressed not only through language but also through visual scenes and everyday social practices. Yet existing evaluations of cu

Vectorizing Projection in Manifold-Constrained Motion Planning for Real-Time Whole-Body Control

ApplicationsDGX agent

arXiv:2604.13323v1 Announce Type: new Abstract: Many robot planning tasks require satisfaction of one or more constraints throughout the entire trajectory. For geometric constraints, manifold-constrai

VGGT-Segmentor: Geometry-Enhanced Cross-View Segmentation

Model ReleasesDGX agent

arXiv:2604.13596v1 Announce Type: new Abstract: Instance-level object segmentation across disparate egocentric and exocentric views is a fundamental challenge in visual understanding, critical for app

VibeFlow: Versatile Video Chroma-Lux Editing through Self-Supervised Learning

ResearchDGX agent

arXiv:2604.13425v1 Announce Type: new Abstract: Video chroma-lux editing, which aims to modify illumination and color while preserving structural and temporal fidelity, remains a significant challenge

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body

Model ReleasesDGX agent

arXiv:2512.14234v2 Announce Type: replace Abstract: Human communication is inherently multimodal and social: words, prosody, and body language jointly carry intent. Yet most prior systems model human

VIGILant: an automatic classification pipeline for glitches in the Virgo detector

ResearchDGX agent

arXiv:2604.13687v1 Announce Type: cross Abstract: Glitches frequently contaminate data in gravitational-wave detectors, complicating the observation and analysis of astrophysical signals. This work in

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap

AgentsDGX agent

arXiv:2604.13654v1 Announce Type: new Abstract: Vision-and-Language Navigation for Unmanned Aerial Vehicles (UAV-VLN) represents a pivotal challenge in embodied artificial intelligence, focused on ena

Visual Self-Fulfilling Alignment: Shaping Safety-Oriented Personas via Threat-Related Images

SafetyDGX agent

arXiv:2603.08486v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) face safety misalignment, where visual inputs enable harmful outputs. To address this, existing methods req

Visual Sparse Steering (VS2): Unsupervised Adaptation for Image Classification using Sparsity-Guided Steering Vectors

ResearchDGX agent

arXiv:2506.01247v2 Announce Type: replace Abstract: Steering vision foundation models at test time, without updating foundation-model weights or using labeled target data, is a desirable yet challengi

VLMs Need Words: Vision Language Models Ignore Visual Detail In Favor of Semantic Anchors

ResearchDGX agent

arXiv:2604.02486v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have achieved impressive performance across a wide range of multimodal tasks. However, they often fail on tasks

VRAG-DFD: Verifiable Retrieval-Augmentation for MLLM-based Deepfake Detection

SafetyDGX agent

arXiv:2604.13660v1 Announce Type: new Abstract: In Deepfake Detection (DFD) tasks, researchers proposed two types of MLLM-based methods: complementary combination with small DFD detectors, or static f

Weakly-supervised Learning for Physics-informed Neural Motion Planning via Sparse Roadmap

Local AiDGX agent

arXiv:2604.13204v1 Announce Type: new Abstract: The motion planning problem requires finding a collision-free path between start and goal configurations in high-dimensional, cluttered spaces. Recent l

WebXSkill: Skill Learning for Autonomous Web Agents

AgentsDGX agent

arXiv:2604.13318v1 Announce Type: cross Abstract: Autonomous web agents powered by large language models (LLMs) have shown promise in completing complex browser tasks, yet they still struggle with lon

What Are We Really Measuring? Rethinking Dataset Bias in Web-Scale Natural Image Collections via Unsupervised Semantic Clustering

SafetyDGX agent

arXiv:2604.13610v1 Announce Type: new Abstract: In computer vision, a prevailing method for quantifying dataset bias is to train a model to distinguish between datasets. High classification accuracy i

When Less Latent Leads to Better Relay: Information-Preserving Compression for Latent Multi-Agent LLM Collaboration

AgentsDGX agent

arXiv:2604.13349v1 Announce Type: new Abstract: Communication in Large Language Model (LLM)-based multi-agent systems is moving beyond discrete tokens to preserve richer context. Recent work such as L

When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?

Model ReleasesDGX agent

arXiv:2503.23137v2 Announce Type: replace-cross Abstract: Understanding humor-particularly when it involves complex, contradictory narratives that require comparative reasoning-remains a significant c

Who Gets Flagged? The Pluralistic Evaluation Gap in AI Content Watermarking

SafetyDGX agent

arXiv:2604.13776v1 Announce Type: cross Abstract: Watermarking is becoming the default mechanism for AI content authentication, with governance policies and frameworks referencing it as infrastructure

Why MLLMs Struggle to Determine Object Orientations

SafetyDGX agent

arXiv:2604.13321v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) struggle with tasks that require reasoning about 2D object orientation in images, as documented in prior work.

Why Multimodal In-Context Learning Lags Behind? Unveiling the Inner Mechanisms and Bottlenecks

SafetyDGX agent

arXiv:2604.13403v1 Announce Type: new Abstract: In-context learning (ICL) enables models to adapt to new tasks via inference-time demonstrations. Despite its success in large language models, the exte

WIN-U: Woodbury-Informed Newton-Unlearning as a retain-free Machine Unlearning Framework

ResearchDGX agent

arXiv:2604.13438v1 Announce Type: new Abstract: Privacy concerns in LLMs have led to the rapidly growing need to enforce a data's 'right to be forgotten'. Machine unlearning addresses precisely this t

Wireless bioelectronic control architectures for biohybrid robotic systems

AgentsDGX agent

arXiv:2603.24959v2 Announce Type: replace Abstract: Wireless bioelectronic interfaces are increasingly used to control tissue-engineered biohybrid robotic systems. However, a unifying engineering fram

Working Notes on Late Interaction Dynamics: Analyzing Targeted Behaviors of Late Interaction Models

Model ReleasesDGX agent

arXiv:2603.26259v2 Announce Type: replace-cross Abstract: While Late Interaction models exhibit strong retrieval performance, many of their underlying dynamics remain understudied, potentially hiding

WorkRB: A Community-Driven Evaluation Framework for AI in the Work Domain

Model ReleasesDGX agent

arXiv:2604.13055v1 Announce Type: new Abstract: Today's evolving labor markets rely increasingly on recommender systems for hiring, talent management, and workforce analytics, with natural language pr

X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations

TutorialsDGX agent

arXiv:2511.04671v2 Announce Type: replace-cross Abstract: Human videos are a scalable source of training data for robot learning. However, humans and robots significantly differ in embodiment, making

YOCO++: Enhancing YOCO with KV Residual Connections for Efficient LLM Inference

ResearchDGX agent

arXiv:2604.13556v1 Announce Type: new Abstract: Cross-layer key-value (KV) compression has been found to be effective in efficient inference of large language models (LLMs). Although they reduce the m

Zero-Shot Function Encoder-Based Differentiable Predictive Control

ResearchDGX agent

arXiv:2511.05757v3 Announce Type: replace-cross Abstract: We introduce a differentiable framework for zero-shot adaptive control over parametric families of nonlinear dynamical systems. Our approach i

ZK-APEX: Zero-Knowledge Approximate Personalized Unlearning with Executable Proofs

SafetyDGX agent

arXiv:2512.09953v2 Announce Type: replace-cross Abstract: Machine unlearning aims to remove the influence of specific data points from a trained model to satisfy privacy, copyright, and safety require

ZoomSpec: A Physics-Guided Coarse-to-Fine Framework for Wideband Spectrum Sensing

ApplicationsDGX agent

arXiv:2604.13568v1 Announce Type: new Abstract: Wideband spectrum sensing for low-altitude monitoring is critical yet challenging due to heterogeneous protocols,large bandwidths, and non-stationary SN

15 Apr 2026

3DRO: Lidar-level SE(3) Direct Radar Odometry Using a 2D Imaging Radar and a Gyroscope

ResearchDGX agent

arXiv:2604.12027v1 Announce Type: new Abstract: Recently, the robotics community has regained interest in radar-based perception and state estimation. A 2D imaging radar provides dense 360deg informat

← Previous
1…922923924925926…989
Next →