AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
64,475 results
Agents

Staying with the Uncertainty: Uncertainty-Scaffolding Strategies for Artificial Moral Advisors in LLM-to-LLM Simulated Conversations

DGX agent

arXiv:2606.05890v1 Announce Type: new Abstract: LLMs are increasingly deployed as Artificial Moral Advisors (AMA) in a variety of contexts: what kind of conversational patterns should they display? In

agentsarxiv-cs-cl
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

StoryVideoQA: Scaling Deep Video Understanding with a Large-Scale, Multi-Genre and Auto-Generated Dataset

DGX agent

arXiv:2606.06338v1 Announce Type: new Abstract: Video question answering (VideoQA) aims to answer questions about given videos. While existing approaches excel on factoid VideoQA, they struggle with d

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

SubtleMemory: A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horizon AI Agents

DGX agent

arXiv:2606.05761v1 Announce Type: cross Abstract: Persistent AI assistants, such as OpenClaw, accumulate large collections of related memories over long-term interactions. As these memories grow, they

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

Symb-xMIL: Symbolic Explanations for Multiple Instance Learning in Digital Pathology

DGX agent

arXiv:2606.06224v1 Announce Type: new Abstract: Explanations of multiple instance learning (MIL) models are widely used for validation and discovery in digital histopathology. Existing methods primari

safetyarxiv-cs-cv
5 Jun 2026
Applications

Synthetic Data Generation and Vision-based Wrinkle and Keypoint Detection for Bimanual Cloth Manipulation

DGX agent

arXiv:2606.06292v1 Announce Type: new Abstract: Robotic manipulation of textiles remains challenging because continuous deformation and self-occlusions hinder the robust visual perception required to

applicationsarxiv-cs-cv
5 Jun 2026
Local Ai

T-FunS3D: Task-Driven Hierarchical Open-Vocabulary 3D Functionality Segmentation

DGX agent

arXiv:2606.05975v1 Announce Type: new Abstract: Open-vocabulary 3D functionality segmentation enables robots to localize functional object components in 3D scenes. It is a challenging task that requir

local-aiarxiv-cs-cv
5 Jun 2026
Local Ai

T-SAR-JEPA: Self-Supervised Temporal Anomaly Detection in SAR Amplitude Stacks via Latent Prediction

DGX agent

arXiv:2606.05700v1 Announce Type: new Abstract: We present T-SAR-JEPA, a self-supervised framework for temporal anomaly detection in SAR amplitude stacks via latent prediction. A ViT-Base/16 encoder f

local-aiarxiv-cs-cv
5 Jun 2026
Local Ai

TAGA: Terrain-aware Active Gaze Learning for Generalizable Agile Humanoid Locomotion

DGX agent

arXiv:2606.05880v1 Announce Type: new Abstract: Agile humanoid locomotion across diverse challenging terrain demands both wide perceptual coverage and precise local geometry understanding. Motivated b

local-aiarxiv-cs-ro
5 Jun 2026
Safety

TAM: Torque Adaptation Module for Robust Motion Transfer in Manipulation

DGX agent

arXiv:2606.06218v1 Announce Type: new Abstract: A policy tuned for one robot often behaves differently on another, whether due to the sim-to-real gap, unknown payloads, or the differing dynamics of tw

safetyarxiv-cs-ro
5 Jun 2026
Research

Tamaththul3D: High-Fidelity 3D Saudi Sign Language Avatars from Monocular Video

DGX agent

arXiv:2605.05367v2 Announce Type: replace Abstract: Existing 3D sign language avatar reconstruction methods are developed and evaluated exclusively on Western sign languages, and no 3D parametric anno

researcharxiv-cs-cv
5 Jun 2026
Model Releases

TARPO: Token-Wise Latent-Explicit Reasoning via Action-Routing Policy Optimization

DGX agent

arXiv:2606.05859v1 Announce Type: new Abstract: Latent reasoning has emerged as a promising alternative to discrete Chain-of-Thought (CoT) in large language models (LLMs), enabling more expressive rea

model-releasesarxiv-cs-cl
5 Jun 2026
Local Ai

Temporal Preference Concepts and their Functions in a Large Language Model

DGX agent

arXiv:2606.05194v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly being deployed to make decisions that require trading off near-term gains against long-term consequences

local-aiarxiv-cs-cl
5 Jun 2026
Safety

TempoVLA: Learning Speed-Controllable Vision-Language-Action Policies

DGX agent

arXiv:2606.06491v1 Announce Type: new Abstract: Robot manipulation alternates between low-risk transit phases that call for fast execution and high-risk contact stages that demand slow, precise motion

safetyarxiv-cs-ro
5 Jun 2026
Model Releases

Ten Headache Specialists versus Artificial Intelligence for Clinical Literature Summarization: A Critical Evaluation and Comparison

DGX agent

arXiv:2606.05436v1 Announce Type: cross Abstract: Summarizing the latest medical literature to guide clinical decision-making is essential for evidence-based medicine and high-quality patient care. Ye

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

TensorBench: Benchmarking Coding Agents on a Compiler-Based Tensor Framework

DGX agent

arXiv:2606.05570v1 Announce Type: new Abstract: Repository-level coding benchmarks face a trade-off between task difficulty and evaluation reliability: tasks that challenge frontier models often invol

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

Texture-preserving implicit neural representation for Cone beam CT truncated reconstruction

DGX agent

arXiv:2606.06039v1 Announce Type: new Abstract: Cone-beam computed tomography (CBCT) frequently suffers from data truncation, which introduces severe artifacts and limits the effective field of view (

safetyarxiv-cs-cv
5 Jun 2026
Model Releases

TextWand: A Unified Framework for Scene Text Editing

DGX agent

arXiv:2606.05730v1 Announce Type: new Abstract: We propose TextWand, a general-purpose framework that unifies scene text removal, generation, and replacement into a single model. By decomposing comple

model-releasesarxiv-cs-cv
5 Jun 2026
Applications

The Generator-Eraser Paradox: Community Guidelines for Responsible LLM-Assisted Dialect Resource Creation

DGX agent

arXiv:2606.06004v1 Announce Type: new Abstract: Dialect resources occupy a unique position at the intersection of scientific description, cultural preservation, and computational infrastructure. Large

applicationsarxiv-cs-cl
5 Jun 2026
Model Releases

The Granularity Gap: A Multi-Dimensional Longitudinal Audit of Sycophancy in Gemini Models

DGX agent

arXiv:2606.05183v1 Announce Type: new Abstract: Large language models are increasingly deployed as high-stakes advisors, yet standard alignment benchmarks treat sycophancy as a binary failure mode. We

model-releasesarxiv-cs-cl
5 Jun 2026
Research

The Invisible Hand of Physics: When Video Diffusion Models Know More Than They Show

DGX agent

arXiv:2606.05328v1 Announce Type: cross Abstract: Modern video diffusion models generate increasingly realistic and temporally coherent videos, motivating their use as candidate world simulators. Yet

researcharxiv-cs-cv
5 Jun 2026
Model Releases

The Mirage of Performance Gains: Why Contrastive Decoding Fails to Mitigate Object Hallucinations in MLLMs?

DGX agent

arXiv:2504.10020v4 Announce Type: replace Abstract: Contrastive decoding strategies are widely used to reduce object hallucinations in multimodal large language models (MLLMs). These methods work by c

model-releasesarxiv-cs-cl
5 Jun 2026
Applications

The Prosody of Emojis

DGX agent

arXiv:2508.00537v2 Announce Type: replace Abstract: Prosodic features such as pitch, timing, and intonation are central to spoken communication, conveying emotion, intent, and discourse structure. In

applicationsarxiv-cs-cl
5 Jun 2026
Agents

The Self-Correction Illusion: LLMs Correct Others but Not Themselves

DGX agent

arXiv:2606.05976v1 Announce Type: cross Abstract: Recent work shows that LLM agents struggle to correct errors in their own reasoning traces yet show markedly higher correction rates when identical cl

agentsarxiv-cs-cl
5 Jun 2026
Research

The Tell-Tale Norm: ell_2 Magnitude as a Signal for Reasoning Dynamics in Large Language Models

DGX agent

arXiv:2606.06188v1 Announce Type: new Abstract: Recent work has sought to understand Large Language Models (LLMs) reasoning, yet a principled, model-intrinsic signal that captures its layer-wise reaso

researcharxiv-cs-cl
5 Jun 2026
Model Releases

Thinking with Imagination: Agentic Visual Spatial Reasoning with World Simulators

DGX agent

arXiv:2606.06476v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have shown strong visual reasoning capabilities, their spatial reasoning abilities remain largely constrained to the

model-releasesarxiv-cs-cv
5 Jun 2026
Research

Three-Dimensional Retinal Microvasculature Restoration in OCT Angiography

DGX agent

arXiv:2606.05375v1 Announce Type: new Abstract: Optical coherence tomographic angiography (OCTA) is a powerful technique for imaging retinal microvasculature. However, acquiring reliable quantificatio

researcharxiv-cs-cv
5 Jun 2026
Applications

To Be Multimodal or Not to Be: Query-Adaptive Audio-Visual Person Retrieval via Active Modality Detection

DGX agent

arXiv:2606.05931v1 Announce Type: new Abstract: When retrieving a person from a video archive by voice and face, should the system be multimodal or not? In real-world broadcast archives, unlike curate

applicationsarxiv-cs-cl
5 Jun 2026
Model Releases

TopoPult-SSL: Gland-Mask-Free Cross-Device Meibomian Gland Segmentation via Self-Distilled Weak Clinical Priors

DGX agent

arXiv:2606.05347v1 Announce Type: new Abstract: Every new clinical imaging device creates a domain shift where dense gland masks are expensive yet cheap clinical signals -- eyelid outlines, Pult grade

model-releasesarxiv-cs-cv
5 Jun 2026
Safety

Toward Culturally Aligned LLMs through Ontology-Guided Multi-Agent Reasoning

DGX agent

arXiv:2601.21700v3 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly support culturally sensitive decision making, yet often exhibit misalignment due to skewed pretraining dat

safetyarxiv-cs-cl
5 Jun 2026
Safety

Towards a Data Flywheel for Embodied Intelligence in Logistics

DGX agent

arXiv:2606.05960v1 Announce Type: new Abstract: Embodied intelligence is moving from laboratory demonstrations toward industrial deployment, with the logistics industry serving as a key application sc

safetyarxiv-cs-ro
5 Jun 2026
Model Releases

Towards Accurate Heart Rate Measurement from Ultra-Short Video Clips via Periodicity-Guided rPPG Estimation and Signal Reconstruction

DGX agent

arXiv:2506.22078v2 Announce Type: replace Abstract: Many remote Heart Rate (HR) measurement methods focus on estimating remote photoplethysmography (rPPG) signals from video clips lasting around 10 se

model-releasesarxiv-cs-cv
5 Jun 2026
Applications

Towards Label-Noise Resistant Learning via Optimal Brain Damage Masking

DGX agent

arXiv:2508.09697v3 Announce Type: replace-cross Abstract: Noisy labels are inevitable in real-world scenarios. Due to the strong capacity of deep neural networks to memorize corrupted labels, these no

applicationsarxiv-cs-cv
5 Jun 2026
Model Releases

Towards One-to-Many Temporal Grounding

DGX agent

arXiv:2606.06294v1 Announce Type: new Abstract: Temporal Grounding (TG) aims to localize video segments corresponding to a textual query. Prior research predominantly focuses on single-segment retriev

model-releasesarxiv-cs-cv
5 Jun 2026
Hardware

Towards Realistic 3D Sonar Simulation

DGX agent

arXiv:2606.06130v1 Announce Type: new Abstract: As underwater robotics research increasingly addresses complex 3D perception and autonomous navigation, the fidelity of sonar simulation has become a ke

hardwarearxiv-cs-ro
5 Jun 2026
Research

Towards Truly Multilingual ASR: Generalizing Code-Switching ASR to Unseen Language Pairs

DGX agent

arXiv:2606.05846v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) has become a key technology for human--AI interaction. However, code-switching ASR (CS-ASR) remains particularly chal

researcharxiv-cs-cl
5 Jun 2026
Local Ai

Trajectory Dynamics in Language Model Hidden States Predict Human Processing Costs Beyond Surprisal

DGX agent

arXiv:2606.05346v1 Announce Type: new Abstract: Human language comprehension unfolds sequentially: each word is processed in the context of those that came before, and the interpretation builds increm

local-aiarxiv-cs-cl
5 Jun 2026
Safety

Two-Way Is Better Than One: Bidirectional Alignment with Cycle Consistency for Exemplar-Free Class-Incremental Learning

DGX agent

arXiv:2606.05675v1 Announce Type: cross Abstract: Continual learning (CL) seeks models that acquire new skills without erasing prior knowledge. In exemplar-free class-incremental learning (EFCIL), thi

safetyarxiv-cs-cv
5 Jun 2026
Model Releases

UltraVR: A Diagnostic Ultra-Resolution Image-VQA Benchmark for Evidence-Grounded Reasoning

DGX agent

arXiv:2606.05576v1 Announce Type: new Abstract: Vision-language models (VLMs) excel on visual question answering and multimodal reasoning benchmarks. Yet their capability on ultra-resolution images -

model-releasesarxiv-cs-cv
5 Jun 2026
Hardware

Uncertainty-Aware Adaptive Sensor Fusion for Autonomous Navigation

DGX agent

arXiv:2606.05437v1 Announce Type: cross Abstract: This work introduces a hybrid deep learning approach integrated with an Unscented Kalman Filter (UKF) to enhance pose estimation accuracy in Visual-In

hardwarearxiv-cs-cv
5 Jun 2026
Research

UnHype: CLIP-Guided Hypernetworks for Dynamic LoRA Unlearning

DGX agent

arXiv:2602.03410v2 Announce Type: replace Abstract: Recent advances in large-scale diffusion models have intensified concerns about their potential misuse, particularly in generating realistic yet har

researcharxiv-cs-cv
5 Jun 2026
Model Releases

Unifying Dataset Pruning and Distillation for Efficient Large-scale Compression

DGX agent

arXiv:2502.06434v2 Announce Type: replace Abstract: Dataset pruning (DP) and dataset distillation (DD) fundamentally differ in their outputs: DP selects original image subsets, while DD generates synt

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

UniPixie: Unified and Probabilistic 3D Physics Learning via Flow Matching

DGX agent

arXiv:2606.05399v1 Announce Type: new Abstract: Existing feed-forward networks excel at predicting a single set of physical properties from visual appearance, but this point-estimate paradigm fundamen

model-releasesarxiv-cs-cv
5 Jun 2026
Safety

UNIVID: Unified Vision-Language Model for Video Moderation

DGX agent

arXiv:2606.05748v1 Announce Type: cross Abstract: Global-scale video moderation faces a dual challenge: the need for fine-grained multi-modal reasoning and the demand for interpretable outputs to supp

safetyarxiv-cs-cl
5 Jun 2026
Safety

Unpaired RGB-Thermal Gaussian-Splatting Using Visual Geometric Transformers

DGX agent

arXiv:2606.05491v1 Announce Type: new Abstract: Multi-modal novel view synthesis (NVS) combining RGB and thermal imagery enables precise 3D scene reconstruction with visual and thermal information. Ho

safetyarxiv-cs-cv
5 Jun 2026
Research

Unsupervised Monocular 3D Keypoint Discovery from Multi-View Diffusion Priors

DGX agent

arXiv:2507.12336v2 Announce Type: replace Abstract: Most existing 3D keypoint estimation methods rely on manual annotations or calibrated multi-view images, both of which are expensive to collect. Thi

researcharxiv-cs-cv
5 Jun 2026
Agents

Unsupervised Skill Discovery for Agentic Data Analysis

DGX agent

arXiv:2606.06416v1 Announce Type: cross Abstract: Inference-time skill augmentation provides a lightweight way to improve data-analytic agents by injecting reusable procedural knowledge without updati

agentsarxiv-cs-cl
5 Jun 2026
Safety

Unveiling the Unknown: Open Vocabulary Object Detection with Scene Graphs

DGX agent

arXiv:2606.05916v1 Announce Type: new Abstract: Open-vocabulary object detection seeks to identify novel object categories that were not part of the training data. Many knowledge distillation-based ap

safetyarxiv-cs-cv
5 Jun 2026
Research

USAD 2.0: Scaling Representation Distillation for Universal Audio Understanding

DGX agent

arXiv:2606.06444v1 Announce Type: cross Abstract: Audio encoders are critical to modern audio applications as large language models (LLMs) increasingly rely on a single encoder for diverse inputs. Whi

researcharxiv-cs-cl
5 Jun 2026
← Previous
1…629630631632633…1344
Next →