AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
Human
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
1 Jul 2026

Think While You Map: Asynchronous Vision-Language Agents for Incremental 3D Scene Graphs

ResearchDGX agent

arXiv:2606.31471v1 Announce Type: new Abstract: Open-vocabulary 3D scene graph methods typically operate in two stages: first reconstruct, then enrich with vision-language models, leaving the graph un

Thinking Before Retrieving: Robust Zero-Shot Composed Image Retrieval via Strategic Planning and Self-Criticism

ResearchDGX agent

arXiv:2606.31222v1 Announce Type: new Abstract: Composed image retrieval requires identifying a target image from a gallery by integrating a reference image with a textual modification instruction. In

Token-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement Learning

SafetyDGX agent

arXiv:2606.31599v1 Announce Type: cross Abstract: Vision-language models (VLMs) combining reinforcement learning (RL) ignite remarkable progress in multimodal reasoning, yet still struggle with medica

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Tone-Conditioned Curriculum Learning for Low-Resource Bantu Speech Recognition

ApplicationsDGX agent

arXiv:2606.31642v1 Announce Type: new Abstract: Southern Bantu languages are spoken by over 80 million people, yet current foundation ASR models still produce zero-shot WER above 100%, which limits pr

TORA: Topological Representation Alignment for 3D Shape Assembly

SafetyDGX agent

arXiv:2604.04050v2 Announce Type: replace Abstract: Flow-matching methods for 3D shape assembly learn point-wise velocity fields that transport parts toward assembled configurations, yet they receive

TotalFM: An Organ-Separated 3D-CT Foundation Model Leveraging Large-Scale Routine Clinical Radiology Data

ResearchDGX agent

arXiv:2601.00260v2 Announce Type: replace Abstract: While foundation models in radiology are expected to be applied to various clinical tasks, computational cost constraints remain a major challenge w

Toward AI-Resilient Assessment in Computer Science Courses in an AI-Native World

TutorialsDGX agent

arXiv:2606.30655v1 Announce Type: cross Abstract: AI-native course assessments in senior computer science courses and related fields should grade students by AI-resilient skill: the ability to achieve

Towards a foundational model for recognising diastematic Gregorian notation

ResearchDGX agent

arXiv:2606.31454v1 Announce Type: new Abstract: Optical recognition of Gregorian notation has recently been attempted with end-to-end methods, with four datasets introduced. However, each of these dat

Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation

ResearchDGX agent

arXiv:2606.31088v1 Announce Type: new Abstract: Conversational talking face generation has recently attracted increasing attention, aiming to synthesize interactive talking videos where characters spe

Towards Inclusive Mobility Modeling: Characterizing and Evaluating Elderly Trajectory Patterns in Urban Systems

Local AiDGX agent

arXiv:2606.31207v1 Announce Type: new Abstract: The rapid advance of smart cities increasingly depends on trajectory data mining, yet underrepresented demographic groups, particularly the elderly, are

Towards Voxel Spacing Consistency for Medical Image Segmentation

ResearchDGX agent

arXiv:2606.31839v1 Announce Type: new Abstract: Volumetric medical image segmentation is essential for both preoperative diagnosis and intraoperative guidance. While recent years have witnessed rapid

Toxicity Assessment in Preclinical Histopathology via Class-Aware Mahalanobis Distance for Known and Novel Anomalies

SafetyDGX agent

arXiv:2602.02124v2 Announce Type: replace-cross Abstract: Drug-induced toxicity is a leading cause of preclinical and early-clinical failure, making early detection critical. Histopathology is the gol

TraCeS: Learning Per-Timestep Constraint-Violation Credit from Sparse Trajectory-Level Labels

SafetyDGX agent

arXiv:2504.12557v3 Announce Type: replace-cross Abstract: Ensuring safe behavior in reinforcement learning (RL) is challenging when safety constraints are implicit and cannot be densely measured. In m

Training Therapeutic Judges and Multi-Agent Systems for Human-Aligned Mental Health Support

SafetyDGX agent

arXiv:2606.30887v1 Announce Type: cross Abstract: Large language models show promise for mental health support, yet therapeutic quality improves only when evaluation functions as an actionable control

Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation

SafetyDGX agent

arXiv:2606.31184v1 Announce Type: cross Abstract: Adaptive experiments for average treatment effects (ATE) require randomized allocations balancing valid inference with statistical efficiency. The ora

TreeAgent: A Generalizable Multi-Agent Framework for Automated Bias Labeling in Forestry via Compiled Expert Rules and Vision-Language Models

SafetyDGX agent

arXiv:2606.31976v1 Announce Type: new Abstract: Human-labeled data are widely used as reference annotations in ML, despite known variability across annotators in many expert-driven domains. In additio

TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.32017v1 Announce Type: cross Abstract: Agentic reinforcement learning requires assigning credit to environment-facing actions such as searches, clicks, edits, navigation commands, and objec

Triospect: A Three-Dimensional Framework for Robust Statistical AI-Generated Text Detection Against Diverse Attacks

ResearchDGX agent

arXiv:2606.31074v1 Announce Type: cross Abstract: Existing AI-generated text detectors are vulnerable to attacks that manipulate textual characteristics. In this study, we propose a novel Triospect De

Truth or Sophistry? LoFa: A Benchmark for LLM Robustness Against Logical Fallacies

Model ReleasesDGX agent

arXiv:2606.31039v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit strong semantic capabilities, yet their resilience to manipulative linguistic patterns such as logical fallacies re

TSHA: A Benchmark for Visual Language Models in Trustworthy Safety Hazard Assessment Scenarios

Model ReleasesDGX agent

arXiv:2603.29759v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have accelerated their application to indoor safety hazards assessment. However, existing ben

UHD-MFF: Shattering Barriers in Multi-Focus Ultra-High-Definition Image Fusion via Learnable Lookup Tables

ResearchDGX agent

arXiv:2606.31242v1 Announce Type: new Abstract: With the advancement of imaging technology, ultra-high-definition images have become increasingly essential in modern visual applications. However, exis

Understanding and Evaluating Claw-like Agent Security Through a Computer-Systems Lens

Model ReleasesDGX agent

arXiv:2606.30755v1 Announce Type: cross Abstract: Claw-like AI agents (e.g., OpenClaw) are always-on processes with persistent access to credentials, files, tools, and external services. They take on

UniCoder: Unified Visual-to-Code Generation via Symbolic Rewards and Reference-Guided Code Optimization

Model ReleasesDGX agent

arXiv:2606.31732v1 Announce Type: new Abstract: Visual-to-Code generation, which transforms scientific plots, vector graphics, and webpages into executable scripts, demands a level of pixel-precise al

Unified Structural-Hydrodynamic Modeling of Underwater Underactuated Mechanisms and Soft Robots

Model ReleasesDGX agent

arXiv:2603.07939v2 Announce Type: replace Abstract: Underwater robots are widely deployed for ocean exploration and manipulation. Underactuated mechanisms are advantageous in aquatic environments beca

UniSAE: Unified Speech Attribute Editing on Speaker, Emotion and Low-Level Content via Discrete Phonetic Posteriorgram Modelling

ResearchDGX agent

arXiv:2606.31128v1 Announce Type: cross Abstract: Speech editing aims to modify specific portions of an utterance while preserving the remaining speech. Existing approaches primarily focus on word-lev

UniTac: A Unified Multimodal Model for Cross-Sensor Tactile Understanding and Generation

SafetyDGX agent

arXiv:2606.31451v1 Announce Type: cross Abstract: Unified multimodal models (UMMs) have shown great promise in integrating understanding and generation across diverse modalities. However, existing res

UniTacVLA: Unified Tactile Understanding and Prediction in Vision Language Action Models

ApplicationsDGX agent

arXiv:2606.31723v1 Announce Type: new Abstract: Vision-language-action (VLA) models have achieved strong performance in many robotic manipulation tasks, yet remain limited in contact-rich dexterous ma

Unsupervised Data-Efficient Cross-Modal Retrieval with Global-Neighborhood Alignment Hashing

SafetyDGX agent

arXiv:2606.31517v1 Announce Type: cross Abstract: Compared to supervised cross-modal hashing (CMH), unsupervised CMH reduces the reliance on manual labeling by learning binary codes from unlabeled ima

Unsupervised Thermodynamics of Molecular Diffusion Models: Action-Operator Semantics and Auditable Free-Energy Readout

ResearchDGX agent

arXiv:2606.30687v1 Announce Type: cross Abstract: Diffusion models are increasingly utilized for modeling molecular structures and conformational ensembles, yet the thermodynamic meaning of their lear

Unveiling Transferability in Trajectory Prediction via Latent Scene Embeddings

AgentsDGX agent

arXiv:2606.30777v1 Announce Type: new Abstract: The growing availability of trajectory datasets has fueled major advances in data-driven motion prediction. Yet, models trained on one dataset often fai

Usage frequency and application variety of research methods in library and information science: Continuous investigation from 1991 to 2021

ResearchDGX agent

arXiv:2606.31081v1 Announce Type: cross Abstract: The present study analyzed over 26,000 research articles published between 1991 and 2021 in twenty-one major LIS (Library and Information Science) jou

Using AI Agents to Automate Black-Box Audits of Personalization Algorithms at Scale

AgentsDGX agent

arXiv:2606.30801v1 Announce Type: new Abstract: Personalization algorithms determine what content users encounter on online platforms. Auditing these systems is difficult because independent auditors

Verification-Gated Agentic Mission-State Governance for Intelligent Industrial Multi-Robot Systems

SafetyDGX agent

arXiv:2606.31339v1 Announce Type: new Abstract: Agentic artificial intelligence is increasingly used to decompose industrial tasks, propose robot actions, and adapt execution plans in dynamic cyber-ph

Verify when Uncertain: Beyond Self-Consistency in Black Box Hallucination Detection

ResearchDGX agent

arXiv:2502.15845v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) often hallucinate, limiting their reliability in sensitive applications. In black-box settings, several self-cons

VertiAdaptor: Online Kinodynamics Adaptation for Vertically Challenging Terrain

AgentsDGX agent

arXiv:2603.06887v3 Announce Type: replace Abstract: Autonomous driving in off-road environments presents significant challenges due to the dynamic and unpredictable nature of unstructured terrain. Tra

VIGOR: VIdeo Geometry-Oriented Reward for Temporal Generative Alignment

SafetyDGX agent

arXiv:2603.16271v3 Announce Type: replace Abstract: Video diffusion models lack explicit geometric supervision during training, leading to inconsistency artifacts such as object deformation, spatial d

Vision-Language Procedural Reasoning for Context-Aware Reward Modeling of Robotic Endovascular Guidewire Navigation

SafetyDGX agent

arXiv:2606.30698v1 Announce Type: new Abstract: Robotic-assisted endovascular interventions demand accurate, stable, and context-aware guidewire navigation in complex and patient-specific vascular ana

Visual Prompt Discovery via Semantic Exploration

AgentsDGX agent

arXiv:2603.16250v2 Announce Type: replace-cross Abstract: LVLMs encounter significant challenges in image understanding and visual reasoning, leading to critical perception failures. Visual prompts, w

Visual Semantic Entropy: Do Vision Language Models Recognize Visual Ambiguity?

ResearchDGX agent

arXiv:2606.31407v1 Announce Type: cross Abstract: Vision-language models can produce confident answers on visually ambiguous inputs, resulting in biased predictions. Common entropy-based methods, such

Visualizing High-Dimensional Graph Embeddings via Informed Multi-View Projections

ResearchDGX agent

arXiv:2606.31119v1 Announce Type: new Abstract: Graphs are commonly visualized in 2D, where humans readily interpret spatial relationships, yet such layouts often distort higher-dimensional structure.

ViTL: Temporal Logic-Guided Zero-Shot Natural Language Navigation via Vision-Language Models

TutorialsDGX agent

arXiv:2606.30696v1 Announce Type: cross Abstract: Enabling robots to follow natural language commands to complete zero-shot long-horizon tasks remains challenging. It requires extracting implicit temp

Von Mises Based Uncertainty Quantification for Closely Spaced Automotive Radar Targets

ResearchDGX agent

arXiv:2606.31473v1 Announce Type: cross Abstract: This work investigates uncertainty-aware deep learning approaches for direction of arrival (DOA) estimation in automotive radar, focusing on probabili

VS3R: Robust Full-frame Video Stabilization via Deep 3D Reconstruction

ResearchDGX agent

arXiv:2603.05851v2 Announce Type: replace Abstract: Video stabilization aims to mitigate camera shake but faces a fundamental trade-off between geometric robustness and full-frame consistency. While 2

WAFT-Stereo: Warping-Alone Field Transforms for Stereo Matching

ResearchDGX agent

arXiv:2603.24836v3 Announce Type: replace Abstract: We introduce WAFT-Stereo, a simple and effective warping-based method for stereo matching. WAFT-Stereo demonstrates that cost volumes, a common desi

Wait, am I Being Fair? Characterizing Deductive Stereotyping and Mitigating It with Fair-GCG

SafetyDGX agent

arXiv:2606.30989v1 Announce Type: cross Abstract: Warning: This paper contains several toxic and offensive statements. While reasoning generally improves fairness in recent large language models (LLMs

Warp RL: Reshaping Base Policy Distributions for Dynamics Adaptation

SafetyDGX agent

arXiv:2606.31043v1 Announce Type: new Abstract: Residual reinforcement learning adapts a pretrained robot policy by learning an additive correction to its actions. While effective when adaptation amou

WarpHammer: Densifying Scene Warps with 3D Object Priors for Extreme View Synthesis

ResearchDGX agent

arXiv:2606.31258v1 Announce Type: new Abstract: Projection-conditioned novel view synthesis (NVS) warps an explicit 3D reconstruction of the input view into the target camera and conditions a generato

WarpI2I: Image Warping for Image-to-Image Translation

ResearchDGX agent

arXiv:2606.31018v1 Announce Type: new Abstract: Image-to-image (I2I) translation has achieved strong results in tasks like human relighting and driving scene translation using latent diffusion models

WaterGen: Decoupling Scene and Medium in Underwater Image Generation

ResearchDGX agent

arXiv:2606.31147v1 Announce Type: new Abstract: Underwater computer vision tasks, such as detection, restoration, and segmentation, are limited by the scarcity of large-scale and diverse training data

Wavelet-Optimized Pseudo-3D Accelerated Diffusion Model for Truncated Computed Laminography

ApplicationsDGX agent

arXiv:2606.31318v1 Announce Type: new Abstract: Computed Laminography (CL) is a key technology for the nondestructive testing of large plate-shaped objects. However, field-of-view (FOV) limitations in

What Counts as an Error? Dual-Reference Benchmarking for Atypical ASR

Model ReleasesDGX agent

arXiv:2606.31112v1 Announce Type: new Abstract: ASR systems have been often reported to underperform on atypical speech. An often conflated compounding factor is the existence of two valid transcripti

What Drives Interactive Improvement from Feedback?

AgentsDGX agent

arXiv:2606.30774v1 Announce Type: new Abstract: We study when natural-language feedback produces improvement beyond the gains obtainable from repeated attempts alone. In multi-turn language agent sett

What If We Allocate Test-Time Compute Adaptively?

Model ReleasesDGX agent

arXiv:2602.01070v5 Announce Type: replace Abstract: Test-time compute scaling allocates inference computation uniformly, uses fixed sampling strategies, and applies verification only for reranking. In

What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States

Model ReleasesDGX agent

arXiv:2606.31612v1 Announce Type: new Abstract: Mobile GUI agents increasingly face long-horizon tasks that require reading, updating, and reusing task-relevant data across pages and applications. Exi

What Probing Reveals about Autonomous Driving: Linking Internal Prediction Errors to Ego Planning

SafetyDGX agent

arXiv:2606.31106v1 Announce Type: cross Abstract: Large-scale datasets and fast simulators have enabled improvements in driving policies that appear safe and robust, yet strong performance in nominal

When Calibration Rankings Reverse: Accuracy-Controlled Evaluation for Fair Comparison of LLMs

ResearchDGX agent

arXiv:2606.30814v1 Announce Type: new Abstract: Calibration evaluates whether a model confidence aligns with its empirical accuracy. Existing studies often compare the calibration of different large l

When Does Learning to Stop Help? A Cost-Aware Study of Early Exits in Reasoning Models

Model ReleasesDGX agent

arXiv:2606.30852v1 Announce Type: new Abstract: Reasoning models spend different amounts of useful computation across instances, but it remains unclear when a learned stopping rule improves over simpl

When few labeled target data suffice: a theory of semi-supervised domain adaptation via fine-tuning from multiple adaptive starts

ResearchDGX agent

arXiv:2507.14661v2 Announce Type: replace-cross Abstract: Semi-supervised domain adaptation (SSDA) seeks to achieve accurate predictions in a target domain with limited labeled target data by exploiti

When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors

Model ReleasesDGX agent

arXiv:2606.32029v1 Announce Type: cross Abstract: While large language models (LLMs) perform well on table tasks, they still make data referencing errors (DREs), i.e., incorrectly citing or omitting t

When Regulation Has Memory: Hysteresis and Control Burden in Artificial Agency

AgentsDGX agent

arXiv:2606.30975v1 Announce Type: new Abstract: Adaptive agents are usually judged by what they do, but an agent can appear stable while the internal effort required to keep it stable is increasing. T

← Previous
1…306307308309310…1025
Next →