AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,360 results
Research

Sparsity-Inducing Divergence Losses for Biometric Verification

DGX agent

arXiv:2606.31664v1 Announce Type: cross Abstract: Performance in face and speaker verification is largely driven by margin-penalty softmax losses such as CosFace and ArcFace. Recently introduced alpha

researcharxiv-cs-ai
1 Jul 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SPFSplatV2: Efficient Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views

DGX agent

arXiv:2509.17246v2 Announce Type: replace Abstract: We introduce SPFSplatV2, an efficient feed-forward framework for 3D Gaussian splatting from sparse multi-view images, requiring no ground-truth pose

researcharxiv-cs-cv
1 Jul 2026
Research

SpheRoPE: Zero-Shot Optimization-Free 360 Panorama Generation with Spherical RoPE

DGX agent

arXiv:2606.32033v1 Announce Type: new Abstract: We present a zero-shot, training-free and optimization-free framework for generating 360 panoramic images and videos by directly injecting spherical pri

researcharxiv-cs-cv
1 Jul 2026
Research

Stage-Transition Dense Reward Modeling for Reinforcement Learning

DGX agent

arXiv:2606.31377v1 Announce Type: cross Abstract: Reinforcement learning for long-horizon robotic manipulation is often limited by sparse and delayed rewards, while manually designing dense shaping si

researcharxiv-cs-ai
1 Jul 2026
Research

StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation

DGX agent

arXiv:2602.23721v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models integrate visual observations and language instructions to predict robot actions, demonstrating promising

researcharxiv-cs-cv
1 Jul 2026
Research

Streaming Gaussian Encoding for 4D Panoptic Occupancy Tracking

DGX agent

arXiv:2606.30754v1 Announce Type: new Abstract: Camera-based 4D panoptic occupancy tracking (4D-POT) is a promising paradigm for holistic scene understanding from multi-view imagery, enabling joint re

researcharxiv-cs-cv
1 Jul 2026
Research

Structured SIR: Efficient and Expressive Importance-Weighted Inference for High-Dimensional Image Registration

DGX agent

arXiv:2603.17415v2 Announce Type: replace-cross Abstract: Image registration is an ill-posed dense vision task, where multiple solutions achieve similar loss values, motivating probabilistic inference

researcharxiv-cs-cv
1 Jul 2026
Research

Surprise as a Signal for Plasticity and Metacognition

DGX agent

arXiv:2606.31495v1 Announce Type: new Abstract: We study a single idea across two settings: that a prediction-error signal, computed by a small predictor over the latent space of a frozen encoder, can

researcharxiv-cs-ai
1 Jul 2026
Research

SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation

DGX agent

arXiv:2606.31259v1 Announce Type: cross Abstract: Diffusion-based text-to-audio (TTA) models achieve impressive synthesis quality but suffer from high inference latency due to iterative multi-step den

researcharxiv-cs-ai
1 Jul 2026
Research

Symmetry in language statistics shapes the geometry of model representations

DGX agent

arXiv:2602.15029v3 Announce Type: replace-cross Abstract: The internal representations learned by language models consistently exhibit striking geometric structure: calendar months organize into a cir

researcharxiv-cs-cl
1 Jul 2026
Research

T-QPM: Enabling Temporal Out-Of-Distribution Detection and Domain Generalization for Vision-Language Models in Open-World

DGX agent

arXiv:2603.18481v2 Announce Type: replace Abstract: Out-of-distribution (OOD) detection remains a critical challenge in open-world learning, where models must adapt to evolving data distributions. Whi

researcharxiv-cs-cv
1 Jul 2026
Research

TabPATE: Differentially Private Tabular In-Context Learning Without Public Data

DGX agent

arXiv:2606.31474v1 Announce Type: new Abstract: Tabular foundation models enable accurate in-context learning (ICL) from small labeled datasets, but the private records placed in context can leak thro

researcharxiv-cs-lg
1 Jul 2026
Research

Teaching Models to Teach Themselves: Reasoning at the Edge of Learnability

DGX agent

arXiv:2601.18778v3 Announce Type: replace-cross Abstract: RL methods for scaling large reasoning models stall on datasets with low initial success rates, and thus little training signal. We investigat

researcharxiv-cs-cl
1 Jul 2026
Research

Team MKC at CLPsych 2026: Capturing and Characterizing Mental Health Changes through Social Media Timeline Dynamics

DGX agent

arXiv:2606.31464v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have motivated their adoption across a wide range of domains, including Artificial Intelligence (AI) f

researcharxiv-cs-ai
1 Jul 2026
Research

Technical Report of RoboSpatial Challenge at CVPR 2026: Selective Reasoning Activation and Reference-Frame Disambiguation for Embodied Spatial Reasoning

DGX agent

arXiv:2606.31645v1 Announce Type: new Abstract: Vision-language models achieve strong general perception but often struggle with the spatial reasoning required for embodied tasks. We present RoboSpati

researcharxiv-cs-cv
1 Jul 2026
Research

Temporal Preservation over Processing: Diagnosing and Designing Spatiotemporal Single-Stage Video Detectors

DGX agent

arXiv:2606.31421v1 Announce Type: cross Abstract: Single-stage video object detectors are increasingly deployed in time-critical applications, yet it remains unclear whether these models genuinely rea

researcharxiv-cs-ai
1 Jul 2026
Research

Temporal Training Strategies for Left Atrium and Left Atrial Appendage Segmentation in Dynamic Contrast 4DCT

DGX agent

arXiv:2606.31444v1 Announce Type: new Abstract: Dynamic contrast-enhanced cardiac CT enables time-resolved analysis of contrast filling and washout in the left atrium (LA) and left atrial appendage (L

researcharxiv-cs-cv
1 Jul 2026
Research

The Geometry of Efficient Nonconvex Sampling

DGX agent

arXiv:2603.25622v2 Announce Type: replace-cross Abstract: We present an efficient algorithm for uniformly sampling from an arbitrary compact body X subset R^n from a warm start under isoperimetry and

researcharxiv-cs-lg
1 Jul 2026
Research

The Impact of Dimensionality on the Stability of Node Embeddings

DGX agent

arXiv:2604.08492v2 Announce Type: replace Abstract: Previous work has shown that node embedding methods can produce different representations and downstream predictions across repeated training runs,

researcharxiv-cs-lg
1 Jul 2026
Research

The Label Imitation Game: Turing Test Network for Zero-Shot Pseudo-Label Pruning

DGX agent

arXiv:2606.30875v1 Announce Type: cross Abstract: Foundation model pseudo-labeling - labeling data strictly via zero-shot inference - enables massive scale, but performance is undermined by hallucinat

researcharxiv-cs-ai
1 Jul 2026
Research

The Quadruped Soft Tail: Compliant Grasping and Swabbing for Contamination Surveys in Harsh Environments

DGX agent

arXiv:2606.30900v1 Announce Type: new Abstract: Beryllium contamination surveys in radioactive areas are challenging for robots in environments cluttered with cables and electronics. To address this p

researcharxiv-cs-ro
1 Jul 2026
Research

Think While You Map: Asynchronous Vision-Language Agents for Incremental 3D Scene Graphs

DGX agent

arXiv:2606.31471v1 Announce Type: new Abstract: Open-vocabulary 3D scene graph methods typically operate in two stages: first reconstruct, then enrich with vision-language models, leaving the graph un

researcharxiv-cs-cv
1 Jul 2026
Research

Thinking Before Retrieving: Robust Zero-Shot Composed Image Retrieval via Strategic Planning and Self-Criticism

DGX agent

arXiv:2606.31222v1 Announce Type: new Abstract: Composed image retrieval requires identifying a target image from a gallery by integrating a reference image with a textual modification instruction. In

researcharxiv-cs-ai
1 Jul 2026
Research

TotalFM: An Organ-Separated 3D-CT Foundation Model Leveraging Large-Scale Routine Clinical Radiology Data

DGX agent

arXiv:2601.00260v2 Announce Type: replace Abstract: While foundation models in radiology are expected to be applied to various clinical tasks, computational cost constraints remain a major challenge w

researcharxiv-cs-cv
1 Jul 2026
Research

Towards a foundational model for recognising diastematic Gregorian notation

DGX agent

arXiv:2606.31454v1 Announce Type: new Abstract: Optical recognition of Gregorian notation has recently been attempted with end-to-end methods, with four datasets introduced. However, each of these dat

researcharxiv-cs-cv
1 Jul 2026
Research

Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation

DGX agent

arXiv:2606.31088v1 Announce Type: new Abstract: Conversational talking face generation has recently attracted increasing attention, aiming to synthesize interactive talking videos where characters spe

researcharxiv-cs-cv
1 Jul 2026
Research

Towards Voxel Spacing Consistency for Medical Image Segmentation

DGX agent

arXiv:2606.31839v1 Announce Type: new Abstract: Volumetric medical image segmentation is essential for both preoperative diagnosis and intraoperative guidance. While recent years have witnessed rapid

researcharxiv-cs-cv
1 Jul 2026
Research

Triospect: A Three-Dimensional Framework for Robust Statistical AI-Generated Text Detection Against Diverse Attacks

DGX agent

arXiv:2606.31074v1 Announce Type: cross Abstract: Existing AI-generated text detectors are vulnerable to attacks that manipulate textual characteristics. In this study, we propose a novel Triospect De

researcharxiv-cs-ai
1 Jul 2026
Research

UHD-MFF: Shattering Barriers in Multi-Focus Ultra-High-Definition Image Fusion via Learnable Lookup Tables

DGX agent

arXiv:2606.31242v1 Announce Type: new Abstract: With the advancement of imaging technology, ultra-high-definition images have become increasingly essential in modern visual applications. However, exis

researcharxiv-cs-cv
1 Jul 2026
Research

UniSAE: Unified Speech Attribute Editing on Speaker, Emotion and Low-Level Content via Discrete Phonetic Posteriorgram Modelling

DGX agent

arXiv:2606.31128v1 Announce Type: cross Abstract: Speech editing aims to modify specific portions of an utterance while preserving the remaining speech. Existing approaches primarily focus on word-lev

researcharxiv-cs-ai
1 Jul 2026
Research

Unsupervised Thermodynamics of Molecular Diffusion Models: Action-Operator Semantics and Auditable Free-Energy Readout

DGX agent

arXiv:2606.30687v1 Announce Type: cross Abstract: Diffusion models are increasingly utilized for modeling molecular structures and conformational ensembles, yet the thermodynamic meaning of their lear

researcharxiv-cs-ai
1 Jul 2026
Research

Verify when Uncertain: Beyond Self-Consistency in Black Box Hallucination Detection

DGX agent

arXiv:2502.15845v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) often hallucinate, limiting their reliability in sensitive applications. In black-box settings, several self-cons

researcharxiv-cs-ai
1 Jul 2026
Research

Visual Semantic Entropy: Do Vision Language Models Recognize Visual Ambiguity?

DGX agent

arXiv:2606.31407v1 Announce Type: cross Abstract: Vision-language models can produce confident answers on visually ambiguous inputs, resulting in biased predictions. Common entropy-based methods, such

researcharxiv-cs-ai
1 Jul 2026
Research

Visualizing High-Dimensional Graph Embeddings via Informed Multi-View Projections

DGX agent

arXiv:2606.31119v1 Announce Type: new Abstract: Graphs are commonly visualized in 2D, where humans readily interpret spatial relationships, yet such layouts often distort higher-dimensional structure.

researcharxiv-cs-lg
1 Jul 2026
Research

Von Mises Based Uncertainty Quantification for Closely Spaced Automotive Radar Targets

DGX agent

arXiv:2606.31473v1 Announce Type: cross Abstract: This work investigates uncertainty-aware deep learning approaches for direction of arrival (DOA) estimation in automotive radar, focusing on probabili

researcharxiv-cs-ai
1 Jul 2026
Research

VS3R: Robust Full-frame Video Stabilization via Deep 3D Reconstruction

DGX agent

arXiv:2603.05851v2 Announce Type: replace Abstract: Video stabilization aims to mitigate camera shake but faces a fundamental trade-off between geometric robustness and full-frame consistency. While 2

researcharxiv-cs-cv
1 Jul 2026
Research

WAFT-Stereo: Warping-Alone Field Transforms for Stereo Matching

DGX agent

arXiv:2603.24836v3 Announce Type: replace Abstract: We introduce WAFT-Stereo, a simple and effective warping-based method for stereo matching. WAFT-Stereo demonstrates that cost volumes, a common desi

researcharxiv-cs-cv
1 Jul 2026
Research

WarpHammer: Densifying Scene Warps with 3D Object Priors for Extreme View Synthesis

DGX agent

arXiv:2606.31258v1 Announce Type: new Abstract: Projection-conditioned novel view synthesis (NVS) warps an explicit 3D reconstruction of the input view into the target camera and conditions a generato

researcharxiv-cs-cv
1 Jul 2026
Research

WarpI2I: Image Warping for Image-to-Image Translation

DGX agent

arXiv:2606.31018v1 Announce Type: new Abstract: Image-to-image (I2I) translation has achieved strong results in tasks like human relighting and driving scene translation using latent diffusion models

researcharxiv-cs-cv
1 Jul 2026
Research

WaterGen: Decoupling Scene and Medium in Underwater Image Generation

DGX agent

arXiv:2606.31147v1 Announce Type: new Abstract: Underwater computer vision tasks, such as detection, restoration, and segmentation, are limited by the scarcity of large-scale and diverse training data

researcharxiv-cs-cv
1 Jul 2026
Research

When Calibration Rankings Reverse: Accuracy-Controlled Evaluation for Fair Comparison of LLMs

DGX agent

arXiv:2606.30814v1 Announce Type: new Abstract: Calibration evaluates whether a model confidence aligns with its empirical accuracy. Existing studies often compare the calibration of different large l

researcharxiv-cs-cl
1 Jul 2026
Research

When few labeled target data suffice: a theory of semi-supervised domain adaptation via fine-tuning from multiple adaptive starts

DGX agent

arXiv:2507.14661v2 Announce Type: replace-cross Abstract: Semi-supervised domain adaptation (SSDA) seeks to achieve accurate predictions in a target domain with limited labeled target data by exploiti

researcharxiv-cs-lg
1 Jul 2026
Research

When Reranking Hurts: Uncertainty-Based Gating for Few-Shot Reranking

DGX agent

arXiv:2606.31087v1 Announce Type: cross Abstract: Few-shot selection typically assumes that reranking retrieved examples always improves performance. We challenge this view by identifying that the exp

researcharxiv-cs-ai
1 Jul 2026
Research

When Sinks Help or Hurt: Unified Framework for Attention Sink in Large Vision-Language Models

DGX agent

arXiv:2604.03316v2 Announce Type: replace Abstract: Attention sinks are defined as tokens that attract disproportionate attention. While these have been studied in single modality transformers, their

researcharxiv-cs-cv
1 Jul 2026
Research

When to Truncate a Feature Ranking: A Residual-Overlap Stopping Rule for Subset Selection

DGX agent

arXiv:2606.31686v1 Announce Type: cross Abstract: Feature rankings are widely used in supervised feature selection because they are simple, scalable and easy to interpret. Variables are first ranked b

researcharxiv-cs-ai
1 Jul 2026
Research

Who Determines the Meaning of an Emotion? Affective Sovereignty as an Epistemic Consequence of Measurement Limits

DGX agent

arXiv:2606.31442v1 Announce Type: new Abstract: Emotion-sensing AI is rapidly becoming embedded in vehicles, home appliances, dialogue agents, and social infrastructure, giving rise to a sphere in whi

researcharxiv-cs-ai
1 Jul 2026
Research

Why Do Few-Step Text Latents Fail When Image Latents Work? Non-Commitment at Sharp Categorical Readouts

DGX agent

arXiv:2606.30705v1 Announce Type: cross Abstract: Deterministic few-step generation succeeds on continuous image latents but collapses to incoherent text on continuous text latents, and we show the ca

researcharxiv-cs-ai
1 Jul 2026
Research

WildProp: Visual Estimation of Wildlife Body Proportions at Scale

DGX agent

arXiv:2606.31125v1 Announce Type: new Abstract: Population-level morphometric measurements underpin ecological and evolutionary studies but traditionally require controlled imaging or physical specime

researcharxiv-cs-cv
1 Jul 2026
← Previous
1…135136137138139…466
Next →