AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
Human
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
7 Jul 2026

Defending from GeoLocalization through Adversarial Road Trips

SafetyDGX agent

arXiv:2607.03277v1 Announce Type: new Abstract: Retrieval-based image geolocalization has emerged as a powerful technique for determining the location of a query image by matching it against a large,

Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models

Model ReleasesDGX agent

arXiv:2607.05390v1 Announce Type: cross Abstract: Predicting object dynamics (i.e., world modeling) is a fundamental challenge for robotic manipulation, and modeling deformable objects presents a part

DeGenseGS: Geometrically and Semantically Decoupled Surgical Scene Understanding in 4D Gaussian Splatting

SafetyDGX agent

arXiv:2607.04761v1 Announce Type: new Abstract: Real-time, text-promptable 4D reconstruction is indispensable for autonomous surgical interaction. Severe misalignment between semantic meaning and phys

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DELTA-TTS: Adapting Autoregressive Model into Diffusion Language Model for Text-to-Speech

Local AiDGX agent

arXiv:2607.04140v1 Announce Type: cross Abstract: Autoregressive (AR) text-to-speech (TTS) models generate discrete speech tokens sequentially, which makes inference slow and can degrade robustness by

DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences

Local AiDGX agent

arXiv:2607.02551v1 Announce Type: cross Abstract: Video multimodal large language models have made strong progress on open-ended video understanding, but they still lack precise local spatiotemporal p

Demonstrating Generalization Failures via Mixtures of Conditional Policies

SafetyDGX agent

arXiv:2607.03478v1 Announce Type: new Abstract: Post-training of frontier language models is conducted on curated task suites, and inevitably leaves a distribution shift between training and deploymen

Denoised Conformal Alignment for Reliable Selection of Conditional Average Treatment Effect Predictions

SafetyDGX agent

arXiv:2607.03161v1 Announce Type: cross Abstract: In selective deployment, practitioners act only on a model-chosen subset of individuals based on predicted conditional average treatment effects, but

Derivations of Error-State Kalman Filter Kinematics for Globally Applicable Aided Inertial Navigation Systems

ResearchDGX agent

arXiv:2607.03211v1 Announce Type: new Abstract: Global navigation systems require state estimation algorithms that handle Earth's curvature, Earth's rotation, and gravitational variations. These facto

Deriving Benchmarking Datasets from Long-Form Recordings: Challenges and Opportunities

ApplicationsDGX agent

arXiv:2607.03201v1 Announce Type: cross Abstract: Long-form recordings (LFRs) of child-centered audio are ecologically valid sources for studying early language development, but three problems limit t

Deriving Neural Scaling Laws from the statistics of natural language

Model ReleasesDGX agent

arXiv:2602.07488v3 Announce Type: replace-cross Abstract: Despite the fact that experimental neural scaling laws have substantially guided empirical progress in large-scale machine learning, no existi

Designing Touch for Trauma-Informed Social Robots: A Design Space for Direct and Indirect Actuation

SafetyDGX agent

arXiv:2607.04981v1 Announce Type: new Abstract: Touch is a fundamental communication modality in human-robot interaction and may support grounding, emotional regulation, and stress reduction in therap

DETECT-3B-Omni is Agnostic of Content and Demographics

ResearchDGX agent

arXiv:2607.03418v1 Announce Type: cross Abstract: A trustworthy and GDPR-compliant deepfake audio detector must base its decisions on acoustic artifacts, not on what is being said or who is speaking.

Detecting Answer-Driven Reasoning in LLM-Based Educational Tutors via Truncated Chain-of-Thought Auditing

TutorialsDGX agent

arXiv:2607.04572v1 Announce Type: new Abstract: Large language model (LLM) tutors often produce fluent step-by-step explanations, but a correct and pedagogically formatted response does not guarantee

Detecting Architectural Drift in Safety-Critical Firmware through Runtime Trace Analysis

SafetyDGX agent

arXiv:2607.03135v1 Announce Type: cross Abstract: Maintaining consistency between architectural design and runtime-observed behavior is challenging in long-lived safety-critical firmware. This paper p

Detecting Hallucinations in Retrieval-Augmented Generation through Grounding-Aware Sensitivity by Perturbation (GASP)

ResearchDGX agent

arXiv:2607.04223v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) reduces but does not eliminate hallucination, and existing detectors return a single answer-level score that does

Determinants and Limits of LLM Security-Tool Orchestration: A Study with HexStrike-AI

Model ReleasesDGX agent

arXiv:2607.02873v1 Announce Type: cross Abstract: Large language model agents driving security tool suites over the Model Context Protocol are increasingly common. Yet the factors that bound their cap

Deterministic Envelopes for Tamed SGLD: Decoupling Stochastic Gradient Noise and Localizing Taming

SafetyDGX agent

arXiv:2606.05242v2 Announce Type: replace-cross Abstract: Stochastic gradient Langevin algorithms often use tamed denominators to stabilize superlinear drifts. This paper shows that when the denominat

Developing an LLM-Based Feedback System Grounded in Evidence-Centered Design to Support Physics Problem Solving

ResearchDGX agent

arXiv:2512.10785v3 Announce Type: replace-cross Abstract: Generative AI offers new opportunities for individualized and adaptive learning, e.g., through large language model (LLM)-based feedback syste

DGSeg: Dynamic Gating of Semantic-Spatial Guided Predictions for Reasoning Segmentation

TutorialsDGX agent

arXiv:2607.04779v1 Announce Type: new Abstract: Reasoning segmentation aims to predict pixel-wise masks for targets given complex language queries. Existing approaches leverage Multimodal Large Langua

Diagnosing Aerial-View Object Detectors with Foundational Image Generative Models

TutorialsDGX agent

arXiv:2607.02718v1 Announce Type: cross Abstract: Recent advances in large-scale image generative models enable photorealistic scene synthesis with controllable attributes. Beyond data augmentation, t

DiCE-CIR: Direct Composition Learning for Efficient Zero-Shot Composed Image Retrieval

SafetyDGX agent

arXiv:2607.04665v1 Announce Type: new Abstract: Zero-shot composed image retrieval (ZS-CIR) aims to retrieve a target image from a multimodal query consisting of a reference image and an edit text des

DICT: Data Injection and Contrastive Trajectory Refinement for Conditional Image Generation with Diffusion Models

SafetyDGX agent

arXiv:2607.03899v1 Announce Type: new Abstract: Diffusion models have become a dominant paradigm for conditional image generation, yet existing approaches generally follow two directions: task-specifi

Dictionaries, Not Darwin: Set-Level Selection Beats LLM Evolution in Scientific Equation Discovery

Model ReleasesDGX agent

arXiv:2607.04108v1 Announce Type: new Abstract: Large language models are increasingly used as evolutionary engines for scientific discovery: generate candidates, select winners, feed them back as par

Differential Amplifier-Inspired AmpAttention for Multi-View Robotic Manipulation

ApplicationsDGX agent

arXiv:2607.02845v1 Announce Type: cross Abstract: Multi-view robotic manipulation methods with the attention mechanism have recently achieved significant progress in both training efficiency and task

Differentiate the Evaluator, Not the Program: An Efficient Runtime Representation for Neuro-Symbolic Learning

Model ReleasesDGX agent

arXiv:2607.03574v1 Announce Type: cross Abstract: AI systems increasingly propose executable scientific models whose value depends on both their symbolic structure and their fitted continuous paramete

Diffusion-Guided Uncertainty-Aware Delayed Policy Optimization

SafetyDGX agent

arXiv:2607.05064v1 Announce Type: new Abstract: Reinforcement learning in real world environments often suffers from severe performance degradation due to delayed feedback. Existing approaches typical

Diffusion learning reveals viable parameter manifolds and compensation geometry in biological dynamical systems

Model ReleasesDGX agent

arXiv:2607.03671v1 Announce Type: cross Abstract: Models of complex systems often have many parameters, yet are constrained by far fewer experimentally accessible observables: similar activity can eme

Diffusion Models are Open-World Affordance Learners: Leveraging Generative Priors for 3D Affordance Learning

Model ReleasesDGX agent

arXiv:2508.01651v2 Announce Type: replace Abstract: 3D affordance grounding aims to understand how diverse objects can be manipulated, making it a cornerstone of embodied interaction. However, prior w

Dimension Reduction for Curves: Simplified and Generalized

ResearchDGX agent

arXiv:2607.03112v1 Announce Type: cross Abstract: We revisit random projections for reducing the dimension of high-dimensional polygonal curves. Drawing from the toolbox of randomized linear algebra,

DIRA-SS:Dynamic Domain Incremental Regularised Adaptation -- Self-Supervised

AgentsDGX agent

arXiv:2311.07461v3 Announce Type: replace Abstract: Autonomous systems (AS) often rely on Deep Neural Network (DNN) classifiers to operate in complex and dynamically changing environments. However, du

Direct Time-of-Flight Measurement Accuracy Improvement With Perimeter-Gated SPADs

ResearchDGX agent

arXiv:2607.02546v1 Announce Type: cross Abstract: Direct time of flight (dToF) measurements are susceptible to errors because of system-level and circuit-level timing jitters. In addition, device-leve

Directional Curvature from Armijo Backtracking: A Low-Cost Sharpness Probe and a Calibration-Free Learning-Rate Safeguard for Adam

SafetyDGX agent

arXiv:2607.03998v1 Announce Type: new Abstract: The local sharpness of the loss, the top Hessian eigenvalue lambda_1, determines the largest stable gradient step, but measuring it normally requires La

Discrete distributions are learnable from metastable samples

Model ReleasesDGX agent

arXiv:2410.13800v4 Announce Type: replace-cross Abstract: Physically motivated stochastic dynamics are widely used to sample from high-dimensional distributions. However, such samplers often get trapp

Displacement Preserving Relational Distillation for Robust Medical Segmentation

SafetyDGX agent

arXiv:2607.04599v1 Announce Type: new Abstract: Accurate 3D medical segmentation is limited by anatomical variability and high computational costs. While knowledge distillation (KD) offers a route for

Distill Where the Student Goes: Teacher-Regularized RL for English-Evidence Cross-Lingual RAG

SafetyDGX agent

arXiv:2607.02966v1 Announce Type: new Abstract: Cross-lingual retrieval-augmented generation (RAG) is often deployed in an English-evidence regime, where users query in diverse languages but retrieved

DistillH-Mamba: A Hypergraph-Mamba-Based Knowledge Distillation Model for Efficient Impact Fall Detection

ResearchDGX agent

arXiv:2607.03156v1 Announce Type: new Abstract: Falls among the elderly represent a significant public health concern due to their prevalence, consequences, and societal burden. While deep learning ha

Distribution-free Deviation Bounds and The Role of Domain Knowledge in Learning via Model Selection with Cross-validation Risk Estimation

SafetyDGX agent

arXiv:2303.08777v3 Announce Type: replace-cross Abstract: Cross-validation is one of the most widely used tools for risk estimation and model selection in statistics and machine learning, yet its theo

Distribution Matching Distillation Meets Reinforcement Learning

ResearchDGX agent

arXiv:2511.13649v5 Announce Type: replace Abstract: Distribution Matching Distillation (DMD) facilitates efficient inference by distilling multi-step diffusion models into few-step variants. Concurren

Diverse Normal Prototypes-Guided Contrastive Reconstruction for Medical Anomaly Detection

SafetyDGX agent

arXiv:2508.19573v2 Announce Type: replace Abstract: Anomaly detection in medical images is challenging due to limited annotations and the domain gap. Existing reconstruction-based methods often rely o

Diversity-aware View Partitioning for Scalable VGGT

ResearchDGX agent

arXiv:2607.01885v2 Announce Type: replace Abstract: Geometry transformers such as VGGT achieve strong performance by jointly reasoning over multiple views with global attention. However, scaling them

DIVO: Continuous-time DVL-Inertial-Visual Odometry for Unmanned Underwater Vehicles

ApplicationsDGX agent

arXiv:2607.04615v1 Announce Type: new Abstract: This paper presents a novel acoustic-visual-inertial odometry solution leveraging a continuous-time trajectory estimation framework for unmanned underwa

Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval

ResearchDGX agent

arXiv:2607.04605v1 Announce Type: cross Abstract: Multi-vector vision-language retrieval preserves fine-grained visual evidence through maximum-similarity late interaction, but dense image-side tokens

Do Diabetic Foot Ulcer Segmentation Models Generalize? A Cross-Dataset Benchmark of CNN and Transformer Architectures

Model ReleasesDGX agent

arXiv:2607.02555v1 Announce Type: new Abstract: Deep learning models for diabetic foot ulcer (DFU) segmentation routinely report high accuracy, but they are almost always trained and tested on the sam

Do ECG Foundation Models Transfer to Rare Cardiac Diseases? Evidence from Brugada Syndrome Detection

SafetyDGX agent

arXiv:2607.03009v1 Announce Type: new Abstract: Background: Foundation models (FMs) trained on large-scale unlabeled physiological data have emerged as a promising paradigm for medical artificial inte

Do Flat Minima Improve Sparse Novel View Synthesis?

ResearchDGX agent

arXiv:2511.17918v2 Announce Type: replace Abstract: Despite the success of recent novel view synthesis methods, they tend to struggle in sparse-view settings. This poor generalization to unseen viewpo

Do GUI Agents Believe Their Eyes? Diagnosing State-Belief Reliance on Pixels versus Structure

AgentsDGX agent

arXiv:2607.04334v1 Announce Type: new Abstract: Multimodal GUI agents read an interface through two redundant channels: the rendered pixels of a screenshot and a serialized structure such as a DOM or

Do Medical Vision Language Models Actually See? A Counterfactual Grounding Framework and Hard-Negative Contrastive Training for Visually-Reliant Medical VLMs

Model ReleasesDGX agent

arXiv:2607.03647v1 Announce Type: new Abstract: Large vision language models (VLMs) report strong accuracy on medical question-answering, yet it remains unclear whether they reason from visual evidenc

Do Vision-Language-Action Models Mean What They Say? On the Role of Faithfulness in Embodied Reasoning

SafetyDGX agent

arXiv:2607.04681v1 Announce Type: cross Abstract: Embodied Chain-of-Thought has emerged as a promising mechanism to enhance robot decision-making and interpretability in black-box Vision-Language Acti

Does It Fail to See or Fail to Know? Attributing Errors in Vision-Language Models

ResearchDGX agent

arXiv:2607.04683v1 Announce Type: cross Abstract: Vision-language models (VLMs) perform well on visual question answering with high-quality images but struggle when questions require knowledge beyond

Domain Knowledge-Informed Self-Supervised Representations for Workout Form Assessment

TutorialsDGX agent

arXiv:2202.14019v3 Announce Type: replace-cross Abstract: Maintaining proper form while exercising is important for preventing injuries and maximizing muscle mass gains. Detecting errors in workout fo

Don't Blame the Large Language Model: How Scaffolding Evolution Shapes Coding Agent Quality

Model ReleasesDGX agent

arXiv:2607.03691v1 Announce Type: cross Abstract: Coding agents, autonomous systems that use large language models (LLMs) to resolve software engineering tasks, rely on agentic scaffolding: a middlewa

Don't Commit Alone: Joint Token Commitment in Diffusion Large Language Models

ResearchDGX agent

arXiv:2607.04469v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) commit multiple tokens per denoising step by decoding each selected position independently from the shared conte

Don't Wait to Reply: Towards Responsive yet Thoughtful Dialogue through Proactive Thinking

ResearchDGX agent

arXiv:2607.03093v1 Announce Type: cross Abstract: Thinking has emerged as a critical capability for Large Language Models (LLMs) tackling complex tasks. However, its reactive nature, where reasoning i

dOPSD: On-Policy Self-Distillation for Diffusion Language Models

SafetyDGX agent

arXiv:2607.04428v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising a masked sequence, offering a parallel alternative to autoregressive mo

DOSE-I: A Multimodal Biosignal Dataset of Procedural Sedation for Endoscopy -- Technical Report

ResearchDGX agent

arXiv:2607.02570v1 Announce Type: cross Abstract: In this document, we describe characteristics and technical details of the multimodal biosignal dataset DOSE-I of procedural sedation for endoscopy pu

Double Fuzzy Probabilistic Interval Linguistic Term Set and a Dynamic Fuzzy Decision Making Model based on Markov Process with tts Application in Multiple Criteria Group Decision Making

ResearchDGX agent

arXiv:2111.15255v2 Announce Type: replace-cross Abstract: The probabilistic linguistic term has been proposed to deal with probability distributions in provided linguistic evaluations. However, becaus

Double-Helix Active Geometry: LiDAR-Anchored Multi-View Depth with Selective Abstention

SafetyDGX agent

arXiv:2607.02561v1 Announce Type: cross Abstract: Consumer depth sensors such as the LiDAR scanner on recent iPhones provide metric range, but their useful range is short and their returns are sparse.

DRBA: Dynamic Robotic Balance Assistant -- An assist-as-needed gait and balance rehabilitation robot for versatile training

ResearchDGX agent

arXiv:2607.03027v1 Announce Type: new Abstract: The decline of human balance control due to aging and pathological conditions increases fall risk, a major concern in geriatric care and rehabilitation.

DREAMSTEER: Latent World Models Can Steer VLA Policies During Deployment Without Any Finetuning

Model ReleasesDGX agent

arXiv:2607.02865v1 Announce Type: new Abstract: Pretrained vision-language-action (VLA) policies show promising zero-shot generalization, but often fail under deployment-time distribution shift, leadi

DriftST: One-Step Generative Inference of Spatial Transcriptomics from H&E Histology

ResearchDGX agent

arXiv:2607.04740v1 Announce Type: new Abstract: Spatial Transcriptomics (ST) measures gene expression while preserving spatial context, but its high cost and low throughput leave public datasets small

← Previous
1…257258259260261…1025
Next →