AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
Human
88,356Total entries
1Added by human
88,355Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
2 Jun 2026

Design-MLLM: A Reinforcement Alignment Framework for Verifiable and Aesthetic Interior Design

Model ReleasesDGX agent

arXiv:2603.13312v2 Announce Type: replace-cross Abstract: Interior design is a requirements-to-visual-plan generation process that must simultaneously satisfy verifiable spatial feasibility and compar

Design Space Exploration of DMA based Finer-Grain Compute Communication Overlap

HardwareDGX agent

arXiv:2512.10236v2 Announce Type: replace-cross Abstract: Modern ML workloads demand distributing training and inference across multiple GPUs. However, these parallelization techniques often suffer fr

DeSQ: Decomposition-based SPARQL Query Generation

ResearchDGX agent

arXiv:2606.00203v1 Announce Type: new Abstract: Dominant approaches to Knowledge Base Question Answering (KBQA) fall into two categories. First is the generation of a formal query that suffers from br

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?

Model ReleasesDGX agent

arXiv:2505.16915v3 Announce Type: replace-cross Abstract: While recent Text-to-Image (T2I) models show impressive capabilities in synthesizing images from brief descriptions, they struggle with the lo

Detect Before You Leap: Mirage Detection in Vision-Language Models

SafetyDGX agent

arXiv:2606.00435v1 Announce Type: cross Abstract: Vision-language models (VLMs) can produce confident visual answers even when the required visual evidence is missing, blank, or unrelated to the quest

Detecting Pen-In-Air States from Video: A Proof-of-Concept Toward Complementary Handwriting Analysis

ResearchDGX agent

arXiv:2606.02342v1 Announce Type: new Abstract: Dynamic aspects of handwriting are critical for assessing developmental disorders such as dysgraphia and are typically captured using digitizing tablets

Detection vs. Execution: Single-Bucket Probes Miss Half the Mamba-2 State Sink

ResearchDGX agent

arXiv:2606.00930v1 Announce Type: cross Abstract: Mechanistic interpretability often assumes that probes identifying a representational signature also identify the circuit executing the corresponding

Detector-Evasive LLM Paraphrasing via Constrained Policy Optimization

SafetyDGX agent

arXiv:2606.00392v1 Announce Type: cross Abstract: AI-text detectors are vulnerable to paraphrasing and detector-guided paraphrasing attacks, but existing detector-evasion methods often lack precise co

Dexterity-BEV: Aligning 3D World and Actions for Generalizable Robot Policies Learning

SafetyDGX agent

arXiv:2606.02274v1 Announce Type: new Abstract: End-to-end manipulation policies, combined with web-scale pretrained Vision-Language Models (VLMs), show the promise for generalizable and dexterous rob

DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding

ResearchDGX agent

arXiv:2606.02091v1 Announce Type: new Abstract: Block diffusion speculative decoding accelerates LLM inference by predicting all tokens within a block simultaneously for the target model to verify in

Diagnosing LLM Arbitration Behavior over Pre-evidence Epistemic States in RAG-based Fact-Checking

Model ReleasesDGX agent

arXiv:2606.01120v1 Announce Type: new Abstract: In RAG-based fact-checking, LLMs are increasingly used as verifiers to check given claims against retrieved evidence. Their parametric knowledge can ind

Dialectics of Alignment: Harnessing Unsafe Knowledge for Dynamic Safety Routing

SafetyDGX agent

arXiv:2606.00686v1 Announce Type: new Abstract: The prevailing paradigm in large language model (LLM) alignment operates via erasure, filtering unsafe data or training models to strictly refuse harmfu

Diamonds in the Sky: Pareidolic Animals in Clouds

ResearchDGX agent

arXiv:2606.01361v1 Announce Type: new Abstract: People often see animal shapes in clouds, a phenomenon known as pareidolia. We propose an AI-based method that aims to predict which animals people are

DiffCrossGait: Trajectory-Level Alignment for 2D-3D Cross-Modal Gait Recognition via Latent Diffusion

SafetyDGX agent

arXiv:2606.00153v1 Announce Type: cross Abstract: Cross-modal 2D-3D gait recognition is impeded by inherent domain discrepancies between 2D silhouette and 3D LiDAR range-view representations. While pr

Differentially Private Datastore Generation for Retrieval-Augmented Inference

Model ReleasesDGX agent

arXiv:2606.01413v1 Announce Type: cross Abstract: It is crucial for modern on-device AI systems that rely on retrieval-augmented inference to release and share datastores without compromising individu

Differing Roles of Leisure and Productivity in GDP - A Machine Learning based comparative analysis of Germany and USA

ResearchDGX agent

arXiv:2606.01234v1 Announce Type: cross Abstract: The GDP of a country is modelled as the relative interaction between two agents - working hours, reflecting the social choice of a population, and Tot

DiffuSent: Towards a Unified Diffusion Framework for Aspect-Based Sentiment Analysis

ResearchDGX agent

arXiv:2606.01323v1 Announce Type: cross Abstract: Aspect-Based Sentiment Analysis (ABSA) encompasses seven distinct subtasks, each focusing on different extracted elements. Despite the proven success

Diffusion Image Generation with Explicit Modeling of Data Manifold Geometry

SafetyDGX agent

arXiv:2606.00094v1 Announce Type: cross Abstract: Image generative models aim to sample data points from the underlying data manifold, a task that requires learning and decoding a dense, low-dimension

Diffusion Models for Hyperspectral Image Analysis: A Comprehensive Review

ResearchDGX agent

arXiv:2505.11158v4 Announce Type: replace-cross Abstract: Hyperspectral image (HSI) analysis plays a critical role in remote sensing, agriculture, and environmental monitoring. However, traditional me

Digging Up Citations: FOSSIL, a Dataset and Workflow for Reference Extraction in Law and the Humanities

ResearchDGX agent

arXiv:2606.01109v1 Announce Type: cross Abstract: Citation extraction tools are designed for the structured end-of-document bibliographies of the natural sciences, but law and humanities scholarship c

Digital-to-Physical Transfer of Adversarial Patches for Aerial Vehicle Detection

ApplicationsDGX agent

arXiv:2606.00159v1 Announce Type: cross Abstract: Deep neural network (DNN)-based object detectors are widely used for analyzing aerial and satellite imagery in applications such as environmental moni

Digital Twin-Assisted Adaptive Multi-Agent DRL for Intelligent Spectrum and Resource Management in Open-RAN UAV-Enabled 6G Networks

AgentsDGX agent

arXiv:2606.01324v1 Announce Type: cross Abstract: The evolution toward 6G wireless networks envisions a seamlessly intelligent, Open-RAN-enabled architecture where unmanned aerial vehicles (UAVs) play

Dimension Reduction via Sum-of-Squares and Improved Clustering Algorithms for Non-Spherical Mixtures

ResearchDGX agent

arXiv:2411.12438v2 Announce Type: replace-cross Abstract: We develop a new approach for clustering non-spherical (i.e., arbitrary component covariances) Gaussian mixture models via a subroutine, based

DINO-GFSA: Geo-Localization via Semantic Gated Fusion and Mamba-based Sequential Aggregation

Model ReleasesDGX agent

arXiv:2606.00784v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL) is critical for Unmanned Aerial Vehicle (UAV) self-positioning and target localization in GNSS-denied environments. H

DIPOLE: Fusing Vision and Geometry for Robust Visuomotor Generalization

SafetyDGX agent

arXiv:2511.22445v2 Announce Type: replace Abstract: Imitation learning has emerged as a crucial approach for acquiring visuomotor skills from demonstrations, where designing effective observation enco

Directed Distance Fields for Constant-Time Ray Queries on Gaussian Splatting

ResearchDGX agent

arXiv:2606.00817v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) renders new views of a scene in real time. Like every rasterizer, it answers only primary rays, the rays from the camera

DiscourseFlip: An Oblique Discourse-Level Opinion Manipulation Attack against Black-box Retrieval-Augmented Generation

Local AiDGX agent

arXiv:2606.01212v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems are widely deployed and increasingly influential, but their reliance on external corpora exposes new secu

Discovering Nonlinear Static Relationships in Unlabeled Dataset using Autoencoder with Ordered Variance

ApplicationsDGX agent

arXiv:2402.14031v2 Announce Type: replace-cross Abstract: This paper presents an autoencoder with ordered variance (AEO), in which the conventional reconstruction loss is augmented by a variance-based

Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies

ResearchDGX agent

arXiv:2508.20072v4 Announce Type: replace Abstract: Vision-Language-Action (VLA) models adapt large vision-language backbones to map images and instructions into robot actions. However, prevailing VLA

Disentanglement-Based Equivariant Learning for Compositional VQA

Model ReleasesDGX agent

arXiv:2606.02168v1 Announce Type: new Abstract: Compositional visual question answering (VQA) represents a challenging yet fundamental task that requires models to comprehend novel combinations of pre

Disentangling Similarity and Relatedness in Topic Models

Model ReleasesDGX agent

arXiv:2603.10619v2 Announce Type: replace Abstract: The recent success of large pre-trained language models (PLMs) has motivated their integration into topic modeling. However, PLM-augmented topic mod

DisFlow: Scene Flow from Distance Field for Object Pose, Velocity Tracking, and Dynamic Object Reconstruction

ResearchDGX agent

arXiv:2606.01824v1 Announce Type: new Abstract: We present DisFlow, a novel framework for online scene flow estimation from distance field that enables 6DoF dynamic object pose estimation, motion trac

Distillation of Large Language Models via Concrete Score Matching

Model ReleasesDGX agent

arXiv:2509.25837v3 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver remarkable performance but are costly to deploy, motivating knowledge distillation (KD) for efficient inf

Distilling Neuro-Symbolic Programs into 3D Multi-modal LLMs

SafetyDGX agent

arXiv:2606.01215v1 Announce Type: cross Abstract: Current 3D spatial reasoning methods face a fundamental trade-off: neuro-symbolic 3D (NS3D) concept learners achieve interpretable reasoning through c

DistMatch: Adaptive Binning via Distribution Matching for Robust Sequential Conformal Prediction

Local AiDGX agent

arXiv:2606.00690v1 Announce Type: new Abstract: Sequential conformal prediction (CP) provides valid uncertainty quantification under the assumption of residual exchangeability. However, this assumptio

Distortion-Aware Fusion of Statistical and Vision-Language Features for Blind Image Quality Assessment

TutorialsDGX agent

arXiv:2606.02002v1 Announce Type: new Abstract: Blind image quality assessment (BIQA) aims to predict perceived image quality without access to a reference image. Classical natural scene statistics (N

Distributed GNEP Algorithms without Multiplier Sharing and Applications to Multi-Robot Coordination and Contextual Bandit-Based Active Learning

ApplicationsDGX agent

arXiv:2606.00759v1 Announce Type: new Abstract: Recent advances in artificial intelligence have expanded the focus from classical optimization to include equilibrium analysis in noncooperative games.

Distribution-free changepoint localization after sequential change detection

ResearchDGX agent

arXiv:2606.01256v1 Announce Type: cross Abstract: This paper introduces a distribution-free framework for constructing post-detection confidence sets for changepoints after stopping a sequential chang

Dive into Ambiguity: A*-Inspired Multi-Agents Commonsense Obfuscation Attack on LLM Prompts

SafetyDGX agent

arXiv:2606.01441v1 Announce Type: new Abstract: Large language models (LLMs) excel in reasoning and knowledge-intensive tasks but remain vulnerable to prompt-level adversarial attacks that preserve in

Dive into Waves: Morlet Spectral Transformer for Cross-Subject Emotion Decoding from EEG

ResearchDGX agent

arXiv:2606.00884v1 Announce Type: cross Abstract: We study cross-subject emotion recognition from EEG, a practically important yet challenging problem in brain-computer interfaces. Unlike tasks with c

Diversity Over Frequency: Rethinking Tool Use in Visual Chain-of-Thought Agents

AgentsDGX agent

arXiv:2606.00096v1 Announce Type: cross Abstract: Visual agents employ external visual tools within visual chains of thought to incorporate fine-grained evidence. While prior work has mainly studied t

Divide and Conquer: Reliable Multi-View Evidential Learning for Deepfake Detection

ResearchDGX agent

arXiv:2606.01885v1 Announce Type: new Abstract: With the evolution of generative models, deepfakes have achieved near-perfect semantic realism, leaving forensic traces only in subtle structural anomal

DLLM-JEPA: Joint Embedding Predictive Architectures for Masked Diffusion Language Models

Model ReleasesDGX agent

arXiv:2606.00091v1 Announce Type: cross Abstract: Joint Embedding Predictive Architectures (JEPAs) have reshaped self-supervised representation learning in vision. The recent LLM-JEPA ported JEPA to a

Do Gender Cues Affect LLM Value Trade-offs? Evidence from a Controlled Decision Benchmark

Model ReleasesDGX agent

arXiv:2606.02214v1 Announce Type: new Abstract: Large language models are increasingly used in value-sensitive decision settings, where irrelevant demographic cues should not alter judgments. We const

Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains

Model ReleasesDGX agent

arXiv:2606.02357v1 Announce Type: cross Abstract: Tool-augmented multimodal agents show strong benchmark gains, often taken as evidence that agents have learned to use tools. We argue that this interp

'Do Not Mention This to the User': Detecting and Understanding Malicious Agent Skills

AgentsDGX agent

arXiv:2602.06547v3 Announce Type: replace-cross Abstract: LLM-based coding agents increasingly rely on third-party extensions called skills, which bundle natural language instructions and helper scrip

Do Text Edits Generalize to Visual Generation? Benchmarking Cross-Modal Knowledge Editing in UMMs

Model ReleasesDGX agent

arXiv:2606.00477v1 Announce Type: new Abstract: Unified multimodal models (UMMs) have emerged as a promising paradigm for general-purpose multimodal intelligence. As they are deployed in real-world ap

Does Compression Preserve Uncertainty? A Unified Benchmark for Quantized and Sparse LLMs via Conformal Prediction

Model ReleasesDGX agent

arXiv:2606.01850v1 Announce Type: new Abstract: Model compression techniques such as quantization and pruning are widely used to reduce the deployment cost of large language models (LLMs), with existi

Doing well with less! On Sampling Techniques for Empirical Pairwise Loss Estimation/Minimization

ResearchDGX agent

arXiv:2606.02345v1 Announce Type: cross Abstract: Many machine learning problems, including similarity learning, ranking, and clustering, rely on empirical pairwise loss functions whose quadratic comp

Doing What They Say, Not What They Reason: Locating the Faithfulness Gap in LLM Agents

ResearchDGX agent

arXiv:2606.00476v1 Announce Type: new Abstract: Do LLM agents act on the reasoning they state? This question of process fidelity is central to using LLMs in social simulation, yet it is hard to measur

Domain Adaptation with a Single Vision-Language Embedding

AgentsDGX agent

arXiv:2410.21361v2 Announce Type: replace Abstract: Domain adaptation has been extensively investigated in computer vision but still requires access to target data at the training time, which might be

Domain-Shift-Aware Conformal Prediction for Large Language Models

Model ReleasesDGX agent

arXiv:2510.05566v2 Announce Type: replace-cross Abstract: Large language models have achieved impressive performance across diverse tasks. However, their tendency to produce overconfident and factuall

Don't Ask the LLM to Track Freshness: A Deterministic Recipe for Memory Conflict Resolution

AgentsDGX agent

arXiv:2606.01435v1 Announce Type: new Abstract: LLM-based memory systems increasingly maintain facts that evolve over time, where a recurring failure is conflict resolution: when a fact has multiple c

Don't Let a Few Network Failures Slow the Entire AllReduce

HardwareDGX agent

arXiv:2606.01680v1 Announce Type: cross Abstract: Network failures are among the most frequent hardware faults in large-scale GPU clusters and a leading cause of training-job interruptions. Modern col

Don't Read Everything: A Curvature-Conditioned Query for Linear Attention

Local AiDGX agent

arXiv:2606.01294v1 Announce Type: new Abstract: Linear attention reduces the quadratic cost of softmax attention by maintaining a recurrent fast-weight state, but it consistently lags on in-context re

DOT-MoE: Differentiable Optimal Transport for MoEfication

SafetyDGX agent

arXiv:2606.01666v1 Announce Type: cross Abstract: The scaling of Large Language Models (LLMs) has driven significant performance gains but created substantial challenges in inference efficiency. While

DPsurv: Dual-Prototype Evidential Fusion for Uncertainty-Aware and Interpretable Whole-Slide Image Survival Prediction

ResearchDGX agent

arXiv:2510.00053v2 Announce Type: replace-cross Abstract: Pathology whole-slide images (WSIs) are widely used for cancer survival analysis because of their comprehensive histopathological information

Dr. DocBench: A Comprehensive Benchmark for Expert-Level and Difficult Document Parsing

Model ReleasesDGX agent

arXiv:2606.01393v1 Announce Type: cross Abstract: Document parsing and recognition are fundamental capabilities for vision-language models (VLMs) and document processing systems. However, existing Opt

DraDDP: A Multimodal Multi-Party Dialogue Discourse Parsing Dataset

ResearchDGX agent

arXiv:2606.00012v1 Announce Type: cross Abstract: Multi-party dialogue discourse parsing aims to identify dependency structures and relation types between utterances in conversations. Previous studies

DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation

ResearchDGX agent

arXiv:2606.00535v1 Announce Type: new Abstract: Speculative decoding (SD) has proven to be an effective technique for accelerating autoregressive generation in large language models (LLMs) however, it

← Previous
1…511512513514515…1049
Next →