AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
Human
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
58,761 results
10 Apr 2026

DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation

ResearchDGX agent

arXiv:2511.19365v2 Announce Type: replace-cross Abstract: Pixel diffusion aims to generate images directly in pixel space in an end-to-end fashion. This approach avoids the limitations of VAE in the t

Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs

SafetyDGX agent

arXiv:2604.07518v1 Announce Type: new Abstract: Vision-Language Models often struggle with complex visual reasoning due to the visual information loss in textual CoT. Existing methods either add the c

Deep Learning-Powered Visual SLAM Aimed at Assisting Visually Impaired Navigation

SafetyDGX agent

arXiv:2510.20549v2 Announce Type: replace Abstract: Despite advancements in SLAM technologies, robust operation under challenging conditions such as low-texture, motion-blur, or challenging lighting r

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models

SafetyDGX agent

arXiv:2604.08527v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models under their own induced distribution while leveraging supervision from stronger teachers. We identify

Density-Driven Optimal Control: Convergence Guarantees for Stochastic LTI Multi-Agent Systems

AgentsDGX agent

arXiv:2604.08495v1 Announce Type: cross Abstract: This paper addresses the decentralized non-uniform area coverage problem for multi-agent systems, a critical task in missions with high spatial priori

Depression Detection at the Point of Care: Automated Analysis of Linguistic Signals from Routine Primary Care Encounters

ResearchDGX agent

arXiv:2604.06193v1 Announce Type: cross Abstract: Depression is underdiagnosed in primary care, yet timely identification remains critical. Recorded clinical encounters, increasingly common with digit

Designing Safe and Accountable GenAI as a Learning Companion with Women Banned from Formal Education

SafetyDGX agent

arXiv:2604.07253v1 Announce Type: cross Abstract: In gender-restrictive and surveilled contexts, where access to formal education may be restricted for women, pursuing education involves safety and pr

Detecting HIV-Related Stigma in Clinical Narratives Using Large Language Models

Model ReleasesDGX agent

arXiv:2604.07717v1 Announce Type: new Abstract: Human immunodeficiency virus (HIV)-related stigma is a critical psychosocial determinant of health for people living with HIV (PLWH), influencing mental

Development of ML model for triboelectric nanogenerator based sign language detection system

ResearchDGX agent

arXiv:2604.06220v1 Announce Type: cross Abstract: Sign language recognition (SLR) is vital for bridging communication gaps between deaf and hearing communities. Vision-based approaches suffer from occ

DHFP-PE: Dual-Precision Hybrid Floating Point Processing Element for AI Acceleration

ResearchDGX agent

arXiv:2604.04507v2 Announce Type: replace-cross Abstract: The rapid adoption of low-precision arithmetic in artificial intelligence and edge computing has created a strong demand for energy-efficient

Diagnosing and Mitigating Sycophancy and Skepticism in LLM Causal Judgment

Model ReleasesDGX agent

arXiv:2601.08258v3 Announce Type: replace Abstract: Large language models increasingly fail in a way that scalar accuracy cannot diagnose: they produce a sound reasoning trace and then abandon it unde

DietDelta: A Vision-Language Approach for Dietary Assessment via Before-and-After Images

ResearchDGX agent

arXiv:2604.06352v1 Announce Type: cross Abstract: Accurate dietary assessment is critical for precision nutrition, yet most image-based methods rely on a single pre-consumption image and provide only

Differentially Private Best-Arm Identification

Local AiDGX agent

arXiv:2406.06408v2 Announce Type: replace-cross Abstract: Best Arm Identification (BAI) problems are progressively used for data-sensitive applications, such as designing adaptive clinical trials, tun

Differentially Private Language Generation and Identification in the Limit

ResearchDGX agent

arXiv:2604.08504v1 Announce Type: cross Abstract: We initiate the study of language generation in the limit, a model recently introduced by Kleinberg and Mullainathan [KM24], under the constraint of d

DiffSketcher: Text Guided Vector Sketch Synthesis through Latent Diffusion Models

TutorialsDGX agent

arXiv:2306.14685v5 Announce Type: replace-cross Abstract: We demonstrate that pre-trained text-to-image diffusion models, despite being trained on raster images, possess a remarkable capacity to guide

Diffusion Language Models Know the Answer Before Decoding

ResearchDGX agent

arXiv:2508.19982v5 Announce Type: replace Abstract: Diffusion language models (DLMs) have recently emerged as an alternative to autoregressive approaches, offering parallel sequence generation and fle

Diffusion Processes on Implicit Manifolds

TutorialsDGX agent

arXiv:2604.07213v1 Announce Type: new Abstract: High-dimensional data are often modeled as lying near a low-dimensional manifold. We study how to construct diffusion processes on this data manifold in

DiffVC: A Non-autoregressive Framework Based on Diffusion Model for Video Captioning

ResearchDGX agent

arXiv:2604.08084v1 Announce Type: new Abstract: Current video captioning methods usually use an encoder-decoder structure to generate text autoregressively. However, autoregressive methods have inhere

Digital Skin, Digital Bias: Uncovering Tone-Based Biases in LLMs and Emoji Embeddings

Model ReleasesDGX agent

arXiv:2604.06863v1 Announce Type: cross Abstract: Skin-toned emojis are crucial for fostering personal identity and social inclusion in online communication. As AI models, particularly Large Language

DINO-QPM: Adapting Visual Foundation Models for Globally Interpretable Image Classification

ResearchDGX agent

arXiv:2604.07166v1 Announce Type: cross Abstract: Although visual foundation models like DINOv2 provide state-of-the-art performance as feature extractors, their complex, high-dimensional representati

DinoRADE: Full Spectral Radar-Camera Fusion with Vision Foundation Model Features for Multi-class Object Detection in Adverse Weather

AgentsDGX agent

arXiv:2604.08074v1 Announce Type: new Abstract: Reliable and weather-robust perception systems are essential for safe autonomous driving and typically employ multi-modal sensor configurations to achie

Direct Segmentation without Logits Optimization for Training-Free Open-Vocabulary Semantic Segmentation

Model ReleasesDGX agent

arXiv:2604.07723v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation (OVSS) aims to segment arbitrary category regions in images using open-vocabulary prompts, necessitating that exis

DisCEdge: Distributed Context Management for Large Language Models at the Edge

ResearchDGX agent

arXiv:2511.22599v2 Announce Type: replace-cross Abstract: Deploying Large Language Model (LLM) services at the edge benefits latency-sensitive and privacy-aware applications. However, the stateless na

Discrete Flow Matching Policy Optimization

SafetyDGX agent

arXiv:2604.06491v1 Announce Type: cross Abstract: We introduce Discrete flow Matching policy Optimization (DoMinO), a unified framework for Reinforcement Learning (RL) fine-tuning Discrete Flow Matchi

DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs

Model ReleasesDGX agent

arXiv:2604.06250v1 Announce Type: cross Abstract: When asked to describe a molecular diagram, a Vision-Language Model correctly identifies ``a benzene ring with an -OH group.'' When asked to reason ab

Distilling Specialized Orders for Visual Generation

ResearchDGX agent

arXiv:2504.17069v2 Announce Type: replace Abstract: Autoregressive (AR) image generators are becoming increasingly popular due to their ability to produce high-quality images and their scalability. Ty

Distributed Interpretability and Control for Large Language Models

Model ReleasesDGX agent

arXiv:2604.06483v1 Announce Type: cross Abstract: Large language models that require multiple GPU cards to host are usually the most capable models. It is necessary to understand and steer these model

Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models

Model ReleasesDGX agent

arXiv:2604.08284v1 Announce Type: new Abstract: Large language models store not only isolated facts but also rules that support reasoning across symbolic expressions, natural language explanations, an

Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook

SafetyDGX agent

arXiv:2604.06210v2 Announce Type: cross Abstract: As LLMs are globally deployed, aligning their cultural value orientations is critical for safety and user engagement. However, existing benchmarks fac

DIVERSED: Relaxed Speculative Decoding via Dynamic Ensemble Verification

ResearchDGX agent

arXiv:2604.07622v1 Announce Type: new Abstract: Speculative decoding is an effective technique for accelerating large language model inference by drafting multiple tokens in parallel. In practice, its

DMin: Scalable Training Data Influence Estimation for Diffusion Models

ResearchDGX agent

arXiv:2412.08637v4 Announce Type: replace Abstract: Identifying the training data samples that most influence a generated image is a critical task in understanding diffusion models (DMs), yet existing

Do MLLMs Really Understand Space? A Mathematical Reasoning Evaluation

Model ReleasesDGX agent

arXiv:2602.11635v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have achieved strong performance on perception-oriented tasks, yet their ability to perform mathematical sp

Do We Need Distinct Representations for Every Speech Token? Unveiling and Exploiting Redundancy in Large Speech Language Models

ResearchDGX agent

arXiv:2604.06871v1 Announce Type: cross Abstract: Large Speech Language Models (LSLMs) typically operate at high token rates (tokens/s) to ensure acoustic fidelity, yet this results in sequence length

Domain-Contextualized Inference: A Computable Graph Architecture for Explicit-Domain Reasoning

Model ReleasesDGX agent

arXiv:2604.04344v2 Announce Type: replace Abstract: We establish a computation-substrate-agnostic inference architecture in which domain is an explicit first-class computational parameter. This produc

'Don't Be Afraid, Just Learn': Insights from Industry Practitioners to Prepare Software Engineers in the Age of Generative AI

TutorialsDGX agent

arXiv:2604.06342v1 Announce Type: cross Abstract: Although tension between university curricula and industry expectations has existed in some form for decades, the rapid integration of generative AI (

'Don't Do That!': Guiding Embodied Systems through Large Language Model-based Constraint Generation

ResearchDGX agent

arXiv:2506.04500v3 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have spurred interest in robotic navigation that incorporates complex spatial, mathematica

Don't Label Twice: Quantity Beats Quality when Comparing Binary Classifiers on a Budget

TutorialsDGX agent

arXiv:2402.02249v3 Announce Type: replace Abstract: We study how to best spend a budget of noisy labels to compare the accuracy of two binary classifiers. It's common practice to collect and aggregate

Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agents

Model ReleasesDGX agent

arXiv:2604.08369v1 Announce Type: cross Abstract: Inference-time compute scaling has emerged as a powerful technique for improving the reliability of large language model (LLM) agents, but existing me

DosimeTron: Automating Personalized Monte Carlo Radiation Dosimetry in PET/CT with Agentic AI

Model ReleasesDGX agent

arXiv:2604.06280v1 Announce Type: cross Abstract: Purpose: To develop and evaluate DosimeTron, an agentic AI system for automated patient-specific MC internal radiation dosimetry in PET/CT examination

Double: Breaking the Acceleration Limit via Double Retrieval Speculative Parallelism

ResearchDGX agent

arXiv:2601.05524v2 Announce Type: replace Abstract: Parallel Speculative Decoding (PSD) accelerates traditional Speculative Decoding (SD) by overlapping draft generation with verification. However, it

DP-DeGauss: Dynamic Probabilistic Gaussian Decomposition for Egocentric 4D Scene Reconstruction

ResearchDGX agent

arXiv:2604.07986v1 Announce Type: new Abstract: Egocentric video is crucial for next-generation 4D scene reconstruction, with applications in AR/VR and embodied AI. However, reconstructing dynamic fir

DQA: Diagnostic Question Answering for IT Support

ApplicationsDGX agent

arXiv:2604.05350v2 Announce Type: replace Abstract: Enterprise IT support interactions are fundamentally diagnostic: effective resolution requires iterative evidence gathering from ambiguous user repo

Draw-In-Mind: Rebalancing Designer-Painter Roles in Unified Multimodal Models Benefits Image Editing

Model ReleasesDGX agent

arXiv:2509.01986v4 Announce Type: replace-cross Abstract: In recent years, integrating multimodal understanding and generation into a single unified model has emerged as a promising paradigm. While th

Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control

SafetyDGX agent

arXiv:2604.03540v2 Announce Type: replace Abstract: Although multi-step generative policies achieve strong performance in robotic manipulation by modeling multimodal action distributions, they require

Drifting Fields are not Conservative

ResearchDGX agent

arXiv:2604.06333v1 Announce Type: new Abstract: Drifting models generate high-quality samples in a single forward pass by transporting generated samples toward the data distribution using a vector val

DROP: Distributional and Regular Optimism and Pessimism for Reinforcement Learning

Model ReleasesDGX agent

arXiv:2410.17473v2 Announce Type: replace Abstract: In reinforcement learning (RL), temporal difference (TD) error is known to be related to the firing rate of dopamine neurons. It has been observed t

DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM Editing

Model ReleasesDGX agent

arXiv:2604.07965v1 Announce Type: new Abstract: Model editing aims to update knowledge to add new concepts and change relevant information without retraining. Lifelong editing is a challenging task, p

Dual-level Modality Debiasing Learning for Unsupervised Visible-Infrared Person Re-Identification

Model ReleasesDGX agent

arXiv:2512.03745v2 Announce Type: replace Abstract: Two-stage learning pipeline has achieved promising results in unsupervised visible-infrared person re-identification (USL-VI-ReID). It first perform

Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving

Model ReleasesDGX agent

arXiv:2604.08075v1 Announce Type: new Abstract: Production vLLM fleets typically provision each instance for the worst-case context length, leading to substantial KV-cache over-allocation and under-ut

DYCP: Dynamic Context Pruning for Long-Form Dialogue with LLMs

ResearchDGX agent

arXiv:2601.07994v4 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly operate over long-form dialogues with frequent topic shifts. While recent LLMs support extended context wi

Dynamic Context Evolution for Scalable Synthetic Data Generation

Model ReleasesDGX agent

arXiv:2604.07147v1 Announce Type: cross Abstract: Large language models produce repetitive output when prompted independently across many batches, a phenomenon we term cross-batch mode collapse: the p

DynLP: Parallel Dynamic Batch Update for Label Propagation in Semi-Supervised Learning

HardwareDGX agent

arXiv:2604.06596v1 Announce Type: cross Abstract: Semi-supervised learning aims to infer class labels using only a small fraction of labeled data. In graph-based semi-supervised learning, this is typi

E-3DPSM: A State Machine for Event-Based Egocentric 3D Human Pose Estimation

ResearchDGX agent

arXiv:2604.08543v1 Announce Type: new Abstract: Event cameras offer multiple advantages in monocular egocentric 3D human pose estimation from head-mounted devices, such as millisecond temporal resolut

E2Edev: Benchmarking Large Language Models in End-to-End Software Development Task

Model ReleasesDGX agent

arXiv:2510.14509v3 Announce Type: replace-cross Abstract: The rapid advancement in large language models (LLMs) has demonstrated significant potential in End-to-End Software Development (E2ESD). Howev

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation

SafetyDGX agent

arXiv:2602.13669v4 Announce Type: replace Abstract: Recent multi-modal video generation models have achieved high visual quality, but their prohibitive latency and limited temporal stability hinder re

ECLipsE-Gen-Local: Efficient Compositional Local Lipschitz Estimates for Deep Neural Networks

ResearchDGX agent

arXiv:2510.05261v2 Announce Type: replace Abstract: The Lipschitz constant is a key measure for certifying the robustness of neural networks to input perturbations. However, computing the exact consta

Ecological Legacies of Pre-Columbian Settlements Evident in Palm Clusters of Neotropical Mountain Forests

ResearchDGX agent

arXiv:2507.06949v3 Announce Type: replace Abstract: Ancient populations inhabited and transformed neotropical forests, yet the spatial extent of their ecological influence remains underexplored at hig

EditCaption: Human-Aligned Instruction Synthesis for Image Editing via Supervised Fine-Tuning and Direct Preference Optimization

Model ReleasesDGX agent

arXiv:2604.08213v1 Announce Type: new Abstract: High-quality training triplets (source-target image pairs with precise editing instructions) are a critical bottleneck for scaling instruction-guided im

EEG2Vision: A Multimodal EEG-Based Framework for 2D Visual Reconstruction in Cognitive Neuroscience

ResearchDGX agent

arXiv:2604.08063v1 Announce Type: new Abstract: Reconstructing visual stimuli from non-invasive electroencephalography (EEG) remains challenging due to its low spatial resolution and high noise, parti

Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction

Model ReleasesDGX agent

arXiv:2604.07659v1 Announce Type: new Abstract: Large language models (LLMs) hold significant promise for healthcare, yet their reliability in high-stakes clinical settings is often compromised by hal

← Previous
1…966967968969970…980
Next →