AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
Human
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
10 Apr 2026

Discrete Flow Matching Policy Optimization

SafetyDGX agent

arXiv:2604.06491v1 Announce Type: cross Abstract: We introduce Discrete flow Matching policy Optimization (DoMinO), a unified framework for Reinforcement Learning (RL) fine-tuning Discrete Flow Matchi

DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs

Model ReleasesDGX agent

arXiv:2604.06250v1 Announce Type: cross Abstract: When asked to describe a molecular diagram, a Vision-Language Model correctly identifies ``a benzene ring with an -OH group.'' When asked to reason ab

Distilling Specialized Orders for Visual Generation

ResearchDGX agent

arXiv:2504.17069v2 Announce Type: replace Abstract: Autoregressive (AR) image generators are becoming increasingly popular due to their ability to produce high-quality images and their scalability. Ty

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Distributed Interpretability and Control for Large Language Models

Model ReleasesDGX agent

arXiv:2604.06483v1 Announce Type: cross Abstract: Large language models that require multiple GPU cards to host are usually the most capable models. It is necessary to understand and steer these model

Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models

Model ReleasesDGX agent

arXiv:2604.08284v1 Announce Type: new Abstract: Large language models store not only isolated facts but also rules that support reasoning across symbolic expressions, natural language explanations, an

Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook

SafetyDGX agent

arXiv:2604.06210v2 Announce Type: cross Abstract: As LLMs are globally deployed, aligning their cultural value orientations is critical for safety and user engagement. However, existing benchmarks fac

DIVERSED: Relaxed Speculative Decoding via Dynamic Ensemble Verification

ResearchDGX agent

arXiv:2604.07622v1 Announce Type: new Abstract: Speculative decoding is an effective technique for accelerating large language model inference by drafting multiple tokens in parallel. In practice, its

DMin: Scalable Training Data Influence Estimation for Diffusion Models

ResearchDGX agent

arXiv:2412.08637v4 Announce Type: replace Abstract: Identifying the training data samples that most influence a generated image is a critical task in understanding diffusion models (DMs), yet existing

Do MLLMs Really Understand Space? A Mathematical Reasoning Evaluation

Model ReleasesDGX agent

arXiv:2602.11635v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have achieved strong performance on perception-oriented tasks, yet their ability to perform mathematical sp

Do We Need Distinct Representations for Every Speech Token? Unveiling and Exploiting Redundancy in Large Speech Language Models

ResearchDGX agent

arXiv:2604.06871v1 Announce Type: cross Abstract: Large Speech Language Models (LSLMs) typically operate at high token rates (tokens/s) to ensure acoustic fidelity, yet this results in sequence length

Domain-Contextualized Inference: A Computable Graph Architecture for Explicit-Domain Reasoning

Model ReleasesDGX agent

arXiv:2604.04344v2 Announce Type: replace Abstract: We establish a computation-substrate-agnostic inference architecture in which domain is an explicit first-class computational parameter. This produc

'Don't Be Afraid, Just Learn': Insights from Industry Practitioners to Prepare Software Engineers in the Age of Generative AI

TutorialsDGX agent

arXiv:2604.06342v1 Announce Type: cross Abstract: Although tension between university curricula and industry expectations has existed in some form for decades, the rapid integration of generative AI (

'Don't Do That!': Guiding Embodied Systems through Large Language Model-based Constraint Generation

ResearchDGX agent

arXiv:2506.04500v3 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have spurred interest in robotic navigation that incorporates complex spatial, mathematica

Don't Label Twice: Quantity Beats Quality when Comparing Binary Classifiers on a Budget

TutorialsDGX agent

arXiv:2402.02249v3 Announce Type: replace Abstract: We study how to best spend a budget of noisy labels to compare the accuracy of two binary classifiers. It's common practice to collect and aggregate

Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agents

Model ReleasesDGX agent

arXiv:2604.08369v1 Announce Type: cross Abstract: Inference-time compute scaling has emerged as a powerful technique for improving the reliability of large language model (LLM) agents, but existing me

DosimeTron: Automating Personalized Monte Carlo Radiation Dosimetry in PET/CT with Agentic AI

Model ReleasesDGX agent

arXiv:2604.06280v1 Announce Type: cross Abstract: Purpose: To develop and evaluate DosimeTron, an agentic AI system for automated patient-specific MC internal radiation dosimetry in PET/CT examination

Double: Breaking the Acceleration Limit via Double Retrieval Speculative Parallelism

ResearchDGX agent

arXiv:2601.05524v2 Announce Type: replace Abstract: Parallel Speculative Decoding (PSD) accelerates traditional Speculative Decoding (SD) by overlapping draft generation with verification. However, it

DP-DeGauss: Dynamic Probabilistic Gaussian Decomposition for Egocentric 4D Scene Reconstruction

ResearchDGX agent

arXiv:2604.07986v1 Announce Type: new Abstract: Egocentric video is crucial for next-generation 4D scene reconstruction, with applications in AR/VR and embodied AI. However, reconstructing dynamic fir

DQA: Diagnostic Question Answering for IT Support

ApplicationsDGX agent

arXiv:2604.05350v2 Announce Type: replace Abstract: Enterprise IT support interactions are fundamentally diagnostic: effective resolution requires iterative evidence gathering from ambiguous user repo

Draw-In-Mind: Rebalancing Designer-Painter Roles in Unified Multimodal Models Benefits Image Editing

Model ReleasesDGX agent

arXiv:2509.01986v4 Announce Type: replace-cross Abstract: In recent years, integrating multimodal understanding and generation into a single unified model has emerged as a promising paradigm. While th

Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control

SafetyDGX agent

arXiv:2604.03540v2 Announce Type: replace Abstract: Although multi-step generative policies achieve strong performance in robotic manipulation by modeling multimodal action distributions, they require

Drifting Fields are not Conservative

ResearchDGX agent

arXiv:2604.06333v1 Announce Type: new Abstract: Drifting models generate high-quality samples in a single forward pass by transporting generated samples toward the data distribution using a vector val

DROP: Distributional and Regular Optimism and Pessimism for Reinforcement Learning

Model ReleasesDGX agent

arXiv:2410.17473v2 Announce Type: replace Abstract: In reinforcement learning (RL), temporal difference (TD) error is known to be related to the firing rate of dopamine neurons. It has been observed t

DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM Editing

Model ReleasesDGX agent

arXiv:2604.07965v1 Announce Type: new Abstract: Model editing aims to update knowledge to add new concepts and change relevant information without retraining. Lifelong editing is a challenging task, p

Dual-level Modality Debiasing Learning for Unsupervised Visible-Infrared Person Re-Identification

Model ReleasesDGX agent

arXiv:2512.03745v2 Announce Type: replace Abstract: Two-stage learning pipeline has achieved promising results in unsupervised visible-infrared person re-identification (USL-VI-ReID). It first perform

Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving

Model ReleasesDGX agent

arXiv:2604.08075v1 Announce Type: new Abstract: Production vLLM fleets typically provision each instance for the worst-case context length, leading to substantial KV-cache over-allocation and under-ut

DYCP: Dynamic Context Pruning for Long-Form Dialogue with LLMs

ResearchDGX agent

arXiv:2601.07994v4 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly operate over long-form dialogues with frequent topic shifts. While recent LLMs support extended context wi

Dynamic Context Evolution for Scalable Synthetic Data Generation

Model ReleasesDGX agent

arXiv:2604.07147v1 Announce Type: cross Abstract: Large language models produce repetitive output when prompted independently across many batches, a phenomenon we term cross-batch mode collapse: the p

DynLP: Parallel Dynamic Batch Update for Label Propagation in Semi-Supervised Learning

HardwareDGX agent

arXiv:2604.06596v1 Announce Type: cross Abstract: Semi-supervised learning aims to infer class labels using only a small fraction of labeled data. In graph-based semi-supervised learning, this is typi

E-3DPSM: A State Machine for Event-Based Egocentric 3D Human Pose Estimation

ResearchDGX agent

arXiv:2604.08543v1 Announce Type: new Abstract: Event cameras offer multiple advantages in monocular egocentric 3D human pose estimation from head-mounted devices, such as millisecond temporal resolut

E2Edev: Benchmarking Large Language Models in End-to-End Software Development Task

Model ReleasesDGX agent

arXiv:2510.14509v3 Announce Type: replace-cross Abstract: The rapid advancement in large language models (LLMs) has demonstrated significant potential in End-to-End Software Development (E2ESD). Howev

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation

SafetyDGX agent

arXiv:2602.13669v4 Announce Type: replace Abstract: Recent multi-modal video generation models have achieved high visual quality, but their prohibitive latency and limited temporal stability hinder re

ECLipsE-Gen-Local: Efficient Compositional Local Lipschitz Estimates for Deep Neural Networks

ResearchDGX agent

arXiv:2510.05261v2 Announce Type: replace Abstract: The Lipschitz constant is a key measure for certifying the robustness of neural networks to input perturbations. However, computing the exact consta

Ecological Legacies of Pre-Columbian Settlements Evident in Palm Clusters of Neotropical Mountain Forests

ResearchDGX agent

arXiv:2507.06949v3 Announce Type: replace Abstract: Ancient populations inhabited and transformed neotropical forests, yet the spatial extent of their ecological influence remains underexplored at hig

EditCaption: Human-Aligned Instruction Synthesis for Image Editing via Supervised Fine-Tuning and Direct Preference Optimization

Model ReleasesDGX agent

arXiv:2604.08213v1 Announce Type: new Abstract: High-quality training triplets (source-target image pairs with precise editing instructions) are a critical bottleneck for scaling instruction-guided im

EEG2Vision: A Multimodal EEG-Based Framework for 2D Visual Reconstruction in Cognitive Neuroscience

ResearchDGX agent

arXiv:2604.08063v1 Announce Type: new Abstract: Reconstructing visual stimuli from non-invasive electroencephalography (EEG) remains challenging due to its low spatial resolution and high noise, parti

Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction

Model ReleasesDGX agent

arXiv:2604.07659v1 Announce Type: new Abstract: Large language models (LLMs) hold significant promise for healthcare, yet their reliability in high-stakes clinical settings is often compromised by hal

Efficient Learned Data Compression via Dual-Stream Feature Decoupling

ResearchDGX agent

arXiv:2604.07239v1 Announce Type: cross Abstract: While Learned Data Compression (LDC) has achieved superior compression ratios, balancing precise probability modeling with system efficiency remains c

Efficient PRM Training Data Synthesis via Formal Verification

ResearchDGX agent

arXiv:2505.15960v3 Announce Type: replace Abstract: Process Reward Models (PRMs) have emerged as a promising approach for improving LLM reasoning capabilities by providing process supervision over rea

Efficient Provably Secure Linguistic Steganography via Range Coding

ResearchDGX agent

arXiv:2604.08052v1 Announce Type: new Abstract: Linguistic steganography involves embedding secret messages within seemingly innocuous texts to enable covert communication. Provable security, which is

Efficient Quantization of Mixture-of-Experts with Theoretical Generalization Guarantees

ResearchDGX agent

arXiv:2604.06515v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) allows scaling of language and vision models efficiently by activating only a small subset of experts per input. While

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World

SafetyDGX agent

arXiv:2604.07607v1 Announce Type: cross Abstract: Robot learning increasingly depends on large and diverse data, yet robot data collection remains expensive and difficult to scale. Egocentric human da

ELC: Evidential Lifelong Classifier for Uncertainty Aware Radar Pulse Classification

TutorialsDGX agent

arXiv:2604.06958v1 Announce Type: cross Abstract: Reliable radar pulse classification is essential in Electromagnetic Warfare for situational awareness and decision support. Deep Neural Networks have

EMMa: End-Effector Stability-Oriented Mobile Manipulation for Tracked Rescue Robots

SafetyDGX agent

arXiv:2604.08292v1 Announce Type: new Abstract: The autonomous operation of tracked mobile manipulators in rescue missions requires not only ensuring the reachability and safety of robot motion but al

EmoMAS: Emotion-Aware Multi-Agent System for High-Stakes Edge-Deployable Negotiation with Bayesian Orchestration

Local AiDGX agent

arXiv:2604.07003v1 Announce Type: new Abstract: Large language models (LLMs) has been widely used for automated negotiation, but their high computational cost and privacy risks limit deployment in pri

Emotion Concepts and their Function in a Large Language Model

Model ReleasesDGX agent

arXiv:2604.07729v1 Announce Type: cross Abstract: Large language models (LLMs) sometimes appear to exhibit emotional reactions. We investigate why this is the case in Claude Sonnet 4.5 and explore imp

EMSDialog: Synthetic Multi-person Emergency Medical Service Dialogue Generation from Electronic Patient Care Reports via Multi-LLM Agents

AgentsDGX agent

arXiv:2604.07549v1 Announce Type: new Abstract: Conversational diagnosis prediction requires models to track evolving evidence in streaming clinical conversations and decide when to commit to a diagno

Enabling Intrinsic Reasoning over Dense Geospatial Embeddings with DFR-Gemma

Model ReleasesDGX agent

arXiv:2604.07490v1 Announce Type: new Abstract: Representation learning for geospatial and spatio-temporal data plays a critical role in enabling general-purpose geospatial intelligence. Recent geospa

Energy-based Tissue Manifolds for Longitudinal Multiparametric MRI Analysis

TutorialsDGX agent

arXiv:2604.07180v1 Announce Type: cross Abstract: We propose a geometric framework for longitudinal multi-parametric MRI analysis based on patient-specific energy modelling in sequence space. Rather t

Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models

ResearchDGX agent

arXiv:2604.06893v1 Announce Type: cross Abstract: Deep convolutional neural networks achieve remarkable performance by exhaustively processing dense spatial feature maps, yet this brute-force strategy

Energy Saving for Cell-Free Massive MIMO Networks: A Multi-Agent Deep Reinforcement Learning Approach

AgentsDGX agent

arXiv:2604.07133v1 Announce Type: cross Abstract: This paper focuses on energy savings in downlink operation of cell-free massive MIMO (CF mMIMO) networks under dynamic traffic conditions. We propose

Entropy After </Think> for reasoning model early exiting

Model ReleasesDGX agent

arXiv:2509.26522v3 Announce Type: replace Abstract: Reasoning LLMs show improved performance with longer chains of thought. However, recent work has highlighted their tendency to overthink, continuing

Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models

ResearchDGX agent

arXiv:2604.08456v1 Announce Type: cross Abstract: Despite rapid progress, pretrained vision-language models still struggle when answers depend on tiny visual details or on combining clues spread acros

Environmental, Social and Governance Sentiment Analysis on Slovene News: A Novel Dataset and Models

ApplicationsDGX agent

arXiv:2604.06826v1 Announce Type: cross Abstract: Environmental, Social, and Governance (ESG) considerations are increasingly integral to assessing corporate performance, reputation, and long-term sus

EPIR: An Efficient Patch Tokenization, Integration and Representation Framework for Micro-expression Recognition

TutorialsDGX agent

arXiv:2604.08106v1 Announce Type: new Abstract: Micro-expression recognition can obtain the real emotion of the individual at the current moment. Although deep learning-based methods, especially Trans

Epistemic Robust Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.07072v1 Announce Type: new Abstract: Offline reinforcement learning learns policies from fixed datasets without further environment interaction. A key challenge in this setting is epistemic

Equivariant Multi-agent Reinforcement Learning for Multimodal Vehicle-to-Infrastructure Systems

SafetyDGX agent

arXiv:2604.06914v1 Announce Type: new Abstract: In this paper, we study a vehicle-to-infrastructure (V2I) system where distributed base stations (BSs) acting as road-side units (RSUs) collect multimod

ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions

Model ReleasesDGX agent

arXiv:2604.07772v1 Announce Type: new Abstract: Open-world video anomaly detection (OWVAD) aims to detect and explain abnormal events under different anomaly definitions, which is important for applic

ETCH-X: Robustify Expressive Body Fitting to Clothed Humans with Composable Datasets

Model ReleasesDGX agent

arXiv:2604.08548v1 Announce Type: new Abstract: Human body fitting, which aligns parametric body models such as SMPL to raw 3D point clouds of clothed humans, serves as a crucial first step for downst

Evaluating In-Context Translation with Synchronous Context-Free Grammar Transduction

ResearchDGX agent

arXiv:2604.07320v1 Announce Type: cross Abstract: Low-resource languages pose a challenge for machine translation with large language models (LLMs), which require large amounts of training data. One p

← Previous
1…10121013101410151016…1025
Next →