AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
Human
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
21 Apr 2026

Data Compressibility Quantifies LLM Memorization

ResearchDGX agent

arXiv:2507.06056v4 Announce Type: replace Abstract: Large Language Models (LLMs) are known to memorize portions of their training data, sometimes even reproduce content verbatim when prompted appropri

Data Mixing for Large Language Models Pretraining: A Survey and Outlook

ResearchDGX agent

arXiv:2604.16380v1 Announce Type: new Abstract: Large language models (LLMs) rely on pretraining on massive and heterogeneous corpora, where training data composition has a decisive impact on training

Debate as Reward: A Multi-Agent Reward System for Scientific Ideation via RL Post-Training

SafetyDGX agent

arXiv:2604.16723v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated potential in automating scientific ideation, yet current approaches relying on iterative prompting or c

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Decision-Aware Attention Propagation for Vision Transformer Explainability

Local AiDGX agent

arXiv:2604.18094v1 Announce Type: new Abstract: Vision Transformers (ViTs) have become a dominant architecture in computer vision, yet their prediction process remains difficult to interpret because i

Decisive: Guiding User Decisions with Optimal Preference Elicitation from Unstructured Documents

ResearchDGX agent

arXiv:2604.18122v1 Announce Type: new Abstract: Decision-making is a cognitively intensive task that requires synthesizing relevant information from multiple unstructured sources, weighing competing f

Decoding AI Tutor Effects for Educational Measurement: Temporal, Multi-Outcome, and Behavior-Cognitive Analysis

SafetyDGX agent

arXiv:2604.16366v1 Announce Type: cross Abstract: Artificial intelligence (AI) tutors have become increasingly popular in learning environments. In this study, we propose an AI agent prototype framewo

Decoding RWA Tokenized U.S. Treasuries: Functional Dissection and Address Role Inference

ApplicationsDGX agent

arXiv:2507.14808v3 Announce Type: replace-cross Abstract: Tokenized U.S. Treasuries have emerged as a prominent subclass of real-world assets (RWAs), offering cryptographically secured, yield-bearing

Decomposing the Depth Profile of Fine-Tuning

ResearchDGX agent

arXiv:2604.17177v1 Announce Type: new Abstract: Fine-tuning adapts pretrained networks to new objectives. Whether the resulting depth profile of representational change reflects an intrinsic property

Decoupling the Effect of Chain-of-Thought Reasoning: A Human Label Variation Perspective

SafetyDGX agent

arXiv:2601.03154v2 Announce Type: replace Abstract: Reasoning-tuned LLMs utilizing long Chain-of-Thought (CoT) excel at single-answer tasks, yet their ability to model Human Label Variation--which req

Deep Hierarchical Knowledge Loss for Fault Intensity Diagnosis

ApplicationsDGX agent

arXiv:2604.16459v1 Announce Type: cross Abstract: Fault intensity diagnosis (FID) plays a pivotal role in intelligent manufacturing while neglecting dependencies among target classes hinders its pract

Deep learning based Non-Rigid Volume-to-Surface Registration for Brain Shift compensation Using Point Cloud

SafetyDGX agent

arXiv:2604.17389v1 Announce Type: new Abstract: Soft-tissue deformation remains a major limitation in image-guided neurosurgery, where intra-operative anatomy can deviate substantially from pre-operat

Deep Learning-Enhanced Calibration of the Heston Model: A Unified Framework

Model ReleasesDGX agent

arXiv:2510.24074v2 Announce Type: replace-cross Abstract: The Heston stochastic volatility model is a widely used tool in financial mathematics for pricing European options. However, its calibration r

Deep Learning for Virtual Reality User Identification: A Benchmark

Model ReleasesDGX agent

arXiv:2604.16341v1 Announce Type: cross Abstract: Virtual Reality (VR) applications require robust user identification systems to ensure secure access to equipment and protect worker identities. Motio

DeepDetect: Learning All-in-One Dense Keypoints

ResearchDGX agent

arXiv:2510.17422v4 Announce Type: replace Abstract: Keypoint detection is the foundation of many computer vision tasks, including image registration, structure-from-motion, 3D reconstruction, visual o

DeepRitzSplit Neural Operator for Phase-Field Models via Energy Splitting

ResearchDGX agent

arXiv:2604.18261v1 Announce Type: cross Abstract: The multi-scale and non-linear nature of phase-field models of solidification requires fine spatial and temporal discretization, leading to long compu

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models

SafetyDGX agent

arXiv:2511.15669v2 Announce Type: replace Abstract: Does Chain-of-Thought (CoT) reasoning genuinely improve Vision-Language-Action (VLA) models, or does it merely add overhead? Existing CoT-VLA system

Defragmenting Language Models: An Interpretability-based Approach for Vocabulary Expansion

ResearchDGX agent

arXiv:2604.16656v1 Announce Type: new Abstract: All languages are equal; when it comes to tokenization, some are more equal than others. Tokens are the hidden currency that dictate the cost and latenc

DeInfer: Efficient Parallel Inferencing for Decomposed Large Language Models

ResearchDGX agent

arXiv:2604.17709v1 Announce Type: new Abstract: Existing works on large language model (LLM) decomposition mainly focus on improving performance on downstream tasks, but they ignore the poor parallel

DEM Refinement and Validation on the Lunar Surface Using Shape-from-Shading with Chandrayaan-2 OHRC Imagery

Model ReleasesDGX agent

arXiv:2604.17436v1 Announce Type: new Abstract: This study presents a Shape from Shading (SfS) framework to enhance sub-metre resolution lunar digital elevation models (DEMs) using imagery from the Or

Demonstrating Real Advantage of Machine-Learning-Enhanced Monte Carlo for Combinatorial Optimization

ResearchDGX agent

arXiv:2510.19544v2 Announce Type: replace-cross Abstract: Combinatorial optimization problems are central to both practical applications and the development of optimization methods. While classical an

Demystifying the unreasonable effectiveness of online alignment methods

SafetyDGX agent

arXiv:2604.17207v1 Announce Type: cross Abstract: Iterative alignment methods based on purely greedy updates are remarkably effective in practice, yet existing theoretical guarantees of (O(log T)) KL-

Denoise and Align: Diffusion-Driven Foreground Knowledge Prompting for Open-Vocabulary Temporal Action Detection

Local AiDGX agent

arXiv:2604.18313v1 Announce Type: new Abstract: Open-Vocabulary Temporal Action Detection (OV-TAD) aims to localize and classify action segments of unseen categories in untrimmed videos, where effecti

Densemarks: Learning Canonical Embeddings for Human Heads Images via Point Tracks

TutorialsDGX agent

arXiv:2511.02830v2 Announce Type: replace Abstract: We propose DenseMarks - a new learned representation for human heads, enabling high-quality dense correspondences of human head images. For a 2D ima

Depth Adaptive Efficient Visual Autoregressive Modeling

ResearchDGX agent

arXiv:2604.17286v1 Announce Type: new Abstract: Visual Autoregressive (VAR) modeling inefficiently applies a fixed computational depth to each position when generating high-resolution images. While ex

Depth Registers Unlock W4A4 on SwiGLU: A Reader/Generator Decomposition

Model ReleasesDGX agent

arXiv:2604.18128v1 Announce Type: new Abstract: We study post-training W4A4 quantization in a controlled 300M-parameter SwiGLU decoder-only language model trained on 5B tokens of FineWeb-Edu, and ask

Designing Explainable Conversational Agentic Systems for Guarani Speakers

AgentsDGX agent

arXiv:2603.05743v3 Announce Type: replace Abstract: Although artificial intelligence (AI) and Human-Computer Interaction (HCI) systems are often presented as universal solutions, their design remains

Detecting Alarming Student Verbal Responses using Text and Audio Classifier

SafetyDGX agent

arXiv:2604.16717v1 Announce Type: new Abstract: This paper addresses a critical safety gap in the use Automated Verbal Response Scoring (AVRS). We present a novel hybrid framework for troubled student

Detecting LLM-Generated Spam Reviews by Integrating Language Model Embeddings and Graph Neural Network

Model ReleasesDGX agent

arXiv:2510.01801v2 Announce Type: replace Abstract: The rise of large language models (LLMs) has enabled the generation of highly persuasive spam reviews that closely mimic human writing. These review

DexWorldModel: Causal Latent World Modeling towards Automated Learning of Embodied Tasks

ApplicationsDGX agent

arXiv:2604.16484v1 Announce Type: new Abstract: Deploying generative World-Action Models for manipulation is severely bottlenecked by redundant pixel-level reconstruction, O(T) memory scaling, and seq

DFedReweighting: A Unified Framework for Objective-Oriented Reweighting in Decentralized Federated Learning

SafetyDGX agent

arXiv:2512.12022v2 Announce Type: replace Abstract: Decentralized federated learning (DFL) has emerged as a promising paradigm that enables multiple clients to collaboratively train machine learning m

DGSSM: Diffusion guided state-space models for multimodal salient object detection

ResearchDGX agent

arXiv:2604.17585v1 Announce Type: new Abstract: Salient object detection (SOD) requires modeling both long-range contextual dependencies and fine-grained structural details, which remains challenging

Diagnosing LLM-based Rerankers in Cold-Start Recommender Systems: Coverage, Exposure and Practical Mitigations

SafetyDGX agent

arXiv:2604.16318v1 Announce Type: cross Abstract: Large language models (LLMs) and cross-encoder rerankers have gained attention for improving recommender systems, particularly in cold-start scenarios

DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs

SafetyDGX agent

arXiv:2601.03559v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning improves multi-step mathematical problem solving in large language models but remains vulnerable to exposure bias a

Different Paths to Harmful Compliance: Behavioral Side Effects and Mechanistic Divergence Across LLM Jailbreaks

Local AiDGX agent

arXiv:2604.18510v1 Announce Type: cross Abstract: Open-weight language models can be rendered unsafe through several distinct interventions, but the resulting models may differ substantially in capabi

Differential Privacy in Two-Layer Networks: How DP-SGD Harms Fairness and Robustness

SafetyDGX agent

arXiv:2603.04881v2 Announce Type: replace Abstract: Differentially private learning is essential for training models on sensitive data, but empirical studies consistently show that it can degrade perf

DifFoundMAD: Foundation Models meet Differential Morphing Attack Detection

Model ReleasesDGX agent

arXiv:2604.17961v1 Announce Type: new Abstract: In this work, we introduce DifFoundMAD, a parameter-efficient D-MAD framework that exploits the generalisation capabilities of vision foundation models

DiffuSAM: Diffusion Guided Zero-Shot Object Grounding for Remote Sensing Imagery

ResearchDGX agent

arXiv:2604.18201v1 Announce Type: new Abstract: Diffusion models have emerged as powerful tools for a wide range of vision tasks, including text-guided image generation and editing. In this work, we e

Diffusion-Based Optimization for Accelerated Convergence of Redundant Dual-Arm Minimum Time Problems

ResearchDGX agent

arXiv:2604.16670v1 Announce Type: new Abstract: We present a framework leveraging a novel variant of the model-based diffusion algorithm to minimize the time required for a redundant dual-arm robot co

Dimensional Criticality at Grokking Across MLPs and Transformers

ResearchDGX agent

arXiv:2604.16431v1 Announce Type: new Abstract: Abrupt transitions between distinct dynamical regimes are a hallmark of complex systems. Grokking in deep neural networks provides a striking example --

DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching

ResearchDGX agent

arXiv:2602.05449v3 Announce Type: replace Abstract: While diffusion models have achieved great success in the field of video generation, this progress is accompanied by a rapidly escalating computatio

Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining

Model ReleasesDGX agent

arXiv:2604.16391v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have shown great potential in building generalist robots, but still face a dilemma-misalignment of 2D image foreca

Dissipative Latent Residual Physics-Informed Neural Networks for Modeling and Identification of Electromechanical Systems

TutorialsDGX agent

arXiv:2604.18277v1 Announce Type: new Abstract: Accurate dynamical modeling is essential for simulation and control of embodied systems, yet first-principles models of electromechanical systems often

Distributional Off-Policy Evaluation with Deep Quantile Process Regression

SafetyDGX agent

arXiv:2604.18143v1 Announce Type: cross Abstract: This paper investigates the off-policy evaluation (OPE) problem from a distributional perspective. Rather than focusing solely on the expectation of t

Distributionally Robust Regret Optimal Control Under Moment-Based Ambiguity Sets

ResearchDGX agent

arXiv:2512.10906v2 Announce Type: replace-cross Abstract: We consider a class of finite-horizon, linear-quadratic stochastic control problems, where the probability distribution governing the noise pr

Diverse Dictionary Learning

SafetyDGX agent

arXiv:2604.17568v1 Announce Type: new Abstract: Given only observational data X = g(Z), where both the latent variables Z and the generating process g are unknown, recovering Z is ill-posed without ad

Diversity Collapse in Multi-Agent LLM Systems: Structural Coupling and Collective Failure in Open-Ended Idea Generation

AgentsDGX agent

arXiv:2604.18005v1 Announce Type: cross Abstract: Multi-agent systems (MAS) are increasingly used for open-ended idea generation, driven by the expectation that collective interaction will broaden the

DMax: Aggressive Parallel Decoding for dLLMs

SafetyDGX agent

arXiv:2604.08302v2 Announce Type: replace Abstract: We present DMax, a new paradigm for efficient diffusion language models (dLLMs). It mitigates error accumulation in parallel decoding, enabling aggr

Do LLM-derived graph priors improve multi-agent coordination?

Model ReleasesDGX agent

arXiv:2604.17191v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is crucial for AI systems that operate collaboratively in distributed and adversarial settings, particularly i

Do LLMs Encode Functional Importance of Reasoning Tokens?

ResearchDGX agent

arXiv:2601.03066v2 Announce Type: replace Abstract: Large language models solve complex tasks by generating long reasoning chains, achieving higher accuracy at the cost of increased computational cost

Do LLMs Use Cultural Knowledge Without Being Told? A Multilingual Evaluation of Implicit Pragmatic Adaptation

SafetyDGX agent

arXiv:2604.17718v1 Announce Type: new Abstract: Many benchmarks show that large language models can answer direct questions about culture. We study a different question: do they also change how they s

DocQAC: Adaptive Trie-Guided Decoding for Effective In-Document Query Auto-Completion

Model ReleasesDGX agent

arXiv:2604.18257v1 Announce Type: cross Abstract: Query auto-completion (QAC) has been widely studied in the context of web search, yet remains underexplored for in-document search, which we term DocQ

Document-as-Image Representations Fall Short for Scientific Retrieval

Model ReleasesDGX agent

arXiv:2604.18508v1 Announce Type: cross Abstract: Many recent document embedding models are trained on document-as-image representations, embedding rendered pages as images rather than the underlying

Does AI See like Art Historians? Interpreting How Vision Language Models Recognize Artistic Style

ResearchDGX agent

arXiv:2603.11024v2 Announce Type: replace Abstract: VLMs have become increasingly proficient at a range of computer vision tasks, such as visual question answering and object detection. This includes

Does 'Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients?

SafetyDGX agent

arXiv:2604.18161v1 Announce Type: new Abstract: In policy gradient reinforcement learning, access to a differentiable model enables 1st-order gradient estimation that accelerates learning compared to

Does Welsh media need a review? Detecting bias in Nation.Cymru's political reporting

SafetyDGX agent

arXiv:2604.17628v1 Announce Type: new Abstract: Wales' political landscape has been marked by growing accusations of bias in Welsh media. This paper takes the first computational step toward testing t

Domain-oriented RAG Assessment (DoRA): Synthetic Benchmarking for RAG-based Question Answering on Defense Documents

Model ReleasesDGX agent

arXiv:2604.17943v1 Announce Type: new Abstract: Open-domain RAG benchmarks over public corpora can overestimate deployment performance due to pretraining overlap and weak attribution requirements. We

Domain-Specialized Object Detection via Model-Level Mixtures of Experts

ResearchDGX agent

arXiv:2604.18256v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models provide a structured approach to combining specialized neural networks and offer greater interpretability than conventio

Don't Adapt Small Language Models for Tools; Adapt Tool Schemas to the Models

AgentsDGX agent

arXiv:2510.07248v3 Announce Type: replace Abstract: Small language models (SLMs) enable scalable tool-augmented multi-agent systems where multiple SLMs handle subtasks orchestrated by a powerful coord

DORA Explorer: Improving the Exploration Ability of LLMs Without Training

Model ReleasesDGX agent

arXiv:2604.17244v1 Announce Type: new Abstract: Despite the rapid progress, LLMs for sequential decision-making (i.e., LLM agents) still struggle to produce diverse outputs. This leads to insufficient

DOSE: Data Selection for Multi-Modal LLMs via Off-the-Shelf Models

SafetyDGX agent

arXiv:2604.16979v1 Announce Type: cross Abstract: High-quality and diverse multimodal data are essential for improving vision-language models (VLMs), yet existing datasets often contain noisy, redunda

← Previous
1…885886887888889…998
Next →