AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
Human
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
18 May 2026

Decomposed Vision-Language Alignment for Fine-Grained Open-Vocabulary Segmentation

SafetyDGX agent

arXiv:2605.15942v1 Announce Type: cross Abstract: Open-vocabulary segmentation models often struggle to generalize to unseen combinations of object categories and attributes, because fine-grained desc

Decouple Searching from Training: Scaling Data Mixing via Model Merging for Large Language Model Pre-training

Model ReleasesDGX agent

arXiv:2602.00747v2 Announce Type: replace-cross Abstract: Determining an effective data mixture is a key factor in Large Language Model (LLM) pre-training, where models must balance general competence

Deep Double Q-learning

SafetyDGX agent

arXiv:2507.00275v2 Announce Type: replace-cross Abstract: Double Q-learning is a classical control algorithm that mitigates the maximization bias of Q-learning. To do so, it explicitly trains two inde

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Deep Learning Alternatives of the Kolmogorov Superposition Theorem

TutorialsDGX agent

arXiv:2410.01990v3 Announce Type: replace Abstract: This paper explores alternative formulations of the Kolmogorov Superposition Theorem (KST) as a foundation for neural network design. The original K

Deep Pre-Alignment for VLMs

Model ReleasesDGX agent

arXiv:2605.15300v1 Announce Type: new Abstract: Most Vision Language Models (VLMs) directly map outputs from ViT encoders to the LLM via a lightweight projector. While effective, recent analysis sugge

DeepSlide: From Artifacts to Presentation Delivery

Model ReleasesDGX agent

arXiv:2605.15202v1 Announce Type: new Abstract: Presentations are a primary medium for scholarly communication, yet most AI slide generators optimize the artifact (a visually plausible deck) while und

Defining Cultural Capabilities for AI Evaluation: A Taxonomy Grounded in Intercultural Communication Theory

ApplicationsDGX agent

arXiv:2605.15990v1 Announce Type: new Abstract: Tremendous efforts have been put into evaluating the inclusivity and effectiveness of AI systems across cultures. However, the cultural capabilities con

Degradation-Aware Blur-Segmentation of Brain Tumor

ResearchDGX agent

arXiv:2605.15671v1 Announce Type: cross Abstract: Multimodal 3D MRI brain tumor segmentation is a pivotal step in radiotherapy target delineation, surgical planning and post-treatment assessment. Exis

DeltaPrompts: Escaping the Zero-Delta Trap in Multimodal Distillation

SafetyDGX agent

arXiv:2605.15532v1 Announce Type: cross Abstract: Distillation enables compact Vision-Language Models (VLMs) to obtain strong reasoning capabilities, yet the prompts driving this process are typically

Density Estimation via Binless Multidimensional Integration

ResearchDGX agent

arXiv:2407.08094v3 Announce Type: replace-cross Abstract: We introduce the Binless Multidimensional Thermodynamic Integration (BMTI) method for nonparametric, robust, and data-efficient density estima

Designing Datacenter Power Delivery Hierarchies for the AI Era

SafetyDGX agent

arXiv:2605.16255v1 Announce Type: cross Abstract: Demand for AI accelerators is rapidly increasing rack power density, with projections approaching 1MW per deployment by 2027. This poses a major chall

Designing for Robot Wranglers: A Synthesis of Literature and Practice

ResearchDGX agent

arXiv:2605.15892v1 Announce Type: new Abstract: Robots are increasingly present in human spaces, such as for conducting deliveries in hospitals, interacting with visitors at museums, and stocking item

Detecting Heel Strike and toe off Events Using Kinematic Methods and LSTM Models

SafetyDGX agent

arXiv:2503.00794v2 Announce Type: replace Abstract: Accurate gait event detection is crucial for gait analysis, rehabilitation, and assistive technology, particularly in exoskeleton control, where pre

Detecting Localized Density Anomalies in Multivariate Data via Coin-Flip Statistics

Local AiDGX agent

arXiv:2503.23927v3 Announce Type: replace-cross Abstract: Detecting localized differences between two samples is a central task in scientific data analysis, required for the identification of signal e

Detecting Privilege Escalation in Polyglot Microservices via Agentic Program Analysis

AgentsDGX agent

arXiv:2605.15569v1 Announce Type: cross Abstract: Microservices are widely adopted in modern cloud systems due to their scalability and fault tolerance. However, microservice architectures introduce s

DetectRL-X: Towards Reliable Multilingual and Real-World LLM-Generated Text Detection

Model ReleasesDGX agent

arXiv:2605.15518v1 Announce Type: new Abstract: The effective detection and governance of Large Language Model (LLM) generated content has become increasingly critical due to the growing risk of misus

Deterministic Coreset for Lp Subspace

Model ReleasesDGX agent

arXiv:2601.00361v3 Announce Type: replace-cross Abstract: We introduce the first iterative algorithm for constructing a arepsilon-coreset that guarantees deterministic ell_p subspace embedding for any

Deterministic Event-Graph Substrates as World Models for Counterfactual Reasoning

Model ReleasesDGX agent

arXiv:2605.15967v1 Announce Type: new Abstract: We study event-graph substrates: a class of world models that represent agent state as an append-only log of typed RDF triples and answer counterfactual

DexJoCo: A Benchmark and Toolkit for Task-Oriented Dexterous Manipulation on MuJoCo

Model ReleasesDGX agent

arXiv:2605.16257v1 Announce Type: new Abstract: Achieving human-level manipulation requires dexterous robotic hands capable of complex object interactions. Advancing such capabilities further demands

Diagonal Adaptive Non-local Observables on Quantum Neural Networks

Model ReleasesDGX agent

arXiv:2605.15410v1 Announce Type: cross Abstract: Adaptive Non-local Observables (ANOs) have shown that making quantum observables dynamic can substantially enlarge the function space of Variational Q

Differentiable Mixture-of-Agents Incentivizes Swarm Intelligence of Large Language Models

AgentsDGX agent

arXiv:2605.15706v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have catalyzed the development of multi-agent systems (MAS) for complex reasoning tasks. However, existi

Differentially Private Motif-Preserving Multi-modal Hashing

SafetyDGX agent

arXiv:2605.15460v1 Announce Type: cross Abstract: Cross-modal hashing enables efficient retrieval by encoding images and text into compact binary codes. State-of-the-art methods rely on semantic simil

Diffusion Policy for Coordinated Control of a Nonholonomic Mobile Base and Dual Arms in Door Opening and Passing

SafetyDGX agent

arXiv:2605.15352v1 Announce Type: new Abstract: Opening heavy, self closing doors, especially those that require pulling remains a long standing challenge in robotics. Humans naturally employ both arm

DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments

SafetyDGX agent

arXiv:2605.15519v1 Announce Type: cross Abstract: Visual active search (VAS) has been introduced as a modeling framework that leverages visual cues to direct aerial (e.g., UAV-based) exploration and p

DiLA: Disentangled Latent Action World Models

ResearchDGX agent

arXiv:2605.15725v1 Announce Type: cross Abstract: Latent Action Models (LAMs) enable the learning of world models from unlabeled video by inferring abstract actions between consecutive frames. However

DimMem: Dimensional Structuring for Efficient Long-Term Agent Memory

Model ReleasesDGX agent

arXiv:2605.15759v1 Announce Type: new Abstract: Large language model (LLM) agents require long-term memory to leverage information from past interactions. However, existing memory systems often face a

DIPA: Distilled Preconditioned Algorithms for Solving Imaging Inverse Problems

ResearchDGX agent

arXiv:2605.15456v1 Announce Type: cross Abstract: Solving imaging inverse problems has usually been addressed by designing proper prior models of the underlying signal. However, minimizing the data fi

DiscoExplorer: An Open Interface for the Study of Multilingual Discourse Relations

ResearchDGX agent

arXiv:2605.15304v1 Announce Type: new Abstract: The relations connecting propositions in discourse such as cause (A because B) or concession (A although B) are a subject of intense interest in Computa

Discretizing Group-Convolutional Neural Networks for 3D Geometry in Feature Space

SafetyDGX agent

arXiv:2605.15368v1 Announce Type: new Abstract: Group-convolutional neural networks (GCNNs) are among the most important methods for introducing symmetry as an inductive bias in deep learning: In each

DiscussLLM: Teaching Large Language Models When to Speak

TutorialsDGX agent

arXiv:2508.18167v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in understanding and generating human-like text, yet they largely operate as

Do Biological Structural Guarantees Earn Their Complexity?

AgentsDGX agent

arXiv:2605.15225v1 Announce Type: cross Abstract: Biologically-inspired AI agent frameworks claim reliability benefits through structural guarantees adapted from gene regulatory networks, immune syste

Do Chinese models speak Chinese languages?

ResearchDGX agent

arXiv:2504.00289v3 Announce Type: replace-cross Abstract: The release of top-performing open-weight LLMs has cemented China's role as a leading force in AI development. Do these models support languag

Do Less, Achieve More: Do We Need Every-Step Optimization for RL Fine-tuning of Diffusion Models?

SafetyDGX agent

arXiv:2605.15855v1 Announce Type: new Abstract: Despite strong image-generation performance, diffusion models' reconstruction objectives limit alignment with human preferences. RL enables such alignme

Does Theory of Mind Improvement Really Benefit Human-AI Interactions? Empirical Findings from Interactive Evaluations

ApplicationsDGX agent

arXiv:2605.15205v1 Announce Type: new Abstract: Improving the Theory of Mind (ToM) capability of Large Language Models (LLMs) is crucial for effective social interactions between these AI models and h

Domain-Independent Game Abstraction using Word Embedding Techniques

ApplicationsDGX agent

arXiv:2605.15543v1 Announce Type: cross Abstract: Many games of interest in the real world are often intractably large, thereby necessitating the use of game abstraction to shrink them in size, typica

Don't Stop Me Yet: Sampling Loss Minima via Dissipative Riemannian Mechanics

Local AiDGX agent

arXiv:2605.15459v1 Announce Type: new Abstract: The minima of modern neural network loss functions are typically not isolated, rather they form connected components of reparameterization invariant sol

DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research

SafetyDGX agent

arXiv:2511.19399v3 Announce Type: replace-cross Abstract: Deep research agents perform multi-step research to produce long-form, well-attributed answers. However, most open deep research agents are tr

Drawback of Enforcing Equivariance and its Compensation via the Lens of Expressive Power

SafetyDGX agent

arXiv:2512.09673v3 Announce Type: replace-cross Abstract: Equivariant neural networks encode the intrinsic symmetry of data as an inductive bias, which has achieved impressive performance in wide doma

DreamSR: Towards Ultra-High-Resolution Image Super-Resolution via a Receptive-Field Enhanced Diffusion Transformer

Local AiDGX agent

arXiv:2605.15682v1 Announce Type: new Abstract: Large-scale pre-trained diffusion models have been extensively adopted for real-world image Super-Resolution because of their powerful generative priors

Driving Through the Network: Performance and Workload Under Latency and Video Impairment

SafetyDGX agent

arXiv:2605.15952v1 Announce Type: cross Abstract: Teleoperation promises to extend the operational envelope of automated vehicles, yet it critically depends on network latency and video quality. We re

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding

ResearchDGX agent

arXiv:2605.15542v1 Announce Type: new Abstract: GUI agents powered by Multimodal Large Language Models (MLLMs) have demonstrated impressive capability in understanding and executing user instructions.

DrugSAGE:Self-evolving Agent Experience for Efficient State-of-the-Art Drug Discovery

AgentsDGX agent

arXiv:2605.15461v1 Announce Type: cross Abstract: Building state-of-the-art (SOTA) predictive models for drug discovery requires expensive search over tools, architectures, and training strategies. Cu

DualKV: Shared-Prompt Flash Attention for Efficient RL Training with Large Rollouts and Long Contexts

SafetyDGX agent

arXiv:2605.15422v1 Announce Type: new Abstract: Modern RL post-training methods such as GRPO and DAPO train on N response sequences of R tokens sampled from a shared prompt of P tokens, but standard F

DualReg: Dual-Space Filtering and Reinforcement for Rigid Registration

SafetyDGX agent

arXiv:2508.17034v2 Announce Type: cross Abstract: Noisy, partially overlapping data and the need for real-time processing pose major challenges for rigid registration. Considering that feature-based m

Dynamic Chunking for Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.15676v1 Announce Type: new Abstract: Block discrete diffusion language models factorize a sequence autoregressively over fixed-size positional blocks, decoupling within-block parallel denoi

Dynamic Plasma Shape Control with Arbitrary Sensor Subsets

SafetyDGX agent

arXiv:2605.15935v1 Announce Type: new Abstract: Plasma shape control in tokamaks requires a real-time controller that tracks dynamically changing shape targets while tolerating diagnostic failures. Cl

Dynamic-TreeRPO: Breaking the Independent Trajectory Bottleneck with Structured Sampling

SafetyDGX agent

arXiv:2509.23352v3 Announce Type: replace-cross Abstract: The integration of Reinforcement Learning (RL) into flow matching models for text-to-image (T2I) generation has driven substantial advances in

Dynamics-Level Watermarking of Flow Matching Models with Random Codes

ResearchDGX agent

arXiv:2605.16239v1 Announce Type: new Abstract: We introduce a dynamics-level approach to watermarking generative models. Rather than embedding signals into model weights or outputs, we embed the wate

Echo-Forcing: A Scene Memory Framework for Interactive Long Video Generation

Model ReleasesDGX agent

arXiv:2605.16003v1 Announce Type: new Abstract: Autoregressive video diffusion models enable open-ended generation through local attention and KV caching. However, existing training-free long-video op

EduVQA: Towards Concept-Aware Assessment of Educational AI-Generated Videos

Model ReleasesDGX agent

arXiv:2603.03066v2 Announce Type: replace Abstract: Existing AI-generated video quality assessment (AIGVQA) methods mainly focus on global perceptual realism and coarse text-video alignment, while ove

Effective Harness Engineering for Algorithm Discovery with Coding Agents

ResearchDGX agent

arXiv:2605.15221v1 Announce Type: cross Abstract: AlphaEvolve and FunSearch have demonstrated the potential of combining large language models (LLMs) with evolutionary search for automated algorithm d

Efficient Image Synthesis with Sphere Latent Encoder

ResearchDGX agent

arXiv:2605.15592v1 Announce Type: new Abstract: Few-step image generation has seen rapid progress, with consistency and meanflow-based methods significantly reducing the number of sampling steps. Desp

Efficiently Solving Mixed-Hierarchy Games with Quasi-Policy Approximations

SafetyDGX agent

arXiv:2602.01568v2 Announce Type: replace-cross Abstract: Multi-robot coordination often exhibits hierarchical structure, with some robots' decisions depending on the planned behaviors of others. Whil

Egalitarian Gradient Descent: A Simple Approach to Accelerated Grokking

ResearchDGX agent

arXiv:2510.04930v2 Announce Type: replace Abstract: Grokking is the phenomenon whereby, unlike the training performance, which peaks early in the training process, the test/generalization performance

EgoExo-WM: Unlocking Exo Video for Ego World Models

SafetyDGX agent

arXiv:2605.15477v1 Announce Type: new Abstract: Egocentric world models present a promising direction for enabling agents to predict and plan, but their performance is constrained by the limited avail

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices

SafetyDGX agent

arXiv:2605.15684v1 Announce Type: new Abstract: The Diffusion Transformer (DiT) architecture is the state-of-the-art paradigm for high-fidelity image generation, underpinning models like Stable Diffus

ELDOR: A Dataset and Benchmark for Illegal Gold Mining in the Amazon Rainforest

Model ReleasesDGX agent

arXiv:2605.15397v1 Announce Type: new Abstract: Illegal gold mining in the Amazon rainforest causes deforestation, water contamination, and long-term ecosystem disruption, yet remains difficult to mon

Embedding-perturbed Exploration Preference Optimization for Flow Models

SafetyDGX agent

arXiv:2605.15803v1 Announce Type: new Abstract: Recent advancements have established Reinforcement Learning (RL) as a pivotal paradigm for aligning generative models with human intent. However, group-

Embracing Biased Transition Matrices for Complementary-Label Learning with Many Classes

SafetyDGX agent

arXiv:2605.15586v1 Announce Type: cross Abstract: Complementary-label learning (CLL) is a weakly supervised paradigm where instances are labeled with classes they do not belong to. Despite a decade of

EMFusion: An Uncertainty-Aware Conditional Diffusion Framework for Frequency-Selective EMF Forecasting in Wireless Networks

TutorialsDGX agent

arXiv:2512.15067v3 Announce Type: replace-cross Abstract: The rapid growth in wireless infrastructure has increased the need to accurately estimate and forecast electromagnetic field (EMF) levels to e

← Previous
1…673674675676677…1025
Next →