AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
Human
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
4 May 2026

Learning from Supervision with Semantic and Episodic Memory: A Reflective Approach to Agent Adaptation

Model ReleasesDGX agent

arXiv:2510.19897v2 Announce Type: replace Abstract: We investigate how agents built on pretrained large language models (LLMs) can learn target classification functions from labeled examples without p

Learning from the Unseen: Generative Data Augmentation for Geometric-Semantic Accident Anticipation

Model ReleasesDGX agent

arXiv:2605.00051v1 Announce Type: new Abstract: Anticipating traffic accidents is a critical yet unresolved problem for autonomous driving, hindered by the inherent complexity of modeling interactions

Learning How and What to Memorize: Cognition-Inspired Two-Stage Optimization for Evolving Memory

SafetyDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.00702v1 Announce Type: new Abstract: Large language model (LLM) agents require long-term user memory for consistent personalization, but limited context windows hinder tracking evolving pre

Learning Locally, Revising Globally: Global Reviser for Federated Learning with Noisy Labels

Model ReleasesDGX agent

arXiv:2412.00452v2 Announce Type: replace-cross Abstract: Conventioanl federated learning (FL) heavily depends on high-quality labels, which are often impractical in the real world, leading to the fed

Learning Multimodal Energy-Based Model with Multimodal Variational Auto-Encoder via MCMC Revision

ResearchDGX agent

arXiv:2605.00644v1 Announce Type: new Abstract: Energy-based models (EBMs) are a flexible class of deep generative models and are well-suited to capture complex dependencies in multimodal data. Howeve

Learning physically grounded traffic accident reconstruction from public accident reports

SafetyDGX agent

arXiv:2605.00050v1 Announce Type: cross Abstract: Traffic accidents are routinely documented in textual reports, yet physically grounded accident reconstruction remains difficult because detailed scen

Learning the Helmholtz equation operator with DeepONet for non-parametric 2D geometries

Local AiDGX agent

arXiv:2605.00760v1 Announce Type: new Abstract: This paper deals with solving the 2D Helmholtz equation on non-parametric domains, leveraging a physics-informed neural operator network based on the De

Learning while Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies

SafetyDGX agent

arXiv:2605.00416v1 Announce Type: new Abstract: Generalist robot policies increasingly benefit from large-scale pretraining, but offline data alone is insufficient for robust real-world deployment. De

Let ViT Speak: Generative Language-Image Pre-training

ResearchDGX agent

arXiv:2605.00809v1 Announce Type: new Abstract: In this paper, we present extbf{Gen}erative extbf{L}anguage-extbf{I}mage extbf{P}re-training (GenLIP), a minimalist generative pretraining framework for

Leveraging Vision-Language Models as Weak Annotators in Active Learning

ResearchDGX agent

arXiv:2605.00480v1 Announce Type: new Abstract: Active learning aims to reduce annotation cost by selectively querying informative samples for supervision under a limited labeling budget. In this work

Lightweight Domain Adaptation of a Large Language Model for Legal Assistance in the Indian Context

Model ReleasesDGX agent

arXiv:2505.22003v2 Announce Type: replace Abstract: In India, access to legal assistance for the general public has been observed to have a critical gap, as many citizens are not able to take full adv

LIMSSR: LLM-Driven Sequence-to-Score Reasoning under Training-Time Incomplete Multimodal Observations

ApplicationsDGX agent

arXiv:2605.00434v1 Announce Type: new Abstract: Real-world multimodal learning is often hindered by missing modalities. While Incomplete Multimodal Learning (IML) has gained traction, existing methods

Linking Behaviour and Perception to Evaluate Meaningful Human Control over Partially Automated Driving

SafetyDGX agent

arXiv:2605.00556v1 Announce Type: cross Abstract: Partial driving automation creates a tension: drivers remain legally responsible for vehicle behaviour, yet their active control is significantly redu

LLM DNA: Tracing Model Evolution via Functional Representations

ResearchDGX agent

arXiv:2509.24496v3 Announce Type: replace Abstract: The explosive growth of large language models (LLMs) has created a vast but opaque landscape: millions of models exist, yet their evolutionary relat

LLM-Oriented Information Retrieval: A Denoising-First Perspective

Model ReleasesDGX agent

arXiv:2605.00505v1 Announce Type: cross Abstract: Modern information retrieval (IR) is no longer consumed primarily by humans but increasingly by large language models (LLMs) via retrieval-augmented g

Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism

AgentsDGX agent

arXiv:2512.04341v3 Announce Type: replace Abstract: Popular offline reinforcement learning (RL) methods rely on explicit conservatism, penalizing out-of-dataset actions or restricting rollout horizons

Lost in State Space: Probing Frozen Mamba Representations

ResearchDGX agent

arXiv:2605.00253v1 Announce Type: new Abstract: Mamba's recurrent state h_t is, by construction, a compressed summary of every token seen so far. This raises a tempting hypothesis: if we extract token

Lucid-XR: An Extended-Reality Data Engine for Robotic Manipulation

Local AiDGX agent

arXiv:2605.00244v1 Announce Type: cross Abstract: We introduce Lucid-XR, a generative data engine for creating diverse and realistic-looking multi-modal data to train real-world robotic systems. At th

M-CaStLe: Uncovering Local Causal Structures in Multivariate Space-Time Gridded Data

Model ReleasesDGX agent

arXiv:2605.00398v1 Announce Type: new Abstract: Causal graph discovery for space-time systems is challenging in high-dimensional gridded data, which often has many more grid cells than temporal observ

MAEPose: Self-Supervised Spatiotemporal Learning for Human Pose Estimation on mmWave Video

TutorialsDGX agent

arXiv:2605.00242v1 Announce Type: new Abstract: Millimetre-wave (mmWave) radar offers a more privacy-preserving alternative to RGB-based human pose estimation. However, existing methods typically rely

Make Your LVLM KV Cache More Lightweight

Model ReleasesDGX agent

arXiv:2605.00789v1 Announce Type: new Abstract: Key-Value (KV) cache has become a de facto component of modern Large Vision-Language Models (LVLMs) for inference. While it enhances decoding efficiency

Making Every Verified Token Count: Adaptive Verification for MoE Speculative Decoding

ResearchDGX agent

arXiv:2605.00342v1 Announce Type: new Abstract: Tree-based speculative decoding accelerates autoregressive generation by verifying multiple draft candidates in parallel, but this advantage weakens for

Map2World: Segment Map Conditioned Text to 3D World Generation

AgentsDGX agent

arXiv:2605.00781v1 Announce Type: new Abstract: 3D world generation is essential for applications such as immersive content creation or autonomous driving simulation. Recent advances in 3D world gener

Matroid Algorithms Under Size-Sensitive Independence Oracles

ResearchDGX agent

arXiv:2605.00201v1 Announce Type: cross Abstract: The standard oracle model for matroid algorithms assumes that each independence query can be answered in constant time, regardless of the size of the

Mean-field limit from general mixtures of experts to quantum neural networks

ResearchDGX agent

arXiv:2501.14660v2 Announce Type: replace-cross Abstract: In this work, we study the asymptotic behavior of Mixture of Experts (MoE) trained via gradient flow on supervised learning problems. Our main

Memory in the LLM Era: Modular Architectures and Strategies in a Unified Framework

AgentsDGX agent

arXiv:2604.01707v2 Announce Type: replace Abstract: Memory emerges as the core module in the large language model (LLM)-based agents for long-horizon complex tasks (e.g., multi-turn dialogue, game pla

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Model ReleasesDGX agent

arXiv:2510.17281v5 Announce Type: replace Abstract: Scaling up data, parameters, and test-time computation has been the mainstream methods to improve LLM systems (LLMsys), but their upper bounds are a

MemRouter: Memory-as-Embedding Routing for Long-Term Conversational Agents

SafetyDGX agent

arXiv:2605.00356v1 Announce Type: new Abstract: Long-term conversational agents must decide which turns to store in external memory, yet recent systems rely on autoregressive LLM generation at every t

Meritocratic Fairness in Budgeted Combinatorial Multi-armed Bandits via Shapley Values

SafetyDGX agent

arXiv:2605.00762v1 Announce Type: new Abstract: We propose a new framework for meritocratic fairness in budgeted combinatorial multi-armed bandits with full-bandit feedback (BCMAB-FBF). Unlike semi-ba

Mesh Field Theory: Port-Hamiltonian Formulation of Mesh-Based Physics

SafetyDGX agent

arXiv:2605.00394v1 Announce Type: new Abstract: We present Mesh Field Theory (MeshFT) and its neural realization, MeshFT-Net: a structure-preserving framework for mesh-based continuum physics that cle

Minimizing Human Intervention in Online Classification

Model ReleasesDGX agent

arXiv:2510.23557v2 Announce Type: replace-cross Abstract: Training or fine-tuning large language model (LLM)-based systems often requires costly human feedback, yet there is limited understanding of h

MiniVLA-Nav v1: A Multi-Scene Simulation Dataset for Language-Conditioned Robot Navigation

HardwareDGX agent

arXiv:2605.00397v1 Announce Type: new Abstract: We present MiniVLA-Nav v1, a simulation dataset for Language-Conditioned Object Approach (LCOA) navigation: given a short natural-language instruction,

ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering

Model ReleasesDGX agent

arXiv:2505.23723v2 Announce Type: replace Abstract: The emergence of large language model (LLM)-based agents has significantly advanced the development of autonomous machine learning (ML) engineering.

ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models

Model ReleasesDGX agent

arXiv:2605.00689v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in cross-linguistic contexts, ensuring safety in diverse regulatory and cultural environments

MMAudio-LABEL: Audio Event Labeling via Audio Generation for Silent Video

ApplicationsDGX agent

arXiv:2605.00495v1 Announce Type: cross Abstract: Recent advances in multimodal generation have enabled high-quality audio generation from silent videos. Practical applications, such as sound producti

MMAudioReverbs: Video-Guided Acoustic Modeling for Dereverberation and Room Impulse Response Estimation

ResearchDGX agent

arXiv:2605.00431v1 Announce Type: cross Abstract: Although recent video-to-audio (V2A) models excelled at synthesizing semantically plausible sounds from visual inputs, they do not explicitly model ro

MoDAl: Self-Supervised Neural Modality Discovery via Decorrelation for Speech Neuroprosthesis

Model ReleasesDGX agent

arXiv:2605.00025v1 Announce Type: cross Abstract: Speech neuroprosthesis systems decode intended speech from neural activity in the absence of audible output, offering a path to restoring communicatio

Model-Based Reinforcement Learning with Double Oracle Efficiency in Policy Optimization and Offline Estimation

SafetyDGX agent

arXiv:2605.00393v1 Announce Type: new Abstract: Reinforcement learning (RL) in large environments often suffers from severe computational bottlenecks, as conventional regret minimization algorithms re

Modeling Subjective Urban Perception with Human Gaze

ResearchDGX agent

arXiv:2605.00764v1 Announce Type: new Abstract: Urban perception describes how people subjectively evaluate urban environments, shaping how cities are experienced and understood. Existing computationa

MSACT: Multistage Spatial Alignment for Stable Low-Latency Fine Manipulation

Local AiDGX agent

arXiv:2605.00475v1 Announce Type: cross Abstract: Real-world fine manipulation, particularly in bimanual manipulation, typically requires low-latency control and stable visual localization, while coll

Multi-frame Restoration for High-rate Lissajous Confocal Laser Endomicroscopy

Model ReleasesDGX agent

arXiv:2605.00527v1 Announce Type: cross Abstract: Lissajous confocal laser endomicroscopy (CLE) is a promising solution for high speed in vivo optical biopsy for handheld scenarios. However, Lissajous

Mutatis Mutandis: Revisiting the Comparator in Discrimination Testing

ApplicationsDGX agent

arXiv:2405.13693v4 Announce Type: replace Abstract: Testing for individual discrimination involves deriving a profile, the comparator, similar to the one making the discrimination claim, the complaina

Near-optimal and Efficient First-Order Algorithm for Multi-Task Learning with Shared Linear Representation

ResearchDGX agent

arXiv:2605.00473v1 Announce Type: new Abstract: Multi-task learning (MTL) has emerged as a pivotal paradigm in machine learning by leveraging shared structures across multiple related tasks. Despite i

Network Digital Untwinning: Towards Backward Optimization of Digital Twins

ApplicationsDGX agent

arXiv:2605.00169v1 Announce Type: cross Abstract: Network digital twins (NDTs) are transforming network management by offering precise virtual replicas of physical network systems. However, their reli

NLPOpt-Net: A Learning Method for Nonlinear Optimization with Feasibility Guarantees

Local AiDGX agent

arXiv:2605.00260v1 Announce Type: new Abstract: Nonlinear Parametric Optimization Network (NLPOpt-Net) is an unsupervised learning architecture to solve constrained nonlinear programs (NLP). Given the

NonZero: Interaction-Guided Exploration for Multi-Agent Monte Carlo Tree Search

Local AiDGX agent

arXiv:2605.00751v1 Announce Type: new Abstract: Monte Carlo Tree Search (MCTS) scales poorly in cooperative multi-agent domains because expansion must consider an exponentially large set of joint acti

NorBERTo: A ModernBERT Model Trained for Portuguese with 331 Billion Tokens Corpus

Model ReleasesDGX agent

arXiv:2605.00086v1 Announce Type: new Abstract: High-quality corpora are essential for advancing Natural Language Processing (NLP) in Portuguese. Building on previous encoder-only models such as BERTi

NRGPT: An Energy-based Alternative for GPT

ResearchDGX agent

arXiv:2512.16762v3 Announce Type: replace Abstract: Generative Pre-trained Transformer (GPT) architectures are the most popular design for language modeling. Energy-based modeling is a different parad

Observable Performance Does Not Fully Reflect System Organization: A Multi-Level Analysis of Gait Dynamics Under Occlusal Constraint

ResearchDGX agent

arXiv:2605.00778v1 Announce Type: new Abstract: In biomechanical systems, observable performance is often used as a proxy for underlying system organization. However, this assumption implicitly presum

Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning

ResearchDGX agent

arXiv:2605.00347v1 Announce Type: cross Abstract: Given the rapidly growing capabilities of vision-language models (VLMs), extending them to interactive decision-making tasks such as video games has e

On the Expressive Power of Contextual Relations in Transformers

ResearchDGX agent

arXiv:2603.25860v2 Announce Type: replace-cross Abstract: Transformer architectures have achieved remarkable empirical success in modeling contextual relations, yet a clear understanding of their expr

On the Role of Artificial Intelligence in Human-Machine Symbiosis

AgentsDGX agent

arXiv:2605.00440v1 Announce Type: cross Abstract: The evolution of artificial intelligence (AI) has rendered the boundary between humanity and computational machinery increasingly ambiguous. In the pr

One Good Source is All You Need: Near-Optimal Regret for Bandits under Heterogeneous Noise

ApplicationsDGX agent

arXiv:2602.14474v2 Announce Type: replace Abstract: We study K-armed Multiarmed Bandit (MAB) problem with M heterogeneous data sources, each exhibiting unknown and distinct noise variances {sigma_j^2}

Online Self-Calibration Against Hallucination in Vision-Language Models

SafetyDGX agent

arXiv:2605.00323v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) often suffer from hallucinations, generating descriptions that include visual details absent from the input image.

Optimal hypersurface decision trees

ResearchDGX agent

arXiv:2509.12057v3 Announce Type: replace Abstract: The study of optimal decision trees has gained increasing attention in recent years; however, despite substantial progress, it still suffers from tw

Optimal Spatio-Temporal Decoupling for Bayesian Conformal Prediction

SafetyDGX agent

arXiv:2605.00432v1 Announce Type: new Abstract: Online Conformal Prediction (CP) struggles to balance temporal adaptability and structural stability. Feedback-driven methods (e.g., Adaptive Conformal

Optimizing Resource-Constrained Non-Pharmaceutical Interventions for Multi-Cluster Outbreak Control Using Hierarchical Reinforcement Learning

SafetyDGX agent

arXiv:2603.19397v2 Announce Type: replace Abstract: Non-pharmaceutical interventions (NPIs), such as diagnostic testing and quarantine, are crucial for controlling infectious disease outbreaks but are

OTSS: Output-Targeted Soft Segmentation for Contextual Decision-Weight Learning

Model ReleasesDGX agent

arXiv:2605.00193v1 Announce Type: new Abstract: Many machine learning systems make constrained decisions by optimizing factorized objectives, but the context-specific objective is often treated as fix

Paired-CSLiDAR: Height-Stratified Registration for Cross-Source Aerial-Ground LiDAR Pose Refinement

Model ReleasesDGX agent

arXiv:2605.00634v1 Announce Type: cross Abstract: We introduce Paired-CSLiDAR (CSLiDAR), a cross-source aerial-ground LiDAR benchmark for single-scan pose refinement: refining a ground-scan pose withi

PAMod: Modeling Cyclical Shifts via Phase-Amplitude Modulation for Non-stationary Time Series Forecasting

ApplicationsDGX agent

arXiv:2605.00466v1 Announce Type: new Abstract: Real-world time series forecasting faces the fundamental challenge of non-stationary statistical properties, including shifts in mean and variance over

← Previous
1…792793794795796…998
Next →