AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
Human
88,343Total entries
1Added by human
88,342Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
9 Jun 2026

Multi-Turn Evaluation of Deep Research Agents Under Process-Level Feedback

AgentsDGX agent

arXiv:2606.09748v1 Announce Type: new Abstract: Existing benchmarks for deep research agents (DRAs) assess only single-shot outputs, ignoring a key question: can DRAs improve their reports when guided

Multi-View Speech Representation Learning for Parkinson's Disease Detection Using Context-guided Cross-modal Attention

ApplicationsDGX agent

arXiv:2606.09271v1 Announce Type: cross Abstract: Parkinson's disease (PD) is a progressive neurodegenerative disorder that frequently causes speech impairments associated with hypokinetic dysarthria.

Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers

SafetyDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.12263v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) integrate visual and textual knowledge into unified representations that increasingly underpin modern retrieval

Multimodal Group Emotion Recognition In-the-Wild Towards a Privacy-Safe Non-Individual Approach

ApplicationsDGX agent

arXiv:2606.07585v1 Announce Type: cross Abstract: This thesis addresses group emotion recognition (GER) in-the-wild with a focus on privacy preservation. Unlike traditional emotion recognition methods

Multimodal Large Language Models as Synthetic Participants in Video-Based Studies: An Evaluation

Model ReleasesDGX agent

arXiv:2606.07541v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have shown strong performance on objective tasks such as video understanding and reasoning. However, it remai

Muon Learns More Robust and Transferable Features than Adam

ResearchDGX agent

arXiv:2606.09658v1 Announce Type: cross Abstract: Muon has recently emerged as a state-of-the-art optimizer for pretraining Large Language Models (LLMs) and vision classifiers. Despite its efficiency

Muses: Designing, Composing, Generating Nonexistent Fantasy 3D Creatures without Training

SafetyDGX agent

arXiv:2601.03256v2 Announce Type: replace Abstract: We present Muses, the first training-free method for fantastic 3D creature generation in a feed-forward paradigm. Previous methods, which rely on pa

Need We Teach Foundation Models What is a Generative Image? Gradient-Free Generative Artifact Detection via Analytic Spectral Adaptation

Local AiDGX agent

arXiv:2606.07660v1 Announce Type: new Abstract: Adapting foundation models to detect generative artifacts via gradient-based updates compromises their intrinsic representations. Under optimization on

Neural Field Tokenizations with Hierarchy and Spatial Locality Priors

TutorialsDGX agent

arXiv:2606.08204v1 Announce Type: cross Abstract: Neural fields parameterize data as functions from coordinates to values, providing a unified framework for representation learning across modalities.

Neural Legendre-Fenchel transform with Hessian Preconditioning

TutorialsDGX agent

arXiv:2606.09077v1 Announce Type: new Abstract: The Legendre-Fenchel (LF) transform is a fundamental tool in convex analysis and machine learning that maps lower semi-continuous functions to their con

Neuro-Symbolic Injection of LTLf Constraints in Autoregressive Reinforcement Learning Policies

SafetyDGX agent

arXiv:2606.08312v1 Announce Type: new Abstract: In this work we study offline reinforcement learning (RL) under temporally extended task constraints expressed in Linear Temporal Logic over finite trac

NeuroAlign: Hierarchical Multimodal Fusion of Dynamic and Structural Neuroimaging for MCI Analysis

SafetyDGX agent

arXiv:2606.07635v1 Announce Type: cross Abstract: Multimodal neuroimaging fusion of functional MRI (fMRI) and diffusion tensor imaging (DTI) provides complementary information for cognitive impairment

Neutrality Bites: Gender Representation in AI-Generated Animal Stories

SafetyDGX agent

arXiv:2606.07969v1 Announce Type: cross Abstract: Gender bias in AI-generated stories is a well-documented problem. While much attention has been paid to reducing or mitigating this bias, it is not al

New Fractional Ambiguity Function Integrated with CNN-Based Machine Learning for Signal Classification

ResearchDGX agent

arXiv:2606.08110v1 Announce Type: cross Abstract: A new fractional ambiguity function (NFrAF) derived from the fractional Fourier transform is introduced as a generalization of the classical ambiguity

Next-Token Prediction Learns Generalisable Representations of Sleep Physiology

ApplicationsDGX agent

arXiv:2606.09605v1 Announce Type: new Abstract: Foundation models offer a promising route to compress multi-modal physiological signals into compact representations of human health, with broad applica

NGram-MoSE: Efficient Remote Sensing Super-Resolution via N-Gram Context and Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2606.08535v1 Announce Type: new Abstract: Remote sensing applications for environmental monitoring and disaster management are frequently constrained by a spatial--temporal trade-off: imagery wi

No Free Lunch for Synthetic Images under Data Scarcity Conditions

ResearchDGX agent

arXiv:2606.07640v1 Announce Type: cross Abstract: This study investigates the trade-offs between fidelity, privacy, and utility in synthetic data generation under conditions of data scarcity and priva

No Modality Left Behind: Adapting to Missing Modalities via Knowledge Distillation for Brain Tumor Segmentation

SafetyDGX agent

arXiv:2509.15017v2 Announce Type: replace Abstract: Accurate brain tumor segmentation is essential for preoperative evaluation and personalized treatment. Multi-modal MRI is widely used due to its abi

Noise-Adaptive High-Probability Regret Bounds for Online Convex Optimization

ResearchDGX agent

arXiv:2606.08028v1 Announce Type: new Abstract: We study high-probability regret bounds for online convex optimization (OCO) with strongly convex losses and establish three results that resolve open q

Non-Archimedean Polydisc Spaces and Applications to Optimisation

ResearchDGX agent

arXiv:2606.07782v1 Announce Type: cross Abstract: We propose a new framework for optimisation over non-Archimedean spaces inspired by Berkovich geometry. Specifically, we introduce polydisc spaces, wh

Nonparametric LLM Evaluation from Preference Data

ApplicationsDGX agent

arXiv:2601.21816v2 Announce Type: replace Abstract: Evaluating the performance of large language models (LLMs) from human preference data is crucial for obtaining LLM leaderboards. However, many exist

NoRD: A Data-Efficient Vision-Language-Action Model that Drives without Reasoning

SafetyDGX agent

arXiv:2602.21172v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are advancing autonomous driving by replacing modular pipelines with unified end-to-end architectures. However,

Normality Calibration in Semi-supervised Graph Anomaly Detection

SafetyDGX agent

arXiv:2510.02014v3 Announce Type: replace Abstract: Graph anomaly detection (GAD) has attracted growing interest for its crucial ability to uncover irregular patterns in broad applications. Semi-super

Not Just After One: Sleep-Inspired Replay Prevents Catastrophic Forgetting After Sequential Tasks

TutorialsDGX agent

arXiv:2606.08447v1 Announce Type: cross Abstract: One of the critical limitations of artificial neural networks is their lack of ability to continually learn: training on new tasks often leads to inte

Now You (Still) See Me: Detecting Evasive Steganographic Payloads in LLMs

Model ReleasesDGX agent

arXiv:2606.09411v1 Announce Type: cross Abstract: Large language models can be fine-tuned to encode prompt-borne secrets into fluent, seemingly benign outputs. This creates a steganographic exfiltrati

NutriMLLM: Multimodal Large Language Models for Dietary Micronutrient Analysis

Model ReleasesDGX agent

arXiv:2606.08948v1 Announce Type: cross Abstract: Comprehensive estimation of dietary micronutrients from food images could improve clinical nutrition care, but training such models requires large mul

OASIS: From Simulation Data Collection to Real-World Humanoid Loco-Manipulation

SafetyDGX agent

arXiv:2606.08548v1 Announce Type: new Abstract: Recent progress in robot manipulation has been largely driven by learning from large-scale demonstrations. For humanoid robot loco-manipulation tasks, h

Observability for Delegated Execution in Agentic AI Systems

AgentsDGX agent

arXiv:2606.09692v1 Announce Type: cross Abstract: Delegation-scoped execution is not identifiable from standard observables: audit logs and execution traces can be identical under multiple incompatibl

Observation-driven correction of numerical weather prediction for marine winds

Local AiDGX agent

arXiv:2512.03606v2 Announce Type: replace Abstract: Accurate marine wind forecasts are essential for safe navigation, ship routing, and energy operations, yet they remain challenging because observati

OctaOctree Neural Radiosity for Real-time Glossy Material Rendering

ResearchDGX agent

arXiv:2606.08469v1 Announce Type: cross Abstract: Modeling high-frequency outgoing radiance distributions remains a fundamental challenge in global illumination, especially for glossy and specular mat

Offline Reinforcement Learning for Plasma Control in Nuclear Fusion: Codebase and Benchmark

Model ReleasesDGX agent

arXiv:2606.07550v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) offers a promising route for developing plasma controllers from historical tokamak data, since online trial-and-er

omega-EVA: Envision, Verify, and Act with Latent Interactive World Models

SafetyDGX agent

arXiv:2606.09457v1 Announce Type: new Abstract: Embodied policies typically map current observations directly to actions, leaving candidate-action consequences implicit. World models provide predictiv

OmniCap-IF: Benchmarking and Improving Instruction Following Abilities for Omni-Video Captioning

Model ReleasesDGX agent

arXiv:2606.08572v1 Announce Type: new Abstract: While Omni-modal Large Language Models (OLLMs) have demonstrated impressive capabilities in jointly processing audio and visual streams, their ability t

OmniFaceRig: Fully Automatic Inner-Mouth-Aware Face Rigging Across Diverse 3D Character Topologies

Model ReleasesDGX agent

arXiv:2606.08043v1 Announce Type: cross Abstract: Facial rigging - creating FACS-based blendshapes together with inner-mouth geometry (teeth, gums, and tongue) - remains a major bottleneck in 3D chara

OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics

Model ReleasesDGX agent

arXiv:2606.09826v1 Announce Type: cross Abstract: Vision-language model (VLM) agents are increasingly deployed in interactive game environments. Yet game benchmarks for VLM agents typically report a s

OmniGen-AR: AutoRegressive Any-to-Image Generation

Model ReleasesDGX agent

arXiv:2606.09156v1 Announce Type: new Abstract: Autoregressive (AR) models have demonstrated strong potential in visual generation, offering superior performance with simple architectures and optimiza

OmniMem: Perturbation-aware Memory Compression for Streaming Audio-Visual LLMs

Model ReleasesDGX agent

arXiv:2606.07577v1 Announce Type: new Abstract: Audio-visual large language models (LLMs) hold strong promise for long-form video understanding, yet their long-video inference is fundamentally limited

OmniTryOn: Video Try-On Anything at Once!

Model ReleasesDGX agent

arXiv:2606.08514v1 Announce Type: new Abstract: Although video virtual try-on (VVT) has achieved significant progress, existing methods still exhibit two fundamental limitations: first, they are restr

On Choosing the mu Parameter in Gaussian Differential Privacy

Model ReleasesDGX agent

arXiv:2606.09582v1 Announce Type: new Abstract: Recent work argues for using Gaussian differential privacy (GDP) to report the privacy guarantees in privacy-preserving machine learning. We provide pri

On solving symmetric multi-type orthogonal non-negative matrix tri-factorization problem

ResearchDGX agent

arXiv:2606.08291v1 Announce Type: new Abstract: We study the symmetric multi-type orthogonal non-negative matrix tri-factorization problem, where several symmetric non-negative matrices are simultaneo

On the Complexity of Offline Reinforcement Learning with Q^star-Approximation and Partial Coverage

ResearchDGX agent

arXiv:2602.12107v2 Announce Type: replace-cross Abstract: We study offline reinforcement learning under Q^star-approximation and partial coverage, a setting that motivates practical algorithms such as

On-the-fly hand-eye calibration for the da Vinci surgical robot

SafetyDGX agent

arXiv:2601.14871v2 Announce Type: replace Abstract: In Robot-Assisted Minimally Invasive Surgery (RMIS), accurate tool localization is crucial to ensure patient safety and successful task execution. H

On the Superlinear Relationship between SGD Noise Covariance and Loss Landscape Curvature

ResearchDGX agent

arXiv:2602.05600v2 Announce Type: replace Abstract: Stochastic Gradient Descent (SGD) introduces anisotropic noise that is correlated with the local curvature of the loss landscape, thereby biasing op

On the Wasserstein Geodesic Principal Component Analysis of probability measures

TutorialsDGX agent

arXiv:2506.04480v2 Announce Type: replace-cross Abstract: This paper focuses on Geodesic Principal Component Analysis (GPCA) on a collection of probability distributions using the Otto-Wasserstein geo

One if by Land, Two if by Sea, Three if by Four Seas, and More to Come -- Values of Perception, Prediction, Communication, and Common Sense in Decision Making

AgentsDGX agent

arXiv:2601.06077v2 Announce Type: replace-cross Abstract: This work aims to rigorously define the values of perception, prediction, communication, and common sense in decision making. The defined quan

One Stone, Three Birds: Self-adaptive Optimal Transport for Multi-VLM Selection, Adaptation, and Ensembling

ResearchDGX agent

arXiv:2606.08126v1 Announce Type: new Abstract: Vision-language models (VLMs) enable visual recognition from semantic class descriptions, which makes them attractive when target annotations are scarce

Online Agent-as-a-Judge: Situation-Generating Evaluation for Interactive Agents

AgentsDGX agent

arXiv:2606.08200v1 Announce Type: new Abstract: Evaluating LLM-powered interactive social agents is challenging because socially relevant behaviors depend not only on isolated outputs, but also on pri

Online Learning for Supervisory Switching Control

ApplicationsDGX agent

arXiv:2603.14762v3 Announce Type: replace-cross Abstract: We study supervisory switching control for partially-observed linear dynamical systems. The objective is to identify and deploy a suitable con

Online Learning with Recency: Algorithms for Sliding-window Streaming Multi-armed Bandits

Model ReleasesDGX agent

arXiv:2606.08977v1 Announce Type: new Abstract: Motivated by the recency effect in online learning, we study algorithms for single-pass *sliding-window streaming multi-armed bandits (MABs)* in this pa

OnlyDense: Reduced-Order Modeling for Lagrangian simulation

ResearchDGX agent

arXiv:2606.09065v1 Announce Type: cross Abstract: In science and engineering, Lagrangian simulation methods such as Smooth Particle Hydrodynamics (SPH) or Material Point Method (MPM) are often employe

Operationalising the Superficial Alignment Hypothesis via Task Complexity

SafetyDGX agent

arXiv:2602.15829v2 Announce Type: replace Abstract: The superficial alignment hypothesis (SAH) posits that large language models learn most of their knowledge during pre-training, and that post-traini

Operator learning for solving Fokker-Planck equations with various initial conditions

ResearchDGX agent

arXiv:2606.09434v1 Announce Type: new Abstract: The Fokker-Planck equation (FPE) plays a pivotal role in describing the time evolution of probability density functions (PDFs) for systems governed by s

Operator learning for the 2D incompressible Navier-Stokes equations: a conformal prediction approach in the data-scarce regime

Model ReleasesDGX agent

arXiv:2606.08654v1 Announce Type: new Abstract: In this paper, we propose a perturbation-based conformal prediction framework for uncertainty quantification in operator learning, with a focus on the 2

Optical Music Recognition for Real-World Manuscripts with Synthetic Data

ApplicationsDGX agent

arXiv:2606.09479v1 Announce Type: new Abstract: Optical Music Recognition (OMR) has seen major progress in model design, with end-to-end methods now capable of recognising notation at all levels of co

Optical Reasoning: Rethinking Images as an Expressive Reasoning Medium Beyond Text

ResearchDGX agent

arXiv:2606.09585v1 Announce Type: new Abstract: Chain-of-Thought (CoT) improves the performance of Large Language Models (LLMs) and has been extended to Multimodal Large Language Models (MLLMs). More

Optimal and Provable Calibration in High-Dimensional Binary Classification: Angular Calibration and Platt Scaling

ResearchDGX agent

arXiv:2502.15131v4 Announce Type: replace-cross Abstract: We study the fundamental problem of calibrating a linear binary classifier of the form sigma(hat{w}^op x), where the feature vector x is Gauss

Optimal Fair Aggregation of Crowdsourced Noisy Labels using Demographic Parity Constraints

SafetyDGX agent

arXiv:2601.23221v2 Announce Type: replace Abstract: As acquiring reliable ground-truth labels is usually costly, or infeasible, crowdsourcing and aggregation of noisy human annotations is the typical

Optimality of Sequential Filtering Under Independent Cost and Selectivity Models

ResearchDGX agent

arXiv:2606.07589v1 Announce Type: new Abstract: Sequential filtering pipelines are a common design pattern in large-scale systems, where a large population of items is progressively reduced by a seque

Optimizing Energy-based Neural Network Training with Coherent Ising Machine

ResearchDGX agent

arXiv:2606.09117v1 Announce Type: cross Abstract: While Ising machines serve as advanced physical solvers for the Ising model,enabling applications in combinatorial optimization and neural network tra

Optimizing Few-Step Generation with Adaptive Matching Distillation

ResearchDGX agent

arXiv:2602.07345v2 Announce Type: replace Abstract: Distribution Matching Distillation (DMD) is a powerful acceleration paradigm, yet its stability is often compromised in Forbidden Zone, regions wher

← Previous
1…449450451452453…1049
Next →