AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
Human
86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
2 Jul 2026

MetaHOPE: A Metaphor-Oriented Evaluation Framework for Analysing MT and LLM Translation Errors

ResearchDGX agent

arXiv:2607.00848v1 Announce Type: new Abstract: In this opinion paper, we propose MetaHOPE, an error severity-aware annotation framework for evaluating metaphor translations. Metaphors present challen

MetaOthello: A Controlled Study of Multiple World Models in Transformers

TutorialsDGX agent

arXiv:2602.23164v2 Announce Type: replace Abstract: Foundation models must handle multiple generative processes, yet mechanistic interpretability largely studies capabilities in isolation; it remains

MG-RWKV: Multi-Grained Context-Aware RWKV for Temporal Forgery Localization

ResearchDGX agent

arXiv:2607.00902v1 Announce Type: new Abstract: Driven by Artificial Intelligence-Generated Content (AIGC), the authenticity of audio-visual content is facing severe challenges. Temporal Forgery Local

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MG-SpaIR: Multi-grade Sparse-guided Implicit Representation for Training-Data-Free Image Restoration

ResearchDGX agent

arXiv:2607.00138v1 Announce Type: new Abstract: MG-SpaIR is a training-data-free framework for restoring a clean image from a single observation corrupted by a mixture of blur, downsampling, noise, an

MindAU: EEG-Conditioned Facial Action Unit Editing via Dual-Stream Manifold Alignment

Model ReleasesDGX agent

arXiv:2607.00410v1 Announce Type: new Abstract: Recent brain decoding studies have made substantial progress in reconstructing externally perceived visual content from neural signals. However, using e

MindEdit-Bench: Benchmarking Object-Level Counterfactual Spatial Reasoning in VLMs from In-the-Wild Photos

Model ReleasesDGX agent

arXiv:2607.00491v1 Announce Type: cross Abstract: Benchmarks for vision-language models (VLMs) mostly test observational spatial reasoning: models describe relations already visible in the input. Exis

MineRobot: An Actuator-Centered Kinematic Modeling and Solving Framework for Underground Mining Robots

ResearchDGX agent

arXiv:2603.22055v2 Announce Type: replace-cross Abstract: Underground mining robots are increasingly modeled for planning, operator training, and digital-twin workflows, where reliable actuator-level

Mirror-Fusion Attention for Reflection-Aware Self-Supervised Representation Learning

SafetyDGX agent

arXiv:2607.00850v1 Announce Type: new Abstract: Most self-supervised learning (SSL) methods encourage invariance across augmentations, but strict flip invariance can suppress informative left--right c

Mixture of Distributions Matters: Dynamic Sparse Attention for Efficient Video Diffusion Transformers

ResearchDGX agent

arXiv:2601.11641v3 Announce Type: replace Abstract: While Diffusion Transformers (DiTs) have achieved notable progress in video generation, this long-sequence generation task remains constrained by th

MMLoP: Multi-Modal Low-Rank Prompting for Efficient Vision-Language Adaptation

Model ReleasesDGX agent

arXiv:2602.21397v2 Announce Type: replace Abstract: Prompt learning has become a dominant paradigm for adapting vision-language models (VLMs) such as CLIP to downstream tasks without modifying pretrai

Mnemosyne: Agentic Transaction Processing for Validating and Repairing AI-generated Workflows

SafetyDGX agent

arXiv:2607.00269v1 Announce Type: new Abstract: LLMs, solvers, and agent teams increasingly generate workflow actions, repairs, and plans, but a generated action may be syntactically valid yet stale,

MoHallBench: A Benchmark for Motion Hallucination in Video Large Language Models

Model ReleasesDGX agent

arXiv:2607.01117v1 Announce Type: new Abstract: Video Large Language Models (VideoLLMs) have shown strong progress in video understanding, yet they still suffer from hallucinations that are inconsiste

Moire Video Authentication: A Physical Signature Against AI Video Generation

ResearchDGX agent

arXiv:2604.01654v2 Announce Type: replace-cross Abstract: Recent advances in video generation have made AI-synthesized content increasingly difficult to distinguish from real footage. We propose a phy

MolSafeEval: A Benchmark for Uncovering Safety Risks in AI-Generated Molecules

Model ReleasesDGX agent

arXiv:2607.00464v1 Announce Type: cross Abstract: Current molecular generation benchmarks emphasize task complexity, molecule novelty, and property alignment; they largely overlook a critical concern:

MonoMSK: Monocular 3D Musculoskeletal Dynamics Estimation

ResearchDGX agent

arXiv:2511.19326v2 Announce Type: replace Abstract: Reconstructing biomechanically realistic 3D human motion - recovering both kinematics (motion) and kinetics (forces) - is a critical challenge. Whil

MosaicKV: Serving Long-Context LLM with Dynamic Two-D KV Cache Compression

HardwareDGX agent

arXiv:2607.00760v1 Announce Type: new Abstract: Long-context LLM services now sustain prompts with hundreds of thousands to millions of tokens, making the key-value (KV) cache a first-order serving co

MoVA: Learning Asymmetric Dual Projections for Modular Long Video-Text Alignment

SafetyDGX agent

arXiv:2607.00858v1 Announce Type: new Abstract: Contrastive pre-training has propelled video-text alignment, yet models often inherit the critical limitations of their image-text predecessors like CLI

MSQA: A Natively Sourced Multilingual and Multicultural SimpleQA Benchmark

Model ReleasesDGX agent

arXiv:2607.00724v1 Announce Type: new Abstract: Multilingual fluency often invites a stronger assumption: a model that can speak a user's language must also understand the culture encoded by that lang

Multi-Embodiment Robotic Retargeting via Guided Diffusion Model

ResearchDGX agent

arXiv:2505.20857v2 Announce Type: replace Abstract: Motion retargeting for specific robot from existing motion datasets is one critical step in transferring motion patterns from human behaviors to and

Multi-Hypothesis Test-Time Adaptation to Mitigate Underspecification

Model ReleasesDGX agent

arXiv:2607.00259v1 Announce Type: cross Abstract: Test-Time Adaptation (TTA) seeks to improve model robustness under distribution shifts by adapting parameters using unlabeled target data. However, in

Multi-Label Node Classification with Label Influence Propagation

Model ReleasesDGX agent

arXiv:2607.00671v1 Announce Type: cross Abstract: Graphs are a complex and versatile data structure used across various domains, with possibly multi-label nodes playing a particularly crucial role. Ex

Multi-scale Mixture of World Models for Embodied Agents in Evolving Environments

SafetyDGX agent

arXiv:2607.00457v1 Announce Type: new Abstract: Embodied agents operating in the real world require multi-scale reasoning and knowledge adaptation as conditions change. We identify two challenges in a

Multi-Turn Agentic Scientific Literature Search via Workflow Induction

AgentsDGX agent

arXiv:2607.00597v1 Announce Type: new Abstract: Scientific literature search often requires more than retrieving papers from a single query: users' intents are underspecified, preference-dependent, an

Multimodal Continuous Reasoning via Asymmetric Mutual Variational Learning

Model ReleasesDGX agent

arXiv:2607.00461v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are often constrained by a language-space bottleneck, forcing complex visual reasoning into discrete tokens whi

Multiplicity is an Inevitable and Inherent Challenge in Multimodal Learning

SafetyDGX agent

arXiv:2505.19614v2 Announce Type: replace-cross Abstract: Multimodal learning has seen remarkable progress, particularly with large-scale pre-training across various modalities. Most current approache

MultiSynt/MT: Trillion-Token Multi-Parallel Pre-Training Data Translated Across 36 Languages

Model ReleasesDGX agent

arXiv:2607.00890v1 Announce Type: new Abstract: Open web-scale pre-training corpora remain concentrated in English, limiting multilingual LLM development. We introduce MultiSynt/MT, an open synthetic

Muon as a Residual Connection

TutorialsDGX agent

arXiv:2607.01124v1 Announce Type: cross Abstract: Muon has recently emerged as one of the most effective optimizers for training large neural networks, yet its empirical success has been explained fro

MVDGC: Joint 3D and 2D Multi-view Pedestrian Detection via Dual Geometric Constraints

ResearchDGX agent

arXiv:2607.00273v1 Announce Type: new Abstract: The core challenge in multi-view pedestrian detection (MVPD) lies in effective aggregation of visual features from different viewpoints for robust occlu

NeHMO: Neural Hamilton-Jacobi Reachability Learning for Decentralized Safe Multi-Arm Motion Planning

SafetyDGX agent

arXiv:2607.00326v1 Announce Type: new Abstract: Safe multi-arm motion planning is a challenging problem in robotics due to its high dimensionality, coupled configuration space, and complex collision c

Neural Certificate Pricing for Combinatorial Optimization Problems

ResearchDGX agent

arXiv:2607.01185v1 Announce Type: new Abstract: Combinatorial optimization (CO) problems are difficult because certifiable discrete structure induces exponential search. One needs to search over the s

Neural Network-Based Estimation of Time-Dependent Parameters in AR(p) Processes

Model ReleasesDGX agent

arXiv:2607.00470v1 Announce Type: cross Abstract: We investigate a forecasting framework based on a simple discrete-time dynamic model with coefficients varying in time. The parameters of the model ar

Neural Surface and Reflectance Modelling from 3D Radar Data

AgentsDGX agent

arXiv:2603.25623v2 Announce Type: replace Abstract: Robust scene representation is essential for autonomous systems to safely operate in challenging low-visibility environments. In these conditions, r

NeuroCogMap Reveals Cognitive Organization of Large Language Models

SafetyDGX agent

arXiv:2607.00397v1 Announce Type: cross Abstract: Understanding how complex cognitive functions are organized within artificial systems is central to interpreting large language models (LLMs) and rela

NeuroFilter: Activation-Based Guardrails for Privacy-Conscious LLM Agents

AgentsDGX agent

arXiv:2601.14660v2 Announce Type: replace-cross Abstract: Agentic Large Language Models (LLMs) are models able to reason, plan, and execute tools over unstructured data. These abilities are enabling t

Next-Frame Decoding for Ultra-Low-Bitrate Image Compression with Video Diffusion Priors

Model ReleasesDGX agent

arXiv:2603.15129v3 Announce Type: replace Abstract: We present a novel paradigm for ultra-low-bitrate image compression (ULB-IC) that exploits the ``temporal'' evolution in generative image compressio

NI-Tex: Non-isometric Image-based Garment Texture Generation

ApplicationsDGX agent

arXiv:2511.18765v3 Announce Type: replace-cross Abstract: Existing industrial 3D garment meshes already cover most real-world clothing geometries, yet their texture diversity remains limited. To acqui

NoPA: Non-Parametric Online 3D Scene Graph Generation

ResearchDGX agent

arXiv:2607.00529v1 Announce Type: new Abstract: Classic 3D scene graph generation approaches fail to work in real-time due to the heavy computational cost of environment mapping and the need to genera

Not All Prediction Targets Keep Training-Free Diffusion Guidance on the Manifold

Model ReleasesDGX agent

arXiv:2607.00647v1 Announce Type: new Abstract: Training-free guidance (TFG) steers a pretrained diffusion model toward a desired attribute at inference. To be effective, this guidance must be applied

NOVA: Next-step Open-Vocabulary Autoregression for 3D Multi-Object Tracking in Autonomous Driving

AgentsDGX agent

arXiv:2603.06254v2 Announce Type: replace Abstract: Generalizing across unknown targets is critical for open-world perception, yet existing 3D Multi-Object Tracking (3D MOT) pipelines remain limited b

OmniFall: From Staged Through Synthetic to Wild, A Unified Multi-Domain Dataset for Robust Fall Detection

Model ReleasesDGX agent

arXiv:2505.19889v3 Announce Type: replace Abstract: Visual fall detection models are usually trained on small, staged datasets. Their real-world utility remains unclear; such data lacks diversity and

OmniMoE: An Efficient MoE by Orchestrating Atomic Experts at Scale

Model ReleasesDGX agent

arXiv:2602.05711v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) architectures are evolving towards finer granularity to improve parameter efficiency. However, existing MoE designs f

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping

SafetyDGX agent

arXiv:2607.00881v1 Announce Type: new Abstract: Spatial intelligence remains a persistent challenge for Multimodal Large Language Models (MLLMs), as it requires coherent spatial scene representations

On the Reliability of Cue Conflict and Beyond

Model ReleasesDGX agent

arXiv:2603.10834v4 Announce Type: replace-cross Abstract: Understanding how neural networks rely on visual cues offers a human-interpretable view of their internal decision processes. The cue-conflict

OnPoint: Offline-to-Online Multi-Level Distillation for Point-Supervised Online Temporal Action Localization

ResearchDGX agent

arXiv:2607.00289v1 Announce Type: new Abstract: Temporal Action Localization (TAL) typically relies on segment annotations or offline access to full videos, limiting scalability and online use. We int

OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning

SafetyDGX agent

arXiv:2510.24636v3 Announce Type: replace Abstract: Reward models (RMs) have become essential for aligning large language models (LLMs), serving as scalable proxies for human evaluation in both traini

OpFML: Pipeline for ML-based Operational Inference

ResearchDGX agent

arXiv:2601.11046v2 Announce Type: replace Abstract: Machine learning models for climate and Earth science are becoming increasingly capable, yet model deployment into operational use remains a largely

Optimal any-angle path planning in static and dynamic environments

ResearchDGX agent

arXiv:2607.00065v1 Announce Type: cross Abstract: Any-angle path planning extends traditional graph-based path planning by allowing movement between any pair of vertices, rather than being restricted

Optimal Resource Utilization for Autonomous Laboratory Orchestrators

AgentsDGX agent

arXiv:2607.01188v1 Announce Type: new Abstract: In autonomous laboratories, AI agents suggest the next batch of experiments to do. However, planning and executing those tasks taking full advantage of

Optimal scaling of MCMC algorithms: exploiting the symmetry of the Metropolis-Hastings formula

TutorialsDGX agent

arXiv:2607.00586v1 Announce Type: cross Abstract: We present a simple, yet general approach to study the scaling properties as the dimensionality of Metropolised MCMC sampling algorithms increases. Th

Optimization on the Oblique Manifold for Sparse Simplex Constraints via Multiplicative Updates

ResearchDGX agent

arXiv:2503.24075v4 Announce Type: replace-cross Abstract: Low-rank optimization problems with sparse simplex constraints involve variables that must satisfy nonnegativity, sparsity, and sum-to-1 condi

OSCAR: Occupancy-based Shape Completion via Acoustic Neural Implicit Representations

Model ReleasesDGX agent

arXiv:2603.08279v2 Announce Type: replace Abstract: Accurate 3D reconstruction of vertebral anatomy from ultrasound is important for guiding minimally invasive spine interventions, but it remains chal

Pano2World: End-to-End 3D Generation via Unified Multi-View Sequences

Model ReleasesDGX agent

arXiv:2607.00832v1 Announce Type: cross Abstract: A single panorama captures the full visual sphere from one camera center, yet confines users to looking around in place without enabling true scene ex

PanoGrounder: Bridging 2D and 3D with Panoramic Scene Representations for VLM-based 3D Visual Grounding

ResearchDGX agent

arXiv:2512.20907v2 Announce Type: replace Abstract: 3D Visual Grounding (3DVG) is a critical bridge from vision-language perception to robotics, requiring both language understanding and 3D scene reas

PAPA: Online Personalized Active Preference Alignment

SafetyDGX agent

arXiv:2607.00486v1 Announce Type: cross Abstract: Diffusion models are highly effective at modeling complex data distributions, including images and text. However, in applications like personalized re

Partial Skeleton Visibility for Action Recognition: A Constrained Field-of-View Approach

ApplicationsDGX agent

arXiv:2607.00716v1 Announce Type: cross Abstract: Skeleton-based action recognition has achieved remarkable success by exploiting joint coordinates and their topological connections, yet prevailing me

Path Planning in Physically Viable World Models

ResearchDGX agent

arXiv:2607.00673v1 Announce Type: new Abstract: Robots deployed in unstructured outdoor environments often plan from scene reconstructions collected before deployment because operators cannot remap la

PedNStream: Scalable Network Flow Simulation for Pedestrian Traffic Management

ApplicationsDGX agent

arXiv:2607.01021v1 Announce Type: new Abstract: Large-scale crowd management requires pedestrian simulations that are both computationally efficient and compatible with feedback-based control. However

Perceive-to-Reason: Decoupling Perception and Reasoning for Fine-Grained Visual Reasoning

Local AiDGX agent

arXiv:2607.01191v1 Announce Type: new Abstract: Fine-grained visual reasoning remains challenging for vision-language models, especially when small but critical visual cues are buried in high-resoluti

Persona Non Grata: LLM Persona-Driven Generations in MCQA are Unstable in Distinct Dimensions

ResearchDGX agent

arXiv:2607.00937v1 Announce Type: new Abstract: Persona-driven generations (PDGs) have seen prolific use in research and industry applications, where a large language model (LLM) takes on a 'persona'

Persona Without Substrate: Regime-Dependence and the LLM Individuation Problem

Model ReleasesDGX agent

arXiv:2607.00006v1 Announce Type: cross Abstract: Beckmann & Butlin's (2026) ontological framework for the LLM individuation problem inherits an unargued cross-regime co-reference assumption from the

← Previous
1…291292293294295…1025
Next →