AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
Human
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
1 Jul 2026

MECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied Environments

Model ReleasesDGX agent

arXiv:2606.31966v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have strong potential as embodied agents, but their ability to collaborate in visually grounded enviro

Medical Image Spatial Grounding with Semantic Sampling

Model ReleasesDGX agent

arXiv:2603.14579v3 Announce Type: replace Abstract: Vision language models (VLMs) have shown significant promise in visual grounding for images as well as videos. In medical imaging research, VLMs rep

MediEncoder: Nonlinear Representation Learning for High-Dimensional Causal Mediation Analysis

ResearchDGX agent

arXiv:2606.30648v1 Announce Type: cross Abstract: Causal mediation analysis decomposes a treatment effect into indirect pathways through mediators and direct pathways not operating through them. Moder

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MemLearner: Learning to Query Context memory for Video World Models

ApplicationsDGX agent

arXiv:2606.31734v1 Announce Type: new Abstract: Video World Models are interactive video generation models that predict future world states based on user actions and history video frames. A critical c

Mesh BDF: Barycentric Dominance Field for 3D Native Mesh Generation

ResearchDGX agent

arXiv:2606.31777v1 Announce Type: new Abstract: Autoregressive (AR) modeling has recently achieved remarkable progress in native 3D mesh generation, largely due to its natural ability to handle variab

MetricHMSR:Metric Human Mesh and Scene Recovery from Monocular Images

Local AiDGX agent

arXiv:2506.09919v4 Announce Type: replace Abstract: We introduce MetricHMSR, a novel framework for recovering metric human meshes and 3D scenes from a single monocular image. Existing methods struggle

Mind the Residual Gap: Probabilistic Downscaling under Real-World Bias

Model ReleasesDGX agent

arXiv:2606.30821v1 Announce Type: new Abstract: Probabilistic downscaling is the task of modeling the conditional distribution of high-resolution fields given coarse inputs, and is a central challenge

Minimizing Quantized Semantic Age of Information (QSAoI) in Foundation Model-Based Semantic Communications

ResearchDGX agent

arXiv:2606.31303v1 Announce Type: cross Abstract: The emerging techniques of semantic communications and edge computing in 6G networks necessitate a paradigm shift toward co-designed semantic-aware an

MIRTH: Mutual-Information Reasoning with Temporal Hubs for Vision-Language-Action Agents

Model ReleasesDGX agent

arXiv:2606.31167v1 Announce Type: cross Abstract: VLA models have emerged as a powerful paradigm for transferring semantic knowledge from web-scale data to physical robotic control. However, current s

Mitigating Positional Leakage in 3D Masked Autoencoders for Robust Representation Learning

ResearchDGX agent

arXiv:2606.31570v1 Announce Type: cross Abstract: Masked autoencoding has emerged as a prominent paradigm for self-supervised learning on 3D point clouds, achieving competitive performance across down

Mixture-of-Control: State-Aware Fine-Tuning for Transformer-based Models

Model ReleasesDGX agent

arXiv:2606.31397v1 Announce Type: cross Abstract: State-based fine-tuning has emerged as a compelling alternative to weight-based adaptation for transformers, updating lightweight controls into states

MNAR-k-means: A k-means Clustering for Data Missing Not at Random with Magnitude-Decaying Probability

SafetyDGX agent

arXiv:2606.31253v1 Announce Type: cross Abstract: The classical k-means clustering, based on distances computed from all data features, cannot be directly applied to incomplete data with missing value

Modal CEGAR-tableaux with RECAR and resolution-based SAT-shortcuts

ResearchDGX agent

arXiv:2606.31878v1 Announce Type: cross Abstract: We investigate two approaches for extending CEGAR-tableaux with SAT-shortcuts using a previously known approach called RECAR but also a totally new ap

Modality-Driven Search with Holistic Trace Judging for ARC-AGI-2

Model ReleasesDGX agent

arXiv:2606.31543v1 Announce Type: new Abstract: Large language models can produce fluent, internally coherent reasoning traces for abstract reasoning tasks while still being confidently wrong - making

Modeling Cell-Cycle-Aware Single-Cell Drug Perturbation Responses

Model ReleasesDGX agent

arXiv:2606.30695v1 Announce Type: cross Abstract: Single-cell drug perturbation models should predict not only transcriptional response magnitude, but also whether a treatment alters the proliferative

Moral Safety in LLMs: Exposing Performative Compliance with Puzzled Cues

Model ReleasesDGX agent

arXiv:2606.31644v1 Announce Type: new Abstract: As large language models take on morally consequential roles in healthcare, legal, and hiring contexts, we need to examine whether their ethical behavio

Motion Planning in Compressed Representation Spaces

AgentsDGX agent

arXiv:2606.30940v1 Announce Type: cross Abstract: Deep learning methods have vastly expanded the capabilities of motion planning in robotics applications, as learning priors from large-scale data has

MS-Resampler: Multi-Scope Visual Resampling for Efficient Multimodal LLMs

ResearchDGX agent

arXiv:2606.31383v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) typically employ resampling-based projectors to transform dense visual features into a compact token sequence f

MSNN-LINet: Cross-Modal Learning via Continuous Linear Integration

ResearchDGX agent

arXiv:2606.31135v1 Announce Type: new Abstract: We present LINet (Linear Integration Network), a Multi-Stream Neural Network (MSNN) for RGB-D scene classification. Current multi-modal architectures tr

Multi-Channel Uncertainty-Weighted Score Matching for Conditional Diffusion in Medical UDA

ResearchDGX agent

arXiv:2509.22476v2 Announce Type: replace Abstract: Robust medical image segmentation across modalities remains challenging due to severe domain shifts and the lack of target-domain labels. While diff

Multi-Robot Coordination for Planning under Context Uncertainty

TutorialsDGX agent

arXiv:2603.13748v3 Announce Type: replace Abstract: Real-world robots often operate in settings where objective priorities depend on the underlying context of operation. When the underlying context is

Multilingual Polarization Detection Using Transformer-Based Models with Class Weighting and Threshold Tuning

ResearchDGX agent

arXiv:2606.30857v1 Announce Type: new Abstract: This paper describes our submission to SemEval-2026 Task 9 on detecting multilingual, multicultural, and multievent online polarization. We address all

Multimodal Benchmark for Safety Assessment in Industrial Inspection Scenarios

Model ReleasesDGX agent

arXiv:2601.21173v2 Announce Type: replace-cross Abstract: With the rapid development of industrial intelligence and unmanned inspection, reliable perception and safety assessment for AI systems in com

Multiple Testing of Linear Forms for Noisy Matrix Completion

SafetyDGX agent

arXiv:2312.00305v3 Announce Type: replace-cross Abstract: Many important tasks of large-scale recommender systems can be naturally cast as testing multiple linear forms for noisy matrix completion. Th

Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining

Model ReleasesDGX agent

arXiv:2509.10406v4 Announce Type: replace Abstract: Pretraining transformers on long sequences (entire code repositories, collections of related documents) is bottlenecked by quadratic attention costs

Multisensory Continual Learning: Adapting Pretrained Visuomotor Policies to Force

SafetyDGX agent

arXiv:2606.30988v1 Announce Type: new Abstract: Robot manipulation often relies on sensory feedback beyond vision, particularly in contact-rich settings where force, tactile, or audio signals reveal i

Multistage Defer Trees for Hybrid Interpretability: If at First You Can't Succeed, Tree Again

ResearchDGX agent

arXiv:2606.30995v1 Announce Type: new Abstract: Recent work has shown that well-optimized individual decision trees can match complex black box models in some settings, primarily in noisy domains. For

MultiUAV-Plat: An LLM-Oriented Platform, Benchmark and Framework for Multi-UAV Collaborative Task Planning

Model ReleasesDGX agent

arXiv:2606.31073v1 Announce Type: new Abstract: Large language models (LLMs) provide a promising interface for high-level robotic task planning, but their use in multi-UAV collaboration remains diffic

MuSViT: A Foundation Vision Model for Sheet Music Representation

ApplicationsDGX agent

arXiv:2606.31811v1 Announce Type: new Abstract: Foundation models have transformed vision and language processing by providing rich, reusable representations that transfer across diverse tasks. Sheet

MV-GEL: Language-Driven Multi-View Geometric Entity Localization on Meshes

ResearchDGX agent

arXiv:2606.31533v1 Announce Type: new Abstract: Identifying and grounding precise geometric entities, such as edges, planar regions, and curved surfaces within 3D objects, is foundational to computer-

MVP-Nav: Multi-layer Value Map Planner Navigator

TutorialsDGX agent

arXiv:2606.31919v1 Announce Type: cross Abstract: Zero-shot Object Goal Navigation (ZSON) with RGB-only perception poses a fundamental challenge for embodied agents, as the absence of explicit depth i

Nazrin: An Atomic Neural Proof Automation Tactic in Lean 4

AgentsDGX agent

arXiv:2602.18767v3 Announce Type: replace-cross Abstract: In Machine-Assisted Theorem Proving, a theorem proving agent searches for a sequence of expressions and tactics that can prove a statement in

Neuro-Bayesian-Symbolic Residual Attention Shallow Network: Explainable Deep Learning for Cybersecurity Risk Assessment

ResearchDGX agent

arXiv:2606.30953v1 Announce Type: new Abstract: We introduce the Neuro-Bayesian-Symbolic Residual Attention Shallow Network (NBS-RASN), a hybrid neural architecture for explainable cybersecurity risk

No Adaptation Without Observation: Observability-Constrained Test-Time Prompt Tuning for LiDAR Semantic Segmentation

SafetyDGX agent

arXiv:2606.30937v1 Announce Type: new Abstract: LiDAR semantic segmentation often degrades under real-world deployment due to evolving sensing conditions, while collecting new annotations for retraini

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs

Model ReleasesDGX agent

arXiv:2606.31933v1 Announce Type: new Abstract: We introduce VidPair-Halluc, a new benchmark for evaluating video hallucination in large video models (LVMs) under rigorous and controlled conditions. U

No Prompt, No Leaks: A Robust Generative Steganography Framework via Prompt-Free Diffusion

ResearchDGX agent

arXiv:2606.31427v1 Announce Type: new Abstract: Generative image steganography synthesizes stego images directly from secret information to achieve inherent security advantages. Latent Diffusion Model

Nonlinearity-Aware LoRA: Structured Gate Adaptation under Low-Rank Constraints

Model ReleasesDGX agent

arXiv:2606.31717v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) is commonly viewed as an update-space approximation to full fine-tuning, yet this view is incomplete for self-gated Transform

Not Every Time and Frequency Need to Be Forgotten in Diffusion Unlearning

ResearchDGX agent

arXiv:2510.17917v2 Announce Type: replace-cross Abstract: Data unlearning aims to remove the influence of specific training samples from a trained model. In fine-tuning methods, data unlearning relies

NURBS Splatting: A Unified Differentiable Rendering Framework for Vector Graphics

Model ReleasesDGX agent

arXiv:2606.31764v1 Announce Type: cross Abstract: Differentiable rendering of planar rational splines remains largely underexplored, despite their widespread use in vector graphics and design. Existin

Off the Rails: Hijacking the Scoring Head in Generative End-to-End Driving Planners with Safety-Violating Adversarial Perturbations

SafetyDGX agent

arXiv:2606.30807v1 Announce Type: cross Abstract: Generative models have recently seen rapid adoption in End-to-End (E2E) autonomous driving (AD), with diffusion-based denoising and vocabulary-based r

Offline Reinforcement Learning for Fluid Controls: Data-based Multi-observational Policy Extraction

SafetyDGX agent

arXiv:2606.31025v1 Announce Type: new Abstract: Active flow control is a fundamental application in engineering. Recent advances in deep reinforcement learning have made progress in this field. Howeve

On Optimal Data Splitting for Split Conformal Prediction

ApplicationsDGX agent

arXiv:2606.31600v1 Announce Type: cross Abstract: Conformal prediction and its variants, including the split conformal prediction, provide a distribution-free framework for uncertainty quantification

On Optimizing Multimodal Jailbreaks for Spoken Language Models

SafetyDGX agent

arXiv:2603.19127v2 Announce Type: replace Abstract: As Spoken Language Models (SLMs) integrate speech and text modalities, they inherit the safety vulnerabilities of their LLM backbone while introduci

On the Convergence of Self-Improving Online LLM Alignment

Model ReleasesDGX agent

arXiv:2606.31524v1 Announce Type: cross Abstract: The Self-Improving Alignment (SAIL) algorithm addresses distribution shift by reducing a bilevel formulation of the problem to an efficient, single-le

On the Role of Rotation Equivariance in Monocular 2D-to-3D Human Pose Lifting

ResearchDGX agent

arXiv:2601.13913v2 Announce Type: replace Abstract: Estimating 3D from 2D is one of the central tasks in computer vision. In this work, we consider the monocular setting, i.e. single-view input, for 3

One Reflection Is Not Enough: Self-Correcting Autonomous Research via Multi-Hypothesis Failure Attribution

Model ReleasesDGX agent

arXiv:2606.31478v1 Announce Type: new Abstract: Autonomous research agents can now draft hypotheses, write code, run experiments, and produce papers, but they remain brittle when experiments fail. Und

One Retrieval to Cover Them All: Co-occurrence-Aware Knowledge Base Reorganization for Session-Level RAG

ApplicationsDGX agent

arXiv:2606.31156v1 Announce Type: cross Abstract: RAG systems retrieve documents optimized for answering one query at a time. Yet enterprise users arrive with sessions, that is, coherent episodes of r

One Shot vs. Iterative: Rethinking Pruning Strategies for Model Compression

Model ReleasesDGX agent

arXiv:2508.13836v2 Announce Type: replace-cross Abstract: Pruning is a core technique for compressing neural networks to improve computational efficiency. This process is typically approached in two w

One Video, One World: Turning Monocular Video into Physical 4D Scenes

Model ReleasesDGX agent

arXiv:2606.31388v1 Announce Type: new Abstract: We introduce extbf{OVOW}, the first training-free system that reconstructs instance-level, simulation-ready 4D mesh scenes from a single monocular video

Online Generation of Collision-Free Trajectories in Dynamic Environments

SafetyDGX agent

arXiv:2603.00759v2 Announce Type: replace Abstract: In this paper, we present an online method for converting an arbitrary geometric path, represented by a sequence of states, and generated by any pla

Online TT-ALS for Streaming Tensor Decomposition with Incremental Orthogonalization

ResearchDGX agent

arXiv:2606.31061v1 Announce Type: cross Abstract: Tensor Train (TT) decomposition is a powerful technique for analyzing high-dimensional data. Existing algorithms for computing TT decompositions can b

OopsieVerse: A Safety Benchmark with Damage-Aware Simulation for Robot Manipulation

Model ReleasesDGX agent

arXiv:2606.31993v1 Announce Type: new Abstract: While robotic manipulation capabilities have advanced rapidly, physical safety remains a major barrier to deploying household robots: task success is in

OpenLife: Toward Open-World Artificial Life with Autonomous LLM Agents

AgentsDGX agent

arXiv:2606.31046v1 Announce Type: new Abstract: Artificial life has explored life-like behavior on many computational substrates, but mostly in researcher-designed closed worlds. We argue that large l

Optimal Self-Consistency for Efficient Reasoning with Large Language Models

ResearchDGX agent

arXiv:2511.12309v2 Announce Type: replace-cross Abstract: Self-consistency (SC) is a widely used test-time inference technique for improving performance in chain-of-thought reasoning. It consists of g

Optimization Algorithms for Joint OFDM Waveform Design and RIS Configuration in 6G Networks: From Convex Relaxation to Foundation Models

Model ReleasesDGX agent

arXiv:2606.31334v1 Announce Type: new Abstract: Joint OFDM-RIS optimization for 6G is a mixed-integer nonlinear programming (MINLP) problem covering sum-rate maximization, energy efficiency, max-min f

OTCache: Optimal Transport for Geometry-Aware Caching in Diffusion Models

Model ReleasesDGX agent

arXiv:2606.31026v1 Announce Type: cross Abstract: We propose OTCache, a training-free framework for accelerating diffusion sampling via caching schedule prediction. Existing graph-based caching method

Overview of the TalentCLEF 2026: Skill and Job Title Intelligence for Human Capital Management

ResearchDGX agent

arXiv:2606.31692v1 Announce Type: new Abstract: This paper presents an overview of the second edition of the TalentCLEF challenge, organized as a Lab at the Conference and Labs of the Evaluation Forum

PA-VAD: Diffusion-Based Pseudo-Only Video Anomaly Detection via Domain-Aligned Memory Updates

SafetyDGX agent

arXiv:2512.06845v2 Announce Type: replace Abstract: Deploying video anomaly detection (VAD) in the real world is often constrained by the scarcity, privacy, and cost of collecting real abnormal footag

Pano3D: Unified 3D Reconstruction and Panoptic Segmentation

ResearchDGX agent

arXiv:2606.14307v2 Announce Type: replace Abstract: Recent advances in 3D feedforward reconstruction neural networks have achieved remarkable success in dense reconstruction from images without any ca

Paper2Rebuttal: A Multi-Agent Framework for Transparent Author Response Assistance

SafetyDGX agent

arXiv:2601.14171v2 Announce Type: replace Abstract: Writing effective rebuttals is a high-stakes task that demands more than linguistic fluency, as it requires precise alignment between reviewer inten

← Previous
1…302303304305306…1025
Next →