AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
22 Apr 2026

Multi-Cycle Spatio-Temporal Adaptation in Human-Robot Teaming

TutorialsDGX agent

arXiv:2604.19670v1 Announce Type: cross Abstract: Effective human-robot teaming is crucial for the practical deployment of robots in human workspaces. However, optimizing joint human-robot plans remai

Multi-Domain Learning with Global Expert Mapping

Model ReleasesDGX agent

arXiv:2604.18842v1 Announce Type: new Abstract: Human perception generalizes well across different domains, but most vision models struggle beyond their training data. This gap motivates multi-dataset

Multi-Gait Learning for Humanoid Robots Using Reinforcement Learning with Selective Adversarial Motion Prior

SafetyDGX agent

arXiv:2604.19102v1 Announce Type: cross Abstract: Learning diverse locomotion skills for humanoid robots in a unified reinforcement learning framework remains challenging due to the conflicting requir

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Multi-Level Temporal Graph Networks with Local-Global Fusion for Industrial Fault Diagnosis

Local AiDGX agent

arXiv:2604.18765v1 Announce Type: cross Abstract: Fault detection and diagnosis are critical for the optimal and safe operation of industrial processes. The correlations among sensors often display no

Multi-modal Reasoning with LLMs for Visual Semantic Arithmetic

SafetyDGX agent

arXiv:2604.19567v1 Announce Type: new Abstract: Reinforcement learning (RL) as post-training is crucial for enhancing the reasoning ability of large language models (LLMs) in coding and math. However,

Multi-modal Test-time Adaptation via Adaptive Probabilistic Gaussian Calibration

Model ReleasesDGX agent

arXiv:2604.19093v1 Announce Type: cross Abstract: Multi-modal test-time adaptation (TTA) enhances the resilience of benchmark multi-modal models against distribution shifts by leveraging the unlabeled

Multi-Step Gaussian Process Propagation for Adaptive Path Planning

AgentsDGX agent

arXiv:2604.19148v1 Announce Type: new Abstract: Efficient and robust path planning hinges on combining all accessible information sources. In particular, the task of path planning for robotic environm

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

SafetyDGX agent

arXiv:2603.11665v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have been widely adopted as MLLM-as-a-Judges due to their strong alignment with human judgment across vario

Multi-view Crowd Tracking Transformer with View-Ground Interactions Under Large Real-World Scenes

ApplicationsDGX agent

arXiv:2604.19318v1 Announce Type: new Abstract: Multi-view crowd tracking estimates each person's tracking trajectories on the ground of the scene. Recent research works mainly rely on CNNs-based mult

Multiclass Local Calibration with the Jensen-Shannon Distance

SafetyDGX agent

arXiv:2510.26566v2 Announce Type: replace-cross Abstract: Developing trustworthy Machine Learning (ML) models requires their predicted probabilities to be well-calibrated, meaning they should reflect

Multilingual Language Models Encode Script Over Linguistic Structure

Model ReleasesDGX agent

arXiv:2604.05090v2 Announce Type: replace Abstract: Multilingual language models (LMs) organize representations for typologically and orthographically diverse languages into a shared parameter space,

Multimodal embodiment-aware navigation transformer

SafetyDGX agent

arXiv:2604.19267v1 Announce Type: new Abstract: Goal-conditioned navigation models for ground robots trained using supervised learning show promising zero-shot transfer, but their collision-avoidance

Multimodal Transformer for Sample-Aware Prediction of Metal-Organic Framework Properties

TutorialsDGX agent

arXiv:2604.19383v1 Announce Type: cross Abstract: Metal-organic frameworks (MOFs) are a major target of machine-learning-based property prediction, yet most models assume that a single framework repre

NemeSys: Toward Online Underwater Exploration with Remote Operator-in-the-loop Adaptive Autonomy

Model ReleasesDGX agent

arXiv:2507.11889v2 Announce Type: replace Abstract: Adaptive mission control and dynamic parameter reconfiguration are essential for autonomous underwater vehicles (AUVs) operating in GPS-denied, comm

NeuroAI and Beyond: Bridging Between Advances in Neuroscience and ArtificialIntelligence

ResearchDGX agent

arXiv:2604.18637v1 Announce Type: cross Abstract: Neuroscience and Artificial Intelligence (AI) have made impressive progress in recent years but remain only loosely interconnected. Based on a worksho

Neuromorphic Continual Learning for Sequential Deployment of Nuclear Plant Monitoring Systems

SafetyDGX agent

arXiv:2604.18611v1 Announce Type: cross Abstract: Anomaly detection in nuclear industrial control systems (ICS) requires continuous, energy-efficient monitoring across multiple subsystems that are oft

Nexusformer: Nonlinear Attention Expansion for Stable and Inheritable Transformer Scaling

ResearchDGX agent

arXiv:2604.19147v1 Announce Type: cross Abstract: Scaling Transformers typically necessitates training larger models from scratch, as standard architectures struggle to expand without discarding learn

Nonmonotone subgradient methods based on a local descent lemma

ResearchDGX agent

arXiv:2510.19341v2 Announce Type: replace-cross Abstract: In this paper we present a nonmonotone line search subgradient algorithm tailored to upper-C^2 functions. This is a family of nonsmooth and no

ODMA: On-Demand Memory Allocation Strategy for LLM Serving on LPDDR-Class Accelerators

Model ReleasesDGX agent

arXiv:2512.09427v5 Announce Type: replace-cross Abstract: Existing memory management techniques severely hinder efficient Large Language Model serving on accelerators constrained by poor random-access

OLLM: Options-based Large Language Models

Model ReleasesDGX agent

arXiv:2604.19087v1 Announce Type: new Abstract: We introduce Options LLM (OLLM), a simple, general method that replaces the single next-token prediction of standard LLMs with a extit{set of learned op

OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration

AgentsDGX agent

arXiv:2505.11765v3 Announce Type: replace-cross Abstract: Agents powered by advanced large language models (LLMs) have demonstrated impressive capabilities across diverse complex applications. Recentl

OmniGen2: Towards Instruction-Aligned Multimodal Generation

Model ReleasesDGX agent

arXiv:2506.18871v4 Announce Type: replace-cross Abstract: In this work, we introduce OmniGen2, a versatile and open-source generative model designed to provide a unified solution for diverse generatio

OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural Tokens

Model ReleasesDGX agent

arXiv:2604.18827v1 Announce Type: cross Abstract: Scaling data and artificial neural networks has transformed AI, driving breakthroughs in language and vision. Whether similar principles apply to mode

OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models

ResearchDGX agent

arXiv:2502.16161v2 Announce Type: replace-cross Abstract: Visually-situated text parsing (VsTP) has recently seen notable advancements, driven by the growing demand for automated document understandin

OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models

ResearchDGX agent

arXiv:2604.00688v3 Announce Type: replace Abstract: We present OmniVoice, a massively multilingual zero-shot text-to-speech (TTS) model that scales to over 600 languages. At its core is a novel diffus

On Accelerating Grounded Code Development for Research

ResearchDGX agent

arXiv:2604.19022v1 Announce Type: new Abstract: A major challenge for niche scientific and technical domains in leveraging coding agents is the lack of access to up-to-date, domain- specific knowledge

On Solving the Multiple Variable Gapped Longest Common Subsequence Problem

ResearchDGX agent

arXiv:2604.18645v1 Announce Type: new Abstract: This paper addresses the Variable Gapped Longest Common Subsequence (VGLCS) problem, a generalization of the classical LCS problem involving flexible ga

On Temperature-Constrained Non-Deterministic Machine Translation: Potential and Evaluation

ApplicationsDGX agent

arXiv:2601.13729v2 Announce Type: replace Abstract: In recent years, the non-deterministic properties of language models have garnered considerable attention and have shown a significant influence on

On the Conditioning Consistency Gap in Conditional Neural Processes

ResearchDGX agent

arXiv:2604.19312v1 Announce Type: new Abstract: Neural processes are meta-learning models that map context sets to predictive distributions. While inspired by stochastic processes, NPs do not generall

On the Derivation of Tightly-Coupled LiDAR-Inertial Odometry with VoxelMap

ResearchDGX agent

arXiv:2603.15471v2 Announce Type: replace Abstract: This note presents a concise mathematical formulation of tightly-coupled LiDAR-Inertial Odometry within an iterated error-state Kalman filter framew

On the Generalizability of Foundation Models for Crop Type Mapping

SafetyDGX agent

arXiv:2409.09451v5 Announce Type: replace Abstract: Foundation models pre-trained using self-supervised learning have shown powerful transfer learning capabilities on various downstream tasks, includi

On the Spatiotemporal Dynamics of Generalization in Neural Networks

Local AiDGX agent

arXiv:2602.01651v2 Announce Type: replace-cross Abstract: Why do neural networks fail to generalize addition from 16-digit to 32-digit numbers, while a child who learns the rule can apply it to arbitr

On two ways to use determinantal point processes for Monte Carlo integration

ResearchDGX agent

arXiv:2604.19698v1 Announce Type: new Abstract: The standard Monte Carlo estimator widehat{I}_N^{MC} of int fdomega relies on independent samples from omega and has variance of order 1/N. Replacing th

Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models

Model ReleasesDGX agent

arXiv:2602.05437v2 Announce Type: replace Abstract: Vision-language models (VLMs) can achieve high accuracy while still accepting culturally plausible but visually incorrect interpretations. Existing

One Persona, Many Cues, Different Results: How Sociodemographic Cues Impact LLM Personalization

SafetyDGX agent

arXiv:2601.18572v2 Announce Type: replace Abstract: Personalization of LLMs by sociodemographic subgroup often improves user experience, but can also introduce or amplify biases and unfair outcomes ac

One Step Forward and K Steps Back: Better Reasoning with Denoising Recursion Models

Model ReleasesDGX agent

arXiv:2604.18839v1 Announce Type: cross Abstract: Looped transformers scale computational depth without increasing parameter count by repeatedly applying a shared transformer block and can be used for

Online Learning of Whittle Indices for Restless Bandits with Non-Stationary Transition Kernels

SafetyDGX agent

arXiv:2506.18186v3 Announce Type: replace Abstract: The restless multi-armed bandit (RMAB) framework is a popular approach to solving resource allocation problems in networked systems. In this paper,

Ontology-Constrained Neural Reasoning in Enterprise Agentic Systems: A Neurosymbolic Architecture for Domain-Grounded AI Agents

Local AiDGX agent

arXiv:2604.00555v2 Announce Type: replace Abstract: Enterprise adoption of Large Language Models (LLMs) is constrained by hallucination, domain drift, and the inability to enforce regulatory complianc

Opinion de-polarization in social networks with GNNs

ResearchDGX agent

arXiv:2412.09404v3 Announce Type: replace-cross Abstract: Nowadays, social media is the ground for political debate and exchange of opinions. There is a significant amount of research that suggests th

Optimal Exploration of New Products under Assortment Decisions

TutorialsDGX agent

arXiv:2604.18800v1 Announce Type: cross Abstract: We study online learning for new products on a platform that makes capacity-constrained assortment decisions on which products to offer. For a newly l

Optimal Routing for Federated Learning over Dynamic Satellite Networks: Tractable or Not?

Local AiDGX agent

arXiv:2604.19399v1 Announce Type: new Abstract: Federated learning (FL) is a key paradigm for distributed model learning across decentralized data sources. Communication in each FL round typically con

Optimized Architectures for Kolmogorov-Arnold Networks

TutorialsDGX agent

arXiv:2512.12448v2 Announce Type: replace Abstract: Efforts to improve Kolmogorov--Arnold networks (KANs) with architectural enhancements have been stymied by the complexity those enhancements bring,

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models

Model ReleasesDGX agent

arXiv:2509.15435v2 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) exhibit strong multimodal capabilities but remain vulnerable to hallucinations from intrinsic errors and

Owner-Harm: A Missing Threat Model for AI Agent Safety

Model ReleasesDGX agent

arXiv:2604.18658v1 Announce Type: cross Abstract: Existing AI agent safety benchmarks focus on generic criminal harm (cybercrime, harassment, weapon synthesis), leaving a systematic blind spot for a d

PanDA: Unsupervised Domain Adaptation for Multimodal 3D Panoptic Segmentation in Autonomous Driving

AgentsDGX agent

arXiv:2604.19379v1 Announce Type: new Abstract: This paper presents the first study on Unsupervised Domain Adaptation (UDA) for multimodal 3D panoptic segmentation (mm-3DPS), aiming to improve general

Paparazzo: Active Mapping of Moving 3D Objects

Model ReleasesDGX agent

arXiv:2604.19556v1 Announce Type: new Abstract: Current 3D mapping pipelines generally assume static environments, which limits their ability to accurately capture and reconstruct moving objects. To a

ParamBoost: Gradient Boosted Piecewise Cubic Polynomials

ApplicationsDGX agent

arXiv:2604.18864v1 Announce Type: new Abstract: Generalized Additive Models (GAMs) can be used to create non-linear glass-box (i.e. explicitly interpretable) models, where the predictive function is f

Pause or Fabricate? Training Language Models for Grounded Reasoning

ResearchDGX agent

arXiv:2604.19656v1 Announce Type: new Abstract: Large language models have achieved remarkable progress on complex reasoning tasks. However, they often implicitly fabricate information when inputs are

PC2Model: ISPRS benchmark on 3D point cloud to model registration

Model ReleasesDGX agent

arXiv:2604.19596v1 Announce Type: new Abstract: Point cloud registration involves aligning one point cloud with another or with a three-dimensional (3D) model, enabling the integration of multimodal d

Personalized Benchmarking: Evaluating LLMs by Individual Preferences

SafetyDGX agent

arXiv:2604.18943v1 Announce Type: new Abstract: With the rise in capabilities of large language models (LLMs) and their deployment in real-world tasks, evaluating LLM alignment with human preferences

Personalized Embodied Navigation for Portable Object Finding

AgentsDGX agent

arXiv:2403.09905v5 Announce Type: replace-cross Abstract: Embodied navigation methods commonly operate in static environments with stationary objects. In this work, we present approaches for tackling

Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications

SafetyDGX agent

arXiv:2411.06837v2 Announce Type: replace Abstract: The rapid rise of Large Language Models (LLMs) has created new disruptive possibilities for persuasive communication, enabling fully-automated, pers

Phase-Aware Policy Learning for Skateboard Riding of Quadruped Robots via Feature-wise Linear Modulation

SafetyDGX agent

arXiv:2602.09370v2 Announce Type: replace Abstract: Skateboards offer a compact and efficient means of transportation as a type of personal mobility device. However, controlling them with legged robot

Phase Transitions in the Fluctuations of Functionals of Random Neural Networks

ResearchDGX agent

arXiv:2604.19738v1 Announce Type: cross Abstract: We establish central and non-central limit theorems for sequences of functionals of the Gaussian output of an infinitely-wide random neural network on

PhotoFramer: Multi-modal Image Composition Instruction

TutorialsDGX agent

arXiv:2512.00993v2 Announce Type: replace Abstract: Composition matters during the photo-taking process, yet many casual users struggle to frame well-composed images. To provide composition guidance,

Physics-Informed Neural Operators for Cardiac Electrophysiology

TutorialsDGX agent

arXiv:2511.08418v2 Announce Type: replace Abstract: Accurately simulating systems governed by PDEs, such as voltage fields in cardiac electrophysiology (EP) modelling, remains a significant modelling

PhysMem: Scaling Test-time Physical Memory for Robot Manipulation

TutorialsDGX agent

arXiv:2602.20323v5 Announce Type: replace-cross Abstract: Reliable object manipulation requires understanding physical properties that vary across objects and environments. Vision-language model (VLM)

Pixels or Positions? Benchmarking Modalities in Group Activity Recognition

Model ReleasesDGX agent

arXiv:2511.12606v3 Announce Type: replace Abstract: Group Activity Recognition (GAR) is well studied on the video modality for surveillance and indoor team sports (e.g., volleyball, basketball). Yet,

PLaMo 2.1-VL Technical Report

AgentsDGX agent

arXiv:2604.19324v1 Announce Type: cross Abstract: We introduce PLaMo 2.1-VL, a lightweight Vision Language Model (VLM) for autonomous devices, available in 8B and 2B variants and designed for local an

Planning in entropy-regularized Markov decision processes and games

ResearchDGX agent

arXiv:2604.19695v1 Announce Type: new Abstract: We propose SmoothCruiser, a new planning algorithm for estimating the value function in entropy-regularized Markov decision processes and two-player gam

← Previous
1…875876877878879…998
Next →