AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
Human
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
21 Apr 2026

MESA: A Training-Free Multi-Exemplar Deep Framework for Restoring Ancient Inscription Textures

SafetyDGX agent

arXiv:2604.17390v1 Announce Type: new Abstract: Ancient inscriptions frequently suffer missing or corrupted regions from fragmentation, erosion, or other damage, hindering reading, and analysis. We re

MeSH: Memory-as-State-Highways for Recursive Transformers

Model ReleasesDGX agent

arXiv:2510.07739v2 Announce Type: replace Abstract: Recursive transformers reuse parameters and iterate over hidden states multiple times, decoupling compute depth from parameter depth. However, under

MetaCloak-JPEG: JPEG-Robust Adversarial Perturbation for Preventing Unauthorized DreamBooth-Based Deepfake Generation

ResearchDGX agent

arXiv:2604.18537v1 Announce Type: new Abstract: The rapid progress of subject-driven text-to-image synthesis, and in particular DreamBooth, has enabled a consent-free deepfake pipeline: an adversary n

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MetaLint: Easy-to-Hard Generalization for Code Linting

Model ReleasesDGX agent

arXiv:2507.11687v4 Announce Type: replace-cross Abstract: Large language models excel at code generation but struggle with code linting, particularly in generalizing to unseen or evolving best practic

MetaMem: Evolving Meta-Memory for Knowledge Utilization through Self-Reflective Symbolic Optimization

TutorialsDGX agent

arXiv:2602.11182v2 Announce Type: replace Abstract: Existing memory systems enable Large Language Models (LLMs) to support long-horizon human-LLM interactions by persisting historical interactions bey

Method for Aggregating Unstructured Data Using Large Language Models

Model ReleasesDGX agent

arXiv:2604.16425v1 Announce Type: cross Abstract: This paper presents a method for the automated collection and aggregation of unstructured data from diverse web sources, utilizing Large Language Mode

MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models

SafetyDGX agent

arXiv:2604.17730v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored as scalable tools for mental health counseling, yet evaluating their safety remains challenging d

Migrant Voices, Local News: Insights on Bridging Community Needs with Media Content

ResearchDGX agent

arXiv:2604.16651v1 Announce Type: new Abstract: Research shows news consumption differs across demographics, yet little is known about non-mainstream audiences, especially in relation to local media.

Mind the Way You Select Negative Texts: Pursuing the Distance Consistency in OOD Detection with VLMs

Model ReleasesDGX agent

arXiv:2603.02618v3 Announce Type: replace Abstract: Out-of-distribution (OOD) detection seeks to identify samples from unknown classes, a critical capability for deploying machine learning models in o

Mira-Embeddings-V1: Domain-Adapted Semantic Reranking for Recruitment via LLM-Synthesized Data

Local AiDGX agent

arXiv:2604.17738v1 Announce Type: new Abstract: Candidate sourcing for recruiters is best viewed as a two-stage retrieval and reranking pipeline with recall as the primary objective under a limited re

Missing-by-Design: Certifiable Modality Deletion for Revocable Multimodal Sentiment Analysis

Model ReleasesDGX agent

arXiv:2602.16144v3 Announce Type: replace Abstract: As multimodal systems increasingly process sensitive personal data, the ability to selectively revoke specific data modalities has become a critical

Missing Pattern Tree based Decision Grouping and Ensemble for Enhancing Pair Utilization in Deep Incomplete Multi-View Clustering

Model ReleasesDGX agent

arXiv:2512.21510v2 Announce Type: replace-cross Abstract: Real-world multi-view data often exhibit highly inconsistent missing patterns, posing significant challenges for incomplete multi-view cluster

Mitigating Multimodal Hallucination via Phase-wise Self-reward

ResearchDGX agent

arXiv:2604.17982v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) still struggle with vision hallucination, where generated responses are inconsistent with the visual input. Exist

Mix and Match: Context Pairing for Scalable Topic-Controlled Educational Summarisation

SafetyDGX agent

arXiv:2604.18087v1 Announce Type: new Abstract: Topic-controlled summarisation enables users to generate summaries focused on specific aspects of source documents. This paper investigates a data augme

MLE-UVAD: Minimal Latent Entropy Autoencoder for Fully Unsupervised Video Anomaly Detection

ResearchDGX agent

arXiv:2603.23868v2 Announce Type: replace Abstract: In this paper, we address the challenging problem of single-scene, fully unsupervised video anomaly detection (VAD), where raw videos containing bot

mlr3torch: A Deep Learning Framework in R based on mlr3 and torch

TutorialsDGX agent

arXiv:2604.18152v1 Announce Type: cross Abstract: Deep learning (DL) has become a cornerstone of modern machine learning (ML) praxis. We introduce the R package mlr3torch, which is an extensible DL fr

MM-Hand: A 21-DOF Multi-modal Modular Dexterous Robotic Hand with Remote Actuation

TutorialsDGX agent

arXiv:2604.17245v1 Announce Type: new Abstract: High-DOF dexterous hands require compact actuation, rich sensing, and reliable thermal behavior, but conventional designs often occupy valuable in-hand

MM-JudgeBias: A Benchmark for Evaluating Compositional Biases in MLLM-as-a-Judge

Model ReleasesDGX agent

arXiv:2604.18164v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have been increasingly used as automatic evaluators-a paradigm known as MLLM-as-a-Judge. However, their reliabi

MMErroR: A Benchmark for Erroneous Reasoning in Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.03331v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have improved performance in multi-modal learning, raising the question of whether these models tru

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation

Model ReleasesDGX agent

arXiv:2604.16943v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown impressive capabilities, yet they often struggle to effectively capture the fine-grained textual inf

MobileAgeNet: Lightweight Facial Age Estimation for Mobile Deployment

Model ReleasesDGX agent

arXiv:2604.17007v1 Announce Type: new Abstract: Mobile deployment of facial age estimation requires models that balance predictive accuracy with low latency and compact size. In this work, we present

MoCo: A One-Stop Shop for Model Collaboration Research

SafetyDGX agent

arXiv:2601.21257v2 Announce Type: replace Abstract: Advancing beyond single monolithic language models (LMs), recent research increasingly recognizes the importance of model collaboration, where multi

Model in Distress: Sentiment Analysis on French Synthetic Social Media

Model ReleasesDGX agent

arXiv:2604.18226v1 Announce Type: new Abstract: Automated analysis of customer feedback on social media is hindered by three challenges: the high cost of annotated training data, the scarcity of evalu

Modeling Biomechanical Constraint Violations for Language-Agnostic Lip-Sync Deepfake Detection

ResearchDGX agent

arXiv:2604.16808v1 Announce Type: new Abstract: Current lip-sync deepfake detectors rely on pixel-level artifacts or audio-visual correspondence, failing to generalize across languages because these c

Modeling, Control and Self-sensing of Dielectric Elastomer Soft Actuators: A Review

ResearchDGX agent

arXiv:2604.17199v1 Announce Type: new Abstract: Dielectric elastomer actuators (DEAs) have garnered extensive attention especially in soft robotic applications over the past few decades owing to the a

Modeling Higher-Order Brain Interactions via a Multi-View Information Bottleneck Framework for fMRI-based Psychiatric Diagnosis

Model ReleasesDGX agent

arXiv:2604.17713v1 Announce Type: new Abstract: Resting-state functional magnetic resonance imaging (fMRI) has emerged as a cornerstone for psychiatric diagnosis, yet most approaches rely on pairwise

Modeling Human Perspectives with Socio-Demographic Representations

ApplicationsDGX agent

arXiv:2604.18069v1 Announce Type: new Abstract: Humans often hold different perspectives on the same issues. In many NLP tasks, annotation disagreement can reflect valid subjective perspectives. Model

Modeling Multi-Dimensional Cognitive States in Large Language Models under Cognitive Crowding

Model ReleasesDGX agent

arXiv:2604.17174v1 Announce Type: new Abstract: Modeling human cognitive states is essential for advanced artificial intelligence. Existing Large Language Models (LLMs) mainly address isolated tasks s

Modeling Multiple Support Strategies within a Single Turn for Emotional Support Conversations

Model ReleasesDGX agent

arXiv:2604.17972v1 Announce Type: new Abstract: Emotional Support Conversation (ESC) aims to assist individuals experiencing distress by generating empathetic and supportive dialogue. While prior work

Modeling User Exploration Saturation: When Recommender Systems Should Stop Pushing Novelty

SafetyDGX agent

arXiv:2604.16419v1 Announce Type: cross Abstract: Fairness-aware recommender systems often mitigate bias by increasing exposure to under-represented or long-tail content, commonly through mechanisms t

Modelling Gas-Phase Reaction Kinetics with Guided Particle Diffusion Sampling

Model ReleasesDGX agent

arXiv:2604.16461v1 Announce Type: cross Abstract: Physics-guided sampling with diffusion priors has recently shown strong performance in solving complex systems of partial differential equations (PDEs

MODEST: Multi-Optics Depth-of-Field Stereo Dataset

AgentsDGX agent

arXiv:2511.20853v3 Announce Type: replace Abstract: Reliable depth estimation under real optical conditions remains a core challenge for camera vision in systems such as autonomous robotics and augmen

Modular Representation Compression: Adapting LLMs for Efficient and Effective Recommendations

TutorialsDGX agent

arXiv:2604.18146v1 Announce Type: cross Abstract: Recently, large language models (LLMs) have advanced recommendation systems (RSs), and recent works have begun to explore how to integrate LLMs into i

MoE-nD: Per-Layer Mixture-of-Experts Routing for Multi-Axis KV Cache Compression

ResearchDGX agent

arXiv:2604.17695v1 Announce Type: cross Abstract: KV cache memory is the dominant bottleneck for long-context LLM inference. Existing compression methods each act on a single axis of the four-dimensio

MoGERNN: An Inductive Traffic Predictor for Unobserved Locations

ApplicationsDGX agent

arXiv:2501.12281v2 Announce Type: replace Abstract: Given a partially observed road network, how can we predict the traffic state of interested unobserved locations? Traffic prediction is crucial for

More Than Meets the Eye: Measuring the Semiotic Gap in Vision-Language Models via Semantic Anchorage

Model ReleasesDGX agent

arXiv:2604.17354v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at photorealistic generation, yet often struggle to represent abstract meaning such as idiomatic interpretations of

More Than Sum of Its Parts: Deciphering Intent Shifts in Multimodal Hate Speech Detection

Model ReleasesDGX agent

arXiv:2603.21298v2 Announce Type: replace Abstract: Combating hate speech on social media is critical for securing cyberspace, yet relies heavily on the efficacy of automated detection systems. As con

MoRI: Learning Motivation-Grounded Reasoning for Scientific Ideation in Large Language Models

AgentsDGX agent

arXiv:2603.19044v2 Announce Type: replace Abstract: Scientific ideation aims to propose novel solutions within a given scientific context. Existing LLM-based agentic approaches emulate human research

Motif-Video 2B: Technical Report

Model ReleasesDGX agent

arXiv:2604.16503v1 Announce Type: new Abstract: Training strong video generation models usually requires massive datasets, large parameter counts, and substantial compute. In this work, we ask whether

Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition

SafetyDGX agent

arXiv:2604.17062v1 Announce Type: new Abstract: Zero-shot action recognition is challenging due to the semantic gap between seen and unseen classes. We present a novel framework that enhances CLIP wit

MoVE: Translating Laughter and Tears via Mixture of Vocalization Experts in Speech-to-Speech Translation

ApplicationsDGX agent

arXiv:2604.17435v1 Announce Type: new Abstract: Recent Speech-to-Speech Translation (S2ST) systems achieve strong semantic accuracy yet consistently strip away non-verbal vocalizations (NVs), such as

MTSQL-R1: Towards Long-Horizon Multi-Turn Text-to-SQL via Agentic Training

Model ReleasesDGX agent

arXiv:2510.12831v3 Announce Type: replace Abstract: Multi-turn Text-to-SQL aims to translate a user's conversational utterances into executable SQL while preserving dialogue coherence and grounding to

MU-GeNeRF: Multi-view Uncertainty-guided Generalizable Neural Radiance Fields for Distractor-aware Scene

ApplicationsDGX agent

arXiv:2604.17965v1 Announce Type: new Abstract: Generalizable Neural Radiance Fields (GeNeRFs) enable high-quality scene reconstruction from sparse views and can generalize to unseen scenes. However,

MUA: Mobile Ultra-detailed Animatable Avatars

Local AiDGX agent

arXiv:2604.18583v1 Announce Type: new Abstract: Building photorealistic, animatable full-body digital humans remains a longstanding challenge in computer graphics and vision. Recent advances in animat

Multi-Beholder: Biomarker Prediction for Low-Grade Glioma with Multiple Instance Learning and One-Class Classification

ResearchDGX agent

arXiv:2310.07464v2 Announce Type: replace-cross Abstract: Biomarker detection is an indispensable part of the diagnosis and treatment of low-grade glioma (LGG). However, current LGG biomarker detectio

Multi-Camera Self-Calibration in Sports Motion Capture: Leveraging Human and Stick Poses

Model ReleasesDGX agent

arXiv:2604.17567v1 Announce Type: new Abstract: Multi-camera systems are widely employed in sports to capture the 3D motion of athletes and equipment, yet calibrating their extrinsic parameters remain

Multi-Label Phase Diagram Prediction in Complex Alloys via Physics-Informed Graph Attention Networks

ResearchDGX agent

arXiv:2604.16468v1 Announce Type: new Abstract: Accurate phase equilibria are foundational to alloy design because they encode the underlying thermodynamics governing stability, transformations, and p

Multi-Scale Reversible Chaos Game Representation: A Unified Framework for Sequence Classification

ResearchDGX agent

arXiv:2604.18477v1 Announce Type: new Abstract: Biological classification with interpretability remains a challenging task. For this, we introduce a novel encoding framework, Multi-Scale Reversible Ch

Multi-stage Planning for Multi-target Surveillance using Aircrafts Equipped with Synthetic Aperture Radars Aware of Target Visibility

ResearchDGX agent

arXiv:2604.16962v1 Announce Type: new Abstract: Generating trajectories for synthetic aperture radar (SAR)-equipped aircraft poses significant challenges due to terrain constraints, and the need for s

Multi-View Hierarchical Graph Neural Network for Sketch-Based 3D Shape Retrieval

Local AiDGX agent

arXiv:2604.18019v1 Announce Type: new Abstract: Sketch-based 3D shape retrieval (SBSR) aims to retrieve 3D shapes that are consistent with the category of the input hand-drawn sketch. The core challen

Multilevel neural networks with dual-stage feature fusion for human activity recognition

Model ReleasesDGX agent

arXiv:2604.16577v1 Announce Type: new Abstract: Human activity recognition (HAR) refers to the process of identifying human actions and activities using data collected from sensors. Neural networks, s

Multilingual Training and Evaluation Resources for Vision-Language Models

ResearchDGX agent

arXiv:2604.18347v1 Announce Type: new Abstract: Vision Language Models (VLMs) achieved rapid progress in the recent years. However, despite their growth, VLMs development is heavily grounded on Englis

Multimodal Claim Extraction for Fact-Checking

Model ReleasesDGX agent

arXiv:2604.16311v1 Announce Type: new Abstract: Automated Fact-Checking (AFC) relies on claim extraction as a first step, yet existing methods largely overlook the multimodal nature of today's misinfo

Multimodal Fusion of Histopathology Images and Electronic Health Records for Early Breast Cancer Diagnosis

ResearchDGX agent

arXiv:2604.17122v1 Announce Type: new Abstract: Breast cancer is a leading cause of cancer-related mortality worldwide, and timely accurate diagnosis is critical to improving survival outcomes. While

Multimodal In-context Learning for ASR of Low-resource Languages

Model ReleasesDGX agent

arXiv:2601.05707v2 Announce Type: replace Abstract: Automatic speech recognition (ASR) still covers only a small fraction of the world's languages, mainly due to supervised data scarcity. In-context l

Multimodal Policy Internalization for Conversational Agents

SafetyDGX agent

arXiv:2510.09474v2 Announce Type: replace Abstract: Modern conversational agents like ChatGPT and Alexa+ rely on predefined policies specifying metadata, response styles, and tool-usage rules. As thes

Multimodal Sentiment Analysis with Missing Modality: A Knowledge-Transfer Approach

ResearchDGX agent

arXiv:2401.10747v5 Announce Type: replace-cross Abstract: Multimodal sentiment analysis aims to identify the emotions expressed by individuals through visual, language, and acoustic cues. However, mos

Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs

Model ReleasesDGX agent

arXiv:2604.18203v1 Announce Type: new Abstract: Multimodal LLMs can accurately perceive numerical content across modalities yet fail to perform exact multi-digit multiplication when the identical unde

MultiWorld: Scalable Multi-Agent Multi-View Video World Models

AgentsDGX agent

arXiv:2604.18564v1 Announce Type: new Abstract: Video world models have achieved remarkable success in simulating environmental dynamics in response to actions by users or agents. They are modeled as

Muscle-inspired magnetic actuators that push, pull, crawl, and grasp

HardwareDGX agent

arXiv:2604.18090v1 Announce Type: new Abstract: Functional magnetic composites capable of large deformation, load bearing, and multifunctional motion are essential for next-generation adaptive soft ro

← Previous
1…893894895896897…998
Next →