AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
Human
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
5 May 2026

Medmarks: A Comprehensive Open-Source LLM Benchmark Suite for Medical Tasks

Model ReleasesDGX agent

arXiv:2605.01417v1 Announce Type: new Abstract: Evaluating large language models (LLMs) for medical applications remains challenging due to benchmark saturation, limited data accessibility, and insuff

MedMosaic: A Challenging Large Scale Benchmark of Diverse Medical Audio

Model ReleasesDGX agent

arXiv:2605.00969v1 Announce Type: cross Abstract: We present MedMosaic, a medical audio question-answering dataset designed to benchmark language and audio reasoning models under realistic clinical co

MedScribe: Clinically Grounded CT Reporting through Agentic Workflows

AgentsDGX agent

arXiv:2605.01779v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown potential for automated radiology report generation, yet existing approaches rely on global embedding compressi

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MemeLens: Multilingual Multitask VLMs for Memes

ResearchDGX agent

arXiv:2601.12539v3 Announce Type: replace-cross Abstract: Memes are a dominant medium for online communication and manipulation because meaning emerges from interactions between embedded text, imagery

MemORAI: Memory Organization and Retrieval via Adaptive Graph Intelligence for LLM Conversational Agents

ResearchDGX agent

arXiv:2605.01386v1 Announce Type: new Abstract: Large Language Models (LLMs) lack persistent memory for long-term personalized conversations. Existing graph-based memory systems suffer from informatio

MER-DG: Modality-Entropy Regularization for Multimodal Domain Generalization

ApplicationsDGX agent

arXiv:2605.01967v1 Announce Type: cross Abstract: Deploying multimodal models in real-world scenarios requires generalization to new environments where recording conditions differ from training, a cha

Mesh Based Simulations with Spatial and Temporal awareness

ResearchDGX agent

arXiv:2605.01542v1 Announce Type: new Abstract: Machine Learning surrogates for Computational Fluid Dynamics (CFD), particularly Graph Neural Networks (GNNs) and Transformers, have become a new import

Meta-learning Structure-Preserving Dynamics

Model ReleasesDGX agent

arXiv:2508.11205v2 Announce Type: replace Abstract: Structure-preserving approaches to dynamics discovery have demonstrated great potential for modeling physical systems due to their use of strong ind

Methods, Data, and Conceptual Change: Reflections from Two Quantitative Diachronic Case Studies

ResearchDGX agent

arXiv:2605.02052v1 Announce Type: new Abstract: This discussion paper reflects on how quantitative approaches to historical linguistics interact with dataset properties. Drawing on two worked examples

Metric-Normalized Posterior Leakage (mPL): Attacker-Aligned Privacy for Joint Consumption

Local AiDGX agent

arXiv:2605.01137v1 Announce Type: new Abstract: Metric differential privacy (mDP) strengthens local differential privacy (LDP) by scaling noise to semantic distance, but many machine learning (ML) sys

Metric Unreliability in Multimodal Machine Unlearning: A Systematic Analysis and Principled Unified Score

Model ReleasesDGX agent

arXiv:2605.02206v1 Announce Type: new Abstract: Machine unlearning in Vision-Language Models (VLMs) is required for compliance with the General Data Protection Regulation (GDPR), yet current evaluatio

Mextsuperscript{4}Fuse: Lightweight State-Space MoE with a Cross-Scale Gating Bridge for Brain Tumor Segmentation

Model ReleasesDGX agent

arXiv:2605.02444v1 Announce Type: new Abstract: Encoder-decoder imbalance and the reliance on large input volumes make many 3D brain tumor segmentation models both compute-heavy and brittle. We presen

Middle-mile logistics through the lens of goal-conditioned reinforcement learning

ResearchDGX agent

arXiv:2605.02461v1 Announce Type: cross Abstract: Middle-mile logistics describes the problem of routing parcels through a network of hubs linked by trucks with finite capacity. We rephrase this as a

Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models

Model ReleasesDGX agent

arXiv:2605.00123v1 Announce Type: new Abstract: Safety trained large language models (LLMs) can often be induced to answer harmful requests through jailbreak prompts. Because we lack a robust understa

Minimizing Collateral Damage in Activation Steering

SafetyDGX agent

arXiv:2605.01167v1 Announce Type: new Abstract: Activation steering is a method for controlling Large Language Model (LLM) behavior by intervening in its internal representations to increase the align

Minimum Specification Perturbation: Robustness as Distance-to-Falsification in Causal Inference

Model ReleasesDGX agent

arXiv:2605.01579v1 Announce Type: cross Abstract: Empirical causal claims depend on many analyst decisions, from selecting covariates to choosing estimators. Existing robustness tools summarize how re

MIRA: A Score for Conditional Distribution Accuracy and Model Comparison

SafetyDGX agent

arXiv:2605.02014v1 Announce Type: cross Abstract: We introduce Mira, a sample-based score for assessing the accuracy of a candidate conditional distribution using only joint samples from the true data

MIRL: Mutual Information-Guided Reinforcement Learning for Vision-Language Models

ResearchDGX agent

arXiv:2605.01520v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) frequently suffer from visual perception errors and hallucinations that compromise answer accuracy in complex reasoning

Misclassification Rate and Privacy-Utility Trade-offs in Graph Convolutional Networks via Subsampling Stability

ResearchDGX agent

arXiv:2605.01987v1 Announce Type: new Abstract: We study differential privacy (DP) in Graph Convolutional Networks (GCNs) through the framework of extit{subsampling stability}. We derive upper bounds

Missingness-aware Data Imputation via AI-powered Bayesian Generative Modeling

ResearchDGX agent

arXiv:2605.01676v1 Announce Type: cross Abstract: Missing data imputation remains a fundamental challenge in modern data science, especially when uncertainty quantification is essential. In this work,

Mitigating Misalignment Contagion by Steering with Implicit Traits

SafetyDGX agent

arXiv:2605.02751v1 Announce Type: cross Abstract: Language models (LMs) are increasingly used in high-stakes, multi-agent settings, where following instructions and maintaining value alignment are cri

Mitigating Multimodal LLMs Hallucinations via Relevance Propagation at Inference Time

ResearchDGX agent

arXiv:2605.01766v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have revolutionized the landscape of AI, demonstrating impressive capabilities in tackling complex vision and

Mixture Prototype Flow Matching for Open-Set Supervised Anomaly Detection

ResearchDGX agent

arXiv:2605.02438v1 Announce Type: new Abstract: Open-set supervised anomaly detection (OSAD) aims to identify unseen anomalies using limited anomalous supervision. However, existing prototype-based me

MOC-3D: Manifold-Order Consistency for Text-to-3D Generation

SafetyDGX agent

arXiv:2605.01743v1 Announce Type: new Abstract: With the burgeoning development of fields such as the Metaverse, Virtual Reality (VR), and Digital Twins, text-to-3D generation has emerged as a researc

Model-Based Proactive Cost Generation for Learning Safe Policies Offline with Limited Violation Data

SafetyDGX agent

arXiv:2605.01356v1 Announce Type: new Abstract: Learning constraint-satisfying policies from offline data without risky online interaction is crucial for safety-critical decision making. Conventional

Model-Dowser: Data-Free Importance Probing to Mitigate Catastrophic Forgetting in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2602.04509v4 Announce Type: replace Abstract: Fine-tuning Multimodal Large Language Models (MLLMs) on task-specific data is an effective way to improve performance on downstream applications. Ho

Model Merging: Foundations and Algorithms

Model ReleasesDGX agent

arXiv:2605.01580v1 Announce Type: new Abstract: Modern deep learning usually treats models as separate artifacts: trained independently, specialized for particular purposes, and replaced when improved

Model Organisms Are Leaky: Perplexity Differencing Often Reveals Finetuning Objectives

ResearchDGX agent

arXiv:2605.00994v1 Announce Type: new Abstract: Finetuning can significantly modify the behavior of large language models, including introducing harmful or unsafe behaviors. To study these risks, rese

MOGO: Residual Quantized Hierarchical Causal Transformer for High-Quality and Real-Time 3D Human Motion Generation

Model ReleasesDGX agent

arXiv:2506.05952v4 Announce Type: replace Abstract: Recent advances in transformer-based text-to-motion generation have led to impressive progress in synthesizing high-quality human motion. Neverthele

Moira: Language-driven Hierarchical Reinforcement Learning for Pair Trading

ApplicationsDGX agent

arXiv:2605.01954v1 Announce Type: cross Abstract: Many sequential decision-making problems exhibit hierarchical structure, where high-level semantic choices constrain downstream actions and feedback i

Molecular Representations for Large Language Models

Model ReleasesDGX agent

arXiv:2605.01822v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly being used to support scientific discovery. In chemistry, tasks such as reaction prediction and structure

MolmoAct2: Action Reasoning Models for Real-world Deployment

Model ReleasesDGX agent

arXiv:2605.02881v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models aim to provide a single generalist controller for robots, but today's systems fall short on the criteria that matter

MolViBench: Evaluating LLMs on Molecular Vibe Coding

Model ReleasesDGX agent

arXiv:2605.02351v1 Announce Type: new Abstract: Molecular Vibe Coding, a paradigm where chemists interact with LLMs to generate executable programs for molecular tasks, has emerged as a flexible alter

Momentum-Anchored Multi-Scale Fusion Model for Long-Tailed Chest X-Ray Classification

SafetyDGX agent

arXiv:2605.02292v1 Announce Type: new Abstract: Chest X-ray classification suffers from severe class imbalance where gradient updates bias toward majority classes, causing feature drift and poor perfo

MooD: An Efficient VA-Driven Affective Image Editing Framework via Fine-Grained Semantic Control

ResearchDGX agent

arXiv:2605.02521v1 Announce Type: new Abstract: Affective image editing (AIE) aims to edit visual content to evoke target emotions. However, existing methods often overlook inference efficiency and pr

MorphIt: Flexible Spherical Approximation of Robot Morphology for Representation-driven Adaptation

Model ReleasesDGX agent

arXiv:2507.14061v2 Announce Type: replace Abstract: What if a robot could rethink its own morphological representation to better meet the demands of diverse tasks? Most robotic systems today treat the

MOSAIC: Multi-agent Orchestration for Task-Intelligent Scientific Coding

Model ReleasesDGX agent

arXiv:2510.08804v3 Announce Type: replace Abstract: We present MOSAIC, a multi-agent Large Language Model (LLM) framework for solving challenging scientific coding tasks. Unlike general-purpose coding

Motion-Aware Caching for Efficient Autoregressive Video Generation

ResearchDGX agent

arXiv:2605.01725v1 Announce Type: new Abstract: Autoregressive video generation paradigms offer theoretical promise for long video synthesis, yet their practical deployment is hindered by the computat

MPCS: Neuroplastic Continual Learning via Multi-Component Plasticity and Topology-Aware EWC

Model ReleasesDGX agent

arXiv:2605.02509v1 Announce Type: new Abstract: Continual learning systems face a fundamental tension between plasticity -- acquiring new knowledge -- and stability -- retaining prior knowledge. We in

MSMixer: Learned Multi-Scale Temporal Mixing with Complementary Linear Shortcut for Long-Term Time Series Forecasting

ResearchDGX agent

arXiv:2605.02689v1 Announce Type: new Abstract: Long-term time series forecasting requires models that simultaneously capture rapid oscillations, medium-range periodicities, and slowly evolving macro-

MTA: Multi-Granular Trajectory Alignment for Large Language Model Distillation

SafetyDGX agent

arXiv:2605.01374v1 Announce Type: new Abstract: Knowledge distillation is a key technique for compressing large language models (LLMs), but most existing methods align representations at fixed layers

MU-SHOT-Fi: Self-Supervised Multi-User Wi-Fi Sensing with Source-free Unsupervised Domain Adaptation

TutorialsDGX agent

arXiv:2605.01369v1 Announce Type: cross Abstract: Deep learning has been widely adopted for WiFi CSI-based human activity recognition (HAR) due to its ability to learn spatio-temporal features in a pr

Multi-Branch Non-Homogeneous Image Dehazing via Concentration Partitioning and Image Fusion

ResearchDGX agent

arXiv:2605.00885v1 Announce Type: new Abstract: Existing single image dehazing methods have demonstrated satisfactory performance on homogeneous thin-haze images; however, they often struggle with non

Multi-Dataset Cross-Domain Knowledge Distillation for Unified Medical Image Segmentation, Classification, and Detection

ResearchDGX agent

arXiv:2605.01563v1 Announce Type: new Abstract: We propose a unified cross-domain transfer learning framework that leverages knowledge from multiple heterogeneous medical imaging datasets to improve p

Multi-fidelity surrogates for mechanics of composites: from co-kriging to multi-fidelity neural networks

Model ReleasesDGX agent

arXiv:2605.02871v1 Announce Type: cross Abstract: Composite materials exhibit strongly hierarchical and anisotropic properties governed by coupled mechanisms spanning constituents, plies, laminates, s

Multi-Objective Instruction-Aware Representation Learning in Procedural Content Generation RL

ResearchDGX agent

arXiv:2508.09193v2 Announce Type: replace Abstract: Recent advancements in generative modeling emphasize the importance of natural language as a highly expressive and accessible modality for controlli

Multi-Perspective Transformers in ARC-AGI-2 Challenge

Model ReleasesDGX agent

arXiv:2605.01154v1 Announce Type: new Abstract: ARC-AGI-2 is a benchmark of human-intuitive visual puzzles that measures a machine's ability to generalize from limited examples, interpret symbolic mea

Multi-Rater Calibrated Segmentation Models

ResearchDGX agent

arXiv:2605.02437v1 Announce Type: new Abstract: Objective: Accurate probability estimates are essential for the safe deployment of medical image segmentation models in clinical decision-making. Howeve

Multi-Scale Gaussian-Language Map for Zero-shot Embodied Navigation and Reasoning

SafetyDGX agent

arXiv:2605.01736v1 Announce Type: new Abstract: Understanding the geometric and semantic structure of environments is essential for embodied navigation and reasoning. Existing semantic mapping methods

Multi-User Dueling Bandits: A Fair Approach using Nash Social Welfare

SafetyDGX agent

arXiv:2605.01961v1 Announce Type: new Abstract: Learning from human preference data is becoming a useful tool, from fine-tuning large language models to training reinforcement learning agents. However

Multi-View Hierarchical Representation Learning of Fetal Hemodynamics for Maternal Hypertension Detection at the Edge

SafetyDGX agent

arXiv:2605.00872v1 Announce Type: cross Abstract: Hypertensive disorders of pregnancy remain a leading cause of maternal and fetal morbidity worldwide, yet diagnosis relies on intermittent cuff-based

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety

Model ReleasesDGX agent

arXiv:2605.01687v1 Announce Type: new Abstract: We present MultiBreak, a scalable and diverse multi-turn jailbreak benchmark to evaluate large language model (LLM) safety. Multi-turn jailbreaks mimic

Multimodal Confidence Modeling in Audio-Visual Quality Assessment

ApplicationsDGX agent

arXiv:2605.01219v1 Announce Type: cross Abstract: Audio-visual quality assessment (AVQA) is essential for streaming, teleconferencing, and immersive media. In realistic streaming scenarios, distortion

Multimodal Data Curation Through Ranked Retrieval

SafetyDGX agent

arXiv:2605.01163v1 Announce Type: cross Abstract: Shared embedding spaces are widely used for multimodal search and data curation. In practice, two problems often limit how well this works. First, emb

Multiple Choice Questions: Reasoning Makes Large Language Models (LLMs) More Self-Confident, Especially When They are Wrong

Model ReleasesDGX agent

arXiv:2501.09775v3 Announce Type: replace Abstract: Multiple Choice Question (MCQ) tests are among the most used methods for evaluating large language models (LLMs). Besides checking the correctness o

MultiSense-Pneumo: A Multimodal Learning Framework for Pneumonia Screening in Resource-Constrained Settings

ApplicationsDGX agent

arXiv:2605.02207v1 Announce Type: new Abstract: Pneumonia remains a leading global cause of morbidity and mortality, particularly in low resource settings where access to imaging, laboratory testing,

Multispectral Blind Image Super-Resolution for Standing Dead Tree Segmentation

TutorialsDGX agent

arXiv:2605.02471v1 Announce Type: new Abstract: Mapping standing dead trees is crucial for acquiring information on the effects of climate change on forests and forest biodiversity. However, leveragin

MusicInfuser: Making Video Diffusion Listen and Dance

HardwareDGX agent

arXiv:2503.14505v3 Announce Type: replace Abstract: We introduce MusicInfuser, an approach that aligns pre-trained text-to-video diffusion models to generate high-quality dance videos synchronized wit

MV-S2V: Multi-View Subject-Consistent Video Generation

ApplicationsDGX agent

arXiv:2601.17756v3 Announce Type: replace Abstract: Existing Subject-to-Video Generation (S2V) methods have achieved high-fidelity and subject-consistent video generation, yet remain constrained to si

MVP-LAM: Learning Action-Centric Latent Action via Cross-Viewpoint Reconstruction

ResearchDGX agent

arXiv:2602.03668v2 Announce Type: replace-cross Abstract: Latent actions learned from diverse human videos serve as pseudo-labels for vision-language-action (VLA) pretraining, but provide effective su

← Previous
1…780781782783784…998
Next →