AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

MatSciBench: Benchmarking the Reasoning Ability of Large Language Models in Materials Science

DGX agent

arXiv:2510.12171v2 Announce Type: replace Abstract: Large Language Models have shown strong scientific reasoning ability, but their performance on materials science problems remains less studied. To f

model-releasesarxiv-cs-ai
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MBABench: Evaluating LLM Agents on End-to-End Spreadsheet Tasks in Finance

DGX agent

arXiv:2605.22664v2 Announce Type: replace Abstract: LLM agents are increasingly expected to carry out end-to-end workflows, producing complete artifacts from high-level user instructions. To meet ente

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

MC-CPO: Mastery-Conditioned Constrained Policy Optimization for Pedagogically Safe Intelligent Tutoring Systems

DGX agent

arXiv:2604.04251v2 Announce Type: replace Abstract: Intelligent tutoring systems increasingly rely on reinforcement learning to personalise instruction, yet optimising for observable engagement signal

safetyarxiv-cs-ai
9 Jun 2026
Safety

MC-PDD: Masked Corpus-Level Pretraining Data Detection for Black-Box Large Language Models

DGX agent

arXiv:2606.07996v1 Announce Type: cross Abstract: Pretraining is fundamental to the development of Large Language Models (LLMs), yet the opacity of pretraining data complicates model analysis and rais

safetyarxiv-cs-ai
9 Jun 2026
Research

Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units

DGX agent

arXiv:2601.21996v2 Announce Type: replace-cross Abstract: While Mechanistic Interpretability has identified interpretable circuits in LLMs, their causal origins in training data remain elusive. We int

researcharxiv-cs-ai
9 Jun 2026
Research

MeCo: One-Step MeanFlow-based Corrector for Multi-Channel Speech Separation

DGX agent

arXiv:2606.09677v1 Announce Type: cross Abstract: While discriminative models for multi-channel speech separation excel in reference-based metrics, they often exhibit suboptimal human listening qualit

researcharxiv-cs-ai
9 Jun 2026
Applications

MedicalRec: Medical recommender system for image classification without retraining

DGX agent

arXiv:2606.07553v1 Announce Type: cross Abstract: The emergence of machine learning and deep learning has revolutionized the efficiency of diagnostic, therapeutic, and administrative systems in health

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

MedVision: Benchmarking Quantitative Medical Image Analysis

DGX agent

arXiv:2511.18676v2 Announce Type: replace-cross Abstract: Current vision-language models (VLMs) in medicine are primarily designed for categorical question answering (e.g., 'Is this normal or abnormal

model-releasesarxiv-cs-ai
9 Jun 2026
Hardware

Meeting SLOs, Slashing Hours: Automated Enterprise LLM Optimization with OptiKIT

DGX agent

arXiv:2601.20408v2 Announce Type: replace-cross Abstract: Enterprise LLM deployment faces a critical scalability challenge: organizations must optimize models systematically to scale AI initiatives wi

hardwarearxiv-cs-ai
9 Jun 2026
Safety

Memetic Capture: A Pluralistic Policy Framework for Governing AI-Driven Cultural Disempowerment

DGX agent

arXiv:2606.07802v1 Announce Type: cross Abstract: Culture is the most insidious vector of gradual human disempowerment by AI: unlike economic or political displacement, cultural displacement attacks t

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Memory Beyond Recall: A Dual-Process Cognitive Memory System for Self-Evolving LLM Agents

DGX agent

arXiv:2606.09483v1 Announce Type: cross Abstract: Long-term memory for an LLM agent is more than retrieving the right passage at the right time. Current memory systems collapse belief revision, causal

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

MemoVAD: Resource-Efficient Video Anomaly Detection via Dynamic Semantic Memory in Edge Computing Scenarios

DGX agent

arXiv:2606.07669v1 Announce Type: cross Abstract: Deploying Video Anomaly Detection (VAD) in real-world surveillance faces a fundamental tension between the demand for high-level semantics to ensure e

local-aiarxiv-cs-ai
9 Jun 2026
Agents

MemToolAgent overview with a simple restaurant booking scenario where the agent retrieves similar memories, receives feedback on an invalid time format, and generates a reflection to update its memory

DGX agent

arXiv:2606.07909v1 Announce Type: new Abstract: Modern large language model (LLM) agents can use external tools to help users solve complex tasks. However, for problems that require learning from long

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

MEnvAgent: Scalable Polyglot Environment Construction for Verifiable Software Engineering

DGX agent

arXiv:2601.22859v3 Announce Type: replace-cross Abstract: The evolution of Large Language Model (LLM) agents for software engineering (SWE) is constrained by the scarcity of verifiable datasets, a bot

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

MetaEvo: A Meta-Optimization Framework for Experience-Driven Agent Evolution

DGX agent

arXiv:2606.07603v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong reasoning capabilities, yet most LLM-based agents are statically deployed and unable to improve through ta

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Minibatch Selection via Partition Matroid Constrained Gradient Matching

DGX agent

arXiv:2606.07954v1 Announce Type: cross Abstract: Training large language models (LLMs) on heterogeneous data requires selecting minibatches that balance convergence speed with coverage across domains

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

MIRAGE: Metadata-Integrated Repository Analysis and Guided Enhancement for MSR Datasets

DGX agent

arXiv:2606.07611v1 Announce Type: cross Abstract: This paper proposes an improved approach to the analysis of Mining Software Repositories (MSR) datasets via metadata enrichment, FAIRness assessment,

safetyarxiv-cs-ai
9 Jun 2026
Tutorials

MixReasoning: Switching Modes to Think

DGX agent

arXiv:2510.06052v2 Announce Type: replace Abstract: Reasoning models enhance performance by tackling problems in a step-by-step manner, decomposing them into sub-problems and exploring long chains of

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models

DGX agent

arXiv:2606.07706v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated strong performance across multimodal tasks, yet their safety robustness remains an open challenge. Whi

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

mllm-shap: A Shapley Value Explainability Platform for Text-Audio Multimodal Large Language Models

DGX agent

arXiv:2606.07531v1 Announce Type: cross Abstract: We introduce mllm-shap, an open-source Python framework designed to extend Shapley Value (SV) explainability from text-only Large Language Models to M

safetyarxiv-cs-ai
9 Jun 2026
Research

MM-Matryoshka: Towards Budget-Elastic Visual Document Retrieval via a 2D Multimodal Matryoshka Training Framework

DGX agent

arXiv:2606.07654v1 Announce Type: cross Abstract: Multi-vector visual document retrievers achieve strong fine-grained matching by representing each page with multiple vectors from deep Vision-Language

researcharxiv-cs-ai
9 Jun 2026
Model Releases

MMR-GRPO: Accelerating GRPO-Style Training through Diversity-Aware Reward Reweighting

DGX agent

arXiv:2601.09085v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) has become a standard approach for training mathematical reasoning models; however, its reliance on

model-releasesarxiv-cs-ai
9 Jun 2026
Tutorials

Mobility-Embedded POIs: Learning What A Place Is and How It Is Used from Human Movement

DGX agent

arXiv:2601.21149v3 Announce Type: replace-cross Abstract: Recent progress in geospatial foundation models highlights the importance of learning general-purpose representations for real-world locations

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

Model Multiplicity for Adversarial Detection in Small Language Model Training on Edge Devices

DGX agent

arXiv:2606.07857v1 Announce Type: cross Abstract: The rise of edge-based machine learning has enabled distributed adaptation of language models across mobile and IoT devices, offering privacy preserva

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

Model Poisoning Against Federated Model Adaptation with Chain of Bit-Flips

DGX agent

arXiv:2606.09548v1 Announce Type: cross Abstract: Federated Learning (FL) allows a set of clients to collectively train a global model without sharing local training data. Giving the responsibility of

local-aiarxiv-cs-ai
9 Jun 2026
Applications

Modeling the Diachronic Evolution of Legal Norms: An LRMoo-Based, Component-Level, Event-Centric Approach to Legal Knowledge Graphs

DGX agent

arXiv:2506.07853v5 Announce Type: replace Abstract: Representing the temporal evolution of legal norms is a critical challenge for automated processing. While foundational frameworks exist, they lack

applicationsarxiv-cs-ai
9 Jun 2026
Safety

Momentum for Reasoning: Dense Intrinsic Signals in Policy Optimization

DGX agent

arXiv:2606.08815v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a powerful paradigm for eliciting long-chain reasoning in large language models. Ho

safetyarxiv-cs-ai
9 Jun 2026
Research

More Bang for the Buck: Improving the Inference of Large Language Models at a Fixed Budget using Reset and Discard (ReD)

DGX agent

arXiv:2601.21522v2 Announce Type: replace-cross Abstract: The performance of large language models (LLMs) on verifiable tasks is usually measured by pass@k, the probability of answering a question cor

researcharxiv-cs-ai
9 Jun 2026
Research

More Yap Less Meaning: Uncovering Self-Improvement Behavior in SLMs

DGX agent

arXiv:2606.08471v1 Announce Type: cross Abstract: Recently, language models have made rapid progress across various domains and applications. However, their capability for self-improvement, i.e., whet

researcharxiv-cs-ai
9 Jun 2026
Research

MOSS-Video-Preview: Toward Real-Time Video Understanding via Cross-Attention

DGX agent

arXiv:2606.07639v1 Announce Type: cross Abstract: Video understanding is shifting from the offline paradigm -- taking a fully recorded video as input and producing a single answer after it ends -- tow

researcharxiv-cs-ai
9 Jun 2026
Research

Multi-planar 2D-U-Net Segmentation of 3D-CT Abdominal Organs augmented by Spatial Occurrence Maps

DGX agent

arXiv:2606.07717v1 Announce Type: cross Abstract: This work proposes a lightweight 2D-U-Net-based framework for segmenting five abdominal organs in large field-of-view 3D CT scans. The method combines

researcharxiv-cs-ai
9 Jun 2026
Agents

Multi-Turn Evaluation of Deep Research Agents Under Process-Level Feedback

DGX agent

arXiv:2606.09748v1 Announce Type: new Abstract: Existing benchmarks for deep research agents (DRAs) assess only single-shot outputs, ignoring a key question: can DRAs improve their reports when guided

agentsarxiv-cs-ai
9 Jun 2026
Safety

Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers

DGX agent

arXiv:2601.12263v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) integrate visual and textual knowledge into unified representations that increasingly underpin modern retrieval

safetyarxiv-cs-ai
9 Jun 2026
Applications

Multimodal Group Emotion Recognition In-the-Wild Towards a Privacy-Safe Non-Individual Approach

DGX agent

arXiv:2606.07585v1 Announce Type: cross Abstract: This thesis addresses group emotion recognition (GER) in-the-wild with a focus on privacy preservation. Unlike traditional emotion recognition methods

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

Multimodal Large Language Models as Synthetic Participants in Video-Based Studies: An Evaluation

DGX agent

arXiv:2606.07541v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have shown strong performance on objective tasks such as video understanding and reasoning. However, it remai

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Muon Learns More Robust and Transferable Features than Adam

DGX agent

arXiv:2606.09658v1 Announce Type: cross Abstract: Muon has recently emerged as a state-of-the-art optimizer for pretraining Large Language Models (LLMs) and vision classifiers. Despite its efficiency

researcharxiv-cs-ai
9 Jun 2026
Safety

Neuro-Symbolic Injection of LTLf Constraints in Autoregressive Reinforcement Learning Policies

DGX agent

arXiv:2606.08312v1 Announce Type: new Abstract: In this work we study offline reinforcement learning (RL) under temporally extended task constraints expressed in Linear Temporal Logic over finite trac

safetyarxiv-cs-ai
9 Jun 2026
Safety

NeuroAlign: Hierarchical Multimodal Fusion of Dynamic and Structural Neuroimaging for MCI Analysis

DGX agent

arXiv:2606.07635v1 Announce Type: cross Abstract: Multimodal neuroimaging fusion of functional MRI (fMRI) and diffusion tensor imaging (DTI) provides complementary information for cognitive impairment

safetyarxiv-cs-ai
9 Jun 2026
Safety

Neutrality Bites: Gender Representation in AI-Generated Animal Stories

DGX agent

arXiv:2606.07969v1 Announce Type: cross Abstract: Gender bias in AI-generated stories is a well-documented problem. While much attention has been paid to reducing or mitigating this bias, it is not al

safetyarxiv-cs-ai
9 Jun 2026
Applications

Next-Token Prediction Learns Generalisable Representations of Sleep Physiology

DGX agent

arXiv:2606.09605v1 Announce Type: new Abstract: Foundation models offer a promising route to compress multi-modal physiological signals into compact representations of human health, with broad applica

applicationsarxiv-cs-ai
9 Jun 2026
Research

No Free Lunch for Synthetic Images under Data Scarcity Conditions

DGX agent

arXiv:2606.07640v1 Announce Type: cross Abstract: This study investigates the trade-offs between fidelity, privacy, and utility in synthetic data generation under conditions of data scarcity and priva

researcharxiv-cs-ai
9 Jun 2026
Safety

NoRD: A Data-Efficient Vision-Language-Action Model that Drives without Reasoning

DGX agent

arXiv:2602.21172v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are advancing autonomous driving by replacing modular pipelines with unified end-to-end architectures. However,

safetyarxiv-cs-ai
9 Jun 2026
Tutorials

Not Just After One: Sleep-Inspired Replay Prevents Catastrophic Forgetting After Sequential Tasks

DGX agent

arXiv:2606.08447v1 Announce Type: cross Abstract: One of the critical limitations of artificial neural networks is their lack of ability to continually learn: training on new tasks often leads to inte

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

NutriMLLM: Multimodal Large Language Models for Dietary Micronutrient Analysis

DGX agent

arXiv:2606.08948v1 Announce Type: cross Abstract: Comprehensive estimation of dietary micronutrients from food images could improve clinical nutrition care, but training such models requires large mul

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Observability for Delegated Execution in Agentic AI Systems

DGX agent

arXiv:2606.09692v1 Announce Type: cross Abstract: Delegation-scoped execution is not identifiable from standard observables: audit logs and execution traces can be identical under multiple incompatibl

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Offline Reinforcement Learning for Plasma Control in Nuclear Fusion: Codebase and Benchmark

DGX agent

arXiv:2606.07550v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) offers a promising route for developing plasma controllers from historical tokamak data, since online trial-and-er

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics

DGX agent

arXiv:2606.09826v1 Announce Type: cross Abstract: Vision-language model (VLM) agents are increasingly deployed in interactive game environments. Yet game benchmarks for VLM agents typically report a s

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

OmniMem: Perturbation-aware Memory Compression for Streaming Audio-Visual LLMs

DGX agent

arXiv:2606.07577v1 Announce Type: new Abstract: Audio-visual large language models (LLMs) hold strong promise for long-form video understanding, yet their long-video inference is fundamentally limited

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…182183184185186…448
Next →