AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
Applications

Length-MAX Tokenizer for Language Models

DGX agent

arXiv:2511.20849v2 Announce Type: replace-cross Abstract: We introduce a new tokenizer for language models that minimizes the average tokens per character, thereby reducing the number of tokens needed

applicationsarxiv-cs-ai
11 Aug 2026
Safety

LF{}^{2}AR: Accounting for Layerwise Dynamics to Improve Multimodal Adaptation of Language Models

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2503.06211v3 Announce Type: replace-cross Abstract: Text-pretrained language models (LMs) encode rich world knowledge, but adapting them to process and generate perceptual modalities such as aud

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

LGNNIC: Acceleration of Large-Scale GNN Training using SmartNICs

DGX agent

arXiv:2608.07733v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) are widely used across domains such as natural sciences, social network analysis, chip design, and recommendation systems

model-releasesarxiv-cs-ai
11 Aug 2026
Research

LibraSpec: Dynamic Diffusion-Based Speculative Decoding via Marginal-Gain-Driven Optimization

DGX agent

arXiv:2608.08721v1 Announce Type: cross Abstract: Speculative decoding accelerates large language model inference by drafting multiple tokens for parallel verification, with efficiency critically dete

researcharxiv-cs-ai
11 Aug 2026
Research

Linearized 2-Simplicial Attention

DGX agent

arXiv:2608.09307v1 Announce Type: new Abstract: We present a linearized form of 2-simplicial attention by rewriting the trilinear score as an inner product between a composite query and a key, so that

researcharxiv-cs-ai
11 Aug 2026
Agents

Lingjing: A Simulation Testbed for Multi-Agent Embodied Tasks in Open-Ended Cities

DGX agent

arXiv:2608.08045v1 Announce Type: new Abstract: Urban embodied intelligence requires coordination among heterogeneous agents (e.g., UAVs, ground robots, and autonomous vehicles) in dynamic cities. Sim

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

Listen, See and Track: Spatio-Temporal Audio-Visual Sound Event Reasoning for Omni-Modal Language Models

DGX agent

arXiv:2608.09435v1 Announce Type: new Abstract: Understanding dynamic sound sources requires jointly determining what produces a sound, where the source is located, and how it moves over time. Yet exi

model-releasesarxiv-cs-ai
11 Aug 2026
Research

LITEWAY: LIghtweight HAR via Temporal Efficient highWAY

DGX agent

arXiv:2608.09421v1 Announce Type: cross Abstract: Wearable human activity recognition (HAR) remains challenging due to the computational and energy constraints of deep learning models on resource-limi

researcharxiv-cs-ai
11 Aug 2026
Model Releases

LLM-Driven AutoML for Cross-Lingual Handwritten OCR: Closed-Loop Neural Architecture Search with GPT-5, GPT-4o, and Claude Sonnet 4

DGX agent

arXiv:2607.15509v2 Announce Type: replace-cross Abstract: We present a fully automated closed-loop AutoML framework that uses GPT-5, GPT-4o, and Claude Sonnet 4 as autonomous neural architecture desig

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

LLM-Guided Heuristic Design from Simulation Traces: A Case Study in Dynamic Production and AGV Scheduling

DGX agent

arXiv:2608.09343v1 Announce Type: new Abstract: Simulation-based optimization (SBO) evaluates executable policies under stochastic dynamics, but most methods treat the simulator as a black box: aggreg

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

LLM Reasoning for Subjective Tasks: Failure Modes, Mitigation, and Dynamic Reasoning Routing

DGX agent

arXiv:2608.08889v1 Announce Type: new Abstract: Recommendation systems thrive on personalization, where ''correctness'' is rarely a binary truth but a matter of subjective human preference. As Large L

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

LLM within MCP Matters: Measuring Inefficient Resource Utilization Driven by LLMs

DGX agent

arXiv:2608.08467v1 Announce Type: new Abstract: The Model Context Protocol (MCP) standardizes how servers expose data and tools to Large Language Models (LLMs). A common server design embeds frequentl

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

LLMs Remember First, Forget Last: Dual-Process Interference in Large Language Models

DGX agent

arXiv:2603.00270v3 Announce Type: replace-cross Abstract: Large language models can process millions of tokens, yet how they handle conflicting information within context remains poorly understood. Fr

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

LLMVisor: A Real-Time Latency Attribution Model for Multi-Tenant LLM Serving

DGX agent

arXiv:2608.08382v1 Announce Type: new Abstract: As LLM inference shifts to multi-tenant GPU clusters, co-batching improves throughput but obscures per-tenant usage and limits control. Enabling fractio

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Locating Failure in Multi-Page Visually Rich Document Understanding: An Empirical Attribution

DGX agent

arXiv:2608.07943v1 Announce Type: new Abstract: Multi-page visually-rich document understanding (MP-VRDU) requires managing evidence that is sparse, spread across pages, and often exceeds a model's co

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Long SKILL Compliance as Logical Reasoning: Closure-Grounded Detection with Scaling-Guided On-Policy Distillation

DGX agent

arXiv:2608.08146v1 Announce Type: new Abstract: The increasing complexity of enterprise business scenarios has promoted the widespread adoption of long SKILL documents in agent systems, posing new cha

model-releasesarxiv-cs-ai
11 Aug 2026
Research

LookME: Lookup-Based Multimodal Embeddings for Layer Injection in Vision-Language Models

DGX agent

arXiv:2607.16305v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have achieved strong progress in multimodal understanding. However, scaling dense or sparse Mixture-of-Experts (

researcharxiv-cs-ai
11 Aug 2026
Model Releases

LoRSA: Toward Generalizable Parameter-Efficient Fine-Tuning for Biomedical Downstream Tasks

DGX agent

arXiv:2608.07749v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning enables the adaptation of vision foundation models to biomedical tasks under limited computational resources, but a si

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

M^3Prune: Hierarchical Communication Graph Pruning for Efficient Multi-Modal Multi-Agent Retrieval-Augmented Generation

DGX agent

arXiv:2511.19969v2 Announce Type: replace Abstract: Recent advancements in multi-modal retrieval-augmented generation (mRAG), which enhance multi-modal large language models (MLLMs) with external know

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

MADBench: A Benchmark for Modality-Aware Audio Deepfake Detection

DGX agent

arXiv:2608.09593v1 Announce Type: cross Abstract: Recent advances in speech synthesis and audio generation have made high-fidelity acoustic forgery low-cost and difficult to attribute, enabling a real

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

MARA: Flow-Matching-Guided Multi-Agent Resource Allocation for Computational Resource Efficient Learning

DGX agent

arXiv:2608.09130v1 Announce Type: cross Abstract: Allocating limited computation among concurrent learning tasks is difficult when each task must reach a target loss before a deadline but its required

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

MasDrift: Benchmarking Authorization Preservation Across Multi-Agent Architectures

DGX agent

arXiv:2608.07556v1 Announce Type: cross Abstract: Multi-agent systems (MAS) decompose long-horizon tasks across supervisors and subagents, but delegated goals do not necessarily carry their original a

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Matching Supervision to the Student's Learning Capacity: A Unified Framework for On-Policy Self-Distillation

DGX agent

arXiv:2608.08176v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) improves the reasoning abilities of LLMs by internalizing privileged context into model parameters through self-disti

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks

DGX agent

arXiv:2507.03162v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has transformed various domains, particularly computer science (CS) education. These mod

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MathShikkha: A Controlled Study of Answer-Only and Chain-of-Thought Supervision for Bangla Mathematical Reasoning in Small Language Models

DGX agent

arXiv:2608.08503v1 Announce Type: new Abstract: Mathematical reasoning remains challenging in low-resource languages such as Bangla. We study whether teacher-generated Bangla Chain-of-Thought (CoT) su

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Matryoshka Language Model Suites

DGX agent

arXiv:2608.09703v1 Announce Type: new Abstract: Training a language model suite classically requires training each model separately and serving them independently. We improve both training and inferen

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MCIF: Multimodal Crosslingual Instruction-Following Benchmark from Scientific Talks

DGX agent

arXiv:2507.19634v4 Announce Type: replace-cross Abstract: Recent advances in large language models have laid the foundation for multimodal LLMs (MLLMs), which unify text, speech, and vision within a s

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Measuring the Wrong Thing: Internal Harmfulness Scores Anti-Rank Successful Jailbreaks

DGX agent

arXiv:2608.09624v1 Announce Type: cross Abstract: Internal safety scores judge a prompt before any text is generated, and they are validated by how well they separate harmful prompts from benign ones.

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Med-CRAFT: An Information System for Explainable and Configurable Construction of Multimodal Medical QA Datasets

DGX agent

arXiv:2512.01045v2 Announce Type: replace Abstract: Data-intensive artificial intelligence applications increasingly rely on large-scale, high-quality, explainable, and reproducible datasets, yet the

researcharxiv-cs-ai
11 Aug 2026
Safety

MedCalc-R1: Knowledge-Guided Reward Framework for Medical Mathematical Reasoning

DGX agent

arXiv:2608.08623v1 Announce Type: new Abstract: In Reinforcement Learning with Verifiable Rewards (RLVR) frameworks for mathematical reasoning tasks, floating-point results are typically evaluated usi

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

MedPixel: A Unified Pixel-Language Model for Medical Reasoning and Segmentation

DGX agent

arXiv:2608.09818v1 Announce Type: cross Abstract: Reliable medical image understanding requires models to connect clinical language and visual reasoning with pixel-level grounding. Yet medical vision-

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MELLON - Multimodal Enhanced LLM for Online Navigation

DGX agent

arXiv:2608.09121v1 Announce Type: new Abstract: Web navigation agents are capable of addressing various types of tasks on different websites. Current baselines on web navigation are either unimodal or

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

Mendel Godel Machine: Recursive Self-Improving Coding Agents via Comparative Evolution

DGX agent

arXiv:2608.07645v1 Announce Type: new Abstract: Self-improving coding agents that iteratively rewrite their own source code have demonstrated impressive performance on coding tasks. However, existing

agentsarxiv-cs-ai
11 Aug 2026
Hardware

Mesh-Attention: A New Communication-Efficient Distributed Attention with Improved Data Locality

DGX agent

arXiv:2512.20968v2 Announce Type: replace-cross Abstract: Distributed attention is essential for scaling large language models (LLMs) to long contexts, yet existing methods either have limited paralle

hardwarearxiv-cs-ai
11 Aug 2026
Applications

Metadata Reconstruction from Values Alone: Recovering Column Semantics in Undocumented Warehouses

DGX agent

arXiv:2608.07946v1 Announce Type: cross Abstract: Text-to-SQL benchmarks ship schemas whose column names already say what the columns mean. Production warehouses are the inverse: cryptic identifiers,

applicationsarxiv-cs-ai
11 Aug 2026
Safety

Metanormative Theory for RL-Based Moral Agents

DGX agent

arXiv:2608.08220v1 Announce Type: new Abstract: The overlapping disciplines of machine ethics and value alignment are concerned with designing artificial agents that are aligned with human values and

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

MetaSpace: Metamorphic Testing for Spatial Cognition in Embodied Agents

DGX agent

arXiv:2608.07533v1 Announce Type: new Abstract: An embodied agent is an intelligent entity that interacts with its environment through a physical body. Currently, the evaluation of embodied agents pri

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Mismatch Matters: On-Policy Distillation Beyond Token Agreement

DGX agent

arXiv:2608.09836v1 Announce Type: new Abstract: On-policy distillation (OPD) has emerged as a core component of modern LLM post-training pipelines, yet we reveal a failure mode: degenerate agreement,

safetyarxiv-cs-ai
11 Aug 2026
Safety

Mitigating Gender Bias in English to Romanian Machine Translation

DGX agent

arXiv:2608.08606v1 Announce Type: cross Abstract: Machine translation (MT) systems often fail to correctly translate gender, especially when converting from a gender-neutral language like English to a

safetyarxiv-cs-ai
11 Aug 2026
Research

Mitigating Over-Personalization in LLMs via Structured Memory

DGX agent

arXiv:2608.08300v1 Announce Type: new Abstract: Conversational assistants increasingly rely on persistent long-term memory to personalize responses across sessions. However, when stored user informati

researcharxiv-cs-ai
11 Aug 2026
Research

MixFormer: Linear Transformer with Mixture of Memory Experts

DGX agent

arXiv:2608.09468v1 Announce Type: cross Abstract: State Space Models (SSMs), as a mainstream research direction of linear Transformers, aim to achieve higher efficiency than standard Transformers in l

researcharxiv-cs-ai
11 Aug 2026
Model Releases

MMArch: Benchmarking Multimodal Reasoning Grounded in Architectural Evidence

DGX agent

arXiv:2608.09281v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) perform strongly on engineering imagery, yet existing benchmarks mostly test drawing recognition, information e

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models

DGX agent

arXiv:2608.09696v1 Announce Type: new Abstract: Predicting the answer to interventional ``what if'' questions --- the outcome of an action never taken --- requires a mechanistic, causal model, not a c

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Modern Backbones Improve Multi-task DETR for Mammography Classification and Lesion Localization

DGX agent

arXiv:2608.09801v1 Announce Type: cross Abstract: Joint exam-level prediction and candidate-region localization may improve the usefulness of AI support in mammography. We study this setting using a m

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

MonitorBench: A Comprehensive Benchmark for Chain-of-Thought Monitorability in Large Language Models

DGX agent

arXiv:2603.28590v3 Announce Type: replace Abstract: Large language models (LLMs) can generate chains of thought (CoTs) that are not always causally responsible for their final outputs. When such a mis

model-releasesarxiv-cs-ai
11 Aug 2026
Research

MoNo: Multiscale Optimal Transport Neural Operator for Solving PDEs on General Geometries

DGX agent

arXiv:2608.09764v1 Announce Type: cross Abstract: Transformer-based neural operators have achieved substantial progress in solving Partial Differential Equations (PDEs) by projecting spatial observati

researcharxiv-cs-ai
11 Aug 2026
Research

Monotonicity-Guided Bottom-Up Petri Net Discovery: The SPECpp Framework

DGX agent

arXiv:2608.09398v1 Announce Type: cross Abstract: Process discovery is one of the central challenges in process mining. Petri nets are particularly attractive because simple local constructs can expre

researcharxiv-cs-ai
11 Aug 2026
Model Releases

MoRSE: Task-Oriented Multi-Agent System with Mixture of Role-Subtask Experts

DGX agent

arXiv:2608.09251v1 Announce Type: cross Abstract: Large language model-based multi-agent systems have recently shown strong potential for complex, long-horizon tasks. However, existing methods mainly

model-releasesarxiv-cs-ai
11 Aug 2026
← Previous
1…1112131415…438
Next →