AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Agents

MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks

DGX agent

arXiv:2603.02630v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved great success in many real-world applications, especially the one serving as the cognitive backbone

agentsarxiv-cs-ai
1 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MAVEN: Improving Generalization in Agentic Tool Calling

DGX agent

arXiv:2605.30738v1 Announce Type: new Abstract: Generalization across agentic tool-calling environments remains a central challenge for reliable agentic reasoning systems. Although large language mode

model-releasesarxiv-cs-ai
1 Jun 2026
Research

Mechanistic Interpretability as Statistical Estimation: A Variance Analysis

DGX agent

arXiv:2510.00845v4 Announce Type: replace-cross Abstract: Mechanistic Interpretability (MI) aims to reverse-engineer model behaviors by identifying functional sub-networks. Yet, the scientific validit

researcharxiv-cs-ai
1 Jun 2026
Applications

MechVQA: Benchmarking and Enhancing Multimodal LLMs on Comprehensive Mechanical Drawing Understanding

DGX agent

arXiv:2605.30794v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated significant achievements in general visual question answering (VQA) tasks. However, they re

applicationsarxiv-cs-ai
1 Jun 2026
Agents

MedCoG: Maximizing LLM Inference Density in Medical Reasoning via Meta-Cognitive Regulation

DGX agent

arXiv:2602.07905v2 Announce Type: replace Abstract: Large Language Models (LLMs) have shown strong potential in complex medical reasoning yet face diminishing gains under inference scaling laws. While

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts

DGX agent

arXiv:2509.12440v3 Announce Type: replace-cross Abstract: Deploying Large Language Models (LLMs) in medical applications requires fact-checking capabilities to ensure patient safety and regulatory com

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode

DGX agent

arXiv:2605.30571v1 Announce Type: cross Abstract: Physical AI systems, including robots, autonomous vehicles, embodied agents and edge copilots, often run a different inference workload from cloud LLM

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

Mental Damage: Caption Poisoning Attacks on Retrieval-Augmented Text-to-Music Generation

DGX agent

arXiv:2605.30365v1 Announce Type: cross Abstract: Retrieval-augmented text-to-music (TTM) systems augment underspecified user prompts using captions retrieved from a music caption dataset. This design

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

MIMO: Multilingual Information Retrieval via Monolingual Objectives

DGX agent

arXiv:2605.31171v1 Announce Type: cross Abstract: Multilingual Information Retrieval (MLIR) reflects real-world search environments in which queries and relevant documents may appear in different lang

model-releasesarxiv-cs-ai
1 Jun 2026
Research

MindVoice: Reconstructing Intelligible Speech from Non-invasive Neural Signals with Pretrained Priors

DGX agent

arXiv:2605.31173v1 Announce Type: cross Abstract: Reconstructing continuous speech from non-invasive neural recordings is a fundamental problem for probing human auditory perception and building safe,

researcharxiv-cs-ai
1 Jun 2026
Research

Mixture of Concept Bottleneck Experts

DGX agent

arXiv:2602.02886v2 Announce Type: replace-cross Abstract: Concept Bottleneck Models (CBMs) promote interpretability by grounding predictions in human-understandable concepts. However, existing CBMs ty

researcharxiv-cs-ai
1 Jun 2026
Local Ai

Mixture of Horizons in Action Chunking

DGX agent

arXiv:2511.19433v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models have shown remarkable capabilities in robotic manipulation, but their performance is sensitive to the extb

local-aiarxiv-cs-ai
1 Jun 2026
Safety

Multi-Agent Teams Hold Experts Back

DGX agent

arXiv:2602.01011v4 Announce Type: replace-cross Abstract: Multi-agent LLM systems are increasingly deployed as autonomous collaborators, where agents interact freely rather than execute fixed, pre-spe

safetyarxiv-cs-ai
1 Jun 2026
Applications

Neither Replacement nor Panacea: Comparing LLM-Based Conversational and Graphical Decision Support in Industrial Tasks

DGX agent

arXiv:2605.31287v1 Announce Type: cross Abstract: Managers in manufacturing settings rely on digital interfaces to interpret operational data for decision-making, but growing data volume and complexit

applicationsarxiv-cs-ai
1 Jun 2026
Agents

NEMO: Execution-Aware Optimization Modeling via Autonomous Coding Agents

DGX agent

arXiv:2601.21372v2 Announce Type: replace Abstract: We present NEMO, a system that translates Natural-language descriptions of decision problems into formal Executable Mathematical Optimization implem

agentsarxiv-cs-ai
1 Jun 2026
Local Ai

Neuro-Symbolic Predictive Process Monitoring

DGX agent

arXiv:2509.00834v2 Announce Type: replace Abstract: This paper addresses the problem of suffix prediction in Business Process Management (BPM) by proposing a Neuro-Symbolic Predictive Process Monitori

local-aiarxiv-cs-ai
1 Jun 2026
Model Releases

Neuro-symbolic Syntactic Parsing: Shaping a Neural Network with the CYK Algorithm

DGX agent

arXiv:2605.31421v1 Announce Type: cross Abstract: In this paper, we show the possibility of a direct injection of algorithms into neural network architecture. We focus on a complex algorithm, that is,

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

NGDBench: Towards Neural Graph Data Management

DGX agent

arXiv:2603.05529v2 Announce Type: replace-cross Abstract: Data critical to real-world decision-making is increasingly found within organizations. Such data is heterogeneous, constantly evolving, and o

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Not All Synthetic Data Is Yours to Learn From

DGX agent

arXiv:2605.31126v1 Announce Type: cross Abstract: Can a language model improve from plain text sampled from itself, with no prompts, no teacher, no verifier, and no reward model? Yes, but only when th

model-releasesarxiv-cs-ai
1 Jun 2026
Applications

NumLeak: Public Numeric Benchmarks as Latent Labels in Foundation Models

DGX agent

arXiv:2605.30393v1 Announce Type: cross Abstract: Public numeric benchmarks appear in pretraining, so an evaluation that conditions on a date may be measuring memorized recall rather than out-of-sampl

applicationsarxiv-cs-ai
1 Jun 2026
Model Releases

OBCache: Optimal Brain KV Cache Pruning for Efficient Long-Context LLM Inference

DGX agent

arXiv:2510.07651v2 Announce Type: replace-cross Abstract: Large language models (LLMs) with extended context windows enable powerful applications but impose significant memory overhead, as caching all

model-releasesarxiv-cs-ai
1 Jun 2026
Research

OBLIQ-Bench: Exposing Overlooked Bottlenecks in Modern Retrievers with Latent and Implicit Queries

DGX agent

arXiv:2605.06235v2 Announce Type: replace-cross Abstract: Retrieval benchmarks are increasingly saturating, but we argue that efficient search is far from a solved problem. We identify a class of quer

researcharxiv-cs-ai
1 Jun 2026
Applications

OLG++: A Semantic Extension of Obligation Logic Graph

DGX agent

arXiv:2507.05488v2 Announce Type: replace Abstract: We present OLG++, a semantic extension of the Obligation Logic Graph (OLG) for modeling regulatory and legal rules in municipal and interjurisdictio

applicationsarxiv-cs-ai
1 Jun 2026
Hardware

On Efficient Scaling of GNNs via IO-Aware Layers Implementations

DGX agent

arXiv:2605.31500v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) are bottlenecked by sparse, irregular memory access. Popular frameworks such as DGL and PyTorch Geometric support general

hardwarearxiv-cs-ai
1 Jun 2026
Research

On Revisiting Entropy for Identifying Mislabeled Images

DGX agent

arXiv:2605.31090v1 Announce Type: cross Abstract: Mislabeled samples in training datasets severely degrade the performance of deep networks, as overparameterized models tend to memorize erroneous labe

researcharxiv-cs-ai
1 Jun 2026
Research

On the impact of retrieved content representations in RAG Pipelines

DGX agent

arXiv:2605.30790v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) supplements a language model's input with retrieved documents, yet most RAG pipelines inherit retrieval component

researcharxiv-cs-ai
1 Jun 2026
Model Releases

On the Robustness of Multilingual Text Embedding Rankings Across Learning Tasks, Languages, and Benchmark Datasets

DGX agent

arXiv:2605.31142v1 Announce Type: cross Abstract: Large-scale multilingual text embedding models play crucial role in both research and industry, yet their behavior in language-specific, multi-task se

model-releasesarxiv-cs-ai
1 Jun 2026
Research

OpenSTBench: Beyond Semantic Evaluation for Speech Translation

DGX agent

arXiv:2605.30792v1 Announce Type: cross Abstract: Speech translation systems increasingly span speech-to-text translation (S2TT), speech-to-speech translation (S2ST), offline translation, and streamin

researcharxiv-cs-ai
1 Jun 2026
Applications

OrcaRouter: A Production-Oriented LLM Router with Hybrid Offline-Online Learning

DGX agent

arXiv:2605.30736v1 Announce Type: cross Abstract: The rapid development of large language models, each with distinct capabilities and inference costs, raises a practical deployment question: given an

applicationsarxiv-cs-ai
1 Jun 2026
Safety

Organizational Adaptation to Generative AI in Cybersecurity

DGX agent

arXiv:2506.12060v2 Announce Type: replace-cross Abstract: Cybersecurity organizations are adapting to GenAI integration through modified frameworks and hybrid operational processes, with success influ

safetyarxiv-cs-ai
1 Jun 2026
Safety

PAC-Bayesian Reinforcement Learning Trains Generalizable Policies

DGX agent

arXiv:2510.10544v3 Announce Type: replace-cross Abstract: We derive a novel PAC-Bayesian generalization bound for reinforcement learning that explicitly accounts for Markov dependencies in the data, t

safetyarxiv-cs-ai
1 Jun 2026
Research

ParalESN: Enabling parallel information processing in Reservoir Computing

DGX agent

arXiv:2601.22296v2 Announce Type: replace-cross Abstract: Reservoir Computing (RC) has established itself as an efficient paradigm for temporal processing. However, its scalability remains severely co

researcharxiv-cs-ai
1 Jun 2026
Safety

PASTA: A Scalable Framework for Multi-Policy AI Compliance Evaluation

DGX agent

arXiv:2601.11702v3 Announce Type: replace-cross Abstract: AI compliance is becoming increasingly critical as AI systems grow more powerful and pervasive. Yet the rapid expansion of AI policies creates

safetyarxiv-cs-ai
1 Jun 2026
Local Ai

PatchWorld: Gradient-Free Optimization of Executable World Models

DGX agent

arXiv:2605.30880v1 Announce Type: cross Abstract: Text-agent environments are typically modeled as partially observable Markov decision processes (POMDPs), assuming that the simulator's latent state a

local-aiarxiv-cs-ai
1 Jun 2026
Applications

Performance and Complexity Trade-off Optimization of Speech Models During Training

DGX agent

arXiv:2601.13704v3 Announce Type: replace-cross Abstract: In speech machine learning, neural network models are typically designed by choosing an architecture with fixed layer sizes and structure. The

applicationsarxiv-cs-ai
1 Jun 2026
Research

Personalized to Persuade: The Effects of Contextualization and Warmth on Trust and Reliance in Conversational AI

DGX agent

arXiv:2605.31275v1 Announce Type: cross Abstract: Artificial Intelligence (AI) agents personalize their responses by tailoring explanations to users' backgrounds, interests, and prior interactions, re

researcharxiv-cs-ai
1 Jun 2026
Model Releases

PhyDrawGen: Physically Grounded Diagram Generation from Natural Language

DGX agent

arXiv:2605.30512v1 Announce Type: new Abstract: Generating physics diagrams from text requires strict adherence to physical laws. While current generative models produce visually plausible outputs, th

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Physically Viable World Models: A Case for Query-Conditioned Embodied AI

DGX agent

arXiv:2605.30542v1 Announce Type: new Abstract: World models for embodied AI must be physically viable: constructed to answer intervention queries by representing the physical structure governing acti

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

PictSure: Pretraining Embeddings Matters for In-Context Learning Image Classifiers

DGX agent

arXiv:2506.14842v2 Announce Type: replace-cross Abstract: Building image classification models remains cumbersome in data-scarce domains, where collecting large labeled datasets is impractical. In-con

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

PInVerify: An Offline Embodied Benchmark for Active Instance Verification

DGX agent

arXiv:2605.30639v1 Announce Type: cross Abstract: Embodied agents have made strong progress in navigating to target objects, but reaching the goal vicinity does not guarantee that the agent has found

model-releasesarxiv-cs-ai
1 Jun 2026
Hardware

PithTrain: A Compact and Agent-Native MoE Training System

DGX agent

arXiv:2605.31463v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for frontier language models. To meet this demand, production frameworks have built opti

hardwarearxiv-cs-ai
1 Jun 2026
Model Releases

Plain Transformers are Surprisingly Powerful Link Predictors

DGX agent

arXiv:2602.01553v2 Announce Type: replace-cross Abstract: Link prediction is a core challenge in graph machine learning, demanding models that capture rich and complex topological dependencies. While

model-releasesarxiv-cs-ai
1 Jun 2026
Research

Planner-Centric Reinforcement Learning for Deep Research with Structure-Aware Reward

DGX agent

arXiv:2605.30824v1 Announce Type: new Abstract: Deep research tasks require LLMs to plan what to investigate, retrieve evidence, and synthesize long-form answers across multiple branches of inquiry. E

researcharxiv-cs-ai
1 Jun 2026
Research

Position: Evaluation of ECG Representations Must Be Fixed

DGX agent

arXiv:2602.17531v2 Announce Type: replace-cross Abstract: This position paper argues that current benchmarking practice in 12-lead ECG representation learning must be fixed to ensure progress is relia

researcharxiv-cs-ai
1 Jun 2026
Applications

Positional versus Symbolic Attention Heads: Learning Dynamics, RoPE Geometry, and Length Generalization

DGX agent

arXiv:2605.31558v1 Announce Type: cross Abstract: Transformer-based language models are widespread in today's society. As such, understanding the mechanisms by which they solve structured tasks and pr

applicationsarxiv-cs-ai
1 Jun 2026
Research

Post-Training LLMs as Better Decision-Making Agents: A Regret-Minimization Approach

DGX agent

arXiv:2511.04393v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as 'agents' for decision-making (DM) in interactive and dynamic environments. Yet, since they

researcharxiv-cs-ai
1 Jun 2026
Tutorials

Practical Cross-Band Channel Prediction for AI-RAN via Physics-Guided Deep Unfolding

DGX agent

arXiv:2605.31279v1 Announce Type: cross Abstract: To make cross-band channel prediction practical for AI-native RAN, algorithms must generalize across diverse environments and support real-time infere

tutorialsarxiv-cs-ai
1 Jun 2026
Safety

PReMISE: Policy Rubrics as Measurement Specifications for LLM Judges

DGX agent

arXiv:2605.30803v1 Announce Type: new Abstract: LLM judges are increasingly used to evaluate open-ended responses, but their scores depend strongly on the rubrics that condition them. A vague rubric a

safetyarxiv-cs-ai
1 Jun 2026
← Previous
1…236237238239240…452
Next →