AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
1 Jun 2026

Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models

TutorialsDGX agent

arXiv:2605.31603v1 Announce Type: cross Abstract: Connector-based video unified models have demonstrated strong capability in instruction-grounded video synthesis, but integrating a large high-fidelit

MAECO-Lite: Modular Ontology for Dynamic Malware Analysis

ResearchDGX agent

arXiv:2605.31199v1 Announce Type: cross Abstract: Capturing dynamic malware behavior in a practical but still semantically precise manner remains a significant challenge in cyber threat intelligence.

MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks

AgentsDGX agent

arXiv:2603.02630v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved great success in many real-world applications, especially the one serving as the cognitive backbone


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MAVEN: Improving Generalization in Agentic Tool Calling

Model ReleasesDGX agent

arXiv:2605.30738v1 Announce Type: new Abstract: Generalization across agentic tool-calling environments remains a central challenge for reliable agentic reasoning systems. Although large language mode

Mechanistic Interpretability as Statistical Estimation: A Variance Analysis

ResearchDGX agent

arXiv:2510.00845v4 Announce Type: replace-cross Abstract: Mechanistic Interpretability (MI) aims to reverse-engineer model behaviors by identifying functional sub-networks. Yet, the scientific validit

MechVQA: Benchmarking and Enhancing Multimodal LLMs on Comprehensive Mechanical Drawing Understanding

ApplicationsDGX agent

arXiv:2605.30794v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated significant achievements in general visual question answering (VQA) tasks. However, they re

MedCoG: Maximizing LLM Inference Density in Medical Reasoning via Meta-Cognitive Regulation

AgentsDGX agent

arXiv:2602.07905v2 Announce Type: replace Abstract: Large Language Models (LLMs) have shown strong potential in complex medical reasoning yet face diminishing gains under inference scaling laws. While

MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts

Model ReleasesDGX agent

arXiv:2509.12440v3 Announce Type: replace-cross Abstract: Deploying Large Language Models (LLMs) in medical applications requires fact-checking capabilities to ensure patient safety and regulatory com

Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode

Model ReleasesDGX agent

arXiv:2605.30571v1 Announce Type: cross Abstract: Physical AI systems, including robots, autonomous vehicles, embodied agents and edge copilots, often run a different inference workload from cloud LLM

Mental Damage: Caption Poisoning Attacks on Retrieval-Augmented Text-to-Music Generation

SafetyDGX agent

arXiv:2605.30365v1 Announce Type: cross Abstract: Retrieval-augmented text-to-music (TTM) systems augment underspecified user prompts using captions retrieved from a music caption dataset. This design

MIMO: Multilingual Information Retrieval via Monolingual Objectives

Model ReleasesDGX agent

arXiv:2605.31171v1 Announce Type: cross Abstract: Multilingual Information Retrieval (MLIR) reflects real-world search environments in which queries and relevant documents may appear in different lang

MindVoice: Reconstructing Intelligible Speech from Non-invasive Neural Signals with Pretrained Priors

ResearchDGX agent

arXiv:2605.31173v1 Announce Type: cross Abstract: Reconstructing continuous speech from non-invasive neural recordings is a fundamental problem for probing human auditory perception and building safe,

Mixture of Concept Bottleneck Experts

ResearchDGX agent

arXiv:2602.02886v2 Announce Type: replace-cross Abstract: Concept Bottleneck Models (CBMs) promote interpretability by grounding predictions in human-understandable concepts. However, existing CBMs ty

Mixture of Horizons in Action Chunking

Local AiDGX agent

arXiv:2511.19433v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models have shown remarkable capabilities in robotic manipulation, but their performance is sensitive to the extb

Multi-Agent Teams Hold Experts Back

SafetyDGX agent

arXiv:2602.01011v4 Announce Type: replace-cross Abstract: Multi-agent LLM systems are increasingly deployed as autonomous collaborators, where agents interact freely rather than execute fixed, pre-spe

Neither Replacement nor Panacea: Comparing LLM-Based Conversational and Graphical Decision Support in Industrial Tasks

ApplicationsDGX agent

arXiv:2605.31287v1 Announce Type: cross Abstract: Managers in manufacturing settings rely on digital interfaces to interpret operational data for decision-making, but growing data volume and complexit

NEMO: Execution-Aware Optimization Modeling via Autonomous Coding Agents

AgentsDGX agent

arXiv:2601.21372v2 Announce Type: replace Abstract: We present NEMO, a system that translates Natural-language descriptions of decision problems into formal Executable Mathematical Optimization implem

Neuro-Symbolic Predictive Process Monitoring

Local AiDGX agent

arXiv:2509.00834v2 Announce Type: replace Abstract: This paper addresses the problem of suffix prediction in Business Process Management (BPM) by proposing a Neuro-Symbolic Predictive Process Monitori

Neuro-symbolic Syntactic Parsing: Shaping a Neural Network with the CYK Algorithm

Model ReleasesDGX agent

arXiv:2605.31421v1 Announce Type: cross Abstract: In this paper, we show the possibility of a direct injection of algorithms into neural network architecture. We focus on a complex algorithm, that is,

NGDBench: Towards Neural Graph Data Management

Model ReleasesDGX agent

arXiv:2603.05529v2 Announce Type: replace-cross Abstract: Data critical to real-world decision-making is increasingly found within organizations. Such data is heterogeneous, constantly evolving, and o

Not All Synthetic Data Is Yours to Learn From

Model ReleasesDGX agent

arXiv:2605.31126v1 Announce Type: cross Abstract: Can a language model improve from plain text sampled from itself, with no prompts, no teacher, no verifier, and no reward model? Yes, but only when th

NumLeak: Public Numeric Benchmarks as Latent Labels in Foundation Models

ApplicationsDGX agent

arXiv:2605.30393v1 Announce Type: cross Abstract: Public numeric benchmarks appear in pretraining, so an evaluation that conditions on a date may be measuring memorized recall rather than out-of-sampl

OBCache: Optimal Brain KV Cache Pruning for Efficient Long-Context LLM Inference

Model ReleasesDGX agent

arXiv:2510.07651v2 Announce Type: replace-cross Abstract: Large language models (LLMs) with extended context windows enable powerful applications but impose significant memory overhead, as caching all

OBLIQ-Bench: Exposing Overlooked Bottlenecks in Modern Retrievers with Latent and Implicit Queries

ResearchDGX agent

arXiv:2605.06235v2 Announce Type: replace-cross Abstract: Retrieval benchmarks are increasingly saturating, but we argue that efficient search is far from a solved problem. We identify a class of quer

OLG++: A Semantic Extension of Obligation Logic Graph

ApplicationsDGX agent

arXiv:2507.05488v2 Announce Type: replace Abstract: We present OLG++, a semantic extension of the Obligation Logic Graph (OLG) for modeling regulatory and legal rules in municipal and interjurisdictio

On Efficient Scaling of GNNs via IO-Aware Layers Implementations

HardwareDGX agent

arXiv:2605.31500v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) are bottlenecked by sparse, irregular memory access. Popular frameworks such as DGL and PyTorch Geometric support general

On Revisiting Entropy for Identifying Mislabeled Images

ResearchDGX agent

arXiv:2605.31090v1 Announce Type: cross Abstract: Mislabeled samples in training datasets severely degrade the performance of deep networks, as overparameterized models tend to memorize erroneous labe

On the impact of retrieved content representations in RAG Pipelines

ResearchDGX agent

arXiv:2605.30790v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) supplements a language model's input with retrieved documents, yet most RAG pipelines inherit retrieval component

On the Robustness of Multilingual Text Embedding Rankings Across Learning Tasks, Languages, and Benchmark Datasets

Model ReleasesDGX agent

arXiv:2605.31142v1 Announce Type: cross Abstract: Large-scale multilingual text embedding models play crucial role in both research and industry, yet their behavior in language-specific, multi-task se

OpenSTBench: Beyond Semantic Evaluation for Speech Translation

ResearchDGX agent

arXiv:2605.30792v1 Announce Type: cross Abstract: Speech translation systems increasingly span speech-to-text translation (S2TT), speech-to-speech translation (S2ST), offline translation, and streamin

OrcaRouter: A Production-Oriented LLM Router with Hybrid Offline-Online Learning

ApplicationsDGX agent

arXiv:2605.30736v1 Announce Type: cross Abstract: The rapid development of large language models, each with distinct capabilities and inference costs, raises a practical deployment question: given an

Organizational Adaptation to Generative AI in Cybersecurity

SafetyDGX agent

arXiv:2506.12060v2 Announce Type: replace-cross Abstract: Cybersecurity organizations are adapting to GenAI integration through modified frameworks and hybrid operational processes, with success influ

PAC-Bayesian Reinforcement Learning Trains Generalizable Policies

SafetyDGX agent

arXiv:2510.10544v3 Announce Type: replace-cross Abstract: We derive a novel PAC-Bayesian generalization bound for reinforcement learning that explicitly accounts for Markov dependencies in the data, t

ParalESN: Enabling parallel information processing in Reservoir Computing

ResearchDGX agent

arXiv:2601.22296v2 Announce Type: replace-cross Abstract: Reservoir Computing (RC) has established itself as an efficient paradigm for temporal processing. However, its scalability remains severely co

PASTA: A Scalable Framework for Multi-Policy AI Compliance Evaluation

SafetyDGX agent

arXiv:2601.11702v3 Announce Type: replace-cross Abstract: AI compliance is becoming increasingly critical as AI systems grow more powerful and pervasive. Yet the rapid expansion of AI policies creates

PatchWorld: Gradient-Free Optimization of Executable World Models

Local AiDGX agent

arXiv:2605.30880v1 Announce Type: cross Abstract: Text-agent environments are typically modeled as partially observable Markov decision processes (POMDPs), assuming that the simulator's latent state a

Performance and Complexity Trade-off Optimization of Speech Models During Training

ApplicationsDGX agent

arXiv:2601.13704v3 Announce Type: replace-cross Abstract: In speech machine learning, neural network models are typically designed by choosing an architecture with fixed layer sizes and structure. The

Personalized to Persuade: The Effects of Contextualization and Warmth on Trust and Reliance in Conversational AI

ResearchDGX agent

arXiv:2605.31275v1 Announce Type: cross Abstract: Artificial Intelligence (AI) agents personalize their responses by tailoring explanations to users' backgrounds, interests, and prior interactions, re

PhyDrawGen: Physically Grounded Diagram Generation from Natural Language

Model ReleasesDGX agent

arXiv:2605.30512v1 Announce Type: new Abstract: Generating physics diagrams from text requires strict adherence to physical laws. While current generative models produce visually plausible outputs, th

Physically Viable World Models: A Case for Query-Conditioned Embodied AI

Model ReleasesDGX agent

arXiv:2605.30542v1 Announce Type: new Abstract: World models for embodied AI must be physically viable: constructed to answer intervention queries by representing the physical structure governing acti

PictSure: Pretraining Embeddings Matters for In-Context Learning Image Classifiers

AgentsDGX agent

arXiv:2506.14842v2 Announce Type: replace-cross Abstract: Building image classification models remains cumbersome in data-scarce domains, where collecting large labeled datasets is impractical. In-con

PInVerify: An Offline Embodied Benchmark for Active Instance Verification

Model ReleasesDGX agent

arXiv:2605.30639v1 Announce Type: cross Abstract: Embodied agents have made strong progress in navigating to target objects, but reaching the goal vicinity does not guarantee that the agent has found

PithTrain: A Compact and Agent-Native MoE Training System

HardwareDGX agent

arXiv:2605.31463v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for frontier language models. To meet this demand, production frameworks have built opti

Plain Transformers are Surprisingly Powerful Link Predictors

Model ReleasesDGX agent

arXiv:2602.01553v2 Announce Type: replace-cross Abstract: Link prediction is a core challenge in graph machine learning, demanding models that capture rich and complex topological dependencies. While

Planner-Centric Reinforcement Learning for Deep Research with Structure-Aware Reward

ResearchDGX agent

arXiv:2605.30824v1 Announce Type: new Abstract: Deep research tasks require LLMs to plan what to investigate, retrieve evidence, and synthesize long-form answers across multiple branches of inquiry. E

Position: Evaluation of ECG Representations Must Be Fixed

ResearchDGX agent

arXiv:2602.17531v2 Announce Type: replace-cross Abstract: This position paper argues that current benchmarking practice in 12-lead ECG representation learning must be fixed to ensure progress is relia

Positional versus Symbolic Attention Heads: Learning Dynamics, RoPE Geometry, and Length Generalization

ApplicationsDGX agent

arXiv:2605.31558v1 Announce Type: cross Abstract: Transformer-based language models are widespread in today's society. As such, understanding the mechanisms by which they solve structured tasks and pr

Post-Training LLMs as Better Decision-Making Agents: A Regret-Minimization Approach

ResearchDGX agent

arXiv:2511.04393v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as 'agents' for decision-making (DM) in interactive and dynamic environments. Yet, since they

Practical Cross-Band Channel Prediction for AI-RAN via Physics-Guided Deep Unfolding

TutorialsDGX agent

arXiv:2605.31279v1 Announce Type: cross Abstract: To make cross-band channel prediction practical for AI-native RAN, algorithms must generalize across diverse environments and support real-time infere

PReMISE: Policy Rubrics as Measurement Specifications for LLM Judges

SafetyDGX agent

arXiv:2605.30803v1 Announce Type: new Abstract: LLM judges are increasingly used to evaluate open-ended responses, but their scores depend strongly on the rubrics that condition them. A vague rubric a

Prior Availability in Industrial Visual Sim-to-Real: A Review of CAD-Guided and CAD-Unavailable Regimes

ApplicationsDGX agent

arXiv:2605.30581v1 Announce Type: cross Abstract: Industrial visual sim-to-real is often described as transferring from synthetic images to real images, but industrial deployment usually involves a br

PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection

ApplicationsDGX agent

arXiv:2502.12119v4 Announce Type: replace-cross Abstract: Visual instruction tuning adapts pre-trained Multimodal Large Language Models (MLLMs) to follow human instructions for real-world applications

Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration

Model ReleasesDGX agent

arXiv:2605.31196v1 Announce Type: cross Abstract: Safe human--robot collaboration requires more than visual description: a monitor must determine whether the robot body is safely separated, already co

Procedural Generation of First Person Shooter Maps using Map-Elites

TutorialsDGX agent

arXiv:2605.30570v1 Announce Type: new Abstract: We investigate the application of MAP-Elites (a well-known quality diversity algorithm) to design levels for First-Person Shooter (FPS) games. We consid

ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving

ResearchDGX agent

arXiv:2502.04671v3 Announce Type: replace Abstract: Neural approaches to theorem proving require robust infrastructure for interfacing with interactive theorem provers (ITPs), extracting structured pr

Pull Requests as a Training Signal for Repo-Level Code Editing

AgentsDGX agent

arXiv:2602.07457v2 Announce Type: replace-cross Abstract: Repository-level code editing requires models to understand complex dependencies and execute precise multi-file modifications across a large c

Rank-Factorized Implicit Neural Bias: Scaling Super-Resolution Transformer with FlashAttention

Local AiDGX agent

arXiv:2603.06738v2 Announce Type: replace-cross Abstract: Recent Super-Resolution~(SR) methods mainly adopt Transformers for their strong long-range modeling capability and exceptional representationa

Rationalize: Shared Semantic Reasoning for Human-AI Alignment

SafetyDGX agent

arXiv:2605.30632v1 Announce Type: cross Abstract: We introduce Rationalize, a role-pair framework for shared semantic reasoning between humans and AI models in data-driven sensemaking. Building on ide

RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video

ApplicationsDGX agent

arXiv:2605.31535v1 Announce Type: cross Abstract: Self-supervised novel view synthesis (NVS) remains challenging to scale, despite the abundance of video data, largely due to the brittleness of traini

Reading Between the Citations: A Typed Claim Network for Scientific Literature

SafetyDGX agent

arXiv:2605.30966v1 Announce Type: cross Abstract: Knowledge graphs over corpora of inter-referencing documents - scholarly papers, legal opinions, policy briefs - encode the topology of reference but

← Previous
1…185186187188189…358
Next →