AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Model Releases

POIROT: Interrogating Agents for Failure Detection in Multi-Agent Systems

DGX agent

arXiv:2606.02282v1 Announce Type: new Abstract: Orchestrating Large Language Models into Multi-Agent Systems (LLM-MAS) has unlocked remarkable reasoning capabilities, yet emergent failures and halluci

model-releasesarxiv-cs-ai
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

PolarMem: A Training-Free Polarized Latent Graph Memory for Verifiable Vision-Language Models

DGX agent

arXiv:2602.00415v2 Announce Type: replace Abstract: Memory is not merely a storage mechanism for intelligent systems, but a structure for organizing evidence and constraining belief. This is especiall

researcharxiv-cs-ai
2 Jun 2026
Safety

Policy and World Modeling Co-Training for Language Agents

DGX agent

arXiv:2606.02388v1 Announce Type: cross Abstract: Reinforcement learning (RL) improves large language model (LLM) agents by teaching them which actions lead to high rewards, but provides little superv

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

PolySpeech-100: A Large-Scale Benchmark for Speech Understanding Across 100+ Languages and Dialects

DGX agent

arXiv:2606.01016v1 Announce Type: cross Abstract: While End-to-End (E2E) Speech-Large Language Models (Speech-LLMs) are rapidly evolving, their evaluation methodologies remain limited to the era of si

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice via Social Determinants

DGX agent

arXiv:2508.08337v3 Announce Type: replace-cross Abstract: Algorithmic fairness research has largely framed unfairness as discrimination along sensitive attributes. However, this approach limits visibi

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Position Paper: Post-Solve Robustness in Decision Engines: Feasible Regions and Smoothness Under Perturbations

DGX agent

arXiv:2606.00002v1 Announce Type: new Abstract: Mixed-Integer Linear Programming (MILP) decision engines routinely output nominally optimal plans for high-stakes industrial systems. Yet deployment rar

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Post-Deterministic Distributed Systems: A New Foundation for Trustworthy Autonomous Infrastructure

DGX agent

arXiv:2606.01722v1 Announce Type: cross Abstract: For decades, distributed systems have typically assumed that correct participants execute protocol-specified behavior with stable, externally defined,

safetyarxiv-cs-ai
2 Jun 2026
Safety

PR2: Predictive Routing Replay for MoE-Based LLM Reinforcement Learning

DGX agent

arXiv:2606.00395v1 Announce Type: cross Abstract: Mixture of Experts (MoE) Large Language Models (LLMs) achieve strong performance at scale. However, reinforcement learning (RL) on MoE-based LLMs ofte

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Pre-Deployment Robustness Stress Testing for CT Segmentation Systems Using Clinically Motivated Multi-Corruption Augmentation

DGX agent

arXiv:2606.00491v1 Announce Type: cross Abstract: Deep learning-based CT segmentation systems often achieve high accuracy on clean benchmark images, but their performance may degrade under heterogeneo

model-releasesarxiv-cs-ai
2 Jun 2026
Hardware

Predicting Future Utility: Global Combinatorial Optimization for Task-Agnostic KV Cache Eviction

DGX agent

arXiv:2602.08585v2 Announce Type: replace-cross Abstract: Given the quadratic complexity of attention, KV cache eviction is vital to accelerate model inference. Current KV cache eviction methods typic

hardwarearxiv-cs-ai
2 Jun 2026
Applications

Predicting the risk of colorectal anastomotic leak based on preoperative mapping of the blood supply of the bowel

DGX agent

arXiv:2606.02156v1 Announce Type: cross Abstract: Anastomotic leak remains one of the most serious complications following colorectal cancer surgery, substantially affecting patient outcomes, recovery

applicationsarxiv-cs-ai
2 Jun 2026
Research

Principle-Evolvable Scientific Discovery via Uncertainty Minimization

DGX agent

arXiv:2602.06448v2 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based scientific agents have accelerated scientific discovery, yet they often suffer from significant inefficiencie

researcharxiv-cs-ai
2 Jun 2026
Safety

Principled Synthetic Data Enables the First Scaling Laws for LLMs in Recommendation

DGX agent

arXiv:2602.07298v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) represent a promising frontier for recommender systems, yet their development has been impeded by the absence of

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

PrivacyPeek: Auditing What LLM-Based Agents Acquire, Not Just What They Say

DGX agent

arXiv:2606.00152v1 Announce Type: cross Abstract: LLM-based agents are rapidly advancing, autonomously invoking external tools to complete multi-step tasks for users. However, agents often acquire mor

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning

DGX agent

arXiv:2602.02098v2 Announce Type: replace-cross Abstract: Multi-task reinforcement learning trains generalist policies that can execute multiple tasks. While recent years have seen significant progres

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Probe Before You Edit: Probing-Guided Molecular Optimization for LLM Agents in Structure-Based Drug Design

DGX agent

arXiv:2606.00555v1 Announce Type: new Abstract: Structure-based drug design increasingly employs LLM agents to iteratively refine ligands against a target pocket, yet a viable ligand must satisfy two

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ProbeScale: Probing Analysis to Optimize Neural Scaling Laws for Efficient Small Language Model Inference

DGX agent

arXiv:2606.01806v1 Announce Type: cross Abstract: Small Language Models (SLMs) offer a balance between capability and computational feasibility. Neural scaling laws inform their optimal training, sugg

model-releasesarxiv-cs-ai
2 Jun 2026
Research

ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts

DGX agent

arXiv:2606.01509v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models scale by activating only a small subset of experts per token. However, training such models remains challenging becaus

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Product-Aware Deep Autoencoders for Robust Process Monitoring in Multi-Product Cyber-Physical Systems

DGX agent

arXiv:2606.00052v1 Announce Type: new Abstract: As Industry 4.0 accelerates the integration of Cyber-Physical Systems (CPS) in manufacturing, robust anomaly detection has become critical for ensuring

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ProductWebGen: Benchmarking Multimodal Product Webpage Generation

DGX agent

arXiv:2606.01022v1 Announce Type: cross Abstract: Crafting a product display webpage from a source product image, along with layout and visual content instructions, holds significant practical value f

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

Project SPARROW and the Future of Conservation Technology

DGX agent

arXiv:2606.00108v1 Announce Type: cross Abstract: Global biodiversity is declining at unprecedented rates, yet the tools available to monitor and protect ecosystems remain limited by constraints in po

local-aiarxiv-cs-ai
2 Jun 2026
Applications

Property Prediction of Stacked Bilayer Materials: A Multimodal Learning Approach

DGX agent

arXiv:2606.01012v1 Announce Type: new Abstract: AI for materials science is a critical topic within AI for science, aiming to accelerate materials discovery and produce accurate property predictions.

applicationsarxiv-cs-ai
2 Jun 2026
Local Ai

PropLLM: Propagation-Aware Scene Reconstruction for Network Fault Diagnosis

DGX agent

arXiv:2606.00582v1 Announce Type: new Abstract: Network faults propagate layer by layer along topology and protocol dependencies, yet operations systems typically observe only symptomatic alerts at th

local-aiarxiv-cs-ai
2 Jun 2026
Safety

Prospect-Theory Behavior from Bellman Optimality in MDPs with Catastrophic States

DGX agent

arXiv:2606.00970v1 Announce Type: new Abstract: We study risk-neutral control in Markov decision processes with an absorbing catastrophic state. Even though rewards are linear and the agent has no uti

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Prototype Transformer: Towards Language Model Architectures Interpretable by Design

DGX agent

arXiv:2602.11852v2 Announce Type: replace Abstract: While state-of-the-art language models (LMs) surpass most humans in certain domains, their reasoning remains largely opaque, reducing trust and incr

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Prototypicality Bias Reveals Blindspots in Multimodal Evaluation Metrics

DGX agent

arXiv:2601.04946v3 Announce Type: replace-cross Abstract: Automatic metrics are widely used to evaluate text-to-image models, often replacing human judgment in benchmarking, model selection, and large

model-releasesarxiv-cs-ai
2 Jun 2026
Research

PSG-Nav: Probabilistic Scene Graph Navigation via Multiverse Decision Making

DGX agent

arXiv:2606.01313v1 Announce Type: cross Abstract: Open-vocabulary navigation requires embodied agents to manage significant perception uncertainty stemming from semantic ambiguity and model errors. Ho

researcharxiv-cs-ai
2 Jun 2026
Safety

Quantitative Movement Testing: Measuring Patient Movements from a Single Smartphone Video

DGX agent

arXiv:2606.02301v1 Announce Type: cross Abstract: Chronic pain diminishes quality of life by decreasing functional ability, yet objectively measuring this functional impact remains challenging in real

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Quantum Algorithm for Distributed Reduction of Entanglements (QADR): A Trainable and Simulation-Efficient QML Framework

DGX agent

arXiv:2606.01291v1 Announce Type: cross Abstract: Training Variational Quantum Circuits (VQCs) under Noisy Intermediate-Scale Quantum (NISQ) constraints introduces severe computational limitations: cl

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Quantum Tunneling-Aware Machine Learning: Physics-Derived Noise Models for Robust Deployment

DGX agent

arXiv:2606.00741v1 Announce Type: cross Abstract: Transistor scaling is approaching a quantum-mechanical limit, as thin gate oxides induce electron leakage through quantum tunneling. Unlike convention

researcharxiv-cs-ai
2 Jun 2026
Local Ai

Query Circuits: Explaining How Language Models Answer User Prompts

DGX agent

arXiv:2509.24808v2 Announce Type: replace Abstract: Explaining why a language model produces a particular output requires local, input-level explanations. Existing methods uncover global capability ci

local-aiarxiv-cs-ai
2 Jun 2026
Local Ai

RA-LWLM: Retrieval-Augmented In-Context Localization with Wireless Foundation Models

DGX agent

arXiv:2606.01899v1 Announce Type: cross Abstract: Wireless localization is a fundamental capability of sixth-generation (6G) networks. Conventional model-based methods require accurate modeling of the

local-aiarxiv-cs-ai
2 Jun 2026
Agents

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography

DGX agent

arXiv:2604.15231v2 Announce Type: replace Abstract: Vision-language models (VLM) have markedly advanced AI-driven interpretation and reporting of complex medical imaging, such as computed tomography (

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

RadioMaster: Multi-Agent System for Autonomous Radio Signal Generation

DGX agent

arXiv:2606.01862v1 Announce Type: cross Abstract: Translating user intents into physical radio signals represents the critical yet notoriously tedious final step in wireless prototyping, as it require

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

RAFT: Data Refinement and Adaptive Distillation for Domain Fine-Tuning with Alleviated Forgetting

DGX agent

arXiv:2606.00147v1 Announce Type: cross Abstract: Domain-specific supervised fine-tuning (SFT) often improves in-domain performance at the cost of degrading a model's general capabilities. We view thi

safetyarxiv-cs-ai
2 Jun 2026
Applications

Rank-Constrained Deep Matrix Completion for Group Recommendation

DGX agent

arXiv:2606.01948v1 Announce Type: cross Abstract: The growing popularity of group activities has increased the need for methods that provide recommendations to groups of users given their individual p

applicationsarxiv-cs-ai
2 Jun 2026
Research

Ranking vs. Assignment: The Metric Mismatch in Multi-View Object Association

DGX agent

arXiv:2606.02022v1 Announce Type: cross Abstract: Multi-view object association is an important computer vision problem that underlies many multi-camera perception tasks. While this task is naturally

researcharxiv-cs-ai
2 Jun 2026
Research

Rare Events, Real Signals: Functional Ensembles as Units of Computation in Deep Spiking Networks

DGX agent

arXiv:2606.00073v1 Announce Type: cross Abstract: We investigate how internal representations emerge across hierarchical processing systems by introducing a neuroscience-inspired framework for analyzi

researcharxiv-cs-ai
2 Jun 2026
Research

RASER: Recoverability-Aware Selective Escalation Router for Multi-Hop Question Answering

DGX agent

arXiv:2606.02488v1 Announce Type: new Abstract: Multi-hop question-answering systems often use expensive retrieval on every question. They may decompose the question, run several retrieval rounds, or

researcharxiv-cs-ai
2 Jun 2026
Agents

Rashomon Memory: Towards Argumentation-Driven Retrieval for Multi-Perspective Agent Memory

DGX agent

arXiv:2604.03588v3 Announce Type: replace Abstract: AI agents operating over extended time horizons accumulate experiences that serve multiple concurrent goals, and must often maintain conflicting int

agentsarxiv-cs-ai
2 Jun 2026
Safety

REAL: Resolving Knowledge Conflicts in Knowledge-Intensive Visual Question Answering via Reasoning-Pivot Alignment

DGX agent

arXiv:2602.14065v2 Announce Type: replace Abstract: Knowledge-intensive Visual Question Answering (KI-VQA) frequently suffers from severe knowledge conflicts caused by the inherent limitations of open

safetyarxiv-cs-ai
2 Jun 2026
Research

Real2SAM2Real: Generative 3D Caches as Complementary Context for Video Diffusion

DGX agent

arXiv:2606.00299v1 Announce Type: cross Abstract: While Video Diffusion Models (VDMs) excel at synthesizing high-fidelity videos, enabling precise camera and scene control remains challenging. Existin

researcharxiv-cs-ai
2 Jun 2026
Model Releases

ReasonBENCH: Benchmarking the (In)Stability of LLM Reasoning

DGX agent

arXiv:2512.07795v2 Announce Type: replace Abstract: Benchmark scores for LLM reasoning systems are reported as single numbers, yet the same model, strategy, and task can produce meaningfully different

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches

DGX agent

arXiv:2606.01145v1 Announce Type: new Abstract: While Reasoning Language Models (RLMs) are rapidly emerging as powerful tools for scientific research, their impact is primarily concentrated in 'hard s

researcharxiv-cs-ai
2 Jun 2026
Safety

REBot: From RAG to CatRAG with Semantic Enrichment and Graph Routing

DGX agent

arXiv:2510.01800v3 Announce Type: replace Abstract: Academic regulation advising is essential for helping students interpret and comply with institutional policies, yet building effective systems requ

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Recent Advances in Multi-modal 3D Intelligence: A Comprehensive Survey and Evaluation

DGX agent

arXiv:2310.15676v2 Announce Type: replace-cross Abstract: Multi-modal 3D Intelligence has gained considerable attention due to its wide applications in autonomous driving and world simulation, etc. Co

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Recognize Your Orchestrator: An Entropy Dynamics Perspective for LLM Multi-Agent Systems

DGX agent

arXiv:2606.01351v1 Announce Type: new Abstract: The transition from single-turn models to Multi-Agent Systems (MAS) promises enhanced problem-solving capabilities, yet the centralized orchestration to

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

RefDiffNet: Learning to Expose Subtle PCB Defects Before Detection

DGX agent

arXiv:2606.00852v1 Announce Type: cross Abstract: Printed circuit board (PCB) defect detection is challenging because many defects are small and difficult to distinguish from complex background patter

model-releasesarxiv-cs-ai
2 Jun 2026
← Previous
1…227228229230231…452
Next →