AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
2 Jun 2026

Predicting the risk of colorectal anastomotic leak based on preoperative mapping of the blood supply of the bowel

ApplicationsDGX agent

arXiv:2606.02156v1 Announce Type: cross Abstract: Anastomotic leak remains one of the most serious complications following colorectal cancer surgery, substantially affecting patient outcomes, recovery

Principle-Evolvable Scientific Discovery via Uncertainty Minimization

ResearchDGX agent

arXiv:2602.06448v2 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based scientific agents have accelerated scientific discovery, yet they often suffer from significant inefficiencie

Principled Synthetic Data Enables the First Scaling Laws for LLMs in Recommendation

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.07298v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) represent a promising frontier for recommender systems, yet their development has been impeded by the absence of

PrivacyPeek: Auditing What LLM-Based Agents Acquire, Not Just What They Say

Model ReleasesDGX agent

arXiv:2606.00152v1 Announce Type: cross Abstract: LLM-based agents are rapidly advancing, autonomously invoking external tools to complete multi-step tasks for users. However, agents often acquire mor

Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning

SafetyDGX agent

arXiv:2602.02098v2 Announce Type: replace-cross Abstract: Multi-task reinforcement learning trains generalist policies that can execute multiple tasks. While recent years have seen significant progres

Probe Before You Edit: Probing-Guided Molecular Optimization for LLM Agents in Structure-Based Drug Design

Model ReleasesDGX agent

arXiv:2606.00555v1 Announce Type: new Abstract: Structure-based drug design increasingly employs LLM agents to iteratively refine ligands against a target pocket, yet a viable ligand must satisfy two

ProbeScale: Probing Analysis to Optimize Neural Scaling Laws for Efficient Small Language Model Inference

Model ReleasesDGX agent

arXiv:2606.01806v1 Announce Type: cross Abstract: Small Language Models (SLMs) offer a balance between capability and computational feasibility. Neural scaling laws inform their optimal training, sugg

ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts

ResearchDGX agent

arXiv:2606.01509v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models scale by activating only a small subset of experts per token. However, training such models remains challenging becaus

Product-Aware Deep Autoencoders for Robust Process Monitoring in Multi-Product Cyber-Physical Systems

Model ReleasesDGX agent

arXiv:2606.00052v1 Announce Type: new Abstract: As Industry 4.0 accelerates the integration of Cyber-Physical Systems (CPS) in manufacturing, robust anomaly detection has become critical for ensuring

ProductWebGen: Benchmarking Multimodal Product Webpage Generation

Model ReleasesDGX agent

arXiv:2606.01022v1 Announce Type: cross Abstract: Crafting a product display webpage from a source product image, along with layout and visual content instructions, holds significant practical value f

Project SPARROW and the Future of Conservation Technology

Local AiDGX agent

arXiv:2606.00108v1 Announce Type: cross Abstract: Global biodiversity is declining at unprecedented rates, yet the tools available to monitor and protect ecosystems remain limited by constraints in po

Property Prediction of Stacked Bilayer Materials: A Multimodal Learning Approach

ApplicationsDGX agent

arXiv:2606.01012v1 Announce Type: new Abstract: AI for materials science is a critical topic within AI for science, aiming to accelerate materials discovery and produce accurate property predictions.

PropLLM: Propagation-Aware Scene Reconstruction for Network Fault Diagnosis

Local AiDGX agent

arXiv:2606.00582v1 Announce Type: new Abstract: Network faults propagate layer by layer along topology and protocol dependencies, yet operations systems typically observe only symptomatic alerts at th

Prospect-Theory Behavior from Bellman Optimality in MDPs with Catastrophic States

SafetyDGX agent

arXiv:2606.00970v1 Announce Type: new Abstract: We study risk-neutral control in Markov decision processes with an absorbing catastrophic state. Even though rewards are linear and the agent has no uti

Prototype Transformer: Towards Language Model Architectures Interpretable by Design

Model ReleasesDGX agent

arXiv:2602.11852v2 Announce Type: replace Abstract: While state-of-the-art language models (LMs) surpass most humans in certain domains, their reasoning remains largely opaque, reducing trust and incr

Prototypicality Bias Reveals Blindspots in Multimodal Evaluation Metrics

Model ReleasesDGX agent

arXiv:2601.04946v3 Announce Type: replace-cross Abstract: Automatic metrics are widely used to evaluate text-to-image models, often replacing human judgment in benchmarking, model selection, and large

PSG-Nav: Probabilistic Scene Graph Navigation via Multiverse Decision Making

ResearchDGX agent

arXiv:2606.01313v1 Announce Type: cross Abstract: Open-vocabulary navigation requires embodied agents to manage significant perception uncertainty stemming from semantic ambiguity and model errors. Ho

Quantitative Movement Testing: Measuring Patient Movements from a Single Smartphone Video

SafetyDGX agent

arXiv:2606.02301v1 Announce Type: cross Abstract: Chronic pain diminishes quality of life by decreasing functional ability, yet objectively measuring this functional impact remains challenging in real

Quantum Algorithm for Distributed Reduction of Entanglements (QADR): A Trainable and Simulation-Efficient QML Framework

Model ReleasesDGX agent

arXiv:2606.01291v1 Announce Type: cross Abstract: Training Variational Quantum Circuits (VQCs) under Noisy Intermediate-Scale Quantum (NISQ) constraints introduces severe computational limitations: cl

Quantum Tunneling-Aware Machine Learning: Physics-Derived Noise Models for Robust Deployment

ResearchDGX agent

arXiv:2606.00741v1 Announce Type: cross Abstract: Transistor scaling is approaching a quantum-mechanical limit, as thin gate oxides induce electron leakage through quantum tunneling. Unlike convention

Query Circuits: Explaining How Language Models Answer User Prompts

Local AiDGX agent

arXiv:2509.24808v2 Announce Type: replace Abstract: Explaining why a language model produces a particular output requires local, input-level explanations. Existing methods uncover global capability ci

RA-LWLM: Retrieval-Augmented In-Context Localization with Wireless Foundation Models

Local AiDGX agent

arXiv:2606.01899v1 Announce Type: cross Abstract: Wireless localization is a fundamental capability of sixth-generation (6G) networks. Conventional model-based methods require accurate modeling of the

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography

AgentsDGX agent

arXiv:2604.15231v2 Announce Type: replace Abstract: Vision-language models (VLM) have markedly advanced AI-driven interpretation and reporting of complex medical imaging, such as computed tomography (

RadioMaster: Multi-Agent System for Autonomous Radio Signal Generation

Model ReleasesDGX agent

arXiv:2606.01862v1 Announce Type: cross Abstract: Translating user intents into physical radio signals represents the critical yet notoriously tedious final step in wireless prototyping, as it require

RAFT: Data Refinement and Adaptive Distillation for Domain Fine-Tuning with Alleviated Forgetting

SafetyDGX agent

arXiv:2606.00147v1 Announce Type: cross Abstract: Domain-specific supervised fine-tuning (SFT) often improves in-domain performance at the cost of degrading a model's general capabilities. We view thi

Rank-Constrained Deep Matrix Completion for Group Recommendation

ApplicationsDGX agent

arXiv:2606.01948v1 Announce Type: cross Abstract: The growing popularity of group activities has increased the need for methods that provide recommendations to groups of users given their individual p

Ranking vs. Assignment: The Metric Mismatch in Multi-View Object Association

ResearchDGX agent

arXiv:2606.02022v1 Announce Type: cross Abstract: Multi-view object association is an important computer vision problem that underlies many multi-camera perception tasks. While this task is naturally

Rare Events, Real Signals: Functional Ensembles as Units of Computation in Deep Spiking Networks

ResearchDGX agent

arXiv:2606.00073v1 Announce Type: cross Abstract: We investigate how internal representations emerge across hierarchical processing systems by introducing a neuroscience-inspired framework for analyzi

RASER: Recoverability-Aware Selective Escalation Router for Multi-Hop Question Answering

ResearchDGX agent

arXiv:2606.02488v1 Announce Type: new Abstract: Multi-hop question-answering systems often use expensive retrieval on every question. They may decompose the question, run several retrieval rounds, or

Rashomon Memory: Towards Argumentation-Driven Retrieval for Multi-Perspective Agent Memory

AgentsDGX agent

arXiv:2604.03588v3 Announce Type: replace Abstract: AI agents operating over extended time horizons accumulate experiences that serve multiple concurrent goals, and must often maintain conflicting int

REAL: Resolving Knowledge Conflicts in Knowledge-Intensive Visual Question Answering via Reasoning-Pivot Alignment

SafetyDGX agent

arXiv:2602.14065v2 Announce Type: replace Abstract: Knowledge-intensive Visual Question Answering (KI-VQA) frequently suffers from severe knowledge conflicts caused by the inherent limitations of open

Real2SAM2Real: Generative 3D Caches as Complementary Context for Video Diffusion

ResearchDGX agent

arXiv:2606.00299v1 Announce Type: cross Abstract: While Video Diffusion Models (VDMs) excel at synthesizing high-fidelity videos, enabling precise camera and scene control remains challenging. Existin

ReasonBENCH: Benchmarking the (In)Stability of LLM Reasoning

Model ReleasesDGX agent

arXiv:2512.07795v2 Announce Type: replace Abstract: Benchmark scores for LLM reasoning systems are reported as single numbers, yet the same model, strategy, and task can produce meaningfully different

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches

ResearchDGX agent

arXiv:2606.01145v1 Announce Type: new Abstract: While Reasoning Language Models (RLMs) are rapidly emerging as powerful tools for scientific research, their impact is primarily concentrated in 'hard s

REBot: From RAG to CatRAG with Semantic Enrichment and Graph Routing

SafetyDGX agent

arXiv:2510.01800v3 Announce Type: replace Abstract: Academic regulation advising is essential for helping students interpret and comply with institutional policies, yet building effective systems requ

Recent Advances in Multi-modal 3D Intelligence: A Comprehensive Survey and Evaluation

Model ReleasesDGX agent

arXiv:2310.15676v2 Announce Type: replace-cross Abstract: Multi-modal 3D Intelligence has gained considerable attention due to its wide applications in autonomous driving and world simulation, etc. Co

Recognize Your Orchestrator: An Entropy Dynamics Perspective for LLM Multi-Agent Systems

AgentsDGX agent

arXiv:2606.01351v1 Announce Type: new Abstract: The transition from single-turn models to Multi-Agent Systems (MAS) promises enhanced problem-solving capabilities, yet the centralized orchestration to

RefDiffNet: Learning to Expose Subtle PCB Defects Before Detection

Model ReleasesDGX agent

arXiv:2606.00852v1 Announce Type: cross Abstract: Printed circuit board (PCB) defect detection is challenging because many defects are small and difficult to distinguish from complex background patter

Regime-Adaptive Continual Learning for Portfolio Management

SafetyDGX agent

arXiv:2606.00143v1 Announce Type: cross Abstract: Financial markets are inherently non-stationary, exhibiting frequent regime shifts and structural changes that render traditional Portfolio Management

Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief

SafetyDGX agent

arXiv:2606.00680v1 Announce Type: new Abstract: Offline reinforcement learning (RL) aims to optimize policies from pre-collected datasets. A bottleneck of this paradigm is managing epistemic uncertain

Reinforcement Learning Position Control of a Quadrotor Using Soft Actor-Critic (SAC)

SafetyDGX agent

arXiv:2512.18333v2 Announce Type: replace-cross Abstract: This paper proposes a new Reinforcement Learning (RL) based control architecture for quadrotors. With the literature focusing on controlling t

Reinforcement Learning with Pairwise Preferences in Long-Term Decision Problems

SafetyDGX agent

arXiv:2606.00367v1 Announce Type: cross Abstract: Reinforcement learning problems typically define the goal as maximizing the expected value of a scalar reward function. But, pairwise preferences are

Relational Intervention During Functional Collapse in Large Language Models: A Lexical-Statistical Ablation and a Structure x Register Factorial

Local AiDGX agent

arXiv:2606.00935v1 Announce Type: new Abstract: We test whether a relational-style intervention delivered during functional collapse in a small language model produces post-collapse behavior distingui

Repair Before Veto: Repair-Augmented Constraint Learning for Contextual Decisions

TutorialsDGX agent

arXiv:2606.02326v1 Announce Type: new Abstract: Hard constraints are usually treated as terminal vetoes: once a candidate violates a requirement, the learned rule rejects it and any repair is handled

Repurposing Adversarial Perturbations for Continual Learning: From Defense to Active Alignment

SafetyDGX agent

arXiv:2606.02322v1 Announce Type: cross Abstract: In dynamic environments, large language models need to keep adapting to new tasks, but continual learning often suffers from forgetting, limited trans

ReSkill: Reconciling Skill Creation with Policy Optimization in Agentic RL

SafetyDGX agent

arXiv:2606.01619v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) enables LLM agents to improve continuously from environment rewards, yet the resulting policies do not systematicall

ResNet-34 with Lightweight Decoder for Accurate and Efficient Segmentation of Fetal Brain MRI

Model ReleasesDGX agent

arXiv:2606.01293v1 Announce Type: cross Abstract: Accurate segmentation of fetal brain tissues in Magnetic Resonance Imaging (MRI) is critical for early diagnosis of congenital abnormalities and impro

Rethinking Evaluation Paradigms in IBP-based Certified Training

ResearchDGX agent

arXiv:2606.02134v1 Announce Type: cross Abstract: Deep neural networks achieve strong performance on many supervised learning tasks but remain vulnerable to adversarial perturbations. Neural network v

Rethinking RL Evaluation: Can Benchmarks Truly Reveal Failures of RL Methods?

Model ReleasesDGX agent

arXiv:2510.10541v2 Announce Type: replace-cross Abstract: Current benchmarks are inadequate for evaluating progress in reinforcement learning (RL) for large language models (LLMs).Despite recent bench

Rethinking Scientific Modeling: Toward Physically Consistent and Simulation-Executable Programmatic Generation

Model ReleasesDGX agent

arXiv:2602.07083v2 Announce Type: replace-cross Abstract: Structural modeling is a fundamental component of computational engineering science, in which even minor physical inconsistencies or specifica

Rethinking the Role of Temperature in Large Language Model Distillation

ResearchDGX agent

arXiv:2606.00306v1 Announce Type: cross Abstract: Reverse Kullback-Leibler (RKL) divergence is widely favored over forward KL (FKL) in large language models (LLM) distillation, yet this preference is

Retrieval-aligned Tabular Foundation Models Enable Robust Clinical Risk Prediction in Electronic Health Records Under Real-world Constraints

Model ReleasesDGX agent

arXiv:2604.01841v2 Announce Type: replace Abstract: Clinical prediction from structured electronic health records (EHRs) is challenging due to high dimensionality, heterogeneity, class imbalance, and

Revisiting Parameter-Based Knowledge Editing in Large Language Models: Theoretical Limits and Empirical Evidence

Model ReleasesDGX agent

arXiv:2606.00570v1 Announce Type: cross Abstract: Parameter-based knowledge editing updates the internal knowledge of large language models (LLMs) via localized weight modifications and has attracted

Revisiting Ripple Effects in Knowledge Editing through Pressure-Aware Joint Neighborhood Optimization

Model ReleasesDGX agent

arXiv:2606.01610v1 Announce Type: new Abstract: Single-edit updates in large language models can trigger ripple effects across local knowledge neighborhoods: desirable propagation to related facts and

Richer Representations for Neural Algorithmic Reasoning via Auxiliary Reconstruction

TutorialsDGX agent

arXiv:2606.00559v1 Announce Type: cross Abstract: Neural algorithmic reasoning has emerged as a popular research direction. It aims to train neural networks to mimic the step-by-step behavior of class

RL-ACRGNet: Reinforcement Learning-Based Chest Radiology Report Generation Network

SafetyDGX agent

arXiv:2606.02035v1 Announce Type: new Abstract: Medical imaging interpretation is a foundational pillar of modern clinical diagnostics, yet the manual generation of radiology reports remains a time-co

RLVR without Ineffective Samples: Group Prioritized Off-Policy Optimization for LLM Reasoning

SafetyDGX agent

arXiv:2606.01281v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a powerful paradigm for enhancing the reasoning capabilities of large language mo

RoboBenchMart: Benchmarking Robots in Retail Environment

Model ReleasesDGX agent

arXiv:2511.10276v2 Announce Type: replace-cross Abstract: Most existing robotic manipulation benchmarks focus on tabletop or household scenarios. While these setups have driven impressive progress, it

Robust Shielding for Safe Reinforcement Learning

SafetyDGX agent

arXiv:2606.00270v1 Announce Type: new Abstract: Shielding is an effective approach to formally guarantee the safety of reinforcement learning agents in Markov decision processes (MDPs). However, exist

ROGUE: Misaligned Agent Behavior Arising from Ordinary Computer Use

Model ReleasesDGX agent

arXiv:2606.00341v1 Announce Type: cross Abstract: As AI agents are increasingly deployed in real personal and corporate settings (email accounts, development workflows, company databases, etc.), safet

← Previous
1…178179180181182…358
Next →