AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
30 Jun 2026

Propagation of~Interval Belief Structures and~Imprecise Copulas for~Neural Network Verification

SafetyDGX agent

arXiv:2606.30105v1 Announce Type: new Abstract: Quantitative verification of neural networks requires reasoning about probabilities under substantial uncertainty in both input distributions and their

ProSpec RL: Plan Ahead, then Execute

SafetyDGX agent

arXiv:2407.21359v2 Announce Type: replace-cross Abstract: Imagining potential outcomes of actions before execution helps agents make more informed decisions, a prospective thinking ability fundamental

Proteus: Automated Adversarial Robustness Testing for Audio Deepfake Detectors

Model ReleasesDGX agent

arXiv:2606.29544v1 Announce Type: cross Abstract: We present Proteus, a framework developed at Resemble AI for automated robustness testing of our audio deepfake detection system. Given a detector, Pr


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PS-PPO: Prefix-Sampling PPO for Critic-Free RLHF

SafetyDGX agent

arXiv:2606.29758v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) for Large Language Models increasingly relies on critic-free methods as a practical alternative to a

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization

Model ReleasesDGX agent

arXiv:2602.11351v2 Announce Type: replace Abstract: Proactive large language model (LLM) agents aim to actively plan, query, and interact over multiple turns, enabling efficient task completion beyond

Query-Aware Spreading Activation for Multi-Hop Retrieval over Knowledge Graphs

ResearchDGX agent

arXiv:2606.30133v1 Announce Type: cross Abstract: Retrieval-augmented generation built on knowledge graphs (Graph RAG) outperforms flat passage retrieval on multi-hop question answering by leveraging

RADIANT-PET: Reasoning-Augmented PET/CT Lesion Segmentation with Large Language Models and Reinforcement Learning

Local AiDGX agent

arXiv:2606.28392v1 Announce Type: cross Abstract: Accurate lesion segmentation in PET/CT is critical for oncology, yet remains challenging because physiologic tracer uptake and artifacts can mimic mal

Rank-Aware Hyperbolic Alignment for Vision-Language Dataset Distillation

SafetyDGX agent

arXiv:2606.29464v1 Announce Type: cross Abstract: Vision-language dataset distillation (VLDD) compresses a large image-text paired dataset into a small set of synthetic pairs that can efficiently trai

RankGraph-2: Lifecycle Co-Design for Billion-Node Graph Learning in Recommendation

Model ReleasesDGX agent

arXiv:2606.18379v2 Announce Type: replace-cross Abstract: Graph-based retrieval at billion-node scale requires jointly solving three tightly coupled problems -- graph construction, representation lear

ReactiveBFM: Reactive Closed-Loop Motion Planning Towards Universal Humanoid Whole-Body Control

SafetyDGX agent

arXiv:2606.30362v1 Announce Type: cross Abstract: While current Behavior Foundation Models (BFMs) provide robust control priors for humanoids, they only execute pre-defined reference motions. As a res

ReasonRec: A Reasoning-Augmented Multimodal Agent for Unified Recommendation

AgentsDGX agent

arXiv:2606.28357v1 Announce Type: cross Abstract: Recent advances in multimodal recommenders excel at feature fusion but remain opaque and inefficient decision-makers, lacking explicit reasoning and s

Reconsidering Overthinking: Penalizing Internal and External Redundancy in CoT Reasoning

ResearchDGX agent

arXiv:2508.02178v3 Announce Type: replace Abstract: Large reasoning models (LRMs) often exhibit overthinking, producing verbose Chain-of-Thought (CoT) traces that increase inference cost and obscure t

Recursive Self-Evolving Agents via Held-Out Selection

Model ReleasesDGX agent

arXiv:2606.28374v1 Announce Type: new Abstract: LLM agents are increasingly improved without weight updates by evolving a natural-language artifact, such as reflections, workflows, playbooks, cheatshe

Redefining Maritime Anomaly Detection via Equation-Grounded Synthetic Anomalies

Model ReleasesDGX agent

arXiv:2606.29721v1 Announce Type: cross Abstract: Maritime anomaly detection is essential for ensuring maritime safety, security, and efficient traffic management at sea, with Automatic Identification

Rehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question Answering

Model ReleasesDGX agent

arXiv:2606.30294v1 Announce Type: new Abstract: Live product demonstrations are a recurring, high-cost activity in software organizations: a human presenter must select features, dispatch the correspo

Reinforcement Learning for Software Vulnerability Analysis: A Systematic Review with Emphasis on C/C++ Source Code and Static Analysis

AgentsDGX agent

arXiv:2606.28403v1 Announce Type: cross Abstract: Vulnerability detection in C/C++ software remains a major security challenge due to code complexity, manual memory management, and the limitations of

Relevance Is Not Permission: Warranted Attention for Value Contributions

Local AiDGX agent

arXiv:2606.30139v1 Announce Type: new Abstract: Relevance is not permission. Attention lets a model read key-value items related to the current query, but it does not guarantee that the value contribu

ReMAP-PET: Beyond Visual Understanding -- Learning Region-Guided Metabolic Alignment Semantics from Brain PET

SafetyDGX agent

arXiv:2606.29577v1 Announce Type: cross Abstract: Positron Emission Tomography (PET) reveals brain metabolism and is clinically central to neurodegenerative disease assessment, yet existing 3D brain f

Reported Confidence in LLMs Tracks Commitment More Than Correctness

Model ReleasesDGX agent

arXiv:2606.29490v1 Announce Type: cross Abstract: Confidence is an estimate of the probability that a chosen answer is correct. Verbal confidence reports are widely used as uncertainty measures in lar

Representation Learning for Equivariant Inference with Guarantees

ApplicationsDGX agent

arXiv:2505.19809v3 Announce Type: replace-cross Abstract: In many real-world applications of regression, conditional probability estimation, and uncertainty quantification, exploiting symmetries roote

Research Entity Extraction and Topic Detection from UKRI Grant Proposals

Model ReleasesDGX agent

arXiv:2606.30304v1 Announce Type: cross Abstract: This paper presents preliminary findings from a UKRI-funded Metascience project comparing three LLM-based approaches, GPT-4o, Mistral, and a bespoke a

Residual-Guided Expert Specialization for Incomplete Multimodal Learning

TutorialsDGX agent

arXiv:2606.30355v1 Announce Type: cross Abstract: As real-world prediction systems often face missing modalities at inference, incomplete multimodal learning (IML) remains a practical challenge. While

Resonant Brane Splatting for Arbitrary-Scale Super-Resolution

ResearchDGX agent

arXiv:2606.29453v1 Announce Type: cross Abstract: Arbitrary-Scale Super-Resolution (ASR) reconstructs images at continuous magnification factors. Recent methods accelerate inference by replacing compu

RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources

AgentsDGX agent

arXiv:2606.29538v1 Announce Type: cross Abstract: Skills are a useful abstraction for software agents, turning human and agent experience into reusable procedural knowledge. Yet existing skill librari

Rethinking Generative Reconstruction Attacks against Graph Neural Network Models

Model ReleasesDGX agent

arXiv:2606.29748v1 Announce Type: new Abstract: The application of graph data in numerous disciplines raises the need for gathering and analyzing huge volumes of data, some of which is private and sen

Rethinking Role-Playing Evaluation: Anonymous Benchmarking and a Systematic Study of Personality Effects

AgentsDGX agent

arXiv:2603.03915v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown remarkable potential in developing role-playing agents (RPAs). However, current evaluation frameworks

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation

Model ReleasesDGX agent

arXiv:2606.28998v1 Announce Type: cross Abstract: Large Language Model (LLM) alignment trains an LLM using preference data to produce outputs that better meet established quality standards. While LLM

RGLD: Randomized Global-Local Density Estimation for Tabular Anomaly Detection

TutorialsDGX agent

arXiv:2606.28970v1 Announce Type: cross Abstract: Unsupervised tabular anomaly detection requires methods that are accurate, robust across heterogeneous datasets, and computationally efficient. Classi

RIPA: Sensory-Vector Prompt Injection Attacks on LLM-Controlled ROS 2 Robots

Model ReleasesDGX agent

arXiv:2606.28649v1 Announce Type: cross Abstract: We present RIPA, the first systematic multi-channel empirical study of prompt injection attacks delivered through the sensory pipeline of a ROS 2-base

RiverONE: Generating Knowledge-Intensive VLM by Simulated Quantum Machines

Model ReleasesDGX agent

arXiv:2606.29966v1 Announce Type: cross Abstract: Quantum computing provides a powerful paradigm for representing and transforming high-dimensional information through superposition, entanglement, and

RoAd-RL: A Unified Library and Benchmark for Robust Adversarial Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.29867v1 Announce Type: cross Abstract: Deep Reinforcement Learning (DRL) has achieved significant success in robotics and autonomous systems, yet remains vulnerable to adversarial perturbat

RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis

Model ReleasesDGX agent

arXiv:2606.28385v1 Announce Type: cross Abstract: Recent advances in robot world models enable synthetic video generation for embodied prediction and planning. However, evaluating these videos is chal

RoboPIN: Grounded Embodied Reasoning via Pinned Chain-of-Thought

SafetyDGX agent

arXiv:2606.15753v2 Announce Type: replace Abstract: Embodied reasoning requires models to perceive task-relevant objects and spaces in physical environments and maintain consistent visual grounding th

RSGPNet: Geometric Prompting for Remote Sensing Open-Vocabulary Semantic Segmentation

Model ReleasesDGX agent

arXiv:2606.28410v1 Announce Type: cross Abstract: Open-vocabulary semantic segmentation (OVSS) enables text-guided segmentation of unseen objects, breaking fixed-class limitations to achieve open-worl

S-GAI: Spectral Geometry-Aware Initialization for Sigmoidal MLPs -- From Dataset Geometry to Network Weights

ResearchDGX agent

arXiv:2606.28444v1 Announce Type: cross Abstract: Classical universal approximation theorems establish the expressive power of sigmoidal multilayer perceptrons, but they do not prescribe how initial w

SA-VLA: State-aware tokenizer for improving Vision-Language-Action Models' performance

SafetyDGX agent

arXiv:2606.30113v1 Announce Type: cross Abstract: Discrete action tokenization provides a compact interface for autoregressive VLA policies, but accurately recovering continuous robot actions from dis

SABER-Math: Automated Benchmark for Information Retrieval Evaluation in Mathematics

Model ReleasesDGX agent

arXiv:2606.29894v1 Announce Type: cross Abstract: As agentic AI systems tackle more complex mathematical tasks, they increasingly rely on information retrieval (IR) to search problem databases, theore

SafeGEO: Understanding Generative Engine Optimization Risks in Recommendation Agents

AgentsDGX agent

arXiv:2606.28356v1 Announce Type: cross Abstract: Generative Engine Optimization (GEO) lets content owners rewrite web content to increase their visibility in generative systems. In recommendation age

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing

Model ReleasesDGX agent

arXiv:2606.29887v1 Announce Type: new Abstract: In real-world applications, guardrails are often expected to identify unsafe user-model interactions according to application-specific safety policies,

Safety from Honesty in a Disinterested AI Predictor

SafetyDGX agent

arXiv:2606.29657v1 Announce Type: new Abstract: As AI systems become more capable, training procedures that optimize for downstream outcomes risk introducing implicit agency: goal-directed behavior th

SAGA: Scene-Aware, Goal-Evolving Agents for Long-Horizon CivRealm Strategy Planning

AgentsDGX agent

arXiv:2606.29932v1 Announce Type: new Abstract: Long-horizon strategic planning in complex strategy games demands concurrent reasoning across multiple decision domains under imperfect information and

SAKE: Software Architectural Knowledge Evaluation Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2606.29520v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used as assistants across the software development lifecycle, yet their ability to reason about software

Sample-Efficient Learning of Probabilistic Causes for Reachability in Markov Decision Processes with Probabilistic Guarantees

ResearchDGX agent

arXiv:2606.29681v1 Announce Type: new Abstract: Probabilistic model checking for Markov decision processes (MDPs) provides quantitative guarantees, but often offers limited insight into why undesired

SARLO-80: Worldwide Slant SAR Language Optic Dataset 80cm

SafetyDGX agent

arXiv:2606.20523v2 Announce Type: replace-cross Abstract: Multimodal foundation models have advanced rapidly thanks to large optical benchmarks, but comparable resources for synthetic aperture radar (

SAT-RTS: A systematic framework for tactical knowledge extraction and visualization-based analysis in real-time strategy games

ResearchDGX agent

arXiv:2606.30090v1 Announce Type: new Abstract: Efficient tactical knowledge extraction and analysis in real-time strategy (RTS) games micromanagement are constrained by the high-dimensional coupled s

Scalable Synthesis of distributed LLM workloads through Symbolic Tensor Graphs

ResearchDGX agent

arXiv:2511.10480v3 Announce Type: replace-cross Abstract: Optimizing the performance of large language models (LLMs) on large-scale AI training and inference systems requires a scalable and expressive

ScAle: Attention Head Scaling as a Minimal Adapter for Spatial Reasoning in Vision Language Models

Model ReleasesDGX agent

arXiv:2606.29579v1 Announce Type: cross Abstract: Spatial reasoning remains a persistent challenge for many vision language models (VLMs), and improving it typically requires fine-tuning with substant

Scaling Textual Gradients via Sampling-Based Momentum

Model ReleasesDGX agent

arXiv:2506.00400v4 Announce Type: replace-cross Abstract: LLM-based prompt optimization, which uses LLM-provided ``textual gradients'' (feedback) to refine prompts, has emerged as an effective method

SCARCE: Scalable Cascade Analysis for Rare-event Characterisation via Embeddings

Model ReleasesDGX agent

arXiv:2606.29623v1 Announce Type: new Abstract: Rare events govern the safety profile of modern AI systems, yet their probabilities are extremely difficult to estimate: direct Monte Carlo requires pro

Schema-First Retrieval: Embedding Catalogs for Natural Language Analytics

ApplicationsDGX agent

arXiv:2606.28387v1 Announce Type: cross Abstract: Enterprise text-to-SQL systems often fail before SQL is generated: the model receives the wrong schema context. Modern warehouses contain thousands of

SciVisAgentBench: A Benchmark for Evaluating Scientific Data Analysis and Visualization Agents

Model ReleasesDGX agent

arXiv:2603.29139v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled agentic systems to translate natural-language intent into executable scientific visuali

Search for Truth from Reasoning: A Dynamic Representation Editing Framework for Steering LLM Trajectories

Model ReleasesDGX agent

arXiv:2606.28589v1 Announce Type: new Abstract: Current approaches to enhance Large Language Model (LLM) reasoning, such as Chain-of-Thought and 'Wait' prompts, primarily encourage models to think mor

SEATauBench: Adapting Tool-Agent-User Evaluation Into Low-Resource Southeast Asian Languages

Model ReleasesDGX agent

arXiv:2606.28715v1 Announce Type: cross Abstract: While AI development and evaluation for Southeast Asia (SEA) has grown rapidly, agent capabilities in regional languages are still poorly understood d

Selective Memory Retention for Long-Horizon LLM Agents

Model ReleasesDGX agent

arXiv:2606.29178v1 Announce Type: new Abstract: When does retention matter for memory-augmented LLM agents? We study this with TraceRetain, a lightweight framework for bounded external memory in froze

Self-Evolving World Models for LLM Agent Planning

AgentsDGX agent

arXiv:2606.30639v1 Announce Type: new Abstract: World models offer a principled way to equip long-horizon LLM agents with foresight: predictions of action consequences before execution. However, unrel

Self-Organized Conformal Prediction: Reducing Regional Coverage Gaps with Unsupervised Group Discovery

SafetyDGX agent

arXiv:2606.29403v1 Announce Type: cross Abstract: Conformal prediction guarantees marginal coverage, but pooled calibration averages over heterogeneous regions and can mask regional undercoverage in s

Self-Supervised Theorem Discovery in a Formal Axiomatic System

Model ReleasesDGX agent

arXiv:2606.28747v1 Announce Type: new Abstract: Recent artificial intelligence (AI) systems have shown remarkable progress in mathematical reasoning. Many existing approaches, including large language

Semantic-Aware, Physics-Informed, Geometry-Grounded Weather Video Synthesis

AgentsDGX agent

arXiv:2606.29020v1 Announce Type: cross Abstract: Weather synthesis aims to add weather effects to input videos while preserving scene identity, structure, and motion. The key limitation of existing m

SemDynReg: Semantics-Guided Deformation Regularization for Dynamic 3D Gaussian Splatting

TutorialsDGX agent

arXiv:2606.28656v1 Announce Type: cross Abstract: Deformable 3D Gaussian Splatting (3DGS) has emerged as an efficient approach for rendering dynamic scenes in a wide range of 3D applications. However,

SemFlowRAG: Directed Semantic Flow from Abstraction to Evidence for Complex Reasoning

ResearchDGX agent

arXiv:2606.28447v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhanced by Knowledge Graphs has shown promise in complex multi-hop reasoning tasks. However, existing graph-base

← Previous
1…114115116117118…358
Next →