AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Q-Delta: Beyond Key-Value Associative State Evolution

DGX agent

arXiv:2606.08804v1 Announce Type: new Abstract: Linear attention reformulates sequence modeling as recurrent state evolution, enabling efficient linear-time inference. Under the key-value associative

researcharxiv-cs-ai
9 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Quantitative Promise Theory: Intentionality and Inference in Autonomous Agents

DGX agent

arXiv:2606.08552v1 Announce Type: new Abstract: I discuss some quantitative representations of Promise Theory for processes involving autonomous agents. Agent models are common in software systems, ma

safetyarxiv-cs-ai
9 Jun 2026
Research

Quantum-Enhanced Similarity Measures for Polarimetric Materials Classification

DGX agent

arXiv:2606.07766v1 Announce Type: cross Abstract: We present a quantum--classical hybrid pipeline for polarimetric material classification that casts this as a point-matching problem. Voxel cubes, con

researcharxiv-cs-ai
9 Jun 2026
Research

Query Lens: Interpreting Sparse Key-Value Features with Indirect Effects

DGX agent

arXiv:2606.07617v1 Announce Type: cross Abstract: While sparse autoencoders provide features more interpretable than individual neurons, reliably characterizing them remains challenging. We propose Qu

researcharxiv-cs-ai
9 Jun 2026
Research

RadOT-Eval: Auditable Structured-Evidence Transport for Radiology Report Evaluation

DGX agent

arXiv:2606.08769v1 Announce Type: cross Abstract: Automatic evaluation is critical for high-stakes text generation, where errors often involve omitted findings, hallucinated content, polarity reversal

researcharxiv-cs-ai
9 Jun 2026
Agents

RAILS: Verification-Native Clearing For Agentic Commerce

DGX agent

arXiv:2606.08790v1 Announce Type: new Abstract: Autonomous agents negotiate, purchase, deploy code, and move funds, but no neutral mechanism determines whether they met their delegated obligation, who

agentsarxiv-cs-ai
9 Jun 2026
Research

RAPID: Layer-Wise Redundancy-Aware Pruning and Importance-Driven Token Merging for Efficient ViT

DGX agent

arXiv:2606.08156v1 Announce Type: cross Abstract: Vision Transformers (ViTs) achieve strong performance but suffer from high computational costs due to quadratic self-attention complexity. Although to

researcharxiv-cs-ai
9 Jun 2026
Research

Reachability and asymptotics of Gaussian Transformer dynamics

DGX agent

arXiv:2606.07600v1 Announce Type: cross Abstract: We formulate data propagation through the Transformer, the machine learning architecture powering large language models, as a nonlinear control system

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Real-time body pose non-verbal communication with a consistency-based reliability measure

DGX agent

arXiv:2606.09390v1 Announce Type: cross Abstract: Body movement communicates intent at distances and in conditions where neither the face, nor speech can be captured. We study the recognition of commu

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Reasoning Arena: Trace Tournaments When Verifiable Rewards Fall Short

DGX agent

arXiv:2606.09380v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a leading paradigm for improving the reasoning ability of large language models throu

researcharxiv-cs-ai
9 Jun 2026
Research

Reconstructing and forecasting disease trajectories of patients with Alzheimer's disease using routine data in resource-constrained settings

DGX agent

arXiv:2606.07798v1 Announce Type: new Abstract: Alzheimer's disease is a progressive neurodegenerative disorder, and its progression varies substantially across patients. Existing work aims to forecas

researcharxiv-cs-ai
9 Jun 2026
Research

Reconstructing Synthetic SDO/AIA 193 A EUV Images from He I 10830 A Observations with Diffusion Model Translator

DGX agent

arXiv:2606.08652v1 Announce Type: cross Abstract: Routine full-disk EUV imaging has been available only since the modern era, such as SOHO and SDO. To extend EUV coronal context into earlier periods,

researcharxiv-cs-ai
9 Jun 2026
Safety

ReCoVLA: VLM-Guided Reward Compilation for Failure Recovery in Vision-Language-Action Policies

DGX agent

arXiv:2606.09630v1 Announce Type: cross Abstract: Vision-language-action (VLA) policies provide strong priors for language-conditioned manipulation, but remain brittle in off-nominal states requiring

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks

DGX agent

arXiv:2606.07968v1 Announce Type: cross Abstract: Reasoning-capable large language models can be induced to spend their generation budget on injected decoy tasks rather than answering the user's quest

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

REFLECT: Intervention-Supported Error Attribution for Silent Failures in LLM Agent Traces

DGX agent

arXiv:2606.09071v1 Announce Type: new Abstract: Large language model (LLM) agents now solve complex tasks through long plan-and-execution traces, yet the ability to locate errors in a completed traces

agentsarxiv-cs-ai
9 Jun 2026
Agents

Reflection in the Dark: Exposing and Escaping the Black Box in Reflective Prompt Optimization

DGX agent

arXiv:2603.18388v2 Announce Type: replace Abstract: Automatic prompt optimization (APO) has emerged as a powerful paradigm for improving LLM performance without manual prompt engineering. Reflective A

agentsarxiv-cs-ai
9 Jun 2026
Safety

Reinforcement Learning for Flow-Matching Policies with Density Transport

DGX agent

arXiv:2606.08602v1 Announce Type: cross Abstract: We present an online reinforcement learning (RL) algorithm for fine-tuning flow-matching policies in continuous-control problems. Our key insight is t

safetyarxiv-cs-ai
9 Jun 2026
Safety

Reliable to Expressive: A Curriculum for Rubric-Following Safety Judges

DGX agent

arXiv:2606.09165v1 Announce Type: new Abstract: Safety judges are increasingly deployed to evaluate model outputs against evolving criteria, yet recent meta-evaluation work shows they remain brittle u

safetyarxiv-cs-ai
9 Jun 2026
Safety

Repair Before Veto, When Repair Is Hidden: Quantum-Accessible Features for Repair-Augmented Constraint Learning

DGX agent

arXiv:2606.08020v1 Announce Type: cross Abstract: Hard-constraint decision systems usually veto infeasible candidates. This is too rigid when the system can act: if a known affordable repair would mak

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them

DGX agent

arXiv:2606.07597v1 Announce Type: cross Abstract: Pre-training data mixtures are commonly tuned by running small-scale experiments and extrapolating to the target training budget. When high-quality da

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Report on CHIIR 2026 Workshop on Generative AI and Academic Search (GAI&AS)

DGX agent

arXiv:2606.08936v1 Announce Type: cross Abstract: This report summarizes the CHIIR 2026 Workshop on Generative AI and Academic Search (GAI&AS), which examined how GenAI is reshaping academic search sy

researcharxiv-cs-ai
9 Jun 2026
Model Releases

ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research

DGX agent

arXiv:2606.07591v1 Announce Type: cross Abstract: AI coding agents are increasingly used for scientific work, but their end-to-end autonomous research capability remains difficult to verify. We presen

model-releasesarxiv-cs-ai
9 Jun 2026
Hardware

Resource-aware Computation-Communication Overlap for multi-GPU ML Workloads

DGX agent

arXiv:2606.09200v1 Announce Type: cross Abstract: The rapid growth of large-scale machine learning (ML) has made distributed training across multiple GPUs a fundamental component of modern ML systems.

hardwarearxiv-cs-ai
9 Jun 2026
Applications

Retrieval Augmented Generation Framework for the Nepali Legal Domain Question Answering

DGX agent

arXiv:2606.07523v1 Announce Type: cross Abstract: Legal domains in high-resource languages like English have widely adopted artificial intelligence for legal question answering. However, data scarcity

applicationsarxiv-cs-ai
9 Jun 2026
Research

RetroReasoner: A Reasoning LLM for Strategic Retrosynthesis Prediction

DGX agent

arXiv:2603.12666v2 Announce Type: replace-cross Abstract: Retrosynthesis prediction aims to identify reactants that can synthesize a given product molecule. Although molecular large language models (L

researcharxiv-cs-ai
9 Jun 2026
Safety

Revisiting the shutdown problem

DGX agent

arXiv:2606.08296v1 Announce Type: new Abstract: A key premise in leading arguments for existential risk from artificial intelligence is that malfunctioning artificial agents could not be easily shut d

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Revisiting Training Scale: An Empirical Study of Token Count, Power Consumption, and Parameter Efficiency

DGX agent

arXiv:2601.06649v2 Announce Type: replace-cross Abstract: Research in machine learning has questioned whether increases in training token counts reliably produce proportional performance gains in larg

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective

DGX agent

arXiv:2602.02572v2 Announce Type: replace-cross Abstract: Existing alignment methods directly use the reward model learned from user preference data to optimize an LLM policy, subject to KL regulariza

safetyarxiv-cs-ai
9 Jun 2026
Research

Rewrite to Translate, Translate to Reward: Reinforcement Learning for Source Rewriting in Machine Translation

DGX agent

arXiv:2606.08011v1 Announce Type: cross Abstract: Although directly prompting off-the-shelf Large Language Models (LLMs) to generate meaning-preserving source rewrites can effectively enhance Machine

researcharxiv-cs-ai
9 Jun 2026
Model Releases

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations

DGX agent

arXiv:2606.08376v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across socially consequential domains, reports of AI-related harms and failures have

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Robust Renal Mass Segmentation on CT: A Validation Study of an AI-Based Framework

DGX agent

arXiv:2505.07573v2 Announce Type: replace-cross Abstract: Renal mass segmentation has important potential to enhance the clinical workflow, especially in settings requiring quantitative assessments. K

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Robust-U1: Can MLLMs Self-Recover Corrupted Visual Content for Robust Understanding?

DGX agent

arXiv:2606.08063v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in visual understanding, yet their performance degrades significantly un

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Rosetta Memory: Adaptive Memory for Cross-LLM Agents

DGX agent

arXiv:2606.07711v1 Announce Type: cross Abstract: Memory is the key component for transforming a stateless LLM into a persistent, evolving agent through experience accumulation, long-horizon planning,

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

RTL-BenchLS: A Large-Scale Benchmark for RTL Reasoning and Generation with Large Language Models

DGX agent

arXiv:2606.08976v1 Announce Type: new Abstract: LLM-based RTL generation and reasoning is a promising direction for hardware design automation. High-quality benchmarks are critical infrastructure for

model-releasesarxiv-cs-ai
9 Jun 2026
Applications

Rule-based autocorrection of Piping and Instrumentation Diagrams (P&IDs) on graphs

DGX agent

arXiv:2502.18493v2 Announce Type: replace-cross Abstract: A piping and instrumentation diagram (P&ID) is a central reference document in chemical process engineering. Currently, chemical engineers man

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

RunAgent SuperBrowser: A Theory of Autonomous Web Navigation Grounded in Human Browsing Behaviour

DGX agent

arXiv:2606.09399v1 Announce Type: new Abstract: We present SUPERBROWSER, an autonomous web-navigation agent designed against a single guiding hypothesis: a web agent should browse the way a person bro

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Safe-RULE: Safe Reinforcement UnLEarning

DGX agent

arXiv:2606.09559v1 Announce Type: cross Abstract: Offline safe reinforcement learning (Safe RL) enables policy learning without online interactions, making it suitable for safety-critical systems such

model-releasesarxiv-cs-ai
9 Jun 2026
Research

SafeECGMatch: Calibration-Aware Joint Frequency and Time Space Semi-Supervised Learning for Open-Set ECG Classification

DGX agent

arXiv:2606.08037v1 Announce Type: cross Abstract: Electrocardiogram (ECG) classification models often suffer from severe label scarcity, making semi-supervised learning (SSL) an attractive strategy fo

researcharxiv-cs-ai
9 Jun 2026
Model Releases

SafeRun: Enabling Determinism in LLM Planning for Running

DGX agent

arXiv:2606.09027v1 Announce Type: cross Abstract: Large Language Models enable flexible natural-language planning but remain unreliable in determinism-critical domains due to their probabilistic natur

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Safety is Contextual, LLM-Judges Are Not: Navigating the Rigid Priors of Evaluators

DGX agent

arXiv:2606.07874v1 Announce Type: new Abstract: LLMs-as-judges are the only way to evaluate safety at scale. Despite their importance, LLM-judges themselves are rarely evaluated beyond human agreement

safetyarxiv-cs-ai
9 Jun 2026
Agents

SAGE: An LLM-driven Self Reflective Agentic Framework for Fraud Detection

DGX agent

arXiv:2606.08146v1 Announce Type: new Abstract: Fraud detection in payment, e-commerce, and telecommunications systems requires accuracy at the individual level, robustness under severe class imbalanc

agentsarxiv-cs-ai
9 Jun 2026
Research

SAGE: Shape-Adapting Gated Experts for Adaptive Histopathology Image Segmentation

DGX agent

arXiv:2511.18493v4 Announce Type: replace-cross Abstract: The significant variability in cell size and shape continues to pose a major obstacle in computer-assisted cancer detection on gigapixel Whole

researcharxiv-cs-ai
9 Jun 2026
Local Ai

SAILS: Surrogate-based Analysis of Interactions via Local Effect Smooths

DGX agent

arXiv:2606.09404v1 Announce Type: cross Abstract: Feature interactions drive much of the predictive power of machine learning models, yet existing explanation methods only detect and quantify interact

local-aiarxiv-cs-ai
9 Jun 2026
Tutorials

Sample-Efficient LLM-Based Detection of Malicious Web Server Logs with Forensically Explainable Reasoning

DGX agent

arXiv:2606.08649v1 Announce Type: cross Abstract: Forensic analysis of web server logs demands both accurate detection and human-readable explanations that can satisfy legal requirements. We present C

tutorialsarxiv-cs-ai
9 Jun 2026
Safety

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning

DGX agent

arXiv:2606.07602v1 Announce Type: cross Abstract: LLM-based LEGO assembly generation requires both semantic grounding and physical feasibility. We identify a data-induced failure mode, PhysHack, in wh

safetyarxiv-cs-ai
9 Jun 2026
Safety

SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models

DGX agent

arXiv:2606.07705v1 Announce Type: cross Abstract: Although multi-objective reinforcement learning (MORL) is central to aligning large language models with complex human preferences, the prevailing pra

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Scaffold Effects on GAIA: A Controlled Comparison

DGX agent

arXiv:2606.08529v1 Announce Type: new Abstract: Published agent capability scores conflate what a model can do with what its scaffold lets it do, and the magnitude of this elicitation gap is not well

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ScaleSweep: Accurate NVFP4 Post-Training Quantization of LLMs via Block Scale Initialization

DGX agent

arXiv:2606.07618v1 Announce Type: cross Abstract: NVFP4 is a recently introduced hardware-supported FP4 format that improves the fidelity of 4-bit quantization through fine-grained block scales. Howev

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…184185186187188…448
Next →