AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

An AI system to help scientists write expert-level empirical software

DGX agent

arXiv:2509.06503v2 Announce Type: replace Abstract: The cycle of scientific discovery is frequently bottlenecked by the slow, manual creation of software to support computational experimentsite{hannay

researcharxiv-cs-ai
19 May 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

An Assessment of Human vs. Model Uncertainty in Soft-Label Learning and Calibration

DGX agent

arXiv:2605.18648v1 Announce Type: cross Abstract: Central to human-aligned AI is understanding the benefits of human-elicited labels over synthetic alternatives. While human soft-labels improve calibr

safetyarxiv-cs-ai
19 May 2026
Safety

An Empirical Study of Privacy Leakage Chains via Prompt Injection in Black-Box Chatbot Environments

DGX agent

arXiv:2605.18133v1 Announce Type: cross Abstract: LLM-based chatbot agents increasingly process user requests by combining natural-language reasoning with external tools such as web browsing. These ca

safetyarxiv-cs-ai
19 May 2026
Research

An Information-Theoretic Criterion for Efficient Data Synthesis

DGX agent

arXiv:2605.16379v1 Announce Type: cross Abstract: Synthetic data becomes crucial for large language model training, but its effectiveness is highly inconsistent. We provide an information-theoretic ac

researcharxiv-cs-ai
19 May 2026
Tutorials

An Interpretable Closed-Loop Intelligent Tutoring System for Multimodal Affective Feedback in Asynchronous Presentation Training

DGX agent

arXiv:2605.17468v1 Announce Type: cross Abstract: This paper presents an interpretable closed-loop Intelligent Tutoring System (ITS) that supports feedback-guided practice for developing on-camera ora

tutorialsarxiv-cs-ai
19 May 2026
Safety

AnchorDiff: Topology-Aware Masked Diffusion with Confidence-based Rewriting for Radiology Report Generation

DGX agent

arXiv:2605.17071v1 Announce Type: new Abstract: Radiology report generation (RRG) aims to automatically produce clinically accurate textual reports from medical images. Existing methods predominantly

safetyarxiv-cs-ai
19 May 2026
Local Ai

ANNEAL: Adapting LLM Agents via Governed Symbolic Patch Learning

DGX agent

arXiv:2605.16309v1 Announce Type: new Abstract: LLM-based agents can recover from individual execution errors, yet they repeatedly fail on the same fault when the underlying process knowledge--operato

local-aiarxiv-cs-ai
19 May 2026
Tutorials

ANVIL: Analogies and Videos for Lecturers

DGX agent

arXiv:2605.16295v1 Announce Type: cross Abstract: We present ANVIL, a multimodal generative system that automates the production of analogy-based instructional animations for computer science topics.

tutorialsarxiv-cs-ai
19 May 2026
Safety

Are Multimodal LLMs Ready for Surveillance? A Reality Check on Zero-Shot Anomaly Detection in the Wild

DGX agent

arXiv:2603.04727v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have demonstrated impressive general competence in video understanding, yet their reliability for rea

safetyarxiv-cs-ai
19 May 2026
Research

Are Researchers Being Replaced by Artificial Intelligence?

DGX agent

arXiv:2605.16294v1 Announce Type: cross Abstract: A Nature survey from 2023 involving 1,600 researchers shows that scientists are ``concerned, as well as excited, by the increasing use of artificial-i

researcharxiv-cs-ai
19 May 2026
Research

Are Sparse Autoencoder Benchmarks Reliable?

DGX agent

arXiv:2605.18229v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) are a core interpretability tool for large language models, and progress on SAE architectures depends on benchmarks that re

researcharxiv-cs-ai
19 May 2026
Safety

ARROW: Augmented Replay for RObust World models

DGX agent

arXiv:2603.11395v2 Announce Type: replace-cross Abstract: Continual reinforcement learning challenges agents to acquire new skills while retaining previously learned ones with the goal of improving pe

safetyarxiv-cs-ai
19 May 2026
Model Releases

Artificial Adaptive Intelligence: The Missing Stage Between Narrow and General Intelligence

DGX agent

arXiv:2605.16844v1 Announce Type: new Abstract: Between the narrow systems we deploy and the general intelligence we speculate about lies an entire regime of machine behavior that has never received i

model-releasesarxiv-cs-ai
19 May 2026
Research

Artificial Intelligence can Recognize Whether a Job Applicant is Selling and/or Lying According to Facial Expressions and Head Movements Much More Correctly Than Human Interviewers

DGX agent

arXiv:2605.17461v1 Announce Type: cross Abstract: Whether an interviewee's honest and deceptive responses can be detected by facial expression signals in videos has been debated and requires further r

researcharxiv-cs-ai
19 May 2026
Model Releases

AscendOptimizer: Episodic Agent for Ascend NPU Operator Optimization

DGX agent

arXiv:2603.23566v2 Announce Type: replace-cross Abstract: Optimizing AscendC (Ascend C) operators for Ascend NPUs is difficult for two reasons. First, unlike CUDA, the ecosystem offers few public kern

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Asking Back: Interaction-Layer Antidistillation Watermarks

DGX agent

arXiv:2605.16462v1 Announce Type: cross Abstract: Detecting unauthorized knowledge distillation from a deployed LLM API is hard because the defender controls neither the attacker's training pipeline n

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents

DGX agent

arXiv:2605.17324v1 Announce Type: cross Abstract: Clarification-seeking behavior is widely regarded as a desirable property of LLM agents, enabling them to resolve ambiguity before acting on underspec

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Attention-Guided Fusion of 1D and 2D CNNs for Robust ECG-Based Biometric Recognition

DGX agent

arXiv:2605.17685v1 Announce Type: cross Abstract: Electrocardiogram (ECG)-based biometric recognition has emerged as a promising solution for secure authentication and liveness detection. However, mos

model-releasesarxiv-cs-ai
19 May 2026
Research

Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models

DGX agent

arXiv:2605.17310v1 Announce Type: cross Abstract: Existing adversarial attacks on vision-language models (VLMs) can steer model outputs toward attacker-specified target responses, but their effectiven

researcharxiv-cs-ai
19 May 2026
Applications

Attention Sinks and Outliers in Attention Residuals

DGX agent

arXiv:2605.17887v1 Announce Type: cross Abstract: We propose OASIS, an outlier- and sink-aware technique built on inter-layer null signaling. As AttnResidual architectures introduce an additional dept

applicationsarxiv-cs-ai
19 May 2026
Safety

Augmenting Human Evaluation with LLM Judges: How Many Human Reviews Do You Need?

DGX agent

arXiv:2605.16354v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as automated evaluators of AI systems, including in high-stakes applications. In this role, LLMs ar

safetyarxiv-cs-ai
19 May 2026
Model Releases

AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing

DGX agent

arXiv:2603.23069v2 Announce Type: replace-cross Abstract: The task of authorship style transfer involves rewriting text in the style of a target author while preserving the meaning of the original tex

model-releasesarxiv-cs-ai
19 May 2026
Research

Automated Coding of Communication Data Using ChatGPT: Consistency Across Subgroups

DGX agent

arXiv:2510.20584v3 Announce Type: replace-cross Abstract: Assessing communication and collaboration at scale depends on a labor-intensive task of coding communication data into categories according to

researcharxiv-cs-ai
19 May 2026
Tutorials

Automated Knowledge Component Generation for Interpretable Knowledge Tracing in Coding Problems

DGX agent

arXiv:2502.18632v4 Announce Type: replace Abstract: Knowledge components (KCs) mapped to problems help model student learning, tracking their mastery levels on fine-grained skills thereby facilitating

tutorialsarxiv-cs-ai
19 May 2026
Model Releases

Automated Root-Cause Subclassification and No-Code Fix Generation for Invalid Bug Reports

DGX agent

arXiv:2605.17561v1 Announce Type: cross Abstract: Issues faced when using software are reported in the form of bug reports. However, many bug reports are invalid, meaning they do not require code chan

model-releasesarxiv-cs-ai
19 May 2026
Safety

Automatic Generation of High-Performance RL Environments

DGX agent

arXiv:2603.12145v2 Announce Type: replace-cross Abstract: Translating complex reinforcement learning (RL) environments into high-performance implementations has traditionally required months of specia

safetyarxiv-cs-ai
19 May 2026
Applications

Automatic Unsupervised Ensemble Outlier Model Selection--Extended Version

DGX agent

arXiv:2605.16567v1 Announce Type: cross Abstract: Unsupervised outlier detection is attractive because it eliminates the need for labeled data. Moreover, forming multi-model ensembles can improve dete

applicationsarxiv-cs-ai
19 May 2026
Safety

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment

DGX agent

arXiv:2605.17602v1 Announce Type: new Abstract: Aligning Text-to-Image (T2I) generation models with human preferences increasingly relies on image reward models that score or rank generated images acc

safetyarxiv-cs-ai
19 May 2026
Safety

Avoiding Structural Failure Modes in Tabular Fair SSL: Online Primal-Dual Allocation under Confidence Gating

DGX agent

arXiv:2605.16446v1 Announce Type: cross Abstract: Semi-supervised learning (SSL) enables prediction with limited labels, but high-stakes tabular applications (medical, credit, recidivism) require stat

safetyarxiv-cs-ai
19 May 2026
Agents

Baba in Wonderland: Online Self-Supervised Dynamics Discovery for Executable World Models

DGX agent

arXiv:2605.16725v1 Announce Type: new Abstract: Executable world models can be read, edited, executed, and reused for planning, but only if the program captures the environment's transition law rather

agentsarxiv-cs-ai
19 May 2026
Model Releases

Babel: Jailbreaking Safety Attention via Obfuscation Distribution Optimized Sampling

DGX agent

arXiv:2605.17971v1 Announce Type: cross Abstract: Despite rigorous safety alignment, Large Language Models (LLMs) remain vulnerable to jailbreak attacks. Existing black-box methods often rely on heuri

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

BacktestBench: Benchmarking Large Language Models for Automated Quantitative Strategy Backtesting

DGX agent

arXiv:2605.17937v1 Announce Type: cross Abstract: Quantitative backtesting is essential for evaluating trading strategies but remains hampered by high technical barriers and limited scalability. While

model-releasesarxiv-cs-ai
19 May 2026
Research

Balancing Knowledge Distillation for Imbalance Learning with Bilevel Optimization

DGX agent

arXiv:2605.17839v1 Announce Type: cross Abstract: Knowledge distillation transfers knowledge from a high capacity teacher to a compact student using a mixture of hard and soft losses. On imbalanced da

researcharxiv-cs-ai
19 May 2026
Model Releases

Barriers for Learning in an Evolving World: Mathematical Understanding of Loss of Plasticity

DGX agent

arXiv:2510.00304v3 Announce Type: replace-cross Abstract: Deep learning models excel in stationary data but struggle in non-stationary environments due to a phenomenon known as loss of plasticity (LoP

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Bayesian-Monte Carlo Schedule Updating for Construction Digital Twins: A Probabilistic Framework for Dynamic Project Forecasting

DGX agent

arXiv:2605.17608v1 Announce Type: cross Abstract: Construction projects frequently experience schedule delays and forecasting uncertainty due to variability in labor productivity, material availabilit

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Beacon: Single-Turn Diagnosis and Mitigation of Latent Sycophancy in Large Language Models

DGX agent

arXiv:2510.16727v2 Announce Type: replace-cross Abstract: Large language models internalize a structural trade-off between truthfulness and obsequious flattery, emerging from reward optimization that

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Benchmarking Mythos-Linked Bug Rediscovery

DGX agent

arXiv:2605.17416v1 Announce Type: cross Abstract: Anthropic's April 2026 Mythos materials combine benchmark claims with concrete bug-finding stories across OpenBSD, FreeBSD, Linux, FFmpeg, and browser

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

BESplit: Bias-Compensated Split Federated Learning with Evidential Aggregation

DGX agent

arXiv:2605.17508v1 Announce Type: cross Abstract: Split Federated Learning (SFL) enables privacy-preserving collaborative training by partitioning models between clients and a server. However, under n

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs

DGX agent

arXiv:2602.09805v2 Announce Type: replace-cross Abstract: As reasoning LLMs increasingly trade tokens for accuracy through deliberation, search, and self-correction, a single accuracy score can no lon

model-releasesarxiv-cs-ai
19 May 2026
Research

Beyond Accuracy: Robustness, Interpretability and Expressiveness of EEG Foundation Models

DGX agent

arXiv:2605.17562v1 Announce Type: cross Abstract: EEG foundation models (EEG-FMs) have been evaluated predominantly on clean, in-distribution accuracy, leaving their robustness, interpretability and r

researcharxiv-cs-ai
19 May 2026
Applications

Beyond Catalogue Counts: the Dataset Visibility Asymmetry in Low-Resource Multilingual NLP

DGX agent

arXiv:2605.17442v1 Announce Type: cross Abstract: Multilingual NLP often relies on dataset counts from centralized catalogues to characterize which languages are resource-rich or resource-poor. Howeve

applicationsarxiv-cs-ai
19 May 2026
Safety

Beyond Compliance: How AI Could Help Creative Writers by Refusing Them

DGX agent

arXiv:2605.16272v1 Announce Type: cross Abstract: Mainstream creativity support design prioritizes compliant AI for seamless writing interactions, but concerns over inappropriate AI reliance highlight

safetyarxiv-cs-ai
19 May 2026
Research

Beyond Correctness: Harmonizing Process and Outcome Rewards through RL Training

DGX agent

arXiv:2509.03403v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves final-answer accuracy on reasoning tasks, but it does not reliably improve reas

researcharxiv-cs-ai
19 May 2026
Research

Beyond Execution: Static-Analysis Rewards and Hint-Conditioned Diffusion RL for Code Generation

DGX agent

arXiv:2605.17174v1 Announce Type: cross Abstract: Reinforcement Learning (RL) is an important paradigm for aligning Diffusion Language Models (DLMs) toward functional correctness in code generation. H

researcharxiv-cs-ai
19 May 2026
Applications

Beyond Imperfect Alternatives with Rulemapping: A Neuro-Symbolic Case Study on Online Hate Speech

DGX agent

arXiv:2605.16280v1 Announce Type: cross Abstract: Automating legal reasoning forces a choice between imperfect alternatives: symbolic systems offer transparency but struggle with ambiguity, whereas ne

applicationsarxiv-cs-ai
19 May 2026
Local Ai

Beyond Inference-Time Search: Reinforcement Learning Synthesizes Reusable Solvers

DGX agent

arXiv:2605.18374v1 Announce Type: cross Abstract: Large language models (LLMs) typically approach combinatorial optimization as an inference-time procedure, solving each instance separately through sa

local-aiarxiv-cs-ai
19 May 2026
Research

Beyond Linear Superposition: Discovering Climate Features in AI Weather Models with KAN-SAE

DGX agent

arXiv:2605.17493v1 Announce Type: cross Abstract: Deep learning weather prediction models achieve remarkable predictive skill yet remain largely opaque: we know little about how they represent physica

researcharxiv-cs-ai
19 May 2026
Research

Beyond Morphology: Quantifying the Diagnostic Power of Color Features in Cancer Classification

DGX agent

arXiv:2605.18522v1 Announce Type: cross Abstract: In histopathology, human experts primarily rely on color as a means of enhancing contrast to interpret tissue morphology, whereas machine vision model

researcharxiv-cs-ai
19 May 2026
← Previous
1…291292293294295…448
Next →