AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Research

Metamorphic Testing with the Rashomon Set: Explanation Faithfulness in Machine Learning

DGX agent

arXiv:2606.06056v1 Announce Type: cross Abstract: Multiple machine learning models can achieve near-equivalent predictive performance on the same task, yet provide divergent feature-based explanations

researcharxiv-cs-ai
6 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

Microskill Architecture: A Modular Skill-Driven Framework for AI-Native Code Generation

DGX agent

arXiv:2606.05720v1 Announce Type: cross Abstract: Large language models and AI coding agents have reshaped software development, but the path to fully AI-native systems faces structural challenges. Ch

agentsarxiv-cs-ai
6 Jun 2026
Model Releases

Minimizing the Hidden Cost of Scales: Graph-Guided Ultra-Low-Bit Quantization for Large Language Models

DGX agent

arXiv:2606.05429v1 Announce Type: new Abstract: Post-training quantization (PTQ) is critical for the efficient deployment of large language models (LLMs). Recent ultra-low-bit PTQ methods rely on rigi

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Multi-ResNets for Subspace Preconditioning in Constrained Optimization

DGX agent

arXiv:2606.06300v1 Announce Type: new Abstract: We propose MResOpt, a staged residual neural network architecture for constrained optimization problems. Our architecture fits within predict-complete-c

researcharxiv-cs-ai
6 Jun 2026
Model Releases

Multilingual Fine-Tuning via Localized Gradient Conflict Resolution

DGX agent

arXiv:2606.05613v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) has established cross-lingual versatility as a defining feature of modern systems. However, fine-tun

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Mutation Without Variation: Convergence Dynamics in LLM-Driven Program Evolution

DGX agent

arXiv:2606.05408v1 Announce Type: new Abstract: When an LLM repeatedly mutates a program, does it explore new forms or circle back to the same ones? We study this question by analyzing LLM-driven muta

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

No Need to Train Your RDB Foundation Model

DGX agent

arXiv:2602.13697v2 Announce Type: replace Abstract: Relational databases (RDBs) contain vast amounts of heterogeneous tabular information that can be exploited for predictive modeling purposes. But si

model-releasesarxiv-cs-ai
6 Jun 2026
Local Ai

Ontology-constrained multi-LLM scoring of hypothesis support in the predictive processing literature

DGX agent

arXiv:2606.05206v1 Announce Type: cross Abstract: Fragmentation is common in interdisciplinary fields with diverse methods and theoretical commitments. Predictive coding neuroscience is a clear exampl

local-aiarxiv-cs-ai
6 Jun 2026
Model Releases

OPRD: On-Policy Representation Distillation

DGX agent

arXiv:2606.06021v1 Announce Type: cross Abstract: On-policy distillation (OPD) supervises the student only in output space by matching next-token probabilities. This output-only paradigm has two limit

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Output Type Before Quality: A Standards-Derived XAI Admissibility Rubric for Autonomous-Driving Safety

DGX agent

arXiv:2606.05461v1 Announce Type: new Abstract: Safety standards for ML-based autonomous driving specify the kind of evidence an assurance case must contain (directed cause-and-effect chains, quantifi

safetyarxiv-cs-ai
6 Jun 2026
Applications

PAMF: Prior-Aware Multimodal Fusion for Incomplete Time Series Data

DGX agent

arXiv:2606.06328v1 Announce Type: cross Abstract: In healthcare, multimodal time series tasks often operate on incomplete observations in practice, for example when ECG segments are lost because elect

applicationsarxiv-cs-ai
6 Jun 2026
Research

Pattern Selectivity is Not Task-Causal Structure: A Cross-Architecture Mechanistic Study of Composed-Task Circuits in 1B-Class Language Models

DGX agent

arXiv:2606.05378v1 Announce Type: cross Abstract: We test whether a single screen-and-ablate recipe -- identify attention-head circuits by task-pattern selectivity, then verify by causal ablation agai

researcharxiv-cs-ai
6 Jun 2026
Model Releases

PC Layer: Polynomial Weight Preconditioning for Improving LLM Pre-Training

DGX agent

arXiv:2606.06470v1 Announce Type: cross Abstract: We propose a preconditioning (PC) layer, a weight parameterization via polynomial preconditioner that ensures stable weight conditioning throughout LL

model-releasesarxiv-cs-ai
6 Jun 2026
Research

PerceptUI: LLM Agents as Human-Aligned Synthetic Users for UI/UX Evaluation

DGX agent

arXiv:2606.05697v1 Announce Type: new Abstract: User interface (UI) and user experience (UX) evaluation is central to product development, yet reliable feedback still relies on recruiting human partic

researcharxiv-cs-ai
6 Jun 2026
Research

Plug-and-Play Guidance for Discrete Diffusion Models via Gradient-Informed Logit Correction

DGX agent

arXiv:2606.06303v1 Announce Type: cross Abstract: Controllable generation with discrete diffusion models is often hindered by high computational overhead or the need for retraining. In this paper, we

researcharxiv-cs-ai
6 Jun 2026
Safety

Policy-Conditioned Counterfactual Credit for Verifiable Reinforcement Learning of Long-Horizon Language Agents

DGX agent

arXiv:2606.05263v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards improves reasoning and tool use, yet long-horizon language agents still learn unsupported evidence chai

safetyarxiv-cs-ai
6 Jun 2026
Tutorials

Pretraining Recurrent Networks without Recurrence

DGX agent

arXiv:2606.06479v1 Announce Type: cross Abstract: Training recurrent neural networks (RNNs) requires assigning credit across long sequences of computations. Standard backpropagation through time (BPTT

tutorialsarxiv-cs-ai
6 Jun 2026
Model Releases

ProfiliTable: Profiling-Driven Tabular Data Processing via Agentic Workflows

DGX agent

arXiv:2605.12376v2 Announce Type: replace Abstract: Table processing-including cleaning, transformation, augmentation, and matching-is a foundational yet error-prone stage in real-world data pipelines

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

PSEBench: A Controllable and Verifiable Benchmark for Evaluating LLMs in Patient Safety Event Triage

DGX agent

arXiv:2606.05463v1 Announce Type: new Abstract: Patient safety event triage, determining whether a clinical event is reportable under jurisdiction-specific policy, is a high-stakes task typically perf

model-releasesarxiv-cs-ai
6 Jun 2026
Research

QCFuse: Query-Aware Cache Fusion via Compressed View for Efficient RAG Serving

DGX agent

arXiv:2606.05875v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves large language model (LLM) answer quality by grounding generation in external evidence, but processing ret

researcharxiv-cs-ai
6 Jun 2026
Research

Quantum enhanced rare event discovery and sampling

DGX agent

arXiv:2606.06316v1 Announce Type: cross Abstract: Financial crashes, cascading failures in infrastructure, and critical errors in AI systems are frequently triggered by events that occur with extremel

researcharxiv-cs-ai
6 Jun 2026
Applications

RAG Security and Privacy: Formalizing the Threat Model and Attack Surface

DGX agent

arXiv:2509.20324v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) is an emerging approach in natural language processing that combines large language models (LLMs) with ex

applicationsarxiv-cs-ai
6 Jun 2026
Agents

RAINO: Anchoring Agents in Reality, A Systematic Review and Conceptual Framework for Realism in Agent-Based Modelling

DGX agent

arXiv:2606.05167v1 Announce Type: cross Abstract: Realism is a central yet seemingly under-theorized concept in Agent-Based Modelling. This paper presents a Systematic Literature Review, aiming to ide

agentsarxiv-cs-ai
6 Jun 2026
Hardware

RedKnot: Efficient Long-Context LLM Serving with Head-Aware KV Reuse and SegPagedAttention

DGX agent

arXiv:2606.06256v1 Announce Type: new Abstract: As the input length of large language model (LLM) serving continues to grow, the KV cache has become a dominant bottleneck in AI infrastructure. It limi

hardwarearxiv-cs-ai
6 Jun 2026
Research

Reformulating Neural Operators in d+1 Dimensions for Embedding Evolution

DGX agent

arXiv:2505.11766v4 Announce Type: replace-cross Abstract: Neural Operators (NOs) are powerful architectures for learning mappings between function spaces. While most advances focus on refining kernel

researcharxiv-cs-ai
6 Jun 2026
Safety

Regret Minimization with Adaptive Opponents in Repeated Games

DGX agent

arXiv:2606.06486v1 Announce Type: cross Abstract: In this paper, we study regret minimization in repeated games with adaptive opponents who can respond based on histories of play. The standard metric

safetyarxiv-cs-ai
6 Jun 2026
Research

Representation Learning Enables Scalable Multitask Deep Reinforcement Learning

DGX agent

arXiv:2606.05555v1 Announce Type: cross Abstract: Scaling reinforcement learning (RL) to diverse multitask settings remains a central challenge. While recent advances in model-based RL achieve strong

researcharxiv-cs-ai
6 Jun 2026
Safety

Residual Modeling for High-Fidelity Learned Compression of Scientific Data

DGX agent

arXiv:2606.05389v1 Announce Type: new Abstract: Lossy compression is essential for massive spatiotemporal data from scientific simulations. Learned compressors can achieve high compression ratios at m

safetyarxiv-cs-ai
6 Jun 2026
Applications

Rethinking Infrastructure Inspection as Image Difference Classification: A Traffic Sign Case Study

DGX agent

arXiv:2606.06375v1 Announce Type: new Abstract: Digital twins (DTs) allow the digitalization of road infrastructure inspection, though this is hindered by limited annotated data. This work exploits th

applicationsarxiv-cs-ai
6 Jun 2026
Model Releases

Retry Policy Gradients in Continuous Action Spaces

DGX agent

arXiv:2606.05888v1 Announce Type: new Abstract: Retry-based objectives such as pass@K and max@K optimize the best return obtained from multiple sampled trajectories, and recent work has shown that the

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Reward-Decomposed Reinforcement Learning for Immersive Video Role-Playing

DGX agent

arXiv:2605.04733v2 Announce Type: replace Abstract: Text-based role-playing models can imitate character styles, but often fail to capture scene atmosphere and evolving tension, which are crucial for

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Reward Learning through Ranking Mean Squared Error

DGX agent

arXiv:2601.09236v3 Announce Type: replace-cross Abstract: Reward design remains a significant bottleneck in applying reinforcement learning (RL) to real-world problems. A popular alternative is reward

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Risk Assessment of Autonomous Driving: Integrating Technical Failures, Ethical Dilemmas, and Policy Frameworks

DGX agent

arXiv:2606.06396v1 Announce Type: new Abstract: Autonomous driving technology has the potential to reduce the large number of road traffic accidents caused by human error each year, but it also brings

safetyarxiv-cs-ai
6 Jun 2026
Safety

RREDCoT: Segment-Level Reward Redistribution for Reasoning Models

DGX agent

arXiv:2606.06475v1 Announce Type: cross Abstract: Recent advancements in reasoning language models have been driven by Reinforcement Learning (RL) fine-tuning. Most often, these rely on the Group Rela

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

Safety Paradox: How Enhanced Safety Awareness Leaves LLMs Vulnerable to Posterior Attack

DGX agent

arXiv:2606.05614v1 Announce Type: new Abstract: Large language models (LLMs) are rigorously aligned to refuse harmful requests, a process that inherently cultivates a latent capacity to evaluate and r

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

SAGE: Scalable AI Governance & Evaluation

DGX agent

arXiv:2602.07840v3 Announce Type: replace-cross Abstract: Evaluating relevance in large-scale search systems is fundamentally constrained by the governance gap between nuanced, resource-constrained hu

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

SagnacAssisted Enhanced OTDR for Distributed Acoustic Sensing: A Standardized Benchmark and Engineering Evaluation Framework

DGX agent

arXiv:2606.05754v1 Announce Type: cross Abstract: Phase-sensitive optical time-domain reflectometry (phi-OTDR) is widely used in large-scale distributed acoustic sensing (DAS) because it provides dist

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Scaling Laws and Spectra of Shallow Neural Networks in the Feature Learning Regime

DGX agent

arXiv:2509.24882v2 Announce Type: replace-cross Abstract: Neural scaling laws underlie many of the recent advances in deep learning, yet their theoretical understanding remains largely confined to lin

researcharxiv-cs-ai
6 Jun 2026
Model Releases

SciVisAgentSkills: Design and Evaluation of Agent Skills for Scientific Data Analysis and Visualization

DGX agent

arXiv:2606.05525v1 Announce Type: new Abstract: Recent advances in agentic visualization have enabled the translation of natural language into executable scientific visualization (SciVis) workflows. W

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Search-Time Contamination in Deep Research Agents: Measuring Performance Inflation in Public Benchmark Evaluation

DGX agent

arXiv:2606.05241v1 Announce Type: cross Abstract: Public benchmarks enable fair and reproducible evaluation of LLM reasoning, but they become fragile for deep research agents that actively search the

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Selective-Advantage Entropy-Adaptive Horizon GRPO: Asymmetric Token-Level Discounting for Efficient Reinforcement Learning of Language Models

DGX agent

arXiv:2606.05434v1 Announce Type: cross Abstract: Group Relative Policy Optimisation (GRPO) has emerged as an effective reinforcement-learning algorithm for aligning language models on reasoning tasks

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Self-Commitment Latency: A Reward-Free Probe for Prompted Implicit Hacking

DGX agent

arXiv:2606.05625v1 Announce Type: new Abstract: Implicit reward hacking is hard to audit when a language model's chain of thought appears benign: a final answer may be anchored by a prompt shortcut wh

researcharxiv-cs-ai
6 Jun 2026
Research

Semantic Partial Grounding via LLMs

DGX agent

arXiv:2602.22067v2 Announce Type: replace Abstract: Grounding is a critical step in classical planning, yet it often becomes a computational bottleneck due to the exponential growth in grounded action

researcharxiv-cs-ai
6 Jun 2026
Model Releases

SentinelBench: A Benchmark for Long-Running Monitoring Agents

DGX agent

arXiv:2606.05342v1 Announce Type: new Abstract: AI agents are increasingly asked to carry out work that spans minutes, hours, or longer. Yet the default model of agent behavior is continuous action: i

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Separation Power of Equivariant Neural Networks

DGX agent

arXiv:2406.08966v3 Announce Type: replace-cross Abstract: The separation power of a machine learning model refers to its ability to distinguish between different inputs and is often used as a proxy fo

researcharxiv-cs-ai
6 Jun 2026
Research

Severity-Aware Curriculum Learning with Multi-Model Response Selection for Medical Text Generation

DGX agent

arXiv:2606.05510v1 Announce Type: new Abstract: Telehealth systems have become increasingly important for delivering accessible and timely medical information. Existing large language models often str

researcharxiv-cs-ai
6 Jun 2026
Safety

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks

DGX agent

arXiv:2606.05609v1 Announce Type: cross Abstract: As large language models (LLMs) are widely deployed, identifying their vulnerability through jailbreak attacks becomes increasingly critical. Optimiza

safetyarxiv-cs-ai
6 Jun 2026
Safety

Soft Sequence Policy Optimization

DGX agent

arXiv:2602.19327v3 Announce Type: replace-cross Abstract: A significant portion of recent research on Large Language Model (LLM) alignment focuses on developing new policy optimization methods based o

safetyarxiv-cs-ai
6 Jun 2026
← Previous
1…200201202203204…452
Next →