AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Agents

RAINO: Anchoring Agents in Reality, A Systematic Review and Conceptual Framework for Realism in Agent-Based Modelling

DGX agent

arXiv:2606.05167v1 Announce Type: cross Abstract: Realism is a central yet seemingly under-theorized concept in Agent-Based Modelling. This paper presents a Systematic Literature Review, aiming to ide

agentsarxiv-cs-ai
6 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

RedKnot: Efficient Long-Context LLM Serving with Head-Aware KV Reuse and SegPagedAttention

DGX agent

arXiv:2606.06256v1 Announce Type: new Abstract: As the input length of large language model (LLM) serving continues to grow, the KV cache has become a dominant bottleneck in AI infrastructure. It limi

hardwarearxiv-cs-ai
6 Jun 2026
Research

Reformulating Neural Operators in d+1 Dimensions for Embedding Evolution

DGX agent

arXiv:2505.11766v4 Announce Type: replace-cross Abstract: Neural Operators (NOs) are powerful architectures for learning mappings between function spaces. While most advances focus on refining kernel

researcharxiv-cs-ai
6 Jun 2026
Safety

Regret Minimization with Adaptive Opponents in Repeated Games

DGX agent

arXiv:2606.06486v1 Announce Type: cross Abstract: In this paper, we study regret minimization in repeated games with adaptive opponents who can respond based on histories of play. The standard metric

safetyarxiv-cs-ai
6 Jun 2026
Research

Representation Learning Enables Scalable Multitask Deep Reinforcement Learning

DGX agent

arXiv:2606.05555v1 Announce Type: cross Abstract: Scaling reinforcement learning (RL) to diverse multitask settings remains a central challenge. While recent advances in model-based RL achieve strong

researcharxiv-cs-ai
6 Jun 2026
Safety

Residual Modeling for High-Fidelity Learned Compression of Scientific Data

DGX agent

arXiv:2606.05389v1 Announce Type: new Abstract: Lossy compression is essential for massive spatiotemporal data from scientific simulations. Learned compressors can achieve high compression ratios at m

safetyarxiv-cs-ai
6 Jun 2026
Applications

Rethinking Infrastructure Inspection as Image Difference Classification: A Traffic Sign Case Study

DGX agent

arXiv:2606.06375v1 Announce Type: new Abstract: Digital twins (DTs) allow the digitalization of road infrastructure inspection, though this is hindered by limited annotated data. This work exploits th

applicationsarxiv-cs-ai
6 Jun 2026
Model Releases

Retry Policy Gradients in Continuous Action Spaces

DGX agent

arXiv:2606.05888v1 Announce Type: new Abstract: Retry-based objectives such as pass@K and max@K optimize the best return obtained from multiple sampled trajectories, and recent work has shown that the

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Reward-Decomposed Reinforcement Learning for Immersive Video Role-Playing

DGX agent

arXiv:2605.04733v2 Announce Type: replace Abstract: Text-based role-playing models can imitate character styles, but often fail to capture scene atmosphere and evolving tension, which are crucial for

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Reward Learning through Ranking Mean Squared Error

DGX agent

arXiv:2601.09236v3 Announce Type: replace-cross Abstract: Reward design remains a significant bottleneck in applying reinforcement learning (RL) to real-world problems. A popular alternative is reward

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Risk Assessment of Autonomous Driving: Integrating Technical Failures, Ethical Dilemmas, and Policy Frameworks

DGX agent

arXiv:2606.06396v1 Announce Type: new Abstract: Autonomous driving technology has the potential to reduce the large number of road traffic accidents caused by human error each year, but it also brings

safetyarxiv-cs-ai
6 Jun 2026
Safety

RREDCoT: Segment-Level Reward Redistribution for Reasoning Models

DGX agent

arXiv:2606.06475v1 Announce Type: cross Abstract: Recent advancements in reasoning language models have been driven by Reinforcement Learning (RL) fine-tuning. Most often, these rely on the Group Rela

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

Safety Paradox: How Enhanced Safety Awareness Leaves LLMs Vulnerable to Posterior Attack

DGX agent

arXiv:2606.05614v1 Announce Type: new Abstract: Large language models (LLMs) are rigorously aligned to refuse harmful requests, a process that inherently cultivates a latent capacity to evaluate and r

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

SAGE: Scalable AI Governance & Evaluation

DGX agent

arXiv:2602.07840v3 Announce Type: replace-cross Abstract: Evaluating relevance in large-scale search systems is fundamentally constrained by the governance gap between nuanced, resource-constrained hu

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

SagnacAssisted Enhanced OTDR for Distributed Acoustic Sensing: A Standardized Benchmark and Engineering Evaluation Framework

DGX agent

arXiv:2606.05754v1 Announce Type: cross Abstract: Phase-sensitive optical time-domain reflectometry (phi-OTDR) is widely used in large-scale distributed acoustic sensing (DAS) because it provides dist

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Scaling Laws and Spectra of Shallow Neural Networks in the Feature Learning Regime

DGX agent

arXiv:2509.24882v2 Announce Type: replace-cross Abstract: Neural scaling laws underlie many of the recent advances in deep learning, yet their theoretical understanding remains largely confined to lin

researcharxiv-cs-ai
6 Jun 2026
Model Releases

SciVisAgentSkills: Design and Evaluation of Agent Skills for Scientific Data Analysis and Visualization

DGX agent

arXiv:2606.05525v1 Announce Type: new Abstract: Recent advances in agentic visualization have enabled the translation of natural language into executable scientific visualization (SciVis) workflows. W

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Search-Time Contamination in Deep Research Agents: Measuring Performance Inflation in Public Benchmark Evaluation

DGX agent

arXiv:2606.05241v1 Announce Type: cross Abstract: Public benchmarks enable fair and reproducible evaluation of LLM reasoning, but they become fragile for deep research agents that actively search the

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Selective-Advantage Entropy-Adaptive Horizon GRPO: Asymmetric Token-Level Discounting for Efficient Reinforcement Learning of Language Models

DGX agent

arXiv:2606.05434v1 Announce Type: cross Abstract: Group Relative Policy Optimisation (GRPO) has emerged as an effective reinforcement-learning algorithm for aligning language models on reasoning tasks

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Self-Commitment Latency: A Reward-Free Probe for Prompted Implicit Hacking

DGX agent

arXiv:2606.05625v1 Announce Type: new Abstract: Implicit reward hacking is hard to audit when a language model's chain of thought appears benign: a final answer may be anchored by a prompt shortcut wh

researcharxiv-cs-ai
6 Jun 2026
Research

Semantic Partial Grounding via LLMs

DGX agent

arXiv:2602.22067v2 Announce Type: replace Abstract: Grounding is a critical step in classical planning, yet it often becomes a computational bottleneck due to the exponential growth in grounded action

researcharxiv-cs-ai
6 Jun 2026
Model Releases

SentinelBench: A Benchmark for Long-Running Monitoring Agents

DGX agent

arXiv:2606.05342v1 Announce Type: new Abstract: AI agents are increasingly asked to carry out work that spans minutes, hours, or longer. Yet the default model of agent behavior is continuous action: i

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Separation Power of Equivariant Neural Networks

DGX agent

arXiv:2406.08966v3 Announce Type: replace-cross Abstract: The separation power of a machine learning model refers to its ability to distinguish between different inputs and is often used as a proxy fo

researcharxiv-cs-ai
6 Jun 2026
Research

Severity-Aware Curriculum Learning with Multi-Model Response Selection for Medical Text Generation

DGX agent

arXiv:2606.05510v1 Announce Type: new Abstract: Telehealth systems have become increasingly important for delivering accessible and timely medical information. Existing large language models often str

researcharxiv-cs-ai
6 Jun 2026
Safety

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks

DGX agent

arXiv:2606.05609v1 Announce Type: cross Abstract: As large language models (LLMs) are widely deployed, identifying their vulnerability through jailbreak attacks becomes increasingly critical. Optimiza

safetyarxiv-cs-ai
6 Jun 2026
Safety

Soft Sequence Policy Optimization

DGX agent

arXiv:2602.19327v3 Announce Type: replace-cross Abstract: A significant portion of recent research on Large Language Model (LLM) alignment focuses on developing new policy optimization methods based o

safetyarxiv-cs-ai
6 Jun 2026
Research

Stable Deep Reinforcement Learning via Isotropic Gaussian Representations

DGX agent

arXiv:2602.19373v3 Announce Type: replace-cross Abstract: Deep reinforcement learning systems often suffer from unstable training dynamics due to non-stationarity, where learning objectives and data d

researcharxiv-cs-ai
6 Jun 2026
Research

Step-adaptive multimodal fusion network with multi-scale cloud feature learning for ultra-short-term solar irradiance forecasting

DGX agent

arXiv:2606.06102v1 Announce Type: new Abstract: Ultra-short-term solar irradiance prediction is critical for photovoltaic system dispatch and power grid stability. Existing approaches suffer from thre

researcharxiv-cs-ai
6 Jun 2026
Model Releases

Step-by-Step Optimization-like Reasoning in LLMs over Expanding Search Spaces

DGX agent

arXiv:2606.05464v1 Announce Type: new Abstract: Verifiable reward training has improved mathematical and coding reasoning, but these domains capture only part of step-by-step decision making. Many rea

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Subspace-Aware Sparse Autoencoders for Effective Mechanistic Interpretability

DGX agent

arXiv:2606.06333v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) are widely used for mechanistic interpretability in large language models, yet their formulation assigns each latent featur

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Synapse: Federated Tool Routing via Typed Compendium Artifacts

DGX agent

arXiv:2602.00911v2 Announce Type: replace Abstract: The unit of collaboration in federated learning determines what guarantees are even expressible. Flat units like weights, prompts, raw examples, car

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Synthetic Contrastive Reasoning for Multi-Table Q&A

DGX agent

arXiv:2606.05382v1 Announce Type: new Abstract: Multi-table question answering requires models to retrieve relevant evidence, link schemas, and perform compositional reasoning across relational tables

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

TAPO: Tool-Aware Policy Optimization via Credit Transfer for Multimodal Search Agents

DGX agent

arXiv:2606.05784v1 Announce Type: new Abstract: We identify and formally characterize credit misassignment as a systematic failure mode of GRPO in tool-augmented multimodal search agents: its uniform

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

The End of Software Engineering: How AI Agents Are Fundamentally Restructuring the Software Paradigm

DGX agent

arXiv:2606.05608v1 Announce Type: cross Abstract: For over half a century, software engineering has operated on a foundational premise: human engineers decompose problems, encode decision logic into s

model-releasesarxiv-cs-ai
6 Jun 2026
Tutorials

The Role of Instructional Guidance in Generative AI-Assisted Learning: Empirical Evidence from Construction Engineering Education

DGX agent

arXiv:2606.05509v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) is increasingly used to support self-directed learning, yet student interaction with such systems often remain

tutorialsarxiv-cs-ai
6 Jun 2026
Research

The Score Hamiltonian: Mapping Diffusion Models to Adiabatic Transport

DGX agent

arXiv:2606.05217v1 Announce Type: cross Abstract: We exhibit an exact correspondence between sampling with score-based diffusion models and adiabatic transport of ground states for a family of Schrodi

researcharxiv-cs-ai
6 Jun 2026
Agents

The Virtual Roundtable: Multi-Agent Personas Simulating the Dynamics of Human Brainstorming

DGX agent

arXiv:2606.05178v1 Announce Type: cross Abstract: As AI-driven product development accelerates, the bottleneck is shifting from how we build to what we build. Traditional human brainstorming faces cha

agentsarxiv-cs-ai
6 Jun 2026
Agents

TinyML-Driven Cybersecurity for Autonomous Spacecraft: Latency-Accuracy Analysis for SPARTA RF and Cyber Threat Detection

DGX agent

arXiv:2606.05779v1 Announce Type: cross Abstract: Autonomous spacecraft require rapid, lightweight, and reliable onboard detection of cyber-RF threats. Using the SPARTA attack model, we analyze the la

agentsarxiv-cs-ai
6 Jun 2026
Model Releases

TLA-Prover: Verifiable TLA+ Specification Synthesis via Preference-Optimized Low-Rank Adaptation

DGX agent

arXiv:2606.06133v1 Announce Type: cross Abstract: TLA+ is a formal specification language for verifying distributed systems and safety-critical protocols. Large language models (LLMs) frequently produ

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

TokenMizer: Graph-Structured Session Memory for Long-Horizon LLM Context Management

DGX agent

arXiv:2606.06337v1 Announce Type: new Abstract: Large language model (LLM) deployments for long-horizon tasks face a fundamental constraint: context windows are finite while productive work sessions a

model-releasesarxiv-cs-ai
6 Jun 2026
Local Ai

TOKI: A Bitemporal Operator Algebra for Contradiction Resolution in LLM-Agent Persistent Memory

DGX agent

arXiv:2606.06240v1 Announce Type: cross Abstract: Persistent memory for an LLM agent is a write-heavy substrate: every belief update is a versioned write, and a new claim may contradict a stored one.

local-aiarxiv-cs-ai
6 Jun 2026
Model Releases

ToolChoiceConfusion: Causal Minimal Tool Filtering for Reliable LLM Agents

DGX agent

arXiv:2606.06284v1 Announce Type: new Abstract: Large language model agents increasingly rely on external tools, but larger tool menus can reduce reliability and efficiency by increasing wrong-tool ca

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Towards AI epidemiology: a measurement standardisation framework for prospective risk detection

DGX agent

arXiv:2512.15783v3 Announce Type: replace Abstract: This paper proposes a measurement standardisation framework that compresses expert-AI interactions into structured, comparable fields for prospectiv

safetyarxiv-cs-ai
6 Jun 2026
Safety

Towards Healthy Evolution: Exploring the Role and Mechanisms of Human-Agent Interaction in Self-Evolving Systems

DGX agent

arXiv:2606.06114v1 Announce Type: new Abstract: Self-evolving agents improve through continual self-play and self-generated learning signals, but autonomous evolution can also cause capability degrada

safetyarxiv-cs-ai
6 Jun 2026
Research

Towards the Readability of LLM-Generated Codes through Multitask Representation Engineering

DGX agent

arXiv:2606.06214v1 Announce Type: cross Abstract: Correctness and readability are key measures of code quality, respectively ensuring functional fidelity and ease of comprehension. While most existing

researcharxiv-cs-ai
6 Jun 2026
Research

Towards Unified and Data-Efficient Prognostics and Health Management with Tabular Foundation Models

DGX agent

arXiv:2606.05481v1 Announce Type: cross Abstract: Data-driven Prognostics and Health Management (PHM) uses time-varying condition-monitoring data to diagnose system states and estimate remaining usefu

researcharxiv-cs-ai
6 Jun 2026
Safety

Towards World Models in Biomedical Research

DGX agent

arXiv:2606.05925v1 Announce Type: new Abstract: A central goal of biomedicine is to understand, predict and ultimately control the dynamic mechanisms by which biological systems respond to perturbatio

safetyarxiv-cs-ai
6 Jun 2026
Tutorials

TRACE: A Temporal Conditional Estimation for Multimodal Time Series Foundation Models

DGX agent

arXiv:2606.06285v1 Announce Type: new Abstract: Time series foundation models (TS-FMs) aim to learn generalizable temporal representations that can be adapted to a wide range of downstream tasks. In r

tutorialsarxiv-cs-ai
6 Jun 2026
← Previous
1…196197198199200…448
Next →