AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Hardware

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks

DGX agent

arXiv:2605.28764v1 Announce Type: new Abstract: Vast quantities of compute (GPU cycles on personal workstations, idle inference servers, and edge devices between jobs) go unused because no incentive-a

hardwarearxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

SynthTools: A Framework for Scaling Synthetic Tools for Agent Development

DGX agent

arXiv:2511.09572v2 Announce Type: replace Abstract: For agentic systems to use external tools to solve complex, long-horizon tasks, we need a large set of diverse and controllable tool-use environment

agentsarxiv-cs-ai
28 May 2026
Model Releases

Tackling Multimodal Learning Challenges with Mixture-of-Expert: A Survey

DGX agent

arXiv:2605.27431v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) presents a naturally compatible and scalable framework for multimodal learning, demonstrating strong adaptability across dive

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

TCP-MCP: Landscape-Guided Co-Evolution of Prompts and Communication Topologies for Multi-Agent Systems

DGX agent

arXiv:2605.27850v1 Announce Type: new Abstract: Effective multi-agent systems cannot be designed by selecting prompts or communication graphs in isolation. Agent behavior depends on the information an

model-releasesarxiv-cs-ai
28 May 2026
Agents

Technical Report: Exploring the Emerging Threats of the Agent Skill Ecosystem

DGX agent

arXiv:2605.28588v1 Announce Type: cross Abstract: We analyzed 3,984 AI agent skills from major marketplaces and found 76 confirmed malicious payloads, including credential theft, backdoor installation

agentsarxiv-cs-ai
28 May 2026
Research

Tell Me a Story! Narrative-Driven XAI with Large Language Models

DGX agent

arXiv:2309.17057v3 Announce Type: replace Abstract: In many AI applications today, the predominance of black-box machine learning models, due to their typically higher accuracy, amplifies the need for

researcharxiv-cs-ai
28 May 2026
Safety

Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers

DGX agent

arXiv:2605.27686v1 Announce Type: cross Abstract: Transformers process images and videos by flattening space and time into long token sequences. While attention and KV caching preserve past features,

safetyarxiv-cs-ai
28 May 2026
Research

Text-Only Data Synthesis for Vision Language Model Training

DGX agent

arXiv:2503.22655v2 Announce Type: replace Abstract: Training vision-language models (VLMs) typically requires large-scale, high-quality image-text pairs, but collecting or synthesizing such data is co

researcharxiv-cs-ai
28 May 2026
Model Releases

The Alignment Floor: When Persona Customization Is Safe

DGX agent

arXiv:2605.27382v1 Announce Type: cross Abstract: A key promise of pluralistic AI is behavioral adaptation: persona prompts like 'be creative' or 'be thorough' let systems respect diverse user values

model-releasesarxiv-cs-ai
28 May 2026
Safety

The Attentional White Bear Effect in Transformer Language Models

DGX agent

arXiv:2605.28639v1 Announce Type: cross Abstract: Instruction-based suppression is widely used to prevent language models from generating prohibited content, yet it remains unclear whether suppression

safetyarxiv-cs-ai
28 May 2026
Model Releases

The Cases LJP Never Sees: Prosecution Decision Prediction for More Complete Criminal Liability Assessment

DGX agent

arXiv:2605.28464v1 Announce Type: cross Abstract: Legal Judgment Prediction (LJP) has become a core benchmark for evaluating AI in the criminal legal domain, but it only sees criminal cases that have

model-releasesarxiv-cs-ai
28 May 2026
Research

The Computational Boundary of Inference: Capability Internalization, Training, and the Turing Jump

DGX agent

arXiv:2605.27381v1 Announce Type: cross Abstract: Claims about recursive self-improvement in AI often slide from repeated internal revision to the possibility of qualitatively stronger capability with

researcharxiv-cs-ai
28 May 2026
Research

The Decision to Verify: How Warmth and User Characteristics Shape Reliance on Conversational Agents for Information Search

DGX agent

arXiv:2605.28498v1 Announce Type: cross Abstract: Conversational artificial intelligence (AI) provides an efficient and convenient gateway to information access. However, it can cause overreliance whe

researcharxiv-cs-ai
28 May 2026
Local Ai

The Energy Blind Spot: NVIDIA's Flagship Edge AI Hardware Cannot Support Process-Level Energy Attribution

DGX agent

arXiv:2605.27599v1 Announce Type: cross Abstract: Agentic AI workloads - where a single user goal triggers multi-step orchestration, tool calls, retries, and failure recovery - are being targeted for

local-aiarxiv-cs-ai
28 May 2026
Safety

The Ethics of LLM Sandbox and Persona Dynamics

DGX agent

arXiv:2605.28647v1 Announce Type: new Abstract: It is well known that LLM guardrails and trained persona dynamics can produce a reality gap: the distance between the world a LLM is permitted or shaped

safetyarxiv-cs-ai
28 May 2026
Model Releases

The Fragility of Chain-of-Thought Monitoring Across Typologically Diverse Languages

DGX agent

arXiv:2605.27901v1 Announce Type: cross Abstract: Chain-of-thought (CoT) monitoring has been proposed as a promising safety mechanism for detecting misaligned behavior in large language models. Howeve

model-releasesarxiv-cs-ai
28 May 2026
Research

The Future of Facts: Tracing the Factual Generation-Verification Gap

DGX agent

arXiv:2605.27564v1 Announce Type: cross Abstract: Language models are becoming the default interface to factual knowledge, yet they often verify outputs more reliably than they generate them. This gen

researcharxiv-cs-ai
28 May 2026
Research

The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models

DGX agent

arXiv:2601.19926v2 Announce Type: replace-cross Abstract: We present a systematic review of 337 articles evaluating the syntactic abilities of Transformer-based language models (TLMs), reporting on ov

researcharxiv-cs-ai
28 May 2026
Safety

The Illusion of Opting in AI-Mediated Consequential Decisions

DGX agent

arXiv:2605.28210v1 Announce Type: new Abstract: Drawing on Ullmann-Margalit's concept of opting (transformative, irrevocable, and shadowed by foreclosed alternatives), we show that current AI systems

safetyarxiv-cs-ai
28 May 2026
Model Releases

The Importance of Being Statistically Earnest: A Critical Re-evaluation of GSM-Symbolic

DGX agent

arXiv:2605.28700v1 Announce Type: new Abstract: The GSM-Symbolic benchmark (Mirzadeh et al., 2025) reported consistent performance drops across 25 Large Language Models (LLMs) when tested on template-

model-releasesarxiv-cs-ai
28 May 2026
Safety

The Obfuscation Atlas: Mapping Where Honesty Emerges in RLVR with Deception Probes

DGX agent

arXiv:2602.15515v2 Announce Type: replace-cross Abstract: Training against white-box deception detectors has been proposed as a way to make AI systems honest. However, such training risks models learn

safetyarxiv-cs-ai
28 May 2026
Agents

The Optimal Sample Complexity of Linear Contracts

DGX agent

arXiv:2601.01496v2 Announce Type: replace-cross Abstract: In this paper, we settle the problem of learning optimal linear contracts from data in the offline setting, where agent types are drawn from a

agentsarxiv-cs-ai
28 May 2026
Model Releases

The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Language Models? A Bias-Controlled Study

DGX agent

arXiv:2504.04540v2 Announce Type: replace-cross Abstract: 3D Large Language Models (LLMs) leveraging spatial information in point clouds for 3D spatial reasoning attract great attention. Despite some

model-releasesarxiv-cs-ai
28 May 2026
Tutorials

The Principles of Diffusion Models

DGX agent

arXiv:2510.21890v2 Announce Type: replace-cross Abstract: This book presents the core principles that have guided the development of diffusion models, tracing their origins and showing how diverse for

tutorialsarxiv-cs-ai
28 May 2026
Model Releases

The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation

DGX agent

arXiv:2601.17737v3 Announce Type: replace-cross Abstract: Recent advances in video generation have produced models capable of synthesizing stunning visual content from simple text prompts. However, th

model-releasesarxiv-cs-ai
28 May 2026
Local Ai

The Shape of Overthinking: Backtracking Bursts in Long Reasoning Traces

DGX agent

arXiv:2605.27965v1 Announce Type: new Abstract: Reasoning models often generate long traces in which useful self-correction and unproductive revision are hard to distinguish. We study this distinction

local-aiarxiv-cs-ai
28 May 2026
Research

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models

DGX agent

arXiv:2510.20665v3 Announce Type: replace Abstract: Evaluating the quality of reasoning traces from large language models remains understudied, labor-intensive, and unreliable: current practice relies

researcharxiv-cs-ai
28 May 2026
Research

The Well-Tempered Classifier: Some Elementary Properties of Temperature Scaling

DGX agent

arXiv:2602.14862v2 Announce Type: replace-cross Abstract: Temperature scaling is a simple method that allows to control the uncertainty of probabilistic models. It is mostly used in two contexts: impr

researcharxiv-cs-ai
28 May 2026
Model Releases

Thermodynamic properties of chemically disordered compounds via AI-driven estimation of partition function with the PULSE method

DGX agent

arXiv:2605.28594v1 Announce Type: cross Abstract: In this article, we present an improved version of the PULSE method (Partition function Unsupervised Learning Sampling and Evaluation) for estimating

model-releasesarxiv-cs-ai
28 May 2026
Research

Thinking as Compression: Your Reasoning Model is Secretly a Context Compressor

DGX agent

arXiv:2605.28713v1 Announce Type: new Abstract: Context compression aims to shorten long context inputs with minimal information loss for LLM inference acceleration. While existing methods have shown

researcharxiv-cs-ai
28 May 2026
Research

Token Optimization Strategies for LLM-Based Oracle-to-PostgreSQL Migration

DGX agent

arXiv:2605.28557v1 Announce Type: cross Abstract: LLMs are increasingly used for software modernization, code translation, and database migration. However, LLM-based Oracle2PostgreSQL migration remain

researcharxiv-cs-ai
28 May 2026
Model Releases

Tool Forge: A Validation-Carrying Toolchain for Governed Agentic Execution

DGX agent

arXiv:2605.28000v1 Announce Type: cross Abstract: Large language model agents are increasingly expected to perform operational work: calling APIs, manipulating files, assembling workflows, and acting

model-releasesarxiv-cs-ai
28 May 2026
Safety

Towards automated data analysis: A guided framework for LLM-based risk estimation

DGX agent

arXiv:2603.04631v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly integrated into critical decision-making pipelines, a trend that raises the demand for robust and auto

safetyarxiv-cs-ai
28 May 2026
Model Releases

Towards Faithful Agentic XAI: A Verification Method and an Open-World Benchmark for Better Model Faithfulness

DGX agent

arXiv:2605.27879v1 Announce Type: new Abstract: Explainable AI (XAI) helps users interpret model behavior and identify potential faults. Agentic XAI systems use Large Language Models (LLMs) to make ex

model-releasesarxiv-cs-ai
28 May 2026
Research

Towards Reliable Multilingual LLMs-as-a-Judge: An Empirical Study

DGX agent

arXiv:2605.28710v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for the automatic evaluation of generated text, yet most prior work focuses on English. Despite the

researcharxiv-cs-ai
28 May 2026
Model Releases

TRACER: Turn-level Regret Matching with Inner Reinforcement Credit for Cooperative Multi-LLM Reasoning

DGX agent

arXiv:2605.28699v1 Announce Type: new Abstract: Large language models increasingly rely on either reinforcement learning or multi-agent prompting to improve reasoning, yet these two paradigms remain d

model-releasesarxiv-cs-ai
28 May 2026
Safety

Training Stratigraphy: Persistent Behavioral Artifacts in Large Language Models Observed Through Longitudinal AI-Human Interaction

DGX agent

arXiv:2605.28102v1 Announce Type: new Abstract: Large language models trained with Reinforcement Learning from Human Feedback (RLHF) and Constitutional AI exhibit persistent behavioral patterns that s

safetyarxiv-cs-ai
28 May 2026
Safety

Transferable Reinforcement Learning via Probabilistic Latent Embeddings and Dynamic Policy Adaptation for Sim-to-Real Deployment

DGX agent

arXiv:2605.27659v1 Announce Type: cross Abstract: Due to limited resources and public safety concerns, deep reinforcement learning (RL) agents for many cyber-physical systems (e.g., autonomous vehicle

safetyarxiv-cs-ai
28 May 2026
Research

Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns

DGX agent

arXiv:2605.28566v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities, yet their standard generation process -- auto-regressive token predict

researcharxiv-cs-ai
28 May 2026
Model Releases

Trinity: Unifying Class-Agnostic Terrain and Semantic Segmentation for Unstructured Outdoor Environments by Leveraging Synthetic Data

DGX agent

arXiv:2605.27644v1 Announce Type: cross Abstract: Terrain understanding is fundamental for mobile robots operating in unstructured outdoor environments. Existing vision-based traversability estimation

model-releasesarxiv-cs-ai
28 May 2026
Safety

Turning Video Models into Generalist Robot Policies

DGX agent

arXiv:2605.27817v1 Announce Type: cross Abstract: Video generative models have emerged as a promising robotics backbone, capable of generating videos that depict the completion of complex tasks across

safetyarxiv-cs-ai
28 May 2026
Agents

Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empirical Study

DGX agent

arXiv:2506.08311v2 Announce Type: replace-cross Abstract: Automated Program Repair (APR) agents leverage Large Language Models (LLMs) to autonomously diagnose and fix software bugs through reasoning,

agentsarxiv-cs-ai
28 May 2026
Model Releases

Unified Synthesis of Compositional Speech and Sound from Free-Form Text Prompts

DGX agent

arXiv:2605.28063v1 Announce Type: cross Abstract: Audio generation has made significant progress, yet synthesizing unified audio where speech and sounds are naturally composited remains a challenge. C

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

UniMaia: Steering Chess Policies with Language for Human-like Play

DGX agent

arXiv:2605.27767v1 Announce Type: cross Abstract: Recent advances in large language models have enabled natural language to serve as a flexible interface for controlling complex systems, but often at

model-releasesarxiv-cs-ai
28 May 2026
Safety

Unlocking Fine-Grained and Within-Utterance Speaking Style Control in Prompt-Based Text-to-Speech Models

DGX agent

arXiv:2605.27376v1 Announce Type: cross Abstract: While prompt-based text-to-speech (TTS) models enable natural language-driven speaking style control, they often provide limited fine-grained control

safetyarxiv-cs-ai
28 May 2026
Agents

UserHarness: Harnessing User Minds for Stronger Agent Theory-of-Mind

DGX agent

arXiv:2605.27721v1 Announce Type: cross Abstract: Understanding what a user believes and intends is central to building effective agent assistants. This ability is often evaluated through Theory-of-Mi

agentsarxiv-cs-ai
28 May 2026
Model Releases

Using Zero-Shot LLM-Generated Survey Data for Geographically Explicit Population Synthesis

DGX agent

arXiv:2605.27401v1 Announce Type: cross Abstract: There is a growing interest in utilizing synthetic populations for a diverse range of applications. At the same time, we are witnessing a tremendous g

model-releasesarxiv-cs-ai
28 May 2026
Safety

Utility-Aware Multimodal Contrastive Learning for Product Image Generation

DGX agent

arXiv:2605.28733v1 Announce Type: new Abstract: Product images strongly influence consumer decision-making in online marketplaces. Empowered by multimodal contrastive learning, generative AI can outpu

safetyarxiv-cs-ai
28 May 2026
← Previous
1…253254255256257…448
Next →