AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
28 May 2026

Speaking of Language: Reflections on Metalanguage Research in NLP

ResearchDGX agent

arXiv:2604.02645v2 Announce Type: replace-cross Abstract: This work aims to shine a spotlight on the topic of metalanguage. We first define metalanguage, link it to NLP and LLMs, and then discuss our

SSR3D-LLM: Structured Spatial Reasoning via Latent Steps for Fine-Grained Grounding in Unified 3D-LLMs

Model ReleasesDGX agent

arXiv:2605.28490v1 Announce Type: cross Abstract: 3D object grounding localizes referred objects in a 3D scene from natural language. Unified instance-centric 3D-LLMs aim to solve grounding together w

STAB: Specification-driven Testing for Algorithmic Bottlenecks

ResearchDGX agent

arXiv:2605.27981v1 Announce Type: new Abstract: Evaluating the efficiency of algorithmic code requires test cases that expose runtime bottlenecks. Previous methods generate efficiency test cases eithe


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

STARS: Spike Tail-Aware Relational Synthesis for ANN-to-SNN Data-Free Knowledge Distillation

SafetyDGX agent

arXiv:2605.27409v1 Announce Type: cross Abstract: SNNs promise energy-efficient and low-latency inference, but their performance still trails that of ANNs. ANN-to-SNN knowledge distillation helps narr

STFlow: Data-Coupled Flow Matching for Geometric Trajectory Simulation

ResearchDGX agent

arXiv:2505.18647v3 Announce Type: replace-cross Abstract: Simulating trajectories of dynamical systems is a fundamental problem in a wide range of fields such as molecular dynamics, biochemistry, and

Stochastic Gradient Descent with Momentum is Algorithmically Stable

Model ReleasesDGX agent

arXiv:2605.28517v1 Announce Type: cross Abstract: Stochastic gradient descent with momentum (SGDM) is one of the most widely used optimization algorithms in machine learning. While optimization proper

StoryLens: Preference-Aligned Story Rewriting via Context-Aware Narrative Enrichment

Model ReleasesDGX agent

arXiv:2605.28073v1 Announce Type: cross Abstract: Story rewriting aims to adapt existing narratives to diverse reader preferences while preserving plot consistency and narrative coherence. Unlike conv

StoryMI: Steerable Multi-Agent Therapeutic Dialogue Generation

Model ReleasesDGX agent

arXiv:2605.27393v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent dialogue, but prior works lack situational grounding, dynamic strategy control, and evaluation aligne

Structured Agent Distillation for Large Language Model

SafetyDGX agent

arXiv:2505.13820v5 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong capabilities as decision-making agents by interleaving reasoning and actions, as seen in ReAct-sty

Structured Belief State and the First Precision-Aware Benchmark for LLM Memory Retrieval

Model ReleasesDGX agent

arXiv:2605.11325v2 Announce Type: replace-cross Abstract: Every major benchmark for LLM memory systems, LoCoMo foremost, measures whether a model answered correctly, not whether the memory system retr

SuiChat-CN: Benchmarking Contextual Suicide Risk Assessment in Chinese Group Chats

Model ReleasesDGX agent

arXiv:2605.27911v1 Announce Type: new Abstract: Suicide is a critical global public health challenge, causing approximately 720,000 deaths each year and calling for timely, effective prevention strate

Supervised Distributional Reduction via Optimal Transport and Dependence Maximization

SafetyDGX agent

arXiv:2605.27619v1 Announce Type: cross Abstract: Learning representations that capture both intrinsic data geometry and target-relevant structure remains a fundamental challenge, particularly in sett

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks

HardwareDGX agent

arXiv:2605.28764v1 Announce Type: new Abstract: Vast quantities of compute (GPU cycles on personal workstations, idle inference servers, and edge devices between jobs) go unused because no incentive-a

SynthTools: A Framework for Scaling Synthetic Tools for Agent Development

AgentsDGX agent

arXiv:2511.09572v2 Announce Type: replace Abstract: For agentic systems to use external tools to solve complex, long-horizon tasks, we need a large set of diverse and controllable tool-use environment

Tackling Multimodal Learning Challenges with Mixture-of-Expert: A Survey

Model ReleasesDGX agent

arXiv:2605.27431v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) presents a naturally compatible and scalable framework for multimodal learning, demonstrating strong adaptability across dive

TCP-MCP: Landscape-Guided Co-Evolution of Prompts and Communication Topologies for Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.27850v1 Announce Type: new Abstract: Effective multi-agent systems cannot be designed by selecting prompts or communication graphs in isolation. Agent behavior depends on the information an

Technical Report: Exploring the Emerging Threats of the Agent Skill Ecosystem

AgentsDGX agent

arXiv:2605.28588v1 Announce Type: cross Abstract: We analyzed 3,984 AI agent skills from major marketplaces and found 76 confirmed malicious payloads, including credential theft, backdoor installation

Tell Me a Story! Narrative-Driven XAI with Large Language Models

ResearchDGX agent

arXiv:2309.17057v3 Announce Type: replace Abstract: In many AI applications today, the predominance of black-box machine learning models, due to their typically higher accuracy, amplifies the need for

Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers

SafetyDGX agent

arXiv:2605.27686v1 Announce Type: cross Abstract: Transformers process images and videos by flattening space and time into long token sequences. While attention and KV caching preserve past features,

Text-Only Data Synthesis for Vision Language Model Training

ResearchDGX agent

arXiv:2503.22655v2 Announce Type: replace Abstract: Training vision-language models (VLMs) typically requires large-scale, high-quality image-text pairs, but collecting or synthesizing such data is co

The Alignment Floor: When Persona Customization Is Safe

Model ReleasesDGX agent

arXiv:2605.27382v1 Announce Type: cross Abstract: A key promise of pluralistic AI is behavioral adaptation: persona prompts like 'be creative' or 'be thorough' let systems respect diverse user values

The Attentional White Bear Effect in Transformer Language Models

SafetyDGX agent

arXiv:2605.28639v1 Announce Type: cross Abstract: Instruction-based suppression is widely used to prevent language models from generating prohibited content, yet it remains unclear whether suppression

The Cases LJP Never Sees: Prosecution Decision Prediction for More Complete Criminal Liability Assessment

Model ReleasesDGX agent

arXiv:2605.28464v1 Announce Type: cross Abstract: Legal Judgment Prediction (LJP) has become a core benchmark for evaluating AI in the criminal legal domain, but it only sees criminal cases that have

The Computational Boundary of Inference: Capability Internalization, Training, and the Turing Jump

ResearchDGX agent

arXiv:2605.27381v1 Announce Type: cross Abstract: Claims about recursive self-improvement in AI often slide from repeated internal revision to the possibility of qualitatively stronger capability with

The Decision to Verify: How Warmth and User Characteristics Shape Reliance on Conversational Agents for Information Search

ResearchDGX agent

arXiv:2605.28498v1 Announce Type: cross Abstract: Conversational artificial intelligence (AI) provides an efficient and convenient gateway to information access. However, it can cause overreliance whe

The Energy Blind Spot: NVIDIA's Flagship Edge AI Hardware Cannot Support Process-Level Energy Attribution

Local AiDGX agent

arXiv:2605.27599v1 Announce Type: cross Abstract: Agentic AI workloads - where a single user goal triggers multi-step orchestration, tool calls, retries, and failure recovery - are being targeted for

The Ethics of LLM Sandbox and Persona Dynamics

SafetyDGX agent

arXiv:2605.28647v1 Announce Type: new Abstract: It is well known that LLM guardrails and trained persona dynamics can produce a reality gap: the distance between the world a LLM is permitted or shaped

The Fragility of Chain-of-Thought Monitoring Across Typologically Diverse Languages

Model ReleasesDGX agent

arXiv:2605.27901v1 Announce Type: cross Abstract: Chain-of-thought (CoT) monitoring has been proposed as a promising safety mechanism for detecting misaligned behavior in large language models. Howeve

The Future of Facts: Tracing the Factual Generation-Verification Gap

ResearchDGX agent

arXiv:2605.27564v1 Announce Type: cross Abstract: Language models are becoming the default interface to factual knowledge, yet they often verify outputs more reliably than they generate them. This gen

The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models

ResearchDGX agent

arXiv:2601.19926v2 Announce Type: replace-cross Abstract: We present a systematic review of 337 articles evaluating the syntactic abilities of Transformer-based language models (TLMs), reporting on ov

The Illusion of Opting in AI-Mediated Consequential Decisions

SafetyDGX agent

arXiv:2605.28210v1 Announce Type: new Abstract: Drawing on Ullmann-Margalit's concept of opting (transformative, irrevocable, and shadowed by foreclosed alternatives), we show that current AI systems

The Importance of Being Statistically Earnest: A Critical Re-evaluation of GSM-Symbolic

Model ReleasesDGX agent

arXiv:2605.28700v1 Announce Type: new Abstract: The GSM-Symbolic benchmark (Mirzadeh et al., 2025) reported consistent performance drops across 25 Large Language Models (LLMs) when tested on template-

The Obfuscation Atlas: Mapping Where Honesty Emerges in RLVR with Deception Probes

SafetyDGX agent

arXiv:2602.15515v2 Announce Type: replace-cross Abstract: Training against white-box deception detectors has been proposed as a way to make AI systems honest. However, such training risks models learn

The Optimal Sample Complexity of Linear Contracts

AgentsDGX agent

arXiv:2601.01496v2 Announce Type: replace-cross Abstract: In this paper, we settle the problem of learning optimal linear contracts from data in the offline setting, where agent types are drawn from a

The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Language Models? A Bias-Controlled Study

Model ReleasesDGX agent

arXiv:2504.04540v2 Announce Type: replace-cross Abstract: 3D Large Language Models (LLMs) leveraging spatial information in point clouds for 3D spatial reasoning attract great attention. Despite some

The Principles of Diffusion Models

TutorialsDGX agent

arXiv:2510.21890v2 Announce Type: replace-cross Abstract: This book presents the core principles that have guided the development of diffusion models, tracing their origins and showing how diverse for

The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation

Model ReleasesDGX agent

arXiv:2601.17737v3 Announce Type: replace-cross Abstract: Recent advances in video generation have produced models capable of synthesizing stunning visual content from simple text prompts. However, th

The Shape of Overthinking: Backtracking Bursts in Long Reasoning Traces

Local AiDGX agent

arXiv:2605.27965v1 Announce Type: new Abstract: Reasoning models often generate long traces in which useful self-correction and unproductive revision are hard to distinguish. We study this distinction

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models

ResearchDGX agent

arXiv:2510.20665v3 Announce Type: replace Abstract: Evaluating the quality of reasoning traces from large language models remains understudied, labor-intensive, and unreliable: current practice relies

The Well-Tempered Classifier: Some Elementary Properties of Temperature Scaling

ResearchDGX agent

arXiv:2602.14862v2 Announce Type: replace-cross Abstract: Temperature scaling is a simple method that allows to control the uncertainty of probabilistic models. It is mostly used in two contexts: impr

Thermodynamic properties of chemically disordered compounds via AI-driven estimation of partition function with the PULSE method

Model ReleasesDGX agent

arXiv:2605.28594v1 Announce Type: cross Abstract: In this article, we present an improved version of the PULSE method (Partition function Unsupervised Learning Sampling and Evaluation) for estimating

Thinking as Compression: Your Reasoning Model is Secretly a Context Compressor

ResearchDGX agent

arXiv:2605.28713v1 Announce Type: new Abstract: Context compression aims to shorten long context inputs with minimal information loss for LLM inference acceleration. While existing methods have shown

Token Optimization Strategies for LLM-Based Oracle-to-PostgreSQL Migration

ResearchDGX agent

arXiv:2605.28557v1 Announce Type: cross Abstract: LLMs are increasingly used for software modernization, code translation, and database migration. However, LLM-based Oracle2PostgreSQL migration remain

Tool Forge: A Validation-Carrying Toolchain for Governed Agentic Execution

Model ReleasesDGX agent

arXiv:2605.28000v1 Announce Type: cross Abstract: Large language model agents are increasingly expected to perform operational work: calling APIs, manipulating files, assembling workflows, and acting

Towards automated data analysis: A guided framework for LLM-based risk estimation

SafetyDGX agent

arXiv:2603.04631v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly integrated into critical decision-making pipelines, a trend that raises the demand for robust and auto

Towards Faithful Agentic XAI: A Verification Method and an Open-World Benchmark for Better Model Faithfulness

Model ReleasesDGX agent

arXiv:2605.27879v1 Announce Type: new Abstract: Explainable AI (XAI) helps users interpret model behavior and identify potential faults. Agentic XAI systems use Large Language Models (LLMs) to make ex

Towards Reliable Multilingual LLMs-as-a-Judge: An Empirical Study

ResearchDGX agent

arXiv:2605.28710v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for the automatic evaluation of generated text, yet most prior work focuses on English. Despite the

TRACER: Turn-level Regret Matching with Inner Reinforcement Credit for Cooperative Multi-LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.28699v1 Announce Type: new Abstract: Large language models increasingly rely on either reinforcement learning or multi-agent prompting to improve reasoning, yet these two paradigms remain d

Training Stratigraphy: Persistent Behavioral Artifacts in Large Language Models Observed Through Longitudinal AI-Human Interaction

SafetyDGX agent

arXiv:2605.28102v1 Announce Type: new Abstract: Large language models trained with Reinforcement Learning from Human Feedback (RLHF) and Constitutional AI exhibit persistent behavioral patterns that s

Transferable Reinforcement Learning via Probabilistic Latent Embeddings and Dynamic Policy Adaptation for Sim-to-Real Deployment

SafetyDGX agent

arXiv:2605.27659v1 Announce Type: cross Abstract: Due to limited resources and public safety concerns, deep reinforcement learning (RL) agents for many cyber-physical systems (e.g., autonomous vehicle

Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns

ResearchDGX agent

arXiv:2605.28566v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities, yet their standard generation process -- auto-regressive token predict

Trinity: Unifying Class-Agnostic Terrain and Semantic Segmentation for Unstructured Outdoor Environments by Leveraging Synthetic Data

Model ReleasesDGX agent

arXiv:2605.27644v1 Announce Type: cross Abstract: Terrain understanding is fundamental for mobile robots operating in unstructured outdoor environments. Existing vision-based traversability estimation

Turning Video Models into Generalist Robot Policies

SafetyDGX agent

arXiv:2605.27817v1 Announce Type: cross Abstract: Video generative models have emerged as a promising robotics backbone, capable of generating videos that depict the completion of complex tasks across

Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empirical Study

AgentsDGX agent

arXiv:2506.08311v2 Announce Type: replace-cross Abstract: Automated Program Repair (APR) agents leverage Large Language Models (LLMs) to autonomously diagnose and fix software bugs through reasoning,

Unified Synthesis of Compositional Speech and Sound from Free-Form Text Prompts

Model ReleasesDGX agent

arXiv:2605.28063v1 Announce Type: cross Abstract: Audio generation has made significant progress, yet synthesizing unified audio where speech and sounds are naturally composited remains a challenge. C

UniMaia: Steering Chess Policies with Language for Human-like Play

Model ReleasesDGX agent

arXiv:2605.27767v1 Announce Type: cross Abstract: Recent advances in large language models have enabled natural language to serve as a flexible interface for controlling complex systems, but often at

Unlocking Fine-Grained and Within-Utterance Speaking Style Control in Prompt-Based Text-to-Speech Models

SafetyDGX agent

arXiv:2605.27376v1 Announce Type: cross Abstract: While prompt-based text-to-speech (TTS) models enable natural language-driven speaking style control, they often provide limited fine-grained control

UserHarness: Harnessing User Minds for Stronger Agent Theory-of-Mind

AgentsDGX agent

arXiv:2605.27721v1 Announce Type: cross Abstract: Understanding what a user believes and intends is central to building effective agent assistants. This ability is often evaluated through Theory-of-Mi

Using Zero-Shot LLM-Generated Survey Data for Geographically Explicit Population Synthesis

Model ReleasesDGX agent

arXiv:2605.27401v1 Announce Type: cross Abstract: There is a growing interest in utilizing synthetic populations for a diverse range of applications. At the same time, we are witnessing a tremendous g

Utility-Aware Multimodal Contrastive Learning for Product Image Generation

SafetyDGX agent

arXiv:2605.28733v1 Announce Type: new Abstract: Product images strongly influence consumer decision-making in online marketplaces. Empowered by multimodal contrastive learning, generative AI can outpu

← Previous
1…202203204205206…358
Next →