AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
4 Jun 2026

The Saturation Trap and the Subjectivity of Intervention Timing: Why Affect-Based Triggers and LLM Judges Fail to Time Interventions on Autonomous Agents

Model ReleasesDGX agent

arXiv:2606.04296v1 Announce Type: new Abstract: As autonomous AI agents move from conversational systems to long-horizon software execution, runtime safety layers that decide when to interrupt an agen

The Variance Brain Foundation Models Forgot: Third-Order Statistics Predict Cognition Where Billion-Parameter Models Fail

Model ReleasesDGX agent

arXiv:2606.04010v1 Announce Type: cross Abstract: Brain foundation models (BFMs) are self-supervised Transformers pretrained on fMRI data. We posit that these models should capture each subject's cogn

Thinking Through Signs: PEEL as a Semiotic Scaffolding for Epistemically Accountable AI-Enabled Research


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2606.04152v1 Announce Type: new Abstract: Large language models are reshaping research practice while quietly eroding researchers epistemic accountability. This commentary introduces PEEL - Prot

TIDE: Proactive Multi-Problem Discovery via Template-Guided Iteration

AgentsDGX agent

arXiv:2606.04743v1 Announce Type: cross Abstract: Agents are widely deployed as assistants over documents, tools, and code. However, they typically act only on explicit user requests, which surface on

TITAN-FedAnil+: Trust-Based Adaptive Blockchain Federated Learning for Resource-Constrained Intelligent Enterprises

HardwareDGX agent

arXiv:2606.04388v1 Announce Type: cross Abstract: Federated Learning (FL) has emerged as an effective paradigm for collaborative intelligence while preserving data privacy. However, data heterogeneity

Token Rankings are Unforgeable Language Model Signatures

ResearchDGX agent

arXiv:2606.04459v1 Announce Type: cross Abstract: Language model parameters are known to impose unique (to each model) geometric constraints on their logit outputs, which serves as a signature that id

Tomography by Design: An Algebraic Approach to Low-Rank Quantum States

ResearchDGX agent

arXiv:2602.15202v2 Announce Type: replace-cross Abstract: We present an algebraic algorithm for quantum state tomography that leverages measurements of certain observables to estimate structured entri

Topology Matters: Measuring Memory Leakage in Multi-Agent LLMs

AgentsDGX agent

arXiv:2512.04668v4 Announce Type: replace-cross Abstract: Graph topology is a fundamental determinant of memory leakage in multi-agent LLM systems, yet its effects remain poorly quantified. We introdu

Toward Autonomous O-RAN: A Multi-Scale Agentic AI Framework for Real-Time Network Control and Management

AgentsDGX agent

arXiv:2602.14117v2 Announce Type: replace-cross Abstract: Open Radio Access Networks (O-RAN) promise flexible 6G network access through disaggregated, software-driven components and open interfaces, b

Toward Pre-Deployment Assurance for Enterprise AI Agents: Ontology-Grounded Simulation and Trust Certification

Model ReleasesDGX agent

arXiv:2606.04037v1 Announce Type: new Abstract: Pre-deployment verification of enterprise artificial intelligence (AI) agents remains a critical gap between large language model (LLM) capability bench

Towards Efficient and Evidence-grounded Mobility Prediction with LLM-Driven Agent

Model ReleasesDGX agent

arXiv:2606.05130v1 Announce Type: cross Abstract: Individual-level mobility prediction is central to urban simulation, transportation planning, and policy analysis. Supervised sequence models achieve

TPA-AD: A Two-Stage Pseudo Anomaly-Guided Method for Bearing Time-Series Anomaly Detection

ResearchDGX agent

arXiv:2606.04073v1 Announce Type: cross Abstract: This paper proposes a two-stage pseudo anomaly-guided anomaly detection method (extbf{T}wo-stage extbf{P}seudo extbf{A}nomaly-guided extbf{A}nomaly ex

Trace-Mediated Peak Bias: Bridging Temporal Credit Assignment and Cognitive Heuristics in Deep Reinforcement Learning

SafetyDGX agent

arXiv:2606.04735v1 Announce Type: cross Abstract: Temporal credit assignment is central to both biological and artificial intelligence, yet its interaction with non-linear function approximation is po

Treat Traffic Like Trees: A Semantic-Preserving Hierarchical Graph-Based Expert Framework for Encrypted Traffic Analysis

Model ReleasesDGX agent

arXiv:2606.04517v1 Announce Type: cross Abstract: Graph-based deep learning methods have been widely employed in encrypted traffic analysis to exploit latent correlations across different granularitie

Tree-Based Formalization of Multi-Agent Complementarity in Human-AI Interactions

Model ReleasesDGX agent

arXiv:2606.04779v1 Announce Type: new Abstract: Complementarity is the case in which a human--AI interaction (HAI) outperforms the best prediction benchmark available among its members. Although this

Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers

AgentsDGX agent

arXiv:2606.04421v1 Announce Type: new Abstract: Many current agentic systems and LLM pipelines correct mistakes by optimizing outcome reward. This addresses only the what of failure: when an outcome d

Tuning the Implicit Regularizer of Masked Diffusion Language Models: Enhancing Generalization via Insights from k-Parity

Model ReleasesDGX agent

arXiv:2601.22450v2 Announce Type: replace-cross Abstract: Masked Diffusion Language Models have recently emerged as a powerful generative paradigm, yet their generalization properties remain understud

Uncertainty-Aware End-to-End Co-Design of Neural Network Processors: From Training and Mapping to Fabrication

ApplicationsDGX agent

arXiv:2606.04850v1 Announce Type: cross Abstract: Designing a neural network processor is an end-to-end co-design problem: network architecture and training budget determine the inference workload; ha

Uncertainty Estimation using Variance-Gated Distributions

ResearchDGX agent

arXiv:2509.08846v2 Announce Type: replace-cross Abstract: Evaluation of per-sample uncertainty quantification from neural networks is essential for decision-making involving high-risk applications. A

UniCAD: A Unified Benchmark and Universal Model for Multi-Modal Multi-Task CAD

Model ReleasesDGX agent

arXiv:2606.05058v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) underpins modern engineering and manufacturing by enabling the creation of precise, editable 3D models. However, CAD resea

Unifying Model-Free Efficiency and Model-Based Representations via Latent Dynamics

Model ReleasesDGX agent

arXiv:2602.12643v2 Announce Type: replace-cross Abstract: We present Unified Latent Dynamics (ULD), a novel reinforcement learning algorithm that unifies the efficiency of model-free methods with the

Unlocking Feature Learning in Gated Delta Networks at Scale

ResearchDGX agent

arXiv:2606.04048v1 Announce Type: cross Abstract: Training and scaling Large Language Models demand enormous computational resources, motivating both efficient sub-quadratic architectures and principl

Unlocking Proactivity in Task-Oriented Dialogue

SafetyDGX agent

arXiv:2605.22240v2 Announce Type: replace Abstract: Proactive task-oriented dialogue (TOD), such as outbound sales, demands a persuasive agent that actively probes the user's concerns and steers the c

Unpredictable Safety: Domain-Dependent Compliance and the Transparency Gap in Open-Weight LLMs

Model ReleasesDGX agent

arXiv:2606.04035v1 Announce Type: cross Abstract: We present a systematic study of domain-dependent safety behavior in open-weight LLMs: 7 standardized experiments across 7 ethical domains, testing 5

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models

SafetyDGX agent

arXiv:2602.19101v2 Announce Type: replace-cross Abstract: Value alignment of Large Language Models (LLMs) requires us to empirically measure these models' actual, acquired representation of value. Amo

VAMPS: Visual-Assisted Mathematical Problem Solving Benchmark

Model ReleasesDGX agent

arXiv:2606.04244v1 Announce Type: new Abstract: Multimodal large language models are increasingly capable of complex reasoning, yet their performance often degrades when they must externalize a proble

Vectorized Online POMDP Planning

AgentsDGX agent

arXiv:2510.27191v5 Announce Type: replace-cross Abstract: Planning under partial observability is an essential capability of autonomous robots. The Partially Observable Markov Decision Process (POMDP)

VGGSounder: Audio-Visual Evaluations for Foundation Models

Model ReleasesDGX agent

arXiv:2508.08237v4 Announce Type: replace-cross Abstract: The emergence of audio-visual foundation models underscores the importance of reliably assessing their multi-modal understanding. The VGGSound

VISTA: Vision-Grounded and Physics-Validated Adaptation of UMI data for VLA Training

Local AiDGX agent

arXiv:2606.04708v1 Announce Type: cross Abstract: Universal Manipulation Interface (UMI) enables scalable real-world robot data collection without hardware-specific teleoperation, yet leveraging UMI d

What If Prompt Injection Never Left? Exploring Cross-Session Stored Prompt Injection in Agentic Systems

Model ReleasesDGX agent

arXiv:2606.04425v1 Announce Type: cross Abstract: Modern agentic systems transform LLMs from session-bounded assistants into stateful systems that persist and evolve shared world state across sessions

What Type of Inference is Active Inference?

SafetyDGX agent

arXiv:2606.04935v1 Announce Type: new Abstract: Active inference casts decision-making as inference, with the Expected Free Energy (EFE) unifying goal-directed and information-seeking behavior. Recent

Who Needs Labels? Adapting Vision Foundation Models With the Metadata You Already Have

ResearchDGX agent

arXiv:2606.05107v1 Announce Type: cross Abstract: We propose a label-free approach to adapt powerful but generic vision foundation models to specialized scientific domains. Standard supervised fine-tu

Why Muon Outperforms Adam: A Curvature Perspective

Local AiDGX agent

arXiv:2606.04662v1 Announce Type: cross Abstract: Muon improves training efficiency over Adam in large language-model training by about two times, but the local geometric source of this advantage rema

You Only Train Once: Differentiable Subset Selection for Omics Data

ResearchDGX agent

arXiv:2512.17678v2 Announce Type: replace-cross Abstract: Selecting compact and informative gene subsets from single-cell transcriptomic data is essential for biomarker discovery, improving interpreta

'Your AI Text is not Mine': Redefining and Evaluating AI-generated Text Detection under Realistic Assumptions

Model ReleasesDGX agent

arXiv:2606.04906v1 Announce Type: cross Abstract: Although it is generally agreed that AI-generated text poses a broad societal risk, there is no common understanding in the AI-generated text detectio

ZeroWBC: Learning Natural Whole-Body Humanoid Interaction from Human Egocentric Data

SafetyDGX agent

arXiv:2603.09170v2 Announce Type: replace-cross Abstract: Achieving versatile and natural whole-body humanoid interaction control remains challenging due to the high cost of whole-body teleoperation d

3 Jun 2026

A Close Look At World Model Recovery In Supervised Fine-Tuned LLM Planners

TutorialsDGX agent

arXiv:2606.03685v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) improves end-to-end classical planning in large language models (LLMs), but do these models also learn to represent and r

A formal definition and meta-model for a machine theory of mind

ResearchDGX agent

arXiv:2606.03471v1 Announce Type: new Abstract: This paper proposes, for the first time, a rigorous formal definition of the concept of Machine Theory of Mind, based on principles supported by evidenc

A Hybrid Approach For Malware Classification Using Secondary Features Fusion

ResearchDGX agent

arXiv:2606.03432v1 Announce Type: cross Abstract: The number of malware (either variant or novel) is rapidly increasing, making malware detection and mitigation a complex problem. One approach to impr

A Negative Result on Cross-Model Activation Transfer in a Pythia Multi-Hop Setting

SafetyDGX agent

arXiv:2606.03280v1 Announce Type: new Abstract: Recent work shows that language models can transmit behavioural traits through hidden signals in generated data during training. We ask whether a more d

A New Framework for Cybersecurity Refusals in AI Agents

Model ReleasesDGX agent

arXiv:2606.02644v1 Announce Type: cross Abstract: Agentic scaffolds have dramatically improved LLM performance on complex, long-horizon tasks, yielding both broad benefits and amplified risks in domai

A Robust and Explainable Transformer-Based Framework for Phishing Email Detection

Local AiDGX agent

arXiv:2511.12085v3 Announce Type: replace-cross Abstract: Phishing and related cyber threats are becoming increasingly sophisticated, with email-based phishing remaining the most persistent attack vec

A Scoping Review of the Ethical Perspectives on Anthropomorphising Large Language Model-Based Conversational Agents

ResearchDGX agent

arXiv:2601.09869v2 Announce Type: replace Abstract: Anthropomorphisation -- the phenomenon whereby non-human entities are ascribed human-like qualities -- has become increasingly salient with the rise

A Training-Free Mixture-of-Agents Framework for Multi-Document Summarization using LLMs and Knowledge Graphs

AgentsDGX agent

arXiv:2606.03867v1 Announce Type: cross Abstract: Multi-Document Summarization (MDS) plays a critical role in distilling essential information from collections of textual data. Existing approaches oft

Acceptance-Test-Driven Evaluation Protocols for Business-Centric LLM Systems

Model ReleasesDGX agent

arXiv:2606.02755v1 Announce Type: cross Abstract: Large language model (LLM) applications are increasingly expected to satisfy deterministic institutional requirements while relying on probabilistic g

Adaptive Latent Agentic Reasoning

AgentsDGX agent

arXiv:2606.02871v1 Announce Type: cross Abstract: Large reasoning models improve performance by generating extended chain-of-thought (CoT) reasoning, but this behavior becomes inefficient when applied

AdaWeather: Adaptively Mixing Probabilistic Weather Forecasts with Logarithmic Regret

ResearchDGX agent

arXiv:2606.02663v1 Announce Type: cross Abstract: Recent advances in machine learning have produced probabilistic weather forecasting models comparable to state-of-the-art numerical weather predictors

Agent libOS: A Library-OS-Inspired Runtime for Long-Running, Capability-Controlled LLM Agents

Local AiDGX agent

arXiv:2606.03895v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from request-response assistants into long-running software actors: they maintain state across model ca

Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path Forward

Model ReleasesDGX agent

arXiv:2602.12430v4 Announce Type: replace-cross Abstract: The transition from monolithic language models to modular, skill-equipped agents marks a defining shift in how large language models (LLMs) ar

Agentic Chain-of-Thought Steering for Efficient and Controllable LLM Reasoning

AgentsDGX agent

arXiv:2606.03965v1 Announce Type: cross Abstract: Large language models improve final-answer accuracy through extended chain-of-thought reasoning, but often spend tokens inefficiently and offer little

AI Agents Enable Adaptive Computer Worms

SafetyDGX agent

arXiv:2606.03811v1 Announce Type: cross Abstract: A computer worm is malware that spreads on a network by replicating itself from one machine to another. Traditional worms, like WannaCry, exploited pr

AI-Generated Traces for Novice Programmers: Learning Effects and Learner Differences in a Multi-Institutional Study

TutorialsDGX agent

arXiv:2606.03288v1 Announce Type: cross Abstract: Introductory programming (CS1) courses often struggle to support students' understanding of program execution. While visualizations can make execution

AI Model Extraction Attacks: Bypassing Single-Client Assumptions in Defenses

ResearchDGX agent

arXiv:2606.03381v1 Announce Type: cross Abstract: Ensuring the protection of Artificial Intelligence (AI) models deployed in military Command and Control (C2) systems and critical infrastructure is es

AI Rater Discrimination Depends on Scoring Protocol in Complex Clinical Decision-Making

ResearchDGX agent

arXiv:2606.03198v1 Announce Type: cross Abstract: Clinical AI evaluation increasingly delegates scoring to large language models (LLMs) acting as AI raters, yet their scoring behavior across evaluatio

AirDreamer: Generalist Drone Navigation with World Models

SafetyDGX agent

arXiv:2606.03252v1 Announce Type: cross Abstract: Navigating a drone in unseen and cluttered environments requires reliable generalization to unseen scene layouts and understanding of environmental st

Aletheia: What Makes RLVR For Code Verifiers Tick?

SafetyDGX agent

arXiv:2601.12186v3 Announce Type: replace-cross Abstract: Multi-domain thinking verifiers trained via Reinforcement Learning with Verifiable Rewards (RLVR) are a cornerstone of modern post-training. H

Align-KD: Distilling Cross-Modal Alignment Knowledge for Mobile Vision-Language Model Enhancement

SafetyDGX agent

arXiv:2412.01282v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) bring powerful understanding and reasoning capabilities to multimodal tasks. Meanwhile, the great need for capab

AlignAtt4LLM: Fast AlignAtt for Decoder-Only LLMs at IWSLT 2026 Simultaneous Speech Translation Task

Model ReleasesDGX agent

arXiv:2606.03967v1 Announce Type: cross Abstract: We describe AlignAtt4LLM, an IWSLT 2026 simultaneous speech translation system for English to German, Italian, and Chinese. The system is a synchronou

Aligning Data-Driven Predictors with Allocation: A Decision-Focused Approach to Survival Analysis

SafetyDGX agent

arXiv:2606.02671v1 Announce Type: cross Abstract: Machine learning predictors have become essential tools for guiding automated decision making. However, a major misalignment persists: predictive mode

AlphaEval: A Comprehensive and Efficient Evaluation Framework for Formula Alpha Mining

ResearchDGX agent

arXiv:2508.13174v2 Announce Type: replace Abstract: Formula alpha mining, which generates predictive signals from financial data, is critical for quantitative investment. Although various algorithmic

← Previous
1…162163164165166…358
Next →