AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Safety

When Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents Across 67 Frontier Models

DGX agent

arXiv:2606.27288v1 Announce Type: new Abstract: Multi-model LLM systems such as routing, voting, cascades, fusion, and mixture-of-agents are used to beat single-model accuracy. We show that their gain

safetyarxiv-cs-ai
26 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Where Do CoT Training Gains Land in LLM based Agents?

DGX agent

arXiv:2606.26935v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning is widely used in language-model agents, but prior work has shown that verbalized CoT is not always faithful and may in

researcharxiv-cs-ai
26 Jun 2026
Model Releases

Where Do Models Find Happiness? Emotion Vectors in Open-Source LLMs

DGX agent

arXiv:2606.26987v1 Announce Type: cross Abstract: Recent work identified emotion vectors in Claude Sonnet 4.5, which are internal representations that encode emotion concepts, causally influence behav

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

XMSE-Aware Adaptive Empirical Bayes Estimation

DGX agent

arXiv:2606.26975v1 Announce Type: cross Abstract: Empirical Bayes (EB) estimators can match the first-order asymptotic risk of maximum likelihood (ML) while behaving very differently at second order:

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Zero-Shot Size Transfer for Neural ODEs on Sparse Random Graphs: Graphon Limits and Adjoint Convergence

DGX agent

arXiv:2606.26662v1 Announce Type: cross Abstract: Graph Neural Differential Equations (GNDEs) model continuous-time graph dynamics by parameterizing Neural ODE velocity fields with Graph Neural Networ

model-releasesarxiv-cs-ai
26 Jun 2026
Applications

A Marketplace for AI-Generated Adult Content and Deepfakes

DGX agent

arXiv:2601.09117v3 Announce Type: replace-cross Abstract: Generative AI systems increasingly enable the production of highly realistic synthetic media. Civitai, a popular community-driven platform for

applicationsarxiv-cs-ai
25 Jun 2026
Research

A Neuromorphic Trigger for Efficient Audio Event Detection

DGX agent

arXiv:2606.17775v2 Announce Type: replace-cross Abstract: Efficient processing of continuous audio streams remains a key challenge for real-time and resource-constrained systems. This paper introduces

researcharxiv-cs-ai
25 Jun 2026
Model Releases

Agent-as-a-Router: Agentic Model Routing for Coding Tasks

DGX agent

arXiv:2606.22902v2 Announce Type: replace Abstract: Real-world users typically have access to multiple Large Language Models (LLMs) from different providers, and these LLMs often excel at distinct dom

model-releasesarxiv-cs-ai
25 Jun 2026
Agents

Agentic Knowledge Tracing: A Multi-Agent LLM Architecture for Stealth Assessment of Financial Literacy in Serious Games

DGX agent

arXiv:2606.25358v1 Announce Type: new Abstract: Assessing financial literacy during gameplay without disrupting the learning experience remains a key challenge in serious games for education. We prese

agentsarxiv-cs-ai
25 Jun 2026
Agents

Agentic Software Engineering: Foundational Pillars and a Research Roadmap

DGX agent

arXiv:2509.06216v3 Announce Type: replace-cross Abstract: Agentic Software Engineering (SE 3.0) represents a new era where intelligent agents are tasked not with simple code generation, but with achie

agentsarxiv-cs-ai
25 Jun 2026
Agents

Agentic System as Compressor: Quantifying System Intelligence in Bits

DGX agent

arXiv:2606.25960v1 Announce Type: new Abstract: Large language models are turning from isolated predictors into agentic systems: they call tools, retrieve evidence, obey environment constraints, use v

agentsarxiv-cs-ai
25 Jun 2026
Research

AI-Assisted Computational Reproducibility on the FABRIC Testbed

DGX agent

arXiv:2606.25879v1 Announce Type: cross Abstract: Computational reproducibility remains difficult despite being central to scientific research. In this paper, we show how the international FABRIC test

researcharxiv-cs-ai
25 Jun 2026
Agents

AI Snitches Get Glitches: Towards Evading Agentic Surveillance

DGX agent

arXiv:2606.25836v1 Announce Type: new Abstract: To better assist users with completing challenging tasks, AI agents mediate communications, access data, and interact with different APIs. Many employer

agentsarxiv-cs-ai
25 Jun 2026
Model Releases

AIChilles: Automatically Uncovering Hidden Weaknesses in AI-Evolved Systems

DGX agent

arXiv:2606.15834v2 Announce Type: replace Abstract: The computer systems community has recently seen growing interest in AI-driven system evolution, where AI agents iteratively rewrite systems. Framew

model-releasesarxiv-cs-ai
25 Jun 2026
Research

An Approach for a Supporting Multi-LLM System for Automated Certification Based on the German IT-Grundschutz

DGX agent

arXiv:2606.25608v1 Announce Type: cross Abstract: This paper presents a novel approach to perform semi-automated BSI IT-Grundschutz certification using a MultiLarge Language Model system (MLS) with Hy

researcharxiv-cs-ai
25 Jun 2026
Research

Attractive and Repulsive Pattern Control in Sequence Generation

DGX agent

arXiv:2606.24911v1 Announce Type: cross Abstract: Variable-order Markov models preserve local symbolic syntax by adapting context length, but long continuations can enter recurring high-order 'tunnels

researcharxiv-cs-ai
25 Jun 2026
Applications

AutoRelAnnotator: Calibrated Model Cascades for Cost-Efficient Relevance Evaluation in Sponsored Search

DGX agent

arXiv:2606.25871v1 Announce Type: cross Abstract: How can we generate high-quality relevance annotations at scale without the cost and delays of human labeling? Relevance annotations are the backbone

applicationsarxiv-cs-ai
25 Jun 2026
Research

BCoughBench: Benchmarking Respiratory Acoustic Foundation Models Under Body-Coupled Wearable Sensor Conditions

DGX agent

arXiv:2606.25116v1 Announce Type: cross Abstract: Respiratory acoustic foundation models (FMs) are benchmarked exclusively on smartphone recordings, yet clinical deployment increasingly targets body-c

researcharxiv-cs-ai
25 Jun 2026
Research

Beyond Shapley: Efficient Computation of Asymmetric Shapley Values

DGX agent

arXiv:2606.25103v1 Announce Type: new Abstract: We address the problem of explainability in machine learning models through feature attribution methods. In particular, we consider a variant of Shapley

researcharxiv-cs-ai
25 Jun 2026
Model Releases

BrainAgent: A Large Language Model-Driven Multi-Agent Framework for Autonomous Brain Signal Understanding

DGX agent

arXiv:2606.25400v1 Announce Type: new Abstract: Brain-Computer Interfaces (BCIs) and brain signal understanding are pivotal for clinical health and next-generation interactions. Despite this significa

model-releasesarxiv-cs-ai
25 Jun 2026
Agents

Can Trustless Agents Be Trusted? An Empirical Study of the ERC-8004 Decentralized AI Agent Ecosystem

DGX agent

arXiv:2606.26028v1 Announce Type: cross Abstract: As autonomous AI agents increasingly transact across organizational boundaries, a fundamental trust challenge emerges: how can an agent assess whether

agentsarxiv-cs-ai
25 Jun 2026
Model Releases

CausalRAG2: Hierarchical Causal Knowledge Graph Design for RAG

DGX agent

arXiv:2602.05143v2 Announce Type: replace Abstract: Retrieval augmented generation (RAG) has enhanced large language models by enabling access to external knowledge, with graph-based RAG emerging as a

model-releasesarxiv-cs-ai
25 Jun 2026
Research

Confidence Sequences for Online Statistical Model Checking of Markov Decision Processes

DGX agent

arXiv:2606.25797v1 Announce Type: new Abstract: Markov decision processes (MDPs) are a classic model of decision making under uncertainty, exhibiting both non-deterministic choice as well as probabili

researcharxiv-cs-ai
25 Jun 2026
Safety

Conformal Recovery-Deadline Certificates for Runtime Assurance of Adapting Controllers

DGX agent

arXiv:2606.25371v1 Announce Type: cross Abstract: Runtime assurance (RTA) protects a safety-critical system by switching from an advanced controller to a verified safe controller when a monitored cond

safetyarxiv-cs-ai
25 Jun 2026
Research

CrossAccent-TTS: Cross-Lingual Accent-Intensity Controllable Text-to-Speech via Disentangled Speaker and Accent Representations

DGX agent

arXiv:2606.25403v1 Announce Type: cross Abstract: Accent conversion and controllability remain fundamental challenges in cross-lingual text-to-speech (TTS), particularly for low-resource and phonetica

researcharxiv-cs-ai
25 Jun 2026
Agents

Decoupling Reconnaissance and Exploitation: Measuring the Capability Boundaries of LLM-Based Web Penetration Testing

DGX agent

arXiv:2606.25332v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promise for automated penetration testing, yet existing end-to-end black-box evaluations are highly susceptibl

agentsarxiv-cs-ai
25 Jun 2026
Agents

Domain-Specific Agents for Cherenkov Telescope Array Control Software and Gamma-Ray Data Analysis

DGX agent

arXiv:2510.01299v3 Announce Type: replace-cross Abstract: We present domain-adapted large language model agents designed to support Cherenkov Telescope Array operation and data analysis. The agents co

agentsarxiv-cs-ai
25 Jun 2026
Model Releases

Elo-Disentangled Player-Style Embeddings for Human Chess via Rating-Conditioned Residual Move Model

DGX agent

arXiv:2606.25176v1 Announce Type: new Abstract: We study representation learning for individual human chess style: a per-player embedding learned from a player's move history such that inner products

model-releasesarxiv-cs-ai
25 Jun 2026
Local Ai

EmotionAI: A Privacy-Preserving Computational Intelligence Pipeline for Speech-Emotion-Grounded Conversational Analysis

DGX agent

arXiv:2606.24941v1 Announce Type: cross Abstract: Reviewing recorded interviews for affective cues such as composure, hesitation and agitation is slow and subjective, and cloud services that could aut

local-aiarxiv-cs-ai
25 Jun 2026
Research

End-to-End Voice Intent Recognition for Spontaneous Human-Drone Interaction with Naive Users

DGX agent

arXiv:2606.24910v1 Announce Type: cross Abstract: Voice control offers an intuitive alternative to manual drone piloting, yet most existing systems rely on rigid command vocabularies that fail to hand

researcharxiv-cs-ai
25 Jun 2026
Model Releases

Epistemic Bias Injection: Manipulating LLM Opinion via Selective Context Retrieval

DGX agent

arXiv:2512.00804v3 Announce Type: replace-cross Abstract: When answering user queries, LLMs often retrieve knowledge from external sources stored in retrieval-augmented generation (RAG) databases. The

model-releasesarxiv-cs-ai
25 Jun 2026
Agents

Evaluating AGENTS.md: Are Repository-Level Context Files Helpful for Coding Agents?

DGX agent

arXiv:2602.11988v2 Announce Type: replace-cross Abstract: A widespread practice in software development is to tailor coding agents to repositories using context files, such as AGENTS.md. Although this

agentsarxiv-cs-ai
25 Jun 2026
Local Ai

Explainable Control Framework (XCF) based on Fuzzy Model-Agnostic Explanation and LLM Agent-Supported Interface

DGX agent

arXiv:2606.25941v1 Announce Type: cross Abstract: Increasing demand for precise and reliable control in complex scenarios has led to the development of increasingly sophisticated controllers, includin

local-aiarxiv-cs-ai
25 Jun 2026
Model Releases

Exploring Information Seeking Agent Consolidation

DGX agent

arXiv:2602.00585v2 Announce Type: replace Abstract: Information-seeking agents have emerged as a powerful paradigm for knowledge-intensive tasks, yet today's systems remain specialized for the open we

model-releasesarxiv-cs-ai
25 Jun 2026
Research

Externalizing Research Synthesis and Validation in AI Scientists through a Research Harness

DGX agent

arXiv:2606.18874v2 Announce Type: replace Abstract: AI systems can increasingly automate scientific workflows, but the reasoning that links prior evidence, generated ideas, experiments and final claim

researcharxiv-cs-ai
25 Jun 2026
Model Releases

Failure Modes of Large Language Models on Research-Level Mathematics: A Taxonomy and an Empirical Characterisation

DGX agent

arXiv:2606.24902v1 Announce Type: cross Abstract: The 'First Proof' benchmark [1] posed ten research-level mathematics questions to the strongest publicly available LLMs and found them consistently wr

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming

DGX agent

arXiv:2606.19887v2 Announce Type: replace-cross Abstract: Existing safety benchmarks target general adversarial scenarios but miss finance-specific risks. Financial LLMs face regulatory compliance vio

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models

DGX agent

arXiv:2606.25391v1 Announce Type: cross Abstract: Recent Large Audio Language Models (LALMs) have achieved remarkable progress in audio perceptual tasks across individual acoustic layers, including sp

model-releasesarxiv-cs-ai
25 Jun 2026
Research

Fuzzy Quantification over OWL Ontologies and Knowledge Graphs

DGX agent

arXiv:2606.25778v1 Announce Type: new Abstract: This paper presents a versatile framework for evaluating fuzzy quantification queries over both standard and fuzzy ontologies as well as knowledge graph

researcharxiv-cs-ai
25 Jun 2026
Model Releases

Geometry-Aware Online Scheduling for LLM Serving: From Theoretical Bound to System Practice

DGX agent

arXiv:2606.22327v2 Announce Type: replace Abstract: The explosive demand for interactive Large Language Model serving has highlighted the management of the Key-Value cache's dynamic memory footprint a

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

GUI agent: Guided Exploration of User-Sensitive Screens

DGX agent

arXiv:2606.25705v1 Announce Type: new Abstract: LLM agents are increasingly being used to automate tasks for users within an open GUI environment. They inevitably encounter screens containing user-sen

safetyarxiv-cs-ai
25 Jun 2026
Applications

Helpful or Harmful? Evaluating LLM-Assisted Vulnerability Patching via a Human Study

DGX agent

arXiv:2606.25973v1 Announce Type: cross Abstract: Software vulnerability remediation is a cognitively demanding task that requires specialized security expertise often lacking in general developers. I

applicationsarxiv-cs-ai
25 Jun 2026
Safety

Heuresis: Search Strategies for Autonomous AI Research Agents Across Quality, Diversity and Novelty

DGX agent

arXiv:2606.25198v1 Announce Type: new Abstract: Autonomous AI Research promises to accelerate the scientific progress of machine learning. To realise this goal, current Large Language Model (LLM)-base

safetyarxiv-cs-ai
25 Jun 2026
Model Releases

How Small Can 6G Reason? Scaling Tiny-to-Small Language Models for AI-Native Networks

DGX agent

arXiv:2603.02156v2 Announce Type: replace-cross Abstract: Emerging 6G visions, reflected in ongoing standardization efforts within 3GPP, IETF, ETSI, ITU-T, and the O-RAN Alliance, increasingly charact

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Improving Zero-Shot Offline RL via Behavioral Task Sampling

DGX agent

arXiv:2604.25496v2 Announce Type: replace Abstract: Offline zero-shot reinforcement learning (RL) aims to learn agents that optimize unseen reward functions without additional environment interaction.

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

LibEvoBench: Probing Temporal Knowledge Stratification in Code Generation Models

DGX agent

arXiv:2606.25402v1 Announce Type: cross Abstract: Large software projects often depend on older versions of libraries, even as APIs continue to evolve across releases. This creates a challenge for LLM

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

Lightweight PCGAE-Net: Parallel CrossGate Attention and Bottleneck AutoEncoder for Efficient 5G Channel Prediction

DGX agent

arXiv:2606.25401v1 Announce Type: cross Abstract: Accurate channel state information (CSI) prediction is essential for proactive beamforming and resource management in 5G massive MIMO systems, yet the

safetyarxiv-cs-ai
25 Jun 2026
Safety

Long-Term Simulation Exposes Cognitive-Developmental Risks in AI Companions

DGX agent

arXiv:2606.25396v1 Announce Type: new Abstract: AI companions powered by large language models increasingly interact with cognition-developing users, including children and adolescents, creating risks

safetyarxiv-cs-ai
25 Jun 2026
← Previous
1…155156157158159…448
Next →