AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
10 Aug 2026

Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts

Model ReleasesDGX agent

arXiv:2608.06770v1 Announce Type: new Abstract: Controllable surgical world models can provide a generative foundation for surgical artificial intelligence and simulation by synthesizing realistic ins

SyncSBC: Decentralized Swarm Behavior Prediction for Synchronized Autonomous Control

AgentsDGX agent

arXiv:2608.06587v1 Announce Type: cross Abstract: Robot swarms utilize many independent limited-sensing agents to produce complex emergent behaviors without requiring centralized control. However, lit

TaskSense: Focusing on What Matters in World Models

TutorialsDGX agent

arXiv:2608.06544v1 Announce Type: new Abstract: World models for visual control typically learn compact latent states by reconstructing observations, implicitly encouraging representations to preserve


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Taxonomy-Driven Analysis of Open-Source AI Risk Mitigation Tools

ApplicationsDGX agent

arXiv:2608.07446v1 Announce Type: cross Abstract: Rapid adoption of large language models (LLMs) in enterprise settings has introduced operational, security, and governance risks. As generative AI app

TEPA: Revoking Stale Memories for Conflict-Robust Language Agents

ResearchDGX agent

arXiv:2608.07429v1 Announce Type: new Abstract: Long-term memory enables language agents to reuse past facts, preferences, and task experience. Persistence also creates a central falsifiability proble

TEXAS: Task-Expert-Aware Supervision for Downstream Mixture-of-Experts LLM Adaptation

ResearchDGX agent

arXiv:2608.06396v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) language models route each token through a small subset of experts, making routing patterns useful for identifying task-relev

The Optimizer Is the Agent: Reasoning-Driven Search across Prompts, Programs, and ML Workflows

SafetyDGX agent

arXiv:2608.06714v1 Announce Type: new Abstract: Recent systems for optimizing prompts, programs, and ML workflows typically rely on explicit outer-loop controllers such as evolutionary search, bandits

The Perils of Agency: How Developers Perceive, Prioritize, and Address Risks in Agentic AI Products

AgentsDGX agent

arXiv:2606.15485v2 Announce Type: replace-cross Abstract: Agentic AI systems act autonomously, use tools, adapt to context, and operate in complex real-world environments. However, these same characte

TOFD: Target-Oriented Feature Decoupling against Poisoning Attacks in Split Federated Learning

ResearchDGX agent

arXiv:2608.07274v1 Announce Type: cross Abstract: Split Federated Learning (SFL) facilitates privacy-preserving collaborative training with reduced client-side overhead. However, its split architectur

Toward a Causal Data Management Ecosystem for Decision Making and Agentic AI

AgentsDGX agent

arXiv:2608.07214v1 Announce Type: cross Abstract: Modern AI is no longer a single model but an ecosystem: classical ML predictors, deep and multimodal models, large language models, and agents, each t

Towards a Theoretical Understanding of Two Tower Recommendation Models

ApplicationsDGX agent

arXiv:2403.00802v2 Announce Type: replace-cross Abstract: Production-grade recommender systems rely heavily on a large-scale corpus used by online media services, including Netflix, Pinterest, and Ama

Towards Assurance Closure in AI-Native Large-Scale Agile Software Development

AgentsDGX agent

arXiv:2608.07317v1 Announce Type: cross Abstract: The AI-Native Manifesto envisions large-scale agile software development in which humans increasingly govern intent, risk, and exceptions while agents

Towards Multi-Label Graph Foundation Models: from Single-Vector Representation Learning to Multi-Semantic Basis Learning

ResearchDGX agent

arXiv:2608.06394v1 Announce Type: new Abstract: Multi-label node classification is an important yet challenging task in graph learning, where nodes exhibit multiple semantics simultaneously. Existing

TRACE: A Multi-Layer Benchmark for Human AI Controller Coordination Under Drift and Failure

Model ReleasesDGX agent

arXiv:2608.06657v1 Announce Type: new Abstract: Modern cyber-physical and AI-assisted systems couple human operators, AI decision modules, and automated controllers in a single control loop, so trustw

TradeVerse: A Longitudinal Benchmark of Political Negotiation in International Trade

Model ReleasesDGX agent

arXiv:2608.06549v1 Announce Type: cross Abstract: LLMs are increasingly being applied to tasks involving institutional and political texts, but existing benchmarks evaluate them on isolated documents

Transformers Struggle to Use Their Emergent World Models: Revisiting the Tower of Hanoi, and the Illusion of Thinking

Model ReleasesDGX agent

arXiv:2608.07077v1 Announce Type: new Abstract: The Tower of Hanoi is a simple planning puzzle that in prior work has proven challenging for large reasoning models (LRMs). Current models solve the sta

TransSLR: A Lightweight Transformer for Sign Language Recognition

Model ReleasesDGX agent

arXiv:2608.06407v1 Announce Type: cross Abstract: Automated Sign Language Recognition for under-represented languages remains a largely unsolved problem. Central African Sign Language (CASL) exemplifi

TRIBE: Predicting Team Performance via Communication Behavior Ensembles

Model ReleasesDGX agent

arXiv:2608.06926v1 Announce Type: new Abstract: Designing autonomous agents that effectively assist human teams hinges on understanding team dynamics, often without task specific knowledge. We present

Unsupervised Adaptation of PDE Foundation Models

ResearchDGX agent

arXiv:2608.07053v1 Announce Type: new Abstract: Pretrained partial differential equation (PDE) foundation models can generalize across different equations, but adapting them to unseen PDE systems typi

Vehicle routing problem using deep reinforcement learning - A case study about truck planning in the industry

AgentsDGX agent

arXiv:2608.06668v1 Announce Type: new Abstract: As an important component of the supply chain industry, transportation has experienced rapid development in the past decade with the assistance of digit

WebGrader: Training LLMs for Web Development with Self-Evolving Programmatic Grader

Model ReleasesDGX agent

arXiv:2608.06474v1 Announce Type: new Abstract: Large language models increasingly generate complete websites from natural-language descriptions, and reinforcement learning has become a central approa

WebRider: Persona-Conditioned Intent Controllers for Live-Web Assistance

Model ReleasesDGX agent

arXiv:2608.06704v1 Announce Type: new Abstract: Delegating a web task involves more than asking a question; it requires transferring a policy: what to verify, how to handle uncertainty, which preferen

Winning by Peeking: Unenforced Budgets and Test-Set Selection Inflate Short-Budget AutoML Comparisons

Model ReleasesDGX agent

arXiv:2608.07303v1 Announce Type: new Abstract: Comparisons between AutoML systems at short time budgets -- tens of seconds rather than hours -- are common in tool READMEs and workshop papers, and the

WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN

SafetyDGX agent

arXiv:2608.07267v1 Announce Type: new Abstract: Recent vision-language navigation (VLN) systems increasingly adapt pretrained vision-language models (VLMs) into vision-language-action (VLA) policies t

WorldMark: A Plug-and-Play World Knowledge Interface for Cross-Host Language Model Watermarking

Model ReleasesDGX agent

arXiv:2608.06416v1 Announce Type: cross Abstract: Watermarking traces the provenance of text produced by large language models by embedding statistically detectable signals during decoding. Existing s

Zero Gap Is Not Restoration: Stratified Per-Question Probability Evaluation and Step-wise Mitigation of Benchmark Contamination

Model ReleasesDGX agent

arXiv:2608.07341v1 Announce Type: cross Abstract: Test data from public benchmarks inevitably leaks into pretraining corpora, inflating evaluation scores once memorized. extbf{Contamination mitigation

ZIPBrain: Can EEG Foundation Models Be Faster, Locally Deployable, but Accurate?

Local AiDGX agent

arXiv:2608.07033v1 Announce Type: new Abstract: This work investigates whether Electroencephalograph (EEG) foundation models (EFMs) can be made faster and locally deployable without sacrificing accura

7 Aug 2026

A Lexical Analysis of online Reviews on Human-AI Interactions

ResearchDGX agent

arXiv:2511.13480v2 Announce Type: replace-cross Abstract: This study focuses on understanding the complex dynamics between humans and AI systems by analyzing user reviews. While previous research has

A note on conditional PAC-efficient reasoning in large language model routing

ResearchDGX agent

arXiv:2512.03057v2 Announce Type: replace-cross Abstract: We study distribution-free risk control for model routing, motivated by large language model reasoning. We formalize pointwise conditional eff

A Study of ASR Adaptation and Representation Dimensionality Reduction in Persian Speech Emotion Recognition Using Whisper

ResearchDGX agent

arXiv:2608.05165v1 Announce Type: cross Abstract: Speech Emotion Recognition (SER) in low-resource languages remains a challenging problem due to limited labeled data. In this work, we study the use o

A Two-Tier Perspective on Inference-Time Parallelism in Multi-Agent LLM Systems

Model ReleasesDGX agent

arXiv:2608.05791v1 Announce Type: cross Abstract: Large language model (LLM)-driven multi-agent systems typically require multiple model invocations and complex coordination during inference, and thei

A Unified Framework for Trajectory Prediction with Explicit Planning and Reaction Decomposition

ResearchDGX agent

arXiv:2608.05673v1 Announce Type: new Abstract: Trajectory prediction has shifted toward structured formulations with explicit social modeling. However, existing methods inadequately distinguish the f

ABC: Numerical Data Collection under Local Differential Privacy without Prior Knowledge

ResearchDGX agent

arXiv:2608.05737v1 Announce Type: cross Abstract: Local Differential Privacy (LDP) provides strong privacy guarantees for collecting numerical data. A fundamental challenge, however, is that existing

Abstract Event Causal Rules: Induction and Application

Model ReleasesDGX agent

arXiv:2608.05205v1 Announce Type: new Abstract: Event-centric intelligent analytical systems heavily depend on explicit causal event knowledge for risk early warning, decision-making support and narra

Activity Frames: Deterministic Screen-Activity Compilation for Agent Memory and Replay

Local AiDGX agent

arXiv:2608.05784v1 Announce Type: new Abstract: Computer-use agents pay full frontier inference to re-derive routines their user has already performed, because an agent's memory today records what the

Adaptive Arena-based Contestable Argumentative Network-of-Experts for Open-Ended Care Plan Coordination

SafetyDGX agent

arXiv:2608.05391v1 Announce Type: new Abstract: Care plan coordination demands synthesizing heterogeneous clinical, functional, and psychosocial information across multiple professional disciplines, w

AegisShield: Democratizing Cyber Threat Modeling with Generative AI

ResearchDGX agent

arXiv:2509.10482v2 Announce Type: replace-cross Abstract: The increasing sophistication of technology systems makes traditional threat modeling hard to scale, especially for small organizations with l

Agentic Nesting: A New Methodology for Existing Enterprise Application Integration and Services

AgentsDGX agent

arXiv:2608.05159v1 Announce Type: new Abstract: Enterprise operations extensively rely on multiple heterogeneous business systems and information applications, which also result in severe data silos a

Agentic self-driving microscopy benchmarks support qualification but do not necessarily generalize to unseen tasks

Model ReleasesDGX agent

arXiv:2608.05266v1 Announce Type: new Abstract: Large language model agents are increasingly being developed to control a wide range of scientific characterization tools including microscopes and sync

Agentic Software Issue Resolution with Large Language Models: A Survey

AgentsDGX agent

arXiv:2512.22256v2 Announce Type: replace-cross Abstract: Software issue resolution aims to address real-world issues in software repositories based on natural language descriptions provided by users,

AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2608.05987v1 Announce Type: new Abstract: Reinforcement learning (RL) with verifiable rewards constructs trajectory-level advantage estimates, yet it often fails to credit the few pivotal decisi

AI Playing Business Games: Benchmarking Large Language Models on Managerial Decision-Making in Dynamic Simulations

Model ReleasesDGX agent

arXiv:2509.26331v2 Announce Type: replace Abstract: The rapid advancement of LLMs sparked significant interest in their potential to augment or automate managerial functions. One of the most recent tr

All-Quadrant Bounded Clipping GRPO: Closing the Unbounded Blind Spot for Stable and Generalizable Training

SafetyDGX agent

arXiv:2601.03895v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as a popular algorithm for reinforcement learning with large language models (LLMs). How

An Axiomatic Benchmark for Evaluation of Scientific Novelty Metrics

Model ReleasesDGX agent

arXiv:2604.15145v2 Announce Type: replace Abstract: The rigorous evaluation of the novelty of a scientific paper is, even for human scientists, a challenging task. With the increasing interest in AI s

An Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goals

SafetyDGX agent

arXiv:2608.05255v1 Announce Type: cross Abstract: Retail investors lack access to the kind of personalized, tax-aware portfolio management that institutional clients take for granted -- existing robo-

An Optimal Agnostic PAC Algorithm

ResearchDGX agent

arXiv:2608.06363v1 Announce Type: cross Abstract: Let Hsubseteq{-1,+1}^X be a class of finite VC dimension dge1. Writing L for the binary risk and L^*=min_{hin H}L(h), we construct a learner achieving

Analogy as Nonparametric Bayesian Inference over Relational Systems

TutorialsDGX agent

arXiv:2006.04156v2 Announce Type: replace Abstract: Our inferences in the real world are rarely naive - we acquire experiences through our lifetime that can help us more quickly understand the structu

Answer First, Reason Later: Commitment Order in Diffusion LLMs

ResearchDGX agent

arXiv:2608.05687v1 Announce Type: cross Abstract: Masked diffusion language models (dLLMs) can commit tokens in any order -- a freedom marketed as their core advantage over autoregressive decoding. We

AppDeltaWorld: Transition-Grounded Delta Code World Model for Mobile GUI Agents

SafetyDGX agent

arXiv:2608.05891v1 Announce Type: new Abstract: Mobile GUI agents can operate apps through pixel perception and touch actions, making them a promising interface for collecting and improving long-horiz

APQF: Agentic Profiling-Guided Structured Pruning and Mixed-Precision Quantization with Adaptive Fine-Tuning

AgentsDGX agent

arXiv:2608.05499v1 Announce Type: cross Abstract: Modern deep neural networks achieve strong performance, but their scale makes them costly and slow, especially on resource-constrained edge devices. P

As You Wish: Mission Planning with Formal Verification using LLMs in Precision Agriculture

SafetyDGX agent

arXiv:2606.18519v2 Announce Type: replace-cross Abstract: Though robotic systems are now being commercialized and deployed in various industries, many of these systems are highly specialized and often

ASAT: Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection

SafetyDGX agent

arXiv:2505.02299v2 Announce Type: replace-cross Abstract: Machine Learning (ML) models are trained on in-distribution (ID) data but often encounter out-of-distribution (OOD) inputs during deployment--

ASTELD: A Six-Axis Classification Framework for Autonomous AI Agents - Design, Evaluation, and an OpenClaw Case Study

AgentsDGX agent

arXiv:2608.05201v1 Announce Type: cross Abstract: Autonomous AI agent platforms differ substantially in architecture, security, tool integration, execution, autonomy, and deployment, yet the field lac

Audio-to-Score Transcription using Pre-trained Features, Data Augmentation, and the New SheetSage-A2S Dataset

Model ReleasesDGX agent

arXiv:2608.06165v1 Announce Type: cross Abstract: Existing audio-to-score (A2S) systems primarily focus on classical music, and the application to popular music remains underexplored. This paper first

Automatic Detection of Deaths from Social Networking Sites

ResearchDGX agent

arXiv:2608.05183v1 Announce Type: cross Abstract: This dissertation analysed and discussed the differences in linguistic characteristics between pre-mortem and post-mortem social media content, and re

Autonomous Learning From Success and Failure: Goal-Conditioned Supervised Learning with Negative Feedback

SafetyDGX agent

arXiv:2509.03206v2 Announce Type: replace-cross Abstract: Learning from reward functions and imitation learning of demonstrations are the two principal approaches for training autonomous systems that

Autonomous Research Agents: A Survey of AI Scientists and the Verification Gap

Model ReleasesDGX agent

arXiv:2608.05179v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly used across the scientific research lifecycle: ideation, literature search, experiment design and e

AV-AIVAT: 74x Cheaper Agent Evaluation with Certified Anytime-Valid Stopping in Imperfect-Information Games

AgentsDGX agent

arXiv:2608.06362v1 Announce Type: cross Abstract: Deciding which of two agents is stronger means playing games until skill outweighs luck, and every game costs money, model inference, or expert time.

BaKron: Efficient Quantization with Kronecker-Factored Hessians

ResearchDGX agent

arXiv:2608.06291v1 Announce Type: cross Abstract: We accelerate a family of algorithms for neural network quantization whose geometry is informed by any Kronecker-factored approximation of the Hessian

BALANCE: Hybrid Autoregressive-Speculative LLM Inference in Wireless Edge Networks

ResearchDGX agent

arXiv:2608.05926v1 Announce Type: cross Abstract: Edge inference is a promising paradigm to provide large language model (LLM) inference services in next-generation mobile networks. LLM inference main

← Previous
1…1718192021…350
Next →