AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
7 Aug 2026

An Optimal Agnostic PAC Algorithm

ResearchDGX agent

arXiv:2608.06363v1 Announce Type: cross Abstract: Let Hsubseteq{-1,+1}^X be a class of finite VC dimension dge1. Writing L for the binary risk and L^*=min_{hin H}L(h), we construct a learner achieving

Analogy as Nonparametric Bayesian Inference over Relational Systems

TutorialsDGX agent

arXiv:2006.04156v2 Announce Type: replace Abstract: Our inferences in the real world are rarely naive - we acquire experiences through our lifetime that can help us more quickly understand the structu

Answer First, Reason Later: Commitment Order in Diffusion LLMs

ResearchDGX agent

arXiv:2608.05687v1 Announce Type: cross Abstract: Masked diffusion language models (dLLMs) can commit tokens in any order -- a freedom marketed as their core advantage over autoregressive decoding. We


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

AppDeltaWorld: Transition-Grounded Delta Code World Model for Mobile GUI Agents

SafetyDGX agent

arXiv:2608.05891v1 Announce Type: new Abstract: Mobile GUI agents can operate apps through pixel perception and touch actions, making them a promising interface for collecting and improving long-horiz

APQF: Agentic Profiling-Guided Structured Pruning and Mixed-Precision Quantization with Adaptive Fine-Tuning

AgentsDGX agent

arXiv:2608.05499v1 Announce Type: cross Abstract: Modern deep neural networks achieve strong performance, but their scale makes them costly and slow, especially on resource-constrained edge devices. P

As You Wish: Mission Planning with Formal Verification using LLMs in Precision Agriculture

SafetyDGX agent

arXiv:2606.18519v2 Announce Type: replace-cross Abstract: Though robotic systems are now being commercialized and deployed in various industries, many of these systems are highly specialized and often

ASAT: Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection

SafetyDGX agent

arXiv:2505.02299v2 Announce Type: replace-cross Abstract: Machine Learning (ML) models are trained on in-distribution (ID) data but often encounter out-of-distribution (OOD) inputs during deployment--

ASTELD: A Six-Axis Classification Framework for Autonomous AI Agents - Design, Evaluation, and an OpenClaw Case Study

AgentsDGX agent

arXiv:2608.05201v1 Announce Type: cross Abstract: Autonomous AI agent platforms differ substantially in architecture, security, tool integration, execution, autonomy, and deployment, yet the field lac

Audio-to-Score Transcription using Pre-trained Features, Data Augmentation, and the New SheetSage-A2S Dataset

Model ReleasesDGX agent

arXiv:2608.06165v1 Announce Type: cross Abstract: Existing audio-to-score (A2S) systems primarily focus on classical music, and the application to popular music remains underexplored. This paper first

Automatic Detection of Deaths from Social Networking Sites

ResearchDGX agent

arXiv:2608.05183v1 Announce Type: cross Abstract: This dissertation analysed and discussed the differences in linguistic characteristics between pre-mortem and post-mortem social media content, and re

Autonomous Learning From Success and Failure: Goal-Conditioned Supervised Learning with Negative Feedback

SafetyDGX agent

arXiv:2509.03206v2 Announce Type: replace-cross Abstract: Learning from reward functions and imitation learning of demonstrations are the two principal approaches for training autonomous systems that

Autonomous Research Agents: A Survey of AI Scientists and the Verification Gap

Model ReleasesDGX agent

arXiv:2608.05179v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly used across the scientific research lifecycle: ideation, literature search, experiment design and e

AV-AIVAT: 74x Cheaper Agent Evaluation with Certified Anytime-Valid Stopping in Imperfect-Information Games

AgentsDGX agent

arXiv:2608.06362v1 Announce Type: cross Abstract: Deciding which of two agents is stronger means playing games until skill outweighs luck, and every game costs money, model inference, or expert time.

BaKron: Efficient Quantization with Kronecker-Factored Hessians

ResearchDGX agent

arXiv:2608.06291v1 Announce Type: cross Abstract: We accelerate a family of algorithms for neural network quantization whose geometry is informed by any Kronecker-factored approximation of the Hessian

BALANCE: Hybrid Autoregressive-Speculative LLM Inference in Wireless Edge Networks

ResearchDGX agent

arXiv:2608.05926v1 Announce Type: cross Abstract: Edge inference is a promising paradigm to provide large language model (LLM) inference services in next-generation mobile networks. LLM inference main

Bayesian Expected Uncertainty Reduction (B-EUR) Model: A Computational Account of What Makes Design Options Worth Trying

ResearchDGX agent

arXiv:2608.05642v1 Announce Type: new Abstract: This paper proposes the Bayesian Expected Uncertainty Reduction (B-EUR) model, which formalizes the value of trying a candidate design action as its exp

Benchmarking the Benchmarks: Evaluating Benchmarks for Conversational Agents

Model ReleasesDGX agent

arXiv:2608.06329v1 Announce Type: cross Abstract: Task-oriented conversational agents are evaluated using curated or automatically generated benchmarks, yet benchmark quality is rarely assessed. Poor

Berkeley and Heiserman as an Unexhausted Architecture for Embodied Machine Intelligence

ResearchDGX agent

arXiv:2607.16465v2 Announce Type: replace Abstract: Edmund C. Berkeley is usually remembered as a mediator between symbolic logic and early computing, yet that standard description understates the sco

Beyond Feature Importance: A Comparative Analysis of Pattern Detection Methods in Cluster Interpretation

Local AiDGX agent

arXiv:2608.05880v1 Announce Type: cross Abstract: Interpreting clustering outcomes remains a fundamental challenge in data analysis, particularly in domains such as healthcare where meaningful pattern

Beyond Information Retrieval: Generative AI as an Epistemic Arbiter to Enhance Collaborative Problem-Solving

ResearchDGX agent

arXiv:2608.05171v1 Announce Type: cross Abstract: Generative AI (GAI) creates new opportunities for collaborative problem-solving (CPS), yet its role in shaping student interaction remains unclear. To

Beyond Sentiment: Comparing Traditional NLP and LLM-Based Multi-Dimensional Analysis for Political News Evaluation

SafetyDGX agent

arXiv:2608.05155v1 Announce Type: cross Abstract: Traditional sentiment analysis (SA) models, while effective for polarity classification, provide limited insight into the rhetorical, ideological, and

Beyond Sequence Order: Syntax-Informed Positional Embeddings for Transformers

Model ReleasesDGX agent

arXiv:2608.06111v1 Announce Type: cross Abstract: Positional embeddings (PE) in Transformers encode token distance and order but are largely agnostic to extit{syntactic structure}. We introduce extbf{

Beyond Top-K: Replacing Black-Box Retrieval with Interpretable Agentic Operations

AgentsDGX agent

arXiv:2608.06305v1 Announce Type: new Abstract: Retrieval-augmented generation over long documents is dominated by one design: chunk the text, embed the chunks, and surface the top-k nearest neighbour

Beyond Weights and Gradients: A Taxonomy of Federated Learning Messages

ResearchDGX agent

arXiv:2606.16891v2 Announce Type: replace-cross Abstract: Federated Learning is rapidly evolving beyond the exchange of traditional model weights and gradients, yet existing definitions fail to captur

Bias Analysis of L2 Speaking Assessment Systems Using Concept Activation Vectors

SafetyDGX agent

arXiv:2608.06300v1 Announce Type: new Abstract: Automatic speaking assessment systems are increasingly deployed in high-stakes settings to mark second language (L2) learners' speaking tests, making it

Big, Bright, or Invisible: A Frozen-Feature Benchmark of 3D CT Foundation Models

Model ReleasesDGX agent

arXiv:2608.05960v1 Announce Type: cross Abstract: Routine CT interpretation is inherently comprehensive, capturing incidental findings across the entire scan volume. 3D CT foundation models could assi

BioAgent Bench: An AI Agent Evaluation Suite for Bioinformatics

AgentsDGX agent

arXiv:2601.21800v4 Announce Type: replace Abstract: We introduce BioAgent Bench, an evaluation suite designed for measuring the performance and robustness of AI agents in common bioinformatics tasks.

BlockPython: A Process-Aware Agent-Supported Platform for the Transition from Block-Based to Python Programming

AgentsDGX agent

arXiv:2608.05716v1 Announce Type: new Abstract: The transition from block-based to text-based programming requires learners to convert visible program structures into abstract textual expressions, whi

C^3PO: Evaluating Cross-Modal Composition and Counterfactual Performance in Omnimodal Models

Model ReleasesDGX agent

arXiv:2608.05381v1 Announce Type: new Abstract: Current Multimodal Large Language Models (MLLMs) can process diverse sensory inputs, yet their reasoning remains heavily biased toward a dominant modali

CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction

Model ReleasesDGX agent

arXiv:2608.05359v1 Announce Type: new Abstract: CASCADE is an agentic framework that predicts downstream transcriptional effects of gene perturbation from precomputed ARACNe regulatory networks, expos

Cautious Context Steering for Language Model Personalization

TutorialsDGX agent

arXiv:2608.05813v1 Announce Type: new Abstract: Personalizing language models (LMs) to individual user preferences is essential for aligning responses with diverse goals and backgrounds. Existing meth

ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution

Model ReleasesDGX agent

arXiv:2608.05790v1 Announce Type: new Abstract: General-purpose large language model agents have achieved strong performance on tool-augmented tasks, yet they rely on assumptions break down in blockch

Challenges for Musical Education in the Age of AI and Digital Transformation

ApplicationsDGX agent

arXiv:2608.05176v1 Announce Type: cross Abstract: Music education has never been a static discipline. Each major technological shift has forced educators and institutions to reconsider what they teach

Challenges in Evaluating Explanation Methods for Static and Evolving Data

SafetyDGX agent

arXiv:2608.06351v1 Announce Type: new Abstract: This paper addresses the limitations of Explainable Artificial Intelligence (XAI) with respect to insufficient evaluation. They are illustrated through

CoCo: Code as CoT for Text-to-Image Preview and Rare Concept Generation

ResearchDGX agent

arXiv:2603.08652v2 Announce Type: replace Abstract: Recent advancements in Unified Multimodal Models (UMMs) have significantly advanced text-to-image (T2I) generation, particularly through the integra

CodeGrep: An RL-Trained Retrieval Agent for LLM Coding Agents

Model ReleasesDGX agent

arXiv:2608.05886v1 Announce Type: cross Abstract: Modern LLM coding agents such as Claude Code and OpenHands share a common inefficiency: they spend much of their token budget finding the file to patc

CogVis: Must Open-Vocabulary Change Detection Perceive the Scene Anew for Every Query?

Local AiDGX agent

arXiv:2608.06150v1 Announce Type: new Abstract: Earth-surface monitoring requires change detection models capable of recognizing arbitrary semantic categories. Open-Vocabulary Change Detection (OVCD)

Coherence-Oriented Dream Scene Visualisation

ResearchDGX agent

arXiv:2608.05233v1 Announce Type: new Abstract: Dreams can be emotionally intense but difficult to communicate. We describe the Dream Scene Visualiser (DSV) system which turns written dream descriptio

Comparative Approaches to Agent Retrieval over Large Skill Libraries

AgentsDGX agent

arXiv:2608.06196v1 Announce Type: new Abstract: Agents backed by large skill libraries must decide which skills to load and in what order. Loading the entire library into context is expensive and prov

Contextual Information Policy Optimization for Search Agents

SafetyDGX agent

arXiv:2608.06128v1 Announce Type: new Abstract: Search agents extend large language models beyond static parametric memory by enabling them to acquire and use ex ternal evidence during multi-step reas

Continual Learning in Transition

Model ReleasesDGX agent

arXiv:2608.06216v1 Announce Type: cross Abstract: Classical continual learning (CL) has primarily focused on enabling models to update and retain knowledge through parameter-centric mechanisms, e.g.,

Counterfactual Analysis via Large Language Models

ResearchDGX agent

arXiv:2608.05367v1 Announce Type: new Abstract: Counterfactual analysis aims to predict potential outcomes under hypothetical scenarios, offering valuable insights for decision-making. This paper inve

CourseGraph: Finding overlaps and differences in Computer Science courses across universities

SafetyDGX agent

arXiv:2608.05910v1 Announce Type: new Abstract: Student mobility programs such as Erasmus+ enable students to take courses at other universities, broadening their academic and cultural horizons. Howev

CREBench: Evaluating Large Language Models in Cryptographic Binary Reverse Engineering

Model ReleasesDGX agent

arXiv:2604.03750v2 Announce Type: replace-cross Abstract: Reverse engineering (RE) is central to software security, particularly for cryptographic programs that handle sensitive data and are highly pr

CRINN: Contrastive Reinforcement Learning for Approximate Nearest Neighbor Search

Model ReleasesDGX agent

arXiv:2508.02091v4 Announce Type: replace-cross Abstract: Approximate nearest-neighbor search (ANNS) algorithms have become increasingly critical for recent AI applications, particularly in retrieval-

CUDA-L2: Surpassing cuBLAS Performance for Matrix Multiplication through Reinforcement Learning

HardwareDGX agent

arXiv:2512.02551v4 Announce Type: replace-cross Abstract: In this paper, we propose CUDA-L2, a system that combines large language models (LLMs) and reinforcement learning (RL) to automatically optimi

D-CLOT: Double Closed Loop Optimal Transport for Unsupervised Action Segmentation

Model ReleasesDGX agent

arXiv:2608.05877v1 Announce Type: cross Abstract: Optimal transport (OT) has emerged as an effective framework for unsupervised action segmentation. Yet, in existing OT-based methods, the latent actio

d3LLM: Ultra-Fast Diffusion LLM using Pseudo-Trajectory Distillation

ResearchDGX agent

arXiv:2601.07568v3 Announce Type: replace-cross Abstract: Diffusion large language models (dLLMs) offer capabilities beyond those of autoregressive (AR) LLMs, such as parallel decoding and random-orde

DASH: Decoupled Adaptive Surrogate - Acquisition Harness for Automated Bayesian Optimization

Model ReleasesDGX agent

arXiv:2608.00641v2 Announce Type: replace Abstract: Bayesian optimization (BO) relies on a surrogate model and an acquisition function, yet the most suitable choices vary across tasks and optimization

DASH: Divergence-Adaptive Supervision Horizons for On-Policy Self-Distillation of Reasoning Models

Model ReleasesDGX agent

arXiv:2608.06243v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves the reasoning capabilities of large language models using automatically verifiable outcom

DeepForgeSeal: Latent Space-Driven Semi-Fragile Watermarking for Deepfake Detection Using Adversarial Reinforcement Learning

AgentsDGX agent

arXiv:2511.04949v2 Announce Type: replace-cross Abstract: Rapid advances in generative AI have led to increasingly realistic deepfakes, posing growing challenges for law enforcement and public trust.

Depth-Guided Video Object Counting in Crowded Scenes

ResearchDGX agent

arXiv:2608.06236v1 Announce Type: cross Abstract: Our primary objective is to advance video object counting in crowded scenes, aiming to robustly count all instances of a target category based on give

DistMedVL: Distributional Vision-Language Alignment for Uncertainty-Aware Medical Image Segmentation

SafetyDGX agent

arXiv:2608.05683v1 Announce Type: cross Abstract: Cross-modal alignment of visual and textual representations is fundamental to multimodal medical image understanding, yet remains hindered by uncertai

DoctorAgents: an agentic framework to iteratively refine AutoML pipeline for small clinical temporal data

AgentsDGX agent

arXiv:2608.05375v1 Announce Type: new Abstract: Clinical machine learning (ML) has the potential to support high-stakes medical decision-making, but reliable deployment is often constrained by scarce,

Does FLAIR super-resolution erase or hallucinate small white-matter lesions?

ResearchDGX agent

arXiv:2608.06311v1 Announce Type: cross Abstract: White matter hyperintensities (WMH), bright regions on Fluid-attenuated Inversion Recovery (FLAIR) scans are associated with cerebrovascular pathology

Does Latent Context Help? A Controlled Evaluation of Inverse Reinforcement Learning in Arctic Shipping

SafetyDGX agent

arXiv:2608.06105v1 Announce Type: cross Abstract: Artificial Intelligence (AI)-assisted navigation can help Arctic shipping adapt to rapidly changing sea-ice conditions, but reliable deployment requir

Domain-Grounded Candidate Selection for Agentic Image Editing: A Shadow Removal Case

Model ReleasesDGX agent

arXiv:2608.06075v1 Announce Type: cross Abstract: Commercial vision-language models are reshaping computer vision, with visual priors broad enough to rival task-specific systems. This raises a natural

DREAM: LLM-based Dynamic Role-playing via Event-Aware Memory Graph

Model ReleasesDGX agent

arXiv:2608.05170v1 Announce Type: cross Abstract: Role-playing agents (RPAs) have emerged as a key application of large language models, enabling immersive and high-fidelity character simulation. Accu

DreamGuard: Efficient Runtime Guardrail for LLM Agents via Risk-Aware World Model

SafetyDGX agent

arXiv:2608.05695v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly invoke external tools and interact with real-world systems, unsafe actions may cause irreversible cons

ECG-LENS: Lead-Aware Clinical Context Enriched ECG Report Generation and Evaluation

Local AiDGX agent

arXiv:2608.05893v1 Announce Type: new Abstract: Electrocardiography (ECG) is one of the most widely used non-invasive tools for diagnosing cardiovascular disease, but transforming multi-lead ECG recor

← Previous
1…2223242526…354
Next →