AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
10 Aug 2026

Do AI Personas Grow? Analyzing and Benchmarking Personality Evolution in LLM Agents After Life Events

Model ReleasesDGX agent

arXiv:2608.06485v1 Announce Type: cross Abstract: Personality-conditioned LLM agents (PC-Agents) are increasingly used in emotional support, social simulation, and role-playing, motivating the develop

DocMemo: Dynamic Evidence Discovery via Probabilistic Memory-Guided Retrieval for Multi-Modal Document Understanding

ResearchDGX agent

arXiv:2608.07067v1 Announce Type: new Abstract: Long-document understanding requires locating sparse and heterogeneous evidence across hundreds of pages, yet existing systems remain limited by static

Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
SafetyDGX agent

arXiv:2608.06949v1 Announce Type: new Abstract: Prior benchmarking work has shown that a single large language model (LLM), forced to make life-or-death resource-allocation decisions, exhibits measura

Dueling World Models: Advantage-Style Action Channels for Common-Mode Distractor Rejection

AgentsDGX agent

arXiv:2608.06706v1 Announce Type: cross Abstract: Latent world models plan by predicting future states from an action, but when a scene contains motion the agent does not control, they quietly go acti

ED-CSP: Crystal Structure Prediction from Electron Diffraction

Model ReleasesDGX agent

arXiv:2608.06448v1 Announce Type: cross Abstract: Recovering a periodic 3D crystal structure from sparse, unindexed electron diffraction (ED) observations is a challenging generative inverse problem.

EliSeg: Verified Target Construction for Report-Grounded Abnormality Segmentation

ResearchDGX agent

arXiv:2608.07299v1 Announce Type: cross Abstract: Radiology reports describe clinical observations but do not specify executable segmentation targets. They may contain present, negated, prior,uncertai

EMAS: Stabilizing Multi-Agent System Evolution through Evidence-Guided Revision

Model ReleasesDGX agent

arXiv:2608.07196v1 Announce Type: new Abstract: Many methods for automated multi-agent system design optimize prompts and topologies during an initial design stage and then deploy the resulting system

EntropyMoE: Entropy-Aware Sparse Expert Routing for Tokenizer-Free LLMs

ResearchDGX agent

arXiv:2608.06398v1 Announce Type: new Abstract: Recent byte-level large language models (LLMs) have made tokenizer-free modeling increasingly competitive by grouping bytes into dynamically sized patch

Evaluating Useful Surrogate Models for Configuration Tuning Beyond Accuracy: A Fitness Landscape Analysis Perspective

ResearchDGX agent

arXiv:2509.21945v2 Announce Type: replace-cross Abstract: To efficiently tune configuration for better software system performance (e.g., latency) at the deployment and maintenance stage, many tuners

Evaluating XAI Support From A Hierarchical Reinforcement Learning Policy in Human-Agent Collaboration

Model ReleasesDGX agent

arXiv:2608.06381v1 Announce Type: cross Abstract: Explainable AI (XAI) has shown promise for human-agent collaboration, yet results rely on hand-crafted policies in custom environments, limiting gener

Evolving Parallel Algorithm Portfolios via Potential-Aware Instance Generation with LLMs

Local AiDGX agent

arXiv:2608.06808v1 Announce Type: new Abstract: The Automatic Construction of Portfolios via Large Language Models (LLM-ACP) suffers from poor generalization in practical few-shot scenarios when solvi

Explicit, Not Longer: What Makes Epistemic Stance Survive Memory Compression

AgentsDGX agent

arXiv:2608.06953v1 Announce Type: cross Abstract: Agent memory systems compress what they store, and compression is built to drop qualifiers, so a claim's epistemic standing tends not to survive being

Factorized Hypothesis Search for Evidence-to-Taxonomy Retrieval

ResearchDGX agent

arXiv:2608.06614v1 Announce Type: cross Abstract: Large-taxonomy retrieval often assumes that the input already expresses the target concept. In many settings, however, the input is indirect evidence,

Fast LapSum: Exact Differentiable Top-k at Million Scale

HardwareDGX agent

arXiv:2608.06912v1 Announce Type: new Abstract: The top-k operation is a fundamental building block of modern sparse computation, enabling token routing, expert activation, memory selection, and atten

FedLBW: A Loss-Based Weighting Strategy for Federated Learning on Non-IID Data in Wireless Networks

ResearchDGX agent

arXiv:2608.07007v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative machine learning (ML) across distributed clients while preserving privacy. However, efficient model conver

FedVAR: Prototype-Aligned Federated Framework for Video Anomaly Recognition

SafetyDGX agent

arXiv:2608.06876v1 Announce Type: cross Abstract: In the era of Industrial Internet of Things (IIoT) and Cyber-Physical Systems (CPS), Federated Learning (FL) offers a promising decentralized intellig

Finding Usable Weight Mechanisms with Tiled SVD

Model ReleasesDGX agent

arXiv:2608.06969v1 Announce Type: new Abstract: The dominant approach to mechanistic interpretability trains proxy dictionaries such as sparse autoencoders and labels features from max-activating text

FinRank: An Evidence-Grounded Benchmark for Financial Question Answering and Retrieval over SEC Filings

Model ReleasesDGX agent

arXiv:2608.07400v1 Announce Type: new Abstract: Financial question answering is typically evaluated by answer correctness, yet in SEC filings a plausible and even numerically correct answer can be gro

Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing

Model ReleasesDGX agent

arXiv:2608.07437v1 Announce Type: new Abstract: Reliable hypothesis testing is the foundation of many empirical scientific claims. Large language model (LLM) agents are increasingly used to automate t

Flowing Through States: Neural ODE Regularization for Reinforcement Learning

ResearchDGX agent

arXiv:2608.06595v1 Announce Type: cross Abstract: Neural networks applied to sequential decision-making tasks typically rely on latent representations of environment states. While environment dynamics

Fluid-DiT: Graph-Free Diffusion Transformers for Fluid Flow Simulations Learning

ResearchDGX agent

arXiv:2608.07161v1 Announce Type: cross Abstract: Simulating complex fluid flows requires capturing full equilibrium distributions rather than just mean trajectories, yet high-fidelity solvers remain

From Cheap Fakes to Pure Synthesis: Addressing the New Era of T2V Fake News Videos

SafetyDGX agent

arXiv:2608.06732v1 Announce Type: new Abstract: Recent text-to-video (T2V) generation models enable fake news videos to be synthesized from scratch, shifting the threat beyond cheap fakes assembled fr

From Points to Edges: Edge-Conditioned Spectral Operators for Physics-Sensitive PDE Learning

ResearchDGX agent

arXiv:2608.06894v1 Announce Type: new Abstract: Neural operators have become a central tool for solving partial differential equations (PDEs), with spectral operators offering efficient global mixing

From probability to causality in probabilistic logic programming

ResearchDGX agent

arXiv:2608.07230v1 Announce Type: new Abstract: Probabilistic logic programming is a formalism of statistical relational artificial intelligence that supports causal queries, including interventions f

FUSE: Feature-Wise Unified Specialization with Cross-Column Exchange for Mixed-Type Tabular Flow Matching

ResearchDGX agent

arXiv:2608.07294v1 Announce Type: cross Abstract: Generating mixed-type tabular data requires jointly modeling diverse feature distributions and their complex cross-column dependencies. Variational fl

FutureBridge: Token Selection Beyond Local Preference in Collaborative Decoding

Local AiDGX agent

arXiv:2608.06819v1 Announce Type: cross Abstract: Token-level collaboration allows a large language model (LLM) to assist a small language model (SLM) when their predictions diverge. Existing methods

Gated-BEPO: Confidence-Gated Bellman Credit Assignment for Large Language Model Agents

SafetyDGX agent

arXiv:2608.06861v1 Announce Type: new Abstract: Training large language model agents in long-horizon environments requires assigning credit from sparse terminal outcomes to individual actions. Existin

Genotypic Triggers: Exposing Pharmacogenomic Blind Spots via Host-Specific Backdoors in Generative Antimicrobial Peptide Models

SafetyDGX agent

arXiv:2608.06779v1 Announce Type: cross Abstract: Large Language Models (LLMs) have accelerated drug discovery, particularly in the automated design of antimicrobial peptides (AMPs). However, current

Geo-Spatial Concept Probing of Large Language Models: Abstraction, Compositionality, and Grounding

Model ReleasesDGX agent

arXiv:2608.07353v1 Announce Type: cross Abstract: Understanding concepts is fundamental to generalization. Despite their impressive performance on a wide range of tasks, Large Language Models (LLMs) s

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks

Model ReleasesDGX agent

arXiv:2608.07411v1 Announce Type: new Abstract: In the context of geodata, existing Large Language Models have often been studied in a homogeneous setting, which has considerably limited insights into

GeoDistill-Refine: Silhouette-First Geometry Distillation for Annotation-Free Spacecraft Segmentation

ResearchDGX agent

arXiv:2608.07405v1 Announce Type: cross Abstract: Foundation segmentation models can provide supervision for spacecraft imagery without manual training masks, but their predictions vary with textual p

Geometry-Aware Camera Localization for Bronchoscopy

Local AiDGX agent

arXiv:2608.07116v1 Announce Type: cross Abstract: Camera localization in bronchoscopy remains a challenging problem due to stringent accuracy requirements, real-time constraints, and limited training

Georeferencing Non-Gazetteered Place Names using Biological Specimen Records

Model ReleasesDGX agent

arXiv:2608.06884v1 Announce Type: cross Abstract: Biological specimen records collected by natural history institutions constitute a rich source of temporal geographic knowledge, capturing biodiversit

GPTKB 2.0: Browsing, Querying, and Auditing a Disambiguated LLM-Derived Knowledge Base

ResearchDGX agent

arXiv:2608.06992v1 Announce Type: cross Abstract: We present a web demo for exploring a large-scale disambiguated knowledge base (KB) materialized from a large language model (LLM). GPTKB 2.0 contains

H2AL: Hyperbolic Hierarchy-aware Aggregative Learning for Registration-based Few-shot Medical Image Segmentation

TutorialsDGX agent

arXiv:2608.07340v1 Announce Type: cross Abstract: Registration-based Few-shot medical image segmentation (RFMIS) aims to generate pseudo-labels for unlabeled images by warping a labeled image through

Harnessing the Synergy between LLM Agents and Knowledge Graphs for Urban Socioeconomic Prediction

AgentsDGX agent

arXiv:2411.00028v3 Announce Type: replace-cross Abstract: Socioeconomic prediction aims to leverage various urban data to predict the socioeconomic indicators of regions such as population and commerc

HarnessSafe: Evaluating Safety Across Persistent Carriers in Agent Harnesses

Model ReleasesDGX agent

arXiv:2608.06984v1 Announce Type: cross Abstract: Modern agent harnesses persist state across tasks and sessions through persistent carriers like memory, skills, tools, and shared artifacts. However,

Hidden Gauge Controls Feature Specialization in ReLU Networks

Model ReleasesDGX agent

arXiv:2608.06766v1 Announce Type: cross Abstract: Training changes a network's predictions while allocating task-relevant structure across its internal units. In an overparameterized ReLU network, sev

HLSmith: An Expert-Guided Agentic Framework for C/C++-to-HLS Translation

Model ReleasesDGX agent

arXiv:2608.06791v1 Announce Type: cross Abstract: Application-specific FPGA accelerators offer substantial performance and energy-efficiency gains across many application domains, but developing them

Homebot: A Personal AI Agent for Conversational Home Assistance and Automation

Local AiDGX agent

arXiv:2608.02254v2 Announce Type: replace Abstract: exttt{Homebot} is a locally deployable AI agent for conversational household assistance and automation. It accepts voice and instant-messaging reque

How Much AI Is in This Track? Quantifying the Proportion of AI-Generated Stems in Hybrid Music Mixtures

ApplicationsDGX agent

arXiv:2608.07285v1 Announce Type: cross Abstract: AI-generated music is increasingly used at the stem level, with producers integrating synthetic drums, basslines, or vocals alongside human-performed

How Much, Then Where: Credit-Conserving Action-to-Token Allocation for Multi-Turn Agent Reinforcement Learning

SafetyDGX agent

arXiv:2608.07118v1 Announce Type: new Abstract: Credit assignment in multi-turn agent reinforcement learning operates at two levels: assigning trajectory-level credit to actions and distributing each

Human-Centered Explainable AI for TinyML Edge Devices: A Pareto-Based Selection Framework with LLM-Guided Design

Local AiDGX agent

arXiv:2608.07091v1 Announce Type: cross Abstract: Edge Artificial Intelligence (Edge AI) enables the deployment of AI models directly on local edge devices, while such deployments are subject to stric

I Seek You in Videos: Identity-Conditioned Queries for Person-Centric Video Reasoning

Model ReleasesDGX agent

arXiv:2608.07417v1 Announce Type: cross Abstract: Real-world video reasoning often involves multimodal, multi-source inputs, whereas existing video reasoning tasks typically assume a simplified video-

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents

Model ReleasesDGX agent

arXiv:2608.06735v1 Announce Type: new Abstract: Reinforcement learning (RL) has achieved strong results in improving large language models (LLMs) on tasks with stationary, verifiable rewards, such as

Improving Attributed Long-form Question Answering with Intent Awareness

TutorialsDGX agent

arXiv:2603.27435v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly being used to generate comprehensive, knowledge-intensive reports. However, while these models a

Interaction Creates Dynamical AI Behavior Absent in Isolation

ResearchDGX agent

arXiv:2608.07457v1 Announce Type: new Abstract: What will happen when AI agents interact in daily life, e.g. when one AI starts bossing another around? We find a counterintuitive answer that opens new

International Transfer of Stochastic Cortical Self-Reconstruction

ResearchDGX agent

arXiv:2608.07092v1 Announce Type: cross Abstract: Stochastic cortical self-reconstruction (SCSR) enables personalized mapping of gray matter atrophy, a hallmark of neurodegenerative disorders such as

Interpretable reinforcement learning with decision-tree pruning

SafetyDGX agent

arXiv:2608.07151v1 Announce Type: cross Abstract: Reinforcement learning policies are difficult to inspect, but interpreting them is a prerequisite for trustworthiness. Converting a trained policy int

Interpretable Unsupervised Community Detection with LLM-Symbolized Structured Processes

Local AiDGX agent

arXiv:2608.06402v1 Announce Type: new Abstract: Community detection is a fundamental task in graph analytics that aims to identify cohesive groups of entities with similar behaviors or interests. Clas

INTRYGUE: Induction-Aware Entropy Gating for Reliable RAG Uncertainty Estimation

ResearchDGX agent

arXiv:2603.21607v2 Announce Type: replace Abstract: While retrieval-augmented generation (RAG) significantly improves the factual reliability of LLMs, it does not eliminate hallucinations, so robust u

Investigating Quantum-Embedded Transformers on Classical Datasets for Cross-Modality Classification

ResearchDGX agent

arXiv:2608.06846v1 Announce Type: cross Abstract: We test whether a parameterized quantum circuit (PQC) improves a hybrid quantum-classical model's performance on classical datasets, using an interfac

Kimi K2.5: Visual Agentic Intelligence

AgentsDGX agent

arXiv:2602.02276v2 Announce Type: replace-cross Abstract: We introduce Kimi K2.5, an open-source multimodal agentic model designed to advance general agentic intelligence. K2.5 emphasizes the joint op

KNOWPLAN: Knowledge-Driven AI Agents for Smart Degree Pathway Planning

ApplicationsDGX agent

arXiv:2608.06530v1 Announce Type: new Abstract: Planning a degree from official university sources requires solving two problems in order. The institution's curriculum must first be reconstructed from

KReF: Training-Free Retrieval for Long-Term Time-Series Forecasting and Predictive Uncertainty

SafetyDGX agent

arXiv:2608.06748v1 Announce Type: cross Abstract: Probabilistic long-term time-series forecasting commonly relies on trained models. Training-free conformal methods typically construct intervals aroun

Learning in Deep Networks under Dale's Constraint

Model ReleasesDGX agent

arXiv:2608.06963v1 Announce Type: new Abstract: Biologically plausible learning models aim to explain how neural circuits can implement effective learning under the constraints of real neurons. Althou

Learning to Predict Middle-Layer Attention in MLLMs for Visual Token Prunin

TutorialsDGX agent

arXiv:2608.06411v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) achieve strong performance across diverse vision-language tasks, but their efficiency is limited by the cost of

Learning to Walk With Less: A Dyna-Style Approach to Quadrupedal Locomotion

SafetyDGX agent

arXiv:2509.06296v2 Announce Type: replace-cross Abstract: Traditional on-policy reinforcement learning (RL) controllers for quadrupedal locomotion often suffer from low data efficiency, requiring mill

LifelongCrossNav: Persistent 3D Semantic Memory for Cross-Floor Multi-Object Navigation

Model ReleasesDGX agent

arXiv:2608.07079v1 Announce Type: cross Abstract: Object-goal navigation has made substantial progress in semantic perception and exploration, yet persistent memory for multi-object navigation and cro

LiFTER: A Grounded Neuro-Symbolic Microscope for Continuous-Time Dynamic Graph Forecasting

ResearchDGX agent

arXiv:2608.06765v1 Announce Type: new Abstract: Continuous-time dynamic graph models predict future links by compressing past interactions into neural states. Although effective for forecasting, this

← Previous
1…1920212223…354
Next →