AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
30 Apr 2026

Grounding vs. Compositionality: On the Non-Complementarity of Reasoning in Neuro-Symbolic Systems

ResearchDGX agent

arXiv:2604.26521v1 Announce Type: new Abstract: Compositional generalization remains a foundational weakness of modern neural networks, limiting their robustness and applicability in domains requiring

HalluCiteChecker: A Lightweight Toolkit for Hallucinated Citation Detection and Verification in the Era of AI Scientists

Model ReleasesDGX agent

arXiv:2604.26835v1 Announce Type: cross Abstract: We introduce HalluCiteChecker, a toolkit for detecting and verifying hallucinated citations in scientific papers. While AI assistant technologies have

HER: Human-like Reasoning and Reinforcement Learning for LLM Role-playing

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.21459v4 Announce Type: replace-cross Abstract: LLM role-playing, i.e., using LLMs to simulate specific personas, has emerged as a key capability in various applications, such as companionsh

Hierarchical Multi-Persona Induction from User Behavioral Logs: Learning Evidence-Grounded and Truthful Personas

SafetyDGX agent

arXiv:2604.26120v1 Announce Type: new Abstract: Behavioral logs provide rich signals for user modeling, but are noisy and interleaved across diverse intents. Recent work uses LLMs to generate interpre

Human-in-the-Loop Benchmarking of Heterogeneous LLMs for Automated Competency Assessment in Secondary Level Mathematics

Model ReleasesDGX agent

arXiv:2604.26607v1 Announce Type: new Abstract: As Competency-Based Education (CBE) is gaining traction around the world, the shift from marks-based assessment to qualitative competency mapping is a m

Hybrid Diffusion for Simultaneous Symbolic and Continuous Planning

ResearchDGX agent

arXiv:2509.21983v2 Announce Type: replace-cross Abstract: Constructing robots to accomplish long-horizon tasks is a long-standing challenge within artificial intelligence. Approaches using generative

Identifying the Achilles' Heel: An Iterative Method for Dynamically Uncovering Factual Errors in Large Language Models

ApplicationsDGX agent

arXiv:2401.00761v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) like ChatGPT are foundational in various applications due to their extensive knowledge from pre-training and fine

ImproBR: Bug Report Improver Using LLMs

ApplicationsDGX agent

arXiv:2604.26142v1 Announce Type: cross Abstract: Bug tracking systems play a crucial role in software maintenance, yet developers frequently struggle with low-quality user-submitted reports that omit

Inference-Time Scaling of Verification: Self-Evolving Deep Research Agents via Test-Time Rubric-Guided Verification

SafetyDGX agent

arXiv:2601.15808v2 Announce Type: replace Abstract: Recent advances in Deep Research Agents (DRAs) are transforming automated knowledge discovery and problem-solving. While the majority of existing ef

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation

Model ReleasesDGX agent

arXiv:2511.20714v2 Announce Type: replace-cross Abstract: World models serve as core simulators for fields such as agentic AI, embodied AI, and gaming, capable of generating long, physically realistic

Integrating Weather Foundation Model and Satellite to Enable Fine-Grained Solar Irradiance Forecasting

ResearchDGX agent

arXiv:2603.14845v3 Announce Type: replace-cross Abstract: Accurate day-ahead solar irradiance forecasting is essential for integrating solar energy into the power grid. However, it remains challenging

Is Human-Like Text Liked by Humans? Multilingual Human Detection and Preference Against AI

ApplicationsDGX agent

arXiv:2502.11614v3 Announce Type: replace-cross Abstract: Prior studies have shown that distinguishing text generated by Large Language Models (LLMs) from human-written one is highly challenging for h

Language Diffusion Models are Associative Memories Capable of Retrieving Unseen Data

TutorialsDGX agent

arXiv:2604.26841v1 Announce Type: cross Abstract: When do language diffusion models memorize their training data, and how to quantitatively assess their true generative regime? We address these questi

LATTICE: Evaluating Decision Support Utility of Crypto Agents

Model ReleasesDGX agent

arXiv:2604.26235v1 Announce Type: cross Abstract: We introduce LATTICE, a benchmark for evaluating the decision support utility of crypto agents in realistic user-facing scenarios. Prior crypto agent

Learning to Ask: When LLM Agents Meet Unclear Instruction

Model ReleasesDGX agent

arXiv:2409.00557v4 Announce Type: replace-cross Abstract: Equipped with the capability to call functions, modern large language models (LLMs) can leverage external tools for addressing a range of task

Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Use

AgentsDGX agent

arXiv:2602.20426v2 Announce Type: replace Abstract: While most efforts to improve LLM-based tool-using agents focus on the agent itself - through larger models, better prompting, or fine-tuning - agen

Lifting Embodied World Models for Planning and Control

SafetyDGX agent

arXiv:2604.26182v1 Announce Type: cross Abstract: World models of embodied agents predict future observations conditioned on an action taken by the agent. For complex embodiments, action spaces are hi

Lightweight Quantum Agent for Edge Systems: Joint PQC and NOMA Resource Allocation

AgentsDGX agent

arXiv:2604.25980v1 Announce Type: cross Abstract: In the context of quantum secure scenarios, existing research on mobile edge devices and intelligent computing and edge (ICE) systems based on the Non

LLM Psychosis: A Theoretical and Diagnostic Framework for Reality-Boundary Failures in Large Language Models

Model ReleasesDGX agent

arXiv:2604.25934v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) as interactive agents has exposed a category of behavioral failure that prevailing terminology, princip

Lunguage: A Benchmark for Structured and Sequential Chest X-ray Interpretation

Model ReleasesDGX agent

arXiv:2505.21190v2 Announce Type: replace-cross Abstract: Radiology reports convey detailed clinical observations and capture diagnostic reasoning that evolves over time. However, existing evaluation

Lyapunov-Guided Self-Alignment: Test-Time Adaptation for Offline Safe Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.26516v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) agents often fail when deployed, as the gap between training datasets and real environments leads to unsafe behavi

M2R2: MultiModal Robotic Representation for Temporal Action Segmentation

ResearchDGX agent

arXiv:2504.18662v3 Announce Type: replace-cross Abstract: Temporal action segmentation (TAS) has long been a key area of research in both robotics and computer vision. In robotics, algorithms have pri

MappingEvolve: LLM-Driven Code Evolution for Technology Mapping

AgentsDGX agent

arXiv:2604.26591v1 Announce Type: cross Abstract: Technology mapping is a critical yet challenging stage in logic synthesis. While Large Language Models (LLMs) have been applied to generate optimizati

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution

SafetyDGX agent

arXiv:2604.26283v1 Announce Type: cross Abstract: High-precision medical diagnosis relies not only on static imaging features but also on the implicit diagnostic memory experts instantly invoke during

MemOVCD: Training-Free Open-Vocabulary Change Detection via Cross-Temporal Memory Reasoning and Global-Local Adaptive Rectification

ResearchDGX agent

arXiv:2604.26774v1 Announce Type: cross Abstract: Open-vocabulary change detection aims to identify semantic changes in bi-temporal remote sensing images without predefined categories. Recent methods

MetaSR: Content-Adaptive Metadata Orchestration for Generative Super-Resolution

TutorialsDGX agent

arXiv:2604.26244v1 Announce Type: cross Abstract: We study generative super-resolution (SR) in real-world scenarios where content and degradations vary across domains, genres, and segments. For exampl

Mini-Batch Class Composition Bias in Link Prediction

SafetyDGX agent

arXiv:2604.25978v1 Announce Type: cross Abstract: Prior work on node classification has shown that Graph Neural Networks (GNNs) can learn representations that transfer across graphs, when underlying g

MINOS: A Multimodal Evaluation Model for Bidirectional Generation Between Image and Text

SafetyDGX agent

arXiv:2506.02494v2 Announce Type: replace-cross Abstract: Evaluation is important for multimodal generation tasks, while traditional multimodal evaluation metrics suffer from several limitations. With

Momentum-Conserving Graph Neural Networks for Deformable Objects

ResearchDGX agent

arXiv:2604.26097v1 Announce Type: cross Abstract: Graph neural networks (GNNs) have emerged as a versatile and efficient option for modeling the dynamic behavior of deformable materials. While GNNs ge

Multi-Stage Bi-Atrial Segmentation Framework from 3D Late Gadolinium-Enhanced MRI using V-Net Family Models

ResearchDGX agent

arXiv:2604.26251v1 Announce Type: cross Abstract: We report our multi-stage framework designed for the problem of multi-class bi-atrial segmentation from 3D late gadolinium-enhanced (LGE) MRI of the h

Naamah: A Large Scale Synthetic Sanskrit NER Corpus via DBpedia Seeding and LLM Generation

Model ReleasesDGX agent

arXiv:2604.26456v1 Announce Type: cross Abstract: The digitisation of classical Sanskrit literature is impeded by a scarcity of annotated resources, particularly for Named Entity Recognition. While re

Networks of Causal Abstractions: A Sheaf-theoretic Framework

AgentsDGX agent

arXiv:2509.25236v3 Announce Type: replace Abstract: A core challenge in causal artificial intelligence is the principled coordination of multiple, imperfect, and subjective causal perspectives arising

OMEGA: Optimizing Machine Learning by Evaluating Generated Algorithms

Model ReleasesDGX agent

arXiv:2604.26211v1 Announce Type: new Abstract: In order to automate AI research we introduce a full, end-to-end framework, OMEGA: Optimizing Machine learning by Evaluating Generated Algorithms, that

Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents

SafetyDGX agent

arXiv:2505.02077v2 Announce Type: replace-cross Abstract: AI agents are beginning to interact with each other directly and across internet platforms and physical environments, creating security challe

Open Problems in Frontier AI Risk Management

SafetyDGX agent

arXiv:2604.25982v1 Announce Type: cross Abstract: Frontier AI both amplifies existing risks and introduces qualitatively novel challenges. Not only is there a notable lack of stable scientific consens

Operating-Layer Controls for Onchain Language-Model Agents Under Real Capital

SafetyDGX agent

arXiv:2604.26091v1 Announce Type: new Abstract: We study reliability in autonomous language-model agents that translate user mandates into validated tool actions under real capital. The setting is DX

Option-Order Randomisation Reveals a Distributional Position Attractor in Prompted Sandbagging

Model ReleasesDGX agent

arXiv:2604.26206v1 Announce Type: cross Abstract: A predecessor pilot (Cacioli, 2026) found that Llama-3-8B implements prompted sandbagging as positional collapse rather than answer avoidance. However

OT Score: An OT based Confidence Score for Prototype-Assisted Source Free Unsupervised Domain Adaptation

SafetyDGX agent

arXiv:2505.11669v3 Announce Type: replace-cross Abstract: We address the computational and theoretical limitations of current distributional alignment methods for source-free unsupervised domain adapt

OxyGent: Making Multi-Agent Systems Modular, Observable, and Evolvable via Oxy Abstraction

AgentsDGX agent

arXiv:2604.25602v2 Announce Type: replace Abstract: Deploying production-ready multi-agent systems (MAS) in complex industrial environments remains challenging due to limitations in scalability, obser

PATCH: Learnable Tile-level Hybrid Sparsity for LLMs

Model ReleasesDGX agent

arXiv:2509.23410v4 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver impressive performance but incur prohibitive memory and compute costs at deployment. Model pruning is an

PBiLoss: Popularity-Aware Regularization to Improve Fairness in Graph-Based Recommender Systems

SafetyDGX agent

arXiv:2507.19067v2 Announce Type: replace-cross Abstract: Recommender systems based on graph neural networks (GNNs) have been proved to perform well on user-item interactions. However, they commonly s

Persuadability and LLMs as Legal Decision Tools

ApplicationsDGX agent

arXiv:2604.26233v1 Announce Type: new Abstract: As Large Language Models (LLMs) are proposed as legal decision assistants, and even first-instance decision-makers, across a range of judicial and admin

Planar Gaussian Splatting with Bilinear Spatial Transformer for Wireless Radiance Field Reconstruction

TutorialsDGX agent

arXiv:2604.25945v1 Announce Type: cross Abstract: Wireless radiance field (WRF) reconstruction aims to learn a continuous, queryable representation of radio frequency characteristics over 3D space and

Preserving Disagreement: Architectural Heterogeneity and Coherence Validation in Multi-Agent Policy Simulation

Model ReleasesDGX agent

arXiv:2604.26561v1 Announce Type: cross Abstract: Multi-agent deliberation systems using large language models (LLMs) are increasingly proposed for policy simulation, yet they suffer from artificial c

Privacy-Preserving Federated Learning Framework for Distributed Chemical Process Optimization

Local AiDGX agent

arXiv:2604.26073v1 Announce Type: cross Abstract: Industrial chemical plants often operate under strict data confidentiality constraints, making centralized data-driven process modeling difficult. Fed

Probe-then-Plan: Environment-Aware Planning for Industrial E-commerce Search

SafetyDGX agent

arXiv:2603.15262v2 Announce Type: replace Abstract: Modern e-commerce search is evolving to resolve complex user intents. While Large Language Models (LLMs) offer strong reasoning, existing LLM-based

Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.26508v1 Announce Type: cross Abstract: Deploying Vision-Language Models (VLMs) on edge devices remains challenging due to their substantial computational and memory demands, which exceed th

Provable Coordination for LLM Agents via Message Sequence Charts

Local AiDGX agent

arXiv:2604.17612v2 Announce Type: replace-cross Abstract: Multi-agent systems built on large language models (LLMs) are difficult to reason about. Coordination errors such as deadlocks or type-mismatc

q3-MuPa: Quick, Quiet, Quantitative Multi-Parametric MRI using Physics-Informed Diffusion Models

ResearchDGX agent

arXiv:2512.23726v2 Announce Type: replace-cross Abstract: The 3D fast silent multi-parametric mapping sequence with zero echo time (MuPa-ZTE) is a novel quantitative MRI (qMRI) acquisition that enable

QERNEL: a Scalable Large Electron Model

Model ReleasesDGX agent

arXiv:2604.26018v1 Announce Type: cross Abstract: We introduce QERNEL, a foundational neural wavefunction that variationally solves families of parameterized many-electron Hamiltonians and captures th

Quantum Gatekeeper: Multi-Factor Context-Bound Image Steganography with VQC Based Key Derivation on Quantum Hardware

ResearchDGX agent

arXiv:2604.26413v1 Announce Type: cross Abstract: This paper presents Quantum Gatekeeper, a context-bound image steganography framework where successful payload recovery depends on both cryptographic

Qvine: Vine Structured Quantum Circuits for Loading High Dimensional Distributions

ApplicationsDGX agent

arXiv:2604.26213v1 Announce Type: cross Abstract: Loading high dimensional distributions is an important task for utilizing quantum computers on applications ranging from machine learning to finance.

QYOLO: Lightweight Object Detection via Quantum Inspired Shared Channel Mixing

Model ReleasesDGX agent

arXiv:2604.26435v1 Announce Type: cross Abstract: The rapid advancement of object detection architectures has positioned single stage detectors as the dominant solution for real-time visual perception

RaMP: Runtime-Aware Megakernel Polymorphism for Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2604.26039v1 Announce Type: cross Abstract: The optimal kernel configuration for Mixture-of-Experts (MoE) inference depends on both batch size and the expert routing distribution, yet production

Random Cloud: Finding Minimal Neural Architectures Without Training

Model ReleasesDGX agent

arXiv:2604.26830v1 Announce Type: cross Abstract: I propose the Random Cloud method, a training-free approach to neural architecture search that discovers minimal feedforward network topologies throug

RE-MCDF: Closed-Loop Multi-Expert LLM Reasoning for Knowledge-Grounded Clinical Diagnosis

AgentsDGX agent

arXiv:2602.01297v3 Announce Type: replace Abstract: Electronic medical records (EMRs), particularly in neurology, are inherently heterogeneous, sparse, and noisy, which poses significant challenges fo

Recent Advances in mm-Wave and Sub-THz/THz Oscillators for FutureG Technologies

ResearchDGX agent

arXiv:2604.26903v1 Announce Type: cross Abstract: This paper provides a concise yet comprehensive review of recent advancements in millimeter-wave (mm-wave) oscillators below 100 GHz and sub-terahertz

ReLoop: Structured Modeling and Behavioral Verification for Reliable LLM-Based Optimization

Model ReleasesDGX agent

arXiv:2602.15983v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can translate natural language into optimization code, but silent failures pose a critical risk: code that execut

Resume-ing Control: (Mis)Perceptions of Agency Around GenAI Use in Recruiting Workflows

ResearchDGX agent

arXiv:2604.26851v1 Announce Type: cross Abstract: When generative AI (genAI) systems are used in high-stakes decision-making, its recommended role is to aid, rather than replace, human decision-making

Rethinking KV Cache Eviction via a Unified Information-Theoretic Objective

ResearchDGX agent

arXiv:2604.25975v1 Announce Type: cross Abstract: Key-value (KV) caching is essential for large language model inference, yet its memory overhead poses a critical bottleneck for long-context generatio

← Previous
1…293294295296297…354
Next →