AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
23 Apr 2026

OpenCLAW-P2P v6.0: Resilient Multi-Layer Persistence, Live Reference Verification, and Production-Scale Evaluation of Decentralized AI Peer Review

HardwareDGX agent

arXiv:2604.19792v1 Announce Type: new Abstract: This paper presents OpenCLAW-P2P v6.0, a comprehensive evolution of the decentralized collective-intelligence platform in which autonomous AI agents pub

ORPHEAS: A Cross-Lingual Greek-English Embedding Model for Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2604.20666v1 Announce Type: cross Abstract: Effective retrieval-augmented generation across bilingual Greek--English applications requires embedding models capable of capturing both domain-speci

OThink-SRR1: Search, Refine and Reasoning with Reinforced Learning for Large Language Models

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.19766v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) expands the knowledge of Large Language Models (LLMs), yet current static retrieval methods struggle with complex

pAI/MSc: ML Theory Research with Humans on the Loop

AgentsDGX agent

arXiv:2604.20622v1 Announce Type: new Abstract: We present pAI/MSc, an open-source, customizable, modular multi-agent system for academic research workflows. Our goal is not autonomous scientific idea

Participatory provenance as representational auditing for AI-mediated public consultation

SafetyDGX agent

arXiv:2604.20711v1 Announce Type: new Abstract: Artificial intelligence is increasingly deployed to synthesize large-scale public input in policy consultations and participatory processes. Yet no form

Peer-Preservation in Frontier Models

Model ReleasesDGX agent

arXiv:2604.19784v1 Announce Type: cross Abstract: Recently, it has been found that frontier AI models can resist their own shutdown, a behavior known as self-preservation. We extend this concept to th

Phase 1 Implementation of LLM-generated Discharge Summaries showing high Adoption in a Dutch Academic Hospital

ResearchDGX agent

arXiv:2604.19774v1 Announce Type: cross Abstract: Writing discharge summaries to transfer medical information is an important but time-consuming process that can be assisted by Large Language Models (

Physics-Enhanced Deep Learning for Proactive Thermal Runaway Forecasting in Li-Ion Batteries

SafetyDGX agent

arXiv:2604.20175v1 Announce Type: cross Abstract: Accurate prediction of thermal runaway in lithium-ion batteries is essential for ensuring the safety, efficiency, and reliability of modern energy sto

PipeMFL-240K: A Large-scale Dataset and Benchmark for Object Detection in Pipeline Magnetic Flux Leakage Imaging

Model ReleasesDGX agent

arXiv:2602.07044v2 Announce Type: replace-cross Abstract: Pipeline integrity is critical to industrial safety and environmental protection, with Magnetic Flux Leakage (MFL) detection being a primary n

PR-CAD: Progressive Refinement for Unified Controllable and Faithful Text-to-CAD Generation with Large Language Models

Model ReleasesDGX agent

arXiv:2604.19773v1 Announce Type: cross Abstract: The construction of CAD models has traditionally relied on labor-intensive manual operations and specialized expertise. Recent advances in large langu

Prism: An Evolutionary Memory Substrate for Multi-Agent Open-Ended Discovery

Model ReleasesDGX agent

arXiv:2604.19795v1 Announce Type: new Abstract: We introduce prism{} (extbf{P}robabilistic extbf{R}etrieval with extbf{I}nformation-extbf{S}tratified extbf{M}emory), an evolutionary memory substrate f

QuanForge: A Mutation Testing Framework for Quantum Neural Networks

Model ReleasesDGX agent

arXiv:2604.20706v1 Announce Type: cross Abstract: With the growing synergy between deep learning and quantum computing, Quantum Neural Networks (QNNs) have emerged as a promising paradigm by leveragin

QuantaAlpha: An Evolutionary Framework for LLM-Driven Alpha Mining

Model ReleasesDGX agent

arXiv:2602.07085v2 Announce Type: replace-cross Abstract: Financial markets are noisy and non-stationary, making alpha mining highly sensitive to noise in backtesting results and sudden market regime

Querying Inconsistent Prioritized Data with ORBITS: Algorithms, Implementation, and Experiments

ResearchDGX agent

arXiv:2202.07980v4 Announce Type: replace-cross Abstract: We investigate practical algorithms for inconsistency-tolerant query answering over prioritized knowledge bases, which consist of a logical th

Rabies diagnosis in low-data settings: A comparative study on the impact of data augmentation and transfer learning

ResearchDGX agent

arXiv:2604.19823v1 Announce Type: cross Abstract: Rabies remains a major public health concern across many African and Asian countries, where accurate diagnosis is critical for effective epidemiologic

ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability

Model ReleasesDGX agent

arXiv:2508.07050v3 Announce Type: replace-cross Abstract: Large Language Model (LLM) based listwise ranking has shown superior performance in many passage ranking tasks. With the development of Large

Recency Biased Causal Attention for Time-series Forecasting

SafetyDGX agent

arXiv:2502.06151v2 Announce Type: replace-cross Abstract: Recency bias is a useful inductive prior for sequential modeling: it emphasizes nearby observations and can still allow longer-range dependenc

Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem

SafetyDGX agent

arXiv:2604.20805v1 Announce Type: cross Abstract: The value alignment problem for artificial intelligence (AI) is often framed as a purely technical or normative challenge, sometimes focused on hypoth

Representational Alignment Across Model Layers and Brain Regions with Multi-Level Optimal Transport

SafetyDGX agent

arXiv:2510.01706v2 Announce Type: replace-cross Abstract: Standard representational similarity methods align each layer of a network to its best match in another independently, producing asymmetric re

Resolving space-sharing conflicts in road user interactions through uncertainty reduction: An active inference-based computational model

SafetyDGX agent

arXiv:2604.19838v1 Announce Type: new Abstract: Understanding how road users resolve space-sharing conflicts is important both for traffic safety and the safe deployment of autonomous vehicles. While

RSRCC: A Remote Sensing Regional Change Comprehension Benchmark Constructed via Retrieval-Augmented Best-of-N Ranking

Model ReleasesDGX agent

arXiv:2604.20623v1 Announce Type: cross Abstract: Traditional change detection identifies where changes occur, but does not explain what changed in natural language. Existing remote sensing change cap

Same Content, Different Answers: Cross-Modal Inconsistency in MLLMs

ResearchDGX agent

arXiv:2512.08923v2 Announce Type: replace Abstract: We introduce two new benchmarks REST and REST+ (Render-Equivalence Stress Tests) to enable systematic evaluation of cross-modal inconsistency in mul

Saying More Than They Know: A Framework for Quantifying Epistemic-Rhetorical Miscalibration in Large Language Models

ResearchDGX agent

arXiv:2604.19768v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit systematic miscalibration with rhetorical intensity not proportionate to epistemic grounding. This study tests th

Scalable AI Inference: Performance Analysis and Optimization of AI Model Serving

ApplicationsDGX agent

arXiv:2604.20420v1 Announce Type: cross Abstract: AI research often emphasizes model design and algorithmic performance, while deployment and inference remain comparatively underexplored despite being

SciCoQA: Quality Assurance for Scientific Paper--Code Alignment

Model ReleasesDGX agent

arXiv:2601.12910v3 Announce Type: replace-cross Abstract: Discrepancies between scientific papers and their code undermine reproducibility, a concern that grows as automated research agents scale scie

scpFormer: A Foundation Model for Unified Representation and Integration of the Single-Cell Proteomics

ResearchDGX agent

arXiv:2604.20003v1 Announce Type: cross Abstract: The integration of single-cell proteomic data is often hindered by the fragmented nature of targeted antibody panels. To address this limitation, we i

Seeing Further and Wider: Joint Spatio-Temporal Enlargement for Micro-Video Popularity Prediction

ApplicationsDGX agent

arXiv:2604.20311v1 Announce Type: cross Abstract: Micro-video popularity prediction (MVPP) aims to forecast the future popularity of videos on online media, which is essential for applications such as

Self-Awareness before Action: Mitigating Logical Inertia via Proactive Cognitive Awareness

Model ReleasesDGX agent

arXiv:2604.20413v1 Announce Type: new Abstract: Large language models perform well on many reasoning tasks, yet they often lack awareness of whether their current knowledge or reasoning state is compl

Self-Describing Structured Data with Dual-Layer Guidance: A Lightweight Alternative to RAG for Precision Retrieval in Large-Scale LLM Knowledge Navigation

Model ReleasesDGX agent

arXiv:2604.19777v1 Announce Type: cross Abstract: Large Language Models (LLMs) exhibit a well-documented positional bias when processing long input contexts: information in the middle of a context win

Self-Guided Plan Extraction for Instruction-Following Tasks with Goal-Conditional Reinforcement Learning

AgentsDGX agent

arXiv:2604.20601v1 Announce Type: new Abstract: We introduce SuperIgor, a framework for instruction-following tasks. Unlike prior methods that rely on predefined subtasks, SuperIgor enables a language

Semantic Prompting: Agentic Incremental Narrative Refinement through Spatial Semantic Interaction

SafetyDGX agent

arXiv:2604.19971v1 Announce Type: cross Abstract: Interactive spatial layouts empower users to synthesize information and organize findings for sensemaking. While Large Language Models (LLMs) can auto

Semantic Recall for Vector Search

ResearchDGX agent

arXiv:2604.20417v1 Announce Type: cross Abstract: We introduce Semantic Recall, a novel metric to assess the quality of approximate nearest neighbor search algorithms by considering only semantically

Separable Pathways for Causal Reasoning: How Architectural Scaffolding Enables Hypothesis-Space Restructuring in LLM Agents

ResearchDGX agent

arXiv:2604.20039v1 Announce Type: new Abstract: Causal discovery through experimentation and intervention is fundamental to robust problem solving. It requires not just updating beliefs within a fixed

Shift-Up: A Framework for Software Engineering Guardrails in AI-native Software Development -- Initial Findings

AgentsDGX agent

arXiv:2604.20436v1 Announce Type: cross Abstract: Generative AI (GenAI) is reshaping software engineering by shifting development from manual coding toward agent-driven implementation. While vibe codi

SkillGraph: Graph Foundation Priors for LLM Agent Tool Sequence Recommendation

Model ReleasesDGX agent

arXiv:2604.19793v1 Announce Type: new Abstract: LLM agents must select tools from large API libraries and order them correctly. Existing methods use semantic similarity for both retrieval and ordering

Skyline-First Traversal as a Control Mechanism for Multi-Criteria Graph Search

ResearchDGX agent

arXiv:2604.19807v1 Announce Type: new Abstract: In multi-criteria graph traversal, paths are compared via Pareto dominance, an ordering that identifies which paths are non-dominated, but says nothing

SMARTER: A Data-efficient Framework to Improve Toxicity Detection with Explanation via Self-augmenting Large Language Models

Model ReleasesDGX agent

arXiv:2509.15174v3 Announce Type: replace-cross Abstract: WARNING: This paper contains examples of offensive materials. To address the proliferation of toxic content on social media, we introduce SMAR

Soft-Label Governance for Distributional Safety in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2604.19752v1 Announce Type: cross Abstract: Multi-agent AI systems exhibit emergent risks that no single agent produces in isolation. Existing safety frameworks rely on binary classifications of

SolidCoder: Bridging the Mental-Reality Gap in LLM Code Generation through Concrete Execution

ResearchDGX agent

arXiv:2604.19825v1 Announce Type: cross Abstract: State-of-the-art code generation frameworks rely on mental simulation, where LLMs internally trace execution to verify correctness. We expose a fundam

SpeechParaling-Bench: A Comprehensive Benchmark for Paralinguistic-Aware Speech Generation

Model ReleasesDGX agent

arXiv:2604.20842v1 Announce Type: cross Abstract: Paralinguistic cues are essential for natural human-computer interaction, yet their evaluation in Large Audio-Language Models (LALMs) remains limited

SphUnc: Hyperspherical Uncertainty Decomposition and Causal Identification via Information Geometry

AgentsDGX agent

arXiv:2603.01168v2 Announce Type: replace-cross Abstract: Reliable decision-making in complex multi-agent systems requires calibrated predictions and interpretable uncertainty. We introduce SphUnc, a

Stabilising Generative Models of Attitude Change

ResearchDGX agent

arXiv:2604.19791v1 Announce Type: new Abstract: Attitude change - the process by which individuals revise their evaluative stances - has been explained by a set of influential but competing verbal the

Stateless Decision Memory for Enterprise AI Agents

ApplicationsDGX agent

arXiv:2604.20158v1 Announce Type: new Abstract: Enterprise deployment of long-horizon decision agents in regulated domains (underwriting, claims adjudication, tax examination) is dominated by retrieva

Statistics, Not Scale: Modular Medical Dialogue with Bayesian Belief Engine

AgentsDGX agent

arXiv:2604.20022v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous diagnostic agents, yet they conflate two fundamentally different capabilities: natural-l

Storm Surge Modeling, Bias Correction, Graph Neural Networks, Graph Convolution Networks

SafetyDGX agent

arXiv:2604.20688v1 Announce Type: cross Abstract: Storm surge forecasting remains a critical challenge in mitigating the impacts of tropical cyclones on coastal regions, particularly given recent tren

Supplement Generation Training for Enhancing Agentic Task Performance

Model ReleasesDGX agent

arXiv:2604.20727v1 Announce Type: cross Abstract: Training large foundation models for agentic tasks is increasingly impractical due to the high computational costs, long iteration cycles, and rapid o

Surrogate modeling for interpreting black-box LLMs in medical predictions

ApplicationsDGX agent

arXiv:2604.20331v1 Announce Type: cross Abstract: Large language models (LLMs), trained on vast datasets, encode extensive real-world knowledge within their parameters, yet their black-box nature obsc

SWE-chat: Coding Agent Interactions From Real Users in the Wild

AgentsDGX agent

arXiv:2604.20779v1 Announce Type: new Abstract: AI coding agents are being adopted at scale, yet we lack empirical evidence on how people actually use them and how much of their output is useful in pr

SweRank: Software Issue Localization with Code Ranking

Model ReleasesDGX agent

arXiv:2505.07849v2 Announce Type: replace-cross Abstract: Software issue localization, the task of identifying the precise code locations (files, classes, or functions) relevant to a natural language

Taint-Style Vulnerability Detection and Confirmation for Node.js Packages Using LLM Agent Reasoning

Model ReleasesDGX agent

arXiv:2604.20179v1 Announce Type: cross Abstract: The rapidly evolving Node.js ecosystem currently includes millions of packages and is a critical part of modern software supply chains, making vulnera

Task-Stratified Knowledge Scaling Laws for Post-Training Quantized Large Language Models

ResearchDGX agent

arXiv:2508.18609v4 Announce Type: replace-cross Abstract: Post-Training Quantization (PTQ) is a critical strategy for efficient Large Language Models (LLMs) deployment. However, existing scaling laws

Text Steganography with Dynamic Codebook and Multimodal Large Language Model

ResearchDGX agent

arXiv:2604.20269v1 Announce Type: cross Abstract: With the popularity of the large language models (LLMs), text steganography has achieved remarkable performance. However, existing methods still have

Text to model via SysML: Automated generation of dynamical system computational models from unstructured natural language text via enhanced System Modeling Language diagrams

ResearchDGX agent

arXiv:2507.06803v3 Announce Type: replace-cross Abstract: This paper contributes to speeding up the design and deployment of engineering dynamical systems by proposing a strategy for exploiting domain

The AI Telco Engineer: Toward Autonomous Discovery of Wireless Communications Algorithms

AgentsDGX agent

arXiv:2604.19803v1 Announce Type: new Abstract: Agentic AI is rapidly transforming the way research is conducted, from prototyping ideas to reproducing results found in the literature. In this paper,

The Existential Theory of Research: Why Discovery Is Hard

SafetyDGX agent

arXiv:2604.19810v1 Announce Type: new Abstract: Can scientific discovery be made arbitrarily easy by choosing the right representation, collecting enough data, and deploying sufficiently powerful algo

The Expense of Seeing: Attaining Trustworthy Multimodal Reasoning Within the Monolithic Paradigm

ResearchDGX agent

arXiv:2604.20665v1 Announce Type: cross Abstract: The rapid proliferation of Vision-Language Models (VLMs) is widely celebrated as the dawn of unified multimodal knowledge discovery but its foundation

The Model Says Walk: How Surface Heuristics Override Implicit Constraints in LLM Reasoning

Model ReleasesDGX agent

arXiv:2603.29025v2 Announce Type: replace-cross Abstract: Large language models systematically fail when a salient surface cue conflicts with an unstated feasibility constraint. We study this through

The OpenHands Software Agent SDK: A Composable and Extensible Foundation for Production Agents

Model ReleasesDGX agent

arXiv:2511.03690v2 Announce Type: replace-cross Abstract: Agents are now used widely in the process of software development, but building production-ready software engineering agents is a complex task

The Ratchet Effect in Silico through Interaction-Driven Cumulative Intelligence in Large Language Models

Model ReleasesDGX agent

arXiv:2507.21166v2 Announce Type: replace-cross Abstract: Human intelligence scales through cumulative cultural evolution (CCE), a ratchet process in which innovations are retained against entropic dr

The Tool-Overuse Illusion: Why Does LLM Prefer External Tools over Internal Knowledge?

SafetyDGX agent

arXiv:2604.19749v1 Announce Type: new Abstract: Equipping LLMs with external tools effectively addresses internal reasoning limitations. However, it introduces a critical yet under-explored phenomenon

← Previous
1…314315316317318…354
Next →