AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
27 May 2026

Model Merging on Loss Landscape: A Geometry Perspective

ResearchDGX agent

arXiv:2605.26693v1 Announce Type: cross Abstract: Model merging offers a promising avenue for knowledge integration and parallel development without retraining. Yet, existing methods either ignore the

Modeling Agentic Technical Debt and Stochastic Tax: A Standalone Framework for Measurement, Simulation, and Dashboarding

AgentsDGX agent

arXiv:2605.27320v1 Announce Type: new Abstract: Agentic AI systems combine probabilistic reasoning with delegated action through tools, context, memory, orchestration, and external workflow integratio

Modeling Dynamic Mixtures of Time-Delay Systems from Streaming Time Series

Model ReleasesDGX agent

arXiv:2605.26191v1 Announce Type: cross Abstract: This research addresses the problem of adaptive modeling in time-series data streams with clear input-output relationships. This problem is challengin


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Modernising Reinforcement Learning-Based Navigation for Embodied Semantic Scene Graph Generation

SafetyDGX agent

arXiv:2603.25415v2 Announce Type: replace Abstract: Semantic world models enable embodied agents to reason about objects, relations, and spatial context beyond purely geometric representations. In Org

Monte Carlo Permutation Search

SafetyDGX agent

arXiv:2510.06381v2 Announce Type: replace-cross Abstract: We propose Monte Carlo Permutation Search (MCPS), a general-purpose Monte Carlo Tree Search (MCTS) algorithm that improves upon the GRAVE algo

More Expressive Feedforward Layers: Part I. Token-Adaptive Mixing of Activations

Model ReleasesDGX agent

arXiv:2605.26647v1 Announce Type: cross Abstract: Feedforward network (FFN) layers account for a large fraction of parameters and nonlinear expressivity in Transformer-based large language models (LLM

Multi-Agent Causal Discovery Using Large Language Models

Model ReleasesDGX agent

arXiv:2407.15073v4 Announce Type: replace Abstract: Causal discovery aims to identify causal relationships between variables and is a fundamental problem across the sciences. Traditional statistical c

Multi-Stakeholder LLM Alignment: Decomposing Estimation from Aggregation

SafetyDGX agent

arXiv:2605.26878v1 Announce Type: new Abstract: Multi-stakeholder tasks require one output to satisfy users with conflicting preferences. Holistic LLM judges conflate utility estimation and utility ag

MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation

AgentsDGX agent

arXiv:2605.27366v1 Announce Type: new Abstract: Large language model (LLM) agents rely on reusable skills to solve complex tasks. However, existing skill creation approaches treat skills as isolated a

Natural Language Query to Configuration for Retrieval Agents

ResearchDGX agent

arXiv:2605.27361v1 Announce Type: new Abstract: Modern retrieval agents expose many configuration choices -- LLM, retriever, number of documents, number of hops, and synthesis strategy -- each shaping

Negligible in Size, Significant in Effect: On Scale Vectors in Large Language Models

Model ReleasesDGX agent

arXiv:2605.26895v1 Announce Type: cross Abstract: Normalization layers in modern large language models (LLMs) consist of a deterministic normalization operation and a learnable scale vector. While the

Neuro-Symbolic Verification of LLM Outputs for Data-Sensitive Domains (extended preprint)

SafetyDGX agent

arXiv:2605.26942v1 Announce Type: new Abstract: LLMs deployed in high-stakes domains face fundamental reliability challenges: hallucinations, inconsistencies, and privacy vulnerabilities introduce una

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning

Model ReleasesDGX agent

arXiv:2505.17163v2 Announce Type: replace-cross Abstract: Recent advancements in multimodal slow-thinking systems have demonstrated remarkable performance across various visual reasoning tasks. Howeve

Olaf-World: Orienting Latent Actions for Video World Modeling

SafetyDGX agent

arXiv:2602.10104v2 Announce Type: replace-cross Abstract: Scaling action-controllable world models is limited by the scarcity of action labels. While latent action learning promises to extract control

Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2603.16654v2 Announce Type: replace-cross Abstract: Evaluating the reasoning abilities of large language models (LLMs) solely from final answers can obscure failures in intermediate steps, espec

OMD-GraphRAG: Enhancing GraphRAG with Ontology-Guided Extraction, Multi-Dimensional Clustering and Dual-Channel Fusion

Model ReleasesDGX agent

arXiv:2603.25152v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems face significant challenges in complex reasoning, multi-hop queries, and domain-specific QA. While exis

OmniToM: Benchmarking Theory of Mind in LLMs via Explicit Belief Modeling

Model ReleasesDGX agent

arXiv:2605.26322v1 Announce Type: new Abstract: Theory of Mind (ToM), the ability to infer others' knowledge, intentions, and emotions, is commonly evaluated in large language models (LLMs) using end-

On the Detection of Commutative Factors in Factor Graphs: Necessary and Sufficient Conditions

ResearchDGX agent

arXiv:2605.26908v1 Announce Type: new Abstract: Exploiting the indistinguishability of objects in a probabilistic graphical model such as a factor graph is key to lifted probabilistic inference algori

On the Error-Correcting Effects of Stochasticity in Discrete Diffusion

ResearchDGX agent

arXiv:2605.26582v1 Announce Type: cross Abstract: Discrete diffusion models achieve strong performance in text and image generation, but their inference remains slow and must inherently balance sampli

On the Push-Based Asynchronous Federated Learning: A Bias-Correction Aggregation Approach

SafetyDGX agent

arXiv:2605.26162v1 Announce Type: cross Abstract: Asynchronous decentralized federated learning (ADFL) eliminates central coordination and global synchronization, making it attractive for large-scale

ORCA: An End-to-End Interactive Copilot for Optimized Root Cause Analysis

TutorialsDGX agent

arXiv:2605.27022v1 Announce Type: new Abstract: Causal analysis is a crucial task in many domains, including manufacturing, social science, and medicine. However, despite recent progress, the conceptu

ORLoopBench: Solver-in-the-Loop Benchmarks for Self-Correction and Behavioral Rationality in Operations Research

Model ReleasesDGX agent

arXiv:2601.21008v3 Announce Type: replace-cross Abstract: Operations Research practitioners debug infeasible models through an iterative process: inspecting Irreducible Infeasible Subsystems ( IIS), i

Pair-In, Pair-Out: Latent Multi-Token Prediction for Efficient LLMs

SafetyDGX agent

arXiv:2605.27255v1 Announce Type: cross Abstract: Long chain-of-thought reasoning has made autoregressive decoding the dominant inference cost of modern large language models. Existing methods target

ParsVoice: A Large-Scale Multi-Speaker Persian Speech Corpus for Text-to-Speech Synthesis

ResearchDGX agent

arXiv:2510.10774v3 Announce Type: replace-cross Abstract: Persian remains substantially underrepresented in open speech-text resources, limiting progress in multi-speaker text-to-speech (TTS), speech-

PaTAS: A Framework for Trust Propagation in Neural Networks Using Subjective Logic

Model ReleasesDGX agent

arXiv:2511.20586v4 Announce Type: replace Abstract: Trustworthiness has become a key requirement for the deployment of artificial intelligence systems in safety-critical applications. Conventional eva

Periodic Topological Deep Learning for Polymer Design and Discovery

Model ReleasesDGX agent

arXiv:2605.26833v1 Announce Type: cross Abstract: Polymers underpin applications across energy, healthcare, and materials science, yet their vast chemical space makes systematic discovery challenging.

Persistent AI Agents in Academic Research: A Single-Investigator Implementation Case Study

Local AiDGX agent

arXiv:2605.26870v1 Announce Type: cross Abstract: Background: Large language models are typically evaluated as models, benchmarks, or short conversational episodes. Less is known about what happens wh

Persona Generators: Generating Diverse Synthetic Personas for Arbitrary Contexts

AgentsDGX agent

arXiv:2602.03545v2 Announce Type: replace Abstract: Evaluating AI systems that interact with humans requires understanding their behavior across diverse user populations, but collecting representative

Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History

Model ReleasesDGX agent

arXiv:2602.17003v2 Announce Type: replace-cross Abstract: Large language models have advanced web agents, yet current agents lack personalization capabilities. Since users rarely specify every detail

Personalized Generative Models for Contextual Debiasing

TutorialsDGX agent

arXiv:2605.26353v1 Announce Type: cross Abstract: Different visual patterns appear with different frequencies in the world: e.g., beach balls appear on sand more often than they do on a road. These st

Personalizing Embodied Multimodal Large Language Model Agents over Long-term User Interactions

AgentsDGX agent

arXiv:2605.26256v1 Announce Type: new Abstract: Multimodal large language model (MLLM)-based embodied agents have shown strong potential for solving complex tasks in physical environments. However, pe

Phase-Type Variational Autoencoders for Heavy-Tailed Data

ApplicationsDGX agent

arXiv:2603.01800v2 Announce Type: replace-cross Abstract: Heavy-tailed distributions are ubiquitous in real-world data, where rare but extreme events dominate risk and variability. However, standard V

'PhyWorldBench': A Comprehensive Evaluation of Physical Realism in Text-to-Video Models

Model ReleasesDGX agent

arXiv:2507.13428v3 Announce Type: replace-cross Abstract: Video generation models have achieved remarkable progress in creating high-quality, photorealistic content. However, their ability to accurate

PICACO: Pluralistic In-Context Value Alignment of LLMs via Total Correlation Optimization

SafetyDGX agent

arXiv:2507.16679v3 Announce Type: replace-cross Abstract: In-Context Learning has shown great potential for aligning Large Language Models (LLMs) with human values, helping reduce harmful outputs and

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis

Model ReleasesDGX agent

arXiv:2605.27258v1 Announce Type: cross Abstract: Building state-of-the-art text-to-speech (TTS) systems typically demands millions of hours of proprietary data and complex multi-stage architectures,

PitchBench: Measuring Pitch Hearing in Audio-Language Models

Model ReleasesDGX agent

arXiv:2605.26176v1 Announce Type: cross Abstract: Audio-language models (ALMs) are increasingly used in real-world applications that require understanding music, from music tutoring and transcription

Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning

Local AiDGX agent

arXiv:2510.01833v2 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong reasoning abilities via Chain-of-Thought (CoT), but their token-level generation encourages local de

Planning Neural Dynamics with Lie Group Embedding through Supervised Projective Manifold Learning

TutorialsDGX agent

arXiv:2605.26167v1 Announce Type: cross Abstract: We propose Lie group embedded dynamical neural networks (LieEDNN) and the corresponding learning algorithms based on gradient descent and metric proje

Plans for Evaluating Structured Generative Search Summaries

ResearchDGX agent

arXiv:2605.26400v1 Announce Type: cross Abstract: We propose a framework for evaluating structured generative search summaries that are placed atop organic web search results. A structured summary, ge

PolyFusionAgent: A Multimodal Foundation Model and Autonomous AI Assistant for Polymer Property Prediction and Inverse Design

AgentsDGX agent

arXiv:2605.26543v1 Announce Type: new Abstract: Polymer discovery is central to fields ranging from energy storage to biomedicine, but it is hindered by an astronomically large chemical design space a

Position: AI Safety Requires Effective Controllability

Model ReleasesDGX agent

arXiv:2605.27117v1 Announce Type: new Abstract: AI safety is still largely framed as alignment: training models to follow human preferences, safety policies, and normative constraints. That framing ha

Practical Anonymous Two-Party Gradient Boosting Decision Tree

SafetyDGX agent

arXiv:2605.26903v1 Announce Type: cross Abstract: Structured data is well handled by gradient-boosted decision trees (GBDT), which are usually trained on vertically partitioned features across mutuall

Pretraining Data Exposure in Large Language Models: A Survey of Membership Inference, Data Contamination, and Security Implications

ResearchDGX agent

arXiv:2605.26133v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become the predominant paradigm in NLP, advancing both research and industry. As model sizes and pretraining data gr

Prospective evaluation of multimodal respiratory failure prediction: Do chest X-rays improve performance beyond EHR signals?

ResearchDGX agent

arXiv:2605.26255v1 Announce Type: cross Abstract: Early prediction of respiratory failure is critical for timely clinical intervention in intensive care units. Existing electronic health record (EHR)-

Qiskit QuantumKatas: Adapting Microsoft's Quantum Computing exercises for LLM evaluation

Model ReleasesDGX agent

arXiv:2605.27210v1 Announce Type: cross Abstract: We adapt Microsoft's QuantumKatas -- a well-established quantum computing curriculum -- from Q# to Qiskit, the most widely-adopted quantum computing f

Qrita: High-performance Top-k and Top-p using Pivot-based Truncation and Selection

HardwareDGX agent

arXiv:2602.01518v2 Announce Type: replace Abstract: Despite their importance in model sampling, efficient implementation of Top-k and Top-p algorithms for large vocabularies remains a significant chal

QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents

AgentsDGX agent

arXiv:2605.27068v1 Announce Type: cross Abstract: Social deduction games have become a popular testbed for probing reasoning, deception, coordination, and belief modeling in Large Language Model (LLM)

Quantized Keys Steal Attention: Bias Correction for KV-Cache Compression in Video Diffusion

SafetyDGX agent

arXiv:2605.26266v1 Announce Type: cross Abstract: Chunk-wise autoregressive video diffusion models rely on a KV cache of previously generated chunks to avoid redundant computation, but this cache quic

Query Symbolically or Retrieve Semantically? A Dataset and Method for Semi-Structured Question Answering

Model ReleasesDGX agent

arXiv:2605.27164v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems for question answering typically retrieve evidence by semantic similarity between the query and document ch

Querying and Repairing Inconsistent Prioritized Knowledge Bases: Complexity Analysis and Links with Abstract Argumentation

ResearchDGX agent

arXiv:2003.05746v4 Announce Type: replace-cross Abstract: In this paper, we explore the issue of inconsistency handling over prioritized knowledge bases (KBs), which consist of an ontology, a set of f

RAGEAR: Retrieval-Augmented Graph-Enhanced Academic Recommender

TutorialsDGX agent

arXiv:2605.26819v1 Announce Type: cross Abstract: We present RAGEAR (Retrieval-Augmented Graph-Enhanced Academic Recommender), a neurosymbolic recommender system for academic course recommendation. RA

Ratio-Variance Regularized Policy Optimization

Local AiDGX agent

arXiv:2605.26784v1 Announce Type: cross Abstract: Standard on-policy reinforcement learning relies on heuristic clipping to enforce trust regions, but this mechanism imposes a severe cost by indiscrim

Real-Time Progress Prediction in Reasoning Language Models

AgentsDGX agent

arXiv:2506.23274v4 Announce Type: replace-cross Abstract: Recent reasoning language models, particularly those that employ long latent chains of thought, achieve strong performance on complex agentic

Reasoning, Code, or Both? How Large Language Models Handle Variations in Math Questions

Model ReleasesDGX agent

arXiv:2605.26414v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve impressive accuracy on mathematical reasoning benchmarks, yet their performance drops when problems are modified wi

Reasoning Depth and Environment Complexity: A Controlled Study of RLVR Data Allocation across Logical Reasoning Tasks

ApplicationsDGX agent

arXiv:2605.26934v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become central to post-training reasoning models, yet a key limitation of existing studies i

ReasonOps: A Unified Operational Paradigm for Trustworthy Verified LLM Reasoning

SafetyDGX agent

arXiv:2605.27014v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence from primarily generative systems into increasingly capable reasoning agents. Re

ReCA: Multi-Shot Long Video Extrapolation via Recursive Context Allocation

ResearchDGX agent

arXiv:2605.26525v1 Announce Type: cross Abstract: Minute-scale cinematic video generation is a central challenge for generative video models. Existing paradigms address only fragments of this challeng

Recon: Reconstruction-Guided Reasoning Synthesis for User Modeling

ResearchDGX agent

arXiv:2605.26969v1 Announce Type: cross Abstract: User modeling aims to use language models (LMs) to mimic an individual's behavior from a corpus of past context-action pairs (e.g., conversation turns

Reconstructing Multi-Scale Physical Fields from Extremely Sparse Measurements with an Autoencoder-Diffusion Cascade

ResearchDGX agent

arXiv:2512.01572v3 Announce Type: replace-cross Abstract: Extreme sensor sparsity makes full-field reconstruction a fundamentally ill-posed problem in scientific sensing,where the goal is to infer phy

Recursive Flow Matching

ResearchDGX agent

arXiv:2605.26535v1 Announce Type: cross Abstract: Generative models have emerged as a powerful paradigm for solving physics systems and modeling complex spatiotemporal dynamics. However, achieving hig

← Previous
1…207208209210211…358
Next →