AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlog
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Model Merging on Loss Landscape: A Geometry Perspective

DGX agent

arXiv:2605.26693v1 Announce Type: cross Abstract: Model merging offers a promising avenue for knowledge integration and parallel development without retraining. Yet, existing methods either ignore the

researcharxiv-cs-ai
27 May 2026
Agents
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Modeling Agentic Technical Debt and Stochastic Tax: A Standalone Framework for Measurement, Simulation, and Dashboarding

DGX agent

arXiv:2605.27320v1 Announce Type: new Abstract: Agentic AI systems combine probabilistic reasoning with delegated action through tools, context, memory, orchestration, and external workflow integratio

agentsarxiv-cs-ai
27 May 2026
Model Releases

Modeling Dynamic Mixtures of Time-Delay Systems from Streaming Time Series

DGX agent

arXiv:2605.26191v1 Announce Type: cross Abstract: This research addresses the problem of adaptive modeling in time-series data streams with clear input-output relationships. This problem is challengin

model-releasesarxiv-cs-ai
27 May 2026
Safety

Modernising Reinforcement Learning-Based Navigation for Embodied Semantic Scene Graph Generation

DGX agent

arXiv:2603.25415v2 Announce Type: replace Abstract: Semantic world models enable embodied agents to reason about objects, relations, and spatial context beyond purely geometric representations. In Org

safetyarxiv-cs-ai
27 May 2026
Safety

Monte Carlo Permutation Search

DGX agent

arXiv:2510.06381v2 Announce Type: replace-cross Abstract: We propose Monte Carlo Permutation Search (MCPS), a general-purpose Monte Carlo Tree Search (MCTS) algorithm that improves upon the GRAVE algo

safetyarxiv-cs-ai
27 May 2026
Model Releases

More Expressive Feedforward Layers: Part I. Token-Adaptive Mixing of Activations

DGX agent

arXiv:2605.26647v1 Announce Type: cross Abstract: Feedforward network (FFN) layers account for a large fraction of parameters and nonlinear expressivity in Transformer-based large language models (LLM

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Multi-Agent Causal Discovery Using Large Language Models

DGX agent

arXiv:2407.15073v4 Announce Type: replace Abstract: Causal discovery aims to identify causal relationships between variables and is a fundamental problem across the sciences. Traditional statistical c

model-releasesarxiv-cs-ai
27 May 2026
Safety

Multi-Stakeholder LLM Alignment: Decomposing Estimation from Aggregation

DGX agent

arXiv:2605.26878v1 Announce Type: new Abstract: Multi-stakeholder tasks require one output to satisfy users with conflicting preferences. Holistic LLM judges conflate utility estimation and utility ag

safetyarxiv-cs-ai
27 May 2026
Agents

MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation

DGX agent

arXiv:2605.27366v1 Announce Type: new Abstract: Large language model (LLM) agents rely on reusable skills to solve complex tasks. However, existing skill creation approaches treat skills as isolated a

agentsarxiv-cs-ai
27 May 2026
Research

Natural Language Query to Configuration for Retrieval Agents

DGX agent

arXiv:2605.27361v1 Announce Type: new Abstract: Modern retrieval agents expose many configuration choices -- LLM, retriever, number of documents, number of hops, and synthesis strategy -- each shaping

researcharxiv-cs-ai
27 May 2026
Model Releases

Negligible in Size, Significant in Effect: On Scale Vectors in Large Language Models

DGX agent

arXiv:2605.26895v1 Announce Type: cross Abstract: Normalization layers in modern large language models (LLMs) consist of a deterministic normalization operation and a learnable scale vector. While the

model-releasesarxiv-cs-ai
27 May 2026
Safety

Neuro-Symbolic Verification of LLM Outputs for Data-Sensitive Domains (extended preprint)

DGX agent

arXiv:2605.26942v1 Announce Type: new Abstract: LLMs deployed in high-stakes domains face fundamental reliability challenges: hallucinations, inconsistencies, and privacy vulnerabilities introduce una

safetyarxiv-cs-ai
27 May 2026
Model Releases

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning

DGX agent

arXiv:2505.17163v2 Announce Type: replace-cross Abstract: Recent advancements in multimodal slow-thinking systems have demonstrated remarkable performance across various visual reasoning tasks. Howeve

model-releasesarxiv-cs-ai
27 May 2026
Safety

Olaf-World: Orienting Latent Actions for Video World Modeling

DGX agent

arXiv:2602.10104v2 Announce Type: replace-cross Abstract: Scaling action-controllable world models is limited by the scarcity of action labels. While latent action learning promises to extract control

safetyarxiv-cs-ai
27 May 2026
Model Releases

Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models

DGX agent

arXiv:2603.16654v2 Announce Type: replace-cross Abstract: Evaluating the reasoning abilities of large language models (LLMs) solely from final answers can obscure failures in intermediate steps, espec

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

OMD-GraphRAG: Enhancing GraphRAG with Ontology-Guided Extraction, Multi-Dimensional Clustering and Dual-Channel Fusion

DGX agent

arXiv:2603.25152v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems face significant challenges in complex reasoning, multi-hop queries, and domain-specific QA. While exis

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

OmniToM: Benchmarking Theory of Mind in LLMs via Explicit Belief Modeling

DGX agent

arXiv:2605.26322v1 Announce Type: new Abstract: Theory of Mind (ToM), the ability to infer others' knowledge, intentions, and emotions, is commonly evaluated in large language models (LLMs) using end-

model-releasesarxiv-cs-ai
27 May 2026
Research

On the Detection of Commutative Factors in Factor Graphs: Necessary and Sufficient Conditions

DGX agent

arXiv:2605.26908v1 Announce Type: new Abstract: Exploiting the indistinguishability of objects in a probabilistic graphical model such as a factor graph is key to lifted probabilistic inference algori

researcharxiv-cs-ai
27 May 2026
Research

On the Error-Correcting Effects of Stochasticity in Discrete Diffusion

DGX agent

arXiv:2605.26582v1 Announce Type: cross Abstract: Discrete diffusion models achieve strong performance in text and image generation, but their inference remains slow and must inherently balance sampli

researcharxiv-cs-ai
27 May 2026
Safety

On the Push-Based Asynchronous Federated Learning: A Bias-Correction Aggregation Approach

DGX agent

arXiv:2605.26162v1 Announce Type: cross Abstract: Asynchronous decentralized federated learning (ADFL) eliminates central coordination and global synchronization, making it attractive for large-scale

safetyarxiv-cs-ai
27 May 2026
Tutorials

ORCA: An End-to-End Interactive Copilot for Optimized Root Cause Analysis

DGX agent

arXiv:2605.27022v1 Announce Type: new Abstract: Causal analysis is a crucial task in many domains, including manufacturing, social science, and medicine. However, despite recent progress, the conceptu

tutorialsarxiv-cs-ai
27 May 2026
Model Releases

ORLoopBench: Solver-in-the-Loop Benchmarks for Self-Correction and Behavioral Rationality in Operations Research

DGX agent

arXiv:2601.21008v3 Announce Type: replace-cross Abstract: Operations Research practitioners debug infeasible models through an iterative process: inspecting Irreducible Infeasible Subsystems ( IIS), i

model-releasesarxiv-cs-ai
27 May 2026
Safety

Pair-In, Pair-Out: Latent Multi-Token Prediction for Efficient LLMs

DGX agent

arXiv:2605.27255v1 Announce Type: cross Abstract: Long chain-of-thought reasoning has made autoregressive decoding the dominant inference cost of modern large language models. Existing methods target

safetyarxiv-cs-ai
27 May 2026
Research

ParsVoice: A Large-Scale Multi-Speaker Persian Speech Corpus for Text-to-Speech Synthesis

DGX agent

arXiv:2510.10774v3 Announce Type: replace-cross Abstract: Persian remains substantially underrepresented in open speech-text resources, limiting progress in multi-speaker text-to-speech (TTS), speech-

researcharxiv-cs-ai
27 May 2026
Model Releases

PaTAS: A Framework for Trust Propagation in Neural Networks Using Subjective Logic

DGX agent

arXiv:2511.20586v4 Announce Type: replace Abstract: Trustworthiness has become a key requirement for the deployment of artificial intelligence systems in safety-critical applications. Conventional eva

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Periodic Topological Deep Learning for Polymer Design and Discovery

DGX agent

arXiv:2605.26833v1 Announce Type: cross Abstract: Polymers underpin applications across energy, healthcare, and materials science, yet their vast chemical space makes systematic discovery challenging.

model-releasesarxiv-cs-ai
27 May 2026
Local Ai

Persistent AI Agents in Academic Research: A Single-Investigator Implementation Case Study

DGX agent

arXiv:2605.26870v1 Announce Type: cross Abstract: Background: Large language models are typically evaluated as models, benchmarks, or short conversational episodes. Less is known about what happens wh

local-aiarxiv-cs-ai
27 May 2026
Agents

Persona Generators: Generating Diverse Synthetic Personas for Arbitrary Contexts

DGX agent

arXiv:2602.03545v2 Announce Type: replace Abstract: Evaluating AI systems that interact with humans requires understanding their behavior across diverse user populations, but collecting representative

agentsarxiv-cs-ai
27 May 2026
Model Releases

Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History

DGX agent

arXiv:2602.17003v2 Announce Type: replace-cross Abstract: Large language models have advanced web agents, yet current agents lack personalization capabilities. Since users rarely specify every detail

model-releasesarxiv-cs-ai
27 May 2026
Tutorials

Personalized Generative Models for Contextual Debiasing

DGX agent

arXiv:2605.26353v1 Announce Type: cross Abstract: Different visual patterns appear with different frequencies in the world: e.g., beach balls appear on sand more often than they do on a road. These st

tutorialsarxiv-cs-ai
27 May 2026
Agents

Personalizing Embodied Multimodal Large Language Model Agents over Long-term User Interactions

DGX agent

arXiv:2605.26256v1 Announce Type: new Abstract: Multimodal large language model (MLLM)-based embodied agents have shown strong potential for solving complex tasks in physical environments. However, pe

agentsarxiv-cs-ai
27 May 2026
Applications

Phase-Type Variational Autoencoders for Heavy-Tailed Data

DGX agent

arXiv:2603.01800v2 Announce Type: replace-cross Abstract: Heavy-tailed distributions are ubiquitous in real-world data, where rare but extreme events dominate risk and variability. However, standard V

applicationsarxiv-cs-ai
27 May 2026
Model Releases

'PhyWorldBench': A Comprehensive Evaluation of Physical Realism in Text-to-Video Models

DGX agent

arXiv:2507.13428v3 Announce Type: replace-cross Abstract: Video generation models have achieved remarkable progress in creating high-quality, photorealistic content. However, their ability to accurate

model-releasesarxiv-cs-ai
27 May 2026
Safety

PICACO: Pluralistic In-Context Value Alignment of LLMs via Total Correlation Optimization

DGX agent

arXiv:2507.16679v3 Announce Type: replace-cross Abstract: In-Context Learning has shown great potential for aligning Large Language Models (LLMs) with human values, helping reduce harmful outputs and

safetyarxiv-cs-ai
27 May 2026
Model Releases

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis

DGX agent

arXiv:2605.27258v1 Announce Type: cross Abstract: Building state-of-the-art text-to-speech (TTS) systems typically demands millions of hours of proprietary data and complex multi-stage architectures,

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

PitchBench: Measuring Pitch Hearing in Audio-Language Models

DGX agent

arXiv:2605.26176v1 Announce Type: cross Abstract: Audio-language models (ALMs) are increasingly used in real-world applications that require understanding music, from music tutoring and transcription

model-releasesarxiv-cs-ai
27 May 2026
Local Ai

Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning

DGX agent

arXiv:2510.01833v2 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong reasoning abilities via Chain-of-Thought (CoT), but their token-level generation encourages local de

local-aiarxiv-cs-ai
27 May 2026
Tutorials

Planning Neural Dynamics with Lie Group Embedding through Supervised Projective Manifold Learning

DGX agent

arXiv:2605.26167v1 Announce Type: cross Abstract: We propose Lie group embedded dynamical neural networks (LieEDNN) and the corresponding learning algorithms based on gradient descent and metric proje

tutorialsarxiv-cs-ai
27 May 2026
Research

Plans for Evaluating Structured Generative Search Summaries

DGX agent

arXiv:2605.26400v1 Announce Type: cross Abstract: We propose a framework for evaluating structured generative search summaries that are placed atop organic web search results. A structured summary, ge

researcharxiv-cs-ai
27 May 2026
Agents

PolyFusionAgent: A Multimodal Foundation Model and Autonomous AI Assistant for Polymer Property Prediction and Inverse Design

DGX agent

arXiv:2605.26543v1 Announce Type: new Abstract: Polymer discovery is central to fields ranging from energy storage to biomedicine, but it is hindered by an astronomically large chemical design space a

agentsarxiv-cs-ai
27 May 2026
Model Releases

Position: AI Safety Requires Effective Controllability

DGX agent

arXiv:2605.27117v1 Announce Type: new Abstract: AI safety is still largely framed as alignment: training models to follow human preferences, safety policies, and normative constraints. That framing ha

model-releasesarxiv-cs-ai
27 May 2026
Safety

Practical Anonymous Two-Party Gradient Boosting Decision Tree

DGX agent

arXiv:2605.26903v1 Announce Type: cross Abstract: Structured data is well handled by gradient-boosted decision trees (GBDT), which are usually trained on vertically partitioned features across mutuall

safetyarxiv-cs-ai
27 May 2026
Research

Pretraining Data Exposure in Large Language Models: A Survey of Membership Inference, Data Contamination, and Security Implications

DGX agent

arXiv:2605.26133v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become the predominant paradigm in NLP, advancing both research and industry. As model sizes and pretraining data gr

researcharxiv-cs-ai
27 May 2026
Research

Prospective evaluation of multimodal respiratory failure prediction: Do chest X-rays improve performance beyond EHR signals?

DGX agent

arXiv:2605.26255v1 Announce Type: cross Abstract: Early prediction of respiratory failure is critical for timely clinical intervention in intensive care units. Existing electronic health record (EHR)-

researcharxiv-cs-ai
27 May 2026
Model Releases

Qiskit QuantumKatas: Adapting Microsoft's Quantum Computing exercises for LLM evaluation

DGX agent

arXiv:2605.27210v1 Announce Type: cross Abstract: We adapt Microsoft's QuantumKatas -- a well-established quantum computing curriculum -- from Q# to Qiskit, the most widely-adopted quantum computing f

model-releasesarxiv-cs-ai
27 May 2026
Hardware

Qrita: High-performance Top-k and Top-p using Pivot-based Truncation and Selection

DGX agent

arXiv:2602.01518v2 Announce Type: replace Abstract: Despite their importance in model sampling, efficient implementation of Top-k and Top-p algorithms for large vocabularies remains a significant chal

hardwarearxiv-cs-ai
27 May 2026
Agents

QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents

DGX agent

arXiv:2605.27068v1 Announce Type: cross Abstract: Social deduction games have become a popular testbed for probing reasoning, deception, coordination, and belief modeling in Large Language Model (LLM)

agentsarxiv-cs-ai
27 May 2026
Safety

Quantized Keys Steal Attention: Bias Correction for KV-Cache Compression in Video Diffusion

DGX agent

arXiv:2605.26266v1 Announce Type: cross Abstract: Chunk-wise autoregressive video diffusion models rely on a KV cache of previously generated chunks to avoid redundant computation, but this cache quic

safetyarxiv-cs-ai
27 May 2026
← Previous
1…259260261262263…448
Next →