AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Safety

Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning

DGX agent

arXiv:2512.02019v3 Announce Type: replace-cross Abstract: Diffusion models excel at sampling from complex, unnormalized distributions. In this work, we extend Maximum Entropy Reinforcement Learning (M

safetyarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Diffusion-Based Ukrainian Handwritten Text Generation with Cross-Domain Style Transfer

DGX agent

arXiv:2605.27487v1 Announce Type: cross Abstract: Handwritten text generation (HTG) conditioned on writer style has been widely studied for Latin scripts, but remains underexplored for low-resource an

model-releasesarxiv-cs-ai
28 May 2026
Safety

Diffusion Large Language Models for Visual Speech Recognition

DGX agent

arXiv:2605.28456v1 Announce Type: new Abstract: Existing Visual Speech Recognition (VSR) systems commonly rely on left-to-right autoregressive decoding, which can force premature decisions on visually

safetyarxiv-cs-ai
28 May 2026
Agents

DIG to Heal: Scaling General-purpose Agent Collaboration via Explainable Dynamic Decision Paths

DGX agent

arXiv:2603.00309v2 Announce Type: replace Abstract: The increasingly popular agentic AI paradigm promises to harness the power of multiple, general-purpose large language model (LLM) agents to collabo

agentsarxiv-cs-ai
28 May 2026
Agents

Discovery Agents for Real-Time Analytics: Toward Proactive Insight Systems

DGX agent

arXiv:2605.27571v1 Announce Type: new Abstract: Modern analytics systems are fundamentally reactive, requiring users to define queries over increasingly complex and continuously evolving data. In real

agentsarxiv-cs-ai
28 May 2026
Safety

Disentangling Adversarial Prompts: A Semantic-Graph Defense for Robust LLM Security

DGX agent

arXiv:2605.27823v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly vulnerable to adversarial prompts that exploit semantic ambiguities to bypass safety mechanisms, resulti

safetyarxiv-cs-ai
28 May 2026
Agents

Do Agents Know What They Can't Do? Evaluating Feasibility Awareness in Tool-Using Agents

DGX agent

arXiv:2605.28532v1 Announce Type: new Abstract: Tool-using agents often incur substantial computational cost due to long reasoning chains and iterative tool usage. In practical scenarios, many tasks b

agentsarxiv-cs-ai
28 May 2026
Agents

Do Agents Need Semantic Metadata? A Comparative Study in Agentic Data Retrieval

DGX agent

arXiv:2605.28787v1 Announce Type: cross Abstract: In the era of autonomous agents, machine-actionable data is critical for data-driven workflows. For more than a decade, semantic metadata like schema.

agentsarxiv-cs-ai
28 May 2026
Model Releases

Do Agents Think Deeper? A Mechanistic Investigation of Layer-Wise Dynamics in Sequential Planning

DGX agent

arXiv:2605.27935v1 Announce Type: new Abstract: Recent mechanistic studies suggest that large language models (LLMs) may utilize their depth inefficiently in standard single-turn tasks. Whether this s

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do Clinical Models Change Treatment Decisions?

DGX agent

arXiv:2605.28129v1 Announce Type: new Abstract: Clinical foundation models are evaluated with factual or exam-style medical QA, but treatment decisions must change when patient context changes. We int

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do LLMs Build World Models From Text? A Multilingual Diagnostic of Spatial Reasoning

DGX agent

arXiv:2605.28277v1 Announce Type: new Abstract: Whether large language models (LLMs) construct internal spatial world models from pure-text descriptions remains contested, and whether such capabilitie

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do LLMs Favor Their Providers? Measuring Vertical Integration Bias in Code Generation

DGX agent

arXiv:2605.28515v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of software development, especially with the advent of agentic capabilities. Yet, many front

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do Models Know Why They Changed Their Mind? Interpretability and Faithfulness of Chain-of-Thought Under Knowledge Conflict

DGX agent

arXiv:2605.27773v1 Announce Type: cross Abstract: When a language model sees a document contradicting its training knowledge, it must choose: follow the document or trust itself. Prior work proved thi

model-releasesarxiv-cs-ai
28 May 2026
Applications

Do readers prefer AI-generated Italian short stories?

DGX agent

arXiv:2601.17363v2 Announce Type: replace-cross Abstract: This study investigates whether readers prefer AI-generated short stories in Italian over one written by a renowned Italian author. In a blind

applicationsarxiv-cs-ai
28 May 2026
Model Releases

Do We Really Need Quantum Machine Learning?: A Multidimensional Empirical Study

DGX agent

arXiv:2605.27923v1 Announce Type: cross Abstract: The rapid growth of computer vision and increasingly complex image recognition tasks has exposed fundamental computational limitations of classical ma

model-releasesarxiv-cs-ai
28 May 2026
Research

Domain size asymptotics for Markov logic networks

DGX agent

arXiv:2509.04192v2 Announce Type: replace Abstract: A Markov logic network (MLN) M determines a probability distribution P_n^M on the set mathbf{W}_n of structures, or ``possible worlds'', with domain

researcharxiv-cs-ai
28 May 2026
Model Releases

Dr-CiK: A Testbed for Foresight-Driven Agents

DGX agent

arXiv:2605.27904v1 Announce Type: new Abstract: Time series forecasting in real-world settings often depends not only on historical observations, but also on external context that must be actively dis

model-releasesarxiv-cs-ai
28 May 2026
Safety

DREAM-R: Multimodal Speculative Reasoning with RL-Based Refined Drafting, Precise Verification, and Fully Parallel Execution

DGX agent

arXiv:2605.28678v1 Announce Type: new Abstract: Speculative reasoning has recently been proposed as a means to accelerate reasoning-intensive generation in large multimodal models, but its effectivene

safetyarxiv-cs-ai
28 May 2026
Agents

DSSE: a drone swarm search environment

DGX agent

arXiv:2307.06240v2 Announce Type: replace-cross Abstract: The Drone Swarm Search project is an environment, based on extsc{PettingZoo}, that is to be used in conjunction with multi-agent (or single-ag

agentsarxiv-cs-ai
28 May 2026
Model Releases

DynaSchedBench: Calibrated Dynamic Scheduling Benchmarks and Observability Paradox in LLM-based Scheduling Agents

DGX agent

arXiv:2605.27566v1 Announce Type: new Abstract: Progress in neural combinatorial optimization for Dynamic Flexible Job Shop Scheduling Problem (DFJSP) is currently hindered by a methodological tension

model-releasesarxiv-cs-ai
28 May 2026
Research

EAGer: Entropy-Aware GEneRation for Adaptive Inference-Time Scaling

DGX agent

arXiv:2510.11170v2 Announce Type: replace-cross Abstract: With the rise of reasoning language models and test-time scaling methods as a paradigm for improving model performance, substantial computatio

researcharxiv-cs-ai
28 May 2026
Safety

EAPO: Entropy-Driven Adaptive Positive-Negative Sample Weighting for Policy Optimization in Open-Ended QA

DGX agent

arXiv:2605.27846v1 Announce Type: new Abstract: Large Reasoning Models are typically trained via reinforcement learning from verifiable rewards (RLVR). However, existing approaches adopt fixed weights

safetyarxiv-cs-ai
28 May 2026
Safety

ECHO: Entropy-Confidence Hybrid Optimization for Test-Time Reinforcement Learning

DGX agent

arXiv:2602.02150v2 Announce Type: replace-cross Abstract: Test-time reinforcement learning generates multiple candidate answers via repeated rollouts and performs online updates using pseudo-labels co

safetyarxiv-cs-ai
28 May 2026
Model Releases

Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets

DGX agent

arXiv:2605.28510v1 Announce Type: cross Abstract: Large language models (LLMs) for code completion and generation are increasingly used in software development, yet they may reproduce training example

model-releasesarxiv-cs-ai
28 May 2026
Research

Efficient Post-training of LLMs for Code Generation With Offline Reinforcement Learning

DGX agent

arXiv:2605.28409v1 Announce Type: new Abstract: Post-training using online reinforcement learning (RL) is an important training step for LLMs, including code-generating models. However, online RL for

researcharxiv-cs-ai
28 May 2026
Model Releases

Efficient Pre-Training of LLMs through Truncated SVD Layers

DGX agent

arXiv:2605.28573v1 Announce Type: cross Abstract: The massive scaling of Large Language Models (LLMs) has made pretraining increasingly cost-prohibitive. While low-rank representation and orthonormal

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

EgoBench: An Interactive Egocentric Multimodal Benchmark for Tool-Using Agents

DGX agent

arXiv:2605.27820v1 Announce Type: new Abstract: As AI agents increasingly operate in open, real-world environments, they require a deep synergy of multimodal perception, tool invocation with multi-hop

model-releasesarxiv-cs-ai
28 May 2026
Applications

EigeNet: Geometry-Informed Multi-Modal Learning for Few-shot Novel View RIR Prediction

DGX agent

arXiv:2605.28101v1 Announce Type: cross Abstract: Predicting spatially varying Room Impulse Response (RIR) from sparse observations is a critical but highly challenging inverse problem for immersive s

applicationsarxiv-cs-ai
28 May 2026
Research

Eliot: Interactively nderline{E}xploring Fast-Changing Scientific nderline{Li}terature Trends with nderline{O}nline Danderline{t}a and Learning

DGX agent

arXiv:2605.27610v1 Announce Type: cross Abstract: The rapid growth of scientific publishing has made it increasingly difficult to track how fast-moving areas evolve. Search engines and LLM-based assis

researcharxiv-cs-ai
28 May 2026
Safety

Emerging Extrinsic Dexterity in Cluttered Scenes via Dynamics-aware Policy Learning

DGX agent

arXiv:2603.09882v2 Announce Type: replace-cross Abstract: Extrinsic dexterity leverages environmental contact to overcome the limitations of prehensile manipulation. However, achieving such dexterity

safetyarxiv-cs-ai
28 May 2026
Model Releases

Energy-Structured Low-Rank Adaptation for Continual Learning

DGX agent

arXiv:2605.27482v1 Announce Type: cross Abstract: While orthogonal subspace methods try to mitigate task interference in Continual Learning (CL), they often suffer from energy diffusion across the bas

model-releasesarxiv-cs-ai
28 May 2026
Research

Entropy-aware Masking for Masked Language Modeling

DGX agent

arXiv:2605.28526v1 Announce Type: new Abstract: Masked language modeling has become a standard pretraining objective for training encoder-based language models. In this approach, certain tokens in the

researcharxiv-cs-ai
28 May 2026
Research

Entropy Distribution as a Fingerprint for Hallucinations in Generative Models

DGX agent

arXiv:2605.28264v1 Announce Type: new Abstract: Large Language Models (LLMs) often generate factually incorrect outputs, commonly termed hallucinations, that undermine trust and limit deployment in hi

researcharxiv-cs-ai
28 May 2026
Model Releases

ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations

DGX agent

arXiv:2605.27908v1 Announce Type: cross Abstract: Existing emotional support conversation (ESC) systems mainly rely on end-to-end response generation or coarse strategy supervision, offering limited i

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection

DGX agent

arXiv:2505.17654v4 Announce Type: replace-cross Abstract: E-commerce platforms increasingly rely on Large Language Models (LLMs) and Vision Language Models (VLMs) to detect illicit or misleading produ

model-releasesarxiv-cs-ai
28 May 2026
Safety

Evaluating the Realism of LLM-powered Social Agents: A Case Study of Reactions to Spanish Online News

DGX agent

arXiv:2605.28598v1 Announce Type: cross Abstract: LLM-powered social agents are increasingly used to simulate online social behavior, yet their realism remains difficult to validate. Existing work has

safetyarxiv-cs-ai
28 May 2026
Tutorials

Evaluation of AI Ethics Tools in Language Models: A Developers' Perspective Case Study

DGX agent

arXiv:2512.15791v2 Announce Type: replace-cross Abstract: In Artificial Intelligence (AI), language models have gained significant importance due to the widespread adoption of systems capable of simul

tutorialsarxiv-cs-ai
28 May 2026
Model Releases

EvoSpec: Evolving Speculative Decoding via Real-Time Vocabulary and Parameter AdaptationTarget

DGX agent

arXiv:2605.27390v1 Announce Type: cross Abstract: Speculative decoding accelerates Large Language Model inference via a draft-then-verify paradigm, yet the output projection layer becomes a bottleneck

model-releasesarxiv-cs-ai
28 May 2026
Safety

Examining Agents' Bias Amplification versus Suppression in Multi-Agent Systems

DGX agent

arXiv:2605.28098v1 Announce Type: new Abstract: Multi-agent systems are increasingly deployed to support various tasks where agents interact to achieve individual and collective objectives. Although t

safetyarxiv-cs-ai
28 May 2026
Research

Explaining is Harder Than Predicting Alone: Evaluating Concept-based Explanations of MLLMs as ICL Visual Classifiers

DGX agent

arXiv:2605.28215v1 Announce Type: new Abstract: In-context learning (ICL) enables multimodal large language models (MLLMs) to classify images from a few labelled examples. Yet, how these models use th

researcharxiv-cs-ai
28 May 2026
Research

Extracting Small Translation Specialists from LLMs by Aggressively Pruning Experts

DGX agent

arXiv:2605.28042v1 Announce Type: cross Abstract: Modern large language models (LLMs) achieve state-of-the-art machine translation performance, but they do so as broad generalists largely trained for

researcharxiv-cs-ai
28 May 2026
Agents

Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL

DGX agent

arXiv:2605.28751v1 Announce Type: cross Abstract: Linear interpolation between fine-tuned checkpoints has been shown to trace the Pareto front between competing objectives, but whether extrapolative w

agentsarxiv-cs-ai
28 May 2026
Model Releases

FactReview: Evidence-Grounded Peer Review with Execution-Based Claim Verification

DGX agent

arXiv:2604.04074v3 Announce Type: replace Abstract: LLM-based reviewing systems typically take only the manuscript as input, leaving literature and code-based claims hard to verify. We present FactRev

model-releasesarxiv-cs-ai
28 May 2026
Local Ai

FD-RAG: Federated Dual-System Retrieval-Augmented Generation

DGX agent

arXiv:2605.27432v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) has emerged as a paradigm for grounding large language models in external knowledge, yet most existing RAG system

local-aiarxiv-cs-ai
28 May 2026
Model Releases

FedMPT: Federated Multi-label Prompt Tuning of Vision-Language Models

DGX agent

arXiv:2605.28347v1 Announce Type: new Abstract: Multi-Label Recognition (MLR) based on Vision-Language Models (VLMs) aims to leverage their pre-trained knowledge to better adapt complex recognition sc

model-releasesarxiv-cs-ai
28 May 2026
Applications

Fine-Tuned LLM as a Complementary Predictor Improving Ads System

DGX agent

arXiv:2605.27856v1 Announce Type: cross Abstract: Recommendation systems power engagement and monetization across feeds, ads, and short-video platforms, but translating the latest advances in Large La

applicationsarxiv-cs-ai
28 May 2026
Research

FinTexTS: Financial Text-Paired Time-Series Dataset via Semantic-Based and Multi-Level Pairing

DGX agent

arXiv:2603.02702v3 Announce Type: replace Abstract: The financial domain involves a variety of important time-series problems. Recently, time-series analysis methods that jointly leverage textual and

researcharxiv-cs-ai
28 May 2026
Model Releases

FLORO: A Multimodal Geospatial Foundation Model for Ecological Remote Sensing Across Sensors and Scales

DGX agent

arXiv:2605.28174v1 Announce Type: cross Abstract: Foundation models offer a promising route to transferable remote sensing representations, but many current approaches depend on very large pretraining

model-releasesarxiv-cs-ai
28 May 2026
← Previous
1…247248249250251…448
Next →