AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
28 May 2026

CubePart: An Open-Vocabulary Part-Controllable 3D Generator

ResearchDGX agent

arXiv:2605.28763v1 Announce Type: new Abstract: Interactive 3D assets used in games and simulation are typically decomposed into specific semantic parts to support animation, physics, and scripted beh

Cultural Binding Heads in Language Models

Model ReleasesDGX agent

arXiv:2605.28543v1 Announce Type: new Abstract: LLMs often default to equal treatment across cultural groups, even though context warrants differentiation: this is a lack of difference awareness. Usin

Cultural Fidelity in English-to-Hindi Translation: A Preservation-Fluency Frontier for Gender Recoverability

Model ReleasesDGX agent

arXiv:2605.27654v1 Announce Type: cross Abstract: Generative translation systems are cultural technologies because they decide how socially meaningful cues are rendered within culturally specific gram


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cyberbullying Governance on Social Media: A Unified Framework from Content Identification to Intervention

SafetyDGX agent

arXiv:2605.27584v1 Announce Type: new Abstract: The proliferation of social media platforms and online communities has inadvertently catalyzed the spread of cyberbullying, hate speech, and other forms

CyberJurors: A Multi-Agent Simulation Task for E-Commerce Disputes Verdict

Model ReleasesDGX agent

arXiv:2605.28369v1 Announce Type: new Abstract: E-commerce platforms have begun recruiting crowdsourced jurors to adjudicate massive volumes of transaction disputes. Unlike formal legal judgment, E-co

Data-Efficient On-Policy Distillation for Automatic Speech Recognition

Model ReleasesDGX agent

arXiv:2605.28139v1 Announce Type: new Abstract: Building competitive automatic speech recognition (ASR) models usually requires large-scale au- dio supervision, which makes reproduction and specializa

Debate Helps Weak Judges Reward Stronger Models

ResearchDGX agent

arXiv:2605.27483v1 Announce Type: cross Abstract: Despite theoretical promise, debate as a scalable oversight protocol has produced mixed empirical results: gains in some settings, and null effects in

Debate with Images: Detecting Deceptive Behaviors in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2512.00349v2 Announce Type: replace Abstract: Are frontier AI systems becoming more capable? Certainly. Yet such progress is not an unalloyed blessing but rather a Trojan horse: behind their per

DecomposeRL: Learning to Ask Useful, Informative, and Diverse Questions for Semi-Supervised, Traceable Claim Verification

Model ReleasesDGX agent

arXiv:2605.27858v1 Announce Type: cross Abstract: Claim verification splits between end-to-end classifiers that are accurate but yields no inspectable traces, and decomposition-based methods produce i

Deconstructing Spatial Complexity: Hierarchical Decomposition for LLM Spatial Reasoning

SafetyDGX agent

arXiv:2605.28144v1 Announce Type: new Abstract: LLMs have shown remarkable proficiency in general language understanding and reasoning. However, they consistently underperform in spatial reasoning tha

Deep Learning Strain Estimation: Is Physics-Based Simulation the Solution?

ResearchDGX agent

arXiv:2605.28697v1 Announce Type: cross Abstract: Speckle tracking echocardiography (STE) is the clinical standard for myocardial strain estimation. Despite good performance on global strain (GLS), it

Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024

Model ReleasesDGX agent

arXiv:2503.02857v5 Announce Type: replace-cross Abstract: In the age of increasingly realistic generative AI, robust deepfake detection is essential for mitigating fraud and disinformation. While many

DeepSciVerify: Verifying Scientific Claim--Citation Alignment via LLM-Driven Evidence Escalation

Model ReleasesDGX agent

arXiv:2605.27710v1 Announce Type: new Abstract: Misalignment between claims and their cited evidence is a common failure mode in reports generated by large language models, limiting their reliability

Defending LLM-based Multi-Agent Systems Against Cooperative Attacks with Sentence-Level Rectification

AgentsDGX agent

arXiv:2605.28104v1 Announce Type: new Abstract: Recent years have witnessed the rapid development of Large Language Model-based Multi-Agent Systems (MAS), which excel at collaborative decision-making

Delay-Aware Reinforcement Learning for Highway On-Ramp Merging under Stochastic Communication Latency

SafetyDGX agent

arXiv:2403.11852v5 Announce Type: replace-cross Abstract: Delayed and partially observable state information poses significant challenges for reinforcement learning (RL)-based control in real-world au

DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

AgentsDGX agent

arXiv:2605.28148v1 Announce Type: cross Abstract: The rapid development of LLMs coupled with the introduction of Model Context Protocol (MCP) has revolutionized how intelligent agents interact with AP

DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes

SafetyDGX agent

arXiv:2605.28421v1 Announce Type: new Abstract: Reinforcement learning has become a central paradigm for advancing reasoning in large language models, yet most existing methods still depend on stronge

DEPART: DEcomposing PARiTy across Multilingual LLMs

Model ReleasesDGX agent

arXiv:2605.28163v1 Announce Type: cross Abstract: Multilingual Large Language Models (mLLMs) leaderboards report per-language accuracy but rarely explain why disparities emerge, leaving systemic biase

Detect by Yourself: Self-Designing Agentic Workflows for Few-Shot Graph Anomaly Detection

AgentsDGX agent

arXiv:2605.27470v1 Announce Type: cross Abstract: Graph anomaly detection aims to identify anomaly nodes in attributed graphs and plays an important role in real-world applications. However, existing

Detection Without Correction: A Two-Parameter Decomposition of Multi-Stage LLM Pipelines

Model ReleasesDGX agent

arXiv:2605.27559v1 Announce Type: cross Abstract: Multi-stage LLM pipelines that perform multi-agent debate, intrinsic self-correction, or retrieval-augmented verification exhibit puzzling aggregate b

Developing an Intelligent Job Recommendation System Using Semantic Retrieval and Explainable AI Techniques

ResearchDGX agent

arXiv:2605.27656v1 Announce Type: cross Abstract: Online recruitment platforms require recommendation methods capable of retrieving relevant job opportunities from large and heterogeneous collections

Diagnosing Live Within-Policy Instruction Conflicts in LLM Agents with Witnessed Resolution Profiles

SafetyDGX agent

arXiv:2605.27784v1 Announce Type: new Abstract: LLM agents are governed by long-lived natural-language prompt policies, but individually reasonable standing rules can interact in uninspected ways. We

DiagramRAG: A Lightweight Framework to Retrieve Scientific Diagram for Figure Generation

TutorialsDGX agent

arXiv:2605.27931v1 Announce Type: new Abstract: Scientific diagrams are essential for communicating complex methodologies in academic papers. A natural way for researchers to specify such diagrams is

Differential syntactic and semantic encoding in LLMs

Model ReleasesDGX agent

arXiv:2601.04765v4 Announce Type: replace-cross Abstract: We study how syntactic and semantic information is encoded in inner layer representations of Large Language Models (LLMs), focusing on the ver

Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning

SafetyDGX agent

arXiv:2512.02019v3 Announce Type: replace-cross Abstract: Diffusion models excel at sampling from complex, unnormalized distributions. In this work, we extend Maximum Entropy Reinforcement Learning (M

Diffusion-Based Ukrainian Handwritten Text Generation with Cross-Domain Style Transfer

Model ReleasesDGX agent

arXiv:2605.27487v1 Announce Type: cross Abstract: Handwritten text generation (HTG) conditioned on writer style has been widely studied for Latin scripts, but remains underexplored for low-resource an

Diffusion Large Language Models for Visual Speech Recognition

SafetyDGX agent

arXiv:2605.28456v1 Announce Type: new Abstract: Existing Visual Speech Recognition (VSR) systems commonly rely on left-to-right autoregressive decoding, which can force premature decisions on visually

DIG to Heal: Scaling General-purpose Agent Collaboration via Explainable Dynamic Decision Paths

AgentsDGX agent

arXiv:2603.00309v2 Announce Type: replace Abstract: The increasingly popular agentic AI paradigm promises to harness the power of multiple, general-purpose large language model (LLM) agents to collabo

Discovery Agents for Real-Time Analytics: Toward Proactive Insight Systems

AgentsDGX agent

arXiv:2605.27571v1 Announce Type: new Abstract: Modern analytics systems are fundamentally reactive, requiring users to define queries over increasingly complex and continuously evolving data. In real

Disentangling Adversarial Prompts: A Semantic-Graph Defense for Robust LLM Security

SafetyDGX agent

arXiv:2605.27823v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly vulnerable to adversarial prompts that exploit semantic ambiguities to bypass safety mechanisms, resulti

Do Agents Know What They Can't Do? Evaluating Feasibility Awareness in Tool-Using Agents

AgentsDGX agent

arXiv:2605.28532v1 Announce Type: new Abstract: Tool-using agents often incur substantial computational cost due to long reasoning chains and iterative tool usage. In practical scenarios, many tasks b

Do Agents Need Semantic Metadata? A Comparative Study in Agentic Data Retrieval

AgentsDGX agent

arXiv:2605.28787v1 Announce Type: cross Abstract: In the era of autonomous agents, machine-actionable data is critical for data-driven workflows. For more than a decade, semantic metadata like schema.

Do Agents Think Deeper? A Mechanistic Investigation of Layer-Wise Dynamics in Sequential Planning

Model ReleasesDGX agent

arXiv:2605.27935v1 Announce Type: new Abstract: Recent mechanistic studies suggest that large language models (LLMs) may utilize their depth inefficiently in standard single-turn tasks. Whether this s

Do Clinical Models Change Treatment Decisions?

Model ReleasesDGX agent

arXiv:2605.28129v1 Announce Type: new Abstract: Clinical foundation models are evaluated with factual or exam-style medical QA, but treatment decisions must change when patient context changes. We int

Do LLMs Build World Models From Text? A Multilingual Diagnostic of Spatial Reasoning

Model ReleasesDGX agent

arXiv:2605.28277v1 Announce Type: new Abstract: Whether large language models (LLMs) construct internal spatial world models from pure-text descriptions remains contested, and whether such capabilitie

Do LLMs Favor Their Providers? Measuring Vertical Integration Bias in Code Generation

Model ReleasesDGX agent

arXiv:2605.28515v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of software development, especially with the advent of agentic capabilities. Yet, many front

Do Models Know Why They Changed Their Mind? Interpretability and Faithfulness of Chain-of-Thought Under Knowledge Conflict

Model ReleasesDGX agent

arXiv:2605.27773v1 Announce Type: cross Abstract: When a language model sees a document contradicting its training knowledge, it must choose: follow the document or trust itself. Prior work proved thi

Do readers prefer AI-generated Italian short stories?

ApplicationsDGX agent

arXiv:2601.17363v2 Announce Type: replace-cross Abstract: This study investigates whether readers prefer AI-generated short stories in Italian over one written by a renowned Italian author. In a blind

Do We Really Need Quantum Machine Learning?: A Multidimensional Empirical Study

Model ReleasesDGX agent

arXiv:2605.27923v1 Announce Type: cross Abstract: The rapid growth of computer vision and increasingly complex image recognition tasks has exposed fundamental computational limitations of classical ma

Domain size asymptotics for Markov logic networks

ResearchDGX agent

arXiv:2509.04192v2 Announce Type: replace Abstract: A Markov logic network (MLN) M determines a probability distribution P_n^M on the set mathbf{W}_n of structures, or ``possible worlds'', with domain

Dr-CiK: A Testbed for Foresight-Driven Agents

Model ReleasesDGX agent

arXiv:2605.27904v1 Announce Type: new Abstract: Time series forecasting in real-world settings often depends not only on historical observations, but also on external context that must be actively dis

DREAM-R: Multimodal Speculative Reasoning with RL-Based Refined Drafting, Precise Verification, and Fully Parallel Execution

SafetyDGX agent

arXiv:2605.28678v1 Announce Type: new Abstract: Speculative reasoning has recently been proposed as a means to accelerate reasoning-intensive generation in large multimodal models, but its effectivene

DSSE: a drone swarm search environment

AgentsDGX agent

arXiv:2307.06240v2 Announce Type: replace-cross Abstract: The Drone Swarm Search project is an environment, based on extsc{PettingZoo}, that is to be used in conjunction with multi-agent (or single-ag

DynaSchedBench: Calibrated Dynamic Scheduling Benchmarks and Observability Paradox in LLM-based Scheduling Agents

Model ReleasesDGX agent

arXiv:2605.27566v1 Announce Type: new Abstract: Progress in neural combinatorial optimization for Dynamic Flexible Job Shop Scheduling Problem (DFJSP) is currently hindered by a methodological tension

EAGer: Entropy-Aware GEneRation for Adaptive Inference-Time Scaling

ResearchDGX agent

arXiv:2510.11170v2 Announce Type: replace-cross Abstract: With the rise of reasoning language models and test-time scaling methods as a paradigm for improving model performance, substantial computatio

EAPO: Entropy-Driven Adaptive Positive-Negative Sample Weighting for Policy Optimization in Open-Ended QA

SafetyDGX agent

arXiv:2605.27846v1 Announce Type: new Abstract: Large Reasoning Models are typically trained via reinforcement learning from verifiable rewards (RLVR). However, existing approaches adopt fixed weights

ECHO: Entropy-Confidence Hybrid Optimization for Test-Time Reinforcement Learning

SafetyDGX agent

arXiv:2602.02150v2 Announce Type: replace-cross Abstract: Test-time reinforcement learning generates multiple candidate answers via repeated rollouts and performs online updates using pseudo-labels co

Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets

Model ReleasesDGX agent

arXiv:2605.28510v1 Announce Type: cross Abstract: Large language models (LLMs) for code completion and generation are increasingly used in software development, yet they may reproduce training example

Efficient Post-training of LLMs for Code Generation With Offline Reinforcement Learning

ResearchDGX agent

arXiv:2605.28409v1 Announce Type: new Abstract: Post-training using online reinforcement learning (RL) is an important training step for LLMs, including code-generating models. However, online RL for

Efficient Pre-Training of LLMs through Truncated SVD Layers

Model ReleasesDGX agent

arXiv:2605.28573v1 Announce Type: cross Abstract: The massive scaling of Large Language Models (LLMs) has made pretraining increasingly cost-prohibitive. While low-rank representation and orthonormal

EgoBench: An Interactive Egocentric Multimodal Benchmark for Tool-Using Agents

Model ReleasesDGX agent

arXiv:2605.27820v1 Announce Type: new Abstract: As AI agents increasingly operate in open, real-world environments, they require a deep synergy of multimodal perception, tool invocation with multi-hop

EigeNet: Geometry-Informed Multi-Modal Learning for Few-shot Novel View RIR Prediction

ApplicationsDGX agent

arXiv:2605.28101v1 Announce Type: cross Abstract: Predicting spatially varying Room Impulse Response (RIR) from sparse observations is a critical but highly challenging inverse problem for immersive s

Eliot: Interactively nderline{E}xploring Fast-Changing Scientific nderline{Li}terature Trends with nderline{O}nline Danderline{t}a and Learning

ResearchDGX agent

arXiv:2605.27610v1 Announce Type: cross Abstract: The rapid growth of scientific publishing has made it increasingly difficult to track how fast-moving areas evolve. Search engines and LLM-based assis

Emerging Extrinsic Dexterity in Cluttered Scenes via Dynamics-aware Policy Learning

SafetyDGX agent

arXiv:2603.09882v2 Announce Type: replace-cross Abstract: Extrinsic dexterity leverages environmental contact to overcome the limitations of prehensile manipulation. However, achieving such dexterity

Energy-Structured Low-Rank Adaptation for Continual Learning

Model ReleasesDGX agent

arXiv:2605.27482v1 Announce Type: cross Abstract: While orthogonal subspace methods try to mitigate task interference in Continual Learning (CL), they often suffer from energy diffusion across the bas

Entropy-aware Masking for Masked Language Modeling

ResearchDGX agent

arXiv:2605.28526v1 Announce Type: new Abstract: Masked language modeling has become a standard pretraining objective for training encoder-based language models. In this approach, certain tokens in the

Entropy Distribution as a Fingerprint for Hallucinations in Generative Models

ResearchDGX agent

arXiv:2605.28264v1 Announce Type: new Abstract: Large Language Models (LLMs) often generate factually incorrect outputs, commonly termed hallucinations, that undermine trust and limit deployment in hi

ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations

Model ReleasesDGX agent

arXiv:2605.27908v1 Announce Type: cross Abstract: Existing emotional support conversation (ESC) systems mainly rely on end-to-end response generation or coarse strategy supervision, offering limited i

EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection

Model ReleasesDGX agent

arXiv:2505.17654v4 Announce Type: replace-cross Abstract: E-commerce platforms increasingly rely on Large Language Models (LLMs) and Vision Language Models (VLMs) to detect illicit or misleading produ

Evaluating the Realism of LLM-powered Social Agents: A Case Study of Reactions to Spanish Online News

SafetyDGX agent

arXiv:2605.28598v1 Announce Type: cross Abstract: LLM-powered social agents are increasingly used to simulate online social behavior, yet their realism remains difficult to validate. Existing work has

← Previous
1…197198199200201…358
Next →