AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
3 Jun 2026

Local Guidance, Global Impact: Gaussian-Reshaped Trust Region Unlocks Behavior Transitions

Local AiDGX agent

arXiv:2606.03382v1 Announce Type: cross Abstract: While Proximal Policy Optimization (PPO) demonstrates strong performance in stationary settings, we show that its standard optimization paradigm strug

Margin Play: A Multi-Agent System For Public Policy Analysis In The Brazilian Equatorial Margin

SafetyDGX agent

arXiv:2606.02614v1 Announce Type: cross Abstract: The Brazilian Equatorial Margin (BEM) is Brazil's next offshore oil frontier, with operations expected to begin in 2026 in the Foz do Amazonas basin.

Measuring Weak-to-Strong Legibility of Reasoning Models

SafetyDGX agent

arXiv:2603.20508v2 Announce Type: replace-cross Abstract: Reasoning language models (RLMs) and the intermediate chains of thought they emit play an increasingly central role in multi-agent setups such


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MedCUA-Bench: A Screenshot-Only Benchmark for Clinical Computer-Use Agents

Model ReleasesDGX agent

arXiv:2606.03203v1 Announce Type: new Abstract: Computer-use agents could automate repetitive screen-based clinical work, but their reliability in medical graphical user interfaces remains largely unv

MemVerse: Multimodal Memory for Lifelong Learning Agents

ResearchDGX agent

arXiv:2512.03627v2 Announce Type: replace Abstract: Despite rapid progress in large-scale language and vision models, AI agents still suffer from a fundamental limitation: they cannot remember. Withou

Merit or networks? What decides where research is published

ApplicationsDGX agent

arXiv:2606.03763v1 Announce Type: cross Abstract: Does scientific publishing reward the quality of ideas or the advantage of connections? The question is universal to prestige-driven science, yet it h

Message Tuning Outshines Graph Prompt Tuning: A Prismatic Space Perspective

Model ReleasesDGX agent

arXiv:2606.03290v1 Announce Type: cross Abstract: Graph Foundation Models (GFMs), built upon the Pre-training and Adaptation paradigm, have emerged as a research hotspot in graph learning. For GNN-bas

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data

SafetyDGX agent

arXiv:2606.02753v1 Announce Type: cross Abstract: Video world models are a foundational generative technology for embodied AI and the Metaverse, yet existing approaches are inherently limited to a sin

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models

SafetyDGX agent

arXiv:2512.05530v2 Announce Type: replace Abstract: Recently, multimodal large language models (MLLMs) have been widely applied to reasoning tasks. However, they suffer from limited multi-rationale se

Multi-Modal Graph Neural Network with Transformer-Guided Adaptive Diffusion for Preclinical Alzheimer Classification

ResearchDGX agent

arXiv:2606.03322v1 Announce Type: cross Abstract: The graphical representation of the brain offers critical insights into diagnosing and prognosing neurodegenerative disease via relationships between

Multiple Choice Learning of Low-Rank Adapters for Language Modeling

ResearchDGX agent

arXiv:2507.10419v3 Announce Type: replace-cross Abstract: We propose LoRA-MCL, a training scheme that extends next-token prediction in language models with a method designed to decode diverse, plausib

MultiTurnPSB: Evaluating Multi-Turn Jailbreak Attacks an dClassifier-Based Defenses for Medical AI Safety

Model ReleasesDGX agent

arXiv:2606.02630v1 Announce Type: cross Abstract: Patient-facing medical chatbots are commonly evaluated on single-turn prompts, yet real users push back after refusals, add urgency, and invoke author

MUSE: A Unified Agentic Harness for MLLMs

AgentsDGX agent

arXiv:2606.03005v1 Announce Type: cross Abstract: Despite rapid progress, multimodal large language models (MLLMs) still fail on tasks that humans solve effortlessly, such as navigating a grid maze fr

NetKV: Network-Aware Decode Instance Selection for Disaggregated LLM Inference

Local AiDGX agent

arXiv:2606.03910v1 Announce Type: cross Abstract: Disaggregated LLM inference forces the KV cache to traverse the datacenter network before decoding begins, so transfer time enters directly into the T

NeuroArmor: Safe-Variant-Guided Representation Consistency for Selective Re-Anchoring in Jailbreak Defense

Model ReleasesDGX agent

arXiv:2606.03486v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak attacks that hide harmful intent behind seemingly ordinary requests such as role-play, translatio

Non-Identical Diffusion Models in MIMO-OFDM Channel Generation

ResearchDGX agent

arXiv:2509.01641v3 Announce Type: replace-cross Abstract: We propose a novel diffusion model, termed the non-identical diffusion model, and investigate its application to wireless orthogonal frequency

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation

SafetyDGX agent

arXiv:2606.03159v1 Announce Type: cross Abstract: As autonomous vehicle capabilities advance, the safe evaluation of driving policies in long-tail scenarios remains a critical bottleneck. In closed-lo

OpenAgenet/OAN: Open Infrastructure for Trusted Agent Interconnection

AgentsDGX agent

arXiv:2606.03161v1 Announce Type: cross Abstract: OpenAgenet, abbreviated as OAN, is an open infrastructure project for trusted Agent interconnection. It addresses a problem that becomes visible when

OpenAgenet/OAN: Technical Architecture for Trust-Governed Agent Identity and Discovery

AgentsDGX agent

arXiv:2606.03163v1 Announce Type: cross Abstract: This paper describes the technical architecture of OpenAgenet / OAN. OAN is a protocol-neutral trust layer for open Agent interconnection. It specifie

Optimizing Explicit Unit-Distance Lower-Bound Certificates

Model ReleasesDGX agent

arXiv:2606.03419v1 Announce Type: cross Abstract: The 2026 disproof of Erdos's unit-distance conjecture and Sawin's subsequent explicit quantitative refinement show that the maximum number u(n) of uni

Oscillatory State-Space Models as Inductive Biases for Physics-Informed Neural PDE Solvers

TutorialsDGX agent

arXiv:2606.02623v1 Announce Type: cross Abstract: Solving time-dependent partial differential equations (PDEs) is an important problem in computational science and engineering. Physics-informed neural

Overlaying Governance: A Compositional Authorization Framework for Delegation and Scope in Agentic AI

AgentsDGX agent

arXiv:2606.03518v1 Announce Type: new Abstract: As AI systems evolve from passive models into autonomous active agents capable of initiating actions, collaborating, and delegating tasks, the tradition

PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification

SafetyDGX agent

arXiv:2602.07768v3 Announce Type: replace-cross Abstract: Distilling knowledge from large Vision-Language Models (VLMs) into lightweight networks is crucial yet challenging in Fine-Grained Visual Clas

Patcher: Post-Hoc Patching of Backdoored Large Language Models

Local AiDGX agent

arXiv:2606.02995v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak backdoor attacks, where adversaries poison safety alignment data to embed hidden triggers that by

Perceive Before Reasoning: A Pre-Reasoning Perception Framework for Efficient and Reliable Proactive Mobile Agents

Model ReleasesDGX agent

arXiv:2606.03236v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have substantially advanced mobile agents, yet proactive mobile assistance remains challenging because agents m

Pextsuperscript{2}-DPO: Grounding Hallucination in Perceptual Processing via Calibration Direct Preference Optimization

SafetyDGX agent

arXiv:2606.03376v1 Announce Type: cross Abstract: Hallucination has recently garnered significant research attention in Large Vision-Language Models (LVLMs). Direct Preference Optimization (DPO) aims

Phantom Transfer: Data Poisoning can Survive Data-Level Defences

ApplicationsDGX agent

arXiv:2602.04899v2 Announce Type: replace-cross Abstract: We present a data poisoning attack -- Phantom Transfer -- with the property that, even if you know precisely how the poison was placed into an

PHASE: Physiology-Aware Hyperspectral Reconstruction via Object-to-Human Domain Adaptation

SafetyDGX agent

arXiv:2511.13020v2 Announce Type: replace-cross Abstract: Although hyperspectral imaging offers unparalleled non-invasive physiological insight, its bulky hardware, slow acquisition, and regulatory bu

PHASER: Phase-Aware and Semantic Experience Replay for Vision-Language-Action Models

AgentsDGX agent

arXiv:2606.03598v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved remarkable success in language-conditioned robotic manipulation. However, deploying these models in

PhotoCraft: Agentic Reasoning with Hierarchical Self-Evolving Memory for Deep Image Search

AgentsDGX agent

arXiv:2606.03099v1 Announce Type: cross Abstract: Deep Image Search requires multi-step reasoning over rich contextual cues, such as time, location, and event relations. However, most existing LLM-bas

Physics-Guided Policy Optimization with Self-Distillation

SafetyDGX agent

arXiv:2606.03620v1 Announce Type: cross Abstract: Self-distilled policy optimization (SDPO) has become a popular paradigm for LLM post-training, where a model learns from its own predictions condition

Physics-informed diffusion models in spectral space

TutorialsDGX agent

arXiv:2602.09708v2 Announce Type: replace-cross Abstract: We propose physics-informed spectral diffusion (PISD), a methodology that combines generative latent diffusion models with physics-informed ma

PieArena: Ranking and Profiling Language Agents in Realistic Negotiation Scenarios

Model ReleasesDGX agent

arXiv:2602.05302v3 Announce Type: replace Abstract: We present an in-depth evaluation of LLMs' ability to negotiate, a central business task requiring strategic reasoning, theory of mind, and economic

PINNfluence: Interpreting PINNs through Influence Functions

Model ReleasesDGX agent

arXiv:2409.08958v3 Announce Type: replace-cross Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful deep learning approach for solving partial differential equations (PDEs) i

Plan, Verify and Fill: A Structured Parallel Decoding Approach for Diffusion Language Models

Model ReleasesDGX agent

arXiv:2601.12247v3 Announce Type: replace-cross Abstract: Diffusion Language Models (DLMs) present a promising non-sequential paradigm for text generation, distinct from standard autoregressive (AR) a

Plan2Map: A Multimodal Benchmark for Document-Grounded Geospatial Boundary Reconstruction from Planning Records

Model ReleasesDGX agent

arXiv:2606.02747v1 Announce Type: cross Abstract: Planning records define restrictions over geographic areas, but their source documents often provide only indirect spatial evidence rather than machin

Planning with Uncertainty: Symmetries, Policy Inference, and Solution Compression

SafetyDGX agent

arXiv:2403.19883v2 Announce Type: replace Abstract: Fully-observable non-deterministic (FOND) planning is at the core of artificial intelligence planning with uncertainty. It models uncertainty throug

Position: Prioritize Identifying Structure, Not Complex Models, for Scientific Discovery

ResearchDGX agent

arXiv:2606.02632v1 Announce Type: cross Abstract: Modern Machine Learning (ML) and Artificial Intelligence (AI) models, especially large language models (LLMs), are increasingly used to generate scien

Post-Hoc Robustness for Model-Based Reinforcement Learning

SafetyDGX agent

arXiv:2606.03521v1 Announce Type: cross Abstract: To improve the real-world applicability of reinforcement learning (RL), the field of adversarially robust RL studies how to train agents under adversa

Pretraining Language Models on Historical Text

Model ReleasesDGX agent

arXiv:2606.02991v1 Announce Type: cross Abstract: We introduce TypewriterLM, a 7.24B History language model (LM) trained exclusively on English text predating 1913. Developing History LMs requires add

PrimeSVT: An Automated Memory-aware Pruning Framework with Prioritized Compression Policy for Spiking Vision Transformers

SafetyDGX agent

arXiv:2606.03428v1 Announce Type: cross Abstract: The large sizes of Spiking Vision Transformers (SViTs) still hinder their embedded implementation, highlighting the need for model compression. State-

PRISM: Synergizing Vision Foundation Models via Self-organized Expert Specialization

ResearchDGX agent

arXiv:2606.03444v1 Announce Type: cross Abstract: Unifying the complementary strengths of diverse Vision Foundation Models (VFMs) into a single efficient model is highly desirable but challenged by th

Proof-Refactor: Refactoring Generated Formal Proofs into Modular Artifacts

Model ReleasesDGX agent

arXiv:2606.03743v1 Announce Type: new Abstract: While Large Language Models (LLMs) have shown strong performance in generating formal proofs, their outputs often remain less readable, modular, maintai

ProtocolBench: Which LLM MultiAgent Protocol to Choose?

Model ReleasesDGX agent

arXiv:2510.17149v3 Announce Type: replace Abstract: As large-scale multi-agent systems evolve, the communication protocol layer has become a critical yet under-evaluated factor shaping performance and

PSViT: A Methodology for Structurally Pruning Spiking Vision Transformers

ResearchDGX agent

arXiv:2606.03257v1 Announce Type: cross Abstract: Spiking Vision Transformer (SViT) models are promising low-power ViT models for solving vision-based tasks with state-of-the-art performance. However,

PURGE: Projected Unlearning via Retain-Guided Erasure

TutorialsDGX agent

arXiv:2606.03808v1 Announce Type: cross Abstract: We propose PURGE, a machine unlearning algorithm built on a simple but an under-exploited observation: continual learning (CL) and machine unlearning

PyraMathBench: Evaluating and Improving Mathematical Capability in Large Language Models

Model ReleasesDGX agent

arXiv:2606.03858v1 Announce Type: new Abstract: Despite the pivotal role of numerical reasoning as the cornerstone of mathematical capabilities in large language models (LLMs) across applications, few

q0: Primitives for Hyper-Epoch Pretraining

Model ReleasesDGX agent

arXiv:2606.03938v1 Announce Type: cross Abstract: Multi-epoch training is becoming the standard now that compute is growing faster than the supply of high-quality text. But pretraining a single model

Quantifying Faithful Confidence Expression in Large Reasoning Models

SafetyDGX agent

arXiv:2606.03969v1 Announce Type: cross Abstract: Reliable uncertainty communication is critical to the trustworthiness of LLMs, yet faithful calibration (FC)--the alignment between models' intrinsic

QUBRIC: Co-Designing Queries and Rubrics for RL Beyond Verifiable Rewards

SafetyDGX agent

arXiv:2606.03968v1 Announce Type: cross Abstract: Rubric-based RL is a promising route for extending reinforcement learning beyond verifiable rewards, yet existing methods optimize rubrics while treat

Qwen-Image-Flash: Beyond Objective Design

Model ReleasesDGX agent

arXiv:2606.03746v1 Announce Type: cross Abstract: Few-step distillation has become an effective strategy for accelerating advanced visual generative models, yet prior work has largely focused on disti

R^{2k} is Theoretically Large Enough for Embedding-based Top-k Retrieval

ResearchDGX agent

arXiv:2601.20844v3 Announce Type: replace-cross Abstract: This paper studies the Minimal Embeddable Dimension (MED): the least dimension in which there exists a configuration of m object vectors so th

Re-Evaluating Continual Learning with Few-Shot Adaptation

TutorialsDGX agent

arXiv:2606.03843v1 Announce Type: cross Abstract: Continual learning methods aim to maximize the stability and plasticity of machine learning models that are trained on a sequence of tasks. The standa

ReaLM: Residual Quantization Bridging Knowledge Graph Embeddings and Large Language Models

Model ReleasesDGX agent

arXiv:2510.09711v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have recently emerged as a powerful paradigm for Knowledge Graph Completion (KGC), offering strong reasoning and

Reasoning Structure of Large Language Models

Model ReleasesDGX agent

arXiv:2606.03883v1 Announce Type: new Abstract: Large reasoning models (LRMs) are often evaluated using metrics such as final-answer accuracy or token count. However, identical scores on these metrics

Ref-DGS: Reflective Dual Gaussian Splatting

ResearchDGX agent

arXiv:2603.07664v3 Announce Type: replace-cross Abstract: The reflective appearance, especially strong and typically near-field specular reflections, poses a fundamental challenge for accurate surface

Regret Pre-training: Bridging Prior and Posterior Views for Enhanced Knowledge Grounding

Local AiDGX agent

arXiv:2606.03080v1 Announce Type: cross Abstract: Causal language models factorize sequence probabilities using only preceding context, leaving future information unexploited during training despite i

Reinforcement Learning from Cross-domain Videos with Video Prediction Model

AgentsDGX agent

arXiv:2606.03201v1 Announce Type: cross Abstract: Reinforcement learning from expert videos across visually distinct domains is challenging due to the absence of reward signals and the presence of dom

Relational Linearity is a Predictor of Hallucinations

Model ReleasesDGX agent

arXiv:2601.11429v2 Announce Type: replace-cross Abstract: Hallucination is a central failure mode of language models (LMs). We focus on hallucinations in response to questions like: 'Which instrument

RelGT-AC: A Relational Graph Transformer for Autocomplete Tasks in Relational Databases

ApplicationsDGX agent

arXiv:2606.03040v1 Announce Type: new Abstract: Relational databases underpin modern enterprise, scientific, and healthcare systems, yet predictive machine learning on such data remains challenging du

← Previous
1…166167168169170…358
Next →