AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
9 Jun 2026

Neuro-Symbolic Injection of LTLf Constraints in Autoregressive Reinforcement Learning Policies

SafetyDGX agent

arXiv:2606.08312v1 Announce Type: new Abstract: In this work we study offline reinforcement learning (RL) under temporally extended task constraints expressed in Linear Temporal Logic over finite trac

NeuroAlign: Hierarchical Multimodal Fusion of Dynamic and Structural Neuroimaging for MCI Analysis

SafetyDGX agent

arXiv:2606.07635v1 Announce Type: cross Abstract: Multimodal neuroimaging fusion of functional MRI (fMRI) and diffusion tensor imaging (DTI) provides complementary information for cognitive impairment

Neutrality Bites: Gender Representation in AI-Generated Animal Stories

SafetyDGX agent

arXiv:2606.07969v1 Announce Type: cross Abstract: Gender bias in AI-generated stories is a well-documented problem. While much attention has been paid to reducing or mitigating this bias, it is not al


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Next-Token Prediction Learns Generalisable Representations of Sleep Physiology

ApplicationsDGX agent

arXiv:2606.09605v1 Announce Type: new Abstract: Foundation models offer a promising route to compress multi-modal physiological signals into compact representations of human health, with broad applica

No Free Lunch for Synthetic Images under Data Scarcity Conditions

ResearchDGX agent

arXiv:2606.07640v1 Announce Type: cross Abstract: This study investigates the trade-offs between fidelity, privacy, and utility in synthetic data generation under conditions of data scarcity and priva

NoRD: A Data-Efficient Vision-Language-Action Model that Drives without Reasoning

SafetyDGX agent

arXiv:2602.21172v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are advancing autonomous driving by replacing modular pipelines with unified end-to-end architectures. However,

Not Just After One: Sleep-Inspired Replay Prevents Catastrophic Forgetting After Sequential Tasks

TutorialsDGX agent

arXiv:2606.08447v1 Announce Type: cross Abstract: One of the critical limitations of artificial neural networks is their lack of ability to continually learn: training on new tasks often leads to inte

NutriMLLM: Multimodal Large Language Models for Dietary Micronutrient Analysis

Model ReleasesDGX agent

arXiv:2606.08948v1 Announce Type: cross Abstract: Comprehensive estimation of dietary micronutrients from food images could improve clinical nutrition care, but training such models requires large mul

Observability for Delegated Execution in Agentic AI Systems

AgentsDGX agent

arXiv:2606.09692v1 Announce Type: cross Abstract: Delegation-scoped execution is not identifiable from standard observables: audit logs and execution traces can be identical under multiple incompatibl

Offline Reinforcement Learning for Plasma Control in Nuclear Fusion: Codebase and Benchmark

Model ReleasesDGX agent

arXiv:2606.07550v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) offers a promising route for developing plasma controllers from historical tokamak data, since online trial-and-er

OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics

Model ReleasesDGX agent

arXiv:2606.09826v1 Announce Type: cross Abstract: Vision-language model (VLM) agents are increasingly deployed in interactive game environments. Yet game benchmarks for VLM agents typically report a s

OmniMem: Perturbation-aware Memory Compression for Streaming Audio-Visual LLMs

Model ReleasesDGX agent

arXiv:2606.07577v1 Announce Type: new Abstract: Audio-visual large language models (LLMs) hold strong promise for long-form video understanding, yet their long-video inference is fundamentally limited

On the Complexity of Offline Reinforcement Learning with Q^star-Approximation and Partial Coverage

ResearchDGX agent

arXiv:2602.12107v2 Announce Type: replace-cross Abstract: We study offline reinforcement learning under Q^star-approximation and partial coverage, a setting that motivates practical algorithms such as

One if by Land, Two if by Sea, Three if by Four Seas, and More to Come -- Values of Perception, Prediction, Communication, and Common Sense in Decision Making

AgentsDGX agent

arXiv:2601.06077v2 Announce Type: replace-cross Abstract: This work aims to rigorously define the values of perception, prediction, communication, and common sense in decision making. The defined quan

Online Agent-as-a-Judge: Situation-Generating Evaluation for Interactive Agents

AgentsDGX agent

arXiv:2606.08200v1 Announce Type: new Abstract: Evaluating LLM-powered interactive social agents is challenging because socially relevant behaviors depend not only on isolated outputs, but also on pri

OnlyDense: Reduced-Order Modeling for Lagrangian simulation

ResearchDGX agent

arXiv:2606.09065v1 Announce Type: cross Abstract: In science and engineering, Lagrangian simulation methods such as Smooth Particle Hydrodynamics (SPH) or Material Point Method (MPM) are often employe

Optical Reasoning: Rethinking Images as an Expressive Reasoning Medium Beyond Text

ResearchDGX agent

arXiv:2606.09585v1 Announce Type: new Abstract: Chain-of-Thought (CoT) improves the performance of Large Language Models (LLMs) and has been extended to Multimodal Large Language Models (MLLMs). More

Optimizing Energy-based Neural Network Training with Coherent Ising Machine

ResearchDGX agent

arXiv:2606.09117v1 Announce Type: cross Abstract: While Ising machines serve as advanced physical solvers for the Ising model,enabling applications in combinatorial optimization and neural network tra

Order Matters: Unveiling the Hidden Impact of Macro Placement Sequences via Proxy-Guided LLM Evolution

ResearchDGX agent

arXiv:2606.08904v1 Announce Type: new Abstract: Macro placement is a fundamental step in modern chip physical design, playing a crucial role in determining the solution quality of high-dimensional com

OSMGraphCLIP: Learning Global Location Representations from OpenStreetMap Graphs

Local AiDGX agent

arXiv:2606.08046v1 Announce Type: new Abstract: We present OSMGraphCLIP, a CLIP-style geospatial representation model that learns global location embeddings from freely available OpenStreetMap (OSM) d

Outage Detection in Self-Healing Smart Grids Using Reinforcement Learning with Spectral Graph Neural Networks

SafetyDGX agent

arXiv:2606.07583v1 Announce Type: cross Abstract: Self-healing smart grids can quickly adjust their network configuration during outages to minimize power disruptions. During an outage, several action

Overcoming the Regulatory Bottleneck via Agent-to-Agent Protocols: A Nuclear Case Study

SafetyDGX agent

arXiv:2606.07866v1 Announce Type: new Abstract: Regulatory review of advanced nuclear reactor designs routinely spans more than three years and consumes hundreds of millions of dollars in combined reg

Oversight Has a Capacity: Calibrating Agent Guards to a Subjective, Fatiguing Human

SafetyDGX agent

arXiv:2606.08919v1 Announce Type: new Abstract: As LLM agents begin to take real, irreversible actions (shell commands, file edits, deploys), the standard safety pattern is a human-in-the-loop approva

PACE: Anytime-Valid Acceptance Tests for Self-Evolving Agents

AgentsDGX agent

arXiv:2606.08106v1 Announce Type: new Abstract: Self-evolving agents improve by repeatedly proposing changes to their own prompts, skills, or workflows and keeping those that score higher on a small h

PACT: Learning Diverse Diagnostic Strategies via Privileged Synthesis and Branch Consensus

Model ReleasesDGX agent

arXiv:2606.08938v1 Announce Type: cross Abstract: Clinical diagnosis requires flexible use of multiple reasoning paradigms under incomplete patient information. Existing LLM-based medical agents show

PACT: Self-Evolving Physical Safety Alignment for Diffusion Policies in Embodied Manipulation

SafetyDGX agent

arXiv:2606.08414v1 Announce Type: cross Abstract: Diffusion policies have achieved remarkable success in robotic manipulation, yet they often fail to satisfy strict physical constraints required for s

PAEC: Position-Aware Entropy Calibration for LLM Reasoning in RLVR

SafetyDGX agent

arXiv:2606.08543v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves large language model reasoning but often suffers from rapid policy-entropy collapse, wher

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling

SafetyDGX agent

arXiv:2606.07988v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on reward models to align their outputs with diverse user preferences. While personalized reward models a

Page image classifier fine-tuned on century-spanning archives of scanned documents for further content-specific processing

ResearchDGX agent

arXiv:2606.07558v1 Announce Type: cross Abstract: Purpose: Digitization projects in the humanities produce vast, heterogeneous archives of historical documents, making manual sorting impractical at sc

PAI: Preserving Amplitude Information in Representation-Based Time-Series Anomaly Detection

Local AiDGX agent

arXiv:2606.08935v1 Announce Type: cross Abstract: Representation-based time-series anomaly detection algorithms significantly outperform other methods on diverse anomaly detection tasks. However, we n

PathoSage: Towards Multi-Source Evidence Adjudication in Pathology via Experience-Aware Agentic Workflow

SafetyDGX agent

arXiv:2606.07549v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and agent workflows have shown strong promise for computational pathology, yet reliable patc

Payoff scaling shapes cooperation in LLM agents across languages

SafetyDGX agent

arXiv:2601.19082v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that negotiate, coordinate, and act on behalf of users. Whether they coo

Performative Learning Theory

TutorialsDGX agent

arXiv:2602.04402v3 Announce Type: replace-cross Abstract: Performative predictions influence the very outcomes they aim to forecast. We study performative predictions that affect a sample (e.g., only

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs

Model ReleasesDGX agent

arXiv:2606.09038v1 Announce Type: new Abstract: Large Language Models (LLMs) have enabled increasingly personalized interactions by adapting to users' preferences, contexts, and long-term histories. H

Phantom transitions in language model fine-tuning

Model ReleasesDGX agent

arXiv:2606.07559v1 Announce Type: cross Abstract: Fine-tuning a language model on contexts whose correct completion has a near-synonym competitor often fails silently. The cross-entropy loss decreases

Pharmacogenomic Knowledge Graph Augmentation for Graph Neural Network-Based Drug-Drug Interaction Prediction

Model ReleasesDGX agent

arXiv:2606.07698v1 Announce Type: cross Abstract: Graph neural networks (GNNs) applied to drug-drug interaction (DDI) prediction rely exclusively on molecular structure encoded as SMILES-derived graph

Physics-Guided Sequence-Based Generative Framework for Acoustic Metamaterial Inverse Design

ResearchDGX agent

arXiv:2606.09266v1 Announce Type: cross Abstract: Acoustic metamaterial (AMM) inverse design is particularly challenging for broadband target responses due to acoustic dispersion: a structure that mat

PhysScene: A Scene Graph Dataset for Scientific Visual Reasoning in Physics Experiments

ResearchDGX agent

arXiv:2606.09368v1 Announce Type: cross Abstract: Scene Graphs (SGs) provide structured representations of visual scenes by modeling objects and their pairwise relationships. Despite recent progress,

PIPE-Cypher: Automatic Enterprise Benchmark Generation for Text-to-Cypher Systems

Model ReleasesDGX agent

arXiv:2606.08481v1 Announce Type: cross Abstract: Enterprise property graphs vary widely in schema structure, internal terminology, domain assumptions, governance constraints, and user interaction pat

PLAGUE: Plug-and-play framework for Lifelong Adaptive Generation of Multi-turn Exploits

Model ReleasesDGX agent

arXiv:2510.17947v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are improving at an exceptional rate. With the advent of agentic workflows, multi-turn dialogue has become the de

POET-X: Memory-efficient LLM Training by Scaling Orthogonal Transformation

Model ReleasesDGX agent

arXiv:2603.05500v2 Announce Type: replace-cross Abstract: Efficient and stable training of large language models (LLMs) remains a core challenge in modern machine learning systems. To address this cha

POISE: Position-Aware Undetectable Skill Injection on LLM Agents

Model ReleasesDGX agent

arXiv:2606.07943v1 Announce Type: cross Abstract: Agent skills provide a lightweight mechanism for extending general-purpose agents, but their open format exposes them to skill-poisoning attacks. A pr

PolyBuild: An End-to-End Method for Polygonal Building Contour Extraction from High-Resolution Remote Sensing Images

ResearchDGX agent

arXiv:2606.08920v1 Announce Type: cross Abstract: Extracting building polygon contours from high-resolution remote sensing images is a fundamental task for various mapping applications. However, the p

Position: Anthropomorphic Misalignment Research Needs Stronger Evidence

SafetyDGX agent

arXiv:2606.07612v1 Announce Type: cross Abstract: We argue that many Anthropomorphic Misalignment Research (AMR) studies need stronger evidence to ensure that they can provide a robust foundation for

Post-AGI Economies: Superposition and the Second Fundamental Theorem of Welfare Economics

ResearchDGX agent

arXiv:2606.08267v1 Announce Type: cross Abstract: The classical Second Welfare Theorem decentralizes any Pareto efficient allocation through prices and transfers under convexity and regularity. In pos

Post-training is (Massive) Supervised Learning

TutorialsDGX agent

arXiv:2606.07527v1 Announce Type: cross Abstract: The prevailing paradigm for training LLMs has evolved to rely on a massive post-training phase consisting of SFT and RL. In this position paper, we ar

Powering the Future of AI: Navigating the Trade-offs for Europe's Energy Transition and Net-Zero Goals

ResearchDGX agent

arXiv:2606.09617v1 Announce Type: cross Abstract: The rapid expansion of AI globally has led to the proliferation of energy-intensive hyperscale data centres (DCs), making them as a structurally chall

Pre-Intervention Prediction of Sparse Autoencoder Steering Side Effects

Model ReleasesDGX agent

arXiv:2606.08365v1 Announce Type: cross Abstract: Sparse autoencoder (SAE) features are increasingly used to steer language models, but feature steering is rarely clean: the same intervention can beha

Prescriptive Scaling Reveals the Evolution of Language Model Capabilities

Model ReleasesDGX agent

arXiv:2602.15327v2 Announce Type: replace-cross Abstract: Machine learning model performance improvements tend to arise from competition and application. For deployment, we consider prescriptive scali

Preserving Plasticity in Continual Learning via Dynamical Isometry

ResearchDGX agent

arXiv:2606.09762v1 Announce Type: cross Abstract: Continual training of deep neural networks under non-stationarity often leads to a progressive loss of plasticity, eventually limiting further learnin

Pretrained, Frozen, Still Leaking: Auditing Cross-Encoder Attribute Transfer in EEG Foundation Models

Model ReleasesDGX agent

arXiv:2606.09189v1 Announce Type: cross Abstract: EEG foundation-model releases are usually audited one endpoint at a time: raw-reconstruction, membership inference, identity linkage, or DP-SGD on the

Principled Agent Debate: Adversarial Arbitration for Sycophancy Reduction in Large Language Models

Model ReleasesDGX agent

arXiv:2606.07532v1 Announce Type: cross Abstract: RLHF-trained models are systematically biased toward agreement over accuracy, a structural property of the training process. We present Principled Age

PRISM: PRior-guided Imagination Sampling in world Models

Model ReleasesDGX agent

arXiv:2606.07974v1 Announce Type: cross Abstract: A learned world model provides a powerful physical intuition for evaluating future states. But its effectiveness in continuous control also depends cr

PRISM: Recovering Instruction Sets from Language Model Activations

AgentsDGX agent

arXiv:2606.09563v1 Announce Type: new Abstract: As LLMs are deployed as agents, reliable monitoring requires knowing not only what they output, but which instructions are steering their behavior. This

Projecting the Emerging Mindset of SWE Agent by Launching a Wild Code Understanding Journey

AgentsDGX agent

arXiv:2606.08500v1 Announce Type: cross Abstract: Software engineering agents (SWE agents) increasingly work through tool-mediated trajectories in real repositories, yet their behavior remains difficu

Projection and Quantisation: A Unifying View of Learning to Hash, from Random Projections to the RAG Era

Model ReleasesDGX agent

arXiv:2510.04127v2 Announce Type: replace-cross Abstract: Approximate nearest neighbour (ANN) search underpins large-scale retrieval, increasingly within the retrieval-augmented generation pipelines t

Proposal Refinement for Few-Shot Object Detection

ResearchDGX agent

arXiv:2606.09245v1 Announce Type: cross Abstract: Few-shot object detection has gained widely attention in recent years. Some excellent algorithms have been proposed to handle this task. However, most

Provably Efficient Personalized Multi-Objective Bandits with Proactive Conversational Queries

ResearchDGX agent

arXiv:2606.08410v1 Announce Type: cross Abstract: Personalized decision-making in multi-objective bandits requires learning user-specific trade-offs among competing objectives. Since arm utility depen

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization

SafetyDGX agent

arXiv:2606.09711v1 Announce Type: new Abstract: Reward hacking is usually studied after it becomes visible, once a model earns high proxy reward while failing the intended task. We instead study what

PTL-Diffusion: Manifold-Aware Diffusion with Periodic Terminal Laws

Local AiDGX agent

arXiv:2606.09816v1 Announce Type: cross Abstract: Standard diffusion models typically use a single time-homogeneous Gaussian terminal distribution as the reference law for generation. While this choic

← Previous
1…146147148149150…358
Next →