AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlog
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

On the Complexity of Offline Reinforcement Learning with Q^star-Approximation and Partial Coverage

DGX agent

arXiv:2602.12107v2 Announce Type: replace-cross Abstract: We study offline reinforcement learning under Q^star-approximation and partial coverage, a setting that motivates practical algorithms such as

researcharxiv-cs-ai
9 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

One if by Land, Two if by Sea, Three if by Four Seas, and More to Come -- Values of Perception, Prediction, Communication, and Common Sense in Decision Making

DGX agent

arXiv:2601.06077v2 Announce Type: replace-cross Abstract: This work aims to rigorously define the values of perception, prediction, communication, and common sense in decision making. The defined quan

agentsarxiv-cs-ai
9 Jun 2026
Agents

Online Agent-as-a-Judge: Situation-Generating Evaluation for Interactive Agents

DGX agent

arXiv:2606.08200v1 Announce Type: new Abstract: Evaluating LLM-powered interactive social agents is challenging because socially relevant behaviors depend not only on isolated outputs, but also on pri

agentsarxiv-cs-ai
9 Jun 2026
Research

OnlyDense: Reduced-Order Modeling for Lagrangian simulation

DGX agent

arXiv:2606.09065v1 Announce Type: cross Abstract: In science and engineering, Lagrangian simulation methods such as Smooth Particle Hydrodynamics (SPH) or Material Point Method (MPM) are often employe

researcharxiv-cs-ai
9 Jun 2026
Research

Optical Reasoning: Rethinking Images as an Expressive Reasoning Medium Beyond Text

DGX agent

arXiv:2606.09585v1 Announce Type: new Abstract: Chain-of-Thought (CoT) improves the performance of Large Language Models (LLMs) and has been extended to Multimodal Large Language Models (MLLMs). More

researcharxiv-cs-ai
9 Jun 2026
Research

Optimizing Energy-based Neural Network Training with Coherent Ising Machine

DGX agent

arXiv:2606.09117v1 Announce Type: cross Abstract: While Ising machines serve as advanced physical solvers for the Ising model,enabling applications in combinatorial optimization and neural network tra

researcharxiv-cs-ai
9 Jun 2026
Research

Order Matters: Unveiling the Hidden Impact of Macro Placement Sequences via Proxy-Guided LLM Evolution

DGX agent

arXiv:2606.08904v1 Announce Type: new Abstract: Macro placement is a fundamental step in modern chip physical design, playing a crucial role in determining the solution quality of high-dimensional com

researcharxiv-cs-ai
9 Jun 2026
Local Ai

OSMGraphCLIP: Learning Global Location Representations from OpenStreetMap Graphs

DGX agent

arXiv:2606.08046v1 Announce Type: new Abstract: We present OSMGraphCLIP, a CLIP-style geospatial representation model that learns global location embeddings from freely available OpenStreetMap (OSM) d

local-aiarxiv-cs-ai
9 Jun 2026
Safety

Outage Detection in Self-Healing Smart Grids Using Reinforcement Learning with Spectral Graph Neural Networks

DGX agent

arXiv:2606.07583v1 Announce Type: cross Abstract: Self-healing smart grids can quickly adjust their network configuration during outages to minimize power disruptions. During an outage, several action

safetyarxiv-cs-ai
9 Jun 2026
Safety

Overcoming the Regulatory Bottleneck via Agent-to-Agent Protocols: A Nuclear Case Study

DGX agent

arXiv:2606.07866v1 Announce Type: new Abstract: Regulatory review of advanced nuclear reactor designs routinely spans more than three years and consumes hundreds of millions of dollars in combined reg

safetyarxiv-cs-ai
9 Jun 2026
Safety

Oversight Has a Capacity: Calibrating Agent Guards to a Subjective, Fatiguing Human

DGX agent

arXiv:2606.08919v1 Announce Type: new Abstract: As LLM agents begin to take real, irreversible actions (shell commands, file edits, deploys), the standard safety pattern is a human-in-the-loop approva

safetyarxiv-cs-ai
9 Jun 2026
Agents

PACE: Anytime-Valid Acceptance Tests for Self-Evolving Agents

DGX agent

arXiv:2606.08106v1 Announce Type: new Abstract: Self-evolving agents improve by repeatedly proposing changes to their own prompts, skills, or workflows and keeping those that score higher on a small h

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

PACT: Learning Diverse Diagnostic Strategies via Privileged Synthesis and Branch Consensus

DGX agent

arXiv:2606.08938v1 Announce Type: cross Abstract: Clinical diagnosis requires flexible use of multiple reasoning paradigms under incomplete patient information. Existing LLM-based medical agents show

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

PACT: Self-Evolving Physical Safety Alignment for Diffusion Policies in Embodied Manipulation

DGX agent

arXiv:2606.08414v1 Announce Type: cross Abstract: Diffusion policies have achieved remarkable success in robotic manipulation, yet they often fail to satisfy strict physical constraints required for s

safetyarxiv-cs-ai
9 Jun 2026
Safety

PAEC: Position-Aware Entropy Calibration for LLM Reasoning in RLVR

DGX agent

arXiv:2606.08543v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves large language model reasoning but often suffers from rapid policy-entropy collapse, wher

safetyarxiv-cs-ai
9 Jun 2026
Safety

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling

DGX agent

arXiv:2606.07988v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on reward models to align their outputs with diverse user preferences. While personalized reward models a

safetyarxiv-cs-ai
9 Jun 2026
Research

Page image classifier fine-tuned on century-spanning archives of scanned documents for further content-specific processing

DGX agent

arXiv:2606.07558v1 Announce Type: cross Abstract: Purpose: Digitization projects in the humanities produce vast, heterogeneous archives of historical documents, making manual sorting impractical at sc

researcharxiv-cs-ai
9 Jun 2026
Local Ai

PAI: Preserving Amplitude Information in Representation-Based Time-Series Anomaly Detection

DGX agent

arXiv:2606.08935v1 Announce Type: cross Abstract: Representation-based time-series anomaly detection algorithms significantly outperform other methods on diverse anomaly detection tasks. However, we n

local-aiarxiv-cs-ai
9 Jun 2026
Safety

PathoSage: Towards Multi-Source Evidence Adjudication in Pathology via Experience-Aware Agentic Workflow

DGX agent

arXiv:2606.07549v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and agent workflows have shown strong promise for computational pathology, yet reliable patc

safetyarxiv-cs-ai
9 Jun 2026
Safety

Payoff scaling shapes cooperation in LLM agents across languages

DGX agent

arXiv:2601.19082v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that negotiate, coordinate, and act on behalf of users. Whether they coo

safetyarxiv-cs-ai
9 Jun 2026
Tutorials

Performative Learning Theory

DGX agent

arXiv:2602.04402v3 Announce Type: replace-cross Abstract: Performative predictions influence the very outcomes they aim to forecast. We study performative predictions that affect a sample (e.g., only

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs

DGX agent

arXiv:2606.09038v1 Announce Type: new Abstract: Large Language Models (LLMs) have enabled increasingly personalized interactions by adapting to users' preferences, contexts, and long-term histories. H

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Phantom transitions in language model fine-tuning

DGX agent

arXiv:2606.07559v1 Announce Type: cross Abstract: Fine-tuning a language model on contexts whose correct completion has a near-synonym competitor often fails silently. The cross-entropy loss decreases

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Pharmacogenomic Knowledge Graph Augmentation for Graph Neural Network-Based Drug-Drug Interaction Prediction

DGX agent

arXiv:2606.07698v1 Announce Type: cross Abstract: Graph neural networks (GNNs) applied to drug-drug interaction (DDI) prediction rely exclusively on molecular structure encoded as SMILES-derived graph

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Physics-Guided Sequence-Based Generative Framework for Acoustic Metamaterial Inverse Design

DGX agent

arXiv:2606.09266v1 Announce Type: cross Abstract: Acoustic metamaterial (AMM) inverse design is particularly challenging for broadband target responses due to acoustic dispersion: a structure that mat

researcharxiv-cs-ai
9 Jun 2026
Research

PhysScene: A Scene Graph Dataset for Scientific Visual Reasoning in Physics Experiments

DGX agent

arXiv:2606.09368v1 Announce Type: cross Abstract: Scene Graphs (SGs) provide structured representations of visual scenes by modeling objects and their pairwise relationships. Despite recent progress,

researcharxiv-cs-ai
9 Jun 2026
Model Releases

PIPE-Cypher: Automatic Enterprise Benchmark Generation for Text-to-Cypher Systems

DGX agent

arXiv:2606.08481v1 Announce Type: cross Abstract: Enterprise property graphs vary widely in schema structure, internal terminology, domain assumptions, governance constraints, and user interaction pat

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

PLAGUE: Plug-and-play framework for Lifelong Adaptive Generation of Multi-turn Exploits

DGX agent

arXiv:2510.17947v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are improving at an exceptional rate. With the advent of agentic workflows, multi-turn dialogue has become the de

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

POET-X: Memory-efficient LLM Training by Scaling Orthogonal Transformation

DGX agent

arXiv:2603.05500v2 Announce Type: replace-cross Abstract: Efficient and stable training of large language models (LLMs) remains a core challenge in modern machine learning systems. To address this cha

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

POISE: Position-Aware Undetectable Skill Injection on LLM Agents

DGX agent

arXiv:2606.07943v1 Announce Type: cross Abstract: Agent skills provide a lightweight mechanism for extending general-purpose agents, but their open format exposes them to skill-poisoning attacks. A pr

model-releasesarxiv-cs-ai
9 Jun 2026
Research

PolyBuild: An End-to-End Method for Polygonal Building Contour Extraction from High-Resolution Remote Sensing Images

DGX agent

arXiv:2606.08920v1 Announce Type: cross Abstract: Extracting building polygon contours from high-resolution remote sensing images is a fundamental task for various mapping applications. However, the p

researcharxiv-cs-ai
9 Jun 2026
Safety

Position: Anthropomorphic Misalignment Research Needs Stronger Evidence

DGX agent

arXiv:2606.07612v1 Announce Type: cross Abstract: We argue that many Anthropomorphic Misalignment Research (AMR) studies need stronger evidence to ensure that they can provide a robust foundation for

safetyarxiv-cs-ai
9 Jun 2026
Research

Post-AGI Economies: Superposition and the Second Fundamental Theorem of Welfare Economics

DGX agent

arXiv:2606.08267v1 Announce Type: cross Abstract: The classical Second Welfare Theorem decentralizes any Pareto efficient allocation through prices and transfers under convexity and regularity. In pos

researcharxiv-cs-ai
9 Jun 2026
Tutorials

Post-training is (Massive) Supervised Learning

DGX agent

arXiv:2606.07527v1 Announce Type: cross Abstract: The prevailing paradigm for training LLMs has evolved to rely on a massive post-training phase consisting of SFT and RL. In this position paper, we ar

tutorialsarxiv-cs-ai
9 Jun 2026
Research

Powering the Future of AI: Navigating the Trade-offs for Europe's Energy Transition and Net-Zero Goals

DGX agent

arXiv:2606.09617v1 Announce Type: cross Abstract: The rapid expansion of AI globally has led to the proliferation of energy-intensive hyperscale data centres (DCs), making them as a structurally chall

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Pre-Intervention Prediction of Sparse Autoencoder Steering Side Effects

DGX agent

arXiv:2606.08365v1 Announce Type: cross Abstract: Sparse autoencoder (SAE) features are increasingly used to steer language models, but feature steering is rarely clean: the same intervention can beha

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Prescriptive Scaling Reveals the Evolution of Language Model Capabilities

DGX agent

arXiv:2602.15327v2 Announce Type: replace-cross Abstract: Machine learning model performance improvements tend to arise from competition and application. For deployment, we consider prescriptive scali

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Preserving Plasticity in Continual Learning via Dynamical Isometry

DGX agent

arXiv:2606.09762v1 Announce Type: cross Abstract: Continual training of deep neural networks under non-stationarity often leads to a progressive loss of plasticity, eventually limiting further learnin

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Pretrained, Frozen, Still Leaking: Auditing Cross-Encoder Attribute Transfer in EEG Foundation Models

DGX agent

arXiv:2606.09189v1 Announce Type: cross Abstract: EEG foundation-model releases are usually audited one endpoint at a time: raw-reconstruction, membership inference, identity linkage, or DP-SGD on the

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Principled Agent Debate: Adversarial Arbitration for Sycophancy Reduction in Large Language Models

DGX agent

arXiv:2606.07532v1 Announce Type: cross Abstract: RLHF-trained models are systematically biased toward agreement over accuracy, a structural property of the training process. We present Principled Age

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

PRISM: PRior-guided Imagination Sampling in world Models

DGX agent

arXiv:2606.07974v1 Announce Type: cross Abstract: A learned world model provides a powerful physical intuition for evaluating future states. But its effectiveness in continuous control also depends cr

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

PRISM: Recovering Instruction Sets from Language Model Activations

DGX agent

arXiv:2606.09563v1 Announce Type: new Abstract: As LLMs are deployed as agents, reliable monitoring requires knowing not only what they output, but which instructions are steering their behavior. This

agentsarxiv-cs-ai
9 Jun 2026
Agents

Projecting the Emerging Mindset of SWE Agent by Launching a Wild Code Understanding Journey

DGX agent

arXiv:2606.08500v1 Announce Type: cross Abstract: Software engineering agents (SWE agents) increasingly work through tool-mediated trajectories in real repositories, yet their behavior remains difficu

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Projection and Quantisation: A Unifying View of Learning to Hash, from Random Projections to the RAG Era

DGX agent

arXiv:2510.04127v2 Announce Type: replace-cross Abstract: Approximate nearest neighbour (ANN) search underpins large-scale retrieval, increasingly within the retrieval-augmented generation pipelines t

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Proposal Refinement for Few-Shot Object Detection

DGX agent

arXiv:2606.09245v1 Announce Type: cross Abstract: Few-shot object detection has gained widely attention in recent years. Some excellent algorithms have been proposed to handle this task. However, most

researcharxiv-cs-ai
9 Jun 2026
Research

Provably Efficient Personalized Multi-Objective Bandits with Proactive Conversational Queries

DGX agent

arXiv:2606.08410v1 Announce Type: cross Abstract: Personalized decision-making in multi-objective bandits requires learning user-specific trade-offs among competing objectives. Since arm utility depen

researcharxiv-cs-ai
9 Jun 2026
Safety

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization

DGX agent

arXiv:2606.09711v1 Announce Type: new Abstract: Reward hacking is usually studied after it becomes visible, once a model earns high proxy reward while failing the intended task. We instead study what

safetyarxiv-cs-ai
9 Jun 2026
Local Ai

PTL-Diffusion: Manifold-Aware Diffusion with Periodic Terminal Laws

DGX agent

arXiv:2606.09816v1 Announce Type: cross Abstract: Standard diffusion models typically use a single time-homogeneous Gaussian terminal distribution as the reference law for generation. While this choic

local-aiarxiv-cs-ai
9 Jun 2026
← Previous
1…183184185186187…448
Next →