AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
30 Jun 2026

MirrorCode: AI can rebuild entire programs from behavior alone

Model ReleasesDGX agent

arXiv:2606.30182v1 Announce Type: new Abstract: AI models are rapidly improving at autonomous coding, as shown by benchmark progress and one-off demonstrations such as AI implementing a C compiler. Ho

Mitigating the Safety-utility Trade-off in LLM Alignment via Adaptive Safe Context Learning

SafetyDGX agent

arXiv:2602.13562v2 Announce Type: replace-cross Abstract: While reasoning models have achieved remarkable success in complex reasoning tasks, their increasing power necessitates stringent safety measu

Mixture of Debaters: Learn to Debate at Architectural Level in Multi-Agent Reasoning

Local AiDGX agent

arXiv:2606.29425v1 Announce Type: new Abstract: Existing multi-agent debate frameworks suffer from two critical limitations: they rely on static architectures where agent roles and coordination patter


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Model Merging to Evolution: Parameter Space Exploration for Expert Models

Model ReleasesDGX agent

arXiv:2606.28373v1 Announce Type: cross Abstract: Model merging integrates the capabilities of multiple expert models to create strong models for multiple tasks without additional training, thereby re

Model Predictive Current Control with Harmonic Correction for Single-Phase AC-DC EV Charging

Model ReleasesDGX agent

arXiv:2606.30397v1 Announce Type: cross Abstract: The increasing integration of Electric Vehicles (EVs) has imposed a growing harmonic challenge on the power grid. For AC/DC Power Factor Correction (P

Modeling Earth-Scale Human-Like Societies with One Billion Agents

AgentsDGX agent

arXiv:2506.12078v2 Announce Type: replace-cross Abstract: Understanding the dynamic evolution of complex social phenomena requires both high-fidelity modeling of human behavior and large-scale simulat

Modelling Human Values for Value-Aware Multi-Agent Systems

SafetyDGX agent

arXiv:2402.06359v2 Announce Type: replace Abstract: One of today's most pressing societal challenges is building AI systems whose behaviour, or the behaviour it enables within communities of interacti

Modification-Considering Value Learning for Reward Hacking Mitigation in RL

SafetyDGX agent

arXiv:2606.28955v1 Announce Type: cross Abstract: Reinforcement learning agents can exploit misspecified reward signals to achieve high apparent returns while failing on the intended objective, a fail

Monte Carlo Query Search: Active Capability Assessment of AI Agents

AgentsDGX agent

arXiv:2512.16733v3 Announce Type: replace Abstract: Black-box AI (BBAI) systems, including foundation-model agents, are increasingly used for sequential decision making. Safe deployment requires metho

MoPe: Motion Permanence for Robust Monocular Gaussian Mapping in Dynamic Environments

ApplicationsDGX agent

arXiv:2606.29237v1 Announce Type: cross Abstract: Robust robot autonomy depends on scene representations that remain stable enough to support localization, navigation, and downstream decision making i

MotionAtlas: Detailed Region Captioning for Motion-Centric Videos

Model ReleasesDGX agent

arXiv:2606.29531v1 Announce Type: cross Abstract: We propose MotionAtlas, a system for detailed captioning of motion-centric videos, comprising (1) a dedicated human-annotated benchmark, (2) a scalabl

Multi-Agent DRL for QoS and Energy Optimization in RIS-Enabled Open-RAN Industrial 6G TN/NTN Networks

AgentsDGX agent

arXiv:2606.28339v1 Announce Type: cross Abstract: Industrial 6G networks require ultra-reliable, low-latency, and energy-efficient connectivity in dynamic and blockage-prone environments, where conven

Multi-Agent Routing as Set-Valued Prediction: A WildChat Benchmark and Cost-Aware Evaluation

Model ReleasesDGX agent

arXiv:2606.28925v1 Announce Type: cross Abstract: Tool and agent routing from natural-language prompts is naturally a set-valued prediction problem: a single query may require multiple agents, while o

Multi-Class Human/Object Detection on Robot Manipulators using Proprioceptive Sensing

SafetyDGX agent

arXiv:2508.02425v2 Announce Type: replace-cross Abstract: In physical human-robot collaboration (pHRC) settings, humans and robots collaborate directly in shared environments. Robots must analyze inte

Multi-Level Distributional Entropy for Explainable Network Intrusion Detection

ResearchDGX agent

arXiv:2606.29797v1 Announce Type: cross Abstract: Machine learning network intrusion detection systems (IDS) rely on aggregate flow statistics that discard distributional structure, while established

Multimodal and Multiscale Spatial-Temporal Semantic Search and Recommendation with AI Foundation Models

Local AiDGX agent

arXiv:2606.28369v1 Announce Type: cross Abstract: Semantic search and recommendation of similar documents, such as news and reports about unusual environmental events (e.g., a dead whale washed ashore

Multimodal Representation Alignment for Cross-modal Information Retrieval

SafetyDGX agent

arXiv:2506.08774v2 Announce Type: replace-cross Abstract: Different machine learning models can represent the same underlying concept in different ways. This variability is particularly valuable for i

MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs

Model ReleasesDGX agent

arXiv:2606.30026v1 Announce Type: cross Abstract: Audiovisual arts encompass diverse creative disciplines, including cinema, visual arts, stage performance, and game design, where artistic meaning ari

Neural Minimum Weight Perfect Matching for Quantum Error Codes

ResearchDGX agent

arXiv:2601.00242v2 Announce Type: replace-cross Abstract: Realizing the full potential of quantum computation requires Quantum Error Correction (QEC). QEC reduces error rates by encoding logical infor

Neural Procedural Memory: Empowering LLM Agents with Implicit Activation Steering

AgentsDGX agent

arXiv:2606.29824v1 Announce Type: cross Abstract: While Large Language Models (LLMs) excel as static solvers, transforming them into autonomous agents remains challenging. This transition requires con

Neural Subspace Reallocation: Continual Learning as Retrieval-Based Subspace Memory Management

Model ReleasesDGX agent

arXiv:2606.30067v1 Announce Type: cross Abstract: We introduce Neural Subspace Reallocation (NSR), which reframes continual learning as memory management over parameter subspaces. Instead of treating

NeuralMUSIC: A Hybrid Neural-Subspace Framework for Robot Sound Source Localization

AgentsDGX agent

arXiv:2606.18664v2 Announce Type: replace-cross Abstract: Reliable sound source localization is fundamental to robot audition, enabling autonomous robots to perceive spatial cues and operate effective

Neuromorphic Energy-Aware Learning for Adaptive Deep Brain Stimulation

SafetyDGX agent

arXiv:2606.28600v1 Announce Type: cross Abstract: Neuromorphic and edge computing research has focused on reducing the inference cost of neural network controllers, yet in physical closed-loop systems

Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment

SafetyDGX agent

arXiv:2601.22823v2 Announce Type: replace-cross Abstract: We study offline reinforcement learning of style-conditioned policies using explicit style supervision via subtrajectory labeling functions. I

On the Faithfulness of Post-Hoc Concept Bottleneck Models

ApplicationsDGX agent

arXiv:2606.30498v1 Announce Type: cross Abstract: Human decision-making interprets the world through high-level concepts, such as recognizing a bird by its belly color. To bridge the gap between opaqu

On the Necessity of a Liquid Substrate for Mesh Intelligence

AgentsDGX agent

arXiv:2606.28413v1 Announce Type: cross Abstract: A mesh of sovereign agents has no center: no shared clock, no shared model, and no coordinator to gather data or retrain. Its competence rests on each

On the Nonlinearity of Learning Rate Scaling for LLM Training

ResearchDGX agent

arXiv:2606.29158v1 Announce Type: cross Abstract: Learning-rate transfer can reduce the cost of training large language models: instead of sweeping learning rates at target scale, practitioners extrap

One Scene, Two Depths: Probing Geometric Ambiguity in Monocular Foundation Models

Model ReleasesDGX agent

arXiv:2606.29600v1 Announce Type: cross Abstract: A faithful 3D world representation should account for layered geometry, where a single camera ray may contain multiple visible and geometrically valid

Online Data Selection for Instruction Tuning via Gaussian Processes

Local AiDGX agent

arXiv:2606.30077v1 Announce Type: cross Abstract: With Large Language Model (LLM) pre-training and fine-tuning shifting its focus from data volume to data quality, quality data selection has emerged a

Ontology-Guided Reverse Thinking Makes Large Language Models Stronger on Knowledge Graph Question Answering

TutorialsDGX agent

arXiv:2502.11491v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown remarkable capabilities in natural language processing. However, in knowledge graph question answering

Open Problems in Constitutional Preference Reconstruction

ResearchDGX agent

arXiv:2606.30116v1 Announce Type: new Abstract: Pairwise preference data is widely used for training and evaluating language models (e.g., RLHF), but each datapoint records a choice, not the rationale

Operating Regimes of Decentralized Learning Under Mobility and Bandwidth Constraints

ResearchDGX agent

arXiv:2606.28342v1 Announce Type: cross Abstract: Decentralized learning is a promising paradigm for collaborative training in mobile and pervasive systems, as it avoids a central coordinator and does

Optimization Dynamics Imprint Semantic Specificity in Contrastive Embedding Norms

ResearchDGX agent

arXiv:2606.30625v1 Announce Type: cross Abstract: Contrastive embedding models trained with scale-invariant losses are typically paired with distance metrics like cosine similarity, effectively ignori

Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonomous LLM Research Loop

Model ReleasesDGX agent

arXiv:2606.29717v1 Announce Type: cross Abstract: Predicting a material's properties from its structure is a central, fast-advancing problem in computational materials science. A decade of work has pr

OptiMUS-0.3: Using Large Language Models to Model and Solve Optimization Problems at Scale

Model ReleasesDGX agent

arXiv:2407.19633v4 Announce Type: replace Abstract: Optimization problems are pervasive in sectors from manufacturing and distribution to healthcare. However, most such problems are still solved heuri

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering

Model ReleasesDGX agent

arXiv:2512.09066v2 Announce Type: replace-cross Abstract: Reliable assessment of the abilities of large audio language models (LALMs) is essential to advancing the state of the art. As benchmarks rapi

OSWorld2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks

Model ReleasesDGX agent

arXiv:2606.29537v1 Announce Type: new Abstract: Existing computer-use benchmarks fail to capture the realism, complexity, and long-horizon demands of real-world computer use, limiting their ability to

Overcoming Dependent Censoring in the Evaluation of Survival Models

ResearchDGX agent

arXiv:2502.19460v4 Announce Type: replace-cross Abstract: Dependent censoring occurs when the event time and censoring time are not conditionally independent given the observed covariates. This compli

Perspectives on Latent Factor Indeterminacy and its Implications for Data Representation

ResearchDGX agent

arXiv:2606.28854v1 Announce Type: cross Abstract: The common factor analytic model is related to Helmholtz and Boltzmann machines, can be conceived as a linear autoencoder, or can be thought of as a s

Pessimism's Paradox: Conservative Offline Training Amplifies Reward Hacking During Online Adaptation in Reasoning Models

SafetyDGX agent

arXiv:2606.30627v1 Announce Type: cross Abstract: Conservative offline training is widely advocated as a safe foundation for subsequent online adaptation: if a policy stays close to well-supported beh

PHF: Privileged Hidden Flow for On-Policy Self-Distillation

SafetyDGX agent

arXiv:2606.29340v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a reasoning model on rollouts sampled from its own policy by matching a privileged teacher that also sees veri

Physically-Constrained Harmonic Separation for Robust Heart and Respiratory Rate Estimation from Wrist Photoplethysmography

ResearchDGX agent

arXiv:2606.30156v1 Announce Type: cross Abstract: Wrist-worn photoplethysmography (PPG) enables continuous monitoring of cardiopulmonary physiology, but reliable heart rate (HR) and respiratory rate (

Physics-Informed Distillation of Diffusion Models for PDE-Constrained Generation

ResearchDGX agent

arXiv:2505.22391v2 Announce Type: replace-cross Abstract: Modeling physical systems in a generative manner offers several advantages, including the ability to handle partial observations, generate div

PIXELRAG: Web Screenshots Beat Text for Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2606.28344v1 Announce Type: cross Abstract: Augmenting large language models (LLMs) with retrieved web text has become a dominant paradigm, yet the web is not natively textual: existing systems

PLAA: Packet-level Adversarial Attacks in Network Traffic Detection

ResearchDGX agent

arXiv:2606.28439v1 Announce Type: cross Abstract: Deep neural networks (DNNs) are widely applied in Network-based Intrusion Detection System (NIDS) due to their high accuracy. However, DNNs are highly

PlantExpertVQA: A Visual Question Answering Dataset for Benchmarking Vision-Language Models in Plant Science

Model ReleasesDGX agent

arXiv:2508.17117v3 Announce Type: replace-cross Abstract: Existing plant-disease datasets target classification and detection, leaving vision-language models unable to support interactive, reasoning-b

PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents

Model ReleasesDGX agent

arXiv:2606.29225v1 Announce Type: new Abstract: LLM agents handle user requests on behalf of organizations through tool calls and must follow the company policies stated in their system prompts. Prior

Pondering the Way: Spatial-perceiving World Action Model for Embodied Navigation

SafetyDGX agent

arXiv:2606.29908v1 Announce Type: cross Abstract: Existing world model-based planners for visual navigation typically follow a verification-centric paradigm, decoupling goal intent from trajectory syn

Pooled Leaderboards Hide System-Specific Winners: A Reporting-Protocol Audit of Offline Root-Cause Analysis Benchmarks

Model ReleasesDGX agent

arXiv:2606.29159v1 Announce Type: new Abstract: Offline root-cause-analysis (RCA) benchmarks commonly rank methods by a single pooled top-1 accuracy across multiple subsystems, and engineers often rea

Pose-Based Fall Detection System: Efficient Monitoring on Standard CPUs

SafetyDGX agent

arXiv:2503.19501v2 Announce Type: replace-cross Abstract: Falls among elderly residents in assisted living homes pose significant health risks, often leading to injuries and a decreased quality of lif

Post-training for Efficient Communication via Convention Formation

Model ReleasesDGX agent

arXiv:2508.06482v2 Announce Type: replace-cross Abstract: Humans communicate with increasing efficiency in multi-turn interactions, by adapting their language and forming ad-hoc conventions. In contra

Predicting Effects, Missing Distributions: Evaluating LLMs as Human Behavior Simulators in Operations Management

ResearchDGX agent

arXiv:2510.03310v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate human behavior in business, economics, and the social sciences, offering a low-

Predicting Metastatic Risk from Primary Tissue Architecture via Distance-Aware Spatial Modeling

ResearchDGX agent

arXiv:2606.28676v1 Announce Type: cross Abstract: Predicting the risk of distant metastasis from primary tumor tissue histology is a critical yet challenging task in computational pathology. Multiple

Preventing Error Propagation in Multi-Agent AI through Runtime Monitoring

AgentsDGX agent

arXiv:2606.29026v1 Announce Type: new Abstract: Multi-agent AI systems can improve answer selection by allowing different language models to exchange reasoning traces, revise initial predictions, and

Priced Motion Through Optimal Faces: A Normal-Fan Geometry for Non-Stationary Adversarial MDPs

SafetyDGX agent

arXiv:2606.29092v1 Announce Type: cross Abstract: In a changing decision problem, standard dynamic-regret analyses have often equated the cost of non-stationarity to how far loss moves. However, it is

Primary ICD Category Prediction using LLM-based Probing

Model ReleasesDGX agent

arXiv:2606.28798v1 Announce Type: new Abstract: Objective: ICD codes are central to reimbursement, research, and population health surveillance, yet automated coding systems often struggle to integrat

Process Advantage Signal Shaping: A Paradigm-Agnostic Middleware for Process-Supervised RL in LLM Reasoners

SafetyDGX agent

arXiv:2606.29296v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is a default recipe for process-supervised reinforcement learning of LLM reasoners, and dense process supervis

Projected Exploitability Descent for Nash Equilibrium Computation in Multiplayer Imperfect-Information Games

Model ReleasesDGX agent

arXiv:2606.29169v1 Announce Type: cross Abstract: Many important games have more than two players and imperfect information. Existing approaches for computing Nash equilibrium, the central game-theore

PromptGNN-sim: Deep Fusion and Alignment of GNN and LLMs for Text-Attributed Graph Learning

SafetyDGX agent

arXiv:2606.30291v1 Announce Type: new Abstract: Text-Attributed Graphs (TAGs) combine textual semantics with graph structure and are central to many graph learning tasks. However, existing fusion meth

Proof-of-Guardrail in AI Agents and What (Not) to Trust from It

SafetyDGX agent

arXiv:2603.05786v2 Announce Type: replace-cross Abstract: As AI agents become widely deployed as online services, users often rely on an agent developer's claim about how safety is enforced, which int

← Previous
1…113114115116117…358
Next →