AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Applications

On the Faithfulness of Post-Hoc Concept Bottleneck Models

DGX agent

arXiv:2606.30498v1 Announce Type: cross Abstract: Human decision-making interprets the world through high-level concepts, such as recognizing a bird by its belly color. To bridge the gap between opaqu

applicationsarxiv-cs-ai
30 Jun 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

On the Necessity of a Liquid Substrate for Mesh Intelligence

DGX agent

arXiv:2606.28413v1 Announce Type: cross Abstract: A mesh of sovereign agents has no center: no shared clock, no shared model, and no coordinator to gather data or retrain. Its competence rests on each

agentsarxiv-cs-ai
30 Jun 2026
Research

On the Nonlinearity of Learning Rate Scaling for LLM Training

DGX agent

arXiv:2606.29158v1 Announce Type: cross Abstract: Learning-rate transfer can reduce the cost of training large language models: instead of sweeping learning rates at target scale, practitioners extrap

researcharxiv-cs-ai
30 Jun 2026
Model Releases

One Scene, Two Depths: Probing Geometric Ambiguity in Monocular Foundation Models

DGX agent

arXiv:2606.29600v1 Announce Type: cross Abstract: A faithful 3D world representation should account for layered geometry, where a single camera ray may contain multiple visible and geometrically valid

model-releasesarxiv-cs-ai
30 Jun 2026
Local Ai

Online Data Selection for Instruction Tuning via Gaussian Processes

DGX agent

arXiv:2606.30077v1 Announce Type: cross Abstract: With Large Language Model (LLM) pre-training and fine-tuning shifting its focus from data volume to data quality, quality data selection has emerged a

local-aiarxiv-cs-ai
30 Jun 2026
Tutorials

Ontology-Guided Reverse Thinking Makes Large Language Models Stronger on Knowledge Graph Question Answering

DGX agent

arXiv:2502.11491v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown remarkable capabilities in natural language processing. However, in knowledge graph question answering

tutorialsarxiv-cs-ai
30 Jun 2026
Research

Open Problems in Constitutional Preference Reconstruction

DGX agent

arXiv:2606.30116v1 Announce Type: new Abstract: Pairwise preference data is widely used for training and evaluating language models (e.g., RLHF), but each datapoint records a choice, not the rationale

researcharxiv-cs-ai
30 Jun 2026
Research

Operating Regimes of Decentralized Learning Under Mobility and Bandwidth Constraints

DGX agent

arXiv:2606.28342v1 Announce Type: cross Abstract: Decentralized learning is a promising paradigm for collaborative training in mobile and pervasive systems, as it avoids a central coordinator and does

researcharxiv-cs-ai
30 Jun 2026
Research

Optimization Dynamics Imprint Semantic Specificity in Contrastive Embedding Norms

DGX agent

arXiv:2606.30625v1 Announce Type: cross Abstract: Contrastive embedding models trained with scale-invariant losses are typically paired with distance metrics like cosine similarity, effectively ignori

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonomous LLM Research Loop

DGX agent

arXiv:2606.29717v1 Announce Type: cross Abstract: Predicting a material's properties from its structure is a central, fast-advancing problem in computational materials science. A decade of work has pr

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

OptiMUS-0.3: Using Large Language Models to Model and Solve Optimization Problems at Scale

DGX agent

arXiv:2407.19633v4 Announce Type: replace Abstract: Optimization problems are pervasive in sectors from manufacturing and distribution to healthcare. However, most such problems are still solved heuri

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering

DGX agent

arXiv:2512.09066v2 Announce Type: replace-cross Abstract: Reliable assessment of the abilities of large audio language models (LALMs) is essential to advancing the state of the art. As benchmarks rapi

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

OSWorld2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks

DGX agent

arXiv:2606.29537v1 Announce Type: new Abstract: Existing computer-use benchmarks fail to capture the realism, complexity, and long-horizon demands of real-world computer use, limiting their ability to

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Overcoming Dependent Censoring in the Evaluation of Survival Models

DGX agent

arXiv:2502.19460v4 Announce Type: replace-cross Abstract: Dependent censoring occurs when the event time and censoring time are not conditionally independent given the observed covariates. This compli

researcharxiv-cs-ai
30 Jun 2026
Research

Perspectives on Latent Factor Indeterminacy and its Implications for Data Representation

DGX agent

arXiv:2606.28854v1 Announce Type: cross Abstract: The common factor analytic model is related to Helmholtz and Boltzmann machines, can be conceived as a linear autoencoder, or can be thought of as a s

researcharxiv-cs-ai
30 Jun 2026
Safety

Pessimism's Paradox: Conservative Offline Training Amplifies Reward Hacking During Online Adaptation in Reasoning Models

DGX agent

arXiv:2606.30627v1 Announce Type: cross Abstract: Conservative offline training is widely advocated as a safe foundation for subsequent online adaptation: if a policy stays close to well-supported beh

safetyarxiv-cs-ai
30 Jun 2026
Safety

PHF: Privileged Hidden Flow for On-Policy Self-Distillation

DGX agent

arXiv:2606.29340v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a reasoning model on rollouts sampled from its own policy by matching a privileged teacher that also sees veri

safetyarxiv-cs-ai
30 Jun 2026
Research

Physically-Constrained Harmonic Separation for Robust Heart and Respiratory Rate Estimation from Wrist Photoplethysmography

DGX agent

arXiv:2606.30156v1 Announce Type: cross Abstract: Wrist-worn photoplethysmography (PPG) enables continuous monitoring of cardiopulmonary physiology, but reliable heart rate (HR) and respiratory rate (

researcharxiv-cs-ai
30 Jun 2026
Research

Physics-Informed Distillation of Diffusion Models for PDE-Constrained Generation

DGX agent

arXiv:2505.22391v2 Announce Type: replace-cross Abstract: Modeling physical systems in a generative manner offers several advantages, including the ability to handle partial observations, generate div

researcharxiv-cs-ai
30 Jun 2026
Agents

PIXELRAG: Web Screenshots Beat Text for Retrieval-Augmented Generation

DGX agent

arXiv:2606.28344v1 Announce Type: cross Abstract: Augmenting large language models (LLMs) with retrieved web text has become a dominant paradigm, yet the web is not natively textual: existing systems

agentsarxiv-cs-ai
30 Jun 2026
Research

PLAA: Packet-level Adversarial Attacks in Network Traffic Detection

DGX agent

arXiv:2606.28439v1 Announce Type: cross Abstract: Deep neural networks (DNNs) are widely applied in Network-based Intrusion Detection System (NIDS) due to their high accuracy. However, DNNs are highly

researcharxiv-cs-ai
30 Jun 2026
Model Releases

PlantExpertVQA: A Visual Question Answering Dataset for Benchmarking Vision-Language Models in Plant Science

DGX agent

arXiv:2508.17117v3 Announce Type: replace-cross Abstract: Existing plant-disease datasets target classification and detection, leaving vision-language models unable to support interactive, reasoning-b

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents

DGX agent

arXiv:2606.29225v1 Announce Type: new Abstract: LLM agents handle user requests on behalf of organizations through tool calls and must follow the company policies stated in their system prompts. Prior

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Pondering the Way: Spatial-perceiving World Action Model for Embodied Navigation

DGX agent

arXiv:2606.29908v1 Announce Type: cross Abstract: Existing world model-based planners for visual navigation typically follow a verification-centric paradigm, decoupling goal intent from trajectory syn

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Pooled Leaderboards Hide System-Specific Winners: A Reporting-Protocol Audit of Offline Root-Cause Analysis Benchmarks

DGX agent

arXiv:2606.29159v1 Announce Type: new Abstract: Offline root-cause-analysis (RCA) benchmarks commonly rank methods by a single pooled top-1 accuracy across multiple subsystems, and engineers often rea

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Pose-Based Fall Detection System: Efficient Monitoring on Standard CPUs

DGX agent

arXiv:2503.19501v2 Announce Type: replace-cross Abstract: Falls among elderly residents in assisted living homes pose significant health risks, often leading to injuries and a decreased quality of lif

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Post-training for Efficient Communication via Convention Formation

DGX agent

arXiv:2508.06482v2 Announce Type: replace-cross Abstract: Humans communicate with increasing efficiency in multi-turn interactions, by adapting their language and forming ad-hoc conventions. In contra

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Predicting Effects, Missing Distributions: Evaluating LLMs as Human Behavior Simulators in Operations Management

DGX agent

arXiv:2510.03310v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate human behavior in business, economics, and the social sciences, offering a low-

researcharxiv-cs-ai
30 Jun 2026
Research

Predicting Metastatic Risk from Primary Tissue Architecture via Distance-Aware Spatial Modeling

DGX agent

arXiv:2606.28676v1 Announce Type: cross Abstract: Predicting the risk of distant metastasis from primary tumor tissue histology is a critical yet challenging task in computational pathology. Multiple

researcharxiv-cs-ai
30 Jun 2026
Agents

Preventing Error Propagation in Multi-Agent AI through Runtime Monitoring

DGX agent

arXiv:2606.29026v1 Announce Type: new Abstract: Multi-agent AI systems can improve answer selection by allowing different language models to exchange reasoning traces, revise initial predictions, and

agentsarxiv-cs-ai
30 Jun 2026
Safety

Priced Motion Through Optimal Faces: A Normal-Fan Geometry for Non-Stationary Adversarial MDPs

DGX agent

arXiv:2606.29092v1 Announce Type: cross Abstract: In a changing decision problem, standard dynamic-regret analyses have often equated the cost of non-stationarity to how far loss moves. However, it is

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Primary ICD Category Prediction using LLM-based Probing

DGX agent

arXiv:2606.28798v1 Announce Type: new Abstract: Objective: ICD codes are central to reimbursement, research, and population health surveillance, yet automated coding systems often struggle to integrat

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Process Advantage Signal Shaping: A Paradigm-Agnostic Middleware for Process-Supervised RL in LLM Reasoners

DGX agent

arXiv:2606.29296v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is a default recipe for process-supervised reinforcement learning of LLM reasoners, and dense process supervis

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Projected Exploitability Descent for Nash Equilibrium Computation in Multiplayer Imperfect-Information Games

DGX agent

arXiv:2606.29169v1 Announce Type: cross Abstract: Many important games have more than two players and imperfect information. Existing approaches for computing Nash equilibrium, the central game-theore

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

PromptGNN-sim: Deep Fusion and Alignment of GNN and LLMs for Text-Attributed Graph Learning

DGX agent

arXiv:2606.30291v1 Announce Type: new Abstract: Text-Attributed Graphs (TAGs) combine textual semantics with graph structure and are central to many graph learning tasks. However, existing fusion meth

safetyarxiv-cs-ai
30 Jun 2026
Safety

Proof-of-Guardrail in AI Agents and What (Not) to Trust from It

DGX agent

arXiv:2603.05786v2 Announce Type: replace-cross Abstract: As AI agents become widely deployed as online services, users often rely on an agent developer's claim about how safety is enforced, which int

safetyarxiv-cs-ai
30 Jun 2026
Safety

Propagation of~Interval Belief Structures and~Imprecise Copulas for~Neural Network Verification

DGX agent

arXiv:2606.30105v1 Announce Type: new Abstract: Quantitative verification of neural networks requires reasoning about probabilities under substantial uncertainty in both input distributions and their

safetyarxiv-cs-ai
30 Jun 2026
Safety

ProSpec RL: Plan Ahead, then Execute

DGX agent

arXiv:2407.21359v2 Announce Type: replace-cross Abstract: Imagining potential outcomes of actions before execution helps agents make more informed decisions, a prospective thinking ability fundamental

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Proteus: Automated Adversarial Robustness Testing for Audio Deepfake Detectors

DGX agent

arXiv:2606.29544v1 Announce Type: cross Abstract: We present Proteus, a framework developed at Resemble AI for automated robustness testing of our audio deepfake detection system. Given a detector, Pr

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

PS-PPO: Prefix-Sampling PPO for Critic-Free RLHF

DGX agent

arXiv:2606.29758v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) for Large Language Models increasingly relies on critic-free methods as a practical alternative to a

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization

DGX agent

arXiv:2602.11351v2 Announce Type: replace Abstract: Proactive large language model (LLM) agents aim to actively plan, query, and interact over multiple turns, enabling efficient task completion beyond

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Query-Aware Spreading Activation for Multi-Hop Retrieval over Knowledge Graphs

DGX agent

arXiv:2606.30133v1 Announce Type: cross Abstract: Retrieval-augmented generation built on knowledge graphs (Graph RAG) outperforms flat passage retrieval on multi-hop question answering by leveraging

researcharxiv-cs-ai
30 Jun 2026
Local Ai

RADIANT-PET: Reasoning-Augmented PET/CT Lesion Segmentation with Large Language Models and Reinforcement Learning

DGX agent

arXiv:2606.28392v1 Announce Type: cross Abstract: Accurate lesion segmentation in PET/CT is critical for oncology, yet remains challenging because physiologic tracer uptake and artifacts can mimic mal

local-aiarxiv-cs-ai
30 Jun 2026
Safety

Rank-Aware Hyperbolic Alignment for Vision-Language Dataset Distillation

DGX agent

arXiv:2606.29464v1 Announce Type: cross Abstract: Vision-language dataset distillation (VLDD) compresses a large image-text paired dataset into a small set of synthetic pairs that can efficiently trai

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

RankGraph-2: Lifecycle Co-Design for Billion-Node Graph Learning in Recommendation

DGX agent

arXiv:2606.18379v2 Announce Type: replace-cross Abstract: Graph-based retrieval at billion-node scale requires jointly solving three tightly coupled problems -- graph construction, representation lear

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

ReactiveBFM: Reactive Closed-Loop Motion Planning Towards Universal Humanoid Whole-Body Control

DGX agent

arXiv:2606.30362v1 Announce Type: cross Abstract: While current Behavior Foundation Models (BFMs) provide robust control priors for humanoids, they only execute pre-defined reference motions. As a res

safetyarxiv-cs-ai
30 Jun 2026
Agents

ReasonRec: A Reasoning-Augmented Multimodal Agent for Unified Recommendation

DGX agent

arXiv:2606.28357v1 Announce Type: cross Abstract: Recent advances in multimodal recommenders excel at feature fusion but remain opaque and inefficient decision-makers, lacking explicit reasoning and s

agentsarxiv-cs-ai
30 Jun 2026
Research

Reconsidering Overthinking: Penalizing Internal and External Redundancy in CoT Reasoning

DGX agent

arXiv:2508.02178v3 Announce Type: replace Abstract: Large reasoning models (LRMs) often exhibit overthinking, producing verbose Chain-of-Thought (CoT) traces that increase inference cost and obscure t

researcharxiv-cs-ai
30 Jun 2026
← Previous
1…142143144145146…448
Next →