AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Local Ai

On-Device Continual Learning with Dual-Stage Buffer and Dynamic Loss for Point-of-Care Pneumonia Diagnosis

DGX agent

arXiv:2605.19201v1 Announce Type: cross Abstract: Deep learning models detect pneumonia from chest X-rays with high accuracy, but the performance declines under domain shifts caused by differences in

local-aiarxiv-cs-ai
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Open-Set Domain Adaptation Under Background Distribution Shift: Challenges and A Provably Efficient Solution

DGX agent

arXiv:2512.01152v4 Announce Type: replace-cross Abstract: As we deploy machine learning systems in the real world, a core challenge is to maintain a model that is performant even as the data shifts. S

applicationsarxiv-cs-ai
20 May 2026
Research

OpenComputer: Verifiable Software Worlds for Computer-Use Agents

DGX agent

arXiv:2605.19769v1 Announce Type: new Abstract: We present OpenComputer, a verifier-grounded framework for constructing verifiable software worlds for computer-use agents. OpenComputer integrates four

researcharxiv-cs-ai
20 May 2026
Agents

Operationalising Artificial Intelligence Bills of Materials (AIBOMs) for Verifiable AI Provenance and Lifecycle Assurance

DGX agent

arXiv:2605.19755v1 Announce Type: cross Abstract: Artificial Intelligence (AI) systems are increasingly dependent on complex, multi-layered software supply chains that introduce challenges for reprodu

agentsarxiv-cs-ai
20 May 2026
Model Releases

Operationalizing Document AI: A Microservice Architecture for OCR and LLM Pipelines in Production

DGX agent

arXiv:2605.18818v1 Announce Type: new Abstract: Academic research tends to focus on new models for document understanding creating a wide gap in the literature between model definition and running mod

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

optimize_anything: A Universal API for Optimizing any Text Parameter

DGX agent

arXiv:2605.19633v1 Announce Type: cross Abstract: Can a single LLM-based optimization system match specialized tools across fundamentally different domains? We show that when optimization problems are

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

P2DNav: Panorama-to-Downview Reasoning for Zero-shot Vision-and-Language Navigation

DGX agent

arXiv:2605.19634v1 Announce Type: cross Abstract: Vision-and-language navigation (VLN) requires an embodied agent to ground natural-language instructions into executable navigation actions in unseen e

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Passive Construction Site Safety Monitoring via Persona-Scaffolded Adversarial Chain-of-Thought VLM Verification

DGX agent

arXiv:2605.19869v1 Announce Type: cross Abstract: Construction remains the deadliest industry sector in the United States, with 1,055 fatal worker injuries recorded in 2023, and the majority preventab

model-releasesarxiv-cs-ai
20 May 2026
Agents

PAVE: A Cognitive Architecture for Legitimate Violation in Generative Agent Societies

DGX agent

arXiv:2605.19351v1 Announce Type: cross Abstract: Generative agents based on large language models reproduce believable human behavior in cooperative settings, but how they should reason in situations

agentsarxiv-cs-ai
20 May 2026
Safety

PEEK: Context Map as an Orientation Cache for Long-Context LLM Agents

DGX agent

arXiv:2605.19932v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate over long and recurring external contexts, like document corpora and code repositories. Across in

safetyarxiv-cs-ai
20 May 2026
Safety

Phase-Aware Mixture of Experts for Agentic Reinforcement Learning

DGX agent

arXiv:2602.17038v3 Announce Type: replace Abstract: Reinforcement learning (RL) has equipped LLM agents with a strong ability to solve complex tasks. However, existing RL methods normally use a single

safetyarxiv-cs-ai
20 May 2026
Model Releases

PhyWorld: Physics-Faithful World Model for Video Generation

DGX agent

arXiv:2605.19242v1 Announce Type: cross Abstract: World simulators can provide safe and scalable environments for training Physical AI systems before real-world deployment. Large video generation mode

model-releasesarxiv-cs-ai
20 May 2026
Hardware

PiKV: KV Cache Management System for Mixture of Experts

DGX agent

arXiv:2508.06526v3 Announce Type: replace-cross Abstract: As large-scale language models continue to scale up in both size and context length, the memory and communication cost of key-value (KV) cache

hardwarearxiv-cs-ai
20 May 2026
Research

Planner-Admissible Graph-PDE Value Extensions for Sparse Goal-Conditioned Planning

DGX agent

arXiv:2605.19185v1 Announce Type: cross Abstract: Sparse goal-conditioned planning with few cost-to-go labels can be viewed as a graph-PDE Dirichlet extension problem: extend sparse labels on a goal-d

researcharxiv-cs-ai
20 May 2026
Model Releases

PlantTraitNet: An Uncertainty-Aware Multimodal Framework for Global-Scale Plant Trait Inference from Citizen Science Data

DGX agent

arXiv:2511.06943v3 Announce Type: replace-cross Abstract: Global plant maps of plant traits, such as leaf nitrogen or plant height, are essential for understanding ecosystem processes, including the c

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

POLAR-Bench: A Diagnostic Benchmark for Privacy-Utility Trade-offs in LLM Agents

DGX agent

arXiv:2605.19127v1 Announce Type: new Abstract: LLM agents increasingly have access to private user data and act on the user's behalf when interacting with third-party systems. The user defines what m

model-releasesarxiv-cs-ai
20 May 2026
Safety

Position: Let's Develop Data Probes to Fundamentally Understand How Data Affects LLM Performance

DGX agent

arXiv:2605.18801v1 Announce Type: new Abstract: Data is fundamental to large language models (LLMs). However, understanding of what makes certain data useful for different stages of an LLM workflow, i

safetyarxiv-cs-ai
20 May 2026
Model Releases

Position: The Turing-Completeness of Real-World Autoregressive Transformers Relies Heavily on Context Management

DGX agent

arXiv:2605.19514v1 Announce Type: new Abstract: Many works make the eye-catching claim that Transformers are Turing-complete. However, the literature often conflates two distinct settings: (i) a fixed

model-releasesarxiv-cs-ai
20 May 2026
Safety

Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering

DGX agent

arXiv:2605.19220v1 Announce Type: cross Abstract: Uncertainty Quantification (UQ) is widely regarded as the primary safeguard for deploying Large Language Models (LLMs) in high-stakes domains. However

safetyarxiv-cs-ai
20 May 2026
Agents

PragLocker: Protecting Agent Intellectual Property in Untrusted Deployments via Non-Portable Prompts

DGX agent

arXiv:2605.05974v2 Announce Type: replace-cross Abstract: LLM agents rely on prompts to implement task-specific capabilities based on foundation LLMs, making agent prompts valuable intellectual proper

agentsarxiv-cs-ai
20 May 2026
Model Releases

Precision Tracked Transformer via Kalman Filtering, Kriging and Process Noise

DGX agent

arXiv:2605.18832v1 Announce Type: cross Abstract: The Transformer is the foundational building block of modern AI, yet offers no principled handling of uncertainty, which is prevalent in real applicat

model-releasesarxiv-cs-ai
20 May 2026
Safety

Prediction Is Not Physics: Learning and Evaluating Conserved Quantities in Neural Simulators

DGX agent

arXiv:2605.18883v1 Announce Type: cross Abstract: A diffusion model trained on Hamiltonian trajectories can achieve rollout MSE near 10^{-3}, but the standard deviation of its energy over time is betw

safetyarxiv-cs-ai
20 May 2026
Hardware

Prior Knowledge or Search? A Study of LLM Agents in Hardware-Aware Code Optimization

DGX agent

arXiv:2605.19782v1 Announce Type: new Abstract: LLM discovery and optimization systems are increasingly applied across domains, implementing a common propose-evaluate-revise loop. Such optimization or

hardwarearxiv-cs-ai
20 May 2026
Model Releases

PRISM: A Benchmark for Programmatic Spatial-Temporal Reasoning

DGX agent

arXiv:2605.19382v1 Announce Type: new Abstract: Programmatic video generation through code offers geometric precision and temporal coherence beyond pixel-level diffusion models, yet rigorously evaluat

model-releasesarxiv-cs-ai
20 May 2026
Research

Probabilistic Tiny Recursive Model

DGX agent

arXiv:2605.19943v1 Announce Type: new Abstract: Tiny Recursive Models (TRM) solve complex reasoning tasks with a fraction of the parameters of modern large language models (LLMs) by iteratively refini

researcharxiv-cs-ai
20 May 2026
Research

Probability-Conserving Flow Guidance

DGX agent

arXiv:2605.20079v1 Announce Type: cross Abstract: Diffusion and flow-based generative models dominate visual synthesis, with guidance aligning samples to user input and improving perceptual quality. H

researcharxiv-cs-ai
20 May 2026
Agents

Probing Embodied LLMs: When Higher Observation Fidelity Hurts Problem Solving

DGX agent

arXiv:2605.20072v1 Announce Type: new Abstract: Large Language Models are increasingly proposed as cognitive components for robotic systems, yet their opaque decision processes make it difficult to ex

agentsarxiv-cs-ai
20 May 2026
Safety

Progressive Autonomy as Preference Learning: A Formalization of Trust Calibration for Agentic Tool Use

DGX agent

arXiv:2605.19151v1 Announce Type: new Abstract: We formalize trust calibration for agentic tool use (deciding when an automated agent's proposed action may execute autonomously versus require human ap

safetyarxiv-cs-ai
20 May 2026
Research

Projecting Latent RL Actions: Towards Generalizable and Scalable Graph Combinatorial Optimization

DGX agent

arXiv:2605.19721v1 Announce Type: new Abstract: Graph combinatorial optimization (GCO) has attracted growing interest, as many NP-hard problems naturally admit graph formulations, yet their combinator

researcharxiv-cs-ai
20 May 2026
Model Releases

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling

DGX agent

arXiv:2605.20052v1 Announce Type: cross Abstract: Automatic report labeling facilitates the identification of clinical findings from unstructured text and enables large-scale annotation for medical im

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Protein Autoregressive Modeling via Multiscale Structure Generation

DGX agent

arXiv:2602.04883v2 Announce Type: replace-cross Abstract: We present protein autoregressive modeling (PAR), the first multi-scale autoregressive framework for protein backbone generation via coarse-to

model-releasesarxiv-cs-ai
20 May 2026
Safety

PROWL: Prioritized Regret-Driven Optimization for World Model Learning

DGX agent

arXiv:2605.18803v1 Announce Type: cross Abstract: Modern action-conditioned video world models achieve strong short-horizon visual realism, yet remain unreliable on rare, interaction-critical transiti

safetyarxiv-cs-ai
20 May 2026
Research

Proximal Diffusion Neural Sampler

DGX agent

arXiv:2510.03824v2 Announce Type: replace-cross Abstract: The task of learning a diffusion-based neural sampler for drawing samples from an unnormalized target distribution can be viewed as a stochast

researcharxiv-cs-ai
20 May 2026
Safety

Pseudocode-Guided Structured Reasoning for Automating Reliable Inference in Vision-Language Models

DGX agent

arXiv:2605.19663v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are becoming the cornerstone of high-level reasoning for robotic automation, enabling robots to parse natural language com

safetyarxiv-cs-ai
20 May 2026
Model Releases

Quantifying the Generalization Gap in Seizure Detection: A Large-Scale Empirical Benchmark via the SzCORE Challenge

DGX agent

arXiv:2505.18191v2 Announce Type: replace-cross Abstract: Reliable automatic seizure detection from long-term electroencephalography (EEG) remains an unsolved challenge, as current models often fail t

model-releasesarxiv-cs-ai
20 May 2026
Safety

Quantifying the Pre-training Dividend: Generative versus Latent Self-Supervised Learning for Time Series Foundation Models

DGX agent

arXiv:2605.19462v1 Announce Type: cross Abstract: The success of self-supervised learning (SSL) in vision and NLP has motivated its rapid adoption for time series. However, research has focused primar

safetyarxiv-cs-ai
20 May 2026
Applications

Quantized Machine Learning Models for Medical Imaging in Low-Resource Healthcare Settings

DGX agent

arXiv:2605.19207v1 Announce Type: cross Abstract: Deep learning models have shown strong performance in medical image analysis, but deploying them in low-resource clinical environments remains difficu

applicationsarxiv-cs-ai
20 May 2026
Research

Query-Aware Flow Diffusion for Graph-Based RAG with Retrieval Guarantees

DGX agent

arXiv:2605.18775v1 Announce Type: cross Abstract: Graph-based Retrieval-Augmented Generation (RAG) systems leverage interconnected knowledge structures to capture complex relationships that flat retri

researcharxiv-cs-ai
20 May 2026
Local Ai

Query-Conditioned Graph Retrieval for Contextualized LLM Reasoning in Personalized Wearable Data

DGX agent

arXiv:2605.18763v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly applied to analyzing wearable sensing data, which are long-term, multimodal, and highly personalized. A

local-aiarxiv-cs-ai
20 May 2026
Model Releases

RE-VLM: Event-Augmented Vision-Language Model for Scene Understanding

DGX agent

arXiv:2605.19329v1 Announce Type: cross Abstract: Conventional vision-language models (VLMs) struggle to interpret scenes captured under adverse conditions (e.g., low light, high dynamic range, or fas

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

ReacTOD: Bounded Neuro-Symbolic Agentic NLU for Zero-Shot Dialogue State Tracking

DGX agent

arXiv:2605.19077v1 Announce Type: cross Abstract: Task-oriented dialogue systems -- handling transactions, reservations, and service requests -- require predictable behavior, yet the moderately-sized

model-releasesarxiv-cs-ai
20 May 2026
Local Ai

Real-Time Parallel Counterfactual Regret Minimization

DGX agent

arXiv:2605.19928v1 Announce Type: cross Abstract: Counterfactual Regret Minimization (CFR) is the dominant algorithmic family for solving large imperfect-information games, underpinning breakthroughs

local-aiarxiv-cs-ai
20 May 2026
Model Releases

Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models

DGX agent

arXiv:2605.19398v1 Announce Type: cross Abstract: Image-to-video models often generate videos that remain overly static, compared to text-to-video models. While prior approaches mitigate this issue by

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

RecoAtlas: From Semantic Plausibility to Set-Level Utility in LLM Recommendation Agents

DGX agent

arXiv:2605.18805v1 Announce Type: cross Abstract: LLM recommendation agents increasingly produce structured recommendation reports: sets of items accompanied by natural-language justifications. Yet ex

model-releasesarxiv-cs-ai
20 May 2026
Research

ReCrit: Transition-Aware Reinforcement Learning for Scientific Critic Reasoning

DGX agent

arXiv:2605.18799v1 Announce Type: cross Abstract: Large language models can fail in critic interaction not only by answering incorrectly, but also by abandoning an initially correct scientific solutio

researcharxiv-cs-ai
20 May 2026
Model Releases

Recursive Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model

DGX agent

arXiv:2506.00286v3 Announce Type: replace-cross Abstract: We study risk-sensitive reinforcement learning in finite discounted MDPs with recursive entropic risk measures (ERM), where the risk parameter

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Resilient Byzantine Agreement with Predictions

DGX agent

arXiv:2605.19452v1 Announce Type: cross Abstract: This paper studies the Byzantine Agreement problem where the nodes have access to a predictor that flags nodes for suspicion of faulty (Byzantine) beh

model-releasesarxiv-cs-ai
20 May 2026
Safety

Rethinking the Design Space of Reinforcement Learning for Diffusion Models: On the Importance of Likelihood Estimation Beyond Loss Design

DGX agent

arXiv:2602.04663v2 Announce Type: replace-cross Abstract: Reinforcement learning has been widely applied to diffusion and flow models for visual tasks such as text-to-image generation. However, these

safetyarxiv-cs-ai
20 May 2026
← Previous
1…287288289290291…448
Next →