AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Safety

Multi-Environment POMDPs with Finite-Horizon Objectives

DGX agent

arXiv:2605.07537v1 Announce Type: new Abstract: Partially Observable Markov Decision Processes (POMDPs) are systems in which one agent interacts with a stochastic environment, and receives only partia

safetyarxiv-cs-ai
11 May 2026
Safety

Multi-Modal Multi-Agent Reinforcement Learning for Radiology Report Generation

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2603.16876v2 Announce Type: replace-cross Abstract: We propose MARL-Rad, a multi-modal multi-agent reinforcement learning framework for radiology report generation that trains the entire agentic

safetyarxiv-cs-ai
11 May 2026
Model Releases

Multi-Objective Constraint Inference using Inverse reinforcement learning

DGX agent

arXiv:2605.06951v1 Announce Type: new Abstract: Constraint inference is widely considered essential to align reinforcement learning agents with safety boundaries and operational guidelines by observin

model-releasesarxiv-cs-ai
11 May 2026
Applications

Multimodal synthesis of MRI and tabular data with diffusion in a joint latent space via cross-attention

DGX agent

arXiv:2605.06699v1 Announce Type: cross Abstract: We propose a multimodal latent diffusion model that jointly synthesizes volumetric magnetic resonance imaging (MRI) and tabular clinical data within a

applicationsarxiv-cs-ai
11 May 2026
Model Releases

Muon Dynamics as a Spectral Wasserstein Flow

DGX agent

arXiv:2604.04891v2 Announce Type: replace-cross Abstract: Gradient normalization stabilizes deep-learning optimization, and spectral normalizations are especially natural for matrix-shaped parameter b

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Narrow Secret Loyalty Dodges Black-Box Audits

DGX agent

arXiv:2605.06846v1 Announce Type: cross Abstract: Recent work identifies secret loyalties as a distinct threat from standard backdoors. A secret loyalty causes a model to covertly advance the interest

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

NavOne: One-Step Global Planning for Vision-Language Navigation on Top-Down Maps

DGX agent

arXiv:2605.06317v2 Announce Type: replace-cross Abstract: Existing Vision-Language Navigation (VLN) methods typically adopt an egocentric, step-by-step paradigm, which struggles with error accumulatio

tutorialsarxiv-cs-ai
11 May 2026
Model Releases

Neural Operators as Efficient Function Interpolators

DGX agent

arXiv:2605.07792v1 Announce Type: cross Abstract: Neural operators (NOs) are designed to learn maps between infinite-dimensional function spaces. We propose a novel reframing of their use. By introduc

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

Neurosymbolic Framework for Concept-Driven Logical Reasoning in Skeleton-Based Human Action Recognition

DGX agent

arXiv:2605.07140v1 Announce Type: cross Abstract: Skeleton-based human activity recognition has achieved strong empirical performance, yet most existing models remain black boxes and difficult to inte

tutorialsarxiv-cs-ai
11 May 2026
Research

Nurnberg NLP at PsyDefDetect: Multi-Axis Voter Ensembles for Psychological Defence Mechanism Classification

DGX agent

arXiv:2605.07606v1 Announce Type: cross Abstract: Detecting levels of psychological defence mechanisms in supportive conversations is inherently ambiguous. In the PsyDefDetect shared task at BioNLP 20

researcharxiv-cs-ai
11 May 2026
Safety

OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search

DGX agent

arXiv:2604.03675v2 Announce Type: replace Abstract: Agentic search enables language models to solve knowledge-intensive tasks by adaptively acquiring external evidence over multiple steps. Reinforceme

safetyarxiv-cs-ai
11 May 2026
Safety

Offline Policy Optimization with Posterior Sampling

DGX agent

arXiv:2605.07393v1 Announce Type: new Abstract: A fundamental challenge in model-based offline reinforcement learning (RL) lies in the trade-off between generalization and robustness against exploitat

safetyarxiv-cs-ai
11 May 2026
Model Releases

OmicsLM: A Multimodal Large Language Model for Multi-Sample Omics Reasoning

DGX agent

arXiv:2605.06728v1 Announce Type: cross Abstract: Interpreting transcriptomic data is one of the most common analytical tasks in modern biology. Yet most current models either consume expression profi

model-releasesarxiv-cs-ai
11 May 2026
Research

On Privacy Leakage in Tabular Diffusion Models: Influential Factors, Attacker Knowledge, and Metrics

DGX agent

arXiv:2605.06835v1 Announce Type: cross Abstract: Tabular data plays an important role in many fields and industries, including those with elevated privacy considerations and risks. As such, there is

researcharxiv-cs-ai
11 May 2026
Local Ai

On the Tradeoffs of On-Device Generative Models in Federated Predictive Maintenance Systems

DGX agent

arXiv:2605.07860v1 Announce Type: cross Abstract: Federated Learning (FL) has emerged as a promising paradigm for preserving client data ownership and control over distributed Internet of Things (IoT)

local-aiarxiv-cs-ai
11 May 2026
Agents

On Time, Within Budget: Constraint-Driven Online Resource Allocation for Agentic Workflows

DGX agent

arXiv:2605.06110v2 Announce Type: replace Abstract: Agentic systems increasingly solve complex user requests by executing orchestrated workflows, where subtasks are assigned to specialized models or t

agentsarxiv-cs-ai
11 May 2026
Safety

One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy

DGX agent

arXiv:2605.07931v1 Announce Type: cross Abstract: Vision-language-action (VLA) models increasingly rely on auxiliary world modules to plan over long horizons, yet how such modules should be parameteri

safetyarxiv-cs-ai
11 May 2026
Safety

Online Allocation with Unknown Shared Supply

DGX agent

arXiv:2605.07080v1 Announce Type: new Abstract: Many real-world resource allocation systems, such as humanitarian logistics and vaccine distribution, must preposition limited supply across multiple lo

safetyarxiv-cs-ai
11 May 2026
Research

Online Goal Recognition using Path Signature and Dynamic Time Warping

DGX agent

arXiv:2605.07736v1 Announce Type: new Abstract: Online goal recognition in continuous domains poses two central challenges: efficiently encoding large trajectories and effectively comparing them. Rece

researcharxiv-cs-ai
11 May 2026
Tutorials

Open-Ended Task Discovery via Bayesian Optimization

DGX agent

arXiv:2605.07572v1 Announce Type: new Abstract: When applying Bayesian optimization (BO) to scientific workflow, a major yet often overlooked source of uncertainty is the task itself -- namely, what t

tutorialsarxiv-cs-ai
11 May 2026
Safety

Operating Within the Operational Design Domain: Zero-Shot Perception with Vision-Language Models

DGX agent

arXiv:2605.07649v1 Announce Type: cross Abstract: Over the last few years, research on autonomous systems has matured to such a degree that the field is increasingly well-positioned to translate resea

safetyarxiv-cs-ai
11 May 2026
Model Releases

Optimal Experiments for Partial Causal Effect Identification

DGX agent

arXiv:2605.06993v1 Announce Type: new Abstract: Causal queries are often only partially identifiable from observational data, and experiments that could tighten the resulting bounds are typically cost

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Optimizing Language Models for Crosslingual Knowledge Consistency

DGX agent

arXiv:2603.04678v2 Announce Type: replace-cross Abstract: Large language models are known to often exhibit inconsistent knowledge. This is particularly problematic in multilingual scenarios, where mod

model-releasesarxiv-cs-ai
11 May 2026
Safety

OrchJail: Jailbreaking Tool-Calling Text-to-Image Agents by Orchestration-Guided Fuzzing

DGX agent

arXiv:2605.07414v1 Announce Type: cross Abstract: Tool-calling text-to-image (T2I) agents can plan and execute multi-step tool chains to accomplish complex generation and editing queries. However, thi

safetyarxiv-cs-ai
11 May 2026
Local Ai

Overcoming data scarcity through multi-center federated learning for organs-at-risk segmentation in pediatric upper abdominal radiotherapy

DGX agent

arXiv:2605.06820v1 Announce Type: cross Abstract: Deep learning-based organs/structures-at-risk(OARs) auto-contouring models can improve radiotherapy workflows, but models trained on adult data often

local-aiarxiv-cs-ai
11 May 2026
Model Releases

PAIR-Former: Budgeted Relational Multi-Instance Learning for Functional miRNA Target Prediction

DGX agent

arXiv:2602.00465v3 Announce Type: replace-cross Abstract: Functional miRNA--mRNA targeting is a large-bag prediction problem where each transcript yields a heavy-tailed pool of candidate target sites

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

PAMPOS: Causal Transformer-based Trajectory Prediction for Attack-Agnostic Misbehavior Detection in V2X Networks

DGX agent

arXiv:2605.06833v1 Announce Type: cross Abstract: Misbehavior detection in Vehicle-to-Everything (V2X) networks is a second line of defense against insider falsification attacks that cryptographic mec

tutorialsarxiv-cs-ai
11 May 2026
Safety

Pan-FM: A Pan-Organ Foundation Model with Saliency-Guided Masking for Missing Robustness

DGX agent

arXiv:2605.07055v1 Announce Type: cross Abstract: Foundation models (FMs) have shown great promise in medical imaging, but most FMs are trained on unimodal data within isolated domains, such as brain

safetyarxiv-cs-ai
11 May 2026
Research

Parallel Lifted Planning via Semi-Naive Datalog Evaluation

DGX agent

arXiv:2605.07584v1 Announce Type: new Abstract: Lifted classical planners operate directly on first-order planning tasks to avoid the computationally demanding grounding step. However, lifted planning

researcharxiv-cs-ai
11 May 2026
Model Releases

PerfCoder: Large Language Models for Interpretable Code Performance Optimization

DGX agent

arXiv:2512.14018v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable progress in automatic code generation, yet their ability to produce high-performance cod

model-releasesarxiv-cs-ai
11 May 2026
Safety

Physical Simulators as Do-Operators: Causal Discovery under Latent Confounders for AI-for-Science

DGX agent

arXiv:2605.07467v1 Announce Type: cross Abstract: Existing interventional causal discovery methods -- IGSP, DCDI, ENCO -- assume causal sufficiency (no latent confounders) and rely on virtual interven

safetyarxiv-cs-ai
11 May 2026
Safety

Physics-Based Benchmarking Metrics for Multimodal Synthetic Images

DGX agent

arXiv:2511.15204v3 Announce Type: replace-cross Abstract: Current state of the art measures like BLEU, CIDEr, VQA score, SigLIP-2 and CLIPScore are often unable to capture semantic or structural accur

safetyarxiv-cs-ai
11 May 2026
Safety

PLOT: Progressive Localization via Optimal Transport in Neural Causal Abstraction

DGX agent

arXiv:2605.06979v1 Announce Type: cross Abstract: Causal abstraction offers a principled framework for mechanistic interpretability, aligning a high-level causal model with the low-level computation r

safetyarxiv-cs-ai
11 May 2026
Safety

POETS: Uncertainty-Aware LLM Optimization via Compute-Efficient Policy Ensembles

DGX agent

arXiv:2605.07775v1 Announce Type: cross Abstract: Balancing exploration and exploitation is a core challenge in sequential decision-making and black-box optimization. We introduce POETS (extbf{Po}licy

safetyarxiv-cs-ai
11 May 2026
Safety

Position: Mechanistic Interpretability Must Disclose Identification Assumptions for Causal Claims

DGX agent

arXiv:2605.08012v1 Announce Type: cross Abstract: Mechanistic interpretability papers increasingly use causal vocabulary: circuits, mediators, causal abstraction, monosemanticity. Such claims require

safetyarxiv-cs-ai
11 May 2026
Safety

Post-training makes large language models less human-like

DGX agent

arXiv:2605.07632v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as surrogates for human participants, but it remains unclear which models best capture human behavi

safetyarxiv-cs-ai
11 May 2026
Research

PPI-Net connects molecular protein interactions to functional processes in disease

DGX agent

arXiv:2605.07838v1 Announce Type: cross Abstract: Understanding how molecular alterations propagate across biological systems to drive disease remains a central challenge. Although high-throughput pro

researcharxiv-cs-ai
11 May 2026
Local Ai

Predictive but Not Plannable: RC-aux for Latent World Models

DGX agent

arXiv:2605.07278v1 Announce Type: cross Abstract: A latent world model may achieve accurate short-horizon prediction while still inducing a latent space that is poorly aligned with planning. A key iss

local-aiarxiv-cs-ai
11 May 2026
Research

Pretraining a Foundation Model for Small-Molecule Natural Products

DGX agent

arXiv:2503.17656v4 Announce Type: replace-cross Abstract: Natural products, as metabolites from microorganisms, animals, or plants, exhibit diverse biological activities, making them crucial for drug

researcharxiv-cs-ai
11 May 2026
Agents

Proactive Instance Navigation with Comparative Judgment for Ambiguous User Queries

DGX agent

arXiv:2605.06223v2 Announce Type: replace Abstract: Natural-language instance navigation becomes challenging when the initial user request does not uniquely specify the target instance. A practical ag

agentsarxiv-cs-ai
11 May 2026
Model Releases

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices

DGX agent

arXiv:2602.21858v4 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have made significant progress in mobile agent development, yet their capabilities are predominantly confin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study

DGX agent

arXiv:2605.07422v1 Announce Type: cross Abstract: Qualitative analysis plays a pivotal role in understanding the human and social aspects of software engineering. However, it remains a demanding proce

model-releasesarxiv-cs-ai
11 May 2026
Research

ProteinJEPA: Latent prediction complements protein language models

DGX agent

arXiv:2605.07554v1 Announce Type: cross Abstract: Protein language models are trained primarily with masked language modeling (MLM), which predicts amino-acid identities at masked positions. We ask wh

researcharxiv-cs-ai
11 May 2026
Safety

Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning

DGX agent

arXiv:2605.07804v1 Announce Type: cross Abstract: On-policy distillation (OPD) leverages dense teacher rewards to enhance reasoning models. However, scaling OPD to long-horizon tasks exposes a critica

safetyarxiv-cs-ai
11 May 2026
Model Releases

PSK@EEUCA 2026: Fine-Tuning Large Language Models with Synthetic Data Augmentation for Multi-Class Toxicity Detection in Gaming Chat

DGX agent

arXiv:2605.07201v1 Announce Type: cross Abstract: This paper describes our system for the EEUCA 2026 Shared Task on Understanding Toxic Behavior in Gaming Communities. The task involves classifying Wo

model-releasesarxiv-cs-ai
11 May 2026
Safety

Q-MMR: Off-Policy Evaluation via Recursive Reweighting and Moment Matching

DGX agent

arXiv:2605.06474v2 Announce Type: replace-cross Abstract: We present a novel theoretical framework, Q-MMR, for off-policy evaluation in finite-horizon MDPs. Q-MMR learns a set of scalar weights, one f

safetyarxiv-cs-ai
11 May 2026
Model Releases

Quality-Conditioned Agreement in Automated Short Answer Scoring: Mid-Range Degradation and the Impact of Task-Specific Adaptation

DGX agent

arXiv:2605.07647v1 Announce Type: cross Abstract: Automated short answer scoring (ASAS) is shifting from discriminative, fine-tuned models to large language models (LLMs) used in few-shot settings. Th

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Query-efficient model evaluation using cached responses

DGX agent

arXiv:2605.07096v1 Announce Type: cross Abstract: Evaluating a new model on an existing benchmark is often necessary to understand its behavior before deployment. For modern evaluation frameworks, gen

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…352353354355356…448
Next →