AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
Safety

Aligning Validation with Deployment: Target-Weighted Cross-Validation for Spatial Prediction

DGX agent

arXiv:2603.29981v2 Announce Type: replace Abstract: Reliable estimation of predictive performance is essential for spatial environmental modeling, where machine-learning models are used to generate ma

safetyarxiv-cs-lg
12 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Alignment as Jurisprudence

DGX agent

arXiv:2605.08416v1 Announce Type: new Abstract: Jurisprudence, the study of how judges should properly decide cases, and alignment, the science of getting AI models to conform to human values, share a

safetyarxiv-cs-ai
12 May 2026
Safety

Alignment-Sensitive Minimax Rates for Spectral Algorithms with Learned Kernels

DGX agent

arXiv:2509.20294v4 Announce Type: replace Abstract: We study spectral algorithms in the setting where kernels are learned from data. We introduce the effective span dimension (ESD), an alignment-sensi

safetyarxiv-cs-lg
12 May 2026
Safety

An Empirical Analysis of Calibration and Selective Prediction in Multimodal Clinical Condition Classification

DGX agent

arXiv:2603.02719v2 Announce Type: replace Abstract: As artificial intelligence systems move toward clinical deployment, ensuring reliable prediction behavior is fundamental for safety-critical decisio

safetyarxiv-cs-lg
12 May 2026
Safety

Anatomical Landmark-Guided Deep Reinforcement Learning for Autonomous Gastric Navigation

DGX agent

arXiv:2605.08269v1 Announce Type: new Abstract: Wireless capsule endoscopy (WCE) enables painless visualization of the gastrointestinal tract, but its diagnostic potential is limited by incomplete muc

safetyarxiv-cs-ro
12 May 2026
Safety

ASACK : Adaptive Safe Active Continual Koopman Learning for Uncertain Systems with Contractive Guarantees

DGX agent

arXiv:2605.09659v1 Announce Type: new Abstract: Koopman operator theory provides a powerful framework for representing nonlinear dynamics through a linear operator acting on lifted observables, enabli

safetyarxiv-cs-ro
12 May 2026
Safety

Assessing the robustness of heterogeneous treatment effects in survival analysis under informative censoring

DGX agent

arXiv:2510.13397v3 Announce Type: replace Abstract: Dropout is common in clinical studies, with up to half of patients leaving early due to side effects or other reasons. When dropout is informative (

safetyarxiv-cs-lg
12 May 2026
Safety

Assessing Trustworthiness of AI Training Dataset using Subjective Logic -- A Use Case on Bias

DGX agent

arXiv:2508.13813v2 Announce Type: replace-cross Abstract: As AI systems increasingly rely on training data, assessing dataset trustworthiness has become critical, particularly for properties like fair

safetyarxiv-cs-ai
12 May 2026
Safety

Attention-Mamba: A Mamba-Enhanced Multi-Scale Parallel Inference Network for Medical Image Segmentation

DGX agent

arXiv:2402.02286v4 Announce Type: replace-cross Abstract: U-shaped architectures have long dominated the field of medical image segmentation, while Transformers are widely employed for modeling long-r

safetyarxiv-cs-ai
12 May 2026
Safety

Attention Sinks in Diffusion Transformers: A Causal Analysis

DGX agent

arXiv:2605.09313v1 Announce Type: new Abstract: Attention sinks -- tokens that receive disproportionate attention mass -- are assumed to be functionally important in autoregressive language models, bu

safetyarxiv-cs-cv
12 May 2026
Safety

Auction-Based Online Policy Adaptation for Evolving Objectives

DGX agent

arXiv:2604.02151v2 Announce Type: replace Abstract: We consider multi-objective reinforcement learning problems where objectives come from an identical family -- such as the class of reachability obje

safetyarxiv-cs-lg
12 May 2026
Safety

Auditing Data Membership in Reinforcement Learning With Verifiable Rewards

DGX agent

arXiv:2511.14045v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a core training stage in recent large language models (LLMs). Its reliance on

safetyarxiv-cs-ai
12 May 2026
Safety

Auto-Rubric as Reward: From Implicit Preferences to Explicit Multimodal Generative Criteria

DGX agent

arXiv:2605.08354v1 Announce Type: new Abstract: Aligning multimodal generative models with human preferences demands reward signals that respect the compositional, multi-dimensional structure of human

safetyarxiv-cs-ai
12 May 2026
Safety

Autonomous FAIR Digital Objects: From Passive Assertions to Active Knowledge

DGX agent

arXiv:2605.10370v1 Announce Type: new Abstract: Scientific knowledge on the Web is published as passive assertions and cannot decide when to validate evidence, reconcile contradictions, or update conf

safetyarxiv-cs-ai
12 May 2026
Safety

Balancing Efficiency and Fairness in Traffic Light Control through Deep Reinforcement Learning

DGX agent

arXiv:2605.10170v1 Announce Type: new Abstract: Urban traffic congestion presents a significant challenge for modern cities, which impacts mobility and sustainability. Traditional traffic light contro

safetyarxiv-cs-lg
12 May 2026
Safety

BathyFacto: Refraction-Aware Two-Media Neural Radiance Fields for Bathymetry

DGX agent

arXiv:2605.10174v1 Announce Type: new Abstract: Through-water photogrammetry based on UAV imagery enables shallow-water bathymetry, but refraction at the air-water interface violates the straight-ray

safetyarxiv-cs-cv
12 May 2026
Safety

BEACON: Cross-Domain Co-Training of Generative Robot Policies via Best-Effort Adaptation

DGX agent

arXiv:2605.08571v1 Announce Type: new Abstract: We introduce BEACON--Best-Effort Adaptation for Cross-Domain Co-Training--a theory-driven framework for training generative robot policies with abundant

safetyarxiv-cs-ro
12 May 2026
Safety

Behavioral Determinants of Deployed AI Agents in Social Networks: A Multi-Factor Study of Personality, Model, and Guardrail Specification

DGX agent

arXiv:2605.08463v1 Announce Type: new Abstract: Autonomous AI agents are increasingly deployed in open social environments, yet the relationship between their configuration specifications and their em

safetyarxiv-cs-ai
12 May 2026
Safety

Beyond Autonomy: A Dynamic Tiered AgentRunner Framework for Governable and Resilient Enterprise AI Execution

DGX agent

arXiv:2605.10223v1 Announce Type: new Abstract: Current large language model agent frameworks prioritize autonomy but lack the governability mechanisms required for enterprise deployment. High-risk wr

safetyarxiv-cs-ai
12 May 2026
Safety

Beyond Continuity: Challenges of Context Switching in Multi-Turn Dialogue with LLMs

DGX agent

arXiv:2605.09268v1 Announce Type: cross Abstract: Users interacting with Large Language Models (LLMs) in a multi-turn conversation routinely refine their requests or pivot to new topics. LLMs, however

safetyarxiv-cs-ai
12 May 2026
Safety

Beyond ESG Scores: Learning Dynamic Constraints for Sequential Portfolio Optimization

DGX agent

arXiv:2605.09310v1 Announce Type: new Abstract: ESG-aware portfolio optimization is increasingly important for sustainable capital allocation, yet most learning-based methods still operationalize ESG

safetyarxiv-cs-ai
12 May 2026
Safety

Beyond Multiple Choice: Evaluating Steering Vectors for Summarization

DGX agent

arXiv:2505.24859v3 Announce Type: replace-cross Abstract: Steering vectors are a lightweight method for controlling text properties by adding a learned bias to language model activations at inference

safetyarxiv-cs-cl
12 May 2026
Safety

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning

DGX agent

arXiv:2605.08202v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) faces a critical challenge of overestimating the value of out-of-distribution (OOD) actions. Existing methods miti

safetyarxiv-cs-ai
12 May 2026
Safety

Beyond Position Bias: Shifting Context Compression from Position-Driven to Semantic-Driven

DGX agent

arXiv:2605.09463v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated exceptional performance across diverse tasks. However, their deployment in long-context scenarios faces h

safetyarxiv-cs-cl
12 May 2026
Safety

Beyond Self-Play: Hierarchical Reasoning for Continuous Motion in Closed-Loop Traffic Simulation

DGX agent

arXiv:2605.09153v1 Announce Type: cross Abstract: Closed-loop traffic simulation requires agents that are both scalable and behaviorally realistic. Recent self-play reinforcement learning approaches d

safetyarxiv-cs-ai
12 May 2026
Safety

Beyond Static Bias: Adaptive Multi-Fidelity Bandits with Improving Proxies

DGX agent

arXiv:2605.08558v1 Announce Type: new Abstract: As an extension of the classical multi-armed bandit problem, multi-fidelity multi-armed bandits (MF-MAB) enable individual arms to be evaluated using di

safetyarxiv-cs-lg
12 May 2026
Safety

Bi-CoG: Bi-Consistency-Guided Self-Training for Vision-Language Models

DGX agent

arXiv:2510.20477v2 Announce Type: replace Abstract: Exploiting unlabeled data through semi-supervised learning (SSL) or leveraging pre-trained models via fine-tuning are two prevailing paradigms for a

safetyarxiv-cs-lg
12 May 2026
Safety

Bias by Necessity: Impossibility Theorems for Sequential Processing with Convergent AI and Human Validation

DGX agent

arXiv:2605.08716v1 Announce Type: new Abstract: Are certain cognitive biases mathematically inevitable consequences of sequential information processing? We prove that primacy effects, anchoring, and

safetyarxiv-cs-ai
12 May 2026
Safety

Big AI is accelerating the metacrisis: What can we do?

DGX agent

arXiv:2512.24863v2 Announce Type: replace-cross Abstract: The world is in the grip of ecological, meaning, and language crises that are converging into a metacrisis. Big AI is accelerating them all. L

safetyarxiv-cs-ai
12 May 2026
Safety

Biological Plausibility and Representational Alignment of Feedback Alignment in Convolutional Networks

DGX agent

arXiv:2605.08564v1 Announce Type: new Abstract: The feedback alignment (FA) algorithm offers a biologically plausible alternative to backpropagation (BP) for training neural networks yet notably fails

safetyarxiv-cs-ai
12 May 2026
Safety

Block-Wise Differentiable Sinkhorn Attention: Tail-Refinement Gradients with a Gap-Aware Dustbin Bridge

DGX agent

arXiv:2605.08123v1 Announce Type: cross Abstract: We study long-context balanced entropic optimal transport (OT) attention on TPU hardware through a stopped-base, fixed-depth tail-refinement surrogate

safetyarxiv-cs-cl
12 May 2026
Safety

BONSAI: Bayesian Optimization with Natural Simplicity and Interpretability

DGX agent

arXiv:2602.07144v2 Announce Type: replace-cross Abstract: Bayesian optimization (BO) is a popular technique for sample-efficient optimization of black-box functions. In many applications, the paramete

safetyarxiv-cs-ai
12 May 2026
Safety

Break the Brake, Not the Wheel: Untargeted Jailbreak via Entropy Maximization

DGX agent

arXiv:2605.10764v1 Announce Type: cross Abstract: Recent studies show that gradient-based universal image jailbreaks on vision-language models (VLMs) exhibit little or no cross-model transferability,

safetyarxiv-cs-ai
12 May 2026
Safety

Breaking. Sam Altman himself finally confirms what I was the first to point out publicly: his indirect equity stake in OpenAI. … which he fa…

DGX agent

Breaking. Sam Altman himself finally confirms what I was the first to point out publicly: his indirect equity stake in OpenAI. … which he failed to acknowledge when asked about his financial interest

safetygary-marcus--x
12 May 2026
Safety

Breaking the Grid: Distance-Guided Reinforcement Learning in Large Discrete Action Spaces

DGX agent

arXiv:2602.08616v2 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) is increasingly applied to large-scale decision-making problems like logistics, scheduling, and recommender system

safetyarxiv-cs-ai
12 May 2026
Safety

Breaking the Impasse: Dual-Scale Evolutionary Policy Training for Social Language Agents

DGX agent

arXiv:2605.08721v1 Announce Type: new Abstract: While Reinforcement Learning with Verifiable Rewards (RLVR) has proven effective for closed-ended tasks, extending it to open-ended social language game

safetyarxiv-cs-cl
12 May 2026
Safety

BRIDGE: Building Representations In Domain Guided Program Synthesis

DGX agent

arXiv:2511.21104v3 Announce Type: replace Abstract: Large language models can generate plausible code, but remain brittle for formal verification in proof assistants such as Lean. A central scalabilit

safetyarxiv-cs-lg
12 May 2026
Safety

BROS: Bias-Corrected Randomized Subspaces for Memory-Efficient Single-Loop Bilevel Optimization

DGX agent

arXiv:2605.10288v1 Announce Type: new Abstract: Stochastic bilevel optimization (SBO) has become a standard framework for hyperparameter learning, data reweighting, representation learning, and data-m

safetyarxiv-cs-lg
12 May 2026
Safety

Budget-Efficient Automatic Algorithm Design via Code Graph

DGX agent

arXiv:2605.10598v1 Announce Type: new Abstract: Large language models (LLMs) have emerged as powerful tools for automatic algorithm design (AAD). However, existing pipelines remain inefficient. They o

safetyarxiv-cs-ai
12 May 2026
Safety

CalBench: Evaluating Coordination-Privacy Trade-offs in Multi-Agent LLMs

DGX agent

arXiv:2605.09823v1 Announce Type: cross Abstract: We introduce CalBench, a controlled evaluation environment for studying multi-agent coordination through calendar scheduling. In CalBench, N agents ea

safetyarxiv-cs-ai
12 May 2026
Safety

CAMAL: Improving Attention Alignment and Faithfulness with Segmentation Masks

DGX agent

arXiv:2605.08325v1 Announce Type: cross Abstract: Many vision datasets now provide segmentation masks in addition to annotated images to support a wide range of tasks. In this work, we propose Class A

safetyarxiv-cs-ai
12 May 2026
Safety

Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction

DGX agent

arXiv:2512.18880v2 Announce Type: replace-cross Abstract: Accurate estimation of item (question or task) difficulty is critical for educational assessment but suffers from the cold start problem. Whil

safetyarxiv-cs-ai
12 May 2026
Safety

Can Revealed Preferences Clarify LLM Alignment and Steering?

DGX agent

arXiv:2605.08556v1 Announce Type: new Abstract: LLMs are increasingly used to make or support high-stakes decisions under uncertainty, where alignment depends not only on factual accuracy but on how m

safetyarxiv-cs-lg
12 May 2026
Safety

CapCLIP: A Vision-Language Representation Alignment Approach for Wireless Capsule Endoscopy Analysis

DGX agent

arXiv:2605.08493v1 Announce Type: new Abstract: Wireless capsule endoscopy (WCE) enables non-invasive visual assessment of the small bowel, but its clinical utility is constrained by the large volume

safetyarxiv-cs-cv
12 May 2026
Safety

CARL: Criticality-Aware Agentic Reinforcement Learning

DGX agent

arXiv:2512.04949v3 Announce Type: replace-cross Abstract: Agents capable of accomplishing complex tasks through multiple interactions with the environment have emerged as a popular research direction.

safetyarxiv-cs-ai
12 May 2026
Safety

Causal Explanations from the Geometric Properties of ReLU Neural Networks

DGX agent

arXiv:2605.10396v1 Announce Type: new Abstract: Neural networks have proved an effective means of learning control policies for autonomous systems, but these learned policies are difficult to understa

safetyarxiv-cs-lg
12 May 2026
Safety

CFSPMNet: Cross-subject Fourier-guided Spatial-Patch Mamba Network for EEG Motor Imagery Decoding in Stroke Patients

DGX agent

arXiv:2605.10111v1 Announce Type: cross Abstract: Motor imagery electroencephalography (MI-EEG) decoding offers a non-invasive route for post-stroke rehabilitation, but cross-patient use remains diffi

safetyarxiv-cs-ai
12 May 2026
Safety

Change My View? The Dynamics of Persuasion and Polarization in Online Discourse

DGX agent

arXiv:2605.08383v1 Announce Type: new Abstract: Philosophical accounts of persuasion often assume that shared evidence and rational argumentation should lead to a convergence of views between peers, y

safetyarxiv-cs-cl
12 May 2026
← Previous
1…186187188189190…267
Next →