AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,816 results
Safety

Can Aerial VLA Models Cooperate? Evaluating Closed-Loop Air-Ground Coordination with CARLA-Air

DGX agent

arXiv:2605.31066v1 Announce Type: new Abstract: Recent aerial vision-language-action (VLA) models show promising single-UAV capabilities, such as tracking moving objects and navigating to language-spe

safetyarxiv-cs-ro
1 Jun 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Causal Evaluation of Membership Inference Attacks

DGX agent

arXiv:2602.02819v3 Announce Type: replace Abstract: Membership Inference Attacks (MIAs) aim to distinguish training points (members) from unseen data (non-members), and are widely used to quantify mem

safetyarxiv-cs-lg
1 Jun 2026
Safety

CellBRIDGE: Learning Cellular Trajectories via Interaction-Aware Alignment

DGX agent

arXiv:2605.30635v1 Announce Type: new Abstract: Inferring dynamics from population snapshots is a fundamental challenge in machine learning and biology. In scRNA-sequencing (scRNA-seq), destructive me

safetyarxiv-cs-lg
1 Jun 2026
Safety

COFT: Counterfactual-Conformal Decoding for Fair Chain-of-Thought Reasoning in Large Language Models

DGX agent

arXiv:2605.30641v1 Announce Type: cross Abstract: Large language models (LLMs) can reveal and amplify societal biases during chain-of-thought (CoT) generation. We present COFT (Chain of Fair Thought),

safetyarxiv-cs-ai
1 Jun 2026
Safety

COMPASS: Cognitive MCTS-Guided Process Alignment for Safe Search Agents

DGX agent

arXiv:2605.30838v1 Announce Type: new Abstract: LLM-powered search agents enable multi-step reasoning and tool use. However, these capabilities introduce retrieval-induced safety degradation, as harmf

safetyarxiv-cs-ai
1 Jun 2026
Safety

Configurable Reward Model for Balanced Safety Alignment

DGX agent

arXiv:2605.30487v1 Announce Type: new Abstract: Aligning large language models (LLMs) to heterogeneous and rapidly evolving safety requirements remains a critical challenge. Existing instruction-tuned

safetyarxiv-cs-cl
1 Jun 2026
Safety

ConSensus: Multi-Agent Collaboration for Multimodal Sensing

DGX agent

arXiv:2601.06453v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly grounded in sensor data to perceive and reason about human physiology and the physical world. However,

safetyarxiv-cs-ai
1 Jun 2026
Safety

ConsisGuard: Aligning Safety Deliberation with Policy Enforcement in LLM Guardrails

DGX agent

arXiv:2605.31073v1 Announce Type: new Abstract: Reasoning-based LLM guardrails improve safety moderation by generating explicit rationales before issuing final decisions. However, their rationales do

safetyarxiv-cs-cl
1 Jun 2026
Safety

Constrained Multi-Objective Reinforcement Learning with Max-Min Criterion

DGX agent

arXiv:2605.31388v1 Announce Type: new Abstract: Multi-Objective Reinforcement Learning (MORL) extends standard RL by optimizing policies with respect to multiple, often conflicting, objectives. While

safetyarxiv-cs-lg
1 Jun 2026
Safety

Contextual Scalarisation Thompson Sampling for multi-objective decisions in public media

DGX agent

arXiv:2605.31291v1 Announce Type: cross Abstract: Recommender systems may operate under multiple, competing objectives. For example, audience reach, cultural values, public service mandate, and operat

safetyarxiv-cs-lg
1 Jun 2026
Safety

Cost-aware Stopping for Bayesian Optimization

DGX agent

arXiv:2507.12453v5 Announce Type: replace Abstract: In automated machine learning, scientific discovery, and other applications of Bayesian optimization, deciding when to stop evaluating expensive bla

safetyarxiv-cs-lg
1 Jun 2026
Safety

Counterfactual Evaluation Reveals Hidden Capability Profiles in Clinical LLMs and Agents

DGX agent

arXiv:2605.30590v1 Announce Type: cross Abstract: Two clinical AI systems can score nearly identically on coverage-based rubrics yet behave radically differently when their patient inputs change: one

safetyarxiv-cs-ai
1 Jun 2026
Safety

Cross-Modal Attention Calibration for LVLM Hallucination Mitigation

DGX agent

arXiv:2501.01926v3 Announce Type: replace-cross Abstract: Large vision-language models (LVLMs) have shown remarkable capabilities in visual-language understanding. Despite their success, LVLMs still s

safetyarxiv-cs-ai
1 Jun 2026
Safety

Current AI is bandaids all the way down; LLMs can’t play nicely with basic tools like databases and knowledge graphs, and you never know wha…

DGX agent

Current AI is bandaids all the way down; LLMs can’t play nicely with basic tools like databases and knowledge graphs, and you never know what you are going to get. It’s long time to face facts: LLMs h

safetygary-marcus--x
1 Jun 2026
Safety

dashi: A Python library for Dataset Shift Characterization to Support Trustworthy AI Development and Deployment

DGX agent

arXiv:2605.31360v1 Announce Type: cross Abstract: The Artificial Intelligence (AI) life cycle requires a thorough understanding of the underlying data dynamics for robust, safe and cost-effective AI d

safetyarxiv-cs-ai
1 Jun 2026
Safety

Decoding the Surgical Scene: A Scoping Review of Scene Graphs in Surgery

DGX agent

arXiv:2509.20941v2 Announce Type: replace Abstract: As surgical AI transitions from pixel-level detection to complex reasoning, Scene Graphs (SGs) offer the structured, relational representations nece

safetyarxiv-cs-cv
1 Jun 2026
Safety

Detect in Any Scene: An Agentic Framework for Object Detection with Experience-Aware Reasoning

DGX agent

arXiv:2605.31174v1 Announce Type: new Abstract: Object detection in real-world scenarios remains challenging due to diverse image degradations and heterogeneous object distributions, which significant

safetyarxiv-cs-cv
1 Jun 2026
Safety

Diagnosing Failure Modes of Shared-State Collaboration in Resource-Constrained Visual Agents

DGX agent

arXiv:2605.31354v1 Announce Type: new Abstract: Modular visual reasoning systems increasingly rely on shared working memory for multi-step collaboration, yet the failure dynamics of intermediate state

safetyarxiv-cs-ai
1 Jun 2026
Safety

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory

DGX agent

arXiv:2602.00521v2 Announce Type: replace Abstract: While LLM-as-a-Judge is widely used in automated evaluation, existing validation practices primarily operate at the level of observed outputs, offer

safetyarxiv-cs-ai
1 Jun 2026
Safety

Differentially Private Preference Data Synthesis for Large Language Model Alignment

DGX agent

arXiv:2605.30808v1 Announce Type: cross Abstract: Preference alignment is a crucial post-training step for large language models (LLMs) to ensure their outputs align with human values. However, post-t

safetyarxiv-cs-ai
1 Jun 2026
Safety

DISCO: Mitigating Bias in Deep Learning with Conditional Distance Correlation

DGX agent

arXiv:2506.11653v3 Announce Type: replace-cross Abstract: Dataset bias often leads deep learning models to exploit spurious correlations instead of task-relevant signals. We introduce the Standard Ant

safetyarxiv-cs-ai
1 Jun 2026
Safety

Distilling LLM Feedback for Lean Theorem Proving

DGX agent

arXiv:2605.30861v1 Announce Type: new Abstract: Post-training for reasoning models typically combines supervised fine-tuning with reinforcement learning from verifiable rewards, most commonly with GRP

safetyarxiv-cs-ai
1 Jun 2026
Safety

DiTTo: Scalable Order-aware All-in-One Image Restoration Agent

DGX agent

arXiv:2605.30915v1 Announce Type: new Abstract: Real-world images rarely suffer from a single degradation, and the order in which degradations are removed substantially affects the final restoration q

safetyarxiv-cs-cv
1 Jun 2026
Safety

DOA: Training-Free Decoder-Only Attention Policy for Long-Form Simultaneous Translation with SpeechLLMs

DGX agent

arXiv:2605.31432v1 Announce Type: cross Abstract: Simultaneous speech-to-text translation (SimulST) generates translations while speech is still unfolding, requiring a streaming policy that decides wh

safetyarxiv-cs-ai
1 Jun 2026
Safety

Does Visual Information Play a Decisive Role in Vision-Language-Action Model Driving Behavior?

DGX agent

arXiv:2605.31041v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated promising capability in autonomous driving, highlighting the potential of unified multimodal arc

safetyarxiv-cs-ai
1 Jun 2026
Safety

DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization

DGX agent

arXiv:2605.31455v1 Announce Type: cross Abstract: Large language models are increasingly deployed in multi-turn interactive settings where users or environments can iteratively provide lightweight fee

safetyarxiv-cs-cl
1 Jun 2026
Safety

Dual Mechanisms of Value Expression: Intrinsic vs. Prompted Values in Large Language Models

DGX agent

arXiv:2509.24319v4 Announce Type: replace-cross Abstract: Large language models can express values in two main ways: (1) intrinsic expression, reflecting the model's inherent values learned during tra

safetyarxiv-cs-ai
1 Jun 2026
Safety

EchoRL: Reinforcement Learning via Rollout Echoing

DGX agent

arXiv:2605.31228v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards is an effective route for post-training to strengthen the reasoning capability of large language models

safetyarxiv-cs-ai
1 Jun 2026
Safety

Efficient and Uncertainty-Aware Diffusion Framework for Offline-to-Online Reinforcement Learning

DGX agent

arXiv:2605.30776v1 Announce Type: new Abstract: Offline-to-Online Reinforcement Learning (O2O-RL) leverages an offline, pre-trained policy to minimize costly online interactions. Although data-efficie

safetyarxiv-cs-lg
1 Jun 2026
Safety

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation

DGX agent

arXiv:2605.30484v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown promise for robotic manipulation, yet most existing policies operate reactively by directly regressing ac

safetyarxiv-cs-ro
1 Jun 2026
Safety

Elon was right that what OpenAI did in its for profit turn was shameful. But what he is doing with the SpaceX, at the likely expense of many…

DGX agent

Elon was right that what OpenAI did in its for profit turn was shameful. But what he is doing with the SpaceX, at the likely expense of many people’s retirement funds, is just as shameful. SpaceX bein

safetygary-marcus--x
1 Jun 2026
Safety

Enhancing Regime Shift Detection Using Unstructured Data: A Study on the Treasury Market

DGX agent

arXiv:2605.30363v1 Announce Type: cross Abstract: Regime shifts in financial markets reorganise the joint dynamics of asset prices and macro variables, breaking any single-regime calibration. They are

safetyarxiv-cs-ai
1 Jun 2026
Safety

Entropic Projection Alignment: Estimating, Explaining, and Improving Model Performance Under Distribution Shift

DGX agent

arXiv:2605.31250v1 Announce Type: cross Abstract: We propose a unified framework for addressing three key challenges of distribution shift: (1) estimating a model's performance on an unlabeled target

safetyarxiv-cs-ai
1 Jun 2026
Safety

Envisioning Beyond the Few: Disentangled Semantics and Primitives for Few-Shot Atypical Layout-to-Image Generation

DGX agent

arXiv:2605.31266v1 Announce Type: cross Abstract: The layout-to-image (L2I) task enables fine-grained control over image generation via object categories and spatial layouts. However, existing L2I met

safetyarxiv-cs-ai
1 Jun 2026
Safety

Equivariant Latent Alignment via Flow Matching under Group Symmetries

DGX agent

arXiv:2605.30705v1 Announce Type: new Abstract: Geometry-aware generative models and novel view synthesis approaches have shown strong potential in visual fidelity and consistency. In parallel, equiva

safetyarxiv-cs-cv
1 Jun 2026
Safety

Exploiting Chordal Sparsity for Globally Optimal Estimation with Factor Graphs

DGX agent

arXiv:2605.30617v1 Announce Type: new Abstract: Robust and efficient state estimation is crucial for perception, navigation, and control in robotics. State estimation problems are conveniently modeled

safetyarxiv-cs-ro
1 Jun 2026
Safety

Extending the UXR Point of View Pyramid: A Generative AI-Augmented Methodology for Human-Centred AI Systems

DGX agent

arXiv:2605.31143v1 Announce Type: cross Abstract: Rising household debt and cost-of-living pressures in the United Kingdom have intensified the role of AI-driven financial technologies in mediating cr

safetyarxiv-cs-ai
1 Jun 2026
Safety

Fair Decisions from Calibrated Scores: Achieving Optimal Classification While Satisfying Sufficiency

DGX agent

arXiv:2602.07285v2 Announce Type: replace Abstract: Binary classification based on predicted probabilities (scores) is a fundamental task in supervised machine learning. While thresholding scores is B

safetyarxiv-cs-lg
1 Jun 2026
Safety

Feat2Go: Visual Feature-Grounded Value Estimation for Embodied Reinforcement Learning

DGX agent

arXiv:2605.30795v1 Announce Type: new Abstract: Reinforcement learning is a promising approach for improving the capabilities of vision-language-action (VLA) models while avoiding the heavy data requi

safetyarxiv-cs-ro
1 Jun 2026
Safety

Feature-Optimized Vision for Adaptive 3D Scene Reconstruction

DGX agent

arXiv:2605.31534v1 Announce Type: cross Abstract: Three-dimensional scene reconstruction depends on local image evidence that is both visually discriminative and geometrically useful. Fixed feature th

safetyarxiv-cs-ai
1 Jun 2026
Safety

FLAG: Flow Policy MaxEnt-RL by Latent Augmented Guidance

DGX agent

arXiv:2605.30749v1 Announce Type: new Abstract: Maximum entropy reinforcement learning (MaxEnt-RL) enables robust exploration, yet practical implementations often restrict policies to simple Gaussians

safetyarxiv-cs-lg
1 Jun 2026
Safety

Florida AG James Uthmeier sues OpenAI and Sam Altman, seeking to hold Altman personally liable for deceptive trade practices, negligence, and public nuisance (NBC News)

DGX agent

NBC News: Florida AG James Uthmeier sues OpenAI and Sam Altman, seeking to hold Altman personally liable for deceptive trade practices, negligence, and public nuisance — Florida's seeks to hold Sam Al

safetytechmeme
1 Jun 2026
Safety

Forecasting with Hyper-Trees

DGX agent

arXiv:2405.07836v5 Announce Type: replace Abstract: We introduce Hyper-Trees as a novel framework for modeling time series data using gradient boosted trees. Unlike conventional tree-based approaches

safetyarxiv-cs-lg
1 Jun 2026
Safety

From Evidence to Design: Developing an AI-Augmented UX Research Point of View for Digital Wellbeing in Emergency and Public Safety Contexts

DGX agent

arXiv:2605.31146v1 Announce Type: cross Abstract: This paper investigates how User Experience Research (UXR) methods can be combined with AI-supported analysis to develop clearer design direction for

safetyarxiv-cs-ai
1 Jun 2026
Safety

From Internal Diagnosis to External Auditing: A VLM-Driven Paradigm for Data-Free Online Backdoor Defense

DGX agent

arXiv:2601.19448v2 Announce Type: replace Abstract: Deep Neural Networks remain inherently vulnerable to backdoor attacks. Traditional test-time defenses largely operate under the paradigm of internal

safetyarxiv-cs-lg
1 Jun 2026
Safety

From Out-of-Distribution Detection to Hallucination Detection: A Geometric View

DGX agent

arXiv:2602.07253v2 Announce Type: replace Abstract: Detecting hallucinations in large language models is a critical open problem with significant implications for safety and reliability. While existin

safetyarxiv-cs-ai
1 Jun 2026
Safety

Geometry-Aware Control Barrier Functions for Collision Avoidance via Bernstein Polynomial Approximations

DGX agent

arXiv:2605.30696v1 Announce Type: new Abstract: Safe navigation often relies on well-defined conditions based on the shape of robots and obstacles, and can be challenging when they have irregular geom

safetyarxiv-cs-ro
1 Jun 2026
Safety

GlucoFM: A Dual-Stream Foundation Model for Continuous Glucose Monitoring

DGX agent

arXiv:2605.30865v1 Announce Type: new Abstract: Continuous glucose monitoring (CGM) provides a dense view of daily metabolic physiology, yet existing generic time-series and CGM-specific foundation mo

safetyarxiv-cs-lg
1 Jun 2026
← Previous
1…127128129130131…267
Next →