AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
8 Jul 2026

Our approach to government and national security partnerships

SafetyDGX agent

OpenAI outlines its strategy for collaborating with government agencies and national security institutions, likely covering principles for responsible AI deployment in defense and security contexts. T

Platonic Representations for Poverty Mapping: Unified Vision-Language Codes or Agent-Induced Novelty?

SafetyDGX agent

arXiv:2508.01109v3 Announce Type: replace Abstract: We investigate whether socioeconomic indicators, like household wealth, leave recoverable informational imprints in both satellite imagery (capturin

PORTS: Preference-Optimized Retrievers for Tool Selection with Large Language Models

SafetyDGX agent

arXiv:2607.05441v1 Announce Type: cross Abstract: Integrating external tools with Large Language Models (LLMs) has emerged as a promising paradigm for accomplishing complex tasks. Since LLMs still str


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Position: EU AI Act's Research Exemptions Can Break the Publication Norms of Major AI Conferences

SafetyDGX agent

arXiv:2506.03218v2 Announce Type: replace-cross Abstract: The EU has become one of the vanguards in regulating the digital age. A particularly important regulation in the Artificial Intelligence (AI)

Position: Preventing AI-Generated CSAM Necessitates New Approaches to AI Safety

SafetyDGX agent

arXiv:2607.05407v1 Announce Type: cross Abstract: Modern artificial intelligence (AI) systems present profound new risks to child safety. AI is increasingly being misused to create AI-generated child

Property-Driven Synthetic Data Engineering for Data-Scarce Software Systems: Reflections from the Breast Cancer Domain

SafetyDGX agent

arXiv:2607.06133v1 Announce Type: cross Abstract: Modern software systems increasingly depend on data for analysis, prediction, testing, and decision-making. Yet many important domains, including medi

Quantifying Retriever-Generator Alignment in RAG with Local Explanations

SafetyDGX agent

arXiv:2601.21803v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems combine dense retrievers and language models to ground their outputs in external documents. However, th

Realistic Compound-Lens Defocus Blur Synthesis

SafetyDGX agent

arXiv:2607.05837v1 Announce Type: new Abstract: Defocus blur degrades fine image structures and limits visual perception, which can adversely affect downstream vision tasks. Although recent deep learn

Recovering Cloud Microstructures with Cascaded Diffusion Inversion

SafetyDGX agent

arXiv:2607.05637v1 Announce Type: new Abstract: High-resolution satellite imagery is critical for observing fine-scale cloud structures that inform weather modification strategies like cloud seeding f

RoboTALES: Learning Reasoning-Guided Robot Policies via Task-Aligned Simulated Futures

SafetyDGX agent

arXiv:2607.06018v1 Announce Type: new Abstract: Pretrained video generative models are promising backbones for visuomotor control, but their imagined futures often drift from task intent and are not r

RynnWorld-4D: 4D Embodied World Models for Robotic Manipulation

SafetyDGX agent

arXiv:2607.06559v1 Announce Type: new Abstract: Robotic manipulation in the open world requires not only recognizing what a scene looks like, but also anticipating how its 3D structure moves under int

Safe Bayesian Optimization with Counterfactual Policies

SafetyDGX agent

arXiv:2607.05620v1 Announce Type: cross Abstract: In many decision-making settings, new interventions are acceptable only if they do not reduce outcomes below some established threshold. For example,

SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation

SafetyDGX agent

arXiv:2607.05943v1 Announce Type: new Abstract: Training multimodal search agents to perform multi-hop reasoning remains challenging due to a fundamental structural disconnect: existing pipelines cons

Securing Amazon Bedrock AgentCore Runtime with AWS WAF

SafetyDGX agent

This post shows you two architecture patterns that address this problem. Both use an internet-facing ALB with AWS WAF and route traffic through a VPC Interface Endpoint to AgentCore Runtime. Pattern 1

SpatialFly: Implicit 3D Prior-Guided Visual Reparameterization for Continuous UAV Vision-and-Language Navigation

SafetyDGX agent

arXiv:2603.21046v2 Announce Type: replace-cross Abstract: UAVs play an important role in applications such as autonomous exploration, disaster response, and infrastructure inspection. However, UAV VLN

Stability Annealing Selects the Implicit Bias of Smoothed Sign Descent: A Rate-Indexed Barrier Path on Separable Data

SafetyDGX agent

arXiv:2607.06013v1 Announce Type: new Abstract: Adaptive gradient methods can favor max-margin separators that differ from gradient descent, yet a fixed positive numerical stability constant eventuall

Statistical Adversaries: Natural Backdoor-like Features in Vision Datasets

SafetyDGX agent

arXiv:2607.05516v1 Announce Type: cross Abstract: Model-specific adversarial attacks have been extensively studied. We study a different failure mode: naturally occurring statistical signals in vision

Straight-Path Flow Matching for Incomplete Multi-View Clustering

SafetyDGX agent

arXiv:2607.06281v1 Announce Type: new Abstract: Incomplete Multi-View Clustering addresses the problem of clustering multi-modal data when certain views are missing. Recent end-to-end generative appro

Structured-Condensed Prompt Tuning in Vision-Language Models for Fine-grained Image Recognition

SafetyDGX agent

arXiv:2607.06185v1 Announce Type: new Abstract: Fine-grained image recognition poses a significant challenge due to the substantial expertise and effort required for manual annotation. Vision-language

Superman: Unifying Skeleton and Vision for Human Motion Perception and Generation

SafetyDGX agent

arXiv:2602.02401v2 Announce Type: replace Abstract: Human motion analysis tasks, such as temporal 3D pose estimation, motion prediction, and motion in-betweening, play an essential role in computer vi

Synthetic-to-Real Translation for Class-Agnostic Motion Prediction

SafetyDGX agent

arXiv:2607.06319v1 Announce Type: new Abstract: Motion understanding is critical for ensuring safety and robustness in autonomous driving systems, driving increasing interest in motion prediction. A k

The Jagged Global Economy: Frontier AI Unevenly Exposes National Economies

SafetyDGX agent

arXiv:2607.05404v1 Announce Type: cross Abstract: Frontier AI's labor-market effects matter to workers, firms, and policymakers, but current evidence generally comes from a handful of high-income econ

The Large Cancer Assistant (LCA): A Model-Agnostic Orchestration Framework for Scalable Clinical Decision Support in Oncology

SafetyDGX agent

arXiv:2607.06531v1 Announce Type: new Abstract: - Objective: Multimodal deep learning models in oncology are currently limited by monolithic designs that rigidly couple data ingestion, clinical routin

TILDE: TILt-based Distributional Erasure for Concept Unlearning

SafetyDGX agent

arXiv:2607.06432v1 Announce Type: cross Abstract: Concept unlearning in text-to-image diffusion models is critical for safe and practical deployment: with rising privacy concerns, copyright disputes,

To Retain or to Adapt? Generalizing Continual Learning

SafetyDGX agent

arXiv:2607.05609v1 Announce Type: cross Abstract: The Continual Learning (CL) literature has long been driven by the goal of mitigating catastrophic forgetting. This objective rests on a pervasive, of

Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations

SafetyDGX agent

arXiv:2504.05294v3 Announce Type: replace Abstract: Chain-of-thought explanations are widely used to inspect the decision process of large language models (LLMs) and to evaluate the trustworthiness of

TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training

SafetyDGX agent

arXiv:2607.05804v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student policy by matching a stronger teacher on the student's own trajectories, offering a promising framework fo

UniField: A Unified Field-Aware MRI Enhancement Framework

SafetyDGX agent

arXiv:2603.09223v2 Announce Type: replace Abstract: Magnetic Resonance Imaging (MRI) field-strength enhancement holds immense value for both clinical diagnostics and advanced research. However, existi

VEIL: How Visual Encoding Hijacking Induces Bias In Vision Models

SafetyDGX agent

arXiv:2607.05641v1 Announce Type: new Abstract: Rendering time series as chart images for CNN-based classification has become increasingly common in time-series classification (TSC). However, it remai

Volumetric Directional Diffusion: Anchoring Uncertainty Quantification in Anatomical Consensus for Ambiguous Medical Image Segmentation

SafetyDGX agent

arXiv:2603.04024v2 Announce Type: replace-cross Abstract: Ambiguous 3D medical image segmentation often involves boundaries where different expert delineations are non-identical yet clinically plausib

When Lower Privileges Suffice: Investigating Over-Privileged Tool Selection in LLM Agents

SafetyDGX agent

arXiv:2606.20023v2 Announce Type: replace-cross Abstract: As LLM agents increasingly select tools autonomously, their choices among tools with different privileges become safety-relevant. However, pri

Whose fairness? Structural concentration in AI bias research

SafetyDGX agent

arXiv:2607.05574v1 Announce Type: cross Abstract: Artificial intelligence increasingly mediates consequential decisions in healthcare, law, and public services, and the field has responded with an ext

Width-Robust Learnability in Mean-Field Bayesian Neural Networks

SafetyDGX agent

arXiv:2607.05735v1 Announce Type: cross Abstract: Infinite-width limits are a standard way to reason about neural networks, but it is not automatic that the limiting learner has the same complexity-th

WordVoice: Explicit and Decoupled Multi-Dimensional Word-Level Control for LLM-Based TTS

SafetyDGX agent

arXiv:2607.06461v1 Announce Type: cross Abstract: While recent Large Language Model (LLM)-based Text-to-Speech (TTS) systems have achieved remarkable naturalness, they predominantly rely on implicit e

X-FEMR: A Token-level Explainable Approach for Electronic Health Records Foundation Models using Transformer-based Models

SafetyDGX agent

arXiv:2607.06163v1 Announce Type: cross Abstract: Foundation Models for Electronic Health Records (FEMRs) are pretrained on large-scale structured patient data, enabling them to convert longitudinal p

7 Jul 2026

A Bayesian Framework for Evaluating Scenario Compatibility in Generative Population Synthesis

SafetyDGX agent

arXiv:2607.03190v1 Announce Type: cross Abstract: Scenario-based transportation analysis specifies future assumptions through aggregate population targets, whereas generative population synthesis mode

A Few Teacher Steps Go a Long Way: Cost-Efficient On-Policy Data Augmentation for Agent Post-Training

SafetyDGX agent

arXiv:2607.04574v1 Announce Type: cross Abstract: For LLM agents, supervised fine-tuning is not only about teacher labels' quality, but also about which interaction contexts those labels condition on.

A Graph-Based Reinforcement Learning Approach with Frontier Potential Based Reward for Safe Cluttered Environment Exploration

SafetyDGX agent

arXiv:2504.11907v3 Announce Type: replace Abstract: Autonomous exploration of cluttered environments requires efficient exploration strategies that guarantee safety against potential collisions with u

A Hierarchy of Policy Learning Problems

SafetyDGX agent

arXiv:2607.03385v1 Announce Type: cross Abstract: Policy learning has received substantial attention with the goal of learning policies from observational data for decision-making. A majority of work

A Mathematical Theory of Value: a synthesis on goal-directed agency under resource constraints

SafetyDGX agent

arXiv:2606.12502v2 Announce Type: replace-cross Abstract: We propose that value -- the quantity goal-directed agents create, destroy, and exchange -- is a lawful structural quantity in the same catego

A Perception-Manipulation Robotics System for Food Cutting

SafetyDGX agent

arXiv:2607.04367v1 Announce Type: new Abstract: In the development of cooking robots, mastering the task of cutting is crucial. A significant challenge lies in the diverse properties of food, which ne

A Policy Decomposition Framework for Dynamic Order Fulfillment Operations

SafetyDGX agent

arXiv:2607.04056v1 Announce Type: cross Abstract: Modern supply chains span diverse operational environments, ranging from e-commerce distribution networks to customized production-to-order manufactur

A Precedent-Guided Co-Scientist for Side-Effect-Aware Drug Redesign

SafetyDGX agent

arXiv:2607.02944v1 Announce Type: cross Abstract: We propose PRECEDE, a precedent-guided co-scientist for side-effect-aware drug redesign that revises a parent compound to mitigate a specified side ef

A Unified Causal-Origin Taxonomy of Distributional Shifts in Reinforcement Learning

SafetyDGX agent

arXiv:2606.16933v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) systems often degrade when operating conditions differ from those previously encountered, reflecting distributiona

A User-driven Design Framework for Robotaxi

SafetyDGX agent

arXiv:2602.19107v3 Announce Type: replace Abstract: Robotaxis are emerging as a promising form of urban mobility, but removing human drivers fundamentally reshapes passenger-vehicle interaction and ra

ACE: Agentic Control for Embodied Manipulation via Zero-shot Workflow Reasoning

SafetyDGX agent

arXiv:2607.04162v1 Announce Type: cross Abstract: Open-ended tabletop manipulation requires agents to not only understand natural language but also adapt to dynamic environments and execution failures

ACPO: Adaptive Credit Policy Optimization via Fine-Grained Surrogate Entropy

SafetyDGX agent

arXiv:2607.03126v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has substantially improved the reasoning ability of large language models (LLMs), but sparse outcome rewards still make to

Adaptive Entropy-Driven Sensor Selection in a Camera-LiDAR Particle Filter for Single-Vessel Tracking

SafetyDGX agent

arXiv:2603.08457v2 Announce Type: replace-cross Abstract: Robust single-vessel tracking from fixed coastal platforms is hindered by modality-specific degradations: cameras suffer from illumination and

Adaptive Inference Batching using Policy Gradients

SafetyDGX agent

arXiv:2607.05272v1 Announce Type: cross Abstract: Inference serving systems must balance throughput and latency under bursty, heterogeneous workloads, yet the industry standard remains static batching

Adaptive Margin RLHF via Preference over Preferences

SafetyDGX agent

arXiv:2509.22851v4 Announce Type: replace-cross Abstract: Margin-based optimization is fundamental to improving generalization and robustness in classification tasks. In the context of reward model le

Adaptive Partitioning and Learning for Stochastic Control of Diffusion Processes

SafetyDGX agent

arXiv:2512.14991v2 Announce Type: replace Abstract: We study reinforcement learning for controlled diffusion processes with unbounded continuous state spaces, bounded continuous actions, and polynomia

Additive Causal Construction for Transferable and Reconfigurable Cross-System Learning in Multi-Source Image Fusion

SafetyDGX agent

arXiv:2607.02572v1 Announce Type: cross Abstract: In multi-source image fusion scenarios, heterogeneous inputs are typically driven by distinct generative mechanisms and can be viewed as a composition

ADP: Adversarial Dynamics Priors for Physically Grounded Humanoid Locomotion

SafetyDGX agent

arXiv:2607.03454v1 Announce Type: cross Abstract: In this paper, we propose Adversarial Dynamics Priors (ADP) for perturbation-resilient humanoid locomotion control. Existing motion prior-based method

Agentic Artificial Intelligence for Multistage Physics Experiments at a Large-Scale User Facility Particle Accelerator

SafetyDGX agent

arXiv:2509.17255v2 Announce Type: replace-cross Abstract: We present the first language-model-driven agentic artificial intelligence (AI) system to autonomously execute multi-stage physics experiments

AGL-1: The Enterprise AI Governance Layer as a Control Plane for Trusted Enterprise Intelligence

SafetyDGX agent

arXiv:2607.03516v1 Announce Type: cross Abstract: Enterprise artificial intelligence is moving from isolated experimentation toward operational dependency across copilots, retrieval-augmented generati

Aligning Language Models with Selective Prediction

SafetyDGX agent

arXiv:2607.03528v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as critical decision-making components in high-stakes real-world AI systems, rendering LLM reli

Alignment-Guided Largest Table Overlap Size Estimation

SafetyDGX agent

arXiv:2607.03049v1 Announce Type: new Abstract: Fast estimation of the size of the largest overlap between tables enables blocking and query-by-table retrieval in large table repositories. The first a

AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis

SafetyDGX agent

arXiv:2512.11797v2 Announce Type: replace-cross Abstract: The collection of large-scale and diverse robot demonstrations remains a major bottleneck for imitation learning, as real-world data acquisiti

Anticipatory Reinforcement Learning for Trajectory Tracking

SafetyDGX agent

arXiv:2607.03132v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) in industrial control often suffers from lag and overshoot due to purely reactive control based on the current trackin

AquaStereo: Enabling Underwater Stereo Matching via Depth-Conditioned Diffusion and Geometry Self-Distillation

SafetyDGX agent

arXiv:2607.04303v1 Announce Type: new Abstract: Learning-based stereo matching models struggle in underwater environments due to scarce in-domain data and the difficulty of extracting discriminative c

← Previous
1…4041424344…212
Next →