AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
7 Jul 2026

ARLArena: A Unified Framework for Stable Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2602.21534v3 Announce Type: replace Abstract: Agentic reinforcement learning (ARL) has rapidly gained attention as a promising paradigm for training agents to solve complex, multi-step interacti

As the @UN’s Global Dialogue on AI Governance wraps up today, I’ve been encouraged by the discussions surrounding AI and its implications fo…

SafetyDGX agent

As the @UN’s Global Dialogue on AI Governance wraps up today, I’ve been encouraged by the discussions surrounding AI and its implications for global collaboration and policymaking. I’m hopeful that we

ASK in the Dark: Uncertainty-Gated LLM Assistance under Partial Observability

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.02686v1 Announce Type: new Abstract: Reinforcement learning agents operating under partial observability must act on incomplete information, making them natural candidates for guidance from

At UMA, we build and own the full stack, from hardware to software. This allows us to bake safety in at all levels, rather than bolt it on a…

SafetyDGX agent

At UMA, we build and own the full stack, from hardware to software. This allows us to bake safety in at all levels, rather than bolt it on as an afterthought. Because trust is everything when robots s

Athena-WBC: Capability-Aligned Policy Experts for Long-Tail Humanoid Whole-Body Control

SafetyDGX agent

arXiv:2607.04837v1 Announce Type: new Abstract: Large-scale humanoid motion-tracking controllers are commonly improved by reallocating training effort: difficult motions are sampled more often, isolat

Attention Limited Reward Learning

SafetyDGX agent

arXiv:2607.04590v1 Announce Type: new Abstract: Pairwise human comparisons are a primary interface through which modern AI systems learn human preferences. RLHF and related alignment pipelines typical

Attributing Emergence in Million-Agent Systems

SafetyDGX agent

arXiv:2605.11404v2 Announce Type: replace Abstract: Large language models (LLMs) can simulate human-like reasoning and decision-making in individual agents. LLM-powered multi-agent systems (MAS) combi

Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment

SafetyDGX agent

arXiv:2607.04311v1 Announce Type: new Abstract: Subject-driven and multi-element video generation are central to controllable video synthesis, but existing methods still struggle to preserve identity

Autonomous Information Seeking: A Roadmap for Agentic Recommender Systems

SafetyDGX agent

arXiv:2607.04433v1 Announce Type: cross Abstract: The rapid integration of large language model-based agents into recommender systems has driven a shift from static, ranking-based pipelines toward aut

AViS-Mamba: Adaptive Visual Steering of Audio State-Space Dynamics for Violence Detection

SafetyDGX agent

arXiv:2604.03329v2 Announce Type: replace-cross Abstract: Automatic violence detection from video is challenging because violent interactions may be distant, occluded, or only partially visible. Audio

Benign Overfitting Does Not Occur in Diffusion Models

SafetyDGX agent

arXiv:2607.02671v1 Announce Type: cross Abstract: Benign overfitting and double descent have come to shape our understanding of generalization in deep learning, establishing that overfitting is not on

Best-of-Better-N: Generating Pre-Aligned Responses with In-Context Learning

SafetyDGX agent

arXiv:2607.03453v1 Announce Type: cross Abstract: Inference-time alignment methods, such as Best-of-N, offer a flexible alternative to training-based alignment by using reward models to select high-qu

BEVLM: Distilling Semantic Knowledge from LLMs into Bird's-Eye View Representations

SafetyDGX agent

arXiv:2603.06576v2 Announce Type: replace-cross Abstract: The integration of Large Language Models (LLMs) into autonomous driving has attracted growing interest for their strong reasoning and semantic

Beyond Heuristics: A Standardized Real2Sim Pipeline for Physical Human Robot Interaction in Human-in-the-Loop Simulation

SafetyDGX agent

arXiv:2607.03017v1 Announce Type: new Abstract: The aging global population drives demand for assistive robots, yet the safety risks and costs of physical testing make Human-in-the-Loop (HITL) simulat

Beyond Independent Labels: Schwartz-Geometry Decoding for Human Value Detection

SafetyDGX agent

arXiv:2607.05052v1 Announce Type: cross Abstract: Human value detection is commonly formulated as sentence-level multi-label classification over the 19 refined Schwartz values, typically predicted as

Beyond Point-Attached Semantics: Object-Centric Semantic Fields for Generalizable Manipulation

SafetyDGX agent

arXiv:2607.03163v1 Announce Type: new Abstract: Generalizable robot manipulation requires stable 3D understanding of functional object parts, such as handles, tool heads, openings, and graspable regio

Beyond Random Sampling: Distribution-Aware Alignment for Semi-Supervised Medical Image Segmentation

SafetyDGX agent

arXiv:2607.04249v1 Announce Type: new Abstract: Precise medical image segmentation is crucial for clinical diagnosis and treatment planning, yet relies heavily on expensive expert annotations. Semi-su

BGP route policies: Top 3 use cases by customer demand

SafetyDGX agent

When we first made BGP route policies for Cloud Router generally available over a year ago, our goal was to give network administrators deep, programmable control over how network paths are evaluated

BiSLW: Bi-Spectral Latent Watermarking for Generative Diffusion Models

SafetyDGX agent

arXiv:2607.02643v1 Announce Type: new Abstract: Diffusion-based generative models have transformed visual content synthesis, yet they remain vulnerable to unauthorized usage and lack reliable attribut

Bootstrap Flow-Map Tree Sampling Enables Online Feedback Driven Search

SafetyDGX agent

arXiv:2607.02915v1 Announce Type: cross Abstract: In many scientific and engineering domains, maximizing discovery within a limited sampling budget demands strategic, observation-guided exploration. W

BoRP: Bootstrapped Regression Probing for Scalable and Human-Aligned LLM Evaluation

SafetyDGX agent

arXiv:2601.18253v2 Announce Type: replace-cross Abstract: Accurate evaluation of user satisfaction is critical for iterative development of conversational AI. However, for open-ended assistants, tradi

Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process

SafetyDGX agent

arXiv:2607.03748v1 Announce Type: new Abstract: Unified multi-modal models (UMMs) have shown promising interleaved text-image reasoning capabilities, yet effectively optimizing such multi-turn generat

CAGE-1: Control, Assurance, and Governance Evaluation for Enterprise Agentic AI

SafetyDGX agent

arXiv:2607.03510v1 Announce Type: cross Abstract: Enterprise artificial intelligence is moving from experimentation into operational workflows. Early programs focused on model access and retrieval-aug

Can temporal article-level credibility signals improve domain-level credibility prediction?

SafetyDGX agent

arXiv:2607.04560v1 Announce Type: new Abstract: Web domain credibility evaluation is vital for combating misinformation. It is conducted by examining factors such as domain type, transparency, and ove

CanniUplift: A Holistic Framework for Mitigating Seller and Incentive Cannibalization in E-commerce Uplift Modeling

SafetyDGX agent

arXiv:2607.05242v1 Announce Type: cross Abstract: Personalized incentive allocation is vital for e-commerce, where uplift modeling is the standard for estimating Individual Treatment Effects (ITE). Ho

Causal Mechanism Reduction: Mechanism Replacement for Neural Network Pruning and Abstraction

SafetyDGX agent

arXiv:2602.24266v2 Announce Type: replace-cross Abstract: Which internal mechanisms of a neural network can be replaced while preserving the computation it performs? Structured pruning asks for smalle

CCFM: Collision-Constrained Flow Matching for Safety-Critical Scenario Generation

SafetyDGX agent

arXiv:2607.04451v1 Announce Type: new Abstract: Evaluation of autonomous vehicle (AV) planners in safety-critical closed-loop simulation is essential for real-world deployment. However, generating con

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning

SafetyDGX agent

arXiv:2607.03903v1 Announce Type: new Abstract: Multi-task offline safe reinforcement learning (RL) promises to learn a shared optimal safe policy from offline data across multiple tasks. This paradig

Channel-Adaptive Robust Aggregation for Over-the-Air Federated Learning in Heterogeneous Networks

SafetyDGX agent

arXiv:2607.04218v1 Announce Type: new Abstract: The growing demand for privacy-preserving, data-intensive applications such as IoT, augmented reality, and autonomous systems positions Federated Learni

Claim-Level Rubric Rewards for Video Caption Reinforcement Learning

SafetyDGX agent

arXiv:2607.05150v1 Announce Type: new Abstract: In this paper, we introduce Claim-Level Rubric Rewards (CuRe), a structured reward framework designed to address the reward-design bottleneck in reinfor

CLEAR: Closed-Loop Reinforcement Learning at Scale for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2607.02841v1 Announce Type: cross Abstract: End-to-end autonomous driving (E2E-AD) aims to directly map raw sensor information to driving actions. Recently, with the rapid advancement of multi-m

Closed-loop vs. Open-loop Kalman Filter Architectures in Airborne Aided Inertial Navigation

SafetyDGX agent

arXiv:2607.03338v1 Announce Type: new Abstract: Closed-loop (or feedback) error-state Kalman filters with their relatives and offspring are the state-of-the-art in modern aided inertial navigation res

CN-CBF: Composite Neural Control Barrier Function for Robot Navigation in Dynamic Environments

SafetyDGX agent

arXiv:2603.06921v2 Announce Type: replace-cross Abstract: Safe navigation of autonomous robots remains one of the core challenges in the field, especially in dynamic and uncertain environments. One pr

CoDoL: Conditional Domain Prompt Learning for Out-of-Distribution Generalization

SafetyDGX agent

arXiv:2509.15330v2 Announce Type: replace Abstract: Recent advances in pre-training vision-language models (VLMs), e.g., contrastive language-image pre-training (CLIP) methods, have shown great potent

Cohort-attention Evaluation Metrics for Tied Data

SafetyDGX agent

arXiv:2503.12755v3 Announce Type: replace Abstract: Artificial intelligence (AI) has significantly improved medical screening accuracy, particularly in cancer detection and risk assessment. However, t

CONFLUX: A Latent Diusion Model for 3D Chest-CT Synthesis with RL Post-Training

SafetyDGX agent

arXiv:2607.02998v1 Announce Type: cross Abstract: Controllable generative models of 3D medical images can synthesize volumes with specified clinical attributes, but this demands samples that are simul

Context Misleads LLMs: The Role of Context Filtering in Maintaining Safe Alignment of LLMs

SafetyDGX agent

arXiv:2508.10031v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) have shown significant advancements in performance, various jailbreak attacks have posed growing safety and

CONTRA: Red-Teaming Configurations of Personalizable Agents

SafetyDGX agent

arXiv:2607.03220v1 Announce Type: cross Abstract: Recent tools such as OpenClaw have extended the capabilities of LLM-based agents from simple dialog-based systems to fully autonomous agents. These sy

CoRE-VLA: Towards Scalable and Robust Vision-Language-Action Modeling via Conditional Routing of Experts

SafetyDGX agent

arXiv:2607.03693v1 Announce Type: new Abstract: Vision-language-action (VLA) models have advanced generalist robotic manipulation, yet real-world deployment reveals a fundamental challenge: robots are

Counterfactual Methods for Detecting Unfairness in Anti-Money Laundering Algorithms

SafetyDGX agent

arXiv:2607.05101v1 Announce Type: new Abstract: The application of machine learning-based predictive algorithms to Anti-Money Laundering (AML) has grown rapidly, driven by the vast volume of financial

Court filing: Meta says four US states seek 1.4T over claims it designed Facebook and Instagram to addict youth and misled the public; its market cap is ~1.5T (Diana Novak Jones/Reuters)

SafetyDGX agent

Diana Novak Jones / Reuters: Court filing: Meta says four US states seek 1.4T over claims it designed Facebook and Instagram to addict youth and misled the public; its market cap is ~1.5T — Meta Platf

Coverage-Controlled Preference Mining from Noisy Claim Verification for Evidence-Grounded Generation

SafetyDGX agent

arXiv:2603.10494v2 Announce Type: replace Abstract: Evidence-grounded generation produces summaries whose claims should be supported by supplied evidence, but claim-level verifiers provide noisy feedb

Covert Trait Propagation Is Representation Alignment: Mechanistic Evidence from Hidden-Channel Distillation

SafetyDGX agent

arXiv:2607.04432v1 Announce Type: cross Abstract: A student model trained on pure uniform noise can still inherit its teacher's digit-classification ability, provided the two share initialization. Pre

CRODA-ST: Single-Target Cross-Receiver Open-Set Radio Fingerprint Recognition

SafetyDGX agent

arXiv:2607.02567v1 Announce Type: cross Abstract: Radio frequency fingerprint identification (RFFI) provides a physical-layer credential for Internet of Things devices, but open-set decisions become f

Cross-Embodiment Robot Manipulation via a Unified Hand Action Space

SafetyDGX agent

arXiv:2607.03570v1 Announce Type: new Abstract: Robot manipulation policies are typically tied to specific robotic hand embodiments, limiting the transfer of learned behaviors across platforms with di

CrossHallu: Do Hallucination Signals Generalize Across Languages and Domains in Large Language Model's Internals?

SafetyDGX agent

arXiv:2607.04029v1 Announce Type: new Abstract: Recent hallucination detection techniques in large language models (LLMs) focus on directly extracting features from a model's internal representations

CRRL: A Causality-Based Reinforcement Learning Framework for Autonomous System Recovery

SafetyDGX agent

arXiv:2607.03177v1 Announce Type: cross Abstract: Traditional reinforcement learning (RL) for recovery in autonomous systems lacks causal understanding and generalizes poorly to novel failure scenario

Current as Touch: Proprioceptive Contact Feedback for Compliant Dexterous Manipulation

SafetyDGX agent

arXiv:2607.03529v1 Announce Type: new Abstract: Compliance is essential for dexterous manipulation, yet existing solutions often rely on external tactile or force sensors that are costly, fragile, and

CV-DCLR: Causal-Visual Dynamic Label Refinement for Robust Zero-Shot Learning

SafetyDGX agent

arXiv:2607.02601v1 Announce Type: new Abstract: Zero-Shot Learning (ZSL) facilitates knowledge transfer via shared semantic spaces. However, a critical bottleneck in this paradigm is Semantic Entangle

Decentralized Aggregation of LLM Predictions via Wagering Mechanisms

SafetyDGX agent

arXiv:2607.04389v1 Announce Type: new Abstract: It is increasingly common to aggregate predictions from multiple LLMs, each with domain expertise or access to private tools and data, to improve collec

Decomposed Prompting Does Not Fix Knowledge Gaps, But Helps Models Say 'I Don't Know'

SafetyDGX agent

arXiv:2602.04853v2 Announce Type: replace Abstract: Large language models often struggle to recognize their knowledge limits in closed-book question answering, leading to confident hallucinations. Whi

Deep Learning for Dynamic Programming with Recursive Utility

SafetyDGX agent

arXiv:2607.04278v1 Announce Type: cross Abstract: We propose the first deep learning algorithm, the Certainty Equivalent Learning (CEL) algorithm, for solving high-dimensional discrete-time dynamic pr

Defending from GeoLocalization through Adversarial Road Trips

SafetyDGX agent

arXiv:2607.03277v1 Announce Type: new Abstract: Retrieval-based image geolocalization has emerged as a powerful technique for determining the location of a query image by matching it against a large,

DeGenseGS: Geometrically and Semantically Decoupled Surgical Scene Understanding in 4D Gaussian Splatting

SafetyDGX agent

arXiv:2607.04761v1 Announce Type: new Abstract: Real-time, text-promptable 4D reconstruction is indispensable for autonomous surgical interaction. Severe misalignment between semantic meaning and phys

Demonstrating Generalization Failures via Mixtures of Conditional Policies

SafetyDGX agent

arXiv:2607.03478v1 Announce Type: new Abstract: Post-training of frontier language models is conducted on curated task suites, and inevitably leaves a distribution shift between training and deploymen

Denoised Conformal Alignment for Reliable Selection of Conditional Average Treatment Effect Predictions

SafetyDGX agent

arXiv:2607.03161v1 Announce Type: cross Abstract: In selective deployment, practitioners act only on a model-chosen subset of individuals based on predicted conditional average treatment effects, but

Designing Touch for Trauma-Informed Social Robots: A Design Space for Direct and Indirect Actuation

SafetyDGX agent

arXiv:2607.04981v1 Announce Type: new Abstract: Touch is a fundamental communication modality in human-robot interaction and may support grounding, emotional regulation, and stress reduction in therap

Detecting Architectural Drift in Safety-Critical Firmware through Runtime Trace Analysis

SafetyDGX agent

arXiv:2607.03135v1 Announce Type: cross Abstract: Maintaining consistency between architectural design and runtime-observed behavior is challenging in long-lived safety-critical firmware. This paper p

Deterministic Envelopes for Tamed SGLD: Decoupling Stochastic Gradient Noise and Localizing Taming

SafetyDGX agent

arXiv:2606.05242v2 Announce Type: replace-cross Abstract: Stochastic gradient Langevin algorithms often use tamed denominators to stabilize superlinear drifts. This paper shows that when the denominat

DiCE-CIR: Direct Composition Learning for Efficient Zero-Shot Composed Image Retrieval

SafetyDGX agent

arXiv:2607.04665v1 Announce Type: new Abstract: Zero-shot composed image retrieval (ZS-CIR) aims to retrieve a target image from a multimodal query consisting of a reference image and an edit text des

← Previous
1…4142434445…212
Next →