AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,600 results
6 Aug 2026

Memorization in Large Language Models in Medicine: Prevalence, Characteristics, and Implications

SafetyDGX agent

arXiv:2509.08604v5 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated significant potential in medicine, with many studies adapting them through continued pre-traini

MGSB: Manifold Gated Signature Branch Pressure-Domain Baseline Architecture for Two-Phase Pipeline Flows Under Distributional Shift

SafetyDGX agent

arXiv:2608.04805v1 Announce Type: new Abstract: Leak detection models for multiphase pipelines often degrade when deployed under flow regimes that differ from training. Existing evaluations typically

Multi-Objective Ranking for Live-Streaming: Balancing Fresh and Delayed Signals with Segment-Aware Targeting

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.04455v1 Announce Type: cross Abstract: One of the most challenging problems entertainment live-streaming services face in recommendation systems is that user behaviors are sparse and delaye

Multicalibration Yields Better Matchings

SafetyDGX agent

arXiv:2511.11413v2 Announce Type: replace Abstract: Consider the problem of finding the best matching in a weighted graph where we only have access to predictions of the actual stochastic weights, bas

Multimodal Alignment Through Joint Kernel Entropic Gromov--Wasserstein Optimal Transport

SafetyDGX agent

arXiv:2608.04234v1 Announce Type: cross Abstract: We study the problem of aligning data from multiple modalities into a shared representation space, focusing on settings where strong pretrained unimod

muSync-GS: Physics-Synchronized Driving Video Synthesis for Weather and Geometric Road Hazards

SafetyDGX agent

arXiv:2608.04412v1 Announce Type: new Abstract: High-quality driving data are essential for autonomous-driving systems and generative world models. However, rare and safety-critical scenarios involvin

Non-Stationary Inventory Control with Lead Times

SafetyDGX agent

arXiv:2602.05799v2 Announce Type: replace-cross Abstract: We study non-stationary single-item, periodic-review inventory control problems in which the demand distribution is unknown and may change ove

Not All Redundant Tokens Are Alike: Analyzing Visual Token Pruning through Token Roles

SafetyDGX agent

arXiv:2608.04483v1 Announce Type: new Abstract: Vision-language models (VLMs) process an image as a sequence of visual tokens, which creates a substantial computational bottleneck during inference. Re

Not Every Divergence Should Be Suppressed: Counterfactual Recoverability in On-Policy Distillation

SafetyDGX agent

arXiv:2608.04408v1 Announce Type: cross Abstract: On-policy distillation (OPD) supervises student-visited trajectories, yet divergence-based rules cannot determine whether an erroneous prefix remains

NSF-HRPT: Neural Semantic Field meets Hierarchical Risk Perception Tree for Safety-Critical Scenario Assessment

SafetyDGX agent

arXiv:2608.04776v1 Announce Type: new Abstract: The ability to accurately assess and anticipate risks in safety-critical scenarios is crucial for autonomous driving systems. While existing research ha

ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance

SafetyDGX agent

arXiv:2608.04524v1 Announce Type: new Abstract: Synthetic generation of Cognitive Behavioral Therapy (CBT) sessions is challenged by two competing demands: adhering to strict therapeutic structure whi

OMG i wrote some of the original work on what is to be neurosymbolic in 2001 and dude who probably hasn’t read that work is trying to school…

SafetyDGX agent

OMG i wrote some of the original work on what is to be neurosymbolic in 2001 and dude who probably hasn’t read that work is trying to school me on the definition 🤦‍♂️ coding harness and tools calls ar

On July 24, the Black at Mila community hosted Prof. Vukosi Marivate @vukosi (University of Pretoria, co-founder of Masakhane, Lelapa AI, an…

SafetyDGX agent

On July 24, the Black at Mila community hosted Prof. Vukosi Marivate @vukosi (University of Pretoria, co-founder of Masakhane, Lelapa AI, and Deep Learning Indaba) for a powerful session on the future

OPD-V: Visual On-Policy Self-Distillation with Modality Balance

SafetyDGX agent

arXiv:2608.05131v1 Announce Type: cross Abstract: On-Policy Self-Distillation (OPSD) has become a standard post-training approach for improving visual reasoning in multimodal large language models (ML

Optimal Constrained sc-LTL Planning in MDPs via Switching Policies

SafetyDGX agent

arXiv:2608.05021v1 Announce Type: new Abstract: We study the synthesis of optimal policies for planning problems on Markov decision processes with both objectives and safety constraints specified in c

Optimizing What Policies Learn From: Recoverability-aware Rollout Intervention Learning

SafetyDGX agent

arXiv:2608.05080v1 Announce Type: cross Abstract: Critic-free group-based reinforcement learning has become a scalable approach for post-training large language models. However, most existing methods

OutLangSplat: 3D Language Gaussian Splatting for UAV Outdoor Scenes

SafetyDGX agent

arXiv:2608.04560v1 Announce Type: new Abstract: 3D Language Gaussian Splatting embeds open-vocabulary language features into 3D Gaussian Splatting, providing an efficient explicit representation for t

Overcoming Statistical Bias in Action-Controllable World Models

SafetyDGX agent

arXiv:2608.04653v1 Announce Type: new Abstract: Action-conditioned world models aim to predict how visual environments evolve under an agent's actions. Yet future frames are often highly predictable f

Poly-OPD: Heterogeneous Multi-Teacher On-Policy Distillation for Capability-Selectable Flow Models

SafetyDGX agent

arXiv:2608.04349v1 Announce Type: new Abstract: Leading open text-to-image models often carry complementary strengths: one may lead on preference-aligned aesthetics while another follows compositional

PRIMAL3: Pathfinding via Reinforcement and Imitation Multi-Agent Learning - Leveraging LaCAM3

SafetyDGX agent

arXiv:2608.04905v1 Announce Type: new Abstract: We present PRIMAL3, an ultra-large-scale learning-based framework for multi-agent pathfinding (MAPF) that integrates reinforcement learning, topology-aw

Privileged, but Biased: How PI-Conditioned Teachers Break Self-Distillation

SafetyDGX agent

arXiv:2608.04794v1 Announce Type: new Abstract: Self-distillation (SD) has emerged as a compute-efficient alternative to reinforcement learning with verifiable rewards: a self-teacher, conditioned on

Q-CueGraph: Query-Conditioned Visual Evidence Graphs for Multimodal Reasoning

SafetyDGX agent

arXiv:2608.04452v1 Announce Type: cross Abstract: High-resolution pixels and crop or zoom tools give multimodal large language models the ability to inspect an image, but they do not provide a reliabl

Real-time probabilistic tsunami forecasting via generative AI

SafetyDGX agent

arXiv:2608.04327v1 Announce Type: new Abstract: Explicit onshore tsunami inundation forecasting can improve public risk awareness, but deterministically predicted inundation boundaries under highly un

Reward Structure Shapes the Interaction Between Episodic Exploration and Neural Memory in Reinforcement Learning

SafetyDGX agent

arXiv:2608.05111v1 Announce Type: new Abstract: In partially observable reinforcement learning, agents face a dual bottleneck: they must explore to encounter rewarding states and retain that experienc

SafeCommit: Certifying When Memory-Grounded Agents May Safely Act

SafetyDGX agent

arXiv:2608.04289v1 Announce Type: new Abstract: Long-horizon agents increasingly use persistent memory and tools to take actions with external side effects. A central failure mode is premature commitm

SafeLand: Safe Autonomous Landing in Unknown Environments with Bayesian Semantic Mapping

SafetyDGX agent

arXiv:2603.17430v2 Announce Type: replace Abstract: Autonomous landing of uncrewed aerial vehicles (UAVs) in unknown, dynamic environments poses significant safety challenges, particularly near people

SCOPE: Field-of-View-Aware Path Planning in Unknown 3D Environments via Safety-Volume Certification

SafetyDGX agent

arXiv:2608.04420v1 Announce Type: new Abstract: Safe navigation with a body-mounted limited-field-of-view sensor requires the complete robot-inflated volume of an intended motion to be observed and ve

Sequential Kernel-based Conditional Independence Testing via Adaptive Betting

SafetyDGX agent

arXiv:2606.18993v2 Announce Type: replace-cross Abstract: Testing conditional independence is fundamental yet intrinsically difficult: without additional assumptions, Type I error control is impossibl

Social Pressure Breaks Majority Voting in LLM Safety Panels

SafetyDGX agent

arXiv:2608.04415v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to detect unsafe content. A common approach is to combine judgments from a panel of models to correct

SpecRoll: Fast-Slow Verifier-Feedback Adaptation for Speculative Reinforcement Learning Rollouts

SafetyDGX agent

arXiv:2608.04962v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training improves the reasoning capabilities of large language models, but autoregressive rollout generation remains

SpikingNav: Robust Embodied Navigation with Spiking Neural Policies

SafetyDGX agent

arXiv:2608.05078v1 Announce Type: new Abstract: Embodied navigation requires an agent to make sequential decisions from egocentric observations in a physical environment. Existing Artificial Neural Ne

Splat-Based Metal Artifact Reduction in Cone-Beam CT via Compact Attenuation Modeling

SafetyDGX agent

arXiv:2608.04764v1 Announce Type: new Abstract: X-ray computed tomography (CT) suffers from severe metal artifacts when high-attenuation objects such as dental fillings or orthopedic implants are pres

SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation

SafetyDGX agent

arXiv:2608.04419v1 Announce Type: cross Abstract: On-policy distillation (OPD) provides dense teacher supervision on student-generated trajectories, but standard reverse-KL training can assign insuffi

SSC: A Verifiable Structured Representation for Bimanual Manipulation Labelling

SafetyDGX agent

arXiv:2608.04425v1 Announce Type: new Abstract: Subtask labels decompose a long-horizon manipulation demonstration into shorter semantic segments for policy training and evaluation. Natural language d

STEP-OPD: Rethinking Output Targets and Internal Dynamics in On-Policy Distillation for Diffusion Models

SafetyDGX agent

arXiv:2608.04887v1 Announce Type: new Abstract: On-policy distillation (OPD) has become an effective approach for consolidating multiple task-specialized image generation models into a single student.

Structured LLM Reasoning for Zero-Shot Human--Robot Coordination Under Hidden Goals

SafetyDGX agent

arXiv:2608.04309v1 Announce Type: new Abstract: We present a structured large-language-model (LLM) architecture for zero-shot human--robot coordination in a cooperative construction task with private

Temporal Context Awareness: A Defense Framework Against Multi-turn Manipulation Attacks on Large Language Models

SafetyDGX agent

arXiv:2503.15560v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly vulnerable to sophisticated multi-turn manipulation attacks, where adversaries strategically build conte

The Fairness Collapse Phenomenon: Bias Amplification in Language Models Trained on Synthetic Data

SafetyDGX agent

arXiv:2608.04268v1 Announce Type: new Abstract: Generative models trained on artificially generated data have been shown to exhibit model collapse, resulting in significant performance degradation. As

the smartest young people in sf were working on agi/alignment 5-10y ago. they are working on bci today. naomi is a total star and is one of …

SafetyDGX agent

the smartest young people in sf were working on agi/alignment 5-10y ago. they are working on bci today. naomi is a total star and is one of an explosion of young talent into the bci space recently. i’

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation

SafetyDGX agent

arXiv:2608.04436v1 Announce Type: new Abstract: Text-to-image (T2I) models can produce visually compelling images, yet they remain limited on open-world tasks that require complex semantic understandi

Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes

SafetyDGX agent

arXiv:2608.05000v1 Announce Type: new Abstract: Vision offers a critical axis for advancing foundation models, driving a shift towards natively unified multimodal pretraining. Despite this momentum, t

Traceable LLM-Generated Hazard Scenarios for Operational Safety Analysis of Aviation Systems Using ASRS Reports

SafetyDGX agent

arXiv:2608.04697v1 Announce Type: new Abstract: Operational hazard analysis of aviation system operations must consider interactions among weather, ATC actions, airspace constraints, aircraft operatio

TriCLE: Tri-Modal Vision-Language Reasoning for Edge-Deployed Fine-Grained Clustering

SafetyDGX agent

arXiv:2608.04175v1 Announce Type: new Abstract: Edge platforms used for aerial observation must interpret aircraft imagery under limited memory, limited compute, and intermittent connectivity. This se

Unscented KalmanNet: a hybrid deep learning filter with calibrated posterior covariance for nonlinear state estimation

SafetyDGX agent

arXiv:2608.04201v1 Announce Type: new Abstract: State estimation for nonlinear dynamical systems is commonly performed with the Unscented Kalman filter (UKF), which propagates the state moments throug

VoxStruct3D: Structure-Leading Flow Matching for Voxel-Space 3D MRI Synthesis

SafetyDGX agent

arXiv:2608.04557v1 Announce Type: new Abstract: High-fidelity 3D MRI synthesis requires both globally coherent anatomy and fine-grained voxel-level detail. Although latent diffusion makes volumetric g

War in the Abstract: The Rise and Consequences of Militarized Language in Scientific Communication

SafetyDGX agent

arXiv:2606.23462v2 Announce Type: replace-cross Abstract: Scientists do not, by profession, wage war. Yet warfare's vocabulary consistently appears in their abstracts. To quantify the extent to which

When Memory Lies: An Empirical Study of Spatial Memory Staleness in VLM Agents

SafetyDGX agent

arXiv:2608.04574v1 Announce Type: new Abstract: Memory-augmented VLM agents act on persistent spatial knowledge, yet that knowledge silently goes stale as the environment changes. We ask what happens

When Modalities Remember: Continual Learning for Multimodal Knowledge Graphs

SafetyDGX agent

arXiv:2604.02778v2 Announce Type: replace Abstract: Real-world multimodal knowledge graphs (MMKGs) are dynamic, with new entities, relations, and multimodal knowledge emerging over time. Existing cont

YOLOv14:Unified Cross-Domain Real-Time Object Detectionwith Adaptive Multi-View Representation

SafetyDGX agent

arXiv:2608.04720v1 Announce Type: new Abstract: Real-time object detectors achieve remarkable accuracy under controlled conditions, yet degrade sharply on non-ideal inputs: fisheye distortion, game-re

YouTube-Occ: Learning Indoor 3D Semantic Occupancy Prediction from YouTube Videos

SafetyDGX agent

arXiv:2506.18266v2 Announce Type: replace Abstract: 3D semantic occupancy prediction is crucial for fine-grained scene understanding, yet its advancement in privacy-sensitive indoor environments is fu

5 Aug 2026

A Blind Spot in Alignment: Quantifying Biosecurity Risks in Large Language Models

SafetyDGX agent

arXiv:2608.02684v1 Announce Type: cross Abstract: Large Language Models (LLMs) are accelerating biological research, yet this same capability poses a critical biosecurity threat: models that assist in

A game theory for foundation models shows new paths to rational cooperation through similarity inference

SafetyDGX agent

arXiv:2608.03958v1 Announce Type: new Abstract: As autonomous agents powered by foundation models are increasingly integrated into social and economic systems, understanding the principles governing t

A Hierarchical Approach to Imitation Learning for Manipulation Tasks Requiring Time Varying Forces

SafetyDGX agent

arXiv:2608.03103v1 Announce Type: cross Abstract: Diffusion policies have shown strong performance in learning complex, multi-modal behaviors for robotic manipulation. However, their application to co

A Physics-Flavored Transformer Network for Parametrizing Contraction Dynamics of Engineered Skeletal Muscle Tissues

SafetyDGX agent

arXiv:2608.03927v1 Announce Type: new Abstract: Engineered Skeletal Muscle Tissues (ESMs) have become a key structure for biomedical disease modeling and pharmacological screening, yet their functiona

A Security-Oriented Lifecycle Model for Large Language Model Systems

SafetyDGX agent

arXiv:2608.03626v1 Announce Type: cross Abstract: Large language models are being integrated into critical infrastructure and enterprise workflows at unprecedented scale,yet the lifecycle frameworks g

A Wearable Stiffness-Rendering Haptic Device with a Honeycomb Jamming Mechanism for Bilateral Teleoperation

SafetyDGX agent

arXiv:2608.03002v1 Announce Type: new Abstract: This paper addresses the challenge of providing kinesthetic feedback in bilateral teleoperation by designing a wearable, lightweight (20 g), and compact

Accelerating Human-Aware Robot Trajectory Generation via Diffusion and Consistency Distillation

SafetyDGX agent

arXiv:2608.03159v1 Announce Type: new Abstract: This research proposes a constrained motion planning framework for robot manipulators in human-robot interaction (HRI). For a non-redundant manipulator

ADMITBench: A Safety-Governed Reference Framework for Evaluating the Admissibility of Industrial LLM Advisories

SafetyDGX agent

arXiv:2608.03866v1 Announce Type: new Abstract: This white paper presents ADMITBench, a reference framework for evaluating industrial LLM advisories at the level of the proposed action. The framework

Agentic Reinforcement Learning with Self-Distilled Reward Shaping

SafetyDGX agent

arXiv:2608.03223v1 Announce Type: cross Abstract: Agentic reinforcement learning enables LLM agents to learn through interaction, but sparse trajectory-level rewards reveal success without identifying

AI Alignment and Fiduciary Obligation

SafetyDGX agent

arXiv:2608.02660v1 Announce Type: cross Abstract: Advanced AI assistants engage users in extended interactions across a widening range of roles, including advice, decision support, collaboration, lear

← Previous
1…910111213…210
Next →