AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
24 Apr 2026

Language-Conditioned Safe Trajectory Generation for Spacecraft Rendezvous

SafetyDGX agent

arXiv:2512.09111v3 Announce Type: replace-cross Abstract: Reliable real-time trajectory generation is essential for future autonomous spacecraft. While recent progress in nonconvex guidance and contro

Lawmakers and lobbyists say the Trump administration has lobbied against legislation that would regulate AI in at least six Republican-led states (Amrith Ramkumar/Wall Street Journal)

SafetyDGX agent

Amrith Ramkumar / Wall Street Journal: Lawmakers and lobbyists say the Trump administration has lobbied against legislation that would regulate AI in at least six Republican-led states — ‘I am disappo

Learning Dynamic Representations and Policies from Multimodal Clinical Time-Series with Informative Missingness

Safety

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.21235v1 Announce Type: cross Abstract: Multimodal clinical records contain structured measurements and clinical notes recorded over time, offering rich temporal information about the evolut

Learning Physics from Pretrained Video Models: A Multimodal Continuous and Sequential World Interaction Models for Robotic Manipulation

SafetyDGX agent

arXiv:2603.00110v2 Announce Type: replace Abstract: The scarcity of large-scale robotic data has motivated the repurposing of foundation models from other modalities for policy learning. In this work,

Local Neighborhood Instability in Parametric Projections: Quantitative and Visual Analysis

SafetyDGX agent

arXiv:2604.21617v1 Announce Type: new Abstract: Parametric projections let analysts embed new points in real time, but input variations from measurement noise or data drift can produce unpredictable s

Locating acts of mechanistic reasoning in student team conversations with mechanistic machine learning

SafetyDGX agent

arXiv:2604.21870v1 Announce Type: cross Abstract: STEM education researchers are often interested in identifying moments of students' mechanistic reasoning for deeper analysis, but have limited capaci

Logic Jailbreak: Efficiently Unlocking LLM Safety Restrictions Through Formal Logical Expression

SafetyDGX agent

arXiv:2505.13527v3 Announce Type: replace-cross Abstract: Despite substantial advancements in aligning large language models (LLMs) with human values, current safety mechanisms remain susceptible to j

Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions

SafetyDGX agent

arXiv:2604.21871v1 Announce Type: new Abstract: Human moral judgment is context-dependent and modulated by interpersonal relationships. As large language models (LLMs) increasingly function as decisio

Measuring and Exploiting Contextual Bias in LLM-Assisted Security Code Review

SafetyDGX agent

arXiv:2603.18740v2 Announce Type: replace-cross Abstract: Automated Code Review (ACR) systems integrating Large Language Models (LLMs) are increasingly adopted in software development workflows, rangi

Mind the Prompt: Self-adaptive Generation of Task Plan Explanations via LLMs

SafetyDGX agent

arXiv:2604.21092v1 Announce Type: new Abstract: Integrating Large Language Models (LLMs) into complex software systems enables the generation of human-understandable explanations of opaque AI processe

Mitigating Lost in Multi-turn Conversation via Curriculum RL with Verifiable Accuracy and Abstention Rewards

SafetyDGX agent

arXiv:2510.18731v2 Announce Type: replace-cross Abstract: Large Language Models demonstrate strong capabilities in single-turn instruction following but suffer from Lost-in-Conversation (LiC), a degra

Modulating Cross-Modal Convergence with Single-Stimulus, Intra-Modal Dispersion

SafetyDGX agent

arXiv:2604.21836v1 Announce Type: cross Abstract: Neural networks exhibit a remarkable degree of representational convergence across diverse architectures, training objectives, and even data modalitie

Multimodal Protein Language Models for Enzyme Kinetic Parameters: From Substrate Recognition to Conformational Adaptation

SafetyDGX agent

arXiv:2603.12845v2 Announce Type: replace Abstract: Predicting enzyme kinetic parameters quantifies how efficiently an enzyme catalyzes a specific substrate under defined biochemical conditions. Canon

New episode of The Information Bottleneck is out, this time with @liuzhuang1234 (Princeton). We talked about ConvNeXt and whether architectu…

SafetyDGX agent

New episode of The Information Bottleneck is out, this time with @liuzhuang1234 (Princeton). We talked about ConvNeXt and whether architecture still matters; dataset bias and what 'good data' actually

Participation and Representation in Local Government Speech

SafetyDGX agent

arXiv:2604.21202v1 Announce Type: cross Abstract: Local government meetings are the most common formal channel through which residents speak directly with elected officials, contest policies, and shap

Preserving Decision Sovereignty in Military AI: A Trade-Secret-Safe Architectural Framework for Model Replaceability, Human Authority, and State Control

SafetyDGX agent

arXiv:2604.20867v1 Announce Type: cross Abstract: Recent events surrounding the relationship between frontier AI suppliers and national-security customers have made a structural problem newly visible:

Probabilistic Verification of Neural Networks via Efficient Probabilistic Hull Generation

SafetyDGX agent

arXiv:2604.21556v1 Announce Type: new Abstract: The problem of probabilistic verification of a neural network investigates the probability of satisfying the safe constraints in the output space when t

Quotient-Space Diffusion Models

SafetyDGX agent

arXiv:2604.21809v1 Announce Type: cross Abstract: Diffusion-based generative models have reformed generative AI, and have enabled new capabilities in the science domain, for example, generating 3D str

Ramen: Robust Test-Time Adaptation of Vision-Language Models with Active Sample Selection

SafetyDGX agent

arXiv:2604.21728v1 Announce Type: new Abstract: Pretrained vision-language models such as CLIP exhibit strong zero-shot generalization but remain sensitive to distribution shifts. Test-time adaptation

Refining Covariance Matrix Estimation in Stochastic Gradient Descent Through Bias Reduction

SafetyDGX agent

arXiv:2604.21203v1 Announce Type: cross Abstract: We study online inference and asymptotic covariance estimation for the stochastic gradient descent (SGD) algorithm. While classical methods (such as p

Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own

SafetyDGX agent

arXiv:2310.02635v5 Announce Type: replace-cross Abstract: Reinforcement learning (RL) is a promising approach for solving robotic manipulation tasks. However, it is challenging to apply the RL algorit

RELOOP: Recursive Retrieval with Multi-Hop Reasoner and Planners for Heterogeneous QA

SafetyDGX agent

arXiv:2510.20505v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) remains brittle on multi-step questions and heterogeneous evidence sources, trading accuracy against late

Rethinking Cross-Domain Evaluation for Face Forgery Detection with Semantic Fine-grained Alignment and Mixture-of-Experts

SafetyDGX agent

arXiv:2604.21478v1 Announce Type: new Abstract: Nowadays, visual data forgery detection plays an increasingly important role in social and economic security with the rapid development of generative mo

RIFT: Repurposing Negative Samples via Reward-Informed Fine-Tuning

SafetyDGX agent

arXiv:2601.09253v2 Announce Type: replace-cross Abstract: While Supervised Fine-Tuning (SFT) and Rejection Sampling Fine-Tuning (RFT) are standard for LLM alignment, they either rely on costly expert

Robustness Analysis of POMDP Policies to Observation Perturbations

SafetyDGX agent

arXiv:2604.21256v1 Announce Type: new Abstract: Policies for Partially Observable Markov Decision Processes (POMDPs) are often designed using a nominal system model. In practice, this model can deviat

RPG: Robust Policy Gating for Smooth Multi-Skill Transitions in Humanoid Fighting

SafetyDGX agent

arXiv:2604.21355v1 Announce Type: new Abstract: Humanoid robots have demonstrated impressive motor skills in a wide range of tasks, yet whole-body control for humanlike long-time, dynamic fighting rem

SafeMERGE: Preserving Safety Alignment in Fine-Tuned Large Language Models via Selective Layer-Wise Model Merging

SafetyDGX agent

arXiv:2503.17239v3 Announce Type: replace-cross Abstract: Fine-tuning large language models (LLMs) is a common practice to adapt generalist models to specialized domains. However, recent studies show

SafeRedirect: Defeating Internal Safety Collapse via Task-Completion Redirection in Frontier LLMs

SafetyDGX agent

arXiv:2604.20930v1 Announce Type: cross Abstract: Internal Safety Collapse (ISC) is a failure mode in which frontier LLMs, when executing legitimate professional tasks whose correct completion structu

Seeing Without Eyes: 4D Human-Scene Understanding from Wearable IMUs

SafetyDGX agent

arXiv:2604.21926v1 Announce Type: new Abstract: Understanding human activities and their surrounding environments typically relies on visual perception, yet cameras pose persistent challenges in priva

Self-Predictive Representation for Autonomous UAV Object-Goal Navigation

SafetyDGX agent

arXiv:2604.21130v1 Announce Type: new Abstract: Autonomous Unmanned Aerial Vehicles (UAVs) have revolutionized industries through their versatility with applications including aerial surveillance, sea

SemaPop: Semantic-Persona Conditioned and Controllable Population Synthesis

SafetyDGX agent

arXiv:2602.11569v2 Announce Type: replace Abstract: Population synthesis is essential for individual-level simulation in transport planning and socio-economic analysis, yet remains challenging due to

SGG-R^{rm 3}: From Next-Token Prediction to End-to-End Unbiased Scene Graph Generation

SafetyDGX agent

arXiv:2603.07961v3 Announce Type: replace Abstract: Scene Graph Generation (SGG) structures visual scenes as graphs of objects and their relations. While Multimodal Large Language Models (MLLMs) have

Strategic Polysemy in AI Discourse: A Philosophical Analysis of Language, Hype, and Power

SafetyDGX agent

arXiv:2604.21043v1 Announce Type: cross Abstract: This paper examines the strategic use of language in contemporary artificial intelligence (AI) discourse, focusing on the widespread adoption of metap

StyleVAR: Controllable Image Style Transfer via Visual Autoregressive Modeling

SafetyDGX agent

arXiv:2604.21052v1 Announce Type: cross Abstract: We build on the Visual Autoregressive Modeling (VAR) framework and formulate style transfer as conditional discrete sequence modeling in a learned lat

Supervised Learning Has a Necessary Geometric Blind Spot: Theory, Consequences, and Minimal Repair

SafetyDGX agent

arXiv:2604.21395v1 Announce Type: cross Abstract: We prove that empirical risk minimisation (ERM) imposes a necessary geometric constraint on learned representations: any encoder that minimises superv

Survey on Evaluation of LLM-based Agents

SafetyDGX agent

arXiv:2503.16416v2 Announce Type: replace Abstract: LLM-based agents represent a paradigm shift in AI, enabling autonomous systems to plan, reason, and use tools while interacting with dynamic environ

Task-specific Subnetwork Discovery in Reinforcement Learning for Autonomous Underwater Navigation

SafetyDGX agent

arXiv:2604.21640v1 Announce Type: cross Abstract: Autonomous underwater vehicles are required to perform multiple tasks adaptively and in an explainable manner under dynamic, uncertain conditions and

Tempered Sequential Monte Carlo for Trajectory and Policy Optimization with Differentiable Dynamics

SafetyDGX agent

arXiv:2604.21456v1 Announce Type: new Abstract: We propose a sampling-based framework for finite-horizon trajectory and policy optimization under differentiable dynamics by casting controller design a

Temporal Prototyping and Hierarchical Alignment for Unsupervised Video-based Visible-Infrared Person Re-Identification

SafetyDGX agent

arXiv:2604.21324v1 Announce Type: new Abstract: Visible-infrared person re-identification (VI-ReID) enables cross-modality identity matching for all-day surveillance, yet existing methods predominantl

The Economics of p(doom): Scenarios of Existential Risk and Economic Growth in the Age of Transformative AI

SafetyDGX agent

arXiv:2503.07341v2 Announce Type: replace-cross Abstract: Recent advances in artificial intelligence (AI) have led to a wide range of predictions about its long-term impact on humanity. A central focu

The Effect of Idea Elaboration on the Automatic Assessment of Idea Originality

SafetyDGX agent

arXiv:2604.20569v1 Announce Type: cross Abstract: Automatic systems are increasingly used to assess the originality of responses in creative tasks. They offer a potential solution to key limitations o

'This Wasn't Made for Me': Recentering User Experience and Emotional Impact in the Evaluation of ASR Bias

SafetyDGX agent

arXiv:2604.21148v1 Announce Type: new Abstract: Studies on bias in Automatic Speech Recognition (ASR) tend to focus on reporting error rates for speakers of underrepresented dialects, yet less researc

Time, Causality, and Observability Failures in Distributed AI Inference Systems

SafetyDGX agent

arXiv:2604.21361v1 Announce Type: new Abstract: Distributed AI inference pipelines rely heavily on timestamp-based observability to understand system behavior. This work demonstrates that even small c

Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build …

SafetyDGX agent

Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build upon them and use them to evaluate the monitorability of the

Towards a Systematic Risk Assessment of Deep Neural Network Limitations in Autonomous Driving Perception

SafetyDGX agent

arXiv:2604.20895v1 Announce Type: cross Abstract: Safety and security are essential for the admission and acceptance of automated and autonomous vehicles. Deep neural networks (DNNs) are widely used f

TraceScope: Interactive URL Triage via Decoupled Checklist Adjudication

SafetyDGX agent

arXiv:2604.21840v1 Announce Type: cross Abstract: Modern phishing campaigns increasingly evade snapshot-based URL classifiers using interaction gates (e.g., checkbox/slider challenges), delayed conten

Trust-SSL: Additive-Residual Selective Invariance for Robust Aerial Self-Supervised Learning

SafetyDGX agent

arXiv:2604.21349v1 Announce Type: cross Abstract: Self-supervised learning (SSL) is a standard approach for representation learning in aerial imagery. Existing methods enforce invariance between augme

Tumor-anchored deep feature random forests for out-of-distribution detection in lung cancer segmentation

SafetyDGX agent

arXiv:2512.08216v3 Announce Type: replace-cross Abstract: Accurate segmentation of lung tumors from 3D computed tomography (CT) scans is essential for automated treatment planning and response assessm

Unbiased Prevalence Estimation with Multicalibrated LLMs

SafetyDGX agent

arXiv:2604.21549v1 Announce Type: new Abstract: Estimating the prevalence of a category in a population using imperfect measurement devices (diagnostic tests, classifiers, or large language models) is

UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection

SafetyDGX agent

arXiv:2604.21904v1 Announce Type: new Abstract: In recent years, significant progress has been made in both image generation and generated image detection. Despite their rapid, yet largely independent

Value-Conflict Diagnostics Reveal Widespread Alignment Faking in Language Models

SafetyDGX agent

arXiv:2604.20995v1 Announce Type: new Abstract: Alignment faking, where a model behaves aligned with developer policy when monitored but reverts to its own preferences when unobserved, is a concerning

VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models

SafetyDGX agent

arXiv:2510.18457v3 Announce Type: replace Abstract: The performance of Latent Diffusion Models (LDMs) is critically dependent on the quality of their visual tokenizers. While recent works have explore

VFM^{4}SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection

SafetyDGX agent

arXiv:2604.21502v1 Announce Type: new Abstract: In real-world scenarios, continual changes in weather, illumination, and imaging conditions cause significant domain shifts, leading detectors trained o

VLA-Forget: Vision-Language-Action Unlearning for Embodied Foundation Models

SafetyDGX agent

arXiv:2604.03956v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models are emerging as embodied foundation models for robotic manipulation, but their deployment introduces a new

When Bigger Isn't Better: A Comprehensive Fairness Evaluation of Political Bias in Multi-News Summarisation

SafetyDGX agent

arXiv:2604.21309v1 Announce Type: new Abstract: Multi-document news summarisation systems are increasingly adopted for their convenience in processing vast daily news content, making fairness across d

Who Defines Fairness? Target-Based Prompting for Demographic Representation in Generative Models

SafetyDGX agent

arXiv:2604.21036v1 Announce Type: new Abstract: Text-to-image(T2I) models like Stable Diffusion and DALL-E have made generative AI widely accessible, yet recent studies reveal that these systems often

Why are all LLMs Obsessed with Japanese Culture? On the Hidden Cultural and Regional Biases of LLMs

SafetyDGX agent

arXiv:2604.21751v1 Announce Type: cross Abstract: LLMs have been showing limitations when it comes to cultural coverage and competence, and in some cases show regional biases such as amplifying Wester

Why Do Language Model Agents Whistleblow?

SafetyDGX agent

arXiv:2511.17085v3 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) as tool-using agents causes their alignment training to manifest in new ways. Recent work finds

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

SafetyDGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

23 Apr 2026

A Hough transform approach to safety-aware scalar field mapping using Gaussian Processes

SafetyDGX agent

arXiv:2604.20799v1 Announce Type: new Abstract: This paper presents a framework for mapping unknown scalar fields using a sensor-equipped autonomous robot operating in unsafe environments. The unsafe

← Previous
1…181182183184185…212
Next →