AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
6 May 2026

Instance-Level Costs for Nuanced Classifier Evaluation

SafetyDGX agent

arXiv:2605.03135v1 Announce Type: new Abstract: Standard classification treats all errors equally, but in content moderation, medical screening, and safety-critical applications, mistakes on clear-cut

Intervention Complexity as a Canonical Reward and a Measure of Intelligence

SafetyDGX agent

arXiv:2605.02175v1 Announce Type: new Abstract: The Legg--Hutter universal intelligence measure provides a rigorous scalar assessment of general intelligence as expected reward across all computable e

just want to go back to how much intense pushback i got on this story at all levels of the company at the time, and how people speak very di…

SafetyDGX agent

just want to go back to how much intense pushback i got on this story at all levels of the company at the time, and how people speak very differently when under the threat of perjury lot of names etch


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Khala: Scaling Acoustic Token Language Models Toward High-Fidelity Music Generation

SafetyDGX agent

arXiv:2605.01790v1 Announce Type: cross Abstract: A common design pattern in high-quality music generation is to handle structure and fidelity in different representation spaces: a generator first mod

Large Language Models are Universal Reasoners for Visual Generation

SafetyDGX agent

arXiv:2605.04040v1 Announce Type: new Abstract: Text-to-image generation has advanced rapidly with diffusion models, progressing from CLIP and T5 conditioning to unified systems where a single LLM bac

Learning Reactive Dexterous Grasping via Hierarchical Task-Space RL Planning and Joint-Space QP Control

SafetyDGX agent

arXiv:2605.03363v1 Announce Type: new Abstract: In this work, we propose a hybrid hierarchical control framework for reactive dexterous grasping that explicitly decouples high-level spatial intent fro

Like tricksters, LLMs have perfected the art of plausibility, says Tim Harford: https://ft.trib.al/5Foo2YD

SafetyDGX agent

Tim Harford compares large language models to tricksters, arguing that LLMs excel at generating plausible-sounding text without necessarily ensuring accuracy or truthfulness. The article likely explor

LLM-XTM: Enhancing Cross-Lingual Topic Models with Large Language Models

SafetyDGX agent

arXiv:2605.03299v1 Announce Type: new Abstract: Cross-lingual topic modeling aims to discover shared semantic structures across languages, yet existing models depend on sparse bilingual resources and

Logic-Constrained Shortest Paths for Flight Planning

SafetyDGX agent

arXiv:2412.13235v4 Announce Type: replace Abstract: The logic-constrained shortest path problem (LCSPP) combines a one-to-one shortest path problem with satisfiability constraints imposed on the routi

MAGE: Safeguarding LLM Agents against Long-Horizon Threats via Shadow Memory

SafetyDGX agent

arXiv:2605.03228v1 Announce Type: cross Abstract: As large language model (LLM)-powered agents are increasingly deployed to perform complex, real-world tasks, they face a growing class of attacks that

Many trials feel like Rashomon with different witnesses. The amazing thing about Musk-OpenAI is how much agreement there has been (at least …

SafetyDGX agent

Many trials feel like Rashomon with different witnesses. The amazing thing about Musk-OpenAI is how much agreement there has been (at least so far) on the facts. The question is really whether what Op

Memorization In Stable Diffusion Is Unexpectedly Driven by CLIP Embeddings

SafetyDGX agent

arXiv:2605.02908v1 Announce Type: new Abstract: Understanding how textual embeddings contribute to memorization in text-to-image diffusion models is crucial for both interpretability and safety. This

MILD: Mediator Agent System with Bidirectional Perception and Multi-Layered Alignment for Human-Vehicle Collaboration

SafetyDGX agent

arXiv:2605.01507v1 Announce Type: new Abstract: Prior studies report that partial driving automation can increase the cognitive demands on human drivers. This effect largely arises from human drivers'

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation

SafetyDGX agent

arXiv:2602.05048v2 Announce Type: replace Abstract: Joint planning through language-based interactions is a key area of human-AI teaming. Planning problems in the open world often involve various aspe

Mira Murati tells the court that she couldn’t trust Sam Altman’s words

SafetyDGX agent

Mira Murati, OpenAI's former CTO, has testified under oath that CEO Sam Altman lied to her about the safety standards for a new AI model. In a video deposition shown during the ongoing Musk v. Altman

Mira Murati’s testimony is gripping – and what it makes absolutely clear is how utterly wrong most of Twitter was about why Sam was fired. –…

SafetyDGX agent

Mira Murati’s testimony is gripping – and what it makes absolutely clear is how utterly wrong most of Twitter was about why Sam was fired. – It had nothing per se to do with AI safety - It had nothing

Mix3R: Mixing Feed-forward Reconstruction and Generative 3D Priors for Joint Multi-view Aligned 3D Reconstruction and Pose Estimation

SafetyDGX agent

arXiv:2605.03359v1 Announce Type: new Abstract: Recent trends in sparse-view 3D reconstruction have taken two different paths: feed-forward reconstruction that predicts pixel-aligned point maps withou

Model Routing as a Trust Problem: Route Receipts for Adaptive AI Systems

SafetyDGX agent

arXiv:2605.01710v1 Announce Type: new Abstract: AI products often route requests through version aliases, service tiers, tool choices, regional endpoints, fallback rules, or safety handling before res

Model Spec Midtraining: Improving How Alignment Training Generalizes

SafetyDGX agent

arXiv:2605.02087v1 Announce Type: new Abstract: Some frontier AI developers aim to align language models to a Model Spec or Constitution that describes the intended model behavior. However, standard a

Multilingual Safety Alignment via Self-Distillation

SafetyDGX agent

arXiv:2605.02971v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit severe multilingual safety misalignment: they possess strong safeguards in high-resource languages but remain hig

Musk v. Altman: Mira Murati testifies that Sam Altman lied to her about the safety standards for a new OpenAI model and that he made her work more difficult (Jay Peters/The Verge)

SafetyDGX agent

Jay Peters / The Verge: Musk v. Altman: Mira Murati testifies that Sam Altman lied to her about the safety standards for a new OpenAI model and that he made her work more difficult — OpenAI's former C

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation

SafetyDGX agent

arXiv:2605.03058v1 Announce Type: new Abstract: A key goal of explainable AI (XAI) is to express the decision logic of large language models (LLMs) in symbolic form and link it to internal mechanisms.

NORA: A Harness-Engineered Autonomous Research Agent for End-to-End Spatial Data Science

SafetyDGX agent

arXiv:2605.02092v1 Announce Type: new Abstract: The automation of scientific research workflows has emerged as a transformative frontier in artificial intelligence, yet existing autonomous research ag

Nora: Normalized Orthogonal Row Alignment for Scalable Matrix Optimizer

SafetyDGX agent

arXiv:2605.03769v1 Announce Type: new Abstract: Matrix-based optimizers have demonstrated immense potential in training Large Language Models (LLMs), however, designing an ideal optimizer remains a fo

Normalized Matching Transformer

SafetyDGX agent

arXiv:2503.17715v3 Announce Type: replace Abstract: We introduce the Normalized Matching Transformer (NMT), a deep learning approach for efficient and accurate sparse semantic keypoint matching betwee

OGPO: Sample Efficient Full-Finetuning of Generative Control Policies

SafetyDGX agent

arXiv:2605.03065v1 Announce Type: new Abstract: Generative control policies (GCPs), such as diffusion- and flow-based control policies, have emerged as effective parameterizations for robot learning.

On Surprising Effects of Risk-Aware Domain Randomization for Contact-Rich Sampling-based Predictive Control

SafetyDGX agent

arXiv:2605.03290v1 Announce Type: new Abstract: Domain randomization (DR) is widely used in policy learning to improve robustness to modeling error, but remains underexplored in contact-rich sampling-

OpenAI violated Canadian privacy laws in developing first ChatGPT model, probe finds https://www.theglobeandmail.com/business/article-openai…

SafetyDGX agent

OpenAI violated Canadian privacy laws in developing first ChatGPT model, probe finds https://www.theglobeandmail.com/business/article-openai-chatgpt-violated-canadian-privacy-laws-watchdogs-report/?ut

Optimal Posterior Sampling for Policy Identification in Tabular Markov Decision Processes

SafetyDGX agent

arXiv:2605.03921v1 Announce Type: new Abstract: We study the (arepsilon, elta)-PAC policy identification problem in finite-horizon episodic Markov Decision Processes. Existing approaches provide finit

Orientation-Aware Unsupervised Domain Adaptation for Brain Tumor Classification Across Multi-Modal MRI

SafetyDGX agent

arXiv:2605.03490v1 Announce Type: new Abstract: The clinical integration of deep learning models for brain tumor diagnosis in neuro-oncology is severely constrained by limited expert-annotated MRI dat

Poly-EPO: Training Exploratory Reasoning Models

SafetyDGX agent

arXiv:2604.17654v3 Announce Type: replace Abstract: Exploration is a cornerstone of learning from experience: it enables agents to find solutions to complex problems, generalize to novel ones, and sca

Population-Aware Imitation Learning in Mean-field Games with Common Noise

SafetyDGX agent

arXiv:2605.03357v1 Announce Type: new Abstract: Mean Field Games (MFGs) provide a powerful framework for modeling the collective behavior of large populations of interacting agents. In this paper, we

Position: Safety and Fairness in Agentic AI Depend on Interaction Topology, Not on Model Scale or Alignment

SafetyDGX agent

arXiv:2605.01147v1 Announce Type: new Abstract: As large language models are increasingly deployed as interacting agents in high-stakes decisions, the AI safety community assumes that safety propertie

Power-Softmax: Towards Secure LLM Inference over Encrypted Data

SafetyDGX agent

arXiv:2410.09457v2 Announce Type: replace Abstract: Modern cryptographic methods for implementing privacy-preserving LLMs such as gls{HE} require the LLMs to have a polynomial form. Forming such a rep

Predicting missing values: A good idea?

SafetyDGX agent

arXiv:2605.03733v1 Announce Type: cross Abstract: Minimizing the Mean Squared Error (MSE) is a key objective in machine learning and is commonly used for imputing missing values. While this approach p

Privacy Preserving Machine Learning Workflow: from Anonymization to Personalized Differential Privacy Budgets in Federated Learning

SafetyDGX agent

arXiv:2605.02372v1 Announce Type: cross Abstract: The growing development of artificial intelligence based solutions, together with privacy legislation, has driven the rise of the so-called privacy pr

Pseudo-differential-enhanced physics-informed neural networks

SafetyDGX agent

arXiv:2602.14663v2 Announce Type: replace Abstract: We present pseudo-differential enhanced physics-informed neural networks (PINNs), an extension of gradient enhancement but in Fourier space. Gradien

Reinforcement Learning Trained Observer Control for Bearings-Only Tracking

SafetyDGX agent

arXiv:2605.02120v1 Announce Type: new Abstract: This paper develops a deep reinforcement learning based observer control policy for autonomous bearings-only tracking of a moving target. The observer m

Resource-Efficient Reinforcement for Reasoning Large Language Models via Dynamic One-Shot Policy Refinement

SafetyDGX agent

arXiv:2602.00815v2 Announce Type: replace Abstract: Large language models (LLMs) have exhibited remarkable performance on complex reasoning tasks, with reinforcement learning under verifiable rewards

Rethinking the Rank Threshold for LoRA Fine-Tuning

SafetyDGX agent

arXiv:2605.03724v1 Announce Type: new Abstract: A recent landscape analysis of LoRA fine-tuning in the neural tangent kernel regime establishes a sufficient condition r(r+1)/2 > KN on the LoRA rank r

RLDX-1 Technical Report

SafetyDGX agent

arXiv:2605.03269v1 Announce Type: cross Abstract: While Vision-Language-Action models (VLAs) have shown remarkable progress toward human-like generalist robotic policies through the versatile intellig

Safety-critical Control Under Partial Observability: Reach-Avoid POMDP meets Belief Space Control

SafetyDGX agent

arXiv:2603.10572v2 Announce Type: replace Abstract: Partially Observable Markov Decision Processes (POMDPs) provide a principled framework for robot decision-making under uncertainty. Solving reach-av

Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses

SafetyDGX agent

arXiv:2605.02900v1 Announce Type: cross Abstract: Embodied Artificial Intelligence (Embodied AI) integrates perception, cognition, planning, and interaction into agents that operate in open-world, saf

Sample-Efficient Optimization over Generative Priors via Coarse Learnability

SafetyDGX agent

arXiv:2503.06917v5 Announce Type: replace Abstract: We study zeroth-order optimization where solutions must minimize a cost d(s) while maintaining high probability under a complex generative prior L(s

SCION: Size-aware Policy Orchestration for Nonstationary Object Caches (Long Paper Version)

SafetyDGX agent

arXiv:2605.01055v1 Announce Type: cross Abstract: Object caches underpin cloud and edge services, but production workloads are heterogeneous, nonstationary, and throughput-constrained. Recent simple n

SERE: Structural Example Retrieval for Enhancing LLMs in Event Causality Identification

SafetyDGX agent

arXiv:2605.03701v1 Announce Type: new Abstract: Event Causality Identification (ECI) requires models to determine whether a given pair of events in a context exhibits a causal relationship. While Larg

Set-Based Training of Neural Barrier Certificates for Safety Verification of Dynamical Systems

SafetyDGX agent

arXiv:2605.02526v1 Announce Type: cross Abstract: Barrier certificates are scalar functions over the state space of dynamical systems that separate all unsafe states from all reachable states. The exi

SigLoMa: Learning Open-World Quadrupedal Loco-Manipulation from Ego-Centric Vision

SafetyDGX agent

arXiv:2605.03846v1 Announce Type: new Abstract: Designing an open-world quadrupedal loco-manipulation system is highly challenging. Traditional reinforcement learning frameworks utilizing exteroceptio

SMoE: An Algorithm-System Co-Design for Pushing MoE to the Edge via Expert Substitution

SafetyDGX agent

arXiv:2508.18983v3 Announce Type: replace Abstract: The Mixture of Experts (MoE) architecture has emerged as a key technique for scaling Large Language Models by activating only a subset of experts pe

SoDa2: Single-Stage Open-Set Domain Adaptation via Decoupled Alignment for Cross-Scene Hyperspectral Image Classification

SafetyDGX agent

arXiv:2605.03371v1 Announce Type: new Abstract: Cross-scene hyperspectral image (HSI) classification stands as a fundamental research topic in remote sensing, with extensive applications spanning vari

Steerable Adversarial Scenario Generation through Test-Time Preference Alignment

SafetyDGX agent

arXiv:2509.20102v2 Announce Type: cross Abstract: Adversarial scenario generation is a cost-effective approach for safety assessment of autonomous driving systems. However, existing methods are often

Stream-R1: Reliability-Perplexity Aware Reward Distillation for Streaming Video Generation

SafetyDGX agent

arXiv:2605.03849v1 Announce Type: new Abstract: Distillation-based acceleration has become foundational for making autoregressive streaming video diffusion models practical, with distribution matching

Structured Diffusion Bridges: Inductive Bias for Denoising Diffusion Bridges

SafetyDGX agent

arXiv:2605.02973v1 Announce Type: new Abstract: Modality translation is inherently under-constrained, as multiple cross-modal mappings may yield the same marginals. Recent work has shown that diffusio

T^2PO: Uncertainty-Guided Exploration Control for Stable Multi-Turn Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2605.02178v1 Announce Type: new Abstract: Recent progress in multi-turn reinforcement learning (RL) has significantly improved reasoning LLMs' performances on complex interactive tasks. Despite

Talk is Cheap, Communication is Hard: Dynamic Grounding Failures and Repair in Multi-Agent Negotiation

SafetyDGX agent

arXiv:2605.01750v1 Announce Type: cross Abstract: Grounding is the collaborative process of establishing mutual belief sufficient for the current communicative purpose. While static grounding maps lan

TeamUp: Semantic Project Matching and Team Formation for Learning at Scale

SafetyDGX agent

arXiv:2605.03237v1 Announce Type: cross Abstract: Project-based learning improves student engagement and learning outcomes, yet allocating students to appropriately challenging projects while forming

The AI risk repository: A meta-review, database, and taxonomy of risks from artificial intelligence

SafetyDGX agent

arXiv:2408.12622v3 Announce Type: replace-cross Abstract: Artificial intelligence (AI) is reshaping society, from video generation to medical diagnosis, coding agents to autonomous vehicles. Yet resea

The Design and Composition of Structural Causal Decision Processes

SafetyDGX agent

arXiv:2605.02681v1 Announce Type: cross Abstract: We present two new classes of causal models of decision-making agents. Our approach is motivated by the needs of modeling the economics of computing s

The Garden of Forking Paths: Narrative Arc-Conditioned Gameplay Planning

SafetyDGX agent

arXiv:2605.01245v1 Announce Type: cross Abstract: Narrative archetypes (e.g., Hero's Journey, Three-act structure) provide universal story structures that resonate across cultures and media and are im

Towards Safer Large Reasoning Models by Promoting Safety Decision-Making before Chain-of-Thought Generation

SafetyDGX agent

arXiv:2603.17368v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieved remarkable performance via chain-of-thought (CoT), but recent studies showed that such enhanced reasoning cap

← Previous
1…162163164165166…212
Next →