AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
7 May 2026

Efficiency of Parallel and Restart Exploration Strategies in Model Free Stochastic Simulations

SafetyDGX agent

arXiv:2503.03565v3 Announce Type: replace-cross Abstract: We analyze the efficiency of parallelization and restart mechanisms for stochastic simulations in model-free settings, where the underlying sy

Efficient Geometry-Controlled High-Resolution Satellite Image Synthesis

SafetyDGX agent

arXiv:2605.04557v1 Announce Type: new Abstract: High-resolution satellite images are often scarce and costly, especially for remote areas or infrequent events. This shortage hampers the development an

Efficient Model-Based Reinforcement Learning for Robot Control via Online Optimization

SafetyDGX agent

arXiv:2510.18518v2 Announce Type: replace Abstract: We present an online model-based reinforcement learning algorithm suitable for controlling complex robotic systems directly in the real world. Unlik


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Efficiently Aligning Language Models with Online Natural Language Feedback

SafetyDGX agent

arXiv:2605.04356v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has been used to elicit impressive performance from language models in many domains. But, broadly benefic

Elicitation Matters: How Prompts and Query Protocols Shape LLM Surrogates under Sparse Observations

SafetyDGX agent

arXiv:2605.04764v1 Announce Type: new Abstract: Large language models are increasingly used as surrogate models for low-data optimization, but their optimizer-facing prediction and its uncertainty rem

Elon Musk’s DOGE “blatantly used” race, gender and other protected characteristics to execute the largest mass termination of federal grants…

SafetyDGX agent

Elon Musk’s DOGE “blatantly used” race, gender and other protected characteristics to execute the largest mass termination of federal grants in the history of the National Endowment for the Humanities

Encoding Predictability and Legibility for Style-Conditioned Diffusion Policy

SafetyDGX agent

arXiv:2603.16368v2 Announce Type: replace-cross Abstract: Striking a balance between efficiency and transparent motion is a core challenge in human-robot collaboration, as highly expressive movements

Enhancing the interpretability of spatially variable N2O model predictions with soft sensors during wastewater treatment

SafetyDGX agent

arXiv:2605.04082v1 Announce Type: new Abstract: Model-based solutions for nitrous oxide (N2O) emissions from wastewater treatment plants (WWTP) are informed by operational datasets designed to control

EO-1 is now available directly in LeRobot! You can train, evaluate, and deploy EO-1 through the standard LeRobot policy interface, making it…

SafetyDGX agent

EO-1 is now available directly in LeRobot! You can train, evaluate, and deploy EO-1 through the standard LeRobot policy interface, making it much easier to try EO-1 on robot-control workflows. Huge th

EP-GRPO: Entropy-Progress Aligned Group Relative Policy Optimization with Implicit Process Guidance

SafetyDGX agent

arXiv:2605.04960v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR), particularly Group Relative Policy Optimization (GRPO), has advanced LLM reasoning. However, GRPO

Evaluating Patient Safety Risks in Generative AI: Development and Validation of a FMECA Framework for Generated Clinical Content

SafetyDGX agent

arXiv:2605.04085v1 Announce Type: cross Abstract: Objectives: Large language models (LLMs) are increasingly used for clinical text summarization, yet structured methods to assess associated patient sa

Evaluating Semantic Fragility in Text-to-Audio Generation Systems Under Controlled Prompt Perturbations

SafetyDGX agent

arXiv:2603.13824v2 Announce Type: replace-cross Abstract: Recent advances in text-to-audio generation enable models to translate natural-language descriptions into diverse musical output. However, the

Every Step Counts: Step-Level Credit Assignment for Tool-Integrated Text-to-SQL

SafetyDGX agent

arXiv:2605.04719v1 Announce Type: new Abstract: Tool-integrated Text-to-SQL parsing has emerged as a promising paradigm, framing SQL generation as a sequential decision-making process interleaved with

Extending Differential Temporal Difference Methods for Episodic Problems

SafetyDGX agent

arXiv:2605.04368v1 Announce Type: new Abstract: Differential temporal difference (TD) methods are value-based reinforcement learning algorithms that have been proposed for infinite-horizon problems. T

FairEnc: A Fair Vision-Language Model with Fair Vision and Text Encoders for Glaucoma Detection

SafetyDGX agent

arXiv:2605.04882v1 Announce Type: new Abstract: Automated glaucoma detection is critical for preventing irreversible vision loss and reducing the burden on healthcare systems. However, ensuring fairne

Fascinating: People who get advice from chatbots often take that advice – but rarely end up feeling any better.

SafetyDGX agent

Research indicates that people frequently follow advice provided by chatbots, yet this compliance with AI recommendations does not correlate with improved outcomes or subjective well-being. This sugge

First person to update this chart with no mistakes using an AI system and a general prompt (rather than one than that handholds the system e…

SafetyDGX agent

First person to update this chart with no mistakes using an AI system and a general prompt (rather than one than that handholds the system every step of the win) wins a signed book! Please suppy your

Fixed-Length Dense Fingerprint Representation with Alignment and Robust Enhancement

SafetyDGX agent

arXiv:2505.03597v2 Announce Type: replace Abstract: Fixed-length fingerprint representations, which map each fingerprint to a compact and fixed-size feature vector, are computationally efficient and w

FLUID: Continuous-Time Hyperconnected Sparse Transformer for Sink-Free Learning

SafetyDGX agent

arXiv:2605.04421v1 Announce Type: new Abstract: Continuous-time (CT) Transformers improve irregular and long-range modeling over CT-RNNs by exploiting inputs or outputs embeddings with continuous dyna

Former OpenAI employee Rosie Campbell is the FIRST PERSON in the trial to point out that OpenAI's mission does NOT REQUIRE building AGI Unde…

SafetyDGX agent

Former OpenAI employee Rosie Campbell is the FIRST PERSON in the trial to point out that OpenAI's mission does NOT REQUIRE building AGI Undermines the entire justification for creating the for-profit

From Language to Logic: A Theoretical Architecture for VLM-Grounded Safe Navigation

SafetyDGX agent

arXiv:2605.04327v1 Announce Type: new Abstract: We propose an architecture for integrating high-level, human-provided safety rules and operator-aligned semantic preferences into autonomous robot navig

From Reach to Insert: Tactile-Augmented Precision Assembly under Sub-Millimeter Tolerances

SafetyDGX agent

arXiv:2605.04649v1 Announce Type: new Abstract: High-precision assembly frequently involves tight-tolerance insertions, where even slight pose errors can cause jamming or excessive interaction forces,

@GaryMarcus is on fire lately... follow him for #AI

SafetyDGX agent

Gary Marcus is an AI researcher and public intellectual who frequently shares commentary and insights about artificial intelligence developments on social media. His posts on X (formerly Twitter) cove

Globally Solving Unbalanced Optimal Transport and Density Control for Gaussian Distributions

SafetyDGX agent

arXiv:2605.04246v1 Announce Type: cross Abstract: In this article, we study unbalanced optimal transport (UOT) and establish a control-theoretic dynamical extension, which we call the unbalanced densi

Graph-Augmented LLMs for Swiss MP Ideology Prediction

SafetyDGX agent

arXiv:2605.04643v1 Announce Type: new Abstract: Approximating the ideological position of Members of Parliament (MPs) is a fundamental task in political science, helping researchers understand legisla

Graph-SND: Sparse Aggregation for Behavioral Diversity in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.05020v1 Announce Type: new Abstract: System Neural Diversity (SND) measures behavioral heterogeneity in multi-agent reinforcement learning by averaging pairwise distances over all inom{n}{2

Helen Toner's deposition in Musk v Altman includes some striking quotes about Mira Murati's involvement with Altman's ouster. She said Mira …

SafetyDGX agent

Helen Toner's deposition in Musk v Altman includes some striking quotes about Mira Murati's involvement with Altman's ouster. She said Mira was 'totally uninterested in telling her team that her conve

HeterSEED: Semantics-Structure Decoupling for Heterogeneous Graph Learning under Heterophily

SafetyDGX agent

arXiv:2605.04594v1 Announce Type: new Abstract: Many real-world heterogeneous graphs exhibit pronounced heterophily, where connected nodes often have dissimilar labels or play different semantic roles

Hey @bloomberg could you update this? It’s already way out of date. 🙏

SafetyDGX agent

Gary Marcus posted on X (formerly Twitter) requesting that Bloomberg update outdated information, likely regarding AI safety, regulation, or related technological topics that Marcus frequently discuss

Hierarchical Support Vector State Partitioning for Distilling Black Box Reinforcement Learning Policies

SafetyDGX agent

arXiv:2605.04254v1 Announce Type: new Abstract: We introduce State Vector Space Partitioning (SVSP), a novel method to mimic a black box reinforcement learning policy using a set of human-interpretabl

High-Fidelity Single-Image Head Modeling with Industry-Grade Topology

SafetyDGX agent

arXiv:2605.04524v1 Announce Type: new Abstract: We present a single-image head mesh reconstruction framework that addresses the longstanding challenge of simultaneously preserving facial identity and

How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models?

SafetyDGX agent

arXiv:2602.02924v2 Announce Type: replace Abstract: Diffusion policy sampling enables reinforcement learning (RL) to represent multimodal action distributions beyond suboptimal unimodal Gaussian polic

Hybrid Congestion Classification Framework Using Flow-Guided Attention and Empirical Mode Decomposition

SafetyDGX agent

arXiv:2605.04752v1 Announce Type: new Abstract: Accurate traffic congestion classification requires models that jointly capture roadway scene context and non-stationary traffic motion, yet most prior

I really enjoyed chatting with @mattturck, was a great discussion.

SafetyDGX agent

I really enjoyed chatting with @mattturck, was a great discussion. Deeply thoughtful conversation with @zicokolter, board member at @OpenAI and head of the machine learning department at @CarnegieMell

If point #4 is true, things are gonna get really wild.

SafetyDGX agent

If point #4 is true, things are gonna get really wild. Hot take on Elon’s surprise decision to rent 30 megawatts of compute to Anthropic: 1. It’s a tacit concession that xAI is not all that close to A

I'm really excited about this as a new tool in our interpretability tool kit

SafetyDGX agent

I'm really excited about this as a new tool in our interpretability tool kit In a new paper, we present NLAs, an unsupervised method for converting an LLM's internal state into human-readable text. I'

Improving Bias Correction Standards by Quantifying its Effects on Treatment Outcomes

SafetyDGX agent

arXiv:2407.14861v3 Announce Type: replace-cross Abstract: With the growing access to administrative health databases, retrospective studies have become crucial evidence for medical treatments. Yet, no

Improving Medical VQA through Trajectory-Aware Process Supervision

SafetyDGX agent

arXiv:2605.04064v1 Announce Type: cross Abstract: Reasoning capabilities are crucial for reliable medical visual question answering (VQA); however, existing datasets rarely include reasoning explanati

InterFuserDVS: Event-Enhanced Sensor Fusion for Safe RL-Based Decision Making

SafetyDGX agent

arXiv:2605.04355v1 Announce Type: new Abstract: Autonomous driving systems rely heavily on robust sensor fusion to perceive complex envi- ronments. Traditional setups using RGB cameras and LiDAR often

Introducing Trusted Contact in ChatGPT

SafetyDGX agent

OpenAI introduced a 'Trusted Contact' feature in ChatGPT that allows users to designate emergency contacts who can request access to their account in case of incapacity or death. This feature provides

Investigating Trustworthiness of Nonparametric Deep Survival Models for Alzheimer's Disease Progression Analysis

SafetyDGX agent

arXiv:2605.04063v1 Announce Type: new Abstract: Alzheimer's Dementia (AD) is a progressive neurodegenerative disease marked by irreversible decline, making reliable modeling of its progression essenti

It'll be 'impossible to slow down the ASI race' until it (very) suddenly isn't.

SafetyDGX agent

This post discusses the dynamics of artificial superintelligence (ASI) development as a competitive race, arguing that competitive pressures make it difficult to slow progress until a critical inflect

Joint Semantic Token Selection and Prompt Optimization for Interpretable Prompt Learning

SafetyDGX agent

arXiv:2605.04425v1 Announce Type: new Abstract: Vision-language models such as CLIP achieve strong visual-textual alignment, but often suffer from overfitting and limited interpretability when adapted

🚨Just published! Earth revolves around sun! Follow my account for more breaking news and analysis!

SafetyDGX agent

This post appears to be a satirical commentary on social media sensationalism, using the absurd claim that Earth's revolution around the sun is 'breaking news' to critique how trivial or well-establis

Lightweight Cross-Spectral Face Recognition via Contrastive Alignment and Distillation

SafetyDGX agent

arXiv:2605.04769v1 Announce Type: new Abstract: Heterogeneous Face Recognition (HFR) aims at matching face images captured across different sensing modalities, such as thermal-to-visible or near-infra

LineRides: Line-Guided Reinforcement Learning for Bicycle Robot Stunts

SafetyDGX agent

arXiv:2605.05110v1 Announce Type: new Abstract: Designing reward functions for agile robotic maneuvers in reinforcement learning remains difficult, and demonstration-based approaches often require ref

LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey

SafetyDGX agent

arXiv:2505.00753v5 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have sparked growing interest in building fully autonomous agents. However, fully autonomous LLM-bas

LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy

SafetyDGX agent

arXiv:2605.04295v1 Announce Type: new Abstract: LLMs' overconfidence, particularly when hallucinating, poses a significant challenge for the deployment of the models in safety-critical settings and ma

Look Once, Beam Twice: Camera-Primed Real-Time Double-Directional mmWave Beam Management for Vehicular Connectivity

SafetyDGX agent

arXiv:2605.05071v1 Announce Type: cross Abstract: Millimeter-wave (mmWave) frequencies promise multi-gigabit connectivity for vehicle-to-everything (V2X) networks, but face challenges in terms of seve

Many, many OpenAI employees quit over safety concerns, including @DKokotajlo, William Saunders, @sjgadler, etc as well @Janleike. The founde…

SafetyDGX agent

Many, many OpenAI employees quit over safety concerns, including @DKokotajlo, William Saunders, @sjgadler, etc as well @Janleike. The founders of Anthropic such as @DarioAmodei and @jackclarkSF may ha

Marcus (rightly) Mocks X influencer accounts 😁

SafetyDGX agent

Gary Marcus criticizes the credibility and practices of X (formerly Twitter) influencer accounts, likely highlighting misleading claims, engagement manipulation, or questionable expertise common among

Mechanical Conscience: A Mathematical Framework for Dependability of Machine Intelligenc

SafetyDGX agent

arXiv:2605.03847v1 Announce Type: new Abstract: Distributed collaborative intelligence (DCI), encompassing edge-to-edge architectures, federated learning, transfer learning, and swarm systems, creates

MedFabric and EtHER: A Data-Centric Framework for Word-Level Fabrication Generation and Detection in Medical LLMs

SafetyDGX agent

arXiv:2605.04180v1 Announce Type: new Abstract: Large Language Models exhibit strong reasoning and semantic understanding capabilities but often hallucinate in domains that require expert knowledge, a

MenuNet: A Strategy-Proof Mechanism for Matching Markets

SafetyDGX agent

arXiv:2605.03216v1 Announce Type: cross Abstract: Strategy-proofness is a fundamental desideratum in mechanism design, ensuring truthful reporting and robust participation. Stability is another centra

Meta challenges Ofcom in UK High Court over the Online Safety Act, which calculates levies based on global, not UK, revenue, in a case scheduled for October (Sam Tobin/Reuters)

SafetyDGX agent

Sam Tobin / Reuters: Meta challenges Ofcom in UK High Court over the Online Safety Act, which calculates levies based on global, not UK, revenue, in a case scheduled for October — Facebook and Instagr

Misaligned by Reward: Socially Undesirable Preferences in LLMs

SafetyDGX agent

arXiv:2605.05003v1 Announce Type: new Abstract: Reward models are a key component of large language model alignment, serving as proxies for human preferences during training. However, existing evaluat

Multi-Level Bidirectional Biomimetic Learning for EEG-Based Visual Decoding

SafetyDGX agent

arXiv:2605.04680v1 Announce Type: new Abstract: EEG-based visual neural decoding aims to align neural responses with visual stimuli for tasks such as image retrieval. However, limited paired data and

Multi-Scale Wavelet Transformers for Operator Learning of Dynamical Systems

SafetyDGX agent

arXiv:2602.01486v2 Announce Type: replace Abstract: Recent years have seen a surge in data-driven surrogates for dynamical systems that can be orders of magnitude faster than numerical solvers. Howeve

Multivariate Time Series Data Imputation via Distributionally Robust Regularization

SafetyDGX agent

arXiv:2602.00844v2 Announce Type: replace-cross Abstract: Multivariate time series imputation is often compromised by mismatch between the observed and true data distributions, a bias induced by the c

NEAT: Neighborhood-Guided, Efficient, Autoregressive Set Transformer for 3D Molecular Generation

SafetyDGX agent

arXiv:2512.05844v3 Announce Type: replace Abstract: Transformer-based autoregressive models offer an efficient alternative to diffusion- and flow-matching-based approaches for generating 3D molecules.

← Previous
1…159160161162163…212
Next →