AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
7 May 2026

CRAFT: Counterfactual-to-Interactive Reinforcement Fine-Tuning for Driving Policies

SafetyDGX agent

arXiv:2605.04470v1 Announce Type: new Abstract: Open-loop imitation learning has advanced modern autonomous driving policy architectures, but closed-loop deployment remains vulnerable to policy-induce

D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models

SafetyDGX agent

arXiv:2605.05204v1 Announce Type: new Abstract: The landscape of high-performance image generation models is currently shifting from the inefficient multi-step ones to the efficient few-step counterpa

Data-dependent Exploration for Online Reinforcement Learning from Human Feedback

SafetyDGX agent

arXiv:2605.04477v1 Announce Type: new Abstract: Online reinforcement learning from human feedback (RLHF) has emerged as a promising paradigm for aligning large language models (LLMs) by continuously c

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Decompose to Understand, Fuse to Detect: Frequency-Decoupled Anomaly Detection for Encrypted Network Traffic

SafetyDGX agent

arXiv:2605.02970v1 Announce Type: cross Abstract: Network traffic anomaly detection represents a critical cybersecurity task, yet widespread encryption makes this task increasingly challenging. In res

DFPO: Scaling Value Modeling via Distributional Flow towards Robust and Generalizable LLM Post-Training

SafetyDGX agent

arXiv:2602.05890v2 Announce Type: replace-cross Abstract: Training reinforcement learning (RL) systems in real-world environments remains challenging due to noisy supervision and poor out-of-domain (O

Direct Product Flow Matching: Decoupling Radial and Angular Dynamics for Few-Shot Adaptation

SafetyDGX agent

arXiv:2605.05054v1 Announce Type: new Abstract: Recent flow matching (FM) methods improve the few-shot adaptation of vision-language models, by modeling cross-modal alignment as a continuous multi-ste

Discovering Sparse Counterfactual Factors via Latent Adjustment for Survey-based Community Intervention

SafetyDGX agent

arXiv:2605.04460v1 Announce Type: new Abstract: Transportation surveys are widely used to understand travel preferences and adoption barriers, yet most survey-based analyses remain descriptive or pred

Distilling Bayesian Belief States into Language Models for Auditable Negotiation

SafetyDGX agent

arXiv:2605.04507v1 Announce Type: new Abstract: Negotiation agents must infer what their counterpart values, update those beliefs over dialogue turns, and choose actions under uncertainty. End-to-end

Dream-MPC: Gradient-Based Model Predictive Control with Latent Imagination

SafetyDGX agent

arXiv:2605.04568v1 Announce Type: new Abstract: State-of-the-art model-based Reinforcement Learning (RL) approaches either use gradient-free, population-based methods for planning, learned policy netw

Dynamic Hyperparameter Importance for Efficient Multi-Objective Optimization

SafetyDGX agent

arXiv:2601.03166v2 Announce Type: replace Abstract: Choosing a suitable ML model is a complex task that can depend on several objectives, e.g., accuracy, fairness, or energy consumption. In practice,

Efficiency of Parallel and Restart Exploration Strategies in Model Free Stochastic Simulations

SafetyDGX agent

arXiv:2503.03565v3 Announce Type: replace-cross Abstract: We analyze the efficiency of parallelization and restart mechanisms for stochastic simulations in model-free settings, where the underlying sy

Efficient Geometry-Controlled High-Resolution Satellite Image Synthesis

SafetyDGX agent

arXiv:2605.04557v1 Announce Type: new Abstract: High-resolution satellite images are often scarce and costly, especially for remote areas or infrequent events. This shortage hampers the development an

Efficient Model-Based Reinforcement Learning for Robot Control via Online Optimization

SafetyDGX agent

arXiv:2510.18518v2 Announce Type: replace Abstract: We present an online model-based reinforcement learning algorithm suitable for controlling complex robotic systems directly in the real world. Unlik

Efficiently Aligning Language Models with Online Natural Language Feedback

SafetyDGX agent

arXiv:2605.04356v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has been used to elicit impressive performance from language models in many domains. But, broadly benefic

Elicitation Matters: How Prompts and Query Protocols Shape LLM Surrogates under Sparse Observations

SafetyDGX agent

arXiv:2605.04764v1 Announce Type: new Abstract: Large language models are increasingly used as surrogate models for low-data optimization, but their optimizer-facing prediction and its uncertainty rem

Elon Musk’s DOGE “blatantly used” race, gender and other protected characteristics to execute the largest mass termination of federal grants…

SafetyDGX agent

Elon Musk’s DOGE “blatantly used” race, gender and other protected characteristics to execute the largest mass termination of federal grants in the history of the National Endowment for the Humanities

Enhancing the interpretability of spatially variable N2O model predictions with soft sensors during wastewater treatment

SafetyDGX agent

arXiv:2605.04082v1 Announce Type: new Abstract: Model-based solutions for nitrous oxide (N2O) emissions from wastewater treatment plants (WWTP) are informed by operational datasets designed to control

EO-1 is now available directly in LeRobot! You can train, evaluate, and deploy EO-1 through the standard LeRobot policy interface, making it…

SafetyDGX agent

EO-1 is now available directly in LeRobot! You can train, evaluate, and deploy EO-1 through the standard LeRobot policy interface, making it much easier to try EO-1 on robot-control workflows. Huge th

EP-GRPO: Entropy-Progress Aligned Group Relative Policy Optimization with Implicit Process Guidance

SafetyDGX agent

arXiv:2605.04960v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR), particularly Group Relative Policy Optimization (GRPO), has advanced LLM reasoning. However, GRPO

Evaluating Semantic Fragility in Text-to-Audio Generation Systems Under Controlled Prompt Perturbations

SafetyDGX agent

arXiv:2603.13824v2 Announce Type: replace-cross Abstract: Recent advances in text-to-audio generation enable models to translate natural-language descriptions into diverse musical output. However, the

Every Step Counts: Step-Level Credit Assignment for Tool-Integrated Text-to-SQL

SafetyDGX agent

arXiv:2605.04719v1 Announce Type: new Abstract: Tool-integrated Text-to-SQL parsing has emerged as a promising paradigm, framing SQL generation as a sequential decision-making process interleaved with

Extending Differential Temporal Difference Methods for Episodic Problems

SafetyDGX agent

arXiv:2605.04368v1 Announce Type: new Abstract: Differential temporal difference (TD) methods are value-based reinforcement learning algorithms that have been proposed for infinite-horizon problems. T

FairEnc: A Fair Vision-Language Model with Fair Vision and Text Encoders for Glaucoma Detection

SafetyDGX agent

arXiv:2605.04882v1 Announce Type: new Abstract: Automated glaucoma detection is critical for preventing irreversible vision loss and reducing the burden on healthcare systems. However, ensuring fairne

Fascinating: People who get advice from chatbots often take that advice – but rarely end up feeling any better.

SafetyDGX agent

Research indicates that people frequently follow advice provided by chatbots, yet this compliance with AI recommendations does not correlate with improved outcomes or subjective well-being. This sugge

First person to update this chart with no mistakes using an AI system and a general prompt (rather than one than that handholds the system e…

SafetyDGX agent

First person to update this chart with no mistakes using an AI system and a general prompt (rather than one than that handholds the system every step of the win) wins a signed book! Please suppy your

Fixed-Length Dense Fingerprint Representation with Alignment and Robust Enhancement

SafetyDGX agent

arXiv:2505.03597v2 Announce Type: replace Abstract: Fixed-length fingerprint representations, which map each fingerprint to a compact and fixed-size feature vector, are computationally efficient and w

FLUID: Continuous-Time Hyperconnected Sparse Transformer for Sink-Free Learning

SafetyDGX agent

arXiv:2605.04421v1 Announce Type: new Abstract: Continuous-time (CT) Transformers improve irregular and long-range modeling over CT-RNNs by exploiting inputs or outputs embeddings with continuous dyna

Former OpenAI employee Rosie Campbell is the FIRST PERSON in the trial to point out that OpenAI's mission does NOT REQUIRE building AGI Unde…

SafetyDGX agent

Former OpenAI employee Rosie Campbell is the FIRST PERSON in the trial to point out that OpenAI's mission does NOT REQUIRE building AGI Undermines the entire justification for creating the for-profit

Globally Solving Unbalanced Optimal Transport and Density Control for Gaussian Distributions

SafetyDGX agent

arXiv:2605.04246v1 Announce Type: cross Abstract: In this article, we study unbalanced optimal transport (UOT) and establish a control-theoretic dynamical extension, which we call the unbalanced densi

Graph-Augmented LLMs for Swiss MP Ideology Prediction

SafetyDGX agent

arXiv:2605.04643v1 Announce Type: new Abstract: Approximating the ideological position of Members of Parliament (MPs) is a fundamental task in political science, helping researchers understand legisla

Graph-SND: Sparse Aggregation for Behavioral Diversity in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.05020v1 Announce Type: new Abstract: System Neural Diversity (SND) measures behavioral heterogeneity in multi-agent reinforcement learning by averaging pairwise distances over all inom{n}{2

Helen Toner's deposition in Musk v Altman includes some striking quotes about Mira Murati's involvement with Altman's ouster. She said Mira …

SafetyDGX agent

Helen Toner's deposition in Musk v Altman includes some striking quotes about Mira Murati's involvement with Altman's ouster. She said Mira was 'totally uninterested in telling her team that her conve

HeterSEED: Semantics-Structure Decoupling for Heterogeneous Graph Learning under Heterophily

SafetyDGX agent

arXiv:2605.04594v1 Announce Type: new Abstract: Many real-world heterogeneous graphs exhibit pronounced heterophily, where connected nodes often have dissimilar labels or play different semantic roles

Hierarchical Support Vector State Partitioning for Distilling Black Box Reinforcement Learning Policies

SafetyDGX agent

arXiv:2605.04254v1 Announce Type: new Abstract: We introduce State Vector Space Partitioning (SVSP), a novel method to mimic a black box reinforcement learning policy using a set of human-interpretabl

High-Fidelity Single-Image Head Modeling with Industry-Grade Topology

SafetyDGX agent

arXiv:2605.04524v1 Announce Type: new Abstract: We present a single-image head mesh reconstruction framework that addresses the longstanding challenge of simultaneously preserving facial identity and

Hybrid Congestion Classification Framework Using Flow-Guided Attention and Empirical Mode Decomposition

SafetyDGX agent

arXiv:2605.04752v1 Announce Type: new Abstract: Accurate traffic congestion classification requires models that jointly capture roadway scene context and non-stationary traffic motion, yet most prior

If point #4 is true, things are gonna get really wild.

SafetyDGX agent

If point #4 is true, things are gonna get really wild. Hot take on Elon’s surprise decision to rent 30 megawatts of compute to Anthropic: 1. It’s a tacit concession that xAI is not all that close to A

Improving Bias Correction Standards by Quantifying its Effects on Treatment Outcomes

SafetyDGX agent

arXiv:2407.14861v3 Announce Type: replace-cross Abstract: With the growing access to administrative health databases, retrospective studies have become crucial evidence for medical treatments. Yet, no

Improving Medical VQA through Trajectory-Aware Process Supervision

SafetyDGX agent

arXiv:2605.04064v1 Announce Type: cross Abstract: Reasoning capabilities are crucial for reliable medical visual question answering (VQA); however, existing datasets rarely include reasoning explanati

Introducing Trusted Contact in ChatGPT

SafetyDGX agent

OpenAI introduced a 'Trusted Contact' feature in ChatGPT that allows users to designate emergency contacts who can request access to their account in case of incapacity or death. This feature provides

Investigating Trustworthiness of Nonparametric Deep Survival Models for Alzheimer's Disease Progression Analysis

SafetyDGX agent

arXiv:2605.04063v1 Announce Type: new Abstract: Alzheimer's Dementia (AD) is a progressive neurodegenerative disease marked by irreversible decline, making reliable modeling of its progression essenti

Joint Semantic Token Selection and Prompt Optimization for Interpretable Prompt Learning

SafetyDGX agent

arXiv:2605.04425v1 Announce Type: new Abstract: Vision-language models such as CLIP achieve strong visual-textual alignment, but often suffer from overfitting and limited interpretability when adapted

🚨Just published! Earth revolves around sun! Follow my account for more breaking news and analysis!

SafetyDGX agent

This post appears to be a satirical commentary on social media sensationalism, using the absurd claim that Earth's revolution around the sun is 'breaking news' to critique how trivial or well-establis

Lightweight Cross-Spectral Face Recognition via Contrastive Alignment and Distillation

SafetyDGX agent

arXiv:2605.04769v1 Announce Type: new Abstract: Heterogeneous Face Recognition (HFR) aims at matching face images captured across different sensing modalities, such as thermal-to-visible or near-infra

LineRides: Line-Guided Reinforcement Learning for Bicycle Robot Stunts

SafetyDGX agent

arXiv:2605.05110v1 Announce Type: new Abstract: Designing reward functions for agile robotic maneuvers in reinforcement learning remains difficult, and demonstration-based approaches often require ref

Look Once, Beam Twice: Camera-Primed Real-Time Double-Directional mmWave Beam Management for Vehicular Connectivity

SafetyDGX agent

arXiv:2605.05071v1 Announce Type: cross Abstract: Millimeter-wave (mmWave) frequencies promise multi-gigabit connectivity for vehicle-to-everything (V2X) networks, but face challenges in terms of seve

Marcus (rightly) Mocks X influencer accounts 😁

SafetyDGX agent

Gary Marcus criticizes the credibility and practices of X (formerly Twitter) influencer accounts, likely highlighting misleading claims, engagement manipulation, or questionable expertise common among

Mechanical Conscience: A Mathematical Framework for Dependability of Machine Intelligenc

SafetyDGX agent

arXiv:2605.03847v1 Announce Type: new Abstract: Distributed collaborative intelligence (DCI), encompassing edge-to-edge architectures, federated learning, transfer learning, and swarm systems, creates

MedFabric and EtHER: A Data-Centric Framework for Word-Level Fabrication Generation and Detection in Medical LLMs

SafetyDGX agent

arXiv:2605.04180v1 Announce Type: new Abstract: Large Language Models exhibit strong reasoning and semantic understanding capabilities but often hallucinate in domains that require expert knowledge, a

MenuNet: A Strategy-Proof Mechanism for Matching Markets

SafetyDGX agent

arXiv:2605.03216v1 Announce Type: cross Abstract: Strategy-proofness is a fundamental desideratum in mechanism design, ensuring truthful reporting and robust participation. Stability is another centra

MOSAIC-Bench: Measuring Compositional Vulnerability Induction in Coding Agents

Model ReleasesDGX agent

arXiv:2605.03952v1 Announce Type: cross Abstract: Coding agents often pass per-prompt safety review yet ship exploitable code when their tasks are decomposed into routine engineering tickets. The chal

Multi-Level Bidirectional Biomimetic Learning for EEG-Based Visual Decoding

SafetyDGX agent

arXiv:2605.04680v1 Announce Type: new Abstract: EEG-based visual neural decoding aims to align neural responses with visual stimuli for tasks such as image retrieval. However, limited paired data and

Multi-Scale Wavelet Transformers for Operator Learning of Dynamical Systems

SafetyDGX agent

arXiv:2602.01486v2 Announce Type: replace Abstract: Recent years have seen a surge in data-driven surrogates for dynamical systems that can be orders of magnitude faster than numerical solvers. Howeve

Multivariate Time Series Data Imputation via Distributionally Robust Regularization

SafetyDGX agent

arXiv:2602.00844v2 Announce Type: replace-cross Abstract: Multivariate time series imputation is often compromised by mismatch between the observed and true data distributions, a bias induced by the c

NEAT: Neighborhood-Guided, Efficient, Autoregressive Set Transformer for 3D Molecular Generation

SafetyDGX agent

arXiv:2512.05844v3 Announce Type: replace Abstract: Transformer-based autoregressive models offer an efficient alternative to diffusion- and flow-matching-based approaches for generating 3D molecules.

New Bigtable in-memory tier for sub-millisecond read latency

SafetyDGX agent

In the high-stakes world of digital infrastructure, speed isn't just a metric — it’s currency. At Google Cloud Next ‘26 we announced the Bigtable in-memory tier, a breakthrough for our fully managed c

On-line Learning in Tree MDPs by Treating Policies as Bandit Arms

SafetyDGX agent

arXiv:2605.04979v1 Announce Type: cross Abstract: A Tree Markov Decision Problem (T-MDP) is a finite-horizon MDP with a starting state s_{1}, in which every state is reachable from s_{1} through exact

On the Hardness of Junking LLMs

SafetyDGX agent

arXiv:2605.05116v1 Announce Type: new Abstract: Large language models (LLMs) are known to be vulnerable to jailbreak attacks, which typically rely on carefully designed prompts containing explicit sem

One of the things that made the Mythos release hard to interpret is that Anthropic held back details on most vulns they found, to give defen…

SafetyDGX agent

One of the things that made the Mythos release hard to interpret is that Anthropic held back details on most vulns they found, to give defenders time to patch. 1 month later, info from orgs with acces

One Pool, Two Caches: Adaptive HBM Partitioning for Accelerating Generative Recommender Serving

SafetyDGX agent

arXiv:2605.04450v1 Announce Type: cross Abstract: Generative Recommender (GR) inference places embedding hot caches (EMB) and KV caches in direct competition for limited GPU HBM: allocating more memor

← Previous
1…181182183184185…240
Next →