AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
10 Jun 2026

experts on german law: can Google appeal, how soon does the new ruling take effect, etc?

SafetyDGX agent

This post likely discusses expert analysis on German legal proceedings involving Google, covering questions about appeal options available to the company, the timeline for implementing a court ruling,

Exploring Accurate and Transparent Domain Adaptation in Predictive Healthcare via Concept-Grounded Orthogonal Inference

SafetyDGX agent

arXiv:2602.12542v2 Announce Type: replace-cross Abstract: Deep learning models for clinical event prediction on electronic health records (EHR) often suffer performance degradation when deployed under

Fast and Highly Expressive Policy Learning for Offline Reinforcement Learning via Bootstrapped Flow Q-Learning

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.10613v1 Announce Type: cross Abstract: Diffusion-based Q-learning has emerged as a powerful paradigm for offline reinforcement learning, but its reliance on multi-step denoising makes both

Flow Control: Steering Vision-Language-Action Models with Simple Real-Time Inputs

SafetyDGX agent

arXiv:2606.10180v1 Announce Type: cross Abstract: We introduce flow control of vision-language-action (VLA) models, a simple and effective way to steer VLA actions in real-time through generic inputs,

Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models

SafetyDGX agent

arXiv:2606.11025v1 Announce Type: new Abstract: Recent work has demonstrated that online reinforcement learning (RL) can substantially improve the quality and alignment of flow matching models for ima

For Robotaxis, Safety Must Be Built In, Not Bolted On

SafetyDGX agent

A car pulls up to the curb. The app says, “Your ride is here.” No one’s in the driver’s seat. For people who live in one of the dozens of cities now hosting robotaxi services, this is already a realit

Generalized-CVO: Fast and Correspondence-Free Local Point Cloud Registration with Second Order Riemannian Optimization

SafetyDGX agent

arXiv:2606.10019v1 Announce Type: cross Abstract: We propose a fast and correspondence-free local point cloud registration method that leverages geometric surface structure and reproducing kernel Hilb

GHOST: Hierarchical Sub-Goal Policies for Generalizing Robot Manipulation

SafetyDGX agent

arXiv:2606.10025v1 Announce Type: cross Abstract: We present GHOST, a framework for learning visuomotor manipulation policies that generalize beyond the training distribution. GHOST factorizes control

Globally Localizing Lunar Rover in Pixels via Graph Alignment

SafetyDGX agent

arXiv:2606.10602v1 Announce Type: new Abstract: Precise rover localization is a prerequisite for autonomous lunar exploration, yet the absence of Global Navigation Satellite System (GNSS) signals and

Going with the Flow: Koopman Behavioral Models as Pseudo Planners for Visuo-Motor Dexterity

SafetyDGX agent

arXiv:2602.07413v3 Announce Type: replace Abstract: Contemporary visuo-motor dexterity models often rely on expressive policy classes with diffusion and transformer backbones to achieve strong perform

Gradient-Guided Reward Optimization for Inference-time Alignment

SafetyDGX agent

arXiv:2606.09635v1 Announce Type: cross Abstract: Ensuring the reliability of Large Language Models (LLMs) under distribution drift requires inference-time adaptation. While inference-time alignment m

GUI-AC: Enhancing Continual Learning in GUI Agents

SafetyDGX agent

arXiv:2606.10522v1 Announce Type: new Abstract: Graphical User Interfaces (GUIs) serve as the dominant medium for human-computer interaction, yet building GUI agents that generalize across the vast di

GUIDE: Goal-Initialized Directional Understanding for End-to-End Visual Navigation

SafetyDGX agent

arXiv:2606.10832v1 Announce Type: new Abstract: Learning-based visual navigation for legged robots typically relies on continuous goal updates from hierarchical state estimation to provide a persisten

GuideWalk: Learning Unified Autonomous Navigation and Locomotion for Humanoid Robots across Versatile Terrains

SafetyDGX agent

arXiv:2606.10449v1 Announce Type: new Abstract: Humanoid robots have achieved strong locomotion capabilities, but reliable navigation on versatile terrains remains challenging because obstacle avoidan

Hand-centric Human-to-Robot Trajectory Transfer from Video Demonstrations via Open-World Contact Localization

SafetyDGX agent

arXiv:2606.10743v1 Announce Type: new Abstract: Learning from human video demonstrations remains challenging due to noisy hand-object interactions, unseen objects with partial observation, and cross-e

Hidden Consensus:Preference-Validity Compression in Human Feedback

SafetyDGX agent

arXiv:2606.10569v1 Announce Type: cross Abstract: Standard RLHF pipelines often reduce heterogeneous human judgments into a single scalar reward target. We argue that this reduction can mis-measure al

Hierarchical Policies from Verbal and Egocentric Human Signals for Natural Human-Robot Interaction

SafetyDGX agent

arXiv:2606.10276v1 Announce Type: cross Abstract: For natural human-robot interaction, a robot must understand human intent expressed not only through language but also through nonverbal signals such

HiGR: Industrial-Scale Hierarchical Generative Slate Recommendation Framework in Tencent

SafetyDGX agent

arXiv:2512.24787v3 Announce Type: replace-cross Abstract: Slate recommendation, which presents users with a ranked item list in a single display, is ubiquitous across mainstream online platforms. Whil

Human-AI Coordination Zones: A Framework for Designing Human-in-the-Loop Experiences with Agentic AI

SafetyDGX agent

arXiv:2606.09848v1 Announce Type: cross Abstract: As generative and agentic AI becomes embedded in everyday products, practitioners face a persistent challenge: how to design human-AI coordination --

I was first to call the death of tokenmaxxing (ahead of @zerohedge), among the first (in 2023) to warn that the economics of Generative AI w…

SafetyDGX agent

I was first to call the death of tokenmaxxing (ahead of @zerohedge), among the first (in 2023) to warn that the economics of Generative AI was problematic, and first (also 2023) to warn that OpenAI co

IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder

SafetyDGX agent

arXiv:2606.11096v1 Announce Type: new Abstract: Built on pretrained vision foundation models (VFMs), representation autoencoders (RAEs) have recently emerged as a promising approach for constructing s

🔔 IF Anthropic is for real about safety, the should pause, for one month, and show leadership. If they won’t pause, even briefly, we have a…

SafetyDGX agent

🔔 IF Anthropic is for real about safety, the should pause, for one month, and show leadership. If they won’t pause, even briefly, we have a different kind of answer about what is real and what is mark

If the new German court decision — to hold LLM companies liable for the false claims their LLMs made — become universal, effectively sidelin…

SafetyDGX agent

If the new German court decision — to hold LLM companies liable for the false claims their LLMs made — become universal, effectively sidelining generative AI until better techniques came along, would

IMPACT: Learning Internal-Model Predictive Control for Forceful Robotic Manipulation

SafetyDGX agent

arXiv:2606.10818v1 Announce Type: cross Abstract: Real-world robotic manipulation tasks often involve forceful interactions with the environment, such as using tools of varying weights, transporting o

Improving Adversarial Transferability on Vision-Language Pre-training Models via Surrogate-Specific Bias Correction

SafetyDGX agent

arXiv:2606.10571v1 Announce Type: cross Abstract: Adversarial examples reveal vulnerabilities in Vision-Language Pre-training (VLP) models and provide insights for improving robustness. A key property

Improving Text-Instance Alignment Of Foreground Conditioned Out-Painting Via Customized Concept Embedding

SafetyDGX agent

arXiv:2606.10892v1 Announce Type: cross Abstract: To showcase products, merchants often incur substantial costs creating high-quality display images. Foreground Conditioned Outpainting (FCO) meets thi

In addition to transparency, I now believe frontier models should face mandatory third-party testing for cyber, bio, and autonomy risks—with…

SafetyDGX agent

In addition to transparency, I now believe frontier models should face mandatory third-party testing for cyber, bio, and autonomy risks—with the power to block or revoke deployment of models that pose

Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access

SafetyDGX agent

arXiv:2509.26000v3 Announce Type: replace Abstract: Asymmetric reinforcement learning leverages privileged information available during training to improve learning under partial observability. Existi

Inverse Probability Weighting and Age-of-Information Aggregation for Decentralized Federated Learning under Partial Reception

SafetyDGX agent

arXiv:2606.10774v1 Announce Type: new Abstract: Decentralized Federated Learning (DFL) over lossy wireless networks faces two key challenges: selection bias, where updates from poor-quality links are

Is Fairness Truly Fair? Towards Reliable Lipschitz Fairness in Multi-Task Learning via Fixed-exorpdfstring{elta}{delta} Alignment

SafetyDGX agent

arXiv:2606.10632v1 Announce Type: cross Abstract: Lipschitz-style individual fairness formalizes the idea that semantically similar examples should receive similar predictions, but its evaluation in m

It is not literally impossible to pause or slow down. It is a choice. But if the “frontier labs” continue to refuse to do, I strongly believ…

SafetyDGX agent

It is not literally impossible to pause or slow down. It is a choice. But if the “frontier labs” continue to refuse to do, I strongly believe it will eventually lead to tremendous harm. (Essentially e

It Takes One to Bias Them All: Breaking Bad with One-Shot GRPO

SafetyDGX agent

arXiv:2606.10931v1 Announce Type: new Abstract: Warning: This paper contains several toxic and offensive statements. Modern large language models (LLMs) are typically aligned through large-scale post-

LAFP: Preserving Latent Action Structure in Latent Policy Learning via Flow Matching

SafetyDGX agent

arXiv:2606.10517v1 Announce Type: new Abstract: Learning high-quality latent actions from large-scale unlabeled videos, coupled with limited real-world interaction data for training an action decoder,

Lightweight Latent Reasoning for Narrative Tasks

SafetyDGX agent

arXiv:2512.02240v2 Announce Type: replace Abstract: Large language models (LLMs) tackle complex tasks by generating long chains of thought or 'reasoning traces' that act as latent variables in the gen

Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization

SafetyDGX agent

arXiv:2606.11180v1 Announce Type: new Abstract: Diffusion-based lip synchronization models achieve strong visual quality and audio-visual alignment, but full-sequence bidirectional attention and many

LLM-Aided Joint Secrecy Precoding and Trajectory for RSMA-Based Heterogeneous UAV Networks

SafetyDGX agent

arXiv:2507.17188v2 Announce Type: replace-cross Abstract: This paper investigates secure communications in rate-splitting multiple access (RSMA) enabled heterogeneous UAV networks, where multiple UAVs

Locomotion analysis of a quadruped interacting with the lunar granular surface

SafetyDGX agent

arXiv:2606.10273v1 Announce Type: new Abstract: Deploying legged robots in extra-terrestrial environments includes many challenges due to complex terrain interactions, energy, and thermal constraints.

ManiSplat: Manipulation Trajectory Synthesis from Monocular Video via Decoupled 3D Gaussian Splatting

SafetyDGX agent

arXiv:2606.10645v1 Announce Type: new Abstract: Reconstructing dynamic and interactive 3D scenes from real-world observations remains a fundamental challenge in computer vision and robotics. While rec

Many of these policy ideas have common-sense appeal across the political spectrum, and the sooner we act on them, the sooner everyone shares…

SafetyDGX agent

Dario Amodei discusses policy proposals that have bipartisan appeal and advocates for prompt action to implement them for broad societal benefit. The post suggests that certain policy ideas transcend

MARCH: Model-Assisted Reinforcement Learning for the Perceptive Control of Humanoids over Sparse Footholds

SafetyDGX agent

arXiv:2606.10288v1 Announce Type: new Abstract: Perceptive bipedal locomotion over sparse terrain remains a difficult challenge: model-based methods are precise but brittle to uncertainty, while model

Masa can hype AI all he likes, but apparently he can’t even get a loan on his OpenAI shares. 🤔 But I actually half agree—and it might surpr…

SafetyDGX agent

Masa can hype AI all he likes, but apparently he can’t even get a loan on his OpenAI shares. 🤔 But I actually half agree—and it might surprise some of my critics to hear this, especially those who hav

Mean Flow Distillation: Robust and Stable Distillation for Flow Matching Models

SafetyDGX agent

arXiv:2606.11155v1 Announce Type: new Abstract: Flow Matching models have demonstrated strong performance across a wide range of generative tasks. However, their reliance on ODE-based iterative sampli

Measuring Human Value Expression in Social Media Texts: Calibrated LLM Annotation and Encoder Transfer

SafetyDGX agent

arXiv:2606.11018v1 Announce Type: new Abstract: Measuring subjective constructs in naturally occurring social media text requires annotation procedures that are theoretically grounded, empirically val

Mechanistic Analysis of Alignment Algorithms in Language Models

SafetyDGX agent

arXiv:2606.09850v1 Announce Type: cross Abstract: Post-training alignment algorithms are predominantly evaluated as black boxes, obscuring how they reshape language models' internal computations. We p

MedFeat: Model-Aware and Explainability-Driven Feature Engineering with LLMs for Clinical Tabular Prediction

SafetyDGX agent

arXiv:2603.02221v2 Announce Type: replace-cross Abstract: In clinical tabular prediction, classical machine learning models with feature engineering often outperform neural methods. LLMs are increasin

MIND-V: Hierarchical World Model for Long-Horizon Robotic Manipulation with RL-based Physical Alignment

SafetyDGX agent

arXiv:2512.06628v3 Announce Type: replace-cross Abstract: Scalable embodied intelligence is constrained by the scarcity of diverse, long-horizon robotic manipulation data. Existing video world models

Mitigating Bias in Low-SNR Financial Reinforcement Learning via Quantum Representations

SafetyDGX agent

arXiv:2606.10448v1 Announce Type: cross Abstract: The financial market is a typical low signal-to-noise ratio (SNR) setting, which often destabilizes off-policy maximum-entropy methods like Soft Actor

Mitigating Manifold Departure: Uncertainty-Aware Subspace Rectification for Trustworthy MLLM Decoding

SafetyDGX agent

arXiv:2606.09859v1 Announce Type: cross Abstract: MLLMs frequently hallucinate objects inconsistent with visual inputs. This issue is typically attributed to the over-reliance on language priors, whic

MMD Guidance: Training-Free Distribution Adaptation for Diffusion Models via Maximum Mean Discrepancy Guidance

SafetyDGX agent

arXiv:2601.08379v2 Announce Type: replace-cross Abstract: Pre-trained diffusion models have emerged as powerful generative priors for both unconditional and conditional sample generation, yet their ou

MODIP: Efficient Model-Based Optimization for Diffusion Policies

SafetyDGX agent

arXiv:2606.10825v1 Announce Type: new Abstract: Diffusion policies (DPs) have emerged as expressive policy representations for robot learning, often used with imitation learning methods such as behavi

Multi-Faceted Interactivity Alignment in Full-Duplex Speech Models

SafetyDGX agent

arXiv:2606.11167v1 Announce Type: new Abstract: Full-duplex spoken dialogue models can listen and speak simultaneously, making them a promising architecture for natural conversation. However, current

Multilingual Word-Level Forced Alignment with Self-Supervised Representations and Learned Dynamic Programming

SafetyDGX agent

arXiv:2606.10675v1 Announce Type: new Abstract: We present a method for accurate multilingual word-level forced alignment, consisting of an alignment encoder and a learned alignment decoder. The encod

New series from The Observer with @Yoshua_Bengio, me, and others on potential catastrophic risks in AI. (link at bottom of thread)

SafetyDGX agent

New series from The Observer with @Yoshua_Bengio, me, and others on potential catastrophic risks in AI. (link at bottom of thread) For the last few months I’ve been investigating how concerned we shou

Offline Reinforcement Learning for Rotation Profile Control in Tokamaks

SafetyDGX agent

arXiv:2605.05857v2 Announce Type: replace Abstract: Tokamaks remain leading candidates for achieving practical fusion energy, yet many important control problems inside these devices are still difficu

On-sky demonstration of reinforcement learning for adaptive optics control

SafetyDGX agent

arXiv:2606.10771v1 Announce Type: cross Abstract: Reinforcement learning (RL)-based algorithms have recently emerged as a promising approach for adaptive optics (AO) control. In simulations and labora

On the Controllability-Fidelity Frontier in Diffusion Editing

SafetyDGX agent

arXiv:2606.09901v1 Announce Type: cross Abstract: Diffusion-based generative models enable powerful image editing capabilities, but achieving precise control while maintaining fidelity and safety rema

One paper from three years ago has been influencing policymakers around the world about labour decisions regarding AI. What happens when the…

SafetyDGX agent

One paper from three years ago has been influencing policymakers around the world about labour decisions regarding AI. What happens when the gap between the evidence and new policy widens? Our report

$ORCL hit hard after hours on plan for raising even more money for AI fairytales. And this is despite the fact that [see CNBC headline on ri…

SafetyDGX agent

Gary Marcus critiques Oracle's after-hours stock decline following announcements of increased capital allocation toward AI initiatives, suggesting skepticism about the company's AI strategy and spendi

PADD: Path-Aligned Decompression Distillation for Non-Router Teacher to Guide MoE Student Learning

SafetyDGX agent

arXiv:2606.10369v1 Announce Type: new Abstract: As large language models (LLMs) continue to scale, it becomes increasingly challenging to grow model capacity under fixed computation budgets. We propos

ParaBridge: Bridging Paralinguistic Perception and Dialogue Behavior in Speech Language Models

SafetyDGX agent

arXiv:2606.10581v1 Announce Type: new Abstract: Speech carries more information than just words: a child's voice, a fearful tone, or a noisy background should all lead a sufficiently competent spoken-

← Previous
1…7677787980…214
Next →