AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
Safety

Do agents listen to you… or themselves? While evaling subagent behavior in deep agent systems, we noticed an interesting quirk in our agents…

DGX agent

Do agents listen to you… or themselves? While evaling subagent behavior in deep agent systems, we noticed an interesting quirk in our agents' alignment with hand-written system prompts vs. the instruc

safetyharrison-chase--x
21 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Enhancing Speech Large Language Models through Reinforced Behavior Alignment

DGX agent

arXiv:2509.03526v2 Announce Type: replace Abstract: The recent advancements of Large Language Models (LLMs) have spurred considerable research interest in extending their linguistic capabilities beyon

safetyarxiv-cs-cl
21 May 2026
Safety

EvalMORAAL: Interpretable Chain-of-Thought and LLM-as-Judge Evaluation for Moral Alignment in Large Language Models

DGX agent

arXiv:2510.05942v3 Announce Type: replace Abstract: We present EvalMORAAL, a transparent chain-of-thought (CoT) framework that uses two scoring methods (log-probabilities and direct ratings) plus a mo

safetyarxiv-cs-cl
21 May 2026
Safety

extit{Stochastic} MeanFlow Policies: One-Step Generative Control with Entropic Mirror Descent

DGX agent

arXiv:2605.21282v1 Announce Type: new Abstract: Online off-policy reinforcement learning (RL) is shaped by two coupled choices: the policy class and the update rule. Gaussian policies are fast and hav

safetyarxiv-cs-lg
21 May 2026
Safety

FEAT: A Linear-Complexity Foundation Model for Extremely Large Structured Data

DGX agent

arXiv:2603.16513v3 Announce Type: replace Abstract: Structured data is widely used in domains such as healthcare, finance, and scientific data management. Recent studies on structured data foundation

safetyarxiv-cs-lg
21 May 2026
Safety

FlowLong: Inference-time Long Video Generation via Manifold-constrained Tweedie Matching

DGX agent

arXiv:2605.20910v1 Announce Type: new Abstract: Extending the generation horizon of video diffusion models to long sequences remains a long-standing and important challenge. Existing training-free app

safetyarxiv-cs-cv
21 May 2026
Safety

Graph Transductive Sharpening: Leveraging Unlabeled Predictions in Node Classification

DGX agent

arXiv:2605.20248v1 Announce Type: new Abstract: In the transductive setting, where the full graph is observed but node labels are only partially available, progress in semi-supervised node classificat

safetyarxiv-cs-lg
21 May 2026
Safety

GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents

DGX agent

arXiv:2605.20246v1 Announce Type: new Abstract: Recently, vision-language model (VLM) agents have shown promising progress in open-world tasks, where successful task completion often requires multiple

safetyarxiv-cs-lg
21 May 2026
Safety

Here's how the corruption works: Thursday: RJ Reynolds donates $5M to Trump Saturday: Trump invites RJR execs to Mar a Lago; execs ask to lo…

DGX agent

Here's how the corruption works: Thursday: RJ Reynolds donates $5M to Trump Saturday: Trump invites RJR execs to Mar a Lago; execs ask to loosen regs on flavored vapes; Trump calls up RFK Jr. and tell

safetyyann-lecun--x
21 May 2026
Safety

HORST: Composing Optimizer Geometries for Sparse Transformer Training

DGX agent

arXiv:2605.21104v1 Announce Type: new Abstract: Sparsifying transformers remains a fundamental challenge, as standard optimizers fail to simultaneously encourage sparsity and maintain training stabili

safetyarxiv-cs-lg
21 May 2026
Safety

I don't think anyone has a good intuitive sense about what this means, and that failure of imagination is a generally bad thing for planning…

DGX agent

I don't think anyone has a good intuitive sense about what this means, and that failure of imagination is a generally bad thing for planning, investment, and policy. I also don't have an easy solution

safetyethan-mollick--x
21 May 2026
Safety

I was happy to join the @ScienceBoard_UN podcast to discuss deception among frontier AI models and the need to manage its potential global i…

DGX agent

I was happy to join the @ScienceBoard_UN podcast to discuss deception among frontier AI models and the need to manage its potential global impacts and risks. 🎙️What happens when AI learns to lie? Scie

safetyyoshua-bengio--x
21 May 2026
Safety

if you own a tesla, you are the product

DGX agent

if you own a tesla, you are the product Astonishing potential, especially if one takes all the galaxies into account and factors them into the calculation. 😃 And this comes from a man who removed Auto

safetygary-marcus--x
21 May 2026
Safety

Improved convergence rate of kNN graph Laplacians: differentiable self-tuned affinity

DGX agent

arXiv:2410.23212v2 Announce Type: replace-cross Abstract: In graph-based data analysis, k-nearest neighbor (kNN) graphs are widely used due to their adaptivity to local data densities. Allowing weight

safetyarxiv-cs-lg
21 May 2026
Safety

Inference Time Policy Optimization for Offline RL with Differentiable World Models

DGX agent

arXiv:2603.22430v2 Announce Type: replace Abstract: Offline Reinforcement Learning (RL) learns optimal policies from fixed datasets, training a policy once and deploying it at inference time without f

safetyarxiv-cs-lg
21 May 2026
Safety

It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs

DGX agent

arXiv:2605.20258v1 Announce Type: new Abstract: Contextual Integrity (CI) defines privacy not merely as keeping information hidden, but as governing information flows according to the norms of a given

safetyarxiv-cs-lg
21 May 2026
Safety

It’s crazy that OpenAI and xAI were built partly as ways to mitigate @demishassabis’s power and how both are more reckless and less focused …

DGX agent

It’s crazy that OpenAI and xAI were built partly as ways to mitigate @demishassabis’s power and how both are more reckless and less focused on the good part, viz AI for science. And then Anthropic was

safetygary-marcus--x
21 May 2026
Safety

Latent Geometry as a Structural Monitor: Eigenspace Alignment for Anomaly Detection in Anonymity Networks

DGX agent

arXiv:2605.20391v1 Announce Type: cross Abstract: Traditional anomaly detection marks events when measured signals cross predefined thresholds. This captures the moment of transition but not the struc

safetyarxiv-cs-lg
21 May 2026
Safety

Learning Robust Dexterous In-Hand Manipulation from Joint Sensors with Proprioceptive Transformer

DGX agent

arXiv:2605.21330v1 Announce Type: new Abstract: In-hand object manipulation is a fundamental yet challenging capability for dexterous robots. Despite significant progress in dexterous manipulation, ex

safetyarxiv-cs-ro
21 May 2026
Safety

Learning to Think in Physics: Breaking Shortcut Learning in Scientific Diffusion via Representation Alignment

DGX agent

arXiv:2605.20780v1 Announce Type: cross Abstract: Physics-informed diffusion models typically enforce PDE constraints only on final outputs, leaving intermediate representations unconstrained and pron

safetyarxiv-cs-cv
21 May 2026
Safety

Letting Trajectories Spread: Quality-Preserving Control for Diverse Flow Matching

DGX agent

arXiv:2510.09060v2 Announce Type: replace-cross Abstract: Flow-based text-to-image models follow deterministic trajectories, making it costly to explore diverse modes under limited sampling budgets. E

safetyarxiv-cs-cv
21 May 2026
Safety

Linear-DPO: Linear Direct Preference Optimization for Diffusion and Flow-Matching Generative Models

DGX agent

arXiv:2605.21123v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is successful for alignment in LLMs but still faces challenges in text-to-image generation. Existing studies are co

safetyarxiv-cs-cv
21 May 2026
Safety

Listwise Policy Optimization: Group-based RLVR as Target-Projection on the LLM Response Simplex

DGX agent

arXiv:2605.06139v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a standard approach for large language models (LLMs) post-training to incentivize r

safetyarxiv-cs-lg
21 May 2026
Safety

LLM Pretraining Shapes a Generalizable Manifold: Insights into Cross-Modal Transfer to Time Series

DGX agent

arXiv:2605.20449v1 Announce Type: new Abstract: Can language-pretrained transformers become effective time-series forecasters, and why? In this paper, we show that cross-modal transfer arises because

safetyarxiv-cs-lg
21 May 2026
Safety

longer comment on yesterday’s stories

DGX agent

longer comment on yesterday’s stories the nuance between yesterday’s stories that X influencers won’t tell you: https://open.substack.com/pub/garymarcus/p/checking-the-math-behind-openai-and?r=8tdk6&u

safetygary-marcus--x
21 May 2026
Safety

Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX

DGX agent

arXiv:2605.20577v1 Announce Type: cross Abstract: Riichi Mahjong is a multi-player, imperfect-information game characterized by stochasticity and high-dimensional state spaces. These attributes presen

safetyarxiv-cs-lg
21 May 2026
Safety

Mind the Sim-to-Real Gap & Think Like a Scientist

DGX agent

arXiv:2605.21458v1 Announce Type: cross Abstract: Suppose a planner has a pre-trained simulator of a sequential decision problem and the option to run real experiments in the field. The simulator is c

safetyarxiv-cs-lg
21 May 2026
Safety

Mitigating Label Bias with Interpretable Rubric Embeddings

DGX agent

arXiv:2605.21455v1 Announce Type: new Abstract: Statistical decision algorithms are increasingly deployed in domains where ground-truth labels are hard to obtain, such as hiring, university admissions

safetyarxiv-cs-lg
21 May 2026
Safety

Mobile UMI: Cross-View Diffusion Policy with Decoupled Kinematics for Mobile Manipulation

DGX agent

arXiv:2605.20894v1 Announce Type: new Abstract: Mobile imitation learning on portable demonstration interfaces faces two coupled bottlenecks: locomotion-contaminated action labels and inference-induce

safetyarxiv-cs-ro
21 May 2026
Safety

Multi-Head Attention as Ensemble Nadaraya-Watson Estimation: Variance Reduction, Decorrelation, and Optimal Head Diversity

DGX agent

arXiv:2605.20271v1 Announce Type: cross Abstract: We develop a rigorous statistical theory of multi-head attention (MHA) as an ensemble of Nadaraya-Watson (NW) kernel regression estimators. Building o

safetyarxiv-cs-lg
21 May 2026
Safety

Multi-Step Likelihood-Ratio Correction for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2605.20865v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) plays a pivotal role in improving the reasoning ability of large language models. However, widely

safetyarxiv-cs-lg
21 May 2026
Safety

Multimodal LLMs under Pairwise Modalities

DGX agent

arXiv:2605.21059v1 Announce Type: new Abstract: Despite the impressive results achieved by multimodal large language models (MLLMs), their training typically relies on jointly curated multimodal data,

safetyarxiv-cs-cv
21 May 2026
Safety

NaP-Control: Navigating Diffusion Prior for Versatile and Fast Character Control

DGX agent

arXiv:2605.20209v1 Announce Type: cross Abstract: Achieving precise, versatile whole-body character control in physics-based animation remains challenging. Recent diffusion-based policies generate ric

safetyarxiv-cs-lg
21 May 2026
Safety

Neural Collapse by Design: Learning Class Prototypes on the Hypersphere

DGX agent

arXiv:2605.20302v1 Announce Type: cross Abstract: Supervised classification has a theoretical optimum, Neural Collapse (NC), yet neither of its two dominant paradigms reaches it in practice. Cross ent

safetyarxiv-cs-cv
21 May 2026
Safety

nice measure of how insane SpaceX’s claims are

DGX agent

nice measure of how insane SpaceX’s claims are I don't offer investment advice, but when I see that SpaceX claims that their total addressable market is the size of the *complete* US economy (with eve

safetygary-marcus--x
21 May 2026
Safety

nor do we know how the (new) model works nor how it does on anything else nor how it was trained. scientists wait for facts; cheerleaders (o…

DGX agent

nor do we know how the (new) model works nor how it does on anything else nor how it was trained. scientists wait for facts; cheerleaders (over and over) rush to judgments that have often been wrong.

safetygary-marcus--x
21 May 2026
Safety

note that @jeffBezos is not betting on chatbots, but rather focusing on domain-specific tools for engineering, which is also something I wou…

DGX agent

note that @jeffBezos is not betting on chatbots, but rather focusing on domain-specific tools for engineering, which is also something I would encourage. the financial story for what he is doing may b

safetygary-marcus--x
21 May 2026
Safety

OcclusionFormer: Arranging Z-Order for Layout-Grounded Image Generation

DGX agent

arXiv:2605.21343v1 Announce Type: new Abstract: Recent layout-to-image models have achieved remarkable progress in spatial controllability. However, they still struggle with inter-object occlusion. Wh

safetyarxiv-cs-cv
21 May 2026
Safety

👀OpenAI Q1 operating margin was -122% even when excluding SBC and other items.

DGX agent

OpenAI reported a negative 122% operating margin in Q1, indicating the company spent significantly more on operations than it generated in revenue, even after excluding stock-based compensation and ot

safetygary-marcus--x
21 May 2026
Safety

our company, which lost money last quarter, is going to take a big slice of $28.5 trillion market that we project will exist, primarily in a…

DGX agent

our company, which lost money last quarter, is going to take a big slice of 28.5 trillion market that we project will exist, primarily in a field in which we have almost no current presence, which is

safetygary-marcus--x
21 May 2026
Safety

Parallel LLM Reasoning for Bias-Resilient, Robust Conceptual Abstraction

DGX agent

arXiv:2605.20194v1 Announce Type: new Abstract: Large language models (LLMs) have been increasingly used to analyze text. However, they are often plagued with contextual reasoning limitations when ana

safetyarxiv-cs-cl
21 May 2026
Safety

Pareto-Enhanced Portrait Generation: Vision-Aligned Text Supervision for Alignment, Realism, and Aesthetics

DGX agent

arXiv:2605.20640v1 Announce Type: new Abstract: Text-to-image diffusion models often face a severe trilemma in human portrait generation: text-image alignment, photorealism, and human-perceived aesthe

safetyarxiv-cs-cv
21 May 2026
Safety

PointACT: Vision-Language-Action Models with Multi-Scale Point-Action Interaction

DGX agent

arXiv:2605.21414v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general-purpose robotic manipulation by leveraging large pretrained vision-languag

safetyarxiv-cs-cv
21 May 2026
Safety

Pretty profound for an ipo filing: “For the entirety of its existence, human civilization has lived on a single celestial body: Earth. The c…

DGX agent

Pretty profound for an ipo filing: “For the entirety of its existence, human civilization has lived on a single celestial body: Earth. The current paradigm, in which human civilization is confined to

safetyelon-musk--x
21 May 2026
Safety

Principled RL for Flow Matching Emerges from the Chunk-level Policy Optimization

DGX agent

arXiv:2510.21583v2 Announce Type: replace Abstract: Recent Progress in post-training flow matching for text-to-image (T2I) generation with Group Relative Policy Optimization (GRPO) has demonstrated st

safetyarxiv-cs-cv
21 May 2026
Safety

Q-SpiRL: Quantum Spiking Reinforcement Learning for Adaptive Robot Navigation

DGX agent

arXiv:2605.20801v1 Announce Type: new Abstract: Adaptive robot navigation in dynamic environments requires policies that can reach the target reliably while producing efficient and stable trajectories

safetyarxiv-cs-ro
21 May 2026
Safety

Quantum End-to-End Learning for Contextual Combinatorial Optimization

DGX agent

arXiv:2605.20222v1 Announce Type: cross Abstract: Contextual combinatorial optimization (CCO) plays a critical role in decision-making under uncertainty, yet remains a significant challenge. We presen

safetyarxiv-cs-lg
21 May 2026
Safety

Regulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography Analysis

DGX agent

arXiv:2605.20277v1 Announce Type: new Abstract: Medical vision-language models (VLMs) have rapidly advanced as general-purpose multimodal assistants, yet their deployment in 3D Computed Tomography (CT

safetyarxiv-cs-cv
21 May 2026
← Previous
1…193194195196197…302
Next →