AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
11 May 2026

Compute is scarce and has to be bought years in advance, before companies know whether revenue will ever catch up. That is exactly the natur…

SafetyDGX agent

Compute is scarce and has to be bought years in advance, before companies know whether revenue will ever catch up. That is exactly the nature of the largest gamble in history. Heaven help the global e

Conditional generation of antibody sequences with classifier-guided germline-absorbing discrete diffusion

SafetyDGX agent

arXiv:2605.06720v1 Announce Type: cross Abstract: Antibody therapeutics are among the most successful modern medicines, yet computationally designing antibodies with desirable binding and developabili

Confidence-Aware Alignment Makes Reasoning LLMs More Reliable

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.07353v1 Announce Type: new Abstract: Large reasoning models often reach correct answers through flawed intermediate steps, creating a gap between final accuracy and reasoning reliability. E

Conformal-Style Quantile Analyses for Stochastic Bandits

SafetyDGX agent

arXiv:2605.07115v1 Announce Type: new Abstract: Stochastic bandit algorithms are usually analyzed under a mean-reward criterion, yet many problems favor arms with strong upper-tail performance, which

Contact-Grounded Policy: Dexterous Visuotactile Policy with Generative Contact Grounding

SafetyDGX agent

arXiv:2603.05687v3 Announce Type: replace Abstract: Contact-rich dexterous manipulation with multi-finger hands remains an open challenge in robotics because task success depends on multi-point contac

Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy

SafetyDGX agent

arXiv:2605.07171v1 Announce Type: new Abstract: The classic multi-armed bandit (MAB) problem tackles the challenge of accruing maximum reward while making decisions under uncertainty. However, in appl

Curated Synthetic Data Doesn't Have to Collapse: A Theoretical Study of Generative Retraining with Pluralistic Preferences

SafetyDGX agent

arXiv:2605.07724v1 Announce Type: cross Abstract: Recursive retraining of generative models poses a critical representation challenge: when synthetic outputs are curated based on a fixed reward signal

DCGL: Dual-Channel Graph Learning with Large Language Models for Knowledge-Aware Recommendation

SafetyDGX agent

arXiv:2605.07314v1 Announce Type: cross Abstract: Knowledge Graphs (KGs) have proven highly effective for recommendation systems by capturing latent item relationships, while recent integration of Lar

Dear @geoffreyhinton, I literally never said that AI systems “JUST regurgitate”; that’s plainly false. I don’t believe it, and I didn’t say …

SafetyDGX agent

Dear @geoffreyhinton, I literally never said that AI systems “JUST regurgitate”; that’s plainly false. I don’t believe it, and I didn’t say it. (They do *sometimes* regurgitate, and the evidence for t

Decentralized Time-Varying Optimization for Streaming Data via Temporal Weighting

SafetyDGX agent

arXiv:2605.06971v1 Announce Type: cross Abstract: Classical optimization theory largely focuses on fixed objective functions, whereas many modern learning systems operate in dynamic environments where

Decoupling Semantics and Fingerprints: A Universal Representation for AI-Generated Image Detection

SafetyDGX agent

arXiv:2605.07074v1 Announce Type: new Abstract: Detecting AI-generated images across unseen architectures remains challenging, as existing models often overfit to generator-specific fingerprints and s

DiffeoMorph: Learning to Morph 3D Shapes Using Differentiable Agent-Based Simulations

SafetyDGX agent

arXiv:2512.17129v2 Announce Type: replace Abstract: Biological systems can form complex three-dimensional structures through the collective behavior of agents that share a common update rule and opera

Differentially Private Auditing Under Strategic Response

SafetyDGX agent

arXiv:2605.07674v1 Announce Type: cross Abstract: Regulatory audits of AI systems increasingly rely on differential privacy (DP) to protect training data and model internals. We study audit design whe

Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers

SafetyDGX agent

arXiv:2605.07503v1 Announce Type: new Abstract: Efficiently aligning large-scale video diffusion models with human intent requires a scalable and trajectory-aware pathway that bridges the inherent dis

Direction-Flipped Influence Audits Reveal Hidden Structure in Moral Choices of LLMs

SafetyDGX agent

arXiv:2602.22831v2 Announce Type: replace-cross Abstract: Moral benchmarks for LLMs typically score models on context-free prompts, implicitly treating the measured choice rate as stable. We test this

Discovering Multiagent Learning Algorithms with Large Language Models

SafetyDGX agent

arXiv:2602.16928v3 Announce Type: replace-cross Abstract: Much of the advancement in Multi-Agent Reinforcement Learning (MARL) for imperfect-information games has historically depended on the manual,

Does Your Neural Network Extrapolate? Feature Engineering as Identifiability Bias for OOD Generalization

SafetyDGX agent

arXiv:2605.07483v1 Announce Type: cross Abstract: Successful deep neural networks discover salient features of data. We show when and why they fail to learn out-of-distribution (OOD)-relevant represen

Don't Retrain, Align: Adapting Autoregressive LMs to Diffusion LMs via Representation Alignment

SafetyDGX agent

arXiv:2605.06885v1 Announce Type: cross Abstract: Diffusion language models (DLMs) have recently demonstrated capabilities that complement standard autoregressive (AR) models, particularly in non-sequ

Dr-BA: Separable Optimization for Direct Radar Bundle Adjustment & Localization

SafetyDGX agent

arXiv:2605.07041v1 Announce Type: cross Abstract: This paper introduces Dr-BA, a first-of-its-kind radar bundle adjustment (BA) framework that operates directly on 2D spinning radar intensity images.

Dr. Post-Training: A Data Regularization Perspective on LLM Post-Training

SafetyDGX agent

arXiv:2605.07063v1 Announce Type: cross Abstract: Data selection methods address a critical challenge in LLM post-training: effectively leveraging scarce, high-fidelity target data alongside abundant

DReS: Dual Reconstruction Smoothing for Functional Regularization

SafetyDGX agent

arXiv:2510.00253v2 Announce Type: replace Abstract: Smoothness is a key inductive bias in machine learning and is closely related to generalization. Existing smoothness-inducing methods typically rely

Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow

SafetyDGX agent

arXiv:2605.07727v1 Announce Type: cross Abstract: We propose Drifting Field Policy (DFP), a non-ODE one-step generative policy built on the drifting model paradigm. We frame the policy update as a rev

Dynamic one-time delivery of critical data by small and sparse UAV swarms: a model problem for MARL scaling studies

SafetyDGX agent

arXiv:2512.09682v2 Announce Type: replace-cross Abstract: This work studies the application of Multi-Agent Reinforcement Learning (MARL) to decentralized control of unmanned aerial vehicles to relay a

Econometric vs. Causal Structure-Learning for Time-Series Policy Decisions: Evidence from the UK COVID-19 Policies

SafetyDGX agent

arXiv:2603.00041v2 Announce Type: replace-cross Abstract: Causal machine learning (ML) recovers graphical structures that inform us about potential cause-and-effect relationships. Most progress has fo

EditTransfer++: Toward Faithful and Efficient Visual-Prompt-Guided Image Editing

SafetyDGX agent

arXiv:2605.07455v1 Announce Type: new Abstract: Visual-prompt-guided edit transfer aims to learn image transformations directly from example pairs, offering more precise and controllable editing than

EGA: Adapting Frozen Encoders for Vector Search with Bounded Out-of-Distribution Degradation

SafetyDGX agent

arXiv:2605.05674v2 Announce Type: replace-cross Abstract: Vector search systems built on frozen vision encoders face queries from unseen classes at deployment, yet existing adapter training collapses

Emergent social transmission of model-based representations without inference

SafetyDGX agent

arXiv:2604.05777v2 Announce Type: replace Abstract: How do people acquire rich, flexible knowledge about their environment from others despite limited cognitive capacity? Humans are often thought to r

Emergent Symbolic Structure in Health Foundation Models: Extraction, Alignment, and Cross-Modal Transfer

SafetyDGX agent

arXiv:2605.07407v1 Announce Type: new Abstract: Health foundation models (FMs) learn useful representations from wearable sensors, but interpreting what they encode and transferring that knowledge acr

Entropy-Regularized Adjoint Matching for Offline Reinforcement Learning

SafetyDGX agent

arXiv:2605.06156v2 Announce Type: replace-cross Abstract: Integrating expressive generative policies, such as flow-matching models, into offline reinforcement learning (RL) allows agents to capture co

Epistemological standards in SF venture capital are so low that you can write this, and everyone will simply nod along. The actual evidence,…

SafetyDGX agent

Epistemological standards in SF venture capital are so low that you can write this, and everyone will simply nod along. The actual evidence, including a recent NBER paper on thousands of firms, finds:

Escaping the Diversity Trap in Robotic Manipulation via Anchor-Centric Adaptation

SafetyDGX agent

arXiv:2605.07381v1 Announce Type: cross Abstract: While Vision-Language-Action (VLA) models offer broad general capabilities, deploying them on specific hardware requires real-world adaptation to brid

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle

SafetyDGX agent

arXiv:2510.16079v2 Announce Type: replace-cross Abstract: Current Large Language Model (LLM) agents show strong performance in tool use, but lack the crucial capability to systematically learn from th

Ex-OpenAI co-founder Ilya Sutskever claims to have collected proof of Sam Altman’s alleged dishonesty https://www.theglobeandmail.com/busine…

SafetyDGX agent

Ex-OpenAI co-founder Ilya Sutskever claims to have collected proof of Sam Altman’s alleged dishonesty https://www.theglobeandmail.com/business/article-openai-trial-sutskever-sam-altman-musk/?utm_sourc

Excellent explainer video by @FryRsquared on the risks of AI agents. She also raises a crucial point: we shouldn’t make the mistake of think…

SafetyDGX agent

Excellent explainer video by @FryRsquared on the risks of AI agents. She also raises a crucial point: we shouldn’t make the mistake of thinking current limitations will necessarily persist. As we’ve s

Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models

SafetyDGX agent

arXiv:2605.07244v1 Announce Type: cross Abstract: We introduce Mutual Reinforcement Learning, a framework for concurrent RL post-training in which heterogeneous LLM policies exchange typed experience

f-Divergence Regularized RLHF: Two Tales of Sampling and Unified Analyses

SafetyDGX agent

arXiv:2605.06977v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) has become a cornerstone technique for post-training large language models. While most existing appr

Flatness and Gradient Alignment Are Both Necessary: Spectral-Aware Gradient-Aligned Exploration for Multi-Distribution Learning

SafetyDGX agent

arXiv:2605.07914v1 Announce Type: cross Abstract: Sharpness-aware and gradient-alignment methods have been shown to improve generalization, however each family of methods targets a single geometric pr

Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective

SafetyDGX agent

arXiv:2602.09782v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a critical method for enhancing the reasoning capabilities of Large Langu

Flow-OPD: On-Policy Distillation for Flow Matching Models

SafetyDGX agent

arXiv:2605.08063v1 Announce Type: cross Abstract: Existing Flow Matching (FM) text-to-image models suffer from two critical bottlenecks under multi-task alignment: the reward sparsity induced by scala

Fortifying Time Series: DTW-Certified Robust Anomaly Detection

SafetyDGX agent

arXiv:2605.07690v1 Announce Type: new Abstract: Time-series anomaly detection is critical for ensuring safety in high-stakes applications, where robustness is a fundamental requirement rather than a m

From 0-Order Selection to 2-Order Judgment: Combinatorial Hardening Exposes Compositional Failures in Frontier LLMs

SafetyDGX agent

arXiv:2605.07268v1 Announce Type: new Abstract: Multiple-choice reasoning benchmarks face dual challenges: rapid saturation from advancing models and data contamination that undermines static evaluati

From Assistance to Agency: Rethinking Autonomy and Control in CI/CD Pipelines

SafetyDGX agent

arXiv:2605.07062v1 Announce Type: cross Abstract: AI agents are assuming active roles in Continuous Integration and Continuous Deployment (CI/CD) workflows, yet the research community lacks a shared v

From Model to Data (M2D): Shifting Complexity from GNNs to Graphs for Transparent Graph Learning

SafetyDGX agent

arXiv:2605.06814v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve high performance but can be opaque to humans, making it difficult to understand and compare the many proposed archi

From Specification to Deployment: Empirical Evidence from a W3C VC + DID Trust Infrastructure for Autonomous Agents

SafetyDGX agent

arXiv:2605.06738v1 Announce Type: cross Abstract: Autonomous AI agents now transact at production scale -- 69,000 bots executing 165 million transactions across 50 million USDC in cumulative volume on

From Time Series Analysis to Question Answering: A Survey in the LLM Era

SafetyDGX agent

arXiv:2506.11512v2 Announce Type: replace-cross Abstract: Recently, Large Language Models (LLMs) have introduced a novel paradigm in Time Series Analysis (TSA), leveraging strong language capabilities

Future-proof your data strategy: AlloyDB adds PostgreSQL 18 and new Extended Support

SafetyDGX agent

As you look out at your 2026 infrastructure roadmap, your goal is to balance the need for rapid innovation with operational stability. You shouldn't have to choose between adopting the latest database

GASim: A Graph-Accelerated Hybrid Framework for Social Simulation

SafetyDGX agent

arXiv:2605.07692v1 Announce Type: new Abstract: Large-scale social simulators are essential for studying complex social patterns. Prior work explores hybrid methods to scale up simulations, combining

Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions

SafetyDGX agent

arXiv:2512.20974v3 Announce Type: replace-cross Abstract: Bayesian Reinforcement Learning (BRL), a subclass of Meta-Reinforcement Learning (Meta-RL), provides a principled framework for generalisation

🤔Going to be in Vancouver tomorrow? 🧐Curious about whether or not scaling AI is a bad bet? 😎Catch TPN founder @zacharykarabell's chat wit…

SafetyDGX agent

🤔Going to be in Vancouver tomorrow? 🧐Curious about whether or not scaling AI is a bad bet? 😎Catch TPN founder @zacharykarabell's chat with @GaryMarcus at @WebSummit at 10:20am. 🎟️Get your tickets here

Gonna keep retweeting, because this flawed move never dies:

SafetyDGX agent

Gonna keep retweeting, because this flawed move never dies: AI hype has become a giant game of bait and switch. the bait: we are going to make an AI that can solve any problem a expert human could sol

GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization

SafetyDGX agent

arXiv:2605.07399v1 Announce Type: new Abstract: Diffusion Vision-Language Models (dVLMs), built upon the non-causal foundations of Diffusion Large Language Models (dLLMs), have demonstrated remarkable

GRAPE: Let GRPO Supervise Query Rewriting by Ranking for Retrieval

SafetyDGX agent

arXiv:2509.23370v2 Announce Type: replace Abstract: The CLIP model has established itself as a cornerstone of large-scale retrieval systems. However, its performance often degrades under distributiona

Guidance Is Not a Hyperparameter: Learning Dynamic Control in Diffusion Language Models

SafetyDGX agent

arXiv:2605.07701v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used mechanism for controlling diffusion-based generative models, yet its guidance scale is typically treated

GustPilot: A Hierarchical DRL-INDI Framework for Wind-Resilient Quadrotor Navigation

SafetyDGX agent

arXiv:2603.19966v2 Announce Type: replace Abstract: Wind disturbances remain a key barrier to reliable autonomous navigation for lightweight quadrotors, where the rapidly varying airflow can destabili

HAIC: Humanoid Agile Object Interaction Control via Dynamics-Aware World Model

SafetyDGX agent

arXiv:2602.11758v2 Announce Type: replace Abstract: Humanoid robots show promise for complex whole-body tasks in unstructured environments. Although Human-Object Interaction (HOI) has advanced, most m

Hard to Read, Easy to Jailbreak: How Visual Degradation Bypasses MLLM Safety Alignment

SafetyDGX agent

arXiv:2605.07250v1 Announce Type: cross Abstract: Recent advancements in visual context compression enable MLLMs to process ultra-long contexts efficiently by rendering text into images. However, we i

HEART: Hyperspherical Embedding Alignment via Kent-Representation Traversal in Diffusion Models

SafetyDGX agent

arXiv:2605.07973v1 Announce Type: new Abstract: Text-to-image diffusion models can generate visually stunning images, yet, controlling what appears and how it appears, remains surprisingly difficult,

Hidden Coalitions in Multi-Agent AI: A Spectral Diagnostic from Internal Representations

SafetyDGX agent

arXiv:2605.06696v1 Announce Type: new Abstract: Collections of interacting AI agents can form coalitions, creating emergent group-level organization that is critical for AI safety and alignment. Howev

How Log-Barrier Helps Exploration in Policy Optimization

SafetyDGX agent

arXiv:2603.15001v2 Announce Type: replace-cross Abstract: Recently, it has been shown that the Stochastic Gradient Bandit (SGB) algorithm converges to a globally optimal policy with a constant learnin

How to Compress KV Cache in RL Post-Training? Shadow Mask Distillation for Memory-Efficient Alignment

SafetyDGX agent

arXiv:2605.06850v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a crucial paradigm for unlocking the advanced reasoning capabilities of Large Language Models (LLMs), encom

← Previous
1…154155156157158…212
Next →