AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,806 results
Safety

Decoupling Semantics and Fingerprints: A Universal Representation for AI-Generated Image Detection

DGX agent

arXiv:2605.07074v1 Announce Type: new Abstract: Detecting AI-generated images across unseen architectures remains challenging, as existing models often overfit to generator-specific fingerprints and s

safetyarxiv-cs-cv
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

DiffeoMorph: Learning to Morph 3D Shapes Using Differentiable Agent-Based Simulations

DGX agent

arXiv:2512.17129v2 Announce Type: replace Abstract: Biological systems can form complex three-dimensional structures through the collective behavior of agents that share a common update rule and opera

safetyarxiv-cs-lg
11 May 2026
Safety

Differentially Private Auditing Under Strategic Response

DGX agent

arXiv:2605.07674v1 Announce Type: cross Abstract: Regulatory audits of AI systems increasingly rely on differential privacy (DP) to protect training data and model internals. We study audit design whe

safetyarxiv-cs-lg
11 May 2026
Safety

Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers

DGX agent

arXiv:2605.07503v1 Announce Type: new Abstract: Efficiently aligning large-scale video diffusion models with human intent requires a scalable and trajectory-aware pathway that bridges the inherent dis

safetyarxiv-cs-cv
11 May 2026
Safety

Direction-Flipped Influence Audits Reveal Hidden Structure in Moral Choices of LLMs

DGX agent

arXiv:2602.22831v2 Announce Type: replace-cross Abstract: Moral benchmarks for LLMs typically score models on context-free prompts, implicitly treating the measured choice rate as stable. We test this

safetyarxiv-cs-ai
11 May 2026
Safety

Discovering Multiagent Learning Algorithms with Large Language Models

DGX agent

arXiv:2602.16928v3 Announce Type: replace-cross Abstract: Much of the advancement in Multi-Agent Reinforcement Learning (MARL) for imperfect-information games has historically depended on the manual,

safetyarxiv-cs-ai
11 May 2026
Safety

Does Your Neural Network Extrapolate? Feature Engineering as Identifiability Bias for OOD Generalization

DGX agent

arXiv:2605.07483v1 Announce Type: cross Abstract: Successful deep neural networks discover salient features of data. We show when and why they fail to learn out-of-distribution (OOD)-relevant represen

safetyarxiv-cs-ai
11 May 2026
Safety

Don't Retrain, Align: Adapting Autoregressive LMs to Diffusion LMs via Representation Alignment

DGX agent

arXiv:2605.06885v1 Announce Type: cross Abstract: Diffusion language models (DLMs) have recently demonstrated capabilities that complement standard autoregressive (AR) models, particularly in non-sequ

safetyarxiv-cs-ai
11 May 2026
Safety

Dr-BA: Separable Optimization for Direct Radar Bundle Adjustment & Localization

DGX agent

arXiv:2605.07041v1 Announce Type: cross Abstract: This paper introduces Dr-BA, a first-of-its-kind radar bundle adjustment (BA) framework that operates directly on 2D spinning radar intensity images.

safetyarxiv-cs-cv
11 May 2026
Safety

Dr. Post-Training: A Data Regularization Perspective on LLM Post-Training

DGX agent

arXiv:2605.07063v1 Announce Type: cross Abstract: Data selection methods address a critical challenge in LLM post-training: effectively leveraging scarce, high-fidelity target data alongside abundant

safetyarxiv-cs-ai
11 May 2026
Safety

DReS: Dual Reconstruction Smoothing for Functional Regularization

DGX agent

arXiv:2510.00253v2 Announce Type: replace Abstract: Smoothness is a key inductive bias in machine learning and is closely related to generalization. Existing smoothness-inducing methods typically rely

safetyarxiv-cs-lg
11 May 2026
Safety

Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow

DGX agent

arXiv:2605.07727v1 Announce Type: cross Abstract: We propose Drifting Field Policy (DFP), a non-ODE one-step generative policy built on the drifting model paradigm. We frame the policy update as a rev

safetyarxiv-cs-ai
11 May 2026
Safety

Dynamic one-time delivery of critical data by small and sparse UAV swarms: a model problem for MARL scaling studies

DGX agent

arXiv:2512.09682v2 Announce Type: replace-cross Abstract: This work studies the application of Multi-Agent Reinforcement Learning (MARL) to decentralized control of unmanned aerial vehicles to relay a

safetyarxiv-cs-ai
11 May 2026
Safety

Econometric vs. Causal Structure-Learning for Time-Series Policy Decisions: Evidence from the UK COVID-19 Policies

DGX agent

arXiv:2603.00041v2 Announce Type: replace-cross Abstract: Causal machine learning (ML) recovers graphical structures that inform us about potential cause-and-effect relationships. Most progress has fo

safetyarxiv-cs-ai
11 May 2026
Safety

EditTransfer++: Toward Faithful and Efficient Visual-Prompt-Guided Image Editing

DGX agent

arXiv:2605.07455v1 Announce Type: new Abstract: Visual-prompt-guided edit transfer aims to learn image transformations directly from example pairs, offering more precise and controllable editing than

safetyarxiv-cs-cv
11 May 2026
Safety

EGA: Adapting Frozen Encoders for Vector Search with Bounded Out-of-Distribution Degradation

DGX agent

arXiv:2605.05674v2 Announce Type: replace-cross Abstract: Vector search systems built on frozen vision encoders face queries from unseen classes at deployment, yet existing adapter training collapses

safetyarxiv-cs-ai
11 May 2026
Safety

Emergent social transmission of model-based representations without inference

DGX agent

arXiv:2604.05777v2 Announce Type: replace Abstract: How do people acquire rich, flexible knowledge about their environment from others despite limited cognitive capacity? Humans are often thought to r

safetyarxiv-cs-ai
11 May 2026
Safety

Emergent Symbolic Structure in Health Foundation Models: Extraction, Alignment, and Cross-Modal Transfer

DGX agent

arXiv:2605.07407v1 Announce Type: new Abstract: Health foundation models (FMs) learn useful representations from wearable sensors, but interpreting what they encode and transferring that knowledge acr

safetyarxiv-cs-lg
11 May 2026
Safety

Entropy-Regularized Adjoint Matching for Offline Reinforcement Learning

DGX agent

arXiv:2605.06156v2 Announce Type: replace-cross Abstract: Integrating expressive generative policies, such as flow-matching models, into offline reinforcement learning (RL) allows agents to capture co

safetyarxiv-cs-ai
11 May 2026
Safety

Epistemological standards in SF venture capital are so low that you can write this, and everyone will simply nod along. The actual evidence,…

DGX agent

Epistemological standards in SF venture capital are so low that you can write this, and everyone will simply nod along. The actual evidence, including a recent NBER paper on thousands of firms, finds:

safetygary-marcus--x
11 May 2026
Safety

Escaping the Diversity Trap in Robotic Manipulation via Anchor-Centric Adaptation

DGX agent

arXiv:2605.07381v1 Announce Type: cross Abstract: While Vision-Language-Action (VLA) models offer broad general capabilities, deploying them on specific hardware requires real-world adaptation to brid

safetyarxiv-cs-ai
11 May 2026
Safety

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle

DGX agent

arXiv:2510.16079v2 Announce Type: replace-cross Abstract: Current Large Language Model (LLM) agents show strong performance in tool use, but lack the crucial capability to systematically learn from th

safetyarxiv-cs-ai
11 May 2026
Safety

Ex-OpenAI co-founder Ilya Sutskever claims to have collected proof of Sam Altman’s alleged dishonesty https://www.theglobeandmail.com/busine…

DGX agent

Ex-OpenAI co-founder Ilya Sutskever claims to have collected proof of Sam Altman’s alleged dishonesty https://www.theglobeandmail.com/business/article-openai-trial-sutskever-sam-altman-musk/?utm_sourc

safetygary-marcus--x
11 May 2026
Safety

Excellent explainer video by @FryRsquared on the risks of AI agents. She also raises a crucial point: we shouldn’t make the mistake of think…

DGX agent

Excellent explainer video by @FryRsquared on the risks of AI agents. She also raises a crucial point: we shouldn’t make the mistake of thinking current limitations will necessarily persist. As we’ve s

safetyyoshua-bengio--x
11 May 2026
Safety

Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models

DGX agent

arXiv:2605.07244v1 Announce Type: cross Abstract: We introduce Mutual Reinforcement Learning, a framework for concurrent RL post-training in which heterogeneous LLM policies exchange typed experience

safetyarxiv-cs-ai
11 May 2026
Safety

f-Divergence Regularized RLHF: Two Tales of Sampling and Unified Analyses

DGX agent

arXiv:2605.06977v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) has become a cornerstone technique for post-training large language models. While most existing appr

safetyarxiv-cs-ai
11 May 2026
Safety

Flatness and Gradient Alignment Are Both Necessary: Spectral-Aware Gradient-Aligned Exploration for Multi-Distribution Learning

DGX agent

arXiv:2605.07914v1 Announce Type: cross Abstract: Sharpness-aware and gradient-alignment methods have been shown to improve generalization, however each family of methods targets a single geometric pr

safetyarxiv-cs-cv
11 May 2026
Safety

Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective

DGX agent

arXiv:2602.09782v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a critical method for enhancing the reasoning capabilities of Large Langu

safetyarxiv-cs-ai
11 May 2026
Safety

Flow-OPD: On-Policy Distillation for Flow Matching Models

DGX agent

arXiv:2605.08063v1 Announce Type: cross Abstract: Existing Flow Matching (FM) text-to-image models suffer from two critical bottlenecks under multi-task alignment: the reward sparsity induced by scala

safetyarxiv-cs-ai
11 May 2026
Safety

Fortifying Time Series: DTW-Certified Robust Anomaly Detection

DGX agent

arXiv:2605.07690v1 Announce Type: new Abstract: Time-series anomaly detection is critical for ensuring safety in high-stakes applications, where robustness is a fundamental requirement rather than a m

safetyarxiv-cs-lg
11 May 2026
Safety

From 0-Order Selection to 2-Order Judgment: Combinatorial Hardening Exposes Compositional Failures in Frontier LLMs

DGX agent

arXiv:2605.07268v1 Announce Type: new Abstract: Multiple-choice reasoning benchmarks face dual challenges: rapid saturation from advancing models and data contamination that undermines static evaluati

safetyarxiv-cs-cl
11 May 2026
Safety

From Assistance to Agency: Rethinking Autonomy and Control in CI/CD Pipelines

DGX agent

arXiv:2605.07062v1 Announce Type: cross Abstract: AI agents are assuming active roles in Continuous Integration and Continuous Deployment (CI/CD) workflows, yet the research community lacks a shared v

safetyarxiv-cs-ai
11 May 2026
Safety

From Model to Data (M2D): Shifting Complexity from GNNs to Graphs for Transparent Graph Learning

DGX agent

arXiv:2605.06814v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve high performance but can be opaque to humans, making it difficult to understand and compare the many proposed archi

safetyarxiv-cs-lg
11 May 2026
Safety

From Specification to Deployment: Empirical Evidence from a W3C VC + DID Trust Infrastructure for Autonomous Agents

DGX agent

arXiv:2605.06738v1 Announce Type: cross Abstract: Autonomous AI agents now transact at production scale -- 69,000 bots executing 165 million transactions across 50 million USDC in cumulative volume on

safetyarxiv-cs-ai
11 May 2026
Safety

From Time Series Analysis to Question Answering: A Survey in the LLM Era

DGX agent

arXiv:2506.11512v2 Announce Type: replace-cross Abstract: Recently, Large Language Models (LLMs) have introduced a novel paradigm in Time Series Analysis (TSA), leveraging strong language capabilities

safetyarxiv-cs-ai
11 May 2026
Safety

Future-proof your data strategy: AlloyDB adds PostgreSQL 18 and new Extended Support

DGX agent

As you look out at your 2026 infrastructure roadmap, your goal is to balance the need for rapid innovation with operational stability. You shouldn't have to choose between adopting the latest database

safetygoogle-cloud-ai
11 May 2026
Safety

GASim: A Graph-Accelerated Hybrid Framework for Social Simulation

DGX agent

arXiv:2605.07692v1 Announce Type: new Abstract: Large-scale social simulators are essential for studying complex social patterns. Prior work explores hybrid methods to scale up simulations, combining

safetyarxiv-cs-ai
11 May 2026
Safety

Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions

DGX agent

arXiv:2512.20974v3 Announce Type: replace-cross Abstract: Bayesian Reinforcement Learning (BRL), a subclass of Meta-Reinforcement Learning (Meta-RL), provides a principled framework for generalisation

safetyarxiv-cs-ai
11 May 2026
Safety

🤔Going to be in Vancouver tomorrow? 🧐Curious about whether or not scaling AI is a bad bet? 😎Catch TPN founder @zacharykarabell's chat wit…

DGX agent

🤔Going to be in Vancouver tomorrow? 🧐Curious about whether or not scaling AI is a bad bet? 😎Catch TPN founder @zacharykarabell's chat with @GaryMarcus at @WebSummit at 10:20am. 🎟️Get your tickets here

safetygary-marcus--x
11 May 2026
Safety

Gonna keep retweeting, because this flawed move never dies:

DGX agent

Gonna keep retweeting, because this flawed move never dies: AI hype has become a giant game of bait and switch. the bait: we are going to make an AI that can solve any problem a expert human could sol

safetygary-marcus--x
11 May 2026
Safety

GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization

DGX agent

arXiv:2605.07399v1 Announce Type: new Abstract: Diffusion Vision-Language Models (dVLMs), built upon the non-causal foundations of Diffusion Large Language Models (dLLMs), have demonstrated remarkable

safetyarxiv-cs-cv
11 May 2026
Safety

GRAPE: Let GRPO Supervise Query Rewriting by Ranking for Retrieval

DGX agent

arXiv:2509.23370v2 Announce Type: replace Abstract: The CLIP model has established itself as a cornerstone of large-scale retrieval systems. However, its performance often degrades under distributiona

safetyarxiv-cs-cv
11 May 2026
Safety

Guidance Is Not a Hyperparameter: Learning Dynamic Control in Diffusion Language Models

DGX agent

arXiv:2605.07701v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used mechanism for controlling diffusion-based generative models, yet its guidance scale is typically treated

safetyarxiv-cs-cl
11 May 2026
Safety

GustPilot: A Hierarchical DRL-INDI Framework for Wind-Resilient Quadrotor Navigation

DGX agent

arXiv:2603.19966v2 Announce Type: replace Abstract: Wind disturbances remain a key barrier to reliable autonomous navigation for lightweight quadrotors, where the rapidly varying airflow can destabili

safetyarxiv-cs-ro
11 May 2026
Safety

HAIC: Humanoid Agile Object Interaction Control via Dynamics-Aware World Model

DGX agent

arXiv:2602.11758v2 Announce Type: replace Abstract: Humanoid robots show promise for complex whole-body tasks in unstructured environments. Although Human-Object Interaction (HOI) has advanced, most m

safetyarxiv-cs-ro
11 May 2026
Safety

Hard to Read, Easy to Jailbreak: How Visual Degradation Bypasses MLLM Safety Alignment

DGX agent

arXiv:2605.07250v1 Announce Type: cross Abstract: Recent advancements in visual context compression enable MLLMs to process ultra-long contexts efficiently by rendering text into images. However, we i

safetyarxiv-cs-ai
11 May 2026
Safety

HEART: Hyperspherical Embedding Alignment via Kent-Representation Traversal in Diffusion Models

DGX agent

arXiv:2605.07973v1 Announce Type: new Abstract: Text-to-image diffusion models can generate visually stunning images, yet, controlling what appears and how it appears, remains surprisingly difficult,

safetyarxiv-cs-cv
11 May 2026
Safety

Hidden Coalitions in Multi-Agent AI: A Spectral Diagnostic from Internal Representations

DGX agent

arXiv:2605.06696v1 Announce Type: new Abstract: Collections of interacting AI agents can form coalitions, creating emergent group-level organization that is critical for AI safety and alignment. Howev

safetyarxiv-cs-ai
11 May 2026
← Previous
1…195196197198199…267
Next →