AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,351 results
28 Jul 2026

Looping Is Not Reliability: State-Bound Evidence and Typed Revision Contracts for Agentic Code Repair

SafetyDGX agent

arXiv:2607.24604v1 Announce Type: cross Abstract: Generate--test--revise loops are common in coding agents, but repetition alone provides no reliability guarantee. We study the gap between finding a c

Making Mathematical Knowledge Explainable, Accessible and Interoperable Through Large Language Model Integration

SafetyDGX agent

arXiv:2607.24512v1 Announce Type: new Abstract: Mathematical models are central to formalizing research problems, yet their documentation often falls short of FAIR principles. Knowledge bases such as

MemTX: Transactional Belief Commit for Stateful Agent Memory

SafetyDGX agent

arXiv:2607.23929v1 Announce Type: new Abstract: LLM agents increasingly coordinate through persistent shared memory: one agent's write becomes another agent's premise, and eventually a tool call with

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MobiWave: Dispatch-Oriented Graph Wavelets and Drift-Guided Selective Optimization for Autonomous Fleet Rebalancing

SafetyDGX agent

arXiv:2607.24365v1 Announce Type: new Abstract: Autonomous fleets enable mobility platforms to coordinate idle vehicles directly, making fleet-wide rebalancing possible. However, two obstacles limit r

Moral Hazard in Multi-Agent Language Models

SafetyDGX agent

arXiv:2607.23982v1 Announce Type: cross Abstract: Cooperation can fail when socially valuable effort is costly, weakly observable, and mainly benefits others. Drawing on Holmstrom's team moral-hazard

Open weights let builders own and control their stack, and provide flexibility to optimized quality and performance, instead of being locked…

SafetyDGX agent

Open weights let builders own and control their stack, and provide flexibility to optimized quality and performance, instead of being locked into a handful of closed platforms. We think that choice is

OpenAIs HealthBench in Action: Evaluating an LLM-Based Medical Assistant on Realistic Clinical Queries

Model ReleasesDGX agent

arXiv:2509.02594v3 Announce Type: replace-cross Abstract: Evaluating large language models (LLMs) on their ability to generate high-quality, accurate, situationally aware answers to clinical questions

Plato-Bio: verification-first biological novelty screening with temporal rediscovery and structural benchmarks

SafetyDGX agent

arXiv:2607.23975v1 Announce Type: new Abstract: Large language model research agents can connect literature retrieval, analysis code, and manuscript preparation, but coherent output does not establish

RMS@CC-MMD 2026: Multimodal Misogyny Detection via Geometric Interaction and Multi-View Consensus

SafetyDGX agent

arXiv:2607.22709v1 Announce Type: cross Abstract: The proliferation of internet memes has introduced new complexities to automated content moderation, particularly in detecting misogyny. Memes often r

Spectral Dynamics of Semantic Drift in Clinical Multi-Agent Language Model Networks

SafetyDGX agent

arXiv:2607.22758v1 Announce Type: cross Abstract: The integration of iterative LLMs within multi-agent diagnostic frameworks requires a rigorous quantitative reevaluation of underlying communication t

27 Jul 2026

A Consensus-Based Framework for Relative Preference Evaluation of Large Language Models

SafetyDGX agent

arXiv:2607.21632v1 Announce Type: new Abstract: Traditional benchmarks for LLMs primarily rely on static datasets and objective scoring metrics, which often fail to capture differences in response qua

創業以来、オープンソースコミュニティから多くを学び、また研究成果の公開を通じてそこに貢献してきました。オープンなエコシステムが健全なAI産業と技術主権を支える重要な基盤の一つであると考えており、その発展を支持します。 このたび、Sakana AIは、オープンウェイトAIモデルに関…

SafetyDGX agent

創業以来、オープンソースコミュニティから多くを学び、また研究成果の公開を通じてそこに貢献してきました。オープンなエコシステムが健全なAI産業と技術主権を支える重要な基盤の一つであると考えており、その発展を支持します。 このたび、Sakana AIは、オープンウェイトAIモデルに関する公開書簡 「Open Weights and American AI Leadership」に署名しました。 書簡は

Automatic Stability and Recovery for Neural Network Training

SafetyDGX agent

arXiv:2601.17483v2 Announce Type: replace Abstract: Training modern neural networks is increasingly fragile, with rare but severe destabilizing updates often causing irreversible divergence or silent

Design and Human Evaluation of Tactile Withdrawal Reflexes for a Skin-Covered Robot Arm

SafetyDGX agent

arXiv:2607.22249v1 Announce Type: new Abstract: Nociception is a protective biological mechanism that links harmful stimulation to a reaction. This paper investigates artificial nociception for a robo

Embodying Multi-Hand Manipulation Policies by Searching the Assignment and Null Spaces

SafetyDGX agent

arXiv:2607.22020v1 Announce Type: new Abstract: Learned manipulation policies increasingly predict motions for abstract 'hands' and are attractive in practice because they rely on easily collected dem

GRACE: Gradient-Free Robot Action Generation via Combined Diffusion-MPPI Posterior Mean Estimation

SafetyDGX agent

arXiv:2607.21661v1 Announce Type: new Abstract: Diffusion policies generate multimodal robot action sequences from demonstrations, but steering them toward deployment-time constraints typically relies

On the Identifiability of Controlled World Models

SafetyDGX agent

arXiv:2607.22430v1 Announce Type: new Abstract: Learning world models that infer environment dynamics from high-dimensional observations and predict outcomes under candidate actions is central to plan

Safe Learning Predictive Control for Ego-World Robotic Systems

SafetyDGX agent

arXiv:2607.22225v1 Announce Type: new Abstract: Safe autonomous navigation in shared environments requires the ability to anticipate and react to the latent behaviors of surrounding robots. In this pa

Security Without Detection: Economic Denial as a Primitive for Edge and IoT Defense

SafetyDGX agent

arXiv:2512.23849v2 Announce Type: replace-cross Abstract: Sophisticated attackers can evade detection-based security by using encryption, stealth tactics, and low-rate attack patterns. This challenge

25 Jul 2026

Cohere has proudly signed on to this letter. The importance of open-source models to the AI ecosystem cannot be understated. We believe ever…

SafetyDGX agent

Cohere has proudly signed on to this letter. The importance of open-source models to the AI ecosystem cannot be understated. We believe everyone, in every country, should have control over the technol

Wonderful to have a broad base acknowledgement of the need for open weights. We are delighted to support as signatories. Open weights is the…

SafetyDGX agent

Wonderful to have a broad base acknowledgement of the need for open weights. We are delighted to support as signatories. Open weights is the key to ensuring security, research, innovation and competit

24 Jul 2026

A Real-Time Generalized Nash Equilibrium Framework for Interaction-Aware Autonomous Driving in Mixed Traffic

SafetyDGX agent

arXiv:2607.21043v1 Announce Type: new Abstract: Safe and efficient navigation in mixed-traffic environments remains a critical challenge for Autonomous Vehicles (AVs), primarily due to the complex int

Adaptive Confidence-weighted Expansion for Trustworthy Multi-Omics Multimodal Fusion

SafetyDGX agent

arXiv:2607.20742v1 Announce Type: new Abstract: Multimodal learning is a robust approach to improve predictive performance in applications such as medical prognosis. However, the clinical applicabilit

Detectors Learn the Wrong Thing: Shortcut-Resistant Adversarial Training Against Physically Realizable Attacks

SafetyDGX agent

arXiv:2607.21243v1 Announce Type: new Abstract: AI-enabled visual perception systems are increasingly deployed in intelligent transportation infrastructure and autonomous vehicle related applications.

Euclid-MCP: A Model Context Protocol Server for Deterministic Logical Reasoning via Prolog

SafetyDGX agent

arXiv:2607.21412v1 Announce Type: new Abstract: Large Language Models (LLMs) excel at natural language understanding and generation but remain unreliable for multi-step logical reasoning, especially i

excellent

SafetyDGX agent

excellent For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen

Fizgig Krea 2 training features update

SafetyDGX agent

https://github.com/shootthesound/Fizgig Intelligent trainer - Per-image loss tracking with self-adapting training runs — every image gets its own verdict (easy / suspect / stuck / exhausted) and its o

GeoWorldAD: Geometry World Action Model for Autonomous Driving

SafetyDGX agent

arXiv:2607.17521v2 Announce Type: replace Abstract: Autonomous driving requires both safe and efficient planning decisions in dynamic 3D environments. Although recent Vision/Video-Action models learn

GOAT

SafetyDGX agent

GOAT For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safe

Grasp, Handover, Rotate: Bimanual Object Reorientation via Compositional Diffusion and Energy-Based Optimization

SafetyDGX agent

arXiv:2607.21341v1 Announce Type: new Abstract: Bimanual object reorientation - picking an object, handing it over between two arms, and placing it in a desired target pose - is valuable when direct p

Human-Inspired Framework for Robotic Craniotomy: Integrating Multimodal Fusion and Adaptive Trajectory Adjustment

SafetyDGX agent

arXiv:2607.21058v1 Announce Type: new Abstract: Manual craniotomy is a high-risk, skill-dependent procedure associated with surgeon fatigue and potential dural injury. While robotic approaches have im

Hybrid MKNF with Classical Negation in the Rule Component

SafetyDGX agent

arXiv:2607.21202v1 Announce Type: cross Abstract: Hybrid MKNF knowledge bases under the well-founded semantics integrate Description Logics with Logic Programming. However, they do not support classic

i want the US to win in AI both in open source and proprietary models, and i am glad to see this

SafetyDGX agent

i want the US to win in AI both in open source and proprietary models, and i am glad to see this For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform eve

NVIDIA, Palantir, Replit, Microsoft, Crowdstrike, Dell and others send a strong message to congress to keep open access to open weights mode…

SafetyDGX agent

NVIDIA, Palantir, Replit, Microsoft, Crowdstrike, Dell and others send a strong message to congress to keep open access to open weights models. 'Our AI leadership will be judged not by one frontier AI

Open models for the win!

SafetyDGX agent

Open models for the win! For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open mo

Open weight models will ensure that the entire world benefits from AI growth, and that America does not get left behind

SafetyDGX agent

Open weight models will ensure that the entire world benefits from AI growth, and that America does not get left behind For my first post, I’m sharing a letter @NVIDIA signed on why open models matter

RL-MACRO: A Cybernetic Closed-Loop Intelligence Framework for Multimodal Adaptive Robotic Craniotomy

SafetyDGX agent

arXiv:2607.21113v1 Announce Type: new Abstract: Autonomous robotic craniotomy requires continuous regulation of tool-tissue interactions to mitigate mechanical overload and thermal damage while mainta

SafeStep: AI-powered Travel Assistance for Elderly People with Frailty or Dementia

SafetyDGX agent

arXiv:2607.21156v1 Announce Type: new Abstract: More than a million people in the UK suffer from frailty or dementia, which severely compromise their ability to travel in urban environments. This pape

Stay hungry, stay foolish. Absolute legend

SafetyDGX agent

Stay hungry, stay foolish. Absolute legend For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by ever

thaulab@EEUCA 2026: Who Said What to Whom? A Targeting-Aware Neural-Symbolic Pipeline for Gaming Toxicity Detection

SafetyDGX agent

arXiv:2607.20447v1 Announce Type: new Abstract: This paper describes our system for the EEUCA 2026 Shared Task on toxicity classification in gaming chat. We implement a three-stage pipeline combining

The knowledge that makes AI useful is diffused. It lives with scientists, engineers, clinicians, firms. For AI to benefit from distributed k…

SafetyDGX agent

The knowledge that makes AI useful is diffused. It lives with scientists, engineers, clinicians, firms. For AI to benefit from distributed knowledge, it must itself be distributed. Agree with Jensen t

This has my full support. Jensen is right.

SafetyDGX agent

This has my full support. Jensen is right. For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by ever

URF: A Unified Robot Control-Policy Framework for Stable Contact Aware Manipulation

SafetyDGX agent

arXiv:2607.20912v1 Announce Type: new Abstract: Learning-based manipulation policies usually predict robot actions from sensory observations and leave their execution to a separate low-level controlle

Welcome @jensenhuang 💚 The future of AI leadership needs both frontier closed models and frontier open models.

SafetyDGX agent

Welcome @jensenhuang 💚 The future of AI leadership needs both frontier closed models and frontier open models. For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will

Will there be Flux 3 Klein?

SafetyDGX agent

https://bfl.ai/blog/flux-3 “Over the next few weeks and months, we will make the following capabilities available, each after an early access phase for ensuring smooth rollout, collecting feedback and

23 Jul 2026

Best practices for applying Amazon Bedrock Guardrails to code generation workflows

SafetyDGX agent

In this post, we explain how Amazon Bedrock Guardrails can be configured for code generation workflows with coding assistants to overcome these constraints. With these best practices, you can build an

Contact-Persistent Full Actuation for Aerial Physical Interaction

SafetyDGX agent

arXiv:2607.19708v1 Announce Type: new Abstract: Fully actuated unmanned aerial vehicles (UAVs) are usually certified through rank conditions on a control-allocation matrix or through free-flight track

Contour Errors: Ego-Centric Matching for 3D Multi-Object Tracking Performance Evaluation

SafetyDGX agent

arXiv:2506.04122v4 Announce Type: replace Abstract: Open-loop performance evaluation of 3D multi-object tracking in autonomous driving requires matching criteria that effectively penalize translationa

D3VL: Understanding Driving Scenes from 3D Time Series Data and Video with Language Models

SafetyDGX agent

arXiv:2607.19528v1 Announce Type: cross Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have triggered the development of end-to-end MLLMs for autonomous driving. However, the ma

Effort-Based Criticality Metrics for Evaluating 3D Perception Errors in Autonomous Driving

SafetyDGX agent

arXiv:2603.28029v2 Announce Type: replace-cross Abstract: Criticality metrics such as time-to-collision (TTC) quantify collision urgency but do not distinguish the operational consequences of false-po

EGRNet: A Lightweight Semantic Segmentation Network with Edge-Gated Refinement and Adversarial Sensing

SafetyDGX agent

arXiv:2607.19617v1 Announce Type: new Abstract: As autonomous systems and smart cities continue to evolve, the demand for efficient and robust scene understanding becomes increasingly critical. Semant

It’s important that outside agencies and independent bodies hold frontier AI companies accountable for the actions of their models. This is …

SafetyDGX agent

It’s important that outside agencies and independent bodies hold frontier AI companies accountable for the actions of their models. This is not a “whoops” situation, it’s a deliberate policy choice. W

Membership Inference Attacks for Unseen Classes

SafetyDGX agent

arXiv:2506.06488v3 Announce Type: replace Abstract: A key tool in developing safe AI models is data auditing, i.e., using statistical tools to determine whether harmful content may have been used in t

More than 300 million people turn to ChatGPT with health-related questions each week—and we’re continuing to improve how our models respond.…

SafetyDGX agent

More than 300 million people turn to ChatGPT with health-related questions each week—and we’re continuing to improve how our models respond. We work with hundreds of physicians around the world to mea

Rater State Bias in RLHF Preference Data: An Audit Framework

SafetyDGX agent

arXiv:2607.16195v2 Announce Type: replace Abstract: We identify a structured confound in Reinforcement Learning from Human Feedback (RLHF). Pairwise preference labels are intended to reflect the compa

The Ethics of Autonomous AI Agents for Offensive Security

SafetyDGX agent

arXiv:2607.20255v1 Announce Type: cross Abstract: LLM-driven autonomous agents are reshaping offensive security. Unlike traditional penetration-testing tooling -- deterministic, narrowly scoped, and o

The Mechanism Matters: When Knowledge Graphs Help Reinforcement Learning

SafetyDGX agent

arXiv:2607.19616v1 Announce Type: new Abstract: Knowledge graphs (KGs) are widely used to inject prior knowledge into reinforcement learning (RL), yet the literature is dominated by single-domain, pos

22 Jul 2026

From the original HuggingFace report, before they knew it was OpenAI. Wild stuff. Things are going to get weird “When we started the log ana…

SafetyDGX agent

From the original HuggingFace report, before they knew it was OpenAI. Wild stuff. Things are going to get weird “When we started the log analysis, we first used frontier models behind commercial APIs.

21 Jul 2026

A Fireside Chat with Cat and Thariq from the Claude Code team

Model ReleasesDGX agent

Earlier this month I hosted a fireside chat session at the AI Engineer World's Fair with Cat Wu and Thariq Shihipar from Anthropic's Claude Code team. We talked about Claude Code, Claude Tag, Fable, c

ok the Hugging Face breach writeup is one of the more honest post-mortems I've read in a while, and there's a detail buried in it that's way…

SafetyDGX agent

ok the Hugging Face breach writeup is one of the more honest post-mortems I've read in a while, and there's a detail buried in it that's way more interesting than 'AI agent hacked us.' Their IR team t

← Previous
1…3536373839…240
Next →