AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,812 results
22 May 2026

Eight key points from the most recent essay in the “AI as Normal Technology” series by @sayashk and me. Do AI Risks Require Extraordinary Go…

SafetyDGX agent

Eight key points from the most recent essay in the “AI as Normal Technology” series by @sayashk and me. Do AI Risks Require Extraordinary Government Intervention? 1. There is general consensus that AI

Energy-Gated Attention: Spectral Salience as an Inductive Bias for Transformer Attention

SafetyDGX agent

arXiv:2605.21842v1 Announce Type: cross Abstract: Standard transformer attention computes pairwise similarity between queries and keys, treating all tokens as equally salient regardless of their intri

Enhancing Multimodal Large Language Models for Safety-Critical Driving Video Analysis

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.22185v1 Announce Type: new Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated impressive capabilities in general visual understanding. However, thei

FlyRoute: Self-Evolving Agent Profiling via Data Flywheel for Adaptive Task Routing

SafetyDGX agent

arXiv:2605.22057v1 Announce Type: new Abstract: Enterprise routers assign queries to expert agents, yet deployed profiles stay static while agents evolve (prompts, tools, models), and developers rarel

Focusing Where Vision Matters: Selective Training for Large Vision Language Models via Visual Information Gain

SafetyDGX agent

arXiv:2602.17186v2 Announce Type: replace Abstract: Large Vision Language Models (LVLMs) have achieved remarkable progress, yet they often suffer from language bias, producing answers without relying

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model

SafetyDGX agent

arXiv:2605.22671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models often suffer from performance degradation under distribution shifts, as they struggle to learn generalized behavior

Fuck this OpenAI employee, seriously fuck him. Also read this quantitative study, which I had nothing to do with, that says my technical pre…

SafetyDGX agent

Fuck this OpenAI employee, seriously fuck him. Also read this quantitative study, which I had nothing to do with, that says my technical predictions have bee largely correct. https://github.com/davego

“Gary Marcus, former CEO of Geometric Intelligence and New York University professor, joins @SquawkStreet @CNBC to discuss why he still beli…

SafetyDGX agent

“Gary Marcus, former CEO of Geometric Intelligence and New York University professor, joins @SquawkStreet @CNBC to discuss why he still believes OpenAI could be the ‘WeWork’ of AI, why he finds Anthro

GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation

SafetyDGX agent

arXiv:2605.21605v1 Announce Type: new Abstract: Open-ended image generation is no longer a simple prompt-to-image problem. High-quality generation often requires an agent to combine a model's internal

GLeVE: Graph-Guided Lesion Grounding with Proposal Verification in 3D CT

SafetyDGX agent

arXiv:2605.22619v1 Announce Type: new Abstract: Grounding radiology report descriptions to 3D CT volumes is essential for verifiable clinical interpretation, yet remains challenging due to the semanti

Google files its appeal of the US federal ruling deeming it an illegal search monopolist, arguing it 'prevailed in the marketplace fair and square' (Lauren Feiner/The Verge)

SafetyDGX agent

Lauren Feiner / The Verge: Google files its appeal of the US federal ruling deeming it an illegal search monopolist, arguing it “prevailed in the marketplace fair and square” — It wants to throw out t

Governance by Construction for Generalist Agents

SafetyDGX agent

arXiv:2605.20874v1 Announce Type: new Abstract: Enterprise agents are increasingly expected to operate autonomously across tools and interfaces, yet production deployments require governance by constr

Governance by Design: Architecting Agentic AI for Organizational Learning and Scalable Autonomy

SafetyDGX agent

arXiv:2605.20210v1 Announce Type: cross Abstract: Agentic AI systems - systems that can pursue goals through multi-step planning and tool-mediated action with limited direct supervision - are moving f

Harder to Defend: Towards Chinese Toxicity Attacks via Implicit Enhancement and Obfuscation Rewriting

SafetyDGX agent

arXiv:2605.22258v1 Announce Type: new Abstract: Large language models (LLMs) require robust toxicity evaluation beyond explicit wording. This setting remains underexplored in Chinese, where toxicity m

Hierarchical Variational Policies for Reward-Guided Diffusion

SafetyDGX agent

arXiv:2605.21661v1 Announce Type: cross Abstract: Adapting pretrained diffusion models to downstream objectives such as inverse problems often requires expensive test-time guidance or optimization. We

How can reasoning capability empower the AI copilot robot in endoscopic surgery

SafetyDGX agent

arXiv:2605.22322v1 Announce Type: new Abstract: Reasoning capability has significantly advanced complex logical inference and robotic decision-making in general domains. However, its potential in the

I just ran into a guy I know in NYC who said he’s doing a temporary gig for OpenAI. According to him, OpenAI is paying hundreds of people an…

SafetyDGX agent

I just ran into a guy I know in NYC who said he’s doing a temporary gig for OpenAI. According to him, OpenAI is paying hundreds of people and families in New York to install 360-degree cameras through

i suspect that this is a wild underestimate, both ignoring the costs in developing the model and ignoring the fact that many questions may w…

SafetyDGX agent

i suspect that this is a wild underestimate, both ignoring the costs in developing the model and ignoring the fact that many questions may well have been posed that didn’t succeed. If this is true, us

If Elon Musk was really serious about his proposal for 'universal high income,' @scottsantens tells me, he could start funding it right now.…

SafetyDGX agent

If Elon Musk was really serious about his proposal for 'universal high income,' @scottsantens tells me, he could start funding it right now. But he isn't doing that. Instead, he's railing about the ve

LACO: Adaptive Latent Communication for Collaborative Driving

SafetyDGX agent

arXiv:2605.22504v1 Announce Type: cross Abstract: Collaborative driving aims to improve safety and efficiency by enabling connected vehicles to coordinate under partial observability. Recent approache

LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance

SafetyDGX agent

arXiv:2605.22567v1 Announce Type: new Abstract: Reinforcement learning has proven effective for enhancing multi-step reasoning in large language models (LLMs), yet its benefits have not fully translat

Learning Altruistic Collaboration in Heterogeneous Multi-Team Systems

SafetyDGX agent

arXiv:2605.21723v1 Announce Type: new Abstract: This paper studies heterogeneous multi-team collaboration through dynamic robot allocation, where robots are treated as transferable resources. Leveragi

Learning to Configure Agentic AI Systems

SafetyDGX agent

arXiv:2602.11574v3 Announce Type: replace Abstract: Configuring LLM-based agent systems involves choosing workflows, tools, token budgets, and prompts from a large combinatorial design space, and is t

Learning to Evolve: Multi-modal Interactive Fields for Robust Humanoid Navigation in Dynamic Environments

SafetyDGX agent

arXiv:2605.21935v1 Announce Type: new Abstract: Safe manipulation-oriented navigation for humanoid robots requires scene memory that remains reliable under locomotion-induced perceptual distortion, en

Making the Discrete Continuous: Synthetic RAW Augmentations for Fine-Grained Evaluation of Person Detection Performance in Low Light

SafetyDGX agent

arXiv:2605.22455v1 Announce Type: new Abstract: Real-world deployment of AI vision models is both fueled and limited by the data available for training and testing. Real datasets are sparse and uneven

Mapping Tomato Cropping Systems in California Using AlphaEarth Geospatial Embeddings and Deep Learning Analysis

SafetyDGX agent

arXiv:2605.21804v1 Announce Type: cross Abstract: Field-scale crop maps support supply-chain forecasting and policy, yet statewide crop identification still often depends on retrospective surveys or r

Modeling Emotional Dynamics in Agent-to-Agent Interactions on Moltbook

SafetyDGX agent

arXiv:2605.20442v1 Announce Type: cross Abstract: Generative AI systems are increasingly deployed as interactive agents in online environments, such as a social network called Moltbook. In Moltbook, l

Modeling Pathology-Like Behavioral Patterns in Language Models Through Behavioral Fine-Tuning

SafetyDGX agent

arXiv:2605.22356v1 Announce Type: new Abstract: Large language models are increasingly used as computational tools for modeling human-like behavior. We introduce a behavioral induction framework that

Moral Semantics Survive Machine Translation: Cross-Lingual Evidence from Moral Foundations Corpora

SafetyDGX agent

arXiv:2605.22660v1 Announce Type: new Abstract: Moral language is subtle and culturally variable, making it difficult to translate faithfully across languages. Idiomatic expressions, slang, and cultur

my private index of desperation is how many people lie about me. we are in yellow alert territory

SafetyDGX agent

Gary Marcus expresses frustration about false claims being made about him, using a 'desperation index' metric based on the volume of lies circulated in his name. He indicates the situation has reached

NaviAgent: Graph-Driven Bilevel Planning for Scalable Tool Orchestration

SafetyDGX agent

arXiv:2506.19500v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) increasingly act as function-call agents that invoke external tools to tackle tasks beyond their static knowledge

Noise-Space Attribution and Control of Chunk-Boundary Artifact

SafetyDGX agent

arXiv:2603.11642v2 Announce Type: replace Abstract: Action chunking is widely used in generative visuomotor policies, yet the recurring execution discontinuities at chunk boundaries still lack a mecha

Non-Contact Vibration-Based Damage Detection of Civil Structures Using a Cost-Effective Autonomous UAV

SafetyDGX agent

arXiv:2605.21914v1 Announce Type: new Abstract: This paper presents a non-contact approach for vibration-based structural damage detection using an autonomous and customized cost-effective unmanned ae

nope you are. because openai will sell your most private data to the government.

SafetyDGX agent

This post appears to express concerns about OpenAI's data privacy practices and potential government data sharing, though the claim lacks specific evidence or context. Gary Marcus, an AI researcher an

Open-source LLMs administer maximum electric shocks in a Milgram-like obedience experiment

SafetyDGX agent

arXiv:2605.21401v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that make sequences of decisions over extended interactions in high-stakes

Optimus: A Robust Defense Framework for Mitigating Toxicity while Fine-Tuning Conversational AI

SafetyDGX agent

arXiv:2507.05660v3 Announce Type: replace-cross Abstract: Customizing Large Language Models (LLMs) on untrusted datasets poses severe risks of injecting toxic behaviors. In this work, we introduce Opt

Parallel OctoMapping: A Scalable Framework for Enhanced Path Planning in Autonomous Navigation

SafetyDGX agent

arXiv:2603.22508v2 Announce Type: replace Abstract: Mapping is essential in robotics and autonomous systems because it provides the spatial foundation for path planning. Efficient mapping enables plan

PGDG: Physically Grounded Data Generation for Robust Bimanual Policy Learning from a Single Demonstration

SafetyDGX agent

arXiv:2605.21710v1 Announce Type: new Abstract: Behavior cloning for contact-rich bimanual manipulation remains challenging because diverse demonstrations are expensive to collect, and even small dist

PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects

SafetyDGX agent

arXiv:2605.21572v1 Announce Type: new Abstract: Simulation-ready physical 3D assets have emerged as a promising direction owing to their broad applicability in downstream tasks. However, most existing

QuantSR+: Pushing the Limit of Quantized Image Super-Resolution Networks

SafetyDGX agent

arXiv:2605.22351v1 Announce Type: new Abstract: Low-bit quantization is widely used to compress super-resolution (SR) models and reduce storage and computation costs for deployment on resource-limited

Reducing Political Manipulation with Consistency Training

SafetyDGX agent

arXiv:2605.22771v1 Announce Type: new Abstract: Large language models (LLMs) exhibit systematic political bias across a variety of sensitive contexts. We find that LLMs handle counterpart topics from

Reinforcing VLAs in Task-Agnostic World Models

SafetyDGX agent

arXiv:2605.12334v2 Announce Type: replace Abstract: Post-training Vision-Language-Action (VLA) models via reinforcement learning (RL) in learned world models has emerged as an effective strategy to ad

Retaining Suboptimal Actions to Follow Shifting Optima in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2602.17062v2 Announce Type: replace Abstract: Value decomposition is a core approach for cooperative multi-agent reinforcement learning (MARL). However, existing methods still rely on a single o

Safe and Steerable Geometric Motion Policies for Robotic Dexterous Manipulation

SafetyDGX agent

arXiv:2605.21811v1 Announce Type: new Abstract: Robotic dexterous manipulation requires continuously reconciling objectives and constraints defined on heterogeneous geometric spaces: a robot controlle

ScenePilot: Controllable Boundary-Driven Critical Scenario Generation for Autonomous Driving

SafetyDGX agent

arXiv:2605.21168v1 Announce Type: new Abstract: Safety-critical scenarios are central to evaluating autonomous driving systems, yet their rarity in naturalistic logs makes simulation-based stress test

Search-E1: Self-Distillation Drives Self-Evolution in Search-Augmented Reasoning

SafetyDGX agent

arXiv:2605.22511v1 Announce Type: cross Abstract: Post-training has become the dominant recipe for turning a language model into a competent search-augmented reasoning agent. A line of recent work pus

Self-Policy Distillation via Capability-Selective Subspace Projection

SafetyDGX agent

arXiv:2605.22675v1 Announce Type: new Abstract: Self-distillation bootstraps large language models (LLMs) by training on their own generations. However, existing methods either rely on external signal

SENIOR: Efficient Query Selection and Preference-Guided Exploration in Preference-based Reinforcement Learning

SafetyDGX agent

arXiv:2506.14648v2 Announce Type: replace Abstract: Preference-based Reinforcement Learning (PbRL) methods provide a solution to avoid reward engineering by learning reward models based on human prefe

Sources: last-minute calls with Elon Musk, Mark Zuckerberg, David Sacks, and others helped persuade President Trump not to sign the highly anticipated AI EO (Washington Post)

SafetyDGX agent

Washington Post: Sources: last-minute calls with Elon Musk, Mark Zuckerberg, David Sacks, and others helped persuade President Trump not to sign the highly anticipated AI EO — Industry leaders warned

Superhuman Safe and Agile Racing through Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.22748v1 Announce Type: new Abstract: Autonomous systems have achieved superhuman performance in isolation or simulation, yet they remain brittle in shared, dynamic real-world spaces. This f

Supervised Classification Heads as Semantic Prototypes: Unlocking Vision-Language Alignment via Weight Recycling

SafetyDGX agent

arXiv:2605.22484v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at tasks like zero-shot classification and cross-modal retrieval by mapping images and text to a shared space, but t

TacO: Benchmarking Tactile Sensors for Object Manipulation

SafetyDGX agent

arXiv:2605.21976v1 Announce Type: new Abstract: Vision-based learning from demonstrations has achieved remarkable success in enabling robots to perform manipulation tasks and high-level semantic reaso

The Double Dilemma in Multi-Task Radiology Report Generation: A Gradient Dynamics Analysis and Solution

SafetyDGX agent

arXiv:2605.22635v1 Announce Type: cross Abstract: While multi-task learning based automatic radiology report generation (RRG) is widely adopted to ensure clinical consistency, most focus on architectu

The Erdős Proof and AI Capabilities

SafetyDGX agent

View the official memo here. An internal model at OpenAI has autonomously disproved a central conjecture in discrete geometry, a mathematical field with applications in cryptography, wireless device c

The new White House policy requiring green card applicants to apply from outside the US is a capricious attack on legal immigration. It will…

SafetyDGX agent

The new White House policy requiring green card applicants to apply from outside the US is a capricious attack on legal immigration. It will hurt families, leave us with fewer doctors, teachers and sc

The state of LLMs in one video #AI

SafetyDGX agent

Gary Marcus discusses the current state of large language models (LLMs), likely covering their capabilities, limitations, and practical applications in AI. The post probably addresses key challenges s

The White House keeps handing Dems opportunities on silver platters to point out how Big Tech CEOs being in bed with lawmakers means the Ame…

SafetyDGX agent

The White House keeps handing Dems opportunities on silver platters to point out how Big Tech CEOs being in bed with lawmakers means the American people lose. New: The AI exec. order was postponed bec

This is the most interesting paper I have read this week. The authors test a wide range of LLMs on a massive dataset of behavioural experime…

SafetyDGX agent

This is the most interesting paper I have read this week. The authors test a wide range of LLMs on a massive dataset of behavioural experiments, with more than 200,000 participants and nearly 26 milli

To folks who claim I am always wrong, I have a few questions: Was I wrong that these systems would continue to hallucinate and be untrustwor…

SafetyDGX agent

To folks who claim I am always wrong, I have a few questions: Was I wrong that these systems would continue to hallucinate and be untrustworthy? That Sam was a liar? That these companies would struggl

TriSweep: A Four-Drone Swarm Framework for Electromagnetic Side-Channel Analysis

SafetyDGX agent

arXiv:2605.22709v1 Announce Type: cross Abstract: Electromagnetic (EM) side-channel analysis traditionally assumes a stationary, close-proximity probe - a threat model that underestimates aerial adver

← Previous
1…124125126127128…214
Next →