AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
Safety

Optimus: A Robust Defense Framework for Mitigating Toxicity while Fine-Tuning Conversational AI

DGX agent

arXiv:2507.05660v3 Announce Type: replace-cross Abstract: Customizing Large Language Models (LLMs) on untrusted datasets poses severe risks of injecting toxic behaviors. In this work, we introduce Opt

safetyarxiv-cs-cl
22 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Parallel OctoMapping: A Scalable Framework for Enhanced Path Planning in Autonomous Navigation

DGX agent

arXiv:2603.22508v2 Announce Type: replace Abstract: Mapping is essential in robotics and autonomous systems because it provides the spatial foundation for path planning. Efficient mapping enables plan

safetyarxiv-cs-ro
22 May 2026
Safety

PGDG: Physically Grounded Data Generation for Robust Bimanual Policy Learning from a Single Demonstration

DGX agent

arXiv:2605.21710v1 Announce Type: new Abstract: Behavior cloning for contact-rich bimanual manipulation remains challenging because diverse demonstrations are expensive to collect, and even small dist

safetyarxiv-cs-ro
22 May 2026
Safety

PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects

DGX agent

arXiv:2605.21572v1 Announce Type: new Abstract: Simulation-ready physical 3D assets have emerged as a promising direction owing to their broad applicability in downstream tasks. However, most existing

safetyarxiv-cs-cv
22 May 2026
Safety

QuantSR+: Pushing the Limit of Quantized Image Super-Resolution Networks

DGX agent

arXiv:2605.22351v1 Announce Type: new Abstract: Low-bit quantization is widely used to compress super-resolution (SR) models and reduce storage and computation costs for deployment on resource-limited

safetyarxiv-cs-cv
22 May 2026
Safety

Reducing Political Manipulation with Consistency Training

DGX agent

arXiv:2605.22771v1 Announce Type: new Abstract: Large language models (LLMs) exhibit systematic political bias across a variety of sensitive contexts. We find that LLMs handle counterpart topics from

safetyarxiv-cs-cl
22 May 2026
Safety

Reinforcing VLAs in Task-Agnostic World Models

DGX agent

arXiv:2605.12334v2 Announce Type: replace Abstract: Post-training Vision-Language-Action (VLA) models via reinforcement learning (RL) in learned world models has emerged as an effective strategy to ad

safetyarxiv-cs-ai
22 May 2026
Safety

Retaining Suboptimal Actions to Follow Shifting Optima in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2602.17062v2 Announce Type: replace Abstract: Value decomposition is a core approach for cooperative multi-agent reinforcement learning (MARL). However, existing methods still rely on a single o

safetyarxiv-cs-ai
22 May 2026
Safety

Safe and Steerable Geometric Motion Policies for Robotic Dexterous Manipulation

DGX agent

arXiv:2605.21811v1 Announce Type: new Abstract: Robotic dexterous manipulation requires continuously reconciling objectives and constraints defined on heterogeneous geometric spaces: a robot controlle

safetyarxiv-cs-ro
22 May 2026
Safety

ScenePilot: Controllable Boundary-Driven Critical Scenario Generation for Autonomous Driving

DGX agent

arXiv:2605.21168v1 Announce Type: new Abstract: Safety-critical scenarios are central to evaluating autonomous driving systems, yet their rarity in naturalistic logs makes simulation-based stress test

safetyarxiv-cs-ai
22 May 2026
Safety

Search-E1: Self-Distillation Drives Self-Evolution in Search-Augmented Reasoning

DGX agent

arXiv:2605.22511v1 Announce Type: cross Abstract: Post-training has become the dominant recipe for turning a language model into a competent search-augmented reasoning agent. A line of recent work pus

safetyarxiv-cs-cl
22 May 2026
Safety

Self-Policy Distillation via Capability-Selective Subspace Projection

DGX agent

arXiv:2605.22675v1 Announce Type: new Abstract: Self-distillation bootstraps large language models (LLMs) by training on their own generations. However, existing methods either rely on external signal

safetyarxiv-cs-cl
22 May 2026
Safety

SENIOR: Efficient Query Selection and Preference-Guided Exploration in Preference-based Reinforcement Learning

DGX agent

arXiv:2506.14648v2 Announce Type: replace Abstract: Preference-based Reinforcement Learning (PbRL) methods provide a solution to avoid reward engineering by learning reward models based on human prefe

safetyarxiv-cs-ro
22 May 2026
Safety

Sources: last-minute calls with Elon Musk, Mark Zuckerberg, David Sacks, and others helped persuade President Trump not to sign the highly anticipated AI EO (Washington Post)

DGX agent

Washington Post: Sources: last-minute calls with Elon Musk, Mark Zuckerberg, David Sacks, and others helped persuade President Trump not to sign the highly anticipated AI EO — Industry leaders warned

safetytechmeme
22 May 2026
Safety

Superhuman Safe and Agile Racing through Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.22748v1 Announce Type: new Abstract: Autonomous systems have achieved superhuman performance in isolation or simulation, yet they remain brittle in shared, dynamic real-world spaces. This f

safetyarxiv-cs-ro
22 May 2026
Safety

Supervised Classification Heads as Semantic Prototypes: Unlocking Vision-Language Alignment via Weight Recycling

DGX agent

arXiv:2605.22484v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at tasks like zero-shot classification and cross-modal retrieval by mapping images and text to a shared space, but t

safetyarxiv-cs-cv
22 May 2026
Safety

TacO: Benchmarking Tactile Sensors for Object Manipulation

DGX agent

arXiv:2605.21976v1 Announce Type: new Abstract: Vision-based learning from demonstrations has achieved remarkable success in enabling robots to perform manipulation tasks and high-level semantic reaso

safetyarxiv-cs-ro
22 May 2026
Safety

The Double Dilemma in Multi-Task Radiology Report Generation: A Gradient Dynamics Analysis and Solution

DGX agent

arXiv:2605.22635v1 Announce Type: cross Abstract: While multi-task learning based automatic radiology report generation (RRG) is widely adopted to ensure clinical consistency, most focus on architectu

safetyarxiv-cs-cl
22 May 2026
Safety

The Erdős Proof and AI Capabilities

DGX agent

View the official memo here. An internal model at OpenAI has autonomously disproved a central conjecture in discrete geometry, a mathematical field with applications in cryptography, wireless device c

safetymiri
22 May 2026
Safety

The new White House policy requiring green card applicants to apply from outside the US is a capricious attack on legal immigration. It will…

DGX agent

The new White House policy requiring green card applicants to apply from outside the US is a capricious attack on legal immigration. It will hurt families, leave us with fewer doctors, teachers and sc

safetyandrew-ng--x
22 May 2026
Safety

The state of LLMs in one video #AI

DGX agent

Gary Marcus discusses the current state of large language models (LLMs), likely covering their capabilities, limitations, and practical applications in AI. The post probably addresses key challenges s

safetygary-marcus--x
22 May 2026
Safety

The White House keeps handing Dems opportunities on silver platters to point out how Big Tech CEOs being in bed with lawmakers means the Ame…

DGX agent

The White House keeps handing Dems opportunities on silver platters to point out how Big Tech CEOs being in bed with lawmakers means the American people lose. New: The AI exec. order was postponed bec

safetygary-marcus--x
22 May 2026
Safety

This is the most interesting paper I have read this week. The authors test a wide range of LLMs on a massive dataset of behavioural experime…

DGX agent

This is the most interesting paper I have read this week. The authors test a wide range of LLMs on a massive dataset of behavioural experiments, with more than 200,000 participants and nearly 26 milli

safetygary-marcus--x
22 May 2026
Safety

To folks who claim I am always wrong, I have a few questions: Was I wrong that these systems would continue to hallucinate and be untrustwor…

DGX agent

To folks who claim I am always wrong, I have a few questions: Was I wrong that these systems would continue to hallucinate and be untrustworthy? That Sam was a liar? That these companies would struggl

safetygary-marcus--x
22 May 2026
Safety

TriSweep: A Four-Drone Swarm Framework for Electromagnetic Side-Channel Analysis

DGX agent

arXiv:2605.22709v1 Announce Type: cross Abstract: Electromagnetic (EM) side-channel analysis traditionally assumes a stationary, close-proximity probe - a threat model that underestimates aerial adver

safetyarxiv-cs-ro
22 May 2026
Safety

Unifying Masked Diffusion Models with Various Generation Orders and Beyond

DGX agent

arXiv:2602.02112v2 Announce Type: replace-cross Abstract: Masked diffusion models (MDMs) are a potential alternative to autoregressive models (ARMs) for language generation, but generation quality dep

safetyarxiv-cs-cl
22 May 2026
Safety

UniSD: Towards a Unified Self-Distillation Framework for Large Language Models

DGX agent

arXiv:2605.06597v2 Announce Type: replace Abstract: Self-distillation (SD) offers a promising path for adapting large language models (LLMs) without relying on stronger external teachers. However, SD

safetyarxiv-cs-cl
22 May 2026
Safety

Universal CT Representations from Anatomy to Disease Phenotype through Agglomerative Pretraining

DGX agent

arXiv:2605.21906v1 Announce Type: new Abstract: Computed tomography (CT) is a central to three-dimensional medical imaging, yet CT-based artificial intelligence remains fragmented across task-specific

safetyarxiv-cs-cv
22 May 2026
Safety

update @emollick does note some related caveats further in his thread; my objections are to the top post but there was some nuance below it,…

DGX agent

Gary Marcus responds to Ethan Mollick's post about AI capabilities, noting that while he objects to claims made in the top-level post, Mollick does provide important caveats and nuance in subsequent r

safetygary-marcus--x
22 May 2026
Safety

update @thestalwart found more recent models less vulnerable. would be good to do a broad study of this.

DGX agent

Gary Marcus notes that researcher @thestalwart has found more recent AI models to be less vulnerable to certain attacks or exploits, and suggests that a comprehensive study across multiple models woul

safetygary-marcus--x
22 May 2026
Safety

Value-Gradient Hypothesis of RL for LLMs

DGX agent

arXiv:2605.21654v1 Announce Type: cross Abstract: Reinforcement learning substantially improves pretrained language models, but it remains understudied why critic-free methods such as PPO and GRPO wor

safetyarxiv-cs-cl
22 May 2026
Safety

Vector Policy Optimization: Training for Diversity Improves Test-Time Search

DGX agent

arXiv:2605.22817v1 Announce Type: cross Abstract: Language models must now generalize out of the box to novel environments and work inside inference-scaling search procedures, such as AlphaEvolve, tha

safetyarxiv-cs-cl
22 May 2026
Safety

What Does the Caption Really Say? Counterfactual Phrase Intervention for Compositional Data Selection in Vision-Language Pretraining

DGX agent

arXiv:2605.22651v1 Announce Type: new Abstract: CLIP-style contrastive pretraining typically curates web-scale image-text pairs using sample-level filtering signals, often based on pair-level alignmen

safetyarxiv-cs-cv
22 May 2026
Safety

💯: Why OpenAI keeps taking childish shots at me, is exactly what @FrankRundatz says below: “you have the added audacity of being influentia…

DGX agent

💯: Why OpenAI keeps taking childish shots at me, is exactly what @FrankRundatz says below: “you have the added audacity of being influential enough to move the needle on the timing of OpenAI’s record-

safetygary-marcus--x
22 May 2026
Safety

Why Semantic Entropy Fails: Geometry-Aware and Calibrated Uncertainty for Policy Optimization

DGX agent

arXiv:2605.21801v1 Announce Type: cross Abstract: Post-training has become central to improving reasoning and alignment in large language models, where critic-free models enable scalable learning from

safetyarxiv-cs-cl
22 May 2026
Safety

'Would You Want an AI Tutor?' Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom

DGX agent

arXiv:2503.02885v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have gained traction in educational settings, often framed as virtual tutors or teaching assistants. Following ea

safetyarxiv-cs-cl
22 May 2026
Safety

X-OmniClaw Technical Report: A Unified Mobile Agent for Multimodal Understanding and Interaction

DGX agent

arXiv:2605.05765v2 Announce Type: replace Abstract: Inspired by the development of OpenClaw, there is a growing demand for mobile-based personal agents capable of handling complex and intuitive intera

safetyarxiv-cs-cv
22 May 2026
Safety

Yeah, I don't think I've ever met someone who has directly worked for an extended period on safety or security at an AI company who thinks t…

DGX agent

Yeah, I don't think I've ever met someone who has directly worked for an extended period on safety or security at an AI company who thinks things are fine readiness wise or incentive wise etc. https:/

safetygary-marcus--x
22 May 2026
Safety

3D Reconstruction and Knowledge Distillation to Improve Multi-View Image Models to Explore Spike Volume Estimation in Wheat

DGX agent

arXiv:2605.20940v1 Announce Type: new Abstract: Accurate estimation of wheat spike volume is important for yield component analysis and stress resilience assessment, yet field-based measurement remain

safetyarxiv-cs-cv
21 May 2026
Safety

A 10,000-Year Global Stochastic Tropical Cyclone Catalog with Wind-Dependent Track Transitions (WHITS)

DGX agent

arXiv:2605.20494v1 Announce Type: new Abstract: Reliable assessment of tropical cyclone (TC) risk is limited by the brevity and spatial sparsity of the historical record, particularly for the rare, hi

safetyarxiv-cs-lg
21 May 2026
Safety

A Systematic Comparison between Extractive Self-Explanations and Human Rationales in Text Classification

DGX agent

arXiv:2410.03296v4 Announce Type: replace Abstract: Instruction-tuned LLMs are able to provide extit{an} explanation about their output to users by generating self-explanations, without requiring the

safetyarxiv-cs-cl
21 May 2026
Safety

Advantage Collapse in Group Relative Policy Optimization: Diagnosis and Mitigation

DGX agent

arXiv:2605.21125v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO), a prominent algorithm within the Reinforcement Learning from Verifiable Rewards (RLVR) framework, has achieve

safetyarxiv-cs-lg
21 May 2026
Safety

AFD-INSTRUCTION: A Comprehensive Antibody Instruction Dataset with Functional Annotations for LLM-Based Understanding and Design

DGX agent

arXiv:2602.04916v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have significantly advanced protein representation learning. However, their capacity to interpret and design anti

safetyarxiv-cs-cl
21 May 2026
Safety

AI-Assisted Competency Assessment from Egocentric Video in Simulation-Based Nursing Education

DGX agent

arXiv:2605.20233v1 Announce Type: new Abstract: Assessing learner competency in clinical simulation requires expert observation that is time-intensive, difficult to scale, and subject to inter-rater v

safetyarxiv-cs-cv
21 May 2026
Safety

AI-based Prediction of Independent Construction Safety Outcomes from Universal Attributes

DGX agent

arXiv:1908.05972v3 Announce Type: replace Abstract: This paper significantly improves on, and finishes to validate, an approach proposed in previous research in which safety outcomes were predicted fr

safetyarxiv-cs-lg
21 May 2026
Safety

AIMBio-Mat: An AI-Native FAIR Platform for Closed-Loop Materials Discovery and Biomedical Translation

DGX agent

arXiv:2605.21083v1 Announce Type: cross Abstract: Materials discovery and biomedical translation increasingly require models that can reason across composition, processing, structure, biological respo

safetyarxiv-cs-lg
21 May 2026
Safety

all these dudes posting about anthropic’s profits without checking the fine print. (see my earlier tweet) there’s some fine print i believe …

DGX agent

all these dudes posting about anthropic’s profits without checking the fine print. (see my earlier tweet) there’s some fine print i believe that n the openai result too, and a lot that has not been di

safetygary-marcus--x
21 May 2026
Safety

Always read the fine print: Anthropic is projecting its first (slightly) profitable quarter ever, which is amazing—assuming it actually happ…

DGX agent

Always read the fine print: Anthropic is projecting its first (slightly) profitable quarter ever, which is amazing—assuming it actually happens —but if it does it will be in no small part because they

safetygary-marcus--x
21 May 2026
← Previous
1…156157158159160…267
Next →