AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,485 results
Safety

On the Efficacy of Self-Supervised Point Cloud Encoders for Efficient 3D Large Language Models

DGX agent

arXiv:2607.29136v1 Announce Type: new Abstract: 3D point cloud-language models (3D-LLMs) enable 3D understanding by pairing point cloud encoders with large language models, but existing methods rely o

safetyarxiv-cs-cv
3 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness

DGX agent

arXiv:2607.29062v1 Announce Type: new Abstract: Model capabilities have improved in large part due to scaling chain of thought. This has been a promising development for AI safety--where models verbal

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

OpenAI distilled Google’s name lol

DGX agent

Gary Marcus commented on Twitter that OpenAI “distilled” Google’s brand name, prompting a humorous response from Tony Carden. Carden’s reply referenced Google DeepMind’s Project Astra and included a s

safetygary-marcus--x
3 Aug 2026
Safety

OpenAI: No Anthropic: Unlikely

DGX agent

Gary Marcus tweeted “OpenAI: No Anthropic: Unlikely,” suggesting that the company does not expect Anthropic to satisfy particular requirements. The tweet references a discussion about the “RPO obligat

safetygary-marcus--x
3 Aug 2026
Safety

Part of the challenge with the AI math advances is the current air of secrecy around them. We don't know how Astra/Fable/Mythos etc are trai…

DGX agent

Part of the challenge with the AI math advances is the current air of secrecy around them. We don't know how Astra/Fable/Mythos etc are trained. We know many mathematicians have been paid to create tr

safetygary-marcus--x
3 Aug 2026
Safety

Persistent Convolution: A Topological Framework for AI Alignment Testing and Semantic Space Characterization

DGX agent

arXiv:2607.29008v1 Announce Type: cross Abstract: Modern opaque AI models prize performance over interpretability, which makes testing difficult. However, formal statistical tests conducted on a model

safetyarxiv-cs-lg
3 Aug 2026
Safety

Physics-Aligned Self-Supervised Learning for Scientific Imaging

DGX agent

arXiv:2607.28868v1 Announce Type: new Abstract: Data augmentations define the invariances learned by self-supervised learning (SSL). Standard augmentation pipelines were designed for natural images, y

safetyarxiv-cs-cv
3 Aug 2026
Safety

“poorly evidenced sensationalism makes for bad policymaking, politics, philanthropy and much else.” beautiful coda for a weekend in which pe…

DGX agent

“poorly evidenced sensationalism makes for bad policymaking, politics, philanthropy and much else.” beautiful coda for a weekend in which people went completely nuts without even asking for a control

safetygary-marcus--x
3 Aug 2026
Safety

QASP: Query-Adaptive Robust Vector Search Policy

DGX agent

arXiv:2607.29606v1 Announce Type: cross Abstract: A fundamental challenge of vector search is achieving consistently high recall while minimizing computational costs. Fixed search parameters cause sig

safetyarxiv-cs-lg
3 Aug 2026
Safety

QR-Structured Thermal Triggers for Targeted Semantic Attacks on Infrared Vision-Language Models

DGX agent

arXiv:2607.29445v1 Announce Type: cross Abstract: Infrared vision-language models (IR-VLMs) extend thermal perception to open-vocabulary classification, image captioning, and visual question answering

safetyarxiv-cs-ai
3 Aug 2026
Safety

RePaCA: Leveraging Reasoning Large Language Models for Static Automated Patch Correctness Assessment

DGX agent

arXiv:2507.22580v2 Announce Type: replace-cross Abstract: Automated Program Repair (APR) seeks to automatically correct software bugs without requiring human intervention. However, existing tools tend

safetyarxiv-cs-ai
3 Aug 2026
Safety

Rethinking Detection Calibration: A Coordinate and Direction Perspective

DGX agent

arXiv:2607.29040v1 Announce Type: new Abstract: Deep learning based object detectors require trustworthiness beyond competitive detection performance, but deep neural networks are prone to overconfide

safetyarxiv-cs-cv
3 Aug 2026
Safety

Robust Bidirectional Associative Memory via Regularization Inspired by the Subspace Rotation Algorithm

DGX agent

arXiv:2511.11902v2 Announce Type: replace-cross Abstract: Bidirectional Associative Memory (BAM) trained with Bidirectional Backpropagation (B-BP) often suffers from poor robustness and high sensitivi

safetyarxiv-cs-ai
3 Aug 2026
Safety

RTLCurator: Label-Efficient Data Curation for RTL Generation

DGX agent

arXiv:2607.29283v1 Announce Type: cross Abstract: Training large language models (LLMs) to write register-transfer level (RTL) requires large corpora of paired specifications and code, and such data i

safetyarxiv-cs-lg
3 Aug 2026
Safety

SAF-OPD: Stable Advantage Fusion for On-Policy Distillation

DGX agent

arXiv:2607.29209v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) broadcasts a single response-level reward to every token, while on-policy distillation (OPD) sco

safetyarxiv-cs-ai
3 Aug 2026
Safety

SAGP: Semantic Affordance-Guided Grasp Planning via Coarse-Zone VLM Reasoning

DGX agent

arXiv:2607.29374v1 Announce Type: new Abstract: Geometry-based grasp planners ensure physically valid grasps but ignore functional semantics, often generating grasps that are antipodal and collision-f

safetyarxiv-cs-ro
3 Aug 2026
Safety

Sample Efficient Hierarchical Reinforcement Learning via Best Policy Identification

DGX agent

arXiv:2607.29294v1 Announce Type: new Abstract: We present HBPI-UCRL, a model-based algorithm for hierarchical reinforcement learning (HRL) that learns high-level and low-level policies in parallel. H

safetyarxiv-cs-lg
3 Aug 2026
Safety

SatEdit: Mask-Conditioned Image Editing via VLM-Guided Segment Annotation

DGX agent

arXiv:2607.29367v1 Announce Type: new Abstract: Satellite image editing requires spatially precise object-level control, but supervised editing datasets for overhead imagery are costly to build becaus

safetyarxiv-cs-cv
3 Aug 2026
Safety

Scaffolding Critical Engagement with GenAI: Transforming Ethnic Minority Preparatory Students' Collaborative Discourse in Prompt Engineering Tasks

DGX agent

arXiv:2607.28630v1 Announce Type: cross Abstract: Generative AI (GenAI) holds significant promise for advancing educational equity among ethnic minority students by broadening access to learning resou

safetyarxiv-cs-ai
3 Aug 2026
Safety

Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Remember

DGX agent

arXiv:2607.29468v1 Announce Type: new Abstract: Self-play agents can generate training problems without questions from target benchmarks, but their curricula lack persistent state: failures affect gra

safetyarxiv-cs-ai
3 Aug 2026
Safety

Simple-regret rates and minimax optimality of fixed-prior expected improvement in Matern and squared-exponential RKHSs

DGX agent

arXiv:2607.29245v1 Announce Type: cross Abstract: We study the expected improvement (EI) policy for minimizing a deterministic objective function f on a nonempty compact set mathcal X subsetmathbb R^d

safetyarxiv-cs-lg
3 Aug 2026
Safety

StaQ: a Finite Memory Approach to Discrete Action Policy Mirror Descent

DGX agent

arXiv:2506.13862v2 Announce Type: replace-cross Abstract: In Reinforcement Learning (RL), regularization with a Kullback-Leibler divergence that penalizes large deviations between successive policies

safetyarxiv-cs-ai
3 Aug 2026
Safety

Stratified Negation in RDF Rules: A Correct Approach (Extended Version)

DGX agent

arXiv:2607.28778v1 Announce Type: cross Abstract: Combining RDF rule languages, such as N3 or SHACL Rules, with default negation is challenging. Existing methods to stratify negation often fail for RD

safetyarxiv-cs-ai
3 Aug 2026
Safety

TELLER: Dual-Path Iterative Preference Optimization for Table Entity Linking

DGX agent

arXiv:2607.28680v1 Announce Type: new Abstract: Entity linking in tables matches short and ambiguous cell mentions to their corresponding knowledge-base entities. Existing approaches typically rely on

safetyarxiv-cs-cl
3 Aug 2026
Safety

Temporal Policy: History-Initialized Action Generation for Robotic Learning from Demonstration

DGX agent

arXiv:2607.29482v1 Announce Type: new Abstract: By relying on independent couplings from uninformative Gaussian priors, standard diffusion and flow matching models are forced to learn complex, high-co

safetyarxiv-cs-ro
3 Aug 2026
Safety

TerraNova: A Foundation Model for the Anthropocene

DGX agent

arXiv:2607.29527v1 Announce Type: cross Abstract: A defining problem of the Anthropocene is to model the physical Earth and human societies as one coupled system, yet no learned representation spans t

safetyarxiv-cs-ai
3 Aug 2026
Safety

The Greedy Advantage in Finite-Horizon Bandits

DGX agent

arXiv:2607.29375v1 Announce Type: cross Abstract: Organizations increasingly rely on sequential experimentation to improve decision-making. While the multi-armed bandit literature has developed algori

safetyarxiv-cs-lg
3 Aug 2026
Safety

The K-Space Signature: Frequency-Domain Representation Learning for Medical Deepfake Detection

DGX agent

arXiv:2607.29541v1 Announce Type: new Abstract: In medical imaging, generative models are increasingly deployed to synthesize realistic data and augment limited datasets. Unfortunately, while benefici

safetyarxiv-cs-cv
3 Aug 2026
Safety

The Theoretical Foundation of Socratic Tests: Dynamic, Multimodal, Conversational Examinations

DGX agent

arXiv:2607.29624v1 Announce Type: cross Abstract: Traditional static assessments rely on a subtractive, deficit-based grading model that often penalizes ambition and obscures diagnostic feedback. Conv

safetyarxiv-cs-ai
3 Aug 2026
Safety

There should be a public investigation into the AI hacking incidents by OpenAI and Anthropic. We deserve to know whether these labs are genu…

DGX agent

There should be a public investigation into the AI hacking incidents by OpenAI and Anthropic. We deserve to know whether these labs are genuinely world-class security organizations facing a novel thre

safetygary-marcus--x
3 Aug 2026
Safety

this is the real reason people from OpenAI etc are desperate to shut me up. Astra (which didn’t even have a control group and is maybe not t…

DGX agent

this is the real reason people from OpenAI etc are desperate to shut me up. Astra (which didn’t even have a control group and is maybe not that much better than Fable and certainly not ASI) was perhap

safetygary-marcus--x
3 Aug 2026
Safety

TRACE: High-Fidelity 3D Scene Editing via Tangible Reconstruction and Geometry-Aligned Contextual Video Masking

DGX agent

arXiv:2604.01207v2 Announce Type: replace Abstract: Existing 3D Gaussian Splatting (3DGS) editing methods primarily focus on appearance modification and often struggle to support flexible geometry edi

safetyarxiv-cs-cv
3 Aug 2026
Safety

TraceViT: Grounded Trace Supervision for Visual Abstract Reasoning

DGX agent

arXiv:2607.29586v1 Announce Type: cross Abstract: The Abstraction and Reasoning Corpus (ARC) tests whether a model can infer an unseen transformation from a few input-output examples and apply it to a

safetyarxiv-cs-ai
3 Aug 2026
Safety

TRACT: Temporally Routed Action Chunks with Chronological Phase Authority for Contact-Rich Manipulation

DGX agent

arXiv:2607.29285v1 Announce Type: new Abstract: Action chunking shortens the effective decision horizon of robot imitation learning by predicting multiple future actions, while conventional phase cond

safetyarxiv-cs-ro
3 Aug 2026
Safety

Understanding Alignment in Multimodal LLMs: A Comprehensive Study

DGX agent

Preference alignment has become a crucial component in enhancing the performance of Large Language Models (LLMs), yet its impact in Multimodal Large Language Models (MLLMs) remains comparatively under

safetyapple-ml-research
3 Aug 2026
Safety

Unified continuous-time q-learning for mean-field game and mean-field control problems

DGX agent

arXiv:2407.04521v3 Announce Type: replace-cross Abstract: This paper studies the continuous-time q-learning in mean-field jump-diffusion models in a setting where the environment simulator does not pr

safetyarxiv-cs-lg
3 Aug 2026
Safety

VFAD: Variational Semantic Prompting Meets Frequency-Adaptive Representation Learning for Zero-Shot Anomaly Detection

DGX agent

arXiv:2607.29370v1 Announce Type: new Abstract: Zero-shot anomaly detection (ZSAD) aims to detect and localize anomalies in unseen categories without access to target-specific training data. Although

safetyarxiv-cs-cv
3 Aug 2026
Safety

ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding

DGX agent

arXiv:2607.28678v1 Announce Type: new Abstract: Multimodal agents operating in long-horizon environments must build and continually update multimedia memories to support entity-consistent, temporally

safetyarxiv-cs-ai
3 Aug 2026
Safety

WaMo: Wavelet-Enhanced Multi-Frequency Trajectory Analysis for Fine-Grained Text-Motion Retrieval

DGX agent

arXiv:2508.03343v2 Announce Type: replace Abstract: Text-Motion Retrieval (TMR) aims to retrieve 3D motion sequences semantically relevant to text descriptions. However, matching 3D motions with text

safetyarxiv-cs-cv
3 Aug 2026
Safety

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning

DGX agent

arXiv:2607.29613v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training of Vision-Language-Action (VLA) models has shown strong promise for robotic manipulation. Among RL methods,

safetyarxiv-cs-cl
3 Aug 2026
Safety

When Does On-Policy Interaction Help? Representational Tradeoffs in Value-Based Imitation Learning

DGX agent

arXiv:2607.29617v1 Announce Type: cross Abstract: Imitation learning (IL)---training an agent to replicate expert behavior from demonstrations---underpins applications from robotics to language model

safetyarxiv-cs-ai
3 Aug 2026
Safety

When Unlearning Fails: Reliable Data Deletion under Post-Training in Agent Networks

DGX agent

arXiv:2607.28829v1 Announce Type: cross Abstract: Self-improving federated agent networks keep training after deployment by collecting new trajectories with the current policy and feeding them back in

safetyarxiv-cs-lg
3 Aug 2026
Safety

Checkmate: you can’t take the harness (which is typically in large part symbolic) away from the neural model without giving up performance. …

DGX agent

Checkmate: you can’t take the harness (which is typically in large part symbolic) away from the neural model without giving up performance. HUGE victory for neurosymbolic AI, straight from @AnthropicA

safetygary-marcus--x
2 Aug 2026
Safety

exactly. math isn’t done. not at all.

DGX agent

exactly. math isn’t done. not at all. I don’t think being critical of the amazing work AI is doing in pure math is fair to @OpenAI until I can start to say why I feel it’s not yet at the level of our

safetygary-marcus--x
2 Aug 2026
Safety

LLMs can know a task is impossible and still optimize it anyway. Ask whether to walk or drive to a car wash 50 meters away, and some models …

DGX agent

LLMs can know a task is impossible and still optimize it anyway. Ask whether to walk or drive to a car wash 50 meters away, and some models focus on distance while missing that the car itself must rea

safetygary-marcus--x
2 Aug 2026
Safety

OpenAI guy lies about my intent. I would absolutely love to see progress in AI for science and medicine. I have said that here, in my books,…

DGX agent

OpenAI guy lies about my intent. I would absolutely love to see progress in AI for science and medicine. I have said that here, in my books, on countless podcasts, in multiple NYT opeds, in the US Sen

safetygary-marcus--x
2 Aug 2026
Safety

This wins the prize for sleazy misrepresentation. @mattShumer took my 2023 argument for hybridizing LLMs with symbolic tools – which is *exa…

DGX agent

This wins the prize for sleazy misrepresentation. @mattShumer took my 2023 argument for hybridizing LLMs with symbolic tools – which is *exactly* what everyone does nowadays – and made it sound like I

safetygary-marcus--x
2 Aug 2026
Safety

Yet another paper argues that LLMs aren’t close to doing real discovery.

DGX agent

Yet another paper argues that LLMs aren’t close to doing real discovery. MIT and Harvard argue LLMs are nowhere near doing real scientific discovery. They published a paper called “Evaluating Large La

safetygary-marcus--x
2 Aug 2026
← Previous
1…8586878889…302
Next →