AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
3 Aug 2026

SAGP: Semantic Affordance-Guided Grasp Planning via Coarse-Zone VLM Reasoning

SafetyDGX agent

arXiv:2607.29374v1 Announce Type: new Abstract: Geometry-based grasp planners ensure physically valid grasps but ignore functional semantics, often generating grasps that are antipodal and collision-f

Sample Efficient Hierarchical Reinforcement Learning via Best Policy Identification

SafetyDGX agent

arXiv:2607.29294v1 Announce Type: new Abstract: We present HBPI-UCRL, a model-based algorithm for hierarchical reinforcement learning (HRL) that learns high-level and low-level policies in parallel. H

SatEdit: Mask-Conditioned Image Editing via VLM-Guided Segment Annotation

SafetyDGX agent

arXiv:2607.29367v1 Announce Type: new Abstract: Satellite image editing requires spatially precise object-level control, but supervised editing datasets for overhead imagery are costly to build becaus

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Scaffolding Critical Engagement with GenAI: Transforming Ethnic Minority Preparatory Students' Collaborative Discourse in Prompt Engineering Tasks

SafetyDGX agent

arXiv:2607.28630v1 Announce Type: cross Abstract: Generative AI (GenAI) holds significant promise for advancing educational equity among ethnic minority students by broadening access to learning resou

Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Remember

SafetyDGX agent

arXiv:2607.29468v1 Announce Type: new Abstract: Self-play agents can generate training problems without questions from target benchmarks, but their curricula lack persistent state: failures affect gra

Simple-regret rates and minimax optimality of fixed-prior expected improvement in Matern and squared-exponential RKHSs

SafetyDGX agent

arXiv:2607.29245v1 Announce Type: cross Abstract: We study the expected improvement (EI) policy for minimizing a deterministic objective function f on a nonempty compact set mathcal X subsetmathbb R^d

StaQ: a Finite Memory Approach to Discrete Action Policy Mirror Descent

SafetyDGX agent

arXiv:2506.13862v2 Announce Type: replace-cross Abstract: In Reinforcement Learning (RL), regularization with a Kullback-Leibler divergence that penalizes large deviations between successive policies

Stratified Negation in RDF Rules: A Correct Approach (Extended Version)

SafetyDGX agent

arXiv:2607.28778v1 Announce Type: cross Abstract: Combining RDF rule languages, such as N3 or SHACL Rules, with default negation is challenging. Existing methods to stratify negation often fail for RD

TELLER: Dual-Path Iterative Preference Optimization for Table Entity Linking

SafetyDGX agent

arXiv:2607.28680v1 Announce Type: new Abstract: Entity linking in tables matches short and ambiguous cell mentions to their corresponding knowledge-base entities. Existing approaches typically rely on

Temporal Policy: History-Initialized Action Generation for Robotic Learning from Demonstration

SafetyDGX agent

arXiv:2607.29482v1 Announce Type: new Abstract: By relying on independent couplings from uninformative Gaussian priors, standard diffusion and flow matching models are forced to learn complex, high-co

TerraNova: A Foundation Model for the Anthropocene

SafetyDGX agent

arXiv:2607.29527v1 Announce Type: cross Abstract: A defining problem of the Anthropocene is to model the physical Earth and human societies as one coupled system, yet no learned representation spans t

The Greedy Advantage in Finite-Horizon Bandits

SafetyDGX agent

arXiv:2607.29375v1 Announce Type: cross Abstract: Organizations increasingly rely on sequential experimentation to improve decision-making. While the multi-armed bandit literature has developed algori

The K-Space Signature: Frequency-Domain Representation Learning for Medical Deepfake Detection

SafetyDGX agent

arXiv:2607.29541v1 Announce Type: new Abstract: In medical imaging, generative models are increasingly deployed to synthesize realistic data and augment limited datasets. Unfortunately, while benefici

The Theoretical Foundation of Socratic Tests: Dynamic, Multimodal, Conversational Examinations

SafetyDGX agent

arXiv:2607.29624v1 Announce Type: cross Abstract: Traditional static assessments rely on a subtractive, deficit-based grading model that often penalizes ambition and obscures diagnostic feedback. Conv

There should be a public investigation into the AI hacking incidents by OpenAI and Anthropic. We deserve to know whether these labs are genu…

SafetyDGX agent

There should be a public investigation into the AI hacking incidents by OpenAI and Anthropic. We deserve to know whether these labs are genuinely world-class security organizations facing a novel thre

this is the real reason people from OpenAI etc are desperate to shut me up. Astra (which didn’t even have a control group and is maybe not t…

SafetyDGX agent

this is the real reason people from OpenAI etc are desperate to shut me up. Astra (which didn’t even have a control group and is maybe not that much better than Fable and certainly not ASI) was perhap

TRACE: High-Fidelity 3D Scene Editing via Tangible Reconstruction and Geometry-Aligned Contextual Video Masking

SafetyDGX agent

arXiv:2604.01207v2 Announce Type: replace Abstract: Existing 3D Gaussian Splatting (3DGS) editing methods primarily focus on appearance modification and often struggle to support flexible geometry edi

TraceViT: Grounded Trace Supervision for Visual Abstract Reasoning

SafetyDGX agent

arXiv:2607.29586v1 Announce Type: cross Abstract: The Abstraction and Reasoning Corpus (ARC) tests whether a model can infer an unseen transformation from a few input-output examples and apply it to a

TRACT: Temporally Routed Action Chunks with Chronological Phase Authority for Contact-Rich Manipulation

SafetyDGX agent

arXiv:2607.29285v1 Announce Type: new Abstract: Action chunking shortens the effective decision horizon of robot imitation learning by predicting multiple future actions, while conventional phase cond

Understanding Alignment in Multimodal LLMs: A Comprehensive Study

SafetyDGX agent

Preference alignment has become a crucial component in enhancing the performance of Large Language Models (LLMs), yet its impact in Multimodal Large Language Models (MLLMs) remains comparatively under

Unified continuous-time q-learning for mean-field game and mean-field control problems

SafetyDGX agent

arXiv:2407.04521v3 Announce Type: replace-cross Abstract: This paper studies the continuous-time q-learning in mean-field jump-diffusion models in a setting where the environment simulator does not pr

VFAD: Variational Semantic Prompting Meets Frequency-Adaptive Representation Learning for Zero-Shot Anomaly Detection

SafetyDGX agent

arXiv:2607.29370v1 Announce Type: new Abstract: Zero-shot anomaly detection (ZSAD) aims to detect and localize anomalies in unseen categories without access to target-specific training data. Although

ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding

SafetyDGX agent

arXiv:2607.28678v1 Announce Type: new Abstract: Multimodal agents operating in long-horizon environments must build and continually update multimedia memories to support entity-consistent, temporally

WaMo: Wavelet-Enhanced Multi-Frequency Trajectory Analysis for Fine-Grained Text-Motion Retrieval

SafetyDGX agent

arXiv:2508.03343v2 Announce Type: replace Abstract: Text-Motion Retrieval (TMR) aims to retrieve 3D motion sequences semantically relevant to text descriptions. However, matching 3D motions with text

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning

SafetyDGX agent

arXiv:2607.29613v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training of Vision-Language-Action (VLA) models has shown strong promise for robotic manipulation. Among RL methods,

When Does On-Policy Interaction Help? Representational Tradeoffs in Value-Based Imitation Learning

SafetyDGX agent

arXiv:2607.29617v1 Announce Type: cross Abstract: Imitation learning (IL)---training an agent to replicate expert behavior from demonstrations---underpins applications from robotics to language model

When Unlearning Fails: Reliable Data Deletion under Post-Training in Agent Networks

SafetyDGX agent

arXiv:2607.28829v1 Announce Type: cross Abstract: Self-improving federated agent networks keep training after deployment by collecting new trajectories with the current policy and feeding them back in

2 Aug 2026

Checkmate: you can’t take the harness (which is typically in large part symbolic) away from the neural model without giving up performance. …

SafetyDGX agent

Checkmate: you can’t take the harness (which is typically in large part symbolic) away from the neural model without giving up performance. HUGE victory for neurosymbolic AI, straight from @AnthropicA

exactly. math isn’t done. not at all.

SafetyDGX agent

exactly. math isn’t done. not at all. I don’t think being critical of the amazing work AI is doing in pure math is fair to @OpenAI until I can start to say why I feel it’s not yet at the level of our

LLMs can know a task is impossible and still optimize it anyway. Ask whether to walk or drive to a car wash 50 meters away, and some models …

SafetyDGX agent

LLMs can know a task is impossible and still optimize it anyway. Ask whether to walk or drive to a car wash 50 meters away, and some models focus on distance while missing that the car itself must rea

OpenAI guy lies about my intent. I would absolutely love to see progress in AI for science and medicine. I have said that here, in my books,…

SafetyDGX agent

OpenAI guy lies about my intent. I would absolutely love to see progress in AI for science and medicine. I have said that here, in my books, on countless podcasts, in multiple NYT opeds, in the US Sen

This wins the prize for sleazy misrepresentation. @mattShumer took my 2023 argument for hybridizing LLMs with symbolic tools – which is *exa…

SafetyDGX agent

This wins the prize for sleazy misrepresentation. @mattShumer took my 2023 argument for hybridizing LLMs with symbolic tools – which is *exactly* what everyone does nowadays – and made it sound like I

Yet another paper argues that LLMs aren’t close to doing real discovery.

SafetyDGX agent

Yet another paper argues that LLMs aren’t close to doing real discovery. MIT and Harvard argue LLMs are nowhere near doing real scientific discovery. They published a paper called “Evaluating Large La

1 Aug 2026

Github repo to learn the OPD/OPSD and how they perform compared to GRPO, on a consumer grade GPU [P]

SafetyDGX agent

I am trying to learn concepts like On Policy Distillation (OPD), On Policy Self Distillation (OPSD) and how do they compare to RL algorithms like GRPO. There are a lot of papers on this, but because o

If Leopold had read this on June 26 and trimmed his bets accordingly, SALP would not have melted down. I laid everything out. https://open.s…

SafetyDGX agent

If Leopold had read this on June 26 and trimmed his bets accordingly, SALP would not have melted down. I laid everything out. https://open.substack.com/pub/garymarcus/p/the-month-generative-ai-lost-it

The “AGI-is-near” community keeps committing the same logical fallacy over and over; I have seen it at least half a dozen times today alone.…

SafetyDGX agent

The “AGI-is-near” community keeps committing the same logical fallacy over and over; I have seen it at least half a dozen times today alone. Every time there’s an advance, I see the same error. Here’s

31 Jul 2026

AI LEGO: Scaffolding Cross-Functional Collaboration in Industrial Responsible AI Practices during Early Design Stages

SafetyDGX agent

arXiv:2505.10300v2 Announce Type: replace-cross Abstract: Responsible AI (RAI) efforts increasingly emphasize the importance of addressing potential harms early in the AI development lifecycle through

AI Security Priorities: A Field-Wide Agenda

SafetyDGX agent

arXiv:2607.26069v1 Announce Type: cross Abstract: As AI systems are rapidly integrated into critical economic, governmental, and national security functions, the gap between AI adoption and AI securit

Anthropic employee vouches for fiancee of Dario’s chief of staff who holds a lot of Anthropic stock… Where do the Anthropic employees get th…

SafetyDGX agent

Anthropic employee vouches for fiancee of Dario’s chief of staff who holds a lot of Anthropic stock… Where do the Anthropic employees get their media training, exactly? Prediction: SALP will be bigger

APO: Unsupervised Atomic Policy Optimization for 3D Structure Prediction of Atomic Systems

SafetyDGX agent

arXiv:2607.28553v1 Announce Type: new Abstract: Predicting the 3D structures of atomic systems is fundamental to advancing material science and drug discovery. While flow-matching models (, FlowDPO) h

Belief-Guided Decision Making with Uncertainty Gating in the Game of Go

SafetyDGX agent

arXiv:2607.26946v1 Announce Type: new Abstract: Recent advancements in Computer Go, driven by AlphaZero and MuZero, rely heavily on Monte Carlo Tree Search (MCTS) to correct the errors of the neural n

Beyond Feeling Better: Capability-Sustaining Emotional Dialogue as a Longitudinal Research Paradigm

SafetyDGX agent

arXiv:2607.27851v1 Announce Type: new Abstract: Emotional dialogue research includes two influential strategy traditions. Empathetic dialogue prioritizes understanding a speaker's emotional experience

BioPro: Towards Difference-Aware Gender Fairness for Vision-Language Models

SafetyDGX agent

arXiv:2512.00807v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) inherit significant social biases from their training data, notably in gender representation. Current fairness interve

Borrowed Strength: Best-of-N Search over a Code EncodingBreaks Self-Check Jailbreak Defenses

SafetyDGX agent

arXiv:2607.26639v1 Announce Type: cross Abstract: A self-check defense asks the target model to assess a request before answering it; SAGE, the strongest published instance, reports an average 99% def

BridgeAlign: Bridging Preference Alignment for Humanities and Social Sciences

SafetyDGX agent

arXiv:2607.27366v1 Announce Type: new Abstract: While data synthesis for large language models (LLMs) is prevalent, it primarily targets domains with verifiable answers, overlooking open-ended humanit

Certifying when decision-time information justifies adaptive experimentation

SafetyDGX agent

arXiv:2607.27651v1 Announce Type: new Abstract: Adaptive laboratories choose measurements during experiments, yet most methods begin after adaptation is permitted. We introduce Opportunity-aware Polic

Class-Aware Reinforcement Learning for Counterfactual Explanation Generation

SafetyDGX agent

arXiv:2607.27905v1 Announce Type: new Abstract: Counterfactual explanations (CFEs) enhance the interpretability of black-box models by generating alternative instances with adjusted feature values tha

Contrastive Reinforced Policy Optimization via Privileged Self-Distillation

SafetyDGX agent

arXiv:2607.28026v1 Announce Type: new Abstract: Recent advances in post-training Large Language Models (LLMs) increasingly rely on Reinforcement Learning with Verifiable Rewards (RLVR) or On-Policy Se

Creative Transformation in Literary Texts: Modelling Change Across Representational Levels

SafetyDGX agent

arXiv:2607.28513v1 Announce Type: new Abstract: Creativity is often framed as the production of novelty, yet many cultural works emerge through transformation of earlier artifacts and not through isol

DAS-PMVC: A Framework for Partial Multi-View Clustering via Dual Alignment and Structure Enhancement

SafetyDGX agent

arXiv:2607.27761v1 Announce Type: cross Abstract: In recent years, multi-view clustering has attracted widespread research interest. However, due to limitations in data collection devices, data across

DECODE: Tackling Representation and Decision Degradation in Continual AI-Generated Image Detection

SafetyDGX agent

arXiv:2607.27882v1 Announce Type: new Abstract: As generative models continue to evolve, AI-generated image detectors must incrementally adapt to emerging generative domains while preserving knowledge

DexDirect: Direct Kinesthetic Arm Guidance for Efficient Dexterous Demonstration Collection

SafetyDGX agent

arXiv:2607.27784v1 Announce Type: new Abstract: Scalable collection of dexterous manipulation demonstrations remains a major bottleneck for robot learning. High-fidelity interfaces often require costl

Digital Harf: A Clinically Integrated Multimodal AI System for Pervasive Arabic Speech and Language Therapy

SafetyDGX agent

arXiv:2607.27212v1 Announce Type: cross Abstract: Children with Autism Spectrum Disorder in Arabic-speaking countries face compounded barriers to effective speech and language therapy: a shortage of q

DualAnchor: Preserving Language Priors and Improving Lexical Fidelity in Gloss-Free Sign Language Translation

SafetyDGX agent

arXiv:2607.27614v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have led sign language translation (SLT), the task of converting sign-language videos into spoken-langua

Dynamic Spectral Filtering for Temporal Graph Learning: Learning Evolving Propagation Operators

SafetyDGX agent

arXiv:2607.27891v1 Announce Type: cross Abstract: Temporal graph learning is commonly organized around the evolution of node states or the encoding of interaction histories. We study an underexplored,

Eco3S: Complex Socio-Economic System Simulation via Agent-Based Models

SafetyDGX agent

arXiv:2607.26588v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has renewed interest in agent-based modeling (ABM). However, current LLM-based ABM research faces

EgoGenesis: Egocentric World-Action Modeling with Online Anchored Projective Memory and Action-3D RoPE

SafetyDGX agent

arXiv:2607.28243v1 Announce Type: new Abstract: Egocentric video offers rich manipulation experience for embodied AI, yet collecting diverse egocentric data across scenes, objects, motions, and embodi

Evaluation Protocols and Cross-Subject Generalization in EEG Emotion Recognition

SafetyDGX agent

arXiv:2607.27655v1 Announce Type: new Abstract: Reported accuracy in electroencephalography (EEG) emotion recognition depends on the complete evaluation procedure, not only the classifier. We separate

Everyone’s going on about how smart Leopold Aschennbrenner is (or was). But 1. He obviously didn’t know anything at all about risk managemen…

SafetyDGX agent

Everyone’s going on about how smart Leopold Aschennbrenner is (or was). But 1. He obviously didn’t know anything at all about risk management. (Or arrogantly chose to disregard whatever he might have

FA-RDP: A Frequency-Adaptive Reactive Diffusion Policy for Contact-Rich Manipulation

SafetyDGX agent

arXiv:2607.28596v1 Announce Type: new Abstract: In contact-rich manipulation, action multimodality and reactivity dominate different stages of a single episode. Before contact, multiple trajectories m

← Previous
1…6869707172…242
Next →