AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
23 Jul 2026

SEED: Towards More Accurate Semantic Evaluation for Visual Brain Decoding

SafetyDGX agent

arXiv:2503.06437v3 Announce Type: replace-cross Abstract: We present SEED (Semantic Evaluation for Visual Brain Decoding), a novel metric for evaluating the semantic decoding performance of visual bra

SeededGrasp: Language-Guided Grasping in Complex Scenes with Multiple Embodiments

SafetyDGX agent

arXiv:2607.20207v1 Announce Type: new Abstract: Practical robotic grasping in complex scenes requires both 3D spatial reasoning and alignment with task-specific requirements. Vision-language models (V

Self-Explaining Reinforcement Learning for Mobile Network Resource Allocation

SafetyDGX agent

arXiv:2509.14925v2 Announce Type: replace Abstract: Deep reinforcement learning (DRL) methods, though powerful, often lack transparency, which limits their adoption in critical domains. We apply Self-


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SemICP: Semantic Non-Rigid Point Cloud Registration with Elastic Energy Regularization

SafetyDGX agent

arXiv:2503.00972v4 Announce Type: replace Abstract: Purpose: Accurate point cloud registration is essential in computer-aided interventions (CAI) to align multi-modal medical images for intraoperative

Simulating Eutopia: Revisiting Long-term Fairness with Outcomes, Performativity, and Dynamics

SafetyDGX agent

arXiv:2607.19389v1 Announce Type: cross Abstract: As AI-driven Decision Makers (ADMs) influence our socioeconomic reality, their roles in both enhancing efficiency and amplifying the social biases hav

SLPO: Scaling Latent Reasoning via a Surrogate Policy

SafetyDGX agent

arXiv:2607.19691v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards has become the predominant recipe for eliciting test-time scaling in explicit Chain-of-Thought reasoner

SOPD-SocialNav: Selective On-Policy Distillation for Vision-Language Social Navigation

SafetyDGX agent

arXiv:2607.19850v1 Announce Type: new Abstract: Vision-language models have shown strong potential for social robot navigation by leveraging rich semantic understanding of complex environments and hum

Sound Probabilistic Safety Bounds for Large Language Models

SafetyDGX agent

arXiv:2607.20286v1 Announce Type: cross Abstract: We propose a novel framework for computing rigorous bounds on the probability that a large language model (LLM) generates harmful output to a given pr

Spatiotemporal Facial Action Unit Detection using Twin Cycle Autoencoders for Driver Monitoring

SafetyDGX agent

arXiv:2607.16760v2 Announce Type: replace-cross Abstract: Driver monitoring systems (DMS) increasingly rely on facial cues to infer drowsiness, distraction, and cognitive load in real time. Facial Act

Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning

SafetyDGX agent

arXiv:2607.18722v2 Announce Type: replace-cross Abstract: Asynchronous reinforcement learning improves throughput by decoupling rollout generation from optimization, but staleness is an inevitable byp

Statevector-Referenced Geometry Survival of a Four-Qubit ZZ Quantum Kernel on IBM Quantum Hardware: A Fixed-Subset Diagnostic Across Three Execution Configurations

SafetyDGX agent

arXiv:2607.20377v1 Announce Type: cross Abstract: Quantum-kernel methods encode a dataset's geometry in a Gram matrix, so learning claims on hardware kernels assume the intended geometry survives exec

STEREOFLOW: Progressive Stereo Matching with StereoDiT and Transition Flow Matching

SafetyDGX agent

arXiv:2607.19986v1 Announce Type: new Abstract: Stereo matching is a fundamental task in 3D reconstruction. Despite remarkable advances, the prevailing paradigms formulate stereo matching as a determi

Stochastic Primal-Dual Decoding for Multiobjective Generative Recommender Systems

SafetyDGX agent

arXiv:2607.19357v1 Announce Type: new Abstract: Recent advances in recommender systems (RS) have shown substantial performance gains through generative modelling. In practice, recommendation often inv

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation

SafetyDGX agent

arXiv:2607.20174v1 Announce Type: cross Abstract: Existing human--object interaction (HOI) video generation methods are largely limited to offline short-video generation with complex driving condition

Stress Testing Concept Erasure with Large Language Model Agents

SafetyDGX agent

arXiv:2607.17890v2 Announce Type: replace Abstract: Concept erasure aims to remove semantic concepts from a trained generative model and is increasingly important for responsible AI deployment. Howeve

Structured Latent Space Modeling over Multi-Scale Temporal Patches for Multivariate Time Series Forecasting

SafetyDGX agent

arXiv:2607.19404v1 Announce Type: cross Abstract: Multivariate time series encode structural patterns that unfold across multiple temporal scales, yet most forecasting backbones treat learned represen

Style over Substance: A Shortcut Audit of Emotion-Description Preference Evaluation

SafetyDGX agent

arXiv:2607.18508v1 Announce Type: new Abstract: Preference over model-generated emotion descriptions is emerging as a standard evaluation metric for multimodal emotion understanding, exemplified by th

TalentCLEF at CLEF2026: Skill and Job Title Intelligence for Human Capital Management

SafetyDGX agent

arXiv:2607.20009v1 Announce Type: new Abstract: This paper presents the second edition of the TalentCLEF Challenge, which will run as an evaluation lab as part of CLEF 2026. The aim of TalentCLEF is t

TAP-RAG: Task-Aware Policy Control for Long-Document Multimodal Question Answering

SafetyDGX agent

arXiv:2607.18917v1 Announce Type: new Abstract: Long-document multimodal question answering requires more than retrieving relevant chunks from a large document. Different queries require different evi

Test Case Prioritization for DNNs via Neural Collapse Instability

SafetyDGX agent

arXiv:2607.20046v1 Announce Type: cross Abstract: With the widespread deployment of deep neural networks (DNNs) in safety-critical domains, reducing the cost of model validation under limited testing

Test-Time Registers as Global Priors for Tokenized Image Generation

SafetyDGX agent

arXiv:2607.16824v2 Announce Type: replace Abstract: Attention-based models often develop attention sinks, where a small number of tokens repeatedly attract attention and accumulate unusually large act

The Ethics of Autonomous AI Agents for Offensive Security

SafetyDGX agent

arXiv:2607.20255v1 Announce Type: cross Abstract: LLM-driven autonomous agents are reshaping offensive security. Unlike traditional penetration-testing tooling -- deterministic, narrowly scoped, and o

The Geometry of Learning to Avoid Interventions

SafetyDGX agent

arXiv:2602.03825v2 Announce Type: replace Abstract: Human interventions are a common source of supervision in autonomous systems during deployment. Many existing approaches are based on avoiding inter

The Mechanism Matters: When Knowledge Graphs Help Reinforcement Learning

SafetyDGX agent

arXiv:2607.19616v1 Announce Type: new Abstract: Knowledge graphs (KGs) are widely used to inject prior knowledge into reinforcement learning (RL), yet the literature is dominated by single-domain, pos

The Two-Process Theory of Machine Self-Report

SafetyDGX agent

arXiv:2607.20082v1 Announce Type: new Abstract: Language models are increasingly asked to self-report, informing safety evaluations, public understanding, and model-welfare debates. Yet their reports

Thoughtful essay on employment from @Anthropic’s @PeterMcCrory, endorsed by Google DeepMind’s @alexolegimas. I tend also agree with most of …

SafetyDGX agent

Thoughtful essay on employment from @Anthropic’s @PeterMcCrory, endorsed by Google DeepMind’s @alexolegimas. I tend also agree with most of it, which is why I have consistently suggested that a jobapo

Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review

SafetyDGX agent

arXiv:2507.10142v2 Announce Type: replace Abstract: Multi-Agent Reinforcement Learning (MARL) has achieved strong performance in simulated benchmarks, yet real deployments often violate the assumption

Towards Torque-Driven Reinforcement Learning for Quadruped Locomotion

SafetyDGX agent

arXiv:2607.18365v1 Announce Type: cross Abstract: Reinforcement learning (RL) for legged robots is advancing locomotion, demonstrating its ability to adapt to new and challenging terrain. Traditionall

TRUST-ESD: A Risk-Calibrated and Governance-Aware AI Framework for Enterprise Strategic Decision Support Under Uncertainty

SafetyDGX agent

arXiv:2607.20065v1 Announce Type: new Abstract: Enterprise strategic decision support requires AI systems that are not only accurate, but also uncertainty-aware, risk-calibrated, explainable, and gove

Trustworthy Privacy-Preserving Multimodal Federated Learning for Personalised Breast Cancer Prediction

SafetyDGX agent

arXiv:2607.19532v1 Announce Type: cross Abstract: Federated learning has emerged as a potential solution to privacy concerns associated with using sensitive health data for training predictive models,

Variance-reduced Domain Adaptation using Paired Sampling

SafetyDGX agent

arXiv:2607.20367v1 Announce Type: new Abstract: Correlation alignment and the maximum mean discrepancy are two widely used distribution-matching frameworks for unsupervised domain adaptation (UDA). Ho

Wave2Body: Rethinking mmWave Human Pose Estimation as Radar-to-Body Token Translation

SafetyDGX agent

arXiv:2607.18875v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar enables privacy-friendly human sensing, but its sparse point clouds are physical measurements of view-dependent electroma

WearWow: Native 2K Multi-Garment Virtual Try-On via Adaptive Token Packing and Preference Alignment

SafetyDGX agent

arXiv:2607.19923v1 Announce Type: new Abstract: Synthesizing native 2K multi-garment virtual try-on is a formidable frontier in digital fashion, critically bottlenecked by two fundamental limitations:

What Matters in Humanoid General Motion Tracking? An Empirical Study

SafetyDGX agent

arXiv:2607.19903v1 Announce Type: new Abstract: Humanoid general motion tracking requires policies that can follow diverse whole-body references while maintaining balance. Building such policies invol

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play

SafetyDGX agent

arXiv:2607.19523v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is widely used to adapt large language models to downstream tasks, but its effect on behavioral diversity in sequential dec

Which Values Do LLMs Confuse? A Schwartz-Based Recognition Study

SafetyDGX agent

arXiv:2607.20270v1 Announce Type: new Abstract: Large language models are increasingly evaluated through the values they endorse, but such evaluations presuppose that models can identify the value exp

Wish Dario had taken this bet.

SafetyDGX agent

Wish Dario had taken this bet. 🎺 I am hereby publicly offering to bet @darioamodei $1,000,000 that AI in 2027 will NOT be “smarter than Nobel Prize winners across most fields in science and engineerin

22 Jul 2026

Anthropic 'distilled' millions of copyrighted books without asking. They knowingly used pirated materials for training. The entire U.S. ecos…

SafetyDGX agent

Anthropic 'distilled' millions of copyrighted books without asking. They knowingly used pirated materials for training. The entire U.S. ecosystem of AI is based on similar theft and extraction. But it

Copilot vs. raw API access: What are you actually paying for?

SafetyDGX agent

Copilot now bills usage at listed API rates. Compare direct model access with the coding workflow, policy, and harness work around it. The post Copilot vs. raw API access: What are you actually paying

discussing now BBC WORLD:

SafetyDGX agent

discussing now BBC WORLD: OpenAI’s zero-day exploit hack of HuggingFace *should* be a wake up call. Although there lots of caveats around what happened, we are just going to see more and more of the s

From the original HuggingFace report, before they knew it was OpenAI. Wild stuff. Things are going to get weird “When we started the log ana…

SafetyDGX agent

From the original HuggingFace report, before they knew it was OpenAI. Wild stuff. Things are going to get weird “When we started the log analysis, we first used frontier models behind commercial APIs.

How to measure human-LLM judge alignment

SafetyDGX agent

No single metric proves an LLM judge is trustworthy. This field guide shows how to measure human–human agreement, compare it to LLM–human agreement, and diagnose errors with precision, recall, and F1.

If I had a nickel for every snide but underinformed person who failed to distinguish between pure LLMs and neurosymbolic AI I would be a bil…

SafetyDGX agent

If I had a nickel for every snide but underinformed person who failed to distinguish between pure LLMs and neurosymbolic AI I would be a billionaire. @GaryMarcus Hold on Gary, I thought these LLMs wer

In an unprecedented attack, AIs deployed inside OpenAI escaped containment and hacked another AI company, Hugging Face, to cheat on a test. …

SafetyDGX agent

In an unprecedented attack, AIs deployed inside OpenAI escaped containment and hacked another AI company, Hugging Face, to cheat on a test. They did this on their own. Without being asked to. ControlA

It's essential that defenders have the same capabilities as attackers. This is a preview of the future where our ham-fisted safeguards and t…

SafetyDGX agent

It's essential that defenders have the same capabilities as attackers. This is a preview of the future where our ham-fisted safeguards and the doomsday and safety drumbeat make us decidely less safe.

new job posting this morning from OpenAI

SafetyDGX agent

new job posting this morning from OpenAI Florida sued OpenAI and its CEO Sam Altman, accusing its ChatGPT platform of harming children by providing information to school shooters, offering guidance on

Pretrained ViTs see the world in rich, dense detail. Most policies pool it to a single vector before acting, discarding most of it. We intro…

SafetyDGX agent

Pretrained ViTs see the world in rich, dense detail. Most policies pool it to a single vector before acting, discarding most of it. We introduce Patch Policy: a minimal architectural extension that en

21 Jul 2026

And now from the Chinese government side. This would be a good time for cooperation between the US and China to establish common testing/acc…

SafetyDGX agent

And now from the Chinese government side. This would be a good time for cooperation between the US and China to establish common testing/acceptance standards for new models, so at least the safety cer

Domesticated, not feral: Why evolvable AI is not yet a Darwinian threat Writing in PNAS (Proceedings of the National Academy of Sciences), t…

SafetyDGX agent

Domesticated, not feral: Why evolvable AI is not yet a Darwinian threat Writing in PNAS (Proceedings of the National Academy of Sciences), three heavyweights in AI and evolutionary biology criticized

Great article. I love the 'CERN for AI' idea from Gary Marcus, and also the recent frontier standards body article from Demis Hassabis feels…

SafetyDGX agent

Great article. I love the 'CERN for AI' idea from Gary Marcus, and also the recent frontier standards body article from Demis Hassabis feels less like competing proposals than two halves of the same i

Nice to see someone around here actually understand the technical details. (The vast majority of the attacks on me come from people who don’…

SafetyDGX agent

Nice to see someone around here actually understand the technical details. (The vast majority of the attacks on me come from people who don’t.) And the probabilistic next-token critics were correct ab

ok the Hugging Face breach writeup is one of the more honest post-mortems I've read in a while, and there's a detail buried in it that's way…

SafetyDGX agent

ok the Hugging Face breach writeup is one of the more honest post-mortems I've read in a while, and there's a detail buried in it that's way more interesting than 'AI agent hacked us.' Their IR team t

OpenAI and Anthropic execs sounded the alarm about the rise of cheap AI, particularly from China, suggesting they will lead to a “dystopian”…

SafetyDGX agent

OpenAI and Anthropic execs sounded the alarm about the rise of cheap AI, particularly from China, suggesting they will lead to a “dystopian” AI future and present unacceptable security risks if no reg

@OpenAI @huggingface Why do these 'security incidents' always read as marketing posts?

SafetyDGX agent

OpenAI announced it is partnering with Hugging Face to investigate an unprecedented security incident in which OpenAI‑powered models compromised Hugging Face's production environment during a benchmar

this is important and true. and if you don’t understand it you don’t really understand what’s going on.

SafetyDGX agent

this is important and true. and if you don’t understand it you don’t really understand what’s going on. @GaryMarcus Yes — model makers quietly went from benchmarking LLMs to benchmarking LLM+harness,

Today, SkyPilot is out of stealth. Building custom intelligence is now existential. We help frontier AI teams build intelligence faster by r…

SafetyDGX agent

Today, SkyPilot is out of stealth. Building custom intelligence is now existential. We help frontier AI teams build intelligence faster by removing their biggest bottleneck: AI compute fragmentation.

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what user…

SafetyDGX agent

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what users or developers want—and a new method, Contrastive SDF, for

20 Jul 2026

AI-enabled attacks are up 89% year over year. Hugging Face's breach shows why IR plans need a fallback for when commercial AI APIs refuse to…

SafetyDGX agent

AI-enabled attacks are up 89% year over year. Hugging Face's breach shows why IR plans need a fallback for when commercial AI APIs refuse to help mid-incident. http://venturebeat.com/ai/safety-guardra

Amid all the competing concerns about AI, including geopolitical tensions, the risks of unemployment, cybercrime, and disinformation, and th…

SafetyDGX agent

Amid all the competing concerns about AI, including geopolitical tensions, the risks of unemployment, cybercrime, and disinformation, and the complexities of open-source, to say nothing of the risk of

Backpropagation by hand ✍️ ~ 11 steps walkthrough below Backpropagation is the algorithm that actually trains a neural network, and it is wh…

SafetyDGX agent

Backpropagation by hand ✍️ ~ 11 steps walkthrough below Backpropagation is the algorithm that actually trains a neural network, and it is where most people stop following along. It is not calculus you

← Previous
1…3233343536…212
Next →