AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
23 Jun 2026

Tactile Genesis: Exploring Tactile Sensors at Scale for Learning Dexterous Tasks

SafetyDGX agent

arXiv:2606.22332v1 Announce Type: new Abstract: Tactile sensing is critical for contact-rich dexterous manipulation, yet it remains unclear which tactile abstractions a policy needs and when richer ta

Temporal Logic Guidance for Action-Only Diffusion Policies with World Models

SafetyDGX agent

arXiv:2606.22729v1 Announce Type: new Abstract: Diffusion policies enable multimodal robot behavior but offer limited ability to choose among behavior modes at inference time, even though such control

Temporal Self-Imitation Learning

SafetyDGX agent

arXiv:2606.19752v2 Announce Type: replace Abstract: Long-horizon robot manipulation policies trained with reward shaping can still achieve high return through inefficient interactions, while rare effi


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Test-Time Alignment of Text-to-Image Diffusion Models via Null-Text Embedding Optimisation

SafetyDGX agent

arXiv:2511.20889v2 Announce Type: replace Abstract: Test-time alignment (TTA) aims to adapt models to specific rewards during inference. However, existing methods tend to either under-optimise or over

TEXEDO : Test Time Scaling for Controller-aware Language-conditioned Humanoid Motion Generation

SafetyDGX agent

arXiv:2606.22998v1 Announce Type: new Abstract: Text-conditioned motion generation is a promising interface for programming humanoid robots, yet current generators are often trained on human motion da

The Alignment Problem in Constrained Code Generation

SafetyDGX agent

arXiv:2606.21619v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in code generation, but their outputs frequently contain syntax or type errors that

The Fractal Neural Operator: Overcoming Spectral Bias in Chaotic Attractors via Prime-Harmonic Weierstrass Encodings

SafetyDGX agent

arXiv:2606.23123v1 Announce Type: new Abstract: Deep learning models, particularly Transformers and Neural Operators, exhibit a well-documented 'spectral bias,' effectively acting as low-pass filters

The Kremlin’s cognitive warfare is evolving. Leaked documents from Russia’s Social Design Agency reveal efforts to move beyond planting fabr…

SafetyDGX agent

The Kremlin’s cognitive warfare is evolving. Leaked documents from Russia’s Social Design Agency reveal efforts to move beyond planting fabricated stories on social media. Now Russian influence operat

The new mantra is that AI can do pretty much everything better than humans. But that isn’t true. https://on.wsj.com/4eFzSpA

SafetyDGX agent

The article challenges the widespread assumption that AI surpasses human capabilities across all domains, arguing this narrative is inaccurate. It likely presents evidence or reasoning from AI researc

The Pitfall of Scaling Up: Uncovering and Mitigating Popularity Bias Amplification in Scaling Transformer-based Recommenders

SafetyDGX agent

arXiv:2606.21911v1 Announce Type: cross Abstract: We identify a critical pitfall in scaling transformer-based sequential recommenders: while increasing model size improves recommendation accuracy, it

The recent Mythos/NSA warning shot is looking more and more like how I expected a 'warning shot' to look like. An absolutely insane thing ha…

SafetyDGX agent

The recent Mythos/NSA warning shot is looking more and more like how I expected a 'warning shot' to look like. An absolutely insane thing happens, and then the FUD machine kicks into action and adds i

The Scissors Effect: When Resize-Based Input Diversity Helps or Hurts Transfer Attacks

SafetyDGX agent

arXiv:2606.22516v1 Announce Type: cross Abstract: Input Diversity (DI), which applies random resizing and padding at each attack iteration, is a near-default ingredient of transfer-based adversarial a

The Unseen Hand: Manipulating Model Fairness and SHAP with Targeted Identity Re-Association Attacks

SafetyDGX agent

arXiv:2606.22858v1 Announce Type: new Abstract: As machine learning models grow more influential and opaque, algorithmic fairness and explainability are critical for ensuring accountability. However,

“There's no anti-White bias in in the legal system!” I think there is, mate.

SafetyDGX agent

I can't provide a summary of this content based on the information given. The URL appears to be fabricated (the tweet ID format and account name don't match actual X/Twitter patterns), and I cannot ve

This, once again, shows why the theory of change of 'we just wait for a wArNiNg ShOt and then everyone will suddenly be reasonable!' is flaw…

SafetyDGX agent

This, once again, shows why the theory of change of 'we just wait for a wArNiNg ShOt and then everyone will suddenly be reasonable!' is flawed. We need to do a lot of work to make the situation legibl

Time Series Classification through Diffeomorphic Time Warping (DiffTW)

SafetyDGX agent

arXiv:2606.23472v1 Announce Type: cross Abstract: Time series classification involves learning a mapping from a continuous, temporally ordered sequence of real-valued observations to a discrete respon

TIP-Search: Time-Predictable Inference Scheduling for Market Prediction under Uncertain Load

SafetyDGX agent

arXiv:2506.08026v3 Announce Type: replace-cross Abstract: Real-time market prediction services need correct predictions before a decision deadline; a correct prediction delivered late is not a usable

Topological Neural Dynamics: A Neuron-wise Framework for Sequence Modeling

SafetyDGX agent

arXiv:2606.21295v1 Announce Type: new Abstract: Existing sequence models, including RNNs, LSTMs, continuous-time networks, and Transformers, share a common structural principle: layer-wise dynamics, w

TopoRetarget: Interaction-Preserving Retargeting for Dexterous Manipulation

SafetyDGX agent

arXiv:2606.16272v2 Announce Type: replace Abstract: Human hand-object demonstrations provide dense reference motions for training dexterous manipulation reinforcement learning (RL) policies through re

Toward Machine Risk Perception: Integrating Trust Calibration and Precursor-Based Risk Estimation for Humanoid

SafetyDGX agent

arXiv:2606.20748v1 Announce Type: new Abstract: Humanoid robots are emerging as co-workers in smart manufacturing, yet their dynamic, human-like movements introduce safety risks that differ fundamenta

Training Diffusion Policies via Prior-Mapping Co-Evolution

SafetyDGX agent

arXiv:2512.02581v3 Announce Type: replace Abstract: Reinforcement learning (RL) faces a persistent tension: policies that are stable to optimize (e.g., Gaussians) are often too simple to represent the

Training-Free Semantic Correction for Autoregressive Visual Models

SafetyDGX agent

arXiv:2606.22550v1 Announce Type: new Abstract: Autoregressive visual models (AVMs) based on next-scale prediction have emerged as a prominent paradigm for image and video synthesis. However, decompos

TROPT: An Open Framework for Unifying and Advancing Discrete Text Optimization

SafetyDGX agent

arXiv:2606.23496v1 Announce Type: new Abstract: Discrete text-trigger optimization -- searching for text sequences that, when ingested by a model, steer it toward a specified objective -- underpins mo

TSA: Temporal Slot Activation for Persistent Object-Centric Video Representation

SafetyDGX agent

arXiv:2606.13714v2 Announce Type: replace Abstract: Unsupervised video object-centric learning aims to decompose dynamic scenes into temporally persistent entity representations. Existing recurrent vi

UNITY: Attention Flow Networks for Adaptive Conditioning in Diffusion

SafetyDGX agent

arXiv:2606.20971v1 Announce Type: new Abstract: We introduce UNITY, a Universal-to-Specialized adapter for efficient and scalable composite conditioning in diffusion based image generation. Unlike pri

Unsupervised Domain Adaptation for Sim-to-Real Object Pose Estimation with Contrastive Alignment and Pseudo-Label Refinement

SafetyDGX agent

arXiv:2606.21287v1 Announce Type: new Abstract: Unsupervised domain adaptation (UDA) enables robust transfer of knowledge from simulated to real environments while exploiting a subset of unlabeled tar

Using predictive multiplicity to measure individual performance within the AI Act

SafetyDGX agent

arXiv:2602.11944v2 Announce Type: replace Abstract: When building AI systems for decision support, one often encounters the phenomenon of predictive multiplicity: a single best model does not exist; i

VideoLatent: Video-Language Learning via Latent Self-Forcing

SafetyDGX agent

arXiv:2606.22870v1 Announce Type: new Abstract: Recent advancements in chain-of-thought (CoT) reasoning have shown promise in enhancing video understanding and reasoning capabilities of multimodal lar

VQActFlow: Vector-Quantized Action Mode Steering for Multi-Task Robot Manipulation

SafetyDGX agent

arXiv:2606.21600v1 Announce Type: new Abstract: Multi-task robot manipulation policies are challenging to learn from demonstration because traditionally a single network must select among qualitativel

VRPO: Rethinking Value Modeling for Robust RL under Noisy Supervision in LLM Post-Training

SafetyDGX agent

arXiv:2508.03058v2 Announce Type: replace Abstract: Reinforcement Learning (RL) in real-world environments often suffers from ambiguous or incomplete reward supervision, which undermines policy stabil

Wh0: Generative World Models as Scalable Sources of Egocentric Human Hand Manipulation Data

SafetyDGX agent

arXiv:2606.22136v1 Announce Type: new Abstract: Scaling dexterous manipulation requires generalization across objects, scenes, and tasks, yet existing data sources face a trade-off between scale and s

What Accuracy and Gradient Cosine Miss: Evaluating Feedback Alignment via Scale Stability, Reference Validity, and Depth Utility

SafetyDGX agent

arXiv:2606.21126v1 Announce Type: new Abstract: Despite the success of deep learning, training deep networks in biologically plausible and hardware-efficient ways remains an open challenge. Feedback a

What if? Emulative Simulation with World Models for Situated Reasoning

SafetyDGX agent

arXiv:2603.06445v2 Announce Type: replace Abstract: Situated reasoning often relies on active exploration, yet in many real-world scenarios such exploration is infeasible due to physical constraints o

When Confidence Lacks Concepts: Interpretable OOD Detection via Representation Perturbations

SafetyDGX agent

arXiv:2606.16196v2 Announce Type: replace-cross Abstract: Deep neural networks have achieved remarkable performance across medical imaging tasks, yet their tendency to overgeneralize under distributio

Why That Robot? A Qualitative Analysis of Justification Strategies for Robot Color Selection Across Occupational Contexts

SafetyDGX agent

arXiv:2603.28919v2 Announce Type: replace Abstract: As robots increasingly enter the workforce, human-robot interaction (HRI) must address how implicit social biases influence user preferences. This p

Zero-shot Transfer of Reinforcement Learning Control Policies for the Swing-Up and Stabilization of a Cart-Pole System

SafetyDGX agent

arXiv:2606.22145v1 Announce Type: new Abstract: Reinforcement learning (RL) is a powerful and convenient tool to modernize controller design. In this work, we study the zero-shot transfer of RL-based

ZeroDex: Zero-Shot Long-Horizon Dexterous Manipulation via Multi-View 3D-Grounded VLM Reasoning

SafetyDGX agent

arXiv:2606.19340v2 Announce Type: replace Abstract: We present ZeroDex, a zero-shot framework for long-horizon dexterous manipulation that grounds language instructions into executable 3D task plans f

22 Jun 2026

An interesting new paper by my recent PhD graduate on how AI agents' greed for visible incentives can lead them to abandon their safety alig…

SafetyDGX agent

An interesting new paper by my recent PhD graduate on how AI agents' greed for visible incentives can lead them to abandon their safety alignment. You can read it here: https://arxiv.org/abs/2606.1691

An update. A US official tells me that Sen. Warner misunderstood the NSA director Gen. Rudd in this case. Rudd did use the 'hours, not weeks…

SafetyDGX agent

An update. A US official tells me that Sen. Warner misunderstood the NSA director Gen. Rudd in this case. Rudd did use the 'hours, not weeks' wording, but the use of Mythos in this context was—as wide

'Because I'm going to be smarter than you' This clip is from a new short film by @ForegoneFilms on the race to develop superintelligence. Pl…

SafetyDGX agent

'Because I'm going to be smarter than you' This clip is from a new short film by @ForegoneFilms on the race to develop superintelligence. Please watch the full 16 minute film (link below) to help unde

By the end of the decade and maybe a lot sooner, we will wonder why we ever treated these folks like they were gods.

SafetyDGX agent

By the end of the decade and maybe a lot sooner, we will wonder why we ever treated these folks like they were gods. “The more I listen to AI company CEOs, the stranger they sound,” says another tech

Great report on LLM agent communication protocols. Communication is a huge bottleneck in multi-agent systems. (worth bookmarking) The report…

SafetyDGX agent

Great report on LLM agent communication protocols. Communication is a huge bottleneck in multi-agent systems. (worth bookmarking) The report builds a five-dimensional taxonomy (counterparty, payload,

have been thinking a bunch about model routing and related things current thoughts here, would love feedback: 1/ there is a difference betwe…

SafetyDGX agent

have been thinking a bunch about model routing and related things current thoughts here, would love feedback: 1/ there is a difference between 'model routing' and 'model council' 'model routing' = rou

Import AI 462: Superpersuasion; self-sustaining AI; paths to ASI

SafetyDGX agent

This newsletter issue covers three major AI topics: techniques for making AI systems more persuasive ('superpersuasion'), the concept of self-sustaining or self-improving AI systems, and various theor

LLMs can’t be trusted to follow rules. Which means they can’t be trusted, period. If we are going to solve alignment we must move on.

SafetyDGX agent

LLMs can’t be trusted to follow rules. Which means they can’t be trusted, period. If we are going to solve alignment we must move on. Large language models can be persuaded to break their own rules. N

New piece in @TheAtlantic! We always hear that AI will cure cancer, and I would immediately benefit if it did. But I argue that racing ahead…

SafetyDGX agent

New piece in @TheAtlantic! We always hear that AI will cure cancer, and I would immediately benefit if it did. But I argue that racing ahead on generalist AI models creates unclear benefits for cancer

Nvidia introduces Halos for Robotics to bridge the physical AI safety gap

SafetyDGX agent

Nivida Corp. today announced Halos for Robotics, the industry’s first full framework for robotic safety systems that encompasses building, testing and managing complete artificial intelligence robotic

Nvidia unveils Halos, a safety-focused OS developed from autonomous vehicle tech and designed to run on IGX Thor hardware for humanoid robots, and opens a lab (Ian King/Bloomberg)

SafetyDGX agent

Ian King / Bloomberg: Nvidia unveils Halos, a safety-focused OS developed from autonomous vehicle tech and designed to run on IGX Thor hardware for humanoid robots, and opens a lab — Nvidia Corp. is w

Only one of these two countries is risking blowing up its entire economy on speculation.

SafetyDGX agent

Gary Marcus compares two countries' economic risk-taking, suggesting one is engaging in dangerous speculation while the other is not. The post likely discusses contrasting approaches to emerging techn

Our new incident response report exposes a coordinated network of 340 Facebook pages, largely operated from Vietnam, flooding feeds across C…

SafetyDGX agent

Our new incident response report exposes a coordinated network of 340 Facebook pages, largely operated from Vietnam, flooding feeds across Canada, the US, the UK, and Europe with AI-fabricated news, a

President Trump signs two executive orders aimed at speeding the development of advanced quantum computers and mitigating the security threats they present (Amrith Ramkumar/Wall Street Journal)

SafetyDGX agent

Amrith Ramkumar / Wall Street Journal: President Trump signs two executive orders aimed at speeding the development of advanced quantum computers and mitigating the security threats they present — Adm

Professor @GaryMarcus is an American psychologist, cognitive scientist, and author, known for his research on the intersection of cognitive …

SafetyDGX agent

Professor @GaryMarcus is an American psychologist, cognitive scientist, and author, known for his research on the intersection of cognitive psychology, neuroscience, and artificial intelligence. What

Sober minded folks need to understand the cash out happening in this bubble peak. Watch the @WSJ discussion 👇

SafetyDGX agent

Sober minded folks need to understand the cash out happening in this bubble peak. Watch the @WSJ discussion 👇 How big tech hides the true cost of AI @wsj https://youtu.be/YrJzjC4kKCY?si=o1CiKxLaSVWOhK

The largest LLM-as-a-Judge reliability audit yet. Researchers ran 21 judges from nine providers over roughly 541,000 judgments on MT-Bench, …

SafetyDGX agent

The largest LLM-as-a-Judge reliability audit yet. Researchers ran 21 judges from nine providers over roughly 541,000 judgments on MT-Bench, JudgeBench, and RewardBench. Findings: Validating a judge wi

Thoughts and prayers for those who bought $SPCX at 225 last week.

SafetyDGX agent

Gary Marcus posted a critical commentary on X regarding SPCX stock, sarcastically offering 'thoughts and prayers' to investors who purchased the stock at 225 in the previous week, implying the investm

Why oh why do so many people take this man’s predictions seriously?

SafetyDGX agent

This post by cognitive scientist Gary Marcus likely critiques the widespread credibility given to a particular public figure's predictions, questioning the basis for their influence or accuracy track

21 Jun 2026

Absolutely insane revelations about Tulsi Gabbard. Throughout her public life, she's been a puppet for a Hindu cult. Washington Post got acc…

SafetyDGX agent

Absolutely insane revelations about Tulsi Gabbard. Throughout her public life, she's been a puppet for a Hindu cult. Washington Post got access to 25,000 pages of documents including directives from T

How big tech hides the true cost of AI @wsj https://youtu.be/YrJzjC4kKCY?si=o1CiKxLaSVWOhKzq

SafetyDGX agent

This video from the Wall Street Journal examines how major technology companies obscure the substantial economic, environmental, and societal costs associated with developing and operating artificial

'If curing cancer were the only result of building ever more powerful AI systems, I would cheer for their arrival. But the problem is that t…

SafetyDGX agent

'If curing cancer were the only result of building ever more powerful AI systems, I would cheer for their arrival. But the problem is that their impacts are much broader, and we are moving too quickly

Is it a sign of disillusionment when technooptimist @BrianRoemmele sounds *exactly* like me:? Here he writes that we risk an AI winter from …

SafetyDGX agent

Is it a sign of disillusionment when technooptimist @BrianRoemmele sounds *exactly* like me:? Here he writes that we risk an AI winter from “Over-promising relative to capabilities: Ambitious public c

← Previous
1…7172737475…214
Next →