AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
23 Jun 2026

Spectral Gating via Damped Oscillations for Adaptive Implicit Neural Representations

SafetyDGX agent

arXiv:2606.23129v1 Announce Type: new Abstract: Implicit Neural Representations (INRs) have been proven successful in encoding continuous signals through coordinate-based networks, yet facing a spectr

SQLConductor: Search-to-Policy Learning for Step-wise Text-to-SQL Orchestration

SafetyDGX agent

arXiv:2606.23537v1 Announce Type: cross Abstract: Text-to-SQL enables users to access relational databases via natural language, but real-world settings remain challenging due to coordinated reasoning

Stabilizing Consistency Training: A Flow Map Analysis and Self-Distillation

SafetyDGX agent

arXiv:2601.22679v2 Announce Type: replace-cross Abstract: Consistency models have been proposed for fast generative modeling, achieving results competitive with diffusion and flow models. However, the

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Stable Transformer-Actor-Critic Model Predictive Control: A Contraction Analysis Approach

SafetyDGX agent

arXiv:2606.20197v2 Announce Type: replace Abstract: Actor-Critic Model Predictive Control (MPC) effectively addresses complex, non-convex control problems, but guaranteeing the closed-loop stability o

Stationary Robust Mean-Field Games under Model Mismatches

SafetyDGX agent

arXiv:2606.22579v1 Announce Type: new Abstract: Deploying multi-agent reinforcement learning (MARL) in the real world is often limited by model mismatches between the training simulators and the true

Statistical Inference for Misspecified Contextual Bandits

SafetyDGX agent

arXiv:2606.22639v1 Announce Type: cross Abstract: Contextual bandit algorithms have transformed modern experimentation by enabling real-time adaptation for personalized treatment. Yet these advantages

Substitution-Based Analysis of Structural Novelty for Generative Models of Materials

SafetyDGX agent

arXiv:2606.23166v1 Announce Type: new Abstract: There has been rapid progress in generative artificial intelligence (AI) models for inorganic crystal design, which can efficiently generate large numbe

Superhuman AI for Generals.io Using Self-Play Reinforcement Learning

SafetyDGX agent

arXiv:2606.23348v1 Announce Type: new Abstract: We present a superhuman AI agent for Generals.io, a real-time strategy game that requires both long-horizon planning and short-term tactics under strong

SurGE: Surrogate Gradient-guided Evolution for Co-design of Legged Robots with Parallel Elasticity

SafetyDGX agent

arXiv:2606.21866v1 Announce Type: new Abstract: Co-design of legged robots with elastic elements is challenging due to the non-differentiability of contact dynamics and mechanism engagement. This pape

Surprise-Guided MergeSort: Budget-Efficient Human-in-the-Loop Ranking via Adaptive Comparison Scheduling

SafetyDGX agent

arXiv:2606.15623v2 Announce Type: replace Abstract: Pairwise comparison is the gold standard for subjective ranking tasks; however, exhaustive annotation requires a massive number of human comparisons

Synergistic Dual-Branch Adaptation for Multi-modal Generalized Category Discovery

SafetyDGX agent

arXiv:2606.21446v1 Announce Type: new Abstract: Generalized Category Discovery (GCD) aims to classify old categories and discover new ones from unlabeled data. Recent multi-modal approaches introduce

TACT-ful: Multi-Channel Terrain Affordance and Compliance Training for Payload-Robust Perceptive Humanoid Locomotion

SafetyDGX agent

arXiv:2606.20645v1 Announce Type: new Abstract: Foothold selection on structured terrain requires explicit reasoning about contact planarity, surface steepness, and kinematic reachability, properties

Tactile Genesis: Exploring Tactile Sensors at Scale for Learning Dexterous Tasks

SafetyDGX agent

arXiv:2606.22332v1 Announce Type: new Abstract: Tactile sensing is critical for contact-rich dexterous manipulation, yet it remains unclear which tactile abstractions a policy needs and when richer ta

Temporal Logic Guidance for Action-Only Diffusion Policies with World Models

SafetyDGX agent

arXiv:2606.22729v1 Announce Type: new Abstract: Diffusion policies enable multimodal robot behavior but offer limited ability to choose among behavior modes at inference time, even though such control

Temporal Self-Imitation Learning

SafetyDGX agent

arXiv:2606.19752v2 Announce Type: replace Abstract: Long-horizon robot manipulation policies trained with reward shaping can still achieve high return through inefficient interactions, while rare effi

Test-Time Alignment of Text-to-Image Diffusion Models via Null-Text Embedding Optimisation

SafetyDGX agent

arXiv:2511.20889v2 Announce Type: replace Abstract: Test-time alignment (TTA) aims to adapt models to specific rewards during inference. However, existing methods tend to either under-optimise or over

TEXEDO : Test Time Scaling for Controller-aware Language-conditioned Humanoid Motion Generation

SafetyDGX agent

arXiv:2606.22998v1 Announce Type: new Abstract: Text-conditioned motion generation is a promising interface for programming humanoid robots, yet current generators are often trained on human motion da

The Alignment Problem in Constrained Code Generation

SafetyDGX agent

arXiv:2606.21619v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in code generation, but their outputs frequently contain syntax or type errors that

The Fractal Neural Operator: Overcoming Spectral Bias in Chaotic Attractors via Prime-Harmonic Weierstrass Encodings

SafetyDGX agent

arXiv:2606.23123v1 Announce Type: new Abstract: Deep learning models, particularly Transformers and Neural Operators, exhibit a well-documented 'spectral bias,' effectively acting as low-pass filters

The Kremlin’s cognitive warfare is evolving. Leaked documents from Russia’s Social Design Agency reveal efforts to move beyond planting fabr…

SafetyDGX agent

The Kremlin’s cognitive warfare is evolving. Leaked documents from Russia’s Social Design Agency reveal efforts to move beyond planting fabricated stories on social media. Now Russian influence operat

The new mantra is that AI can do pretty much everything better than humans. But that isn’t true. https://on.wsj.com/4eFzSpA

SafetyDGX agent

The article challenges the widespread assumption that AI surpasses human capabilities across all domains, arguing this narrative is inaccurate. It likely presents evidence or reasoning from AI researc

The Pitfall of Scaling Up: Uncovering and Mitigating Popularity Bias Amplification in Scaling Transformer-based Recommenders

SafetyDGX agent

arXiv:2606.21911v1 Announce Type: cross Abstract: We identify a critical pitfall in scaling transformer-based sequential recommenders: while increasing model size improves recommendation accuracy, it

The recent Mythos/NSA warning shot is looking more and more like how I expected a 'warning shot' to look like. An absolutely insane thing ha…

SafetyDGX agent

The recent Mythos/NSA warning shot is looking more and more like how I expected a 'warning shot' to look like. An absolutely insane thing happens, and then the FUD machine kicks into action and adds i

The Scissors Effect: When Resize-Based Input Diversity Helps or Hurts Transfer Attacks

SafetyDGX agent

arXiv:2606.22516v1 Announce Type: cross Abstract: Input Diversity (DI), which applies random resizing and padding at each attack iteration, is a near-default ingredient of transfer-based adversarial a

The Unseen Hand: Manipulating Model Fairness and SHAP with Targeted Identity Re-Association Attacks

SafetyDGX agent

arXiv:2606.22858v1 Announce Type: new Abstract: As machine learning models grow more influential and opaque, algorithmic fairness and explainability are critical for ensuring accountability. However,

“There's no anti-White bias in in the legal system!” I think there is, mate.

SafetyDGX agent

I can't provide a summary of this content based on the information given. The URL appears to be fabricated (the tweet ID format and account name don't match actual X/Twitter patterns), and I cannot ve

This, once again, shows why the theory of change of 'we just wait for a wArNiNg ShOt and then everyone will suddenly be reasonable!' is flaw…

SafetyDGX agent

This, once again, shows why the theory of change of 'we just wait for a wArNiNg ShOt and then everyone will suddenly be reasonable!' is flawed. We need to do a lot of work to make the situation legibl

Time Series Classification through Diffeomorphic Time Warping (DiffTW)

SafetyDGX agent

arXiv:2606.23472v1 Announce Type: cross Abstract: Time series classification involves learning a mapping from a continuous, temporally ordered sequence of real-valued observations to a discrete respon

Topological Neural Dynamics: A Neuron-wise Framework for Sequence Modeling

SafetyDGX agent

arXiv:2606.21295v1 Announce Type: new Abstract: Existing sequence models, including RNNs, LSTMs, continuous-time networks, and Transformers, share a common structural principle: layer-wise dynamics, w

TopoRetarget: Interaction-Preserving Retargeting for Dexterous Manipulation

SafetyDGX agent

arXiv:2606.16272v2 Announce Type: replace Abstract: Human hand-object demonstrations provide dense reference motions for training dexterous manipulation reinforcement learning (RL) policies through re

Training Diffusion Policies via Prior-Mapping Co-Evolution

SafetyDGX agent

arXiv:2512.02581v3 Announce Type: replace Abstract: Reinforcement learning (RL) faces a persistent tension: policies that are stable to optimize (e.g., Gaussians) are often too simple to represent the

Training-Free Semantic Correction for Autoregressive Visual Models

SafetyDGX agent

arXiv:2606.22550v1 Announce Type: new Abstract: Autoregressive visual models (AVMs) based on next-scale prediction have emerged as a prominent paradigm for image and video synthesis. However, decompos

TROPT: An Open Framework for Unifying and Advancing Discrete Text Optimization

SafetyDGX agent

arXiv:2606.23496v1 Announce Type: new Abstract: Discrete text-trigger optimization -- searching for text sequences that, when ingested by a model, steer it toward a specified objective -- underpins mo

TSA: Temporal Slot Activation for Persistent Object-Centric Video Representation

SafetyDGX agent

arXiv:2606.13714v2 Announce Type: replace Abstract: Unsupervised video object-centric learning aims to decompose dynamic scenes into temporally persistent entity representations. Existing recurrent vi

UNITY: Attention Flow Networks for Adaptive Conditioning in Diffusion

SafetyDGX agent

arXiv:2606.20971v1 Announce Type: new Abstract: We introduce UNITY, a Universal-to-Specialized adapter for efficient and scalable composite conditioning in diffusion based image generation. Unlike pri

Unsupervised Domain Adaptation for Sim-to-Real Object Pose Estimation with Contrastive Alignment and Pseudo-Label Refinement

SafetyDGX agent

arXiv:2606.21287v1 Announce Type: new Abstract: Unsupervised domain adaptation (UDA) enables robust transfer of knowledge from simulated to real environments while exploiting a subset of unlabeled tar

Using predictive multiplicity to measure individual performance within the AI Act

SafetyDGX agent

arXiv:2602.11944v2 Announce Type: replace Abstract: When building AI systems for decision support, one often encounters the phenomenon of predictive multiplicity: a single best model does not exist; i

VideoLatent: Video-Language Learning via Latent Self-Forcing

SafetyDGX agent

arXiv:2606.22870v1 Announce Type: new Abstract: Recent advancements in chain-of-thought (CoT) reasoning have shown promise in enhancing video understanding and reasoning capabilities of multimodal lar

VQActFlow: Vector-Quantized Action Mode Steering for Multi-Task Robot Manipulation

SafetyDGX agent

arXiv:2606.21600v1 Announce Type: new Abstract: Multi-task robot manipulation policies are challenging to learn from demonstration because traditionally a single network must select among qualitativel

VRPO: Rethinking Value Modeling for Robust RL under Noisy Supervision in LLM Post-Training

SafetyDGX agent

arXiv:2508.03058v2 Announce Type: replace Abstract: Reinforcement Learning (RL) in real-world environments often suffers from ambiguous or incomplete reward supervision, which undermines policy stabil

Wh0: Generative World Models as Scalable Sources of Egocentric Human Hand Manipulation Data

SafetyDGX agent

arXiv:2606.22136v1 Announce Type: new Abstract: Scaling dexterous manipulation requires generalization across objects, scenes, and tasks, yet existing data sources face a trade-off between scale and s

What Accuracy and Gradient Cosine Miss: Evaluating Feedback Alignment via Scale Stability, Reference Validity, and Depth Utility

SafetyDGX agent

arXiv:2606.21126v1 Announce Type: new Abstract: Despite the success of deep learning, training deep networks in biologically plausible and hardware-efficient ways remains an open challenge. Feedback a

Why That Robot? A Qualitative Analysis of Justification Strategies for Robot Color Selection Across Occupational Contexts

SafetyDGX agent

arXiv:2603.28919v2 Announce Type: replace Abstract: As robots increasingly enter the workforce, human-robot interaction (HRI) must address how implicit social biases influence user preferences. This p

Zero-shot Transfer of Reinforcement Learning Control Policies for the Swing-Up and Stabilization of a Cart-Pole System

SafetyDGX agent

arXiv:2606.22145v1 Announce Type: new Abstract: Reinforcement learning (RL) is a powerful and convenient tool to modernize controller design. In this work, we study the zero-shot transfer of RL-based

ZeroDex: Zero-Shot Long-Horizon Dexterous Manipulation via Multi-View 3D-Grounded VLM Reasoning

SafetyDGX agent

arXiv:2606.19340v2 Announce Type: replace Abstract: We present ZeroDex, a zero-shot framework for long-horizon dexterous manipulation that grounds language instructions into executable 3D task plans f

22 Jun 2026

An update. A US official tells me that Sen. Warner misunderstood the NSA director Gen. Rudd in this case. Rudd did use the 'hours, not weeks…

SafetyDGX agent

An update. A US official tells me that Sen. Warner misunderstood the NSA director Gen. Rudd in this case. Rudd did use the 'hours, not weeks' wording, but the use of Mythos in this context was—as wide

'Because I'm going to be smarter than you' This clip is from a new short film by @ForegoneFilms on the race to develop superintelligence. Pl…

SafetyDGX agent

'Because I'm going to be smarter than you' This clip is from a new short film by @ForegoneFilms on the race to develop superintelligence. Please watch the full 16 minute film (link below) to help unde

By the end of the decade and maybe a lot sooner, we will wonder why we ever treated these folks like they were gods.

SafetyDGX agent

By the end of the decade and maybe a lot sooner, we will wonder why we ever treated these folks like they were gods. “The more I listen to AI company CEOs, the stranger they sound,” says another tech

Great report on LLM agent communication protocols. Communication is a huge bottleneck in multi-agent systems. (worth bookmarking) The report…

SafetyDGX agent

Great report on LLM agent communication protocols. Communication is a huge bottleneck in multi-agent systems. (worth bookmarking) The report builds a five-dimensional taxonomy (counterparty, payload,

have been thinking a bunch about model routing and related things current thoughts here, would love feedback: 1/ there is a difference betwe…

SafetyDGX agent

have been thinking a bunch about model routing and related things current thoughts here, would love feedback: 1/ there is a difference between 'model routing' and 'model council' 'model routing' = rou

Import AI 462: Superpersuasion; self-sustaining AI; paths to ASI

SafetyDGX agent

This newsletter issue covers three major AI topics: techniques for making AI systems more persuasive ('superpersuasion'), the concept of self-sustaining or self-improving AI systems, and various theor

LLMs can’t be trusted to follow rules. Which means they can’t be trusted, period. If we are going to solve alignment we must move on.

SafetyDGX agent

LLMs can’t be trusted to follow rules. Which means they can’t be trusted, period. If we are going to solve alignment we must move on. Large language models can be persuaded to break their own rules. N

New piece in @TheAtlantic! We always hear that AI will cure cancer, and I would immediately benefit if it did. But I argue that racing ahead…

SafetyDGX agent

New piece in @TheAtlantic! We always hear that AI will cure cancer, and I would immediately benefit if it did. But I argue that racing ahead on generalist AI models creates unclear benefits for cancer

Only one of these two countries is risking blowing up its entire economy on speculation.

SafetyDGX agent

Gary Marcus compares two countries' economic risk-taking, suggesting one is engaging in dangerous speculation while the other is not. The post likely discusses contrasting approaches to emerging techn

Our new incident response report exposes a coordinated network of 340 Facebook pages, largely operated from Vietnam, flooding feeds across C…

SafetyDGX agent

Our new incident response report exposes a coordinated network of 340 Facebook pages, largely operated from Vietnam, flooding feeds across Canada, the US, the UK, and Europe with AI-fabricated news, a

President Trump signs two executive orders aimed at speeding the development of advanced quantum computers and mitigating the security threats they present (Amrith Ramkumar/Wall Street Journal)

SafetyDGX agent

Amrith Ramkumar / Wall Street Journal: President Trump signs two executive orders aimed at speeding the development of advanced quantum computers and mitigating the security threats they present — Adm

Professor @GaryMarcus is an American psychologist, cognitive scientist, and author, known for his research on the intersection of cognitive …

SafetyDGX agent

Professor @GaryMarcus is an American psychologist, cognitive scientist, and author, known for his research on the intersection of cognitive psychology, neuroscience, and artificial intelligence. What

Sober minded folks need to understand the cash out happening in this bubble peak. Watch the @WSJ discussion 👇

SafetyDGX agent

Sober minded folks need to understand the cash out happening in this bubble peak. Watch the @WSJ discussion 👇 How big tech hides the true cost of AI @wsj https://youtu.be/YrJzjC4kKCY?si=o1CiKxLaSVWOhK

The largest LLM-as-a-Judge reliability audit yet. Researchers ran 21 judges from nine providers over roughly 541,000 judgments on MT-Bench, …

SafetyDGX agent

The largest LLM-as-a-Judge reliability audit yet. Researchers ran 21 judges from nine providers over roughly 541,000 judgments on MT-Bench, JudgeBench, and RewardBench. Findings: Validating a judge wi

Thoughts and prayers for those who bought $SPCX at 225 last week.

SafetyDGX agent

Gary Marcus posted a critical commentary on X regarding SPCX stock, sarcastically offering 'thoughts and prayers' to investors who purchased the stock at 225 in the previous week, implying the investm

← Previous
1…108109110111112…242
Next →