AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
18 May 2026

Reference-Free Reinforcement Learning Fine-Tuning for MT: A Seq2Seq Perspective

SafetyDGX agent

arXiv:2605.15976v1 Announce Type: cross Abstract: Production machine translation relies overwhelmingly on encoder-decoder Seq2Seq models, yet reinforcement learning approaches to MT fine-tuning have l

Reference Games as a Testbed for the Alignment of Model Uncertainty and Clarification Requests

SafetyDGX agent

arXiv:2601.07820v2 Announce Type: replace Abstract: In human conversation, both interlocutors play an active role in maintaining mutual understanding. When listeners are uncertain about what speakers

Res^2CLIP: Few-Shot Generalist Anomaly Detection with Residual-to-Residual Alignment

SafetyDGX agent

arXiv:2605.16171v1 Announce Type: new Abstract: Few-shot Generalist Anomaly Detection requires models to generalize to novel categories without retraining, posing significant challenges in real-world


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Residual Reinforcement Learning for Robot Teleoperation under Stochastic Delays

SafetyDGX agent

arXiv:2605.15480v1 Announce Type: cross Abstract: Stochastic communication delays in teleoperation introduce signal discontinuities that undermine control stability and degrade control performance. Co

Response-Conditioned Parallel-to-Sequential Orchestration for Multi-Agent Systems

SafetyDGX agent

arXiv:2605.15573v1 Announce Type: new Abstract: Multi-agent systems can solve complex tasks through collaboration between multiple Large Language Model agents. Existing collaboration frameworks typica

Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models

SafetyDGX agent

arXiv:2605.15792v1 Announce Type: new Abstract: The long-standing goal of multimodal AI is to build unified models in which visual understanding and visual generation mutually enhance one another. Des

RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably

SafetyDGX agent

arXiv:2605.15514v1 Announce Type: cross Abstract: We identify intrinsic limitations of Rotary Positional Embeddings (RoPE) in Transformer-based long-context language models. Our theoretical analysis a

SAFE Quantum Machine Learning with Variational Quantum Classifiers

SafetyDGX agent

arXiv:2605.16067v1 Announce Type: new Abstract: We propose a variational quantum classifier operating on high dimensional deep representations via amplitude encoding, stabilized by a learnable classic

SafeGPT: Preventing Data Leakage and Unethical Outputs in Enterprise LLM Use

SafetyDGX agent

arXiv:2601.06366v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are transforming enterprise workflows but introduce security and ethics challenges when employees inadvertently s

ScreenSearch: Uncertainty-Aware OS Exploration

SafetyDGX agent

arXiv:2605.16024v1 Announce Type: new Abstract: Desktop GUI agents operate under partial observability: visually similar screens can correspond to different underlying workflow states, so locally plau

Second-Order Multi-Level Variance Correction for Modality Competition in Multimodal Models

SafetyDGX agent

arXiv:2605.16165v1 Announce Type: cross Abstract: Autoregressive next-token training offers a unified formulation for image generation and text understanding, but it also creates strong modality compe

Seeing What Matters: Visual Preference Policy Optimization for Visual Generation

SafetyDGX agent

arXiv:2511.18719v4 Announce Type: replace Abstract: Reinforcement learning (RL) has become a powerful tool for post-training visual generative models, with Group Relative Policy Optimization (GRPO) in

Self-Supervised Learning by Curvature Alignment

SafetyDGX agent

arXiv:2511.17426v2 Announce Type: replace-cross Abstract: Self-supervised learning (SSL) has recently advanced through non-contrastive methods that couple an invariance term with variance, covariance,

Semi-MedRef: Semi-Supervised Medical Referring Image Segmentation with Cross-Modal Alignment

SafetyDGX agent

arXiv:2605.15720v1 Announce Type: new Abstract: Medical referring image segmentation (MRIS) requires pixel-level masks aligned with textual descriptions of anatomical locations, making annotation cost

seriou question: how do you handle an intellectual doppelganger who has systematically started adopting every position you have argued for f…

SafetyDGX agent

seriou question: how do you handle an intellectual doppelganger who has systematically started adopting every position you have argued for for 30 years while presenting each idea as if it were his own

Sign-Separated Finite-Time Error Analysis of Q-Learning

SafetyDGX agent

arXiv:2605.16103v1 Announce Type: new Abstract: This paper develops a sign-separated finite-time error analysis for constant step-size Q-learning. Starting from the switching-system representation, th

SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization

SafetyDGX agent

arXiv:2604.02268v2 Announce Type: replace Abstract: Agent skills, structured packages of procedural knowledge and executable resources that agents dynamically load at inference time, have become a rel

SkiP: When to Skip and When to Refine for Efficient Robot Manipulation

SafetyDGX agent

arXiv:2605.15536v1 Announce Type: cross Abstract: Previous imitation learning policies predict future actions at every control step, whether in smooth motion phases or precise, contact-rich operation

SLIP & ETHICS: Graduated Intervention for AI Emotional Companions

SafetyDGX agent

arXiv:2605.15915v1 Announce Type: cross Abstract: AI emotional companions face a safety-rapport paradox: restrictive safeguards can damage supportive alliance, while permissive systems risk user harm.

STABLE: Simulation-Ready Tabletop Layout Generation via a Semantics-Physics Dual System

SafetyDGX agent

arXiv:2605.16137v1 Announce Type: new Abstract: Generating simulation-ready tabletop scenes from task instructions is an intriguing and promising research direction in the field of Embodied AI. Howeve

Task-Semantic Graph-Driven Distributed Agent Networking for Underwater Target Tracking

SafetyDGX agent

arXiv:2605.15528v1 Announce Type: new Abstract: Autonomous underwater vehicle (AUV) swarms are emerging as intelligent underwater networks, where each node must sense, communicate, process local data,

TemplateRL: Structured Template-Guided Reinforcement Learning for LLM Reasoning

SafetyDGX agent

arXiv:2505.15692v5 Announce Type: replace Abstract: Reinforcement learning (RL) has emerged as an effective paradigm for enhancing model reasoning. However, existing RL methods like GRPO typically rel

Terrain Consistent Reference-Guided RL for Humanoid Navigation Autonomy

SafetyDGX agent

arXiv:2605.15517v1 Announce Type: new Abstract: We present a method for training reference-guided, perceptive reinforcement learning locomotion policies for humanoid robots in which reference trajecto

the AI trial of the century ended with a procedural whimper rather a bang; the jury agreed that Musk was too late but never weighed in on th…

SafetyDGX agent

the AI trial of the century ended with a procedural whimper rather a bang; the jury agreed that Musk was too late but never weighed in on the questions of whether OpenAI did was legitimate. and so we

The pure LLM debate - which I had for many years, here and elsewhere - is indeed no longer relevant. Why? Because I won; nobody uses pure LL…

SafetyDGX agent

The pure LLM debate - which I had for many years, here and elsewhere - is indeed no longer relevant. Why? Because I won; nobody uses pure LLMs anymore. Nowadays all deployed objects are neurosymbolic,

The U.S. faces a chaotic #AI regulatory patchwork with over 1,200 state bills introduced in 2025 and no unified federal framework, researche…

SafetyDGX agent

The U.S. faces a chaotic #AI regulatory patchwork with over 1,200 state bills introduced in 2025 and no unified federal framework, researchers @JeffSonnenfeld and Stephen Henriques of @YaleSOM and @Ga

TopoEvo: A Topology-Aware Self-Evolving Multi-Agent Framework for Root Cause Analysis in Microservices

SafetyDGX agent

arXiv:2605.15611v1 Announce Type: new Abstract: Root cause analysis (RCA) in microservices is challenging due to (i) noisy and heterogeneous multimodal observability (metrics, logs, traces), (ii) casc

Towards a more realistic evaluation of machine learning models for bearing fault diagnosis

SafetyDGX agent

arXiv:2509.22267v4 Announce Type: replace Abstract: Reliable detection of bearing faults is essential for maintaining the safety and operational efficiency of rotating machinery. While recent advances

Towards Code-Oriented LM Embeddings for Surrogate-Assisted Neural Architecture Search

SafetyDGX agent

arXiv:2605.15649v1 Announce Type: new Abstract: Developing effective surrogates (performance predictors) for Neural Architecture Search (NAS) typically requires expensive fine-tuning or the engineerin

Towards Trustworthy and Explainable AI for Perception Models: From Concept to Prototype Vehicle Deployment

SafetyDGX agent

arXiv:2605.16087v1 Announce Type: cross Abstract: Deep Neural Networks have become the dominant solution for Autonomous Driving perception, but their opacity conflicts with emerging Trustworthy AI gui

Trump just shook down the US government for $1.776 billion dollars of taxpayer money to pay off his buddies.

SafetyDGX agent

I can't verify this claim without access to the actual post and current information. The headline uses inflammatory language ('shook down') that suggests opinion rather than neutral reporting. To crea

Unified High-Probability Analysis of Stochastic Variance-Reduced Estimation

SafetyDGX agent

arXiv:2605.15388v1 Announce Type: new Abstract: Stochastic estimators are fundamental to large-scale optimization, where population quantities must be inferred from noisy oracle observations. Although

Unveiling the Black Box: A Multi-Layer Framework for Explaining Reinforcement Learning-Based Cyber Agents

SafetyDGX agent

arXiv:2505.11708v3 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) agents are increasingly used to simulate sophisticated cyberattacks, but their decision-making processes remain op

Verifiable Agentic Infrastructure: Proof-Derived Authorization for Sovereign AI Systems

SafetyDGX agent

arXiv:2605.15228v1 Announce Type: new Abstract: Modern cloud and enterprise systems rely on identity-centric authorization, assuming that callers possessing valid credentials are safe to execute comma

Video Models Can Reason with Verifiable Rewards

SafetyDGX agent

arXiv:2605.15458v1 Announce Type: new Abstract: Video diffusion models have made rapid progress in perceptual realism and temporal coherence, but they remain primarily optimized for plausible generati

VSPO: Vector-Steered Policy Optimization for Behavioral Control

SafetyDGX agent

arXiv:2605.15604v1 Announce Type: cross Abstract: Modern language models often need to optimize a primary accuracy objective while also accommodating secondary behavioral preferences, such as verbosit

What Is Preference Optimization Doing, and Why?

SafetyDGX agent

arXiv:2512.00778v2 Announce Type: replace Abstract: Preference optimization (PO) is indispensable for large language models (LLMs), with methods such as direct preference optimization (DPO) and proxim

When and Why Adversarial Training Improves PINNs: A Neural Tangent Kernel Perspective

SafetyDGX agent

arXiv:2605.15959v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) are powerful surrogates for differential equations but are notoriously difficult to train due to spectral bia

When Importance Sampling Misallocates Credit: Asymmetric Ratios for Outcome-Supervised RL

SafetyDGX agent

arXiv:2510.06062v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown great promise in large language models (LLMs) post-training, which typically rely on token-level clipping to m

When Latent Geometry Is Not Enough: Draft-Conditioned Latent Refinement for Non-Autoregressive Text Generation

SafetyDGX agent

arXiv:2605.15557v1 Announce Type: new Abstract: Continuous diffusion and flow models are attractive for non-autoregressive text generation because they can update all positions in parallel. A major di

Whole-body motion planning and safety-critical control for aerial manipulation

SafetyDGX agent

arXiv:2511.02342v3 Announce Type: replace Abstract: Aerial manipulation combines the maneuverability of multirotors with the dexterity of robotic arms to perform complex tasks in cluttered spaces. Yet

You know how I said yesterday that my feed is littered with people just making stuff up? We didn’t have 71% of Americans opposed to cell pho…

SafetyDGX agent

You know how I said yesterday that my feed is littered with people just making stuff up? We didn’t have 71% of Americans opposed to cell phones. Or ipods. Or laptops. We *do* have 71% of Americans opp

17 May 2026

Bernie Madoff told the WSJ he was averaging annual returns of 16.3%. Sam Altman has promised annual returns of 17.5% History may not repeat …

SafetyDGX agent

Gary Marcus draws a historical parallel between Bernie Madoff's claimed 16.3% annual returns (which proved to be a Ponzi scheme) and Sam Altman's promised 17.5% returns, warning that similarly implaus

🚨Breaking new study: memory in LLM agents still can’t really be trusted, even after over trillion dollars has gone into the development of …

SafetyDGX agent

🚨Breaking new study: memory in LLM agents still can’t really be trusted, even after over trillion dollars has gone into the development of the field. Excited to share our new paper: “Useful Memories B

GDS weighs in on the NHS's decision to retreat from Open Source

SafetyDGX agent

GDS weighs in on the NHS's decision to retreat from Open Source Terence Eden continues his coverage of the NHS' poorly considered decision to close down access to their open source repositories in res

profound comment by Daniel Eth: AI (long term) will probably change things a lot — but in ways that nobody can really predict.

SafetyDGX agent

profound comment by Daniel Eth: AI (long term) will probably change things a lot — but in ways that nobody can really predict. Your kids will not work 3.5 days a week or live to 100. They might work m

the whole point of databases is to keep reliable records of values that change over time. in general, they can be trusted to be stable over …

SafetyDGX agent

the whole point of databases is to keep reliable records of values that change over time. in general, they can be trusted to be stable over time. the whole point of LLMs is to raise money. in general

This is a nice article (not sure how I stumbled upon it a month later) I directionally agree with it in that: ✅ I have a massive bias for sl…

SafetyDGX agent

This is a nice article (not sure how I stumbled upon it a month later) I directionally agree with it in that: ✅ I have a massive bias for slope, grit, and scrappiness in candidates vs. pure experience

This was always and only the sole and exclusive purpose of the UK online censorship act, to help Labour nuke its enemies

SafetyDGX agent

This was always and only the sole and exclusive purpose of the UK online censorship act, to help Labour nuke its enemies 🚨 Labour is using the “Online Safety Act” to silence political opponents, and T

trillions. and i did try to warn them. repeatedly. for many years. 🤷‍♂️

SafetyDGX agent

trillions. and i did try to warn them. repeatedly. for many years. 🤷‍♂️ @GaryMarcus everyone pretended reliability was a detail theyd fix later now we find out billions later it was always the core pr

Trump just got exposed for running the biggest insider trading operation in American history. Nancy Pelosi traded $5 million in stocks and C…

SafetyDGX agent

Trump just got exposed for running the biggest insider trading operation in American history. Nancy Pelosi traded 5 million in stocks and Congress lost its mind. Trump literally executed 750 MILLION w

What I am about to describe ain’t AGI; it’s a sign of a trillion dollar trainwreck. If I had told you in 2022 that the 2026 version of GPT (…

SafetyDGX agent

What I am about to describe ain’t AGI; it’s a sign of a trillion dollar trainwreck. If I had told you in 2022 that the 2026 version of GPT (which by the way would only be GPT 5.5 and not GPT-6 or 7 li

16 May 2026

Dear @geoffreyhinton, You have to stop lying about me. First was the apparently faked quote on your web page that you couldn’t provide a sou…

SafetyDGX agent

Dear @geoffreyhinton, You have to stop lying about me. First was the apparently faked quote on your web page that you couldn’t provide a source for (literally the only source I found was on your own w

Gary Marcus: '...these companies not only have they stolen every book that's ever been written, but they steal from each other'. Read that a…

SafetyDGX agent

Gary Marcus critiques AI companies for training large language models on copyrighted books and other proprietary content without permission or compensation to authors. He argues that these companies e

holy shit lmao @Gavriel_Cohen he's seriously using this thing for conducting the foreign policy/parliamentary affairs of singapore - and sha…

SafetyDGX agent

holy shit lmao @Gavriel_Cohen he's seriously using this thing for conducting the foreign policy/parliamentary affairs of singapore - and sharing his stack on how he is hacking around WhatsApp and doin

Is it me, or are we living in absurdist, Kurt Vonnegut-like parody of how AI (and society) could go wrong?

SafetyDGX agent

Gary Marcus discusses parallels between current AI development and societal trends with absurdist themes reminiscent of Kurt Vonnegut's work, suggesting that real-world outcomes increasingly resemble

LLM-powered AI agents are gonna be great! You should totally trust them!

SafetyDGX agent

Gary Marcus expresses optimism about the potential of LLM-powered AI agents in a post on X (formerly Twitter). The post advocates for confidence in these AI systems, though without access to the full

OpenAI partners with Malta's AI for All initiative to give citizens a free year of ChatGPT Plus if they complete a University of Malta AI literacy course (Cointelegraph)

SafetyDGX agent

Cointelegraph: OpenAI partners with Malta's AI for All initiative to give citizens a free year of ChatGPT Plus if they complete a University of Malta AI literacy course — Cointelegraph is committed to

pleased to be in this and encourage everyone to read:

SafetyDGX agent

pleased to be in this and encourage everyone to read: Awesome new theme issue of Philosophical Transactions: ‘World models in natural and artificial intelligence’ Thank you @adamsafron! https://royals

Real AGI would not do this. Even after a trillion dollars in LLMs still do.

SafetyDGX agent

Real AGI would not do this. Even after a trillion dollars in LLMs still do. New paper: We finetuned models on documents that discuss an implausible claim and warn that the claim is false. Models ended

← Previous
1…138139140141142…214
Next →