AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
18 May 2026

RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably

SafetyDGX agent

arXiv:2605.15514v1 Announce Type: cross Abstract: We identify intrinsic limitations of Rotary Positional Embeddings (RoPE) in Transformer-based long-context language models. Our theoretical analysis a

SafeGPT: Preventing Data Leakage and Unethical Outputs in Enterprise LLM Use

SafetyDGX agent

arXiv:2601.06366v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are transforming enterprise workflows but introduce security and ethics challenges when employees inadvertently s

ScreenSearch: Uncertainty-Aware OS Exploration

SafetyDGX agent

arXiv:2605.16024v1 Announce Type: new Abstract: Desktop GUI agents operate under partial observability: visually similar screens can correspond to different underlying workflow states, so locally plau

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Second-Order Multi-Level Variance Correction for Modality Competition in Multimodal Models

SafetyDGX agent

arXiv:2605.16165v1 Announce Type: cross Abstract: Autoregressive next-token training offers a unified formulation for image generation and text understanding, but it also creates strong modality compe

Seeing What Matters: Visual Preference Policy Optimization for Visual Generation

SafetyDGX agent

arXiv:2511.18719v4 Announce Type: replace Abstract: Reinforcement learning (RL) has become a powerful tool for post-training visual generative models, with Group Relative Policy Optimization (GRPO) in

Self-Supervised Learning by Curvature Alignment

SafetyDGX agent

arXiv:2511.17426v2 Announce Type: replace-cross Abstract: Self-supervised learning (SSL) has recently advanced through non-contrastive methods that couple an invariance term with variance, covariance,

Semi-MedRef: Semi-Supervised Medical Referring Image Segmentation with Cross-Modal Alignment

SafetyDGX agent

arXiv:2605.15720v1 Announce Type: new Abstract: Medical referring image segmentation (MRIS) requires pixel-level masks aligned with textual descriptions of anatomical locations, making annotation cost

seriou question: how do you handle an intellectual doppelganger who has systematically started adopting every position you have argued for f…

SafetyDGX agent

seriou question: how do you handle an intellectual doppelganger who has systematically started adopting every position you have argued for for 30 years while presenting each idea as if it were his own

Sign-Separated Finite-Time Error Analysis of Q-Learning

SafetyDGX agent

arXiv:2605.16103v1 Announce Type: new Abstract: This paper develops a sign-separated finite-time error analysis for constant step-size Q-learning. Starting from the switching-system representation, th

SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization

SafetyDGX agent

arXiv:2604.02268v2 Announce Type: replace Abstract: Agent skills, structured packages of procedural knowledge and executable resources that agents dynamically load at inference time, have become a rel

SkiP: When to Skip and When to Refine for Efficient Robot Manipulation

SafetyDGX agent

arXiv:2605.15536v1 Announce Type: cross Abstract: Previous imitation learning policies predict future actions at every control step, whether in smooth motion phases or precise, contact-rich operation

STABLE: Simulation-Ready Tabletop Layout Generation via a Semantics-Physics Dual System

SafetyDGX agent

arXiv:2605.16137v1 Announce Type: new Abstract: Generating simulation-ready tabletop scenes from task instructions is an intriguing and promising research direction in the field of Embodied AI. Howeve

Task-Semantic Graph-Driven Distributed Agent Networking for Underwater Target Tracking

SafetyDGX agent

arXiv:2605.15528v1 Announce Type: new Abstract: Autonomous underwater vehicle (AUV) swarms are emerging as intelligent underwater networks, where each node must sense, communicate, process local data,

TemplateRL: Structured Template-Guided Reinforcement Learning for LLM Reasoning

SafetyDGX agent

arXiv:2505.15692v5 Announce Type: replace Abstract: Reinforcement learning (RL) has emerged as an effective paradigm for enhancing model reasoning. However, existing RL methods like GRPO typically rel

Terrain Consistent Reference-Guided RL for Humanoid Navigation Autonomy

SafetyDGX agent

arXiv:2605.15517v1 Announce Type: new Abstract: We present a method for training reference-guided, perceptive reinforcement learning locomotion policies for humanoid robots in which reference trajecto

the AI trial of the century ended with a procedural whimper rather a bang; the jury agreed that Musk was too late but never weighed in on th…

SafetyDGX agent

the AI trial of the century ended with a procedural whimper rather a bang; the jury agreed that Musk was too late but never weighed in on the questions of whether OpenAI did was legitimate. and so we

The pure LLM debate - which I had for many years, here and elsewhere - is indeed no longer relevant. Why? Because I won; nobody uses pure LL…

SafetyDGX agent

The pure LLM debate - which I had for many years, here and elsewhere - is indeed no longer relevant. Why? Because I won; nobody uses pure LLMs anymore. Nowadays all deployed objects are neurosymbolic,

The U.S. faces a chaotic #AI regulatory patchwork with over 1,200 state bills introduced in 2025 and no unified federal framework, researche…

SafetyDGX agent

The U.S. faces a chaotic #AI regulatory patchwork with over 1,200 state bills introduced in 2025 and no unified federal framework, researchers @JeffSonnenfeld and Stephen Henriques of @YaleSOM and @Ga

TopoEvo: A Topology-Aware Self-Evolving Multi-Agent Framework for Root Cause Analysis in Microservices

SafetyDGX agent

arXiv:2605.15611v1 Announce Type: new Abstract: Root cause analysis (RCA) in microservices is challenging due to (i) noisy and heterogeneous multimodal observability (metrics, logs, traces), (ii) casc

Towards Code-Oriented LM Embeddings for Surrogate-Assisted Neural Architecture Search

SafetyDGX agent

arXiv:2605.15649v1 Announce Type: new Abstract: Developing effective surrogates (performance predictors) for Neural Architecture Search (NAS) typically requires expensive fine-tuning or the engineerin

Trump just shook down the US government for $1.776 billion dollars of taxpayer money to pay off his buddies.

SafetyDGX agent

I can't verify this claim without access to the actual post and current information. The headline uses inflammatory language ('shook down') that suggests opinion rather than neutral reporting. To crea

Unified High-Probability Analysis of Stochastic Variance-Reduced Estimation

SafetyDGX agent

arXiv:2605.15388v1 Announce Type: new Abstract: Stochastic estimators are fundamental to large-scale optimization, where population quantities must be inferred from noisy oracle observations. Although

Unveiling the Black Box: A Multi-Layer Framework for Explaining Reinforcement Learning-Based Cyber Agents

SafetyDGX agent

arXiv:2505.11708v3 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) agents are increasingly used to simulate sophisticated cyberattacks, but their decision-making processes remain op

Verifiable Agentic Infrastructure: Proof-Derived Authorization for Sovereign AI Systems

SafetyDGX agent

arXiv:2605.15228v1 Announce Type: new Abstract: Modern cloud and enterprise systems rely on identity-centric authorization, assuming that callers possessing valid credentials are safe to execute comma

Video Models Can Reason with Verifiable Rewards

SafetyDGX agent

arXiv:2605.15458v1 Announce Type: new Abstract: Video diffusion models have made rapid progress in perceptual realism and temporal coherence, but they remain primarily optimized for plausible generati

VSPO: Vector-Steered Policy Optimization for Behavioral Control

SafetyDGX agent

arXiv:2605.15604v1 Announce Type: cross Abstract: Modern language models often need to optimize a primary accuracy objective while also accommodating secondary behavioral preferences, such as verbosit

What Is Preference Optimization Doing, and Why?

SafetyDGX agent

arXiv:2512.00778v2 Announce Type: replace Abstract: Preference optimization (PO) is indispensable for large language models (LLMs), with methods such as direct preference optimization (DPO) and proxim

When and Why Adversarial Training Improves PINNs: A Neural Tangent Kernel Perspective

SafetyDGX agent

arXiv:2605.15959v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) are powerful surrogates for differential equations but are notoriously difficult to train due to spectral bia

When Importance Sampling Misallocates Credit: Asymmetric Ratios for Outcome-Supervised RL

SafetyDGX agent

arXiv:2510.06062v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown great promise in large language models (LLMs) post-training, which typically rely on token-level clipping to m

When Latent Geometry Is Not Enough: Draft-Conditioned Latent Refinement for Non-Autoregressive Text Generation

SafetyDGX agent

arXiv:2605.15557v1 Announce Type: new Abstract: Continuous diffusion and flow models are attractive for non-autoregressive text generation because they can update all positions in parallel. A major di

You know how I said yesterday that my feed is littered with people just making stuff up? We didn’t have 71% of Americans opposed to cell pho…

SafetyDGX agent

You know how I said yesterday that my feed is littered with people just making stuff up? We didn’t have 71% of Americans opposed to cell phones. Or ipods. Or laptops. We *do* have 71% of Americans opp

17 May 2026

Bernie Madoff told the WSJ he was averaging annual returns of 16.3%. Sam Altman has promised annual returns of 17.5% History may not repeat …

SafetyDGX agent

Gary Marcus draws a historical parallel between Bernie Madoff's claimed 16.3% annual returns (which proved to be a Ponzi scheme) and Sam Altman's promised 17.5% returns, warning that similarly implaus

🚨Breaking new study: memory in LLM agents still can’t really be trusted, even after over trillion dollars has gone into the development of …

SafetyDGX agent

🚨Breaking new study: memory in LLM agents still can’t really be trusted, even after over trillion dollars has gone into the development of the field. Excited to share our new paper: “Useful Memories B

GDS weighs in on the NHS's decision to retreat from Open Source

SafetyDGX agent

GDS weighs in on the NHS's decision to retreat from Open Source Terence Eden continues his coverage of the NHS' poorly considered decision to close down access to their open source repositories in res

profound comment by Daniel Eth: AI (long term) will probably change things a lot — but in ways that nobody can really predict.

SafetyDGX agent

profound comment by Daniel Eth: AI (long term) will probably change things a lot — but in ways that nobody can really predict. Your kids will not work 3.5 days a week or live to 100. They might work m

the whole point of databases is to keep reliable records of values that change over time. in general, they can be trusted to be stable over …

SafetyDGX agent

the whole point of databases is to keep reliable records of values that change over time. in general, they can be trusted to be stable over time. the whole point of LLMs is to raise money. in general

This is a nice article (not sure how I stumbled upon it a month later) I directionally agree with it in that: ✅ I have a massive bias for sl…

SafetyDGX agent

This is a nice article (not sure how I stumbled upon it a month later) I directionally agree with it in that: ✅ I have a massive bias for slope, grit, and scrappiness in candidates vs. pure experience

trillions. and i did try to warn them. repeatedly. for many years. 🤷‍♂️

SafetyDGX agent

trillions. and i did try to warn them. repeatedly. for many years. 🤷‍♂️ @GaryMarcus everyone pretended reliability was a detail theyd fix later now we find out billions later it was always the core pr

Trump just got exposed for running the biggest insider trading operation in American history. Nancy Pelosi traded $5 million in stocks and C…

SafetyDGX agent

Trump just got exposed for running the biggest insider trading operation in American history. Nancy Pelosi traded 5 million in stocks and Congress lost its mind. Trump literally executed 750 MILLION w

16 May 2026

Dear @geoffreyhinton, You have to stop lying about me. First was the apparently faked quote on your web page that you couldn’t provide a sou…

SafetyDGX agent

Dear @geoffreyhinton, You have to stop lying about me. First was the apparently faked quote on your web page that you couldn’t provide a source for (literally the only source I found was on your own w

Gary Marcus: '...these companies not only have they stolen every book that's ever been written, but they steal from each other'. Read that a…

SafetyDGX agent

Gary Marcus critiques AI companies for training large language models on copyrighted books and other proprietary content without permission or compensation to authors. He argues that these companies e

holy shit lmao @Gavriel_Cohen he's seriously using this thing for conducting the foreign policy/parliamentary affairs of singapore - and sha…

SafetyDGX agent

holy shit lmao @Gavriel_Cohen he's seriously using this thing for conducting the foreign policy/parliamentary affairs of singapore - and sharing his stack on how he is hacking around WhatsApp and doin

Is it me, or are we living in absurdist, Kurt Vonnegut-like parody of how AI (and society) could go wrong?

SafetyDGX agent

Gary Marcus discusses parallels between current AI development and societal trends with absurdist themes reminiscent of Kurt Vonnegut's work, suggesting that real-world outcomes increasingly resemble

LLM-powered AI agents are gonna be great! You should totally trust them!

SafetyDGX agent

Gary Marcus expresses optimism about the potential of LLM-powered AI agents in a post on X (formerly Twitter). The post advocates for confidence in these AI systems, though without access to the full

OpenAI partners with Malta's AI for All initiative to give citizens a free year of ChatGPT Plus if they complete a University of Malta AI literacy course (Cointelegraph)

SafetyDGX agent

Cointelegraph: OpenAI partners with Malta's AI for All initiative to give citizens a free year of ChatGPT Plus if they complete a University of Malta AI literacy course — Cointelegraph is committed to

pleased to be in this and encourage everyone to read:

SafetyDGX agent

pleased to be in this and encourage everyone to read: Awesome new theme issue of Philosophical Transactions: ‘World models in natural and artificial intelligence’ Thank you @adamsafron! https://royals

Real AGI would not do this. Even after a trillion dollars in LLMs still do.

SafetyDGX agent

Real AGI would not do this. Even after a trillion dollars in LLMs still do. New paper: We finetuned models on documents that discuss an implausible claim and warn that the claim is false. Models ended

tesla robotaxis going about as well you might expect.

SafetyDGX agent

Gary Marcus, a prominent AI researcher and critic, comments on Tesla's robotaxi development progress, likely offering skeptical or cautionary observations about the practical challenges and timeline o

Truly an all-star cast, on one of the most important questions in AI. Thrilled to see some many people finally willing to confront the hard …

SafetyDGX agent

Truly an all-star cast, on one of the most important questions in AI. Thrilled to see some many people finally willing to confront the hard questions of how we can move beyond LLMs, and into what worl

💯. Way too much focus on language models.

SafetyDGX agent

💯. Way too much focus on language models. Fei-Fei Li warns that AI may be staring too hard at language models. The world is not just text on a screen. It is physical, visual, spatial, and always chang

15 May 2026

A cross-species neural foundation model for end-to-end speech decoding

SafetyDGX agent

arXiv:2511.21740v5 Announce Type: replace-cross Abstract: Speech brain-computer interfaces (BCIs) aim to restore communication for people with paralysis by translating neural activity into text. Most

A Security Analysis of the OpenClaw AI Agent Framework

SafetyDGX agent

arXiv:2603.27517v3 Announce Type: replace-cross Abstract: AI agent frameworks connecting large language model (LLM) reasoning to host execution surfaces -- shell, filesystem, containers, and messaging

Achieving Approximate Symmetry Is Exponentially Easier than Exact Symmetry

SafetyDGX agent

arXiv:2512.11855v2 Announce Type: replace-cross Abstract: Enforcing exact symmetry in machine learning models often yields significant gains in scientific applications, serving as a powerful inductive

Active Learners as Efficient PRP Rerankers

SafetyDGX agent

arXiv:2605.14236v1 Announce Type: cross Abstract: Pairwise Ranking Prompting (PRP) elicits pairwise preference judgments from an LLM, which are then aggregated into a ranking, usually via classical so

ActivePusher: Active Learning and Planning with Residual Physics for Nonprehensile Manipulation

SafetyDGX agent

arXiv:2506.04646v4 Announce Type: replace-cross Abstract: Planning with learned dynamics models offers a promising approach toward versatile real-world manipulation, particularly in nonprehensile sett

AIS: Adaptive Importance Sampling for Quantized RL

SafetyDGX agent

arXiv:2605.13907v1 Announce Type: cross Abstract: Reinforcement learning (RL) for large language models (LLMs) is dominated by the cost of rollout generation, which has motivated the use of low-precis

Aligning Latent Geometry for Spherical Flow Matching in Image Generation

SafetyDGX agent

arXiv:2605.15193v1 Announce Type: new Abstract: Latent flow matching for image generation usually transports Gaussian noise to variational autoencoder latents along linear paths. Both endpoints, howev

AMiD: Knowledge Distillation for LLMs with alpha-mixture Assistant Distribution

SafetyDGX agent

arXiv:2510.15982v3 Announce Type: replace-cross Abstract: Autoregressive large language models (LLMs) have achieved remarkable improvement across many tasks but incur high computational and memory cos

Anti-Length Shift: Dynamic Outlier Truncation for Training Efficient Reasoning Models

SafetyDGX agent

arXiv:2601.03969v2 Announce Type: replace Abstract: Large reasoning models enhanced by reinforcement learning with verifiable rewards have achieved significant performance gains by extending their cha

AQKA: Active Quantum Kernel Acquisition Under a Shot Budget

SafetyDGX agent

arXiv:2605.14672v1 Announce Type: new Abstract: Estimating an N imes N quantum kernel from circuit fidelities requires Theta(N^2 S) measurement shots, the dominant bottleneck for deployment on near-te

← Previous
1…164165166167168…242
Next →