AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
Safety

In the future you have a choice. Do you engage brain? or Do you cheat? There will be other choices too, such as: Do you go to the casino? Or…

DGX agent

In the future you have a choice. Do you engage brain? or Do you cheat? There will be other choices too, such as: Do you go to the casino? Or to the library or maker space? I fear most will make the ea

safetygary-marcus--x
30 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Inference-Time Scaling of Verification: Self-Evolving Deep Research Agents via Test-Time Rubric-Guided Verification

DGX agent

arXiv:2601.15808v2 Announce Type: replace Abstract: Recent advances in Deep Research Agents (DRAs) are transforming automated knowledge discovery and problem-solving. While the majority of existing ef

safetyarxiv-cs-ai
30 Apr 2026
Safety

John Oliver @LastWeekTonight slays a pile of greedy tech CEOs, gives props to @GaryMarcus 😻 The tides are changing, and other folks like @w…

DGX agent

John Oliver @LastWeekTonight slays a pile of greedy tech CEOs, gives props to @GaryMarcus 😻 The tides are changing, and other folks like @wendyweeww are speaking out in support. Even though I’ve alrea

safetygary-marcus--x
30 Apr 2026
Safety

Learning Vision-Based Omnidirectional Navigation: A Teacher-Student Approach Using Monocular Depth Estimation

DGX agent

arXiv:2603.01999v2 Announce Type: replace-cross Abstract: Reliable obstacle avoidance in industrial settings demands 3D scene understanding, but widely used 2D LiDAR sensors perceive only a single hor

safetyarxiv-cs-cv
30 Apr 2026
Safety

Lifting Embodied World Models for Planning and Control

DGX agent

arXiv:2604.26182v1 Announce Type: cross Abstract: World models of embodied agents predict future observations conditioned on an action taken by the agent. For complex embodiments, action spaces are hi

safetyarxiv-cs-ai
30 Apr 2026
Safety

Lights Out: A Nighttime UAV Localization Framework Using Thermal Imagery and Semantic 3D Maps

DGX agent

arXiv:2604.26201v1 Announce Type: new Abstract: Reliable backup localization for unmanned aerial vehicles (UAVs) operating in GNSS-denied nighttime conditions remains an open challenge due to the seve

safetyarxiv-cs-ro
30 Apr 2026
Safety

Mapping the maturation of TCM as an adjuvant to radiotherapy

DGX agent

arXiv:2601.11923v2 Announce Type: replace Abstract: The integration of complementary medicine into oncology represents a paradigm shift that has seen to increasing adoption of Traditional Chinese Medi

safetyarxiv-cs-cl
30 Apr 2026
Safety

MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution

DGX agent

arXiv:2604.26283v1 Announce Type: cross Abstract: High-precision medical diagnosis relies not only on static imaging features but also on the implicit diagnostic memory experts instantly invoke during

safetyarxiv-cs-ai
30 Apr 2026
Safety

Meta has dropped 10% today.

DGX agent

Meta has dropped 10% today. BREAKING: Meta stock, META, extends losses to over -10% on the day and is now on track for its biggest daily decline since October 2025. Meta has erased -170 billion in mar

safetygary-marcus--x
30 Apr 2026
Safety

Mini-Batch Class Composition Bias in Link Prediction

DGX agent

arXiv:2604.25978v1 Announce Type: cross Abstract: Prior work on node classification has shown that Graph Neural Networks (GNNs) can learn representations that transfer across graphs, when underlying g

safetyarxiv-cs-ai
30 Apr 2026
Safety

MINOS: A Multimodal Evaluation Model for Bidirectional Generation Between Image and Text

DGX agent

arXiv:2506.02494v2 Announce Type: replace-cross Abstract: Evaluation is important for multimodal generation tasks, while traditional multimodal evaluation metrics suffer from several limitations. With

safetyarxiv-cs-ai
30 Apr 2026
Safety

My argument since day one:

DGX agent

My argument since day one: @GaryMarcus Going all-in on LLMs, as a technology meant to be an approach to creating general purpose ML tech whose capabilities could be described as 'AGI,' *could* perhaps

safetygary-marcus--x
30 Apr 2026
Safety

Near-Optimal Cryptographic Hardness of Learning With Homogeneous Halfspaces Under Gaussian Marginals

DGX agent

arXiv:2604.26446v1 Announce Type: new Abstract: We study three problems that involve identifying homogeneous halfspaces under Gaussian distributions: agnostic learning, one-sided reliable learning, an

safetyarxiv-cs-lg
30 Apr 2026
Safety

One underdiscussed part of the Google-DOD deal is how it blindsided the company’s own employees. One told @cogcelia that senior management h…

DGX agent

One underdiscussed part of the Google-DOD deal is how it blindsided the company’s own employees. One told @cogcelia that senior management had repeatedly insisted Google wouldn’t cave to the Pentagon’

safetygary-marcus--x
30 Apr 2026
Safety

One Word at a Time: Incremental Completion Decomposition Breaks LLM Safety

DGX agent

arXiv:2604.25921v1 Announce Type: new Abstract: Large Language Models (LLMs) are trained to refuse harmful requests, yet they remain vulnerable to jailbreak attacks that exploit weaknesses in conversa

safetyarxiv-cs-cl
30 Apr 2026
Safety

Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents

DGX agent

arXiv:2505.02077v2 Announce Type: replace-cross Abstract: AI agents are beginning to interact with each other directly and across internet platforms and physical environments, creating security challe

safetyarxiv-cs-ai
30 Apr 2026
Safety

Open Problems in Frontier AI Risk Management

DGX agent

arXiv:2604.25982v1 Announce Type: cross Abstract: Frontier AI both amplifies existing risks and introduces qualitatively novel challenges. Not only is there a notable lack of stable scientific consens

safetyarxiv-cs-ai
30 Apr 2026
Safety

OpenAI’s lawyer plays dirty. Why I am not surprised?

DGX agent

OpenAI’s lawyer plays dirty. Why I am not surprised? 🚨 OpenAI's lawyer read a 2016 email from Sutskever to Musk to argue Musk knew OpenAI would go closed-source Sutskever: 'As we get closer to buildin

safetygary-marcus--x
30 Apr 2026
Safety

Operating-Layer Controls for Onchain Language-Model Agents Under Real Capital

DGX agent

arXiv:2604.26091v1 Announce Type: new Abstract: We study reliability in autonomous language-model agents that translate user mandates into validated tool actions under real capital. The setting is DX

safetyarxiv-cs-ai
30 Apr 2026
Safety

Oracle has dropped 50% since I declared that the OpenAI deal they made was “peak bubble”. And, as @edzitron lays out below, they may fall a …

DGX agent

Oracle has dropped 50% since I declared that the OpenAI deal they made was “peak bubble”. And, as @edzitron lays out below, they may fall a lot further yet. At which point the entire house of cards co

safetygary-marcus--x
30 Apr 2026
Safety

OT Score: An OT based Confidence Score for Prototype-Assisted Source Free Unsupervised Domain Adaptation

DGX agent

arXiv:2505.11669v3 Announce Type: replace-cross Abstract: We address the computational and theoretical limitations of current distributional alignment methods for source-free unsupervised domain adapt

safetyarxiv-cs-ai
30 Apr 2026
Safety

Ouch! Current AI assistants often corrupt documents. Sounds like an intern you can’t trust — once again. A trillion dollar investment in sca…

DGX agent

Ouch! Current AI assistants often corrupt documents. Sounds like an intern you can’t trust — once again. A trillion dollar investment in scaling hasn’t solved this. New Microsoft paper shows that curr

safetygary-marcus--x
30 Apr 2026
Safety

PAINT: Partial-Solution Adaptive Interpolated Training for Self-Distilled Reasoners

DGX agent

arXiv:2604.26573v1 Announce Type: new Abstract: Improving large language model (LLM) reasoning requires supervision that is both aligned with the model's own test-time states and informative at the to

safetyarxiv-cs-lg
30 Apr 2026
Safety

PBiLoss: Popularity-Aware Regularization to Improve Fairness in Graph-Based Recommender Systems

DGX agent

arXiv:2507.19067v2 Announce Type: replace-cross Abstract: Recommender systems based on graph neural networks (GNNs) have been proved to perform well on user-item interactions. However, they commonly s

safetyarxiv-cs-ai
30 Apr 2026
Safety

Probe-then-Plan: Environment-Aware Planning for Industrial E-commerce Search

DGX agent

arXiv:2603.15262v2 Announce Type: replace Abstract: Modern e-commerce search is evolving to resolve complex user intents. While Large Language Models (LLMs) offer strong reasoning, existing LLM-based

safetyarxiv-cs-ai
30 Apr 2026
Safety

R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation

DGX agent

arXiv:2510.08547v2 Announce Type: replace-cross Abstract: Towards the aim of generalized robotic manipulation, spatial generalization is the most fundamental capability that requires the policy to wor

safetyarxiv-cs-cv
30 Apr 2026
Safety

Recipes for Calibration Checks in Safety-Critical Applications

DGX agent

arXiv:2604.26479v1 Announce Type: cross Abstract: Safety-critical prediction systems, such as autonomous vehicles, weather forecasters, and medical monitors, commonly rely on probabilistic forecasters

safetyarxiv-cs-lg
30 Apr 2026
Safety

Risk Reporting for Developers' Internal AI Model Use

DGX agent

arXiv:2604.24966v1 Announce Type: cross Abstract: Frontier AI companies first deploy their most advanced models internally, for weeks or months of safety testing, evaluation, and iteration, before a p

safetyarxiv-cs-ai
30 Apr 2026
Safety

Roblox reports Q1 bookings up 43% YoY to 1.7B, vs. 1.73B est., and DAUs up 35% to 132M, below analysts' estimates of 143.8M; RBLX drops 16%+ after hours (Cecilia D'Anastasio/Bloomberg)

DGX agent

Cecilia D'Anastasio / Bloomberg: Roblox reports Q1 bookings up 43% YoY to 1.7B, vs. 1.73B est., and DAUs up 35% to 132M, below analysts' estimates of 143.8M; RBLX drops 16%+ after hours — Roblox Corp.

safetytechmeme
30 Apr 2026
Safety

Robust Alignment: Harmonizing Clean Accuracy and Adversarial Robustness in Adversarial Training

DGX agent

arXiv:2604.26496v1 Announce Type: new Abstract: Adversarial Training (AT) is one of the most effective methods for developing robust deep neural networks (DNNs). However, AT faces a trade-off problem

safetyarxiv-cs-cv
30 Apr 2026
Safety

Rule-based High-Level Coaching for Goal-Conditioned Reinforcement Learning in Search-and-Rescue UAV Missions Under Limited-Simulation Training

DGX agent

arXiv:2604.26833v1 Announce Type: cross Abstract: This paper presents a hierarchical decision-making framework for unmanned aerial vehicle (UAV) missions motivated by search-and-rescue (SAR) scenarios

safetyarxiv-cs-ai
30 Apr 2026
Safety

SAGE: A Strategy-Aware Graph-Enhanced Generation Framework For Online Counseling

DGX agent

arXiv:2604.26630v1 Announce Type: new Abstract: Effective mental health counseling is a complex, theory-driven process requiring the simultaneous integration of psychological frameworks, real-time dis

safetyarxiv-cs-cl
30 Apr 2026
Safety

SD2AIL: Adversarial Imitation Learning from Synthetic Demonstrations via Diffusion Models

DGX agent

arXiv:2512.18583v2 Announce Type: replace Abstract: Adversarial Imitation Learning (AIL) is a dominant framework in imitation learning that infers rewards from expert demonstrations to guide policy op

safetyarxiv-cs-lg
30 Apr 2026
Safety

Seeking Consensus: Geometric-Semantic On-the-Fly Recalibration for Open-Vocabulary Remote Sensing Semantic Segmentation

DGX agent

arXiv:2604.26221v1 Announce Type: cross Abstract: Open-vocabulary semantic segmentation (OVSS) in remote sensing images is a promising task that employs textual descriptions for identifying undefined

safetyarxiv-cs-ai
30 Apr 2026
Safety

Sergey Brin-backed Building a Better California says two ballot countermeasures to the proposed wealth tax are on track to qualify for the November ballot (Laura J. Nelson/Wall Street Journal)

DGX agent

Laura J. Nelson / Wall Street Journal: Sergey Brin-backed Building a Better California says two ballot countermeasures to the proposed wealth tax are on track to qualify for the November ballot — Two

safetytechmeme
30 Apr 2026
Safety

“shitting away money at scale”. @mcuban nails what I have been trying to say:

DGX agent

“shitting away money at scale”. @mcuban nails what I have been trying to say: .@mcuban is not impressed with OpenAI's business: 'They're shitting away money at scale,' he told me. Full discussion live

safetygary-marcus--x
30 Apr 2026
Safety

Sociodemographic Biases in Educational Counselling by Large Language Models

DGX agent

arXiv:2604.25932v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly integrated into educational settings, understanding their potential biases is critical. This study ex

safetyarxiv-cs-ai
30 Apr 2026
Safety

Sparsity as a Key: Unlocking New Insights from Latent Structures for Out-of-Distribution Detection

DGX agent

arXiv:2604.26409v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) have demonstrated significant success in interpreting Large Language Models (LLMs) by decomposing dense representations into

safetyarxiv-cs-cv
30 Apr 2026
Safety

STARRY: Spatial-Temporal Action-Centric World Modeling for Robotic Manipulation

DGX agent

arXiv:2604.26848v1 Announce Type: new Abstract: Robotic manipulation critically requires reasoning about future spatial-temporal interactions, yet existing VLA policies and world-model-enhanced polici

safetyarxiv-cs-ro
30 Apr 2026
Safety

Student Guides Teacher: Weak-to-Strong Inference via Spectral Orthogonal Exploration

DGX agent

arXiv:2601.06160v2 Announce Type: replace Abstract: Large Language Models (LLMs) often suffer from ''Reasoning Collapse'' on challenging mathematical reasoning tasks, where stochastic sampling produce

safetyarxiv-cs-ai
30 Apr 2026
Safety

Talent or Luck? Evaluating Attribution Bias in Large Language Models

DGX agent

arXiv:2505.22910v2 Announce Type: replace Abstract: When a student fails an exam, do we tend to blame their effort or the test's difficulty? Attribution, defined as how reasons are assigned to event o

safetyarxiv-cs-cl
30 Apr 2026
Safety

Tatemae: Detecting Alignment Faking via Tool Selection in LLMs

DGX agent

arXiv:2604.26511v1 Announce Type: cross Abstract: Alignment faking (AF) occurs when an LLM strategically complies with training objectives to avoid value modification, reverting to prior preferences o

safetyarxiv-cs-ai
30 Apr 2026
Safety

Teaching LLM to be Persuasive: Reward-Enhanced Policy Optimization for Alignment from Heterogeneous Rewards

DGX agent

arXiv:2510.04214v3 Announce Type: replace Abstract: We deploy large language models (LLMs) as business development (BD) agents for persuasive price negotiation in online travel agencies (OTAs). The ag

safetyarxiv-cs-cl
30 Apr 2026
Safety

Test-Time Safety Alignment

DGX agent

arXiv:2604.26167v1 Announce Type: cross Abstract: Recent work has shown that a model's input word embeddings can serve as effective control variables for steering its behavior toward outputs that sati

safetyarxiv-cs-ai
30 Apr 2026
Safety

Text Style Transfer with Machine Translation for Graphic Designs

DGX agent

arXiv:2604.26361v1 Announce Type: cross Abstract: Globalization of graphic designs such as those used in marketing materials and magazines is increasingly important for communication to broad audience

safetyarxiv-cs-ai
30 Apr 2026
Safety

The Alignment Flywheel: A Governance-Centric Hybrid MAS for Architecture-Agnostic Safety

DGX agent

arXiv:2603.02259v2 Announce Type: replace-cross Abstract: Multi-agent systems provide mature methodologies for role decomposition, coordination, and normative governance, capabilities that remain esse

safetyarxiv-cs-lg
30 Apr 2026
Safety

The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation

DGX agent

arXiv:2604.26347v1 Announce Type: cross Abstract: Objective metrics for emotional expressiveness are vital for speech generation, particularly in expressive synthesis and voice conversion requiring em

safetyarxiv-cs-cl
30 Apr 2026
Safety

The Zig project's rationale for their firm anti-AI contribution policy

DGX agent

Zig has one of the most stringent anti-LLM policies of any major open source project: No LLMs for issues. No LLMs for pull requests. No LLMs for comments on the bug tracker, including translation. Eng

safetysimon-willison
30 Apr 2026
← Previous
1…215216217218219…265
Next →