AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
Safety

Can Semantic Methods Enhance Team Sports Tactics? A Methodology for Football with Broader Applications

DGX agent

arXiv:2601.00421v2 Announce Type: replace Abstract: This paper explores how semantic-space reasoning, traditionally used in computational linguistics, can be extended to tactical decision-making in te

safetyarxiv-cs-ai
6 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Catching the Infection Before It Spreads: Foresight-Guided Defense in Multi-Agent Systems

DGX agent

arXiv:2605.01758v1 Announce Type: new Abstract: Large multimodal model-based Multi-Agent Systems (MASs) enable collaborative complex problem solving through specialized agents. However, MASs are vulne

safetyarxiv-cs-ai
6 May 2026
Safety

Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents

DGX agent

arXiv:2604.06132v2 Announce Type: replace Abstract: Large language models are increasingly deployed as autonomous agents for multi-step workflows in real-world software environments. However, existing

safetyarxiv-cs-ai
6 May 2026
Safety

C’mon BBC. Zilis was sharp as a tack on the stand, on her role in the OpenAI *nonprofit* board, and how she managed conflicts as they began …

DGX agent

C’mon BBC. Zilis was sharp as a tack on the stand, on her role in the OpenAI *nonprofit* board, and how she managed conflicts as they began to develop, and this (which is not really even news since it

safetygary-marcus--x
6 May 2026
Safety

Coherent Hierarchical Multi-Label Learning to Defer for Medical Imaging

DGX agent

arXiv:2605.02734v1 Announce Type: new Abstract: Learning to Defer (L2D) enables a model to predict autonomously or defer to an expert, but prior work largely assumes flat label spaces. We study the fi

safetyarxiv-cs-ai
6 May 2026
Safety

Correction: the jury is advisory only. It’s the judge who decides; if she sees it as I do, OpenAI loses.

DGX agent

Gary Marcus clarifies that in the legal proceeding he's discussing, the jury serves an advisory role while the judge retains decision-making authority on the case outcome. Marcus expresses confidence

safetygary-marcus--x
6 May 2026
Safety

Deciphering Shortcut Learning from an Evolutionary Game Theory Perspective

DGX agent

arXiv:2605.02658v2 Announce Type: new Abstract: Shortcut learning causes deep learning models to rely on non-essential features within the data. However, its formation in deep neural network training

safetyarxiv-cs-ai
6 May 2026
Safety

Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning

DGX agent

arXiv:2602.20078v3 Announce Type: replace-cross Abstract: Scaling cooperative multi-agent reinforcement learning (MARL) is fundamentally limited by cross-agent noise. When agents share a common reward

safetyarxiv-cs-lg
6 May 2026
Safety

DGPO: Distribution Guided Policy Optimization for Fine Grained Credit Assignment

DGX agent

arXiv:2605.03327v1 Announce Type: new Abstract: Reinforcement learning is crucial for aligning large language models to perform complex reasoning tasks. However, current algorithms such as Group Relat

safetyarxiv-cs-lg
6 May 2026
Safety

Discovering Reinforcement Learning Interfaces with Large Language Models

DGX agent

arXiv:2605.03408v1 Announce Type: new Abstract: Reinforcement learning systems rely on environment interfaces that specify observations and reward functions, yet constructing these interfaces for new

safetyarxiv-cs-lg
6 May 2026
Safety

Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment

DGX agent

arXiv:2605.01899v1 Announce Type: new Abstract: The growing capabilities of large language models (LLMs) have driven their widespread deployment across diverse domains, even in potentially high-risk s

safetyarxiv-cs-ai
6 May 2026
Safety

DMGD: Train-Free Dataset Distillation with Semantic-Distribution Matching in Diffusion Models

DGX agent

arXiv:2605.03877v1 Announce Type: new Abstract: Dataset distillation enables efficient training by distilling the information of large-scale datasets into significantly smaller synthetic datasets. Dif

safetyarxiv-cs-cv
6 May 2026
Safety

@Dr_Gingerballs At least actual ponzi schemes don't light their cash on fire... they just cant meet redemptions at the level of their inflat…

DGX agent

@Dr_Gingerballs At least actual ponzi schemes don't light their cash on fire... they just cant meet redemptions at the level of their inflated fake earnings. After all of the hyperscalers burn every l

safetygary-marcus--x
6 May 2026
Safety

Efficient Temporal Datalog Materialisation for Composite Event Recognition

DGX agent

arXiv:2605.02488v1 Announce Type: new Abstract: Several applications demand the timely detection of critical situations, such as threats to safety and transparency, over high-velocity streams of symbo

safetyarxiv-cs-ai
6 May 2026
Safety

EvoJail: Evolutionary Diverse Jailbreak Prompt Generation for Large Language Models

DGX agent

arXiv:2605.02921v1 Announce Type: cross Abstract: As LLMs continue to shape real-world applications, automated jailbreak generation becomes essential to reveal safety weaknesses and guide model improv

safetyarxiv-cs-lg
6 May 2026
Safety

False Friends in the Shell: Unveiling the Emoticon Semantic Confusion in Large Language Models

DGX agent

arXiv:2601.07885v2 Announce Type: replace-cross Abstract: Emoticons are widely used in digital communication to convey affective intent, yet their safety implications for Large Language Models (LLMs)

safetyarxiv-cs-ai
6 May 2026
Safety

FIBER: A Differentially Private Optimizer with Filter-Aware Innovation Bias Correction

DGX agent

arXiv:2605.03425v1 Announce Type: new Abstract: Differentially private (DP) training protects individual examples by adding noise to gradients, but the injected noise interacts nontrivially with adapt

safetyarxiv-cs-lg
6 May 2026
Safety

FINER-SQL: Boosting Small Language Models for Text-to-SQL

DGX agent

arXiv:2605.03465v1 Announce Type: cross Abstract: Large language models have driven major advances in Text-to-SQL generation. However, they suffer from high computational cost, long latency, and data

safetyarxiv-cs-cl
6 May 2026
Safety

FORMULA: FORmation MPC with neUral barrier Learning for safety Assurance

DGX agent

arXiv:2604.04409v2 Announce Type: replace Abstract: Multi-robot systems (MRS) are essential for large-scale applications such as disaster response, material transport, and warehouse logistics, yet ens

safetyarxiv-cs-ro
6 May 2026
Safety

From SFT to RL: Demystifying the Post-Training Pipeline for LLM-based Vulnerability Detection

DGX agent

arXiv:2602.14012v2 Announce Type: replace-cross Abstract: The integration of LLMs into vulnerability detection (VD) has shifted the field toward more interpretable and context-aware analysis. While po

safetyarxiv-cs-ai
6 May 2026
Safety

@GaryMarcus I can't believe this is still a thing that 'experts' haven't caught up with. @GaryMarcus has been saying this forever, and those…

DGX agent

@GaryMarcus I can't believe this is still a thing that 'experts' haven't caught up with. @GaryMarcus has been saying this forever, and those of us who have actually dug into the tech, analyzed it, use

safetygary-marcus--x
6 May 2026
Safety

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling

DGX agent

arXiv:2507.07982v2 Announce Type: replace Abstract: Videos inherently represent 2D projections of a dynamic 3D world. However, our analysis suggests that video diffusion models trained solely on raw v

safetyarxiv-cs-cv
6 May 2026
Safety

Global and Local Topology-Aware Attention with Persistent Homology and Euler Biases for Time-Series Forecasting

DGX agent

arXiv:2605.03163v1 Announce Type: new Abstract: Scientific time series often encode predictive geometric structure, including connectivity, cycles, shell-like geometry, directional changes, and nonlin

safetyarxiv-cs-lg
6 May 2026
Safety

Google, Microsoft and xAI agree to allow government safety checks of their AI models prior to release

DGX agent

Google LLC, Microsoft Corp. and xAI have agreed to share unreleased versions of their artificial intelligence models with the U.S. Department of Commerce to ensure the technologies do not pose a threa

safetysiliconangle
6 May 2026
Safety

Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure

DGX agent

arXiv:2605.01091v1 Announce Type: cross Abstract: When a traffic signal controller adjusts green phases and a grid manager curtails power on the same corridor, each system may comply with its own obli

safetyarxiv-cs-ai
6 May 2026
Safety

GRAFT: Auditing Graph Neural Networks via Global Feature Attribution

DGX agent

arXiv:2605.03377v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve strong performance on node classification tasks but remain difficult to interpret, particularly with respect to whi

safetyarxiv-cs-lg
6 May 2026
Safety

Grounding Multi-Hop Reasoning in Structural Causal Models via Group Relative Policy Optimization

DGX agent

arXiv:2605.01482v1 Announce Type: new Abstract: Multi-Hop Fact Verification (MHFV) necessitates complex reasoning across disparate evidence, posing significant challenges for Large Language Models (LL

safetyarxiv-cs-ai
6 May 2026
Safety

GRPO-TTA: Test-Time Visual Tuning for Vision-Language Models via GRPO-Driven Reinforcement Learning

DGX agent

arXiv:2605.03403v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has recently shown strong performance in post-training large language models and vision-language models. It ra

safetyarxiv-cs-cv
6 May 2026
Safety

Healthcare AI GYM for Medical Agents

DGX agent

arXiv:2605.02943v1 Announce Type: new Abstract: Clinical reasoning demands multi-step interactions -- gathering patient history, ordering tests, interpreting results, and making safe treatment decisio

safetyarxiv-cs-lg
6 May 2026
Safety

Height Control and Optimal Torque Planning for Jumping With Wheeled-Bipedal Robots

DGX agent

arXiv:2605.03302v1 Announce Type: new Abstract: This paper mainly studies the accurate height jumping control of wheeled-bipedal robots based on torque planning and energy consumption optimization. Du

safetyarxiv-cs-ro
6 May 2026
Safety

Heterogeneous Graph Importance Scoring and Clustering with Automated LLM-based Interpretation

DGX agent

arXiv:2605.02919v1 Announce Type: new Abstract: Urban bridge networks are critical infrastructure whose disruption can cascade into severe impacts on transportation, emergency services, and economic a

safetyarxiv-cs-lg
6 May 2026
Safety

HiMAC: Hierarchical Macro-Micro Learning for Long-Horizon LLM Agents

DGX agent

arXiv:2603.00977v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents have recently demonstrated strong capabilities in interactive decision-making, yet they remain fundamentally

safetyarxiv-cs-lg
6 May 2026
Safety

How Sam ('You parachute him onto a cannibal island, and he comes back five years later as king”) Altman operates:

DGX agent

How Sam ('You parachute him onto a cannibal island, and he comes back five years later as king”) Altman operates: counsel: 'by fall of 2023 did you percieve altman was not candid with you? truthful? h

safetygary-marcus--x
6 May 2026
Safety

Human-in-the-Loop Uncertainty Analysis in Self-Adaptive Robots Using LLMs

DGX agent

arXiv:2605.02983v1 Announce Type: new Abstract: Self-adaptive robots operate in dynamic, unpredictable environments where unaddressed uncertainties can lead to safety violations and operational failur

safetyarxiv-cs-ro
6 May 2026
Safety

I repeat, the bubble is in the 'e' not the 'p' in today's PE ratios. The hucksters and talking heads will, as always, fail to realize until …

DGX agent

I repeat, the bubble is in the 'e' not the 'p' in today's PE ratios. The hucksters and talking heads will, as always, fail to realize until it's too late. But it's a very simple set up. Hyperscalers g

safetygary-marcus--x
6 May 2026
Safety

If Forbes had only waited to hear the testimony at this week’s trial Or read @_KarenHao’s book Or @RonanFarrow’s @newyorker investigation Or…

DGX agent

If Forbes had only waited to hear the testimony at this week’s trial Or read @_KarenHao’s book Or @RonanFarrow’s @newyorker investigation Or my own writings since fall 2023 They would have realized ho

safetygary-marcus--x
6 May 2026
Safety

Important nuance: the jury at this Musk-OpenAI trial is an *advisory* jury, hence not binding on the judge, and is only looking at liability…

DGX agent

Important nuance: the jury at this Musk-OpenAI trial is an *advisory* jury, hence not binding on the judge, and is only looking at liability (not damages, if any). Thanks to @bahhradx for correcting a

safetygary-marcus--x
6 May 2026
Safety

Instance-Level Costs for Nuanced Classifier Evaluation

DGX agent

arXiv:2605.03135v1 Announce Type: new Abstract: Standard classification treats all errors equally, but in content moderation, medical screening, and safety-critical applications, mistakes on clear-cut

safetyarxiv-cs-lg
6 May 2026
Safety

Intervention Complexity as a Canonical Reward and a Measure of Intelligence

DGX agent

arXiv:2605.02175v1 Announce Type: new Abstract: The Legg--Hutter universal intelligence measure provides a rigorous scalar assessment of general intelligence as expected reward across all computable e

safetyarxiv-cs-ai
6 May 2026
Safety

just want to go back to how much intense pushback i got on this story at all levels of the company at the time, and how people speak very di…

DGX agent

just want to go back to how much intense pushback i got on this story at all levels of the company at the time, and how people speak very differently when under the threat of perjury lot of names etch

safetygary-marcus--x
6 May 2026
Safety

Khala: Scaling Acoustic Token Language Models Toward High-Fidelity Music Generation

DGX agent

arXiv:2605.01790v1 Announce Type: cross Abstract: A common design pattern in high-quality music generation is to handle structure and fidelity in different representation spaces: a generator first mod

safetyarxiv-cs-ai
6 May 2026
Safety

Large Language Models are Universal Reasoners for Visual Generation

DGX agent

arXiv:2605.04040v1 Announce Type: new Abstract: Text-to-image generation has advanced rapidly with diffusion models, progressing from CLIP and T5 conditioning to unified systems where a single LLM bac

safetyarxiv-cs-cv
6 May 2026
Safety

Learning Reactive Dexterous Grasping via Hierarchical Task-Space RL Planning and Joint-Space QP Control

DGX agent

arXiv:2605.03363v1 Announce Type: new Abstract: In this work, we propose a hybrid hierarchical control framework for reactive dexterous grasping that explicitly decouples high-level spatial intent fro

safetyarxiv-cs-ro
6 May 2026
Safety

Like tricksters, LLMs have perfected the art of plausibility, says Tim Harford: https://ft.trib.al/5Foo2YD

DGX agent

Tim Harford compares large language models to tricksters, arguing that LLMs excel at generating plausible-sounding text without necessarily ensuring accuracy or truthfulness. The article likely explor

safetygary-marcus--x
6 May 2026
Safety

LLM-XTM: Enhancing Cross-Lingual Topic Models with Large Language Models

DGX agent

arXiv:2605.03299v1 Announce Type: new Abstract: Cross-lingual topic modeling aims to discover shared semantic structures across languages, yet existing models depend on sparse bilingual resources and

safetyarxiv-cs-cl
6 May 2026
Safety

Logic-Constrained Shortest Paths for Flight Planning

DGX agent

arXiv:2412.13235v4 Announce Type: replace Abstract: The logic-constrained shortest path problem (LCSPP) combines a one-to-one shortest path problem with satisfiability constraints imposed on the routi

safetyarxiv-cs-ai
6 May 2026
Safety

MAGE: Safeguarding LLM Agents against Long-Horizon Threats via Shadow Memory

DGX agent

arXiv:2605.03228v1 Announce Type: cross Abstract: As large language model (LLM)-powered agents are increasingly deployed to perform complex, real-world tasks, they face a growing class of attacks that

safetyarxiv-cs-cl
6 May 2026
Safety

Many trials feel like Rashomon with different witnesses. The amazing thing about Musk-OpenAI is how much agreement there has been (at least …

DGX agent

Many trials feel like Rashomon with different witnesses. The amazing thing about Musk-OpenAI is how much agreement there has been (at least so far) on the facts. The question is really whether what Op

safetygary-marcus--x
6 May 2026
← Previous
1…202203204205206…265
Next →