AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,814 results
Safety

Distribution-Free Risk-Aware Planning and Control Under Uncertainty Using Conformal Spectral Risk Control

DGX agent

arXiv:2606.04185v1 Announce Type: new Abstract: Safe navigation in dynamic and uncertain environments often relies on accurate estimation of, or assumptions about, the true underlying uncertainty. How

safetyarxiv-cs-ro
4 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

DPM++: Dynamic Masked Metric Learning for Occluded Person Re-identification

DGX agent

arXiv:2605.06637v2 Announce Type: replace Abstract: Although person re-identification has made impressive progress, occlusion caused by obstacles remains an unsettled issue in real applications. The d

safetyarxiv-cs-cv
4 Jun 2026
Safety

DuDi: Dual-Signal Distillation with Cross-Lingual Verbalizer

DGX agent

arXiv:2606.04694v1 Announce Type: new Abstract: Small language models (SLMs) are efficient and scalable, but their multilingual capabilities degrade severely at sub-billion scales, especially for Sout

safetyarxiv-cs-cl
4 Jun 2026
Safety

DVGT: Driving Visual Geometry Transformer

DGX agent

arXiv:2512.16919v2 Announce Type: replace-cross Abstract: Perceiving and reconstructing 3D scene geometry from visual inputs is crucial for autonomous driving. However, there still lacks a driving-tar

safetyarxiv-cs-ai
4 Jun 2026
Safety

Dynamic Multi-Pair Trading Strategy in Cryptocurrency Markets with Deep Reinforcement Learning

DGX agent

arXiv:2606.04574v1 Announce Type: new Abstract: This study aims to determine whether the application of Deep Reinforcement Learning (DRL) as a specialized execution overlay can enhance pair trading in

safetyarxiv-cs-lg
4 Jun 2026
Safety

Dynamic Policy Learning for Legged Robot with Simplified Model Pretraining and Model-Homotopy-Inspired Transfer

DGX agent

arXiv:2512.24698v2 Announce Type: replace Abstract: Generating dynamic motions for legged robots remains a challenging problem. While reinforcement learning has achieved notable success in various leg

safetyarxiv-cs-ro
4 Jun 2026
Safety

Edge of Stability Selectively Shapes Learning Across the Data Distribution

DGX agent

arXiv:2606.04212v1 Announce Type: new Abstract: Existing analyses of the edge of stability (EoS) treat it as a global property of optimization. We show that it is also selective: the stability constra

safetyarxiv-cs-lg
4 Jun 2026
Safety

Efficient Adversarial Attacks on High-dimensional Offline Bandits

DGX agent

arXiv:2602.01658v2 Announce Type: replace-cross Abstract: Bandit algorithms have recently emerged as a powerful tool for evaluating machine learning models, including generative image models and large

safetyarxiv-cs-ai
4 Jun 2026
Safety

Elon Musk petitioned the FTC in May to end its 2022 order restricting Twitter's data use, claiming Twitter no longer exists as X merged with xAI and then SpaceX (Ashley Belanger/Ars Technica)

DGX agent

Ashley Belanger / Ars Technica: Elon Musk petitioned the FTC in May to end its 2022 order restricting Twitter's data use, claiming Twitter no longer exists as X merged with xAI and then SpaceX — Criti

safetytechmeme
4 Jun 2026
Safety

Enhancing the MADDPG Algorithm for Multi-Agent Learning via Action Inference and Importance Sampling

DGX agent

arXiv:2606.05021v1 Announce Type: new Abstract: We investigate multi-agent deep reinforcement learning and propose two enhancements to the Multi-Agent Deep Deterministic Policy Gradient (MADDPG) algor

safetyarxiv-cs-lg
4 Jun 2026
Safety

Expert-Aware Refusal Steering

DGX agent

arXiv:2606.04160v1 Announce Type: new Abstract: Safety alignment in instruction-tuned large language models (LLMs) depends on a model's ability to reliably refuse to respond to harmful or disallowed r

safetyarxiv-cs-cl
4 Jun 2026
Safety

Explainably Safe Reinforcement Learning

DGX agent

arXiv:2606.04634v1 Announce Type: new Abstract: Trust in a decision-making system requires both safety guarantees and the ability to interpret and understand its behavior. This is particularly importa

safetyarxiv-cs-lg
4 Jun 2026
Safety

Extending Fair Null-Space Projections for Continuous Attributes to Kernel Methods

DGX agent

arXiv:2511.03304v2 Announce Type: replace-cross Abstract: With the on-going integration of machine learning systems into the everyday social life of millions the notion of fairness becomes an ever inc

safetyarxiv-cs-ai
4 Jun 2026
Safety

Feels like a good time to resurface this one Mine and @jaswu_'s basic point: cheaper AI complicates the narrative for OpenAI and Anthropic w…

DGX agent

Feels like a good time to resurface this one Mine and @jaswu_'s basic point: cheaper AI complicates the narrative for OpenAI and Anthropic when they eventually try to go public. Could also ripple acro

safetygary-marcus--x
4 Jun 2026
Safety

Few Tokens, Big Leverage: Preserving Safety Alignment by Constraining Safety Tokens during Fine-tuning

DGX agent

arXiv:2603.07445v2 Announce Type: replace Abstract: Large language models (LLMs) often require fine-tuning (FT) to perform well on downstream tasks, but FT can induce safety-alignment drift even when

safetyarxiv-cs-cl
4 Jun 2026
Safety

FLAGG: Flexible Autoregressive Graph Generation

DGX agent

arXiv:2606.05067v1 Announce Type: new Abstract: The Deep Graph Generation's panorama spans two extremes: one-shot and sequential models. The former generates nodes and edges jointly, while the latter

safetyarxiv-cs-lg
4 Jun 2026
Safety

Fog of Love: Engineering Virtuous Agent Behavior with Affinity-based Reinforcement Learning in a Game Environment

DGX agent

arXiv:2606.04750v1 Announce Type: new Abstract: Instilling virtuous behavior in artificial intelligence has seen increasing interest. One of the techniques proposed is known as affinity-based reinforc

safetyarxiv-cs-ai
4 Jun 2026
Safety

Formal Semantics for Agentic Tool Protocols: A Process Calculus Approach

DGX agent

arXiv:2603.24747v2 Announce Type: replace Abstract: The emergence of large language model agents capable of invoking external tools has created urgent need for formal verification of agent protocols.

safetyarxiv-cs-ai
4 Jun 2026
Safety

From Agent Traces to Trust: Evidence Tracing and Execution Provenance in LLM Agents

DGX agent

arXiv:2606.04990v1 Announce Type: cross Abstract: Large language model (LLM)-based agents increasingly solve complex tasks by interacting with external tools, retrieval systems, memory modules, enviro

safetyarxiv-cs-ai
4 Jun 2026
Safety

GARL: Game-Theoretic Reinforcement Learning for Multi-Agent Strategic Prioritisation

DGX agent

arXiv:2606.05002v1 Announce Type: new Abstract: LLM-based multi-agent systems are increasingly used for strategic decision-making tasks. In such settings, performance depends not only on individual mo

safetyarxiv-cs-cl
4 Jun 2026
Safety

Generalizable Multi-Task Learning for Wireless Networks Using Prompt Decision Transformers

DGX agent

arXiv:2606.04328v1 Announce Type: cross Abstract: Future wireless networks demand rapid adaptation to highly heterogeneous environments and dynamic task configurations, necessitating a shift from conv

safetyarxiv-cs-ai
4 Jun 2026
Safety

Generalization of World Models under Environmental Variability for Vision-based Quadrotor Navigation

DGX agent

arXiv:2606.05015v1 Announce Type: new Abstract: World models, learned generative models that predict how an environment evolves, have become a promising tool for sample-efficient robot learning. Yet h

safetyarxiv-cs-ro
4 Jun 2026
Safety

Geometry-Aware Distillation for Prompt Tuning Biomedical Vision-Language Models

DGX agent

arXiv:2606.04922v1 Announce Type: cross Abstract: Current prompt-based and adapter-based tuning of vision-language models (VLMs) is attractive for medical imaging, where clinical data sensitivity favo

safetyarxiv-cs-ai
4 Jun 2026
Safety

Geospatial Foundation Models to Enable Progress on Sustainable Development Goals

DGX agent

arXiv:2505.24528v3 Announce Type: replace Abstract: Foundation Models (FMs) are large-scale, pre-trained artificial intelligence (AI) systems that have revolutionized natural language processing and c

safetyarxiv-cs-cv
4 Jun 2026
Safety

Global Sketch-Based Watermarking for Diffusion Language Models

DGX agent

arXiv:2606.04486v1 Announce Type: cross Abstract: Watermarking methods for language models have been studied extensively in the autoregressive setting, where tokens are generated sequentially. These w

safetyarxiv-cs-cl
4 Jun 2026
Safety

Good Reasoning Makes Good Demonstrations: Implicit Reasoning Quality Supervision via In-Context Reinforcement Learning

DGX agent

arXiv:2603.09803v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves reasoning in large language models but treats all correct solutions equally, potentia

safetyarxiv-cs-lg
4 Jun 2026
Safety

GRAIL: Gradient-Reweighted Advantages for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2606.04889v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (e.g. GRPO) is now a common way to improve mathematical reasoning in Large Language Models (LLMs). Howeve

safetyarxiv-cs-cl
4 Jun 2026
Safety

HapTile: A Haptic-Informed Vision-Tactile-Language-Action Dataset for Contact-Rich Imitation Learning

DGX agent

arXiv:2606.04825v1 Announce Type: new Abstract: Despite the importance of tactile sensing for reliable manipulation, most existing Vision-Language-Action (VLA) datasets remain vision-only, and those t

safetyarxiv-cs-ro
4 Jun 2026
Safety

How does Elon get off on just lying straight up about the mainstream media? Can someone like BBC sue him for defamation? He claimed “Legacy …

DGX agent

How does Elon get off on just lying straight up about the mainstream media? Can someone like BBC sue him for defamation? He claimed “Legacy mainstream media, same ones who wrote about George Floyd mil

safetygary-marcus--x
4 Jun 2026
Safety

Hybrid Adversarial Defence for Natural Language Understanding Tasks

DGX agent

arXiv:2606.04612v1 Announce Type: new Abstract: Large Language Models (LLMs) are vulnerable both to hallucination and adversarial manipulation. Although these problems are closely related, existing de

safetyarxiv-cs-cl
4 Jun 2026
Safety

I think @Levie is overstating the positive case for employment in the (near term) AI era but that most people have overstated the negative c…

DGX agent

I think @Levie is overstating the positive case for employment in the (near term) AI era but that most people have overstated the negative case, and that the truth is somewhere in between. Which is to

safetygary-marcus--x
4 Jun 2026
Safety

If we can’t trust Goldman on which IPO to buy… who can we trust? WeWork. Goldman Sachs marketed the real estate firm at an inflated 96 billi…

DGX agent

If we can’t trust Goldman on which IPO to buy… who can we trust? WeWork. Goldman Sachs marketed the real estate firm at an inflated 96 billion dollar tech valuation. The IPO was pulled after the prosp

safetygary-marcus--x
4 Jun 2026
Safety

If you could buy into exactly one of the three mega IPOs, which would it be?

DGX agent

Gary Marcus poses a hypothetical investment question asking followers to choose among three major IPOs, likely seeking comparative analysis of high-profile public offerings. The post invites discussio

safetygary-marcus--x
4 Jun 2026
Safety

if you had 100k to invest in OpenAI and/or Anthropic IPOs which would you go for?

DGX agent

Gary Marcus discusses investment strategy between potential OpenAI and Anthropic IPOs, likely weighing factors such as the companies' technological capabilities, market positioning, business models, a

safetygary-marcus--x
4 Jun 2026
Safety

If you – or your retirement funds - get taken for a ride on SpaceX blame hype guys like this, who don’t even mention that Goldman is the lea…

DGX agent

If you – or your retirement funds - get taken for a ride on SpaceX blame hype guys like this, who don’t even mention that Goldman is the lead left on the deal. GOLDMAN SEES SPACEX AI REVENUE EXPLODING

safetygary-marcus--x
4 Jun 2026
Safety

Imbuing Large Language Models with Bidirectional Logic for Robust Chain Repair

DGX agent

arXiv:2606.05030v1 Announce Type: new Abstract: Autoregressive chain-of-thought (CoT) reasoning in large language models (LLMs) is fundamentally forward-directed: each step conditions only on prior to

safetyarxiv-cs-cl
4 Jun 2026
Safety

In-Context Graphical Inference

DGX agent

arXiv:2606.05042v1 Announce Type: cross Abstract: Marginal inference in discrete graphical models forces a choice between exactness and scalability: exact algorithms are intractable for high-treewidth

safetyarxiv-cs-cl
4 Jun 2026
Safety

Incredible. Goldman Sachs is the lead left on the SpaceX IPO, and somehow @ft fails to mention this in the headline below 🤦‍♂️

DGX agent

Incredible. Goldman Sachs is the lead left on the SpaceX IPO, and somehow @ft fails to mention this in the headline below 🤦‍♂️ Goldman Sachs expects SpaceX’s AI revenue to surge 100 times by 2030 http

safetygary-marcus--x
4 Jun 2026
Safety

Inference-Time Vulnerability Beyond Shallow Safety: Alignment Along Generation Trajectories

DGX agent

arXiv:2606.04778v1 Announce Type: new Abstract: Safety-aligned Large Language Models (LLMs) remain vulnerable to interventions during inference that redirect generation toward harmful outputs. Recent

safetyarxiv-cs-ai
4 Jun 2026
Safety

Instance-Level Post Hoc Uncertainty Quantification in Object Detection

DGX agent

arXiv:2606.04656v1 Announce Type: cross Abstract: Object detection is a safety-critical component of autonomous driving. It is essential to quantify the uncertainty in bounding-box predictions for saf

safetyarxiv-cs-ai
4 Jun 2026
Safety

Instant-Fold: In-Context Imitation Learning for Deformable Object Manipulation

DGX agent

arXiv:2606.04269v1 Announce Type: cross Abstract: Deformable object manipulation (DOM) is challenging due to high-dimensional, partially observable states that evolve through long-horizon, topology-ch

safetyarxiv-cs-ai
4 Jun 2026
Safety

Inverse Critical Experiment Design via Gradient Optimization and a Multigroup Attention-Based Neural Network Architecture

DGX agent

arXiv:2606.04033v1 Announce Type: new Abstract: The validation of advanced nuclear reactor designs and fuel concepts requires critical experiments with high neutronic similarity to the target technolo

safetyarxiv-cs-lg
4 Jun 2026
Safety

It seems like @GaryMarcus was right: the AI revenue models are imploding

DGX agent

It seems like @GaryMarcus was right: the AI revenue models are imploding Sam Altman said AI budgeting has recently become a 'huge issue' for some companies, something that 'never came up' earlier this

safetygary-marcus--x
4 Jun 2026
Safety

it takes balls to offer an IPO on a trillion dollar valuation when you are just months from bankruptcy. but maybe that’s what is happening. …

DGX agent

Gary Marcus critiques a company's decision to pursue an IPO at a trillion-dollar valuation despite allegedly being on the brink of financial collapse. The post suggests this represents either audaciou

safetygary-marcus--x
4 Jun 2026
Safety

Kinda crazy (but typical, sadly) that some moron just claimed that I was “doomer” who said that all jobs would be replaced when I actually p…

DGX agent

Kinda crazy (but typical, sadly) that some moron just claimed that I was “doomer” who said that all jobs would be replaced when I actually publicly predicted the opposite! Receipt, from my essay “25 p

safetygary-marcus--x
4 Jun 2026
Safety

KODA: Contrastive Representation Comparison and Alignment for Vision-Language Foundation Models

DGX agent

arXiv:2606.04180v1 Announce Type: new Abstract: Vision-language foundation models such as CLIP and SigLIP provide widely used representations for multimodal learning systems. While these models are ty

safetyarxiv-cs-lg
4 Jun 2026
Safety

La stratégie nationale en matière d’IA dévoilée aujourd’hui prône le développement d’une technologie sécuritaire, éthique, digne de confianc…

DGX agent

La stratégie nationale en matière d’IA dévoilée aujourd’hui prône le développement d’une technologie sécuritaire, éthique, digne de confiance, et bénéfique pour l’ensemble de la société — ce sont les

safetyyoshua-bengio--x
4 Jun 2026
Safety

Large Language Models in K-12 Education: Alignment with State Curriculum Standards and Student Personas

DGX agent

arXiv:2606.04846v1 Announce Type: new Abstract: As Large Language Models (LLMs) become increasingly popular in educational settings, they raise important questions about the ethical implications of th

safetyarxiv-cs-cl
4 Jun 2026
← Previous
1…112113114115116…267
Next →