AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,812 results
Safety

Plan-and-Verify Video Reward Reasoning with Spatio-Temporal Scene Graph Grounding

DGX agent

arXiv:2606.11838v1 Announce Type: new Abstract: Reward models for text-to-video (T2V) generation guide post-training but often fail at fine-grained semantic alignment. We trace this to two structural

safetyarxiv-cs-cv
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

ProcessThinker: Enhancing Multi-modal Large Language Models Reasoning via Rollout-based Process Reward

DGX agent

arXiv:2606.11209v1 Announce Type: cross Abstract: Visual question answering increasingly requires multi-step reasoning. Recent post-training with reinforcement learning under verifiable rewards (RLVR)

safetyarxiv-cs-ai
11 Jun 2026
Safety

Redesign Mixture-of-Experts Routers with Manifold Power Iteration

DGX agent

arXiv:2606.12397v1 Announce Type: cross Abstract: Router is the cornerstone component to the Mixture-of-Experts models. Serving as expert proxies, the rows of the router matrix compute their similarit

safetyarxiv-cs-ai
11 Jun 2026
Safety

Reinforcement Learning Disrupts Gradient-Based Adversarial Optimization

DGX agent

arXiv:2606.12251v1 Announce Type: cross Abstract: Gradient-based adversarial attacks remain a dominant threat to deep neural networks (DNNs), as they exploit gradient information to efficiently optimi

safetyarxiv-cs-ai
11 Jun 2026
Safety

Reinforcement Learning with Action-Triggered Observations

DGX agent

arXiv:2510.02149v2 Announce Type: replace Abstract: We introduce Action-Triggered Sporadically Traceable Markov Decision Processes (ATST-MDPs), a reinforcement learning framework for partial observabi

safetyarxiv-cs-lg
11 Jun 2026
Safety

Reverse Flow Matching: A Unified Framework for Online Reinforcement Learning with Diffusion and Flow Policies

DGX agent

arXiv:2601.08136v2 Announce Type: replace Abstract: Diffusion and flow policies are gaining prominence in online reinforcement learning (RL) due to their expressive power, yet training them efficientl

safetyarxiv-cs-lg
11 Jun 2026
Safety

Risk Under Pressure: Compute-Aware Evaluation of Adversarial Robustness in Language Models

DGX agent

arXiv:2606.11409v1 Announce Type: cross Abstract: Adversarial robustness evaluations of large language models (LLMs) typically report attack success rate (ASR) under fixed query budgets, implicitly tr

safetyarxiv-cs-ai
11 Jun 2026
Safety

RLCSD: Reinforcement Learning with Contrastive On-Policy Self-Distillation

DGX agent

arXiv:2606.11709v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) provides dense, token-level supervision for reasoning models by aligning a model's own distribution with the distri

safetyarxiv-cs-cl
11 Jun 2026
Safety

Runtime Enforcement of Hybrid System Properties

DGX agent

arXiv:2606.12022v1 Announce Type: cross Abstract: Runtime enforcement has emerged as a promising approach for ensuring the safety of autonomous and cyber-physical systems operating in uncertain and dy

safetyarxiv-cs-ai
11 Jun 2026
Safety

SAFER-Nav: Enhancing Safety for Visual Robot Navigation via Segmentation-Aware Fine-Tuning

DGX agent

arXiv:2606.11636v1 Announce Type: new Abstract: Vision-based navigation models, particularly foundation models, generate viable trajectories from RGB observations alone. However, even state-of-the-art

safetyarxiv-cs-ro
11 Jun 2026
Safety

SAGE: Answer-Conditioned Uncertainty Targets for Verbal Uncertainty Alignment

DGX agent

arXiv:2606.11512v1 Announce Type: new Abstract: Large language models increasingly express uncertainty through natural-language statements, yet these expressions often fail to reflect the model's samp

safetyarxiv-cs-cl
11 Jun 2026
Safety

Sample-Efficient Hypergradient Estimation for Decentralized Bi-Level Reinforcement Learning

DGX agent

arXiv:2603.14867v4 Announce Type: replace-cross Abstract: Many strategic decision-making problems, such as environment design for warehouse robots, can be naturally formulated as bi-level reinforcemen

safetyarxiv-cs-ai
11 Jun 2026
Safety

Scenario-based Probing and Steering Cultural Values in Large Language Models--Extended Version

DGX agent

arXiv:2606.11399v1 Announce Type: new Abstract: Large Language Models (LLMs) are deployed across cultural contexts but often reflect homogenized values inherited from training data. Evaluations of cul

safetyarxiv-cs-cl
11 Jun 2026
Safety

Schutzen: Evaluating LLM Safety in Bulgarian and German Contexts

DGX agent

arXiv:2606.11316v1 Announce Type: new Abstract: Large language models are increasingly deployed across professional domains, bringing hard-to-predict risks, including the generation of harmful or disr

safetyarxiv-cs-cl
11 Jun 2026
Safety

Seeing Before Colliding: Anticipatory Safe RL with Frozen Vision-Language Models

DGX agent

arXiv:2606.11266v1 Announce Type: new Abstract: The cost signal that constrained-RL algorithms optimize against is almost always reactive: the simulator emits a non-zero cost only after a collision ha

safetyarxiv-cs-lg
11 Jun 2026
Safety

Semantically-Aware Diver Activity Recognition Framework for Effective Underwater Multi-Human-Robot Collaboration

DGX agent

arXiv:2606.12374v1 Announce Type: cross Abstract: Effective multi-human-robot collaboration is essential for expanding human-led operations in the challenging and high-risk underwater environment. For

safetyarxiv-cs-cv
11 Jun 2026
Safety

Signed Compression Progress on a Sealed Audit is Goodhart-Resistant

DGX agent

arXiv:2606.11417v1 Announce Type: cross Abstract: Compression progress is a long-standing proposal for intrinsic motivation: reward an agent when its world model becomes better at predicting or compre

safetyarxiv-cs-ai
11 Jun 2026
Safety

SIL: Symbiotic Interactive Learning for Language-Conditioned Human-Agent Co-Adaptation

DGX agent

arXiv:2511.05203v3 Announce Type: replace Abstract: Today's autonomous agents, largely driven by foundation models (FMs), can understand natural language instructions and solve long-horizon tasks with

safetyarxiv-cs-ro
11 Jun 2026
Safety

Sovereign Assurance Boundary: Certificate-Bound Admission for Agentic Infrastructure

DGX agent

arXiv:2606.11632v1 Announce Type: cross Abstract: Agentic infrastructure introduces a critical control-plane authorization problem: non-deterministic reasoning systems can propose high-stakes mutation

safetyarxiv-cs-ai
11 Jun 2026
Safety

Spectrally Regularized Latent Flow Matching for Turbulence Generation

DGX agent

arXiv:2606.11691v1 Announce Type: new Abstract: Latent diffusion and flow matching have emerged as leading approaches for synthetic turbulence generation, yet they systematically under-represent dissi

safetyarxiv-cs-lg
11 Jun 2026
Safety

Steering Multirobot Behavior via Closed-Loop Affine Activation Editing

DGX agent

arXiv:2606.11489v1 Announce Type: new Abstract: Real-world robots need to adapt their behavior beyond the envelope of their pre-trained policy. Policy finetuning or retraining are options, but they ri

safetyarxiv-cs-ro
11 Jun 2026
Safety

SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning

DGX agent

arXiv:2606.11770v1 Announce Type: new Abstract: Spatial reasoning remains a challenge for Multimodal Large Language Models (MLLMs), as it requires reliable multi-hop inference over both intermediate s

safetyarxiv-cs-ai
11 Jun 2026
Safety

TacCoRL: Integrating Tactile Feedback into VLA via Simulation

DGX agent

arXiv:2606.11743v1 Announce Type: cross Abstract: Vision-language-action (VLA) models provide strong visual, language, and action priors for robot manipulation, but visual observations alone often mis

safetyarxiv-cs-lg
11 Jun 2026
Safety

Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning

DGX agent

arXiv:2606.11853v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) depend on in-context learning (ICL) for rapid task adaptation, but their scalability is severely limited by

safetyarxiv-cs-ai
11 Jun 2026
Safety

The Algorithm Is Not the Behavior: Learned Priors Override Look-Ahead in a Chess-Playing Neural Network

DGX agent

arXiv:2508.21380v3 Announce Type: replace-cross Abstract: Recent mechanistic work has uncovered learned algorithms within neural networks, from modular arithmetic to search and planning in game-playin

safetyarxiv-cs-ai
11 Jun 2026
Safety

The Art of Interrogation: Consistency Amplifies Factuality in Spatial Reasoning

DGX agent

arXiv:2606.11918v1 Announce Type: new Abstract: Current Large Reasoning Models (LRMs) exhibit remarkable general capabilities but significantly underperform in spatial reasoning tasks. Existing approa

safetyarxiv-cs-ai
11 Jun 2026
Safety

The Unreasonable Effectiveness of Discrete-Time Gaussian Process Mixtures for Robot Policy Learning

DGX agent

arXiv:2505.03296v2 Announce Type: replace-cross Abstract: We present Mixture of Discrete-time Gaussian Processes (MiDiGap), a novel approach for flexible policy representation and imitation learning i

safetyarxiv-cs-ai
11 Jun 2026
Safety

This was perhaps the most controversial aspect of the guardrails around Fable, now being rolled back.

DGX agent

This was perhaps the most controversial aspect of the guardrails around Fable, now being rolled back. Very pleased to hear Anthropic have walked back this policy https://simonwillison.net/2026/Jun/11/

safetyethan-mollick--x
11 Jun 2026
Safety

To Intervene or Not: Guiding Inference-time Alignment with Probabilistic Model Blending

DGX agent

arXiv:2606.11201v1 Announce Type: cross Abstract: The wide deployment of LLMs has made model alignment necessary to make newly trained models safely and effectively respond to user instructions. Among

safetyarxiv-cs-ai
11 Jun 2026
Safety

Toward Preference-aligned Large Language Models via Residual-based Model Steering

DGX agent

arXiv:2509.23982v2 Announce Type: replace-cross Abstract: Preference alignment is a critical step in making Large Language Models (LLMs) useful and aligned with (human) preferences. Existing approache

safetyarxiv-cs-ai
11 Jun 2026
Safety

Towards a Bridge Layer Between Bibliographic and Formalized Mathematical Knowledge

DGX agent

arXiv:2606.11430v1 Announce Type: cross Abstract: Mathematical knowledge is split between bibliographic databases (e.g., MathSciNet, zbMATH Open) and formal proof libraries (e.g., Lean mathlib), preve

safetyarxiv-cs-ai
11 Jun 2026
Safety

Towards Conditional Feature Alignment for Cross-Domain Counting

DGX agent

arXiv:2506.17137v3 Announce Type: replace Abstract: Object counting models often degrade under cross-domain deployment because density composition varies across domains and is itself task-relevant. St

safetyarxiv-cs-cv
11 Jun 2026
Safety

Traceable Virtual Sea Trials in the Marine Robotics Unity Simulator for Manoeuvring Assessment of Unmanned Surface Vehicles

DGX agent

arXiv:2606.12349v1 Announce Type: new Abstract: Accurate identification of hydrodynamic derivatives is essential for control and navigation of Unmanned Surface Vehicles (USVs), but high-fidelity manoe

safetyarxiv-cs-ro
11 Jun 2026
Safety

Traits Run Deeper: Trait-Specific Asymmetric Fusion for Personality Assessment

DGX agent

arXiv:2606.11269v1 Announce Type: new Abstract: Personality assessment aims to infer stable personality traits from dynamic behaviors across language, voice, and facial cues. Since different personali

safetyarxiv-cs-cv
11 Jun 2026
Safety

UGV-Conditioned Multi-UAV Informative Planning on a Shared Exposure Belief

DGX agent

arXiv:2606.12306v1 Announce Type: new Abstract: Safe ground navigation in large, threat-augmented environments requires aerial support that actively reduces the risks that a ground vehicle faces along

safetyarxiv-cs-ro
11 Jun 2026
Safety

UniIntervene: Agentic Intervention for Efficient Real-World Reinforcement Learning

DGX agent

arXiv:2606.12372v1 Announce Type: cross Abstract: Human-in-the-loop reinforcement learning (HiL-RL) has emerged as an effective paradigm for real-world robotic manipulation, enabling online policy imp

safetyarxiv-cs-lg
11 Jun 2026
Safety

UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA

DGX agent

arXiv:2606.11740v1 Announce Type: cross Abstract: We study whether grounded reasoning supervision from abundant 2D medical images can improve 3D medical VQA when both input types are aligned through a

safetyarxiv-cs-cl
11 Jun 2026
Safety

UR-BERT: Scaling Text Encoders for Massively Multilingual TTS Through Universal Romanization and Speech Token Prediction

DGX agent

arXiv:2606.11681v1 Announce Type: new Abstract: We propose UR-BERT, a Romanized transcription-based text-to-speech (TTS) encoder for massively multilingual TTS systems. Conventional grapheme-to-phonem

safetyarxiv-cs-cl
11 Jun 2026
Safety

Urban Heat MiniCubes: An AI-Ready dataset for urban heat research

DGX agent

arXiv:2606.11534v1 Announce Type: cross Abstract: Urban heat is amplified by impermeable surfaces and heterogeneous built environments, yet street-level variability remains difficult to quantify becau

safetyarxiv-cs-lg
11 Jun 2026
Safety

Very pleased to hear Anthropic have walked back this policy https://simonwillison.net/2026/Jun/11/anthropic-walks-back-policy/

DGX agent

Very pleased to hear Anthropic have walked back this policy https://simonwillison.net/2026/Jun/11/anthropic-walks-back-policy/ BREAKING NEWS: Anthropic's latest model will NOT help you if it thinks yo

safetyjeremy-howard--x
11 Jun 2026
Safety

ViT-FREE: Efficient Face Recognition via Early Exiting and Synthetic Adaptation

DGX agent

arXiv:2606.12023v1 Announce Type: new Abstract: Vision Transformers (ViTs) have gained significant attention in computer vision and shown strong potential for face recognition (FR). However, their hig

safetyarxiv-cs-cv
11 Jun 2026
Safety

VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving

DGX agent

arXiv:2606.12396v1 Announce Type: new Abstract: Vision-language-action (VLA) models can describe scenes and reason about them in language, yet still struggle to ground their actions in the dense 3D wo

safetyarxiv-cs-cv
11 Jun 2026
Safety

What are smart people are saying about OpenAI's IPO filing? 🤔

DGX agent

Gary Marcus discusses expert commentary and analysis regarding OpenAI's initial public offering filing, likely examining implications for the AI industry, company valuation, and competitive landscape.

safetygary-marcus--x
11 Jun 2026
Safety

When Context Returns: Toward Robust Internalization in On-Policy Distillation

DGX agent

arXiv:2606.11627v1 Announce Type: cross Abstract: Recent work has shown that on-policy distillation can internalize privileged context, such as system prompts or task hints, into a student model so th

safetyarxiv-cs-ai
11 Jun 2026
Safety

When Researchers Say Mental Model/Theory of Mind of AI, What Are They Really Talking About?

DGX agent

arXiv:2510.02660v2 Announce Type: replace-cross Abstract: When researchers claim AI systems possess ToM or mental models, they are fundamentally discussing behavioral predictions and bias corrections

safetyarxiv-cs-ai
11 Jun 2026
Safety

Which Speech Representation Better Matches Text-Native Reasoning? A Study of Speech-Text Alignment on Frame Rate and Representation

DGX agent

arXiv:2606.12199v1 Announce Type: cross Abstract: Spoken dialogue models typically start from text LLM backbones, yet reasoning often degrades when conditioning on speech instead of text. We attribute

safetyarxiv-cs-cl
11 Jun 2026
Safety

Wrong horse and too early. Gary nailed it. Time will show this.

DGX agent

Wrong horse and too early. Gary nailed it. Time will show this. Masa can hype AI all he likes, but apparently he can’t even get a loan on his OpenAI shares. 🤔 But I actually half agree—and it might su

safetygary-marcus--x
11 Jun 2026
Safety

100%, hallucinations are baked in (as I have been saying since 2001) and that is why basically no LLM company afford to operate in Germany n…

DGX agent

100%, hallucinations are baked in (as I have been saying since 2001) and that is why basically no LLM company afford to operate in Germany now. We need a better technology. @GaryMarcus Lol, and Google

safetygary-marcus--x
10 Jun 2026
← Previous
1…9394959697…267
Next →