AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,816 results
Safety

Rationalize: Shared Semantic Reasoning for Human-AI Alignment

DGX agent

arXiv:2605.30632v1 Announce Type: cross Abstract: We introduce Rationalize, a role-pair framework for shared semantic reasoning between humans and AI models in data-driven sensemaking. Building on ide

safetyarxiv-cs-ai
1 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RDGen: Demonstration Generation for High-Quality Robot Learning via Reinforcement Learning

DGX agent

arXiv:2605.30957v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for general-purpose robot control. However, their performance remains fundament

safetyarxiv-cs-ro
1 Jun 2026
Safety

Reading Between the Citations: A Typed Claim Network for Scientific Literature

DGX agent

arXiv:2605.30966v1 Announce Type: cross Abstract: Knowledge graphs over corpora of inter-referencing documents - scholarly papers, legal opinions, policy briefs - encode the topology of reference but

safetyarxiv-cs-ai
1 Jun 2026
Safety

REAL: Regression-Aware Reinforcement Learning for LLM-as-a-Judge

DGX agent

arXiv:2603.17145v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as automated evaluators that assign numeric scores to model outputs, a paradigm known a

safetyarxiv-cs-ai
1 Jun 2026
Safety

Reassessing Extractive QA Datasets at Scale: LLM-as-a-Judge and In-Depth Analyses

DGX agent

arXiv:2504.11972v3 Announce Type: replace Abstract: Extractive QA tasks are commonly evaluated using Exact Match (EM) and F1-score, but these metrics often fail to reflect true model performance. Rece

safetyarxiv-cs-cl
1 Jun 2026
Safety

Reinforced sequential Monte Carlo for amortised sampling

DGX agent

arXiv:2510.11711v2 Announce Type: replace Abstract: This paper proposes a synergy of amortised and particle-based methods for sampling from distributions defined by unnormalised density functions. We

safetyarxiv-cs-lg
1 Jun 2026
Safety

Reinforcement Learning Amplifies Emergent Misalignment from Harmless Rewards

DGX agent

arXiv:2605.31328v1 Announce Type: new Abstract: Emergent misalignment (EM) is the surprising tendency of language models to become broadly misaligned after fine-tuning on narrowly misaligned examples.

safetyarxiv-cs-cl
1 Jun 2026
Safety

Reinterpreting Safety Thresholds as Neuron Spiking Thresholds

DGX agent

arXiv:2605.30368v1 Announce Type: cross Abstract: Surrogate Safety Measures (SSMs) are extensively utilised in the evaluation of traffic risk in automated driving contexts. However, the majority of SS

safetyarxiv-cs-ai
1 Jun 2026
Safety

Representation Collapse in Sequential Post-Training of Large Language Models

DGX agent

arXiv:2605.30524v1 Announce Type: new Abstract: Large language models are now adapted through chains of post-training stages rather than through a single instruction-tuning pass. This paper studies wh

safetyarxiv-cs-lg
1 Jun 2026
Safety

Rethinking Multimodal Few-Shot 3D Point Cloud Segmentation: From Fused Refinement to Decoupled Arbitration

DGX agent

arXiv:2601.01456v2 Announce Type: replace-cross Abstract: In this paper, we revisit multimodal few-shot 3D point cloud semantic segmentation (FS-PCS), identifying a conflict in 'Fuse-then-Refine' para

safetyarxiv-cs-ai
1 Jun 2026
Safety

// Reusable Context Engineering // Context bloat quietly kills long-horizon runs, but you can fix it from the outside without fine-tuning th…

DGX agent

// Reusable Context Engineering // Context bloat quietly kills long-horizon runs, but you can fix it from the outside without fine-tuning the underlying agent. (bookmark this) Context management is us

safetydair-ai--x
1 Jun 2026
Safety

Revisiting Zeroth-Order Hessian Approximation: A Single-Step Policy Optimization Lens

DGX agent

arXiv:2605.30960v1 Announce Type: new Abstract: Accurate Zeroth-Order (ZO) Hessian estimation is a cornerstone of derivative-free methods, essential for tasks such as bilevel optimization, Bayesian in

safetyarxiv-cs-lg
1 Jun 2026
Safety

Routing on the Stiefel Manifold: When Does Adaptive Subspace Selection Help for Cross-Domain EEG Decoding?

DGX agent

arXiv:2605.31043v1 Announce Type: cross Abstract: Cross-domain EEG decoding remains challenging despite advances in Riemannian deep learning: covariance matrices from different subjects occupy systema

safetyarxiv-cs-ai
1 Jun 2026
Safety

Safeguarding Text-to-Image Generation via Inference-Time Prompt-Noise Optimization

DGX agent

arXiv:2412.03876v2 Announce Type: replace Abstract: Text-to-Image (T2I) diffusion models are widely recognized for their ability to generate high-quality and diverse images based on text prompts. Howe

safetyarxiv-cs-cv
1 Jun 2026
Safety

Scalable Constrained Multi-Agent Reinforcement Learning via State Augmentation and Consensus for Separable Dynamics

DGX agent

arXiv:2605.30461v1 Announce Type: cross Abstract: We present a distributed approach for constrained Multi-Agent Reinforcement Learning (MARL) that combines state-augmented policy learning with distrib

safetyarxiv-cs-ai
1 Jun 2026
Safety

Scaling Multi-Agent Environment Co-Design with Diffusion Models

DGX agent

arXiv:2511.03100v2 Announce Type: replace-cross Abstract: The agent-environment co-design paradigm jointly optimises agent policies and environment configurations in search of improved system performa

safetyarxiv-cs-ai
1 Jun 2026
Safety

SCOPE: Selective Conformal Optimized Pairwise LLM Judging

DGX agent

arXiv:2602.13110v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used as scalable judges in pairwise evaluation, but they remain prone to miscalibration and bias

safetyarxiv-cs-ai
1 Jun 2026
Safety

SDM-Q: Cost-Aware Staged Decision-Making for Multi-Omics Classification with Deep Q-Learning

DGX agent

arXiv:2605.31014v1 Announce Type: new Abstract: Multi-omics data provide complementary molecular characterizations of disease phenotypes and play an important role in disease diagnosis and subtype cla

safetyarxiv-cs-lg
1 Jun 2026
Safety

Secure AI agents with Policy and Lambda interceptors in Amazon Bedrock AgentCore gateway

DGX agent

In this post, we use a lakehouse data agent to demonstrate how you can use Policy for deterministic access control and Lambda interceptors for dynamic validation. We then show how to combine Lambda in

safetyaws-ml-blog
1 Jun 2026
Safety

Seeing Before Agreeing: Aligning Multi-Agent Consensus with Visual Evidence

DGX agent

arXiv:2605.30698v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on visual question answering (VQA). To mitigate individual hallucinations and blind spo

safetyarxiv-cs-ai
1 Jun 2026
Safety

Semantic Motion Anchors: Bridging Motion and Meaning in Co-Speech Gestures

DGX agent

arXiv:2605.30608v1 Announce Type: new Abstract: Learning a shared representation between spoken text and gesture is central to co-speech gesture retrieval, synthesis, and understanding, but remains ch

safetyarxiv-cs-cl
1 Jun 2026
Safety

SemStruct: Contextualizing Semantic Embeddings with Structural Information for Schema Matching

DGX agent

arXiv:2605.30729v1 Announce Type: new Abstract: Schema matching is a fundamental step in integrating heterogeneous data sources. While Pre-trained Language Models (PLMs) have revolutionized this task

safetyarxiv-cs-lg
1 Jun 2026
Safety

Simulation of collision avoidance behavior in crowd movement by data-driven approach

DGX agent

arXiv:2605.31210v1 Announce Type: cross Abstract: Crowd movement simulation is essential for pedestrian safety management and facility layout optimization. Data-driven models enhance trajectory predic

safetyarxiv-cs-ai
1 Jun 2026
Safety

Skill is Not One-Size-Fits-All: Model-Aware Skill Alignment for LLM Agents

DGX agent

arXiv:2605.30723v1 Announce Type: new Abstract: LLM agents increasingly retrieve externally curated skills-procedural instructions retrieved at decision time-to improve performance on long-horizon int

safetyarxiv-cs-cl
1 Jun 2026
Safety

Smaller Models are Natural Explorers for Policy-Level Diversity in GRPO

DGX agent

arXiv:2605.30789v1 Announce Type: cross Abstract: We identify a new dimension for enhancing rollout diversity in Group Relative Policy Optimization (GRPO) for LLMs. While GRPO relies on diverse rollou

safetyarxiv-cs-ai
1 Jun 2026
Safety

Softly Constrained Denoisers for Diffusion Models Applied to Partial Differential Equations

DGX agent

arXiv:2512.14980v4 Announce Type: replace Abstract: Diffusion models have become a powerful generative prior for solutions of partial differential equations (PDEs). Existing approaches enforce physica

safetyarxiv-cs-lg
1 Jun 2026
Safety

SpongeBob: Sync-Aware Harmonious Audio-Visual Generative Editing

DGX agent

arXiv:2605.25193v2 Announce Type: replace Abstract: Visual and acoustic events in the physical world are inherently coupled, yet existing video editing methods typically adopt decoupled pipelines, lac

safetyarxiv-cs-cv
1 Jun 2026
Safety

Stateful Online Monitoring Catches Distributed Agent Attacks

DGX agent

arXiv:2605.31593v1 Announce Type: cross Abstract: Language models can find thousands of severe software vulnerabilities, and agents are increasingly being misused for cyberattacks. To avoid detection,

safetyarxiv-cs-ai
1 Jun 2026
Safety

Structural Bias Beyond Homophily: A Study of Fairness in Link Prediction

DGX agent

arXiv:2602.11802v2 Announce Type: replace Abstract: Graph link prediction (LP) plays a critical role in socially impactful applications such as job recommendation and friendship formation, making fair

safetyarxiv-cs-lg
1 Jun 2026
Safety

Structure-Induced Information for Rerooting Levin Tree Search

DGX agent

arXiv:2605.30664v1 Announce Type: new Abstract: Subgoal-based policy tree search, which uses a policy to guide search, is effective for complex single-agent deterministic problems but often relies on

safetyarxiv-cs-ai
1 Jun 2026
Safety

Supervised Learning as Lossy Compression: Characterizing Generalization and Sample Complexity via Finite Blocklength Analysis

DGX agent

arXiv:2602.04107v2 Announce Type: replace Abstract: This paper presents a novel information-theoretic perspective on generalization in machine learning by framing the learning problem within the conte

safetyarxiv-cs-lg
1 Jun 2026
Safety

Supervised Training Rapidly Degrades Early Visual Cortex Alignment Across Biologically Plausible Learning Rules

DGX agent

arXiv:2605.30556v1 Announce Type: new Abstract: Random, untrained neural networks consistently match or exceed trained networks in representational similarity to early visual cortex. This puzzling fin

safetyarxiv-cs-lg
1 Jun 2026
Safety

Surface Constraint Policy for Learning Surface-Constrained and Dynamically Feasible Robot Skills

DGX agent

arXiv:2605.31321v1 Announce Type: new Abstract: Diffusion-based imitation learning methods have driven rapid progress in robot dexterous manipulation tasks. However, they have limitations when applied

safetyarxiv-cs-ro
1 Jun 2026
Safety

Synthetic Stimuli, Real Gains: Rethinking VLM Fine-Tuning Through Fully Controlled Data Generation

DGX agent

arXiv:2511.11440v3 Announce Type: replace-cross Abstract: Performance gains of Vision Language Models (VLMs) obtained by fine-tuning are generally based on ad hoc data collection and annotation of rea

safetyarxiv-cs-cl
1 Jun 2026
Safety

TALON: Token-Aligned Lightweight Adapters for 6-DoF Spacecraft Pose Estimation

DGX agent

arXiv:2605.31217v1 Announce Type: new Abstract: Monocular 6-DoF spacecraft pose estimation methods predominantly process individual frames, discarding the temporal information present in an image sequ

safetyarxiv-cs-cv
1 Jun 2026
Safety

TARIC: Memory-Augmented Traversability-Aware Outdoor VLN under Interrupted Semantic Cues

DGX agent

arXiv:2605.31121v1 Announce Type: cross Abstract: Outdoor vision-language navigation (VLN) in long-range, open-world environments is frequently disrupted by semantic-cue interruptions, where informati

safetyarxiv-cs-ai
1 Jun 2026
Safety

Task-Focused Memorization for Multimodal Agents

DGX agent

arXiv:2605.31075v1 Announce Type: new Abstract: Long-term memory is essential for multimodal agents to build coherent experience, accumulate world knowledge, and achieve continual learning. However, c

safetyarxiv-cs-cv
1 Jun 2026
Safety

THE DEFINITIVE AI CONVERSATION OF THE YEAR. MUST LISTEN. @JG_Nuke @GaryMarcus @MacrostrategyP https://open.substack.com/pub/georgenoble/p/ai…

DGX agent

THE DEFINITIVE AI CONVERSATION OF THE YEAR. MUST LISTEN. @JG_Nuke @GaryMarcus @MacrostrategyP https://open.substack.com/pub/georgenoble/p/ai-the-biggest-capital-misallocation-d18?r=35saq&utm_medium=io

safetygary-marcus--x
1 Jun 2026
Safety

The Flip Side of RLHF: On-Policy Feedback for Reward Model Self-Supervised Improvement

DGX agent

arXiv:2605.30888v1 Announce Type: new Abstract: Building strong reward models (RMs) for language model alignment is bottlenecked by the cost and difficulty of acquiring diverse and reliable preference

safetyarxiv-cs-cl
1 Jun 2026
Safety

The Global Landscape of Environmental AI Regulation: From the Cost of Reasoning to a Right to Green AI

DGX agent

arXiv:2603.00068v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI) systems impose substantial and growing environmental costs, yet transparency about these impacts has declined eve

safetyarxiv-cs-ai
1 Jun 2026
Safety

The OpenAI Foundation is doing a lot of wonderful things. Helping society become resilient to AI is going to be incredibly important. Much m…

DGX agent

The OpenAI Foundation is doing a lot of wonderful things. Helping society become resilient to AI is going to be incredibly important. Much more to come here! AI is advancing quickly. Society’s ability

safetysam-altman--x
1 Jun 2026
Safety

The Refutability Gap: Challenges in Validating Reasoning by Large Language Models

DGX agent

arXiv:2601.02380v4 Announce Type: replace-cross Abstract: Recent reports claim that Large Language Models (LLMs) have achieved the ability to derive new science and exhibit human-level general intelli

safetyarxiv-cs-ai
1 Jun 2026
Safety

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning

DGX agent

arXiv:2605.31404v1 Announce Type: cross Abstract: Large Language Model (LLM)-based navigation systems commonly construct explicit spatial representations (e.g., topological graphs, semantic raster map

safetyarxiv-cs-ai
1 Jun 2026
Safety

“The technology worked. The value didn’t arrive,” Bain concluded in the report. https://www.bloomberg.com/news/newsletters/2026-06-01/bain-s…

DGX agent

“The technology worked. The value didn’t arrive,” Bain concluded in the report. https://www.bloomberg.com/news/newsletters/2026-06-01/bain-survey-ai-delivers-less-cost-reduction-than-many-firms-predic

safetygary-marcus--x
1 Jun 2026
Safety

This was right five years ago, and still is: “Large scale pretrained models are certainly likely to figure prominently in artificial intelli…

DGX agent

This was right five years ago, and still is: “Large scale pretrained models are certainly likely to figure prominently in artificial intelligence for the near future, and play an important role in com

safetygary-marcus--x
1 Jun 2026
Safety

Traceable by Design: An LLM Pipeline and Dashboard for EU Regulatory Consultation Analysis

DGX agent

arXiv:2605.30995v1 Announce Type: cross Abstract: Public consultations generate large volumes of data in the form of stakeholder submissions that are practically unfeasible to analyse manually. We pre

safetyarxiv-cs-cl
1 Jun 2026
Safety

Trust-Region Behavior Blending for On-Policy Distillation

DGX agent

arXiv:2605.31159v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student on prefixes sampled from its own policy while matching a stronger teacher. This addresses the prefix mis

safetyarxiv-cs-ai
1 Jun 2026
Safety

TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation

DGX agent

arXiv:2605.31590v1 Announce Type: cross Abstract: Text-to-video (T2V) generation faces challenging questions when generating videos with long horizons containing multiple events. Inspired by the intri

safetyarxiv-cs-ai
1 Jun 2026
← Previous
1…129130131132133…267
Next →