AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
Safety

On the Role of Computation in Reinforcement Learning

DGX agent

arXiv:2602.05999v4 Announce Type: replace Abstract: How does the amount of compute available to a reinforcement learning (RL) policy affect its learning? Can policies using a fixed amount of parameter

safetyarxiv-cs-lg
3 Jul 2026
Safety

On the Sample Efficiency of Inverse Dynamics Models for Semi-Supervised Imitation Learning

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2602.02762v2 Announce Type: replace Abstract: Semi-supervised imitation learning (SSIL) consists in learning a policy from a small dataset of action-labeled trajectories and a much larger datase

safetyarxiv-cs-lg
3 Jul 2026
Safety

One Demonstration Is Enough for Real-World Robotic Reinforcement Learning

DGX agent

arXiv:2607.01651v1 Announce Type: new Abstract: Learning effective robot control policies on physical hardware is challenging due to costly data collection and the difficulty of reward specification.

safetyarxiv-cs-ro
3 Jul 2026
Safety

Online Resource Allocation with Continuous Random Consumption: Regret under Degeneracy

DGX agent

arXiv:2607.02196v1 Announce Type: new Abstract: We study online resource allocation when both rewards and consumption sizes may be continuously distributed. Requests arrive sequentially and must be ac

safetyarxiv-cs-lg
3 Jul 2026
Safety

Online Safety Monitoring for LLMs

DGX agent

arXiv:2607.02510v1 Announce Type: new Abstract: Despite alignment training, LLMs remain prone to generating unsafe outputs at deployment time. Monitoring outputs online and raising an alarm when safet

safetyarxiv-cs-ai
3 Jul 2026
Safety

OpenAI offers feds a stake, Anthropic gets out of AI model jail and Meta wants to be a neocloud

DGX agent

OpenAI reportedly has floated giving the U.S. government a 5% stake in the company, perhaps the start of a series of such stakes in other AI companies as well. This no doubt has traditional anti-indus

safetysiliconangle
3 Jul 2026
Safety

Optimizing Visual Generative Models via Distribution-wise Rewards

DGX agent

arXiv:2607.02291v1 Announce Type: new Abstract: Conventional reinforcement learning strategies for visual generation typically employ sample-wise reward functions, yet this practice frequently results

safetyarxiv-cs-lg
3 Jul 2026
Safety

Overthink-Triggered Slowdown Attacks on LVLM-Based Robotic Systems

DGX agent

arXiv:2607.01518v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have been increasingly integrated into robotic systems. However, these models may exhibit overthinking behaviors,

safetyarxiv-cs-ro
3 Jul 2026
Safety

Playing 20 Question Game with Policy-Based Reinforcement Learning

DGX agent

arXiv:1808.07645v5 Announce Type: replace-cross Abstract: The 20 Questions (Q20) game is a well known game which encourages deductive reasoning and creativity. In the game, the answerer first thinks o

safetyarxiv-cs-ai
3 Jul 2026
Safety

Predicting Closed-Loop Performance of Latent World Models: Offline Checkpoint Selection for MPC and Model-Based RL Under Non-Markovian Rewards in LunarLander

DGX agent

arXiv:2607.01736v1 Announce Type: cross Abstract: We study how to predict the downstream closed-loop performance of a learned latent world model from validation-time diagnostics alone. Choosing the ri

safetyarxiv-cs-ai
3 Jul 2026
Safety

Prediction Sets for Counterfactual Decisions: Coverage, Optimality, and Conformal Prediction

DGX agent

arXiv:2607.02206v1 Announce Type: cross Abstract: Predictions are increasingly used to guide high-stakes decisions, from treatment selection to policy making. To ensure reliability with imperfect pred

safetyarxiv-cs-lg
3 Jul 2026
Safety

Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

DGX agent

arXiv:2607.02234v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) has emerged as a promising paradigm for improving LLM reasoning, where a privileged teacher with access to reference

safetyarxiv-cs-ai
3 Jul 2026
Safety

Quantifying the Uncertainty of Blindly Estimated Room Embeddings Using a Dispersion-Calibrated Score

DGX agent

arXiv:2607.01527v1 Announce Type: cross Abstract: Room embeddings derived from reverberant speech are often unreliable: speech content and recording degradation can alter the representation even when

safetyarxiv-cs-lg
3 Jul 2026
Safety

Quantum-Inspired Vision: Leveraging Wave-Particle Duality for Low-Illumination Enhancement

DGX agent

arXiv:2607.01731v1 Announce Type: cross Abstract: This study provides a theoretical expansion of the recent Data Relativistic Uncertainty (DRU) framework by formalizing a physics-to-AI paradigm for im

safetyarxiv-cs-lg
3 Jul 2026
Safety

Rank-Then-Act: Reward-Free Control from Frame-Order Progress

DGX agent

arXiv:2607.01897v1 Announce Type: cross Abstract: We introduce Rank-Then-Act (RTA), a framework for learning control policies from expert video demonstrations without environment rewards. RTA trains a

safetyarxiv-cs-ai
3 Jul 2026
Safety

RedCoder: Automated Multi-Turn Red Teaming for Code LLMs

DGX agent

arXiv:2507.22063v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) for code generation (i.e., Code LLMs) have demonstrated impressive capabilities in AI-assisted software developme

safetyarxiv-cs-ai
3 Jul 2026
Safety

Rethinking Post-Hoc Calibration in Semantic Segmentation

DGX agent

arXiv:2607.01902v1 Announce Type: cross Abstract: Reliable confidence estimates are essential in semantic segmentation, especially in safety-critical settings where overconfident errors can mislead do

safetyarxiv-cs-lg
3 Jul 2026
Safety

Risk Architecture for AI-Native Engineering Teams: An Organizational Framework for Agentic System Governance

DGX agent

arXiv:2607.01421v1 Announce Type: cross Abstract: Engineering management research has produced mature frameworks for software risk: ownership by feature, escalation by severity, and assurance by test

safetyarxiv-cs-ai
3 Jul 2026
Safety

SABER: A Semantic-Aligned Brain Network Analysis Framework via Multi-scale Hypergraphs

DGX agent

arXiv:2607.01901v1 Announce Type: cross Abstract: Effective brain disease diagnosis requires the synergy of brain connectivity patterns and high-level semantic knowledge. Existing methods, however, la

safetyarxiv-cs-ai
3 Jul 2026
Safety

Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model

DGX agent

arXiv:2607.01595v1 Announce Type: new Abstract: As the scale and complexity of cloud-based AI systems continue to escalate, ensuring service reliability through rapid fault detection and adaptive reco

safetyarxiv-cs-ai
3 Jul 2026
Safety

Safeguarding LLM Agents from Misalignment through Provenance Analysis

DGX agent

arXiv:2607.01236v1 Announce Type: cross Abstract: As LLM agents gain increasing access to powerful tools, ensuring that their actions are aligned with the user's intent becomes critical. When an agent

safetyarxiv-cs-ai
3 Jul 2026
Safety

Sim2Real-AD: A Modular Sim-to-Real Framework for Deploying VLM-Guided Reinforcement Learning in Real-World Autonomous Driving

DGX agent

arXiv:2604.03497v2 Announce Type: replace-cross Abstract: Vision-language-model (VLM)-guided reinforcement learning (RL) has recently attracted significant attention for it, replacing brittle hand-cra

safetyarxiv-cs-ai
3 Jul 2026
Safety

SPLC: Social Preference Learning for Crowd Robot Navigation

DGX agent

arXiv:2607.01925v1 Announce Type: new Abstract: Offline reinforcement learning (RL) holds significant potential for crowd robot navigation in human-robot coexistence applications. However, the inheren

safetyarxiv-cs-ro
3 Jul 2026
Safety

Structuring the Space of Sociotechnical Alignment

DGX agent

arXiv:2607.01250v1 Announce Type: cross Abstract: Sociotechnical alignment concerns the social desirability of AI behavior and is thus inherently normative, not merely technical. While NLP research in

safetyarxiv-cs-ai
3 Jul 2026
Safety

The Rising Unsustainability of AI Graphics Cards Production

DGX agent

arXiv:2607.01258v1 Announce Type: cross Abstract: The rapid advancement of Artificial Intelligence (AI) has been accompanied by significant increases in computational and environmental costs, driven b

safetyarxiv-cs-ai
3 Jul 2026
Safety

Tight Lower Bounds for the Multi-Secretary Problem via Bellman Certificates

DGX agent

arXiv:2607.02150v1 Announce Type: cross Abstract: This paper studies additive regret in the multi-secretary problem, defined as the gap between the expected offline prophet reward and the reward of th

safetyarxiv-cs-lg
3 Jul 2026
Safety

Towards Learning Representations of Policies in Two-Player Zero-Sum Imperfect-Information Games

DGX agent

arXiv:2607.01498v1 Announce Type: new Abstract: We investigate the problem of learning useful policy representations (embeddings) in two-player zero-sum imperfect-information games. We make three cont

safetyarxiv-cs-lg
3 Jul 2026
Safety

Transformer Geometry Observatory TGO-II: Representational Similarity Observatory

DGX agent

arXiv:2607.02386v1 Announce Type: cross Abstract: While Vision Transformers have achieved remarkable success across computer vision and language applications, the geometric evolution of their internal

safetyarxiv-cs-lg
3 Jul 2026
Safety

Transport Discrepancy as a Reliability Signal for Vision-Language-Action Models

DGX agent

arXiv:2512.01715v2 Announce Type: replace Abstract: Vision-language-action (VLA) models that generate continuous action chunks via flow matching lack an internal signal for judging whether a given pre

safetyarxiv-cs-ro
3 Jul 2026
Safety

VLAFlow: A Unified Training Framework for Vision-Language-Action Models via Co-training and Future Latent Alignment

DGX agent

arXiv:2607.01586v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have recently advanced robotic manipulation, yet the effects of different robot-data pre-training paradigms remai

safetyarxiv-cs-ai
3 Jul 2026
Safety

WaveLander: A Generalizable Hierarchical Control Framework for UAV Landing on Wave-Disturbed Platforms via Reinforcement Learning

DGX agent

arXiv:2607.01281v1 Announce Type: new Abstract: Autonomous landing of unmanned aerial vehicles (UAVs) on wave-disturbed marine platforms remains challenging due to stochastic platform motion, time-var

safetyarxiv-cs-ro
3 Jul 2026
Safety

WBMM: Windowed Batch Matrix Multiplication for Efficient Large Receptive Field Convolution

DGX agent

arXiv:2607.02097v1 Announce Type: cross Abstract: Large kernel depthwise convolutions achieve strong performance but suffer from significant degradation as kernel size grows due to irregular memory ac

safetyarxiv-cs-lg
3 Jul 2026
Safety

What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates

DGX agent

arXiv:2607.02507v1 Announce Type: new Abstract: LLM agents will increasingly act in socially structured settings where role, audience, and relational context can shape what is advantageous or costly t

safetyarxiv-cs-ai
3 Jul 2026
Safety

When Sample Selection Bias Precipitates Model Collapse

DGX agent

arXiv:2606.13732v2 Announce Type: replace Abstract: The proliferation of recursive training on synthetic data can alleviate data scarcity but risks model collapse, where repeated training erodes distr

safetyarxiv-cs-ai
3 Jul 2026
Safety

When Should Service Agents Reconsider? Difficulty-Routed Control in Customer-Service Operations

DGX agent

arXiv:2607.01426v1 Announce Type: new Abstract: Autonomous customer-service agents are shifting from conversational interfaces toward operational execution roles: they retrieve firm records, apply ser

safetyarxiv-cs-ai
3 Jul 2026
Safety

Wind-Aware Reinforcement Learning Control of a Small Quadrotor Using Learned Onboard Wind Estimation in Simulated Atmospheric Turbulence

DGX agent

arXiv:2607.01528v1 Announce Type: new Abstract: Small multirotor aircraft are increasingly tasked with operations in the atmospheric boundary layer, where turbulent winds comparable to the vehicle's a

safetyarxiv-cs-lg
3 Jul 2026
Safety

WorldSample: Closed-loop Real-robot RL with World Modelling

DGX agent

arXiv:2607.02431v1 Announce Type: cross Abstract: Reinforcement learning (RL) can overcome the demonstration-coverage limitation of imitation learning (IL) by allowing robots to improve through trial-

safetyarxiv-cs-ai
3 Jul 2026
Safety

Wow, even I was surprised how high AI ranked! Very good to see!

DGX agent

Wow, even I was surprised how high AI ranked! Very good to see! Interesting poll of Hill staffers from @PunchbowlNews. 250 years is a long time! But interesting to see that 'losing control of AI' is t

safetyconnor-leahy--x
3 Jul 2026
Safety

YuFeng-XGuard: A Reasoning-Centric, Interpretable, and Flexible Guardrail Model for Large Language Models

DGX agent

arXiv:2601.15588v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in real-world applications, safety guardrails are required to go beyond coarse-grained fil

safetyarxiv-cs-cl
3 Jul 2026
Safety

A Category Theory Account of AI Identity

DGX agent

arXiv:2607.00220v1 Announce Type: cross Abstract: Artificial intelligence (AI) systems are routinely modified after deployment through retraining and changes in their environments. These transformatio

safetyarxiv-cs-ai
2 Jul 2026
Safety

A Filtered Mixture-of-Generators for Fully Synthetic Survival Training

DGX agent

arXiv:2607.00127v1 Announce Type: new Abstract: Survival analysis models time-to-event data, but in clinical settings training data are costly and scarce: events accrue over years of follow-up, cohort

safetyarxiv-cs-lg
2 Jul 2026
Safety

A Mechanism-Driven Theory of Phase Transitions in Active Learning

DGX agent

arXiv:2607.00144v1 Announce Type: cross Abstract: Active learning (AL) performance is known to be budget-dependent, yet regimes are typically defined by heuristic label counts that fail to generalize

safetyarxiv-cs-ai
2 Jul 2026
Safety

A Multi-Resolution Finite-Volume Inspired Deep Learning Framework for Spatiotemporal Dynamics Prediction

DGX agent

arXiv:2607.00460v1 Announce Type: cross Abstract: Predicting complex spatiotemporal dynamics in physical processes often demands computationally expensive numerical methods or data-driven neural netwo

safetyarxiv-cs-ai
2 Jul 2026
Safety

A small tax on every token produced could be transformative, without putting the government into bed with a specific company. And because ev…

DGX agent

A small tax on every token produced could be transformative, without putting the government into bed with a specific company. And because every token draws on uncompensated contributions from multiple

safetygary-marcus--x
2 Jul 2026
Safety

Active-GRPO: Adaptive Imitation and Self-Improving Reasoning for Molecular Optimization

DGX agent

arXiv:2607.00531v1 Announce Type: cross Abstract: Scientific reasoning is an increasingly important capability of large language models, yet improving the robustness and efficiency of training such re

safetyarxiv-cs-ai
2 Jul 2026
Safety

Active Spatial Guidance: Eliminating Injected Positional Mechanisms in Vision Transformers

DGX agent

arXiv:2607.00580v1 Announce Type: new Abstract: Vision Transformers (ViTs) commonly rely on injected positional mechanisms to address self-attention's permutation invariance. Motivated by the spatial

safetyarxiv-cs-cv
2 Jul 2026
Safety

AI Native Games: A Survey and Roadmap

DGX agent

arXiv:2607.00527v1 Announce Type: new Abstract: Generative AI now enables games to produce dialogue, quests, characters, images, and worlds at runtime. Yet generation alone does not make a game AI-nat

safetyarxiv-cs-ai
2 Jul 2026
Safety

Aligning Sentence Embeddings to Human Concepts via Sparse Autoencoders

DGX agent

arXiv:2607.00023v1 Announce Type: cross Abstract: Dense sentence embeddings are fundamental to modern Retrieval-Augmented Generation (RAG) systems but suffer from a lack of interpretability due to fea

safetyarxiv-cs-ai
2 Jul 2026
← Previous
1…5960616263…265
Next →