AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Self-Supervised Relevance Modelling in Autonomous Driving via Counterfactual Analysis

DGX agent

arXiv:2606.10688v1 Announce Type: new Abstract: Autonomous driving relies on computationally intensive perception pipelines to continuously detect and track objects in the surrounding environment. Whi

safetyarxiv-cs-ro
10 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration

DGX agent

arXiv:2606.10228v1 Announce Type: cross Abstract: Safe exploration is a prerequisite for deploying reinforcement learning (RL) agents in safety-critical domains. In this paper, we approach safe explor

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

SinkRec: Mitigating Semantic State Sink in Long Sequence Recommendation with Memory-Conditioned Gated Delta Networks

DGX agent

arXiv:2606.09888v1 Announce Type: new Abstract: Linear attention provides an efficient backbone for long-sequence recommendation by avoiding the quadratic cost of standard Transformers, but its compre

safetyarxiv-cs-lg
10 Jun 2026
Safety

Sketch-to-Layout: A Human-Centric Computational Agent for Constraint-Aware Synthesis of Modular Photobioreactors

DGX agent

arXiv:2606.09849v1 Announce Type: cross Abstract: Building-integrated photobioreactors (PBRs) offer a pathway for carbon-neutral architecture, yet deployment is hindered by configuration complexity an

safetyarxiv-cs-cv
10 Jun 2026
Safety

SocraticPO: Policy Optimization via Interactive Guidance

DGX agent

arXiv:2606.09887v1 Announce Type: cross Abstract: Reinforcement learning (RL) for large language models usually supervises reasoning with scalar outcome rewards, such as binary correctness. Such rewar

safetyarxiv-cs-ai
10 Jun 2026
Safety

SoK: Colluding Adversaries in Machine Learning Pipelines

DGX agent

arXiv:2606.10091v1 Announce Type: cross Abstract: Machine learning (ML) models are susceptible to various security, privacy, and fairness risks. Adversaries with different characteristics (i.e., objec

safetyarxiv-cs-lg
10 Jun 2026
Safety

Speaker Group Encoding in Self-supervised Speech Recognition Models

DGX agent

arXiv:2606.10654v1 Announce Type: new Abstract: We investigate what self-supervised speech recognition models (S3Ms) learn about speaker groups (SGs). We examine several states of S3Ms: pretrained, fi

safetyarxiv-cs-cl
10 Jun 2026
Safety

Standard Language Ideology in AI-Generated Language

DGX agent

arXiv:2406.08726v3 Announce Type: replace Abstract: Large language models (LLMs) generate text that reinforces standard language ideology: a bias towards certain language varieties that are granted mo

safetyarxiv-cs-cl
10 Jun 2026
Safety

STEDiff: Strengthening Text Embedding for Text-to-Image Alignment in Diffusion Model

DGX agent

arXiv:2606.10653v1 Announce Type: new Abstract: Although pretrained text-to-image (T2I) generation models can produce high-quality images, they often fail to faithfully reflect the semantic intent of

safetyarxiv-cs-cv
10 Jun 2026
Safety

Structure-Preserving Learning Improves Geometry Generalization in Neural PDEs

DGX agent

arXiv:2602.02788v2 Announce Type: replace-cross Abstract: We aim to develop physics foundation models for science and engineering that provide real-time solutions to Partial Differential Equations (PD

safetyarxiv-cs-ai
10 Jun 2026
Safety

Support sufficiency as action-sufficient compression: a single-cycle rate-regret formulation

DGX agent

arXiv:2606.09858v1 Announce Type: cross Abstract: Robust decision-making requires compression. A system that forms a rich support state cannot usually preserve its full structure at the point of actio

safetyarxiv-cs-ai
10 Jun 2026
Safety

Synthesizable Molecular Generation via Soft-constrained GFlowNets with Rich Chemical Priors

DGX agent

arXiv:2602.04119v2 Announce Type: replace Abstract: The application of generative models for experimental drug discovery campaigns is severely limited by the difficulty of designing molecules de novo

safetyarxiv-cs-lg
10 Jun 2026
Safety

Task Robustness via Re-Labelling Vision-Action Robot Data

DGX agent

arXiv:2606.10918v1 Announce Type: cross Abstract: The recent trend in scaling models for robot learning has resulted in impressive policies that can perform various manipulation tasks and generalize t

safetyarxiv-cs-lg
10 Jun 2026
Safety

TD-Grokking: Learning from Zero-Reward Problems by Training-Time Decomposition

DGX agent

arXiv:2606.09883v1 Announce Type: cross Abstract: Large language models (LLMs) have made remarkable progress in reasoning tasks, largely driven by post-training paradigms, especially reinforcement lea

safetyarxiv-cs-ai
10 Jun 2026
Safety

Test-time Adversarial Takeover: A Real-time Hijacking Interface against Robotic Diffusion Policies

DGX agent

arXiv:2606.10371v1 Announce Type: cross Abstract: Diffusion-based action generation has become a foundational component of embodied AI, but its reliance on visual conditioning leaves deployed visuomot

safetyarxiv-cs-ai
10 Jun 2026
Safety

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning

DGX agent

arXiv:2606.11087v1 Announce Type: cross Abstract: Expressive continuous control policies, such as diffusion and flow models, form the backbone of recent advances in scaling imitation learning for simu

safetyarxiv-cs-ai
10 Jun 2026
Safety

The Role of Feedback Alignment in Self-Distillation

DGX agent

arXiv:2606.11173v1 Announce Type: new Abstract: Conditioning a language model on additional context, such as feedback on a previous attempt, typically improves its response. Self-distillation trains t

safetyarxiv-cs-ai
10 Jun 2026
Safety

The Whale That Outswam Evolution: Swarm Intelligence Maximises Memory in Connectome Reservoirs

DGX agent

arXiv:2606.09902v1 Announce Type: cross Abstract: Reservoir computing exploits the fixed dynamics of a recurrent network for temporal processing, requiring only a trained linear readout. Biological ne

safetyarxiv-cs-ai
10 Jun 2026
Safety

Toward Calibrated, Fair, and accurate Deepfake Detection

DGX agent

arXiv:2606.09881v1 Announce Type: cross Abstract: Deepfake detectors show large performance gaps across demographic groups. Existing fairness approaches require demographic labels, retraining, or sacr

safetyarxiv-cs-cv
10 Jun 2026
Safety

TRACE: A Unified Rollout Budget Allocation Framework for Efficient Agentic Reinforcement Learning

DGX agent

arXiv:2606.11119v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a promising approach for enhancing reasoning and agentic behavior in large language models. H

safetyarxiv-cs-ai
10 Jun 2026
Safety

Trading Utility for Dynamic Fairness in Multiple Resource Division with Sequential Demand

DGX agent

arXiv:2606.10472v1 Announce Type: cross Abstract: Dynamic multi-resource allocation is a central problem in shared computing environments, where users' demands arrive sequentially and resources must b

safetyarxiv-cs-lg
10 Jun 2026
Safety

UniPET: a universal network for high-quality PET image denoising across varied dose reduction factors

DGX agent

arXiv:2606.11131v1 Announce Type: new Abstract: Most existing deep learning-based PET image denoising methods assume a fixed and known dose reduction factor (DRF) for low-dose PET images. However, the

safetyarxiv-cs-cv
10 Jun 2026
Safety

Using Probabilistic Programs to Train Inductive Reasoning in Large Language Models

DGX agent

arXiv:2606.09856v1 Announce Type: cross Abstract: Post-training Large Language Models (LLMs) for reasoning typically focuses on deductive tasks such as mathematics and coding where correctness is veri

safetyarxiv-cs-ai
10 Jun 2026
Safety

Visual-TCAV: Concept-based Attribution and Saliency Maps for Post-hoc Explainability in Image Classification

DGX agent

arXiv:2411.05698v3 Announce Type: replace-cross Abstract: Convolutional Neural Networks (CNNs) have shown remarkable performance in image classification. However, interpreting their predictions is cha

safetyarxiv-cs-ai
10 Jun 2026
Safety

What Should a Skill Remember? Quality--Cost Trade-offs in Cost-Aware Skill Rewriting for Language Model Agents

DGX agent

arXiv:2606.09421v2 Announce Type: replace Abstract: Large language model agents increasingly rely on skills: reusable procedural documents encoding workflows, tool use, implementation patterns, valida

safetyarxiv-cs-cl
10 Jun 2026
Safety

When Distance Distracts: Representation Distance Bias in BT-Loss for Reward Models

DGX agent

arXiv:2512.06343v3 Announce Type: replace-cross Abstract: Reward models are central to Large Language Model (LLM) alignment within the framework of RLHF. The standard objective used in reward modeling

safetyarxiv-cs-ai
10 Jun 2026
Safety

When to Align, When to Predict: A Phase Diagram for Multimodal Learning

DGX agent

arXiv:2606.11190v1 Announce Type: new Abstract: Cross-modal alignment (CA) and cross-modal prediction (CP) are the dominant paradigms for multimodal representation learning, yet there is no systematic

safetyarxiv-cs-lg
10 Jun 2026
Safety

YUBI: Yielding Universal Bidigital Interface for Bimanual Dexterous Manipulation at Scale

DGX agent

arXiv:2606.10244v1 Announce Type: cross Abstract: We introduce Yielding Universal Bidigital Interface (YUBI), a finger-aligned gripper designed to enable intuitive, ergonomic, and scalable data collec

safetyarxiv-cs-ai
10 Jun 2026
Safety

A Finetuned SpeechLLM for Joint Multi-Granular L2 Assessment and Natural-Language Rationales

DGX agent

arXiv:2606.09470v1 Announce Type: cross Abstract: Automated L2 speech assessment can assign proficiency labels, but often lacks interpretability. We propose a rubric-guided SpeechLLM for multi-aspect,

safetyarxiv-cs-ai
9 Jun 2026
Safety

A Geometric Unification of Concept Learning with Concept Cones

DGX agent

arXiv:2512.07355v2 Announce Type: replace Abstract: Two traditions of interpretability have evolved side by side but seldom spoken to each other: Concept Bottleneck Models (CBMs), which prescribe what

safetyarxiv-cs-ai
9 Jun 2026
Safety

A Joint Finite-Sample Certificate for Adaptive Selective Conformal Risk Control

DGX agent

arXiv:2606.08517v1 Announce Type: new Abstract: Selective predictors answer on confident inputs and abstain elsewhere; deploying one safely needs a single finite-sample certificate that simultaneously

safetyarxiv-cs-lg
9 Jun 2026
Safety

A Mixed Diet Makes DINO An Omnivorous Vision Encoder

DGX agent

arXiv:2602.24181v2 Announce Type: replace-cross Abstract: Pre-trained vision encoders like DINOv2 have demonstrated exceptional performance on unimodal tasks. However, we observe that their features a

safetyarxiv-cs-ai
9 Jun 2026
Safety

A Unifying Lens on Reward Uncertainty in RLHF

DGX agent

arXiv:2606.09073v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) is bottlenecked by reward hacking, where the policy exploits errors in a proxy reward model (RM) and

safetyarxiv-cs-ai
9 Jun 2026
Safety

Ablation-Reversible Heads Don't Transfer: A Stress Test for Mechanistic Role Claims in Transformers

DGX agent

arXiv:2606.08292v1 Announce Type: new Abstract: In mechanistic interpretability, attention heads are commonly elevated to role claims (e.g., 'this head represents addition') when they are necessary fo

safetyarxiv-cs-ai
9 Jun 2026
Safety

ActProbe: Action-Space Probe for Early Failure Detection of Generative Robot Policies

DGX agent

arXiv:2606.08508v1 Announce Type: cross Abstract: Generative robot policies fail unpredictably at deployment: they hesitate at critical moments, drift off-task, or commit to unrecoverable actions. Exi

safetyarxiv-cs-ai
9 Jun 2026
Safety

Adaptive Loss Balancing for Noise-Robust GRPO in Generative Recommendation

DGX agent

arXiv:2606.08480v1 Announce Type: cross Abstract: Reinforcement learning (RL) presents a promising avenue for enhancing generative recommendation beyond supervised imitation, leveraging reward signals

safetyarxiv-cs-ai
9 Jun 2026
Safety

Agent Economics: An Entropy-Controlled Pluralistic Alignment Framework for Preventing Artificial Hivemind in Autonomous Agents

DGX agent

arXiv:2606.09039v1 Announce Type: new Abstract: This study proposes the Behavioral Protocol Framework (BPF), an entropy-controlled pluralistic alignment framework designed to address two critical chal

safetyarxiv-cs-ai
9 Jun 2026
Safety

AgriGov: A Structured Multilingual Dataset Curation for Indian Government Schemes for Farmers

DGX agent

arXiv:2606.08272v1 Announce Type: cross Abstract: AgriGov is a curated, trilingual (English-Hindi-Marathi) dataset designed to address the scarcity of domain-grounded multilingual resources for agricu

safetyarxiv-cs-ai
9 Jun 2026
Safety

AHA-WAM:Asynchronous Horizon-Adaptive World-Action Modeling with Observation-Guided Context Routing

DGX agent

arXiv:2606.09811v1 Announce Type: cross Abstract: World-action models have emerged as a promising paradigm for robot manipulation, jointly modeling visual scene dynamics and actions to inject physical

safetyarxiv-cs-ai
9 Jun 2026
Safety

AI Code Sandboxes: A Comparative Security Study. Part 1 of 2 -- Engine-Level Properties (Attack Surface, Leakage, Stackability, CVE History, Patch Cadence, Fuzzing)

DGX agent

arXiv:2606.08433v1 Announce Type: cross Abstract: This paper reads six engine-level measurements together -- 1.1 host attack surface, 1.2 information leakage, 1.3 defense-in-depth stackability, 1.4 pu

safetyarxiv-cs-ai
9 Jun 2026
Safety

AI-Integrated Learning Management System for Middle School: A Longitudinal Study of Learning Outcomes Through High School and Beyond

DGX agent

arXiv:2606.07544v1 Announce Type: cross Abstract: Middle school is a key window for building core academic skills and the learning routines students carry into later grades, yet many students still fa

safetyarxiv-cs-ai
9 Jun 2026
Safety

Aligned but Not Partner-Specific: Distinguishing How Multimodal LLM Agents Succeed in Reference Games Without Human-Like Conventions

DGX agent

arXiv:2606.08081v1 Announce Type: cross Abstract: Repeated reference games test whether interlocutors replace their initially long descriptions with shorter, partner-specific conventions grounded in s

safetyarxiv-cs-ai
9 Jun 2026
Safety

AMix-1: A Pathway to Test-Time Scalable Protein Foundation Model

DGX agent

arXiv:2507.08920v4 Announce Type: replace-cross Abstract: We introduce AMix-1, a powerful protein foundation model built on Bayesian Flow Networks and empowered by a systematic training methodology, e

safetyarxiv-cs-ai
9 Jun 2026
Safety

An Agency-Transferring Model-Free Policy Enhancement Technique

DGX agent

arXiv:2606.09825v1 Announce Type: cross Abstract: Training reinforcement learning (RL) policies from scratch is costly: it requires careful reward and environment design, extensive tuning, and substan

safetyarxiv-cs-ai
9 Jun 2026
Safety

Anchor-Conditioned Compositional Control for Landscape Image Generation

DGX agent

arXiv:2606.07638v1 Announce Type: cross Abstract: Image generative models, though widely used as creative tools, offer limited support for the kind of compositional control that photographers and visu

safetyarxiv-cs-ai
9 Jun 2026
Safety

Autonomous Aerial Manipulation via Contextual Contrastive Meta Reinforcement Learning

DGX agent

arXiv:2606.08533v1 Announce Type: new Abstract: Unmanned aerial vehicles (UAVs) are increasingly being deployed in logistics, service robotics, and other real-world applications, creating a growing de

safetyarxiv-cs-lg
9 Jun 2026
Safety

Autonomous FPV Flight with Translational Optical Flow and Uncertainty Mask

DGX agent

arXiv:2606.09088v1 Announce Type: new Abstract: Autonomous FPV quadrotor flight in complex environments using a monocular RGB camera as the sole exteroceptive sensor remains a fundamental challenge. R

safetyarxiv-cs-ro
9 Jun 2026
Safety

Autonomous Obstacle Removal for Excavators through Policy Learning with Particle Simulation

DGX agent

arXiv:2606.09183v1 Announce Type: new Abstract: Autonomous obstacle removal from the ground is an important earthwork task, but this is difficult to automate because an excavator must adapt its excava

safetyarxiv-cs-ro
9 Jun 2026
← Previous
1…126127128129130…260
Next →