AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

CALIBER: Calibrating Confidence Before and After Reasoning in Language Models

DGX agent

arXiv:2606.24281v1 Announce Type: cross Abstract: Reasoning language models are increasingly asked not only to answer difficult questions, but also to estimate their likelihood of success. Existing me

safetyarxiv-cs-ai
24 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cryptographic certificates of validity for trustworthy AI

DGX agent

arXiv:2606.23768v1 Announce Type: cross Abstract: We propose cryptographic certificates of validity for agentic AI systems. The core idea is to formally specify a correctness or policy condition as a

safetyarxiv-cs-ai
24 Jun 2026
Safety

Decentralized SGD with Controlled Disagreement Finds Flatter Minima

DGX agent

arXiv:2602.02899v2 Announce Type: replace Abstract: Decentralized training is often regarded as inferior to centralized training because the consensus errors between workers are thought to undermine c

safetyarxiv-cs-lg
24 Jun 2026
Safety

DREG: A Layer-Wise Jacobian Regularization as a General-Purpose Penalty

DGX agent

arXiv:2606.23942v1 Announce Type: new Abstract: We present a large-scale empirical study isolating the contributions of the Derivative Regularization penalty (DREG). Across a fully-crossed factorial s

safetyarxiv-cs-lg
24 Jun 2026
Safety

Dynamic Symmetric Point Tracking: Tackling Non-ideal Reference in Analog In-memory Training

DGX agent

arXiv:2602.21321v2 Announce Type: replace Abstract: Analog in-memory computing (AIMC) performs computation directly within resistive crossbar arrays, offering an energy-efficient platform to scale lar

safetyarxiv-cs-lg
24 Jun 2026
Safety

Enabling Robust Cloth Manipulation via Inference-Time Simulator-in-the-Loop Refinement

DGX agent

arXiv:2606.24552v1 Announce Type: new Abstract: Simulator-in-the-loop optimization offers a promising inference-time mechanism for robot manipulation. It uses a physical simulator as a backend rollout

safetyarxiv-cs-ro
24 Jun 2026
Safety

Enforcing Human-like Kinematics in Dexterous Piano Playing via Adversarial Posture Regularization

DGX agent

arXiv:2606.23848v1 Announce Type: new Abstract: Reinforcement learning can train bimanual dexterous hands to play piano in physics simulation with high note accuracy, but for high-DoF dexterous hands,

safetyarxiv-cs-ro
24 Jun 2026
Safety

Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning

DGX agent

arXiv:2606.24428v1 Announce Type: new Abstract: Experience-driven self-evolution is critical for large language model (LLM) agents to improve through open-world interaction. However, existing experien

safetyarxiv-cs-cl
24 Jun 2026
Safety

Evaluating the Interpretability of Sparse Autoencoders with Concept Annotations

DGX agent

arXiv:2606.24716v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) are increasingly used to extract interpretable concepts from vision and vision language models, yet existing evaluation met

safetyarxiv-cs-ai
24 Jun 2026
Safety

EvidenceLens: A Claim-Evidence Matrix for Auditing Financial Question Answering

DGX agent

arXiv:2606.23724v1 Announce Type: cross Abstract: Large language models are increasingly used to answer questions over annual reports, earnings decks, and analyst notes, yet their outputs remain diffi

safetyarxiv-cs-cl
24 Jun 2026
Safety

EXPO-SQL: Execution-based Clause-level Policy Optimization for Text-to-SQL

DGX agent

arXiv:2606.23693v1 Announce Type: new Abstract: Text-to-SQL enables users to query databases using natural language by generating executable SQL queries. Recent methods have increasingly adopted Large

safetyarxiv-cs-cl
24 Jun 2026
Safety

FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection

DGX agent

arXiv:2508.18684v2 Announce Type: replace-cross Abstract: Signature-based Intrusion Detection Systems (IDS) detect malicious activity by matching network or host events against predefined rules. Secur

safetyarxiv-cs-ai
24 Jun 2026
Safety

From 'Aha Moments' to Controllable Thinking: Toward Meta-Cognitive Reasoning in Large Reasoning Models via Decoupled Reasoning and Control

DGX agent

arXiv:2508.04460v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) can exhibit step-by-step reasoning, reflection, and backtracking, but these behaviors are often unregulated, leading t

safetyarxiv-cs-ai
24 Jun 2026
Safety

From Local Corrections to Generalized Skills: Improving Neuro-Symbolic Policies with MEMO

DGX agent

arXiv:2603.04560v2 Announce Type: replace Abstract: Recent works use a neuro-symbolic framework for general manipulation policies. The advantage of this framework is that -- by applying off-the-shelf

safetyarxiv-cs-ro
24 Jun 2026
Safety

G^3VLA: Geometric inductive bias for Vision-Language-Action Models

DGX agent

arXiv:2606.24472v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have made rapid progress in generalist robot manipulation by harnessing semantic knowledge from pretrained vision-

safetyarxiv-cs-ai
24 Jun 2026
Safety

GENA3D: Generative Amodal 3D Modeling by Bridging 2D Priors and 3D Coherence

DGX agent

arXiv:2511.21945v3 Announce Type: replace Abstract: Generating complete 3D objects under partial occlusions (i.e., amodal scenarios) is a practically important yet challenging problem, as large portio

safetyarxiv-cs-cv
24 Jun 2026
Safety

Generative Manifold Distillation: Aligning Restoration Trajectories with Natural Image Prior

DGX agent

arXiv:2512.11121v2 Announce Type: replace Abstract: Pre-trained image restoration models often fail on out-of-distribution (OOD) real-world degradations. Adapting to these domains is challenging as re

safetyarxiv-cs-cv
24 Jun 2026
Safety

Geometric Action Model for Robot Policy Learning

DGX agent

arXiv:2606.17046v2 Announce Type: replace-cross Abstract: Generalist robot policies must follow user instructions while reasoning about how objects, cameras, and robot actions interact in the 3D physi

safetyarxiv-cs-cv
24 Jun 2026
Safety

Governed Shared Memory for Multi-Agent LLM Systems

DGX agent

arXiv:2606.24535v1 Announce Type: new Abstract: Multi-agent LLM environments require robust mechanisms for shared knowledge management. This paper formalizes the fleet-memory problem and identifies fo

safetyarxiv-cs-ai
24 Jun 2026
Safety

Hierarchical Spatial and Channel Aggregation for Cross-domain Few-shot Segmentation

DGX agent

arXiv:2606.24296v1 Announce Type: new Abstract: Cross-domain Few-shot Segmentation (CD-FSS) aims to learn generalizable segmentation capability from abundant annotated samples in the source domain, en

safetyarxiv-cs-cv
24 Jun 2026
Safety

Hybrid Sequence Modeling and Reinforced Verification for Controllable Target-Conditioned Decision Making

DGX agent

arXiv:2508.16420v3 Announce Type: replace Abstract: Target-conditioned sequence models provide a simple interface for controllable offline decision making, but the requested target return can be an un

safetyarxiv-cs-lg
24 Jun 2026
Safety

JEDEL: Zero-Shot DNA-Encoded Library Design for Early-Stage Drug Discovery

DGX agent

arXiv:2606.23745v1 Announce Type: cross Abstract: We present JEDEL, a framework for generating synthesis-ready DNA-encoded libraries (DELs) directly from three-dimensional pharmacophore representation

safetyarxiv-cs-ai
24 Jun 2026
Safety

KLip-PPO: A per-sample KL perspective on PPO-Clip

DGX agent

arXiv:2606.23932v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) is the standard policy-gradient algorithm for on-policy reinforcement learning. The literature presents it in two for

safetyarxiv-cs-lg
24 Jun 2026
Safety

Latent Visual States for Efficient Multimodal Reasoning

DGX agent

arXiv:2606.24233v1 Announce Type: new Abstract: The integration of visual evidence has significantly enhanced the capabilities of large multimodal models. However, this integration predominantly relie

safetyarxiv-cs-cv
24 Jun 2026
Safety

LecturaAgents: A Multi-Agent Framework for Adaptive Personalized AI-Assisted Learning and Embodied Teaching

DGX agent

arXiv:2606.16428v2 Announce Type: replace-cross Abstract: Effective personalized AI-assisted learning demands systems that can not only generate accurate learner-specific educational materials, but al

safetyarxiv-cs-ai
24 Jun 2026
Safety

Less is More: Quality-Aware Training Data Selection for Scientific Summarization

DGX agent

arXiv:2606.24828v1 Announce Type: new Abstract: Scientific long-document summarization datasets commonly treat author-written abstracts as gold reference summaries, although their quality and alignmen

safetyarxiv-cs-cl
24 Jun 2026
Safety

Light-weight Pronunciation Assessment via Discrete Speech Token Surprisal

DGX agent

arXiv:2606.19910v2 Announce Type: replace Abstract: Training automated pronunciation assessment often relies on labeled learner errors or non-native corpora that are costly to collect. We propose a li

safetyarxiv-cs-cl
24 Jun 2026
Safety

Lightweight Test-Time Adaptation for EMG-Based Gesture Recognition

DGX agent

arXiv:2601.04181v2 Announce Type: replace Abstract: Reliable long-term decoding of gestures from surface electromyography (EMG) is hindered by signal drift caused by electrode displacement, muscle fat

safetyarxiv-cs-lg
24 Jun 2026
Safety

Listening makes Vision Clear for VLMs

DGX agent

arXiv:2606.23763v1 Announce Type: cross Abstract: Recent work typically assesses vision--language consistency using attention distributions of answer-side tokens. However, we observe that highest atte

safetyarxiv-cs-ai
24 Jun 2026
Safety

MSPL: Multi-Step Pseudo-Labeling for Open-Vocabulary Object Detection

DGX agent

arXiv:2510.14792v4 Announce Type: replace Abstract: Open-vocabulary object detection (OVD) aims to recognize and localize object categories beyond the training set. Recent approaches leverage vision-l

safetyarxiv-cs-cv
24 Jun 2026
Safety

Multi-agent imitation learning with function approximation: Linear Markov games and beyond

DGX agent

arXiv:2602.22810v2 Announce Type: replace Abstract: In this work, we present the first theoretical analysis of multi-agent imitation learning (MAIL) in linear Markov games where both the transition dy

safetyarxiv-cs-lg
24 Jun 2026
Safety

Offline Reinforcement Learning for Warehouse SLAM Throughput Control

DGX agent

arXiv:2606.23978v1 Announce Type: cross Abstract: We present an offline reinforcement learning (RL) framework for optimizing SLAM throughput control in a warehouse fulfillment environment. SLAM (Scan/

safetyarxiv-cs-ai
24 Jun 2026
Safety

On the Smallness of the Large Language Models Scaling Exponents

DGX agent

arXiv:2606.24504v1 Announce Type: new Abstract: We discuss reasons why the scaling exponents of current Large Language Models (LLMs) applications are indicating an unsustainable regime in terms of ene

safetyarxiv-cs-ai
24 Jun 2026
Safety

Policies Permitting LLM Use for Polishing Peer Reviews Are Currently Not Enforceable

DGX agent

arXiv:2603.20450v2 Announce Type: replace-cross Abstract: A number of scientific conferences and journals have recently enacted policies that prohibit LLM usage by peer reviewers, except for polishing

safetyarxiv-cs-ai
24 Jun 2026
Safety

Policy Gradient with Self-Attention for Model-Free Distributed Nonlinear Multi-Agent Games

DGX agent

arXiv:2509.18371v2 Announce Type: replace-cross Abstract: Multi-agent games in dynamic nonlinear settings are challenging due to the time-varying interactions among the agents and the non-stationarity

safetyarxiv-cs-ro
24 Jun 2026
Safety

Progressive Alignment Objectives for Aligner-Encoder based ASR

DGX agent

arXiv:2606.24147v1 Announce Type: cross Abstract: Aligner-Encoders are recently proposed seq2seq end-to-end ASR models that replace decoder attention by predicting the uth token directly from the u-th

safetyarxiv-cs-cl
24 Jun 2026
Safety

Real vs. Complex Spectral Bases for Neural Operators: The Role of Green's Function Alignment

DGX agent

arXiv:2606.24851v1 Announce Type: new Abstract: Fourier Neural Operators (FNO) learn solution operators of partial differential equations by parameterizing global convolutions in the complex Fourier d

safetyarxiv-cs-lg
24 Jun 2026
Model Releases

REALM: A Unified Red-Teaming Benchmark for Physical-World VLMs

DGX agent

arXiv:2606.23892v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used as perception-reasoning backbones for embodied intelligence in safety-critical physical systems, whe

model-releasesarxiv-cs-cv
24 Jun 2026
Safety

Red-Teaming the Agentic Red-Team

DGX agent

arXiv:2606.24496v1 Announce Type: cross Abstract: The use of agentic systems to perform offensive security operations has moved from a theoretical possibility to a commoditized capability. However, wh

safetyarxiv-cs-ai
24 Jun 2026
Safety

Reinforcement Learning for Computer-Use Agents with Autonomous Evaluation

DGX agent

arXiv:2606.24515v1 Announce Type: new Abstract: Computer-Use Agents (CUAs) execute high-level user goals by perceiving and acting directly within graphical user interfaces. However, reinforcement lear

safetyarxiv-cs-ai
24 Jun 2026
Safety

Reinforcement Learning Towards Broadly and Persistently Beneficial Models

DGX agent

arXiv:2606.24014v1 Announce Type: new Abstract: As AI systems are deployed across increasingly diverse and high-stakes settings, model alignment must generalize beyond the tasks and domains seen durin

safetyarxiv-cs-ai
24 Jun 2026
Safety

Rethinking Structural Anomaly Detection: From Decision Boundaries to Projection Operators

DGX agent

arXiv:2606.15280v2 Announce Type: replace Abstract: Most existing anomaly detection methods rely on estimating a probability density or learning an enclosing decision boundary, implicitly assuming tha

safetyarxiv-cs-lg
24 Jun 2026
Safety

RoBoSR: Structured Scene Representations for Embodied Robotic Reasoning

DGX agent

arXiv:2606.24338v1 Announce Type: new Abstract: Despite rapid progress, embodied reasoning under real-world variability remains challenging. Existing approaches rely on demonstration-driven sequential

safetyarxiv-cs-ro
24 Jun 2026
Safety

RTFF: Random-to-Target Fabric Flattening Policy using Dual-Arm Manipulator

DGX agent

arXiv:2510.00814v2 Announce Type: replace Abstract: Robotic fabric manipulation remains challenging due to fabric deformability and occlusions from wrinkles and the manipulator. This paper defines Ran

safetyarxiv-cs-ro
24 Jun 2026
Safety

SC3-Eval: Evaluating Robot Foundation Models via Self-Consistent Video Generation

DGX agent

arXiv:2606.18610v2 Announce Type: replace-cross Abstract: Evaluating generalist robot manipulation policies in the real world is expensive, slow, and difficult to scale. Action-conditioned video world

safetyarxiv-cs-cv
24 Jun 2026
Safety

ScaleToT: Generalizing Structured LLM Reasoning for Billion-Scale Low-Activity User Modeling

DGX agent

arXiv:2606.24605v1 Announce Type: new Abstract: Accurate user modeling often depends on rich interaction histories, which are unavailable for billions of low-activity users. Large Language Models (LLM

safetyarxiv-cs-ai
24 Jun 2026
Safety

SEAL: Searching Expandable Architectures for Incremental Learning

DGX agent

arXiv:2505.10457v3 Announce Type: replace-cross Abstract: Incremental learning is a machine learning paradigm where a model learns from a sequential stream of tasks. This setting poses a key challenge

safetyarxiv-cs-ai
24 Jun 2026
Safety

Similarity of Neural Network Representations in Superposition

DGX agent

arXiv:2604.00208v2 Announce Type: replace Abstract: Comparing internal representations is a central goal in neuroscience and machine learning, but standard linear alignment metrics (Representational S

safetyarxiv-cs-lg
24 Jun 2026
← Previous
1…116117118119120…260
Next →