AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
Safety

DRL-CLBA: A Clean Label Backdoor Attack for Speech Classification via DDPG Reinforcement Learning

DGX agent

arXiv:2607.01729v1 Announce Type: new Abstract: Deep learning models for speech classification are vulnerable to backdoor attacks, where malicious triggers cause misclassification at inference time. W

safetyarxiv-cs-ai
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Efficient Waste Sorting for Circular Economy: A Confidence-guided comparison between One-Vs-All and One-Vs-Rest Classification Strategies with Human-in-the-Loop for Automated Waste Sorting

DGX agent

arXiv:2607.02230v1 Announce Type: cross Abstract: The complexity of waste disposal regulations across European countries poses significant challenges for the residents and hinders the transition to a

safetyarxiv-cs-ai
3 Jul 2026
Safety

ElephantAgent: Contextual State Continuity in Agentic Systems

DGX agent

arXiv:2607.01919v1 Announce Type: new Abstract: Agentic systems enhance their capabilities by invoking external tools and maintaining persistent memory. However, these external dependencies introduce

safetyarxiv-cs-ai
3 Jul 2026
Safety

Episodic-to-Semantic Consolidation Without Identity Drift

DGX agent

arXiv:2607.01988v1 Announce Type: new Abstract: Long-running adaptive intelligent agents face a structural tension between knowledge consolidation and information integrity. Memory consolidation is co

safetyarxiv-cs-ai
3 Jul 2026
Safety

Exploring Large Language Models for Access Control Policy Synthesis and Summarization

DGX agent

arXiv:2510.20692v2 Announce Type: replace-cross Abstract: Cloud computing is ubiquitous, with a growing number of services being hosted on the cloud every day. Typical cloud compute systems allow admi

safetyarxiv-cs-ai
3 Jul 2026
Safety

From Approximation to Emergence: A Theory of Deep Learning

DGX agent

arXiv:2607.01311v1 Announce Type: new Abstract: Deep learning has outgrown any single mathematical explanation. From Approximation to Emergence develops a unified, proof-oriented account of modern dee

safetyarxiv-cs-lg
3 Jul 2026
Safety

Full Bayesian Reinforcement Learning via LF-IBIS

DGX agent

arXiv:2607.01741v1 Announce Type: cross Abstract: Reinforcement Learning (RL) is a sequential decision-making framework in which an agent learns optimal policies through interaction with an environmen

safetyarxiv-cs-ai
3 Jul 2026
Safety

Generalization in offline RL: The structure is more important than the amount of pessimism

DGX agent

arXiv:2607.02288v1 Announce Type: cross Abstract: While pessimism counteracts overestimation bias in offline reinforcement learning (RL), being overly conservative has been associated with hindering c

safetyarxiv-cs-ai
3 Jul 2026
Safety

GF-DiT: Scheduling Parallelism for Diffusion Transformer Serving

DGX agent

arXiv:2606.13501v2 Announce Type: replace-cross Abstract: Diffusion Transformers (DiTs) have become the dominant architecture for image and video generation, creating growing demand for efficient DiT

safetyarxiv-cs-lg
3 Jul 2026
Safety

Guided Action Flow: Q-Guided Inference for Flow-Matching Vision-Language-Action Policies

DGX agent

arXiv:2607.02092v1 Announce Type: cross Abstract: Flow-matching vision-language-action policies generate robot action chunks through an iterative transport process, creating an opportunity for test-ti

safetyarxiv-cs-ai
3 Jul 2026
Safety

HAL: Inducing Human-likeness in LLMs with Alignment

DGX agent

arXiv:2601.02813v3 Announce Type: replace Abstract: Aligning language models to qualitative behavioral traits, such as human-likeness, remains difficult because they are hard to define, measure, and o

safetyarxiv-cs-ai
3 Jul 2026
Safety

Kara: Efficient Reasoning LLM Serving via Sliding-Window KV Cache Compression

DGX agent

arXiv:2607.01237v1 Announce Type: cross Abstract: Reasoning language models often generate long chain-of-thought (CoT), which accumulates a massive KV cache during the decoding phase and incurs high d

safetyarxiv-cs-ai
3 Jul 2026
Safety

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution

DGX agent

arXiv:2508.05941v2 Announce Type: replace Abstract: Visuomotor policies trained via behavior cloning are vulnerable to covariate shift, where small deviations from expert trajectories can compound int

safetyarxiv-cs-ro
3 Jul 2026
Safety

Learning Agile Intruder Interception using Differentiable Quadrotor Dynamics

DGX agent

arXiv:2607.02472v1 Announce Type: new Abstract: This paper presents a methodology for learning a control policy to intercept an intruder using the 3D direction unit vector to the intruder and the inte

safetyarxiv-cs-ro
3 Jul 2026
Safety

Learning-based Multi-agent Race Strategies in Formula 1

DGX agent

arXiv:2602.23056v2 Announce Type: replace Abstract: In Formula 1, race strategies are adapted according to evolving race conditions and competitors' actions. This paper proposes a reinforcement learni

safetyarxiv-cs-ai
3 Jul 2026
Safety

Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation

DGX agent

arXiv:2512.18368v2 Announce Type: replace Abstract: Scaling imitation learning to diverse multi-task robot manipulation remains challenging due to suboptimal demonstrations, behavioral multi-modality,

safetyarxiv-cs-ro
3 Jul 2026
Safety

MAGIK: Mapping to Analogous Goals via Imagination-enabled Knowledge Transfer

DGX agent

arXiv:2506.01623v4 Announce Type: replace Abstract: Humans excel at analogical reasoning - applying knowledge from one task to a related one with minimal relearning. In contrast, reinforcement learnin

safetyarxiv-cs-ai
3 Jul 2026
Safety

Mean Field Reinforcement Learning

DGX agent

arXiv:2607.01525v1 Announce Type: cross Abstract: This monograph provides an introduction to mean field reinforcement learning through the lens of Markov decision processes arising from large-populati

safetyarxiv-cs-lg
3 Jul 2026
Safety

Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity

DGX agent

arXiv:2602.03315v2 Announce Type: replace Abstract: Agent memory systems must accommodate continuously growing information while supporting efficient, context-aware retrieval for downstream tasks. Abs

safetyarxiv-cs-ai
3 Jul 2026
Safety

MetaTune: Adjoint-based Meta-tuning via Robotic Differentiable Dynamics

DGX agent

arXiv:2603.27313v2 Announce Type: replace Abstract: Disturbance observer-based control has shown promise in robustifying robotic systems against uncertainties. However, tuning such systems remains cha

safetyarxiv-cs-ro
3 Jul 2026
Safety

Mirror Illusion Art

DGX agent

arXiv:2607.02015v1 Announce Type: cross Abstract: Mirror Illusion Art is a novel reflection-conditioned 3D illusion where one object yields two target appearances (front and mirror). The task is formu

safetyarxiv-cs-ai
3 Jul 2026
Safety

MolSight: A Graph-Aware Vision-Language Model for Unified Chemical Image Understanding

DGX agent

arXiv:2607.01982v1 Announce Type: cross Abstract: Using molecular large language models (LLMs) as a unified framework for understanding molecular structures and functions is emerging as a new trend in

safetyarxiv-cs-ai
3 Jul 2026
Safety

Morphology-Aware Sample Assignment: Overcoming IoU Insensitivity for Surface Defect Detection

DGX agent

arXiv:2606.13723v2 Announce Type: replace-cross Abstract: Intersection-over-Union (IoU), as a pivotal metric for evaluating the spatial alignment between candidate proposals and ground-truth annotatio

safetyarxiv-cs-ai
3 Jul 2026
Safety

Multi-Objective Exploration and Preference Optimization via Mutual Information

DGX agent

arXiv:2607.01392v1 Announce Type: new Abstract: Aligning large language models with diverse and heterogeneous human values requires multi-objective alignment methods to effectively trade off conflicti

safetyarxiv-cs-cl
3 Jul 2026
Safety

Multilayer Q-Matrix-Embedded Neural Network for Cognitive Diagnosis (M-QCDNet): Structure-Aware Deep Learning Architecture for Psychometric Interpretability

DGX agent

arXiv:2607.01278v1 Announce Type: new Abstract: The research proposes a multilayer Q-matrix-embedded neural network for cognitive diagnosis (M-QCDNet), which integrates the structural interpretability

safetyarxiv-cs-lg
3 Jul 2026
Safety

Navigating the Alignment-Calibration Trade-off: A Pareto-Superior Frontier via Model Merging

DGX agent

arXiv:2510.17426v3 Announce Type: replace-cross Abstract: The 'alignment tax' of post-training is typically framed as a drop in task accuracy. We show it also involves a severe loss of calibration, ma

safetyarxiv-cs-ai
3 Jul 2026
Safety

NeoMap: Training-free Novel-View Synthesis from Single Images and Videos

DGX agent

arXiv:2607.01962v1 Announce Type: cross Abstract: We study the challenging problem of novel view video synthesis from single images or monocular videos. Existing methods, which operate under the assum

safetyarxiv-cs-ai
3 Jul 2026
Safety

Neuron-Aware Data Selection for Annotation-Free LLM Self-Distillation

DGX agent

arXiv:2607.02460v1 Announce Type: cross Abstract: Post-training large language models (LLMs) without real-world interaction feedback or human-labeled supervision remains challenging, particularly in s

safetyarxiv-cs-ai
3 Jul 2026
Safety

Object Aligner: A Configurable JSON Schema Similarity Score for Graphs, Applied to LLM Prompt Optimization

DGX agent

arXiv:2607.01972v1 Announce Type: cross Abstract: Large language models (LLMs) are often asked to produce JSON conforming to a fixed schema, powering information extraction, tool calling, agentic plan

safetyarxiv-cs-ai
3 Jul 2026
Safety

On the Role of Computation in Reinforcement Learning

DGX agent

arXiv:2602.05999v4 Announce Type: replace Abstract: How does the amount of compute available to a reinforcement learning (RL) policy affect its learning? Can policies using a fixed amount of parameter

safetyarxiv-cs-lg
3 Jul 2026
Safety

On the Sample Efficiency of Inverse Dynamics Models for Semi-Supervised Imitation Learning

DGX agent

arXiv:2602.02762v2 Announce Type: replace Abstract: Semi-supervised imitation learning (SSIL) consists in learning a policy from a small dataset of action-labeled trajectories and a much larger datase

safetyarxiv-cs-lg
3 Jul 2026
Safety

Online Resource Allocation with Continuous Random Consumption: Regret under Degeneracy

DGX agent

arXiv:2607.02196v1 Announce Type: new Abstract: We study online resource allocation when both rewards and consumption sizes may be continuously distributed. Requests arrive sequentially and must be ac

safetyarxiv-cs-lg
3 Jul 2026
Safety

OpenAI offers feds a stake, Anthropic gets out of AI model jail and Meta wants to be a neocloud

DGX agent

OpenAI reportedly has floated giving the U.S. government a 5% stake in the company, perhaps the start of a series of such stakes in other AI companies as well. This no doubt has traditional anti-indus

safetysiliconangle
3 Jul 2026
Safety

Optimizing Visual Generative Models via Distribution-wise Rewards

DGX agent

arXiv:2607.02291v1 Announce Type: new Abstract: Conventional reinforcement learning strategies for visual generation typically employ sample-wise reward functions, yet this practice frequently results

safetyarxiv-cs-lg
3 Jul 2026
Safety

Playing 20 Question Game with Policy-Based Reinforcement Learning

DGX agent

arXiv:1808.07645v5 Announce Type: replace-cross Abstract: The 20 Questions (Q20) game is a well known game which encourages deductive reasoning and creativity. In the game, the answerer first thinks o

safetyarxiv-cs-ai
3 Jul 2026
Safety

Predicting Closed-Loop Performance of Latent World Models: Offline Checkpoint Selection for MPC and Model-Based RL Under Non-Markovian Rewards in LunarLander

DGX agent

arXiv:2607.01736v1 Announce Type: cross Abstract: We study how to predict the downstream closed-loop performance of a learned latent world model from validation-time diagnostics alone. Choosing the ri

safetyarxiv-cs-ai
3 Jul 2026
Safety

Prediction Sets for Counterfactual Decisions: Coverage, Optimality, and Conformal Prediction

DGX agent

arXiv:2607.02206v1 Announce Type: cross Abstract: Predictions are increasingly used to guide high-stakes decisions, from treatment selection to policy making. To ensure reliability with imperfect pred

safetyarxiv-cs-lg
3 Jul 2026
Safety

Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

DGX agent

arXiv:2607.02234v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) has emerged as a promising paradigm for improving LLM reasoning, where a privileged teacher with access to reference

safetyarxiv-cs-ai
3 Jul 2026
Safety

Quantifying the Uncertainty of Blindly Estimated Room Embeddings Using a Dispersion-Calibrated Score

DGX agent

arXiv:2607.01527v1 Announce Type: cross Abstract: Room embeddings derived from reverberant speech are often unreliable: speech content and recording degradation can alter the representation even when

safetyarxiv-cs-lg
3 Jul 2026
Safety

Quantum-Inspired Vision: Leveraging Wave-Particle Duality for Low-Illumination Enhancement

DGX agent

arXiv:2607.01731v1 Announce Type: cross Abstract: This study provides a theoretical expansion of the recent Data Relativistic Uncertainty (DRU) framework by formalizing a physics-to-AI paradigm for im

safetyarxiv-cs-lg
3 Jul 2026
Safety

Rank-Then-Act: Reward-Free Control from Frame-Order Progress

DGX agent

arXiv:2607.01897v1 Announce Type: cross Abstract: We introduce Rank-Then-Act (RTA), a framework for learning control policies from expert video demonstrations without environment rewards. RTA trains a

safetyarxiv-cs-ai
3 Jul 2026
Safety

RedCoder: Automated Multi-Turn Red Teaming for Code LLMs

DGX agent

arXiv:2507.22063v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) for code generation (i.e., Code LLMs) have demonstrated impressive capabilities in AI-assisted software developme

safetyarxiv-cs-ai
3 Jul 2026
Safety

Risk Architecture for AI-Native Engineering Teams: An Organizational Framework for Agentic System Governance

DGX agent

arXiv:2607.01421v1 Announce Type: cross Abstract: Engineering management research has produced mature frameworks for software risk: ownership by feature, escalation by severity, and assurance by test

safetyarxiv-cs-ai
3 Jul 2026
Safety

SABER: A Semantic-Aligned Brain Network Analysis Framework via Multi-scale Hypergraphs

DGX agent

arXiv:2607.01901v1 Announce Type: cross Abstract: Effective brain disease diagnosis requires the synergy of brain connectivity patterns and high-level semantic knowledge. Existing methods, however, la

safetyarxiv-cs-ai
3 Jul 2026
Safety

Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model

DGX agent

arXiv:2607.01595v1 Announce Type: new Abstract: As the scale and complexity of cloud-based AI systems continue to escalate, ensuring service reliability through rapid fault detection and adaptive reco

safetyarxiv-cs-ai
3 Jul 2026
Safety

Safeguarding LLM Agents from Misalignment through Provenance Analysis

DGX agent

arXiv:2607.01236v1 Announce Type: cross Abstract: As LLM agents gain increasing access to powerful tools, ensuring that their actions are aligned with the user's intent becomes critical. When an agent

safetyarxiv-cs-ai
3 Jul 2026
Safety

Sim2Real-AD: A Modular Sim-to-Real Framework for Deploying VLM-Guided Reinforcement Learning in Real-World Autonomous Driving

DGX agent

arXiv:2604.03497v2 Announce Type: replace-cross Abstract: Vision-language-model (VLM)-guided reinforcement learning (RL) has recently attracted significant attention for it, replacing brittle hand-cra

safetyarxiv-cs-ai
3 Jul 2026
Safety

SPLC: Social Preference Learning for Crowd Robot Navigation

DGX agent

arXiv:2607.01925v1 Announce Type: new Abstract: Offline reinforcement learning (RL) holds significant potential for crowd robot navigation in human-robot coexistence applications. However, the inheren

safetyarxiv-cs-ro
3 Jul 2026
← Previous
1…112113114115116…302
Next →