AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
Safety

BIFROST: Bridging Invariant Feature Representation for Observation-space Sim2Real Transfer

DGX agent

arXiv:2607.01410v1 Announce Type: cross Abstract: Sim2real transfer for robot policy learning suffers due to mismatch between simulation and reality. Existing methods typically address each gap in iso

safetyarxiv-cs-lg
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

CALM: Interpretable Cross-Modal Alignment for Biomarker Discovery from Unpaired Data

DGX agent

arXiv:2607.01656v1 Announce Type: new Abstract: The interaction between brain structure and genetic influences is key to understanding neuropsychiatric disorders. However, most large-scale datasets ar

safetyarxiv-cs-lg
3 Jul 2026
Safety

CaP-X: A Framework for Benchmarking and Improving Coding Agents for Robot Manipulation

DGX agent

arXiv:2603.22435v2 Announce Type: replace-cross Abstract: 'Code-as-Policy' considers how executable code can complement data-intensive Vision-Language-Action (VLA) methods, yet their effectiveness as

safetyarxiv-cs-ai
3 Jul 2026
Safety

Conditional Inference Trees and Forests for Feature Selection

DGX agent

arXiv:2607.01417v1 Announce Type: new Abstract: Conditional inference trees (CIT) and conditional inference forests (CIF) reduce split-selection bias by testing features before choosing split threshol

safetyarxiv-cs-lg
3 Jul 2026
Safety

Copewell: A Multi-Agent Swarm Architecture for Equitable Mental Wellness Support

DGX agent

arXiv:2607.02245v1 Announce Type: new Abstract: Mental health disorders affect nearly one billion people globally, yet 75% of individuals in low- and middle-income countries receive no treatment due t

safetyarxiv-cs-ai
3 Jul 2026
Safety

CoRe: Combined Rewards with Vision-Language Model Feedback for Preference-Aligned Reinforcement Learning

DGX agent

arXiv:2607.01721v1 Announce Type: new Abstract: Reward design remains a central challenge in reinforcement learning (RL). Hand-crafted rewards are often difficult to specify and may lead to suboptimal

safetyarxiv-cs-ro
3 Jul 2026
Safety

Criticality-Based Guard Rail Validation for AI Agent Decisions in Autonomous Telecom Networks

DGX agent

arXiv:2607.02210v1 Announce Type: new Abstract: The evolution toward fully autonomous telecommunications networks (Autonomous Network Levels 4-5) requires AI/ML agents to make real-time network decisi

safetyarxiv-cs-ai
3 Jul 2026
Safety

Cross-Platform Control for Autonomous Surface Vehicles via Adaptive Reinforcement Learning

DGX agent

arXiv:2607.02037v1 Announce Type: cross Abstract: Autonomous surface vehicles vary widely in hydrodynamic and actuation characteristics, yet most controllers are designed for single-platform deploymen

safetyarxiv-cs-lg
3 Jul 2026
Safety

DemoPSD: Disagreement-Modulated Policy Self-Distillation

DGX agent

arXiv:2607.02502v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) has emerged as a practical method for training large language models (LLMs) to reason, where a single model acts as

safetyarxiv-cs-ai
3 Jul 2026
Safety

DiPS: Dialogue Policy Selection for High-Stakes Persuasion Agents

DGX agent

arXiv:2607.01557v1 Announce Type: cross Abstract: Large Language Models (LLMs) often struggle with persuasion in high-stakes scenarios. People's individual personalities and concerns require tailored

safetyarxiv-cs-ai
3 Jul 2026
Safety

Don't Let Gains FADE: Breaking Down Policy Gradient Weights in RL

DGX agent

arXiv:2607.01490v1 Announce Type: cross Abstract: Reinforcement learning post-training dramatically improves LLM reasoning, but suffers from training instability and diversity collapse. Advantage func

safetyarxiv-cs-ai
3 Jul 2026
Safety

DriveVLM-RL: Neuroscience-Inspired Reinforcement Learning with Vision-Language Models for Safe and Deployable Autonomous Driving

DGX agent

arXiv:2603.18315v2 Announce Type: replace-cross Abstract: Traditional reinforcement learning (RL) methods rely on manually engineered rewards or sparse collision signals, which fail to capture the ric

safetyarxiv-cs-ai
3 Jul 2026
Safety

DRL-CLBA: A Clean Label Backdoor Attack for Speech Classification via DDPG Reinforcement Learning

DGX agent

arXiv:2607.01729v1 Announce Type: new Abstract: Deep learning models for speech classification are vulnerable to backdoor attacks, where malicious triggers cause misclassification at inference time. W

safetyarxiv-cs-ai
3 Jul 2026
Safety

Efficient Waste Sorting for Circular Economy: A Confidence-guided comparison between One-Vs-All and One-Vs-Rest Classification Strategies with Human-in-the-Loop for Automated Waste Sorting

DGX agent

arXiv:2607.02230v1 Announce Type: cross Abstract: The complexity of waste disposal regulations across European countries poses significant challenges for the residents and hinders the transition to a

safetyarxiv-cs-ai
3 Jul 2026
Safety

ElephantAgent: Contextual State Continuity in Agentic Systems

DGX agent

arXiv:2607.01919v1 Announce Type: new Abstract: Agentic systems enhance their capabilities by invoking external tools and maintaining persistent memory. However, these external dependencies introduce

safetyarxiv-cs-ai
3 Jul 2026
Safety

Episodic-to-Semantic Consolidation Without Identity Drift

DGX agent

arXiv:2607.01988v1 Announce Type: new Abstract: Long-running adaptive intelligent agents face a structural tension between knowledge consolidation and information integrity. Memory consolidation is co

safetyarxiv-cs-ai
3 Jul 2026
Safety

Epistemic Goggles: A Pretrained Module that Induces an Epistemic Frame via Gradient Editing

DGX agent

arXiv:2607.01690v1 Announce Type: new Abstract: Finetuning a language model on documents that are explicitly annotated as fictional results in a model that still actually believes the documents' core

safetyarxiv-cs-ai
3 Jul 2026
Safety

ESC: Emotional Self-Correction for Reliable Vision-Language Models

DGX agent

arXiv:2607.02089v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance across diverse multimodal tasks, yet they remain vulnerable to unreliable reasoning. Ex

safetyarxiv-cs-ai
3 Jul 2026
Safety

Exploring Large Language Models for Access Control Policy Synthesis and Summarization

DGX agent

arXiv:2510.20692v2 Announce Type: replace-cross Abstract: Cloud computing is ubiquitous, with a growing number of services being hosted on the cloud every day. Typical cloud compute systems allow admi

safetyarxiv-cs-ai
3 Jul 2026
Safety

Fast and Accurate Anomaly Detection in Time Series

DGX agent

arXiv:2607.02046v1 Announce Type: new Abstract: Anomaly detection is a critical and evolving field in Machine Learning, with applications targeting different domains such as cybersecurity, finance, he

safetyarxiv-cs-lg
3 Jul 2026
Safety

From Approximation to Emergence: A Theory of Deep Learning

DGX agent

arXiv:2607.01311v1 Announce Type: new Abstract: Deep learning has outgrown any single mathematical explanation. From Approximation to Emergence develops a unified, proof-oriented account of modern dee

safetyarxiv-cs-lg
3 Jul 2026
Safety

Full Bayesian Reinforcement Learning via LF-IBIS

DGX agent

arXiv:2607.01741v1 Announce Type: cross Abstract: Reinforcement Learning (RL) is a sequential decision-making framework in which an agent learns optimal policies through interaction with an environmen

safetyarxiv-cs-ai
3 Jul 2026
Safety

Generalization in offline RL: The structure is more important than the amount of pessimism

DGX agent

arXiv:2607.02288v1 Announce Type: cross Abstract: While pessimism counteracts overestimation bias in offline reinforcement learning (RL), being overly conservative has been associated with hindering c

safetyarxiv-cs-ai
3 Jul 2026
Safety

GF-DiT: Scheduling Parallelism for Diffusion Transformer Serving

DGX agent

arXiv:2606.13501v2 Announce Type: replace-cross Abstract: Diffusion Transformers (DiTs) have become the dominant architecture for image and video generation, creating growing demand for efficient DiT

safetyarxiv-cs-lg
3 Jul 2026
Safety

Guided Action Flow: Q-Guided Inference for Flow-Matching Vision-Language-Action Policies

DGX agent

arXiv:2607.02092v1 Announce Type: cross Abstract: Flow-matching vision-language-action policies generate robot action chunks through an iterative transport process, creating an opportunity for test-ti

safetyarxiv-cs-ai
3 Jul 2026
Safety

HAL: Inducing Human-likeness in LLMs with Alignment

DGX agent

arXiv:2601.02813v3 Announce Type: replace Abstract: Aligning language models to qualitative behavioral traits, such as human-likeness, remains difficult because they are hard to define, measure, and o

safetyarxiv-cs-ai
3 Jul 2026
Safety

Hardware-Enforced Semantic Coordination for Safety-Critical Real-Time Autonomous Systems

DGX agent

arXiv:2607.02376v1 Announce Type: new Abstract: Recent advances in agentic AI are producing increasingly complex autonomous systems that integrate large language models, world models, optimization eng

safetyarxiv-cs-ai
3 Jul 2026
Safety

Kara: Efficient Reasoning LLM Serving via Sliding-Window KV Cache Compression

DGX agent

arXiv:2607.01237v1 Announce Type: cross Abstract: Reasoning language models often generate long chain-of-thought (CoT), which accumulates a massive KV cache during the decoding phase and incurs high d

safetyarxiv-cs-ai
3 Jul 2026
Safety

kNNGuard: Turning LLM Hidden Activations into a Training-Free Configurable Guardrail

DGX agent

arXiv:2607.02072v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in domains requiring guardrails to detect unsafe, off-topic, or adversarial prompts. Existing g

safetyarxiv-cs-ai
3 Jul 2026
Safety

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution

DGX agent

arXiv:2508.05941v2 Announce Type: replace Abstract: Visuomotor policies trained via behavior cloning are vulnerable to covariate shift, where small deviations from expert trajectories can compound int

safetyarxiv-cs-ro
3 Jul 2026
Safety

Learning Agile Intruder Interception using Differentiable Quadrotor Dynamics

DGX agent

arXiv:2607.02472v1 Announce Type: new Abstract: This paper presents a methodology for learning a control policy to intercept an intruder using the 3D direction unit vector to the intruder and the inte

safetyarxiv-cs-ro
3 Jul 2026
Safety

Learning-based Multi-agent Race Strategies in Formula 1

DGX agent

arXiv:2602.23056v2 Announce Type: replace Abstract: In Formula 1, race strategies are adapted according to evolving race conditions and competitors' actions. This paper proposes a reinforcement learni

safetyarxiv-cs-ai
3 Jul 2026
Safety

Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation

DGX agent

arXiv:2512.18368v2 Announce Type: replace Abstract: Scaling imitation learning to diverse multi-task robot manipulation remains challenging due to suboptimal demonstrations, behavioral multi-modality,

safetyarxiv-cs-ro
3 Jul 2026
Safety

Lightweight Safe Reinforcement Learning for End-to-End UAV Navigation

DGX agent

arXiv:2607.01794v1 Announce Type: cross Abstract: With the rapid development of autonomous aerial systems, Unmanned Aerial Vehicles (UAVs) are increasingly deployed in applications such as inspection,

safetyarxiv-cs-ai
3 Jul 2026
Safety

MAGIK: Mapping to Analogous Goals via Imagination-enabled Knowledge Transfer

DGX agent

arXiv:2506.01623v4 Announce Type: replace Abstract: Humans excel at analogical reasoning - applying knowledge from one task to a related one with minimal relearning. In contrast, reinforcement learnin

safetyarxiv-cs-ai
3 Jul 2026
Safety

Mean Field Reinforcement Learning

DGX agent

arXiv:2607.01525v1 Announce Type: cross Abstract: This monograph provides an introduction to mean field reinforcement learning through the lens of Markov decision processes arising from large-populati

safetyarxiv-cs-lg
3 Jul 2026
Safety

Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity

DGX agent

arXiv:2602.03315v2 Announce Type: replace Abstract: Agent memory systems must accommodate continuously growing information while supporting efficient, context-aware retrieval for downstream tasks. Abs

safetyarxiv-cs-ai
3 Jul 2026
Safety

MetaTune: Adjoint-based Meta-tuning via Robotic Differentiable Dynamics

DGX agent

arXiv:2603.27313v2 Announce Type: replace Abstract: Disturbance observer-based control has shown promise in robustifying robotic systems against uncertainties. However, tuning such systems remains cha

safetyarxiv-cs-ro
3 Jul 2026
Safety

Mirror Illusion Art

DGX agent

arXiv:2607.02015v1 Announce Type: cross Abstract: Mirror Illusion Art is a novel reflection-conditioned 3D illusion where one object yields two target appearances (front and mirror). The task is formu

safetyarxiv-cs-ai
3 Jul 2026
Safety

MolSight: A Graph-Aware Vision-Language Model for Unified Chemical Image Understanding

DGX agent

arXiv:2607.01982v1 Announce Type: cross Abstract: Using molecular large language models (LLMs) as a unified framework for understanding molecular structures and functions is emerging as a new trend in

safetyarxiv-cs-ai
3 Jul 2026
Safety

Morphology-Aware Sample Assignment: Overcoming IoU Insensitivity for Surface Defect Detection

DGX agent

arXiv:2606.13723v2 Announce Type: replace-cross Abstract: Intersection-over-Union (IoU), as a pivotal metric for evaluating the spatial alignment between candidate proposals and ground-truth annotatio

safetyarxiv-cs-ai
3 Jul 2026
Safety

Multi-modal Rail Crossing Safety Analysis

DGX agent

arXiv:2607.01365v1 Announce Type: cross Abstract: Given one or more images of a railway crossing, can we leverage visual cues that allow us to robustly estimate how safe it is? Can we improve our abil

safetyarxiv-cs-ai
3 Jul 2026
Safety

Multi-Objective Exploration and Preference Optimization via Mutual Information

DGX agent

arXiv:2607.01392v1 Announce Type: new Abstract: Aligning large language models with diverse and heterogeneous human values requires multi-objective alignment methods to effectively trade off conflicti

safetyarxiv-cs-cl
3 Jul 2026
Safety

Multilayer Q-Matrix-Embedded Neural Network for Cognitive Diagnosis (M-QCDNet): Structure-Aware Deep Learning Architecture for Psychometric Interpretability

DGX agent

arXiv:2607.01278v1 Announce Type: new Abstract: The research proposes a multilayer Q-matrix-embedded neural network for cognitive diagnosis (M-QCDNet), which integrates the structural interpretability

safetyarxiv-cs-lg
3 Jul 2026
Safety

Navigating the Alignment-Calibration Trade-off: A Pareto-Superior Frontier via Model Merging

DGX agent

arXiv:2510.17426v3 Announce Type: replace-cross Abstract: The 'alignment tax' of post-training is typically framed as a drop in task accuracy. We show it also involves a severe loss of calibration, ma

safetyarxiv-cs-ai
3 Jul 2026
Safety

NeoMap: Training-free Novel-View Synthesis from Single Images and Videos

DGX agent

arXiv:2607.01962v1 Announce Type: cross Abstract: We study the challenging problem of novel view video synthesis from single images or monocular videos. Existing methods, which operate under the assum

safetyarxiv-cs-ai
3 Jul 2026
Safety

Neuron-Aware Data Selection for Annotation-Free LLM Self-Distillation

DGX agent

arXiv:2607.02460v1 Announce Type: cross Abstract: Post-training large language models (LLMs) without real-world interaction feedback or human-labeled supervision remains challenging, particularly in s

safetyarxiv-cs-ai
3 Jul 2026
Safety

Object Aligner: A Configurable JSON Schema Similarity Score for Graphs, Applied to LLM Prompt Optimization

DGX agent

arXiv:2607.01972v1 Announce Type: cross Abstract: Large language models (LLMs) are often asked to produce JSON conforming to a fixed schema, powering information extraction, tool calling, agentic plan

safetyarxiv-cs-ai
3 Jul 2026
← Previous
1…5859606162…265
Next →