AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning

DGX agent

arXiv:2607.02137v1 Announce Type: cross Abstract: We study timestep allocation for score-based diffusion sampling, where a learned reverse-time dynamics is discretized on a finite grid. Uniform and ha

safetyarxiv-cs-ai
3 Jul 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

AthDGC: An Open Diachronic Greek Treebank with Indo-European Parallels

DGX agent

arXiv:2606.15510v2 Announce Type: replace Abstract: AthDGC ('Athens-PROIEL') is an open, end-to-end workflow and dataset. It is, to the best of our knowledge, the first openly licensed dependency-pars

safetyarxiv-cs-cl
3 Jul 2026
Safety

Beyond Detection: Redesigning Assessment and Governande of Generative AI at the Universidad Politecnica de Madrid (UPM)

DGX agent

arXiv:2607.01255v1 Announce Type: cross Abstract: Universities have responded to generative artificial intelligence (GenAI) in noticeably different ways, both internationally and within Spain. So far,

safetyarxiv-cs-ai
3 Jul 2026
Safety

Beyond Next-Token Prediction: An RLVR Proof of Concept for Tool-Use Agents on Atlassian Workflows

DGX agent

arXiv:2607.01465v1 Announce Type: new Abstract: Large language models are trained to predict the next token, not to act inside a specific API. In niche enterprise SaaS workflows -- where success means

safetyarxiv-cs-ai
3 Jul 2026
Safety

BIFROST: Bridging Invariant Feature Representation for Observation-space Sim2Real Transfer

DGX agent

arXiv:2607.01410v1 Announce Type: cross Abstract: Sim2real transfer for robot policy learning suffers due to mismatch between simulation and reality. Existing methods typically address each gap in iso

safetyarxiv-cs-lg
3 Jul 2026
Safety

CALM: Interpretable Cross-Modal Alignment for Biomarker Discovery from Unpaired Data

DGX agent

arXiv:2607.01656v1 Announce Type: new Abstract: The interaction between brain structure and genetic influences is key to understanding neuropsychiatric disorders. However, most large-scale datasets ar

safetyarxiv-cs-lg
3 Jul 2026
Safety

CaP-X: A Framework for Benchmarking and Improving Coding Agents for Robot Manipulation

DGX agent

arXiv:2603.22435v2 Announce Type: replace-cross Abstract: 'Code-as-Policy' considers how executable code can complement data-intensive Vision-Language-Action (VLA) methods, yet their effectiveness as

safetyarxiv-cs-ai
3 Jul 2026
Safety

Conditional Inference Trees and Forests for Feature Selection

DGX agent

arXiv:2607.01417v1 Announce Type: new Abstract: Conditional inference trees (CIT) and conditional inference forests (CIF) reduce split-selection bias by testing features before choosing split threshol

safetyarxiv-cs-lg
3 Jul 2026
Safety

CoRe: Combined Rewards with Vision-Language Model Feedback for Preference-Aligned Reinforcement Learning

DGX agent

arXiv:2607.01721v1 Announce Type: new Abstract: Reward design remains a central challenge in reinforcement learning (RL). Hand-crafted rewards are often difficult to specify and may lead to suboptimal

safetyarxiv-cs-ro
3 Jul 2026
Safety

Criticality-Based Guard Rail Validation for AI Agent Decisions in Autonomous Telecom Networks

DGX agent

arXiv:2607.02210v1 Announce Type: new Abstract: The evolution toward fully autonomous telecommunications networks (Autonomous Network Levels 4-5) requires AI/ML agents to make real-time network decisi

safetyarxiv-cs-ai
3 Jul 2026
Safety

Cross-Platform Control for Autonomous Surface Vehicles via Adaptive Reinforcement Learning

DGX agent

arXiv:2607.02037v1 Announce Type: cross Abstract: Autonomous surface vehicles vary widely in hydrodynamic and actuation characteristics, yet most controllers are designed for single-platform deploymen

safetyarxiv-cs-lg
3 Jul 2026
Safety

DemoPSD: Disagreement-Modulated Policy Self-Distillation

DGX agent

arXiv:2607.02502v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) has emerged as a practical method for training large language models (LLMs) to reason, where a single model acts as

safetyarxiv-cs-ai
3 Jul 2026
Safety

DiPS: Dialogue Policy Selection for High-Stakes Persuasion Agents

DGX agent

arXiv:2607.01557v1 Announce Type: cross Abstract: Large Language Models (LLMs) often struggle with persuasion in high-stakes scenarios. People's individual personalities and concerns require tailored

safetyarxiv-cs-ai
3 Jul 2026
Safety

Don't Let Gains FADE: Breaking Down Policy Gradient Weights in RL

DGX agent

arXiv:2607.01490v1 Announce Type: cross Abstract: Reinforcement learning post-training dramatically improves LLM reasoning, but suffers from training instability and diversity collapse. Advantage func

safetyarxiv-cs-ai
3 Jul 2026
Safety

DRL-CLBA: A Clean Label Backdoor Attack for Speech Classification via DDPG Reinforcement Learning

DGX agent

arXiv:2607.01729v1 Announce Type: new Abstract: Deep learning models for speech classification are vulnerable to backdoor attacks, where malicious triggers cause misclassification at inference time. W

safetyarxiv-cs-ai
3 Jul 2026
Safety

Efficient Waste Sorting for Circular Economy: A Confidence-guided comparison between One-Vs-All and One-Vs-Rest Classification Strategies with Human-in-the-Loop for Automated Waste Sorting

DGX agent

arXiv:2607.02230v1 Announce Type: cross Abstract: The complexity of waste disposal regulations across European countries poses significant challenges for the residents and hinders the transition to a

safetyarxiv-cs-ai
3 Jul 2026
Safety

ElephantAgent: Contextual State Continuity in Agentic Systems

DGX agent

arXiv:2607.01919v1 Announce Type: new Abstract: Agentic systems enhance their capabilities by invoking external tools and maintaining persistent memory. However, these external dependencies introduce

safetyarxiv-cs-ai
3 Jul 2026
Safety

Episodic-to-Semantic Consolidation Without Identity Drift

DGX agent

arXiv:2607.01988v1 Announce Type: new Abstract: Long-running adaptive intelligent agents face a structural tension between knowledge consolidation and information integrity. Memory consolidation is co

safetyarxiv-cs-ai
3 Jul 2026
Safety

Exploring Large Language Models for Access Control Policy Synthesis and Summarization

DGX agent

arXiv:2510.20692v2 Announce Type: replace-cross Abstract: Cloud computing is ubiquitous, with a growing number of services being hosted on the cloud every day. Typical cloud compute systems allow admi

safetyarxiv-cs-ai
3 Jul 2026
Safety

From Approximation to Emergence: A Theory of Deep Learning

DGX agent

arXiv:2607.01311v1 Announce Type: new Abstract: Deep learning has outgrown any single mathematical explanation. From Approximation to Emergence develops a unified, proof-oriented account of modern dee

safetyarxiv-cs-lg
3 Jul 2026
Safety

Full Bayesian Reinforcement Learning via LF-IBIS

DGX agent

arXiv:2607.01741v1 Announce Type: cross Abstract: Reinforcement Learning (RL) is a sequential decision-making framework in which an agent learns optimal policies through interaction with an environmen

safetyarxiv-cs-ai
3 Jul 2026
Safety

Generalization in offline RL: The structure is more important than the amount of pessimism

DGX agent

arXiv:2607.02288v1 Announce Type: cross Abstract: While pessimism counteracts overestimation bias in offline reinforcement learning (RL), being overly conservative has been associated with hindering c

safetyarxiv-cs-ai
3 Jul 2026
Safety

GF-DiT: Scheduling Parallelism for Diffusion Transformer Serving

DGX agent

arXiv:2606.13501v2 Announce Type: replace-cross Abstract: Diffusion Transformers (DiTs) have become the dominant architecture for image and video generation, creating growing demand for efficient DiT

safetyarxiv-cs-lg
3 Jul 2026
Safety

Guided Action Flow: Q-Guided Inference for Flow-Matching Vision-Language-Action Policies

DGX agent

arXiv:2607.02092v1 Announce Type: cross Abstract: Flow-matching vision-language-action policies generate robot action chunks through an iterative transport process, creating an opportunity for test-ti

safetyarxiv-cs-ai
3 Jul 2026
Safety

HAL: Inducing Human-likeness in LLMs with Alignment

DGX agent

arXiv:2601.02813v3 Announce Type: replace Abstract: Aligning language models to qualitative behavioral traits, such as human-likeness, remains difficult because they are hard to define, measure, and o

safetyarxiv-cs-ai
3 Jul 2026
Safety

Kara: Efficient Reasoning LLM Serving via Sliding-Window KV Cache Compression

DGX agent

arXiv:2607.01237v1 Announce Type: cross Abstract: Reasoning language models often generate long chain-of-thought (CoT), which accumulates a massive KV cache during the decoding phase and incurs high d

safetyarxiv-cs-ai
3 Jul 2026
Safety

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution

DGX agent

arXiv:2508.05941v2 Announce Type: replace Abstract: Visuomotor policies trained via behavior cloning are vulnerable to covariate shift, where small deviations from expert trajectories can compound int

safetyarxiv-cs-ro
3 Jul 2026
Safety

Learning Agile Intruder Interception using Differentiable Quadrotor Dynamics

DGX agent

arXiv:2607.02472v1 Announce Type: new Abstract: This paper presents a methodology for learning a control policy to intercept an intruder using the 3D direction unit vector to the intruder and the inte

safetyarxiv-cs-ro
3 Jul 2026
Safety

Learning-based Multi-agent Race Strategies in Formula 1

DGX agent

arXiv:2602.23056v2 Announce Type: replace Abstract: In Formula 1, race strategies are adapted according to evolving race conditions and competitors' actions. This paper proposes a reinforcement learni

safetyarxiv-cs-ai
3 Jul 2026
Safety

Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation

DGX agent

arXiv:2512.18368v2 Announce Type: replace Abstract: Scaling imitation learning to diverse multi-task robot manipulation remains challenging due to suboptimal demonstrations, behavioral multi-modality,

safetyarxiv-cs-ro
3 Jul 2026
Safety

MAGIK: Mapping to Analogous Goals via Imagination-enabled Knowledge Transfer

DGX agent

arXiv:2506.01623v4 Announce Type: replace Abstract: Humans excel at analogical reasoning - applying knowledge from one task to a related one with minimal relearning. In contrast, reinforcement learnin

safetyarxiv-cs-ai
3 Jul 2026
Safety

Mean Field Reinforcement Learning

DGX agent

arXiv:2607.01525v1 Announce Type: cross Abstract: This monograph provides an introduction to mean field reinforcement learning through the lens of Markov decision processes arising from large-populati

safetyarxiv-cs-lg
3 Jul 2026
Safety

Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity

DGX agent

arXiv:2602.03315v2 Announce Type: replace Abstract: Agent memory systems must accommodate continuously growing information while supporting efficient, context-aware retrieval for downstream tasks. Abs

safetyarxiv-cs-ai
3 Jul 2026
Safety

MetaTune: Adjoint-based Meta-tuning via Robotic Differentiable Dynamics

DGX agent

arXiv:2603.27313v2 Announce Type: replace Abstract: Disturbance observer-based control has shown promise in robustifying robotic systems against uncertainties. However, tuning such systems remains cha

safetyarxiv-cs-ro
3 Jul 2026
Safety

Mirror Illusion Art

DGX agent

arXiv:2607.02015v1 Announce Type: cross Abstract: Mirror Illusion Art is a novel reflection-conditioned 3D illusion where one object yields two target appearances (front and mirror). The task is formu

safetyarxiv-cs-ai
3 Jul 2026
Safety

MolSight: A Graph-Aware Vision-Language Model for Unified Chemical Image Understanding

DGX agent

arXiv:2607.01982v1 Announce Type: cross Abstract: Using molecular large language models (LLMs) as a unified framework for understanding molecular structures and functions is emerging as a new trend in

safetyarxiv-cs-ai
3 Jul 2026
Safety

Morphology-Aware Sample Assignment: Overcoming IoU Insensitivity for Surface Defect Detection

DGX agent

arXiv:2606.13723v2 Announce Type: replace-cross Abstract: Intersection-over-Union (IoU), as a pivotal metric for evaluating the spatial alignment between candidate proposals and ground-truth annotatio

safetyarxiv-cs-ai
3 Jul 2026
Safety

Multi-Objective Exploration and Preference Optimization via Mutual Information

DGX agent

arXiv:2607.01392v1 Announce Type: new Abstract: Aligning large language models with diverse and heterogeneous human values requires multi-objective alignment methods to effectively trade off conflicti

safetyarxiv-cs-cl
3 Jul 2026
Safety

Multilayer Q-Matrix-Embedded Neural Network for Cognitive Diagnosis (M-QCDNet): Structure-Aware Deep Learning Architecture for Psychometric Interpretability

DGX agent

arXiv:2607.01278v1 Announce Type: new Abstract: The research proposes a multilayer Q-matrix-embedded neural network for cognitive diagnosis (M-QCDNet), which integrates the structural interpretability

safetyarxiv-cs-lg
3 Jul 2026
Safety

Navigating the Alignment-Calibration Trade-off: A Pareto-Superior Frontier via Model Merging

DGX agent

arXiv:2510.17426v3 Announce Type: replace-cross Abstract: The 'alignment tax' of post-training is typically framed as a drop in task accuracy. We show it also involves a severe loss of calibration, ma

safetyarxiv-cs-ai
3 Jul 2026
Safety

NeoMap: Training-free Novel-View Synthesis from Single Images and Videos

DGX agent

arXiv:2607.01962v1 Announce Type: cross Abstract: We study the challenging problem of novel view video synthesis from single images or monocular videos. Existing methods, which operate under the assum

safetyarxiv-cs-ai
3 Jul 2026
Safety

Neuron-Aware Data Selection for Annotation-Free LLM Self-Distillation

DGX agent

arXiv:2607.02460v1 Announce Type: cross Abstract: Post-training large language models (LLMs) without real-world interaction feedback or human-labeled supervision remains challenging, particularly in s

safetyarxiv-cs-ai
3 Jul 2026
Safety

Object Aligner: A Configurable JSON Schema Similarity Score for Graphs, Applied to LLM Prompt Optimization

DGX agent

arXiv:2607.01972v1 Announce Type: cross Abstract: Large language models (LLMs) are often asked to produce JSON conforming to a fixed schema, powering information extraction, tool calling, agentic plan

safetyarxiv-cs-ai
3 Jul 2026
Safety

On the Role of Computation in Reinforcement Learning

DGX agent

arXiv:2602.05999v4 Announce Type: replace Abstract: How does the amount of compute available to a reinforcement learning (RL) policy affect its learning? Can policies using a fixed amount of parameter

safetyarxiv-cs-lg
3 Jul 2026
Safety

On the Sample Efficiency of Inverse Dynamics Models for Semi-Supervised Imitation Learning

DGX agent

arXiv:2602.02762v2 Announce Type: replace Abstract: Semi-supervised imitation learning (SSIL) consists in learning a policy from a small dataset of action-labeled trajectories and a much larger datase

safetyarxiv-cs-lg
3 Jul 2026
Safety

Online Resource Allocation with Continuous Random Consumption: Regret under Degeneracy

DGX agent

arXiv:2607.02196v1 Announce Type: new Abstract: We study online resource allocation when both rewards and consumption sizes may be continuously distributed. Requests arrive sequentially and must be ac

safetyarxiv-cs-lg
3 Jul 2026
Safety

Optimizing Visual Generative Models via Distribution-wise Rewards

DGX agent

arXiv:2607.02291v1 Announce Type: new Abstract: Conventional reinforcement learning strategies for visual generation typically employ sample-wise reward functions, yet this practice frequently results

safetyarxiv-cs-lg
3 Jul 2026
Safety

Playing 20 Question Game with Policy-Based Reinforcement Learning

DGX agent

arXiv:1808.07645v5 Announce Type: replace-cross Abstract: The 20 Questions (Q20) game is a well known game which encourages deductive reasoning and creativity. In the game, the answerer first thinks o

safetyarxiv-cs-ai
3 Jul 2026
← Previous
1…100101102103104…260
Next →