AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
Safety

IROSA: Interactive Robot Skill Adaptation using Natural Language

DGX agent

arXiv:2603.03897v3 Announce Type: replace-cross Abstract: Foundation models have demonstrated impressive capabilities across diverse domains, while imitation learning provides principled methods for r

safetyarxiv-cs-cl
17 Apr 2026
Safety

Language Model as Planner and Formalizer under Constraints

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2510.05486v2 Announce Type: replace Abstract: LLMs have been widely used in planning, either as planners to generate action sequences end-to-end, or as formalizers to represent the planning doma

safetyarxiv-cs-cl
17 Apr 2026
Safety

Language on Demand, Knowledge at Core: Composing LLMs with Encoder-Decoder Translation Models for Extensible Multilinguality

DGX agent

arXiv:2603.17512v4 Announce Type: replace Abstract: Large language models (LLMs) exhibit strong general intelligence, yet their multilingual performance remains highly imbalanced. Although LLMs encode

safetyarxiv-cs-cl
17 Apr 2026
Safety

Neuro-Symbolic AI for Cybersecurity: State of the Art, Challenges, and Opportunities

DGX agent

arXiv:2509.06921v2 Announce Type: replace-cross Abstract: Cybersecurity demands both rapid pattern recognition and deliberative reasoning, yet purely neural or purely symbolic approaches each address

safetyarxiv-cs-ai
17 Apr 2026
Safety

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework

DGX agent

arXiv:2604.15308v1 Announce Type: new Abstract: High-level autonomous driving requires motion planners capable of modeling multimodal future uncertainties while remaining robust in closed-loop interac

safetyarxiv-cs-cv
17 Apr 2026
Safety

Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models

DGX agent

arXiv:2604.14888v1 Announce Type: new Abstract: Recent advances in vision language models (VLMs) offer reasoning capabilities, yet how these unfold and integrate visual and textual information remains

safetyarxiv-cs-cl
17 Apr 2026
Safety

RECOVER: Designing a Large Language Model-based Remote Patient Monitoring System for Postoperative Gastrointestinal Cancer Care

DGX agent

arXiv:2502.05740v2 Announce Type: replace-cross Abstract: Cancer surgery is a key treatment for gastrointestinal (GI) cancers, a group of cancers that account for more than 35% of cancer-related death

safetyarxiv-cs-ai
17 Apr 2026
Safety

Switch: Learning Agile Skills Switching for Humanoid Robots

DGX agent

arXiv:2604.14834v1 Announce Type: new Abstract: Recent advancements in whole-body control through deep reinforcement learning have enabled humanoid robots to achieve remarkable progress in real-world

safetyarxiv-cs-ro
17 Apr 2026
Safety

Trajectory Planning for a Multi-UAV Rigid-Payload Cascaded Transportation System Based on Enhanced Tube-RRT*

DGX agent

arXiv:2604.15074v1 Announce Type: new Abstract: This paper presents a two-stage trajectory planning framework for a multi-UAV rigid-payload cascaded transportation system, aiming to address planning c

safetyarxiv-cs-ro
17 Apr 2026
Safety

TwinOR: Photorealistic Digital Twins of Dynamic Operating Rooms for Embodied AI Research

DGX agent

arXiv:2511.07412v2 Announce Type: replace Abstract: Developing embodied AI for intelligent surgical systems requires safe, controllable environments for continual learning and evaluation. However, saf

safetyarxiv-cs-cv
17 Apr 2026
Safety

Vision-Based Safe Human-Robot Collaboration with Uncertainty Guarantees

DGX agent

arXiv:2604.15221v1 Announce Type: cross Abstract: We propose a framework for vision-based human pose estimation and motion prediction that gives conformal prediction guarantees for certifiably safe hu

safetyarxiv-cs-cv
17 Apr 2026
Safety

A Bayesian Framework for Uncertainty-Aware Explanations in Power Quality Disturbance Classification

DGX agent

arXiv:2604.13658v1 Announce Type: new Abstract: Advanced deep learning methods have shown remarkable success in power quality disturbance (PQD) classification. To enhance model transparency, explainab

safetyarxiv-cs-lg
16 Apr 2026
Safety

Bias at the End of the Score

DGX agent

arXiv:2604.13305v1 Announce Type: new Abstract: Reward models (RMs) are inherently non-neutral value functions designed and trained to encode specific objectives, such as human preferences or text-ima

safetyarxiv-cs-cv
16 Apr 2026
Safety

ChartNet: A Million-Scale, High-Quality Multimodal Dataset for Robust Chart Understanding

DGX agent

arXiv:2603.27064v2 Announce Type: replace-cross Abstract: Understanding charts requires models to jointly reason over geometric visual patterns, structured numerical data, and natural language -- a ca

safetyarxiv-cs-cl
16 Apr 2026
Safety

Golden Handcuffs make safer AI agents

DGX agent

arXiv:2604.13609v1 Announce Type: new Abstract: Reinforcement learners can attain high reward through novel unintended strategies. We study a Bayesian mitigation for general environments: we expand th

safetyarxiv-cs-lg
16 Apr 2026
Safety

Hardware-Efficient Neuro-Symbolic Networks with the Exp-Minus-Log Operator

DGX agent

arXiv:2604.13871v1 Announce Type: new Abstract: Deep neural networks (DNNs) deliver state-of-the-art accuracy on regression and classification tasks, yet two structural deficits persistently obstruct

safetyarxiv-cs-lg
16 Apr 2026
Safety

Maybe because of this paper? https://x.com/emollick/status/1991624198855561508?s=20

DGX agent

Maybe because of this paper? https://x.com/emollick/status/1991624198855561508?s=20 Tell all the truth but tell it slant— Success in Circuit lies Too bright for our infirm Delight The Truth's superb s

safetyethan-mollick--x
16 Apr 2026
Safety

Multi-Dimensional Knowledge Profiling with Large-Scale Literature Database and Hierarchical Retrieval

DGX agent

arXiv:2601.15170v2 Announce Type: replace Abstract: The rapid expansion of research across machine learning, vision, and language has produced a volume of publications that is increasingly difficult t

safetyarxiv-cs-cv
16 Apr 2026
Safety

Rethinking Uncertainty in Segmentation: From Estimation to Decision

DGX agent

arXiv:2604.13262v1 Announce Type: new Abstract: In medical image segmentation, uncertainty estimates are often reported but rarely used to guide decisions. We study the missing step: how uncertainty m

safetyarxiv-cs-cv
16 Apr 2026
Safety

Self-adaptive Multi-Access Edge Architectures: A Robotics Case

DGX agent

arXiv:2604.13542v1 Announce Type: new Abstract: The growth of compute-intensive AI tasks highlights the need to mitigate the processing costs and improve performance and energy efficiency. This necess

safetyarxiv-cs-ro
16 Apr 2026
Safety

A longitudinal health agent framework

DGX agent

arXiv:2604.12019v1 Announce Type: new Abstract: Although artificial intelligence (AI) agents are increasingly proposed to support potentially longitudinal health tasks, such as symptom management, beh

safetyarxiv-cs-ai
15 Apr 2026
Safety

Active Imitation Learning for Thermal- and Kernel-Aware LFM Inference on 3D S-NUCA Many-Cores

DGX agent

arXiv:2604.11948v1 Announce Type: new Abstract: Large Foundation Model (LFM) inference is both memory- and compute-intensive, traditionally relying on GPUs. However, the limited availability and high

safetyarxiv-cs-lg
15 Apr 2026
Safety

Beyond Static Sandboxing: Learned Capability Governance for Autonomous AI Agents

DGX agent

arXiv:2604.11839v1 Announce Type: cross Abstract: Autonomous AI agents built on open-source runtimes such as OpenClaw expose every available tool to every session by default, regardless of the task. A

safetyarxiv-cs-ai
15 Apr 2026
Safety

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs

DGX agent

arXiv:2604.12616v1 Announce Type: new Abstract: The rapid evolution of Vision-Language Models (VLMs) has catalyzed unprecedented capabilities in artificial intelligence; however, this continuous modal

safetyarxiv-cs-ai
15 Apr 2026
Safety

How Transformers Learn to Plan via Multi-Token Prediction

DGX agent

arXiv:2604.11912v1 Announce Type: cross Abstract: While next-token prediction (NTP) has been the standard objective for training language models, it often struggles to capture global structure in reas

safetyarxiv-cs-ai
15 Apr 2026
Safety

Physics-Grounded Monocular Vehicle Distance Estimation Using Standardized License Plate Typography

DGX agent

arXiv:2604.12239v1 Announce Type: new Abstract: Accurate inter-vehicle distance estimation is a cornerstone of Advanced Driver Assistance Systems (ADAS) and autonomous driving. While LiDAR and radar p

safetyarxiv-cs-cv
15 Apr 2026
Safety

🇧🇪 Positive news for FSD Supervised in Belgium! I just received an official response from the cabinet of @MDiependaele , Minister-Presiden…

DGX agent

🇧🇪 Positive news for FSD Supervised in Belgium! I just received an official response from the cabinet of @MDiependaele , Minister-President of the Flemish Government. A few days ago, the Dutch vehicle

safetyelon-musk--x
15 Apr 2026
Safety

Probabilistic Feature Imputation and Uncertainty-Aware Multimodal Federated Aggregation

DGX agent

arXiv:2604.12970v1 Announce Type: cross Abstract: Multimodal federated learning enables privacy-preserving collaborative model training across healthcare institutions. However, a fundamental challenge

safetyarxiv-cs-cv
15 Apr 2026
Safety

Reasoning about Intent for Ambiguous Requests

DGX agent

arXiv:2511.10453v3 Announce Type: replace-cross Abstract: Large language models often respond to ambiguous requests by implicitly committing to one interpretation, frustrating users and creating safet

safetyarxiv-cs-ai
15 Apr 2026
Safety

Reliability-Guided Depth Fusion for Glare-Resilient Navigation Costmaps

DGX agent

arXiv:2604.12753v1 Announce Type: new Abstract: Specular glare on reflective floors and glass surfaces frequently corrupts RGB-D depth measurements, producing holes and spikes that accumulate as persi

safetyarxiv-cs-ro
15 Apr 2026
Safety

Ternary Logic Encodings of Temporal Behavior Trees with Application to Control Synthesis

DGX agent

arXiv:2604.12092v1 Announce Type: new Abstract: Behavior Trees (BTs) provide designers an intuitive graphical interface to construct long-horizon plans for autonomous systems. To ensure their correctn

safetyarxiv-cs-ro
15 Apr 2026
Safety

The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment

DGX agent

arXiv:2604.12116v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as tool-augmented agents capable of executing system-level operations. While existing benchmarks

safetyarxiv-cs-ai
15 Apr 2026
Safety

This is what a black politician in South Africa said… “We will kill white women, we will kill white children, and we will even kill your pet…

DGX agent

This is what a black politician in South Africa said… “We will kill white women, we will kill white children, and we will even kill your pets' Violent language entering politics is a serious warning s

safetyelon-musk--x
15 Apr 2026
Safety

Uncertainty-Aware Image Classification In Biomedical Imaging Using Spectral-normalized Neural Gaussian Processes

DGX agent

arXiv:2602.02370v2 Announce Type: replace Abstract: Accurate histopathologic interpretation is key for clinical decision-making; however, current deep learning models for digital pathology are often o

safetyarxiv-cs-cv
15 Apr 2026
Safety

AI Integrity: A New Paradigm for Verifiable AI Governance

DGX agent

arXiv:2604.11065v1 Announce Type: new Abstract: AI systems increasingly shape high-stakes decisions in healthcare, law, defense, and education, yet existing governance paradigms -- AI Ethics, AI Safet

safetyarxiv-cs-ai
14 Apr 2026
Safety

Awesome work by @jiaxinwen22, @liangqiu_1994, Joe Benton, and @janhkirchner! For more details, check out the blog post 👇 https://anthropic.…

DGX agent

Jan Leike praised collaborative work by researchers Jiaxin Wen, Liang Qiu, Joe Benton, and Jan Kirchner, directing followers to an Anthropic blog post for further details. The post appears to highligh

safetyjan-leike--x
14 Apr 2026
Safety

Backdoors in RLVR: Jailbreak Backdoors in LLMs From Verifiable Reward

DGX agent

arXiv:2604.09748v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an emerging paradigm that significantly boosts a Large Language Model's (LLM's) reasoning abi

safetyarxiv-cs-ai
14 Apr 2026
Safety

Belief-Aware VLM Model for Human-like Reasoning

DGX agent

arXiv:2604.09686v1 Announce Type: new Abstract: Traditional neural network models for intent inference rely heavily on observable states and struggle to generalize across diverse tasks and dynamic env

safetyarxiv-cs-ai
14 Apr 2026
Safety

CAGenMol: Condition-Aware Diffusion Language Model for Goal-Directed Molecular Generation

DGX agent

arXiv:2604.11483v1 Announce Type: new Abstract: Goal-directed molecular generation requires satisfying heterogeneous constraints such as protein--ligand compatibility and multi-objective drug-like pro

safetyarxiv-cs-lg
14 Apr 2026
Safety

CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation

DGX agent

arXiv:2604.09746v1 Announce Type: cross Abstract: As large language models (LLMs) are increasingly deployed as autonomous agents, understanding how strategic behavior emerges in multi-agent environmen

safetyarxiv-cs-ai
14 Apr 2026
Safety

Detection Is Cheap, Routing Is Learned: Why Refusal-Based Alignment Evaluation Fails

DGX agent

arXiv:2603.18280v2 Announce Type: replace-cross Abstract: Current alignment evaluation mostly measures whether models encode dangerous concepts and whether they refuse harmful requests. Both miss the

safetyarxiv-cs-ai
14 Apr 2026
Safety

Distributionally Robust PAC-Bayesian Control

DGX agent

arXiv:2604.10588v1 Announce Type: new Abstract: We present a distributionally robust PAC-Bayesian framework for certifying the performance of learning-based finite-horizon controllers. While existing

safetyarxiv-cs-lg
14 Apr 2026
Safety

Enabling and Inhibitory Pathways of Students' AI Use Concealment Intention in Higher Education: Evidence from SEM and fsQCA

DGX agent

arXiv:2604.10978v1 Announce Type: cross Abstract: This study investigates students' AI use concealment intention in higher education by integrating the cognition-affect-conation (CAC) framework with a

safetyarxiv-cs-ai
14 Apr 2026
Safety

Explainability and Certification of AI-Generated Educational Assessments

DGX agent

arXiv:2604.09622v1 Announce Type: cross Abstract: The rapid adoption of generative artificial intelligence (AI) in educational assessment has created new opportunities for scalable item creation, pers

safetyarxiv-cs-ai
14 Apr 2026
Safety

Explainable Planning for Hybrid Systems

DGX agent

arXiv:2604.09578v1 Announce Type: new Abstract: The recent advancement in artificial intelligence (AI) technologies facilitates a paradigm shift toward automation. Autonomous systems are fully or part

safetyarxiv-cs-ai
14 Apr 2026
Safety

FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning

DGX agent

arXiv:2604.10693v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has improved LLM reasoning, but models often generate explanations that appear coherent while containing unfaithful int

safetyarxiv-cs-ai
14 Apr 2026
Safety

Fake-HR1: Rethinking Reasoning of Vision Language Model for Synthetic Image Detection

DGX agent

arXiv:2602.10042v3 Announce Type: replace-cross Abstract: Recent studies have demonstrated that incorporating Chain-of-Thought (CoT) reasoning into the detection process can enhance a model's ability

safetyarxiv-cs-ai
14 Apr 2026
Safety

From Answers to Arguments: Toward Trustworthy Clinical Diagnostic Reasoning with Toulmin-Guided Curriculum Goal-Conditioned Learning

DGX agent

arXiv:2604.11137v1 Announce Type: new Abstract: The integration of Large Language Models (LLMs) into clinical decision support is critically obstructed by their opaque and often unreliable reasoning.

safetyarxiv-cs-ai
14 Apr 2026
← Previous
1…6162636465…300
Next →