AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

On the Properties of Feature Attribution for Supervised Contrastive Learning

DGX agent

arXiv:2604.22540v1 Announce Type: cross Abstract: Most Neural Networks (NNs) for classification are trained using Cross-Entropy as a loss function. This approach requires the model to have an explicit

safetyarxiv-cs-ai
27 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Reliable Self-Harm Risk Screening via Adaptive Multi-Agent LLM Systems

DGX agent

arXiv:2604.22154v1 Announce Type: cross Abstract: Emerging AI systems in behavioral health and psychiatry use multi-step or multi-agent LLM pipelines for tasks like assessing self-harm risk and screen

safetyarxiv-cs-ai
27 Apr 2026
Safety

Rethinking XAI Evaluation: A Human-Centered Audit of Shapley Benchmarks in High-Stakes Settings

DGX agent

arXiv:2604.22662v1 Announce Type: cross Abstract: Shapley values are a cornerstone of explainable AI, yet their proliferation into competing formulations has created a fragmented landscape with little

safetyarxiv-cs-ai
27 Apr 2026
Safety

A Survey of Legged Robotics in Non-Inertial Environments: Past, Present, and Future

DGX agent

arXiv:2604.20990v1 Announce Type: new Abstract: Legged robots have demonstrated remarkable agility on rigid, stationary ground, but their locomotion reliability remains limited in non-inertial environ

safetyarxiv-cs-ro
24 Apr 2026
Safety

Automated Annotation of Shearographic Measurements Enabling Weakly Supervised Defect Detection

DGX agent

arXiv:2512.06171v2 Announce Type: replace Abstract: Shearography is an interferometric technique sensitive to surface displacement gradients, providing high sensitivity for detecting subsurface defect

safetyarxiv-cs-cv
24 Apr 2026
Safety

CARE: Counselor-Aligned Response Engine for Online Mental-Health Support

DGX agent

arXiv:2604.21352v1 Announce Type: new Abstract: Mental health challenges are increasing worldwide, straining emotional support services and leading to counselor overload. This can result in delayed re

safetyarxiv-cs-cl
24 Apr 2026
Model Releases

Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles

DGX agent

arXiv:2604.21152v1 Announce Type: cross Abstract: As state-of-the-art Large Language Models (LLMs) have become ubiquitous, ensuring equitable performance across diverse demographics is critical. Howev

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

Escaping the Agreement Trap: Defensibility Signals for Evaluating Rule-Governed AI

DGX agent

arXiv:2604.20972v1 Announce Type: new Abstract: Content moderation systems are typically evaluated by measuring agreement with human labels. In rule-governed environments this assumption fails: multip

safetyarxiv-cs-ai
24 Apr 2026
Safety

Fairness Evaluation and Inference Level Mitigation in LLMs

DGX agent

arXiv:2510.18914v4 Announce Type: replace-cross Abstract: Large language models often display undesirable behaviors embedded in their internal representations, undermining fairness, inconsistency drif

safetyarxiv-cs-ai
24 Apr 2026
Safety

Inferring High-Level Events from Timestamped Data: Complexity and Medical Applications

DGX agent

arXiv:2604.21793v1 Announce Type: new Abstract: In this paper, we develop a novel logic-based approach to detecting high-level temporally extended events from timestamped data and background knowledge

safetyarxiv-cs-ai
24 Apr 2026
Safety

Language-Conditioned Safe Trajectory Generation for Spacecraft Rendezvous

DGX agent

arXiv:2512.09111v3 Announce Type: replace-cross Abstract: Reliable real-time trajectory generation is essential for future autonomous spacecraft. While recent progress in nonconvex guidance and contro

safetyarxiv-cs-ai
24 Apr 2026
Safety

Probabilistic Verification of Neural Networks via Efficient Probabilistic Hull Generation

DGX agent

arXiv:2604.21556v1 Announce Type: new Abstract: The problem of probabilistic verification of a neural network investigates the probability of satisfying the safe constraints in the output space when t

safetyarxiv-cs-ai
24 Apr 2026
Safety

Task-specific Subnetwork Discovery in Reinforcement Learning for Autonomous Underwater Navigation

DGX agent

arXiv:2604.21640v1 Announce Type: cross Abstract: Autonomous underwater vehicles are required to perform multiple tasks adaptively and in an explainable manner under dynamic, uncertain conditions and

safetyarxiv-cs-ai
24 Apr 2026
Safety

TraceScope: Interactive URL Triage via Decoupled Checklist Adjudication

DGX agent

arXiv:2604.21840v1 Announce Type: cross Abstract: Modern phishing campaigns increasingly evade snapshot-based URL classifiers using interaction gates (e.g., checkbox/slider challenges), delayed conten

safetyarxiv-cs-ai
24 Apr 2026
Safety

Tumor-anchored deep feature random forests for out-of-distribution detection in lung cancer segmentation

DGX agent

arXiv:2512.08216v3 Announce Type: replace-cross Abstract: Accurate segmentation of lung tumors from 3D computed tomography (CT) scans is essential for automated treatment planning and response assessm

safetyarxiv-cs-cv
24 Apr 2026
Safety

Unbiased Prevalence Estimation with Multicalibrated LLMs

DGX agent

arXiv:2604.21549v1 Announce Type: new Abstract: Estimating the prevalence of a category in a population using imperfect measurement devices (diagnostic tests, classifiers, or large language models) is

safetyarxiv-cs-ai
24 Apr 2026
Safety

Why Do Language Model Agents Whistleblow?

DGX agent

arXiv:2511.17085v3 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) as tool-using agents causes their alignment training to manifest in new ways. Recent work finds

safetyarxiv-cs-ai
24 Apr 2026
Safety

Environmental Understanding Vision-Language Model for Embodied Agent

DGX agent

arXiv:2604.19839v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown strong perception and reasoning abilities for instruction-following embodied agents. However, despite these a

safetyarxiv-cs-ai
23 Apr 2026
Safety

Explainable AML Triage with LLMs: Evidence Retrieval and Counterfactual Checks

DGX agent

arXiv:2604.19755v1 Announce Type: new Abstract: Anti-money laundering (AML) transaction monitoring generates large volumes of alerts that must be rapidly triaged by investigators under strict audit an

safetyarxiv-cs-ai
23 Apr 2026
Safety

From Fuzzy to Formal: Scaling Hospital Quality Improvement with AI

DGX agent

arXiv:2604.20055v1 Announce Type: new Abstract: Hospital Quality Improvement (QI) plays a critical role in optimizing healthcare delivery by translating high-level hospital goals into actionable solut

safetyarxiv-cs-ai
23 Apr 2026
Safety

From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP

DGX agent

arXiv:2510.12817v3 Announce Type: replace-cross Abstract: Human Label Variation (HLV) refers to legitimate disagreement in annotation that reflects the diversity of human perspectives rather than mere

safetyarxiv-cs-ai
23 Apr 2026
Safety

FSFM: A Biologically-Inspired Framework for Selective Forgetting of Agent Memory

DGX agent

arXiv:2604.20300v1 Announce Type: new Abstract: For LLM agents, memory management critically impacts efficiency, quality, and security. While much research focuses on retention, selective forgetting--

safetyarxiv-cs-ai
23 Apr 2026
Safety

Large language models perceive cities through a culturally uneven baseline

DGX agent

arXiv:2604.20048v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to describe, evaluate and interpret places, yet it remains unclear whether they do so from a cultural

safetyarxiv-cs-cl
23 Apr 2026
Safety

MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills

DGX agent

arXiv:2604.20441v1 Announce Type: new Abstract: Background: Agent skills are increasingly deployed as modular, reusable capability units in AI agent systems. Medical research agent skills require safe

safetyarxiv-cs-ai
23 Apr 2026
Safety

NeuroSymActive: Differentiable Neural-Symbolic Reasoning with Active Exploration for Knowledge Graph Question Answering

DGX agent

arXiv:2602.15353v2 Announce Type: replace-cross Abstract: Large pretrained language models and neural reasoning systems have advanced many natural language tasks, yet they remain challenged by knowled

safetyarxiv-cs-ai
23 Apr 2026
Safety

Semantic Prompting: Agentic Incremental Narrative Refinement through Spatial Semantic Interaction

DGX agent

arXiv:2604.19971v1 Announce Type: cross Abstract: Interactive spatial layouts empower users to synthesize information and organize findings for sensemaking. While Large Language Models (LLMs) can auto

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

A Functionality-Grounded Benchmark for Evaluating Web Agents in E-commerce Domains

DGX agent

arXiv:2508.15832v2 Announce Type: replace-cross Abstract: Web agents have shown great promise in performing many tasks on ecommerce website. To assess their capabilities, several benchmarks have been

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

Assessing VLM-Driven Semantic-Affordance Inference for Non-Humanoid Robot Morphologies

DGX agent

arXiv:2604.19509v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in understanding human-object interactions, but their application to robotic sys

safetyarxiv-cs-ro
22 Apr 2026
Safety

ASVSim (AirSim for Surface Vehicles): A High-Fidelity Simulation Framework for Autonomous Surface Vehicle Research

DGX agent

arXiv:2506.22174v2 Announce Type: replace-cross Abstract: The transport industry has recently shown significant interest in unmanned surface vehicles (USVs), specifically for port and inland waterway

safetyarxiv-cs-lg
22 Apr 2026
Safety

CAHAL: Clinically Applicable resolution enHAncement for Low-resolution MRI scans

DGX agent

arXiv:2604.18781v1 Announce Type: new Abstract: Large-scale automated morphometric analysis of brain MRI is limited by the thick-slice, anisotropic acquisitions prevalent in routine clinical practice.

safetyarxiv-cs-cv
22 Apr 2026
Safety

Decomposed Trust: Privacy, Adversarial Robustness, Ethics, and Fairness in Low-Rank LLMs

DGX agent

arXiv:2511.22099v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have driven major advances across domains, yet their massive size hinders deployment in resource-constrained sett

safetyarxiv-cs-ai
22 Apr 2026
Safety

Developing a Robotic Surgery Training System for Wide Accessibility and Research

DGX agent

arXiv:2505.20562v2 Announce Type: replace Abstract: Robotic surgery represents a major breakthrough in medical interventions, which has revolutionized surgical procedures. However, the high cost and l

safetyarxiv-cs-ro
22 Apr 2026
Safety

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling

DGX agent

arXiv:2604.19544v1 Announce Type: new Abstract: Multimodal reward models (MRMs) play a crucial role in aligning Multimodal Large Language Models (MLLMs) with human preferences. Training a good MRM req

safetyarxiv-cs-ai
22 Apr 2026
Safety

Hierarchically Robust Zero-shot Vision-language Models

DGX agent

arXiv:2604.18867v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) can perform zero-shot classification but are susceptible to adversarial attacks. While robust fine-tuning improves their

safetyarxiv-cs-ai
22 Apr 2026
Safety

Symbolic Quantile Regression for the Interpretable Prediction of Conditional Quantiles

DGX agent

arXiv:2508.08080v2 Announce Type: replace Abstract: Symbolic Regression (SR) is a well-established framework for generating interpretable or white-box predictive models. Although SR has been successfu

safetyarxiv-cs-lg
22 Apr 2026
Safety

Task-Adaptive Admittance Control for Human-Quadrotor Cooperative Load Transportation with Dynamic Cable-Length Regulation

DGX agent

arXiv:2604.18905v1 Announce Type: new Abstract: The collaboration between humans and robots is critical in many robotic applications, especially in those requiring physical human-robot interaction (pH

safetyarxiv-cs-ro
22 Apr 2026
Safety

TROJail: Trajectory-Level Optimization for Multi-Turn Large Language Model Jailbreaks with Process Rewards

DGX agent

arXiv:2512.07761v3 Announce Type: replace Abstract: Large language models have seen widespread adoption, yet they remain vulnerable to multi-turn jailbreak attacks, threatening their safe deployment.

safetyarxiv-cs-ai
22 Apr 2026
Safety

User Simulation in the Era of Generative AI: User Modeling, Synthetic Data Generation, and System Evaluation

DGX agent

arXiv:2501.04410v2 Announce Type: replace Abstract: User simulation is an emerging interdisciplinary topic with multiple critical applications in the era of Generative AI. It involves creating an inte

safetyarxiv-cs-ai
22 Apr 2026
Safety

Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition

DGX agent

arXiv:2604.17803v1 Announce Type: cross Abstract: Post-training Large Language Models requires diverse, high-quality data which is rare and costly to obtain, especially in low resource domains and for

safetyarxiv-cs-lg
21 Apr 2026
Safety

Arch: An AI-Native Hardware Description Language for Register-Transfer Clocked Hardware Design

DGX agent

arXiv:2604.05983v2 Announce Type: replace-cross Abstract: We present Arch (AI-native Register-transfer Clocked Hardware), a hardware description language for micro-architecture specification and AI-as

safetyarxiv-cs-cl
21 Apr 2026
Safety

BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories

DGX agent

arXiv:2604.17008v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to generate narrative content, including children's stories, which play an important role in social a

safetyarxiv-cs-cl
21 Apr 2026
Safety

CAPC-CG: A Large-Scale, Expert-Directed LLM-Annotated Corpus of Adaptive Policy Communication in China

DGX agent

arXiv:2510.08986v2 Announce Type: replace Abstract: We introduce CAPC-CG, the Chinese Adaptive Policy Communication (Central Government) Corpus, the first open dataset of Chinese policy directives ann

safetyarxiv-cs-cl
21 Apr 2026
Safety

Deep learning based Non-Rigid Volume-to-Surface Registration for Brain Shift compensation Using Point Cloud

DGX agent

arXiv:2604.17389v1 Announce Type: new Abstract: Soft-tissue deformation remains a major limitation in image-guided neurosurgery, where intra-operative anatomy can deviate substantially from pre-operat

safetyarxiv-cs-cv
21 Apr 2026
Safety

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning

DGX agent

arXiv:2510.00761v5 Announce Type: replace Abstract: Large language model (LLM) unlearning aims to surgically remove the influence of undesired data or knowledge from an existing model while preserving

safetyarxiv-cs-lg
21 Apr 2026
Safety

Dual Alignment Between Language Model Layers and Human Sentence Processing

DGX agent

arXiv:2604.18563v1 Announce Type: new Abstract: A recent study (Kuribayashi et al., 2025) has shown that human sentence processing behavior, typically measured on syntactically unchallenging construct

safetyarxiv-cs-cl
21 Apr 2026
Safety

Emergency Stopping for Liquid-manipulating Robots

DGX agent

arXiv:2604.16667v1 Announce Type: new Abstract: Manipulating open liquid containers is challenging because liquids are highly sensitive to vessel accelerations and jerks. Although spill-free liquid ma

safetyarxiv-cs-ro
21 Apr 2026
Safety

Flow-Opt: Scalable Centralized Multi-Robot Trajectory Optimization with Flow Matching and Differentiable Optimization

DGX agent

arXiv:2510.09204v2 Announce Type: replace-cross Abstract: Centralized trajectory optimization in the joint space of multiple robots allows access to a larger feasible space that can result in smoother

safetyarxiv-cs-lg
21 Apr 2026
Safety

Infrastructure-Centric World Models: Bridging Temporal Depth and Spatial Breadth for Roadside Perception

DGX agent

arXiv:2604.17651v1 Announce Type: new Abstract: World models, generative AI systems that simulate how environments evolve, are transforming autonomous driving, yet all existing approaches adopt an ego

safetyarxiv-cs-cv
21 Apr 2026
← Previous
1…5354555657…257
Next →