AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

CSA-Graphs: A Privacy-Preserving Structural Dataset for Child Sexual Abuse Research

DGX agent

arXiv:2604.07132v1 Announce Type: cross Abstract: Child Sexual Abuse Imagery (CSAI) classification is an important yet challenging problem for computer vision research due to the strict legal and ethi

safetyarxiv-cs-ai
10 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Data Leakage in Automotive Perception: Practitioners' Insights

DGX agent

arXiv:2604.06899v1 Announce Type: cross Abstract: Data leakage is the inadvertent transfer of information between training and evaluation datasets that poses a subtle, yet critical, risk to the reliab

safetyarxiv-cs-lg
10 Apr 2026
Safety

Deep Learning-Powered Visual SLAM Aimed at Assisting Visually Impaired Navigation

DGX agent

arXiv:2510.20549v2 Announce Type: replace Abstract: Despite advancements in SLAM technologies, robust operation under challenging conditions such as low-texture, motion-blur, or challenging lighting r

safetyarxiv-cs-cv
10 Apr 2026
Local Ai

FedDetox: Robust Federated SLM Alignment via On-Device Data Sanitization

DGX agent

arXiv:2604.06833v1 Announce Type: cross Abstract: As high quality public data becomes scarce, Federated Learning (FL) provides a vital pathway to leverage valuable private user data while preserving p

local-aiarxiv-cs-lg
10 Apr 2026
Safety

From experimentation to engagement: on the paradox of participatory AI and power in contexts of forced displacement and humanitarian crises

DGX agent

arXiv:2604.06219v1 Announce Type: cross Abstract: Across the Global North, calls for participatory artificial intelligence (AI) to improve the responsible, safe, and ethical use of AI have increased,

safetyarxiv-cs-ai
10 Apr 2026
Safety

Front-End Ethics for Sensor-Fused Health Conversational Agents: An Ethical Design Space for Biometrics

DGX agent

arXiv:2604.06203v1 Announce Type: cross Abstract: The integration of continuous data from built-in sensors and Large Language Models (LLMs) has fueled a surge of 'Sensor-Fused LLM agents' for personal

safetyarxiv-cs-ai
10 Apr 2026
Safety

Governing frontier general-purpose AI in the public sector: adaptive risk management and policy capacity under uncertainty through 2030

DGX agent

arXiv:2604.06215v1 Announce Type: cross Abstract: The governance of frontier general-purpose artificial intelligence has become a public-sector problem of institutional design, not merely a technical

safetyarxiv-cs-ai
10 Apr 2026
Safety

Harnessing Embodied Agents: Runtime Governance for Policy-Constrained Execution

DGX agent

arXiv:2604.07833v1 Announce Type: new Abstract: Embodied agents are evolving from passive reasoning systems into active executors that interact with tools, robots, and physical environments. Once gran

safetyarxiv-cs-ro
10 Apr 2026
Safety

Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization

DGX agent

arXiv:2604.06285v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have become essential for tasks such as image synthesis, captioning, and retrieval by aligning textual and visual inform

safetyarxiv-cs-ai
10 Apr 2026
Safety

Incorporating Social Awareness into Control of Unknown Multi-Agent Systems: A Real-Time Spatiotemporal Tubes Approach

DGX agent

arXiv:2510.25597v2 Announce Type: replace-cross Abstract: This paper presents a decentralized control framework that incorporates social awareness into multi-agent systems with unknown dynamics to ach

safetyarxiv-cs-ro
10 Apr 2026
Safety

Invisible to Humans, Triggered by Agents: Stealthy Jailbreak Attacks on Mobile Vision-Language Agents

DGX agent

arXiv:2510.07809v4 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) empower autonomous mobile agents, yet their security under realistic mobile deployment constraints remain

safetyarxiv-cs-ai
10 Apr 2026
Safety

Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEM

DGX agent

arXiv:2604.08425v1 Announce Type: cross Abstract: When humans label subjective content, they disagree, and that disagreement is not noise. It reflects genuine differences in perspective shaped by anno

safetyarxiv-cs-cl
10 Apr 2026
Safety

LINE: LLM-based Iterative Neuron Explanations for Vision Models

DGX agent

arXiv:2604.08039v1 Announce Type: new Abstract: Interpreting the concepts encoded by individual neurons in deep neural networks is a crucial step towards understanding their complex decision-making pr

safetyarxiv-cs-cv
10 Apr 2026
Safety

LUMINA: Foundation Models for Topology Transferable ACOPF

DGX agent

arXiv:2603.04300v2 Announce Type: replace Abstract: Foundation models in general promise to accelerate scientific computation by learning reusable representations across problem instances, yet constra

safetyarxiv-cs-lg
10 Apr 2026
Safety

MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning

DGX agent

arXiv:2604.08203v1 Announce Type: new Abstract: Medical Vision-Language Models (VLMs) hold immense promise for complex clinical tasks, but their reasoning capabilities are often constrained by text-on

safetyarxiv-cs-cv
10 Apr 2026
Safety

Rethinking Generalization in Reasoning SFT: A Conditional Analysis on Optimization, Data, and Model Capability

DGX agent

arXiv:2604.06628v1 Announce Type: new Abstract: A prevailing narrative in LLM post-training holds that supervised finetuning (SFT) memorizes while reinforcement learning (RL) generalizes. We revisit t

safetyarxiv-cs-ai
10 Apr 2026
Safety

Safe Large-Scale Robust Nonlinear MPC in Milliseconds via Reachability-Constrained System Level Synthesis on the GPU

DGX agent

arXiv:2604.07644v1 Announce Type: new Abstract: We present GPU-SLS, a GPU-parallelized framework for safe, robust nonlinear model predictive control (MPC) that scales to high-dimensional uncertain rob

safetyarxiv-cs-ro
10 Apr 2026
Safety

Towards the Development of an LLM-Based Methodology for Automated Security Profiling in Compliance with Ukrainian Cybersecurity Regulations

DGX agent

arXiv:2604.06274v1 Announce Type: cross Abstract: In recent years, the pace of development of information technology in various areas has increased drastically, forcing cybersecurity specialists to co

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories

DGX agent

arXiv:2604.07223v1 Announce Type: cross Abstract: As large language models (LLMs) evolve from static chatbots into autonomous agents, the primary vulnerability surface shifts from final outputs to int

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

WebSP-Eval: Evaluating Web Agents on Website Security and Privacy Tasks

DGX agent

arXiv:2604.06367v1 Announce Type: cross Abstract: Web agents automate browser tasks, ranging from simple form completion to complex workflows like ordering groceries. While current benchmarks evaluate

safetyarxiv-cs-ai
10 Apr 2026
Safety

A-3PO: Accelerating Asynchronous LLM Training with Staleness-aware Proximal Policy Approximation

DGX agent

arXiv:2512.06547v4 Announce Type: replace-cross Abstract: Decoupled PPO has been a successful reinforcement learning (RL) algorithm to deal with the high data staleness under the asynchronous RL setti

safetyarxiv-cs-ai
13 Aug 2026
Safety

A method for tissue-mask supported whole-body image registration in the UK Biobank

DGX agent

arXiv:2512.02702v3 Announce Type: replace Abstract: The UK Biobank is a large-scale study collecting imaging and non-imaging health data. Robust and accurate inter-subject image registration of these

safetyarxiv-cs-cv
13 Aug 2026
Safety

A New First-Order Meta-Learning Algorithm with Convergence Guarantees

DGX agent

arXiv:2409.03682v2 Announce Type: replace Abstract: Learning new tasks by leveraging prior experience is a fundamental trait of intelligent systems. While Model-Agnostic Meta-Learning (MAML) is a lead

safetyarxiv-cs-lg
13 Aug 2026
Safety

Accuracy and Order Sensitivity Diverge Under Label-Free Strategies

DGX agent

arXiv:2608.11947v1 Announce Type: cross Abstract: Multiple-choice benchmarks are widely used to evaluate large language models, but MCQ scores conflate knowledge with sensitivity to option order, whic

safetyarxiv-cs-ai
13 Aug 2026
Safety

Adaptation of Generalist Robot Policies with Minimal Data

DGX agent

arXiv:2608.11363v1 Announce Type: cross Abstract: A central goal in robot learning is to move beyond task-specific human data collection toward robots that improve through autonomous interaction. Yet

safetyarxiv-cs-lg
13 Aug 2026
Safety

Alignment of Similarity-Transformed Images Based on Fourier--Mellin Transform Using Auxiliary Function Method

DGX agent

arXiv:2608.11565v1 Announce Type: cross Abstract: This paper proposes an algorithm for estimating the similarity transformation, namely translation, scale, and rotation, between two images with subpix

safetyarxiv-cs-cv
13 Aug 2026
Safety

Anti-Shortcut Distillation via Temporal Negative Knowledge Transfer

DGX agent

arXiv:2608.11789v1 Announce Type: new Abstract: Knowledge distillation (KD) trains a compact student by attracting it towards a converged teacher. It is silent about which directions the teacher itsel

safetyarxiv-cs-cv
13 Aug 2026
Safety

AutoGrable: What Is a Good Graph for a Table?

DGX agent

arXiv:2608.11431v1 Announce Type: new Abstract: Graph learning presupposes a graph, and tables and relational databases do not come with one. Applying a GNN to them requires deciding which entities be

safetyarxiv-cs-lg
13 Aug 2026
Safety

Benchmarking Trustworthiness of SLMs: Pre-trained vs. Compressed

DGX agent

arXiv:2608.11981v1 Announce Type: new Abstract: Small Language Models (SLMs) have emerged as a more efficient alternative to traditional Large Language Models (LLMs), offering promising potential in r

safetyarxiv-cs-cl
13 Aug 2026
Safety

Better Slots, Better Worlds: Representation Quality & Robustness in Object-Centric World Models

DGX agent

arXiv:2608.12078v1 Announce Type: cross Abstract: Learning world models from offline trajectories enables agents to accomplish different tasks through planning. Object-centric (OC) representations, wh

safetyarxiv-cs-ai
13 Aug 2026
Safety

CLAIM: Leading Open-domain Active Clarification of Large Language Models with Uncertainty Measurement

DGX agent

arXiv:2608.11631v1 Announce Type: new Abstract: In open-domain human-computer interaction scenarios, large language models (LLMs) frequently encounter user queries that are ambiguous or incomplete. In

safetyarxiv-cs-ai
13 Aug 2026
Safety

CoAdapt-GUI: Joint Workflow Context and Policy Adaptation for Unseen GUI Applications

DGX agent

arXiv:2608.11588v1 Announce Type: new Abstract: Mobile GUI agents remain brittle when deployed to applications absent from source training. We study novel-app generalization under a limited target int

safetyarxiv-cs-ai
13 Aug 2026
Safety

Continuous-Latent Predictive Modeling with Semantic Alignment for EEG-Language Foundation Models

DGX agent

arXiv:2608.11656v1 Announce Type: new Abstract: Recent advances in EEG foundation models have demonstrated the potential of large-scale pretraining to enable generalizable neural decoding across subje

safetyarxiv-cs-lg
13 Aug 2026
Safety

CookVoice: Unified Framework for Style Controllable Multi-Modal Human Voice Generation

DGX agent

arXiv:2608.11590v1 Announce Type: cross Abstract: Human voice generation has made rapid progress in speech generation, singing voice generation, voice cloning, and voice editing. However, most existin

safetyarxiv-cs-lg
13 Aug 2026
Safety

Defending against Model Extraction for GNNs with Model Reprogramming

DGX agent

arXiv:2608.11495v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) serve as the backbone for high-stakes applications in Machine-Learning-as-a-Service (MLaaS). Still, their black-box deploym

safetyarxiv-cs-lg
13 Aug 2026
Safety

Diagnosis Before Recovery: Turning Agent Failures into Selective Self-Correction

DGX agent

arXiv:2608.11772v1 Announce Type: new Abstract: Self-correction is particularly useful when a failure constrains the next repair. Coding agents benefit from this property because compilers, tests, and

safetyarxiv-cs-cl
13 Aug 2026
Safety

Diffusion-Guided Cooperative Policy Learning for Target Tracking Based on Underwater Mobile Agent Networks

DGX agent

arXiv:2603.29426v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning (MARL) provides a promising solution for cooperative target tracking in networks of autonomous underwater v

safetyarxiv-cs-lg
13 Aug 2026
Safety

Disentangling the Expressivity of RoPE

DGX agent

arXiv:2608.11909v1 Announce Type: new Abstract: Two accounts recur in explanations of the success of rotary position embeddings (RoPE). Expressivity studies associate periodic position information wit

safetyarxiv-cs-lg
13 Aug 2026
Safety

Dual Anchors, Do It Better: Hierarchical Group Merging for Zero-Shot Anomaly Detection

DGX agent

arXiv:2608.11933v1 Announce Type: new Abstract: Zero-shot anomaly detection (ZSAD) aims to identify anomalies in unseen domains, a setting that is particularly critical for industrial and medical appl

safetyarxiv-cs-cv
13 Aug 2026
Safety

Dual-Domain Cross-Modal Decoding for Clinical Text-Guided Medical Image Segmentation

DGX agent

arXiv:2608.11335v1 Announce Type: cross Abstract: Clinical text can narrow down what to segment, but recent text-guided designs emphasize spatial alignment while overlooking frequency content that gov

safetyarxiv-cs-ai
13 Aug 2026
Safety

Dynamic Governance of Multi-LLM Agent Systems for Collaborative Conversational Outcomes

DGX agent

arXiv:2608.11207v1 Announce Type: new Abstract: When two LLM agents with structurally opposed objectives interact across multiple turns, the absence of a shared goal function produces not competition

safetyarxiv-cs-ai
13 Aug 2026
Safety

Enhancing Visual Domain Robustness in Behaviour Cloning via Saliency-Guided Augmentation

DGX agent

arXiv:2608.11870v1 Announce Type: new Abstract: In vision-based behavior cloning (BC), conventional image augmentations such as Random Crop and Color Jitter often fall short under substantial visual d

safetyarxiv-cs-ro
13 Aug 2026
Safety

Epiplexity Guided Data Selection and Generation for Out-of-Distribution Generalization

DGX agent

arXiv:2608.11746v1 Announce Type: cross Abstract: Modern systems are increasingly expected to transfer across tasks not specified during training. What data facilitates generalization in these new, un

safetyarxiv-cs-cl
13 Aug 2026
Safety

FM-LLM: A frequency-enhanced mixture-of-experts framework for adapting LLMs to time series forecasting

DGX agent

arXiv:2608.11623v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have spurred cross-modal solutions for time-series forecasting. However, existing methods rely heavily

safetyarxiv-cs-ai
13 Aug 2026
Safety

Foresight Without Seeing: Latent Futures for World Action Models

DGX agent

arXiv:2608.11605v1 Announce Type: new Abstract: World Action Models (WAMs) couple future visual prediction with robot action generation, enabling policies to model how the physical world evolves durin

safetyarxiv-cs-ai
13 Aug 2026
Safety

FQTree: Fine-grained Quantization and Hardware Generation of Boosted Decision Trees

DGX agent

arXiv:2608.12140v1 Announce Type: cross Abstract: Boosted decision trees (BDTs) are widely used in latency-critical applications, but efficient hardware deployment remains challenging. Existing design

safetyarxiv-cs-lg
13 Aug 2026
Safety

From Prompting to Behavioral Alignment: Personalized LLM Judges for Recommendation Evaluation

DGX agent

arXiv:2608.11493v1 Announce Type: new Abstract: Traditional offline recommendation evaluation relies heavily on complex, manually maintained feature pipelines that are difficult to scale. While Large

safetyarxiv-cs-ai
13 Aug 2026
Safety

GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs

DGX agent

arXiv:2608.11674v1 Announce Type: cross Abstract: On-policy rollout methods such as GRPO are central to post-training of large language models, yet they frequently suffer from training instabilities,

safetyarxiv-cs-ai
13 Aug 2026
← Previous
1…5657585960…257
Next →