AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
Safety

Rethinking Cross-Domain Evaluation for Face Forgery Detection with Semantic Fine-grained Alignment and Mixture-of-Experts

DGX agent

arXiv:2604.21478v1 Announce Type: new Abstract: Nowadays, visual data forgery detection plays an increasingly important role in social and economic security with the rapid development of generative mo

safetyarxiv-cs-cv
24 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

RIFT: Repurposing Negative Samples via Reward-Informed Fine-Tuning

DGX agent

arXiv:2601.09253v2 Announce Type: replace-cross Abstract: While Supervised Fine-Tuning (SFT) and Rejection Sampling Fine-Tuning (RFT) are standard for LLM alignment, they either rely on costly expert

safetyarxiv-cs-ai
24 Apr 2026
Safety

Robustness Analysis of POMDP Policies to Observation Perturbations

DGX agent

arXiv:2604.21256v1 Announce Type: new Abstract: Policies for Partially Observable Markov Decision Processes (POMDPs) are often designed using a nominal system model. In practice, this model can deviat

safetyarxiv-cs-ai
24 Apr 2026
Safety

RPG: Robust Policy Gating for Smooth Multi-Skill Transitions in Humanoid Fighting

DGX agent

arXiv:2604.21355v1 Announce Type: new Abstract: Humanoid robots have demonstrated impressive motor skills in a wide range of tasks, yet whole-body control for humanlike long-time, dynamic fighting rem

safetyarxiv-cs-ro
24 Apr 2026
Safety

SafeMERGE: Preserving Safety Alignment in Fine-Tuned Large Language Models via Selective Layer-Wise Model Merging

DGX agent

arXiv:2503.17239v3 Announce Type: replace-cross Abstract: Fine-tuning large language models (LLMs) is a common practice to adapt generalist models to specialized domains. However, recent studies show

safetyarxiv-cs-ai
24 Apr 2026
Safety

SafeRedirect: Defeating Internal Safety Collapse via Task-Completion Redirection in Frontier LLMs

DGX agent

arXiv:2604.20930v1 Announce Type: cross Abstract: Internal Safety Collapse (ISC) is a failure mode in which frontier LLMs, when executing legitimate professional tasks whose correct completion structu

safetyarxiv-cs-ai
24 Apr 2026
Safety

Seeing Without Eyes: 4D Human-Scene Understanding from Wearable IMUs

DGX agent

arXiv:2604.21926v1 Announce Type: new Abstract: Understanding human activities and their surrounding environments typically relies on visual perception, yet cameras pose persistent challenges in priva

safetyarxiv-cs-cv
24 Apr 2026
Safety

Self-Predictive Representation for Autonomous UAV Object-Goal Navigation

DGX agent

arXiv:2604.21130v1 Announce Type: new Abstract: Autonomous Unmanned Aerial Vehicles (UAVs) have revolutionized industries through their versatility with applications including aerial surveillance, sea

safetyarxiv-cs-ro
24 Apr 2026
Safety

SemaPop: Semantic-Persona Conditioned and Controllable Population Synthesis

DGX agent

arXiv:2602.11569v2 Announce Type: replace Abstract: Population synthesis is essential for individual-level simulation in transport planning and socio-economic analysis, yet remains challenging due to

safetyarxiv-cs-ai
24 Apr 2026
Safety

SGG-R^{rm 3}: From Next-Token Prediction to End-to-End Unbiased Scene Graph Generation

DGX agent

arXiv:2603.07961v3 Announce Type: replace Abstract: Scene Graph Generation (SGG) structures visual scenes as graphs of objects and their relations. While Multimodal Large Language Models (MLLMs) have

safetyarxiv-cs-cv
24 Apr 2026
Safety

Strategic Polysemy in AI Discourse: A Philosophical Analysis of Language, Hype, and Power

DGX agent

arXiv:2604.21043v1 Announce Type: cross Abstract: This paper examines the strategic use of language in contemporary artificial intelligence (AI) discourse, focusing on the widespread adoption of metap

safetyarxiv-cs-ai
24 Apr 2026
Safety

StyleVAR: Controllable Image Style Transfer via Visual Autoregressive Modeling

DGX agent

arXiv:2604.21052v1 Announce Type: cross Abstract: We build on the Visual Autoregressive Modeling (VAR) framework and formulate style transfer as conditional discrete sequence modeling in a learned lat

safetyarxiv-cs-ai
24 Apr 2026
Safety

Supervised Learning Has a Necessary Geometric Blind Spot: Theory, Consequences, and Minimal Repair

DGX agent

arXiv:2604.21395v1 Announce Type: cross Abstract: We prove that empirical risk minimisation (ERM) imposes a necessary geometric constraint on learned representations: any encoder that minimises superv

safetyarxiv-cs-ai
24 Apr 2026
Safety

Survey on Evaluation of LLM-based Agents

DGX agent

arXiv:2503.16416v2 Announce Type: replace Abstract: LLM-based agents represent a paradigm shift in AI, enabling autonomous systems to plan, reason, and use tools while interacting with dynamic environ

safetyarxiv-cs-ai
24 Apr 2026
Safety

Task-specific Subnetwork Discovery in Reinforcement Learning for Autonomous Underwater Navigation

DGX agent

arXiv:2604.21640v1 Announce Type: cross Abstract: Autonomous underwater vehicles are required to perform multiple tasks adaptively and in an explainable manner under dynamic, uncertain conditions and

safetyarxiv-cs-ai
24 Apr 2026
Safety

Tempered Sequential Monte Carlo for Trajectory and Policy Optimization with Differentiable Dynamics

DGX agent

arXiv:2604.21456v1 Announce Type: new Abstract: We propose a sampling-based framework for finite-horizon trajectory and policy optimization under differentiable dynamics by casting controller design a

safetyarxiv-cs-lg
24 Apr 2026
Safety

Temporal Prototyping and Hierarchical Alignment for Unsupervised Video-based Visible-Infrared Person Re-Identification

DGX agent

arXiv:2604.21324v1 Announce Type: new Abstract: Visible-infrared person re-identification (VI-ReID) enables cross-modality identity matching for all-day surveillance, yet existing methods predominantl

safetyarxiv-cs-cv
24 Apr 2026
Safety

The Economics of p(doom): Scenarios of Existential Risk and Economic Growth in the Age of Transformative AI

DGX agent

arXiv:2503.07341v2 Announce Type: replace-cross Abstract: Recent advances in artificial intelligence (AI) have led to a wide range of predictions about its long-term impact on humanity. A central focu

safetyarxiv-cs-ai
24 Apr 2026
Safety

The Effect of Idea Elaboration on the Automatic Assessment of Idea Originality

DGX agent

arXiv:2604.20569v1 Announce Type: cross Abstract: Automatic systems are increasingly used to assess the originality of responses in creative tasks. They offer a potential solution to key limitations o

safetyarxiv-cs-ai
24 Apr 2026
Safety

'This Wasn't Made for Me': Recentering User Experience and Emotional Impact in the Evaluation of ASR Bias

DGX agent

arXiv:2604.21148v1 Announce Type: new Abstract: Studies on bias in Automatic Speech Recognition (ASR) tend to focus on reporting error rates for speakers of underrepresented dialects, yet less researc

safetyarxiv-cs-cl
24 Apr 2026
Safety

Time, Causality, and Observability Failures in Distributed AI Inference Systems

DGX agent

arXiv:2604.21361v1 Announce Type: new Abstract: Distributed AI inference pipelines rely heavily on timestamp-based observability to understand system behavior. This work demonstrates that even small c

safetyarxiv-cs-ai
24 Apr 2026
Safety

Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build …

DGX agent

Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build upon them and use them to evaluate the monitorability of the

safetysam-altman--x
24 Apr 2026
Safety

Towards a Systematic Risk Assessment of Deep Neural Network Limitations in Autonomous Driving Perception

DGX agent

arXiv:2604.20895v1 Announce Type: cross Abstract: Safety and security are essential for the admission and acceptance of automated and autonomous vehicles. Deep neural networks (DNNs) are widely used f

safetyarxiv-cs-lg
24 Apr 2026
Safety

TraceScope: Interactive URL Triage via Decoupled Checklist Adjudication

DGX agent

arXiv:2604.21840v1 Announce Type: cross Abstract: Modern phishing campaigns increasingly evade snapshot-based URL classifiers using interaction gates (e.g., checkbox/slider challenges), delayed conten

safetyarxiv-cs-ai
24 Apr 2026
Safety

Trust-SSL: Additive-Residual Selective Invariance for Robust Aerial Self-Supervised Learning

DGX agent

arXiv:2604.21349v1 Announce Type: cross Abstract: Self-supervised learning (SSL) is a standard approach for representation learning in aerial imagery. Existing methods enforce invariance between augme

safetyarxiv-cs-ai
24 Apr 2026
Safety

Tumor-anchored deep feature random forests for out-of-distribution detection in lung cancer segmentation

DGX agent

arXiv:2512.08216v3 Announce Type: replace-cross Abstract: Accurate segmentation of lung tumors from 3D computed tomography (CT) scans is essential for automated treatment planning and response assessm

safetyarxiv-cs-cv
24 Apr 2026
Safety

Unbiased Prevalence Estimation with Multicalibrated LLMs

DGX agent

arXiv:2604.21549v1 Announce Type: new Abstract: Estimating the prevalence of a category in a population using imperfect measurement devices (diagnostic tests, classifiers, or large language models) is

safetyarxiv-cs-ai
24 Apr 2026
Safety

UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection

DGX agent

arXiv:2604.21904v1 Announce Type: new Abstract: In recent years, significant progress has been made in both image generation and generated image detection. Despite their rapid, yet largely independent

safetyarxiv-cs-cv
24 Apr 2026
Safety

Value-Conflict Diagnostics Reveal Widespread Alignment Faking in Language Models

DGX agent

arXiv:2604.20995v1 Announce Type: new Abstract: Alignment faking, where a model behaves aligned with developer policy when monitored but reverts to its own preferences when unobserved, is a concerning

safetyarxiv-cs-ai
24 Apr 2026
Safety

VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models

DGX agent

arXiv:2510.18457v3 Announce Type: replace Abstract: The performance of Latent Diffusion Models (LDMs) is critically dependent on the quality of their visual tokenizers. While recent works have explore

safetyarxiv-cs-cv
24 Apr 2026
Safety

VFM^{4}SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection

DGX agent

arXiv:2604.21502v1 Announce Type: new Abstract: In real-world scenarios, continual changes in weather, illumination, and imaging conditions cause significant domain shifts, leading detectors trained o

safetyarxiv-cs-cv
24 Apr 2026
Safety

VLA-Forget: Vision-Language-Action Unlearning for Embodied Foundation Models

DGX agent

arXiv:2604.03956v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models are emerging as embodied foundation models for robotic manipulation, but their deployment introduces a new

safetyarxiv-cs-ai
24 Apr 2026
Safety

When Bigger Isn't Better: A Comprehensive Fairness Evaluation of Political Bias in Multi-News Summarisation

DGX agent

arXiv:2604.21309v1 Announce Type: new Abstract: Multi-document news summarisation systems are increasingly adopted for their convenience in processing vast daily news content, making fairness across d

safetyarxiv-cs-cl
24 Apr 2026
Safety

Who Defines Fairness? Target-Based Prompting for Demographic Representation in Generative Models

DGX agent

arXiv:2604.21036v1 Announce Type: new Abstract: Text-to-image(T2I) models like Stable Diffusion and DALL-E have made generative AI widely accessible, yet recent studies reveal that these systems often

safetyarxiv-cs-ai
24 Apr 2026
Safety

Why are all LLMs Obsessed with Japanese Culture? On the Hidden Cultural and Regional Biases of LLMs

DGX agent

arXiv:2604.21751v1 Announce Type: cross Abstract: LLMs have been showing limitations when it comes to cultural coverage and competence, and in some cases show regional biases such as amplifying Wester

safetyarxiv-cs-ai
24 Apr 2026
Safety

Why Do Language Model Agents Whistleblow?

DGX agent

arXiv:2511.17085v3 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) as tool-using agents causes their alignment training to manifest in new ways. Recent work finds

safetyarxiv-cs-ai
24 Apr 2026
Safety

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

DGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

safetyarxiv-cs-cl
24 Apr 2026
Safety

A Hough transform approach to safety-aware scalar field mapping using Gaussian Processes

DGX agent

arXiv:2604.20799v1 Announce Type: new Abstract: This paper presents a framework for mapping unknown scalar fields using a sensor-equipped autonomous robot operating in unsafe environments. The unsafe

safetyarxiv-cs-ro
23 Apr 2026
Safety

A lot of institutions and companies won’t die because of AI. They’ll die because of their own internal inability to learn. Someone recently …

DGX agent

A lot of institutions and companies won’t die because of AI. They’ll die because of their own internal inability to learn. Someone recently told me that at a big company, legal had to review and appro

safetycristobal-valenzuela--x
23 Apr 2026
Safety

A Survey of Scaling in Large Language Model Reasoning

DGX agent

arXiv:2504.02181v2 Announce Type: replace Abstract: The rapid advancements in large Language models (LLMs) have significantly enhanced their reasoning capabilities, driven by various strategies such a

safetyarxiv-cs-ai
23 Apr 2026
Safety

A Synchronized Audio-Visual Multi-View Capture System

DGX agent

arXiv:2603.23089v2 Announce Type: replace Abstract: Multi-view capture systems have been an important tool in research for recording human motion under controlling conditions. Most existing systems ar

safetyarxiv-cs-cv
23 Apr 2026
Safety

AdaTracker: Learning Adaptive In-Context Policy for Cross-Embodiment Active Visual Tracking

DGX agent

arXiv:2604.20305v1 Announce Type: new Abstract: Realizing active visual tracking with a single unified model across diverse robots is challenging, as the physical constraints and motion dynamics vary

safetyarxiv-cs-ro
23 Apr 2026
Safety

AgentSOC: A Multi-Layer Agentic AI Framework for Security Operations Automation

DGX agent

arXiv:2604.20134v1 Announce Type: cross Abstract: Security Operations Centers (SOCs) increasingly encounter difficulties in correlating heterogeneous alerts, interpreting multi-stage attack progressio

safetyarxiv-cs-ai
23 Apr 2026
Safety

AI models of unstable flow exhibit hallucination

DGX agent

arXiv:2604.20372v1 Announce Type: cross Abstract: We report the first systematic evidence of hallucination in AI models of fluid dynamics, demonstrated in the canonical problem of hydrodynamically uns

safetyarxiv-cs-ai
23 Apr 2026
Safety

Aligning Human-AI-Interaction Trust for Mental Health Support: Survey and Position for Multi-Stakeholders

DGX agent

arXiv:2604.20166v1 Announce Type: new Abstract: Building trustworthy AI systems for mental health support is a shared priority across stakeholders from multiple disciplines. However, 'trustworthy' rem

safetyarxiv-cs-cl
23 Apr 2026
Safety

All Languages Matter: Understanding and Mitigating Language Bias in Multilingual RAG

DGX agent

arXiv:2604.20199v1 Announce Type: new Abstract: Multilingual Retrieval-Augmented Generation (mRAG) leverages cross-lingual evidence to ground Large Language Models (LLMs) in global knowledge. However,

safetyarxiv-cs-cl
23 Apr 2026
Safety

Ask Only When Needed: Proactive Retrieval from Memory and Skills for Experience-Driven Lifelong Agents

DGX agent

arXiv:2604.20572v1 Announce Type: new Abstract: Online lifelong learning enables agents to accumulate experience across interactions and continually improve on long-horizon tasks. However, existing me

safetyarxiv-cs-cl
23 Apr 2026
Safety

Atomic Decision Boundaries: A Structural Requirement for Guaranteeing Execution-Time Admissibility in Autonomous Systems

DGX agent

arXiv:2604.17511v2 Announce Type: replace-cross Abstract: Autonomous systems increasingly execute actions that directly modify shared state, creating an urgent need for precise control over which tran

safetyarxiv-cs-ai
23 Apr 2026
← Previous
1…227228229230231…265
Next →