AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
All
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,600 results
Safety

AffordGen: Generating Diverse Demonstrations for Generalizable Object Manipulation with Afford Correspondence

DGX agent

arXiv:2604.10579v1 Announce Type: cross Abstract: Despite the recent success of modern imitation learning methods in robot manipulation, their performance is often constrained by geometric variations

safetyarxiv-cs-ai
14 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Agentic Video Generation: From Text to Executable Event Graphs via Tool-Constrained LLM Planning

DGX agent

arXiv:2604.10383v1 Announce Type: new Abstract: Existing multi-agent video generation systems use LLM agents to orchestrate neural video generators, producing visually impressive but semantically unre

safetyarxiv-cs-cv
14 Apr 2026
Safety

AI Integrity: A New Paradigm for Verifiable AI Governance

DGX agent

arXiv:2604.11065v1 Announce Type: new Abstract: AI systems increasingly shape high-stakes decisions in healthcare, law, defense, and education, yet existing governance paradigms -- AI Ethics, AI Safet

safetyarxiv-cs-ai
14 Apr 2026
Safety

AI Organizations are More Effective but Less Aligned than Individual Agents

DGX agent

arXiv:2604.10290v1 Announce Type: new Abstract: AI is increasingly deployed in multi-agent systems; however, most research considers only the behavior of individual models. We experimentally show that

safetyarxiv-cs-ai
14 Apr 2026
Safety

Anthropic believes that good transparency legislation needs to ensure public safety and accountability for the companies developing this pow…

DGX agent

Anthropic believes that good transparency legislation needs to ensure public safety and accountability for the companies developing this powerful technology, not provide a get-out-of-jail-free card ag

safetygary-marcus--x
14 Apr 2026
Safety

Anthropic details using AI agents to accelerate alignment research on 'weak-to-strong supervision', where a weak model supervises the training of a stronger one (Anthropic)

DGX agent

Anthropic: Anthropic details using AI agents to accelerate alignment research on “weak-to-strong supervision”, where a weak model supervises the training of a stronger one — Large language models' eve

safetytechmeme
14 Apr 2026
Safety

Anthropogenic Regional Adaptation in Multimodal Vision-Language Model

DGX agent

arXiv:2604.11490v1 Announce Type: new Abstract: While the field of vision-language (VL) has achieved remarkable success in integrating visual and textual information across multiple languages and doma

safetyarxiv-cs-ai
14 Apr 2026
Safety

// Artifacts as Memory Beyond the Agent Boundary // An agent doesn't always need a bigger memory buffer. Sometimes the environment itself re…

DGX agent

// Artifacts as Memory Beyond the Agent Boundary // An agent doesn't always need a bigger memory buffer. Sometimes the environment itself remembers on the agent's behalf. New research formalizes this

safetydair-ai--x
14 Apr 2026
Safety

ASPIRin: Action Space Projection for Interactivity-Optimized Reinforcement Learning in Full-Duplex Speech Language Models

DGX agent

arXiv:2604.10065v1 Announce Type: cross Abstract: End-to-end full-duplex Speech Language Models (SLMs) require precise turn-taking for natural interaction. However, optimizing temporal dynamics via st

safetyarxiv-cs-ai
14 Apr 2026
Safety

Assessing Model-Agnostic XAI Methods against EU AI Act Explainability Requirements

DGX agent

arXiv:2604.09628v1 Announce Type: cross Abstract: Explainable AI (XAI) has evolved in response to expectations and regulations, such as the EU AI Act, which introduces regulatory requirements on AI-po

safetyarxiv-cs-ai
14 Apr 2026
Safety

Auto-regressive transformation for image alignment

DGX agent

arXiv:2505.04864v2 Announce Type: replace-cross Abstract: Existing methods for image alignment struggle in cases involving feature-sparse regions, extreme scale and field-of-view differences, and larg

safetyarxiv-cs-ai
14 Apr 2026
Safety

Autonomous Diffractometry Enabled by Visual Reinforcement Learning

DGX agent

arXiv:2604.11773v1 Announce Type: cross Abstract: Automation underpins progress across scientific and industrial disciplines. Yet, automating tasks requiring interpretation of abstract visual informat

safetyarxiv-cs-cv
14 Apr 2026
Safety

AWARE: Adaptive Whole-body Active Rotating Control for Enhanced LiDAR-Inertial Odometry under Human-in-the-Loop Interaction

DGX agent

arXiv:2604.10598v1 Announce Type: new Abstract: Human-in-the-loop (HITL) UAV operation is essential in complex and safety-critical aerial surveying environments, where human operators provide navigati

safetyarxiv-cs-ro
14 Apr 2026
Safety

Awesome work by @jiaxinwen22, @liangqiu_1994, Joe Benton, and @janhkirchner! For more details, check out the blog post 👇 https://anthropic.…

DGX agent

Jan Leike praised collaborative work by researchers Jiaxin Wen, Liang Qiu, Joe Benton, and Jan Kirchner, directing followers to an Anthropic blog post for further details. The post appears to highligh

safetyjan-leike--x
14 Apr 2026
Safety

Backdoors in RLVR: Jailbreak Backdoors in LLMs From Verifiable Reward

DGX agent

arXiv:2604.09748v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an emerging paradigm that significantly boosts a Large Language Model's (LLM's) reasoning abi

safetyarxiv-cs-ai
14 Apr 2026
Safety

Belief-Aware VLM Model for Human-like Reasoning

DGX agent

arXiv:2604.09686v1 Announce Type: new Abstract: Traditional neural network models for intent inference rely heavily on observable states and struggle to generalize across diverse tasks and dynamic env

safetyarxiv-cs-ai
14 Apr 2026
Safety

Belief-State RWKV for Reinforcement Learning under Partial Observability

DGX agent

arXiv:2604.09671v1 Announce Type: new Abstract: We propose a stronger formulation of RL on top of RWKV-style recurrent sequence models, in which the fixed-size recurrent state is explicitly interprete

safetyarxiv-cs-lg
14 Apr 2026
Safety

Below-ground Fungal Biodiversity Can be Monitored Using Self-Supervised Learning Satellite Features

DGX agent

arXiv:2604.09818v1 Announce Type: new Abstract: Mycorrhizal fungi are vital to terrestrial ecosystem functioning. Yet monitoring their biodiversity at landscape scales is often unfeasible due to time

safetyarxiv-cs-lg
14 Apr 2026
Safety

Beyond Compliance: A Resistance-Informed Motivation Reasoning Framework for Challenging Psychological Client Simulation

DGX agent

arXiv:2604.10507v1 Announce Type: new Abstract: Psychological client simulators have emerged as a scalable solution for training and evaluating counselor trainees and psychological LLMs. Yet existing

safetyarxiv-cs-ai
14 Apr 2026
Safety

Beyond Message Passing: A Semantic View of Agent Communication Protocols

DGX agent

arXiv:2604.02369v3 Announce Type: replace-cross Abstract: Agent communication protocols are becoming critical infrastructure for large language model (LLM) systems that must use tools, coordinate with

safetyarxiv-cs-ai
14 Apr 2026
Safety

Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels

DGX agent

arXiv:2604.10367v1 Announce Type: new Abstract: Audio-driven human video generation has achieved remarkable success in monologue scenarios, largely driven by advancements in powerful video generation

safetyarxiv-cs-ai
14 Apr 2026
Safety

Beyond Reconstruction: Reconstruction-to-Vector Diffusion for Hyperspectral Anomaly Detection

DGX agent

arXiv:2604.11390v1 Announce Type: new Abstract: While Hyperspectral Anomaly Detection (HAD) excels at identifying sparse targets in complex scenes, existing models remain trapped in a scalar 'reconstr

safetyarxiv-cs-cv
14 Apr 2026
Safety

Bidirectional Learning of Facial Action Units and Expressions via Structured Semantic Mapping across Heterogeneous Datasets

DGX agent

arXiv:2604.10541v1 Announce Type: new Abstract: Facial action unit (AU) detection and facial expression (FE) recognition can be jointly viewed as affective facial behavior tasks, representing fine-gra

safetyarxiv-cs-cv
14 Apr 2026
Safety

Binary Flow Matching: Prediction-Loss Space Alignment for Robust Learning

DGX agent

arXiv:2602.10420v2 Announce Type: replace Abstract: Flow matching has emerged as a powerful framework for generative modeling, with recent empirical successes highlighting the effectiveness of signal-

safetyarxiv-cs-lg
14 Apr 2026
Safety

bioLeak: Leakage-Aware Modeling and Diagnostics for Machine Learning in R

DGX agent

arXiv:2604.10965v1 Announce Type: cross Abstract: Data leakage remains a recurrent source of optimistic bias in biomedical machine learning studies. Standard row-wise cross-validation and globally est

safetyarxiv-cs-lg
14 Apr 2026
Safety

Brain-Grasp: Graph-based Saliency Priors for Improved fMRI-based Visual Brain Decoding

DGX agent

arXiv:2604.10617v1 Announce Type: cross Abstract: Recent progress in brain-guided image generation has improved the quality of fMRI-based reconstructions; however, fundamental challenges remain in pre

safetyarxiv-cs-cv
14 Apr 2026
Safety

Bridging Linguistic Gaps: Cross-Lingual Mapping in Pre-Training and Dataset for Enhanced Multilingual LLM Performance

DGX agent

arXiv:2604.10590v1 Announce Type: cross Abstract: Multilingual Large Language Models (LLMs) struggle with cross-lingual tasks due to data imbalances between high-resource and low-resource languages, a

safetyarxiv-cs-ai
14 Apr 2026
Safety

Bridging the RGB-IR Gap: Consensus and Discrepancy Modeling for Text-Guided Multispectral Detection

DGX agent

arXiv:2604.11234v1 Announce Type: new Abstract: Text-guided multispectral object detection uses text semantics to guide semantic-aware cross-modal interaction between RGB and IR for more robust percep

safetyarxiv-cs-cv
14 Apr 2026
Safety

Budget-Aware Uncertainty for Radiotherapy Segmentation QA Using nnU-Net

DGX agent

arXiv:2604.11798v1 Announce Type: cross Abstract: Accurate delineation of the Clinical Target Volume (CTV) is essential for radiotherapy planning, yet remains time-consuming and difficult to assess, e

safetyarxiv-cs-ai
14 Apr 2026
Safety

C2F-Thinker: Coarse-to-Fine Reasoning with Hint-Guided Reinforcement Learning for Multimodal Sentiment Analysis

DGX agent

arXiv:2604.00013v2 Announce Type: replace-cross Abstract: Multimodal sentiment analysis aims to integrate textual, acoustic, and visual information for deep emotional understanding. Despite the progre

safetyarxiv-cs-ai
14 Apr 2026
Safety

CAGenMol: Condition-Aware Diffusion Language Model for Goal-Directed Molecular Generation

DGX agent

arXiv:2604.11483v1 Announce Type: new Abstract: Goal-directed molecular generation requires satisfying heterogeneous constraints such as protein--ligand compatibility and multi-objective drug-like pro

safetyarxiv-cs-lg
14 Apr 2026
Safety

Calibration Collapse Under Sycophancy Fine-Tuning: How Reward Hacking Breaks Uncertainty Quantification in LLMs

DGX agent

arXiv:2604.10585v1 Announce Type: cross Abstract: Modern large language models (LLMs) are increasingly fine-tuned via reinforcement learning from human feedback (RLHF) or related reward optimisation s

safetyarxiv-cs-ai
14 Apr 2026
Safety

China leaps ahead of the US in the race to control dangerously anthropomorphic AI.

DGX agent

China leaps ahead of the US in the race to control dangerously anthropomorphic AI. 🚨 BREAKING: China's new law on AI anthropomorphism has been officially enacted, and it is the world's STRICTEST law o

safetygary-marcus--x
14 Apr 2026
Safety

CID-TKG: Collaborative Historical Invariance and Evolutionary Dynamics Learning for Temporal Knowledge Graph Reasoning

DGX agent

arXiv:2604.09600v1 Announce Type: new Abstract: Temporal knowledge graph (TKG) reasoning aims to infer future facts at unseen timestamps from temporally evolving entities and relations. Despite recent

safetyarxiv-cs-ai
14 Apr 2026
Safety

CityGuard: Graph-Aware Private Descriptors for Bias-Resilient Identity Search Across Urban Cameras

DGX agent

arXiv:2602.18047v3 Announce Type: replace Abstract: City-scale person re-identification across distributed cameras must handle severe appearance changes from viewpoint, occlusion, and domain shift whi

safetyarxiv-cs-cv
14 Apr 2026
Safety

Claim2Vec: Embedding Fact-Check Claims for Multilingual Similarity and Clustering

DGX agent

arXiv:2604.09812v1 Announce Type: new Abstract: Recurrent claims present a major challenge for automated fact-checking systems designed to combat misinformation, especially in multilingual settings. W

safetyarxiv-cs-cl
14 Apr 2026
Safety

Closed-Form Concept Erasure via Double Projections

DGX agent

arXiv:2604.10032v1 Announce Type: cross Abstract: While modern generative models such as diffusion-based architectures have enabled impressive creative capabilities, they also raise important safety a

safetyarxiv-cs-ai
14 Apr 2026
Safety

ComSim: Building Scalable Real-World Robot Data Generation via Compositional Simulation

DGX agent

arXiv:2604.11386v1 Announce Type: cross Abstract: Recent advancements in foundational models, such as large language models and world models, have greatly enhanced the capabilities of robotics, enabli

safetyarxiv-cs-cv
14 Apr 2026
Safety

ConfigSpec: Profiling-Based Configuration Selection for Distributed Edge--Cloud Speculative LLM Serving

DGX agent

arXiv:2604.09722v1 Announce Type: cross Abstract: Speculative decoding enables collaborative Large Language Model (LLM) inference across cloud and edge by separating lightweight token drafting from he

safetyarxiv-cs-ai
14 Apr 2026
Safety

CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation

DGX agent

arXiv:2604.09746v1 Announce Type: cross Abstract: As large language models (LLMs) are increasingly deployed as autonomous agents, understanding how strategic behavior emerges in multi-agent environmen

safetyarxiv-cs-ai
14 Apr 2026
Safety

Consolidation or Adaptation? PRISM: Disentangling SFT and RL Data via Gradient Concentration

DGX agent

arXiv:2601.07224v2 Announce Type: replace Abstract: While Hybrid Supervised Fine-Tuning (SFT) followed by Reinforcement Learning (RL) has become the standard paradigm for training LLM agents, effectiv

safetyarxiv-cs-ai
14 Apr 2026
Safety

Context Matters: Vision-Based Depression Detection Comparing Classical and Deep Approaches

DGX agent

arXiv:2604.10344v1 Announce Type: new Abstract: The classical approach to detecting depression from vision emphasizes interpretable features, such as facial expression, and classifiers such as the Sup

safetyarxiv-cs-cv
14 Apr 2026
Safety

Contour Refinement using Discrete Diffusion in Low Data Regime

DGX agent

arXiv:2602.05880v2 Announce Type: replace Abstract: Boundary detection of irregular and translucent objects is an important problem with applications in medical imaging, environmental monitoring and m

safetyarxiv-cs-cv
14 Apr 2026
Safety

CoPS: Conditional Prompt Synthesis for Zero-Shot Anomaly Detection

DGX agent

arXiv:2508.03447v2 Announce Type: replace Abstract: Recently, large pre-trained vision-language models have shown remarkable performance in zero-shot anomaly detection (ZSAD). With fine-tuning on a si

safetyarxiv-cs-cv
14 Apr 2026
Safety

COSMIK-MPPI: Scaling Constrained Model Predictive Control to Collision Avoidance in Close-Proximity Dynamic Human Environments

DGX agent

arXiv:2604.10358v1 Announce Type: new Abstract: Ensuring safe physical interaction between torque-controlled manipulators and humans is essential for deploying robots in everyday environments. Model P

safetyarxiv-cs-ro
14 Apr 2026
Safety

Cost-optimal Sequential Testing via Doubly Robust Q-learning

DGX agent

arXiv:2604.11165v1 Announce Type: cross Abstract: Clinical decision-making often involves selecting tests that are costly, invasive, or time-consuming, motivating individualized, sequential strategies

safetyarxiv-cs-ai
14 Apr 2026
Safety

CoSToM:Causal-oriented Steering for Intrinsic Theory-of-Mind Alignment in Large Language Models

DGX agent

arXiv:2604.10031v1 Announce Type: cross Abstract: Theory of Mind (ToM), the ability to attribute mental states to others, is a hallmark of social intelligence. While large language models (LLMs) demon

safetyarxiv-cs-ai
14 Apr 2026
Safety

COXNet: Cross-Layer Fusion with Adaptive Alignment and Scale Integration for RGBT Tiny Object Detection

DGX agent

arXiv:2508.09533v2 Announce Type: replace-cross Abstract: Detecting tiny objects in multimodal Red-Green-Blue-Thermal (RGBT) imagery is a critical challenge in computer vision, particularly in surveil

safetyarxiv-cs-ai
14 Apr 2026
← Previous
1…247248249250251…263
Next →