AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
Safety

ProPACT: A Proactive AI-Driven Adaptive Collaborative Tutor for Pair Programming

DGX agent

arXiv:2605.02703v1 Announce Type: cross Abstract: Effective pair programming depends on coordination of attention, cognitive effort, and joint regulation over time, yet most adaptive learning systems

safetyarxiv-cs-lg
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

ProtoFair: Fair Self-Supervised Contrastive Learning via Pseudo-Counterfactual Pairs

DGX agent

arXiv:2605.01971v1 Announce Type: new Abstract: Self-supervised learning methods learn high-quality visual representations, yet recent studies show that these representations often capture demographic

safetyarxiv-cs-cv
5 May 2026
Safety

RA-CMF: Region-Adaptive Conditional MeanFlow for CT Image Reconstruction

DGX agent

arXiv:2605.00901v1 Announce Type: new Abstract: The use of CT imaging is important for screening, diagnosis, therapy planning, and prognosis of lung cancers. Unfortunately, due to differences in imagi

safetyarxiv-cs-cv
5 May 2026
Safety

Rationality Measurement and Theory for Reinforcement Learning Agents

DGX agent

arXiv:2602.04737v2 Announce Type: replace Abstract: This paper proposes a suite of rationality measures and associated theory for reinforcement learning agents, a property increasingly critical yet ra

safetyarxiv-cs-lg
5 May 2026
Safety

RECAP: Transparent Inference-Time Emotion Alignment for Medical Dialogue Systems

DGX agent

arXiv:2509.10746v3 Announce Type: replace Abstract: Large language models in healthcare often produce emotionally flat or opaque responses, failing to provide the transparent reasoning required for cl

safetyarxiv-cs-cl
5 May 2026
Safety

Registration-Free Learnable Multi-View Capture of Faces in Dense Semantic Correspondence

DGX agent

arXiv:2605.01450v1 Announce Type: new Abstract: Recent frameworks like ToFu and TEMPEH provide an automated alternative to classical registration pipelines by predicting 3D meshes in dense semantic co

safetyarxiv-cs-cv
5 May 2026
Safety

Remote Action Generation: Remote Control with Minimal Communication

DGX agent

arXiv:2605.01833v1 Announce Type: cross Abstract: We address the challenge of remote control where one or more actors, lacking direct reward access, are steered by a controller over a communication-co

safetyarxiv-cs-lg
5 May 2026
Safety

Representation learning from OCT images

DGX agent

arXiv:2605.02589v1 Announce Type: new Abstract: Optical Coherence Tomography (OCT) has become one of the most used imaging modality in ophthalmology. It provides high-resolution, non-invasive visualiz

safetyarxiv-cs-cv
5 May 2026
Safety

Research on Vision-Language Question Answering Models for Industrial Robots

DGX agent

arXiv:2605.01483v1 Announce Type: new Abstract: A hierarchical cross-modal fusion model is proposed for vision-language question answering (VLQA) in industrial robotics, targeting the challenges of se

safetyarxiv-cs-cv
5 May 2026
Safety

Rethinking Model Selection in VLM Through the Lens of Gromov-Wasserstein Distance

DGX agent

arXiv:2605.01325v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have enhanced traditional LLMs with visual capabilities through the integration of vision encoders. While recent works hav

safetyarxiv-cs-cv
5 May 2026
Safety

Retrieval-Guided Generation for Safer Histopathology Image Captioning

DGX agent

arXiv:2605.00893v1 Announce Type: new Abstract: Generative vision-language models can produce fluent medical image captions but remain prone to hallucination, over-specific diagnostic claims, and fact

safetyarxiv-cs-cv
5 May 2026
Safety

Rhythm: Learning Interactive Whole-Body Control for Dual Humanoids

DGX agent

arXiv:2603.02856v2 Announce Type: replace Abstract: Realizing interactive whole-body control for multi-humanoid systems is critical for unlocking complex collaborative capabilities in shared environme

safetyarxiv-cs-ro
5 May 2026
Safety

Robo3R: Enhancing Robotic Manipulation with Accurate Feed-Forward 3D Reconstruction

DGX agent

arXiv:2602.10101v2 Announce Type: replace Abstract: 3D spatial perception is fundamental to generalizable robotic manipulation, yet obtaining reliable, high-quality 3D geometry remains challenging. De

safetyarxiv-cs-ro
5 May 2026
Safety

Semantic-Contact Fields for Category-Level Generalizable Tactile Tool Manipulation

DGX agent

arXiv:2602.13833v2 Announce Type: replace Abstract: Generalizing tool manipulation requires both semantic planning and precise physical control. Modern generalist robot policies, such as Vision-Langua

safetyarxiv-cs-ro
5 May 2026
Safety

Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts

DGX agent

arXiv:2510.22628v2 Announce Type: replace-cross Abstract: This paper presents a real-time modular defense system named Sentra-Guard. The system detects and mitigates jailbreak and prompt injection att

safetyarxiv-cs-ai
5 May 2026
Safety

Separation Assurance between Heterogeneous Fleets of Small Unmanned Aerial Systems via Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.01041v1 Announce Type: cross Abstract: In the envisioned future dense urban airspace, multiple companies will operate heterogeneous fleets of small unmanned aerial systems (sUASs), where ea

safetyarxiv-cs-lg
5 May 2026
Safety

SIFT-VTON: Geometric Correspondence Supervision on Cross-Attention for Virtual Try-On

DGX agent

arXiv:2605.01296v1 Announce Type: new Abstract: Diffusion-based virtual try-on methods achieve photorealistic synthesis through cross-attention mechanisms that transfer garment features to target body

safetyarxiv-cs-cv
5 May 2026
Safety

SignVerse-2M: A Two-Million-Clip Pose-Native Universe of 25+ Sign Languages

DGX agent

arXiv:2605.01720v1 Announce Type: cross Abstract: Existing large-scale sign language resources typically provide supervision only at the level of raw video-text alignment and are often produced in lab

safetyarxiv-cs-cl
5 May 2026
Safety

Singular Bayesian Neural Networks

DGX agent

arXiv:2602.00387v3 Announce Type: replace-cross Abstract: Bayesian neural networks promise calibrated uncertainty but require O(mn) parameters for standard mean-field Gaussian posteriors. We argue thi

safetyarxiv-cs-lg
5 May 2026
Safety

Skills as Verifiable Artifacts: A Trust Schema and a Biconditional Correctness Criterion for Human-in-the-Loop Agent Runtimes

DGX agent

arXiv:2605.00424v1 Announce Type: cross Abstract: Agent skills -- structured packages of instructions, scripts, and references that augment a large language model (LLM) without modifying the model its

safetyarxiv-cs-ai
5 May 2026
Safety

So let me make sure I’ve got this correct: >Google gives Anthropic cash to pay Google for compute. >anthropic gives compute away for a loss,…

DGX agent

So let me make sure I’ve got this correct: >Google gives Anthropic cash to pay Google for compute. >anthropic gives compute away for a loss, and this discounted compute draws in large user bases to ca

safetygary-marcus--x
5 May 2026
Safety

Sonar-GPS Fusion for Seabed Mapping in Turbid Shallow Waters with an Autonomous Surface Vehicle

DGX agent

arXiv:2605.01949v1 Announce Type: cross Abstract: Accurate seabed mapping is essential for habitat monitoring and infrastructure inspection. In turbid, shallow coastal waters, such as shellfish aquacu

safetyarxiv-cs-cv
5 May 2026
Safety

Sources: the WH is mulling EOs to address security risks from advanced AI, including barring companies from 'interfering' with the government's use of models (Politico)

DGX agent

Politico: Sources: the WH is mulling EOs to address security risks from advanced AI, including barring companies from “interfering” with the government's use of models — A White House spokesperson sai

safetytechmeme
5 May 2026
Safety

SpectraDINO: Bridging the Spectral Gap in Vision Foundation Models via Lightweight Adapters

DGX agent

arXiv:2605.02258v1 Announce Type: new Abstract: Vision Foundation Models (VFMs) pretrained on large-scale RGB data have demonstrated remarkable representation quality, yet their applicability to multi

safetyarxiv-cs-cv
5 May 2026
Safety

SRA: Span Representation Alignment for Large Language Model Distillation

DGX agent

arXiv:2605.01205v1 Announce Type: new Abstract: Cross-Tokenizer Knowledge Distillation (CTKD) enables knowledge transfer between a large language model and a smaller student, even when they employ dif

safetyarxiv-cs-cl
5 May 2026
Model Releases

SRTJ: Self-Evolving Rule-Driven Training-Free LLM Jailbreaking

DGX agent

arXiv:2605.00974v1 Announce Type: cross Abstract: LLMs are increasingly equipped with safety alignment mechanisms, yet recent studies demonstrate that they remain vulnerable to jailbreaking attacks th

model-releasesarxiv-cs-cl
5 May 2026
Safety

StableMind: Source-Free Cross-Subject fMRI Decoding with Regularized Adaptation

DGX agent

arXiv:2605.02586v1 Announce Type: new Abstract: Existing cross-subject fMRI decoding methods typically train a model on multiple scanned subjects and then adapt it to a new subject using substantial p

safetyarxiv-cs-cv
5 May 2026
Safety

STABLEVAL: Disagreement-Aware and Stable Evaluation of AI Systems

DGX agent

arXiv:2605.02122v1 Announce Type: new Abstract: Human evaluation remains the primary standard for assessing modern AI systems, yet annotator disagreement, bias, and variability make system rankings fr

safetyarxiv-cs-lg
5 May 2026
Safety

SwiftPie: Lightning-fast Subject-driven Image Personalization via One step Diffusion

DGX agent

arXiv:2605.01510v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success in high-quality image synthesis, sparking interest in image-guided generation tasks such as subject-dr

safetyarxiv-cs-cv
5 May 2026
Safety

SynPAIN: A Synthetic Dataset of Pain and Non-Pain Facial Expressions

DGX agent

arXiv:2507.19673v3 Announce Type: replace Abstract: Accurate pain assessment in patients with limited ability to communicate, such as older adults with severe dementia, represents a critical healthcar

safetyarxiv-cs-cv
5 May 2026
Safety

Task-Related Token Compression in Multimodal Large Language Models from an Explainability Perspective

DGX agent

arXiv:2506.01097v2 Announce Type: replace Abstract: Existing Multimodal Large Language Models (MLLMs) process a large number of visual tokens, leading to significant computational costs and inefficien

safetyarxiv-cs-cv
5 May 2026
Safety

Thank you @UN Secretary General @antonioguterres for an excellent discussion on the current state of AI capabilities, and the urgent need fo…

DGX agent

Thank you @UN Secretary General @antonioguterres for an excellent discussion on the current state of AI capabilities, and the urgent need for robust global governance. International collaboration is c

safetyyoshua-bengio--x
5 May 2026
Safety

The Case for ESM3 as a General-Purpose AI Model with Systemic Risk Under the EU AI Act

DGX agent

arXiv:2605.01611v1 Announce Type: cross Abstract: Due to ambiguity in the wording of the EU AI Act, we examine the question of to what extent frontier biological foundation models such as ESM3 are sub

safetyarxiv-cs-lg
5 May 2026
Safety

The Geometric Inductive Bias of Grokking: Bypassing Phase Transitions via Architectural Topology

DGX agent

arXiv:2603.05228v3 Announce Type: replace Abstract: Mechanistic interpretability typically relies on post-hoc analysis of trained networks. We instead adopt an interventional approach: testing hypothe

safetyarxiv-cs-lg
5 May 2026
Safety

The Geometric Mechanics of Contrastive Representation Learning: Alignment Potentials, Entropic Dispersion, and Cross-modal Divergence

DGX agent

arXiv:2601.19597v3 Announce Type: replace Abstract: While InfoNCE underlies modern contrastive learning, its geometric mechanisms remain under-characterized beyond the canonical alignment--uniformity

safetyarxiv-cs-lg
5 May 2026
Safety

The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling

DGX agent

arXiv:2605.02427v1 Announce Type: cross Abstract: A recurring pattern in 'reasoning without training' is that base LLMs already assign non-trivial probability mass to correct multi-step solutions; the

safetyarxiv-cs-lg
5 May 2026
Safety

The Partial Testimony of Logs: Evaluation of Language Model Generation under Confounded Model Choice

DGX agent

arXiv:2605.01311v1 Announce Type: new Abstract: Offline evaluation of language models from usage logs is biased when model choice is confounded: the same user-side factors that influence which model i

safetyarxiv-cs-lg
5 May 2026
Safety

The Pragmatic Frames of Spurious Correlations in Machine Learning: Interpreting How and Why They Matter

DGX agent

arXiv:2411.04696v5 Announce Type: replace Abstract: Learning correlations from data forms the foundation of today's machine learning (ML) and artificial intelligence research. While contemporary metho

safetyarxiv-cs-lg
5 May 2026
Safety

Today, the CEO of @coinbase announced a 14% headcount reduction. Let's talk about why. He cites AI as a core reason for the shift, alongside…

DGX agent

Today, the CEO of @coinbase announced a 14% headcount reduction. Let's talk about why. He cites AI as a core reason for the shift, alongside market dynamics - though we don't know whether each of thos

safetyallie-k--miller--x
5 May 2026
Safety

Too early to tell how the trial ends, but if Elon does win, the takeaway might be in part that Greg Brockman made two immense blunders: 1. W…

DGX agent

Too early to tell how the trial ends, but if Elon does win, the takeaway might be in part that Greg Brockman made two immense blunders: 1. Writing down all that stuff in his diary 2. Blowing off Elon’

safetygary-marcus--x
5 May 2026
Safety

Topological Neural Tangent Kernel

DGX agent

arXiv:2605.01110v1 Announce Type: new Abstract: Graph neural tangent kernels give a principled infinite-width theory for graph neural networks, but inherit a basic limitation of graph models: they see

safetyarxiv-cs-lg
5 May 2026
Safety

Toward a Scientific Discovery Engine for Weather and Climate Data: A Visual Analytics Workbench for Embedding-Based Exploration

DGX agent

arXiv:2605.00972v1 Announce Type: cross Abstract: Earth system science is producing increasingly large, high-dimensional datasets from physics based Earth system models to AI-based weather and climate

safetyarxiv-cs-cv
5 May 2026
Safety

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning

DGX agent

arXiv:2605.01663v1 Announce Type: new Abstract: We propose Flow-Anchored Noise-conditioned Q-Learning (FAN), a highly efficient and high-performing offline reinforcement learning (RL) algorithm. Recen

safetyarxiv-cs-lg
5 May 2026
Safety

Towards Improving Speaker Distance Estimation through Generative Impulse Response Augmentation

DGX agent

arXiv:2605.00721v1 Announce Type: cross Abstract: The Room Acoustics and Speaker Distance Estimation (SDE) Challenge at ICASSP 2025 explores the effectiveness of augmented room impulse response (RIR)

safetyarxiv-cs-ai
5 May 2026
Safety

Training Non-Differentiable Networks via Optimal Transport

DGX agent

arXiv:2605.01928v1 Announce Type: new Abstract: Neural networks increasingly embed non-differentiable components (spiking neurons, quantized layers, discrete routing, blackbox simulators, etc.) where

safetyarxiv-cs-lg
5 May 2026
Safety

TRAP: Tail-aware Ranking Attack for World-Model Planning

DGX agent

arXiv:2605.01950v1 Announce Type: new Abstract: World models enable long-horizon planning by internally generating and evaluating imagined trajectories, making them a promising foundation for generali

safetyarxiv-cs-lg
5 May 2026
Safety

TUR-DPO: Topology- and Uncertainty-Aware Direct Preference Optimization

DGX agent

arXiv:2605.00224v1 Announce Type: new Abstract: Aligning large language models (LLMs) with human preferences is commonly done via reinforcement learning from human feedback (RLHF) with Proximal Policy

safetyarxiv-cs-ai
5 May 2026
Safety

Ultrasound Vision-Language Alignment via Contrastive Learning

DGX agent

arXiv:2605.02126v1 Announce Type: new Abstract: Ultrasound foundation models have achieved strong performance on structured prediction tasks but remain exclusively vision-based, limiting zero-shot and

safetyarxiv-cs-cv
5 May 2026
← Previous
1…233234235236237…300
Next →