AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

StableMind: Source-Free Cross-Subject fMRI Decoding with Regularized Adaptation

DGX agent

arXiv:2605.02586v1 Announce Type: new Abstract: Existing cross-subject fMRI decoding methods typically train a model on multiple scanned subjects and then adapt it to a new subject using substantial p

safetyarxiv-cs-cv
5 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

STABLEVAL: Disagreement-Aware and Stable Evaluation of AI Systems

DGX agent

arXiv:2605.02122v1 Announce Type: new Abstract: Human evaluation remains the primary standard for assessing modern AI systems, yet annotator disagreement, bias, and variability make system rankings fr

safetyarxiv-cs-lg
5 May 2026
Safety

SwiftPie: Lightning-fast Subject-driven Image Personalization via One step Diffusion

DGX agent

arXiv:2605.01510v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success in high-quality image synthesis, sparking interest in image-guided generation tasks such as subject-dr

safetyarxiv-cs-cv
5 May 2026
Safety

SynPAIN: A Synthetic Dataset of Pain and Non-Pain Facial Expressions

DGX agent

arXiv:2507.19673v3 Announce Type: replace Abstract: Accurate pain assessment in patients with limited ability to communicate, such as older adults with severe dementia, represents a critical healthcar

safetyarxiv-cs-cv
5 May 2026
Safety

Task-Related Token Compression in Multimodal Large Language Models from an Explainability Perspective

DGX agent

arXiv:2506.01097v2 Announce Type: replace Abstract: Existing Multimodal Large Language Models (MLLMs) process a large number of visual tokens, leading to significant computational costs and inefficien

safetyarxiv-cs-cv
5 May 2026
Safety

The Case for ESM3 as a General-Purpose AI Model with Systemic Risk Under the EU AI Act

DGX agent

arXiv:2605.01611v1 Announce Type: cross Abstract: Due to ambiguity in the wording of the EU AI Act, we examine the question of to what extent frontier biological foundation models such as ESM3 are sub

safetyarxiv-cs-lg
5 May 2026
Safety

The Geometric Inductive Bias of Grokking: Bypassing Phase Transitions via Architectural Topology

DGX agent

arXiv:2603.05228v3 Announce Type: replace Abstract: Mechanistic interpretability typically relies on post-hoc analysis of trained networks. We instead adopt an interventional approach: testing hypothe

safetyarxiv-cs-lg
5 May 2026
Safety

The Geometric Mechanics of Contrastive Representation Learning: Alignment Potentials, Entropic Dispersion, and Cross-modal Divergence

DGX agent

arXiv:2601.19597v3 Announce Type: replace Abstract: While InfoNCE underlies modern contrastive learning, its geometric mechanisms remain under-characterized beyond the canonical alignment--uniformity

safetyarxiv-cs-lg
5 May 2026
Safety

The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling

DGX agent

arXiv:2605.02427v1 Announce Type: cross Abstract: A recurring pattern in 'reasoning without training' is that base LLMs already assign non-trivial probability mass to correct multi-step solutions; the

safetyarxiv-cs-lg
5 May 2026
Safety

The Partial Testimony of Logs: Evaluation of Language Model Generation under Confounded Model Choice

DGX agent

arXiv:2605.01311v1 Announce Type: new Abstract: Offline evaluation of language models from usage logs is biased when model choice is confounded: the same user-side factors that influence which model i

safetyarxiv-cs-lg
5 May 2026
Safety

The Pragmatic Frames of Spurious Correlations in Machine Learning: Interpreting How and Why They Matter

DGX agent

arXiv:2411.04696v5 Announce Type: replace Abstract: Learning correlations from data forms the foundation of today's machine learning (ML) and artificial intelligence research. While contemporary metho

safetyarxiv-cs-lg
5 May 2026
Safety

Topological Neural Tangent Kernel

DGX agent

arXiv:2605.01110v1 Announce Type: new Abstract: Graph neural tangent kernels give a principled infinite-width theory for graph neural networks, but inherit a basic limitation of graph models: they see

safetyarxiv-cs-lg
5 May 2026
Safety

Toward a Scientific Discovery Engine for Weather and Climate Data: A Visual Analytics Workbench for Embedding-Based Exploration

DGX agent

arXiv:2605.00972v1 Announce Type: cross Abstract: Earth system science is producing increasingly large, high-dimensional datasets from physics based Earth system models to AI-based weather and climate

safetyarxiv-cs-cv
5 May 2026
Safety

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning

DGX agent

arXiv:2605.01663v1 Announce Type: new Abstract: We propose Flow-Anchored Noise-conditioned Q-Learning (FAN), a highly efficient and high-performing offline reinforcement learning (RL) algorithm. Recen

safetyarxiv-cs-lg
5 May 2026
Safety

Towards Improving Speaker Distance Estimation through Generative Impulse Response Augmentation

DGX agent

arXiv:2605.00721v1 Announce Type: cross Abstract: The Room Acoustics and Speaker Distance Estimation (SDE) Challenge at ICASSP 2025 explores the effectiveness of augmented room impulse response (RIR)

safetyarxiv-cs-ai
5 May 2026
Safety

Training Non-Differentiable Networks via Optimal Transport

DGX agent

arXiv:2605.01928v1 Announce Type: new Abstract: Neural networks increasingly embed non-differentiable components (spiking neurons, quantized layers, discrete routing, blackbox simulators, etc.) where

safetyarxiv-cs-lg
5 May 2026
Safety

TRAP: Tail-aware Ranking Attack for World-Model Planning

DGX agent

arXiv:2605.01950v1 Announce Type: new Abstract: World models enable long-horizon planning by internally generating and evaluating imagined trajectories, making them a promising foundation for generali

safetyarxiv-cs-lg
5 May 2026
Safety

TUR-DPO: Topology- and Uncertainty-Aware Direct Preference Optimization

DGX agent

arXiv:2605.00224v1 Announce Type: new Abstract: Aligning large language models (LLMs) with human preferences is commonly done via reinforcement learning from human feedback (RLHF) with Proximal Policy

safetyarxiv-cs-ai
5 May 2026
Safety

Ultrasound Vision-Language Alignment via Contrastive Learning

DGX agent

arXiv:2605.02126v1 Announce Type: new Abstract: Ultrasound foundation models have achieved strong performance on structured prediction tasks but remain exclusively vision-based, limiting zero-shot and

safetyarxiv-cs-cv
5 May 2026
Safety

Understanding Adversarial Imitation Learning in Small Sample Regime: A Stage-coupled Analysis

DGX agent

arXiv:2208.01899v2 Announce Type: replace Abstract: Imitation learning learns a policy from expert trajectories. While the expert data is believed to be crucial for imitation quality, it was found tha

safetyarxiv-cs-lg
5 May 2026
Safety

Unified Map Prior Encoder for Mapping and Planning

DGX agent

arXiv:2605.02762v1 Announce Type: new Abstract: Online mapping and end-to-end (E2E) planning in autonomous driving remain largely sensor-centric, leaving rich map priors, including HD/SD vector maps,

safetyarxiv-cs-cv
5 May 2026
Safety

Unsupervised Machine Learning for Detecting Structural Anomalies in European Regional Statistics

DGX agent

arXiv:2605.02884v1 Announce Type: new Abstract: Ensuring the coherence of regional socio-economic statistics is a central task for national statistical institutes. Traditional validation tools, such a

safetyarxiv-cs-lg
5 May 2026
Safety

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation

DGX agent

arXiv:2601.23286v2 Announce Type: replace Abstract: While recent video diffusion models (VDMs) produce visually impressive results, they fundamentally struggle to maintain 3D structural consistency, o

safetyarxiv-cs-cv
5 May 2026
Safety

VOFA: Visual Object Goal Pushing with Force-Adaptive Control for Humanoids

DGX agent

arXiv:2605.01518v1 Announce Type: new Abstract: The ability to push large objects in a goal-directed manner using onboard egocentric perception is an essential skill for humanoid robots to perform com

safetyarxiv-cs-ro
5 May 2026
Safety

When Attention Collapses: Residual Evidence Modeling for Compositional Inference

DGX agent

arXiv:2605.02323v1 Announce Type: new Abstract: Compositional inference - the decomposition of observations into an unknown number of latent components - is central to perception and scientific data a

safetyarxiv-cs-lg
5 May 2026
Safety

Who Decides What Is Harmful? Content Moderation Policy Through A Multi-Agent Personalised Inference Framework

DGX agent

arXiv:2605.01416v1 Announce Type: cross Abstract: The increasing scale and complexity of online platforms raises critical policy questions around harmful content, digital well-being, and user autonomy

safetyarxiv-cs-cl
5 May 2026
Safety

Zero-Shot Adaptation of Behavioral Foundation Models to Unseen Dynamics

DGX agent

arXiv:2505.13150v2 Announce Type: replace Abstract: Behavioral Foundation Models (BFMs) proved successful in producing policies for arbitrary tasks in a zero-shot manner, requiring no test-time traini

safetyarxiv-cs-lg
5 May 2026
Safety

A Policy-Driven DRL Framework for System-Level Tradeoff Control in NR-U/Wi-Fi Coexistence

DGX agent

arXiv:2605.00457v1 Announce Type: cross Abstract: The coexistence of NR-U and Wi-Fi in unlicensed spectrum introduces a system-level resource coordination problem, where heterogeneous channel access m

safetyarxiv-cs-lg
4 May 2026
Safety

A unified perspective on fine-tuning and sampling with diffusion and flow models

DGX agent

arXiv:2605.00229v1 Announce Type: cross Abstract: We study the problem of training diffusion and flow generative models to sample from target distributions defined by an exponential tilting of a base

safetyarxiv-cs-lg
4 May 2026
Safety

Adaptive Equilibrium: Dynamic Weighting Framework for Generalized Interruption of DeepFake Models

DGX agent

arXiv:2605.00443v1 Announce Type: cross Abstract: The advancement of generalized deepfake disruption is constrained by the interruption imbalance, a fundamental bottleneck inherent to the generation o

safetyarxiv-cs-cv
4 May 2026
Safety

Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines

DGX agent

arXiv:2605.00410v1 Announce Type: new Abstract: A multi-agent pipeline with N agents typically issues N LLM calls per run. Merging agents into fewer calls (compound execution) promises token savings,

safetyarxiv-cs-cl
4 May 2026
Safety

Beyond Prompt-Induced Lies: Investigating LLM Deception on Benign Prompts

DGX agent

arXiv:2508.06361v4 Announce Type: replace Abstract: Large Language Models (LLMs) are widely deployed in reasoning, planning, and decision-making tasks, making their trustworthiness critical. A signifi

safetyarxiv-cs-lg
4 May 2026
Safety

Bias in Large Language Models: Origin, Evaluation, and Mitigation

DGX agent

arXiv:2411.10915v2 Announce Type: replace Abstract: Large Language Models (LLMs) have revolutionized natural language processing, but their susceptibility to biases poses significant challenges. This

safetyarxiv-cs-cl
4 May 2026
Safety

BlenderRAG: High-Fidelity 3D Object Generation via Retrieval-Augmented Code Synthesis

DGX agent

arXiv:2605.00632v1 Announce Type: new Abstract: Automatic generation of executable Blender code from natural language remains challenging, with state-of-the-art LLMs producing frequent syntactic error

safetyarxiv-cs-cv
4 May 2026
Safety

BOLT: Online Lightweight Adaptation for Preparation-Free Heterogeneous Cooperative Perception

DGX agent

arXiv:2605.00405v1 Announce Type: new Abstract: Most existing heterogeneous cooperative perception methods depend on prior preparation like offline joint training or tailored collaborator-model adapta

safetyarxiv-cs-cv
4 May 2026
Safety

Can Small Language Models Handle Context-Summarized Multi-Turn Customer-Service QA? A Synthetic Data-Driven Comparative Evaluation

DGX agent

arXiv:2602.00665v3 Announce Type: replace Abstract: Customer-service question answering (QA) systems increasingly rely on conversational language understanding. While Large Language Models (LLMs) achi

safetyarxiv-cs-cl
4 May 2026
Safety

Data Deletion Can Help in Adaptive RL

DGX agent

arXiv:2605.00298v1 Announce Type: new Abstract: Deploying reinforcement learning policies in the real world requires adapting to time-varying environments. We study this problem in the contextual Mark

safetyarxiv-cs-lg
4 May 2026
Safety

Debate-Enhanced Pseudo Labeling and Frequency-Aware Progressive Debiasing for Weakly-Supervised Camouflaged Object Detection with Scribble Annotations

DGX agent

arXiv:2512.20260v5 Announce Type: replace Abstract: Weakly-Supervised Camouflaged Object Detection (WSCOD) aims to locate and segment objects that are visually concealed within their surrounding scene

safetyarxiv-cs-cv
4 May 2026
Safety

Decentralized Proximal Stochastic Gradient Langevin Dynamics

DGX agent

arXiv:2605.00723v1 Announce Type: cross Abstract: We propose Decentralized Proximal Stochastic Gradient Langevin Dynamics (DE-PSGLD), a decentralized Markov chain Monte Carlo (MCMC) algorithm for samp

safetyarxiv-cs-lg
4 May 2026
Model Releases

Do Open-Loop Metrics Predict Closed-Loop Driving? A Cross-Benchmark Correlation Study of NAVSIM and Bench2Drive

DGX agent

arXiv:2605.00066v1 Announce Type: new Abstract: Open-loop evaluation offers fast, reproducible assessment of autonomous driving planners, but its ability to predict real closed-loop driving performanc

model-releasesarxiv-cs-ro
4 May 2026
Safety

Estimating LLM Grading Ability and Response Difficulty in Automatic Short Answer Grading via Item Response Theory

DGX agent

arXiv:2605.00238v1 Announce Type: new Abstract: Automated short answer grading (ASAG) with large language models (LLMs) is commonly evaluated with aggregate metrics such as macro-F1 and Cohen's kappa.

safetyarxiv-cs-cl
4 May 2026
Safety

Fair Dataset Distillation via Cross-Group Barycenter Alignment

DGX agent

arXiv:2605.00185v1 Announce Type: new Abstract: Dataset Distillation aims to compress a large dataset into a small synthetic one while maintaining predictive performance. We show that as different dem

safetyarxiv-cs-lg
4 May 2026
Safety

Fairness of Classifiers in the Presence of Constraints between Features

DGX agent

arXiv:2605.00592v1 Announce Type: new Abstract: In Machine Learning, an accepted definition of fairness of a decision taken by a classifier is that it should not depend on protected features, such as

safetyarxiv-cs-lg
4 May 2026
Safety

Foundation AI Models for Aerosol Optical Depth Estimation from PACE Satellite Data

DGX agent

arXiv:2605.00678v1 Announce Type: new Abstract: Aerosol Optical Depth (AOD) retrieval is essential for Earth observation, supporting applications from air quality monitoring to climate studies. Conven

safetyarxiv-cs-cv
4 May 2026
Safety

FreeRet: MLLMs as Training-Free Retrievers

DGX agent

arXiv:2509.24621v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are emerging as versatile foundations for mixed-modality retrieval. Yet, they often require heavy post-hoc

safetyarxiv-cs-cv
4 May 2026
Safety

How Language Models Process Out-of-Distribution Inputs: A Two-Pathway Framework

DGX agent

arXiv:2605.00269v1 Announce Type: new Abstract: Recent white-box OOD detection methods for LLMs -- including CED, RAUQ, and WildGuard confidence scores -- appear effective, but we show they are struct

safetyarxiv-cs-cl
4 May 2026
Safety

HyCOP: Hybrid Composition Operators for Interpretable Learning of PDEs

DGX agent

arXiv:2605.00820v1 Announce Type: cross Abstract: We introduce HyCOP, a modular framework that learns parametric PDE solution operators by composing simple modules (advection, diffusion, learned closu

safetyarxiv-cs-lg
4 May 2026
Safety

InpaintSLat: Inpainting Structured 3D Latents via Initial Noise Optimization

DGX agent

arXiv:2605.00664v1 Announce Type: new Abstract: We present a training-free approach for controllable 3D inpainting based on initial noise optimization. In the structured 3D latent diffusion framework,

safetyarxiv-cs-cv
4 May 2026
← Previous
1…202203204205206…257
Next →