AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
19 May 2026

Simple Approximation and Derivative Free Inference-Time Scaling for Diffusion Models via Sequential Monte Carlo on Path Measures

SafetyDGX agent

arXiv:2605.17850v1 Announce Type: cross Abstract: iffusion-based generative models increasingly rely on inference-time guidance, adding a drift term or reweighting mixture of experts, to improve sampl

Single-Sample Black-Box Membership Inference Attack against Vision-Language Models via Cross-modal Semantic Alignment

SafetyDGX agent

arXiv:2605.17341v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable success, yet their reliance on massive datasets and unintended memorization of training data ra

SNLP: Layer-Parallel Inference via Structured Newton Corrections

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.17842v1 Announce Type: new Abstract: Autoregressive language models execute Transformer layers sequentially, creating a latency bottleneck that is not removed by conventional tensor or pipe

Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration

SafetyDGX agent

arXiv:2605.17423v1 Announce Type: new Abstract: We study series-level cinematic remaking, a long-horizon video-to-video generation problem that localizes full episodes or films via stylization or acto

Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias

SafetyDGX agent

arXiv:2509.22061v2 Announce Type: replace-cross Abstract: Speech Continuation (SC) is the task of generating a coherent extension of a spoken prompt while preserving both semantic context and speaker

SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language Reasoning

SafetyDGX agent

arXiv:2510.16416v4 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have shown remarkable abilities by integrating large language models with visual inputs. However, they often fai

Stable and Near-Reversible Diffusion ODE Solvers for Image Editing

SafetyDGX agent

arXiv:2605.16399v1 Announce Type: new Abstract: The inversion of diffusion models plays a central role in image editing. Algebraically reversible ODE solvers provide an appealing approach to diffusion

Stable Routing for Mixture-of-Experts in Class-Incremental Learning

SafetyDGX agent

arXiv:2605.17571v1 Announce Type: new Abstract: Class-incremental learning (CIL) requires models to learn new classes sequentially while preserving prior knowledge. Recently, approaches that combine p

StableHand: Quality-Aware Flow Matching for World-Space Dual-Hand Motion Estimation from Egocentric Video

SafetyDGX agent

arXiv:2605.18553v1 Announce Type: cross Abstract: Recovering world space 4D motion of two interacting hands from egocentric video is a fundamental capability for supervising robot policy learning, whe

State-Conditional Adversarial Learning: An Off-Policy Visual Domain Transfer Method for End-to-End Imitation Learning

SafetyDGX agent

arXiv:2512.05335v3 Announce Type: replace Abstract: We study visual domain transfer for end-to-end imitation learning in a realistic and challenging setting where target-domain data are strictly off-p

State Contamination in Memory-Augmented LLM Agents

SafetyDGX agent

arXiv:2605.16746v1 Announce Type: new Abstract: LLM agents increasingly rely on persistent state, including transcripts, summaries, retrieved context, and memory buffers, to support long-horizon inter

Stochastic Penalty-Barrier Methods for Constrained Machine Learning

SafetyDGX agent

arXiv:2605.18618v1 Announce Type: cross Abstract: Constrained machine learning enables fairness-aware training, physics-informed neural networks, and integration of symbolic domain knowledge into stat

Stop When Reasoning Converges: Semantic-Preserving Early Exit for Reasoning Models

SafetyDGX agent

arXiv:2605.17672v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance by generating long chains of thought (CoT), but often overthink, continuing to reason after a s

Structure-Aware Masking for Protein Representation Learning

SafetyDGX agent

arXiv:2605.16581v1 Announce Type: new Abstract: Masked language modeling (MLM) is the standard objective for training protein language models, typically implemented by randomly masking individual resi

Superb conversation between @bgreene and @GaryMarcus Very happy to get my confirmation bias pampered by mentions of interpolation vs extrapo…

SafetyDGX agent

Superb conversation between @bgreene and @GaryMarcus Very happy to get my confirmation bias pampered by mentions of interpolation vs extrapolation, and how having a better physics understanding might

SuReNav: Superpixel Graph-based Constraint Relaxation for Navigation in Over-constrained Environments

SafetyDGX agent

arXiv:2602.06807v2 Announce Type: replace-cross Abstract: We address the over-constrained planning problem in semi-static environments. The planning objective is to find a best-effort solution that av

SURGE: Approximation-free Training Free Particle Filter for Diffusion Surrogate

SafetyDGX agent

arXiv:2605.18745v1 Announce Type: cross Abstract: Diffusion-based generative models increasingly rely on inference-time guidance, adding a drift term or reweighting mixture of experts, to improve samp

Surgical Post-Training: Proximal On-Policy Distillation for Reasoning with Knowledge Retention

SafetyDGX agent

arXiv:2603.01683v2 Announce Type: replace-cross Abstract: Injecting new reasoning knowledge into Large Language Models (LLMs) via post-training often induces catastrophic forgetting. Recent studies em

SutureFormer: Learning Surgical Trajectories via Goal-conditioned Offline RL in Pixel Space

SafetyDGX agent

arXiv:2603.26720v2 Announce Type: replace-cross Abstract: Predicting surgical needle trajectories from endoscopic video is critical for robot-assisted suturing, enabling anticipatory planning, real-ti

SVL: Spike-based Vision-language Pretraining for Efficient 3D Open-world Understanding

SafetyDGX agent

arXiv:2505.17674v2 Announce Type: replace Abstract: Spiking Neural Networks (SNNs) provide an energy-efficient way to extract 3D spatio-temporal features. However, existing SNNs still exhibit a signif

T-FIX: Text-Based Explanations with Features Interpretable to eXperts

SafetyDGX agent

arXiv:2511.04070v3 Announce Type: replace Abstract: As LLMs are deployed in knowledge-intensive settings (e.g., surgery, astronomy, therapy), users are often domain experts who expect not just answers

TacSE3: Equivariant SE(3) Motion Estimation from Low-Texture Visuotactile Images for In-Gripper Tracking and Compensation

SafetyDGX agent

arXiv:2605.17929v1 Announce Type: new Abstract: Robotic in-hand manipulation requires reliable object-motion tracking under frequent visual occlusion, yet low-texture visuotactile images provide few s

Taming 'Zombie'' Agents: A Markov State-Aware Framework for Resilient Multi-Agent Evolution

SafetyDGX agent

arXiv:2605.17348v1 Announce Type: new Abstract: Recent advancements in LLM-based multi-agent systems have demonstrated remarkable collaborative capabilities across complex tasks. To improve overall ef

TClone: Low-Latency Forking of Live GUI Environments for Computer-Use Agents

SafetyDGX agent

arXiv:2605.17320v1 Announce Type: cross Abstract: Computer-use agents increasingly operate inside live personal workspaces, where their actions can modify files, applications, GUI state, credentials,

Temporal Task Diversity: Inductive Biases Under Non-Stationarity in Synthetic Sequence Modelling

SafetyDGX agent

arXiv:2605.18281v1 Announce Type: new Abstract: Modern deep learning science often assumes that neural networks learn from a fixed data distribution. However, many practically important learning probl

terrific deep dive on Mythos from @cloudflare

SafetyDGX agent

terrific deep dive on Mythos from @cloudflare Cloudflare's security team spent the last few weeks testing Anthropic's Mythos against fifty of our own repositories. What we learned about offensive AI,

The Alien Space of Science: Sampling Coherent but Cognitively Unavailable Research Directions

SafetyDGX agent

arXiv:2603.01092v2 Announce Type: replace Abstract: Scientific discovery is constrained not only by what is true, but by what is cognitively available to the researchers currently exploring a field. M

The Bayesian Geometry of Transformer Attention

SafetyDGX agent

arXiv:2512.22471v5 Announce Type: replace-cross Abstract: Transformers often appear to perform Bayesian reasoning in context, but verifying this rigorously has been impossible: natural data lack analy

The Capability Paradox: How Smarter Auditors Make Multi-Agent Systems Less Secure

SafetyDGX agent

arXiv:2605.17480v1 Announce Type: new Abstract: Multi-agent systems extend large language models (LLMs) by decomposing tasks among specialized agents, but their distributed decision process creates ne

The Hidden Cost of Contextual Sycophancy: an AI Literacy Intervention in Human-AI Collaboration

SafetyDGX agent

arXiv:2605.18372v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in educational settings as interactive tools for collaboration. However, their tendency toward syco

The Impact of AI Search on the Online Content Ecosystem: Evidence from Google and Reddit

SafetyDGX agent

arXiv:2605.16428v1 Announce Type: cross Abstract: Search engines traditionally complement online content platforms by directing users seeking information to external websites. The emergence of generat

The Laplacian Keyboard: Beyond the Linear Span

SafetyDGX agent

arXiv:2602.07730v2 Announce Type: replace-cross Abstract: Across scientific disciplines, Laplacian eigenvectors serve as a fundamental basis for simplifying complex systems, from signal processing to

“the uncritical adoption of AI in science is alarming”

SafetyDGX agent

“the uncritical adoption of AI in science is alarming” Artificial intelligence is rapidly accelerating scientific output, but risks narrowing inquiry, weakening judgement and undermining how scientist

Toward Robust Multilingual Adaptation of LLMs for Low-Resource Languages

SafetyDGX agent

arXiv:2510.14466v3 Announce Type: replace-cross Abstract: Large language models (LLMs) continue to struggle with low-resource languages, primarily due to limited training data, translation noise, and

Towards Sustainable Growth: A Multi-Value-Aware Retrieval Framework for E-Commerce Search

SafetyDGX agent

arXiv:2605.17994v1 Announce Type: cross Abstract: New item growth is critical for maintaining a healthy ecosystem in large-scale e-commerce platforms. However, existing systems tend to prioritize pres

Towards Universal Physical Adversarial Attacks via a Joint Multi-Objective and Multi-Model Optimization Framework

SafetyDGX agent

arXiv:2605.17772v1 Announce Type: new Abstract: Physical adversarial attacks often overfit single surrogate models and optimization objectives. While ensemble attacks can mitigate this, existing metho

Transformer-Based MCS Prediction for 5G Multicast-Broadcast Services (MBS)

SafetyDGX agent

arXiv:2605.16735v1 Announce Type: cross Abstract: The deployment of 5G Multicast-Broadcast Services (MBS) is emerging as a critical technology for spectral-efficient UHD content delivery and serving a

Uncertainty-Calibrated Recommendations for Low-Active Users

SafetyDGX agent

arXiv:2605.17788v1 Announce Type: cross Abstract: A fundamental challenge in recommender systems is balancing reliability for Low-Active Users (LAUs) with diversity for High-Active Users (HAUs). The k

Uncertainty Reliability Under Domain Shift: An Investigation for Data-Driven Blood Pressure Estimation in Photoplethysmography

SafetyDGX agent

arXiv:2605.18008v1 Announce Type: new Abstract: Uncertainty quantification (UQ) is critical for safety-critical domains like healthcare, yet it is rarely evaluated under realistic out-of-distribution

UniAlign: A Model-Agnostic Framework for Robust Network Traffic Classification under Distribution Shifts

SafetyDGX agent

arXiv:2605.17575v1 Announce Type: cross Abstract: Network traffic classification (NTC) models often suffer severe performance degradation when deployed in real-world environments due to distribution s

Unified Walking, Running, and Recovery for Humanoids via State-Dependent Adversarial Motion Priors

SafetyDGX agent

arXiv:2605.18611v1 Announce Type: new Abstract: We propose a unified reinforcement learning framework that enables a single policy to perform walking, running, and fall recovery on the Unitree G1 huma

Unifying Contrastive and Generative Objectives for Visual Understanding and Text-to-Image Generation

SafetyDGX agent

arXiv:2603.02667v2 Announce Type: replace Abstract: Unifying text-image contrastive learning and text-to-image (T2I) generation in a single end-to-end model is challenging because the two objectives d

Universal Pose Pretraining for Generalizable Vision-Language-Action Policies

SafetyDGX agent

arXiv:2602.19710v2 Announce Type: replace Abstract: Existing Vision-Language-Action (VLA) models often suffer from feature collapse and low training efficiency because they entangle high-level percept

Unlearning Isn't Deletion: Investigating Reversibility of Machine Unlearning in LLMs

SafetyDGX agent

arXiv:2505.16831v3 Announce Type: replace-cross Abstract: Unlearning in large language models (LLMs) aims to remove specified data, but its efficacy is typically assessed with task-level metrics like

Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2602.22801v2 Announce Type: replace-cross Abstract: Diffusion models have become a popular choice for decision-making tasks in robotics, and more recently, are also being considered for solving

Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning

SafetyDGX agent

arXiv:2510.02590v2 Announce Type: replace Abstract: The use of target networks is a popular approach for estimating value functions in deep Reinforcement Learning (RL). While effective, the target net

Video Reconstruction using Diffusion-based Image-to-Video Generation with Trajectory Guidance

SafetyDGX agent

arXiv:2605.16420v1 Announce Type: new Abstract: This paper addresses the problem of reconstructing missing or dropped frames in top-down drone video of autonomous surface vehicles performing structure

View-Aware Semantic Alignment for Aerial-Ground Person Re-Identification

SafetyDGX agent

arXiv:2605.18192v1 Announce Type: new Abstract: Aerial-Ground Person Re-Identification (AGPReID) remains highly challenging due to drastic viewpoint variations between drones and fixed cameras. Existi

Vision Transformer-Conditioned UNet for Domain-Adaptive Semantic Segmentation

SafetyDGX agent

arXiv:2605.16393v1 Announce Type: cross Abstract: Semantic segmentation is essential for analysing anatomical features in biomedical research, yet a performance gap remains for Vision Transformers (Vi

Visual Sculpting: Visually-Aligned Planning Representations for Long-Horizon Robot Clay Sculpting

SafetyDGX agent

arXiv:2605.17556v1 Announce Type: cross Abstract: Clay sculpting is a nuanced, artistic task involving dexterous manipulation with long-horizon planning to achieve high-level goals. As a robotics prob

VLM-AutoDrive: Post-Training Vision-Language Models for Safety-Critical Autonomous Driving Events

SafetyDGX agent

arXiv:2603.18178v2 Announce Type: replace-cross Abstract: The rapid growth of ego-centric dashcam footage presents a major challenge for detecting safety-critical events such as collisions and near-co

Voices in the Loop: Mapping Participatory AI

SafetyDGX agent

arXiv:2605.16827v1 Announce Type: new Abstract: Participatory approaches to artificial intelligence are increasingly documented across public, civic, and humanitarian settings, but evidence about how

VolTA-3D: Self-Supervised Learning for Brain MRI using 3D Volumetric Token Alignment

SafetyDGX agent

arXiv:2605.16775v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has advanced medical image analysis be enabling learning form large unlabelled data. However, in brain magnetic resonan

Weak-to-Strong Elicitation via Mismatched Wrong Drafts

SafetyDGX agent

arXiv:2605.17314v1 Announce Type: cross Abstract: We consider whether off-policy experience from a smaller, weaker model can elicit capability in a stronger learner that on-policy RL fine-tuning (e.g.

What we find most useful about CNA is that the intervention is simple yet powerful. The steering is a multiplicative ablation on a sparse se…

SafetyDGX agent

What we find most useful about CNA is that the intervention is simple yet powerful. The steering is a multiplicative ablation on a sparse set of MLP neurons, which makes CNA a clean addition on top of

When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited

SafetyDGX agent

arXiv:2605.17017v1 Announce Type: cross Abstract: Behavior Foundation Models (BFMs) enable scalable imitation learning (IL) by pretraining task-agnostic representations that can be rapidly adapted to

When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search

SafetyDGX agent

arXiv:2605.16362v1 Announce Type: cross Abstract: Activation steering offers a lightweight way to control LLMs without retraining, but its effectiveness varies sharply across concepts. Prior work ofte

When Vision Speaks for Sound

SafetyDGX agent

arXiv:2605.16403v1 Announce Type: new Abstract: Despite rapid progress in video-capable MLLMs, we find that their apparent audio understanding in videos is often vision-driven: models rely on visual c

Where Pretraining writes and Alignment reads: the asymmetry of Transformer weight space

SafetyDGX agent

arXiv:2605.16600v1 Announce Type: cross Abstract: Cross-entropy pretraining and preference alignment update the same transformer weights, but leave geometrically distinct traces. We characterise this

Whispers in the Noise: Surrogate-Guided Concept Awakening via a Multi-Agent Framework

SafetyDGX agent

arXiv:2605.18150v1 Announce Type: new Abstract: Diffusion models (DMs) are widely used for text-to-image generation, but their strong generative capabilities also raise concerns about unsafe or undesi

← Previous
1…135136137138139…214
Next →