AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
19 May 2026

Reducing Credit Assignment Variance via Counterfactual Reasoning Paths

SafetyDGX agent

arXiv:2605.16302v1 Announce Type: cross Abstract: Reinforcement learning for multi-step reasoning with large language models (LLMs) often relies on sparse terminal rewards, leading to poor credit assi

Reliability and Effectiveness of Autonomous AI Agents in Supply Chain Management

SafetyDGX agent

arXiv:2605.17036v1 Announce Type: new Abstract: This paper studies autonomous generative AI agents in multi-echelon supply chains using the MIT Beer Game. We identify four inference-time levers that s

Representational Alignment with Chemical Induced Fit for Molecular Relational Learning

SafetyDGX agent

arXiv:2502.07027v2 Announce Type: replace-cross Abstract: Molecular Relational Learning (MRL) is widely applied in natural sciences to predict relationships between molecular pairs by extracting struc

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Resolving Representation Ambiguity in Feedforward Novel View Synthesis Transformer via Semantic-Spatial Decoupling

SafetyDGX agent

arXiv:2605.18599v1 Announce Type: new Abstract: Transformer-based models have advanced feedforward novel view synthesis (NVS). Current architectures such as GS-LRM and LVSM mix semantic information (e

Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review

SafetyDGX agent

arXiv:2605.17548v1 Announce Type: cross Abstract: Code review has evolved for decades, from informal peer checking to today's pull request (PR) workflows, yet it remains a largely manual, uneven, and

Retrieval and competition: how a protein foundation model starts a protein

SafetyDGX agent

arXiv:2605.16331v1 Announce Type: cross Abstract: Protein language models are increasingly used to guide experimental and clinical decisions, yet it is often unclear whether a confident prediction ref

Right Predictions, Misleading Explanations: On the Vulnerability of Vision-Language Model Explanations

SafetyDGX agent

arXiv:2605.16651v1 Announce Type: new Abstract: Explanation mechanisms are increasingly used to support transparency and trust in vision-language models (VLMs), particularly in settings where model de

RL4RLA: Teaching ML to Discover Randomized Linear Algebra Algorithms Through Curriculum Design and Graph-Based Search

SafetyDGX agent

arXiv:2605.18004v1 Announce Type: new Abstract: Randomized linear algebra (RLA) algorithms are a modern class of numerical linear algebra techniques that play an essential role in scientific computing

Running the same CNA search on base models (before instruction tuning) yields a structurally similar set of neurons, but ablating them produ…

SafetyDGX agent

Running the same CNA search on base models (before instruction tuning) yields a structurally similar set of neurons, but ablating them produces almost no behavioral change. We read this as evidence th

S2Aligner: Pair-Efficient and Transferable Pre-Training for Sparse Text-Attributed Graphs

SafetyDGX agent

arXiv:2605.18579v1 Announce Type: new Abstract: Pre-training on text-attributed graphs (TAGs) is central to building transferable graph foundation models, where LLM-as-Aligner methods align graph and

SADP: Subgoal-Aware Diffusion Policy for Explainable Robots Learned from Foundation Model Generated Demonstrations

SafetyDGX agent

arXiv:2605.16871v1 Announce Type: new Abstract: Explainable robots require not only successful task execution but also the ability to expose internal decision-making process in a user-friendly manner.

SAPO: Step-Aligned Policy Optimization for Reasoning-Based Generative Recommendation

SafetyDGX agent

arXiv:2605.17648v1 Announce Type: new Abstract: Generative recommendation treats next-item prediction as autoregressive item-identifier generation. Specifically, items are encoded as semantic identifi

Scalable Bi-causal Optimal Transport via KL Relaxation and Policy Gradients

SafetyDGX agent

arXiv:2605.17271v1 Announce Type: cross Abstract: Bi-causal optimal transport (OT) is a natural framework for comparing and coupling stochastic processes under nonanticipative information constraints,

Scale Determines Whether Language Models Organize Representation Geometry for Prediction

SafetyDGX agent

arXiv:2605.17084v1 Announce Type: cross Abstract: In language models, what a representation encodes is determined by the geometry of its representation space: distances, not activations, carry meaning

SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning

SafetyDGX agent

arXiv:2605.18299v1 Announce Type: new Abstract: Search-augmented reasoning agents interleave internal reasoning with calls to an external retriever, and their performance relies on the quality of each

Self-Evolving Spatial Reasoning in Vision Language Models via Geometric Logic Consistency

SafetyDGX agent

arXiv:2605.18162v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have made striking progress, yet their spatial reasoning remains fragile: models that answer an original input correctly

SGSoft: Learning Fused Semantic-Geometric Features for 3D Shape Correspondence via Template-Guided Soft Signals

SafetyDGX agent

arXiv:2605.18039v1 Announce Type: new Abstract: Learning dense correspondences across deformable 3D shapes remains a long-standing challenge due to structural variability, non-isometric deformation, a

Shared Backbone PPO for Multi-UAV Communication Coverage with Connection Preservation

SafetyDGX agent

arXiv:2605.17999v1 Announce Type: new Abstract: This paper proposes a Shared Backbone Proximal Policy Optimization (Shared Backbone PPO) algorithm. By sharing the base module between the Actor and Cri

SHED: Style-Homogenized Embedding Alignment for Domain Generalization

SafetyDGX agent

arXiv:2605.16973v1 Announce Type: new Abstract: Domain generalization aims to enhance model robustness against unseen domains with embedding distribution shifts. While large-scale vision-language mode

show me your /goal prompt that beats this (i'll share mine below so as not to bias)

SafetyDGX agent

This post appears to be a community prompt-sharing discussion where Swyx invites people to share their '/goal' prompts (likely system prompts or directives designed to optimize AI behavior toward spec

Simple Approximation and Derivative Free Inference-Time Scaling for Diffusion Models via Sequential Monte Carlo on Path Measures

SafetyDGX agent

arXiv:2605.17850v1 Announce Type: cross Abstract: iffusion-based generative models increasingly rely on inference-time guidance, adding a drift term or reweighting mixture of experts, to improve sampl

Single-Sample Black-Box Membership Inference Attack against Vision-Language Models via Cross-modal Semantic Alignment

SafetyDGX agent

arXiv:2605.17341v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable success, yet their reliance on massive datasets and unintended memorization of training data ra

SNLP: Layer-Parallel Inference via Structured Newton Corrections

SafetyDGX agent

arXiv:2605.17842v1 Announce Type: new Abstract: Autoregressive language models execute Transformer layers sequentially, creating a latency bottleneck that is not removed by conventional tensor or pipe

Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration

SafetyDGX agent

arXiv:2605.17423v1 Announce Type: new Abstract: We study series-level cinematic remaking, a long-horizon video-to-video generation problem that localizes full episodes or films via stylization or acto

Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias

SafetyDGX agent

arXiv:2509.22061v2 Announce Type: replace-cross Abstract: Speech Continuation (SC) is the task of generating a coherent extension of a spoken prompt while preserving both semantic context and speaker

SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language Reasoning

SafetyDGX agent

arXiv:2510.16416v4 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have shown remarkable abilities by integrating large language models with visual inputs. However, they often fai

Stable and Near-Reversible Diffusion ODE Solvers for Image Editing

SafetyDGX agent

arXiv:2605.16399v1 Announce Type: new Abstract: The inversion of diffusion models plays a central role in image editing. Algebraically reversible ODE solvers provide an appealing approach to diffusion

Stable Routing for Mixture-of-Experts in Class-Incremental Learning

SafetyDGX agent

arXiv:2605.17571v1 Announce Type: new Abstract: Class-incremental learning (CIL) requires models to learn new classes sequentially while preserving prior knowledge. Recently, approaches that combine p

StableHand: Quality-Aware Flow Matching for World-Space Dual-Hand Motion Estimation from Egocentric Video

SafetyDGX agent

arXiv:2605.18553v1 Announce Type: cross Abstract: Recovering world space 4D motion of two interacting hands from egocentric video is a fundamental capability for supervising robot policy learning, whe

State-Conditional Adversarial Learning: An Off-Policy Visual Domain Transfer Method for End-to-End Imitation Learning

SafetyDGX agent

arXiv:2512.05335v3 Announce Type: replace Abstract: We study visual domain transfer for end-to-end imitation learning in a realistic and challenging setting where target-domain data are strictly off-p

Stochastic Penalty-Barrier Methods for Constrained Machine Learning

SafetyDGX agent

arXiv:2605.18618v1 Announce Type: cross Abstract: Constrained machine learning enables fairness-aware training, physics-informed neural networks, and integration of symbolic domain knowledge into stat

Stop When Reasoning Converges: Semantic-Preserving Early Exit for Reasoning Models

SafetyDGX agent

arXiv:2605.17672v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance by generating long chains of thought (CoT), but often overthink, continuing to reason after a s

Structure-Aware Masking for Protein Representation Learning

SafetyDGX agent

arXiv:2605.16581v1 Announce Type: new Abstract: Masked language modeling (MLM) is the standard objective for training protein language models, typically implemented by randomly masking individual resi

Superb conversation between @bgreene and @GaryMarcus Very happy to get my confirmation bias pampered by mentions of interpolation vs extrapo…

SafetyDGX agent

Superb conversation between @bgreene and @GaryMarcus Very happy to get my confirmation bias pampered by mentions of interpolation vs extrapolation, and how having a better physics understanding might

SURGE: Approximation-free Training Free Particle Filter for Diffusion Surrogate

SafetyDGX agent

arXiv:2605.18745v1 Announce Type: cross Abstract: Diffusion-based generative models increasingly rely on inference-time guidance, adding a drift term or reweighting mixture of experts, to improve samp

Surgical Post-Training: Proximal On-Policy Distillation for Reasoning with Knowledge Retention

SafetyDGX agent

arXiv:2603.01683v2 Announce Type: replace-cross Abstract: Injecting new reasoning knowledge into Large Language Models (LLMs) via post-training often induces catastrophic forgetting. Recent studies em

SutureFormer: Learning Surgical Trajectories via Goal-conditioned Offline RL in Pixel Space

SafetyDGX agent

arXiv:2603.26720v2 Announce Type: replace-cross Abstract: Predicting surgical needle trajectories from endoscopic video is critical for robot-assisted suturing, enabling anticipatory planning, real-ti

SVL: Spike-based Vision-language Pretraining for Efficient 3D Open-world Understanding

SafetyDGX agent

arXiv:2505.17674v2 Announce Type: replace Abstract: Spiking Neural Networks (SNNs) provide an energy-efficient way to extract 3D spatio-temporal features. However, existing SNNs still exhibit a signif

T-FIX: Text-Based Explanations with Features Interpretable to eXperts

SafetyDGX agent

arXiv:2511.04070v3 Announce Type: replace Abstract: As LLMs are deployed in knowledge-intensive settings (e.g., surgery, astronomy, therapy), users are often domain experts who expect not just answers

TacSE3: Equivariant SE(3) Motion Estimation from Low-Texture Visuotactile Images for In-Gripper Tracking and Compensation

SafetyDGX agent

arXiv:2605.17929v1 Announce Type: new Abstract: Robotic in-hand manipulation requires reliable object-motion tracking under frequent visual occlusion, yet low-texture visuotactile images provide few s

Taming 'Zombie'' Agents: A Markov State-Aware Framework for Resilient Multi-Agent Evolution

SafetyDGX agent

arXiv:2605.17348v1 Announce Type: new Abstract: Recent advancements in LLM-based multi-agent systems have demonstrated remarkable collaborative capabilities across complex tasks. To improve overall ef

terrific deep dive on Mythos from @cloudflare

SafetyDGX agent

terrific deep dive on Mythos from @cloudflare Cloudflare's security team spent the last few weeks testing Anthropic's Mythos against fifty of our own repositories. What we learned about offensive AI,

The Alien Space of Science: Sampling Coherent but Cognitively Unavailable Research Directions

SafetyDGX agent

arXiv:2603.01092v2 Announce Type: replace Abstract: Scientific discovery is constrained not only by what is true, but by what is cognitively available to the researchers currently exploring a field. M

The Bayesian Geometry of Transformer Attention

SafetyDGX agent

arXiv:2512.22471v5 Announce Type: replace-cross Abstract: Transformers often appear to perform Bayesian reasoning in context, but verifying this rigorously has been impossible: natural data lack analy

The Hidden Cost of Contextual Sycophancy: an AI Literacy Intervention in Human-AI Collaboration

SafetyDGX agent

arXiv:2605.18372v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in educational settings as interactive tools for collaboration. However, their tendency toward syco

The Impact of AI Search on the Online Content Ecosystem: Evidence from Google and Reddit

SafetyDGX agent

arXiv:2605.16428v1 Announce Type: cross Abstract: Search engines traditionally complement online content platforms by directing users seeking information to external websites. The emergence of generat

The Laplacian Keyboard: Beyond the Linear Span

SafetyDGX agent

arXiv:2602.07730v2 Announce Type: replace-cross Abstract: Across scientific disciplines, Laplacian eigenvectors serve as a fundamental basis for simplifying complex systems, from signal processing to

“the uncritical adoption of AI in science is alarming”

SafetyDGX agent

“the uncritical adoption of AI in science is alarming” Artificial intelligence is rapidly accelerating scientific output, but risks narrowing inquiry, weakening judgement and undermining how scientist

Toward Robust Multilingual Adaptation of LLMs for Low-Resource Languages

SafetyDGX agent

arXiv:2510.14466v3 Announce Type: replace-cross Abstract: Large language models (LLMs) continue to struggle with low-resource languages, primarily due to limited training data, translation noise, and

Towards Sustainable Growth: A Multi-Value-Aware Retrieval Framework for E-Commerce Search

SafetyDGX agent

arXiv:2605.17994v1 Announce Type: cross Abstract: New item growth is critical for maintaining a healthy ecosystem in large-scale e-commerce platforms. However, existing systems tend to prioritize pres

Towards Universal Physical Adversarial Attacks via a Joint Multi-Objective and Multi-Model Optimization Framework

SafetyDGX agent

arXiv:2605.17772v1 Announce Type: new Abstract: Physical adversarial attacks often overfit single surrogate models and optimization objectives. While ensemble attacks can mitigate this, existing metho

Uncertainty-Calibrated Recommendations for Low-Active Users

SafetyDGX agent

arXiv:2605.17788v1 Announce Type: cross Abstract: A fundamental challenge in recommender systems is balancing reliability for Low-Active Users (LAUs) with diversity for High-Active Users (HAUs). The k

UniAlign: A Model-Agnostic Framework for Robust Network Traffic Classification under Distribution Shifts

SafetyDGX agent

arXiv:2605.17575v1 Announce Type: cross Abstract: Network traffic classification (NTC) models often suffer severe performance degradation when deployed in real-world environments due to distribution s

Unified Walking, Running, and Recovery for Humanoids via State-Dependent Adversarial Motion Priors

SafetyDGX agent

arXiv:2605.18611v1 Announce Type: new Abstract: We propose a unified reinforcement learning framework that enables a single policy to perform walking, running, and fall recovery on the Unitree G1 huma

Unifying Contrastive and Generative Objectives for Visual Understanding and Text-to-Image Generation

SafetyDGX agent

arXiv:2603.02667v2 Announce Type: replace Abstract: Unifying text-image contrastive learning and text-to-image (T2I) generation in a single end-to-end model is challenging because the two objectives d

Universal Pose Pretraining for Generalizable Vision-Language-Action Policies

SafetyDGX agent

arXiv:2602.19710v2 Announce Type: replace Abstract: Existing Vision-Language-Action (VLA) models often suffer from feature collapse and low training efficiency because they entangle high-level percept

Unlearning Isn't Deletion: Investigating Reversibility of Machine Unlearning in LLMs

SafetyDGX agent

arXiv:2505.16831v3 Announce Type: replace-cross Abstract: Unlearning in large language models (LLMs) aims to remove specified data, but its efficacy is typically assessed with task-level metrics like

Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning

SafetyDGX agent

arXiv:2510.02590v2 Announce Type: replace Abstract: The use of target networks is a popular approach for estimating value functions in deep Reinforcement Learning (RL). While effective, the target net

Video Reconstruction using Diffusion-based Image-to-Video Generation with Trajectory Guidance

SafetyDGX agent

arXiv:2605.16420v1 Announce Type: new Abstract: This paper addresses the problem of reconstructing missing or dropped frames in top-down drone video of autonomous surface vehicles performing structure

View-Aware Semantic Alignment for Aerial-Ground Person Re-Identification

SafetyDGX agent

arXiv:2605.18192v1 Announce Type: new Abstract: Aerial-Ground Person Re-Identification (AGPReID) remains highly challenging due to drastic viewpoint variations between drones and fixed cameras. Existi

← Previous
1…161162163164165…242
Next →