AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
29 May 2026

EvoMD-LLM: Learning the Language of Species Evolution in Reactive Molecular Dynamics

SafetyDGX agent

arXiv:2605.29394v1 Announce Type: new Abstract: While large language models (LLMs) excel at static scientific reasoning, they struggle to model the temporal structure of dynamic physical processes. We

EvoRubric: Self-Evolving Rubric-Driven RL for Open-Ended Generation

SafetyDGX agent

arXiv:2605.29847v1 Announce Type: new Abstract: Reinforcement Learning (RL) has significantly advanced Large Language Models (LLMs) in verifiable domains, but aligning models for open-ended generation

exactly this.

SafetyDGX agent

exactly this. @Michael14kBall @GaryMarcus @Vivek4real_ He's been publicly trashed by AI boosters this whole time, in dismissive terms. And the thing he's doing, with a number of others, is to try to c

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

FakeVLM-R1: Internalizing Physical Laws via CoT for Synthetic Image Detection

SafetyDGX agent

arXiv:2605.30062v1 Announce Type: new Abstract: The development of generative artificial intelligence technologies has propelled the visual realism of synthetic images to an unprecedented level. Altho

Fewer Steps, Better Performance: Efficient Cross-Modal Clip Trimming for Video Moment Retrieval Using Language

SafetyDGX agent

arXiv:2605.29793v1 Announce Type: new Abstract: Given an untrimmed video and a sentence query, video moment retrieval using language (VMR) aims to locate a target query-relevant moment. Since the untr

Fisher-Preserving Guidance: Training-Free Manifold Constraints for Safe Diffusion Control

SafetyDGX agent

arXiv:2605.29937v1 Announce Type: cross Abstract: Diffusion models are effective for waypoint prediction in visual navigation, but standard sampling and test time guidance can produce unreliable or in

FlowSeg: Dynamic Semantic Guidance for LLM-Conditioned Segmentation

SafetyDGX agent

arXiv:2605.29461v1 Announce Type: new Abstract: LLM-conditioned segmentation has recently advanced rapidly by coupling large language models with iterative mask generation frameworks. However, we iden

Follow-the-Perturbed-Leader for Decoupled Bandits: Best-of-Both-Worlds and Practicality

SafetyDGX agent

arXiv:2510.12152v2 Announce Type: replace-cross Abstract: We study the decoupled multi-armed bandit problem, where the learner separately selects one arm for exploration and one, possibly different, a

From Context Shift to Stylistic Collapse: Why Training Objectives Matter More Than Scale

SafetyDGX agent

arXiv:2605.28826v1 Announce Type: new Abstract: In modern LLMs, linguistic features function not as stylistic artifacts but as probes of probability mass, allocated under training alignment objectives

fully agree!

SafetyDGX agent

fully agree! Artificial intelligences do not undergo experiences, do not possess a body, do not feel joy or pain, do not mature through relationships, and do not know from within what love, work, frie

Future Forcing: Future-aware Training-free KV Cache Policy for Autoregressive Video Generation

SafetyDGX agent

arXiv:2605.30083v1 Announce Type: new Abstract: Autoregressive (AR) video generation has emerged as a promising paradigm for long-horizon video synthesis, where each frame is generated conditioned on

GAP3D: Generative Alignment of VLM Latents to Patch-Level Embeddings for 3D Generation

SafetyDGX agent

arXiv:2605.28995v1 Announce Type: new Abstract: Recent approaches integrating vision-language models (VLMs) as prompt encoders for generative model conditioning typically rely on expensive end-to-end

GAPD: Gold-Action Policy Distillation for Agentic Reinforcement Learning in Knowledge Base Question Answering

SafetyDGX agent

arXiv:2605.29584v1 Announce Type: new Abstract: Reinforcement learning (RL) is a natural fit for agentic knowledge base question answering (KBQA), where a model must issue executable actions, observe

GASS: Geometry-Aware Spherical Sampling for Disentangled Diversity Enhancement in Text-to-Image Generation

SafetyDGX agent

arXiv:2602.17200v2 Announce Type: replace Abstract: Despite high semantic alignment, modern text-to-image (T2I) generative models still struggle to synthesize diverse images from a given prompt. In th

Gaze2Act: Gaze-Conditioned Vision-Language-Action Policies for Interactive Robot Manipulation

SafetyDGX agent

arXiv:2605.30282v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently shown strong potential for robot learning by following language instructions. However, in practice, la

GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models

SafetyDGX agent

arXiv:2605.29398v1 Announce Type: cross Abstract: Reinforcement learning (RL) can be used to improve the policy (denoiser) of diffusion large language models (dLLMs), while being hindered by the intra

Genetically Aligned Patient Representations Improve Hematological Diagnosis

SafetyDGX agent

arXiv:2605.29980v1 Announce Type: cross Abstract: Multimodal alignment of histopathology encoders with transcriptomic and genomic data has been shown to significantly improve performance in downstream

Geometry-Guided Modeling of Foundation Features Enables Generalizable Object Shape Deformation Learning

SafetyDGX agent

arXiv:2605.29661v1 Announce Type: new Abstract: Monocular 3D shape recovery is fundamental to geometric understanding, yet achieving robust generalization across arbitrary viewpoints and unseen object

Grammar-Aware Literate Generative Mathematical Programming with Compiler-in-the-Loop

SafetyDGX agent

arXiv:2601.17670v2 Announce Type: replace-cross Abstract: Mathematical programming is widely employed across various sectors - such as logistics, energy, and workforce planning - to model and solve in

Graph-Enhanced Policy Optimization in LLM Agent Training

SafetyDGX agent

arXiv:2510.26270v2 Announce Type: replace Abstract: Multi-step LLM agents in interactive environments represent a crucial step toward long-horizon decision-making. To train such agents, group-based re

GrepSeek: Training Search Agents for Direct Corpus Interaction

SafetyDGX agent

arXiv:2605.29307v1 Announce Type: cross Abstract: Large Language Model (LLM) search agents have shown strong promise for knowledge-intensive language tasks through multiple rounds of reasoning and inf

Grounded 3D-Aware Spatial Vision-Language Modeling

SafetyDGX agent

arXiv:2605.30307v1 Announce Type: new Abstract: We present GR3D, a spatial vision language model equipped with three complementary grounding capabilities--explicit 2D grounding, implicit 2D grounding,

GRPO is Secretly a Process Reward Model

SafetyDGX agent

arXiv:2509.21154v4 Announce Type: replace-cross Abstract: Process reward models (PRMs) allow for fine-grained credit assignment in reinforcement learning (RL), and seemingly contrast with outcome rewa

GRUFF: LLM Pronoun Fidelity, Reasoning, and Biases in German

SafetyDGX agent

arXiv:2605.30214v1 Announce Type: new Abstract: Third-person singular pronouns have long been used to study stereotypical biases in language models and to test their abilities to reason about referenc

Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization

SafetyDGX agent

arXiv:2605.29198v1 Announce Type: new Abstract: Group-advantage-based reinforcement learning methods, such as GRPO and DAPO, have demonstrated strong performance across diverse domains, including math

Harnessing non-adversarial robustness in large language models

SafetyDGX agent

arXiv:2605.29816v1 Announce Type: new Abstract: The work presents an approach for addressing the challenge of robustness in Large Language Models (LLMs) to alterations and potential errors caused by s

How's it going? Reinforcement learning in language models recruits a functional welfare axis

SafetyDGX agent

arXiv:2605.30232v1 Announce Type: cross Abstract: How does reinforcement learning shape a language model's internal representations? We present evidence that RL recruits a pre-existing representation

HPO: Hysteretic Policy Optimization for Stable and Efficient Training under Sparse-Reward Regime

SafetyDGX agent

arXiv:2605.30201v1 Announce Type: cross Abstract: We investigate a narrow but common failure mode of GRPO-style reinforcement learning in the context of sparse verifiable rewards: early updates contai

i have strong reason to believe this is cope. we may find out soon…

SafetyDGX agent

i have strong reason to believe this is cope. we may find out soon… Im calling BS on this story. 1. That would be 100,000 employees spending 5k/mo each or 10,000 employees averaging 50k/mo each. No wa

Improving CLIP Adaptation by Breaking Tail Alignment for Source-Free Cross-Domain Few-Shot Learning

SafetyDGX agent

arXiv:2605.29776v1 Announce Type: new Abstract: Vision-Language Models (VLMs) such as CLIP demonstrate strong zero-shot generalization, but their performance significantly degrades in cross-domain sce

In-Context Reward Adaptation for Robust Preference Modeling

SafetyDGX agent

arXiv:2605.30323v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) typically relies on static reward models to align Large Language Models with human preferences. Howe

Inferring Code Correctness from Specification

SafetyDGX agent

arXiv:2605.29822v1 Announce Type: cross Abstract: Large language models (LLMs) have become integral to modern software development, enabling automated code generation at scale. However, validating the

Information-Directed Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2605.29405v1 Announce Type: new Abstract: Decision-making from offline datasets typically warm-starts a policy or score model from fixed offline data and then refines it with limited online inte

It might be time to fill one of these out again. Please remit to myself or @GaryMarcus Thank you.

SafetyDGX agent

Gary Marcus is requesting that someone complete a form or document and submit it to him or another person (possibly Gary Marcus himself based on the mention of @GaryMarcus). The post appears to be a r

It’s a good day when the Pope vouches for your recent comment in Nature.

SafetyDGX agent

It’s a good day when the Pope vouches for your recent comment in Nature. The Pope is making exactly our point. LLMs “may imitate or even simulate, but they do not understand.” This is the core epistem

It's a shame what happened to @kevinroose. He used to be a pretty damn good tech reporter. But recently he got one-shotted by spending too m…

SafetyDGX agent

It's a shame what happened to @kevinroose. He used to be a pretty damn good tech reporter. But recently he got one-shotted by spending too much time in/around the big AI Labs (research for the book he

KGEdit: Ambiguity-Aware Knowledge Graphs for Training-Free Precise Video Generation and Editing

SafetyDGX agent

arXiv:2605.29509v1 Announce Type: new Abstract: In recent years, training-free video generation has progressed remarkably. However, when handling complex textual instructions, existing methods still s

Learning A Simulation-based Visual Policy for Real-world Peg In Unseen Holes

SafetyDGX agent

arXiv:2205.04297v2 Announce Type: replace-cross Abstract: This paper proposes a learning-based visual peg-in-hole that enables training with several shapes in simulation, and adapting to arbitrary uns

Learning to Choose: An Empowerment-Guided Multi-Agent System with semantic communication for Adaptive Method Selection

SafetyDGX agent

arXiv:2605.30042v1 Announce Type: new Abstract: Automating scientific computing workflows requires more than generating executable code: autonomous systems must also select appropriate computational s

Lee Kuan Yew abolished trial by jury in Singapore after determining that it was too easy for defence lawyers to appeal to racial and religio…

SafetyDGX agent

Lee Kuan Yew abolished trial by jury in Singapore after determining that it was too easy for defence lawyers to appeal to racial and religious biases of juries in multicultural Singapore. He writes in

Less Is More: Elevating RAG via Performance-Driven Context Compression

SafetyDGX agent

arXiv:2508.19282v4 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has emerged as a promising paradigm for improving the timeliness of knowledge updates and the factual acc

LLM-Guided Future Hypotheses for Horizon-Aware Exploration in Multi-Step Robot Manipulation

SafetyDGX agent

arXiv:2605.29864v1 Announce Type: new Abstract: Multi-step robot manipulation requires acting under uncertainty about how the scene will evolve, making exploration and policy adaptation challenging. W

LoMo: Local Modality Substitution for Deeper Vision-Language Fusion

SafetyDGX agent

arXiv:2605.30265v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved substantial progress across a wide range of understanding and reasoning tasks, driven by large-scale image

Looking Beyond Text: Reducing Language bias in Large Vision-Language Models via Multimodal Dual-Attention and Soft-Image Guidance

SafetyDGX agent

arXiv:2411.14279v2 Announce Type: replace-cross Abstract: Large vision-language models (LVLMs) have achieved impressive results in various vision-language tasks. However, despite showing promising per

MARS Policy: Multimodality Only When It Matters

SafetyDGX agent

arXiv:2605.29766v1 Announce Type: new Abstract: Imitation learning has become a cornerstone for solving complex robotic manipulation tasks. In particular, multimodality, which enables robots to captur

Mask the Target: A Plug-and-Play Regularizer Against LoRA Forgetting

SafetyDGX agent

arXiv:2605.29498v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become one of the most widely used fine-tuning mechanisms for adapting large language models to new domains, tasks, and u

MATANet: A Multi-context Attention and Taxonomy-Aware Network for Fine-Grained Underwater Recognition of Marine Species

SafetyDGX agent

arXiv:2601.03729v2 Announce Type: replace Abstract: Fine-grained recognition of marine organisms is important for ecological research, biodiversity monitoring, habitat conservation, and evidence-based

Mean-Field Diffuser: Scaling Offline MARL to Thousands of Agents

SafetyDGX agent

arXiv:2605.30190v1 Announce Type: new Abstract: Diffusion-based planning has achieved strong results in single-agent offline reinforcement learning, yet scaling to many-agent systems remains intractab

MetaRanker: Human-in-the-loop Active Ranking for Metalens Image Quality

SafetyDGX agent

arXiv:2605.29212v1 Announce Type: new Abstract: Image quality in modern imaging systems emerges from the coupled effects of the sensor, optics, and computational reconstruction. Ultra-thin metalenses

MIC: Maximizing Informational Capacity in Adaptive Representations via Isotropic Subspace Alignment

SafetyDGX agent

arXiv:2605.29987v1 Announce Type: cross Abstract: Although multi-scales representation learning enables elastic-dimension embeddings, nested subspaces often suffer from dimensional redundancy and spec

Mining or Synthesis? Rethinking Exploration Efficiency in Iterative Alignment of Mathematical Reasoning

SafetyDGX agent

arXiv:2602.05370v3 Announce Type: replace Abstract: Iterative Direct Preference Optimization (DPO) has emerged as a widely used paradigm for aligning Large Language Models on reasoning tasks. Existing

Mitigating State Aliasing in Vision-Language-Action Models via Inverse Dynamics Learning

SafetyDGX agent

arXiv:2605.29577v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising framework that unifies perception, reasoning, and control for robot manipulation by adap

Modality Alignment across Trees on Heterogeneous Hyperbolic Manifolds

SafetyDGX agent

arXiv:2510.27391v2 Announce Type: replace Abstract: Modality alignment is critical for vision-language models (VLMs) to effectively integrate information across modalities. However, existing methods e

Model Fusion via Retrofitting

SafetyDGX agent

arXiv:2507.00037v2 Announce Type: replace-cross Abstract: Model fusion seeks to combine independently trained neural networks into a single model without retraining, but is complicated by representati

Modeling Hierarchical Thinking in Large Reasoning Models

SafetyDGX agent

arXiv:2510.22437v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) solve complex tasks by generating long Chain-of-Thought (CoT) sequences; however, the emergent dynamics governing reas

Modularizing Educational LLM-Agency for Fostering Responsible Learning Assistance

SafetyDGX agent

arXiv:2605.30187v1 Announce Type: new Abstract: The widespread adoption of AI chatbots in education will drastically change learning, making responsible deployment a critical concern. While large lang

MonoPhysics: Estimating Geometry, Appearance, and Physical Parameters from Monocular Videos

SafetyDGX agent

arXiv:2605.30320v1 Announce Type: new Abstract: Existing inverse physics methods recover physical parameters from multi-view videos, where geometric constraints across views resolve scale and 3D struc

Multiple executives tell me they’re looking to decrease their AI expenses just as three major AI IPOs are on the horizon. My latest:

SafetyDGX agent

Multiple executives tell me they’re looking to decrease their AI expenses just as three major AI IPOs are on the horizon. My latest: CEOs are bargain hunting for AI https://www.axios.com/2026/05/29/ce

Native Audio-Visual Alignment for Generation

SafetyDGX agent

arXiv:2605.30073v1 Announce Type: new Abstract: Joint audio-video generation aims to synthesize temporally synchronized and semantically coherent visual-acoustic content. However, existing open-source

Neural Operator-Based Surrogate Model for CFD:Helical Coil Steam Generator in Small Modular Reactor

SafetyDGX agent

arXiv:2605.30277v1 Announce Type: new Abstract: Real-time thermal-hydraulic simulation is essential for digital twin (DT) technology that supports the safe and efficient operation of small modular rea

← Previous
1…137138139140141…242
Next →