AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,814 results
29 May 2026

HPO: Hysteretic Policy Optimization for Stable and Efficient Training under Sparse-Reward Regime

SafetyDGX agent

arXiv:2605.30201v1 Announce Type: cross Abstract: We investigate a narrow but common failure mode of GRPO-style reinforcement learning in the context of sparse verifiable rewards: early updates contai

i have strong reason to believe this is cope. we may find out soon…

SafetyDGX agent

i have strong reason to believe this is cope. we may find out soon… Im calling BS on this story. 1. That would be 100,000 employees spending 5k/mo each or 10,000 employees averaging 50k/mo each. No wa

Improving CLIP Adaptation by Breaking Tail Alignment for Source-Free Cross-Domain Few-Shot Learning

SafetyDGX agent

arXiv:2605.29776v1 Announce Type: new Abstract: Vision-Language Models (VLMs) such as CLIP demonstrate strong zero-shot generalization, but their performance significantly degrades in cross-domain sce


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

In-Context Reward Adaptation for Robust Preference Modeling

SafetyDGX agent

arXiv:2605.30323v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) typically relies on static reward models to align Large Language Models with human preferences. Howe

Inferring Code Correctness from Specification

SafetyDGX agent

arXiv:2605.29822v1 Announce Type: cross Abstract: Large language models (LLMs) have become integral to modern software development, enabling automated code generation at scale. However, validating the

Information-Directed Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2605.29405v1 Announce Type: new Abstract: Decision-making from offline datasets typically warm-starts a policy or score model from fixed offline data and then refines it with limited online inte

Intent-aligned Autonomous Spacecraft Guidance via Reasoning Models

SafetyDGX agent

arXiv:2604.17176v2 Announce Type: replace-cross Abstract: Future spacecraft operations require autonomy that can interpret high-level mission intent while preserving safety. However, existing trajecto

It might be time to fill one of these out again. Please remit to myself or @GaryMarcus Thank you.

SafetyDGX agent

Gary Marcus is requesting that someone complete a form or document and submit it to him or another person (possibly Gary Marcus himself based on the mention of @GaryMarcus). The post appears to be a r

It’s a good day when the Pope vouches for your recent comment in Nature.

SafetyDGX agent

It’s a good day when the Pope vouches for your recent comment in Nature. The Pope is making exactly our point. LLMs “may imitate or even simulate, but they do not understand.” This is the core epistem

It's a shame what happened to @kevinroose. He used to be a pretty damn good tech reporter. But recently he got one-shotted by spending too m…

SafetyDGX agent

It's a shame what happened to @kevinroose. He used to be a pretty damn good tech reporter. But recently he got one-shotted by spending too much time in/around the big AI Labs (research for the book he

Jailbreaking and Mitigation of Vulnerabilities in Large Language Models

SafetyDGX agent

arXiv:2410.15236v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence by advancing natural language understanding and generation, enabling app

KGEdit: Ambiguity-Aware Knowledge Graphs for Training-Free Precise Video Generation and Editing

SafetyDGX agent

arXiv:2605.29509v1 Announce Type: new Abstract: In recent years, training-free video generation has progressed remarkably. However, when handling complex textual instructions, existing methods still s

Learning A Simulation-based Visual Policy for Real-world Peg In Unseen Holes

SafetyDGX agent

arXiv:2205.04297v2 Announce Type: replace-cross Abstract: This paper proposes a learning-based visual peg-in-hole that enables training with several shapes in simulation, and adapting to arbitrary uns

Learning to Choose: An Empowerment-Guided Multi-Agent System with semantic communication for Adaptive Method Selection

SafetyDGX agent

arXiv:2605.30042v1 Announce Type: new Abstract: Automating scientific computing workflows requires more than generating executable code: autonomous systems must also select appropriate computational s

Lee Kuan Yew abolished trial by jury in Singapore after determining that it was too easy for defence lawyers to appeal to racial and religio…

SafetyDGX agent

Lee Kuan Yew abolished trial by jury in Singapore after determining that it was too easy for defence lawyers to appeal to racial and religious biases of juries in multicultural Singapore. He writes in

Less Is More: Elevating RAG via Performance-Driven Context Compression

SafetyDGX agent

arXiv:2508.19282v4 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has emerged as a promising paradigm for improving the timeliness of knowledge updates and the factual acc

LLM-Guided Future Hypotheses for Horizon-Aware Exploration in Multi-Step Robot Manipulation

SafetyDGX agent

arXiv:2605.29864v1 Announce Type: new Abstract: Multi-step robot manipulation requires acting under uncertainty about how the scene will evolve, making exploration and policy adaptation challenging. W

LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback

SafetyDGX agent

arXiv:2605.30273v1 Announce Type: cross Abstract: Large language models (LLMs) show promise in generating supportive responses for mental health queries, but improving their usefulness, empathy, and s

LoMo: Local Modality Substitution for Deeper Vision-Language Fusion

SafetyDGX agent

arXiv:2605.30265v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved substantial progress across a wide range of understanding and reasoning tasks, driven by large-scale image

Looking Beyond Text: Reducing Language bias in Large Vision-Language Models via Multimodal Dual-Attention and Soft-Image Guidance

SafetyDGX agent

arXiv:2411.14279v2 Announce Type: replace-cross Abstract: Large vision-language models (LVLMs) have achieved impressive results in various vision-language tasks. However, despite showing promising per

Low-Magnification SEM May Suffice: Interpretable Deep Learning for Multi-Scale Fracture-Cause Classification in Zirconia-Toughened Alumina

SafetyDGX agent

arXiv:2605.29798v1 Announce Type: new Abstract: Reliable identification of fracture origins in alumina matrix composite hip and knee implants is critical for quality assurance and patient safety, yet

MARS Policy: Multimodality Only When It Matters

SafetyDGX agent

arXiv:2605.29766v1 Announce Type: new Abstract: Imitation learning has become a cornerstone for solving complex robotic manipulation tasks. In particular, multimodality, which enables robots to captur

Mask the Target: A Plug-and-Play Regularizer Against LoRA Forgetting

SafetyDGX agent

arXiv:2605.29498v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become one of the most widely used fine-tuning mechanisms for adapting large language models to new domains, tasks, and u

Masked Diffusion Modeling for Anomaly Detection

SafetyDGX agent

arXiv:2605.30046v1 Announce Type: cross Abstract: Anomaly detection aims to identify samples that deviate from the nominal data distribution and is central to many safety-critical applications. Howeve

MATANet: A Multi-context Attention and Taxonomy-Aware Network for Fine-Grained Underwater Recognition of Marine Species

SafetyDGX agent

arXiv:2601.03729v2 Announce Type: replace Abstract: Fine-grained recognition of marine organisms is important for ecological research, biodiversity monitoring, habitat conservation, and evidence-based

Mean-Field Diffuser: Scaling Offline MARL to Thousands of Agents

SafetyDGX agent

arXiv:2605.30190v1 Announce Type: new Abstract: Diffusion-based planning has achieved strong results in single-agent offline reinforcement learning, yet scaling to many-agent systems remains intractab

MetaRanker: Human-in-the-loop Active Ranking for Metalens Image Quality

SafetyDGX agent

arXiv:2605.29212v1 Announce Type: new Abstract: Image quality in modern imaging systems emerges from the coupled effects of the sensor, optics, and computational reconstruction. Ultra-thin metalenses

Metric-Dependent Annotation Saturation for Learning from Label Distributions

SafetyDGX agent

arXiv:2605.29797v1 Announce Type: new Abstract: When annotators disagree on a label, the disagreement itself carries signal -- and the number of annotators needed to capture it depends on the evaluati

MIC: Maximizing Informational Capacity in Adaptive Representations via Isotropic Subspace Alignment

SafetyDGX agent

arXiv:2605.29987v1 Announce Type: cross Abstract: Although multi-scales representation learning enables elastic-dimension embeddings, nested subspaces often suffer from dimensional redundancy and spec

Mining or Synthesis? Rethinking Exploration Efficiency in Iterative Alignment of Mathematical Reasoning

SafetyDGX agent

arXiv:2602.05370v3 Announce Type: replace Abstract: Iterative Direct Preference Optimization (DPO) has emerged as a widely used paradigm for aligning Large Language Models on reasoning tasks. Existing

Mitigating State Aliasing in Vision-Language-Action Models via Inverse Dynamics Learning

SafetyDGX agent

arXiv:2605.29577v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising framework that unifies perception, reasoning, and control for robot manipulation by adap

Modality Alignment across Trees on Heterogeneous Hyperbolic Manifolds

SafetyDGX agent

arXiv:2510.27391v2 Announce Type: replace Abstract: Modality alignment is critical for vision-language models (VLMs) to effectively integrate information across modalities. However, existing methods e

Model Fusion via Retrofitting

SafetyDGX agent

arXiv:2507.00037v2 Announce Type: replace-cross Abstract: Model fusion seeks to combine independently trained neural networks into a single model without retraining, but is complicated by representati

Modeling Hierarchical Thinking in Large Reasoning Models

SafetyDGX agent

arXiv:2510.22437v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) solve complex tasks by generating long Chain-of-Thought (CoT) sequences; however, the emergent dynamics governing reas

Modularizing Educational LLM-Agency for Fostering Responsible Learning Assistance

SafetyDGX agent

arXiv:2605.30187v1 Announce Type: new Abstract: The widespread adoption of AI chatbots in education will drastically change learning, making responsible deployment a critical concern. While large lang

MonoPhysics: Estimating Geometry, Appearance, and Physical Parameters from Monocular Videos

SafetyDGX agent

arXiv:2605.30320v1 Announce Type: new Abstract: Existing inverse physics methods recover physical parameters from multi-view videos, where geometric constraints across views resolve scale and 3D struc

Multi-Resolution End-to-End Deep Neural Network for Optimizing Latency-Accuracy Tradeoff in Autonomous Driving

SafetyDGX agent

arXiv:2605.29138v1 Announce Type: cross Abstract: Latency-accuracy tradeoffs are fundamental in real-time applications of deep neural networks (DNNs) for cyber-physical systems. In autonomous driving,

Multiple executives tell me they’re looking to decrease their AI expenses just as three major AI IPOs are on the horizon. My latest:

SafetyDGX agent

Multiple executives tell me they’re looking to decrease their AI expenses just as three major AI IPOs are on the horizon. My latest: CEOs are bargain hunting for AI https://www.axios.com/2026/05/29/ce

Native Audio-Visual Alignment for Generation

SafetyDGX agent

arXiv:2605.30073v1 Announce Type: new Abstract: Joint audio-video generation aims to synthesize temporally synchronized and semantically coherent visual-acoustic content. However, existing open-source

Neural Network Verification using Partial Multi-Neuron Relaxation

SafetyDGX agent

arXiv:2605.30155v1 Announce Type: cross Abstract: The increasing integration of deep neural networks in critical systems has spawned a theoretical and practical interest in formally guaranteeing safet

Neural Operator-Based Surrogate Model for CFD:Helical Coil Steam Generator in Small Modular Reactor

SafetyDGX agent

arXiv:2605.30277v1 Announce Type: new Abstract: Real-time thermal-hydraulic simulation is essential for digital twin (DT) technology that supports the safe and efficient operation of small modular rea

Nobody knows for sure where employment is going and over what time frame. But one thing I can tell you for sure is that the big AI CEO’s hav…

SafetyDGX agent

Nobody knows for sure where employment is going and over what time frame. But one thing I can tell you for sure is that the big AI CEO’s have started lying about it. When they tell you know “we are ju

Note to all staff: Turns out that super tasty AI candy we've been putting out in large bowls isn't free and it isn't necessarily leading to …

SafetyDGX agent

Note to all staff: Turns out that super tasty AI candy we've been putting out in large bowls isn't free and it isn't necessarily leading to better products for our customers. From now on, all employee

Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition

SafetyDGX agent

arXiv:2505.05968v3 Announce Type: replace Abstract: Offline cooperative multi-agent reinforcement learning (MARL) faces unique challenges due to distributional shifts, particularly stemming from the h

Offline Reinforcement Learning with Generative Trajectory Policies

SafetyDGX agent

arXiv:2510.11499v2 Announce Type: replace-cross Abstract: Generative models have emerged as a powerful class of policies for offline reinforcement learning (RL) due to their ability to capture complex

On the Geometry of Games and their Solvers

SafetyDGX agent

arXiv:2605.29919v1 Announce Type: new Abstract: A central challenge in game theory and learning systems such as GANs is understanding which algorithms can efficiently compute equilibria across the het

Online Fair Division with Additional Information

SafetyDGX agent

arXiv:2505.24503v3 Announce Type: replace-cross Abstract: We study the problem of fairly allocating indivisible goods to agents in an online setting, where goods arrive sequentially and must be alloca

Open Problem: Separating Geometric and Algorithmic Compression via Cayley-Table Completion

SafetyDGX agent

arXiv:2605.29885v1 Announce Type: new Abstract: Modern statistical learning theory and deep learning characterize generalization primarily in terms of continuous capacity control (e.g., norm-based reg

Opus 4.8 is insane, nothing will be the same after this model 💀

SafetyDGX agent

Gary Marcus expresses strong enthusiasm about Opus 4.8, suggesting it represents a significant breakthrough in AI capabilities. The post implies the model introduces substantial improvements or novel

Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation

SafetyDGX agent

arXiv:2605.29390v1 Announce Type: new Abstract: Text-to-image (T2I) models have become increasingly capable of generating high-quality images. Yet, enforcing the explicit absence of a specified object

Paper Agents, Paper Gains: An Empirical Analysis of DeFi Investment Agents

SafetyDGX agent

arXiv:2605.29174v1 Announce Type: new Abstract: DeFi investment agents, systems that use AI for autonomous on-chain trading, have attained over USD 3 billion in combined token valuations since late 20

Path-Space Mirror Descent for On-Policy Reinforcement Learning under the Generalized Schrodinger Bridge

SafetyDGX agent

arXiv:2603.21621v2 Announce Type: replace Abstract: Classical on-policy algorithms such as PPO and mirror descent policy optimization provide stable proximal policy updates through tractable action li

PEARL: Training Socratic Tutors with Pedagogically Aligned Reinforcement Learning

SafetyDGX agent

arXiv:2605.29582v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promise as educational tutors, yet effective tutoring requires more than solving problems: it must provide pro

Permutation-Invariant Spectral Learning via Dyson Diffusion

SafetyDGX agent

arXiv:2510.08535v2 Announce Type: replace-cross Abstract: Diffusion models are central to generative modeling and have been adapted to graphs by diffusing adjacency matrix representations. The challen

PersonaAgent: Bridging Memory and Action for Personalized LLM Agents

SafetyDGX agent

arXiv:2506.06254v2 Announce Type: replace Abstract: Large Language Model (LLM) empowered agents have recently emerged as advanced paradigms that exhibit impressive capabilities in a wide range of doma

Phase-Conditioned Imitation Learning with Autonomous Failure Recovery for Robust Deformable Object Manipulation

SafetyDGX agent

arXiv:2605.29407v1 Announce Type: new Abstract: This paper presents a phase-conditioned, force-aware framework for robust deformable object manipulation. Standard imitation learning policies such as A

Plan, Don't Pose: Long Composite Motion Generation with Text-Aligned BFM

SafetyDGX agent

arXiv:2605.29906v1 Announce Type: new Abstract: Text-to-motion (T2M) generation has broad applications in character animation, virtual avatars, and human-robot interaction. Existing methods typically

Position: Stop Chasing the C-index when Evaluating Survival Analysis Models

SafetyDGX agent

arXiv:2506.02075v2 Announce Type: replace-cross Abstract: The current state of evaluation in survival analysis is plagued by the persistent use of evaluation metrics in ways that are misaligned with t

Practitioner Beliefs and Behaviors in AI-Enhanced Education: DOT Framework Survey Evidence

SafetyDGX agent

arXiv:2605.29041v1 Announce Type: new Abstract: This study reports findings from a cross-sectional survey (n = 72) of higher education practitioners examining beliefs, behaviors, and institutional con

PRO-CUA: Process-Reward Optimization for Computer Use Agents

SafetyDGX agent

arXiv:2605.29119v1 Announce Type: new Abstract: Computer use agents (CUAs) have shown strong potential for automating complex digital workflows, yet their training remains constrained by costly live e

← Previous
1…106107108109110…214
Next →