AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Mining or Synthesis? Rethinking Exploration Efficiency in Iterative Alignment of Mathematical Reasoning

DGX agent

arXiv:2602.05370v3 Announce Type: replace Abstract: Iterative Direct Preference Optimization (DPO) has emerged as a widely used paradigm for aligning Large Language Models on reasoning tasks. Existing

safetyarxiv-cs-cl
29 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Mitigating State Aliasing in Vision-Language-Action Models via Inverse Dynamics Learning

DGX agent

arXiv:2605.29577v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising framework that unifies perception, reasoning, and control for robot manipulation by adap

safetyarxiv-cs-cv
29 May 2026
Safety

Modality Alignment across Trees on Heterogeneous Hyperbolic Manifolds

DGX agent

arXiv:2510.27391v2 Announce Type: replace Abstract: Modality alignment is critical for vision-language models (VLMs) to effectively integrate information across modalities. However, existing methods e

safetyarxiv-cs-cv
29 May 2026
Safety

Model Fusion via Retrofitting

DGX agent

arXiv:2507.00037v2 Announce Type: replace-cross Abstract: Model fusion seeks to combine independently trained neural networks into a single model without retraining, but is complicated by representati

safetyarxiv-cs-ai
29 May 2026
Safety

Modeling Hierarchical Thinking in Large Reasoning Models

DGX agent

arXiv:2510.22437v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) solve complex tasks by generating long Chain-of-Thought (CoT) sequences; however, the emergent dynamics governing reas

safetyarxiv-cs-ai
29 May 2026
Safety

Modularizing Educational LLM-Agency for Fostering Responsible Learning Assistance

DGX agent

arXiv:2605.30187v1 Announce Type: new Abstract: The widespread adoption of AI chatbots in education will drastically change learning, making responsible deployment a critical concern. While large lang

safetyarxiv-cs-ai
29 May 2026
Safety

MonoPhysics: Estimating Geometry, Appearance, and Physical Parameters from Monocular Videos

DGX agent

arXiv:2605.30320v1 Announce Type: new Abstract: Existing inverse physics methods recover physical parameters from multi-view videos, where geometric constraints across views resolve scale and 3D struc

safetyarxiv-cs-cv
29 May 2026
Safety

Native Audio-Visual Alignment for Generation

DGX agent

arXiv:2605.30073v1 Announce Type: new Abstract: Joint audio-video generation aims to synthesize temporally synchronized and semantically coherent visual-acoustic content. However, existing open-source

safetyarxiv-cs-cv
29 May 2026
Safety

Neural Operator-Based Surrogate Model for CFD:Helical Coil Steam Generator in Small Modular Reactor

DGX agent

arXiv:2605.30277v1 Announce Type: new Abstract: Real-time thermal-hydraulic simulation is essential for digital twin (DT) technology that supports the safe and efficient operation of small modular rea

safetyarxiv-cs-lg
29 May 2026
Safety

Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition

DGX agent

arXiv:2505.05968v3 Announce Type: replace Abstract: Offline cooperative multi-agent reinforcement learning (MARL) faces unique challenges due to distributional shifts, particularly stemming from the h

safetyarxiv-cs-lg
29 May 2026
Safety

Offline Reinforcement Learning with Generative Trajectory Policies

DGX agent

arXiv:2510.11499v2 Announce Type: replace-cross Abstract: Generative models have emerged as a powerful class of policies for offline reinforcement learning (RL) due to their ability to capture complex

safetyarxiv-cs-ai
29 May 2026
Safety

On the Geometry of Games and their Solvers

DGX agent

arXiv:2605.29919v1 Announce Type: new Abstract: A central challenge in game theory and learning systems such as GANs is understanding which algorithms can efficiently compute equilibria across the het

safetyarxiv-cs-ai
29 May 2026
Safety

Online Fair Division with Additional Information

DGX agent

arXiv:2505.24503v3 Announce Type: replace-cross Abstract: We study the problem of fairly allocating indivisible goods to agents in an online setting, where goods arrive sequentially and must be alloca

safetyarxiv-cs-ai
29 May 2026
Safety

Open Problem: Separating Geometric and Algorithmic Compression via Cayley-Table Completion

DGX agent

arXiv:2605.29885v1 Announce Type: new Abstract: Modern statistical learning theory and deep learning characterize generalization primarily in terms of continuous capacity control (e.g., norm-based reg

safetyarxiv-cs-lg
29 May 2026
Safety

Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation

DGX agent

arXiv:2605.29390v1 Announce Type: new Abstract: Text-to-image (T2I) models have become increasingly capable of generating high-quality images. Yet, enforcing the explicit absence of a specified object

safetyarxiv-cs-cv
29 May 2026
Safety

Paper Agents, Paper Gains: An Empirical Analysis of DeFi Investment Agents

DGX agent

arXiv:2605.29174v1 Announce Type: new Abstract: DeFi investment agents, systems that use AI for autonomous on-chain trading, have attained over USD 3 billion in combined token valuations since late 20

safetyarxiv-cs-ai
29 May 2026
Safety

Path-Space Mirror Descent for On-Policy Reinforcement Learning under the Generalized Schrodinger Bridge

DGX agent

arXiv:2603.21621v2 Announce Type: replace Abstract: Classical on-policy algorithms such as PPO and mirror descent policy optimization provide stable proximal policy updates through tractable action li

safetyarxiv-cs-lg
29 May 2026
Safety

PEARL: Training Socratic Tutors with Pedagogically Aligned Reinforcement Learning

DGX agent

arXiv:2605.29582v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promise as educational tutors, yet effective tutoring requires more than solving problems: it must provide pro

safetyarxiv-cs-cl
29 May 2026
Safety

Permutation-Invariant Spectral Learning via Dyson Diffusion

DGX agent

arXiv:2510.08535v2 Announce Type: replace-cross Abstract: Diffusion models are central to generative modeling and have been adapted to graphs by diffusing adjacency matrix representations. The challen

safetyarxiv-cs-lg
29 May 2026
Safety

PersonaAgent: Bridging Memory and Action for Personalized LLM Agents

DGX agent

arXiv:2506.06254v2 Announce Type: replace Abstract: Large Language Model (LLM) empowered agents have recently emerged as advanced paradigms that exhibit impressive capabilities in a wide range of doma

safetyarxiv-cs-ai
29 May 2026
Safety

Phase-Conditioned Imitation Learning with Autonomous Failure Recovery for Robust Deformable Object Manipulation

DGX agent

arXiv:2605.29407v1 Announce Type: new Abstract: This paper presents a phase-conditioned, force-aware framework for robust deformable object manipulation. Standard imitation learning policies such as A

safetyarxiv-cs-ro
29 May 2026
Safety

Plan, Don't Pose: Long Composite Motion Generation with Text-Aligned BFM

DGX agent

arXiv:2605.29906v1 Announce Type: new Abstract: Text-to-motion (T2M) generation has broad applications in character animation, virtual avatars, and human-robot interaction. Existing methods typically

safetyarxiv-cs-lg
29 May 2026
Safety

Position: Stop Chasing the C-index when Evaluating Survival Analysis Models

DGX agent

arXiv:2506.02075v2 Announce Type: replace-cross Abstract: The current state of evaluation in survival analysis is plagued by the persistent use of evaluation metrics in ways that are misaligned with t

safetyarxiv-cs-lg
29 May 2026
Safety

Practitioner Beliefs and Behaviors in AI-Enhanced Education: DOT Framework Survey Evidence

DGX agent

arXiv:2605.29041v1 Announce Type: new Abstract: This study reports findings from a cross-sectional survey (n = 72) of higher education practitioners examining beliefs, behaviors, and institutional con

safetyarxiv-cs-ai
29 May 2026
Safety

PRO-CUA: Process-Reward Optimization for Computer Use Agents

DGX agent

arXiv:2605.29119v1 Announce Type: new Abstract: Computer use agents (CUAs) have shown strong potential for automating complex digital workflows, yet their training remains constrained by costly live e

safetyarxiv-cs-ai
29 May 2026
Safety

Promoting Generalization for Exact Solvers via Adversarial Instance Augmentation

DGX agent

arXiv:2310.14161v2 Announce Type: replace Abstract: Machine learning has been successfully applied to improve the efficiency of Mixed-Integer Linear Programming (MILP) solvers. However, the learning-b

safetyarxiv-cs-lg
29 May 2026
Safety

Provably Secure Agent Guardrail

DGX agent

arXiv:2605.29251v1 Announce Type: new Abstract: As large language models transition from bounded generative engines to agents with expansive execution privileges, AI going out of control precipitates

safetyarxiv-cs-ai
29 May 2026
Safety

Quantifying and Optimizing Simplicity via Polynomial Representations

DGX agent

arXiv:2605.29823v1 Announce Type: new Abstract: Deep networks often exhibit a preference for 'simple' solutions, and such a simplicity bias is widely believed to play a key role in generalization. Yet

safetyarxiv-cs-ai
29 May 2026
Safety

Quotient DAGs for Off-Policy Evaluation:Forward-Flow Importance Sampling and Exact Slate Propensities

DGX agent

arXiv:2605.29500v1 Announce Type: cross Abstract: Off-policy evaluation estimates how a target policy would perform using data collected by a different behavior policy, which is crucial when online te

safetyarxiv-cs-ai
29 May 2026
Safety

Real-Time Retargeting Using Controllability Boundary for Chandrayaan-3 Lunar Landing

DGX agent

arXiv:2605.29412v1 Announce Type: cross Abstract: This paper presents the real-time retargeting guidance policy developed for the Chandrayaan-3 lunar landing mission. The baseline guidance generates a

safetyarxiv-cs-lg
29 May 2026
Safety

Reasoning While Asking: Transforming Reasoning Large Language Models from Passive Solvers to Proactive Inquirers

DGX agent

arXiv:2601.22139v2 Announce Type: replace-cross Abstract: Reasoning-oriented Large Language Models (LLMs) have achieved remarkable progress with Chain-of-Thought (CoT) prompting, yet they remain funda

safetyarxiv-cs-ai
29 May 2026
Safety

ReasonLight: A Multimodal Foundation Model-Enhanced Reinforcement Learning Framework for Zero-Shot Traffic Signal Control

DGX agent

arXiv:2605.29425v1 Announce Type: new Abstract: Reinforcement learning (RL) has shown promise in traffic signal control (TSC). However, its reliance on predefined states limits responsiveness to obser

safetyarxiv-cs-ai
29 May 2026
Safety

Recovering Policy-Induced Errors: Benchmarking and Trajectory Synthesis for Robust GUI Agents

DGX agent

arXiv:2605.29447v1 Announce Type: cross Abstract: While GUI agents have advanced rapidly, they often lack the robustness to recover from their own errors, hindering real-world deployment. To bridge th

safetyarxiv-cs-cl
29 May 2026
Safety

Recurrent Structural Policy Gradient for Partially Observable Mean Field Games

DGX agent

arXiv:2602.20141v2 Announce Type: replace Abstract: Mean Field Games (MFGs) provide a principled framework for modelling interactions in large population systems. However, algorithmic progress has bee

safetyarxiv-cs-ai
29 May 2026
Safety

Reinforcing Few-step Generators via Reward-Tilted Distribution Matching

DGX agent

arXiv:2605.26108v2 Announce Type: replace Abstract: Recent advances in few-step diffusion distillation have enabled efficient image generation, yet aligning these models with human preferences remains

safetyarxiv-cs-cv
29 May 2026
Safety

Replicable Simulation-Based Robot Validation through Provenance

DGX agent

arXiv:2605.29973v1 Announce Type: new Abstract: Robot behavior is often validated through simulation-based testing, yet the replicability of such campaigns depends critically on transparent documentat

safetyarxiv-cs-ro
29 May 2026
Safety

Representation Alignment Rests on Linear Structure

DGX agent

arXiv:2605.28870v1 Announce Type: cross Abstract: We investigate the Platonic Representation Hypothesis (PRH) through a tripartite statistical framework of representations: signal, bias, and noise. {1

safetyarxiv-cs-ai
29 May 2026
Safety

Representation Signatures and Risk-Feedback Alignment in LLM Trading Agents

DGX agent

arXiv:2605.28850v1 Announce Type: new Abstract: We study behavioral alignment and representation dynamics of large language model (LLM) agents in financial decision environments. Using TradeArena, an

safetyarxiv-cs-lg
29 May 2026
Safety

Resolution as a Direction: Vector-Panning Feature Alignment for Cross-Resolution Re-Identification

DGX agent

arXiv:2510.00936v2 Announce Type: replace Abstract: Cross-resolution person re-identification (CR-ReID) remains challenging in practical surveillance, where camera quality and capture distance lead to

safetyarxiv-cs-cv
29 May 2026
Safety

Resolving Endpoint Underfitting in Diffusion Bridges via Noise Alignment

DGX agent

arXiv:2605.28962v1 Announce Type: new Abstract: Diffusion bridge models offer a powerful framework for connecting two data distributions, such as in image restoration and translation. Many existing me

safetyarxiv-cs-cv
29 May 2026
Safety

REST3D: Reconstructing Physically Stable 3D Scenes from a Single Image

DGX agent

arXiv:2605.30338v1 Announce Type: new Abstract: Reconstructing physically stable 3D scenes from a single RGB image enables casual images to be converted into simulation-ready digital assets for applic

safetyarxiv-cs-cv
29 May 2026
Safety

Review Arcade: On the Human Alignment and Gameability of LLM Reviews

DGX agent

arXiv:2605.28897v1 Announce Type: new Abstract: LLM-generated reviews for scientific papers are gaining considerable traction and are even being officially piloted by major conferences. We have to ass

safetyarxiv-cs-ai
29 May 2026
Safety

RL2ML: Finite-Rollout Surrogate Objectives from Reinforcement Learning to Maximum Likelihood

DGX agent

arXiv:2605.30154v1 Announce Type: new Abstract: Correctness-based Reinforcement Learning with Verifiable Rewards (RLVR) trains language models from binary feedback on sampled outputs, but the objectiv

safetyarxiv-cs-lg
29 May 2026
Safety

Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training

DGX agent

arXiv:2603.00454v2 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) enable fine-tuning large language models to approximate reward-proportional posteriors, but they remain p

safetyarxiv-cs-ai
29 May 2026
Safety

RUBRIC-ARROW: Alternating Pointwise Rubric Reward Modeling for LLM Post-training in Non-verifiable Domains

DGX agent

arXiv:2605.29156v1 Announce Type: cross Abstract: Pointwise reward modeling offers critical signals for LLM post-training, yet struggles with absolute scoring in subjective, non-verifiable settings. R

safetyarxiv-cs-cl
29 May 2026
Safety

Rubric-Guided Process Reward for Stepwise Model Routing

DGX agent

arXiv:2605.29310v1 Announce Type: new Abstract: Stepwise model routing improves the efficiency of Large Reasoning Models (LRMs) by assigning each reasoning step to a suitable model. Recent methods for

safetyarxiv-cs-ai
29 May 2026
Safety

SAHG: Sector-Anisotropic Hyperbolic Graph Model for Social Bot Detection

DGX agent

arXiv:2605.30166v1 Announce Type: cross Abstract: LLM-driven social bots can generate fluent, human-like text, reducing the discriminative advantage of content-based detection alone. However, coordina

safetyarxiv-cs-lg
29 May 2026
Safety

Same Evidence, Different Answers: Canonical-Context On-Policy Distillation for Multi-Turn Language Models

DGX agent

arXiv:2605.30251v1 Announce Type: cross Abstract: Large language models (LLMs) often solve a task when all instructions are given in a single prompt, but fail when the same information is revealed gra

safetyarxiv-cs-ai
29 May 2026
← Previous
1…151152153154155…260
Next →