AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Incentivizing High-Quality Human Annotations with Golden Questions

DGX agent

arXiv:2505.19134v2 Announce Type: replace-cross Abstract: Human-annotated data plays a vital role in training large language models (LLMs), such as supervised fine-tuning and human preference alignmen

safetyarxiv-cs-lg
15 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Information-Geometric Decomposition of Generalization Error in Unsupervised Learning

DGX agent

arXiv:2604.12340v1 Announce Type: cross Abstract: We decompose the Kullback--Leibler generalization error (GE) -- the expected KL divergence from the data distribution to the trained model -- of unsup

safetyarxiv-cs-lg
15 Apr 2026
Safety

InsightFlow: LLM-Driven Synthesis of Patient Narratives for Mental Health into Causal Models

DGX agent

arXiv:2604.12721v1 Announce Type: new Abstract: Clinical case formulation organizes patient symptoms and psychosocial factors into causal models, often using the 5P framework. However, constructing su

safetyarxiv-cs-cl
15 Apr 2026
Safety

Labeled TrustSet Guided: Batch Active Learning with Reinforcement Learning

DGX agent

arXiv:2604.12303v1 Announce Type: new Abstract: Batch active learning (BAL) is a crucial technique for reducing labeling costs and improving data efficiency in training large-scale deep learning model

safetyarxiv-cs-lg
15 Apr 2026
Safety

Learning step-level dynamic soaring in shear flow

DGX agent

arXiv:2604.12413v1 Announce Type: cross Abstract: Dynamic soaring enables sustained flight by extracting energy from wind shear, yet it is commonly understood as a cycle-level maneuver that assumes st

safetyarxiv-cs-ro
15 Apr 2026
Safety

Learning Versatile Humanoid Manipulation with Touch Dreaming

DGX agent

arXiv:2604.13015v1 Announce Type: new Abstract: Humanoid robots promise general-purpose assistance, yet real-world humanoid loco-manipulation remains challenging because it requires whole-body stabili

safetyarxiv-cs-ro
15 Apr 2026
Safety

Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation

DGX agent

arXiv:2604.13010v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as an efficient post-training paradigm for large language models. However, standard OPD requires a live teach

safetyarxiv-cs-ai
15 Apr 2026
Safety

LiveMoments: Reselected Key Photo Restoration in Live Photos via Reference-guided Diffusion

DGX agent

arXiv:2604.12286v1 Announce Type: new Abstract: Live Photo captures both a high-quality key photo and a short video clip to preserve the precious dynamics around the captured moment. While users may c

safetyarxiv-cs-cv
15 Apr 2026
Safety

Man and machine: artificial intelligence and judicial decision making

DGX agent

arXiv:2603.19042v4 Announce Type: replace Abstract: The integration of artificial intelligence (AI) technologies into judicial decision-making, particularly in pretrial, sentencing, and parole context

safetyarxiv-cs-ai
15 Apr 2026
Safety

Meet Dynamic Individual Preferences: Resolving Conflicting Human Value with Paired Fine-Tuning

DGX agent

arXiv:2604.12479v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have significantly improved the alignment of models with general human preferences. However, a major cha

safetyarxiv-cs-cl
15 Apr 2026
Safety

Models Know Their Shortcuts: Deployment-Time Shortcut Mitigation

DGX agent

arXiv:2604.12277v1 Announce Type: new Abstract: Pretrained language models often rely on superficial features that appear predictive during training yet fail to generalize at test time, a phenomenon k

safetyarxiv-cs-lg
15 Apr 2026
Safety

MODIX: A Training-Free Multimodal Information-Driven Positional Index Scaling for Vision-Language Models

DGX agent

arXiv:2604.12537v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in multimodal understanding, yet their positional encoding mechanisms remain suboptima

safetyarxiv-cs-ai
15 Apr 2026
Safety

MolMem: Memory-Augmented Agentic Reinforcement Learning for Sample-Efficient Molecular Optimization

DGX agent

arXiv:2604.12237v1 Announce Type: cross Abstract: In drug discovery, molecular optimization aims to iteratively refine a lead compound to improve molecular properties while preserving structural simil

safetyarxiv-cs-ai
15 Apr 2026
Safety

Mutual Information Surprise: Rethinking Unexpectedness in Autonomous Systems

DGX agent

arXiv:2508.17403v3 Announce Type: replace Abstract: A community of researchers appears to think that a machine can be surprised and have introduced various surprise measures, principally the Shannon S

safetyarxiv-cs-lg
15 Apr 2026
Safety

No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning

DGX agent

arXiv:2601.06794v2 Announce Type: replace Abstract: Critique-guided reinforcement learning (RL) has emerged as a powerful paradigm for training LLM agents by augmenting sparse outcome rewards with nat

safetyarxiv-cs-ai
15 Apr 2026
Safety

Not All Turns Are Equally Hard: Adaptive Thinking Budgets For Efficient Multi-Turn Reasoning

DGX agent

arXiv:2604.05164v2 Announce Type: replace-cross Abstract: As LLM reasoning performance plateau, improving inference-time compute efficiency is crucial to mitigate overthinking and long thinking traces

safetyarxiv-cs-ai
15 Apr 2026
Safety

Offline-Online Reinforcement Learning for Linear Mixture MDPs

DGX agent

arXiv:2604.11994v1 Announce Type: new Abstract: We study offline-online reinforcement learning in linear mixture Markov decision processes (MDPs) under environment shift. In the offline phase, data ar

safetyarxiv-cs-lg
15 Apr 2026
Safety

PAINT: Partner-Agnostic Intent-Aware Cooperative Transport with Legged Robots

DGX agent

arXiv:2604.12852v1 Announce Type: new Abstract: Collaborative transport requires robots to infer partner intent through physical interaction while maintaining stable loco-manipulation. This becomes pa

safetyarxiv-cs-ro
15 Apr 2026
Safety

Perception-Aware Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2507.06448v5 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has proven to be a highly effective strategy for endowing Large Language Models (LLMs) with ro

safetyarxiv-cs-cl
15 Apr 2026
Safety

PR-MaGIC: Prompt Refinement Via Mask Decoder Gradient Flow For In-Context Segmentation

DGX agent

arXiv:2604.12113v1 Announce Type: cross Abstract: Visual Foundation Models (VFMs) such as the Segment Anything Model (SAM) have significantly advanced broad use of image segmentation. However, SAM and

safetyarxiv-cs-ai
15 Apr 2026
Safety

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation

DGX agent

arXiv:2511.17097v2 Announce Type: replace Abstract: Vision-Language Navigation requires agents to act coherently over long horizons by understanding not only local visual context but also how far they

safetyarxiv-cs-ro
15 Apr 2026
Safety

PubSwap: Public-Data Off-Policy Coordination for Federated RLVR

DGX agent

arXiv:2604.12160v1 Announce Type: new Abstract: Reasoning post-training with reinforcement learning from verifiable rewards (RLVR) is typically studied in centralized settings, yet many realistic appl

safetyarxiv-cs-lg
15 Apr 2026
Safety

Redefining Quality Criteria and Distance-Aware Score Modeling for Image Editing Assessment

DGX agent

arXiv:2604.12175v1 Announce Type: new Abstract: Recent advances in image editing have heightened the need for reliable Image Editing Quality Assessment (IEQA). Unlike traditional methods, IEQA require

safetyarxiv-cs-cv
15 Apr 2026
Safety

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models

DGX agent

arXiv:2604.12582v1 Announce Type: new Abstract: Recent Video Large Language Models (Video-LLMs) have demonstrated strong capability in video understanding, yet they still suffer from hallucinations. E

safetyarxiv-cs-cv
15 Apr 2026
Safety

Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

DGX agent

arXiv:2604.13016v1 Announce Type: cross Abstract: On-policy distillation (OPD) has become a core technique in the post-training of large language models, yet its training dynamics remain poorly unders

safetyarxiv-cs-ai
15 Apr 2026
Safety

Retrieval as a Decision: Training-Free Adaptive Gating for Efficient RAG

DGX agent

arXiv:2511.09803v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) improves factuality but retrieving for every query often hurts quality while inflating tokens and latency. We p

safetyarxiv-cs-cl
15 Apr 2026
Safety

SAM3-I: Segment Anything with Instructions

DGX agent

arXiv:2512.04585v3 Announce Type: replace Abstract: Segment Anything Model 3 (SAM3) advances open-vocabulary segmentation through promptable concept segmentation, enabling users to segment all instanc

safetyarxiv-cs-cv
15 Apr 2026
Safety

Scaffold-Conditioned Preference Triplets for Controllable Molecular Optimization with Large Language Models

DGX agent

arXiv:2604.12350v1 Announce Type: cross Abstract: Molecular property optimization is central to drug discovery, yet many deep learning methods rely on black-box scoring and offer limited control over

safetyarxiv-cs-ai
15 Apr 2026
Safety

Scalable and General Whole-Body Control for Cross-Humanoid Locomotion

DGX agent

arXiv:2602.05791v2 Announce Type: replace Abstract: Learning-based whole-body controllers have become a key driver for humanoid robots, yet most existing approaches require robot-specific training. In

safetyarxiv-cs-ro
15 Apr 2026
Safety

Schema-Adaptive Tabular Representation Learning with LLMs for Generalizable Multimodal Clinical Reasoning

DGX agent

arXiv:2604.11835v1 Announce Type: cross Abstract: Machine learning for tabular data remains constrained by poor schema generalization, a challenge rooted in the lack of semantic understanding of struc

safetyarxiv-cs-ai
15 Apr 2026
Safety

Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision

DGX agent

arXiv:2604.12002v1 Announce Type: new Abstract: Current post-training methods in verifiable settings fall into two categories. Reinforcement learning (RLVR) relies on binary rewards, which are broadly

safetyarxiv-cs-cl
15 Apr 2026
Safety

Simulation as Supervision: Mechanistic Pretraining for Scientific Discovery

DGX agent

arXiv:2507.08977v4 Announce Type: replace-cross Abstract: Scientific modeling faces a tradeoff between the interpretability of mechanistic theory and the predictive power of machine learning. While ex

safetyarxiv-cs-ai
15 Apr 2026
Safety

SOAR: Self-Correction for Optimal Alignment and Refinement in Diffusion Models

DGX agent

arXiv:2604.12617v1 Announce Type: cross Abstract: The post-training pipeline for diffusion models currently has two stages: supervised fine-tuning (SFT) on curated data and reinforcement learning (RL)

safetyarxiv-cs-ai
15 Apr 2026
Safety

StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback

DGX agent

arXiv:2510.20093v2 Announce Type: replace-cross Abstract: Although recent advancements in diffusion models have significantly enriched the quality of generated images, challenges remain in synthesizin

safetyarxiv-cs-ai
15 Apr 2026
Safety

Task Alignment: A simple and effective proxy for model merging in computer vision

DGX agent

arXiv:2604.12935v1 Announce Type: new Abstract: Efficiently merging several models fine-tuned for different tasks, but stemming from the same pretrained base model, is of great practical interest. Des

safetyarxiv-cs-cv
15 Apr 2026
Safety

Teaching LLMs Human-Like Editing of Inappropriate Argumentation via Reinforcement Learning

DGX agent

arXiv:2604.12770v1 Announce Type: new Abstract: Editing human-written text has become a standard use case of large language models (LLMs), for example, to make one's arguments more appropriate for a d

safetyarxiv-cs-cl
15 Apr 2026
Safety

The role of System 1 and System 2 semantic memory structure in human and LLM biases

DGX agent

arXiv:2604.12816v1 Announce Type: new Abstract: Implicit biases in both humans and large language models (LLMs) pose significant societal risks. Dual process theories propose that biases arise primari

safetyarxiv-cs-cl
15 Apr 2026
Safety

The Stackelberg Speaker: Optimizing Persuasive Communication in Social Deduction Games

DGX agent

arXiv:2510.09087v2 Announce Type: replace Abstract: Large language model (LLM) agents have shown remarkable progress in social deduction games (SDGs). However, existing approaches primarily focus on i

safetyarxiv-cs-ai
15 Apr 2026
Safety

Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training

DGX agent

arXiv:2509.25758v2 Announce Type: replace Abstract: The remarkable capabilities of modern large reasoning models are largely unlocked through post-training techniques such as supervised fine-tuning (S

safetyarxiv-cs-ai
15 Apr 2026
Safety

Token-Level Policy Optimization: Linking Group-Level Rewards to Token-Level Aggregation via Sequence-Level Likelihood

DGX agent

arXiv:2604.12736v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has significantly advanced the reasoning ability of large language models (LLMs), particularly in their mathem

safetyarxiv-cs-cl
15 Apr 2026
Safety

Towards Generalized Certified Robustness with Multi-Norm Training

DGX agent

arXiv:2410.03000v3 Announce Type: replace Abstract: Existing certified training methods can only train models to be robust against a certain perturbation type (e.g. l_infty or l_2). However, an l_inft

safetyarxiv-cs-lg
15 Apr 2026
Safety

Towards Platonic Representation for Table Reasoning: A Foundation for Permutation-Invariant Retrieval

DGX agent

arXiv:2604.12133v1 Announce Type: new Abstract: Historical approaches to Table Representation Learning (TRL) have largely adopted the sequential paradigms of Natural Language Processing (NLP). We argu

safetyarxiv-cs-ai
15 Apr 2026
Safety

WebChain: A Large-Scale Human-Annotated Dataset of Real-World Web Interaction Traces

DGX agent

arXiv:2603.05295v3 Announce Type: replace Abstract: We introduce WebChain, the largest open-source dataset of human-annotated trajectories on real-world websites, designed to accelerate reproducible r

safetyarxiv-cs-ai
15 Apr 2026
Safety

Whole-Body Mobile Manipulation using Offline Reinforcement Learning on Sub-optimal Controllers

DGX agent

arXiv:2604.12509v1 Announce Type: cross Abstract: Mobile Manipulation (MoMa) of articulated objects, such as opening doors, drawers, and cupboards, demands simultaneous, whole-body coordination betwee

safetyarxiv-cs-cv
15 Apr 2026
Safety

WiseOWL: A Methodology for Evaluating Ontological Descriptiveness and Semantic Correctness for Ontology Reuse and Ontology Recommendations

DGX agent

arXiv:2604.12025v1 Announce Type: new Abstract: The Semantic Web standardizes concept meaning for humans and machines, enabling machine-operable content and consistent interpretation that improves adv

safetyarxiv-cs-ai
15 Apr 2026
Safety

XRZero-G0: Pushing the Frontier of Dexterous Robotic Manipulation with Interfaces, Quality and Ratios

DGX agent

arXiv:2604.13001v1 Announce Type: new Abstract: The acquisition of high-quality, action-aligned demonstration data remains a fundamental bottleneck in scaling foundation models for dexterous robot man

safetyarxiv-cs-ro
15 Apr 2026
Safety

3D Multi-View Stylization with Pose-Free Correspondences Matching for Robust 3D Geometry Preservation

DGX agent

arXiv:2604.09639v1 Announce Type: new Abstract: Artistic style transfer is well studied for images and videos, but extending it to multi-view 3D scenes remains difficult because stylization can disrup

safetyarxiv-cs-cv
14 Apr 2026
Safety

A Comparative Theoretical Analysis of Entropy Control Methods in Reinforcement Learning

DGX agent

arXiv:2604.09676v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a key approach for enhancing reasoning in large language models (LLMs), yet scalable training is often hindered

safetyarxiv-cs-ai
14 Apr 2026
← Previous
1…226227228229230…257
Next →