AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

BoundRL: Efficient Structured Text Segmentation through Reinforced Boundary Generation

DGX agent

arXiv:2510.20151v2 Announce Type: replace Abstract: Structured texts refer to texts containing structured elements beyond plain texts, such as code snippets and placeholders. Such structured texts inc

safetyarxiv-cs-cl
17 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Calibration-Gated LLM Pseudo-Observations for Online Contextual Bandits

DGX agent

arXiv:2604.14961v1 Announce Type: new Abstract: Contextual bandit algorithms suffer from high regret during cold-start, when the learner has insufficient data to distinguish good arms from bad. We pro

safetyarxiv-cs-lg
17 Apr 2026
Safety

Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs

DGX agent

arXiv:2604.14520v1 Announce Type: new Abstract: Omni-modal Large Language Models (Omni-MLLMs) promise a unified integration of diverse sensory streams. However, recent evaluations reveal a critical pe

safetyarxiv-cs-cv
17 Apr 2026
Safety

ClimateCause: Complex and Implicit Causal Structures in Climate Reports

DGX agent

arXiv:2604.14856v1 Announce Type: new Abstract: Understanding climate change requires reasoning over complex causal networks. Yet, existing causal discovery datasets predominantly capture explicit, di

safetyarxiv-cs-cl
17 Apr 2026
Safety

ConfLayers: Adaptive Confidence-based Layer Skipping for Self-Speculative Decoding

DGX agent

arXiv:2604.14612v1 Announce Type: cross Abstract: Self-speculative decoding is an inference technique for large language models designed to speed up generation without sacrificing output quality. It c

safetyarxiv-cs-cl
17 Apr 2026
Safety

Continuous-time reinforcement learning: ellipticity enables model-free value function approximation

DGX agent

arXiv:2602.06930v2 Announce Type: replace Abstract: We study off-policy reinforcement learning for controlling continuous-time Markov diffusion processes with discrete-time observations and actions. W

safetyarxiv-cs-lg
17 Apr 2026
Safety

Controllable Video Object Insertion via Multiview Priors

DGX agent

arXiv:2604.14556v1 Announce Type: new Abstract: Video object insertion is a critical task for dynamically inserting new objects into existing environments. Previous video generation methods focus prim

safetyarxiv-cs-cv
17 Apr 2026
Safety

Crowdsourcing of Real-world Image Annotation via Visual Properties

DGX agent

arXiv:2604.14449v1 Announce Type: new Abstract: Recent advances in data-centric artificial intelligence highlight inherent limitations in object recognition datasets. One of the primary issues stems f

safetyarxiv-cs-cv
17 Apr 2026
Safety

CURA: Clinical Uncertainty Risk Alignment for Language Model-Based Risk Prediction

DGX agent

arXiv:2604.14651v1 Announce Type: new Abstract: Clinical language models (LMs) are increasingly applied to support clinical risk prediction from free-text notes, yet their uncertainty estimates often

safetyarxiv-cs-cl
17 Apr 2026
Safety

Diagnosing and Improving Diffusion Models by Estimating the Optimal Loss Value

DGX agent

arXiv:2506.13763v2 Announce Type: replace-cross Abstract: Diffusion models have achieved remarkable success in generative modeling. Despite more stable training, the loss of diffusion models is not in

safetyarxiv-cs-cv
17 Apr 2026
Safety

Direct Preference Optimization for Primitive-Enabled Hierarchical RL: A Bilevel Approach

DGX agent

arXiv:2411.00361v4 Announce Type: replace Abstract: Hierarchical reinforcement learning (HRL) enables agents to solve complex, long-horizon tasks by decomposing them into manageable sub-tasks. However

safetyarxiv-cs-lg
17 Apr 2026
Safety

Do Not Step Into the Same River Twice: Learning to Reason from Trial and Error

DGX agent

arXiv:2510.26109v4 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly boosted the reasoning capability of language models (LMs). However, existing

safetyarxiv-cs-lg
17 Apr 2026
Safety

DockAnywhere: Data-Efficient Visuomotor Policy Learning for Mobile Manipulation via Novel Demonstration Generation

DGX agent

arXiv:2604.15023v1 Announce Type: new Abstract: Mobile manipulation is a fundamental capability that enables robots to interact in expansive environments such as homes and factories. Most existing app

safetyarxiv-cs-ro
17 Apr 2026
Safety

Emergent Neural Automaton Policies: Learning Symbolic Structure from Visuomotor Trajectories

DGX agent

arXiv:2603.25903v2 Announce Type: replace Abstract: Scaling robot learning to long-horizon tasks remains a formidable challenge. While end-to-end policies often lack the structural priors needed for e

safetyarxiv-cs-ro
17 Apr 2026
Safety

Enhancing LLM-based Search Agents via Contribution Weighted Group Relative Policy Optimization

DGX agent

arXiv:2604.14267v1 Announce Type: new Abstract: Search agents extend Large Language Models (LLMs) beyond static parametric knowledge by enabling access to up-to-date and long-tail information unavaila

safetyarxiv-cs-lg
17 Apr 2026
Safety

Exploration and Exploitation Errors Are Measurable for Language Model Agents

DGX agent

arXiv:2604.13151v1 Announce Type: new Abstract: Language Model (LM) agents are increasingly used in complex open-ended decision-making tasks, from AI coding to physical AI. A core requirement in these

safetyarxiv-cs-ai
17 Apr 2026
Safety

Filling in the Mechanisms: How do LMs Learn Filler-Gap Dependencies under Developmental Constraints?

DGX agent

arXiv:2604.14459v1 Announce Type: new Abstract: For humans, filler-gap dependencies require a shared representation across different syntactic constructions. Although causal analyses suggest this may

safetyarxiv-cs-cl
17 Apr 2026
Safety

Flow with the Force Field: Learning 3D Compliant Flow Matching Policies from Force and Demonstration-Guided Simulation Data

DGX agent

arXiv:2510.02738v3 Announce Type: replace-cross Abstract: While visuomotor policy has made advancements in recent years, contact-rich tasks still remain a challenge. Robotic manipulation tasks that re

safetyarxiv-cs-lg
17 Apr 2026
Safety

From Boundaries to Semantics: Prompt-Guided Multi-Task Learning for Petrographic Thin-section Segmentation

DGX agent

arXiv:2604.14805v1 Announce Type: new Abstract: Grain-edge segmentation (GES) and lithology semantic segmentation (LSS) are two pivotal tasks for quantifying rock fabric and composition. However, thes

safetyarxiv-cs-cv
17 Apr 2026
Safety

From Plausible to Causal: Counterfactual Semantics for Policy Evaluation in Simulated Online Communities

DGX agent

arXiv:2604.03920v2 Announce Type: replace Abstract: LLM-based social simulations can generate believable community interactions, enabling ``policy wind tunnels'' where governance interventions are tes

safetyarxiv-cs-cl
17 Apr 2026
Safety

GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification

DGX agent

arXiv:2604.14258v1 Announce Type: cross Abstract: Large language models are typically post-trained using supervised fine-tuning (SFT) and reinforcement learning (RL), yet effectively unifying efficien

safetyarxiv-cs-lg
17 Apr 2026
Safety

GSNR: Graph Smooth Null-Space Representation for Inverse Problems

DGX agent

arXiv:2602.20328v2 Announce Type: replace Abstract: Inverse problems in imaging are ill-posed, leading to infinitely many solutions consistent with the measurements due to the non-trivial null-space o

safetyarxiv-cs-cv
17 Apr 2026
Safety

Heat and Matern Kernels on Matchings

DGX agent

arXiv:2604.14331v1 Announce Type: new Abstract: Applying kernel methods to matchings is challenging due to their discrete, non-Euclidean nature. In this paper, we develop a principled framework for co

safetyarxiv-cs-lg
17 Apr 2026
Safety

Hierarchical Retrieval Augmented Generation for Adversarial Technique Annotation in Cyber Threat Intelligence Text

DGX agent

arXiv:2604.14166v1 Announce Type: new Abstract: Mapping Cyber Threat Intelligence (CTI) text to MITRE ATT&CK technique IDs is a critical task for understanding adversary behaviors and automating threa

safetyarxiv-cs-cl
17 Apr 2026
Safety

Hybrid Latents -- Geometry-Appearance-Aware Surfel Splatting

DGX agent

arXiv:2604.14928v1 Announce Type: new Abstract: We introduce a hybrid Gaussian-hash-grid radiance representation for reconstructing 2D Gaussian scene models from multi-view images. Similar to NeST spl

safetyarxiv-cs-cv
17 Apr 2026
Safety

IG-Search: Step-Level Information Gain Rewards for Search-Augmented Reasoning

DGX agent

arXiv:2604.15148v1 Announce Type: cross Abstract: Reinforcement learning has emerged as an effective paradigm for training large language models to perform search-augmented reasoning. However, existin

safetyarxiv-cs-cl
17 Apr 2026
Safety

Implicit Neural Representations: A Signal Processing Perspective

DGX agent

arXiv:2604.15047v1 Announce Type: new Abstract: Implicit neural representations (INRs) mark a fundamental shift in signal modeling, moving from discrete sampled data to continuous functional represent

safetyarxiv-cs-cv
17 Apr 2026
Safety

Improving Machine Learning Performance with Synthetic Augmentation

DGX agent

arXiv:2604.14498v1 Announce Type: cross Abstract: Synthetic augmentation is increasingly used to mitigate data scarcity in financial machine learning, yet its statistical role remains poorly understoo

safetyarxiv-cs-lg
17 Apr 2026
Safety

Language of Thought Shapes Output Diversity in Large Language Models

DGX agent

arXiv:2601.11227v2 Announce Type: replace Abstract: Output diversity is crucial for Large Language Models as it underpins pluralism and creativity. In this work, we reveal that controlling the languag

safetyarxiv-cs-cl
17 Apr 2026
Safety

Layered Mutability: Continuity and Governance in Persistent Self-Modifying Agents

DGX agent

arXiv:2604.14717v1 Announce Type: cross Abstract: Persistent language-model agents increasingly combine tool use, tiered memory, reflective prompting, and runtime adaptation. In such systems, behavior

safetyarxiv-cs-lg
17 Apr 2026
Safety

LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories

DGX agent

arXiv:2604.15311v1 Announce Type: new Abstract: This paper focuses on the alignment of flow matching models with human preferences. A promising way is fine-tuning by directly backpropagating reward gr

safetyarxiv-cs-cv
17 Apr 2026
Safety

Learning Ad Hoc Network Dynamics via Graph-Structured World Models

DGX agent

arXiv:2604.14811v1 Announce Type: new Abstract: Ad hoc wireless networks exhibit complex, innate and coupled dynamics: node mobility, energy depletion and topology change that are difficult to model a

safetyarxiv-cs-lg
17 Apr 2026
Safety

Learning Adaptive Reasoning Paths for Efficient Visual Reasoning

DGX agent

arXiv:2604.14568v1 Announce Type: cross Abstract: Visual reasoning models (VRMs) have recently shown strong cross-modal reasoning capabilities by integrating visual perception with language reasoning.

safetyarxiv-cs-cl
17 Apr 2026
Safety

Learning to Think Like a Cartoon Captionist: Incongruity-Resolution Supervision for Multimodal Humor Understanding

DGX agent

arXiv:2604.15210v1 Announce Type: cross Abstract: Humor is one of the few cognitive tasks where getting the reasoning right matters as much as getting the answer right. While recent work evaluates hum

safetyarxiv-cs-cl
17 Apr 2026
Safety

MARS^2: Scaling Multi-Agent Tree Search via Reinforcement Learning for Code Generation

DGX agent

arXiv:2604.14564v1 Announce Type: cross Abstract: Reinforcement learning (RL) paradigms have demonstrated strong performance on reasoning-intensive tasks such as code generation. However, limited traj

safetyarxiv-cs-cl
17 Apr 2026
Safety

Mean Flow Policy Optimization

DGX agent

arXiv:2604.14698v1 Announce Type: new Abstract: Diffusion models have recently emerged as expressive policy representations for online reinforcement learning (RL). However, their iterative generative

safetyarxiv-cs-lg
17 Apr 2026
Safety

Meituan Merchant Business Diagnosis via Policy-Guided Dual-Process User Simulation

DGX agent

arXiv:2604.15190v1 Announce Type: cross Abstract: Simulating group-level user behavior enables scalable counterfactual evaluation of merchant strategies without costly online experiments. However, bui

safetyarxiv-cs-cl
17 Apr 2026
Safety

Metric-Aware Principal Component Analysis (MAPCA):A Unified Framework for Scale-Invariant Representation Learning

DGX agent

arXiv:2604.14249v1 Announce Type: new Abstract: We introduce Metric-Aware Principal Component Analysis (MAPCA), a unified framework for scale-invariant representation learning based on the generalised

safetyarxiv-cs-lg
17 Apr 2026
Safety

Model-Free Assessment of Simulator Fidelity via Quantile Curves

DGX agent

arXiv:2512.05024v3 Announce Type: replace-cross Abstract: As generative AI models are increasingly used to simulate real-world systems, quantifying the ``sim-to-real'' gap is critical. For each input

safetyarxiv-cs-lg
17 Apr 2026
Safety

Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problem

DGX agent

arXiv:2604.14808v1 Announce Type: new Abstract: Machine unlearning for large language models (LLMs) aims to remove targeted knowledge while preserving general capability. In this paper, we recast LLM

safetyarxiv-cs-cl
17 Apr 2026
Safety

Multi-Modal Manipulation via Multi-Modal Policy Consensus

DGX agent

arXiv:2509.23468v3 Announce Type: replace-cross Abstract: Effectively integrating diverse sensory modalities is crucial for robotic manipulation. However, the typical approach of feature concatenation

safetyarxiv-cs-lg
17 Apr 2026
Safety

Multi-Persona Thinking for Bias Mitigation in Large Language Models

DGX agent

arXiv:2601.15488v2 Announce Type: replace Abstract: Large Language Models (LLMs) exhibit social biases, which can lead to harmful stereotypes and unfair outcomes. We propose extbf{Multi-Persona Thinki

safetyarxiv-cs-cl
17 Apr 2026
Safety

Multi-User mmWave Beam and Rate Adaptation via Combinatorial Satisficing Bandits

DGX agent

arXiv:2604.14908v1 Announce Type: new Abstract: We study downlink beam and rate adaptation in a multi-user mmWave MISO system where multiple base stations (BSs), each using analog beamforming from fin

safetyarxiv-cs-lg
17 Apr 2026
Safety

NG-GS: NeRF-Guided 3D Gaussian Splatting Segmentation

DGX agent

arXiv:2604.14706v1 Announce Type: new Abstract: Recent advances in 3D Gaussian Splatting (3DGS) have enabled highly efficient and photorealistic novel view synthesis. However, segmenting objects accur

safetyarxiv-cs-cv
17 Apr 2026
Safety

NLP needs Diversity outside of 'Diversity'

DGX agent

arXiv:2604.14595v1 Announce Type: new Abstract: This position paper argues that recent progress with diversity in NLP is disproportionately concentrated on a small number of areas surrounding fairness

safetyarxiv-cs-cl
17 Apr 2026
Safety

ORBIT: On-policy Exploration-Exploitation for Controllable Multi-Budget Reasoning

DGX agent

arXiv:2601.08310v2 Announce Type: replace Abstract: Recent Large Reasoning Models (LRMs) achieve strong performance by leveraging long-form Chain-of-Thought (CoT) reasoning, but uniformly applying ove

safetyarxiv-cs-lg
17 Apr 2026
Safety

Practical estimation of the optimal classification error with soft labels and calibration

DGX agent

arXiv:2505.20761v3 Announce Type: replace Abstract: While the performance of machine learning systems has experienced significant improvement in recent years, relatively little attention has been paid

safetyarxiv-cs-lg
17 Apr 2026
Safety

Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation

DGX agent

arXiv:2603.13683v2 Announce Type: replace Abstract: Although debiased large language models (LLMs) excel at handling known or low-bias prompts, they often fail on unfamiliar and high-bias prompts. We

safetyarxiv-cs-cl
17 Apr 2026
← Previous
1…222223224225226…257
Next →