AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,351 results
Safety

Do Not Step Into the Same River Twice: Learning to Reason from Trial and Error

DGX agent

arXiv:2510.26109v4 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly boosted the reasoning capability of language models (LMs). However, existing

safetyarxiv-cs-lg
17 Apr 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

DockAnywhere: Data-Efficient Visuomotor Policy Learning for Mobile Manipulation via Novel Demonstration Generation

DGX agent

arXiv:2604.15023v1 Announce Type: new Abstract: Mobile manipulation is a fundamental capability that enables robots to interact in expansive environments such as homes and factories. Most existing app

safetyarxiv-cs-ro
17 Apr 2026
Safety

Emergent Neural Automaton Policies: Learning Symbolic Structure from Visuomotor Trajectories

DGX agent

arXiv:2603.25903v2 Announce Type: replace Abstract: Scaling robot learning to long-horizon tasks remains a formidable challenge. While end-to-end policies often lack the structural priors needed for e

safetyarxiv-cs-ro
17 Apr 2026
Safety

Enhancing LLM-based Search Agents via Contribution Weighted Group Relative Policy Optimization

DGX agent

arXiv:2604.14267v1 Announce Type: new Abstract: Search agents extend Large Language Models (LLMs) beyond static parametric knowledge by enabling access to up-to-date and long-tail information unavaila

safetyarxiv-cs-lg
17 Apr 2026
Safety

Everyone’s quitting but everything’s going great. AGI will be here next week. Scout’s honor!

DGX agent

Everyone’s quitting but everything’s going great. AGI will be here next week. Scout’s honor! This story has now been updated with more details. Three leaders departed from OpenAI today: - Kevin Weil,

safetygary-marcus--x
17 Apr 2026
Safety

Exploration and Exploitation Errors Are Measurable for Language Model Agents

DGX agent

arXiv:2604.13151v1 Announce Type: new Abstract: Language Model (LM) agents are increasingly used in complex open-ended decision-making tasks, from AI coding to physical AI. A core requirement in these

safetyarxiv-cs-ai
17 Apr 2026
Safety

Filling in the Mechanisms: How do LMs Learn Filler-Gap Dependencies under Developmental Constraints?

DGX agent

arXiv:2604.14459v1 Announce Type: new Abstract: For humans, filler-gap dependencies require a shared representation across different syntactic constructions. Although causal analyses suggest this may

safetyarxiv-cs-cl
17 Apr 2026
Safety

Flow with the Force Field: Learning 3D Compliant Flow Matching Policies from Force and Demonstration-Guided Simulation Data

DGX agent

arXiv:2510.02738v3 Announce Type: replace-cross Abstract: While visuomotor policy has made advancements in recent years, contact-rich tasks still remain a challenge. Robotic manipulation tasks that re

safetyarxiv-cs-lg
17 Apr 2026
Safety

From Boundaries to Semantics: Prompt-Guided Multi-Task Learning for Petrographic Thin-section Segmentation

DGX agent

arXiv:2604.14805v1 Announce Type: new Abstract: Grain-edge segmentation (GES) and lithology semantic segmentation (LSS) are two pivotal tasks for quantifying rock fabric and composition. However, thes

safetyarxiv-cs-cv
17 Apr 2026
Safety

From Plausible to Causal: Counterfactual Semantics for Policy Evaluation in Simulated Online Communities

DGX agent

arXiv:2604.03920v2 Announce Type: replace Abstract: LLM-based social simulations can generate believable community interactions, enabling ``policy wind tunnels'' where governance interventions are tes

safetyarxiv-cs-cl
17 Apr 2026
Safety

GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification

DGX agent

arXiv:2604.14258v1 Announce Type: cross Abstract: Large language models are typically post-trained using supervised fine-tuning (SFT) and reinforcement learning (RL), yet effectively unifying efficien

safetyarxiv-cs-lg
17 Apr 2026
Safety

GSNR: Graph Smooth Null-Space Representation for Inverse Problems

DGX agent

arXiv:2602.20328v2 Announce Type: replace Abstract: Inverse problems in imaging are ill-posed, leading to infinitely many solutions consistent with the measurements due to the non-trivial null-space o

safetyarxiv-cs-cv
17 Apr 2026
Safety

Heat and Matern Kernels on Matchings

DGX agent

arXiv:2604.14331v1 Announce Type: new Abstract: Applying kernel methods to matchings is challenging due to their discrete, non-Euclidean nature. In this paper, we develop a principled framework for co

safetyarxiv-cs-lg
17 Apr 2026
Safety

Hierarchical Retrieval Augmented Generation for Adversarial Technique Annotation in Cyber Threat Intelligence Text

DGX agent

arXiv:2604.14166v1 Announce Type: new Abstract: Mapping Cyber Threat Intelligence (CTI) text to MITRE ATT&CK technique IDs is a critical task for understanding adversary behaviors and automating threa

safetyarxiv-cs-cl
17 Apr 2026
Safety

Hybrid Latents -- Geometry-Appearance-Aware Surfel Splatting

DGX agent

arXiv:2604.14928v1 Announce Type: new Abstract: We introduce a hybrid Gaussian-hash-grid radiance representation for reconstructing 2D Gaussian scene models from multi-view images. Similar to NeST spl

safetyarxiv-cs-cv
17 Apr 2026
Safety

I went on @BBCNewsnight this week to discuss the recent developments in AI's capabilities, as well as the potential harms and concentration …

DGX agent

I went on @BBCNewsnight this week to discuss the recent developments in AI's capabilities, as well as the potential harms and concentration of power they could entail. We need coordinated internationa

safetyyoshua-bengio--x
17 Apr 2026
Safety

IG-Search: Step-Level Information Gain Rewards for Search-Augmented Reasoning

DGX agent

arXiv:2604.15148v1 Announce Type: cross Abstract: Reinforcement learning has emerged as an effective paradigm for training large language models to perform search-augmented reasoning. However, existin

safetyarxiv-cs-cl
17 Apr 2026
Safety

Implicit Neural Representations: A Signal Processing Perspective

DGX agent

arXiv:2604.15047v1 Announce Type: new Abstract: Implicit neural representations (INRs) mark a fundamental shift in signal modeling, moving from discrete sampled data to continuous functional represent

safetyarxiv-cs-cv
17 Apr 2026
Safety

Improving Machine Learning Performance with Synthetic Augmentation

DGX agent

arXiv:2604.14498v1 Announce Type: cross Abstract: Synthetic augmentation is increasingly used to mitigate data scarcity in financial machine learning, yet its statistical role remains poorly understoo

safetyarxiv-cs-lg
17 Apr 2026
Safety

in 1996 mitzenmacher showed that sampling two backends and picking the better one drops max load exponentially vs. random selection one extr…

DGX agent

in 1996 mitzenmacher showed that sampling two backends and picking the better one drops max load exponentially vs. random selection one extra comparison. that's the whole trick we built pinecone assis

safetypinecone--x
17 Apr 2026
Safety

Language of Thought Shapes Output Diversity in Large Language Models

DGX agent

arXiv:2601.11227v2 Announce Type: replace Abstract: Output diversity is crucial for Large Language Models as it underpins pluralism and creativity. In this work, we reveal that controlling the languag

safetyarxiv-cs-cl
17 Apr 2026
Safety

Layered Mutability: Continuity and Governance in Persistent Self-Modifying Agents

DGX agent

arXiv:2604.14717v1 Announce Type: cross Abstract: Persistent language-model agents increasingly combine tool use, tiered memory, reflective prompting, and runtime adaptation. In such systems, behavior

safetyarxiv-cs-lg
17 Apr 2026
Safety

LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories

DGX agent

arXiv:2604.15311v1 Announce Type: new Abstract: This paper focuses on the alignment of flow matching models with human preferences. A promising way is fine-tuning by directly backpropagating reward gr

safetyarxiv-cs-cv
17 Apr 2026
Safety

Learning Ad Hoc Network Dynamics via Graph-Structured World Models

DGX agent

arXiv:2604.14811v1 Announce Type: new Abstract: Ad hoc wireless networks exhibit complex, innate and coupled dynamics: node mobility, energy depletion and topology change that are difficult to model a

safetyarxiv-cs-lg
17 Apr 2026
Safety

Learning Adaptive Reasoning Paths for Efficient Visual Reasoning

DGX agent

arXiv:2604.14568v1 Announce Type: cross Abstract: Visual reasoning models (VRMs) have recently shown strong cross-modal reasoning capabilities by integrating visual perception with language reasoning.

safetyarxiv-cs-cl
17 Apr 2026
Safety

Learning to Think Like a Cartoon Captionist: Incongruity-Resolution Supervision for Multimodal Humor Understanding

DGX agent

arXiv:2604.15210v1 Announce Type: cross Abstract: Humor is one of the few cognitive tasks where getting the reasoning right matters as much as getting the answer right. While recent work evaluates hum

safetyarxiv-cs-cl
17 Apr 2026
Safety

MARS^2: Scaling Multi-Agent Tree Search via Reinforcement Learning for Code Generation

DGX agent

arXiv:2604.14564v1 Announce Type: cross Abstract: Reinforcement learning (RL) paradigms have demonstrated strong performance on reasoning-intensive tasks such as code generation. However, limited traj

safetyarxiv-cs-cl
17 Apr 2026
Safety

Mean Flow Policy Optimization

DGX agent

arXiv:2604.14698v1 Announce Type: new Abstract: Diffusion models have recently emerged as expressive policy representations for online reinforcement learning (RL). However, their iterative generative

safetyarxiv-cs-lg
17 Apr 2026
Safety

Meituan Merchant Business Diagnosis via Policy-Guided Dual-Process User Simulation

DGX agent

arXiv:2604.15190v1 Announce Type: cross Abstract: Simulating group-level user behavior enables scalable counterfactual evaluation of merchant strategies without costly online experiments. However, bui

safetyarxiv-cs-cl
17 Apr 2026
Safety

Metric-Aware Principal Component Analysis (MAPCA):A Unified Framework for Scale-Invariant Representation Learning

DGX agent

arXiv:2604.14249v1 Announce Type: new Abstract: We introduce Metric-Aware Principal Component Analysis (MAPCA), a unified framework for scale-invariant representation learning based on the generalised

safetyarxiv-cs-lg
17 Apr 2026
Safety

Model-Free Assessment of Simulator Fidelity via Quantile Curves

DGX agent

arXiv:2512.05024v3 Announce Type: replace-cross Abstract: As generative AI models are increasingly used to simulate real-world systems, quantifying the ``sim-to-real'' gap is critical. For each input

safetyarxiv-cs-lg
17 Apr 2026
Safety

Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problem

DGX agent

arXiv:2604.14808v1 Announce Type: new Abstract: Machine unlearning for large language models (LLMs) aims to remove targeted knowledge while preserving general capability. In this paper, we recast LLM

safetyarxiv-cs-cl
17 Apr 2026
Safety

Multi-Modal Manipulation via Multi-Modal Policy Consensus

DGX agent

arXiv:2509.23468v3 Announce Type: replace-cross Abstract: Effectively integrating diverse sensory modalities is crucial for robotic manipulation. However, the typical approach of feature concatenation

safetyarxiv-cs-lg
17 Apr 2026
Safety

Multi-Persona Thinking for Bias Mitigation in Large Language Models

DGX agent

arXiv:2601.15488v2 Announce Type: replace Abstract: Large Language Models (LLMs) exhibit social biases, which can lead to harmful stereotypes and unfair outcomes. We propose extbf{Multi-Persona Thinki

safetyarxiv-cs-cl
17 Apr 2026
Safety

Multi-User mmWave Beam and Rate Adaptation via Combinatorial Satisficing Bandits

DGX agent

arXiv:2604.14908v1 Announce Type: new Abstract: We study downlink beam and rate adaptation in a multi-user mmWave MISO system where multiple base stations (BSs), each using analog beamforming from fin

safetyarxiv-cs-lg
17 Apr 2026
Safety

NG-GS: NeRF-Guided 3D Gaussian Splatting Segmentation

DGX agent

arXiv:2604.14706v1 Announce Type: new Abstract: Recent advances in 3D Gaussian Splatting (3DGS) have enabled highly efficient and photorealistic novel view synthesis. However, segmenting objects accur

safetyarxiv-cs-cv
17 Apr 2026
Safety

NLP needs Diversity outside of 'Diversity'

DGX agent

arXiv:2604.14595v1 Announce Type: new Abstract: This position paper argues that recent progress with diversity in NLP is disproportionately concentrated on a small number of areas surrounding fairness

safetyarxiv-cs-cl
17 Apr 2026
Safety

ORBIT: On-policy Exploration-Exploitation for Controllable Multi-Budget Reasoning

DGX agent

arXiv:2601.08310v2 Announce Type: replace Abstract: Recent Large Reasoning Models (LRMs) achieve strong performance by leveraging long-form Chain-of-Thought (CoT) reasoning, but uniformly applying ove

safetyarxiv-cs-lg
17 Apr 2026
Safety

Practical estimation of the optimal classification error with soft labels and calibration

DGX agent

arXiv:2505.20761v3 Announce Type: replace Abstract: While the performance of machine learning systems has experienced significant improvement in recent years, relatively little attention has been paid

safetyarxiv-cs-lg
17 Apr 2026
Safety

Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation

DGX agent

arXiv:2603.13683v2 Announce Type: replace Abstract: Although debiased large language models (LLMs) excel at handling known or low-bias prompts, they often fail on unfamiliar and high-bias prompts. We

safetyarxiv-cs-cl
17 Apr 2026
Safety

PROXIMA: A Reliability Scoring Framework for Proxy Metrics in Online Controlled Experiments

DGX agent

arXiv:2604.14352v1 Announce Type: cross Abstract: Online A/B testing at scale relies on proxy metrics -- short-term, easily-measured signals used in place of slow-moving long-term outcomes. When the p

safetyarxiv-cs-lg
17 Apr 2026
Safety

Pushing the Boundaries of Multiple Choice Evaluation to One Hundred Options

DGX agent

arXiv:2604.14634v1 Announce Type: new Abstract: Multiple choice evaluation is widely used for benchmarking large language models, yet near ceiling accuracy in low option settings can be sustained by s

safetyarxiv-cs-cl
17 Apr 2026
Safety

QU-NLP at ArchEHR-QA 2026: Two-Stage QLoRA Fine-Tuning of Qwen3-4B for Patient-Oriented Clinical Question Answering and Evidence Sentence Alignment

DGX agent

arXiv:2604.14175v1 Announce Type: new Abstract: We present a unified system addressing both Subtask 3 (answer generation) and Subtask 4 (evidence sentence alignment) of the ArchEHR-QA Shared Task. For

safetyarxiv-cs-cl
17 Apr 2026
Safety

R3D: Revisiting 3D Policy Learning

DGX agent

arXiv:2604.15281v1 Announce Type: new Abstract: 3D policy learning promises superior generalization and cross-embodiment transfer, but progress has been hindered by training instabilities and severe o

safetyarxiv-cs-cv
17 Apr 2026
Safety

RaTA-Tool: Retrieval-based Tool Selection with Multimodal Large Language Models

DGX agent

arXiv:2604.14951v1 Announce Type: cross Abstract: Tool learning with foundation models aims to endow AI systems with the ability to invoke external resources -- such as APIs, computational utilities,

safetyarxiv-cs-cl
17 Apr 2026
Safety

Reinforcement Learning via Value Gradient Flow

DGX agent

arXiv:2604.14265v1 Announce Type: new Abstract: We study behavior-regularized reinforcement learning (RL), where regularization toward a reference distribution (the dataset in offline RL or the base m

safetyarxiv-cs-lg
17 Apr 2026
Safety

Reward-Aware Trajectory Shaping for Few-step Visual Generation

DGX agent

arXiv:2604.14910v1 Announce Type: new Abstract: Achieving high-fidelity generation in extremely few sampling steps has long been a central goal of generative modeling. Existing approaches largely rely

safetyarxiv-cs-cv
17 Apr 2026
Safety

RoSLAC: Robust Simultaneous Localization and Calibration of Multiple Magnetometers

DGX agent

arXiv:2604.14353v1 Announce Type: new Abstract: Localization of autonomous mobile robots (AMRs) in enclosed or semi-enclosed environments such as offices, hotels, hospitals, indoor parking facilities,

safetyarxiv-cs-ro
17 Apr 2026
← Previous
1…258259260261262…299
Next →