AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

StructRL: Recovering Dynamic Programming Structure from Learning Dynamics in Distributional Reinforcement Learning

DGX agent

arXiv:2604.08620v1 Announce Type: cross Abstract: Reinforcement learning is typically treated as a uniform, data-driven optimization process, where updates are guided by rewards and temporal-differenc

safetyarxiv-cs-ai
13 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

SubQuad: Near-Quadratic-Free Structure Inference with Distribution-Balanced Objectives in Adaptive Receptor framework

DGX agent

arXiv:2602.17330v3 Announce Type: replace-cross Abstract: Comparative analysis of adaptive immune repertoires at population scale is hampered by two practical bottlenecks: the near-quadratic cost of p

safetyarxiv-cs-ai
13 Apr 2026
Safety

The causal relation between off-street parking and electric vehicle adoption in Scotland

DGX agent

arXiv:2604.09271v1 Announce Type: new Abstract: The transition to electric mobility hinges on maximising aggregate adoption while also facilitating equitable access. This study examines whether the 'c

safetyarxiv-cs-lg
13 Apr 2026
Safety

The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?

DGX agent

arXiv:2601.23045v2 Announce Type: replace Abstract: As AI becomes more capable, we entrust it with more general and consequential tasks. The risks from failure grow more severe with increasing task sc

safetyarxiv-cs-ai
13 Apr 2026
Safety

The Two-Stage Decision-Sampling Hypothesis: Understanding the Emergence of Self-Reflection in RL-Trained LLMs

DGX agent

arXiv:2601.01580v2 Announce Type: replace-cross Abstract: Self-reflection capabilities emerge in Large Language Models after RL post-training, with multi-turn RL achieving substantial gains over SFT c

safetyarxiv-cs-ai
13 Apr 2026
Safety

Think Less, Know More: State-Aware Reasoning Compression with Knowledge Guidance for Efficient Reasoning

DGX agent

arXiv:2604.09150v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance on complex tasks by leveraging long Chain-of-Thought (CoT), but often suffer from overthinking,

safetyarxiv-cs-cl
13 Apr 2026
Safety

Through Their Eyes: Fixation-aligned Tuning for Personalized User Emulation

DGX agent

arXiv:2604.09368v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as scalable user simulators for recommender system evaluation. Yet existing simulators per

safetyarxiv-cs-cv
13 Apr 2026
Safety

TME-PSR: Time-aware, Multi-interest, and Explanation Personalization for Sequential Recommendation

DGX agent

arXiv:2604.09439v1 Announce Type: cross Abstract: In this paper, we propose a sequential recommendation model that integrates Time-aware personalization, Multi-interest personalization, and Explanatio

safetyarxiv-cs-ai
13 Apr 2026
Safety

Tora3: Trajectory-Guided Audio-Video Generation with Physical Coherence

DGX agent

arXiv:2604.09057v1 Announce Type: new Abstract: Audio-video (AV) generation has recently made strong progress in perceptual quality and multimodal coherence, yet generating content with plausible moti

safetyarxiv-cs-cv
13 Apr 2026
Safety

Toward World Models for Epidemiology

DGX agent

arXiv:2604.09519v1 Announce Type: new Abstract: World models have emerged as a unifying paradigm for learning latent dynamics, simulating counterfactual futures, and supporting planning under uncertai

safetyarxiv-cs-lg
13 Apr 2026
Safety

Training event-based neural networks with exact gradients via Differentiable ODE Solving in JAX

DGX agent

arXiv:2603.08146v3 Announce Type: replace Abstract: Existing frameworks for gradient-based training of spiking neural networks face a trade-off: discrete-time methods using surrogate gradients support

safetyarxiv-cs-lg
13 Apr 2026
Safety

Traj2Action: A Co-Denoising Framework for Trajectory-Guided Human-to-Robot Skill Transfer

DGX agent

arXiv:2510.00491v3 Announce Type: replace-cross Abstract: Learning diverse manipulation skills for real-world robots is severely bottlenecked by the reliance on costly and hard-to-scale teleoperated d

safetyarxiv-cs-ai
13 Apr 2026
Safety

Truncated Rectified Flow Policy for Reinforcement Learning with One-Step Sampling

DGX agent

arXiv:2604.09159v1 Announce Type: new Abstract: Maximum entropy reinforcement learning (MaxEnt RL) has become a standard framework for sequential decision making, yet its standard Gaussian policy para

safetyarxiv-cs-lg
13 Apr 2026
Safety

Unbiased Rectification for Sequential Recommender Systems Under Fake Orders

DGX agent

arXiv:2604.08550v1 Announce Type: cross Abstract: Fake orders pose increasing threats to sequential recommender systems by misleading recommendation results through artificially manipulated interactio

safetyarxiv-cs-ai
13 Apr 2026
Safety

UniSemAlign: Text-Prototype Alignment with a Foundation Encoder for Semi-Supervised Histopathology Segmentation

DGX agent

arXiv:2604.09169v1 Announce Type: new Abstract: Semi-supervised semantic segmentation in computational pathology remains challenging due to scarce pixel-level annotations and unreliable pseudo-label s

safetyarxiv-cs-cv
13 Apr 2026
Safety

VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis

DGX agent

arXiv:2604.09330v1 Announce Type: cross Abstract: Recent advances in robot foundation models trained on large-scale human teleoperation data have enabled robots to perform increasingly complex real-wo

safetyarxiv-cs-cv
13 Apr 2026
Safety

VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning

DGX agent

arXiv:2604.09508v1 Announce Type: cross Abstract: Visual Retrieval-Augmented Generation (VRAG) empowers Vision-Language Models to retrieve and reason over visually rich documents. To tackle complex qu

safetyarxiv-cs-ai
13 Apr 2026
Safety

Visually-Guided Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2604.09349v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly advanced the reasoning ability of vision-language models (VLMs). However, the

safetyarxiv-cs-ai
13 Apr 2026
Safety

When & How to Write for Personalized Demand-aware Query Rewriting in Video Search

DGX agent

arXiv:2602.17667v2 Announce Type: replace-cross Abstract: In video search systems, user historical behaviors provide rich context for identifying search intent and resolving ambiguity. However, tradit

safetyarxiv-cs-cv
13 Apr 2026
Safety

Wireless Communication Enhanced Value Decomposition for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.08728v1 Announce Type: new Abstract: Cooperation in multi-agent reinforcement learning (MARL) benefits from inter-agent communication, yet most approaches assume idealized channels and exis

safetyarxiv-cs-lg
13 Apr 2026
Safety

You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector

DGX agent

arXiv:2603.15757v2 Announce Type: replace-cross Abstract: What happens when a pretrained generative robot policy is provided a constant initial noise as input, rather than repeatedly sampling it from

safetyarxiv-cs-ai
13 Apr 2026
Safety

A Clinical Point Cloud Paradigm for In-Hospital Mortality Prediction from Multi-Level Incomplete Multimodal EHRs

DGX agent

arXiv:2604.04614v2 Announce Type: replace-cross Abstract: Deep learning-based modeling of multimodal Electronic Health Records (EHRs) has become an important approach for clinical diagnosis and risk p

safetyarxiv-cs-ai
10 Apr 2026
Safety

A First Guess is Rarely the Final Answer: Learning to Search in the Travelling Salesperson Problem

DGX agent

arXiv:2604.06940v1 Announce Type: cross Abstract: Most neural solvers for the Traveling Salesperson Problem (TSP) are trained to output a single solution, even though practitioners rarely stop there:

safetyarxiv-cs-ai
10 Apr 2026
Safety

A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring

DGX agent

arXiv:2604.07395v1 Announce Type: cross Abstract: Robotic manipulation systems that follow language instructions often execute grasp primitives in a largely single-shot manner: a model proposes an act

safetyarxiv-cs-cv
10 Apr 2026
Safety

A systematic framework for generating novel experimental hypotheses from language models

DGX agent

arXiv:2408.05086v3 Announce Type: replace Abstract: Neural language models (LMs) have been shown to capture complex linguistic patterns, yet their utility in understanding human language and more broa

safetyarxiv-cs-cl
10 Apr 2026
Safety

Active Reward Machine Inference From Raw State Trajectories

DGX agent

arXiv:2604.07480v1 Announce Type: new Abstract: Reward machines are automaton-like structures that capture the memory required to accomplish a multi-stage task. When combined with reinforcement learni

safetyarxiv-cs-ro
10 Apr 2026
Safety

ActiveGlasses: Learning Manipulation with Active Vision from Ego-centric Human Demonstration

DGX agent

arXiv:2604.08534v1 Announce Type: new Abstract: Large-scale real-world robot data collection is a prerequisite for bringing robots into everyday deployment. However, existing pipelines often rely on s

safetyarxiv-cs-ro
10 Apr 2026
Safety

Adaptive Replay Buffer for Offline-to-Online Reinforcement Learning

DGX agent

arXiv:2512.10510v2 Announce Type: replace-cross Abstract: Offline-to-Online Reinforcement Learning (O2O RL) faces a critical dilemma in balancing the use of a fixed offline dataset with newly collecte

safetyarxiv-cs-ai
10 Apr 2026
Safety

AgentCity: Constitutional Governance for Autonomous Agent Economies via Separation of Power

DGX agent

arXiv:2604.07007v1 Announce Type: cross Abstract: Autonomous AI agents are beginning to operate across organizational boundaries on the open internet -- discovering, transacting with, and delegating t

safetyarxiv-cs-ai
10 Apr 2026
Safety

AI-Driven Research for Databases

DGX agent

arXiv:2604.06566v1 Announce Type: cross Abstract: As the complexity of modern workloads and hardware increasingly outpaces human research and engineering capacity, existing methods for database perfor

safetyarxiv-cs-ai
10 Apr 2026
Safety

Alternatives to the Laplacian for Scalable Spectral Clustering with Group Fairness Constraints

DGX agent

arXiv:2510.20220v3 Announce Type: replace Abstract: Recent research has focused on mitigating algorithmic bias in clustering by incorporating fairness constraints into algorithmic design. Notions such

safetyarxiv-cs-lg
10 Apr 2026
Safety

An Agentic Evaluation Architecture for Historical Bias Detection in Educational Textbooks

DGX agent

arXiv:2604.07883v1 Announce Type: cross Abstract: History textbooks often contain implicit biases, nationalist framing, and selective omissions that are difficult to audit at scale. We propose an agen

safetyarxiv-cs-cl
10 Apr 2026
Safety

Android Coach: Improve Online Agentic Training Efficiency with Single State Multiple Actions

DGX agent

arXiv:2604.07277v1 Announce Type: cross Abstract: Online reinforcement learning (RL) serves as an effective method for enhancing the capabilities of Android agents. However, guiding agents to learn th

safetyarxiv-cs-ai
10 Apr 2026
Safety

Are Face Embeddings Compatible Across Deep Neural Network Models?

DGX agent

arXiv:2604.07282v1 Announce Type: cross Abstract: Automated face recognition has made rapid strides over the past decade due to the unprecedented rise of deep neural network (DNN) models that can be t

safetyarxiv-cs-lg
10 Apr 2026
Safety

Attention Flows: Tracing LLM Conceptual Engagement via Story Summaries

DGX agent

arXiv:2604.06416v1 Announce Type: cross Abstract: Although LLM context lengths have grown, there is evidence that their ability to integrate information across long-form texts has not kept pace. We ev

safetyarxiv-cs-ai
10 Apr 2026
Safety

AudioRole: An Audio Dataset for Character Role-Playing in Large Language Models

DGX agent

arXiv:2509.23435v2 Announce Type: replace-cross Abstract: The creation of high-quality multimodal datasets remains fundamental for advancing role-playing capabilities in large language models (LLMs).

safetyarxiv-cs-ai
10 Apr 2026
Safety

Beyond Loss Values: Robust Dynamic Pruning via Loss Trajectory Alignment

DGX agent

arXiv:2604.07306v1 Announce Type: cross Abstract: Existing dynamic data pruning methods often fail under noisy-label settings, as they typically rely on per-sample loss as the ranking criterion. This

safetyarxiv-cs-lg
10 Apr 2026
Safety

Beyond Pessimism: Offline Learning in KL-regularized Games

DGX agent

arXiv:2604.06738v1 Announce Type: cross Abstract: We study offline learning in KL-regularized two-player zero-sum games, where policies are optimized under a KL constraint to a fixed reference policy.

safetyarxiv-cs-lg
10 Apr 2026
Safety

Beyond Surface Judgments: Human-Grounded Risk Evaluation of LLM-Generated Disinformation

DGX agent

arXiv:2604.06820v1 Announce Type: new Abstract: Large language models (LLMs) can generate persuasive narratives at scale, raising concerns about their potential use in disinformation campaigns. Assess

safetyarxiv-cs-ai
10 Apr 2026
Safety

Bias Redistribution in Visual Machine Unlearning: Does Forgetting One Group Harm Another?

DGX agent

arXiv:2604.08111v1 Announce Type: cross Abstract: Machine unlearning enables models to selectively forget training data, driven by privacy regulations such as GDPR and CCPA. However, its fairness impl

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

Blind Refusal: Language Models Refuse to Help Users Evade Unjust, Absurd, and Illegitimate Rules

DGX agent

arXiv:2604.06233v1 Announce Type: new Abstract: Safety-trained language models routinely refuse requests for help circumventing rules. But not all rules deserve compliance. When users ask for help eva

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Brain3D: EEG-to-3D Decoding of Visual Representations via Multimodal Reasoning

DGX agent

arXiv:2604.08068v1 Announce Type: new Abstract: Decoding visual information from electroencephalography (EEG) has recently achieved promising results, primarily focusing on reconstructing two-dimensio

safetyarxiv-cs-cv
10 Apr 2026
Safety

CAFP: A Post-Processing Framework for Group Fairness via Counterfactual Model Averaging

DGX agent

arXiv:2604.07009v1 Announce Type: new Abstract: Ensuring fairness in machine learning predictions is a critical challenge, especially when models are deployed in sensitive domains such as credit scori

safetyarxiv-cs-ai
10 Apr 2026
Safety

CNN-based Surface Temperature Forecasts with Ensemble Numerical Weather Prediction

DGX agent

arXiv:2507.18937v3 Announce Type: replace-cross Abstract: Due to limited computational resources, medium-range temperature forecasts typically rely on low-resolution numerical weather prediction (NWP)

safetyarxiv-cs-ai
10 Apr 2026
Safety

Contextualising (Im)plausible Events Triggers Figurative Language

DGX agent

arXiv:2604.07885v1 Announce Type: new Abstract: This work explores the connection between (non-)literalness and plausibility at the example of subject-verb-object events in English. We design a system

safetyarxiv-cs-cl
10 Apr 2026
Safety

Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs

DGX agent

arXiv:2604.07518v1 Announce Type: new Abstract: Vision-Language Models often struggle with complex visual reasoning due to the visual information loss in textual CoT. Existing methods either add the c

safetyarxiv-cs-cl
10 Apr 2026
Safety

Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models

DGX agent

arXiv:2604.08527v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models under their own induced distribution while leveraging supervision from stronger teachers. We identify

safetyarxiv-cs-cl
10 Apr 2026
Safety

Discrete Flow Matching Policy Optimization

DGX agent

arXiv:2604.06491v1 Announce Type: cross Abstract: We introduce Discrete flow Matching policy Optimization (DoMinO), a unified framework for Reinforcement Learning (RL) fine-tuning Discrete Flow Matchi

safetyarxiv-cs-ai
10 Apr 2026
← Previous
1…233234235236237…257
Next →