AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Advancing Analytic Class-Incremental Learning through Vision-Language Calibration

DGX agent

arXiv:2602.13670v2 Announce Type: replace Abstract: Class-incremental learning (CIL) with pre-trained models (PTMs) faces a critical trade-off between efficient adaptation and long-term stability. Whi

safetyarxiv-cs-lg
7 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Agent-Based Modeling of Low-Emission Fertilizer Adoption for Dairy Farm Decarbonisation using Empirical Farm Data

DGX agent

arXiv:2605.03648v1 Announce Type: new Abstract: To understand complex system dynamics in dairy farming, it is essential to use modeling tools that capture farm heterogeneity, social interactions, and

safetyarxiv-cs-ai
7 May 2026
Safety

Anticipating Innovation Using Large Language Models

DGX agent

arXiv:2605.04875v1 Announce Type: new Abstract: Forecasting innovation, intended as the emergence of new technological combinations, is a fundamental challenge for science and policy. We show that for

safetyarxiv-cs-cl
7 May 2026
Safety

AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation

DGX agent

arXiv:2507.12768v2 Announce Type: replace Abstract: Learning generalizable manipulation policies hinges on data, yet robot manipulation data is scarce and often entangled with specific embodiments, ma

safetyarxiv-cs-cv
7 May 2026
Safety

Balanced Aggregation: Understanding and Fixing Aggregation Bias in GRPO

DGX agent

arXiv:2605.04077v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a central paradigm for improving reasoning and code generation in large language mode

safetyarxiv-cs-cl
7 May 2026
Safety

Beyond Public Access in LLM Pre-Training Data

DGX agent

arXiv:2505.00020v2 Announce Type: replace Abstract: Using a legally obtained dataset of 34 copyrighted O'Reilly Media books, we apply the DE-COP membership inference attack method to investigate wheth

safetyarxiv-cs-cl
7 May 2026
Safety

Causal discovery under mean independence and linearity

DGX agent

arXiv:2605.04381v1 Announce Type: cross Abstract: Causal discovery methods such as LiNGAM identify causal structure from observational data by assuming mutually independent disturbances. This assumpti

safetyarxiv-cs-lg
7 May 2026
Safety

CHE-TKG: Collaborative Historical Evidence and Evolutionary Dynamics Learning for Temporal Knowledge Graph Reasoning

DGX agent

arXiv:2605.04652v1 Announce Type: new Abstract: Temporal knowledge graph (TKG) reasoning aims to predict future events from historical facts. A key challenge lies in jointly capturing two sources of p

safetyarxiv-cs-cl
7 May 2026
Safety

Confronting Label Indeterminacy in Automated Bail Decisions

DGX agent

arXiv:2605.04073v1 Announce Type: new Abstract: Bail decisions present a fundamental challenge for data-driven decision support systems. When bail is denied, the counterfactual outcome of whether the

safetyarxiv-cs-lg
7 May 2026
Safety

Connecting online criminal behavior with machine learning: Using authorship attribution to analyze and link potential online traffickers

DGX agent

arXiv:2605.04080v1 Announce Type: new Abstract: This research investigated how online criminal activities can be better understood and connected using data-driven machine learning methods. Many illega

safetyarxiv-cs-cl
7 May 2026
Safety

Copula-Based Endogeneity Correction for Doubly Robust Estimation of Treatment Effect

DGX agent

arXiv:2605.03278v2 Announce Type: cross Abstract: Doubly Robust (DR) estimation of treatment effect relies on an untestable assumption that is the absence of unobserved confounding. This assumption is

safetyarxiv-cs-ai
7 May 2026
Safety

Counter-Dyna: Data-Efficient RL-Based HVAC Control using Counterfactual Building Models

DGX agent

arXiv:2605.04555v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) offers a promising approach for data-efficient energy management in buildings, combining the strengths of pred

safetyarxiv-cs-lg
7 May 2026
Safety

CRAFT: Counterfactual-to-Interactive Reinforcement Fine-Tuning for Driving Policies

DGX agent

arXiv:2605.04470v1 Announce Type: new Abstract: Open-loop imitation learning has advanced modern autonomous driving policy architectures, but closed-loop deployment remains vulnerable to policy-induce

safetyarxiv-cs-lg
7 May 2026
Safety

D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models

DGX agent

arXiv:2605.05204v1 Announce Type: new Abstract: The landscape of high-performance image generation models is currently shifting from the inefficient multi-step ones to the efficient few-step counterpa

safetyarxiv-cs-cv
7 May 2026
Safety

Data-dependent Exploration for Online Reinforcement Learning from Human Feedback

DGX agent

arXiv:2605.04477v1 Announce Type: new Abstract: Online reinforcement learning from human feedback (RLHF) has emerged as a promising paradigm for aligning large language models (LLMs) by continuously c

safetyarxiv-cs-lg
7 May 2026
Safety

Decompose to Understand, Fuse to Detect: Frequency-Decoupled Anomaly Detection for Encrypted Network Traffic

DGX agent

arXiv:2605.02970v1 Announce Type: cross Abstract: Network traffic anomaly detection represents a critical cybersecurity task, yet widespread encryption makes this task increasingly challenging. In res

safetyarxiv-cs-ai
7 May 2026
Safety

DFPO: Scaling Value Modeling via Distributional Flow towards Robust and Generalizable LLM Post-Training

DGX agent

arXiv:2602.05890v2 Announce Type: replace-cross Abstract: Training reinforcement learning (RL) systems in real-world environments remains challenging due to noisy supervision and poor out-of-domain (O

safetyarxiv-cs-cl
7 May 2026
Safety

Direct Product Flow Matching: Decoupling Radial and Angular Dynamics for Few-Shot Adaptation

DGX agent

arXiv:2605.05054v1 Announce Type: new Abstract: Recent flow matching (FM) methods improve the few-shot adaptation of vision-language models, by modeling cross-modal alignment as a continuous multi-ste

safetyarxiv-cs-cv
7 May 2026
Safety

Discovering Sparse Counterfactual Factors via Latent Adjustment for Survey-based Community Intervention

DGX agent

arXiv:2605.04460v1 Announce Type: new Abstract: Transportation surveys are widely used to understand travel preferences and adoption barriers, yet most survey-based analyses remain descriptive or pred

safetyarxiv-cs-lg
7 May 2026
Safety

Distilling Bayesian Belief States into Language Models for Auditable Negotiation

DGX agent

arXiv:2605.04507v1 Announce Type: new Abstract: Negotiation agents must infer what their counterpart values, update those beliefs over dialogue turns, and choose actions under uncertainty. End-to-end

safetyarxiv-cs-cl
7 May 2026
Safety

Dream-MPC: Gradient-Based Model Predictive Control with Latent Imagination

DGX agent

arXiv:2605.04568v1 Announce Type: new Abstract: State-of-the-art model-based Reinforcement Learning (RL) approaches either use gradient-free, population-based methods for planning, learned policy netw

safetyarxiv-cs-lg
7 May 2026
Safety

Dynamic Hyperparameter Importance for Efficient Multi-Objective Optimization

DGX agent

arXiv:2601.03166v2 Announce Type: replace Abstract: Choosing a suitable ML model is a complex task that can depend on several objectives, e.g., accuracy, fairness, or energy consumption. In practice,

safetyarxiv-cs-lg
7 May 2026
Safety

Efficiency of Parallel and Restart Exploration Strategies in Model Free Stochastic Simulations

DGX agent

arXiv:2503.03565v3 Announce Type: replace-cross Abstract: We analyze the efficiency of parallelization and restart mechanisms for stochastic simulations in model-free settings, where the underlying sy

safetyarxiv-cs-lg
7 May 2026
Safety

Efficient Geometry-Controlled High-Resolution Satellite Image Synthesis

DGX agent

arXiv:2605.04557v1 Announce Type: new Abstract: High-resolution satellite images are often scarce and costly, especially for remote areas or infrequent events. This shortage hampers the development an

safetyarxiv-cs-cv
7 May 2026
Safety

Efficient Model-Based Reinforcement Learning for Robot Control via Online Optimization

DGX agent

arXiv:2510.18518v2 Announce Type: replace Abstract: We present an online model-based reinforcement learning algorithm suitable for controlling complex robotic systems directly in the real world. Unlik

safetyarxiv-cs-ro
7 May 2026
Safety

Efficiently Aligning Language Models with Online Natural Language Feedback

DGX agent

arXiv:2605.04356v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has been used to elicit impressive performance from language models in many domains. But, broadly benefic

safetyarxiv-cs-lg
7 May 2026
Safety

Elicitation Matters: How Prompts and Query Protocols Shape LLM Surrogates under Sparse Observations

DGX agent

arXiv:2605.04764v1 Announce Type: new Abstract: Large language models are increasingly used as surrogate models for low-data optimization, but their optimizer-facing prediction and its uncertainty rem

safetyarxiv-cs-cl
7 May 2026
Safety

Enhancing the interpretability of spatially variable N2O model predictions with soft sensors during wastewater treatment

DGX agent

arXiv:2605.04082v1 Announce Type: new Abstract: Model-based solutions for nitrous oxide (N2O) emissions from wastewater treatment plants (WWTP) are informed by operational datasets designed to control

safetyarxiv-cs-lg
7 May 2026
Safety

EP-GRPO: Entropy-Progress Aligned Group Relative Policy Optimization with Implicit Process Guidance

DGX agent

arXiv:2605.04960v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR), particularly Group Relative Policy Optimization (GRPO), has advanced LLM reasoning. However, GRPO

safetyarxiv-cs-lg
7 May 2026
Safety

Evaluating Semantic Fragility in Text-to-Audio Generation Systems Under Controlled Prompt Perturbations

DGX agent

arXiv:2603.13824v2 Announce Type: replace-cross Abstract: Recent advances in text-to-audio generation enable models to translate natural-language descriptions into diverse musical output. However, the

safetyarxiv-cs-ai
7 May 2026
Safety

Every Step Counts: Step-Level Credit Assignment for Tool-Integrated Text-to-SQL

DGX agent

arXiv:2605.04719v1 Announce Type: new Abstract: Tool-integrated Text-to-SQL parsing has emerged as a promising paradigm, framing SQL generation as a sequential decision-making process interleaved with

safetyarxiv-cs-cl
7 May 2026
Safety

Extending Differential Temporal Difference Methods for Episodic Problems

DGX agent

arXiv:2605.04368v1 Announce Type: new Abstract: Differential temporal difference (TD) methods are value-based reinforcement learning algorithms that have been proposed for infinite-horizon problems. T

safetyarxiv-cs-lg
7 May 2026
Safety

FairEnc: A Fair Vision-Language Model with Fair Vision and Text Encoders for Glaucoma Detection

DGX agent

arXiv:2605.04882v1 Announce Type: new Abstract: Automated glaucoma detection is critical for preventing irreversible vision loss and reducing the burden on healthcare systems. However, ensuring fairne

safetyarxiv-cs-cv
7 May 2026
Safety

Fixed-Length Dense Fingerprint Representation with Alignment and Robust Enhancement

DGX agent

arXiv:2505.03597v2 Announce Type: replace Abstract: Fixed-length fingerprint representations, which map each fingerprint to a compact and fixed-size feature vector, are computationally efficient and w

safetyarxiv-cs-cv
7 May 2026
Safety

FLUID: Continuous-Time Hyperconnected Sparse Transformer for Sink-Free Learning

DGX agent

arXiv:2605.04421v1 Announce Type: new Abstract: Continuous-time (CT) Transformers improve irregular and long-range modeling over CT-RNNs by exploiting inputs or outputs embeddings with continuous dyna

safetyarxiv-cs-lg
7 May 2026
Safety

Globally Solving Unbalanced Optimal Transport and Density Control for Gaussian Distributions

DGX agent

arXiv:2605.04246v1 Announce Type: cross Abstract: In this article, we study unbalanced optimal transport (UOT) and establish a control-theoretic dynamical extension, which we call the unbalanced densi

safetyarxiv-cs-lg
7 May 2026
Safety

Graph-Augmented LLMs for Swiss MP Ideology Prediction

DGX agent

arXiv:2605.04643v1 Announce Type: new Abstract: Approximating the ideological position of Members of Parliament (MPs) is a fundamental task in political science, helping researchers understand legisla

safetyarxiv-cs-cl
7 May 2026
Safety

Graph-SND: Sparse Aggregation for Behavioral Diversity in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.05020v1 Announce Type: new Abstract: System Neural Diversity (SND) measures behavioral heterogeneity in multi-agent reinforcement learning by averaging pairwise distances over all inom{n}{2

safetyarxiv-cs-lg
7 May 2026
Safety

HeterSEED: Semantics-Structure Decoupling for Heterogeneous Graph Learning under Heterophily

DGX agent

arXiv:2605.04594v1 Announce Type: new Abstract: Many real-world heterogeneous graphs exhibit pronounced heterophily, where connected nodes often have dissimilar labels or play different semantic roles

safetyarxiv-cs-lg
7 May 2026
Safety

Hierarchical Support Vector State Partitioning for Distilling Black Box Reinforcement Learning Policies

DGX agent

arXiv:2605.04254v1 Announce Type: new Abstract: We introduce State Vector Space Partitioning (SVSP), a novel method to mimic a black box reinforcement learning policy using a set of human-interpretabl

safetyarxiv-cs-lg
7 May 2026
Safety

High-Fidelity Single-Image Head Modeling with Industry-Grade Topology

DGX agent

arXiv:2605.04524v1 Announce Type: new Abstract: We present a single-image head mesh reconstruction framework that addresses the longstanding challenge of simultaneously preserving facial identity and

safetyarxiv-cs-cv
7 May 2026
Safety

Hybrid Congestion Classification Framework Using Flow-Guided Attention and Empirical Mode Decomposition

DGX agent

arXiv:2605.04752v1 Announce Type: new Abstract: Accurate traffic congestion classification requires models that jointly capture roadway scene context and non-stationary traffic motion, yet most prior

safetyarxiv-cs-cv
7 May 2026
Safety

Improving Bias Correction Standards by Quantifying its Effects on Treatment Outcomes

DGX agent

arXiv:2407.14861v3 Announce Type: replace-cross Abstract: With the growing access to administrative health databases, retrospective studies have become crucial evidence for medical treatments. Yet, no

safetyarxiv-cs-lg
7 May 2026
Safety

Improving Medical VQA through Trajectory-Aware Process Supervision

DGX agent

arXiv:2605.04064v1 Announce Type: cross Abstract: Reasoning capabilities are crucial for reliable medical visual question answering (VQA); however, existing datasets rarely include reasoning explanati

safetyarxiv-cs-cv
7 May 2026
Safety

Investigating Trustworthiness of Nonparametric Deep Survival Models for Alzheimer's Disease Progression Analysis

DGX agent

arXiv:2605.04063v1 Announce Type: new Abstract: Alzheimer's Dementia (AD) is a progressive neurodegenerative disease marked by irreversible decline, making reliable modeling of its progression essenti

safetyarxiv-cs-lg
7 May 2026
Safety

Joint Semantic Token Selection and Prompt Optimization for Interpretable Prompt Learning

DGX agent

arXiv:2605.04425v1 Announce Type: new Abstract: Vision-language models such as CLIP achieve strong visual-textual alignment, but often suffer from overfitting and limited interpretability when adapted

safetyarxiv-cs-cv
7 May 2026
Safety

Lightweight Cross-Spectral Face Recognition via Contrastive Alignment and Distillation

DGX agent

arXiv:2605.04769v1 Announce Type: new Abstract: Heterogeneous Face Recognition (HFR) aims at matching face images captured across different sensing modalities, such as thermal-to-visible or near-infra

safetyarxiv-cs-cv
7 May 2026
Safety

LineRides: Line-Guided Reinforcement Learning for Bicycle Robot Stunts

DGX agent

arXiv:2605.05110v1 Announce Type: new Abstract: Designing reward functions for agile robotic maneuvers in reinforcement learning remains difficult, and demonstration-based approaches often require ref

safetyarxiv-cs-ro
7 May 2026
← Previous
1…196197198199200…257
Next →