AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,600 results
17 Apr 2026

Wasserstein Formulation of Reinforcement Learning. An Optimal Transport Perspective on Policy Optimization

SafetyDGX agent

arXiv:2604.14765v1 Announce Type: new Abstract: We present a geometric framework for Reinforcement Learning (RL) that views policies as maps into the Wasserstein space of action probabilities. First,

What some leaders think of using universal income to mitigate AI-fueled layoffs: Musk calls it the 'best way', OpenAI's policy doc mentions a Public Wealth Fund (Siladitya Ray/Forbes)

SafetyDGX agent

Siladitya Ray / Forbes: What some leaders think of using universal income to mitigate AI-fueled layoffs: Musk calls it the “best way”, OpenAI's policy doc mentions a Public Wealth Fund — Topline — Elo

When Fairness Metrics Disagree: Evaluating the Reliability of Demographic Fairness Assessment in Machine Learning


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety
DGX agent

arXiv:2604.15038v1 Announce Type: cross Abstract: The evaluation of fairness in machine learning systems has become a central concern in high-stakes applications, including biometric recognition, heal

When Missing Becomes Structure: Intent-Preserving Policy Completion from Financial KOL Discourse

SafetyDGX agent

arXiv:2604.14333v1 Announce Type: new Abstract: Key Opinion Leader (KOL) discourse on social media is widely consumed as investment guidance, yet turning it into executable trading strategies without

Why Do Vision Language Models Struggle To Recognize Human Emotions?

SafetyDGX agent

arXiv:2604.15280v1 Announce Type: new Abstract: Understanding emotions is a fundamental ability for intelligent systems to be able to interact with humans. Vision-language models (VLMs) have made trem

Your LLM Agents are Temporally Blind: The Misalignment Between Tool Use Decisions and Human Time Perception

SafetyDGX agent

arXiv:2510.23853v3 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly used to interact with and execute tasks in dynamic environments. However, a critical yet overlook

16 Apr 2026

A Bayesian Framework for Uncertainty-Aware Explanations in Power Quality Disturbance Classification

SafetyDGX agent

arXiv:2604.13658v1 Announce Type: new Abstract: Advanced deep learning methods have shown remarkable success in power quality disturbance (PQD) classification. To enhance model transparency, explainab

A High-Resolution Landscape Dataset for Concept-Based XAI With Application to Species Distribution Models

SafetyDGX agent

arXiv:2604.13240v1 Announce Type: new Abstract: Mapping the spatial distribution of species is essential for conservation policy and invasive species management. Species distribution models (SDMs) are

A Mechanistic Analysis of Sim-and-Real Co-Training in Generative Robot Policies

SafetyDGX agent

arXiv:2604.13645v1 Announce Type: cross Abstract: Co-training, which combines limited in-domain real-world data with abundant surrogate data such as simulation or cross-embodiment robot data, is widel

A Multi-Model Approach to English-Bangla Sentiment Classification of Government Mobile Banking App Reviews

SafetyDGX agent

arXiv:2604.13057v1 Announce Type: new Abstract: For millions of users in developing economies who depend on mobile banking as their primary gateway to financial services, app quality directly shapes f

A 'race' implies that it is something that is worth 'winning', which is not the case for superintelligence. Whoever finished first is just t…

SafetyDGX agent

A 'race' implies that it is something that is worth 'winning', which is not the case for superintelligence. Whoever finished first is just the first to trigger irrecoverable catastrophe. 'There is no

Action Images: End-to-End Policy Learning via Multiview Video Generation

SafetyDGX agent

arXiv:2604.06168v2 Announce Type: replace Abstract: World action models (WAMs) have emerged as a promising direction for robot policy learning, as they can leverage powerful video backbones to model t

Activation-Guided Local Editing for Jailbreaking Attacks

SafetyDGX agent

arXiv:2508.00555v2 Announce Type: replace-cross Abstract: Jailbreaking is an essential adversarial technique for red-teaming these models to uncover and patch security flaws. However, existing jailbre

Ad firms settle with Trump FTC over claims they boycotted conservative media

SafetyDGX agent

Three major advertising companies—Dentsu, Publicis, and WPP—settled with the Federal Trade Commission over allegations they colluded on anti-misinformation policies that reduced ad revenue for conserv

ADP-DiT: Text-Guided Diffusion Transformer for Brain Image Generation in Alzheimer's Disease Progression

SafetyDGX agent

arXiv:2604.13495v1 Announce Type: new Abstract: Alzheimer's disease (AD) progresses heterogeneously across individuals, motivating subject-specific synthesis of follow-up magnetic resonance imaging (M

Alignment as Institutional Design: From Behavioral Correction to Transaction Structure in Intelligent Systems

SafetyDGX agent

arXiv:2604.13079v1 Announce Type: cross Abstract: Current AI alignment paradigms rely on behavioral correction: external supervisors (e.g., RLHF) observe outputs, judge against preferences, and adjust

Asymmetric-Loss-Guided Hybrid CNN-BiLSTM-Attention Model for Industrial RUL Prediction with Interpretable Failure Heatmaps

SafetyDGX agent

arXiv:2604.13459v1 Announce Type: new Abstract: Turbofan engine degradation under sustained operational stress necessitates robust prognostic systems capable of accurately estimating the Remaining Use

Beyond Arrow's Impossibility: Fairness as an Emergent Property of Multi-Agent Collaboration

SafetyDGX agent

arXiv:2604.13705v1 Announce Type: new Abstract: Fairness in language models is typically studied as a property of a single, centrally optimized model. As large language models become increasingly agen

Beyond Conservative Automated Driving in Multi-Agent Scenarios via Coupled Model Predictive Control and Deep Reinforcement Learning

SafetyDGX agent

arXiv:2604.13891v1 Announce Type: new Abstract: Automated driving at unsignalized intersections is challenging due to complex multi-vehicle interactions and the need to balance safety and efficiency.

Beyond State Consistency: Behavior Consistency in Text-Based World Models

SafetyDGX agent

arXiv:2604.13824v1 Announce Type: new Abstract: World models have been emerging as critical components for assessing the consequences of actions generated by interactive agents in online planning and

Bias at the End of the Score

SafetyDGX agent

arXiv:2604.13305v1 Announce Type: new Abstract: Reward models (RMs) are inherently non-neutral value functions designed and trained to encode specific objectives, such as human preferences or text-ima

Bias-Corrected Adaptive Conformal Inference for Multi-Horizon Time Series Forecasting

SafetyDGX agent

arXiv:2604.13253v1 Announce Type: new Abstract: Adaptive Conformal Inference (ACI) provides distribution-free prediction intervals with asymptotic coverage guarantees for time series under distributio

Biased Federated Learning under Wireless Heterogeneity

SafetyDGX agent

arXiv:2503.06078v2 Announce Type: replace Abstract: Federated learning (FL) has emerged as a promising framework for distributed learning, enabling collaborative model training without sharing private

Boundary Sampling to Learn Predictive Safety Filters via Pontryagin's Maximum Principle

SafetyDGX agent

arXiv:2604.13325v1 Announce Type: new Abstract: Safety filters provide a practical approach for enforcing safety constraints in autonomous systems. While learning-based tools scale to high-dimensional

C^2T: Captioning-Structure and LLM-Aligned Common-Sense Reward Learning for Traffic--Vehicle Coordination

SafetyDGX agent

arXiv:2604.13098v1 Announce Type: cross Abstract: State-of-the-art (SOTA) urban traffic control increasingly employs Multi-Agent Reinforcement Learning (MARL) to coordinate Traffic Light Controllers (

Can people please stop taking shots at @sama, and boycott his company instead? Violence is not the answer.

SafetyDGX agent

Can people please stop taking shots at @sama, and boycott his company instead? Violence is not the answer. This is really bad. The scary part in the US is that it doesn’t matter whether you are the CE

Capability-Aware Heterogeneous Control Barrier Functions for Decentralized Multi-Robot Safe Navigation

SafetyDGX agent

arXiv:2604.13245v1 Announce Type: new Abstract: Safe navigation for multi-robot systems requires enforcing safety without sacrificing task efficiency under decentralized decision-making. Existing dece

CausalDisenSeg: A Causality-Guided Disentanglement Framework with Counterfactual Reasoning for Robust Brain Tumor Segmentation Under Missing Modalities

SafetyDGX agent

arXiv:2604.13409v1 Announce Type: new Abstract: In clinical practice, the robustness of deep learning models for multimodal brain tumor segmentation is severely compromised by incomplete MRI data. Thi

Character Beyond Speech: Leveraging Role-Playing Evaluation in Audio Large Language Models via Reinforcement Learning

SafetyDGX agent

arXiv:2604.13804v1 Announce Type: new Abstract: The rapid evolution of multimodal large models has revolutionized the simulation of diverse characters in speech dialogue systems, enabling a novel inte

ChartNet: A Million-Scale, High-Quality Multimodal Dataset for Robust Chart Understanding

SafetyDGX agent

arXiv:2603.27064v2 Announce Type: replace-cross Abstract: Understanding charts requires models to jointly reason over geometric visual patterns, structured numerical data, and natural language -- a ca

CLIP Architecture for Abdominal CT Image-Text Alignment and Zero-Shot Learning: Investigating Batch Composition and Data Scaling

SafetyDGX agent

arXiv:2604.13561v1 Announce Type: new Abstract: Vision-language models trained with contrastive learning on paired medical images and reports show strong zero-shot diagnostic capabilities, yet the eff

Composite Silhouette: A Subsampling-based Aggregation Strategy

SafetyDGX agent

arXiv:2604.13816v1 Announce Type: new Abstract: Determining the number of clusters is a central challenge in unsupervised learning, where ground-truth labels are unavailable. The Silhouette coefficien

Context Sensitivity Improves Human-Machine Visual Alignment

SafetyDGX agent

arXiv:2604.13883v1 Announce Type: new Abstract: Modern machine learning models typically represent inputs as fixed points in a high-dimensional embedding space. While this approach has been proven pow

Data-Efficient RLVR via Off-Policy Influence Guidance

SafetyDGX agent

arXiv:2510.26491v2 Announce Type: replace Abstract: Data selection is a critical aspect of Reinforcement Learning with Verifiable Rewards (RLVR) for enhancing the reasoning capabilities of large langu

Debate to Align: Reliable Entity Alignment through Two-Stage Multi-Agent Debate

SafetyDGX agent

arXiv:2604.13551v1 Announce Type: new Abstract: Entity alignment (EA) aims to identify entities referring to the same real-world object across different knowledge graphs (KGs). Recent approaches based

Depth-Aware Image and Video Orientation Estimation

SafetyDGX agent

arXiv:2604.13995v1 Announce Type: new Abstract: This paper introduces a novel approach for image and video orientation estimation by leveraging depth distribution in natural images. The proposed metho

DiPO: Disentangled Perplexity Policy Optimization for Fine-grained Exploration-Exploitation Trade-Off

SafetyDGX agent

arXiv:2604.13902v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has catalyzed significant advances in the reasoning capabilities of Large Language Models (LLMs).

Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching

SafetyDGX agent

arXiv:2509.21912v2 Announce Type: replace Abstract: Guidance provides a simple and effective framework for posterior sampling by steering the generation process towards the desired distribution. When

Doc-V*:Coarse-to-Fine Interactive Visual Reasoning for Multi-Page Document VQA

SafetyDGX agent

arXiv:2604.13731v1 Announce Type: new Abstract: Multi-page Document Visual Question Answering requires reasoning over semantics, layouts, and visual elements in long, visually dense documents. Existin

Echoes Over Time: Unlocking Length Generalization in Video-to-Audio Generation Models

SafetyDGX agent

arXiv:2602.20981v3 Announce Type: replace Abstract: Scaling multimodal alignment between video and audio is challenging, particularly due to limited data and the mismatch between text descriptions and

Empirical Prediction of Pedestrian Comfort in Mobile Robot Pedestrian Encounters

SafetyDGX agent

arXiv:2604.13677v1 Announce Type: new Abstract: Mobile robots joining public spaces like sidewalks must care for pedestrian comfort. Many studies consider pedestrians' objective safety, for example, b

Enhanced Text-to-Image Generation by Fine-grained Multimodal Reasoning

SafetyDGX agent

arXiv:2604.13491v1 Announce Type: new Abstract: With the rapid progress of Multimodal Large Language Models (MLLMs), unified MLLMs that jointly perform image understanding and generation have advanced

Enhancing Reinforcement Learning for Radiology Report Generation with Evidence-aware Rewards and Self-correcting Preference Learning

SafetyDGX agent

arXiv:2604.13598v1 Announce Type: new Abstract: Recent reinforcement learning (RL) approaches have advanced radiology report generation (RRG), yet two core limitations persist: (1) report-level reward

Estimating Continuous Treatment Effects with Two-Stage Kernel Ridge Regression

SafetyDGX agent

arXiv:2604.13410v1 Announce Type: cross Abstract: We study the problem of estimating the effect function for a continuous treatment, which maps each treatment value to a population-averaged outcome. A

Evolvable Embodied Agent for Robotic Manipulation via Long Short-Term Reflection and Optimization

SafetyDGX agent

arXiv:2604.13533v1 Announce Type: cross Abstract: Achieving general-purpose robotics requires empowering robots to adapt and evolve based on their environment and feedback. Traditional methods face li

Failure Identification in Imitation Learning Via Statistical and Semantic Filtering

SafetyDGX agent

arXiv:2604.13788v1 Announce Type: cross Abstract: Imitation learning (IL) policies in robotics deliver strong performance in controlled settings but remain brittle in real-world deployments: rare even

FCBV-Net: Category-Level Robotic Garment Smoothing via Feature-Conditioned Bimanual Value Prediction

SafetyDGX agent

arXiv:2508.05153v2 Announce Type: replace Abstract: Category-level generalization for robotic garment manipulation, such as bimanual smoothing, remains a significant hurdle due to high dimensionality,

First-See-Then-Design: A Multi-Stakeholder View for Optimal Performance-Fairness Trade-Offs

SafetyDGX agent

arXiv:2604.14035v1 Announce Type: new Abstract: Fairness in algorithmic decision-making is often defined in the predictive space, where predictive performance - used as a proxy for decision-maker (DM)

Foresight Optimization for Strategic Reasoning in Large Language Models

SafetyDGX agent

arXiv:2604.13592v1 Announce Type: new Abstract: Reasoning capabilities in large language models (LLMs) have generally advanced significantly. However, it is still challenging for existing reasoning-ba

From Alignment to Prediction: A Study of Self-Supervised Learning and Predictive Representation Learning

SafetyDGX agent

arXiv:2604.13518v1 Announce Type: new Abstract: Self-supervised learning has emerged as a major technique for the task of learning from unlabeled data, where the current methods mostly revolve around

From Instruction to Event: Sound-Triggered Mobile Manipulation

SafetyDGX agent

arXiv:2601.21667v2 Announce Type: replace-cross Abstract: Current mobile manipulation research predominantly follows an instruction-driven paradigm, where agents rely on predefined textual commands to

From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space

SafetyDGX agent

arXiv:2604.14142v1 Announce Type: cross Abstract: While reinforcement learning with verifiable rewards (RLVR) significantly enhances LLM reasoning by optimizing the conditional distribution P(y|x), it

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction

SafetyDGX agent

arXiv:2604.13067v1 Announce Type: cross Abstract: SpeechLLMs process spoken language directly from audio, but accent and vocal identity cues can lead to biased behaviour. Current bias evaluations ofte

From Where Words Come: Efficient Regularization of Code Tokenizers Through Source Attribution

SafetyDGX agent

arXiv:2604.14053v1 Announce Type: new Abstract: Efficiency and safety of Large Language Models (LLMs), among other factors, rely on the quality of tokenization. A good tokenizer not only improves infe

@GaryMarcus Essentially your prediction that LLMs to scale wouldn’t solve key fundamental problems, reasoning, factual reliability, composit…

SafetyDGX agent

@GaryMarcus Essentially your prediction that LLMs to scale wouldn’t solve key fundamental problems, reasoning, factual reliability, compositionality, grounded understanding etc are now elephantine chi

@GaryMarcus @sama We need to rediscover the power of Satyagraha (principled, ardent, nonviolent resistence)

SafetyDGX agent

Gary Marcus advocates for applying Satyagraha—a principle of nonviolent resistance emphasizing moral conviction—as a response to contemporary challenges, likely in the context of AI development and go

Golden Handcuffs make safer AI agents

SafetyDGX agent

arXiv:2604.13609v1 Announce Type: new Abstract: Reinforcement learners can attain high reward through novel unintended strategies. We study a Bayesian mitigation for general environments: we expand th

GRITS: A Spillage-Aware Guided Diffusion Policy for Robot Food Scooping Tasks

SafetyDGX agent

arXiv:2510.00573v2 Announce Type: replace Abstract: Robotic food scooping is a critical manipulation skill for food preparation and service robots. However, existing robot learning algorithms, especia

HAMLET: Switch your Vision-Language-Action Model into a History-Aware Policy

SafetyDGX agent

arXiv:2510.00695v3 Announce Type: replace-cross Abstract: Inherently, robotic manipulation tasks are history-dependent: leveraging past context could be beneficial. However, most existing Vision-Langu

Hardware-Efficient Neuro-Symbolic Networks with the Exp-Minus-Log Operator

SafetyDGX agent

arXiv:2604.13871v1 Announce Type: new Abstract: Deep neural networks (DNNs) deliver state-of-the-art accuracy on regression and classification tasks, yet two structural deficits persistently obstruct

← Previous
1…193194195196197…210
Next →