AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,351 results
22 Apr 2026

Diff-SBSR: Learning Multimodal Feature-Enhanced Diffusion Models for Zero-Shot Sketch-Based 3D Shape Retrieval

SafetyDGX agent

arXiv:2604.19135v1 Announce Type: new Abstract: This paper presents the first exploration of text-to-image diffusion models for zero-shot sketch-based 3D shape retrieval (ZS-SBSR). Existing sketch-bas

DINO Eats CLIP: Adapting Beyond Knowns for Open-set 3D Object Retrieval

SafetyDGX agent

arXiv:2604.19432v1 Announce Type: new Abstract: Vision foundation models have shown great promise for open-set 3D object retrieval (3DOR) through efficient adaptation to multi-view images. Leveraging

Discovering a Shared Logical Subspace: Steering LLM Logical Reasoning via Alignment of Natural-Language and Symbolic Views

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.19716v1 Announce Type: new Abstract: Large Language Models (LLMs) still struggle with multi-step logical reasoning. Existing approaches either purely refine the reasoning chain in natural l

Do Emotions Influence Moral Judgment in Large Language Models?

SafetyDGX agent

arXiv:2604.19125v1 Announce Type: new Abstract: Large language models have been extensively studied for emotion recognition and moral reasoning as distinct capabilities, yet the extent to which emotio

Dual Triangle Attention: Effective Bidirectional Attention Without Positional Embeddings

SafetyDGX agent

arXiv:2604.18603v1 Announce Type: cross Abstract: Bidirectional transformers are the foundation of many sequence modeling tasks across natural, biological, and chemical language domains, but they are

Efforts to revive chip manufacturing in Pennsylvania have been left in limbo by President Trump's sudden upending of US semiconductor policy over the past year (Michael Acton/Financial Times)

SafetyDGX agent

Michael Acton / Financial Times: Efforts to revive chip manufacturing in Pennsylvania have been left in limbo by President Trump's sudden upending of US semiconductor policy over the past year — High-

Evaluating LLM-Driven Summarisation of Parliamentary Debates with Computational Argumentation

SafetyDGX agent

arXiv:2604.19331v1 Announce Type: new Abstract: Understanding how policy is debated and justified in parliament is a fundamental aspect of the democratic process. However, the volume and complexity of

EVPO: Explained Variance Policy Optimization for Adaptive Critic Utilization in LLM Post-Training

SafetyDGX agent

arXiv:2604.19485v1 Announce Type: cross Abstract: Reinforcement learning (RL) for LLM post-training faces a fundamental design choice: whether to use a learned critic as a baseline for policy optimiza

ExpertGen: Scalable Sim-to-Real Expert Policy Learning from Imperfect Behavior Priors

SafetyDGX agent

arXiv:2603.15956v2 Announce Type: replace-cross Abstract: Learning generalizable and robust behavior cloning policies requires large volumes of high-quality robotics data. While human demonstrations (

Failure Modes in Multi-Hop QA: The Weakest Link Effect and the Recognition Bottleneck

SafetyDGX agent

arXiv:2601.12499v2 Announce Type: replace Abstract: Despite scaling to massive context windows, Large Language Models (LLMs) struggle with multi-hop reasoning due to inherent position bias, which caus

Fairness Audits of Institutional Risk Models in Deployed ML Pipelines

SafetyDGX agent

arXiv:2604.19468v1 Announce Type: cross Abstract: Fairness audits of institutional risk models are critical for understanding how deployed machine learning pipelines allocate resources. Drawing on mul

FairTree: Subgroup Fairness Auditing of Machine Learning Models with Bias-Variance Decomposition

SafetyDGX agent

arXiv:2604.19357v1 Announce Type: new Abstract: The evaluation of machine learning models typically relies mainly on performance metrics based on loss functions, which risk to overlook changes in perf

Fascinating how AI is getting better at diagrams like these (at least for ones that you could easily find on web search) but still making so…

SafetyDGX agent

Fascinating how AI is getting better at diagrams like these (at least for ones that you could easily find on web search) but still making some pretty wacky errors — like confusing where the rear brake

FASE : A Fairness-Aware Spatiotemporal Event Graph Framework for Predictive Policing

SafetyDGX agent

arXiv:2604.18644v1 Announce Type: cross Abstract: Predictive policing systems that allocate patrol resources based solely on predicted crime risk can unintentionally amplify racial disparities through

FASTER: Value-Guided Sampling for Fast RL

SafetyDGX agent

arXiv:2604.19730v1 Announce Type: cross Abstract: Some of the most performant reinforcement learning algorithms today can be prohibitively expensive as they use test-time scaling methods such as sampl

FB-NLL: A Feature-Based Approach to Tackle Noisy Labels in Personalized Federated Learning

SafetyDGX agent

arXiv:2604.19729v1 Announce Type: new Abstract: Personalized Federated Learning (PFL) aims to learn multiple task-specific models rather than a single global model across heterogeneous data distributi

Filing: Tron founder Justin Sun sues the Trump family's World Liberty Financial, alleging it unfairly locked up his WLFI holdings and threatened and defamed him (CoinDesk)

SafetyDGX agent

CoinDesk: Filing: Tron founder Justin Sun sues the Trump family's World Liberty Financial, alleging it unfairly locked up his WLFI holdings and threatened and defamed him — World Liberty unfairly froz

Fitted Q Evaluation Without Bellman Completeness via Stationary Weighting

SafetyDGX agent

arXiv:2512.23805v2 Announce Type: replace-cross Abstract: Fitted Q-evaluation (FQE) is a foundational method for off-policy evaluation in reinforcement learning, but existing theory typically relies o

Framelet-Based Blind Image Restoration with Minimax Concave Regularization

SafetyDGX agent

arXiv:2604.19314v1 Announce Type: new Abstract: Recovering corrupted images is one of the most challenging problems in image processing. Among various restoration tasks, blind image deblurring has bee

Gives new meaning to “Rear Brake Lever”!

SafetyDGX agent

This post likely references a humorous or unexpected use case involving a rear brake lever, possibly demonstrating an unintended design flaw, unconventional application, or double meaning related to b

God these people are annoying. Obnoxious comment and the guy can’t be bothered to notice the front tire that is labeled as a fork 🙄 Or to n…

SafetyDGX agent

God these people are annoying. Obnoxious comment and the guy can’t be bothered to notice the front tire that is labeled as a fork 🙄 Or to notice the front brake that’s lost its cable and is hovering u

GRAIL:Learning to Interact with Large Knowledge Graphs for Retrieval Augmented Reasoning

SafetyDGX agent

arXiv:2508.05498v2 Announce Type: replace Abstract: Large Language Models (LLMs) integrated with Retrieval-Augmented Generation (RAG) techniques have exhibited remarkable performance across a wide ran

Ground-Level Near Real-Time Modeling for PM2.5 Pollution Prediction

SafetyDGX agent

arXiv:2604.18973v1 Announce Type: cross Abstract: Air pollution is a worldwide public health threat that can cause or exacerbate many illnesses, including respiratory disease, cardiovascular disease,

Guiding Distribution Matching Distillation with Gradient-Based Reinforcement Learning

SafetyDGX agent

arXiv:2604.19009v1 Announce Type: cross Abstract: Diffusion distillation, exemplified by Distribution Matching Distillation (DMD), has shown great promise in few-step generation but often sacrifices q

How does the optimizer implicitly bias the model merging loss landscape?

SafetyDGX agent

arXiv:2510.04686v2 Announce Type: replace-cross Abstract: Model merging combines independent solutions with different capabilities into a single one while maintaining the same inference cost. Two popu

Investigating Counterfactual Unfairness in LLMs towards Identities through Humor

SafetyDGX agent

arXiv:2604.18729v1 Announce Type: new Abstract: Humor holds up a mirror to social perception: what we find funny often reflects who we are and how we judge others. When language models engage with hum

Knowledge-Guided Time-Varying Causal Inference for Arctic Sea Ice Dynamics

SafetyDGX agent

arXiv:2601.17647v2 Announce Type: replace-cross Abstract: Quantifying the causal relationship between sea ice thickness and sea surface height (SSH) is essential for understanding the mechanisms drivi

LASER: Learning Active Sensing for Continuum Field Reconstruction

SafetyDGX agent

arXiv:2604.19355v1 Announce Type: cross Abstract: High-fidelity measurements of continuum physical fields are essential for scientific discovery and engineering design but remain challenging under spa

Learning Hybrid-Control Policies for High-Precision In-Contact Manipulation Under Uncertainty

SafetyDGX agent

arXiv:2604.19677v1 Announce Type: cross Abstract: Reinforcement learning-based control policies have been frequently demonstrated to be more effective than analytical techniques for many manipulation

Learning to Credit the Right Steps: Objective-aware Process Optimization for Visual Generation

SafetyDGX agent

arXiv:2604.19234v1 Announce Type: new Abstract: Reinforcement learning, particularly Group Relative Policy Optimization (GRPO), has emerged as an effective framework for post-training visual generativ

Let me say this clearly: LLMs cannot feel emotions. Emotions are evolutionary mechanisms. They push us to avoid danger or approach what is b…

SafetyDGX agent

Let me say this clearly: LLMs cannot feel emotions. Emotions are evolutionary mechanisms. They push us to avoid danger or approach what is beneficial. We experience emotions because we are alive, and

🦤 LeWorldModel: Learning Physics from Pixels — Stable World Models with Just Two Losses World models: 1️⃣ DINO-WM: pretrained ViT encoder (…

SafetyDGX agent

🦤 LeWorldModel: Learning Physics from Pixels — Stable World Models with Just Two Losses World models: 1️⃣ DINO-WM: pretrained ViT encoder (from ImageNet) → features → predictor. But encoder is frozen,

LiteParse: our open-source, layout-aware PDF parser for AI agents. The secret? Grid projection. Instead of heavy ML layout models or flat te…

SafetyDGX agent

LiteParse: our open-source, layout-aware PDF parser for AI agents. The secret? Grid projection. Instead of heavy ML layout models or flat text extraction, it projects text onto a monospace grid so ali

Location Not Found: Exposing Implicit Local and Global Biases in Multilingual LLMs

SafetyDGX agent

arXiv:2604.19292v1 Announce Type: cross Abstract: Multilingual large language models (LLMs) have minimized the fluency gap between languages. This advancement, however, exposes models to the risk of b

LogosKG: Hardware-Optimized Scalable and Interpretable Knowledge Graph Retrieval

SafetyDGX agent

arXiv:2604.18913v1 Announce Type: new Abstract: Knowledge graphs (KGs) are increasingly integrated with large language models (LLMs) to provide structured, verifiable reasoning. A core operation in th

Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning

SafetyDGX agent

arXiv:2604.18978v1 Announce Type: cross Abstract: Scaling critic capacity is a promising direction for enhancing off-policy reinforcement learning (RL). However, larger critics are prone to overfittin

Lyapunov-Certified Direct Switching Theory for Q-Learning

SafetyDGX agent

arXiv:2604.19569v1 Announce Type: cross Abstract: Q-learning is one of the most fundamental algorithms in reinforcement learning. We analyze constant-stepsize Q-learning through a direct stochastic sw

M^{2}GRPO: Mamba-based Multi-Agent Group Relative Policy Optimization for Biomimetic Underwater Robots Pursuit

SafetyDGX agent

arXiv:2604.19404v1 Announce Type: cross Abstract: Traditional policy learning methods in cooperative pursuit face fundamental challenges in biomimetic underwater robots, where long-horizon decision ma

Machine individuality: Separating genuine idiosyncrasy from response bias in large language models

SafetyDGX agent

arXiv:2604.16755v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly integrated into daily life, in roles ranging from high-stakes decision support to companionship, un

MacroNav: Multi-Task Context Representation Learning Enables Efficient Navigation in Unknown Environments

SafetyDGX agent

arXiv:2511.04320v2 Announce Type: replace Abstract: Autonomous navigation in unknown environments requires multi-scale spatial understanding that captures geometric details, topological connectivity,

Mask World Model: Predicting What Matters for Robust Robot Policy Learning

SafetyDGX agent

arXiv:2604.19683v1 Announce Type: new Abstract: World models derived from large-scale video generative pre-training have emerged as a promising paradigm for generalist robot policy learning. However,

Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling

SafetyDGX agent

arXiv:2510.08145v2 Announce Type: replace Abstract: Large Language Models (LLMs) as automatic evaluators, commonly referred to as LLM-as-a-Judge, have also attracted growing attention. This approach p

Mitigating Long-Tail Bias via Prompt-Controlled Diffusion Augmentation

SafetyDGX agent

arXiv:2602.04749v2 Announce Type: replace Abstract: Long-tailed class imbalance remains a fundamental obstacle in semantic segmentation of high-resolution remote-sensing imagery, where dominant classe

Mixture of Predefined Experts: Maximizing Data Usage on Vertical Federated Learning

SafetyDGX agent

arXiv:2602.12708v2 Announce Type: replace Abstract: Vertical Federated Learning (VFL) has emerged as a critical paradigm for collaborative model training in privacy-sensitive domains such as finance a

MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation

SafetyDGX agent

arXiv:2604.19679v1 Announce Type: new Abstract: Recent advances in Diffusion Transformers (DiTs) have enabled high-quality joint audio-video generation, producing videos with synchronized audio within

MOSA: Motion-Guided Semantic Alignment for Dynamic Scene Graph Generation

SafetyDGX agent

arXiv:2604.19631v1 Announce Type: new Abstract: Dynamic Scene Graph Generation (DSGG) aims to structurally model objects and their dynamic interactions in video sequences for high-level semantic under

MRS: Multi-Resolution Skills for HRL Agents

SafetyDGX agent

arXiv:2505.21410v2 Announce Type: replace Abstract: Hierarchical reinforcement learning (HRL) decomposes the policy into a manager and a worker, enabling long-horizon planning but introducing a perfor

Multi-Gait Learning for Humanoid Robots Using Reinforcement Learning with Selective Adversarial Motion Prior

SafetyDGX agent

arXiv:2604.19102v1 Announce Type: cross Abstract: Learning diverse locomotion skills for humanoid robots in a unified reinforcement learning framework remains challenging due to the conflicting requir

Multi-modal Reasoning with LLMs for Visual Semantic Arithmetic

SafetyDGX agent

arXiv:2604.19567v1 Announce Type: new Abstract: Reinforcement learning (RL) as post-training is crucial for enhancing the reasoning ability of large language models (LLMs) in coding and math. However,

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

SafetyDGX agent

arXiv:2603.11665v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have been widely adopted as MLLM-as-a-Judges due to their strong alignment with human judgment across vario

Multiclass Local Calibration with the Jensen-Shannon Distance

SafetyDGX agent

arXiv:2510.26566v2 Announce Type: replace-cross Abstract: Developing trustworthy Machine Learning (ML) models requires their predicted probabilities to be well-calibrated, meaning they should reflect

On the Generalizability of Foundation Models for Crop Type Mapping

SafetyDGX agent

arXiv:2409.09451v5 Announce Type: replace Abstract: Foundation models pre-trained using self-supervised learning have shown powerful transfer learning capabilities on various downstream tasks, includi

Online Learning of Whittle Indices for Restless Bandits with Non-Stationary Transition Kernels

SafetyDGX agent

arXiv:2506.18186v3 Announce Type: replace Abstract: The restless multi-armed bandit (RMAB) framework is a popular approach to solving resource allocation problems in networked systems. In this paper,

Personalized Benchmarking: Evaluating LLMs by Individual Preferences

SafetyDGX agent

arXiv:2604.18943v1 Announce Type: new Abstract: With the rise in capabilities of large language models (LLMs) and their deployment in real-world tasks, evaluating LLM alignment with human preferences

Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications

SafetyDGX agent

arXiv:2411.06837v2 Announce Type: replace Abstract: The rapid rise of Large Language Models (LLMs) has created new disruptive possibilities for persuasive communication, enabling fully-automated, pers

Phase-Aware Policy Learning for Skateboard Riding of Quadruped Robots via Feature-wise Linear Modulation

SafetyDGX agent

arXiv:2602.09370v2 Announce Type: replace Abstract: Skateboards offer a compact and efficient means of transportation as a type of personal mobility device. However, controlling them with legged robot

Policy Gradient Primal-Dual Method for Safe Reinforcement Learning from Human Feedback

SafetyDGX agent

arXiv:2604.19024v1 Announce Type: new Abstract: Safe Reinforcement Learning from Human Feedback (Safe RLHF) has recently achieved empirical success in developing helpful and harmless large language mo

Prioritizing the Best: Incentivizing Reliable Multimodal Reasoning by Rewarding Beyond Answer Correctness

SafetyDGX agent

arXiv:2604.18892v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves multimodal reasoning by rewarding verifiable final answers. Yet answer-correct trajectori

Probing for Reading Times

SafetyDGX agent

arXiv:2604.18712v1 Announce Type: new Abstract: Probing has shown that language model representations encode rich linguistic information, but it remains unclear whether they also capture cognitive sig

Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference

SafetyDGX agent

arXiv:2604.19069v1 Announce Type: cross Abstract: Neural NLI models overfit dataset artifacts instead of truly reasoning. A hypothesis-only model gets 57.7% in SNLI, showing strong spurious correlatio

← Previous
1…200201202203204…240
Next →