AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,356 results
Safety

Location Not Found: Exposing Implicit Local and Global Biases in Multilingual LLMs

DGX agent

arXiv:2604.19292v1 Announce Type: cross Abstract: Multilingual large language models (LLMs) have minimized the fluency gap between languages. This advancement, however, exposes models to the risk of b

safetyarxiv-cs-ai
22 Apr 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

LogosKG: Hardware-Optimized Scalable and Interpretable Knowledge Graph Retrieval

DGX agent

arXiv:2604.18913v1 Announce Type: new Abstract: Knowledge graphs (KGs) are increasingly integrated with large language models (LLMs) to provide structured, verifiable reasoning. A core operation in th

safetyarxiv-cs-cl
22 Apr 2026
Safety

Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning

DGX agent

arXiv:2604.18978v1 Announce Type: cross Abstract: Scaling critic capacity is a promising direction for enhancing off-policy reinforcement learning (RL). However, larger critics are prone to overfittin

safetyarxiv-cs-ai
22 Apr 2026
Safety

Lyapunov-Certified Direct Switching Theory for Q-Learning

DGX agent

arXiv:2604.19569v1 Announce Type: cross Abstract: Q-learning is one of the most fundamental algorithms in reinforcement learning. We analyze constant-stepsize Q-learning through a direct stochastic sw

safetyarxiv-cs-ai
22 Apr 2026
Safety

M^{2}GRPO: Mamba-based Multi-Agent Group Relative Policy Optimization for Biomimetic Underwater Robots Pursuit

DGX agent

arXiv:2604.19404v1 Announce Type: cross Abstract: Traditional policy learning methods in cooperative pursuit face fundamental challenges in biomimetic underwater robots, where long-horizon decision ma

safetyarxiv-cs-ai
22 Apr 2026
Safety

Machine individuality: Separating genuine idiosyncrasy from response bias in large language models

DGX agent

arXiv:2604.16755v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly integrated into daily life, in roles ranging from high-stakes decision support to companionship, un

safetyarxiv-cs-ai
22 Apr 2026
Safety

MacroNav: Multi-Task Context Representation Learning Enables Efficient Navigation in Unknown Environments

DGX agent

arXiv:2511.04320v2 Announce Type: replace Abstract: Autonomous navigation in unknown environments requires multi-scale spatial understanding that captures geometric details, topological connectivity,

safetyarxiv-cs-ro
22 Apr 2026
Safety

Mask World Model: Predicting What Matters for Robust Robot Policy Learning

DGX agent

arXiv:2604.19683v1 Announce Type: new Abstract: World models derived from large-scale video generative pre-training have emerged as a promising paradigm for generalist robot policy learning. However,

safetyarxiv-cs-ro
22 Apr 2026
Safety

Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling

DGX agent

arXiv:2510.08145v2 Announce Type: replace Abstract: Large Language Models (LLMs) as automatic evaluators, commonly referred to as LLM-as-a-Judge, have also attracted growing attention. This approach p

safetyarxiv-cs-cl
22 Apr 2026
Safety

Mitigating Long-Tail Bias via Prompt-Controlled Diffusion Augmentation

DGX agent

arXiv:2602.04749v2 Announce Type: replace Abstract: Long-tailed class imbalance remains a fundamental obstacle in semantic segmentation of high-resolution remote-sensing imagery, where dominant classe

safetyarxiv-cs-cv
22 Apr 2026
Safety

Mixture of Predefined Experts: Maximizing Data Usage on Vertical Federated Learning

DGX agent

arXiv:2602.12708v2 Announce Type: replace Abstract: Vertical Federated Learning (VFL) has emerged as a critical paradigm for collaborative model training in privacy-sensitive domains such as finance a

safetyarxiv-cs-lg
22 Apr 2026
Safety

MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation

DGX agent

arXiv:2604.19679v1 Announce Type: new Abstract: Recent advances in Diffusion Transformers (DiTs) have enabled high-quality joint audio-video generation, producing videos with synchronized audio within

safetyarxiv-cs-cv
22 Apr 2026
Safety

MOSA: Motion-Guided Semantic Alignment for Dynamic Scene Graph Generation

DGX agent

arXiv:2604.19631v1 Announce Type: new Abstract: Dynamic Scene Graph Generation (DSGG) aims to structurally model objects and their dynamic interactions in video sequences for high-level semantic under

safetyarxiv-cs-cv
22 Apr 2026
Safety

MRS: Multi-Resolution Skills for HRL Agents

DGX agent

arXiv:2505.21410v2 Announce Type: replace Abstract: Hierarchical reinforcement learning (HRL) decomposes the policy into a manager and a worker, enabling long-horizon planning but introducing a perfor

safetyarxiv-cs-ai
22 Apr 2026
Safety

Multi-Gait Learning for Humanoid Robots Using Reinforcement Learning with Selective Adversarial Motion Prior

DGX agent

arXiv:2604.19102v1 Announce Type: cross Abstract: Learning diverse locomotion skills for humanoid robots in a unified reinforcement learning framework remains challenging due to the conflicting requir

safetyarxiv-cs-ai
22 Apr 2026
Safety

Multi-modal Reasoning with LLMs for Visual Semantic Arithmetic

DGX agent

arXiv:2604.19567v1 Announce Type: new Abstract: Reinforcement learning (RL) as post-training is crucial for enhancing the reasoning ability of large language models (LLMs) in coding and math. However,

safetyarxiv-cs-ai
22 Apr 2026
Safety

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

DGX agent

arXiv:2603.11665v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have been widely adopted as MLLM-as-a-Judges due to their strong alignment with human judgment across vario

safetyarxiv-cs-cl
22 Apr 2026
Safety

Multiclass Local Calibration with the Jensen-Shannon Distance

DGX agent

arXiv:2510.26566v2 Announce Type: replace-cross Abstract: Developing trustworthy Machine Learning (ML) models requires their predicted probabilities to be well-calibrated, meaning they should reflect

safetyarxiv-cs-ai
22 Apr 2026
Safety

On the Generalizability of Foundation Models for Crop Type Mapping

DGX agent

arXiv:2409.09451v5 Announce Type: replace Abstract: Foundation models pre-trained using self-supervised learning have shown powerful transfer learning capabilities on various downstream tasks, includi

safetyarxiv-cs-cv
22 Apr 2026
Safety

Online Learning of Whittle Indices for Restless Bandits with Non-Stationary Transition Kernels

DGX agent

arXiv:2506.18186v3 Announce Type: replace Abstract: The restless multi-armed bandit (RMAB) framework is a popular approach to solving resource allocation problems in networked systems. In this paper,

safetyarxiv-cs-lg
22 Apr 2026
Safety

Personalized Benchmarking: Evaluating LLMs by Individual Preferences

DGX agent

arXiv:2604.18943v1 Announce Type: new Abstract: With the rise in capabilities of large language models (LLMs) and their deployment in real-world tasks, evaluating LLM alignment with human preferences

safetyarxiv-cs-ai
22 Apr 2026
Safety

Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications

DGX agent

arXiv:2411.06837v2 Announce Type: replace Abstract: The rapid rise of Large Language Models (LLMs) has created new disruptive possibilities for persuasive communication, enabling fully-automated, pers

safetyarxiv-cs-cl
22 Apr 2026
Safety

Phase-Aware Policy Learning for Skateboard Riding of Quadruped Robots via Feature-wise Linear Modulation

DGX agent

arXiv:2602.09370v2 Announce Type: replace Abstract: Skateboards offer a compact and efficient means of transportation as a type of personal mobility device. However, controlling them with legged robot

safetyarxiv-cs-ro
22 Apr 2026
Safety

Policy Gradient Primal-Dual Method for Safe Reinforcement Learning from Human Feedback

DGX agent

arXiv:2604.19024v1 Announce Type: new Abstract: Safe Reinforcement Learning from Human Feedback (Safe RLHF) has recently achieved empirical success in developing helpful and harmless large language mo

safetyarxiv-cs-lg
22 Apr 2026
Safety

Prioritizing the Best: Incentivizing Reliable Multimodal Reasoning by Rewarding Beyond Answer Correctness

DGX agent

arXiv:2604.18892v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves multimodal reasoning by rewarding verifiable final answers. Yet answer-correct trajectori

safetyarxiv-cs-cl
22 Apr 2026
Safety

Probing for Reading Times

DGX agent

arXiv:2604.18712v1 Announce Type: new Abstract: Probing has shown that language model representations encode rich linguistic information, but it remains unclear whether they also capture cognitive sig

safetyarxiv-cs-cl
22 Apr 2026
Safety

Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference

DGX agent

arXiv:2604.19069v1 Announce Type: cross Abstract: Neural NLI models overfit dataset artifacts instead of truly reasoning. A hypothesis-only model gets 57.7% in SNLI, showing strong spurious correlatio

safetyarxiv-cs-ai
22 Apr 2026
Safety

Proposing Topic Models and Evaluation Frameworks for Analyzing Associations with External Outcomes: An Application to Leadership Analysis Using Large-Scale Corporate Review Data

DGX agent

arXiv:2604.18919v1 Announce Type: new Abstract: Analyzing topics extracted from text data in relation to external outcomes is important across fields such as computational social science and organizat

safetyarxiv-cs-cl
22 Apr 2026
Safety

QTMRL: An Agent for Quantitative Trading Decision-Making Based on Multi-Indicator Guided Reinforcement Learning

DGX agent

arXiv:2508.20467v2 Announce Type: replace-cross Abstract: In the highly volatile and uncertain global financial markets, traditional quantitative trading models relying on statistical modeling or empi

safetyarxiv-cs-lg
22 Apr 2026
Safety

Quantifying Data Similarity Using Cross Learning

DGX agent

arXiv:2510.10866v3 Announce Type: replace-cross Abstract: Measuring dataset similarity is fundamental in machine learning, particularly for transfer learning and domain adaptation. In the context of s

safetyarxiv-cs-lg
22 Apr 2026
Safety

Reasoning-Aware AIGC Detection via Alignment and Reinforcement

DGX agent

arXiv:2604.19172v1 Announce Type: new Abstract: The rapid advancement and widespread adoption of Large Language Models (LLMs) have elevated the need for reliable AI-generated content (AIGC) detection,

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

ReflectMT: Internalizing Reflection for Efficient and High-Quality Machine Translation

DGX agent

arXiv:2604.19144v1 Announce Type: new Abstract: Recent years have witnessed growing interest in applying Large Reasoning Models (LRMs) to Machine Translation (MT). Existing approaches predominantly ad

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

Regulating Artificial Intimacy: From Locks and Blocks to Relational Accountability

DGX agent

arXiv:2604.18893v1 Announce Type: cross Abstract: A series of high-profile tragedies involving companion chatbots has triggered an unusually rapid regulatory response. Several jurisdictions, including

safetyarxiv-cs-ai
22 Apr 2026
Safety

Reinforcement Learning Improves LLM Accuracy and Reasoning in Disease Classification from Radiology Reports

DGX agent

arXiv:2604.19060v1 Announce Type: new Abstract: Accurate disease classification from radiology reports is essential for many applications. While supervised fine-tuning (SFT) of lightweight LLMs improv

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

RepIt: Steering Language Models with Concept-Specific Refusal Vectors

DGX agent

arXiv:2509.13281v5 Announce Type: replace Abstract: Current safety evaluations of language models rely on benchmark-based assessments that may miss localized vulnerabilities. We present RepIt, a simpl

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

RESFL: An Uncertainty-Aware Framework for Responsible Federated Learning by Balancing Privacy, Fairness and Utility

DGX agent

arXiv:2503.16251v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has gained prominence in machine learning applications across critical domains by enabling collaborative model trainin

safetyarxiv-cs-cv
22 Apr 2026
Safety

RL-ABC: Reinforcement Learning for Accelerator Beamline Control

DGX agent

arXiv:2604.19146v1 Announce Type: new Abstract: Particle accelerator beamline optimization is a high-dimensional control problem traditionally requiring significant expert intervention. We present RLA

safetyarxiv-cs-lg
22 Apr 2026
Safety

Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring

DGX agent

arXiv:2604.18835v1 Announce Type: cross Abstract: We propose a scalable, multifactorial experimental framework that systematically probes LLM sensitivity to subtle semantic changes in pairwise documen

safetyarxiv-cs-ai
22 Apr 2026
Safety

Sherpa.ai Privacy-Preserving Multi-Party Entity Alignment without Intersection Disclosure for Noisy Identifiers

DGX agent

arXiv:2604.19219v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training among multiple parties without centralizing raw data. There are two main paradigms in FL:

safetyarxiv-cs-ai
22 Apr 2026
Safety

Sources: Micron is pushing the US Congress to pass the 'MATCH Act', which would put new export restrictions on equipment its Chinese rivals use to make chips (Karen Freifeld/Reuters)

DGX agent

Karen Freifeld / Reuters: Sources: Micron is pushing the US Congress to pass the “MATCH Act”, which would put new export restrictions on equipment its Chinese rivals use to make chips — Micron Technol

safetytechmeme
22 Apr 2026
Safety

SpanVLA: Efficient Action Bridging and Learning from Negative-Recovery Samples for Vision-Language-Action Model

DGX agent

arXiv:2604.19710v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising autonomous driving paradigm for leveraging world knowledge and reasoning capabilities, especially

safetyarxiv-cs-cv
22 Apr 2026
Safety

Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented Generation

DGX agent

arXiv:2601.02993v4 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has become a key paradigm for reducing factual hallucinations in Large Language Models (LLMs), yet little is kn

safetyarxiv-cs-cl
22 Apr 2026
Safety

STAR-Teaming: A Strategy-Response Multiplex Network Approach to Automated LLM Red Teaming

DGX agent

arXiv:2604.18976v1 Announce Type: new Abstract: While Large Language Models (LLMs) are widely used, they remain susceptible to jailbreak prompts that can elicit harmful or inappropriate responses. Thi

safetyarxiv-cs-cl
22 Apr 2026
Safety

TEMPO: Scaling Test-time Training for Large Reasoning Models

DGX agent

arXiv:2604.19295v1 Announce Type: new Abstract: Test-time training (TTT) adapts model parameters on unlabeled test instances during inference time, which continuously extends capabilities beyond the r

safetyarxiv-cs-lg
22 Apr 2026
Safety

Terrible

DGX agent

Terrible This is insane… The Virginia redistricting amendment on the ballot today is framed as a vote to 'restore fairness in the upcoming elections.' In reality, it turns a state that Kamala barely w

safetyelon-musk--x
22 Apr 2026
Safety

The Data-Driven Censored Newsvendor Problem

DGX agent

arXiv:2412.01763v3 Announce Type: replace-cross Abstract: We study a censored variant of the data-driven newsvendor problem, where the decision-maker must select an ordering quantity that minimizes ex

safetyarxiv-cs-lg
22 Apr 2026
Safety

The PROPER Approach to Proactivity: Benchmarking and Advancing Knowledge Gap Navigation

DGX agent

arXiv:2601.09926v3 Announce Type: replace Abstract: Current approaches to proactive assistance move beyond the ask-and-respond paradigm by anticipating user needs. In practice, they either burden user

safetyarxiv-cs-lg
22 Apr 2026
Safety

The signal is the ceiling: Measurement limits of LLM-predicted experience ratings from open-ended survey text

DGX agent

arXiv:2604.19645v1 Announce Type: new Abstract: An earlier paper (Hong, Potteiger, and Zapata 2026) established that an unoptimized GPT 4.1 prompt predicts fan-reported experience ratings within one p

safetyarxiv-cs-cl
22 Apr 2026
← Previous
1…251252253254255…300
Next →