AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
Safety

Machine individuality: Separating genuine idiosyncrasy from response bias in large language models

DGX agent

arXiv:2604.16755v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly integrated into daily life, in roles ranging from high-stakes decision support to companionship, un

safetyarxiv-cs-ai
22 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

MacroNav: Multi-Task Context Representation Learning Enables Efficient Navigation in Unknown Environments

DGX agent

arXiv:2511.04320v2 Announce Type: replace Abstract: Autonomous navigation in unknown environments requires multi-scale spatial understanding that captures geometric details, topological connectivity,

safetyarxiv-cs-ro
22 Apr 2026
Safety

Mask World Model: Predicting What Matters for Robust Robot Policy Learning

DGX agent

arXiv:2604.19683v1 Announce Type: new Abstract: World models derived from large-scale video generative pre-training have emerged as a promising paradigm for generalist robot policy learning. However,

safetyarxiv-cs-ro
22 Apr 2026
Safety

Mind2Drive: Predicting Driver Intentions from EEG in Real-world On-Road Driving

DGX agent

arXiv:2604.19368v1 Announce Type: new Abstract: Predicting driver intention from neurophysiological signals offers a promising pathway for enhancing proactive safety in advanced driver assistance syst

safetyarxiv-cs-cv
22 Apr 2026
Safety

Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling

DGX agent

arXiv:2510.08145v2 Announce Type: replace Abstract: Large Language Models (LLMs) as automatic evaluators, commonly referred to as LLM-as-a-Judge, have also attracted growing attention. This approach p

safetyarxiv-cs-cl
22 Apr 2026
Safety

Mitigating Long-Tail Bias via Prompt-Controlled Diffusion Augmentation

DGX agent

arXiv:2602.04749v2 Announce Type: replace Abstract: Long-tailed class imbalance remains a fundamental obstacle in semantic segmentation of high-resolution remote-sensing imagery, where dominant classe

safetyarxiv-cs-cv
22 Apr 2026
Safety

Mixture of Predefined Experts: Maximizing Data Usage on Vertical Federated Learning

DGX agent

arXiv:2602.12708v2 Announce Type: replace Abstract: Vertical Federated Learning (VFL) has emerged as a critical paradigm for collaborative model training in privacy-sensitive domains such as finance a

safetyarxiv-cs-lg
22 Apr 2026
Safety

MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation

DGX agent

arXiv:2604.19679v1 Announce Type: new Abstract: Recent advances in Diffusion Transformers (DiTs) have enabled high-quality joint audio-video generation, producing videos with synchronized audio within

safetyarxiv-cs-cv
22 Apr 2026
Safety

MOSA: Motion-Guided Semantic Alignment for Dynamic Scene Graph Generation

DGX agent

arXiv:2604.19631v1 Announce Type: new Abstract: Dynamic Scene Graph Generation (DSGG) aims to structurally model objects and their dynamic interactions in video sequences for high-level semantic under

safetyarxiv-cs-cv
22 Apr 2026
Safety

MRS: Multi-Resolution Skills for HRL Agents

DGX agent

arXiv:2505.21410v2 Announce Type: replace Abstract: Hierarchical reinforcement learning (HRL) decomposes the policy into a manager and a worker, enabling long-horizon planning but introducing a perfor

safetyarxiv-cs-ai
22 Apr 2026
Safety

Multi-Gait Learning for Humanoid Robots Using Reinforcement Learning with Selective Adversarial Motion Prior

DGX agent

arXiv:2604.19102v1 Announce Type: cross Abstract: Learning diverse locomotion skills for humanoid robots in a unified reinforcement learning framework remains challenging due to the conflicting requir

safetyarxiv-cs-ai
22 Apr 2026
Safety

Multi-modal Reasoning with LLMs for Visual Semantic Arithmetic

DGX agent

arXiv:2604.19567v1 Announce Type: new Abstract: Reinforcement learning (RL) as post-training is crucial for enhancing the reasoning ability of large language models (LLMs) in coding and math. However,

safetyarxiv-cs-ai
22 Apr 2026
Safety

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

DGX agent

arXiv:2603.11665v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have been widely adopted as MLLM-as-a-Judges due to their strong alignment with human judgment across vario

safetyarxiv-cs-cl
22 Apr 2026
Safety

Multiclass Local Calibration with the Jensen-Shannon Distance

DGX agent

arXiv:2510.26566v2 Announce Type: replace-cross Abstract: Developing trustworthy Machine Learning (ML) models requires their predicted probabilities to be well-calibrated, meaning they should reflect

safetyarxiv-cs-ai
22 Apr 2026
Safety

Multimodal embodiment-aware navigation transformer

DGX agent

arXiv:2604.19267v1 Announce Type: new Abstract: Goal-conditioned navigation models for ground robots trained using supervised learning show promising zero-shot transfer, but their collision-avoidance

safetyarxiv-cs-ro
22 Apr 2026
Safety

Neuromorphic Continual Learning for Sequential Deployment of Nuclear Plant Monitoring Systems

DGX agent

arXiv:2604.18611v1 Announce Type: cross Abstract: Anomaly detection in nuclear industrial control systems (ICS) requires continuous, energy-efficient monitoring across multiple subsystems that are oft

safetyarxiv-cs-ai
22 Apr 2026
Safety

On the Generalizability of Foundation Models for Crop Type Mapping

DGX agent

arXiv:2409.09451v5 Announce Type: replace Abstract: Foundation models pre-trained using self-supervised learning have shown powerful transfer learning capabilities on various downstream tasks, includi

safetyarxiv-cs-cv
22 Apr 2026
Safety

One Persona, Many Cues, Different Results: How Sociodemographic Cues Impact LLM Personalization

DGX agent

arXiv:2601.18572v2 Announce Type: replace Abstract: Personalization of LLMs by sociodemographic subgroup often improves user experience, but can also introduce or amplify biases and unfair outcomes ac

safetyarxiv-cs-cl
22 Apr 2026
Safety

Online Learning of Whittle Indices for Restless Bandits with Non-Stationary Transition Kernels

DGX agent

arXiv:2506.18186v3 Announce Type: replace Abstract: The restless multi-armed bandit (RMAB) framework is a popular approach to solving resource allocation problems in networked systems. In this paper,

safetyarxiv-cs-lg
22 Apr 2026
Safety

Our newsroom AI policy

DGX agent

Ars Technica's newsroom AI policy forbids unlabeled AI material in reported stories and requires human confirmation of every quotation's accuracy. The publication does not permit the publication of AI

safetyars-technica
22 Apr 2026
Safety

Personalized Benchmarking: Evaluating LLMs by Individual Preferences

DGX agent

arXiv:2604.18943v1 Announce Type: new Abstract: With the rise in capabilities of large language models (LLMs) and their deployment in real-world tasks, evaluating LLM alignment with human preferences

safetyarxiv-cs-ai
22 Apr 2026
Safety

Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications

DGX agent

arXiv:2411.06837v2 Announce Type: replace Abstract: The rapid rise of Large Language Models (LLMs) has created new disruptive possibilities for persuasive communication, enabling fully-automated, pers

safetyarxiv-cs-cl
22 Apr 2026
Safety

Phase-Aware Policy Learning for Skateboard Riding of Quadruped Robots via Feature-wise Linear Modulation

DGX agent

arXiv:2602.09370v2 Announce Type: replace Abstract: Skateboards offer a compact and efficient means of transportation as a type of personal mobility device. However, controlling them with legged robot

safetyarxiv-cs-ro
22 Apr 2026
Safety

Policy Gradient Primal-Dual Method for Safe Reinforcement Learning from Human Feedback

DGX agent

arXiv:2604.19024v1 Announce Type: new Abstract: Safe Reinforcement Learning from Human Feedback (Safe RLHF) has recently achieved empirical success in developing helpful and harmless large language mo

safetyarxiv-cs-lg
22 Apr 2026
Safety

Prioritizing the Best: Incentivizing Reliable Multimodal Reasoning by Rewarding Beyond Answer Correctness

DGX agent

arXiv:2604.18892v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves multimodal reasoning by rewarding verifiable final answers. Yet answer-correct trajectori

safetyarxiv-cs-cl
22 Apr 2026
Safety

Probing for Reading Times

DGX agent

arXiv:2604.18712v1 Announce Type: new Abstract: Probing has shown that language model representations encode rich linguistic information, but it remains unclear whether they also capture cognitive sig

safetyarxiv-cs-cl
22 Apr 2026
Safety

Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference

DGX agent

arXiv:2604.19069v1 Announce Type: cross Abstract: Neural NLI models overfit dataset artifacts instead of truly reasoning. A hypothesis-only model gets 57.7% in SNLI, showing strong spurious correlatio

safetyarxiv-cs-ai
22 Apr 2026
Safety

ProjLens: Unveiling the Role of Projectors in Multimodal Model Safety

DGX agent

arXiv:2604.19083v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success in cross-modal understanding and generation, yet their deployment is threate

safetyarxiv-cs-ai
22 Apr 2026
Safety

Proposing Topic Models and Evaluation Frameworks for Analyzing Associations with External Outcomes: An Application to Leadership Analysis Using Large-Scale Corporate Review Data

DGX agent

arXiv:2604.18919v1 Announce Type: new Abstract: Analyzing topics extracted from text data in relation to external outcomes is important across fields such as computational social science and organizat

safetyarxiv-cs-cl
22 Apr 2026
Safety

Q1 2026 Shareholder Update https://ir.tesla.com/#quarterly-disclosure We continued to make meaningful progress on the build out of the infra…

DGX agent

Q1 2026 Shareholder Update https://ir.tesla.com/#quarterly-disclosure We continued to make meaningful progress on the build out of the infrastructure & AI software that underpins our Robotaxi & future

safetyelon-musk--x
22 Apr 2026
Safety

QTMRL: An Agent for Quantitative Trading Decision-Making Based on Multi-Indicator Guided Reinforcement Learning

DGX agent

arXiv:2508.20467v2 Announce Type: replace-cross Abstract: In the highly volatile and uncertain global financial markets, traditional quantitative trading models relying on statistical modeling or empi

safetyarxiv-cs-lg
22 Apr 2026
Safety

Quantifying Data Similarity Using Cross Learning

DGX agent

arXiv:2510.10866v3 Announce Type: replace-cross Abstract: Measuring dataset similarity is fundamental in machine learning, particularly for transfer learning and domain adaptation. In the context of s

safetyarxiv-cs-lg
22 Apr 2026
Safety

Reasoning-Aware AIGC Detection via Alignment and Reinforcement

DGX agent

arXiv:2604.19172v1 Announce Type: new Abstract: The rapid advancement and widespread adoption of Large Language Models (LLMs) have elevated the need for reliable AI-generated content (AIGC) detection,

safetyarxiv-cs-ai
22 Apr 2026
Safety

Reasoning Structure Matters for Safety Alignment of Reasoning Models

DGX agent

arXiv:2604.18946v1 Announce Type: new Abstract: Large reasoning models (LRMs) achieve strong performance on complex reasoning tasks but often generate harmful responses to malicious user queries. This

safetyarxiv-cs-ai
22 Apr 2026
Safety

Regulating Artificial Intimacy: From Locks and Blocks to Relational Accountability

DGX agent

arXiv:2604.18893v1 Announce Type: cross Abstract: A series of high-profile tragedies involving companion chatbots has triggered an unusually rapid regulatory response. Several jurisdictions, including

safetyarxiv-cs-ai
22 Apr 2026
Safety

Reinforcement Learning Improves LLM Accuracy and Reasoning in Disease Classification from Radiology Reports

DGX agent

arXiv:2604.19060v1 Announce Type: new Abstract: Accurate disease classification from radiology reports is essential for many applications. While supervised fine-tuning (SFT) of lightweight LLMs improv

safetyarxiv-cs-ai
22 Apr 2026
Safety

RESFL: An Uncertainty-Aware Framework for Responsible Federated Learning by Balancing Privacy, Fairness and Utility

DGX agent

arXiv:2503.16251v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has gained prominence in machine learning applications across critical domains by enabling collaborative model trainin

safetyarxiv-cs-cv
22 Apr 2026
Safety

REVEAL: Multimodal Vision-Language Alignment of Retinal Morphometry and Clinical Risks for Incident AD and Dementia Prediction

DGX agent

arXiv:2604.18757v1 Announce Type: cross Abstract: The retina provides a unique, noninvasive window into Alzheimer's disease (AD) and dementia, capturing early structural changes through morphometric f

safetyarxiv-cs-ai
22 Apr 2026
Safety

Rivian and Volkswagen Group's joint venture put Devin to work on testing and ticket triage across a software platform that will power up to …

DGX agent

Rivian and Volkswagen Group's joint venture put Devin to work on testing and ticket triage across a software platform that will power up to 30M vehicles. Devin handles autonomous ticket triage in Slac

safetycognition-ai--x
22 Apr 2026
Safety

RL-ABC: Reinforcement Learning for Accelerator Beamline Control

DGX agent

arXiv:2604.19146v1 Announce Type: new Abstract: Particle accelerator beamline optimization is a high-dimensional control problem traditionally requiring significant expert intervention. We present RLA

safetyarxiv-cs-lg
22 Apr 2026
Safety

Safety-Critical Contextual Control via Online Riemannian Optimization with World Models

DGX agent

arXiv:2604.19639v1 Announce Type: cross Abstract: Modern world models are becoming too complex to admit explicit dynamical descriptions. We study safety-critical contextual control, where a Planner mu

safetyarxiv-cs-ai
22 Apr 2026
Safety

SEAT: Sparse Entity-Aware Tuning for Knowledge Adaptation while Preserving Epistemic Abstention

DGX agent

arXiv:2506.14387v3 Announce Type: replace Abstract: Adapting LLMs with new knowledge is increasingly important, but standard fine-tuning often erodes aligned epistemic abstention: the ability to ackno

safetyarxiv-cs-ai
22 Apr 2026
Safety

Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring

DGX agent

arXiv:2604.18835v1 Announce Type: cross Abstract: We propose a scalable, multifactorial experimental framework that systematically probes LLM sensitivity to subtle semantic changes in pairwise documen

safetyarxiv-cs-ai
22 Apr 2026
Safety

Sherpa.ai Privacy-Preserving Multi-Party Entity Alignment without Intersection Disclosure for Noisy Identifiers

DGX agent

arXiv:2604.19219v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training among multiple parties without centralizing raw data. There are two main paradigms in FL:

safetyarxiv-cs-ai
22 Apr 2026
Safety

Sources: Micron is pushing the US Congress to pass the 'MATCH Act', which would put new export restrictions on equipment its Chinese rivals use to make chips (Karen Freifeld/Reuters)

DGX agent

Karen Freifeld / Reuters: Sources: Micron is pushing the US Congress to pass the “MATCH Act”, which would put new export restrictions on equipment its Chinese rivals use to make chips — Micron Technol

safetytechmeme
22 Apr 2026
Safety

SpanVLA: Efficient Action Bridging and Learning from Negative-Recovery Samples for Vision-Language-Action Model

DGX agent

arXiv:2604.19710v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising autonomous driving paradigm for leveraging world knowledge and reasoning capabilities, especially

safetyarxiv-cs-cv
22 Apr 2026
Safety

Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented Generation

DGX agent

arXiv:2601.02993v4 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has become a key paradigm for reducing factual hallucinations in Large Language Models (LLMs), yet little is kn

safetyarxiv-cs-cl
22 Apr 2026
Safety

STAR-Teaming: A Strategy-Response Multiplex Network Approach to Automated LLM Red Teaming

DGX agent

arXiv:2604.18976v1 Announce Type: new Abstract: While Large Language Models (LLMs) are widely used, they remain susceptible to jailbreak prompts that can elicit harmful or inappropriate responses. Thi

safetyarxiv-cs-cl
22 Apr 2026
← Previous
1…232233234235236…265
Next →