AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,351 results
22 Apr 2026

Proposing Topic Models and Evaluation Frameworks for Analyzing Associations with External Outcomes: An Application to Leadership Analysis Using Large-Scale Corporate Review Data

SafetyDGX agent

arXiv:2604.18919v1 Announce Type: new Abstract: Analyzing topics extracted from text data in relation to external outcomes is important across fields such as computational social science and organizat

QTMRL: An Agent for Quantitative Trading Decision-Making Based on Multi-Indicator Guided Reinforcement Learning

SafetyDGX agent

arXiv:2508.20467v2 Announce Type: replace-cross Abstract: In the highly volatile and uncertain global financial markets, traditional quantitative trading models relying on statistical modeling or empi

Quantifying Data Similarity Using Cross Learning

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2510.10866v3 Announce Type: replace-cross Abstract: Measuring dataset similarity is fundamental in machine learning, particularly for transfer learning and domain adaptation. In the context of s

Reasoning-Aware AIGC Detection via Alignment and Reinforcement

SafetyDGX agent

arXiv:2604.19172v1 Announce Type: new Abstract: The rapid advancement and widespread adoption of Large Language Models (LLMs) have elevated the need for reliable AI-generated content (AIGC) detection,

ReflectMT: Internalizing Reflection for Efficient and High-Quality Machine Translation

Model ReleasesDGX agent

arXiv:2604.19144v1 Announce Type: new Abstract: Recent years have witnessed growing interest in applying Large Reasoning Models (LRMs) to Machine Translation (MT). Existing approaches predominantly ad

Regulating Artificial Intimacy: From Locks and Blocks to Relational Accountability

SafetyDGX agent

arXiv:2604.18893v1 Announce Type: cross Abstract: A series of high-profile tragedies involving companion chatbots has triggered an unusually rapid regulatory response. Several jurisdictions, including

Reinforcement Learning Improves LLM Accuracy and Reasoning in Disease Classification from Radiology Reports

SafetyDGX agent

arXiv:2604.19060v1 Announce Type: new Abstract: Accurate disease classification from radiology reports is essential for many applications. While supervised fine-tuning (SFT) of lightweight LLMs improv

RepIt: Steering Language Models with Concept-Specific Refusal Vectors

Model ReleasesDGX agent

arXiv:2509.13281v5 Announce Type: replace Abstract: Current safety evaluations of language models rely on benchmark-based assessments that may miss localized vulnerabilities. We present RepIt, a simpl

RESFL: An Uncertainty-Aware Framework for Responsible Federated Learning by Balancing Privacy, Fairness and Utility

SafetyDGX agent

arXiv:2503.16251v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has gained prominence in machine learning applications across critical domains by enabling collaborative model trainin

RL-ABC: Reinforcement Learning for Accelerator Beamline Control

SafetyDGX agent

arXiv:2604.19146v1 Announce Type: new Abstract: Particle accelerator beamline optimization is a high-dimensional control problem traditionally requiring significant expert intervention. We present RLA

Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring

SafetyDGX agent

arXiv:2604.18835v1 Announce Type: cross Abstract: We propose a scalable, multifactorial experimental framework that systematically probes LLM sensitivity to subtle semantic changes in pairwise documen

Sherpa.ai Privacy-Preserving Multi-Party Entity Alignment without Intersection Disclosure for Noisy Identifiers

SafetyDGX agent

arXiv:2604.19219v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training among multiple parties without centralizing raw data. There are two main paradigms in FL:

Sources: Micron is pushing the US Congress to pass the 'MATCH Act', which would put new export restrictions on equipment its Chinese rivals use to make chips (Karen Freifeld/Reuters)

SafetyDGX agent

Karen Freifeld / Reuters: Sources: Micron is pushing the US Congress to pass the “MATCH Act”, which would put new export restrictions on equipment its Chinese rivals use to make chips — Micron Technol

SpanVLA: Efficient Action Bridging and Learning from Negative-Recovery Samples for Vision-Language-Action Model

SafetyDGX agent

arXiv:2604.19710v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising autonomous driving paradigm for leveraging world knowledge and reasoning capabilities, especially

Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2601.02993v4 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has become a key paradigm for reducing factual hallucinations in Large Language Models (LLMs), yet little is kn

STAR-Teaming: A Strategy-Response Multiplex Network Approach to Automated LLM Red Teaming

SafetyDGX agent

arXiv:2604.18976v1 Announce Type: new Abstract: While Large Language Models (LLMs) are widely used, they remain susceptible to jailbreak prompts that can elicit harmful or inappropriate responses. Thi

TEMPO: Scaling Test-time Training for Large Reasoning Models

SafetyDGX agent

arXiv:2604.19295v1 Announce Type: new Abstract: Test-time training (TTT) adapts model parameters on unlabeled test instances during inference time, which continuously extends capabilities beyond the r

Terrible

SafetyDGX agent

Terrible This is insane… The Virginia redistricting amendment on the ballot today is framed as a vote to 'restore fairness in the upcoming elections.' In reality, it turns a state that Kamala barely w

The Data-Driven Censored Newsvendor Problem

SafetyDGX agent

arXiv:2412.01763v3 Announce Type: replace-cross Abstract: We study a censored variant of the data-driven newsvendor problem, where the decision-maker must select an ordering quantity that minimizes ex

The PROPER Approach to Proactivity: Benchmarking and Advancing Knowledge Gap Navigation

SafetyDGX agent

arXiv:2601.09926v3 Announce Type: replace Abstract: Current approaches to proactive assistance move beyond the ask-and-respond paradigm by anticipating user needs. In practice, they either burden user

The signal is the ceiling: Measurement limits of LLM-predicted experience ratings from open-ended survey text

SafetyDGX agent

arXiv:2604.19645v1 Announce Type: new Abstract: An earlier paper (Hong, Potteiger, and Zapata 2026) established that an unoptimized GPT 4.1 prompt predicts fan-reported experience ratings within one p

The Triadic Loop: A Framework for Negotiating Alignment in AI Co-hosted Livestreaming

SafetyDGX agent

arXiv:2604.18850v1 Announce Type: cross Abstract: AI systems are increasingly embedded in multi-user social environments, yet most alignment frameworks conceptualize interaction as a dyadic relationsh

This actually happened. Smart second-graders know better. WATF. 🤯

SafetyDGX agent

Gary Marcus expresses skepticism or concern about an artificial intelligence claim or incident that he finds implausible, suggesting that even young children would recognize the flaw in the reasoning

Toward Clinically Acceptable Chest X-ray Report Generation: A Qualitative Retrospective Pilot Study of CXRMate-2

SafetyDGX agent

arXiv:2604.18967v1 Announce Type: new Abstract: Chest X-ray (CXR) radiology report generation (RRG) models have shown rapid progress, yet their clinical utility remains uncertain due to limited evalua

TRN-R1-Zero: Text-rich Network Reasoning via LLMs with Reinforcement Learning Only

SafetyDGX agent

arXiv:2604.19070v1 Announce Type: new Abstract: Zero-shot reasoning on text-rich networks (TRNs) remains a challenging frontier, as models must integrate textual semantics with relational structure wi

VimRAG: Navigating Massive Visual Context in Retrieval-Augmented Generation via Multimodal Memory Graph

SafetyDGX agent

arXiv:2602.12735v2 Announce Type: replace-cross Abstract: Effectively retrieving, reasoning, and understanding multimodal information remains a critical challenge for agentic systems. Traditional Retr

VoteGCL: Enhancing Graph-based Recommendations with Majority-Voting LLM-Rerank Augmentation

SafetyDGX agent

arXiv:2507.21563v4 Announce Type: replace-cross Abstract: Recommendation systems often suffer from data sparsity caused by limited user-item interactions, which degrade their performance and amplify p

21 Apr 2026

3 new ways Ads Advisor is making Google Ads safer and faster

SafetyDGX agent

Ads Advisor is an agentic conversational experience built with Gemini in Google Ads designed to help maximize performance based on business goals. The tool can help users get personalized answers, und

A High-Accuracy Optical Music Recognition Method Based on Bottleneck Residual Convolutions

SafetyDGX agent

arXiv:2604.16446v1 Announce Type: new Abstract: Optical Music Recognition (OMR) aims to convert printed or handwritten music score images into editable symbolic representations. This paper presents an

A Quasi-Experimental Developer Study of Security Training in LLM-Assisted Web Application Development

SafetyDGX agent

arXiv:2604.17763v1 Announce Type: cross Abstract: This paper presents a controlled quasi-experimental developer study examining whether a layer-based security training package is associated with impro

A Sensitivity Approach to Causal Inference Under Limited Overlap

SafetyDGX agent

arXiv:2511.22003v2 Announce Type: replace-cross Abstract: Limited overlap between treated and control groups is a key challenge in observational analysis. Standard approaches like trimming importance

A Text-To-Text Alignment Algorithm for Better Evaluation of Modern Speech Recognition Systems

SafetyDGX agent

arXiv:2509.24478v2 Announce Type: replace Abstract: Modern neural networks have greatly improved performance across speech recognition benchmarks. However, gains are often driven by frequent words wit

Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations

SafetyDGX agent

arXiv:2510.16458v2 Announce Type: replace Abstract: Natural Language Inference (NLI) datasets often exhibit human label variation. To better understand these variations, explanation-based approaches a

Align Documents to Questions: Question-Oriented Document Rewriting for Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2604.17325v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) enhances the factuality of Large Language Models (LLMs) by incorporating retrieved documents and/or generated conte

Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning

SafetyDGX agent

arXiv:2604.16622v1 Announce Type: new Abstract: Backchannels (e.g., `yeah', `mhm', and `right') are short, non-interruptive feedback signals whose lexical form and prosody jointly convey pragmatic mea

Aligning Language Models for Lyric-to-Melody Generation with Rule-Based Musical Constraints

SafetyDGX agent

arXiv:2604.18489v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise in lyric-to-melody generation, but models trained with Supervised Fine-Tuning (SFT) often produce musically

An `Inverse' Experimental Framework to Estimate Market Efficiency

SafetyDGX agent

arXiv:2604.18130v1 Announce Type: new Abstract: Digital marketplaces processing billions of dollars annually represent critical infrastructure in sociotechnical ecosystems, yet their performance optim

Annotation-Assisted Learning of Treatment Policies From Multimodal Electronic Health Records

SafetyDGX agent

arXiv:2507.20993v3 Announce Type: replace Abstract: We study how to learn treatment policies from multimodal electronic health records (EHRs) that consist of tabular data and clinical text. These poli

ARCS: Autoregressive Circuit Synthesis with Topology-Aware Graph Attention and Spec Conditioning

SafetyDGX agent

arXiv:2603.29068v3 Announce Type: replace Abstract: This paper presents ARCS (Autoregressive Circuit Synthesis), a system for amortized analog circuit generation. ARCS produces complete, SPICE-simulat

Attention-space Contrastive Guidance for Efficient Hallucination Mitigation in LVLMs

SafetyDGX agent

arXiv:2601.13707v2 Announce Type: replace Abstract: Hallucinations in large vision--language models (LVLMs) often arise when language priors dominate over visual evidence, leading to object misidentif

Audio-DeepThinker: Progressive Reasoning-Aware Reinforcement Learning for High-Quality Chain-of-Thought Emergence in Audio Language Models

SafetyDGX agent

arXiv:2604.18187v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have made significant progress in audio understanding, yet they primarily operate as perception-and-answer systems

AutoGraph-R1: End-to-End Reinforcement Learning for Knowledge Graph Construction

SafetyDGX agent

arXiv:2510.15339v3 Announce Type: replace Abstract: Building effective knowledge graphs (KGs) for Retrieval-Augmented Generation (RAG) is pivotal for advancing question answering (QA) systems. However

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale

SafetyDGX agent

arXiv:2604.18572v1 Announce Type: new Abstract: The Platonic Representation Hypothesis suggests that neural networks trained on different modalities (e.g., text and images) align and eventually conver

BASIS: Balanced Activation Sketching with Invariant Scalars for 'Ghost Backpropagation'

SafetyDGX agent

arXiv:2604.16324v1 Announce Type: new Abstract: The activation memory required for exact backpropagation scales linearly with network depth, context length, and feature dimensionality, forming an O(L

Batch-Adaptive Causal Annotations

SafetyDGX agent

arXiv:2502.10605v3 Announce Type: replace-cross Abstract: Estimating the causal effects of interventions is crucial to policy and decision-making, yet outcome data are often missing or subject to non-

Better with Less: Tackling Heterogeneous Multi-Modal Image Joint Pretraining via Conditioned and Degraded Masked Autoencoder

SafetyDGX agent

arXiv:2604.16952v1 Announce Type: new Abstract: Learning robust representations across extremely heterogeneous modalities remains a fundamental challenge in multi-modal vision. As a critical and profo

Beyond Overlap Metrics: Rewarding Reasoning and Preferences for Faithful Multi-Role Dialogue Summarization

SafetyDGX agent

arXiv:2604.17188v1 Announce Type: new Abstract: Multi-role dialogue summarization requires modeling complex interactions among multiple speakers while preserving role-specific information and factual

Bias-constrained multimodal intelligence for equitable and reliable clinical AI

SafetyDGX agent

arXiv:2604.16884v1 Announce Type: new Abstract: The integration of medical imaging and clinical text has enabled the emergence of generalist artificial intelligence (AI) systems for healthcare. Howeve

Boltzmann Machine Learning with a Parallel, Persistent Markov chain Monte Carlo method for Estimating Evolutionary Fields and Couplings from a Protein Multiple Sequence Alignment

SafetyDGX agent

arXiv:2604.18022v1 Announce Type: cross Abstract: The inverse Potts problem for estimating evolutionary single-site fields and pairwise couplings in homologous protein sequences from their single-site

Bounded Ratio Reinforcement Learning

SafetyDGX agent

arXiv:2604.18578v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) has become the predominant algorithm for on-policy reinforcement learning due to its scalability and empirical robust

BRIDGE the Gap: Mitigating Bias Amplification in Automated Scoring of English Language Learners via Inter-group Data Augmentation

SafetyDGX agent

arXiv:2602.23580v2 Announce Type: replace Abstract: In the field of educational assessment, automated scoring systems increasingly rely on deep learning and large language models (LLMs). However, thes

C-GenReg: Training-Free 3D Point Cloud Registration by Multi-View-Consistent Geometry-to-Image Generation with Probabilistic Modalities Fusion

SafetyDGX agent

arXiv:2604.16680v1 Announce Type: new Abstract: We introduce C-GenReg, a training-free framework for 3D point cloud registration that leverages the complementary strengths of world-scale generative pr

Can Explicit Physical Feasibility Benefit VLA Learning? An Empirical Study

SafetyDGX agent

arXiv:2604.17896v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models map multimodal inputs directly to robot actions and are typically trained through large-scale imitation learning. Wh

CAPO: Counterfactual Credit Assignment in Sequential Cooperative Teams

SafetyDGX agent

arXiv:2604.17693v1 Announce Type: new Abstract: In cooperative teams where agents act in a fixed order and share a single team reward, it is hard to know how much each agent contributed, and harder st

CCAR: Intrinsic Robustness as an Emergent Geometric Property

SafetyDGX agent

arXiv:2604.16861v1 Announce Type: cross Abstract: Standard supervised learning optimizes for predictive accuracy but remains agnostic to the internal geometry of learned features, often yielding repre

CGCMA: Conditionally-Gated Cross-Modal Attention for Event-Conditioned Asynchronous Fusion

SafetyDGX agent

arXiv:2604.16411v1 Announce Type: new Abstract: We study asynchronous alignment, a first-class multimodal learning setting in which a dense primary stream must be fused with sporadic external context

Chasing Ghosts: A Simulation-to-Real Olfactory Navigation Stack with Optional Vision Augmentation

SafetyDGX agent

arXiv:2602.19577v2 Announce Type: replace Abstract: Autonomous odor source localization remains a challenging problem for aerial robots due to turbulent airflow, sparse and delayed sensory signals, an

Closing the Modality Reasoning Gap for Speech Large Language Models

SafetyDGX agent

arXiv:2601.05543v2 Announce Type: replace Abstract: Although Speech Large Language Models have achieved notable progress, a substantial modality reasoning gap remains: their reasoning performance on s

COFFAIL: A Dataset of Successful and Anomalous Robot Skill Executions in the Context of Coffee Preparation

SafetyDGX agent

arXiv:2604.18236v1 Announce Type: new Abstract: In the context of robot learning for manipulation, curated datasets are an important resource for advancing the state of the art; however, available dat

Composed Vision-Language Retrieval for Skin Cancer Case Search via Joint Alignment of Global and Local Representations

SafetyDGX agent

arXiv:2603.09108v2 Announce Type: replace Abstract: Medical image retrieval aims to identify clinically relevant lesion cases to support diagnostic decision making, education, and quality control. In

← Previous
1…201202203204205…240
Next →