AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,598 results
14 Apr 2026

MimicLM: Zero-Shot Voice Imitation through Autoregressive Modeling of Pseudo-Parallel Speech Corpora

SafetyDGX agent

arXiv:2604.11552v1 Announce Type: cross Abstract: Voice imitation aims to transform source speech to match a reference speaker's timbre and speaking style while preserving linguistic content. A straig

Minimal Embodiment Enables Efficient Learning of Number Concepts in Robot

SafetyDGX agent

arXiv:2604.11373v1 Announce Type: cross Abstract: Robots are increasingly entering human-interactive scenarios that require understanding of quantity. How intelligent systems acquire abstract numerica

Mixture of Cognitive Reasoners: Modular Reasoning with Brain-Like Specialization

SafetyDGX agent

arXiv:2506.13331v3 Announce Type: replace Abstract: Human cognitive behavior arises from the interaction of specialized brain networks dedicated to distinct functions, such as language, logic, and soc


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MM-LIMA: Less Is More for Alignment in Multi-Modal Datasets

SafetyDGX agent

arXiv:2308.12067v3 Announce Type: replace-cross Abstract: Multimodal large language models are typically trained in two stages: first pre-training on image-text pairs, and then fine-tuning using super

MobiFlow: Real-World Mobile Agent Benchmarking through Trajectory Fusion

SafetyDGX agent

arXiv:2604.09587v1 Announce Type: new Abstract: Mobile agents can autonomously complete user-assigned tasks through GUI interactions. However, existing mainstream evaluation benchmarks, such as Androi

MonoEM-GS: Monocular Expectation-Maximization Gaussian Splatting SLAM

SafetyDGX agent

arXiv:2604.10593v1 Announce Type: new Abstract: Feed-forward geometric foundation models can infer dense point clouds and camera motion directly from RGB streams, providing priors for monocular SLAM.

MoRI: Mixture of RL and IL Experts for Long-Horizon Manipulation Tasks

SafetyDGX agent

arXiv:2604.10165v1 Announce Type: new Abstract: Reinforcement Learning (RL) and Imitation Learning (IL) are the standard frameworks for policy acquisition in manipulation. While IL offers efficient po

MOSAIC: Multi-Domain Orthogonal Session Adaptive Intent Capture for Prescient Recommendations

SafetyDGX agent

arXiv:2604.10147v1 Announce Type: cross Abstract: Capturing user intent across heterogeneous behavioral domains stands as a fundamental challenge in session-based recommender systems. Yet, existing mu

Multi-Frequency Local Plasticity for Visual Representation Learning

SafetyDGX agent

arXiv:2604.09734v1 Announce Type: cross Abstract: We study how far structured architectural bias can compensate for the absence of end-to-end gradient-based representation learning in visual recogniti

Multi-Granularity Reasoning for Image Quality Assessment via Attribute-Aware Reinforcement Learning to Rank

SafetyDGX agent

arXiv:2604.09704v1 Announce Type: new Abstract: Recent advances in reasoning-induced image quality assessment (IQA) have demonstrated the power of reinforcement learning to rank (RL2R) for training vi

Multi-Modal Learning meets Genetic Programming: Analyzing Alignment in Latent Space Optimization

SafetyDGX agent

arXiv:2604.08324v2 Announce Type: replace-cross Abstract: Symbolic regression (SR) aims to discover mathematical expressions from data, a task traditionally tackled using Genetic Programming (GP) thro

Multi-Model Synthetic Training for Mission-Critical Small Language Models

SafetyDGX agent

arXiv:2509.13047v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across many domains, yet their application to specialized fields remain

Multimodal Dataset Normalization and Perceptual Validation for Music-Taste Correspondences

SafetyDGX agent

arXiv:2604.10632v1 Announce Type: cross Abstract: Collecting large, aligned cross-modal datasets for music-flavor research is difficult because perceptual experiments are costly and small by design. W

Naka-GS: A Bionics-inspired Dual-Branch Naka Correction and Progressive Point Pruning for Low-Light 3DGS

SafetyDGX agent

arXiv:2604.11142v1 Announce Type: new Abstract: Low-light conditions severely hinder 3D restoration and reconstruction by degrading image visibility, introducing color distortions, and contaminating g

NameBERT: Scaling Name-Based Nationality Classification with LLM-Augmented Open Academic Data

SafetyDGX agent

arXiv:2604.10401v1 Announce Type: new Abstract: Inferring nationality from personal names is a critical capability for equity and bias monitoring, personalization, and a valuable tool in biomedical an

Normative Common Ground Replication (NormCoRe): Replication-by-Translation for Studying Norms in Multi-Agent AI

SafetyDGX agent

arXiv:2603.11974v2 Announce Type: replace Abstract: In the late 2010s, the fashion trend NormCore framed sameness as a signal of belonging, illustrating how norms emerge through collective coordinatio

NOSE: Neural Olfactory-Semantic Embedding with Tri-Modal Orthogonal Contrastive Learning

SafetyDGX agent

arXiv:2604.10452v1 Announce Type: new Abstract: Olfaction lies at the intersection of chemical structure, neural encoding, and linguistic perception, yet existing representation methods fail to fully

Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning

SafetyDGX agent

arXiv:2504.13818v4 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as the leading approach for enhancing reasoning capabilities in large langua

OmniUMI: Towards Physically Grounded Robot Learning via Human-Aligned Multimodal Interaction

SafetyDGX agent

arXiv:2604.10647v1 Announce Type: new Abstract: UMI-style interfaces enable scalable robot learning, but existing systems remain largely visuomotor, relying primarily on RGB observations and trajector

On the Effectiveness of Textual Prompting with Lightweight Fine-Tuning for SAM3 Remote Sensing Segmentation

SafetyDGX agent

arXiv:2512.15564v2 Announce Type: replace Abstract: Remote sensing (RS) image segmentation is constrained by the limited availability of annotated data and a gap between overhead imagery and natural i

Online Learning-Enhanced High Order Adaptive Safety Control

SafetyDGX agent

arXiv:2511.19651v2 Announce Type: replace Abstract: Control barrier functions (CBFs) are an effective model-based tool to formally certify the safety of a system. With the growing complexity of modern

OOM-RL: Out-of-Money Reinforcement Learning Market-Driven Alignment for LLM-Based Multi-Agent Systems

SafetyDGX agent

arXiv:2604.11477v1 Announce Type: new Abstract: The alignment of Multi-Agent Systems (MAS) for autonomous software engineering is constrained by evaluator epistemic uncertainty. Current paradigms, suc

🦔OpenAI is backing an Illinois state bill that would shield AI labs from liability in cases where their models cause mass casualties or lar…

SafetyDGX agent

🦔OpenAI is backing an Illinois state bill that would shield AI labs from liability in cases where their models cause mass casualties or large-scale financial disasters, defined as death or serious inj

Optimization-Guided Diffusion for Interactive Scene Generation

SafetyDGX agent

arXiv:2512.07661v3 Announce Type: replace Abstract: Realistic and diverse multi-agent driving scenes are crucial for evaluating autonomous vehicles, but safety-critical events which are essential for

Orthogonal machine learning for conditional odds and risk ratios

SafetyDGX agent

arXiv:2604.10412v1 Announce Type: cross Abstract: Conditional effects are commonly used measures for understanding how treatment effects vary across different groups, and are often used to target trea

PACO: Proxy-Task Alignment and Online Calibration for On-the-Fly Category Discovery

SafetyDGX agent

arXiv:2604.11484v1 Announce Type: new Abstract: On-the-Fly Category Discovery (OCD) requires a model, trained on an offline support set, to recognize known classes while discovering new ones from an o

Particle Diffusion Matching: Random Walk Correspondence Search for the Alignment of Standard and Ultra-Widefield Fundus Images

SafetyDGX agent

arXiv:2604.10085v1 Announce Type: new Abstract: We propose a robust alignment technique for Standard Fundus Images (SFIs) and Ultra-Widefield Fundus Images (UWFIs), which are challenging to align due

PAT: Privacy-Preserving Adversarial Transfer for Accurate, Robust and Privacy-Preserving EEG Decoding

SafetyDGX agent

arXiv:2412.11390v3 Announce Type: replace-cross Abstract: An electroencephalogram (EEG)-based brain-computer interface (BCI) enables direct communication between the brain and external devices. Howeve

PEMANT: Persona-Enriched Multi-Agent Negotiation for Travel

SafetyDGX agent

arXiv:2604.10475v1 Announce Type: new Abstract: Modeling household-level trip generation is fundamental to accurate demand forecasting, traffic flow estimation, and urban system planning. Existing stu

Perceptual Inductive Bias Is What You Need Before Contrastive Learning

SafetyDGX agent

arXiv:2506.01201v2 Announce Type: replace Abstract: David Marr's seminal theory of human perception stipulates that visual processing is a multi-stage process, prioritizing the derivation of boundary

PICon: A Multi-Turn Interrogation Framework for Evaluating Persona Agent Consistency

SafetyDGX agent

arXiv:2603.25620v2 Announce Type: replace Abstract: Large language model (LLM)-based persona agents are rapidly being adopted as scalable proxies for human participants across diverse domains. Yet the

Policy Split: Incentivizing Dual-Mode Exploration in LLM Reinforcement with Dual-Mode Entropy Regularization

SafetyDGX agent

arXiv:2604.11510v1 Announce Type: cross Abstract: To encourage diverse exploration in reinforcement learning (RL) for large language models (LLMs) without compromising accuracy, we propose Policy Spli

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2509.21882v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a practical, scalable way to improve large language models on math, code, and other s

Predictions can be weapons of power. They only work if we believe them. -@carissaveliz @TEDTalks 2026

SafetyDGX agent

Carissa Véliz delivered a TED Talk in 2026 arguing that predictions function as instruments of power, shaping behavior and outcomes through the act of belief itself. Her thesis suggests that predictiv

Preference-Agile Multi-Objective Optimization for Real-time Vehicle Dispatching

SafetyDGX agent

arXiv:2604.10664v1 Announce Type: new Abstract: Multi-objective optimization (MOO) has been widely studied in literature because of its versatility in human-centered decision making in real-life appli

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation

SafetyDGX agent

arXiv:2603.20725v2 Announce Type: replace Abstract: Text-to-image generation has advanced rapidly, yet it still struggles to capture the nuanced user preferences. Existing approaches typically rely on

Principles Do Not Apply Themselves: A Hermeneutic Perspective on AI Alignment

SafetyDGX agent

arXiv:2604.10673v1 Announce Type: new Abstract: AI alignment is often framed as the task of ensuring that an AI system follows a set of stated principles or human preferences, but general principles r

PRISM Risk Signal Framework: Hierarchy-Based Red Lines for AI Behavioral Risk

SafetyDGX agent

arXiv:2604.11070v1 Announce Type: new Abstract: Current approaches to AI safety define red lines at the case level: specific prompts, specific outputs, specific harms. This paper argues that red lines

Prompt Injection as Role Confusion

SafetyDGX agent

arXiv:2603.12277v3 Announce Type: replace-cross Abstract: Language models remain vulnerable to prompt injection attacks despite extensive safety training. We trace this failure to role confusion: mode

Prompt Relay: Inference-Time Temporal Control for Multi-Event Video Generation

SafetyDGX agent

arXiv:2604.10030v1 Announce Type: new Abstract: Video diffusion models have achieved remarkable progress in generating high-quality videos. However, these models struggle to represent the temporal suc

ProUIE: A Macro-to-Micro Progressive Learning Method for LLM-based Universal Information Extraction

SafetyDGX agent

arXiv:2604.10633v1 Announce Type: new Abstract: LLM-based universal information extraction (UIE) methods often rely on additional information beyond the original training data, which increases trainin

Proximal Supervised Fine-Tuning

SafetyDGX agent

arXiv:2508.17784v2 Announce Type: replace-cross Abstract: Supervised fine-tuning (SFT) of foundation models often leads to poor generalization, where prior capabilities deteriorate after tuning on new

QFS-Composer: Query-focused summarization pipeline for less resourced languages

SafetyDGX agent

arXiv:2604.10687v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong performance in text summarization, yet their effectiveness drops significantly across languages with res

QShield: Securing Neural Networks Against Adversarial Attacks using Quantum Circuits

SafetyDGX agent

arXiv:2604.10933v1 Announce Type: cross Abstract: Deep neural networks remain highly vulnerable to adversarial perturbations, limiting their reliability in security- and safety-critical applications.

Quantifying the Climate Risk of Generative AI: Region-Aware Carbon Accounting with G-TRACE and the AI Sustainability Pyramid

SafetyDGX agent

arXiv:2511.04776v2 Announce Type: replace-cross Abstract: Generative Artificial Intelligence (GenAI) represents a rapidly expanding digital infrastructure whose energy demand and associated CO2 emissi

RAG-KT: Cross-platform Explainable Knowledge Tracing with Multi-view Fusion Retrieval Generation

SafetyDGX agent

arXiv:2604.10960v1 Announce Type: new Abstract: Knowledge Tracing (KT) infers a student's knowledge state from past interactions to predict future performance. Conventional Deep Learning (DL)-based KT

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought

SafetyDGX agent

arXiv:2506.16796v4 Announce Type: replace Abstract: Real-World Image Super-Resolution is one of the most challenging task in image restoration. However, existing methods struggle with an accurate unde

Reasoning Resides in Layers: Restoring Temporal Reasoning in Video-Language Models with Layer-Selective Merging

SafetyDGX agent

arXiv:2604.11399v1 Announce Type: cross Abstract: Multimodal adaptation equips large language models (LLMs) with perceptual capabilities, but often weakens the reasoning ability inherited from languag

Rebooting Microreboot: Architectural Support for Safe, Parallel Recovery in Microservice Systems

SafetyDGX agent

arXiv:2604.09963v1 Announce Type: cross Abstract: Microreboot enables fast recovery by restarting only the failing component, but in modern microservices naive restarts are unsafe: dense dependencies

Regularized Entropy Information Adaptation with Temporal-Awareness Networks for Simultaneous Speech Translation

SafetyDGX agent

arXiv:2604.09916v1 Announce Type: new Abstract: Simultaneous Speech Translation (SimulST) requires balancing high translation quality with low latency. Recent work introduced REINA, a method that trai

Reinforcement Learning for Intensity Control: An Application to Choice-Based Network Revenue Management

SafetyDGX agent

arXiv:2406.05358v3 Announce Type: replace Abstract: Intensity control is a class of continuous-time dynamic optimization problems with many important applications in Operations Research including queu

Relative Entropy Pathwise Policy Optimization

SafetyDGX agent

arXiv:2507.11019v4 Announce Type: replace Abstract: Score-function based methods for policy learning, such as REINFORCE and PPO, have delivered strong results in game-playing and robotics, yet their h

Reliable and Real-Time Highway Trajectory Planning via Hybrid Learning-Optimization Frameworks

SafetyDGX agent

arXiv:2508.04436v2 Announce Type: replace Abstract: Autonomous highway driving involves high-speed safety risks due to limited reaction time, where rare but dangerous events may lead to severe consequ

Reliable Evaluation Protocol for Low-Precision Retrieval

SafetyDGX agent

arXiv:2508.03306v4 Announce Type: replace-cross Abstract: Lowering the numerical precision of model parameters and computations is widely adopted to improve the efficiency of retrieval systems. Howeve

Resilient Write: A Six-Layer Durable Write Surface for LLM Coding Agents

SafetyDGX agent

arXiv:2604.10842v1 Announce Type: cross Abstract: LLM-powered coding agents increasingly rely on tool-use protocols such as the Model Context Protocol~(MCP) to read and write files on a developer's wo

Rethinking LLM Watermark Detection in Black-Box Settings: A Non-Intrusive Third-Party Framework

SafetyDGX agent

arXiv:2603.14968v2 Announce Type: replace-cross Abstract: While watermarking serves as a critical mechanism for LLM provenance, existing secret-key schemes tightly couple detection with injection, req

Rethinking Token-Level Credit Assignment in RLVR: A Polarity-Entropy Analysis

SafetyDGX agent

arXiv:2604.11056v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has substantially improved the reasoning ability of Large Language Models (LLMs). However, its s

Revisiting Compositionality in Dual-Encoder Vision-Language Models: The Role of Inference

SafetyDGX agent

arXiv:2604.11496v1 Announce Type: cross Abstract: Dual-encoder Vision-Language Models (VLMs) such as CLIP are often characterized as bag-of-words systems due to their poor performance on compositional

Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?

SafetyDGX agent

arXiv:2505.24778v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly used in high-stakes domains, accurately assessing their confidence is crucial. Humans typically exp

Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility

SafetyDGX agent

arXiv:2602.03402v3 Announce Type: replace Abstract: Vision language models (VLMs) extend the reasoning capabilities of large language models (LLMs) to cross-modal settings, yet remain highly vulnerabl

← Previous
1…200201202203204…210
Next →