AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

A Survey of Multimodal Mathematical Reasoning: From Perception, Alignment to Reasoning

DGX agent

arXiv:2603.08291v3 Announce Type: replace Abstract: Multimodal Mathematical Reasoning (MMR) has recently attracted increasing attention for its capability to solve mathematical problems involving both

safetyarxiv-cs-ai
15 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

A Two-Timescale Primal-Dual Framework for Reinforcement Learning via Online Dual Variable Guidance

DGX agent

arXiv:2505.04494v3 Announce Type: replace-cross Abstract: We study reinforcement learning by combining recent advances in regularized linear programming formulations with the classical theory of stoch

safetyarxiv-cs-lg
15 Apr 2026
Safety

AAPO: Enhancing the Reasoning Capabilities of LLMs with Advantage Margin

DGX agent

arXiv:2505.14264v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has emerged as an effective approach for enhancing the reasoning capabilities of large language models (LLMs), esp

safetyarxiv-cs-cl
15 Apr 2026
Safety

Advancing Multi-Agent RAG Systems with Minimalist Reinforcement Learning

DGX agent

arXiv:2505.17086v4 Announce Type: replace Abstract: Large Language Models (LLMs) equipped with modern Retrieval-Augmented Generation (RAG) systems often employ multi-turn interaction pipelines to inte

safetyarxiv-cs-cl
15 Apr 2026
Safety

ART-VITON: Measurement-Guided Latent Diffusion for Artifact-Free Virtual Try-On

DGX agent

arXiv:2509.25749v2 Announce Type: cross Abstract: Virtual try-on (VITON) aims to generate realistic images of a person wearing a target garment, requiring precise garment alignment in try-on regions a

safetyarxiv-cs-ai
15 Apr 2026
Safety

BayMOTH: Bayesian optiMizatiOn with meTa-lookahead -- a simple approacH

DGX agent

arXiv:2604.12005v1 Announce Type: cross Abstract: Bayesian optimization (BO) has for sequential optimization of expensive black-box functions demonstrated practicality and effectiveness in many real-w

safetyarxiv-cs-ai
15 Apr 2026
Safety

Beyond Factual Grounding: The Case for Opinion-Aware Retrieval-Augmented Generation

DGX agent

arXiv:2604.12138v1 Announce Type: new Abstract: RAG systems have transformed how LLMs access external knowledge, but we find that current implementations exhibit a bias toward factual, objective conte

safetyarxiv-cs-ai
15 Apr 2026
Safety

Beyond Transcription: Unified Audio Schema for Perception-Aware AudioLLMs

DGX agent

arXiv:2604.12506v1 Announce Type: new Abstract: Recent Audio Large Language Models (AudioLLMs) exhibit a striking performance inversion: while excelling at complex reasoning tasks, they consistently u

safetyarxiv-cs-cl
15 Apr 2026
Safety

Black-Box Optimization From Small Offline Datasets via Meta Learning with Synthetic Tasks

DGX agent

arXiv:2604.12325v1 Announce Type: cross Abstract: We consider the problem of offline black-box optimization, where the goal is to discover optimal designs (e.g., molecules or materials) from past expe

safetyarxiv-cs-ai
15 Apr 2026
Safety

BRAIN: Bias-Mitigation Continual Learning Approach to Vision-Brain Understanding

DGX agent

arXiv:2508.18187v2 Announce Type: replace-cross Abstract: Memory decay makes it harder for the human brain to recognize visual objects and retain details. Consequently, recorded brain signals become w

safetyarxiv-cs-ai
15 Apr 2026
Safety

Brain-DiT: A Universal Multi-state fMRI Foundation Model with Metadata-Conditioned Pretraining

DGX agent

arXiv:2604.12683v1 Announce Type: new Abstract: Current fMRI foundation models primarily rely on a limited range of brain states and mismatched pretraining tasks, restricting their ability to learn ge

safetyarxiv-cs-cv
15 Apr 2026
Safety

Bridging the Micro--Macro Gap: Frequency-Aware Semantic Alignment for Image Manipulation Localization

DGX agent

arXiv:2604.12341v1 Announce Type: new Abstract: As generative image editing advances, image manipulation localization (IML) must handle both traditional manipulations with conspicuous forensic artifac

safetyarxiv-cs-cv
15 Apr 2026
Safety

Calibration-Aware Policy Optimization for Reasoning LLMs

DGX agent

arXiv:2604.12632v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) enhances LLM reasoning but often induces overconfidence, where incorrect responses yield lower perplexity th

safetyarxiv-cs-ai
15 Apr 2026
Safety

CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning

DGX agent

arXiv:2602.00181v3 Announce Type: replace-cross Abstract: Understanding camera dynamics is a fundamental pillar of video spatial intelligence. However, existing multimodal models predominantly treat t

safetyarxiv-cs-ai
15 Apr 2026
Safety

Causal Diffusion Models for Counterfactual Outcome Distributions in Longitudinal Data

DGX agent

arXiv:2604.12992v1 Announce Type: cross Abstract: Predicting counterfactual outcomes in longitudinal data, where sequential treatment decisions heavily depend on evolving patient states, is critical y

safetyarxiv-cs-lg
15 Apr 2026
Safety

Challenging Vision-Language Models with Physically Deployable Multimodal Semantic Lighting Attacks

DGX agent

arXiv:2604.12833v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown remarkable performance, yet their security remains insufficiently understood. Existing adversarial studies focu

safetyarxiv-cs-cv
15 Apr 2026
Safety

CIA: Inferring the Communication Topology from LLM-based Multi-Agent Systems

DGX agent

arXiv:2604.12461v1 Announce Type: new Abstract: LLM-based Multi-Agent Systems (MAS) have demonstrated remarkable capabilities in solving complex tasks. Central to MAS is the communication topology whi

safetyarxiv-cs-ai
15 Apr 2026
Safety

CLEAR: Cross-Lingual Enhancement in Alignment via Reverse-training

DGX agent

arXiv:2604.05821v2 Announce Type: replace Abstract: Existing multilingual embedding models often encounter challenges in cross-lingual scenarios due to imbalanced linguistic resources and less conside

safetyarxiv-cs-cl
15 Apr 2026
Safety

Combating Pattern and Content Bias: Adversarial Feature Learning for Generalized AI-Generated Image Detection

DGX agent

arXiv:2604.12353v1 Announce Type: new Abstract: In recent years, the rapid development of generative artificial intelligence technology has significantly lowered the barrier to creating high-quality f

safetyarxiv-cs-cv
15 Apr 2026
Safety

Contextual Multi-Task Reinforcement Learning for Autonomous Reef Monitoring

DGX agent

arXiv:2604.12645v1 Announce Type: cross Abstract: Although autonomous underwater vehicles promise the capability of marine ecosystem monitoring, their deployment is fundamentally limited by the diffic

safetyarxiv-cs-ai
15 Apr 2026
Safety

Continuous Knowledge Metabolism: Generating Scientific Hypotheses from Evolving Literature

DGX agent

arXiv:2604.12243v1 Announce Type: cross Abstract: Scientific hypothesis generation requires tracking how knowledge evolves, not just what is currently known. We introduce Continuous Knowledge Metaboli

safetyarxiv-cs-ai
15 Apr 2026
Safety

CoSyncDiT: Cognitive Synchronous Diffusion Transformer for Movie Dubbing

DGX agent

arXiv:2604.12292v1 Announce Type: cross Abstract: Movie dubbing aims to synthesize speech that preserves the vocal identity of a reference audio while synchronizing with the lip movements in a target

safetyarxiv-cs-cv
15 Apr 2026
Safety

Cross-Cultural Simulation of Citizen Emotional Responses to Bureaucratic Red Tape Using LLM Agents

DGX agent

arXiv:2604.12545v1 Announce Type: new Abstract: Improving policymaking is a central concern in public administration. Prior human subject studies reveal substantial cross-cultural differences in citiz

safetyarxiv-cs-ai
15 Apr 2026
Safety

Cross-Modal Knowledge Distillation for PET-Free Amyloid-Beta Detection from MRI

DGX agent

arXiv:2604.12574v1 Announce Type: new Abstract: Detecting amyloid-eta (Aeta) positivity is crucial for early diagnosis of Alzheimer's disease but typically requires PET imaging, which is costly, invas

safetyarxiv-cs-cv
15 Apr 2026
Safety

Cycle-Consistent Search: Question Reconstructability as a Proxy Reward for Search Agent Training

DGX agent

arXiv:2604.12967v1 Announce Type: new Abstract: Reinforcement Learning (RL) has shown strong potential for optimizing search agents in complex information retrieval tasks. However, existing approaches

safetyarxiv-cs-ai
15 Apr 2026
Safety

DBGL: Decay-aware Bipartite Graph Learning for Irregular Medical Time Series Classification

DGX agent

arXiv:2604.11842v1 Announce Type: cross Abstract: Irregular Medical Time Series play a critical role in the clinical domain to better understand the patient's condition. However, inherent irregularity

safetyarxiv-cs-ai
15 Apr 2026
Safety

Designing Reliable LLM-Assisted Rubric Scoring for Constructed Responses: Evidence from Physics Exams

DGX agent

arXiv:2604.12227v1 Announce Type: new Abstract: Student responses in STEM assessments are often handwritten and combine symbolic expressions, calculations, and diagrams, creating substantial variation

safetyarxiv-cs-ai
15 Apr 2026
Safety

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding

DGX agent

arXiv:2604.12812v1 Announce Type: new Abstract: Existing Multimodal Large Language Models (MLLMs) suffer from significant performance degradation on the long document understanding task as document le

safetyarxiv-cs-ai
15 Apr 2026
Safety

Does RLVR Extend Reasoning Boundaries? Investigating Capability Expansion in Vision-Language Models

DGX agent

arXiv:2511.00710v4 Announce Type: replace Abstract: Recent studies posit that Reinforcement Learning with Verifiable Rewards (RLVR) primarily amplifies behaviors inherent to the pre-training distribut

safetyarxiv-cs-ai
15 Apr 2026
Safety

DyBBT: Dynamic Balance via Bandit-inspired Targeting for Dialog Policy with Cognitive Dual-Systems

DGX agent

arXiv:2509.19695v3 Announce Type: replace-cross Abstract: Task oriented dialog systems often rely on static exploration strategies that do not adapt to dynamic dialog contexts, leading to inefficient

safetyarxiv-cs-ai
15 Apr 2026
Safety

Dynamic Multi-Robot Task Allocation under Uncertainty and Communication Constraints: A Game-Theoretic Approach

DGX agent

arXiv:2604.11954v1 Announce Type: cross Abstract: We study dynamic multi-robot task allocation under uncertain task completion, time-window constraints, and incomplete information. Tasks arrive online

safetyarxiv-cs-ro
15 Apr 2026
Safety

E2E-Fly: An Integrated Training-to-Deployment System for End-to-End Quadrotor Autonomy

DGX agent

arXiv:2604.12916v1 Announce Type: new Abstract: Training and transferring learning-based policies for quadrotors from simulation to reality remains challenging due to inefficient visual rendering, phy

safetyarxiv-cs-ro
15 Apr 2026
Safety

Enhancing Agentic Textual Graph Retrieval with Synthetic Stepwise Supervision

DGX agent

arXiv:2510.03323v2 Announce Type: replace Abstract: Integrating textual graphs into Large Language Models (LLMs) is promising for complex graph-based QA. However, a key bottleneck is retrieving inform

safetyarxiv-cs-cl
15 Apr 2026
Safety

Euler-inspired Decoupling Neural Operator for Efficient Pansharpening

DGX agent

arXiv:2604.12463v1 Announce Type: cross Abstract: Pansharpening aims to synthesize high-resolution multispectral (HR-MS) images by fusing the spatial textures of panchromatic (PAN) images with the spe

safetyarxiv-cs-ai
15 Apr 2026
Safety

Evaluating the Limitations of Protein Sequence Representations for Parkinson's Disease Classification

DGX agent

arXiv:2604.11852v1 Announce Type: cross Abstract: The identification of reliable molecular biomarkers for Parkinson's disease remains challenging due to its multifactorial nature. Although protein seq

safetyarxiv-cs-ai
15 Apr 2026
Safety

Evolution-Inspired Sample Competition for Deep Neural Network Optimization

DGX agent

arXiv:2604.12568v1 Announce Type: new Abstract: Conventional deep network training generally optimizes all samples under a largely uniform learning paradigm, without explicitly modeling the heterogene

safetyarxiv-cs-cv
15 Apr 2026
Safety

EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution

DGX agent

arXiv:2604.12776v1 Announce Type: new Abstract: Realizing endogenous narrative evolution in LLM-based multi-agent systems is hindered by the inherent stochasticity of generative emergence. In particul

safetyarxiv-cs-cl
15 Apr 2026
Safety

From Myopic Selection to Long-Horizon Awareness: Sequential LLM Routing for Multi-Turn Dialogue

DGX agent

arXiv:2604.12385v1 Announce Type: new Abstract: Multi-turn dialogue is the predominant form of interaction with large language models (LLMs). While LLM routing is effective in single-turn settings, ex

safetyarxiv-cs-cl
15 Apr 2026
Safety

GeoAlign: Geometric Feature Realignment for MLLM Spatial Reasoning

DGX agent

arXiv:2604.12630v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have exhibited remarkable performance in various visual tasks, yet still struggle with spatial reasoning. Rec

safetyarxiv-cs-cl
15 Apr 2026
Safety

Gradient boundaries through confidence intervals for forced alignment estimates using model ensembles

DGX agent

arXiv:2506.01256v4 Announce Type: replace-cross Abstract: Forced alignment is a common tool to align audio with orthographic and phonetic transcriptions. Most forced alignment tools provide only point

safetyarxiv-cs-cl
15 Apr 2026
Safety

Gradient flow dynamics of shallow ReLU networks for square loss and orthogonal inputs

DGX agent

arXiv:2206.00939v3 Announce Type: replace-cross Abstract: The training of neural networks by gradient descent methods is a cornerstone of the deep learning revolution. Yet, despite some recent progres

safetyarxiv-cs-lg
15 Apr 2026
Safety

Graph-Based Chain-of-Thought Pruning for Reducing Redundant Reflections in Reasoning LLMs

DGX agent

arXiv:2604.05643v2 Announce Type: replace Abstract: Extending CoT through RL has been widely used to enhance the reasoning capabilities of LLMs. However, due to the sparsity of reward signals, it can

safetyarxiv-cs-cl
15 Apr 2026
Safety

Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions

DGX agent

arXiv:2604.12929v1 Announce Type: new Abstract: We present Grasp in Gaussians (GraG), a fast and robust method for reconstructing dynamic 3D hand-object interactions from a single monocular video. Unl

safetyarxiv-cs-cv
15 Apr 2026
Safety

GUIDE: Guided Updates for In-context Decision Evolution in LLM-Driven Spacecraft Operations

DGX agent

arXiv:2603.27306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been proposed as supervisory agents for spacecraft operations, but existing approaches rely on static prompt

safetyarxiv-cs-ai
15 Apr 2026
Safety

Hail to the Thief: Exploring Attacks and Defenses in Decentralised GRPO

DGX agent

arXiv:2511.09780v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has demonstrated wide adoption in the post-training of Large Language Models (LLMs). In GRPO, prompts are

safetyarxiv-cs-lg
15 Apr 2026
Safety

HiCoLoRA: Addressing Context-Prompt Misalignment via Hierarchical Collaborative LoRA for Zero-Shot DST

DGX agent

arXiv:2509.19742v4 Announce Type: replace-cross Abstract: Zero-shot Dialog State Tracking (zs-DST) is essential for enabling Task-Oriented Dialog Systems (TODs) to generalize to new domains without co

safetyarxiv-cs-ai
15 Apr 2026
Safety

Human-Centric Topic Modeling with Goal-Prompted Contrastive Learning and Optimal Transport

DGX agent

arXiv:2604.12663v1 Announce Type: new Abstract: Existing topic modeling methods, from LDA to recent neural and LLM-based approaches, which focus mainly on statistical coherence, often produce redundan

safetyarxiv-cs-ai
15 Apr 2026
Safety

Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance

DGX agent

arXiv:2511.21356v2 Announce Type: replace-cross Abstract: Adversarial Inverse Reinforcement Learning (AIRL) has shown promise in addressing the sparse reward problem in reinforcement learning (RL) by

safetyarxiv-cs-ai
15 Apr 2026
← Previous
1…225226227228229…257
Next →