AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,600 results
15 Apr 2026

ARGen: Affect-Reinforced Generative Augmentation towards Vision-based Dynamic Emotion Perception

SafetyDGX agent

arXiv:2604.12255v1 Announce Type: cross Abstract: Dynamic facial expression recognition in the wild remains challenging due to data scarcity and long-tail distributions, which hinder models from effec

ART-VITON: Measurement-Guided Latent Diffusion for Artifact-Free Virtual Try-On

SafetyDGX agent

arXiv:2509.25749v2 Announce Type: cross Abstract: Virtual try-on (VITON) aims to generate realistic images of a person wearing a target garment, requiring precise garment alignment in try-on regions a

ASGuard: Activation-Scaling Guard to Mitigate Targeted Jailbreaking Attack

SafetyDGX agent

arXiv:2509.25843v2 Announce Type: replace Abstract: Large language models (LLMs), despite being safety-aligned, exhibit brittle refusal behaviors that can be circumvented by simple linguistic changes.


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

BayMOTH: Bayesian optiMizatiOn with meTa-lookahead -- a simple approacH

SafetyDGX agent

arXiv:2604.12005v1 Announce Type: cross Abstract: Bayesian optimization (BO) has for sequential optimization of expensive black-box functions demonstrated practicality and effectiveness in many real-w

Beyond Factual Grounding: The Case for Opinion-Aware Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2604.12138v1 Announce Type: new Abstract: RAG systems have transformed how LLMs access external knowledge, but we find that current implementations exhibit a bias toward factual, objective conte

Beyond Static Sandboxing: Learned Capability Governance for Autonomous AI Agents

SafetyDGX agent

arXiv:2604.11839v1 Announce Type: cross Abstract: Autonomous AI agents built on open-source runtimes such as OpenClaw expose every available tool to every session by default, regardless of the task. A

Beyond Transcription: Unified Audio Schema for Perception-Aware AudioLLMs

SafetyDGX agent

arXiv:2604.12506v1 Announce Type: new Abstract: Recent Audio Large Language Models (AudioLLMs) exhibit a striking performance inversion: while excelling at complex reasoning tasks, they consistently u

Black-Box Optimization From Small Offline Datasets via Meta Learning with Synthetic Tasks

SafetyDGX agent

arXiv:2604.12325v1 Announce Type: cross Abstract: We consider the problem of offline black-box optimization, where the goal is to discover optimal designs (e.g., molecules or materials) from past expe

BRAIN: Bias-Mitigation Continual Learning Approach to Vision-Brain Understanding

SafetyDGX agent

arXiv:2508.18187v2 Announce Type: replace-cross Abstract: Memory decay makes it harder for the human brain to recognize visual objects and retain details. Consequently, recorded brain signals become w

Brain-DiT: A Universal Multi-state fMRI Foundation Model with Metadata-Conditioned Pretraining

SafetyDGX agent

arXiv:2604.12683v1 Announce Type: new Abstract: Current fMRI foundation models primarily rely on a limited range of brain states and mismatched pretraining tasks, restricting their ability to learn ge

Bridging the Micro--Macro Gap: Frequency-Aware Semantic Alignment for Image Manipulation Localization

SafetyDGX agent

arXiv:2604.12341v1 Announce Type: new Abstract: As generative image editing advances, image manipulation localization (IML) must handle both traditional manipulations with conspicuous forensic artifac

Calibration-Aware Policy Optimization for Reasoning LLMs

SafetyDGX agent

arXiv:2604.12632v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) enhances LLM reasoning but often induces overconfidence, where incorrect responses yield lower perplexity th

CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning

SafetyDGX agent

arXiv:2602.00181v3 Announce Type: replace-cross Abstract: Understanding camera dynamics is a fundamental pillar of video spatial intelligence. However, existing multimodal models predominantly treat t

Causal Diffusion Models for Counterfactual Outcome Distributions in Longitudinal Data

SafetyDGX agent

arXiv:2604.12992v1 Announce Type: cross Abstract: Predicting counterfactual outcomes in longitudinal data, where sequential treatment decisions heavily depend on evolving patient states, is critical y

Challenging Vision-Language Models with Physically Deployable Multimodal Semantic Lighting Attacks

SafetyDGX agent

arXiv:2604.12833v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown remarkable performance, yet their security remains insufficiently understood. Existing adversarial studies focu

CIA: Inferring the Communication Topology from LLM-based Multi-Agent Systems

SafetyDGX agent

arXiv:2604.12461v1 Announce Type: new Abstract: LLM-based Multi-Agent Systems (MAS) have demonstrated remarkable capabilities in solving complex tasks. Central to MAS is the communication topology whi

CLEAR: Cross-Lingual Enhancement in Alignment via Reverse-training

SafetyDGX agent

arXiv:2604.05821v2 Announce Type: replace Abstract: Existing multilingual embedding models often encounter challenges in cross-lingual scenarios due to imbalanced linguistic resources and less conside

Cloud CISO Perspectives: How CISOs can pursue technical and cultural resilience (Q&A)

SafetyDGX agent

Welcome to the first Cloud CISO Perspectives for April 2026. Today, Thiébaut Meyer and Lia Wertheimer from Google Cloud’s Office of the CISO share Thiébaut’s conversation with Matt Rowe, chief securit

Combating Pattern and Content Bias: Adversarial Feature Learning for Generalized AI-Generated Image Detection

SafetyDGX agent

arXiv:2604.12353v1 Announce Type: new Abstract: In recent years, the rapid development of generative artificial intelligence technology has significantly lowered the barrier to creating high-quality f

Compiling Activation Steering into Weights via Null-Space Constraints for Stealthy Backdoors

SafetyDGX agent

arXiv:2604.12359v1 Announce Type: cross Abstract: Safety-aligned large language models (LLMs) are increasingly deployed in real-world pipelines, yet this deployment also enlarges the supply-chain atta

ContextLens: Modeling Imperfect Privacy and Safety Context for Legal Compliance

SafetyDGX agent

arXiv:2604.12308v1 Announce Type: new Abstract: Individuals' concerns about data privacy and AI safety are highly contextualized and extend beyond sensitive patterns. Addressing these issues requires

Contextual Multi-Task Reinforcement Learning for Autonomous Reef Monitoring

SafetyDGX agent

arXiv:2604.12645v1 Announce Type: cross Abstract: Although autonomous underwater vehicles promise the capability of marine ecosystem monitoring, their deployment is fundamentally limited by the diffic

Continuous Knowledge Metabolism: Generating Scientific Hypotheses from Evolving Literature

SafetyDGX agent

arXiv:2604.12243v1 Announce Type: cross Abstract: Scientific hypothesis generation requires tracking how knowledge evolves, not just what is currently known. We introduce Continuous Knowledge Metaboli

CoSyncDiT: Cognitive Synchronous Diffusion Transformer for Movie Dubbing

SafetyDGX agent

arXiv:2604.12292v1 Announce Type: cross Abstract: Movie dubbing aims to synthesize speech that preserves the vocal identity of a reference audio while synchronizing with the lip movements in a target

Cross-Cultural Simulation of Citizen Emotional Responses to Bureaucratic Red Tape Using LLM Agents

SafetyDGX agent

arXiv:2604.12545v1 Announce Type: new Abstract: Improving policymaking is a central concern in public administration. Prior human subject studies reveal substantial cross-cultural differences in citiz

Cross-Modal Knowledge Distillation for PET-Free Amyloid-Beta Detection from MRI

SafetyDGX agent

arXiv:2604.12574v1 Announce Type: new Abstract: Detecting amyloid-eta (Aeta) positivity is crucial for early diagnosis of Alzheimer's disease but typically requires PET imaging, which is costly, invas

Cycle-Consistent Search: Question Reconstructability as a Proxy Reward for Search Agent Training

SafetyDGX agent

arXiv:2604.12967v1 Announce Type: new Abstract: Reinforcement Learning (RL) has shown strong potential for optimizing search agents in complex information retrieval tasks. However, existing approaches

Dataset Safety in Autonomous Driving: Requirements, Risks, and Assurance

SafetyDGX agent

arXiv:2511.08439v2 Announce Type: replace Abstract: Dataset integrity is fundamental to the safety and reliability of AI systems, especially in autonomous driving. This paper presents a structured fra

DBGL: Decay-aware Bipartite Graph Learning for Irregular Medical Time Series Classification

SafetyDGX agent

arXiv:2604.11842v1 Announce Type: cross Abstract: Irregular Medical Time Series play a critical role in the clinical domain to better understand the patient's condition. However, inherent irregularity

Deep QP Safety Filter: Model-free Learning for Reachability-based Safety Filter

SafetyDGX agent

arXiv:2601.21297v2 Announce Type: replace Abstract: We introduce Deep QP Safety Filter, a fully data-driven safety layer for black-box dynamical systems. Our method learns a Quadratic-Program (QP) saf

Designing Reliable LLM-Assisted Rubric Scoring for Constructed Responses: Evidence from Physics Exams

SafetyDGX agent

arXiv:2604.12227v1 Announce Type: new Abstract: Student responses in STEM assessments are often handwritten and combine symbolic expressions, calculations, and diagrams, creating substantial variation

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding

SafetyDGX agent

arXiv:2604.12812v1 Announce Type: new Abstract: Existing Multimodal Large Language Models (MLLMs) suffer from significant performance degradation on the long document understanding task as document le

Does RLVR Extend Reasoning Boundaries? Investigating Capability Expansion in Vision-Language Models

SafetyDGX agent

arXiv:2511.00710v4 Announce Type: replace Abstract: Recent studies posit that Reinforcement Learning with Verifiable Rewards (RLVR) primarily amplifies behaviors inherent to the pre-training distribut

DyBBT: Dynamic Balance via Bandit-inspired Targeting for Dialog Policy with Cognitive Dual-Systems

SafetyDGX agent

arXiv:2509.19695v3 Announce Type: replace-cross Abstract: Task oriented dialog systems often rely on static exploration strategies that do not adapt to dynamic dialog contexts, leading to inefficient

Dynamic Multi-Robot Task Allocation under Uncertainty and Communication Constraints: A Game-Theoretic Approach

SafetyDGX agent

arXiv:2604.11954v1 Announce Type: cross Abstract: We study dynamic multi-robot task allocation under uncertain task completion, time-window constraints, and incomplete information. Tasks arrive online

E2E-Fly: An Integrated Training-to-Deployment System for End-to-End Quadrotor Autonomy

SafetyDGX agent

arXiv:2604.12916v1 Announce Type: new Abstract: Training and transferring learning-based policies for quadrotors from simulation to reality remains challenging due to inefficient visual rendering, phy

Enhancing Agentic Textual Graph Retrieval with Synthetic Stepwise Supervision

SafetyDGX agent

arXiv:2510.03323v2 Announce Type: replace Abstract: Integrating textual graphs into Large Language Models (LLMs) is promising for complex graph-based QA. However, a key bottleneck is retrieving inform

Euler-inspired Decoupling Neural Operator for Efficient Pansharpening

SafetyDGX agent

arXiv:2604.12463v1 Announce Type: cross Abstract: Pansharpening aims to synthesize high-resolution multispectral (HR-MS) images by fusing the spatial textures of panchromatic (PAN) images with the spe

Evaluating the Limitations of Protein Sequence Representations for Parkinson's Disease Classification

SafetyDGX agent

arXiv:2604.11852v1 Announce Type: cross Abstract: The identification of reliable molecular biomarkers for Parkinson's disease remains challenging due to its multifactorial nature. Although protein seq

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs

SafetyDGX agent

arXiv:2604.12616v1 Announce Type: new Abstract: The rapid evolution of Vision-Language Models (VLMs) has catalyzed unprecedented capabilities in artificial intelligence; however, this continuous modal

Evolution-Inspired Sample Competition for Deep Neural Network Optimization

SafetyDGX agent

arXiv:2604.12568v1 Announce Type: new Abstract: Conventional deep network training generally optimizes all samples under a largely uniform learning paradigm, without explicitly modeling the heterogene

EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution

SafetyDGX agent

arXiv:2604.12776v1 Announce Type: new Abstract: Realizing endogenous narrative evolution in LLM-based multi-agent systems is hindered by the inherent stochasticity of generative emergence. In particul

Fragile Reconstruction: Adversarial Vulnerability of Reconstruction-Based Detectors for Diffusion-Generated Images

SafetyDGX agent

arXiv:2604.12781v1 Announce Type: new Abstract: Recently, detecting AI-generated images produced by diffusion-based models has attracted increasing attention due to their potential threat to safety. A

From Myopic Selection to Long-Horizon Awareness: Sequential LLM Routing for Multi-Turn Dialogue

SafetyDGX agent

arXiv:2604.12385v1 Announce Type: new Abstract: Multi-turn dialogue is the predominant form of interaction with large language models (LLMs). While LLM routing is effective in single-turn settings, ex

FSD v14.3.1 review after 10 drives and many hours, here are my thoughts (it’s a great one) - v14.3.1 feels like a big jump even compared to …

SafetyDGX agent

FSD v14.3.1 review after 10 drives and many hours, here are my thoughts (it’s a great one) - v14.3.1 feels like a big jump even compared to FSD v14.3, even for “just” a point release build. Everything

Generative AI in nutshell: 'If the technology is as formidable as the [tech CEO's] claim, then they could be leading us toward existential d…

SafetyDGX agent

Generative AI in nutshell: 'If the technology is as formidable as the [tech CEO's] claim, then they could be leading us toward existential disaster; if the technology proves less transformative, and t

GeoAlign: Geometric Feature Realignment for MLLM Spatial Reasoning

SafetyDGX agent

arXiv:2604.12630v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have exhibited remarkable performance in various visual tasks, yet still struggle with spatial reasoning. Rec

Goal-Conditioned Neural ODEs with Guaranteed Safety and Stability for Learning-Based All-Pairs Motion Planning

SafetyDGX agent

arXiv:2604.02821v2 Announce Type: replace Abstract: This paper presents a learning-based approach for all-pairs motion planning, where the initial and goal states are allowed to be arbitrary points in

Gradient boundaries through confidence intervals for forced alignment estimates using model ensembles

SafetyDGX agent

arXiv:2506.01256v4 Announce Type: replace-cross Abstract: Forced alignment is a common tool to align audio with orthographic and phonetic transcriptions. Most forced alignment tools provide only point

Gradient flow dynamics of shallow ReLU networks for square loss and orthogonal inputs

SafetyDGX agent

arXiv:2206.00939v3 Announce Type: replace-cross Abstract: The training of neural networks by gradient descent methods is a cornerstone of the deep learning revolution. Yet, despite some recent progres

Graph-Based Chain-of-Thought Pruning for Reducing Redundant Reflections in Reasoning LLMs

SafetyDGX agent

arXiv:2604.05643v2 Announce Type: replace Abstract: Extending CoT through RL has been widely used to enhance the reasoning capabilities of LLMs. However, due to the sparsity of reward signals, it can

Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions

SafetyDGX agent

arXiv:2604.12929v1 Announce Type: new Abstract: We present Grasp in Gaussians (GraG), a fast and robust method for reconstructing dynamic 3D hand-object interactions from a single monocular video. Unl

Grok just hit its highest monthly traffic EVER: over 326 MILLION visits in March alone That’s a massive 61% jump YoY and up 9.3% just since …

SafetyDGX agent

Grok just hit its highest monthly traffic EVER: over 326 MILLION visits in March alone That’s a massive 61% jump YoY and up 9.3% just since February People love Grok because it's the only AI they can

GUIDE: Guided Updates for In-context Decision Evolution in LLM-Driven Spacecraft Operations

SafetyDGX agent

arXiv:2603.27306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been proposed as supervisory agents for spacecraft operations, but existing approaches rely on static prompt

Hail to the Thief: Exploring Attacks and Defenses in Decentralised GRPO

SafetyDGX agent

arXiv:2511.09780v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has demonstrated wide adoption in the post-training of Large Language Models (LLMs). In GRPO, prompts are

HiCoLoRA: Addressing Context-Prompt Misalignment via Hierarchical Collaborative LoRA for Zero-Shot DST

SafetyDGX agent

arXiv:2509.19742v4 Announce Type: replace-cross Abstract: Zero-shot Dialog State Tracking (zs-DST) is essential for enabling Task-Oriented Dialog Systems (TODs) to generalize to new domains without co

How Transformers Learn to Plan via Multi-Token Prediction

SafetyDGX agent

arXiv:2604.11912v1 Announce Type: cross Abstract: While next-token prediction (NTP) has been the standard objective for training language models, it often struggles to capture global structure in reas

Human-Centric Topic Modeling with Goal-Prompted Contrastive Learning and Optimal Transport

SafetyDGX agent

arXiv:2604.12663v1 Announce Type: new Abstract: Existing topic modeling methods, from LDA to recent neural and LLM-based approaches, which focus mainly on statistical coherence, often produce redundan

Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance

SafetyDGX agent

arXiv:2511.21356v2 Announce Type: replace-cross Abstract: Adversarial Inverse Reinforcement Learning (AIRL) has shown promise in addressing the sparse reward problem in reinforcement learning (RL) by

I find it fascinating that OpenAI's Global Affairs team will denounce AI doomers for being too negative and then, simultaneously, advocate f…

SafetyDGX agent

I find it fascinating that OpenAI's Global Affairs team will denounce AI doomers for being too negative and then, simultaneously, advocate for an Illinois bill that would shield them from liability if

← Previous
1…195196197198199…210
Next →