AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,600 results
14 Apr 2026

Closed-Form Concept Erasure via Double Projections

SafetyDGX agent

arXiv:2604.10032v1 Announce Type: cross Abstract: While modern generative models such as diffusion-based architectures have enabled impressive creative capabilities, they also raise important safety a

ComSim: Building Scalable Real-World Robot Data Generation via Compositional Simulation

SafetyDGX agent

arXiv:2604.11386v1 Announce Type: cross Abstract: Recent advancements in foundational models, such as large language models and world models, have greatly enhanced the capabilities of robotics, enabli

ConfigSpec: Profiling-Based Configuration Selection for Distributed Edge--Cloud Speculative LLM Serving

SafetyDGX agent

arXiv:2604.09722v1 Announce Type: cross Abstract: Speculative decoding enables collaborative Large Language Model (LLM) inference across cloud and edge by separating lightweight token drafting from he


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation

SafetyDGX agent

arXiv:2604.09746v1 Announce Type: cross Abstract: As large language models (LLMs) are increasingly deployed as autonomous agents, understanding how strategic behavior emerges in multi-agent environmen

Consolidation or Adaptation? PRISM: Disentangling SFT and RL Data via Gradient Concentration

SafetyDGX agent

arXiv:2601.07224v2 Announce Type: replace Abstract: While Hybrid Supervised Fine-Tuning (SFT) followed by Reinforcement Learning (RL) has become the standard paradigm for training LLM agents, effectiv

Context Matters: Vision-Based Depression Detection Comparing Classical and Deep Approaches

SafetyDGX agent

arXiv:2604.10344v1 Announce Type: new Abstract: The classical approach to detecting depression from vision emphasizes interpretable features, such as facial expression, and classifiers such as the Sup

Contour Refinement using Discrete Diffusion in Low Data Regime

SafetyDGX agent

arXiv:2602.05880v2 Announce Type: replace Abstract: Boundary detection of irregular and translucent objects is an important problem with applications in medical imaging, environmental monitoring and m

CoPS: Conditional Prompt Synthesis for Zero-Shot Anomaly Detection

SafetyDGX agent

arXiv:2508.03447v2 Announce Type: replace Abstract: Recently, large pre-trained vision-language models have shown remarkable performance in zero-shot anomaly detection (ZSAD). With fine-tuning on a si

COSMIK-MPPI: Scaling Constrained Model Predictive Control to Collision Avoidance in Close-Proximity Dynamic Human Environments

SafetyDGX agent

arXiv:2604.10358v1 Announce Type: new Abstract: Ensuring safe physical interaction between torque-controlled manipulators and humans is essential for deploying robots in everyday environments. Model P

Cost-optimal Sequential Testing via Doubly Robust Q-learning

SafetyDGX agent

arXiv:2604.11165v1 Announce Type: cross Abstract: Clinical decision-making often involves selecting tests that are costly, invasive, or time-consuming, motivating individualized, sequential strategies

CoSToM:Causal-oriented Steering for Intrinsic Theory-of-Mind Alignment in Large Language Models

SafetyDGX agent

arXiv:2604.10031v1 Announce Type: cross Abstract: Theory of Mind (ToM), the ability to attribute mental states to others, is a hallmark of social intelligence. While large language models (LLMs) demon

COXNet: Cross-Layer Fusion with Adaptive Alignment and Scale Integration for RGBT Tiny Object Detection

SafetyDGX agent

arXiv:2508.09533v2 Announce Type: replace-cross Abstract: Detecting tiny objects in multimodal Red-Green-Blue-Thermal (RGBT) imagery is a critical challenge in computer vision, particularly in surveil

CROP: Conservative Reward for Model-based Offline Policy Optimization

SafetyDGX agent

arXiv:2310.17245v2 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL) aims to optimize a policy using collected data without online interactions. Model-based approaches are par

Cross-Cultural Bias in Mel-Scale Representations: Evidence and Alternatives from Speech and Music

SafetyDGX agent

arXiv:2604.10503v1 Announce Type: cross Abstract: Modern audio systems universally employ mel-scale representations derived from 1940s Western psychoacoustic studies, potentially encoding cultural bia

Cross-Cultural Value Awareness in Large Vision-Language Models

SafetyDGX agent

arXiv:2604.09945v1 Announce Type: cross Abstract: The rapid adoption of large vision-language models (LVLMs) in recent years has been accompanied by growing fairness concerns due to their propensity t

CSPO: Alleviating Reward Ambiguity for Structured Table-to-LaTeX Generation

SafetyDGX agent

arXiv:2604.10918v1 Announce Type: new Abstract: Tables contain rich structured information, yet when stored as images their contents remain 'locked' within pixels. Converting table images into LaTeX c

Curriculum-based Sample Efficient Reinforcement Learning for Robust Stabilization of a Quadrotor

SafetyDGX agent

arXiv:2501.18490v3 Announce Type: replace-cross Abstract: This article introduces a novel sample-efficient curriculum learning (CL) approach for training an end-to-end reinforcement learning (RL) poli

DecAlign: Hierarchical Cross-Modal Alignment for Decoupled Multimodal Representation Learning

SafetyDGX agent

arXiv:2503.11892v3 Announce Type: replace Abstract: Multimodal representation learning aims to capture both shared and complementary semantic information across multiple modalities. However, the intri

Decision-Theoretic Safety Assessment of Persona-Driven Multi-Agent Systems in O-RAN

SafetyDGX agent

arXiv:2604.09682v1 Announce Type: cross Abstract: Autonomous network management in Open Radio Access Networks requires intelligent decision making across conflicting objectives, yet existing LLM based

Deep deterministic policy gradient with symmetric data augmentation for lateral attitude tracking control of a fixed-wing aircraft

SafetyDGX agent

arXiv:2407.11077v3 Announce Type: replace-cross Abstract: The symmetry of dynamical systems can be exploited for state-transition prediction and to facilitate control policy optimization. This paper l

DeepFleet: Multi-Agent Foundation Models for Mobile Robots

SafetyDGX agent

arXiv:2508.08574v3 Announce Type: replace Abstract: We introduce DeepFleet, a suite of foundation models designed to support coordination and planning for large-scale mobile robot fleets. These models

Deliberative Alignment is Deep, but Uncertainty Remains: Inference time safety improvement in reasoning via attribution of unsafe behavior to base model

SafetyDGX agent

arXiv:2604.09665v1 Announce Type: cross Abstract: While the wide adoption of refusal training in large language models (LLMs) has showcased improvements in model safety, recent works have highlighted

Demographic and Linguistic Bias Evaluation in Omnimodal Language Models

SafetyDGX agent

arXiv:2604.10014v1 Announce Type: cross Abstract: This paper provides a comprehensive evaluation of demographic and linguistic biases in omnimodal language models that process text, images, audio, and

Descriptor-Injected Cross-Modal Learning: A Systematic Exploration of Audio-MIDI Alignment via Spectral and Melodic Features

SafetyDGX agent

arXiv:2604.10283v1 Announce Type: cross Abstract: Cross-modal retrieval between audio recordings and symbolic music representations (MIDI) remains challenging because continuous waveforms and discrete

Design Experiments to Compare Multi-armed Bandit Algorithms

SafetyDGX agent

arXiv:2603.05919v2 Announce Type: replace Abstract: Online platforms routinely compare multi-armed bandit algorithms, such as UCB and Thompson Sampling, to select the best-performing policy. Unlike st

Designing Adaptive Digital Nudging Systems with LLM-Driven Reasoning

SafetyDGX agent

arXiv:2604.11206v1 Announce Type: cross Abstract: Digital nudging systems lack architectural guidance for translating behavioral science into software design. While research identifies nudge strategie

Detection Is Cheap, Routing Is Learned: Why Refusal-Based Alignment Evaluation Fails

SafetyDGX agent

arXiv:2603.18280v2 Announce Type: replace-cross Abstract: Current alignment evaluation mostly measures whether models encode dangerous concepts and whether they refuse harmful requests. Both miss the

Device-Conditioned Neural Architecture Search for Efficient Robotic Manipulation

SafetyDGX agent

arXiv:2604.10170v1 Announce Type: cross Abstract: The growing complexity of visuomotor policies poses significant challenges for deployment with heterogeneous robotic hardware constraints. However, mo

Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent Debate

SafetyDGX agent

arXiv:2604.11258v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) in healthcare suffer from severe confirmation bias, often hallucinating visual details to support initial, pote

Dialogue based Interactive Explanations for Safety Decisions in Human Robot Collaboration

SafetyDGX agent

arXiv:2604.05896v2 Announce Type: replace Abstract: As robots increasingly operate in shared, safety critical environments, acting safely is no longer sufficient robots must also make their safety dec

Differentiable free energy surface: a variational approach to directly observing rare events using generative deep-learning models

SafetyDGX agent

arXiv:2604.09769v1 Announce Type: cross Abstract: Rare events are central to the evolution of complex many-body systems, characterized as key transitional configurations on the free energy surface (FE

Diffusion-Based Generative Priors for Efficient Beam Alignment in Directional Networks

SafetyDGX agent

arXiv:2604.09653v1 Announce Type: cross Abstract: Beam alignment is a key challenge in directional mmWave and THz systems, where narrow beams require accurate yet low-overhead training. Existing learn

Discrete Flow Maps

SafetyDGX agent

arXiv:2604.09784v1 Announce Type: cross Abstract: The sequential nature of autoregressive next-token prediction imposes a fundamental speed limit on large language models. While continuous flow models

DistDF: Time-Series Forecasting Needs Joint-Distribution Wasserstein Alignment

SafetyDGX agent

arXiv:2510.24574v2 Announce Type: replace-cross Abstract: Training time-series forecasting models requires aligning the conditional distribution of model forecasts with that of the label sequence. The

Distributionally Robust PAC-Bayesian Control

SafetyDGX agent

arXiv:2604.10588v1 Announce Type: new Abstract: We present a distributionally robust PAC-Bayesian framework for certifying the performance of learning-based finite-horizon controllers. While existing

Do LLMs Know Tool Irrelevance? Demystifying Structural Alignment Bias in Tool Invocations

SafetyDGX agent

arXiv:2604.11322v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated impressive capabilities in utilizing external tools. In practice, however, LLMs are often exposed to to

Do Machines Fail Like Humans? A Human-Centred Out-of-Distribution Spectrum for Mapping Error Alignment

SafetyDGX agent

arXiv:2603.07462v2 Announce Type: replace Abstract: Determining whether AI systems process information similarly to humans is central to cognitive science and trustworthy AI. While modern AI models ca

DRIFT: Deep Restoration, ISP Fusion, and Tone-mapping

SafetyDGX agent

arXiv:2604.03402v2 Announce Type: replace-cross Abstract: Smartphone cameras have gained immense popularity with the adoption of high-resolution and high-dynamic range imaging. As a result, high-perfo

dTRPO: Trajectory Reduction in Policy Optimization of Diffusion Large Language Models

SafetyDGX agent

arXiv:2603.18806v2 Announce Type: replace Abstract: Diffusion Large Language Models (dLLMs) introduce a new paradigm for language generation, which in turn presents new challenges for aligning them wi

Dual-Exposure Imaging with Events

SafetyDGX agent

arXiv:2604.10273v1 Announce Type: new Abstract: By combining complementary benefits of short- and long-exposure images, Dual-Exposure Imaging (DEI) enhances image quality in low-light scenarios. Howev

Early Decisions Matter: Proximity Bias and Initial Trajectory Shaping in Non-Autoregressive Diffusion Language Models

SafetyDGX agent

arXiv:2604.10567v1 Announce Type: cross Abstract: Diffusion-based language models (dLLMs) have emerged as a promising alternative to autoregressive language models, offering the potential for parallel

EE-MCP: Self-Evolving MCP-GUI Agents via Automated Environment Generation and Experience Learning

SafetyDGX agent

arXiv:2604.09815v1 Announce Type: new Abstract: Computer-use agents that combine GUI interaction with structured API calls via the Model Context Protocol (MCP) show promise for automating software tas

EEPO: Exploration-Enhanced Policy Optimization via Sample-Then-Forget

SafetyDGX agent

arXiv:2510.05837v2 Announce Type: replace Abstract: Balancing exploration and exploitation remains a central challenge in reinforcement learning with verifiable rewards (RLVR) for large language model

Efficient Disruption of Criminal Networks through Multi-Objective Genetic Algorithms

SafetyDGX agent

arXiv:2604.09647v1 Announce Type: cross Abstract: Criminal networks, such as the Sicilian Mafia, pose substantial threats to public safety, national security, and economic stability. Outdated disrupti

Efficient KernelSHAP Explanations for Patch-based 3D Medical Image Segmentation

SafetyDGX agent

arXiv:2604.11775v1 Announce Type: cross Abstract: Perturbation-based explainability methods such as KernelSHAP provide model-agnostic attributions but are typically impractical for patch-based 3D medi

Efficient Training for Cross-lingual Speech Language Models

SafetyDGX agent

arXiv:2604.11096v1 Announce Type: cross Abstract: Currently, large language models (LLMs) predominantly focus on the text modality. To enable more natural human-AI interaction, speech LLMs are emergin

Elon Musk’s AI software, Grok, continues to generate sexualized images of people without their consent, despite his company’s pledge months …

SafetyDGX agent

Elon Musk’s AI software, Grok, continues to generate sexualized images of people without their consent, despite his company’s pledge months ago to halt abusive deepfakes after a public backlash and go

EmergentBridge: Improving Zero-Shot Cross-Modal Transfer in Unified Multimodal Embedding Models

SafetyDGX agent

arXiv:2604.11043v1 Announce Type: new Abstract: Unified multimodal embedding spaces underpin practical applications such as cross-modal retrieval and zero-shot recognition. In many real deployments, h

Empowering Video Translation using Multimodal Large Language Models

SafetyDGX agent

arXiv:2604.11283v1 Announce Type: new Abstract: Recent developments in video translation have further enhanced cross-lingual access to video content, with multimodal large language models (MLLMs) play

Enabling and Inhibitory Pathways of Students' AI Use Concealment Intention in Higher Education: Evidence from SEM and fsQCA

SafetyDGX agent

arXiv:2604.10978v1 Announce Type: cross Abstract: This study investigates students' AI use concealment intention in higher education by integrating the cognition-affect-conation (CAC) framework with a

End-to-end Contrastive Language-Speech Pretraining Model For Long-form Spoken Question Answering

SafetyDGX agent

arXiv:2511.09282v3 Announce Type: replace-cross Abstract: Significant progress has been made in spoken question answering (SQA) in recent years. However, many existing methods, including large audio l

Endogenous Information in Routing Games: Memory-Constrained Equilibria, Recall Braess Paradoxes, and Memory Design

SafetyDGX agent

arXiv:2604.11733v1 Announce Type: cross Abstract: We study routing games in which travelers optimize over routes that are remembered or surfaced, rather than over a fixed exogenous action set. The pap

Enhancing Fine-Grained Spatial Grounding in 3D CT Report Generation via Discriminative Guidance

SafetyDGX agent

arXiv:2604.10437v1 Announce Type: new Abstract: Vision--language models (VLMs) for radiology report generation (RRG) can produce long-form chest CT reports from volumetric scans and show strong potent

Evaluating the Impact of Medical Image Reconstruction on Downstream AI Fairness and Performance

SafetyDGX agent

arXiv:2604.10904v1 Announce Type: cross Abstract: AI-based image reconstruction models are increasingly deployed in clinical workflows to improve image quality from noisy data, such as low-dose X-rays

Evolutionary Token-Level Prompt Optimization for Diffusion Models

SafetyDGX agent

arXiv:2604.09861v1 Announce Type: new Abstract: Text-to-image diffusion models exhibit strong generative performance but remain highly sensitive to prompt formulation, often requiring extensive manual

EvoNash-MARL: A Closed-Loop Multi-Agent Reinforcement Learning Framework for Medium-Horizon Equity Allocation

SafetyDGX agent

arXiv:2604.10911v1 Announce Type: new Abstract: Medium-to-long-horizon stock allocation presents significant challenges due toveak predictive structures, non-stadonary market regimes, and the degradat

Examining EAP Students' AI Disclosure Intention: A Cognition-Affect-Conation Perspective

SafetyDGX agent

arXiv:2604.10991v1 Announce Type: cross Abstract: The growing use of generative artificial intelligence (AI) in academic writing has raised increasing concerns regarding transparency and academic inte

Exciting news! @theaidocfilm is now available to rent or buy on Apple TV, Prime Video, or wherever you get your movies on demand. If you mis…

SafetyDGX agent

Exciting news! @theaidocfilm is now available to rent or buy on Apple TV, Prime Video, or wherever you get your movies on demand. If you missed the theatrical release, you can now watch this urgent fi

Explainability and Certification of AI-Generated Educational Assessments

SafetyDGX agent

arXiv:2604.09622v1 Announce Type: cross Abstract: The rapid adoption of generative artificial intelligence (AI) in educational assessment has created new opportunities for scalable item creation, pers

Explainable Planning for Hybrid Systems

SafetyDGX agent

arXiv:2604.09578v1 Announce Type: new Abstract: The recent advancement in artificial intelligence (AI) technologies facilitates a paradigm shift toward automation. Autonomous systems are fully or part

← Previous
1…198199200201202…210
Next →