AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
Safety

Poetry overheard on http://B.sky:

DGX agent

Gary Marcus shared an observation or commentary about poetry encountered on a platform or service referenced as 'B.sky' (likely Bluesky, the decentralized social network). The post appears to document

safetygary-marcus--x
21 Apr 2026
Safety

Policy Testing in Markov Decision Processes

DGX agent

arXiv:2505.15342v2 Announce Type: replace-cross Abstract: We study the policy testing problem in discounted Markov decision processes (MDPs) in the fixed-confidence setting under a generative model wi

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
safetyarxiv-cs-lg
21 Apr 2026
Safety

PoliLegalLM: A Technical Report on a Large Language Model for Political and Legal Affairs

DGX agent

arXiv:2604.17543v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable success in general-domain tasks, yet their direct application to the legal domain remains challeng

safetyarxiv-cs-cl
21 Apr 2026
Safety

PowerCLIP: Powerset Alignment for Contrastive Pre-Training

DGX agent

arXiv:2511.23170v5 Announce Type: replace Abstract: Contrastive vision-language pre-training frameworks such as CLIP have demonstrated impressive zero-shot performance across a range of vision-languag

safetyarxiv-cs-cv
21 Apr 2026
Safety

PrinciplismQA: A Philosophy-Grounded Approach to Assessing LLM-Human Clinical Medical Ethics Alignment

DGX agent

arXiv:2508.05132v2 Announce Type: replace Abstract: As medical LLMs transition to clinical deployment, assessing their ethical reasoning capability becomes critical. While achieving high accuracy on k

safetyarxiv-cs-cl
21 Apr 2026
Safety

Privacy Collapse: Benign Fine-Tuning Can Break Contextual Privacy in Language Models

DGX agent

arXiv:2601.15220v2 Announce Type: replace Abstract: We identify a novel phenomenon in language models: benign fine-tuning of frontier models can lead to privacy collapse. We find that diverse, subtle

safetyarxiv-cs-cl
21 Apr 2026
Safety

ProtoCLIP: Prototype-Aligned Latent Refinement for Robust Zero-Shot Chest X-Ray Classification

DGX agent

arXiv:2604.18444v1 Announce Type: cross Abstract: Zero-shot vision-language models (VLMs) have shown promise for chest radiograph classification, but their performance is often limited by confounding

safetyarxiv-cs-cv
21 Apr 2026
Safety

Q-SINDy: Quantum-Kernel Sparse Identification of Nonlinear Dynamics with Provable Coefficient Debiasing

DGX agent

arXiv:2604.16779v1 Announce Type: cross Abstract: Quantum feature maps offer expressive embeddings for classical learning tasks, and augmenting sparse identification of nonlinear dynamics (SINDy) with

safetyarxiv-cs-lg
21 Apr 2026
Safety

R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation

DGX agent

arXiv:2506.07826v2 Announce Type: replace Abstract: Validating autonomous driving (AD) systems requires diverse and safety-critical testing, making photorealistic virtual environments essential. Tradi

safetyarxiv-cs-cv
21 Apr 2026
Safety

RAYEN: Imposition of Hard Convex Constraints on Neural Networks

DGX agent

arXiv:2307.08336v2 Announce Type: replace Abstract: Despite the numerous applications of convex constraints in Robotics, enforcing them within learning-based frameworks remains an open challenge. Exis

safetyarxiv-cs-lg
21 Apr 2026
Safety

Reasoning on the Manifold: Bidirectional Consistency for Self-Verification in Diffusion Language Models

DGX agent

arXiv:2604.16565v1 Announce Type: new Abstract: While Diffusion Large Language Models (dLLMs) offer structural advantages for global planning, efficiently verifying that they arrive at correct answers

safetyarxiv-cs-lg
21 Apr 2026
Safety

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction

DGX agent

arXiv:2506.01770v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved tremendous success in various tasks, yet concerns about their safety and security have emerged. In

safetyarxiv-cs-lg
21 Apr 2026
Safety

RemoteShield: Enable Robust Multimodal Large Language Models for Earth Observation

DGX agent

arXiv:2604.17243v1 Announce Type: new Abstract: A robust Multimodal Large Language Model (MLLM) for Earth Observation should maintain consistent interpretation and reasoning under realistic input vari

safetyarxiv-cs-cv
21 Apr 2026
Safety

Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning

DGX agent

arXiv:2601.14750v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting has achieved remarkable success in unlocking the reasoning capabilities of Large Language Models (LLMs). Although C

safetyarxiv-cs-cl
21 Apr 2026
Safety

Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contrastive Scoring

DGX agent

arXiv:2512.12069v3 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) are vulnerable to a growing array of multimodal jailbreak attacks, necessitating defenses that are both g

safetyarxiv-cs-cl
21 Apr 2026
Safety

Rethinking the Comparison Unit in Sequence-Level Reinforcement Learning: An Equal-Length Paired Training Framework from Loss Correction to Sample Construction

DGX agent

arXiv:2604.17328v1 Announce Type: new Abstract: This paper investigates the length problem in sequence-level relative reinforcement learning. We observe that, although existing methods partially allev

safetyarxiv-cs-lg
21 Apr 2026
Safety

Retrieval-Augmented Multimodal Model for Fake News Detection

DGX agent

arXiv:2604.18112v1 Announce Type: new Abstract: In recent years, multimodal multidomain fake news detection has garnered increasing attention. Nevertheless, this direction presents two significant cha

safetyarxiv-cs-cl
21 Apr 2026
Safety

Reverse Constitutional AI: A Framework for Controllable Toxic Data Generation via Probability-Clamped RLAIF

DGX agent

arXiv:2604.17769v1 Announce Type: new Abstract: Ensuring the safety of large language models (LLMs) requires robust red teaming, yet the systematic synthesis of high-quality toxic data remains under-e

safetyarxiv-cs-cl
21 Apr 2026
Safety

Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models

DGX agent

arXiv:2604.17415v1 Announce Type: cross Abstract: Reward-based fine-tuning aims to steer a pretrained diffusion or flow-based generative model toward higher-reward samples while remaining close to the

safetyarxiv-cs-cv
21 Apr 2026
Safety

REZE: Representation Regularization for Domain-adaptive Text Embedding Pre-finetuning

DGX agent

arXiv:2604.17257v1 Announce Type: new Abstract: Recent text embedding models are often adapted to specialized domains via contrastive pre-finetuning (PFT) on a naive collection of scattered, heterogen

safetyarxiv-cs-cl
21 Apr 2026
Safety

RISC-V Functional Safety for Autonomous Automotive Systems: An Analytical Framework and Research Roadmap for ML-Assisted Certification

DGX agent

arXiv:2604.17391v1 Announce Type: cross Abstract: RISC-V is emerging as a viable platform for automotive-grade embedded computing, with recent ISO 26262 ASIL-D certifications demonstrating readiness f

safetyarxiv-cs-lg
21 Apr 2026
Safety

Robust Tool Use via Fission-GRPO: Learning to Recover from Execution Errors

DGX agent

arXiv:2601.15625v2 Announce Type: replace Abstract: Large language models (LLMs) can call tools effectively, yet they remain brittle in multi-turn execution: after a tool-call error, smaller models of

safetyarxiv-cs-lg
21 Apr 2026
Safety

S-GRPO: Unified Post-Training for Large Vision-Language Models

DGX agent

arXiv:2604.16557v1 Announce Type: cross Abstract: Current post-training methodologies for adapting Large Vision-Language Models (LVLMs) generally fall into two paradigms: Supervised Fine-Tuning (SFT)

safetyarxiv-cs-cl
21 Apr 2026
Safety

SafeLM: Unified Privacy-Aware Optimization for Trustworthy Federated Large Language Models

DGX agent

arXiv:2604.16606v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in high-stakes domains, yet a unified treatment of their overlapping safety challenges remains

safetyarxiv-cs-lg
21 Apr 2026
Safety

SaFeR-Steer: Evolving Multi-Turn MLLMs via Synthetic Bootstrapping and Feedback Dynamics

DGX agent

arXiv:2604.16358v1 Announce Type: cross Abstract: MLLMs are increasingly deployed in multi-turn settings, where attackers can escalate unsafe intent through the evolving visual-text history and exploi

safetyarxiv-cs-cl
21 Apr 2026
Safety

Safer Trajectory Planning with CBF-guided Diffusion Model for Unmanned Aerial Vehicles

DGX agent

arXiv:2604.17527v1 Announce Type: new Abstract: Safe and agile trajectory planning is essential for autonomous systems, especially during complex aerobatic maneuvers. Motivated by the recent success o

safetyarxiv-cs-ro
21 Apr 2026
Safety

Safety, Security, and Cognitive Risks in State-Space Models: A Systematic Threat Analysis with Spectral, Stateful, and Capacity Attacks

DGX agent

arXiv:2604.16424v1 Announce Type: cross Abstract: State-Space Models (SSMs) -- structured SSMs (S4, S4D, DSS, S5), selective SSMs (Mamba, Mamba-2), and hybrid architectures (Jamba) -- are deployed in

safetyarxiv-cs-cl
21 Apr 2026
Safety

Scalable Neighborhood-Based Multi-Agent Actor-Critic

DGX agent

arXiv:2604.18190v1 Announce Type: new Abstract: We propose MADDPG-K, a scalable extension to Multi-Agent Deep Deterministic Policy Gradient (MADDPG) that addresses the computational limitations of cen

safetyarxiv-cs-lg
21 Apr 2026
Safety

Scalable Physics-Informed Neural Differential Equations and Data-Driven Algorithms for HVAC Systems

DGX agent

arXiv:2604.18438v1 Announce Type: new Abstract: We present a scalable, data-driven simulation framework for large-scale heating, ventilation, and air conditioning (HVAC) systems that couples physics-i

safetyarxiv-cs-lg
21 Apr 2026
Safety

See Through the Noise: Improving Domain Generalization in Gaze Estimation

DGX agent

arXiv:2604.16562v1 Announce Type: new Abstract: Generalizable gaze estimation methods have garnered increasing attention due to their critical importance in real-world applications and have achieved s

safetyarxiv-cs-cv
21 Apr 2026
Safety

SemLT3D: Semantic-Guided Expert Distillation for Camera-only Long-Tailed 3D Object Detection

DGX agent

arXiv:2604.18476v1 Announce Type: new Abstract: Camera-only 3D object detection has emerged as a cost-effective and scalable alternative to LiDAR for autonomous driving, yet existing methods primarily

safetyarxiv-cs-cv
21 Apr 2026
Safety

Sharpening Lightweight Models for Generalized Polyp Segmentation: A Boundary Guided Distillation from Foundation Models

DGX agent

arXiv:2604.17865v1 Announce Type: new Abstract: Automated polyp segmentation is critical for early colorectal cancer detection and its prevention, yet remains challenging due to weak boundaries, large

safetyarxiv-cs-cv
21 Apr 2026
Safety

Shepherding UAV Swarm with Action Prediction Based on Movement Constraints

DGX agent

arXiv:2604.17189v1 Announce Type: new Abstract: In this study, we propose a new sheepdog-inspired control method for a swarm of small unmanned aerial vehicles (UAVs), which predicts the swarm behavior

safetyarxiv-cs-ro
21 Apr 2026
Safety

Soft Label Pruning and Quantization for Large-Scale Dataset Distillation

DGX agent

arXiv:2604.18135v1 Announce Type: new Abstract: Large-scale dataset distillation requires storing auxiliary soft labels that can be 30-40x larger on ImageNet-1K and 200x larger on ImageNet-21K than th

safetyarxiv-cs-cv
21 Apr 2026
Safety

Source-Free Domain Adaptation with Vision-Language Prior

DGX agent

arXiv:2604.17748v1 Announce Type: new Abstract: Source-Free Domain Adaptation (SFDA) seeks to adapt a source model, which is pre-trained on a supervised source domain, for a target domain, with only a

safetyarxiv-cs-cv
21 Apr 2026
Safety

(Sparse) Attention to the Details: Preserving Spectral Fidelity in ML-based Weather Forecasting Models

DGX agent

arXiv:2604.16429v1 Announce Type: cross Abstract: We introduce Mosaic, a probabilistic weather forecasting model that addresses two principal sources of spectral degradation in ML-based weather predic

safetyarxiv-cs-cv
21 Apr 2026
Safety

Spectral bandits for smooth graph functions

DGX agent

arXiv:2604.18420v1 Announce Type: cross Abstract: Smooth functions on graphs have wide applications in manifold and semi-supervised learning. In this paper, we study a bandit problem where the payoffs

safetyarxiv-cs-lg
21 Apr 2026
Safety

Speculative Verification: Exploiting Information Gain to Refine Speculative Decoding

DGX agent

arXiv:2509.24328v2 Announce Type: replace Abstract: LLMs have low GPU efficiency and high latency due to autoregressive decoding. Speculative decoding (SD) mitigates this using a small draft model to

safetyarxiv-cs-cl
21 Apr 2026
Safety

SpiralThinker: Latent Reasoning through an Iterative Process with Text-Latent Interleaving

DGX agent

arXiv:2511.08983v2 Announce Type: replace Abstract: Recent advances in large reasoning models have been driven by reinforcement learning and test-time scaling, accompanied by growing interest in laten

safetyarxiv-cs-cl
21 Apr 2026
Safety

SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models

DGX agent

arXiv:2604.16995v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a promising paradigm for training reasoning-oriented models by leveraging rule-based reward signals. However,

safetyarxiv-cs-cl
21 Apr 2026
Safety

StealthGraph: Exposing Domain-Specific Risks in LLMs through Knowledge-Graph-Guided Harmful Prompt Generation

DGX agent

arXiv:2601.04740v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied in specialized domains such as finance and healthcare, where they introduce unique safety risk

safetyarxiv-cs-cl
21 Apr 2026
Safety

STL-Based Motion Planning and Uncertainty-Aware Risk Analysis for Human-Robot Collaboration with a Multi-Rotor Aerial Vehicle

DGX agent

arXiv:2509.10692v3 Announce Type: replace Abstract: This paper presents a motion planning and risk analysis framework for enhancing human-robot collaboration with a Multi-Rotor Aerial Vehicle. The pro

safetyarxiv-cs-ro
21 Apr 2026
Safety

Structure-Aware Diversity Pursuit as an AI Safety Strategy against Homogenization

DGX agent

arXiv:2601.06116v2 Announce Type: replace-cross Abstract: Generative AI models reproduce the biases in the training data and can further amplify them through mode collapse. We refer to the resulting h

safetyarxiv-cs-cl
21 Apr 2026
Safety

Sub-metre Lunar DEM Generation and Validation from Chandrayaan-2 OHRC Multi-View Imagery Using an Open-Source Pipeline

DGX agent

arXiv:2604.01032v3 Announce Type: replace Abstract: High-resolution digital elevation models (DEMs) of the lunar surface are essential for surface mobility planning, landing site characterization, and

safetyarxiv-cs-cv
21 Apr 2026
Safety

Support Sufficiency as Consequence-Sensitive Compression in Belief Arbitration

DGX agent

arXiv:2604.16434v1 Announce Type: cross Abstract: When a system commits to a hypothesis, much of the evidential structure behind that commitment is lost to compression. Standard accounts assume that s

safetyarxiv-cs-lg
21 Apr 2026
Safety

// Survey on Multi-Agent Systems // The paper traces the landscape from classical paradigms (consensus, distributed control, swarm intellige…

DGX agent

// Survey on Multi-Agent Systems // The paper traces the landscape from classical paradigms (consensus, distributed control, swarm intelligence, cooperative learning) to foundation-model-enabled MAS (

safetydair-ai--x
21 Apr 2026
Safety

SynAgent: Generalizable Cooperative Humanoid Manipulation via Solo-to-Cooperative Agent Synergy

DGX agent

arXiv:2604.18557v1 Announce Type: new Abstract: Controllable cooperative humanoid manipulation is a fundamental yet challenging problem for embodied intelligence, due to severe data scarcity, complexi

safetyarxiv-cs-cv
21 Apr 2026
Safety

SynopticBench: Evaluating Vision-Language Models on Generating Weather Forecast Discussions of the Future

DGX agent

arXiv:2604.16451v1 Announce Type: new Abstract: Recent advances in visual-language models (VLMs) have led to significant improvements in a plethora of complex multimodal tasks like image captioning, r

safetyarxiv-cs-cl
21 Apr 2026
← Previous
1…237238239240241…265
Next →