AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

The Labyrinth and the Thread: Rethinking Regularizations in Sequential Knowledge Editing for Large Language Models

DGX agent

arXiv:2605.26670v1 Announce Type: cross Abstract: Sequential editing of structured knowledge in large language models allows targeted factual updates without retraining, yet existing methods often rel

safetyarxiv-cs-ai
27 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

The Rescue Effect: Spatio-Semantic Early Exit Bypasses Quantization Collapse in CLIP

DGX agent

arXiv:2605.26415v1 Announce Type: cross Abstract: Deploying Vision-Language Models on resource-constrained hardware typically requires INT8 quantization, but in joint-embedding architectures such as C

safetyarxiv-cs-ai
27 May 2026
Safety

The Role of Causal Features in Strategic Classification for Robustness and Alignment

DGX agent

arXiv:2605.27163v1 Announce Type: new Abstract: In strategic classification, an institution (e.g., a bank) anticipates adaptation from users who change their features to increase utility in a classifi

safetyarxiv-cs-lg
27 May 2026
Safety

To model human linguistic prediction, make LLMs less superhuman

DGX agent

arXiv:2510.05141v2 Announce Type: replace Abstract: When we read, we make predictions about upcoming words; these predictions influence our reading behavior. The success of large language models (LLMs

safetyarxiv-cs-cl
27 May 2026
Safety

Triadic Dynamics Aware Diffusion Posterior Sampling for Inverse Problems: Optimizing Guidance and Stochasticity Schedules

DGX agent

arXiv:2605.26470v1 Announce Type: new Abstract: Generative posterior sampling using diffusion models has emerged as a dominant paradigm for solving inverse problems in imaging, which usually consists

safetyarxiv-cs-cv
27 May 2026
Safety

Turning Bias into Bugs: Bandit-Guided Style Manipulation Attacks on LLM Judges

DGX agent

arXiv:2605.26156v1 Announce Type: cross Abstract: The known stylistic biases in LLM judges, such as a preference for verbosity or specific sentence structures, present an underexplored security vulner

safetyarxiv-cs-ai
27 May 2026
Safety

UCPO: Uncertainty-Aware Policy Optimization

DGX agent

arXiv:2601.22648v2 Announce Type: replace Abstract: The key to building trustworthy large language models (LLMs) lies in endowing them with inherent uncertainty expression capabilities, thereby mitiga

safetyarxiv-cs-ai
27 May 2026
Safety

Uniboost: Global Coordination with Value Alignment for Fair and Efficient Traffic Allocation

DGX agent

arXiv:2605.26424v1 Announce Type: cross Abstract: With the rapid evolution of internet services, recommendation systems have become indispensable. In particular, the blending (re-ranking) stage plays

safetyarxiv-cs-ai
27 May 2026
Safety

Unique Lives, Shared World: Learning from Single-Life Videos

DGX agent

arXiv:2512.04085v2 Announce Type: replace Abstract: We introduce the 'single-life' learning paradigm, where we train a distinct vision model exclusively on egocentric videos captured by one individual

safetyarxiv-cs-cv
27 May 2026
Safety

V2V3D: View-to-View Denoised 3D Reconstruction for Light-Field Microscopy

DGX agent

arXiv:2504.07853v2 Announce Type: replace Abstract: Light field microscopy (LFM) has gained significant attention due to its ability to capture snapshot-based, large-scale 3D fluorescence images. Howe

safetyarxiv-cs-cv
27 May 2026
Safety

VERA-V: Variational Inference Framework for Jailbreaking Vision-Language Models

DGX agent

arXiv:2510.17759v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) extend large language models with visual reasoning, but their multimodal design also introduces new, underexplor

safetyarxiv-cs-cl
27 May 2026
Safety

VR-DAgger: Immersive VR for Dexterous Data Collection and Uncertainty-Guided On-Policy Correction

DGX agent

arXiv:2605.27114v1 Announce Type: new Abstract: Learning from demonstrations is effective for robotic manipulation, but collecting sufficient task-specific data remains a major bottleneck. Under distr

safetyarxiv-cs-ro
27 May 2026
Safety

When Does LeJEPA Learn a World Model?

DGX agent

arXiv:2605.26379v1 Announce Type: cross Abstract: A representation that scrambles the true degrees of freedom of the world cannot support reliable planning or compositional generalization. We prove th

safetyarxiv-cs-lg
27 May 2026
Safety

When Eyes Betray AI: Social Gaze Consistency as a Semantic Cue for AI-Generated Image Detection

DGX agent

arXiv:2605.27348v1 Announce Type: cross Abstract: Recent generative models have largely closed the gap on low-level artifacts - pixel fingerprints, frequency anomalies, upsampling traces - particularl

safetyarxiv-cs-ai
27 May 2026
Safety

Which Changes Matter? Towards Trustworthy Legal AI via Relevance-Sensitive Evaluation and Solver-Grounded Reasoning

DGX agent

arXiv:2605.26530v1 Announce Type: new Abstract: Legal reasoning requires distinguishing changes that matter from those that do not. Legal AI should remain stable under legally irrelevant perturbations

safetyarxiv-cs-ai
27 May 2026
Safety

A comparative study of accuracy and rollout stability of temporal surrogate models

DGX agent

arXiv:2605.24868v1 Announce Type: new Abstract: Temporal surrogate models are effective for predicting chaotic dynamical systems where computational cost can be prohibitive. Several deep neural networ

safetyarxiv-cs-lg
26 May 2026
Safety

A Decentralized LiDAR-SLAM System with Certifiably Optimal Pose Graph Optimization

DGX agent

arXiv:2605.25051v1 Announce Type: new Abstract: Decentralized multi-robot LiDAR-SLAM is essential for collaborative missions but faces significant challenges in maintaining global consistency. Existin

safetyarxiv-cs-ro
26 May 2026
Safety

A governance horizon for ethical-use constraints in open-weight AI models

DGX agent

arXiv:2605.24383v1 Announce Type: new Abstract: Ethical constraints on open-weight AI models are both a reflection of societal concerns and a foundation for AI governance policy. They are expected to

safetyarxiv-cs-ai
26 May 2026
Safety

A Multimodal 3D Foundation Model for Light Sheet Fluorescence Microscopy Enables Few-Shot Segmentation, Classification, and Deblurring

DGX agent

arXiv:2605.26026v1 Announce Type: cross Abstract: Light sheet fluorescence microscopy (LSM) enables high-resolution, three-dimensional (3D) imaging of biological specimens, providing rich volumetric d

safetyarxiv-cs-ai
26 May 2026
Safety

A Multimodal Framework for Dementia Detection via Linguistic and Acoustic Representation Learning

DGX agent

arXiv:2605.25540v1 Announce Type: cross Abstract: Alzheimer's disease (AD) is a progressive neurodegenerative disorder and the leading cause of dementia, affecting memory, reasoning, communication, an

safetyarxiv-cs-lg
26 May 2026
Safety

A Sober Look at Agentic Misalignment in Automated Workflows

DGX agent

arXiv:2605.24197v1 Announce Type: new Abstract: We study a class of emergent misalignment in multi-agent systems (MAS), with a focus on automated workflows, which we refer to agentic misalignment. Alt

safetyarxiv-cs-ai
26 May 2026
Safety

A Tertiary Review of Large Language Model-Based Code Generating Tasks: Trends, Challenges, and Future Directions

DGX agent

arXiv:2605.25536v1 Announce Type: cross Abstract: Context. Large language models (LLMs) are increasingly applied to code-generating tasks (CGTs) in software engineering. While reported results are pro

safetyarxiv-cs-ai
26 May 2026
Safety

A Unified Python Framework for Direct PPO-based Control of AHUs with Economizer Logic and CO2-Constrained Ventilation

DGX agent

arXiv:2605.24406v1 Announce Type: new Abstract: Optimizing HVAC (Heating, Ventilation and Air Conditioning) can enhance a building's energy efficiency while providing comfort levels for its occupants.

safetyarxiv-cs-lg
26 May 2026
Safety

Active Learning for Stochastic Contextual Linear Bandits

DGX agent

arXiv:2605.24803v1 Announce Type: new Abstract: A key goal in stochastic contextual linear bandits is to efficiently learn a near-optimal policy. Prior algorithms for this problem learn a policy by st

safetyarxiv-cs-lg
26 May 2026
Safety

Adaptive Human-AI Coordination via Hierarchical Action Disentanglement

DGX agent

arXiv:2605.24343v1 Announce Type: new Abstract: Human-AI collaboration requires agents that can adapt to diverse partner behaviors and skill levels while remaining robust to unseen partners. Existing

safetyarxiv-cs-ai
26 May 2026
Safety

Adaptive Preference Optimization with Uncertainty-aware Utility Anchor

DGX agent

arXiv:2509.10515v1 Announce Type: cross Abstract: Offline preference optimization methods are efficient for large language models (LLMs) alignment. Direct Preference optimization (DPO)-like learning,

safetyarxiv-cs-cl
26 May 2026
Safety

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models

DGX agent

arXiv:2605.26013v1 Announce Type: cross Abstract: We introduce AdvantageFlow, a forward-process reinforcement learning algorithm for rectified flow models. Unlike Flow-GRPO, which optimizes the revers

safetyarxiv-cs-ai
26 May 2026
Safety

Adversarial Error Correction for Visual Autoregressive Generation

DGX agent

arXiv:2605.24843v1 Announce Type: cross Abstract: Visual Autoregressive (VAR) models have emerged as a powerful paradigm for image synthesis by performing hierarchical next-scale prediction. However,

safetyarxiv-cs-ai
26 May 2026
Model Releases

AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue

DGX agent

arXiv:2605.23974v1 Announce Type: new Abstract: Current language models create two safety challenges: risk must be detected early enough to avoid exposing harmful continuation, and the harmfulness its

model-releasesarxiv-cs-cl
26 May 2026
Safety

Agent-Centric Social Trajectory Prediction: A Free Energy Principle Perspective

DGX agent

arXiv:2605.25748v1 Announce Type: new Abstract: Trajectory prediction methods have demonstrated remarkable capabilities in capturing complex motion patterns. However, existing methods rely on global s

safetyarxiv-cs-ai
26 May 2026
Safety

Agent-Facing Information Design in LLM Tool Registries

DGX agent

arXiv:2605.23916v1 Announce Type: cross Abstract: LLM tool registries function as unregulated advertising platforms: providers write free-text descriptions that agents use for selection, yet no measur

safetyarxiv-cs-ai
26 May 2026
Safety

Agent Learning via Early Experience

DGX agent

arXiv:2510.08558v3 Announce Type: replace Abstract: A long-term goal of language agents is to learn and improve through their own experience, ultimately outperforming humans in complex, real-world tas

safetyarxiv-cs-ai
26 May 2026
Safety

Agent-ToM: Learning to Monitor Autonomous LLM Agents via Theory-of-Mind Reasoning

DGX agent

arXiv:2605.24216v1 Announce Type: cross Abstract: Monitoring autonomous large language model (LLM) agents for covert malicious behavior is challenging due to delayed, context-dependent, and long-horiz

safetyarxiv-cs-ai
26 May 2026
Safety

AI-Assisted Systematization for Evaluating GenAI Systems

DGX agent

arXiv:2605.26001v1 Announce Type: cross Abstract: Evaluating generative AI (GenAI) systems is challenging because many targets of evaluation are broad, contested concepts, such as 'reasoning,' 'fairne

safetyarxiv-cs-ai
26 May 2026
Safety

AI-Driven Controlled Environment Agriculture as Resilient Infrastructure for U.S. Fresh-Produce Supply Chains

DGX agent

arXiv:2605.23946v1 Announce Type: cross Abstract: Climate volatility, regional production concentration, labor constraints, cyber risk, and dependence on long-distance fresh-produce supply chains expo

safetyarxiv-cs-ai
26 May 2026
Safety

Analogies between Transformer Layers and Power Method

DGX agent

arXiv:2605.25619v1 Announce Type: new Abstract: In the paper we show that there is an analogy between the operations occurring in a layer of a transformer (projections and layer normalizations, disreg

safetyarxiv-cs-lg
26 May 2026
Safety

Anatomy-Anchored Self-Supervision: Distilling Vision Foundation Models for Invariant Ultrasound Representation

DGX agent

arXiv:2605.25402v1 Announce Type: cross Abstract: Self-supervised pre-training paradigm has gained increasing prominence for learning transferable representations in medical imaging, yet existing meth

safetyarxiv-cs-ai
26 May 2026
Safety

Authority Inversion in LLM-Mediated Ubiquitous Systems: When Models Trust Users Over Sensors

DGX agent

arXiv:2605.23938v1 Announce Type: new Abstract: Large language models (LLMs) increasingly fuse heterogeneous inputs in ubiquitous systems. Yet, how LLMs implicitly allocate authority when sensor measu

safetyarxiv-cs-ai
26 May 2026
Safety

Autoregressive Language Models are Secretly Energy-Based Models: Insights into the Lookahead Capabilities of Next-Token Prediction

DGX agent

arXiv:2512.15605v4 Announce Type: replace Abstract: Autoregressive models (ARMs) currently constitute the dominant paradigm for large language models (LLMs). Energy-based models (EBMs) represent anoth

safetyarxiv-cs-lg
26 May 2026
Safety

AvAtar: Learning to Align via Active Optimal Transport

DGX agent

arXiv:2605.24395v1 Announce Type: new Abstract: Alignment plays a fundamental role in many machine learning problems, such as multi-network analysis, multimodal learning, and point cloud registration.

safetyarxiv-cs-lg
26 May 2026
Safety

Balancing Fairness, Privacy, and Accuracy: A Multitask Adversarial Framework for Centralized Data-Driven Systems

DGX agent

arXiv:2605.24458v1 Announce Type: cross Abstract: The integration of fairness and privacy in centralized data-driven applications is critical, especially as these systems increasingly influence sector

safetyarxiv-cs-ai
26 May 2026
Safety

Beyond Killer Robots: General AI Attitudes and Public Support for Military AI in Nine Countries

DGX agent

arXiv:2605.25196v1 Announce Type: cross Abstract: AI-enabled military systems are a fixture of modern military conflict. Applications vary from autonomous drones for surveillance and attack to AI-supp

safetyarxiv-cs-ai
26 May 2026
Safety

Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training

DGX agent

arXiv:2505.20110v3 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) excel at sampling diverse, high-reward objects. In many practical applications where active reward querie

safetyarxiv-cs-ai
26 May 2026
Safety

Beyond the Target: From Imitation to Collaboration in Speculative Decoding

DGX agent

arXiv:2605.24793v1 Announce Type: new Abstract: Speculative decoding (SPD) accelerates large language model (LLM) inference by letting a smaller draft model propose multiple future tokens that are ver

safetyarxiv-cs-cl
26 May 2026
Safety

Blocked Gibbs meets Diffusion Transformers: Unsupervised Learning for Constraint Optimization

DGX agent

arXiv:2605.25129v1 Announce Type: new Abstract: Diffusion models have shown promise in learning to solve constraint optimization problems. However, they are mostly restricted to problems with binary v

safetyarxiv-cs-lg
26 May 2026
Safety

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion

DGX agent

arXiv:2605.24975v1 Announce Type: cross Abstract: Proximal Policy Optimization (PPO) has become the de facto standard for training legged robots, thanks to its robustness and scalability in massively

safetyarxiv-cs-ai
26 May 2026
Safety

Can Large Language Models Resolve Semantic Discrepancy in Self-Destructive Subcultures? Evidence from Jirai Kei

DGX agent

arXiv:2601.05004v2 Announce Type: replace Abstract: Self-destructive behaviors are linked to complex psychological states and can be challenging to diagnose. These behaviors may be even harder to iden

safetyarxiv-cs-cl
26 May 2026
Safety

Causal methods for LLM development and evaluation

DGX agent

arXiv:2605.25998v1 Announce Type: new Abstract: Large language model (LLM) development is currently driven by large-scale empirical iteration over data mixtures, reward models, routing strategies, and

safetyarxiv-cs-lg
26 May 2026
← Previous
1…158159160161162…260
Next →