AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Dual Mechanisms of Value Expression: Intrinsic vs. Prompted Values in Large Language Models

DGX agent

arXiv:2509.24319v4 Announce Type: replace-cross Abstract: Large language models can express values in two main ways: (1) intrinsic expression, reflecting the model's inherent values learned during tra

safetyarxiv-cs-ai
1 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

EchoRL: Reinforcement Learning via Rollout Echoing

DGX agent

arXiv:2605.31228v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards is an effective route for post-training to strengthen the reasoning capability of large language models

safetyarxiv-cs-ai
1 Jun 2026
Safety

Efficient and Uncertainty-Aware Diffusion Framework for Offline-to-Online Reinforcement Learning

DGX agent

arXiv:2605.30776v1 Announce Type: new Abstract: Offline-to-Online Reinforcement Learning (O2O-RL) leverages an offline, pre-trained policy to minimize costly online interactions. Although data-efficie

safetyarxiv-cs-lg
1 Jun 2026
Safety

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation

DGX agent

arXiv:2605.30484v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown promise for robotic manipulation, yet most existing policies operate reactively by directly regressing ac

safetyarxiv-cs-ro
1 Jun 2026
Safety

Enhancing Regime Shift Detection Using Unstructured Data: A Study on the Treasury Market

DGX agent

arXiv:2605.30363v1 Announce Type: cross Abstract: Regime shifts in financial markets reorganise the joint dynamics of asset prices and macro variables, breaking any single-regime calibration. They are

safetyarxiv-cs-ai
1 Jun 2026
Safety

Entropic Projection Alignment: Estimating, Explaining, and Improving Model Performance Under Distribution Shift

DGX agent

arXiv:2605.31250v1 Announce Type: cross Abstract: We propose a unified framework for addressing three key challenges of distribution shift: (1) estimating a model's performance on an unlabeled target

safetyarxiv-cs-ai
1 Jun 2026
Safety

Envisioning Beyond the Few: Disentangled Semantics and Primitives for Few-Shot Atypical Layout-to-Image Generation

DGX agent

arXiv:2605.31266v1 Announce Type: cross Abstract: The layout-to-image (L2I) task enables fine-grained control over image generation via object categories and spatial layouts. However, existing L2I met

safetyarxiv-cs-ai
1 Jun 2026
Safety

Equivariant Latent Alignment via Flow Matching under Group Symmetries

DGX agent

arXiv:2605.30705v1 Announce Type: new Abstract: Geometry-aware generative models and novel view synthesis approaches have shown strong potential in visual fidelity and consistency. In parallel, equiva

safetyarxiv-cs-cv
1 Jun 2026
Safety

Extending the UXR Point of View Pyramid: A Generative AI-Augmented Methodology for Human-Centred AI Systems

DGX agent

arXiv:2605.31143v1 Announce Type: cross Abstract: Rising household debt and cost-of-living pressures in the United Kingdom have intensified the role of AI-driven financial technologies in mediating cr

safetyarxiv-cs-ai
1 Jun 2026
Safety

Fair Decisions from Calibrated Scores: Achieving Optimal Classification While Satisfying Sufficiency

DGX agent

arXiv:2602.07285v2 Announce Type: replace Abstract: Binary classification based on predicted probabilities (scores) is a fundamental task in supervised machine learning. While thresholding scores is B

safetyarxiv-cs-lg
1 Jun 2026
Safety

Feat2Go: Visual Feature-Grounded Value Estimation for Embodied Reinforcement Learning

DGX agent

arXiv:2605.30795v1 Announce Type: new Abstract: Reinforcement learning is a promising approach for improving the capabilities of vision-language-action (VLA) models while avoiding the heavy data requi

safetyarxiv-cs-ro
1 Jun 2026
Safety

Feature-Optimized Vision for Adaptive 3D Scene Reconstruction

DGX agent

arXiv:2605.31534v1 Announce Type: cross Abstract: Three-dimensional scene reconstruction depends on local image evidence that is both visually discriminative and geometrically useful. Fixed feature th

safetyarxiv-cs-ai
1 Jun 2026
Safety

FLAG: Flow Policy MaxEnt-RL by Latent Augmented Guidance

DGX agent

arXiv:2605.30749v1 Announce Type: new Abstract: Maximum entropy reinforcement learning (MaxEnt-RL) enables robust exploration, yet practical implementations often restrict policies to simple Gaussians

safetyarxiv-cs-lg
1 Jun 2026
Safety

Forecasting with Hyper-Trees

DGX agent

arXiv:2405.07836v5 Announce Type: replace Abstract: We introduce Hyper-Trees as a novel framework for modeling time series data using gradient boosted trees. Unlike conventional tree-based approaches

safetyarxiv-cs-lg
1 Jun 2026
Safety

GlucoFM: A Dual-Stream Foundation Model for Continuous Glucose Monitoring

DGX agent

arXiv:2605.30865v1 Announce Type: new Abstract: Continuous glucose monitoring (CGM) provides a dense view of daily metabolic physiology, yet existing generic time-series and CGM-specific foundation mo

safetyarxiv-cs-lg
1 Jun 2026
Safety

Hamiltonian-Inspired Attention Mechanism for Scalable RF Transmitter Fingerprinting

DGX agent

arXiv:2605.30364v1 Announce Type: cross Abstract: Radio-frequency (RF) fingerprinting identifies wire-less transmitters using hardware-induced imperfections present in baseband I/Q signals. However, d

safetyarxiv-cs-ai
1 Jun 2026
Safety

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model

DGX agent

arXiv:2605.31234v1 Announce Type: new Abstract: Learning generalizable vision-language-action (VLA) models from large-scale human videos is promising but challenging due to cross-embodiment discrepanc

safetyarxiv-cs-ro
1 Jun 2026
Safety

Healthcare Mechanisms from Policy-as-Code Search under Strategic Provider Response

DGX agent

arXiv:2605.30680v1 Announce Type: new Abstract: Healthcare mechanisms are inseparable from the strategic provider response they induce: existing healthcare AI benchmarks hold this response fixed and s

safetyarxiv-cs-ai
1 Jun 2026
Safety

HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents

DGX agent

arXiv:2602.16165v2 Announce Type: replace-cross Abstract: Training LLMs as interactive agents for multi-turn decision-making remains challenging, particularly in long-horizon tasks with sparse and del

safetyarxiv-cs-ai
1 Jun 2026
Safety

HQ-JEPA: Hybrid Quantum Joint-Embedding Predictive Architecture for Cross-Modal Remote Sensing Representation Learning

DGX agent

arXiv:2605.31068v1 Announce Type: new Abstract: We introduce HQ-JEPA, a hybrid quantum-classical joint-embedding predictive architecture for cross-modal remote sensing representation learning. The pro

safetyarxiv-cs-cv
1 Jun 2026
Safety

Human-Alignment and Calibration of Inference-Time Uncertainty in Large Language Models

DGX agent

arXiv:2508.08204v2 Announce Type: replace-cross Abstract: There has been much recent interest in evaluating large language models for uncertainty calibration to facilitate model control and modulate u

safetyarxiv-cs-ai
1 Jun 2026
Safety

Human-Alignment, Calibration, and Activation Patterns in Large Language Model Uncertainty

DGX agent

arXiv:2605.30675v1 Announce Type: cross Abstract: Uncertainty Quantification is a large and growing subfield of large language model behavioral analysis. Primarily to recognize and combat hallucinatio

safetyarxiv-cs-ai
1 Jun 2026
Safety

Human Psychometric Questionnaires Mischaracterize LLM Behavior

DGX agent

arXiv:2509.10078v4 Announce Type: replace-cross Abstract: We examine whether human psychometric questionnaires can serve as reliable tools for characterizing and predicting LLM behavior in everyday us

safetyarxiv-cs-ai
1 Jun 2026
Safety

IAPO: Information-Aware Policy Optimization for Token-Efficient Reasoning

DGX agent

arXiv:2602.19049v2 Announce Type: replace Abstract: Large language models increasingly rely on long chains of thought to improve accuracy, yet such gains come with substantial inference-time costs. We

safetyarxiv-cs-cl
1 Jun 2026
Safety

ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and Domain-Specific Representation Alignment

DGX agent

arXiv:2605.30965v1 Announce Type: cross Abstract: Recent advancements in text-guided audio generation have yielded promising results in diverse domains, including sound effects, speech, and music. How

safetyarxiv-cs-ai
1 Jun 2026
Safety

IsoCLIP: Decomposing CLIP Projectors for Efficient Intra-modal Alignment

DGX agent

arXiv:2603.19862v2 Announce Type: replace Abstract: Vision-Language Models like CLIP are extensively used for inter-modal tasks which involve both visual and text modalities. However, when the individ

safetyarxiv-cs-cv
1 Jun 2026
Safety

LARK: Learnability-Grounded Trajectory Selection for Efficient Reasoning Distillation

DGX agent

arXiv:2605.30651v1 Announce Type: cross Abstract: We study trajectory selection for reasoning distillation, where teacher-generated reasoning trajectories are selectively used as supervision for a stu

safetyarxiv-cs-ai
1 Jun 2026
Safety

Learning Controlled Separation of Small Objects Between Two Fingers with a Tactile Skin

DGX agent

arXiv:2605.31486v1 Announce Type: new Abstract: We introduce and solve the novel task of controlled separation of small objects with two fingers of a multi-purpose robotic hand: after grasping into a

safetyarxiv-cs-ro
1 Jun 2026
Safety

Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt

DGX agent

arXiv:2505.20795v2 Announce Type: replace Abstract: Recent robot learning methods commonly rely on imitation learning from massive robotic dataset collected with teleoperation. When facing a new task,

safetyarxiv-cs-ro
1 Jun 2026
Safety

Learning Terrain-Aware Whole-Body Control for Perceptive Legged Loco-Manipulation

DGX agent

arXiv:2605.31343v1 Announce Type: new Abstract: Legged manipulators integrate exceptional terrain adaptability along with mobile manipulation capabilities, which make them highly promising for deploym

safetyarxiv-cs-ro
1 Jun 2026
Safety

Lightweight CNN-Based Anomaly Detection for High Voltage Converter Modulators in the Spallation Neutron Source

DGX agent

arXiv:2605.31259v1 Announce Type: new Abstract: Unscheduled trips of high-power pulsed converters are a leading source of downtime at large accelerator facilities. At the Spallation Neutron Source (SN

safetyarxiv-cs-lg
1 Jun 2026
Safety

LVSA: Training-Free Sparse Attention for Long Video Diffusion

DGX agent

arXiv:2605.31057v1 Announce Type: new Abstract: Dense self-attention is the compute and quality bottleneck of long-video diffusion inference: cost grows quadratically with the sequence length, and bey

safetyarxiv-cs-cv
1 Jun 2026
Safety

Mental Damage: Caption Poisoning Attacks on Retrieval-Augmented Text-to-Music Generation

DGX agent

arXiv:2605.30365v1 Announce Type: cross Abstract: Retrieval-augmented text-to-music (TTM) systems augment underspecified user prompts using captions retrieved from a music caption dataset. This design

safetyarxiv-cs-ai
1 Jun 2026
Safety

MergeTok: Unified Continuous and Discrete Visual Tokenization via Token Merging

DGX agent

arXiv:2605.30904v1 Announce Type: new Abstract: Most visual tokenizers for image generation are bifurcated into two families with complementary limitations: continuous VAEs offer high-fidelity reconst

safetyarxiv-cs-cv
1 Jun 2026
Safety

Multi-Agent Teams Hold Experts Back

DGX agent

arXiv:2602.01011v4 Announce Type: replace-cross Abstract: Multi-agent LLM systems are increasingly deployed as autonomous collaborators, where agents interact freely rather than execute fixed, pre-spe

safetyarxiv-cs-ai
1 Jun 2026
Safety

Neuron-Level Interventions for Gendered and Gender-Neutral Generation in Language Models

DGX agent

arXiv:2605.30717v1 Announce Type: new Abstract: Language models (LMs) can produce gendered language and stereotypes even when given neutral prompts. Most prior work on gender bias in LMs primarily exa

safetyarxiv-cs-cl
1 Jun 2026
Safety

OmniMem: Scalable and Adaptive Memory Retrieval for Long Video Generation

DGX agent

arXiv:2605.30519v1 Announce Type: new Abstract: Autoregressive (AR) video generation extends videos by producing latent chunks sequentially, but scaling to long videos requires repeated access to a gr

safetyarxiv-cs-cv
1 Jun 2026
Safety

On the Illusion of Gender Bias in Face Recognition: Explaining the Fairness Issue Through Non-demographic Attributes

DGX agent

arXiv:2501.12020v2 Announce Type: replace Abstract: Face recognition systems (FRS) exhibit significant accuracy differences based on the user's gender. Since such a gender gap reduces the trustworthin

safetyarxiv-cs-cv
1 Jun 2026
Safety

On the 'Induction Bias' in Sequence Models

DGX agent

arXiv:2602.18333v2 Announce Type: replace-cross Abstract: Despite the remarkable practical success of transformer-based language models, recent work has raised concerns about their ability to perform

safetyarxiv-cs-cl
1 Jun 2026
Safety

On the Relationship Between Activation Outliers and Feature Death in Sparse Autoencoders

DGX agent

arXiv:2605.31518v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) decompose neural network activations into interpretable features, but many learned features never activate, a problem called

safetyarxiv-cs-lg
1 Jun 2026
Safety

Optimizing Rank for High-Fidelity Implicit Neural Representations

DGX agent

arXiv:2512.14366v2 Announce Type: replace Abstract: Implicit Neural Representations (INRs) based on vanilla Multi-Layer Perceptrons (MLPs) are widely believed to be incapable of representing high-freq

safetyarxiv-cs-cv
1 Jun 2026
Safety

Organizational Adaptation to Generative AI in Cybersecurity

DGX agent

arXiv:2506.12060v2 Announce Type: replace-cross Abstract: Cybersecurity organizations are adapting to GenAI integration through modified frameworks and hybrid operational processes, with success influ

safetyarxiv-cs-ai
1 Jun 2026
Safety

PAC-Bayesian Reinforcement Learning Trains Generalizable Policies

DGX agent

arXiv:2510.10544v3 Announce Type: replace-cross Abstract: We derive a novel PAC-Bayesian generalization bound for reinforcement learning that explicitly accounts for Markov dependencies in the data, t

safetyarxiv-cs-ai
1 Jun 2026
Safety

Parallel Tempering Initial Sampling in Inference-Time Reward Alignment

DGX agent

arXiv:2605.30991v1 Announce Type: cross Abstract: Inference-time reward alignment steers pretrained diffusion and flow-based generative models to satisfy user-specified rewards without retraining. Rec

safetyarxiv-cs-cv
1 Jun 2026
Safety

PASTA: A Scalable Framework for Multi-Policy AI Compliance Evaluation

DGX agent

arXiv:2601.11702v3 Announce Type: replace-cross Abstract: AI compliance is becoming increasingly critical as AI systems grow more powerful and pervasive. Yet the rapid expansion of AI policies creates

safetyarxiv-cs-ai
1 Jun 2026
Safety

*-PLUIE: Personalisable metric with Llm Used for Improved Evaluation

DGX agent

arXiv:2602.15778v2 Announce Type: replace Abstract: Evaluating the quality of automatically generated text often relies on LLM-as-a-judge (LLM-judge) methods. While effective, these approaches are com

safetyarxiv-cs-cl
1 Jun 2026
Safety

Population-Free Pareto Tracking for Sample-Efficient Multi-Policy MORL

DGX agent

arXiv:2508.02217v2 Announce Type: replace Abstract: Multi-objective reinforcement learning (MORL) is a fundamental framework for real-world decision-making problems involving multiple conflicting crit

safetyarxiv-cs-lg
1 Jun 2026
Safety

PostCam: Camera-Controllable Novel-View Video Generation with Query-Shared Cross-Attention

DGX agent

arXiv:2511.17185v2 Announce Type: replace Abstract: We propose PostCam, a streamlined framework for novel-view video generation that achieves superior detail preservation and precise camera trajectory

safetyarxiv-cs-cv
1 Jun 2026
← Previous
1…147148149150151…260
Next →