AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,814 results
1 Jun 2026

Feat2Go: Visual Feature-Grounded Value Estimation for Embodied Reinforcement Learning

SafetyDGX agent

arXiv:2605.30795v1 Announce Type: new Abstract: Reinforcement learning is a promising approach for improving the capabilities of vision-language-action (VLA) models while avoiding the heavy data requi

Feature-Optimized Vision for Adaptive 3D Scene Reconstruction

SafetyDGX agent

arXiv:2605.31534v1 Announce Type: cross Abstract: Three-dimensional scene reconstruction depends on local image evidence that is both visually discriminative and geometrically useful. Fixed feature th

FLAG: Flow Policy MaxEnt-RL by Latent Augmented Guidance

SafetyDGX agent

arXiv:2605.30749v1 Announce Type: new Abstract: Maximum entropy reinforcement learning (MaxEnt-RL) enables robust exploration, yet practical implementations often restrict policies to simple Gaussians


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Florida AG James Uthmeier sues OpenAI and Sam Altman, seeking to hold Altman personally liable for deceptive trade practices, negligence, and public nuisance (NBC News)

SafetyDGX agent

NBC News: Florida AG James Uthmeier sues OpenAI and Sam Altman, seeking to hold Altman personally liable for deceptive trade practices, negligence, and public nuisance — Florida's seeks to hold Sam Al

Forecasting with Hyper-Trees

SafetyDGX agent

arXiv:2405.07836v5 Announce Type: replace Abstract: We introduce Hyper-Trees as a novel framework for modeling time series data using gradient boosted trees. Unlike conventional tree-based approaches

From Evidence to Design: Developing an AI-Augmented UX Research Point of View for Digital Wellbeing in Emergency and Public Safety Contexts

SafetyDGX agent

arXiv:2605.31146v1 Announce Type: cross Abstract: This paper investigates how User Experience Research (UXR) methods can be combined with AI-supported analysis to develop clearer design direction for

From Internal Diagnosis to External Auditing: A VLM-Driven Paradigm for Data-Free Online Backdoor Defense

SafetyDGX agent

arXiv:2601.19448v2 Announce Type: replace Abstract: Deep Neural Networks remain inherently vulnerable to backdoor attacks. Traditional test-time defenses largely operate under the paradigm of internal

From Out-of-Distribution Detection to Hallucination Detection: A Geometric View

SafetyDGX agent

arXiv:2602.07253v2 Announce Type: replace Abstract: Detecting hallucinations in large language models is a critical open problem with significant implications for safety and reliability. While existin

Geometry-Aware Control Barrier Functions for Collision Avoidance via Bernstein Polynomial Approximations

SafetyDGX agent

arXiv:2605.30696v1 Announce Type: new Abstract: Safe navigation often relies on well-defined conditions based on the shape of robots and obstacles, and can be challenging when they have irregular geom

GlucoFM: A Dual-Stream Foundation Model for Continuous Glucose Monitoring

SafetyDGX agent

arXiv:2605.30865v1 Announce Type: new Abstract: Continuous glucose monitoring (CGM) provides a dense view of daily metabolic physiology, yet existing generic time-series and CGM-specific foundation mo

GSAM: A Generalizable and Safe Robotic Framework for Articulated Object Manipulation

SafetyDGX agent

arXiv:2605.30740v1 Announce Type: cross Abstract: Articulated object manipulation is a unique challenge for service robots. Existing methods employ end-to-end policy learning, visionmotion planning, a

Hamiltonian-Inspired Attention Mechanism for Scalable RF Transmitter Fingerprinting

SafetyDGX agent

arXiv:2605.30364v1 Announce Type: cross Abstract: Radio-frequency (RF) fingerprinting identifies wire-less transmitters using hardware-induced imperfections present in baseband I/Q signals. However, d

HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model

SafetyDGX agent

arXiv:2605.31234v1 Announce Type: new Abstract: Learning generalizable vision-language-action (VLA) models from large-scale human videos is promising but challenging due to cross-embodiment discrepanc

Healthcare Mechanisms from Policy-as-Code Search under Strategic Provider Response

SafetyDGX agent

arXiv:2605.30680v1 Announce Type: new Abstract: Healthcare mechanisms are inseparable from the strategic provider response they induce: existing healthcare AI benchmarks hold this response fixed and s

HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents

SafetyDGX agent

arXiv:2602.16165v2 Announce Type: replace-cross Abstract: Training LLMs as interactive agents for multi-turn decision-making remains challenging, particularly in long-horizon tasks with sparse and del

HQ-JEPA: Hybrid Quantum Joint-Embedding Predictive Architecture for Cross-Modal Remote Sensing Representation Learning

SafetyDGX agent

arXiv:2605.31068v1 Announce Type: new Abstract: We introduce HQ-JEPA, a hybrid quantum-classical joint-embedding predictive architecture for cross-modal remote sensing representation learning. The pro

Human-Alignment and Calibration of Inference-Time Uncertainty in Large Language Models

SafetyDGX agent

arXiv:2508.08204v2 Announce Type: replace-cross Abstract: There has been much recent interest in evaluating large language models for uncertainty calibration to facilitate model control and modulate u

Human-Alignment, Calibration, and Activation Patterns in Large Language Model Uncertainty

SafetyDGX agent

arXiv:2605.30675v1 Announce Type: cross Abstract: Uncertainty Quantification is a large and growing subfield of large language model behavioral analysis. Primarily to recognize and combat hallucinatio

Human Psychometric Questionnaires Mischaracterize LLM Behavior

SafetyDGX agent

arXiv:2509.10078v4 Announce Type: replace-cross Abstract: We examine whether human psychometric questionnaires can serve as reliable tools for characterizing and predicting LLM behavior in everyday us

Hybrid Energy-Aware Reward Shaping: A Unified Lightweight Physics-Guided Methodology for Policy Optimization

SafetyDGX agent

arXiv:2603.11600v2 Announce Type: replace Abstract: Deep reinforcement learning for continuous control often suffers from high variance, low energy efficiency, and poor generalization under distributi

I would say that it is absurd to say that “AI belongs to the people”, when so many companies have worked so hard to develop it, but it is al…

SafetyDGX agent

I would say that it is absurd to say that “AI belongs to the people”, when so many companies have worked so hard to develop it, but it is also absurd that Generative AI has basically been built on the

IAPO: Information-Aware Policy Optimization for Token-Efficient Reasoning

SafetyDGX agent

arXiv:2602.19049v2 Announce Type: replace Abstract: Large language models increasingly rely on long chains of thought to improve accuracy, yet such gains come with substantial inference-time costs. We

If “Insanity is doing the same thing over and over again and expecting different results”, what the heck is Generative AI?

SafetyDGX agent

Gary Marcus critiques generative AI by applying the 'insanity' aphorism to highlight repetitive patterns in how these systems operate—suggesting they may be fundamentally limited by repeating the same

ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and Domain-Specific Representation Alignment

SafetyDGX agent

arXiv:2605.30965v1 Announce Type: cross Abstract: Recent advancements in text-guided audio generation have yielded promising results in diverse domains, including sound effects, speech, and music. How

Import AI 459: AI oversight is difficult; scaling laws for protein folding models; and pricing the extinction risk of AI systems

SafetyDGX agent

This newsletter issue discusses three key topics in AI development and safety: the challenges involved in overseeing and controlling advanced AI systems, empirical findings about how protein folding A

IsoCLIP: Decomposing CLIP Projectors for Efficient Intra-modal Alignment

SafetyDGX agent

arXiv:2603.19862v2 Announce Type: replace Abstract: Vision-Language Models like CLIP are extensively used for inter-modal tasks which involve both visual and text modalities. However, when the individ

Kiss your pensions goodbye, folks May 20 - SpaceX's (SPCX) S-1 filing Ludicrous targeted valuation: 1.8T despite 4.28B loss over last year…

SafetyDGX agent

Kiss your pensions goodbye, folks May 20 - SpaceX's (SPCX) S-1 filing Ludicrous targeted valuation: 1.8T despite 4.28B loss over last year June 11 - offering price set June 12 - first day of trading /

LARK: Learnability-Grounded Trajectory Selection for Efficient Reasoning Distillation

SafetyDGX agent

arXiv:2605.30651v1 Announce Type: cross Abstract: We study trajectory selection for reasoning distillation, where teacher-generated reasoning trajectories are selectively used as supervision for a stu

Learning Controlled Separation of Small Objects Between Two Fingers with a Tactile Skin

SafetyDGX agent

arXiv:2605.31486v1 Announce Type: new Abstract: We introduce and solve the novel task of controlled separation of small objects with two fingers of a multi-purpose robotic hand: after grasping into a

Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt

SafetyDGX agent

arXiv:2505.20795v2 Announce Type: replace Abstract: Recent robot learning methods commonly rely on imitation learning from massive robotic dataset collected with teleoperation. When facing a new task,

Learning Terrain-Aware Whole-Body Control for Perceptive Legged Loco-Manipulation

SafetyDGX agent

arXiv:2605.31343v1 Announce Type: new Abstract: Legged manipulators integrate exceptional terrain adaptability along with mobile manipulation capabilities, which make them highly promising for deploym

LiftNav: Path Planning via Semantic Lifting in TSDF-Guided Gaussian Splatting

SafetyDGX agent

arXiv:2605.31376v1 Announce Type: cross Abstract: Autonomous robots in unknown indoor environments require both reliable collision avoidance and object-level understanding. Classical representations s

Lightweight CNN-Based Anomaly Detection for High Voltage Converter Modulators in the Spallation Neutron Source

SafetyDGX agent

arXiv:2605.31259v1 Announce Type: new Abstract: Unscheduled trips of high-power pulsed converters are a leading source of downtime at large accelerator facilities. At the Spallation Neutron Source (SN

LLM Judges Inconsistently Disagree Across Safety Criteria and Harm Categories

SafetyDGX agent

arXiv:2605.31381v1 Announce Type: new Abstract: We evaluate the consistency of automated judges in conducting a multi-dimensional safety evaluation in a reference-free setup. Our results indicate that

LVSA: Training-Free Sparse Attention for Long Video Diffusion

SafetyDGX agent

arXiv:2605.31057v1 Announce Type: new Abstract: Dense self-attention is the compute and quality bottleneck of long-video diffusion inference: cost grows quadratically with the sequence length, and bey

Mental Damage: Caption Poisoning Attacks on Retrieval-Augmented Text-to-Music Generation

SafetyDGX agent

arXiv:2605.30365v1 Announce Type: cross Abstract: Retrieval-augmented text-to-music (TTM) systems augment underspecified user prompts using captions retrieved from a music caption dataset. This design

MergeTok: Unified Continuous and Discrete Visual Tokenization via Token Merging

SafetyDGX agent

arXiv:2605.30904v1 Announce Type: new Abstract: Most visual tokenizers for image generation are bifurcated into two families with complementary limitations: continuous VAEs offer high-fidelity reconst

Modeling a digital twin of a food supply chain using BigQuery Graph

SafetyDGX agent

The example of a growing restaurant Imagine you are running a restaurant chain. You just can't physically feel and touch things to know how your business operates. You need tools and a digital replica

Multi-Agent Teams Hold Experts Back

SafetyDGX agent

arXiv:2602.01011v4 Announce Type: replace-cross Abstract: Multi-agent LLM systems are increasingly deployed as autonomous collaborators, where agents interact freely rather than execute fixed, pre-spe

Neuron-Level Interventions for Gendered and Gender-Neutral Generation in Language Models

SafetyDGX agent

arXiv:2605.30717v1 Announce Type: new Abstract: Language models (LMs) can produce gendered language and stereotypes even when given neutral prompts. Most prior work on gender bias in LMs primarily exa

Neurosymbolic rising!

SafetyDGX agent

Neurosymbolic rising! I've wanted to work on deep neurosymbolic integration for a while: if you make an 800k transformer reason *like* a logical solver, you get 100% on sudoku-extreme w. 15m of train

OmniMem: Scalable and Adaptive Memory Retrieval for Long Video Generation

SafetyDGX agent

arXiv:2605.30519v1 Announce Type: new Abstract: Autoregressive (AR) video generation extends videos by producing latent chunks sequentially, but scaling to long videos requires repeated access to a gr

On the Illusion of Gender Bias in Face Recognition: Explaining the Fairness Issue Through Non-demographic Attributes

SafetyDGX agent

arXiv:2501.12020v2 Announce Type: replace Abstract: Face recognition systems (FRS) exhibit significant accuracy differences based on the user's gender. Since such a gender gap reduces the trustworthin

On the 'Induction Bias' in Sequence Models

SafetyDGX agent

arXiv:2602.18333v2 Announce Type: replace-cross Abstract: Despite the remarkable practical success of transformer-based language models, recent work has raised concerns about their ability to perform

On the Relationship Between Activation Outliers and Feature Death in Sparse Autoencoders

SafetyDGX agent

arXiv:2605.31518v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) decompose neural network activations into interpretable features, but many learned features never activate, a problem called

Optimizing Rank for High-Fidelity Implicit Neural Representations

SafetyDGX agent

arXiv:2512.14366v2 Announce Type: replace Abstract: Implicit Neural Representations (INRs) based on vanilla Multi-Layer Perceptrons (MLPs) are widely believed to be incapable of representing high-freq

Organizational Adaptation to Generative AI in Cybersecurity

SafetyDGX agent

arXiv:2506.12060v2 Announce Type: replace-cross Abstract: Cybersecurity organizations are adapting to GenAI integration through modified frameworks and hybrid operational processes, with success influ

Our views on AI policy and political advocacy

SafetyDGX agent

OpenAI outlines its positions on AI policy priorities and explains its approach to political engagement and advocacy efforts. The document likely covers key regulatory areas OpenAI supports, such as s

PAC-Bayesian Reinforcement Learning Trains Generalizable Policies

SafetyDGX agent

arXiv:2510.10544v3 Announce Type: replace-cross Abstract: We derive a novel PAC-Bayesian generalization bound for reinforcement learning that explicitly accounts for Markov dependencies in the data, t

Parallel Tempering Initial Sampling in Inference-Time Reward Alignment

SafetyDGX agent

arXiv:2605.30991v1 Announce Type: cross Abstract: Inference-time reward alignment steers pretrained diffusion and flow-based generative models to satisfy user-specified rewards without retraining. Rec

PASTA: A Scalable Framework for Multi-Policy AI Compliance Evaluation

SafetyDGX agent

arXiv:2601.11702v3 Announce Type: replace-cross Abstract: AI compliance is becoming increasingly critical as AI systems grow more powerful and pervasive. Yet the rapid expansion of AI policies creates

*-PLUIE: Personalisable metric with Llm Used for Improved Evaluation

SafetyDGX agent

arXiv:2602.15778v2 Announce Type: replace Abstract: Evaluating the quality of automatically generated text often relies on LLM-as-a-judge (LLM-judge) methods. While effective, these approaches are com

Population-Free Pareto Tracking for Sample-Efficient Multi-Policy MORL

SafetyDGX agent

arXiv:2508.02217v2 Announce Type: replace Abstract: Multi-objective reinforcement learning (MORL) is a fundamental framework for real-world decision-making problems involving multiple conflicting crit

PostCam: Camera-Controllable Novel-View Video Generation with Query-Shared Cross-Attention

SafetyDGX agent

arXiv:2511.17185v2 Announce Type: replace Abstract: We propose PostCam, a streamlined framework for novel-view video generation that achieves superior detail preservation and precise camera trajectory

Prediction: Nobody knows when this will all collapse, but 2026 will be remembered in hindsight as the year in which retail investors and ind…

SafetyDGX agent

Prediction: Nobody knows when this will all collapse, but 2026 will be remembered in hindsight as the year in which retail investors and index funds were left holding the bag. It's funny how people th

Preference-Aware Rubric Learning for Personalized Evaluation

SafetyDGX agent

arXiv:2605.31545v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve from general-purpose assistants to user-centric agents, personalization has become central to aligning model beha

PReMISE: Policy Rubrics as Measurement Specifications for LLM Judges

SafetyDGX agent

arXiv:2605.30803v1 Announce Type: new Abstract: LLM judges are increasingly used to evaluate open-ended responses, but their scores depend strongly on the rubrics that condition them. A vague rubric a

Primitive Subspaces Mediate Few-Shot Transfer in VLAs

SafetyDGX agent

arXiv:2605.30695v1 Announce Type: new Abstract: Deploying vision-language-action (VLA) policies in industrial environments requires the ability to teach new tasks at low cost, a property current VLAs

Rationalize: Shared Semantic Reasoning for Human-AI Alignment

SafetyDGX agent

arXiv:2605.30632v1 Announce Type: cross Abstract: We introduce Rationalize, a role-pair framework for shared semantic reasoning between humans and AI models in data-driven sensemaking. Building on ide

RDGen: Demonstration Generation for High-Quality Robot Learning via Reinforcement Learning

SafetyDGX agent

arXiv:2605.30957v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for general-purpose robot control. However, their performance remains fundament

← Previous
1…102103104105106…214
Next →