AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
23 Jul 2026

Extreme-RGMT: Continual Learning of Highly Dynamic Skills for Robust Generalist Humanoid Control

SafetyDGX agent

arXiv:2607.20110v1 Announce Type: new Abstract: Humans can progressively acquire highly dynamic motor skills while preserving reliable everyday motor abilities. In contrast, existing humanoid controll

Formal Foundations for Known Good Reliable Die Screening in Chiplet-Based AI Systems-on-Chip

SafetyDGX agent

arXiv:2607.20141v1 Announce Type: cross Abstract: The rapid growth of chiplet-based artificial intelligence systems-on-chip (SoCs) has exposed a fundamental gap in semiconductor test methodology. Exis

From Trajectories to Prefixes: Reusing Teacher Trajectories via Replayed Prefixes and Online Continuation

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.19395v1 Announce Type: cross Abstract: Small language models are attractive backbones for interactive agents, but direct distillation from strong teacher trajectories often turns rich multi

FSDBN: Foreground-Aware EEG-Visual Alignment via Dynamic Brain Networks

SafetyDGX agent

arXiv:2607.18344v2 Announce Type: replace-cross Abstract: EEG-based visual decoding provides a non-invasive pathway for interpreting visual semantics. However, existing methods often overlook the perc

Geometry-Guided Generative Representation for Functional Brain Graphs

SafetyDGX agent

arXiv:2511.04539v2 Announce Type: replace-cross Abstract: In network neuroscience, functional brain systems are often characterized using separate yet related graph-theoretic or spectral descriptors,

Global Building Area Estimation Products: How Accurate Are They?

SafetyDGX agent

arXiv:2607.19766v1 Announce Type: new Abstract: Geo-spatial rasters of building footprint area are useful for a variety of tasks, such as monitoring urbanization, improving energy efficiency, and trac

Great X: A Unified Multi-Modal Simulator Bridging the Sim2Real Gap for 6G

SafetyDGX agent

arXiv:2507.08716v4 Announce Type: replace Abstract: Large-scale, precisely synchronized multi-modal datasets are critical for data-driven sixth-generation (6G) wireless research, yet real-world collec

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents

SafetyDGX agent

arXiv:2607.19449v1 Announce Type: cross Abstract: Evaluation frameworks for tool-augmented LLM agents focus overwhelmingly on capability metrics or explicit tool crashes, leaving silent infrastructure

Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent Triage

SafetyDGX agent

arXiv:2607.19899v1 Announce Type: cross Abstract: Disagreement-triggered escalation can create a structural blind spot in multi-agent arbitration: as base learners improve, they tend to converge, weak

Hazard or Anomaly? Evaluating VLMs for Understanding Dangers and Discrepancies

SafetyDGX agent

arXiv:2607.18325v1 Announce Type: new Abstract: Modern safety-critical systems increasingly rely on human-robot interaction to reduce disaster risk and support decision-making during emergencies. Visi

HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement

SafetyDGX agent

arXiv:2607.18217v2 Announce Type: replace Abstract: Human-object centric video personalization (HOCVP) is a core task within subject-driven video generation. However, existing methods suffer from two

How Fast Can Reward Models Score? A Systems Study of C++ and PyTorch Inference Runtimes for RLHF

SafetyDGX agent

arXiv:2607.19712v1 Announce Type: new Abstract: In RLHF pipelines, reward scoring blocks policy updates. Slow scoring bottlenecks the entire loop, since no update runs until every rollout gets a score

HyGRL: Adaptive Hybrid Graph Reasoning for Multi-Entity Questions

SafetyDGX agent

arXiv:2607.19398v1 Announce Type: new Abstract: Multi-entity compositional questions pose significant challenges to existing retrieval-augmented language models. Conventional methods fall into a dilem

In-Run Data Shapley for Adam Optimizer

SafetyDGX agent

arXiv:2602.00329v4 Announce Type: replace-cross Abstract: Reliable data attribution is essential for mitigating bias and reducing computational waste in modern machine learning, with the Shapley value

Interpretable Fuzzy Rule-Based Regression Extension for Ex-Fuzzy Library

SafetyDGX agent

arXiv:2607.20277v1 Announce Type: new Abstract: Machine learning models achieve high predictive accuracy in regression tasks, but their deployment in safety-critical and regulated domains requires int

Isaac Sim-to-Real: Reinforcement Learning based Locomotion for Quadrupeds

SafetyDGX agent

arXiv:2607.18135v1 Announce Type: cross Abstract: Learning-based approaches to locomotion have risen in popularity in recent years, showing the capability for complex legged locomotion and whole-body

It’s important that outside agencies and independent bodies hold frontier AI companies accountable for the actions of their models. This is …

SafetyDGX agent

It’s important that outside agencies and independent bodies hold frontier AI companies accountable for the actions of their models. This is not a “whoops” situation, it’s a deliberate policy choice. W

it’s just for your safety

SafetyDGX agent

On July 23, 2026, a Polymarket tweet announced that Anthropic donated $20 million to a political nonprofit calling for stricter AI regulation ahead of the U.S. midterm elections. The post, posted at 4

'It's really hard to overstate how crazy it is what happened here.' ControlAI's US Director Connor Leahy (@NPCollapse) speaks with @LindseyR…

SafetyDGX agent

'It's really hard to overstate how crazy it is what happened here.' ControlAI's US Director Connor Leahy (@NPCollapse) speaks with @LindseyReiser on CBS News, after OpenAI's own AI autonomously escape

JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models

SafetyDGX agent

arXiv:2607.17572v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) is a powerful reinforcement learning algorithm for aligning generative models with human preferences. Whil

JANUS: Foreseeing Latent Risk for Long-Horizon Agent Safety

SafetyDGX agent

arXiv:2607.19913v1 Announce Type: new Abstract: Agent safety is moving from content moderation toward preventing operational failures before tool-using agents act. We propose Janus, a foresight-orient

LatentLens: Revealing Highly Interpretable Visual Tokens in LLMs

SafetyDGX agent

arXiv:2602.00462v5 Announce Type: replace-cross Abstract: Transforming a large language model (LLM) into a vision-language model (VLM) can be achieved by mapping the visual tokens from a vision encode

Lean-SAM2: Target-Anchored Memory and Encoder Acceleration for SAM2

SafetyDGX agent

arXiv:2607.19811v1 Announce Type: new Abstract: The Segment Anything Model 2 (SAM2) has advanced temporal promptable segmentation, yet its deployment remains hindered by heavy memory cross-attention o

Learning Encoding-Decoding Direction Pairs to Unveil Concepts of Influence in Deep Vision Networks

SafetyDGX agent

arXiv:2509.23926v4 Announce Type: replace Abstract: Empirical evidence shows that deep vision networks often represent concepts as directions in latent space with concept information written along dir

Learning Personalized Safety Interventions for Haptic Human-Robot Shared Control

SafetyDGX agent

arXiv:2607.19534v1 Announce Type: new Abstract: Haptic feedback provides an implicit channel for communicating safety intentions during human-robot shared control. Existing haptic guidance systems typ

Local Causal Structure Learning in the Presence of Latent Variables and Selection Bias

SafetyDGX agent

arXiv:2607.19866v1 Announce Type: new Abstract: Discovering the direct causes and effects of a target variable from observational data is a fundamental problem in causal discovery, with broad applicat

Masked Visual Actions for Unified World Modeling

SafetyDGX agent

arXiv:2607.19343v1 Announce Type: new Abstract: Video models absorb rich priors over how the visual world moves, interacts, and responds to contact, making them promising substrates for robotic world

Matching Ranks Over Probability Yields Truly Deep Safety Alignment

SafetyDGX agent

arXiv:2512.05518v2 Announce Type: replace-cross Abstract: Open-source Large Language Models (LLMs) play a critical role in the democratization of AI, yet their 'open' nature introduces more avenues fo

MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators

SafetyDGX agent

arXiv:2607.15273v2 Announce Type: replace Abstract: MeanFlow generators achieve fast few-step sampling by predicting average velocities over time intervals, making them attractive for efficient genera

Membership Inference Attacks for Unseen Classes

SafetyDGX agent

arXiv:2506.06488v3 Announce Type: replace Abstract: A key tool in developing safe AI models is data auditing, i.e., using statistical tools to determine whether harmful content may have been used in t

Meta-Learning Preferences for Multilingual LLM Alignment

SafetyDGX agent

arXiv:2607.13315v2 Announce Type: replace Abstract: Unequal availability of human preference data across languages poses a significant challenge for aligning large language models in multilingual sett

Milo, a Fully Autonomous Indoor/Outdoor Robotic Guide Dog

SafetyDGX agent

arXiv:2607.19530v1 Announce Type: new Abstract: Many Blind and Low-Vision (BLV) people rely on guide dogs for moment-to-moment navigation, such as staying on path and avoiding obstacles and pedestrian

Minimize idle accelerators: Native RL job interleaving with co-operative time-slicing in llm-d

SafetyDGX agent

The math behind reinforcement learning (RL) post-training for large language models (LLMs) is notoriously unforgiving. As frontier AI labs push the boundaries of reasoning and coding models using RL p

Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval

SafetyDGX agent

arXiv:2607.19027v1 Announce Type: new Abstract: Zero-shot video moment retrieval aims to overcome the limitations of traditional approaches that require large-scale datasets annotated with text and it

Mitigating Scaffolding Collapse in Socratic Tutors via Representation Alignment

SafetyDGX agent

arXiv:2607.19371v1 Announce Type: new Abstract: Large language model (LLM)-based Socratic tutors increasingly guide students through multi-turn questioning, but they can suffer from scaffolding collap

ModPack: An Extensible Teleoperation Interface for Bimanual Mobile Manipulation

SafetyDGX agent

arXiv:2607.19479v1 Announce Type: cross Abstract: Existing teleoperation systems are often tailored to specific robot hardware and task domains, limiting their scalability and adaptability. We present

MOF-Sleuth: Tool-Grounded Reward Alignment for Explainable Fine-Grained MOF CIF Auditing

SafetyDGX agent

arXiv:2607.19935v1 Announce Type: new Abstract: Large metal-organic framework (MOF) databases support simulation, screening, and machine learning through crystallographic information files (CIFs). Sub

More than 300 million people turn to ChatGPT with health-related questions each week—and we’re continuing to improve how our models respond.…

SafetyDGX agent

More than 300 million people turn to ChatGPT with health-related questions each week—and we’re continuing to improve how our models respond. We work with hundreds of physicians around the world to mea

MTVDiff: Multimodal Conditional Latent Diffusion for Enhanced Thermal-to-Visible Face Translation

SafetyDGX agent

arXiv:2607.19886v1 Announce Type: new Abstract: Thermal-to-visible face translation presents fundamental challenges including geometric discontinuities, semantic attribute mismatches, and identity deg

No Training, Better Flights: Test-Time Scaled VLMs for UAV Navigation

SafetyDGX agent

arXiv:2607.19288v1 Announce Type: new Abstract: Test-time scaling offers a promising method to improve the inference performance of Vision-Language Models (VLMs) without additional training. Existing

Nonlinear Bias-Compensated Adaptive Filter and Its Application for Time-Series Prediction

SafetyDGX agent

arXiv:2607.19902v1 Announce Type: new Abstract: Most existing nonlinear adaptive filtering algorithms only account for output noise, neglecting the fact that input noise is also prevalent in practice.

Norm or Direction? Decoding Vision Mambas for High-Resolution Vision

SafetyDGX agent

arXiv:2607.18625v1 Announce Type: new Abstract: Vision Mamba models replace quadratic self-attention with linear complexity selective state space models (SSMs), emerging as efficient visual backbones.

Now You See the Hate: Adaptive View Retrieval for Hidden Hateful Illusions

SafetyDGX agent

arXiv:2607.19061v2 Announce Type: replace-cross Abstract: Hateful optical illusions expose a serious gap in current multimodal safety systems. On original-view hateful illusions, previous work shows t

One4Many-StablePacker: An Efficient Deep Reinforcement Learning Framework for the 3D Bin Packing Problem

SafetyDGX agent

arXiv:2510.10057v2 Announce Type: replace Abstract: The three-dimensional bin packing problem (3D-BPP) is widely applied in logistics and warehousing. Existing learning-based approaches often neglect

Online Variance Reduction for Domain Adaptation on Streaming Data

SafetyDGX agent

arXiv:2607.20374v1 Announce Type: new Abstract: This paper studies the problem of stochastic variance reduction (SVR) for the maximum mean discrepancy (MMD) and correlation alignment (CORAL) loss func

OpenEvoShield: Dual Non-Stationary Continual Defense for Open-World Multi-Agent System Attacks

SafetyDGX agent

arXiv:2607.19351v1 Announce Type: new Abstract: LLM-based multi-agent systems (LLM-MAS) are increasingly deployed in safety-critical applications, where adversaries inject malicious instructions throu

OPIUM: Mitigating Steering Externalities and Over-Refusal via Dual Objective Latent Optimization

SafetyDGX agent

arXiv:2607.19806v1 Announce Type: cross Abstract: Activation steering provides a lightweight mechanism for controlling large language models at inference time, but steering vectors can have unintended

Pathologist Attention-Aligned Report Generation for Prostate Histopathology

SafetyDGX agent

arXiv:2607.19624v1 Announce Type: new Abstract: The allocation of visual attention by pathologists during cancer diagnosis is a highly selective process that critically shapes the information extracte

PAVXploreRL: Physical-Action-Visual World Model Reinforcement Learning with Action Exploration

SafetyDGX agent

arXiv:2607.16602v2 Announce Type: replace Abstract: Action-conditioned world models are a key component of embodied AI, serving as scalable policy evaluators that reduce reliance on expensive real-wor

PGTT: Phase-Guided Terrain Traversal for Perceptive Legged Locomotion

SafetyDGX agent

arXiv:2510.18348v2 Announce Type: replace-cross Abstract: State-of-the-art perceptive Reinforcement Learning controllers for legged robots typically either (i) impose oscillator-or IK-based gait prior

Predictive Extrema, Unprofitable Policies: An AI-Assisted Audit of Candle-Based Binance Spot Timing Models

SafetyDGX agent

arXiv:2607.19453v1 Announce Type: cross Abstract: We audit whether candle-based machine-learning models can turn predictions of cryptocurrency extrema or short-horizon outcomes into positive Binance S

Privileged Lesion-Context Relational Distillation for Mask-Free Skin Lesion Classification

SafetyDGX agent

arXiv:2607.18773v1 Announce Type: new Abstract: Accurate skin lesion classification can benefit from lesion segmentation masks, but requiring masks or an auxiliary segmentation model during inference

Prompt Programming for Cultural Bias and Alignment of Large Language Models

SafetyDGX agent

arXiv:2603.16827v2 Announce Type: replace Abstract: Culture shapes reasoning, values, prioritization, and strategic decision-making, yet large language models (LLMs) often exhibit cultural biases that

PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference

SafetyDGX agent

arXiv:2607.20327v1 Announce Type: new Abstract: Large language models (LLMs) provide strong reasoning capabilities but are expensive to serve at scale, whereas small language models (SLMs) are cheaper

Rater State Bias in RLHF Preference Data: An Audit Framework

SafetyDGX agent

arXiv:2607.16195v2 Announce Type: replace Abstract: We identify a structured confound in Reinforcement Learning from Human Feedback (RLHF). Pairwise preference labels are intended to reflect the compa

Really insightful Substack post from @GaryMarcus ref the recent OpenAI -HuggingFace incident. It provided much need context around initial m…

SafetyDGX agent

Really insightful Substack post from @GaryMarcus ref the recent OpenAI -HuggingFace incident. It provided much need context around initial media reporting. Two points especially stood out to me: #Tech

REGEN: Replay-recycling for Expert-to-Generalist distillation with Offline Reinforcement Learning

SafetyDGX agent

arXiv:2607.19450v1 Announce Type: cross Abstract: Large-scale online reinforcement learning (RL) is the predominant means of eliciting advanced abilities including long-term reasoning and agentic tool

Rewarding Better Thinking for LLM Preference Alignment

SafetyDGX agent

arXiv:2607.19824v1 Announce Type: new Abstract: LLM preference alignment aims to optimize models toward human preferences across diverse user instructions. Reinforcement learning has become a major po

SafeGen: Goal-Conditioned Video Diffusion of Safety-Critical Scenarios for VLM-Based Autonomous Driving

SafetyDGX agent

arXiv:2607.19701v1 Announce Type: new Abstract: VLMs are increasingly deployed in AD systems, creating an urgent need for rigorous safety evaluation under rare yet safety-critical scenarios. Among the

Scale-Aware Learning of Chaotic Dynamics on Unstructured Meshes via Binned Spectral Losses

SafetyDGX agent

arXiv:2607.19387v1 Announce Type: cross Abstract: Surrogate modeling for high-dimensional nonlinear dynamical systems that exhibit chaos requires mechanisms that preserve not only pointwise accuracy b

← Previous
1…3132333435…212
Next →