AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,598 results
7 Aug 2026

LAWM-3D: Learning 3D-Aware Latent Actions from Human Videos for Generalizable Robot World Models

SafetyDGX agent

arXiv:2608.05706v1 Announce Type: new Abstract: World models enable agents to perform forward rollout and planning without real-world interaction. However, their application in open-world embodied int

LC-GRPO: Bridging Train-Inference Gap for Flow-Based GRPO with Langevin Correction

SafetyDGX agent

arXiv:2608.05600v1 Announce Type: cross Abstract: Flow-based generative models are typically sampled by solving a deterministic ordinary differential equation (ODE), whereas online reinforcement learn

Learning visual representations for compositional analysis of artworks and photographs

SafetyDGX agent

arXiv:2608.06142v1 Announce Type: new Abstract: Composition, the deliberate arrangement of visual elements, is central to how meaning, emotion, and aesthetic quality are conveyed in artwork, yet it re


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MACRO: Markov Chain Routing of Transformer Layers

SafetyDGX agent

arXiv:2608.05872v1 Announce Type: cross Abstract: Standard Large Language Models (LLMs) execute layers sequentially. Dynamic layer routing, i.e. search for a different execution path through layers in

Mapping Patient-Perceived Physician Traits from Nationwide Online Reviews with LLMs

SafetyDGX agent

arXiv:2510.03997v2 Announce Type: replace Abstract: Understanding how patients perceive their physicians is essential to improving trust, communication, and satisfaction. Patients increasingly consult

MapTCL: Temporal Consistency Learning via Bidirectional Alignment for Vectorized HD Map Construction

SafetyDGX agent

arXiv:2608.05209v1 Announce Type: new Abstract: Constructing reliable online HD maps remains challenging in dynamic urban environments due to moving objects and occlusions. While recent works employ f

Menlo Security targets real-time AI agent security with MARS platform

SafetyDGX agent

As AI agents gain access to enterprise systems and sensitive data, security teams need greater visibility into their actions. Real-time monitoring and policy enforcement are becoming essential as a pa

MicroEvo: Knowledge-Guided LLM Sampling for Efficient Microarchitecture Design Space Exploration

SafetyDGX agent

arXiv:2608.06183v1 Announce Type: new Abstract: Microarchitecture design space exploration suffers from expansive search spaces and expensive PPA evaluation, leaving only a small simulation budget for

Minimax Optimal Early-Stopped Gradient Descent for Gaussian Mixture Classification

SafetyDGX agent

arXiv:2608.06250v1 Announce Type: cross Abstract: In overparameterised classification, training data can be linearly separable even when the underlying distribution is not. In this setting, gradient d

MIRA: A Modular Open-Source Micro-UAV for Indoor Research

SafetyDGX agent

arXiv:2607.11785v2 Announce Type: replace Abstract: Indoor robotics research increasingly uses micro-UAV platforms whose airframes, electronics, and control software are open to modification. Off-the-

Mitigating Scoring Bias in LLM-as-a-Judge via Random Number Generation

SafetyDGX agent

arXiv:2608.05726v1 Announce Type: new Abstract: Large Language Models (LLMs) are often used as evaluators of text quality, known as LLM-as-a-Judge, which can outperform conventional automatic evaluati

MM-ISTS: Cooperating Irregularly Sampled Time Series Forecasting with Multimodal Vision-Text LLMs

SafetyDGX agent

arXiv:2603.05997v2 Announce Type: replace-cross Abstract: Irregularly sampled time series (ISTS) are widespread in real-world scenarios, exhibiting asynchronous observations on uneven time intervals a

Mood Matters: How Syntactic Sensitivity Undermines Safety Alignment

SafetyDGX agent

arXiv:2608.05409v1 Announce Type: new Abstract: Large language models typically undergo post-training to align them with safety policies but there exist many sophisticated jailbreaks that sidestep est

Multi-Agent Reinforcement Learning for Online Traffic Scheduling in Time-Sensitive Application

SafetyDGX agent

arXiv:2608.05346v1 Announce Type: cross Abstract: Time-sensitive networking (TSN) is increasingly integrated into mobile edge computing (MEC) to support applications with stringent latency requirement

Negotiating Risk Boundaries in AI for Policing Through Mixed-Stakeholder Deliberation

SafetyDGX agent

arXiv:2608.05418v1 Announce Type: new Abstract: AI tools are being increasingly adopted in policing in the UK and worldwide. Racial bias is a known and well-documented risk, yet representatives of aff

Observation-Grounded Self-Predictive Reinforcement Learning for Visual Continuous Control

SafetyDGX agent

arXiv:2608.05989v1 Announce Type: new Abstract: Sample-efficient policy learning from pixels is a long-standing challenge in reinforcement learning (RL). Recent dynamics-based representation learning

On-Policy Delta Distillation for Multilingual Math Reasoning

SafetyDGX agent

arXiv:2608.05802v1 Announce Type: new Abstract: On-Policy Distillation (OPD) is emerging as a promising alternative to reinforcement learning for LLM post-training, yet its effectiveness in multilingu

On-Policy Self-Distillation without Any Supervision

SafetyDGX agent

arXiv:2608.06296v1 Announce Type: new Abstract: On-policy (Self-)Distillation (OPD / OPSD) has shown strong potential for post-training large language models (LLMs). However, existing methods still re

Ordered Diffusion for 3D Human Registration

SafetyDGX agent

arXiv:2608.05804v1 Announce Type: new Abstract: 3D human registration has historically been treated as a regression task, assuming a unique ground-truth alignment exists between the template and an in

Overcoming Attention Drift: Homogeneity-Heterogeneity Guided Feature Aggregation for Low-Light Remote Sensing Image Enhancement

SafetyDGX agent

arXiv:2608.05843v1 Announce Type: new Abstract: Restoring high-fidelity remote sensing imagery from extreme low-light degradation is indispensable for reliable Earth observation and downstream machine

Path Planning of Cleaning Robot with Reinforcement Learning

SafetyDGX agent

arXiv:2208.08211v2 Announce Type: replace-cross Abstract: Recently, as the demand for cleaning robots has steadily increased, therefore household electricity consumption is also increasing. To solve t

PD-GS: Phoneme-Driven 3DGS for Audio-Driven Talking Heads

SafetyDGX agent

arXiv:2608.05218v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) enables fast, photorealistic talking-head rendering, yet accurate lip articulation remains elusive: mouth motion is often o

Personalized Deep Research Query Refinement with Graph-Scaffolded Evidence Grounding

SafetyDGX agent

arXiv:2608.05876v1 Announce Type: new Abstract: User requests serve as research specifications for deep research agents, shaping what evidence to seek and how to synthesize it. In personalized deep re

PhyLatent: Learning Dynamics-Relevant Representations for JEPA World Models

SafetyDGX agent

arXiv:2608.05720v1 Announce Type: new Abstract: We propose PhyLatent, a dynamics-relevant training objective for JointEmbedding Predictive Architecture (JEPA) world models. Our key observation is that

Poli-Bias: Understanding and Measuring Large Language Model Biases in International Political Conflicts

SafetyDGX agent

arXiv:2608.06123v1 Announce Type: new Abstract: Measuring political bias in large language models (LLMs) remains challenging as it can manifest through subtle differences in framing, argumentation, an

PolyAlign: Conditional Human-Distribution Alignment

SafetyDGX agent

arXiv:2606.13227v2 Announce Type: replace Abstract: Post-training methods such as supervised fine-tuning (SFT) and preference optimization typically align language models toward a single global assist

RA-CAD: Learning Post-Execution Critique for State-Aware Text-to-CAD Generation

SafetyDGX agent

arXiv:2608.05714v1 Announce Type: new Abstract: Text-to-CAD generation translates natural-language design intent into editable and executable parametric computer-aided design (CAD) codes, reducing the

RASP-QAOA: Resource-Aware Per-Instance Selection for Exact QAOA Simulation

SafetyDGX agent

arXiv:2608.05646v1 Announce Type: cross Abstract: Exact QAOA simulation spans several computational representations whose useful regions differ sharply across graph structure, circuit depth, precision

Resourced Authority A Mechanism-Design Model for Participatory Governance of Deployed AI Agents

SafetyDGX agent

arXiv:2608.06353v1 Announce Type: cross Abstract: We give a formal mechanism design model for the continuous participatory governance of a deployed AI agent. The mechanism is built on the principle th

Robust-WAM: Bridging Generative Pretraining and Semantic Foresight in World-Action Models

SafetyDGX agent

arXiv:2608.05903v1 Announce Type: new Abstract: Mainstream World-Action Models (WAMs) adapt pretrained video generation models (VGMs) for robot control, transferring their learned dynamics prior for a

RP-OPSD: Reasoning-Pivot-Guided On-Policy Self-Distillation for Multilingual Reasoning Transfer

SafetyDGX agent

arXiv:2608.06347v1 Announce Type: new Abstract: Multilingual reasoning transfer is crucial for extending reasoning capabilities of large language models (LLMs) beyond high-resource languages. On-polic

Safe Evolution with Circuit Anchors

SafetyDGX agent

arXiv:2608.05158v1 Announce Type: new Abstract: In biological evolution, unconstrained mutation can lead to catastrophic outcomes: organisms may evolve enhanced capabilities while losing essential fun

SAGA: Score-Weighted Adaptive Generation Alignment for Low-Resource Nordic Language Models

SafetyDGX agent

arXiv:2608.06179v1 Announce Type: new Abstract: Preference optimisation has proven effective for improving large language models but typically relies on costly human preference annotations. Extending

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training

SafetyDGX agent

arXiv:2608.06125v1 Announce Type: new Abstract: Latent reward models can supervise visual diffusion models without decoding intermediate states into pixel space. This makes alignment with human prefer

Scalable estimation of VARMA models

SafetyDGX agent

arXiv:2608.06340v1 Announce Type: cross Abstract: Vector autoregressive moving-average (VARMA) models have long been considered impractical beyond moderate dimensions: the likelihood is non-convex, th

SCI-CLIP: Segment-Centric Inference with Reference Memory for Training-Free Open-Vocabulary Segmentation

SafetyDGX agent

arXiv:2608.05627v1 Announce Type: new Abstract: Training-free open-vocabulary segmentation remains limited by a missing inference abstraction. Frozen vision-language features are produced at patch lev

SCP-NL2TL: Selective Conformal Prediction with Semantic Verification for Natural Language to Temporal Logic Specifications

SafetyDGX agent

arXiv:2608.05439v1 Announce Type: new Abstract: Translating natural language instructions into machine-interpretable formal specifications enables robots and autonomous systems to plan, reason, and fo

Search-Aided Joint Agent-Environment Reinforcement Learning for Robust Lifelong Multi-Agent Path Finding with Rotations

SafetyDGX agent

arXiv:2608.05588v1 Announce Type: cross Abstract: Lifelong Multi-Agent Path Finding (LMAPF) requires repeatedly planning collision-free paths for agents that continuously receive new goals upon reachi

Studying People to Study AI: Expert Perspectives on the Epistemic Fit and Barriers of Human Research in AI Safety & Ethics

SafetyDGX agent

arXiv:2608.05656v1 Announce Type: cross Abstract: Safety risks of AI are becoming increasingly evident in human interactions with AI technologies. The prominent approaches to evaluating these risks fa

Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts

SafetyDGX agent

arXiv:2510.14538v3 Announce Type: replace Abstract: Neuro-symbolic (NeSy) AI aims to develop deep neural networks whose predictions comply with prior knowledge encoding, e.g. safety or structural cons

Temporal Bridges for Spatial Resolution: Enhancing Climate Data Super-Resolution with Bidirectional Alignment

SafetyDGX agent

arXiv:2608.05981v1 Announce Type: new Abstract: High-resolution climate data is crucial for meteorological predictions and for informing decision support across diverse domains. However, the acquisiti

Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet

SafetyDGX agent

arXiv:2509.06861v3 Announce Type: replace Abstract: Test-time scaling increases inference-time computation through longer reasoning chains and has shown strong performance gains across many domains. H

Text Generation: A Systematic Literature Review of Tasks, Evaluation, and Challenges

SafetyDGX agent

arXiv:2405.15604v4 Announce Type: replace Abstract: Text generation has become more accessible than ever, and the growing interest in these systems, especially those using large language models, has s

The Closing Window: How Governments Could Lose Their Ability to Restrain Advanced AI

SafetyDGX agent

arXiv:2608.05173v1 Announce Type: cross Abstract: As AI capabilities advance, AI systems will pose greater risks to national security and potentially humanity as a whole. Governments may eventually co

The Illusion of Visual Tool-Use: A Causal Audit of Thinking with Images

SafetyDGX agent

arXiv:2608.06270v1 Announce Type: new Abstract: The 'thinking-with-images' paradigm equips multimodal LLMs with active visual operations such as crop-and-zoom. However, models using these operations o

TRACE: Learned Proprioceptive Odometry for Legged Robots under Unreliable Contact Conditions

SafetyDGX agent

arXiv:2608.05975v1 Announce Type: cross Abstract: In this paper, we present TRACE (Tokenized Robust Attention for Contact-Aware Estimation), an end-to-end learned proprioceptive odometry estimator for

Training a Conditioned Video Game Agent on a VLM Annotated Dataset

SafetyDGX agent

arXiv:2608.05954v1 Announce Type: new Abstract: Reinforcement Learning (RL) is a powerful but far from easy-to-use technique for policy learning. In the specific case of video games, access to the gam

UniVVT: A Unified End-to-End Framework for High-Fidelity Video Virtual Try-on

SafetyDGX agent

arXiv:2608.05745v1 Announce Type: cross Abstract: Video Virtual Try-On (VVT) synthesizes a video of a person wearing a target garment while preserving identity, motion, and scene dynamics. Dominant ap

VIDP: Variable Impedance Diffusion Policy for Compliant Robot Manipulation from Diverse Demonstrations

SafetyDGX agent

arXiv:2608.06210v1 Announce Type: new Abstract: Contact-rich manipulation requires precise tracking and mechanical compliance, where variable impedance control can improve robustness in task success,

Visual Grounding in Zero-Shot Vision-Language Control

SafetyDGX agent

arXiv:2608.06154v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used as zero-shot controllers, but successful trajectories do not necessarily show that decisions are g

When Privileged Guidance Misaligns: State-Matched Routing and Contextualized Self-Distillation for Multi-Turn Agents

SafetyDGX agent

arXiv:2608.05219v1 Announce Type: new Abstract: Privileged on-policy distillation provides dense supervision for multi-turn agents by allowing a synchronized teacher to re-score the student's response

Who Gets Access? Global Region and Academic Status Bias in AI-Generated Academic Gatekeeping Scenarios

SafetyDGX agent

arXiv:2608.05178v1 Announce Type: cross Abstract: Equitable access to scientific knowledge often depends on informal gatekeeping decisions, particularly when resources such as paywalled articles, data

Worst-Case Distance-Aware Error Bounds for Neural Networks

SafetyDGX agent

arXiv:2510.22021v3 Announce Type: replace Abstract: Safety-critical applications of machine learning require uncertainty estimates that support reliable worst-case analysis. Neural networks (NNs) prov

XEWorld: Can Action-Conditioned World Models Generalize to Unseen Robot Embodiments?

SafetyDGX agent

arXiv:2608.05799v1 Announce Type: cross Abstract: Action-conditioned world models are promising learned simulators for robotic manipulation, yet evaluating them exclusively on training robots fails to

6 Aug 2026

A Chain Is Only as Strong as Its Weakest Link: A Scoping Review of System Integration Audits in AI

SafetyDGX agent

arXiv:2608.04921v1 Announce Type: cross Abstract: As AI systems become increasingly integrated into diverse interfaces and applications, model-centric audits are insufficient to address risks arising

A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination

SafetyDGX agent

arXiv:2608.04872v1 Announce Type: cross Abstract: Symbolic regression aims to discover closed-form equations from data, but existing LLM-guided methods often rely on a unified proposal loop that compr

A/B Agent: A Self-Evolving Agent for Strategy Iteration in Industrial A/B Testing

SafetyDGX agent

arXiv:2608.04625v1 Announce Type: new Abstract: Industrial recommendation strategy iteration heavily relies on large-scale A/B experimentation. Traditional tuning requires experts to repeatedly design

Agent Skills for Automated Reasoning policies in Amazon Bedrock

SafetyDGX agent

Learn how to run the full Amazon Bedrock Automated Reasoning policy lifecycle from your coding agent. A suite of open source Agent Skills builds, reviews, tests, debugs, deploys, and validates a custo

Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation

SafetyDGX agent

arXiv:2608.04788v1 Announce Type: cross Abstract: Large language model agents are commonly trained through reinforcement learning with sparse trajectory-level rewards, which offer limited guidance on

Among proponents of neurosymbolic architectures, there had been some debate over the years about whether the outer level would be symbolic (…

SafetyDGX agent

Among proponents of neurosymbolic architectures, there had been some debate over the years about whether the outer level would be symbolic (i.e. a harness that calls neural models) or whether the oute

← Previous
1…7891011…210
Next →