AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,485 results
20 May 2026

CEER: Compliant End-Effector and Root Control as a Unified Interface for Hierarchical Humanoid Loco-Manipulation

SafetyDGX agent

arXiv:2605.19981v1 Announce Type: new Abstract: Humanoid robots have achieved impressive locomotion performance, yet contact-rich and long-horizon manipulation remains a major bottleneck. Manipulation

Certifiable Alignment of GNSS and Local Frames via Lagrangian Duality

SafetyDGX agent

arXiv:2512.20931v2 Announce Type: replace Abstract: Estimating the absolute orientation of a local system relative to a global navigation satellite system (GNSS) reference often suffers from local min

Chessformer: A Unified Architecture for Chess Modeling

SafetyDGX agent

arXiv:2605.19091v1 Announce Type: new Abstract: Chess has long served as a canonical testbed for artificial intelligence, but modeling approaches for its central tasks have diverged. Maximizing playin

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CoLD: Counterfactually-Guided Length Debiasing for Process Reward Models in Mathematical Reasoning

SafetyDGX agent

arXiv:2507.15698v2 Announce Type: replace-cross Abstract: Process Reward Models (PRMs) play a central role in evaluating and guiding multi-step reasoning in large language models (LLMs), especially fo

Concept-Guided Noisy Negative Suppression for Zero-Shot Classification and Grounding of Chest X-Ray Findings

SafetyDGX agent

arXiv:2605.19374v1 Announce Type: cross Abstract: Vision-language alignment using chest X-rays and radiology reports has emerged as an advanced paradigm for zero-shot classification and grounding of c

Context-dependent manifold learning: A neuromodulated constrained autoencoder approach

SafetyDGX agent

arXiv:2603.11673v2 Announce Type: replace Abstract: Many physical systems exhibit a low-dimensional structure that varies with external parameters: link lengths in a robot, forcing constants in a flui

ContextFlow: Hierarchical Task-State Alignment for Long-Horizon Embodied Agents

SafetyDGX agent

arXiv:2605.19314v1 Announce Type: cross Abstract: Long-horizon embodied agents increasingly delegate navigation, search, approach, and manipulation to specialist executors. As these executors become s

CopT: Contrastive On-Policy Thinking with Continuous Spaces for General and Agentic Reasoning

SafetyDGX agent

arXiv:2605.20075v1 Announce Type: cross Abstract: Chain-of-thought (CoT) is a standard approach for eliciting reasoning capabilities from large language models (LLMs). However, the common CoT paradigm

‼️ Could large language models turn out to be the tech industry’s Vietnam? All In’s @jason notes below that today’s students are speaking ou…

SafetyDGX agent

‼️ Could large language models turn out to be the tech industry’s Vietnam? All In’s @jason notes below that today’s students are speaking out against AI, just as students in the 60s and 70s spoke out

CriterAlign: Criterion-Centric Rationale Alignment for Code Preference Judging

SafetyDGX agent

arXiv:2605.19665v1 Announce Type: cross Abstract: Pairwise human preference prediction is central to evaluating code-generation systems, where quality often depends on task-specific trade-offs beyond

Cross-modal Consistency Guidance for Robust Emotion Control in Auto-Regressive TTS Models

SafetyDGX agent

arXiv:2510.13293v3 Announce Type: replace Abstract: While Text-to-Speech (TTS) systems enable emotional control via natural-language instructions, expressiveness, naturalness, and speech quality degra

D-CLING: Prior-Preserving Depth-Conditioned Fine-Tuning for Navigation Foundation Models

SafetyDGX agent

arXiv:2605.19690v1 Announce Type: new Abstract: Navigation Foundation Models (NFMs) trained on large cross-embodied datasets have demonstrated powerful generalizability in various scenarios. Adopting

Data-driven Acceleration of MPC with Guarantees

SafetyDGX agent

arXiv:2511.13588v2 Announce Type: replace-cross Abstract: Model Predictive Control (MPC) is a powerful framework for optimal control but can be too slow for low-latency applications. We present a data

DEFLECT: Delay-Robust Execution via Flow-matching Likelihood-Estimated Counterfactual Tuning for VLA Policies

SafetyDGX agent

arXiv:2605.19294v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies are typically deployed with asynchronous inference: the robot executes a previously predicted action chunk while

Dimensional Balance Improves Large Scale Spatiotemporal Prediction Performance

SafetyDGX agent

arXiv:2605.18793v1 Announce Type: cross Abstract: Accurate spatiotemporal pattern analysis is critical in fields such as urban traffic, meteorology, and public health monitoring. However, existing met

Directed Acyclic Graph Convolutional Networks

SafetyDGX agent

arXiv:2506.12218v2 Announce Type: replace-cross Abstract: Directed acyclic graphs (DAGs) are central to science and engineering applications including causal inference, scheduling, and neural architec

Domain-Adaptive Communication-Rate Optimization for Sim-to-Real Humanoid-Robot Wireless XR Teleoperation

SafetyDGX agent

arXiv:2605.19293v1 Announce Type: cross Abstract: Wireless extended reality (XR) teleoperation provides embodied interaction capability for collecting humanoid robot demonstrations, but the large-scal

Don't Let Bandit Feedback Pull Continual LLM-Recommender Updates Off Target

SafetyDGX agent

arXiv:2605.18899v1 Announce Type: cross Abstract: Generative LLM-based recommenders (LLM-Rec) require continual post-deployment updates, yet deployment logs provide only policy-shaped contextual bandi

Dual-Gated Epistemic Time-Dilation: Autonomous Compute Modulation in Asynchronous MARL

SafetyDGX agent

arXiv:2603.23722v2 Announce Type: replace-cross Abstract: While Multi-Agent Reinforcement Learning (MARL) algorithms achieve unprecedented successes across complex continuous domains, their standard d

DynaSTy: A Framework for SpatioTemporal Node Attribute Prediction in Dynamic Graphs

SafetyDGX agent

arXiv:2601.05391v2 Announce Type: replace Abstract: Accurate multistep forecasting of node-level attributes on dynamic graphs is critical for applications ranging from financial trust networks to biol

DynaTok: Temporally Adaptive and Positional Bias-Aware Token Compression for Video-LLMs

SafetyDGX agent

arXiv:2605.19322v1 Announce Type: new Abstract: Recent advances in Video Large Language Models (Video-LLMs) have greatly expanded multimodal reasoning capabilities. However, the massive number of visu

Efficient Transferable Optimal Transport via Min-Sliced Transport Plans

SafetyDGX agent

arXiv:2511.19741v3 Announce Type: replace Abstract: Optimal Transport (OT) offers a powerful framework for finding correspondences between distributions and addressing matching and alignment problems

Exact Linear Attention

SafetyDGX agent

arXiv:2605.18848v1 Announce Type: cross Abstract: This paper introduces Exact Linear Attention (ELA), a mechanism that achieves linear computational complexity for Transformer attention by leveraging

Extreme Self-Preference in Language Models

SafetyDGX agent

arXiv:2509.26464v2 Announce Type: replace Abstract: Self-preference is a fundamental feature of biological organisms. Since large language models (LLMs) lack sentience, they might be expected to avoid

Feature-Space Smoothing: Certified Robustness of Deep Representations

SafetyDGX agent

arXiv:2601.16200v3 Announce Type: replace-cross Abstract: Modern deep learning models exhibit strong capabilities across diverse applications, yet remain vulnerable to malicious inputs that induce err

Formal Skill: Programmable Runtime Skills for Efficient and Accurate LLM Agents

SafetyDGX agent

arXiv:2605.19604v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly act inside real workspaces, where tools and skills determine whether model reasoning becomes reliable act

GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning

SafetyDGX agent

arXiv:2605.19235v1 Announce Type: new Abstract: Competitive multi-agent reinforcement learning in imperfect-information games requires agents to act under partial observability and against adversarial

Generative-Evaluative Agreement: A Necessary Validity Criterion for LLM-Enabled Adaptive Assessment

SafetyDGX agent

arXiv:2605.19529v1 Announce Type: new Abstract: When the same LLM generates assessment items, simulates student responses, and scores them, the validation loop is self-referential. We introduce Genera

HeadRank: Decoding-Free Passage Reranking via Preference-Aligned Attention Heads

SafetyDGX agent

arXiv:2604.17237v2 Announce Type: replace-cross Abstract: Decoding-free reranking methods that read relevance signals directly from LLM attention weights offer significant latency advantages over auto

HiDe: Rethinking The Zoom-IN method in High Resolution MLLMs via Hierarchical Decoupling

SafetyDGX agent

arXiv:2510.00054v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have made significant strides in visual understanding tasks. However, their performance on high-resol

HOI-PAGE: Zero-Shot Human-Object Interaction Generation with Part Affordance Guidance

SafetyDGX agent

arXiv:2506.07209v2 Announce Type: replace-cross Abstract: We present HOI-PAGE, a new approach that prioritizes part-level affordance reasoning to generate high-fidelity 4D human-object interactions (H

How Do Document Parsers Break? Auditing Structural Vulnerability in Document Intelligence

SafetyDGX agent

arXiv:2605.19309v1 Announce Type: new Abstract: Document Layout Analysis (DLA) pipelines provide structured page representations for retrieval-augmented generation, long-document question answering, a

How does longer temporal context enhance multimodal narrative video processing in the brain?

SafetyDGX agent

arXiv:2602.07570v2 Announce Type: replace-cross Abstract: Understanding how humans and artificial intelligence systems process complex narrative videos is a fundamental challenge at the intersection o

How Does Overparameterization Affect Machine Unlearning of Deep Neural Networks?

SafetyDGX agent

arXiv:2503.08633v2 Announce Type: replace Abstract: Machine unlearning is the task of updating a trained model to forget specific training data without retraining from scratch. In this paper, we inves

Implicit Bias of Mirror Flow in Homogeneous Neural Networks: Sparse and Dense Feature Learning

SafetyDGX agent

arXiv:2605.19458v1 Announce Type: new Abstract: We study the max-margin solutions reached by mirror flow in deep neural networks with homogeneous activation functions. Extending classical results on g

Increasing Missingness to Reduce Bias: Richardson-SGD with Missing Data

SafetyDGX agent

arXiv:2605.19641v1 Announce Type: cross Abstract: Stochastic gradient methods are central to modern large-scale learning, but their use with incomplete covariates remains delicate since imputation sch

Interoceptive Divergence in Aesthetic Evaluation and Implications for Human-AI Alignment

SafetyDGX agent

arXiv:2605.18759v1 Announce Type: cross Abstract: Artificial intelligence (AI), exemplified by large language models (LLMs), is rapidly approaching and in some cases surpassing human performance acros

Inverse Design of Metasurface based Absorbers using Physics Guided Conditional Diffusion Models

SafetyDGX agent

arXiv:2605.19611v1 Announce Type: new Abstract: Inverse design of metasurfaces for specific electromagnetic responses requires generating geometries that satisfy stringent spectral constraints while m

Kickstarter retracts stricter rules on mature content after creator backlash, and says it adopted the tougher rules because of its payment processor Stripe (Mariella Moon/Engadget)

SafetyDGX agent

Mariella Moon / Engadget: Kickstarter retracts stricter rules on mature content after creator backlash, and says it adopted the tougher rules because of its payment processor Stripe — It explained tha

LambdaPO: A Lambda Style Policy Optimization for Reasoning Language Models

SafetyDGX agent

arXiv:2605.19416v1 Announce Type: new Abstract: Group Relative Policy Optimization(GRPO) has become a cornerstone of modern reinforcement learning alignment, prized for its efficacy in foregoing an ex

Learning ORDER-Aware Multimodal Representations for Composite Materials Design

SafetyDGX agent

arXiv:2602.02513v2 Announce Type: replace Abstract: Artificial intelligence has shown remarkable success in materials discovery and property prediction, particularly for crystalline and polymer system

Learning with Foresight: Enhancing Neural Routing Policy via Multi-Node Lookahead Prediction

SafetyDGX agent

arXiv:2605.19975v1 Announce Type: cross Abstract: Neural policies have shown promise in solving vehicle routing problems due to their reduced reliance on handcrafted heuristics. However, current train

LLM agents & memory systems operate in continuously updated environments (Git repos, evolving docs). They must process long contexts, recove…

SafetyDGX agent

LLM agents & memory systems operate in continuously updated environments (Git repos, evolving docs). They must process long contexts, recover earlier information, and reason over many updates that cre

Low-Compute Watermark Removal via Dual-Domain Natural Projection

SafetyDGX agent

arXiv:2510.07538v2 Announce Type: replace Abstract: Effective removal of semantic watermarks requires balancing three competing objectives: high removal success, low perceptual distortion, and low com

Measuring Stereotype and Deviation Biases in Large Language Models

SafetyDGX agent

arXiv:2508.06649v3 Announce Type: replace Abstract: Large language models (LLMs) are widely applied across diverse domains, raising concerns about their limitations and potential risks. In this study,

Mega-ASR: Towards In-the-wild^2 Speech Recognition via Scaling up Real-world Acoustic Simulation

SafetyDGX agent

arXiv:2605.19833v1 Announce Type: cross Abstract: Despite rapid advances in automatic speech recognition (ASR) and large audio-language models, robust recognition in real-world environments remains li

Memory-Augmented Reinforcement Learning Agent for CAD Generation

SafetyDGX agent

arXiv:2605.19748v1 Announce Type: new Abstract: Automatic generation of computer-aided design (CAD) models is a core technology for enabling intelligence in advanced manufacturing. Existing generation

Metric-Gradient Projection for Stable Multi-Agent Policy Learning

SafetyDGX agent

arXiv:2605.18809v1 Announce Type: cross Abstract: General-sum multi-agent learning is often governed by a stacked update field in which each agent's policy update changes the optimization landscape fa

Multi-Session Ground Texture SLAM in Low-Dynamic Environments

SafetyDGX agent

arXiv:2605.19701v1 Announce Type: new Abstract: The simultaneous localization and mapping community has introduced a growing number of systems adapted for multi-session operations where the operationa

Neuron Incidence Redistribution for Fairness in Medical Image Classification

SafetyDGX agent

arXiv:2605.19393v1 Announce Type: new Abstract: Deep learning models for medical image classification are susceptible to subgroup performance disparities across demographic attributes such as age, gen

Noise-corrected GRPO: From Noisy Rewards to Unbiased Gradients

SafetyDGX agent

arXiv:2510.18924v3 Announce Type: replace-cross Abstract: Reinforcement learning from human feedback (RLHF) or verifiable rewards (RLVR), the standard paradigm for aligning LLMs or building recent SOT

Not all uncertainty is alike: volatility, stochasticity, and exploration

SafetyDGX agent

arXiv:2605.19215v1 Announce Type: new Abstract: Adaptive decision-making in biological and artificial intelligence requires balancing the exploitation of known outcomes with the exploration of uncerta

Not Every Rubric Teaches Equally: Policy-Aware Rubric Rewards for RLVR

SafetyDGX agent

arXiv:2605.20164v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has made post-training highly effective when correctness can be checked automatically. However, many impo

One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer

SafetyDGX agent

arXiv:2511.22940v3 Announce Type: replace Abstract: Recent advances in diffusion models have greatly improved pose-driven character animation. However, existing methods are limited to spatially aligne

Optimal Representation Size: High-Dimensional Analysis of Pretraining and Linear Probing

SafetyDGX agent

arXiv:2605.20105v1 Announce Type: new Abstract: Learning to generalise from limited data is a fundamental challenge for both artificial and biological systems. A common strategy is to extract reusable

PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.19580v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models show promising ability in language-guided robotic tasks. However, making VLA policies reliable remains challenging,

patches allow for trust boundaries > patches separate 'system wants to change state' from 'change is now accepted policy determines what hap…

SafetyDGX agent

Patches represent a mechanism for establishing trust boundaries in systems by decoupling the intent to change state from the acceptance and policy-driven implementation of that change. This separation

PEEK: Context Map as an Orientation Cache for Long-Context LLM Agents

SafetyDGX agent

arXiv:2605.19932v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate over long and recurring external contexts, like document corpora and code repositories. Across in

Phase-Aware Mixture of Experts for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2602.17038v3 Announce Type: replace Abstract: Reinforcement learning (RL) has equipped LLM agents with a strong ability to solve complex tasks. However, existing RL methods normally use a single

Physics-informed simulation framework for realistic sonar image generation and statistical validation

SafetyDGX agent

arXiv:2605.19712v1 Announce Type: new Abstract: Synthetic sonar datasets offer a scalable alternative to costly real-world acquisition, yet their utility remains limited by the absence of rigorous qua

← Previous
1…156157158159160…242
Next →