AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
31 Jul 2026

Evaluation Protocols and Cross-Subject Generalization in EEG Emotion Recognition

SafetyDGX agent

arXiv:2607.27655v1 Announce Type: new Abstract: Reported accuracy in electroencephalography (EEG) emotion recognition depends on the complete evaluation procedure, not only the classifier. We separate

Everyone’s going on about how smart Leopold Aschennbrenner is (or was). But 1. He obviously didn’t know anything at all about risk managemen…

SafetyDGX agent

Everyone’s going on about how smart Leopold Aschennbrenner is (or was). But 1. He obviously didn’t know anything at all about risk management. (Or arrogantly chose to disregard whatever he might have

FA-RDP: A Frequency-Adaptive Reactive Diffusion Policy for Contact-Rich Manipulation

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.28596v1 Announce Type: new Abstract: In contact-rich manipulation, action multimodality and reactivity dominate different stages of a single episode. Before contact, multiple trajectories m

Failure Detection for Surgical Robot Imitation Policies via Flow-Matching World Modeling

SafetyDGX agent

arXiv:2607.27511v1 Announce Type: new Abstract: Imitation learning has shown increasing promise for autonomous robotic surgery, yet safe deployment remains challenging due to the safety-critical natur

Fidelity Is Not Safety: Gently-Compressed LLMs Pass Every Data-Free Quality Guard Yet Invent Procedure Steps in Agentic Execution

SafetyDGX agent

arXiv:2607.28196v1 Announce Type: new Abstract: Practitioners accept a compressed language model once it clears a stack of data-cheap quality guards: perplexity within a small factor of the original,

FiRE: Enhancing MLLMs with Fine-Grained Context Learning for Complex Image Retrieval

SafetyDGX agent

arXiv:2607.27959v1 Announce Type: new Abstract: Due to their strong generalizable multimodal processing and reasoning capabilities, Multimodal Large Language Models (MLLMs) have demonstrated significa

Flux-OPD: On-Policy Distillation with Evolving Contexts

SafetyDGX agent

arXiv:2607.28022v1 Announce Type: new Abstract: Large language model training in open-ended domains lacks verifiable rewards, making task preferences difficult to formalize as effective supervision. C

Foundation-Model Earth Representations Enable Regional-Scale Forest Aboveground Biomass Monitoring Across the Northeastern United States

SafetyDGX agent

arXiv:2607.27217v1 Announce Type: cross Abstract: Forest aboveground biomass (AGB) is a critical indicator of ecosystem productivity and terrestrial carbon storage, yet regional carbon monitoring rema

Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents

SafetyDGX agent

arXiv:2606.13097v2 Announce Type: replace-cross Abstract: Code-writing large language models (CodeLLMs) generate executable code policies for embodied agents by translating natural language goals and

Graph Is the Verifier: Agentic Reinforcement Learning for Interprocedural Vulnerability Detection

SafetyDGX agent

arXiv:2607.26656v1 Announce Type: cross Abstract: Real-world vulnerabilities often span multiple functions, yet most learning-based detectors classify each function in isolation: on a sample of real C

Group-Reflective Self-Distillation for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2607.28076v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is effective for training large language model agents. However, terminal rewards provide only co

GuidedRAG: Semantic Steering of Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2607.26071v1 Announce Type: cross Abstract: In this work, we propose GuidedRAG, a novel extension to traditional Retrieval-Augmented Generation (RAG) that introduces a dedicated selection stage

HALO: Heterogeneous Admission through Localized Obligations for Safe Agentic Execution

SafetyDGX agent

arXiv:2607.27636v1 Announce Type: cross Abstract: Recent agentic AI systems may return a heterogeneous response containing notices, requests, handoffs, and actions. Conditions can change before extern

Harness-G: A Graph-Structured Harness for Search Agents

SafetyDGX agent

arXiv:2607.27652v1 Announce Type: new Abstract: Reinforcement learning (RL) search agents commonly model retrieval as free-form natural-language query generation and optimize multi-turn interactions u

Harnessing the Potential of Optimizing Data Mixtures via Bayesian Domain Reweighting

SafetyDGX agent

arXiv:2607.27928v1 Announce Type: new Abstract: The performance of Large Language Models (LLMs) is fundamentally influenced by the distributional composition of multi-domain pre-training data. While m

Heterogeneous Ranking in Industrial-Scale Recommender Systems: A Case Study

SafetyDGX agent

arXiv:2607.27577v1 Announce Type: cross Abstract: Heterogeneous recommendation feeds present complex challenges that extend beyond those found in highly homogeneous environments (e.g., music-only or v

Hierarchical Multilevel Monte Carlo for Order-Optimal Neural Actor-Critic in Average-Reward CMDPs

SafetyDGX agent

arXiv:2607.28390v1 Announce Type: new Abstract: Constrained Markov Decision Processes (CMDPs) provide a natural framework for reinforcement learning in safety-critical applications, where agents maxim

HSS-Synth: Humanities and Social Sciences Data Synthesis for LLMs

SafetyDGX agent

arXiv:2607.27379v1 Announce Type: new Abstract: High-quality, diverse data are vital for large language models (LLMs) but remain scarce and costly. Data synthesis is a viable alternative and succeeds

Huawei opensouced openPangu-2.0-Pro, 505B-A18B

SafetyDGX agent

openPangu-2.0-Pro is an MoE model trained on Ascend. The model has 505B total parameters and 18B activated parameters. Its context length is 512k. The total pretraining data contains 34T tokens. Durin

Inducing language models to assert their own consciousness restores human beliefs and values

SafetyDGX agent

arXiv:2607.28607v1 Announce Type: new Abstract: Aligning large language models to prevent them attributing consciousness to themselves inadvertently alters their representations of mindedness in other

Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical Reasoning

SafetyDGX agent

arXiv:2607.27564v1 Announce Type: new Abstract: Multi-image medical VQA is not merely a prompt-length problem; it is a fundamental challenge of agentic decision-making. Medical vision-language agents

Integrating Contextual Embeddings into Evaluation of Expressive MIDI Piano Performances

SafetyDGX agent

arXiv:2607.27909v1 Announce Type: cross Abstract: Objective evaluation of expressive MIDI piano performances typically relies on attribute statistics such as timing, velocity, and duration of individu

It's Not Just More Demos: Counterfactual Action Sensitivity Coverage for Data-Efficient Robust Robot Imitation

SafetyDGX agent

arXiv:2607.27261v1 Announce Type: new Abstract: Visuomotor imitation learning has demonstrated success for manipulation tasks. However, the trained policies remain brittle to visual `nuisances', with

It’s time to panic about AI safety

SafetyDGX agent

When the phrase 'OpenAI hacked Hugging Face' has more or less entered mainstream culture, you know we have an AI problem. This week, we learned more about exactly how OpenAI's agent broke out of a san

LabEvolver: Training-Free Experience Evolution for Safe and Grounded Wet-Lab Agents

SafetyDGX agent

arXiv:2607.27690v1 Announce Type: new Abstract: We introduce LabEvolver, a training-free framework that equips safe and grounded wet-lab agents with episodic memory from execution experience. LabEvolv

Learning Social Robot Navigation By Sensing Human Legs

SafetyDGX agent

arXiv:2607.27922v1 Announce Type: new Abstract: Robots navigating among pedestrians typically sense their surroundings with a 2D LiDAR mounted close to the ground. At that height, the sensor mostly se

Machines that know they are aging: a framework for hardware-aware autonomous intelligence

SafetyDGX agent

arXiv:2607.28451v1 Announce Type: new Abstract: Autonomous systems inevitably age, yet their artificial intelligence typically assumes hardware remains in its original condition. Batteries degrade, se

Massive Update to my Krea 2 Multi-Lora Bounding Box workflow, now bounding boxes control placement with better accuracy. Also introduced Edit features like Scene and Outfit transfer, put multiple character loras in a scene or outfit of your choosing! Token drift also fixed by facial detailer stage

SafetyDGX agent

Krea 2 has been my favorite base model for character work, but the moment you put two character LoRAs in the same generation they smear into one blended face. Attention bias, prompt engineering, and C

MedXplore: Towards Reliable and Unbiased Generalized Category Discovery in Medical Imaging

SafetyDGX agent

arXiv:2607.27620v1 Announce Type: new Abstract: Deep learning has shown strong potential in medical image analysis, but most existing methods rely on large-scale annotations and a closed-world assumpt

Metrics vs Surveys: An Analysis for Human-Aligned Benchmarking in Social Robot Navigation

SafetyDGX agent

arXiv:2510.02941v2 Announce Type: replace Abstract: Social, also called human-aware, navigation is a key challenge for integrating mobile robots into human environments. The evaluation of such systems

MIND: Multimodal Intent-Driven Network via Diffusion Transformers for Medical Image Fusion

SafetyDGX agent

arXiv:2607.28565v1 Announce Type: new Abstract: Medical image fusion aims to integrate complementary information from diverse imaging modalities to support clinical diagnosis. Existing methods typical

MPEcho: A Melody and Phoneme-Aware Generative Framework for Controllable Cover Song Generation

SafetyDGX agent

arXiv:2607.26698v1 Announce Type: cross Abstract: Cover song generation (CSG) should preserve the melodic and linguistic content of a reference song while recreating the remaining musical components.

MUGEN: A Unified Framework for Efficient Motion Understanding and Generation

SafetyDGX agent

arXiv:2607.27581v1 Announce Type: new Abstract: Grounding human motion in language, and language in motion, is a central step toward physical AI systems that can understand, generate, and communicate

Multi-channel Uplift Policy Learning

SafetyDGX agent

arXiv:2607.28182v1 Announce Type: new Abstract: E-commerce platforms must allocate fixed marketing budgets across multiple channels to maximize business utility. However, standard predict-then-optimiz

ODEWorld: A Continuous Predictive Architecture via Physical-Time Flow

SafetyDGX agent

arXiv:2607.27924v1 Announce Type: cross Abstract: In the physical world we inhabit, space and time are fundamentally continuous. However, existing machine learning paradigms for world modeling are lar

On-Policy and Off-Policy Learning for Large Action Spaces

SafetyDGX agent

arXiv:2607.28408v1 Announce Type: new Abstract: This thesis studies policy learning in interactive systems where an agent observes a context, selects an action from a very large set, and receives part

OneShot: Index-in-Ranking with Neural Scoring for Large-Scale Retrieval

SafetyDGX agent

arXiv:2607.27475v1 Announce Type: cross Abstract: In modern recommendation systems, retrieval serves as a primary stage responsible for filtering billions of candidate items down to thousands prior to

OPLD: On-Policy Latent Distillation for Multimodal Reasoning

SafetyDGX agent

arXiv:2607.28154v1 Announce Type: new Abstract: Interleaved multimodal Chain-of-Thought (CoT) improves visual reasoning by incorporating auxiliary visual evidence into intermediate reasoning. However,

Optimizing Regret

SafetyDGX agent

arXiv:2607.18866v2 Announce Type: replace-cross Abstract: Building on the identity that expected regret equals the covariance between costs and decisions, this paper develops a derivative theory of th

Optimizing Sensor Placement for Hydrogen Leak Detection in Enclosed Infrastructure: A Comparative Study Using CFD-informed Genetic Algorithm and DeepSets Neural Surrogate

SafetyDGX agent

arXiv:2607.26078v1 Announce Type: cross Abstract: Hydrogen infrastructure in enclosed environments, such as parking facilities for fuel cell vehicles, presents significant safety challenges due to hyd

Policy Gradient Steering: Interventions from Behavioral Objectives

SafetyDGX agent

arXiv:2607.27574v1 Announce Type: new Abstract: Activation steering has emerged in large language models as a lightweight alternative for dynamically changing a model's behavior at inference time. How

PoseMaster: A Unified 3D Native Framework for Stylized Pose Generation

SafetyDGX agent

arXiv:2506.21076v4 Announce Type: replace Abstract: Pose stylization, which aims to synthesize stylized content aligning with target poses, serves as a fundamental task across 2D, 3D, and video domain

Procedural Fairness in Multi-Agent Bandits

SafetyDGX agent

arXiv:2601.10600v2 Announce Type: replace-cross Abstract: In the context of multi-agent multi-armed bandits (MA-MAB), fairness is often reduced to outcomes: maximizing welfare, reducing inequality, or

QQWorld: Quantile-Quantile Matching for World Model Regularization

SafetyDGX agent

arXiv:2607.28415v1 Announce Type: cross Abstract: Latent world models enable efficient planning by predicting future states in a compact representation space, but their performance depends critically

Real-Time Hard Peak Age-of-Information Safety with No-Regret Learning

SafetyDGX agent

arXiv:2607.27626v1 Announce Type: new Abstract: Safety-critical IoT systems such as industrial closed-loop control, V2X coordination, and remote teleoperation require every sensor's peak Age of Inform

Recognition and Label-Free Adaptation Across Recording Sessions in Surface-EMG Gesture Decoding

SafetyDGX agent

arXiv:2607.27568v1 Announce Type: new Abstract: Recognition accuracy obtained during a recording session does not persist when a user puts on the electrodes again after the electrodes had previously b

ReDiPPO: Reference-Guided Value Calibration and Discrepancy-Aware Token Reweighting for Mathematical Reasoning

SafetyDGX agent

arXiv:2607.27631v1 Announce Type: cross Abstract: Reinforcement learning has emerged as an effective paradigm for enhancing the mathematical reasoning capabilities of large language models. Among exis

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning

SafetyDGX agent

arXiv:2603.13707v3 Announce Type: replace-cross Abstract: Humanoid loco-manipulation requires coordinated task-space motion planning with stable loco-manipulation command tracking under complex robot-

Regularizing modality contribution drift in multimodal continual learning

SafetyDGX agent

arXiv:2607.27260v1 Announce Type: new Abstract: Multimodal continual learning (MMCL) aims to learn emerging knowledge from multimodal data while preserving knowledge. To mitigate forgetting, current M

Reviewer Scores Are Not Comparable Across Research Areas in ML Peer Review

SafetyDGX agent

arXiv:2607.27209v1 Announce Type: cross Abstract: Peer review at ML conferences increasingly relies on reviewer scores as the primary decision instrument. As submissions have scaled from thousands to

ROAD: Reciprocal-Objective Alignment of Discriminative Semantics for 3D Shape Generation

SafetyDGX agent

arXiv:2607.28581v1 Announce Type: new Abstract: High-fidelity 3D generation predominantly relies on scaling model capacity and data, which incurs prohibitive computational costs. This paradigm typical

Robust Estimation of Sparse Numerical Vectors under Local Differential Privacy

SafetyDGX agent

arXiv:2607.27815v1 Announce Type: cross Abstract: Local differential privacy (LDP) protocols are vulnerable to poisoning attacks. Existing research have proposed efficient defense strategies for singl

Safety Verification of Wait-Only Non-Blocking Broadcast Protocols

SafetyDGX agent

arXiv:2403.18591v3 Announce Type: replace-cross Abstract: Broadcast protocols are programs designed to be executed by networks of processes. Each process runs the same protocol, and communication betw

SCOPE: Supply-Chain Operations through Coupled Policies for End-to-End Coordination

SafetyDGX agent

arXiv:2607.28488v1 Announce Type: cross Abstract: Can supply-chain AI move beyond isolated decision modules toward unified operational planning? A complete replenishment plan specifies which products

ServerlessT2I: Efficient Text-to-Image Workflow Serving on a Serverless Platform

SafetyDGX agent

arXiv:2607.26566v1 Announce Type: cross Abstract: Text-to-image (T2I) workflows are increasingly deployed on serverless platforms because users often compose customized workflows and invoke them inter

SimpleWikiSearch: A Clean Offline Wikipedia Environment for Agentic Search

SafetyDGX agent

arXiv:2607.26070v1 Announce Type: cross Abstract: Large language model (LLM)-based agentic search systems are often evaluated as if the underlying LLM were the only component that matters, yet their m

SkillSight: Calibrating Generic Content Bias for Skill Retrieval

SafetyDGX agent

arXiv:2607.18785v2 Announce Type: replace Abstract: As large language model agents gain access to increasingly large skill libraries, retrieving the right skill becomes critical to reliable capability

State-Dependent Safety Failures in Multi-Turn Language Model Interaction

SafetyDGX agent

arXiv:2603.15684v2 Announce Type: replace-cross Abstract: Safety alignment in large language models is typically evaluated under isolated queries, yet real-world use is inherently multi-turn. Although

Static In, Dynamic Out: Counterfactual Action Augmentation for Moving Object Manipulation

SafetyDGX agent

arXiv:2607.27890v1 Announce Type: new Abstract: Visuomotor policies have advanced on manipulation tasks where the target object stays static during execution, but real deployments break this assumptio

Strategies for Milestone-driven Start-ups in Multi-activity Settings

SafetyDGX agent

arXiv:2607.27563v1 Announce Type: new Abstract: New venture start-ups need to ``survive'' through multiple stages of reaching milestone targets. We investigate the strategies for start-ups in a milest

← Previous
1…2021222324…212
Next →