AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
9 Jul 2026

Fast Rates for Semi-Supervised Learning via Data-Augmentation Graph Regularization

SafetyDGX agent

arXiv:2607.07513v1 Announce Type: new Abstract: Self-supervised learning matches supervised accuracy from a fraction of the labels, but the labeled-sample efficiency behind this has lacked a theoretic

Feynman Kac Reweighted Schrodinger Bridge Matching for Surface-Based Tau PET Harmonization

SafetyDGX agent

arXiv:2606.17420v2 Announce Type: replace-cross Abstract: Tau positron emission tomography (PET) is widely used for the in vivo characterization of disease stage and progression in Alzheimer's disease

Framing Instability in LLM Ethical Stance: Auditing Negation Sensitivity in Moral Dilemmas

SafetyDGX agent

arXiv:2601.21433v2 Announce Type: replace Abstract: Language models are increasingly consulted on ethically consequential questions, yet the stance a model expresses may not survive a change in framin


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

GemNav: Discrete-Token Visual Robot Navigation using a Multimodal Large Language Model

SafetyDGX agent

arXiv:2607.06882v1 Announce Type: cross Abstract: Visual navigation policies built on large pretrained models have so far followed a common recipe: a dedicated visual encoder, a bespoke action head, a

Gen4U: Unifying Video Generation and Understanding via Diffusion

SafetyDGX agent

arXiv:2607.06856v1 Announce Type: new Abstract: Prior work suggests that diffusion representations capture low-level geometry but struggle with high-level semantics. We demonstrate that state-of-the-a

Gradient-Based Speech-to-Text Alignment for Any ASR Model: From CTC to Speech LLMs

SafetyDGX agent

arXiv:2607.06831v1 Announce Type: cross Abstract: Speech-to-text alignment means finding the temporal boundaries of each word in the audio. Some models provide such an alignment directly and others do

HART: High-Resolution Annotation-Free Reasoning Technique through a Closed-loop Framework

SafetyDGX agent

arXiv:2602.23615v3 Announce Type: replace Abstract: Current Large Multimodal Models (LMMs) struggle with high-resolution visual inputs during the reasoning process, as the number of image tokens incre

HiDVFS: Hierarchical Multi-Agent DVFS for Real-Time OpenMP DAG Workloads

SafetyDGX agent

arXiv:2601.06425v2 Announce Type: replace-cross Abstract: Leakage power in multicore embedded systems now rivals dynamic power, so DVFS schedulers must respect deadlines and thermal limits, not just a

HumAIN: Human-Aware Implicit Social Robot Navigation

SafetyDGX agent

arXiv:2607.07357v1 Announce Type: cross Abstract: Effective social robot navigation requires sensitivity to human behavior, often revealed through subtle skeletal cues like gait and orientation. We pr

Incentivizing Temporal-Awareness in Egocentric Video Understanding Models

SafetyDGX agent

Multimodal large language models (MLLMs) have recently shown strong performance in visual understanding, yet they often lack temporal awareness, particularly in egocentric settings where reasoning dep

InductWave: Inductive Multi-Hop Logical Query Answering on Knowledge Graphs

SafetyDGX agent

arXiv:2607.07422v1 Announce Type: new Abstract: Logical Multi-Hop Query Answering over Knowledge Graphs (KGs) can be formulated as querying, with an implicit completeness assumption. Current works mai

Initiation Safety: A Missing Dimension in Generalist-Robot Safety

SafetyDGX agent

arXiv:2607.07420v1 Announce Type: new Abstract: Safety for generalist robots is usually discussed in terms of motion or dialogue. We argue a third question is missing: should the robot take its first

Latent Policy Steering through One-Step Flow Policies

SafetyDGX agent

arXiv:2603.05296v2 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL) allows robots to learn from offline datasets without risky exploration. Yet, offline RL's performance ofte

Learning social norms enhances compatibility in dynamic human-AI coordination

SafetyDGX agent

arXiv:2607.07021v1 Announce Type: new Abstract: Humans continuously coordinate with others in dynamic interactions, often through implicit, hard-to-quantify social norms that act as shared tacit expec

LHM-Humanoid: Long-Horizon Human Motion Control for Continuous Object Transport in Cluttered Scenes

SafetyDGX agent

arXiv:2508.16943v3 Announce Type: replace-cross Abstract: Physics-based human motion control can make a simulated character walk, sit, and manipulate objects with high physical realism. Almost always,

Lipschitz-Regularized Critics Lead to Policy Robustness Against Transition Dynamics Uncertainty

SafetyDGX agent

arXiv:2404.13879v5 Announce Type: replace Abstract: Uncertainties in transition dynamics pose a critical challenge in reinforcement learning (RL), often resulting in performance degradation of trained

LipSSD: Lipschitz-Constrained Single-Shot Detection for Adversarially Robust Object Detection

SafetyDGX agent

arXiv:2607.06592v1 Announce Type: cross Abstract: Object detectors have many applications in safety-critical systems, but they are known to be sensitive to worst-case perturbations such as adversarial

LLM-Guided Task-Semantic Field Factorization for Industrial Process Forecasting

SafetyDGX agent

arXiv:2607.06623v1 Announce Type: cross Abstract: Process industries rely on time-series forecasting and soft sensing to estimate quality variables that are hard to measure online. Labeled data are sc

LLM-powered reasoning in agent-based modeling

SafetyDGX agent

arXiv:2607.06757v1 Announce Type: new Abstract: Agent-based modeling (ABM) has the capability to model millions of individuals and their interactions, which is useful for policy making. However, ABMs

LLMs Silently Correct African American English: Auditing and Mitigating Dialect Bias via Activation Steering

SafetyDGX agent

arXiv:2607.06845v1 Announce Type: new Abstract: African American English (AAE), a rule-governed dialect spoken by over 30 million people, is routinely misinterpreted and 'corrected' by large language

Manual, Joystick, or Haptic Control? An In Vitro Comparison of Navigation Strategies for Robotic Interventional Neuroradiology Procedures

SafetyDGX agent

arXiv:2607.07253v1 Announce Type: new Abstract: Objective: To evaluate robotic controller interfaces for interventional neuroradiology procedures in-vitro incorporating a force-sensing platform to ass

Mathematical methods of reinforcement learning

SafetyDGX agent

arXiv:2607.06935v1 Announce Type: cross Abstract: Reinforcement learning (RL) is increasingly grounded in tools from probability, optimization, and operator theory. This survey organizes the mathemati

Max Out GRPO Signal: Adaptive Trace Prefix Control for Hard Reasoning Problems

SafetyDGX agent

arXiv:2607.07674v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) stalls on a model's hardest problems: when no rollout in a group succeeds, the group-relative advantages van

Mechanistic Interpretability for Neural Networks: Circuits, Sparse Features and Symbolic Reasoning

SafetyDGX agent

arXiv:2607.07316v1 Announce Type: new Abstract: This article offers a comprehensive overview of mechanistic interpretability, an emerging field that seeks to reverse-engineer the internal algorithms o

Microsoft President Brad Smith says the US now has AI 'regulation without transparent or complete rules', and adds that 'without rules, businesses can't plan' (Beatrice Nolan/Fortune)

SafetyDGX agent

Beatrice Nolan / Fortune: Microsoft President Brad Smith says the US now has AI “regulation without transparent or complete rules”, and adds that “without rules, businesses can't plan” — Microsoft Pre

Multi-Agent AI Control: Distributed Attacks Hamper Per-Instance Monitors

SafetyDGX agent

arXiv:2607.07368v1 Announce Type: cross Abstract: AI control is a family of techniques to prevent an AI with malicious goals from subverting its operator's intent. AI Control usually studies a single

NativeMEM: Native Memory Compression for Long-Horizon Robotic Manipulation

SafetyDGX agent

arXiv:2607.06678v1 Announce Type: new Abstract: How can pretrained Vision-Language-Action (VLA) models retain long-horizon visual histories with high-frequency updates without sacrificing efficiency?

NonTextual Target Attack

SafetyDGX agent

arXiv:2510.02999v5 Announce Type: replace-cross Abstract: Existing gradient-based jailbreak attacks on Large Language Models (LLMs) typically optimize adversarial suffixes to align the LLM output with

Online Data Selection Is Implicit Alignment

SafetyDGX agent

arXiv:2607.07023v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is often treated as a capability-adaptation step, while alignment is attributed to later preference optimization or reinfor

Open-Ended Scenario Reasoning for Specialist Model Adaptation

SafetyDGX agent

arXiv:2607.06625v1 Announce Type: cross Abstract: Process industries have accumulated validated specialist models, yet sensor drift, feedstock variation, and regime switching cause these models to deg

ORAN-DEFEND: Subspace Detection and Sanitization of Backdoor DRL xApps in Open RAN

SafetyDGX agent

arXiv:2607.06647v1 Announce Type: cross Abstract: Open Radio Access Networks (O-RAN) increasingly delegate near-real-time control to deep reinforcement learning (DRL) xApps obtained from third-party v

ORCAID: Oblique Rule-Based Continuous-Action Interpretation for Deep RL Policies

SafetyDGX agent

arXiv:2607.07235v1 Announce Type: cross Abstract: Explainability remains a key issue in reinforcement learning (RL). Distilling an interpretable policy from an agent trained in a complex environment i

PB-OEL: A Performance-Bounded Online Ensemble Learning Framework With Mixed Feedback for Real-Time Safety Assessment

SafetyDGX agent

arXiv:2503.15581v2 Announce Type: replace Abstract: Real-time safety assessment is critical for ensuring the reliable operation of complex dynamic systems. However, obtaining full safety labels in rea

Pelican-VLA 0.5: Attending Before Acting Benefits Generalization

SafetyDGX agent

arXiv:2607.06655v1 Announce Type: cross Abstract: In this report, we present Pelican-VLA 0.5, a unified VLA model that integrates vision-language understanding, future-frame generation, and action pre

Progressive Crystallization: Turning Agent Exploration into Deterministic, Lower-Cost Workflows in Production

SafetyDGX agent

arXiv:2607.07052v1 Announce Type: cross Abstract: AI agents deployed for IT operations are typically permanent cost centers because every execution requires full LLM inference, even for previously sol

Provably Efficient Off-Policy Adversarial Imitation Learning with Convergence Guarantees

SafetyDGX agent

arXiv:2405.16668v2 Announce Type: replace Abstract: Adversarial Imitation Learning (AIL) faces challenges with sample inefficiency because of its reliance on sufficient on-policy data to evaluate the

R^3: Advertisement Compliance Rectification via Group-Relative Experience Extractor and Curriculum Reinforcement

SafetyDGX agent

arXiv:2607.07318v1 Announce Type: new Abstract: Rigorous content moderation is crucial for online advertising but leads to millions of daily rejections. This scale renders manual rectification infeasi

Rapidly Learning Soft Robot Control via Implicit Time-Stepping

SafetyDGX agent

arXiv:2511.06667v2 Announce Type: replace-cross Abstract: With the explosive growth of rigid-body simulators, policy learning in simulation has become the de facto standard for most rigid morphologies

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops

SafetyDGX agent

arXiv:2607.07663v1 Announce Type: new Abstract: AI systems increasingly participate in their own improvement: revising their outputs, adapting their own harnesses during deployment, training on data t

Residual-Conservative Model Predictive Path Integral Control

SafetyDGX agent

arXiv:2607.06950v1 Announce Type: cross Abstract: Sampling-based model predictive control methods handle nonlinear dynamics and complex cost landscapes through Monte Carlo rollouts, yet typically empl

RoboSnap: One-Shot Real-to-Sim Scene Generation for Generalizable Robot Learning and Evaluation

SafetyDGX agent

arXiv:2607.06699v1 Announce Type: new Abstract: Recovering real-world scenes as interactive simulation environments can enable generalizable robot learning and reproducible policy evaluation. However,

Safe Reinforcement Learning using Ideas from Model Predictive Control

SafetyDGX agent

arXiv:2607.07252v1 Announce Type: new Abstract: Reinforcement learning (RL) enables the synthesis of control policies directly from data, making it highly appealing for complex cyber-physical systems

Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence

SafetyDGX agent

arXiv:2607.07675v1 Announce Type: new Abstract: Despite the recent promise in robot control, video generative models suffer from a domain mismatch due to their primary focus on content creation. For e

Selective Timestep Weighting and Advantage-Based Replay for Sample-Efficient Diffusion RLHF

SafetyDGX agent

arXiv:2607.07693v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) has emerged as a powerful paradigm for aligning generative models with human preferences. However, a

SHTA: Semantic Hard Token Correction and Center Alignment for Semi-Supervised Medical Image Segmentation

SafetyDGX agent

arXiv:2607.07019v1 Announce Type: new Abstract: Recent advances in semi-supervised medical image segmentation have achieved remarkable performance through prediction consistency, pseudo-label supervis

Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2607.07508v1 Announce Type: cross Abstract: Reinforcement learning (RL) is becoming increasingly important for post-training large language models (LLMs). Previous RL pipelines for LLMs were mos

SOMtime the World Ain't Fair: Violating Fairness Using Self-Organizing Maps

SafetyDGX agent

arXiv:2602.18201v2 Announce Type: replace Abstract: Unsupervised representations are widely assumed to be neutral with respect to sensitive attributes when those attributes are withheld from training.

Spatiotemporal Semantic V2X Framework for Cooperative Collision Prediction

SafetyDGX agent

arXiv:2601.17216v3 Announce Type: replace-cross Abstract: Intelligent Transportation Systems (ITS) demand real-time collision prediction to ensure road safety and reduce accident severity. Conventiona

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs

SafetyDGX agent

arXiv:2511.16107v3 Announce Type: replace-cross Abstract: Visual in-context learning (VICL) solves visual tasks by conditioning on a few input-output demonstrations without any model training. Recent

TACoS: Weakly Supervised Learning of Two-Dimensional Materials from Scribble Annotations to Precise Segmentation

SafetyDGX agent

arXiv:2607.07169v1 Announce Type: new Abstract: The precise pixel-level localization of 2D material flakes is crucial for high-throughput screening. However, traditional fully supervised methods rely

The Blind Curator: How a Biased Judge Silently Disables Skill Retirement in Self-Evolving Agents

SafetyDGX agent

arXiv:2607.07436v1 Announce Type: new Abstract: A self-evolving agent retires its bad skills by watching them fail, so what happens when the judge cannot see the failures? Skill retirement is the stru

The New York Times and several news outlets are accusing OpenAI of hiding its ability to search its AI training data and output logs for mor…

SafetyDGX agent

The New York Times and several news outlets are accusing OpenAI of hiding its ability to search its AI training data and output logs for more than two years during an ongoing copyright lawsuit: • The

The Power of Backdoor Absorption in Community Training

SafetyDGX agent

arXiv:2607.06643v1 Announce Type: cross Abstract: Backdoor attacks severely threaten large-scale AI models. When model owners delegate training to external compute providers within a decentralized tra

Thinking Seeds: Leveraging Historical Diversity for Position-Aware RL in LLMs

SafetyDGX agent

arXiv:2601.21476v2 Announce Type: replace Abstract: On-policy reinforcement learning (RL) for language model post-training suffers from a fundamental tension: as training progresses, policy entropy co

This is a very important observation. AI is not adopting the language and shaping it to fit its known ideas. It has no known ideas. It is re…

SafetyDGX agent

This is a very important observation. AI is not adopting the language and shaping it to fit its known ideas. It has no known ideas. It is repeating patterns found in the language. super interesting -

TIR-Agent: Training an Explorative and Efficient Agent for Image Restoration

SafetyDGX agent

arXiv:2603.27742v2 Announce Type: replace Abstract: Vision-language agents that orchestrate specialized tools for image restoration (IR) have emerged as a promising method, yet most existing framework

Towards a Pseudo-Labeling Workflow for Celltype-Classification from Explanted Brain Slice Recordings

SafetyDGX agent

arXiv:2607.06569v1 Announce Type: cross Abstract: This paper proposes an unsupervised workflow to pseudo-label extracellular spikes from human brain slice MEA recordings into two putative cell types:

Trajectory Inference of Human Aging from Cross-Sectional DNA Methylation Data

SafetyDGX agent

arXiv:2607.06583v1 Announce Type: cross Abstract: DNA methylation (DNAm) serves as one of the most robust molecular biomarkers of biological aging. While conventional epigenetic clocks accurately pred

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why

SafetyDGX agent

On-policy distillation offers dense, per-token supervision for training reasoning models; however, it remains unclear under which conditions this signal is beneficial and under which it is detrimental

UP: Unbounded Positive Asymmetric Optimization for Breaking the Exploration-Stability Dilemma

SafetyDGX agent

arXiv:2607.06987v1 Announce Type: new Abstract: Reinforcement learning (RL) has become the standard paradigm for enhancing the complex reasoning capabilities of large language models (LLMs). To achiev

← Previous
1…3839404142…212
Next →