AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
28 Apr 2026

Preserving Long-Tailed Expert Information in Mixture-of-Experts Tuning

SafetyDGX agent

arXiv:2604.23036v1 Announce Type: cross Abstract: Despite MoE models leading many benchmarks, supervised fine-tuning (SFT) for the MoE architectures remains difficult because its router layers are fra

Probing CLIP's Comprehension of 360-Degree Textual and Visual Semantics

SafetyDGX agent

arXiv:2604.24642v1 Announce Type: new Abstract: The dream of instantly creating rich 360-degree panoramic worlds from text is rapidly becoming a reality, yet a crucial gap exists in our ability to rel

ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation

SafetyDGX agent

arXiv:2604.23099v1 Announce Type: cross Abstract: Evaluating generative AI models is increasingly resource-intensive due to slow inference, expensive raters, and a rapidly growing landscape of models


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Protecting the Trace: A Principled Black-Box Approach Against Distillation Attacks

SafetyDGX agent

arXiv:2604.23238v1 Announce Type: cross Abstract: Frontier models push the boundaries of what is learnable at extreme computational costs, yet distillation via sampling reasoning traces exposes closed

Quantifying and Mitigating Self-Preference Bias of LLM Judges

SafetyDGX agent

arXiv:2604.22891v1 Announce Type: cross Abstract: LLM-as-a-Judge has become a dominant approach in automated evaluation systems, playing critical roles in model alignment, leaderboard construction, qu

Quantifying Divergence in Inter-LLM Communication Through API Retrieval and Ranking

SafetyDGX agent

arXiv:2604.22760v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly operate as autonomous agents that reason over external APIs to perform complex tasks. However, their reliabi

QuietWalk: Physics-Informed Reinforcement Learning for Ground Reaction Force-Aware Humanoid Locomotion Under Diverse Footwear

SafetyDGX agent

arXiv:2604.23702v1 Announce Type: new Abstract: Humanoid robots operating in human-centered environments (e.g., homes, hospitals, and offices) must mitigate foot--ground impact transients, as impact-i

Real-Time Non-Contact Force Compensation for Wrist-Mounted Force/Torque Sensors in Haptic-Enabled Robotic Surgery Training

SafetyDGX agent

arXiv:2604.23696v1 Announce Type: new Abstract: Haptic feedback has been a long-missed feature in robotic-assisted surgery, one that would allow surgeons to perceive tissue properties and apply contro

RecoverFormer: End-to-End Contact-Aware Recovery for Humanoid Robots

SafetyDGX agent

arXiv:2604.22911v1 Announce Type: new Abstract: Humanoid robots operating in unstructured environments must recover from unexpected disturbances-a capability that remains challenging for end-to-end co

Reflective Flow Sampling Enhancement

SafetyDGX agent

arXiv:2603.06165v2 Announce Type: replace-cross Abstract: The growing demand for text-to-image generation has led to rapid advances in generative modeling. Recently, text-to-image diffusion models tra

Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes

SafetyDGX agent

arXiv:2603.25562v2 Announce Type: replace-cross Abstract: On-policy distillation (OPD) is increasingly used in LLM post-training because it can leverage a teacher model to provide dense supervision on

Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis

SafetyDGX agent

arXiv:2604.24198v1 Announce Type: cross Abstract: Process Reward Models (PRMs) have achieved remarkable success in augmenting the reasoning capabilities of Large Language Models (LLMs) within static d

Right-to-Act: A Pre-Execution Non-Compensatory Decision Protocol for AI Systems

SafetyDGX agent

arXiv:2604.24153v1 Announce Type: new Abstract: Current AI systems increasingly operate in contexts where their outputs directly trigger real-world actions. Most existing approaches to AI safety, risk

Risk-Aware Robust Learning: Reducing Clinical Risk under Label Noise in Medical Image Classification

SafetyDGX agent

arXiv:2604.23875v1 Announce Type: cross Abstract: Noisy labels are a pervasive challenge in medical image classification, where annotation errors arise from inter-observer variability and diagnostic a

RL Token: Bootstrapping Online RL with Vision-Language-Action Models

SafetyDGX agent

arXiv:2604.23073v1 Announce Type: new Abstract: Vision-language-action (VLA) models can learn to perform diverse manipulation skills 'out of the box,' but achieving the precision and speed that real-w

RouteGuard: Internal-Signal Detection of Skill Poisoning in LLM Agents

SafetyDGX agent

arXiv:2604.22888v1 Announce Type: cross Abstract: Agent skills introduce a new and more severe form of indirect injection for LLM agents: unlike traditional indirect prompt injection, attackers can hi

Safe Navigation in Unknown and Cluttered Environments via Direction-Aware Convex Free-Region Generation

SafetyDGX agent

arXiv:2604.23648v1 Announce Type: new Abstract: Convex free regions provide a structured and optimization-friendly representation of collision-free space for robot navigation in unknown and cluttered

Safety-aware Goal-oriented Semantic Sensing, Communication, and Control for Robotics

SafetyDGX agent

arXiv:2603.13502v2 Announce Type: replace Abstract: Wirelessly-connected robotic systems empower robots with real-time intelligence by leveraging remote computing resources for decision-making. Howeve

SAGE: Sparse Adaptive Guidance for Dependency-Aware Tabular Data Generation

SafetyDGX agent

arXiv:2604.24368v1 Announce Type: new Abstract: Generating high-fidelity synthetic tabular data remains a critical challenge for enhancing data availability in privacy-sensitive and low-resource domai

SARM: Stage-Aware Reward Modeling for Long Horizon Robot Manipulation

SafetyDGX agent

arXiv:2509.25358v4 Announce Type: replace Abstract: Large-scale robot learning has made progress on complex manipulation tasks, yet long horizon, contact rich problems, especially those involving defo

Scalable Production Scheduling: Linear Complexity via Unified Homogeneous Graphs

SafetyDGX agent

arXiv:2604.23841v1 Announce Type: cross Abstract: Efficiently solving the Job Shop Scheduling Problem in real-world industrial applications requires policies that are both computationally lean and top

SceneSelect: Selective Learning for Trajectory Scene Classification and Expert Scheduling

SafetyDGX agent

arXiv:2604.24514v1 Announce Type: new Abstract: Accurate trajectory prediction is fundamentally challenging due to high scene heterogeneity - the severe variance in motion velocity, spatial density, a

Scheduling Your LLM Reinforcement Learning with Reasoning Trees

SafetyDGX agent

arXiv:2510.24832v2 Announce Type: replace Abstract: Using Reinforcement Learning with Verifiable Rewards (RLVR) to optimize Large Language Models (LLMs) can be conceptualized as progressively editing

Security Considerations for Multi-agent Systems

SafetyDGX agent

arXiv:2603.09002v2 Announce Type: replace-cross Abstract: Multi-agent artificial intelligence systems or MAS are systems of autonomous agents that exercise delegated tool authority, share persistent m

Seer: Language Instructed Video Prediction with Latent Diffusion Models

SafetyDGX agent

arXiv:2303.14897v4 Announce Type: replace Abstract: Imagining the future trajectory is the key for robots to make sound planning and successfully reach their goals. Therefore, text-conditioned video p

Self-Supervised Learning for Android Malware Detection on a Time-Stamped Dataset

SafetyDGX agent

arXiv:2604.23025v1 Announce Type: cross Abstract: Android malware detectors built with machine learning often suffer from temporal bias: models are trained and evaluated without respecting apps' actua

Self-Supervised Representation Learning via Hyperspherical Density Shaping

SafetyDGX agent

arXiv:2604.24498v1 Announce Type: new Abstract: Modern self-supervised representation learning methods often relies on empirical heuristics that are not theoretically grounded. In this study we propos

SemML 2.0: Synthesizing Controllers for LTL

SafetyDGX agent

arXiv:2604.24102v1 Announce Type: new Abstract: Synthesizing a reactive system from specifications given in linear temporal logic (LTL) is a classical problem, finding its applications in safety-criti

ShowFlow: From Robust Single Concept to Condition-Free Multi-Concept Generation

SafetyDGX agent

arXiv:2506.18493v2 Announce Type: replace Abstract: Customizing image generation remains a core challenge in controllable image synthesis. For single-concept generation, maintaining both identity pres

SkyfireAI lands $11M to bring AI autonomy to public safety and defense drones

SafetyDGX agent

Autonomous drone startup SkyfireAI Inc. today announced that it has raised 11 million in new funding to accelerate the development of its dual-use, artificial intelligence-native platform for autonomo

Sliding Mode Control for Safe Trajectory Tracking with Moving Obstacles Avoidance: Experimental Validation on Planar Robots

SafetyDGX agent

arXiv:2604.24518v1 Announce Type: cross Abstract: This paper presents a unified control framework for robust trajectory tracking and moving obstacle avoidance applicable to a broad class of mobile rob

SMP: Reusable Score-Matching Motion Priors for Physics-Based Character Control

SafetyDGX agent

arXiv:2512.03028v3 Announce Type: replace-cross Abstract: Data-driven motion priors that can guide agents toward producing naturalistic behaviors play a pivotal role in creating life-like virtual char

South Africa withdraws its first draft national AI policy after revelations that it contained fictitious sources that appeared to have been AI-generated (Nellie Peyton/Reuters)

SafetyDGX agent

Nellie Peyton / Reuters: South Africa withdraws its first draft national AI policy after revelations that it contained fictitious sources that appeared to have been AI-generated — South Africa has wit

StackFeat RL: Reinforcement Learning over Iterative Dual Criterion Feature Selection for Stable Biomarker Discovery

SafetyDGX agent

arXiv:2604.22892v1 Announce Type: new Abstract: Feature selection in high-dimensional genomic data (d gg n) demands methods that are simultaneously accurate, sparse, and stable. Existing approaches ei

Structural Enforcement of Goal Integrity in AI Agents via Separation-of-Powers Architecture

SafetyDGX agent

arXiv:2604.23646v1 Announce Type: new Abstract: Recent evidence suggests that frontier AI systems can exhibit agentic misalignment, generating and executing harmful actions derived from internally con

Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning

SafetyDGX agent

arXiv:2511.01490v2 Announce Type: replace Abstract: As synthetic data becomes widely used in language model development, understanding its impact on model behavior is crucial. This paper investigates

Talking AI with @MarioNawfal momentarily (9am PT; link info will be at his pinned tweet, or view on YouTube).

SafetyDGX agent

Gary Marcus announced a discussion about AI scheduled for 9am PT, with details to be found either in Mario Nawfal's pinned tweet or on YouTube. The conversation appears to be a live social media event

TCOD: Exploring Temporal Curriculum in On-Policy Distillation for Multi-turn Autonomous Agents

SafetyDGX agent

arXiv:2604.24005v1 Announce Type: cross Abstract: On-policy distillation (OPD) has shown strong potential for transferring reasoning ability from frontier or domain-specific models to smaller students

Text-Guided Multimodal Unified Industrial Anomaly Detection

SafetyDGX agent

arXiv:2604.22899v1 Announce Type: new Abstract: Industrial anomaly detection based on RGB-3D multimodal data has emerged as a mainstream paradigm for intelligent quality inspection. However, existing

The Collapse of Heterogeneity in Silicon Philosophers

SafetyDGX agent

arXiv:2604.23575v1 Announce Type: cross Abstract: Silicon samples are increasingly used as a low-cost substitute for human panels and have been shown to reproduce aggregate human opinion with high fid

The Consensus Trap: Dissecting Subjectivity and the 'Ground Truth' Illusion in Data Annotation

SafetyDGX agent

arXiv:2602.11318v3 Announce Type: replace Abstract: In machine learning, 'ground truth' refers to the assumed correct labels used to train and evaluate models. However, the foundational 'ground truth'

The Imbalanced User-AI Relationships as an Ethical Failure of Front-End Design in Healthcare AI

SafetyDGX agent

arXiv:2604.22767v1 Announce Type: cross Abstract: Ethical discourse on AI in healthcare has focused predominantly on back-end concerns such as bias, fairness and explainability, while the front-end in

The Swarm Intelligence Freeway-Urban Trajectories (SWIFTraj) Dataset -- Part II: A Graph-Based Approach for Trajectory Connection

SafetyDGX agent

arXiv:2602.21954v2 Announce Type: replace-cross Abstract: In Part I of this companion paper series, we introduced SWIFTraj, a new open-source vehicle trajectory dataset collected using a unmanned aeri

Thoughtful book on LLMs and linguistics — by someone who actually knows linguistics.

SafetyDGX agent

Thoughtful book on LLMs and linguistics — by someone who actually knows linguistics. Interested in the LLM vs. linguistic theory debate? My just out book will tell you (almost) all about it. https://m

Time-Series Forecasting in Safety-Critical Environments: An EU-AI-Act-Compliant Open-Source Package / Zeitreihenprognose in sicherheitskritischen Umgebungen: Ein KI-VO-konformes Open-Source-Paket

SafetyDGX agent

arXiv:2604.23859v1 Announce Type: new Abstract: With spotforecast2-safe we present an integrated Compliance-by-Design approach to Python-based point forecasting of time series in safety-critical envir

Towards Any-Quality Image Segmentation via Generative and Adaptive Latent Space Enhancement

SafetyDGX agent

arXiv:2601.02018v2 Announce Type: replace Abstract: Segment Anything Models (SAMs), known for their exceptional zero-shot segmentation performance, have garnered significant attention in the research

Towards Fair and Robust Volumetric CT Classification via KL-Regularised Group Distributionally Robust Optimisation

SafetyDGX agent

arXiv:2603.15941v2 Announce Type: replace Abstract: Automated diagnosis from chest computed tomography (CT) scans faces two persistent challenges in clinical deployment: distribution shift across acqu

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey

SafetyDGX agent

arXiv:2505.15957v4 Announce Type: replace-cross Abstract: With advancements in large audio-language models (LALMs), which enhance large language models (LLMs) with auditory capabilities, these models

Training a General Purpose Automated Red Teaming Model

SafetyDGX agent

arXiv:2604.23067v1 Announce Type: cross Abstract: Automated methods for red teaming LLMs are an important tool to identify LLM vulnerabilities that may not be covered in static benchmarks, allowing fo

Transferable Physical-World Adversarial Patches Against Object Detection in Autonomous Driving

SafetyDGX agent

arXiv:2604.23105v1 Announce Type: new Abstract: Deep learning drives major advances in autonomous driving (AD), where object detectors are central to perception. However, adversarial attacks pose sign

TSAssistant: A Human-in-the-Loop Agentic Framework for Automated Target Safety Assessment

SafetyDGX agent

arXiv:2604.23938v1 Announce Type: new Abstract: Target Safety Assessment (TSA) requires systematic integration of heterogeneous evidence, including genetic, transcriptomic, target homology, pharmacolo

UGAF-ITS: A Standards Harmonization Framework and Validation Tool for Multi-Framework AI Governance in Distributed Intelligent Transportation Systems

SafetyDGX agent

arXiv:2604.22789v1 Announce Type: cross Abstract: Organizations deploying AI-enabled Intelligent Transportation Systems face fragmented governance: ISO/IEC 42001 demands a certifiable management syste

Understanding Representation Gaps Across Scales in Tropical Tree Species Classification from Drone Imagery

SafetyDGX agent

arXiv:2604.23019v1 Announce Type: new Abstract: Accurate classification of tropical tree species from unoccupied aerial vehicle (UAV) imagery remains challenging due to high species diversity and stro

UniAda: Universal Adaptive Multi-objective Adversarial Attack for End-to-End Autonomous Driving Systems

SafetyDGX agent

arXiv:2604.23362v1 Announce Type: cross Abstract: Adversarial attacks play a pivotal role in testing and improving the reliability of deep learning (DL) systems. Existing literature has demonstrated t

UNSEEN: A Cross-Stack LLM Unlearning Defense against AR-LLM Social Engineering Attacks

SafetyDGX agent

arXiv:2604.23141v1 Announce Type: cross Abstract: Emerging AR-LLM-based Social Engineering attack (e.g., SEAR) is at the edge of posing great threats to real-world social life. In such AR-LLM-SE attac

Utility-Aware Data Pricing: Token-Level Quality and Empirical Training Gain for LLMs

SafetyDGX agent

arXiv:2604.22893v1 Announce Type: cross Abstract: Traditional data valuation methods based on ``row-count imes quality coefficient'' paradigms fail to capture the nuanced, nonlinear contributions that

V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think

SafetyDGX agent

arXiv:2604.23380v1 Announce Type: cross Abstract: Aligning denoising generative models with human preferences or verifiable rewards remains a key challenge. While policy-gradient online reinforcement

Value Alignment Tax: Measuring Value Trade-offs in LLM Alignment

SafetyDGX agent

arXiv:2602.12134v2 Announce Type: replace Abstract: Existing work on value alignment typically characterizes value relations statically, ignoring how alignment interventions, such as prompting, fine-t

Verifying Quantized GNNs With Readout Is Decidable But Highly Intractable

SafetyDGX agent

arXiv:2510.08045v2 Announce Type: replace-cross Abstract: We introduce a logical language for reasoning about quantized aggregate-combine graph neural networks with global readout (ACR-GNNs). We provi

Vision-Language-Action in Robotics: A Survey of Datasets, Benchmarks, and Data Engines

SafetyDGX agent

arXiv:2604.23001v1 Announce Type: cross Abstract: Despite remarkable progress in Vision--Language--Action (VLA) models, a central bottleneck remains underexamined: the data infrastructure that underli

← Previous
1…177178179180181…212
Next →