AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,600 results
4 Aug 2026

Semantic Alignment of AI Models: Concept Collapse, Checkpoint Dynamics, and Cross-Lingual Transfer

SafetyDGX agent

arXiv:2608.01585v1 Announce Type: new Abstract: Language model benchmarking is a difficult task. Outcome reasoning alone does not test the model's conceptualization of language and popular open-source

Semantically Calibrated Evidence Composition for CT Vision-Language Learning

SafetyDGX agent

arXiv:2608.00239v1 Announce Type: new Abstract: Learning transferable representations from CT-report pairs requires combining whole-volume context with anatomy-specific evidence. Existing methods typi

Sen-Cap: Sensor-Flexible and Noise-Resilient Human Motion Capture via LiDAR-Camera Integration

SafetyDGX agent

arXiv:2608.02285v1 Announce Type: new Abstract: We propose Sen-Cap, a Sensor-Flexible and Noise-Resilient 3D human motion Capture framework that integrates multi-modal data from LiDAR and camera. Whil


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SG-Layout: Structured Scene Graph-Guided Layout Generation with LLMs

SafetyDGX agent

arXiv:2608.01106v1 Announce Type: new Abstract: Understanding and generating spatially coherent layouts from natural language remains a fundamental yet challenging task for large language models (LLMs

SG-WAM: Self-Guided World Modeling in Geometry-Aware Policy Space

SafetyDGX agent

arXiv:2608.01397v1 Announce Type: cross Abstract: World Action Models (WAMs) couple action generation with prediction of future states. Their effectiveness depends on whether future dynamics are model

SIGMA: Semantic Identifier Grouping for Molecular Autoregression

SafetyDGX agent

arXiv:2603.25062v2 Announce Type: replace Abstract: Autoregressive molecular models assign probability to molecular serializations even though chemical identity is invariant to serialization. Equivale

Slot2Text: Object-Centric Visual Tokenization for Efficient and Spatially Traceable Surgical MLLMs

SafetyDGX agent

arXiv:2608.01473v1 Announce Type: cross Abstract: Multimodal large language models (MLLM) for surgical scene understanding typically inject hundreds of dense visual tokens into a language model, leadi

SPAE: Spectrally Guided Autoencoder for Pretrained Visual Latents

SafetyDGX agent

arXiv:2608.01306v1 Announce Type: new Abstract: Latents from vision foundation models (VFMs) are semantically rich and well suited for visual understanding. Recent representation autoencoder methods s

SPIRIT: Spatio-temporal Pairwise Relational Modeling of Instrument-Tissue Interactions for Surgical Action Triplet Recognition

SafetyDGX agent

arXiv:2608.02188v1 Announce Type: new Abstract: Fine-grained understanding of surgical activity is essential for context-aware assistance in the operating room, including safety monitoring, adverse ev

StableMimic: Smooth Human-Like Recovery for Humanoid Motion Tracking - Learning Beyond the Tracking Distribution for Structured Post-Fall Behavior

SafetyDGX agent

arXiv:2608.02385v1 Announce Type: new Abstract: Humanoid motion trackers perform reliably within learned tracking distributions, but falls can move the robot into low-height, contact-rich states from

Start Classifying: Categorical Critics for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2608.02181v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) for large language models typically trains its critic by mean-squared-error (MSE) regression on scalar value targets.

STEAM:ASpatio-TEmporal Alignment Mixture-of-Experts Model with Hierarchical Pre-training for EEG Decoding

SafetyDGX agent

arXiv:2608.02070v1 Announce Type: new Abstract: Brain-computer interfaces (BCIs) have been widely used in motor rehabilitation, disease diagnosis, and other neural engineering scenarios. However, conv

StochSIPP: Safe Interval Path Planning in Stochastic Dynamic Environments

SafetyDGX agent

arXiv:2608.00792v1 Announce Type: new Abstract: Safe navigation under uncertain time-dependent blockage requires anticipating observations before committing to motion. We present StochSIPP, an exact c

TANGO-VIO: Triangulation-Aware Navigation with Guaranteed Feature-Observability for Visual-Inertial Odometry

SafetyDGX agent

arXiv:2608.02079v1 Announce Type: new Abstract: In vision-aided navigation and visual-inertial odometry, the quality of triangulated three-dimensional feature positions is a fundamental prerequisite f

Teleopit: A Full-Embodiment Humanoid Teleoperation System

SafetyDGX agent

arXiv:2608.01834v1 Announce Type: new Abstract: Humanoid teleoperation for demonstration collection requires coordinated whole-body motion, continuous dexterous hand control, and viewpoint control. Ex

The model takes moderation policy as a plain-language question and returns a calibrated score. Text and images — one interface. Read the ful…

SafetyDGX agent

Mistral AI has released Shieldstral, a 3‑billion‑parameter open‑weight model for content safety that can run locally on device. It interprets plain‑language moderation queries and returns calibrated s

Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and …

SafetyDGX agent

Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and reasons through the complex world - thinks before it acts. I

Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning

SafetyDGX agent

arXiv:2608.01743v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a central paradigm for large language model (LLM) post-training, but optimization toward new objectives can deg

Towards Compact Unified Multimodal Tracking: Synergizing Knowledge Distillation with Structural Pruning

SafetyDGX agent

arXiv:2608.01488v1 Announce Type: new Abstract: Unified multimodal object tracking has achieved remarkable robustness by leveraging complementary sensor data (e.g., RGB, Thermal, Depth), yet the heavy

Towards General Language-Conditioned Latent Safety Filters

SafetyDGX agent

arXiv:2608.00315v1 Announce Type: cross Abstract: Robot policies are becoming increasingly general, with vision-language-action (VLA) models enabling a single policy to execute diverse tasks specified

Track-Guided Hierarchical Reinforcement Learning for Autonomous Vehicle Drifting with Minimum-Lap-Time Planning

SafetyDGX agent

arXiv:2608.00113v1 Announce Type: new Abstract: In Formula 1, drivers optimize racing lines within tire grip limits to minimize lap times; however, in rally racing, drivers intentionally break tractio

Training Small LLMs as Spatial Multi-Agent Policies

SafetyDGX agent

arXiv:2608.01425v1 Announce Type: cross Abstract: Training LLM-based multi-agent systems with multi-agent reinforcement learning is rapidly gaining traction, and a parallel line of work argues that su

TravKAN: Fast and Interpretable Nonlinear Traversability Analysis with Kolmogorov-Arnold Networks

SafetyDGX agent

arXiv:2608.02320v1 Announce Type: cross Abstract: Traversability analysis is a fundamental capability for autonomous mobile robots operating in unstructured environments. While modern machine learning

Trust or Check? Understanding the (Evolutionary) Dynamics of User Trust in AI Systems

SafetyDGX agent

arXiv:2603.24742v2 Announce Type: replace-cross Abstract: As the capabilities and adoption of Artificial Intelligence (AI) systems grow, trust in these AI systems is an increasingly urgent concern. Mu

Trustworthy AI in Digital Health: A Comprehensive Review of Robustness and Explainability

SafetyDGX agent

arXiv:2608.02238v1 Announce Type: cross Abstract: Ensuring trust in AI systems is essential for the safe and ethical integration of machine learning systems into high-stakes domains such as digital he

TS-Reasoner: Aligning Time Series Foundation Models with LLM Reasoning

SafetyDGX agent

arXiv:2510.03519v2 Announce Type: replace Abstract: Time series reasoning is crucial to decision-making in diverse domains, including finance, energy, and scientific discovery. While existing time ser

UAV-Based Environmental Monitoring of Rip-Current Indicators Using Wavelet-Derived Texture Features

SafetyDGX agent

arXiv:2608.02448v1 Announce Type: new Abstract: Rip currents are recurrent coastal natural hazards that threaten beachgoers and create operational challenges for lifeguards and coastal managers. Relia

UDT: Reconciling U-Nets and Diffusion Transformers with Data-Adaptive Token Reduction

SafetyDGX agent

arXiv:2608.01298v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have emerged as a core architecture in generative modeling due to their scalability and adaptability to multimodal tasks.

Uncovering and Mitigating Positional Blind Spots in Vision-Language-Action Models

SafetyDGX agent

arXiv:2608.01573v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) models achieve promising performance in robotic manipulation, typically measured by success rates aggregated over pr

Understanding and Correcting Low-Frequency Bias in EEG Foundation Model

SafetyDGX agent

arXiv:2608.01898v1 Announce Type: new Abstract: Increasing EEG pretraining data scale or model capacity does not consistently improve downstream performance. We identify a persistent low-frequency bia

Understanding and Overcoming Cross-modal Fusion Bias in Multimodal Anomaly Detection From A Fisher Information Perspective

SafetyDGX agent

arXiv:2608.00986v1 Announce Type: new Abstract: Current advancements in Multimodal Anomaly Detection (MAD) are largely driven by enhancing multimodal fusion, particularly through the integration of RG

Unleashing the Potential of Large Language Models: A Blueprint for Real-Time, Enterprise-Ready Deployments

SafetyDGX agent

arXiv:2608.00419v1 Announce Type: cross Abstract: Large language models deployed in real-time, regulated settings face knowledge staleness, catastrophic forgetting, hallucination, and weak feedback lo

Upper-Expectile Multi-Step Q-Learning for Off-Policy Reinforcement Learning

SafetyDGX agent

arXiv:2608.02034v1 Announce Type: new Abstract: Multi-step returns accelerate reward propagation in off-policy reinforcement learning, but couple the evaluation of each decision to the suboptimal logg

USP-Mamba: Unmixing-Derived Spectral and Structural Prompting for Hyperspectral Image Super-Resolution

SafetyDGX agent

arXiv:2608.02401v1 Announce Type: new Abstract: Hyperspectral image super-resolution aims to reconstruct high-resolution imagery while preserving dense spectral information. Recently, Mamba-based mode

Verifier-Induced Support Reshaping in On-Policy Optimization

SafetyDGX agent

arXiv:2608.00220v1 Announce Type: cross Abstract: We show that on-policy reinforcement learning with verifiable rewards (RLVR) can improve the current objective while making successful behaviors for l

VLAGuard: A Framework for Evaluating and Mitigating Physical Attention Hijacking in Vision-Language-Action Robots within Wireless Sensor Networks

SafetyDGX agent

arXiv:2608.01028v1 Announce Type: new Abstract: Deploying Vision-Language-Action (VLA) robots as mobile edge nodes within wireless sensor networks (WSNs) requires robust protection against physical ad

Volcanic Clouds Detection through QCNN and Geostationary Satellite Multispectral Imagery

SafetyDGX agent

arXiv:2608.00072v1 Announce Type: new Abstract: Recent advances in quantum computing are opening new possibilities for Earth Observation (EO) data analysis. Quantum machine learning (QML) approaches o

WAM-Diff2: Hierarchical AR-to-Diffusion Distillation for Highly Efficient Autonomous Driving VLA

SafetyDGX agent

arXiv:2608.01035v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a prominent paradigm for end-to-end autonomous driving; however, their efficient deployment is sev

Wasserstein mixing time of the unadjusted Langevin algorithm

SafetyDGX agent

arXiv:2608.02430v1 Announce Type: cross Abstract: We provide new estimates in Wasserstein distance for the asymptotic bias of the unadjusted Langevin algorithm, in the classical setting of log-smooth

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills

SafetyDGX agent

arXiv:2608.01851v1 Announce Type: new Abstract: Robot learning is splitting into two bets: policies that bake competence into frozen weights (vision-language-action, or VLA, models), and agents that w

When May a Model Replace the Experiment? Audits, Licenses, and the Price of Trust in Surrogate-Driven Design

SafetyDGX agent

arXiv:2608.01378v1 Announce Type: new Abstract: Design campaigns in chemistry, materials science, and machine learning share a bottleneck: determining how good a candidate truly is requires an expensi

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs

SafetyDGX agent

arXiv:2608.00076v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) increasingly support high-stakes decision making by combining complementary information from images and text. W

Who Should Be Generated? Justifying Demographic Targets in Open-Ended Generation

SafetyDGX agent

arXiv:2608.02551v1 Announce Type: cross Abstract: Fairness evaluation concerns not only what a model produces, but also what its outputs ought to be compared against. When a model generates 'a CEO in

Why Does Action Chunking Improve Behavioral Cloning Performance in Robotic Control?

SafetyDGX agent

arXiv:2608.02547v1 Announce Type: new Abstract: Action chunking---predicting and executing multiple actions instead of a single action---has proven to be a critical component for learning effective ro

World Action Models in Real Time: An Empirical Study of Smooth Execution via Asynchronous Deployment

SafetyDGX agent

arXiv:2608.01880v1 Announce Type: new Abstract: World Action Models generate fixed-horizon action chunks through iterative denoising, creating substantial inference latency that can cause pauses, stal

3 Aug 2026

A detailed recap of the real-world target hacks by OpenAI and Anthropic models, exposing failures in AI alignment training and lack of meaningful supervision (Zvi Mowshowitz/Don't Worry About the Vase)

SafetyDGX agent

Zvi Mowshowitz / Don't Worry About the Vase: A detailed recap of the real-world target hacks by OpenAI and Anthropic models, exposing failures in AI alignment training and lack of meaningful supervisi

ActFovea: Runtime Safeguarding for VLA Policies via Spatiotemporal Visual-Action Consistency

SafetyDGX agent

arXiv:2607.29169v1 Announce Type: cross Abstract: Vision-language-action (VLA) policies achieve strong performance in robotic manipulation but remain vulnerable to runtime disturbances that break the

Adaptive Emotional Video Captioning via Affective Heterogeneous Graph Reasoning and Multi-task Joint Learning

SafetyDGX agent

arXiv:2607.29045v1 Announce Type: new Abstract: Emotional video captioning (EVC) aims to describe a video with both factual correctness and affective expressiveness. It requires a model to perceive su

Adaptive FastOPD: Progress-Aware Rollout Horizon Expansion for Efficient On-Policy Distillation

SafetyDGX agent

arXiv:2607.29494v1 Announce Type: new Abstract: On-policy distillation (OPD) provides dense teacher supervision along student-generated trajectories, but its online rollout process incurs substantial

Adjudicated Captioning: Multi-Agent Alignment Scoring and Consensus-Distilled Beam Arbitration for Strict Zero-Shot Image Captioning

SafetyDGX agent

arXiv:2607.28986v1 Announce Type: cross Abstract: Zero-shot image captioning (ZIC) describes images without paired image-caption supervision during captioner training, relying on text-only corpora and

Advances, challenges, and opportunities for legged robots

SafetyDGX agent

arXiv:2607.28952v1 Announce Type: new Abstract: Humanoid and quadrupedal robots have the potential to revolutionize the way we work, interact, and coexist with intelligent machines. To understand thei

Agreement Is Not Quality: Blind Expert Verification of Human and LLM Qualitative Coding When Human Consensus Is Not Ground Truth

SafetyDGX agent

arXiv:2607.28890v1 Announce Type: cross Abstract: Evaluations of LLM-assisted qualitative coding almost universally measure model performance as agreement with human coders, a practice that presumes h

An analysis of machine learning approaches for enhancing decision-making in complex discrete choice tasks

SafetyDGX agent

arXiv:2607.28854v1 Announce Type: new Abstract: Discrete choice modeling is a common tool used for preference elicitation during policy-making, but this is typically done through parametric models. Ma

Automated Reasoning policy refinement in Amazon Bedrock

SafetyDGX agent

Amazon Bedrock now supports automatic Automated Reasoning policy refinement. The refinement engine diagnoses failing tests and proposes formal-logic fixes for rule issues and language issues, and you

Automated Straight-line Sewing of Stretchable Fabrics with Different Lengths

SafetyDGX agent

arXiv:2607.29464v1 Announce Type: new Abstract: Different Length Alignment Sewing (DLAS), which involves stretching the shorter fabric to match the longer one and sewing them together in a straight li

Beyond Component Testing: Validating Agentic AI Systems

SafetyDGX agent

arXiv:2607.29405v1 Announce Type: new Abstract: Agentic AI systems act through multi-step trajectories that combine planning, tool use, memory, interaction, and adaptation. This behavior stretches val

Beyond Feature and Structure Alignment: Learning Transferable Propagation Knowledge for Graph Foundation Models

SafetyDGX agent

arXiv:2607.28980v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) have recently emerged as a promising paradigm for enabling knowledge transfer across diverse domains. Unlike traditional

Bridging the Question-Answer Gap in Retrieval-Augmented Generation: Hypothetical Prompt Embeddings

SafetyDGX agent

arXiv:2607.29402v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems synergize retrieval mechanisms with generative language models to enhance the accuracy and relevance of r

CAGE: Certified Authorization under Typed-Return Uncertainty for Tool-Using Agents

SafetyDGX agent

arXiv:2607.29190v1 Announce Type: new Abstract: Tool-using LLM agents act on typed tool returns, records pairing provenance and categorical fields with numerical values. Runtime permission gates gener

Can Zero-Shot LLMs Predict Child Malnutrition? A Fairness and Temporal Robustness Study

SafetyDGX agent

arXiv:2607.29082v1 Announce Type: new Abstract: Child malnutrition remains a major public health challenge in low- and middle-income countries, particularly in South Asia, where early identification o

← Previous
1…1516171819…210
Next →