AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,600 results
4 Aug 2026

CoWAM: Coordination Contracts for Selective Policy Intervention with WAMs

SafetyDGX agent

arXiv:2608.02578v1 Announce Type: cross Abstract: World Action Models (WAMs) augment robot policies with action-conditioned predicted futures, but a plausible future alone does not justify changing th

CriPO: Enhancing Rubric-based RL via Self-Distillation

SafetyDGX agent

arXiv:2607.18082v3 Announce Type: replace Abstract: Rubric-based RL has recently shown promise in improving LLMs on open-ended tasks. A widely recognized limitation of rubric-based RL is limited explo

CRISP: Critical Step Perception for Training Efficient Deep Search Agents

SafetyDGX agent

arXiv:2608.01867v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly extended into deep search agents that solve complex questions through multi-step interaction with external


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cross-Branch Conflict as a Shield: Safeguarding Facial Identities in Unified Multimodal Image Editing

SafetyDGX agent

arXiv:2607.16898v2 Announce Type: replace-cross Abstract: Unified multimodal models (UMMs) have recently demonstrated powerful instruction-based image editing capabilities, while also raising serious

Cross-Domain Hybrid OPD for Generalizable Search Agents

SafetyDGX agent

arXiv:2608.02101v1 Announce Type: new Abstract: Recent advances in Reinforcement Learning (RL) have substantially improved the capabilities of autonomous search agents, enabling sophisticated planning

CTRAG: An In-Context Retrieval-based Framework for Automated Compliance Checking using LLMs

SafetyDGX agent

arXiv:2608.02472v1 Announce Type: new Abstract: Trust is fundamental in modern regulatory ecosystems, and compliance checking plays a critical role in fostering that trust. Regulatory compliance verif

DAVET: Denoising-Aware Visual Evidence Trajectory Allocation for Diffusion Vision-Language Models

SafetyDGX agent

arXiv:2608.01821v1 Announce Type: new Abstract: Diffusion vision-language models (dVLMs) iteratively denoise masked responses while conditioning each denoising step on visual evidence, making visual c

Deep Research Pretraining via Predictive Navigation

SafetyDGX agent

arXiv:2608.00432v1 Announce Type: new Abstract: Deep research agents are often trained on expensive, environment-grounded tool-use trajectories that require repeated retrieval, document inspection, an

DeepDefense: Robust Learning via Layer-Wise Gradient-Feature Alignment

SafetyDGX agent

arXiv:2511.13749v2 Announce Type: replace Abstract: Deep neural networks are known to be vulnerable to adversarial perturbations, which are small, carefully crafted inputs that lead to incorrect predi

Demystifying When and Why VLAs Fail in Contact-Rich Tasks and How to Fix Them

SafetyDGX agent

arXiv:2608.01402v1 Announce Type: new Abstract: We address the problem of understanding when and why Vision-Language-Action models struggle with contact-rich manipulation tasks that require precise ph

Developing Combined Manipulation and Locomotion Skills with Interaction Representation and Skill Composition

SafetyDGX agent

arXiv:2608.00208v1 Announce Type: new Abstract: This paper addresses how to enable a humanoid robot to learn motion policies based on developmental principles and combine policies to create more sophi

DiffuseAgent-MI: Distributionally-Grounded,Tool-Integrated Self-Evolving Agents for Faithful Visual Reasoning

SafetyDGX agent

arXiv:2608.00540v1 Announce Type: new Abstract: Tool-integrated vision-language agents have made remarkable progress on compositional and multi-step visual reasoning. Yet their outputs frequently exhi

Diffusion Policy with Behavioral Advantage Correction for Offline Reinforcement Learning

SafetyDGX agent

arXiv:2608.02332v1 Announce Type: new Abstract: In offline reinforcement learning (RL), the distribution shift between behavioral data and the learned policy can lead to erroneous Q-value estimation,

Disentangled Contrastive Learning for Zero-Shot Multilingual Dense Retrieval

SafetyDGX agent

arXiv:2608.02189v1 Announce Type: cross Abstract: Multilingual dense retrieval aims to handle queries and documents across different languages based on a unified retriever model. The challenge lies in

Distill What the Student Can See: Fisher-Projected On-Policy Distillation for Vision-Language Models

SafetyDGX agent

arXiv:2608.01263v1 Announce Type: new Abstract: On-policy distillation (OPD) samples trajectories from the current student policy and minimizes token-level divergence between student and teacher next-

Distill Where You Fail: Recovering Learning Signals of Negative RL-Groups from Adaptive Teacher Guidance

SafetyDGX agent

arXiv:2608.00782v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a standard paradigm for post-training large language models (LLMs). While Group Relativ

DocPO: Advancing Document Policy Optimization via Tailored Step-Aware Rewards

SafetyDGX agent

arXiv:2608.00536v1 Announce Type: new Abstract: Reinforcement learning (RL) for document parsing often relies on reference-based rewards rooted in edit distance (e.g., tree edit distance), yet it rema

Domain-Generalized Adaptive Semantic Communication for Collaborative Perception

SafetyDGX agent

arXiv:2608.00056v1 Announce Type: cross Abstract: We propose RSTA, a domain-generalized semantic communication framework enabling source-free V2X collaborative perception under both observation-domain

DPL: Depth-only Perceptive Humanoid Locomotion via Realistic Depth Synthesis and Cross-Attention Terrain Reconstruction

SafetyDGX agent

arXiv:2510.07152v3 Announce Type: replace Abstract: Recent advancements in legged robot perceptive locomotion have shown promising progress. However, terrain-aware humanoid locomotion remains largely

DreamTrajectory: Trajectory-Guided Action Generation with World Model Alignment for Mobile Manipulation

SafetyDGX agent

arXiv:2608.01381v1 Announce Type: new Abstract: Mobile manipulation requires a robot to coordinate base and arm motion under continuously changing viewpoints and contact conditions, within an action s

Driver2Map: Imitating Human Driving for Online High-Definition Map Construction

SafetyDGX agent

arXiv:2608.01338v1 Announce Type: new Abstract: High-definition (HD) maps are essential for autonomous driving systems. In constructing such maps, onboard multi-view camera images, standard-definition

Element-Aware Group Learning for E-Commerce Image Generation

SafetyDGX agent

arXiv:2608.00584v1 Announce Type: new Abstract: Recent advances in image generation and editing have made prompt quality a key bottleneck for e-commerce creatives. Vision-language models (VLMs) can ge

EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation

SafetyDGX agent

arXiv:2608.01221v1 Announce Type: new Abstract: Autonomous endoscopic navigation can reduce clinicians' operational burden, yet robust control remains challenging due to tissue deformation, transient

Ensemble of Unsupervised Deep Learning for Clustering Imbalanced Tabular Data

SafetyDGX agent

arXiv:2608.00346v1 Announce Type: new Abstract: Data imbalance poses a major challenge in supervised classification, where the majority-class bias contributes to false negatives and overestimates clas

Entity-Aware Sequence Transduction for Player-Centric Ball Action Spotting

SafetyDGX agent

arXiv:2608.01696v1 Announce Type: new Abstract: Player-centric ball action spotting requires temporally precise event detection together with actor attribution in crowded, partially observed multi-age

EnvShip: A Unified Framework for Context-Aware and Cross-Region Vessel Trajectory Forecasting

SafetyDGX agent

arXiv:2606.15240v2 Announce Type: replace Abstract: Accurate vessel trajectory forecasting is essential for maritime situational awareness, navigation safety, traffic management, and autonomous naviga

Evolutionary Curriculum Learning Improves Biological Sequence Modeling

SafetyDGX agent

arXiv:2608.00697v1 Announce Type: cross Abstract: Variational autoencoders (VAEs) trained on multiple sequence alignments (MSAs) have emerged as powerful generative models for biological sequences, wi

Experience-Calibrated Contrastive Decoding for Mitigating Hallucinations in LM-Based Text-to-Speech

SafetyDGX agent

arXiv:2608.00722v1 Announce Type: cross Abstract: Language model-based text-to-speech (LM-based TTS) remains vulnerable to speech hallucinations that deviate from the target text. Existing mitigation

Expert-Choice Routing Enables Adaptive Computation in Diffusion Language Models

SafetyDGX agent

arXiv:2604.01622v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) enable parallel, non-autoregressive text generation, yet existing DLM mixture-of-experts (MoE) models inherit

Exploiting Intrinsic Duality for Multi-Hop Question Generation

SafetyDGX agent

arXiv:2608.00712v1 Announce Type: new Abstract: Multi hop question generation (MQG) aims to generate questions from multiple given documents and target answers, whereas question answering (QA) focuses

Fairness Auditing: Lower Bounds on Company Manipulation

SafetyDGX agent

arXiv:2608.00568v1 Announce Type: new Abstract: Fairness audits are increasingly mandated in high-stakes applications such as hiring, lending, and automated decision-making. Recent work has establishe

Fairness in Augmented Graph Learning: A Survey

SafetyDGX agent

arXiv:2504.21296v2 Announce Type: replace Abstract: Graph learning has evolved into Augmented Graph Learning (AGL) by integrating specialized machine learning (ML) techniques. Examples include federat

FineMoLA: Towards Fine-Grained Motion-Language Alignment from Clip-Level Supervision

SafetyDGX agent

arXiv:2608.01392v1 Announce Type: new Abstract: Text-conditioned human motion generation has made rapid progress with the emergence of large-scale motion--language datasets. However, even datasets wit

From AI Weather Prediction to Infrastructure Resilience: A Real-Time Correction-Downscaling Framework for Tropical Cyclone Impact Forecasting

SafetyDGX agent

arXiv:2603.12828v2 Announce Type: replace-cross Abstract: This paper addresses a missing capability in infrastructure resilience: turning fast, global AI weather forecasts into asset-scale, actionable

From Failures to Supervision: DynamicEnvPlan for Robust Long-Horizon Embodied Planning

SafetyDGX agent

arXiv:2608.00613v1 Announce Type: new Abstract: Physical-world interaction is inherently dynamic, as environments can evolve during execution, requiring agents to adapt their plans under non-stationar

From Vessel Trajectories to Safety-Critical Encounter Scenarios: A Generative AI Framework for Autonomous Ship Digital Testing

SafetyDGX agent

arXiv:2603.28067v2 Announce Type: replace Abstract: Digital testing has emerged as a key paradigm for the development and verification of autonomous maritime navigation systems, yet the availability o

FusionRS: A Large-Scale RGB-Infrared-Style Remote Sensing Dataset for Cross-Modal Vision-Language Learning

SafetyDGX agent

arXiv:2606.17020v2 Announce Type: replace Abstract: Remote sensing vision-language models have advanced Earth observation, but available large-scale vision-language resources remain RGB-centered, leav

GAPSL: A Gradient-Aligned Parallel Split Learning over Data-Heterogeneous Edge Computing Systems

SafetyDGX agent

arXiv:2603.18540v2 Announce Type: replace Abstract: The increasing complexity of neural networks poses significant challenges for democratizing federated learning (FL) on resource-constrained edge dev

GenTrack: Physical Alignment for Robot-Native Motion Generation and Zero-Shot Humanoid Tracking

SafetyDGX agent

arXiv:2608.01410v1 Announce Type: cross Abstract: General-purpose humanoid trackers can execute diverse references, but their zero-shot coverage depends on large embodied corpora that are costly to ex

GeoArbiter: Verifiability-Guided Grounding for Remote-Sensing Multimodal LLMs

SafetyDGX agent

arXiv:2608.00877v1 Announce Type: new Abstract: Remote-sensing multimodal large language models (MLLMs) often assert facts that imagery cannot establish, such as a facility's identity or function. Coo

Grounded Vision-Language Interpreter for Long-Horizon Bimanual Task and Motion Planning

SafetyDGX agent

arXiv:2506.03270v3 Announce Type: replace Abstract: While recent advances in vision-language models have accelerated language-guided robot planning, their black-box nature lacks the safety guarantees

HapticVLA: Contact-Rich Manipulation via Vision-Language-Action Model without Inference-Time Tactile Sensing

SafetyDGX agent

arXiv:2603.15257v2 Announce Type: replace Abstract: Tactile sensing is a crucial capability for Vision-Language-Action (VLA) architectures, as it enables dexterous and safe manipulation in contact-ric

Hermite Curves as Trajectory Priors for Vision-Language-Action Models

SafetyDGX agent

arXiv:2608.01265v1 Announce Type: cross Abstract: Despite recent progress in Vision-Language-Action (VLA) models for robotic manipulation, the action chunk remains a weakly structured interface. Exist

Heterogeneous Multi-Agent Reinforcement Learning for Radio Resource Management under Coupled Finite-Horizon Constraints

SafetyDGX agent

arXiv:2608.01745v1 Announce Type: new Abstract: Maximizing throughput under proportional fairness in dense wireless networks requires jointly managing user association, scheduling, base station (BS) a

Hierarchical Pre-Training of Vision Encoders with Large Language Model

SafetyDGX agent

arXiv:2604.00086v2 Announce Type: replace-cross Abstract: The field of computer vision has experienced significant advancements through scalable vision encoders and multimodal pre-training frameworks.

How Deutsche Bank unlocked agility with an API-ready ecosystem

SafetyDGX agent

When people think about digital transformation in banking, they often focus on the visible results: mobile apps and new digital services. But there's an invisible infrastructure making all these servi

How Target is enhancing retail discovery and cutting database maintenance by 50% with Spanner Graph

SafetyDGX agent

In today’s retail environment, shoppers expect highly personalized product discovery experiences and conversational assistance that feels genuine, natural, and genuinely helpful. Today, successful pro

Human-LLM Alignment in Language Attitudes Toward Non-Native Japanese

SafetyDGX agent

arXiv:2608.01629v1 Announce Type: new Abstract: Large language models (LLMs) increasingly evaluate human writing in high-stakes domains such as hiring and academic assessment, putting non-native speak

Hybrid-Domain Posterior Sampling for Inverse Problems via Latent Flow Matching

SafetyDGX agent

arXiv:2608.00537v1 Announce Type: new Abstract: Latent Flow Models have revolutionized compressed-space image synthesis, yet their application to high-fidelity inverse problems remains bottlenecked. I

IACM-RL: Intent-Aware Context Management and Reinforcement Learning for Complex Tool Invocation under Dynamic Intent Fluctuations

SafetyDGX agent

arXiv:2608.02110v1 Announce Type: new Abstract: Executing long-horizon tool invocations in real-world environments is severely challenged by dynamic user intent noise. Existing methods attempt robustn

Inference-Time Policy Alignment for Fair Reinforcement Learning

SafetyDGX agent

arXiv:2608.00175v1 Announce Type: new Abstract: Deep reinforcement learning (RL) agents achieve strong performance by optimizing scalar reward functions. However, once deployed, the policies of these

Instruction-Conditioned Exploration with Asymmetric Reinforcement Learning and Self-Distillation

SafetyDGX agent

arXiv:2608.02087v1 Announce Type: cross Abstract: Post-training Large Language Models (LLMs) with Reinforcement Learning (RL) has become an important tool for improving model capabilities, but the LLM

Interaction Dynamics MPC for Knee Rehabilitation Exoskeletons: A Closed-Loop SEA Outer-Loop Study

SafetyDGX agent

arXiv:2606.13485v3 Announce Type: replace-cross Abstract: Safe rehabilitation is an interaction-dynamics problem: the controller must regulate a prescribed motion while absorbing involuntary spasm, vo

Investigating Social Bias in Narrative Image Generation

SafetyDGX agent

arXiv:2608.01780v1 Announce Type: new Abstract: Text-to-image (T2I) generation models are increasingly embedded in applications such as media content creation and education, raising concerns about how

Invisible Ink Threats: Adversarial Goals Behind Legitimate Tasks in Computer-Use Agents

SafetyDGX agent

arXiv:2608.02018v1 Announce Type: new Abstract: Computer-use agents (CUAs), which empower large language models to autonomously operate operating systems and the web, are increasingly vulnerable to in

LAB-Tab: LLM-Augmented Bayesian Network Adaptation for Few-Shot Tabular Generation

SafetyDGX agent

arXiv:2608.01879v1 Announce Type: new Abstract: Tabular data generation supports analysis and decision-making when target-domain data are scarce, yet collecting complete target samples is often costly

Latent-Regime Bias Auditing for Volatility Forecasting

SafetyDGX agent

arXiv:2608.01599v1 Announce Type: new Abstract: Volatility forecasts are commonly evaluated with aggregate accuracy metrics such as RMSE and MAE, but these metrics can hide conditional failures that m

LEAP: Lean Environment-Feedback via Adaptive Pruning for Code RL in GPU Kernel Generation

SafetyDGX agent

arXiv:2608.01804v1 Announce Type: new Abstract: Post-training large language models (LLMs) via reinforcement learning (RL) has significantly advanced code generation capabilities. To bypass the heavy

Learning-Based Collaborative MEC for LLM Inference with Soft-Deadline Awareness via Transformer-Enhanced PPO

SafetyDGX agent

arXiv:2608.02031v1 Announce Type: cross Abstract: This paper investigates collaborative mobile edge computing (MEC) servers for large language model (LLM) inference under soft deadline constraints. In

Learning-Based Motion Planning for Dynamic Environments: From Foundational Algorithms to Emerging Paradigms

SafetyDGX agent

arXiv:2608.00625v1 Announce Type: new Abstract: Motion planning in dynamic environments is a fundamental problem in robotics, aiming to generate safe and efficient paths, trajectories, or control acti

← Previous
1…1314151617…210
Next →