AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
1 Jul 2026

Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking

SafetyDGX agent

arXiv:2509.12046v2 Announce Type: replace-cross Abstract: Although autoregressive (AR) models have demonstrated remarkable success in image generation, extending these models to layout-conditioned gen

Learning Dexterous Grasping from Sparse Taxonomy Guidance

SafetyDGX agent

arXiv:2604.04138v2 Announce Type: replace-cross Abstract: Dexterous manipulation requires planning a grasp configuration suited to the object and task, which is then executed through coordinated multi

Learning Structurally Consistent Representations for Multi-View Radar Semantic Segmentation

SafetyDGX agent

arXiv:2606.31609v1 Announce Type: cross Abstract: Radar sensors provide reliable perception under adverse weather and lighting conditions, but their sparse, noisy, and weakly semantic measurements mak

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Learning Where to Look: A Reinforcement Learning Framework for Robust Micro-Ultrasound Prostate Cancer Detection

SafetyDGX agent

arXiv:2606.30951v1 Announce Type: cross Abstract: Micro-ultrasound (muUS) is a new, emerging, and promising imaging modality for prostate cancer (PCa) detection, but accurate identification of suspici

Linguistic Bias Mitigation for Spoofing Detection via Gradient Reversal and A Variational Information Bottleneck

SafetyDGX agent

arXiv:2606.31411v1 Announce Type: new Abstract: Rapid advancements in generative speech technology have compromised the reliability of voice biometrics. While current spoofing detectors excel when ass

LLM-Empowered Agentic MAC Protocols: A Dynamic Stackelberg Game Approach

SafetyDGX agent

arXiv:2510.10895v2 Announce Type: replace Abstract: Medium Access Control (MAC) protocols, essential for wireless networks, are typically manually configured. While deep reinforcement learning (DRL)-b

Locker-based Truck-Drone Routing with Integrated Considerations of Pickups, Deliveries, and No-Fly Zones

SafetyDGX agent

arXiv:2606.30680v1 Announce Type: cross Abstract: Truck-drone delivery is an emerging last-mile logistics mode combining the long-haul capacity of trucks with the flexible service capability of drones

Long-term Traffic Simulation via Structured Autoregressive Modeling

SafetyDGX agent

arXiv:2606.31209v1 Announce Type: new Abstract: Interactive traffic simulation is a vital world model for autonomous driving. A central challenge in long-horizon simulation is modeling sustained multi

MNAR-k-means: A k-means Clustering for Data Missing Not at Random with Magnitude-Decaying Probability

SafetyDGX agent

arXiv:2606.31253v1 Announce Type: cross Abstract: The classical k-means clustering, based on distances computed from all data features, cannot be directly applied to incomplete data with missing value

Multiple Testing of Linear Forms for Noisy Matrix Completion

SafetyDGX agent

arXiv:2312.00305v3 Announce Type: replace-cross Abstract: Many important tasks of large-scale recommender systems can be naturally cast as testing multiple linear forms for noisy matrix completion. Th

Multisensory Continual Learning: Adapting Pretrained Visuomotor Policies to Force

SafetyDGX agent

arXiv:2606.30988v1 Announce Type: new Abstract: Robot manipulation often relies on sensory feedback beyond vision, particularly in contact-rich settings where force, tactile, or audio signals reveal i

No Adaptation Without Observation: Observability-Constrained Test-Time Prompt Tuning for LiDAR Semantic Segmentation

SafetyDGX agent

arXiv:2606.30937v1 Announce Type: new Abstract: LiDAR semantic segmentation often degrades under real-world deployment due to evolving sensing conditions, while collecting new annotations for retraini

No country wins a race to superintelligence. An AI vastly smarter than us wouldn't be a weapon we control, we'd be handing control over to i…

SafetyDGX agent

No country wins a race to superintelligence. An AI vastly smarter than us wouldn't be a weapon we control, we'd be handing control over to it. ControlAI's US Director Connor Leahy (@NPCollapse) on why

Offline Reinforcement Learning for Fluid Controls: Data-based Multi-observational Policy Extraction

SafetyDGX agent

arXiv:2606.31025v1 Announce Type: new Abstract: Active flow control is a fundamental application in engineering. Recent advances in deep reinforcement learning have made progress in this field. Howeve

PA-VAD: Diffusion-Based Pseudo-Only Video Anomaly Detection via Domain-Aligned Memory Updates

SafetyDGX agent

arXiv:2512.06845v2 Announce Type: replace Abstract: Deploying video anomaly detection (VAD) in the real world is often constrained by the scarcity, privacy, and cost of collecting real abnormal footag

Paper2Rebuttal: A Multi-Agent Framework for Transparent Author Response Assistance

SafetyDGX agent

arXiv:2601.14171v2 Announce Type: replace Abstract: Writing effective rebuttals is a high-stakes task that demands more than linguistic fluency, as it requires precise alignment between reviewer inten

Personalizing Marketplace Policies with Competing Objectives and Constrained Experiments: Evidence from a Job Marketplace

SafetyDGX agent

arXiv:2606.30932v1 Announce Type: new Abstract: Two-sided marketplaces connect distinct user groups whose interests often conflict -- improving outcomes on one side could degrade the other side's expe

PiLoT v2: Pixel-to-Orthogonal Map Alignment for Free-view UAV Geo-localization

SafetyDGX agent

arXiv:2606.31098v1 Announce Type: new Abstract: Real-time, drift-free UAV geo-localization is essential for autonomous missions in GNSS-denied environments. The pioneering system, PiLoT, achieves high

Policy Optimization Achieves Data-Dependent Regret Bounds in MDPs with Unknown Transitions

SafetyDGX agent

arXiv:2606.31769v1 Announce Type: new Abstract: We study policy optimization for online episodic tabular Markov decision processes with unknown transition kernels, aiming for best-of-both-worlds guara

PolicyGuard: From Organizational Policies to Neuro-SymbolicCompliance Review Engines

SafetyDGX agent

arXiv:2606.32004v1 Announce Type: new Abstract: Policy-grounded document review requires determining whether a target document complies with organization-specific policies, guidelines, or playbooks. W

Pretrained Video Models as Differentiable Physics Simulators for Urban Wind Flows

Model ReleasesDGX agent

arXiv:2603.21210v3 Announce Type: replace Abstract: Designing urban spaces that provide pedestrian wind comfort and safety requires time-resolved Computational Fluid Dynamics (CFD) simulations, but th

Prompting Robot Teams with Natural Language

SafetyDGX agent

arXiv:2509.24575v2 Announce Type: replace-cross Abstract: This paper presents a framework to prompt multi-robot teams with high-level tasks using natural language expressions. Our objective is to use

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

SafetyDGX agent

arXiv:2606.32034v1 Announce Type: cross Abstract: LLM agents increasingly act over long horizons, where a single trajectory can contain hundreds or thousands of actions. In these settings, outcome-onl

ReGRPO: Reflection-Augmented Policy Optimization for Tool-Using Agents

SafetyDGX agent

arXiv:2606.31392v1 Announce Type: new Abstract: Tool-augmented vision-language models (VLMs) can solve multimodal, multi-step tasks by calling external tools, yet they remain fragile in practice. Exis

Reinforcement Learning-Based Control for an Inline Skating Humanoid Robot

SafetyDGX agent

arXiv:2606.31807v1 Announce Type: new Abstract: As humanoid robots become increasingly dynamic, coupling them with reinforcement learning offers a promising approach to solving the complex, underactua

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs

SafetyDGX agent

arXiv:2606.32032v1 Announce Type: cross Abstract: Metacognition is a critical component of intelligence that describes the ability to monitor and regulate one's own cognitive processes. Yet LLMs exhib

Reselling “excess AI processing capacity”: first SPCX, now META. In a rational world these moves would be seen as signs that we have *alre…

SafetyDGX agent

Reselling “excess AI processing capacity”: first SPCX, now META. In a rational world these moves would be seen as signs that we have *already* started to overbuild. Boom. $META is developing a cloud s

Resolving superposition in AI for interpretability and cross-modal alignment in patient-neuronal images

SafetyDGX agent

arXiv:2606.31394v1 Announce Type: cross Abstract: Artificial intelligence is transforming our capability to solve biological challenges. In dimensionality bottleneck regimes exacerbated by high-dimens

Rethinking On-policy Optimization for Query Augmentation

SafetyDGX agent

arXiv:2510.17139v3 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have led to a surge of interest in query augmentation for information retrieval (IR). Two main appro

Revisiting the Volume Hypothesis

SafetyDGX agent

arXiv:2606.31282v1 Announce Type: new Abstract: Modern deep neural networks often contain far more parameters than needed to fit their training data, yet they achieve impressive generalization. A comm

Rhythm-Structured Predictive Learning for Remote Photoplethysmography

SafetyDGX agent

arXiv:2606.31736v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) estimates physiological signals from facial videos by analyzing subtle pulse induced skin color variations. Despite r

Robust Text Watermarking for Large Language Models via Dual Semantic Embeddings

SafetyDGX agent

arXiv:2606.31602v1 Announce Type: new Abstract: This work presents Dual-Embedding Watermarking (DEW), a semantic watermarking scheme for large language models (LLMs) that leverages contextual and toke

Robustness of Robotic Manipulation: Foundations and Frontiers

SafetyDGX agent

arXiv:2606.31494v1 Announce Type: cross Abstract: Humans and animals exhibit remarkable robustness in physical manipulation, yet robots remain far behind. Progress toward human-level manipulation robu

Sampling-Based Coordination-Informed Multi-Objective Multi-Robot Reinforcement Learning

SafetyDGX agent

arXiv:2606.30893v1 Announce Type: new Abstract: Multi-robot systems must simultaneously optimize competing objectives while maintaining coordinated behavior. Existing multi-agent reinforcement learnin

Seeing Is Not Sharing: Some Vision-Language Models Overestimate Common Ground in Asymmetric Dialogue

SafetyDGX agent

arXiv:2606.31719v1 Announce Type: cross Abstract: In collaborative dialogue, shared perception does not guarantee shared interpretation. Mutual understanding must be established through interaction. W

Size Doesn't Matter: Cosine-Scored Sparse Autoencoders

SafetyDGX agent

arXiv:2606.15054v2 Announce Type: replace Abstract: Sparse autoencoders (SAEs) detect features via inner product, so a feature's activation scales with both its directional alignment and the input's n

Smart charging of large fleets of Electric Vehicles: Independent Multi-Agent Reinforcement Learning approaches

SafetyDGX agent

arXiv:2606.31347v1 Announce Type: new Abstract: The electrification of transportation through electric vehicles introduces new challenges for power grid management, such as increased peak demand, volt

SpectralSplats: Robust Differentiable Tracking via Spectral Moment Supervision

SafetyDGX agent

arXiv:2603.24036v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) enables real-time, photorealistic novel view synthesis, making it a highly attractive representation for model-based vi

Stabilization Learning: A Paradigm Transition Bridging Control Theory and Machine Learning

SafetyDGX agent

arXiv:2606.31562v1 Announce Type: new Abstract: Stabilization learning is an interdisciplinary paradigm that bridges control theory and machine learning. Its core idea is to enable systems to adjust t

Structural Preservation and the Logical Expressiveness of Graph Neural Networks

SafetyDGX agent

arXiv:2606.17882v2 Announce Type: replace Abstract: Bridges between graph neural networks (GNNs) and logical formalisms have been established by fixing architectural choices, such as the types of aggr

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation

SafetyDGX agent

arXiv:2606.30849v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have significantly advanced audio-driven portrait animation, but their high computational cost leads to substantial infere

TactX: Learning Shared Tactile Representations Across Diverse Sensors

SafetyDGX agent

arXiv:2606.31236v1 Announce Type: new Abstract: Tactile sensors provide critical information for contact-rich manipulation, yet tactile representations and policies remain tightly coupled to each spec

TDGT: A Tabular Data Generation Toolkit supporting adaptive GPU-accelerated Bayesian mixture models, diffusion-based models, and latent-space generative modeling

SafetyDGX agent

arXiv:2606.31268v1 Announce Type: cross Abstract: The growing demand for privacy-preserving data sharing has positioned synthetic data generation as a critical component of responsible AI workflows. D

Test-Time Verification for Text-to-SQL via Outcome Reward Models

SafetyDGX agent

arXiv:2606.30851v1 Announce Type: cross Abstract: Improving the reliability of large language models (LLMs) at inference time is a central challenge in structured reasoning tasks such as Text-to-SQL.

That day when @elonmusk starts to sound a bit like my 2023 TED Talk, which called for “a global, nonprofit organization to regulate the tech…

SafetyDGX agent

That day when @elonmusk starts to sound a bit like my 2023 TED Talk, which called for “a global, nonprofit organization to regulate the tech for the sake of democracy and our collective future.” Good

The Past Is Prologue: A Plug-in Controller for Selective Updates in Sequentially Evolving LLM Memory

SafetyDGX agent

arXiv:2606.31121v1 Announce Type: new Abstract: Sequentially evolving LLM memory enables agents to reuse past experience, but existing systems usually deploy each locally generated memory update witho

Token-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement Learning

SafetyDGX agent

arXiv:2606.31599v1 Announce Type: cross Abstract: Vision-language models (VLMs) combining reinforcement learning (RL) ignite remarkable progress in multimodal reasoning, yet still struggle with medica

TORA: Topological Representation Alignment for 3D Shape Assembly

SafetyDGX agent

arXiv:2604.04050v2 Announce Type: replace Abstract: Flow-matching methods for 3D shape assembly learn point-wise velocity fields that transport parts toward assembled configurations, yet they receive

Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation

SafetyDGX agent

arXiv:2606.31184v1 Announce Type: cross Abstract: Adaptive experiments for average treatment effects (ATE) require randomized allocations balancing valid inference with statistical efficiency. The ora

TreeAgent: A Generalizable Multi-Agent Framework for Automated Bias Labeling in Forestry via Compiled Expert Rules and Vision-Language Models

SafetyDGX agent

arXiv:2606.31976v1 Announce Type: new Abstract: Human-labeled data are widely used as reference annotations in ML, despite known variability across annotators in many expert-driven domains. In additio

TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.32017v1 Announce Type: cross Abstract: Agentic reinforcement learning requires assigning credit to environment-facing actions such as searches, clicks, edits, navigation commands, and objec

UniTac: A Unified Multimodal Model for Cross-Sensor Tactile Understanding and Generation

SafetyDGX agent

arXiv:2606.31451v1 Announce Type: cross Abstract: Unified multimodal models (UMMs) have shown great promise in integrating understanding and generation across diverse modalities. However, existing res

Unsupervised Data-Efficient Cross-Modal Retrieval with Global-Neighborhood Alignment Hashing

SafetyDGX agent

arXiv:2606.31517v1 Announce Type: cross Abstract: Compared to supervised cross-modal hashing (CMH), unsupervised CMH reduces the reliance on manual labeling by learning binary codes from unlabeled ima

VIGOR: VIdeo Geometry-Oriented Reward for Temporal Generative Alignment

SafetyDGX agent

arXiv:2603.16271v3 Announce Type: replace Abstract: Video diffusion models lack explicit geometric supervision during training, leading to inconsistency artifacts such as object deformation, spatial d

Vision-Language Procedural Reasoning for Context-Aware Reward Modeling of Robotic Endovascular Guidewire Navigation

SafetyDGX agent

arXiv:2606.30698v1 Announce Type: new Abstract: Robotic-assisted endovascular interventions demand accurate, stable, and context-aware guidewire navigation in complex and patient-specific vascular ana

Wait, am I Being Fair? Characterizing Deductive Stereotyping and Mitigating It with Fair-GCG

SafetyDGX agent

arXiv:2606.30989v1 Announce Type: cross Abstract: Warning: This paper contains several toxic and offensive statements. While reasoning generally improves fairness in recent large language models (LLMs

Warp RL: Reshaping Base Policy Distributions for Dynamics Adaptation

SafetyDGX agent

arXiv:2606.31043v1 Announce Type: new Abstract: Residual reinforcement learning adapts a pretrained robot policy by learning an additive correction to its actions. While effective when adaptation amou

What Probing Reveals about Autonomous Driving: Linking Internal Prediction Errors to Ego Planning

SafetyDGX agent

arXiv:2606.31106v1 Announce Type: cross Abstract: Large-scale datasets and fast simulators have enabled improvements in driving policies that appear safe and robust, yet strong performance in nominal

Will the AI bubble pop is the wrong question. We partnered with Damon Cassidy, a video essayist with ~300k subscribers: even if it pops, the…

SafetyDGX agent

Will the AI bubble pop is the wrong question. We partnered with Damon Cassidy, a video essayist with ~300k subscribers: even if it pops, the race to uncontrollable superintelligence doesn't go away. P

30 Jun 2026

A causal modeling perspective on decision theory

SafetyDGX agent

arXiv:2606.29911v1 Announce Type: new Abstract: Decision theory provides a formal framework for how agents should make choices under uncertainty, drawing on ideas from philosophy, probability, and cau

← Previous
1…9394959697…242
Next →