AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
2 Jul 2026

Dataset Biases and Shortcut Learning in Motion-Based AI-Generated Video Detection

SafetyDGX agent

arXiv:2607.00948v1 Announce Type: new Abstract: The visual quality of AI-generated videos has improved drastically in recent years, making it increasingly difficult for humans to distinguish between r

Decentralized Geometric Control for Cable-Suspended Payload Transport with Adaptive Mass Estimation

SafetyDGX agent

arXiv:2607.00024v1 Announce Type: new Abstract: Cooperative aerial transport requires controllers that respect nonlinear manifold geometry, operate without centralized coordination, and respect operat

Diffusion-GR2: Diffusion Generative Reasoning Re-ranker

SafetyDGX agent

arXiv:2607.01170v1 Announce Type: cross Abstract: Generative reasoning re-rankers achieve strong recommendation accuracy by emitting a chain-of-thought before re-ordering a candidate list, but they ar


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Distill to Detect: Exposing Stealth Biases in LLMs through Cartridge Distillation

SafetyDGX agent

arXiv:2607.01208v1 Announce Type: cross Abstract: Language models deployed in high-stakes roles can potentially favor certain entities, brands, or viewpoints, steering user decisions at scale. Such pr

Distributed Multi Robot Lunar Cargo Transportation via Phase Decomposed Reinforcement Learning

SafetyDGX agent

arXiv:2607.00160v1 Announce Type: new Abstract: Modular reconfigurable robotic systems provide a scalable solution for cooperative surface operations in future lunar missions. However, cooperative car

Domain Arithmetic: One-Shot VLA Adaptation under Environmental Shifts

SafetyDGX agent

arXiv:2607.00666v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models often fail to perform the same learned tasks under environmental shifts, such as changes in camera pose and shifts

Dual-Informed Vertical Expansion for Multi-Objective Node Selection in Anytime Conflict-Based Search

SafetyDGX agent

arXiv:2607.00156v1 Announce Type: new Abstract: Conflict-Based Search (CBS) is a leading exact algorithm for Multi-Agent Path Finding (MAPF), but its high-level node-selection rule is usually treated

ECoSim: Data Efficient Fine-Tuning for Controllable Traffic Simulation

SafetyDGX agent

arXiv:2607.00545v1 Announce Type: new Abstract: Controllable traffic simulation is critical for testing autonomous driving systems, yet existing approaches often require retraining large generative mo

ELMP: Efficient Learning for Motion Planning via Analytical Policy Gradients

SafetyDGX agent

arXiv:2607.00215v1 Announce Type: new Abstract: Neural Motion Planners (NMPs) enable fast reactive motion generation, but adapting them to new environments typically requires recollecting large expert

Emails disclosed in a court filing detail the uneasy back-and-forth between Dario Amodei and DOD's Emil Michael and how Anthropic's relationship with DOD soured (Wall Street Journal)

SafetyDGX agent

Wall Street Journal: Emails disclosed in a court filing detail the uneasy back-and-forth between Dario Amodei and DOD's Emil Michael and how Anthropic's relationship with DOD soured — Undersecretary E

Enhancing Hardware Fault Tolerance in Machines with Reinforcement Learning Policy Gradient Algorithms

SafetyDGX agent

arXiv:2407.15283v2 Announce Type: replace-cross Abstract: Industry is moving toward autonomous, network-connected machines that detect and adapt to changing conditions, including hardware faults. Conv

EPO: Boosting 3D Foundation Models with Edge-based Pose Optimization

SafetyDGX agent

arXiv:2607.00579v1 Announce Type: new Abstract: We introduce extbf{Edge-based Pose Optimization (EPO)}, a trackless geometric optimization framework specifically designed to boost the Structure-from-M

EquiSteer: Cross-Attention Steering Towards a Fairer Text-Guided Image Generation

SafetyDGX agent

arXiv:2607.01147v1 Announce Type: new Abstract: Text-to-image diffusion models power everyday creative tasks, but they still reproduce the demographic biases in their training data. On common prompts

Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles

SafetyDGX agent

arXiv:2511.06160v2 Announce Type: replace Abstract: While recent safety guardrails effectively suppress overtly biased outputs, subtler forms of social bias emerge during complex logical reasoning tas

Exploring the Semantic Gap in Agentic Data Systems: A Formative Study of Operationalization Failures in Analytical Workflows

SafetyDGX agent

arXiv:2607.00828v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate queries, invoke tools, and construct analytical workflows. Although recent advances hav

FAR: Failure-Aware Retry for Test-Time Recovery and Continual Policy Improvement

SafetyDGX agent

arXiv:2607.01111v1 Announce Type: cross Abstract: Robot policies inevitably encounter failures when deployed in real environments. Naive retries often repeat the same mistakes, while many existing rec

FastBridge: Closing the Model-Based Realization Gap in Safety Filters on 3D Gaussian Splatting for Fast Quadrotor Flight

SafetyDGX agent

arXiv:2607.01200v1 Announce Type: new Abstract: Fast quadrotor flight requires safe obstacle avoidance under tight onboard compute limits. While 3D Gaussian Splatting (3DGS) provides a continuous, geo

ForAug: Mitigating Biases in Image Classification via Controlled Image Compositions

SafetyDGX agent

arXiv:2503.09399v4 Announce Type: replace-cross Abstract: Large-scale image classification datasets exhibit strong compositional biases: objects tend to be centered, appear at characteristic scales, a

FrameONE: Hierarchical Motion Modeling for Universal Multi-View Echocardiographic Keyframe Detection

SafetyDGX agent

arXiv:2607.00748v1 Announce Type: new Abstract: Accurate detection of end-systole (ES) and end-diastole (ED) frames is fundamental to echocardiographic assessment. Existing methods are typically devel

From Holistic Evaluation to Structured Criteria: Rubrics Across the Evolving LLM Landscape

SafetyDGX agent

arXiv:2606.08625v2 Announce Type: replace Abstract: As Large Language Models (LLMs) advance toward open-ended autonomous agents, the mechanisms used to evaluate and guide their behavior must evolve ac

From Pixels to Temporal Correlations: Learning Informative Representations for Reinforcement Learning Pre-training

SafetyDGX agent

arXiv:2607.00811v1 Announce Type: new Abstract: Unsupervised pre-training on large-scale datasets has demonstrated significant potential for improving the sample efficiency and performance of Reinforc

From Prediction Uncertainty to Conformalized Distance Fields for Safe Motion Planning

SafetyDGX agent

arXiv:2607.00776v1 Announce Type: new Abstract: Safe motion planning in dynamic environments requires reasoning about the uncertainty in predicted obstacle motion without sacrificing real-time perform

From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning

SafetyDGX agent

arXiv:2603.10263v2 Announce Type: replace-cross Abstract: We introduce Distribution Contractive Reinforcement Learning (DICE-RL), a framework that uses reinforcement learning (RL) as a 'distribution c

From Silos to Systems: Process-Oriented Hazard Analysis for AI Systems

SafetyDGX agent

arXiv:2410.22526v2 Announce Type: replace Abstract: To effectively address potential harms from Artificial Intelligence (AI) systems, it is essential to identify and mitigate system-level hazards. Cur

From World Models to World Action Models: A Concise Tutorial for Robotics

SafetyDGX agent

arXiv:2607.00836v1 Announce Type: cross Abstract: World models are increasingly used in embodied intelligence and generative simulation, yet their scope remains ambiguous across communities. This tuto

Gauging, Measuring, and Controlling Critic Complexity in Actor-Critic Reinforcement Learning

SafetyDGX agent

arXiv:2607.00452v1 Announce Type: cross Abstract: Actor-critic methods depend on learned critics, but critic quality is often evaluated only indirectly through return, temporal-difference error, or va

GaussianFusion: Unified 3D Gaussian Representation for Multi-Modal Fusion Perception

SafetyDGX agent

arXiv:2607.00746v1 Announce Type: cross Abstract: The bird's-eye view (BEV) representation enables multi-sensor features to be fused within a unified space, serving as the primary approach for achievi

Google’s Continued Disruption of Malicious Residential Proxy Networks

SafetyDGX agent

Background Today, in coordination with the FBI, Lumen, and others, Google took action against the NetNut residential proxy network, also known as Popa. This action builds on our disruption of the IPID

Graph-Native Reinforcement Learning Enables Traceable Scientific Hypothesis Generation through Conceptual Recombination

SafetyDGX agent

arXiv:2607.00924v1 Announce Type: new Abstract: Accelerating materials discovery requires AI systems that can generate scientifically valid hypotheses through multi-step, domain-grounded reasoning. St

GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation Identity

SafetyDGX agent

arXiv:2607.00152v1 Announce Type: cross Abstract: Three of the most popular methods for training language models to reason look like three different tricks. They are not. All three adjust a single num

HARC: Coupling Harmfulness and Refusal Directions for Robust Safety Alignment

SafetyDGX agent

arXiv:2607.00572v1 Announce Type: new Abstract: Understanding how aligned LLMs internally represent safety is critical for diagnosing alignment vulnerabilities, as it explains why jailbreaks succeed a

How Environment and Urbanization Shape Bird Diversity in Sri Lanka

SafetyDGX agent

arXiv:2607.00582v1 Announce Type: cross Abstract: This study presents a comprehensive analysis of bird diversity across Sri Lanka by integrating spatial, temporal, and environmental data. Bird observa

HyFL-CLIP: Hyperbolic Fine-Tuning of CLIP for Robust Long-Context Understanding

SafetyDGX agent

arXiv:2607.00428v1 Announce Type: new Abstract: CLIP (Contrastive Language-Image Pre-training) has become a de facto paradigm for image-text alignment, but it struggles with long-context descriptions

Identifying Latent Concepts and Structures for Generalized Category Discovery

SafetyDGX agent

arXiv:2607.00620v1 Announce Type: cross Abstract: Generalized Category Discovery (GCD) aims to recognize known classes while autonomously discovering novel ones in open-world settings. However, curren

in case you needed more evidence that tokenmaxxing was a stupid idea

SafetyDGX agent

in case you needed more evidence that tokenmaxxing was a stupid idea Meta burns 2.65B a year on AI tokens. at 300K for a Meta engineer, that's enough to pay ~9,000 engineers for a full year. now ask y

KungfuBot: Physics-Based Humanoid Whole-Body Control for Learning Highly-Dynamic Skills

SafetyDGX agent

arXiv:2506.12851v3 Announce Type: replace-cross Abstract: Humanoid robots are promising to acquire various skills by imitating human behaviors. However, existing algorithms are only capable of trackin

Learning Category-level Last-meter Navigation from RGB Demonstrations of a Single-instance

SafetyDGX agent

arXiv:2512.11173v4 Announce Type: replace Abstract: Achieving precise positioning of the mobile manipulator's base is essential for successful manipulation actions that follow. Most of the RGB-based n

Learning Expert Strategy for Autonomous Robotic Endovascular Intervention via Decoupled Procedural Execution

SafetyDGX agent

arXiv:2607.00066v1 Announce Type: new Abstract: Endovascular interventions are high-stakes procedures requiring precise device operation within complex and tortuous vascular anatomies. Autonomous endo

Learning from Demonstration via Spatiotemporal Tubes for Unknown Euler-Lagrange Systems

SafetyDGX agent

arXiv:2607.00534v1 Announce Type: new Abstract: We present STT-LfD, a unified Learning from Demonstration (LfD) framework that integrates motion learning with control for unknown Euler-Lagrange system

Learning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications

SafetyDGX agent

arXiv:2607.00442v1 Announce Type: cross Abstract: Reinforcement learning (RL) for quadruped locomotion commonly depends on fixed, hand-crafted, and Markovian reward functions that limit both interpret

Learning Generalizable Skill Policy with Data-Efficient Unsupervised RL

SafetyDGX agent

arXiv:2607.00392v1 Announce Type: cross Abstract: Unsupervised Reinforcement Learning (URL) aims to pre-train scalable, skill-conditioned policies without extrinsic rewards, serving as a foundation fo

Learning User-Aware Recall: Personalized Retrieval in Long-Term Conversational Memory

SafetyDGX agent

arXiv:2607.00017v1 Announce Type: cross Abstract: Long-term conversational agents are expected to remember past interactions, but memory is useful only when the right evidence is recalled for the righ

LIST3R: Long-sequence Instance-aware 3D Reconstruction

SafetyDGX agent

arXiv:2607.00375v1 Announce Type: new Abstract: We present LIST3R, an instance-aware framework for long-sequence 3D reconstruction inspired by the way humans organize spatial memory around stable and

Managed Autonomy at Runtime: Gear-Based Safety and Governance for Single- and Multi-Agent Cyber-Physical Systems

SafetyDGX agent

arXiv:2607.00334v1 Announce Type: new Abstract: Autonomous agents, whether LLM-driven software agents or robotic physical agents, face a common class of failure modes when operating without continuous

Manifold-constrained Hamilton-Jacobi Reachability Learning for Decentralized Multi-Agent Motion Planning

SafetyDGX agent

arXiv:2511.03591v2 Announce Type: replace Abstract: Safe multi-agent motion planning (MAMP) under task-induced constraints is a critical challenge in robotics. Many real-world scenarios require robots

Measuring Dead Directions: Decomposing and Classifying Singular Structure off Canonical Alignment

SafetyDGX agent

arXiv:2607.00603v1 Announce Type: new Abstract: We give a descent-free, alignment-free measurement of singular structure on trained networks. At a single frozen checkpoint the read recovers the order

MedCAGD: Context-Aware Gated Decoder for Efficient Medical Image Segmentation

SafetyDGX agent

arXiv:2607.00409v1 Announce Type: new Abstract: Medical image segmentation relies on the ability of encoder-decoder architectures to translate rich feature representations into accurate pixel-level pr

Meta-Transfer Learning for mmWave Beam Alignment

SafetyDGX agent

arXiv:2607.00860v1 Announce Type: cross Abstract: Millimeter-wave (mmWave) beam alignment plays a critical role in next-generation wireless systems, yet its efficient implementation remains challengin

Mirror-Fusion Attention for Reflection-Aware Self-Supervised Representation Learning

SafetyDGX agent

arXiv:2607.00850v1 Announce Type: new Abstract: Most self-supervised learning (SSL) methods encourage invariance across augmentations, but strict flip invariance can suppress informative left--right c

Mnemosyne: Agentic Transaction Processing for Validating and Repairing AI-generated Workflows

SafetyDGX agent

arXiv:2607.00269v1 Announce Type: new Abstract: LLMs, solvers, and agent teams increasingly generate workflow actions, repairs, and plans, but a generated action may be syntactically valid yet stale,

MoVA: Learning Asymmetric Dual Projections for Modular Long Video-Text Alignment

SafetyDGX agent

arXiv:2607.00858v1 Announce Type: new Abstract: Contrastive pre-training has propelled video-text alignment, yet models often inherit the critical limitations of their image-text predecessors like CLI

Multi-scale Mixture of World Models for Embodied Agents in Evolving Environments

SafetyDGX agent

arXiv:2607.00457v1 Announce Type: new Abstract: Embodied agents operating in the real world require multi-scale reasoning and knowledge adaptation as conditions change. We identify two challenges in a

Multiplicity is an Inevitable and Inherent Challenge in Multimodal Learning

SafetyDGX agent

arXiv:2505.19614v2 Announce Type: replace-cross Abstract: Multimodal learning has seen remarkable progress, particularly with large-scale pre-training across various modalities. Most current approache

NeHMO: Neural Hamilton-Jacobi Reachability Learning for Decentralized Safe Multi-Arm Motion Planning

SafetyDGX agent

arXiv:2607.00326v1 Announce Type: new Abstract: Safe multi-arm motion planning is a challenging problem in robotics due to its high dimensionality, coupled configuration space, and complex collision c

NeuroCogMap Reveals Cognitive Organization of Large Language Models

SafetyDGX agent

arXiv:2607.00397v1 Announce Type: cross Abstract: Understanding how complex cognitive functions are organized within artificial systems is central to interpreting large language models (LLMs) and rela

NEW paper from NVIDIA. They discuss robot programming that compounds experience instead of throwing it away. Traditional robot programming f…

SafetyDGX agent

NEW paper from NVIDIA. They discuss robot programming that compounds experience instead of throwing it away. Traditional robot programming forces you to orchestrate perception, contact dynamics, diver

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping

SafetyDGX agent

arXiv:2607.00881v1 Announce Type: new Abstract: Spatial intelligence remains a persistent challenge for Multimodal Large Language Models (MLLMs), as it requires coherent spatial scene representations

OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning

SafetyDGX agent

arXiv:2510.24636v3 Announce Type: replace Abstract: Reward models (RMs) have become essential for aligning large language models (LLMs), serving as scalable proxies for human evaluation in both traini

PAPA: Online Personalized Active Preference Alignment

SafetyDGX agent

arXiv:2607.00486v1 Announce Type: cross Abstract: Diffusion models are highly effective at modeling complex data distributions, including images and text. However, in applications like personalized re

Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising

SafetyDGX agent

arXiv:2607.00407v1 Announce Type: new Abstract: Slide design requires personalizing both deck themes and page layouts. Yet, current AI agent-based methods struggle with fine-grained, page-level design

← Previous
1…4849505152…212
Next →