AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning

DGX agent

arXiv:2606.12195v1 Announce Type: new Abstract: Recent progress in foundation models has shifted toward agentic behavior involving multi-step reasoning and tool use. However, open-source efforts large

safetyarxiv-cs-cv
11 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ISAP-3D: Identity-Slot Aligned Part-Aware 3D Generation

DGX agent

arXiv:2606.12099v1 Announce Type: new Abstract: Part-aware 3D generation aims to synthesize structured objects with semantically meaningful components, yet often suffers from structural ambiguity due

safetyarxiv-cs-cv
11 Jun 2026
Safety

KinematicRL: A Sim-to-Real Reinforcement Learning Framework For Social Navigation With Kinodynamic Feasibility

DGX agent

arXiv:2606.12042v1 Announce Type: new Abstract: Deep Reinforcement Learning (DRL) has shown promise for social navigation, yet its real-world deployment remains hindered by a persistent sim-to-real ga

safetyarxiv-cs-ro
11 Jun 2026
Safety

LAST: Bridging Vision-Language and Action Manifolds via Gromov-Wasserstein Alignment

DGX agent

arXiv:2606.11221v1 Announce Type: new Abstract: We take a Gromov-Wasserstein perspective on Vision-Language-Action (VLA) learning, where the goal is to make the relational geometry of action represent

safetyarxiv-cs-cv
11 Jun 2026
Safety

Latent World Recovery for Multimodal Learning with Missing Modalities

DGX agent

arXiv:2606.12362v1 Announce Type: cross Abstract: We study multimodal learning under missing modalities, with particular motivation from bioscience applications in which heterogeneous modalities are o

safetyarxiv-cs-ai
11 Jun 2026
Safety

Learning Instance-Adaptive Low-Rank Orthogonal Subspaces for Clothes-Changing Person Re-Identification

DGX agent

arXiv:2606.11661v1 Announce Type: new Abstract: Clothes-changing person re-identification (CC-ReID) aims to recognize individuals despite drastic appearance changes caused by clothing variation. While

safetyarxiv-cs-cv
11 Jun 2026
Safety

Learning to Inject: Automated Prompt Injection via Reinforcement Learning

DGX agent

arXiv:2602.05746v2 Announce Type: replace-cross Abstract: Prompt injection is a critical vulnerability in LLM agents, yet the strongest methods still rely on human red-teamers and hand-crafted prompts

safetyarxiv-cs-ai
11 Jun 2026
Safety

Learning What to Say to Your VLA: Mostly Harmless Vision Language Action Model Steering

DGX agent

arXiv:2606.12299v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models provide a natural language interface to robot control, but the mapping from language to behavior is often brittle

safetyarxiv-cs-lg
11 Jun 2026
Safety

LUCID: Learning Embodiment-Agnostic Intent Models from Unstructured Human Videos for Scalable Dexterous Robot Skill Acquisition

DGX agent

arXiv:2606.11628v1 Announce Type: cross Abstract: The most widely-adopted robot learning pipelines today learn skills from robot demonstrations or structured human data, which are expensive to collect

safetyarxiv-cs-ai
11 Jun 2026
Safety

MASK: Multi-Agent Semantic K-Scheduling for Risk-Sensitive 6G Robotics

DGX agent

arXiv:2606.11249v1 Announce Type: cross Abstract: Realizing the vision of 6G connected robotics requires reconciling high-performance collaborative control with the rigid spectral limitations of physi

safetyarxiv-cs-lg
11 Jun 2026
Safety

Measuring Semantic Progress in Multi-turn Dialogue via Information Gain

DGX agent

arXiv:2606.12332v1 Announce Type: new Abstract: Evaluating multi-turn dialogue is challenging because quality emerges across turns rather than within individual responses. We focus on a key dimension

safetyarxiv-cs-cl
11 Jun 2026
Safety

Mirror Descent Beyond Euclidean Stability: An Exponential Separation in Initialization Sensitivity

DGX agent

arXiv:2606.11431v1 Announce Type: new Abstract: Mirror Descent (MD) extends Gradient Descent (GD) beyond Euclidean geometry and has recently reappeared as a lens for KL-regularized policy optimization

safetyarxiv-cs-lg
11 Jun 2026
Safety

Mitigating Disparate Impact of Differentially Private Learning through Bounded Adaptive Clipping

DGX agent

arXiv:2506.01396v2 Announce Type: replace Abstract: Differential privacy (DP) has become an essential framework for privacy-preserving machine learning. Existing DP learning methods, however, often ha

safetyarxiv-cs-lg
11 Jun 2026
Safety

MLT-Dedup: Efficient Large-Scale Online Video Deduplication via Multi-Level Representations and Spatial-Temporal Matching

DGX agent

arXiv:2606.12215v1 Announce Type: new Abstract: The explosive growth of user-generated video content on online platforms is accompanied by the emergence of numerous near-duplicate videos--videos that

safetyarxiv-cs-cv
11 Jun 2026
Safety

Noise-Guided Transport for Imitation Learning

DGX agent

arXiv:2509.26294v2 Announce Type: replace-cross Abstract: We consider imitation learning in the low-data regime, where only a limited number of expert demonstrations are available. In this setting, me

safetyarxiv-cs-ai
11 Jun 2026
Safety

Offline Diffusion Policy for Multi-User Delay-Constrained Scheduling

DGX agent

arXiv:2501.12942v2 Announce Type: replace Abstract: Effective multi-user delay-constrained scheduling is crucial in various real-world applications, including embodied AI, instant messaging, live stre

safetyarxiv-cs-ai
11 Jun 2026
Safety

Open Materials Generation with Inference-Time Reinforcement Learning

DGX agent

arXiv:2602.00424v2 Announce Type: replace Abstract: Continuous-time generative models for crystalline materials enable inverse materials design by learning to predict stable crystal structures, but in

safetyarxiv-cs-lg
11 Jun 2026
Safety

PAWS: Preference Learning with Advantage-Weighted Segments

DGX agent

arXiv:2606.11982v1 Announce Type: new Abstract: Preference-based reinforcement learning (PbRL) learns policies from human trajectory-level comparisons, avoiding explicit reward design and expert demon

safetyarxiv-cs-lg
11 Jun 2026
Safety

Plan-and-Verify Video Reward Reasoning with Spatio-Temporal Scene Graph Grounding

DGX agent

arXiv:2606.11838v1 Announce Type: new Abstract: Reward models for text-to-video (T2V) generation guide post-training but often fail at fine-grained semantic alignment. We trace this to two structural

safetyarxiv-cs-cv
11 Jun 2026
Safety

ProcessThinker: Enhancing Multi-modal Large Language Models Reasoning via Rollout-based Process Reward

DGX agent

arXiv:2606.11209v1 Announce Type: cross Abstract: Visual question answering increasingly requires multi-step reasoning. Recent post-training with reinforcement learning under verifiable rewards (RLVR)

safetyarxiv-cs-ai
11 Jun 2026
Safety

Redesign Mixture-of-Experts Routers with Manifold Power Iteration

DGX agent

arXiv:2606.12397v1 Announce Type: cross Abstract: Router is the cornerstone component to the Mixture-of-Experts models. Serving as expert proxies, the rows of the router matrix compute their similarit

safetyarxiv-cs-ai
11 Jun 2026
Safety

Reinforcement Learning Disrupts Gradient-Based Adversarial Optimization

DGX agent

arXiv:2606.12251v1 Announce Type: cross Abstract: Gradient-based adversarial attacks remain a dominant threat to deep neural networks (DNNs), as they exploit gradient information to efficiently optimi

safetyarxiv-cs-ai
11 Jun 2026
Safety

Reinforcement Learning with Action-Triggered Observations

DGX agent

arXiv:2510.02149v2 Announce Type: replace Abstract: We introduce Action-Triggered Sporadically Traceable Markov Decision Processes (ATST-MDPs), a reinforcement learning framework for partial observabi

safetyarxiv-cs-lg
11 Jun 2026
Safety

Reverse Flow Matching: A Unified Framework for Online Reinforcement Learning with Diffusion and Flow Policies

DGX agent

arXiv:2601.08136v2 Announce Type: replace Abstract: Diffusion and flow policies are gaining prominence in online reinforcement learning (RL) due to their expressive power, yet training them efficientl

safetyarxiv-cs-lg
11 Jun 2026
Safety

RLCSD: Reinforcement Learning with Contrastive On-Policy Self-Distillation

DGX agent

arXiv:2606.11709v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) provides dense, token-level supervision for reasoning models by aligning a model's own distribution with the distri

safetyarxiv-cs-cl
11 Jun 2026
Safety

SAGE: Answer-Conditioned Uncertainty Targets for Verbal Uncertainty Alignment

DGX agent

arXiv:2606.11512v1 Announce Type: new Abstract: Large language models increasingly express uncertainty through natural-language statements, yet these expressions often fail to reflect the model's samp

safetyarxiv-cs-cl
11 Jun 2026
Safety

Sample-Efficient Hypergradient Estimation for Decentralized Bi-Level Reinforcement Learning

DGX agent

arXiv:2603.14867v4 Announce Type: replace-cross Abstract: Many strategic decision-making problems, such as environment design for warehouse robots, can be naturally formulated as bi-level reinforcemen

safetyarxiv-cs-ai
11 Jun 2026
Safety

Signed Compression Progress on a Sealed Audit is Goodhart-Resistant

DGX agent

arXiv:2606.11417v1 Announce Type: cross Abstract: Compression progress is a long-standing proposal for intrinsic motivation: reward an agent when its world model becomes better at predicting or compre

safetyarxiv-cs-ai
11 Jun 2026
Safety

SIL: Symbiotic Interactive Learning for Language-Conditioned Human-Agent Co-Adaptation

DGX agent

arXiv:2511.05203v3 Announce Type: replace Abstract: Today's autonomous agents, largely driven by foundation models (FMs), can understand natural language instructions and solve long-horizon tasks with

safetyarxiv-cs-ro
11 Jun 2026
Safety

Sovereign Assurance Boundary: Certificate-Bound Admission for Agentic Infrastructure

DGX agent

arXiv:2606.11632v1 Announce Type: cross Abstract: Agentic infrastructure introduces a critical control-plane authorization problem: non-deterministic reasoning systems can propose high-stakes mutation

safetyarxiv-cs-ai
11 Jun 2026
Safety

Spectrally Regularized Latent Flow Matching for Turbulence Generation

DGX agent

arXiv:2606.11691v1 Announce Type: new Abstract: Latent diffusion and flow matching have emerged as leading approaches for synthetic turbulence generation, yet they systematically under-represent dissi

safetyarxiv-cs-lg
11 Jun 2026
Safety

Steering Multirobot Behavior via Closed-Loop Affine Activation Editing

DGX agent

arXiv:2606.11489v1 Announce Type: new Abstract: Real-world robots need to adapt their behavior beyond the envelope of their pre-trained policy. Policy finetuning or retraining are options, but they ri

safetyarxiv-cs-ro
11 Jun 2026
Safety

SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning

DGX agent

arXiv:2606.11770v1 Announce Type: new Abstract: Spatial reasoning remains a challenge for Multimodal Large Language Models (MLLMs), as it requires reliable multi-hop inference over both intermediate s

safetyarxiv-cs-ai
11 Jun 2026
Safety

TacCoRL: Integrating Tactile Feedback into VLA via Simulation

DGX agent

arXiv:2606.11743v1 Announce Type: cross Abstract: Vision-language-action (VLA) models provide strong visual, language, and action priors for robot manipulation, but visual observations alone often mis

safetyarxiv-cs-lg
11 Jun 2026
Safety

Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning

DGX agent

arXiv:2606.11853v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) depend on in-context learning (ICL) for rapid task adaptation, but their scalability is severely limited by

safetyarxiv-cs-ai
11 Jun 2026
Safety

The Art of Interrogation: Consistency Amplifies Factuality in Spatial Reasoning

DGX agent

arXiv:2606.11918v1 Announce Type: new Abstract: Current Large Reasoning Models (LRMs) exhibit remarkable general capabilities but significantly underperform in spatial reasoning tasks. Existing approa

safetyarxiv-cs-ai
11 Jun 2026
Safety

The Unreasonable Effectiveness of Discrete-Time Gaussian Process Mixtures for Robot Policy Learning

DGX agent

arXiv:2505.03296v2 Announce Type: replace-cross Abstract: We present Mixture of Discrete-time Gaussian Processes (MiDiGap), a novel approach for flexible policy representation and imitation learning i

safetyarxiv-cs-ai
11 Jun 2026
Safety

To Intervene or Not: Guiding Inference-time Alignment with Probabilistic Model Blending

DGX agent

arXiv:2606.11201v1 Announce Type: cross Abstract: The wide deployment of LLMs has made model alignment necessary to make newly trained models safely and effectively respond to user instructions. Among

safetyarxiv-cs-ai
11 Jun 2026
Safety

Toward Preference-aligned Large Language Models via Residual-based Model Steering

DGX agent

arXiv:2509.23982v2 Announce Type: replace-cross Abstract: Preference alignment is a critical step in making Large Language Models (LLMs) useful and aligned with (human) preferences. Existing approache

safetyarxiv-cs-ai
11 Jun 2026
Safety

Towards a Bridge Layer Between Bibliographic and Formalized Mathematical Knowledge

DGX agent

arXiv:2606.11430v1 Announce Type: cross Abstract: Mathematical knowledge is split between bibliographic databases (e.g., MathSciNet, zbMATH Open) and formal proof libraries (e.g., Lean mathlib), preve

safetyarxiv-cs-ai
11 Jun 2026
Safety

Towards Conditional Feature Alignment for Cross-Domain Counting

DGX agent

arXiv:2506.17137v3 Announce Type: replace Abstract: Object counting models often degrade under cross-domain deployment because density composition varies across domains and is itself task-relevant. St

safetyarxiv-cs-cv
11 Jun 2026
Safety

Traits Run Deeper: Trait-Specific Asymmetric Fusion for Personality Assessment

DGX agent

arXiv:2606.11269v1 Announce Type: new Abstract: Personality assessment aims to infer stable personality traits from dynamic behaviors across language, voice, and facial cues. Since different personali

safetyarxiv-cs-cv
11 Jun 2026
Safety

UniIntervene: Agentic Intervention for Efficient Real-World Reinforcement Learning

DGX agent

arXiv:2606.12372v1 Announce Type: cross Abstract: Human-in-the-loop reinforcement learning (HiL-RL) has emerged as an effective paradigm for real-world robotic manipulation, enabling online policy imp

safetyarxiv-cs-lg
11 Jun 2026
Safety

UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA

DGX agent

arXiv:2606.11740v1 Announce Type: cross Abstract: We study whether grounded reasoning supervision from abundant 2D medical images can improve 3D medical VQA when both input types are aligned through a

safetyarxiv-cs-cl
11 Jun 2026
Safety

UR-BERT: Scaling Text Encoders for Massively Multilingual TTS Through Universal Romanization and Speech Token Prediction

DGX agent

arXiv:2606.11681v1 Announce Type: new Abstract: We propose UR-BERT, a Romanized transcription-based text-to-speech (TTS) encoder for massively multilingual TTS systems. Conventional grapheme-to-phonem

safetyarxiv-cs-cl
11 Jun 2026
Safety

Urban Heat MiniCubes: An AI-Ready dataset for urban heat research

DGX agent

arXiv:2606.11534v1 Announce Type: cross Abstract: Urban heat is amplified by impermeable surfaces and heterogeneous built environments, yet street-level variability remains difficult to quantify becau

safetyarxiv-cs-lg
11 Jun 2026
Safety

ViT-FREE: Efficient Face Recognition via Early Exiting and Synthetic Adaptation

DGX agent

arXiv:2606.12023v1 Announce Type: new Abstract: Vision Transformers (ViTs) have gained significant attention in computer vision and shown strong potential for face recognition (FR). However, their hig

safetyarxiv-cs-cv
11 Jun 2026
Safety

VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving

DGX agent

arXiv:2606.12396v1 Announce Type: new Abstract: Vision-language-action (VLA) models can describe scenes and reason about them in language, yet still struggle to ground their actions in the dense 3D wo

safetyarxiv-cs-cv
11 Jun 2026
← Previous
1…123124125126127…260
Next →