AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

DGX agent

arXiv:2606.32034v1 Announce Type: cross Abstract: LLM agents increasingly act over long horizons, where a single trajectory can contain hundreds or thousands of actions. In these settings, outcome-onl

safetyarxiv-cs-ai
1 Jul 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ReGRPO: Reflection-Augmented Policy Optimization for Tool-Using Agents

DGX agent

arXiv:2606.31392v1 Announce Type: new Abstract: Tool-augmented vision-language models (VLMs) can solve multimodal, multi-step tasks by calling external tools, yet they remain fragile in practice. Exis

safetyarxiv-cs-ai
1 Jul 2026
Safety

Reinforcement Learning-Based Control for an Inline Skating Humanoid Robot

DGX agent

arXiv:2606.31807v1 Announce Type: new Abstract: As humanoid robots become increasingly dynamic, coupling them with reinforcement learning offers a promising approach to solving the complex, underactua

safetyarxiv-cs-ro
1 Jul 2026
Safety

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs

DGX agent

arXiv:2606.32032v1 Announce Type: cross Abstract: Metacognition is a critical component of intelligence that describes the ability to monitor and regulate one's own cognitive processes. Yet LLMs exhib

safetyarxiv-cs-ai
1 Jul 2026
Safety

Resolving superposition in AI for interpretability and cross-modal alignment in patient-neuronal images

DGX agent

arXiv:2606.31394v1 Announce Type: cross Abstract: Artificial intelligence is transforming our capability to solve biological challenges. In dimensionality bottleneck regimes exacerbated by high-dimens

safetyarxiv-cs-ai
1 Jul 2026
Safety

Rethinking On-policy Optimization for Query Augmentation

DGX agent

arXiv:2510.17139v3 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have led to a surge of interest in query augmentation for information retrieval (IR). Two main appro

safetyarxiv-cs-cl
1 Jul 2026
Safety

Revisiting the Volume Hypothesis

DGX agent

arXiv:2606.31282v1 Announce Type: new Abstract: Modern deep neural networks often contain far more parameters than needed to fit their training data, yet they achieve impressive generalization. A comm

safetyarxiv-cs-lg
1 Jul 2026
Safety

Rhythm-Structured Predictive Learning for Remote Photoplethysmography

DGX agent

arXiv:2606.31736v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) estimates physiological signals from facial videos by analyzing subtle pulse induced skin color variations. Despite r

safetyarxiv-cs-cv
1 Jul 2026
Safety

Robust Text Watermarking for Large Language Models via Dual Semantic Embeddings

DGX agent

arXiv:2606.31602v1 Announce Type: new Abstract: This work presents Dual-Embedding Watermarking (DEW), a semantic watermarking scheme for large language models (LLMs) that leverages contextual and toke

safetyarxiv-cs-cl
1 Jul 2026
Safety

Robustness of Robotic Manipulation: Foundations and Frontiers

DGX agent

arXiv:2606.31494v1 Announce Type: cross Abstract: Humans and animals exhibit remarkable robustness in physical manipulation, yet robots remain far behind. Progress toward human-level manipulation robu

safetyarxiv-cs-ai
1 Jul 2026
Safety

Sampling-Based Coordination-Informed Multi-Objective Multi-Robot Reinforcement Learning

DGX agent

arXiv:2606.30893v1 Announce Type: new Abstract: Multi-robot systems must simultaneously optimize competing objectives while maintaining coordinated behavior. Existing multi-agent reinforcement learnin

safetyarxiv-cs-ro
1 Jul 2026
Safety

Seeing Is Not Sharing: Some Vision-Language Models Overestimate Common Ground in Asymmetric Dialogue

DGX agent

arXiv:2606.31719v1 Announce Type: cross Abstract: In collaborative dialogue, shared perception does not guarantee shared interpretation. Mutual understanding must be established through interaction. W

safetyarxiv-cs-ai
1 Jul 2026
Safety

Size Doesn't Matter: Cosine-Scored Sparse Autoencoders

DGX agent

arXiv:2606.15054v2 Announce Type: replace Abstract: Sparse autoencoders (SAEs) detect features via inner product, so a feature's activation scales with both its directional alignment and the input's n

safetyarxiv-cs-lg
1 Jul 2026
Safety

Smart charging of large fleets of Electric Vehicles: Independent Multi-Agent Reinforcement Learning approaches

DGX agent

arXiv:2606.31347v1 Announce Type: new Abstract: The electrification of transportation through electric vehicles introduces new challenges for power grid management, such as increased peak demand, volt

safetyarxiv-cs-ai
1 Jul 2026
Safety

SpectralSplats: Robust Differentiable Tracking via Spectral Moment Supervision

DGX agent

arXiv:2603.24036v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) enables real-time, photorealistic novel view synthesis, making it a highly attractive representation for model-based vi

safetyarxiv-cs-cv
1 Jul 2026
Safety

Stabilization Learning: A Paradigm Transition Bridging Control Theory and Machine Learning

DGX agent

arXiv:2606.31562v1 Announce Type: new Abstract: Stabilization learning is an interdisciplinary paradigm that bridges control theory and machine learning. Its core idea is to enable systems to adjust t

safetyarxiv-cs-ro
1 Jul 2026
Safety

Structural Preservation and the Logical Expressiveness of Graph Neural Networks

DGX agent

arXiv:2606.17882v2 Announce Type: replace Abstract: Bridges between graph neural networks (GNNs) and logical formalisms have been established by fixing architectural choices, such as the types of aggr

safetyarxiv-cs-ai
1 Jul 2026
Safety

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation

DGX agent

arXiv:2606.30849v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have significantly advanced audio-driven portrait animation, but their high computational cost leads to substantial infere

safetyarxiv-cs-cv
1 Jul 2026
Safety

TactX: Learning Shared Tactile Representations Across Diverse Sensors

DGX agent

arXiv:2606.31236v1 Announce Type: new Abstract: Tactile sensors provide critical information for contact-rich manipulation, yet tactile representations and policies remain tightly coupled to each spec

safetyarxiv-cs-ro
1 Jul 2026
Safety

TDGT: A Tabular Data Generation Toolkit supporting adaptive GPU-accelerated Bayesian mixture models, diffusion-based models, and latent-space generative modeling

DGX agent

arXiv:2606.31268v1 Announce Type: cross Abstract: The growing demand for privacy-preserving data sharing has positioned synthetic data generation as a critical component of responsible AI workflows. D

safetyarxiv-cs-ai
1 Jul 2026
Safety

Test-Time Verification for Text-to-SQL via Outcome Reward Models

DGX agent

arXiv:2606.30851v1 Announce Type: cross Abstract: Improving the reliability of large language models (LLMs) at inference time is a central challenge in structured reasoning tasks such as Text-to-SQL.

safetyarxiv-cs-ai
1 Jul 2026
Safety

The Past Is Prologue: A Plug-in Controller for Selective Updates in Sequentially Evolving LLM Memory

DGX agent

arXiv:2606.31121v1 Announce Type: new Abstract: Sequentially evolving LLM memory enables agents to reuse past experience, but existing systems usually deploy each locally generated memory update witho

safetyarxiv-cs-ai
1 Jul 2026
Safety

Token-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement Learning

DGX agent

arXiv:2606.31599v1 Announce Type: cross Abstract: Vision-language models (VLMs) combining reinforcement learning (RL) ignite remarkable progress in multimodal reasoning, yet still struggle with medica

safetyarxiv-cs-ai
1 Jul 2026
Safety

TORA: Topological Representation Alignment for 3D Shape Assembly

DGX agent

arXiv:2604.04050v2 Announce Type: replace Abstract: Flow-matching methods for 3D shape assembly learn point-wise velocity fields that transport parts toward assembled configurations, yet they receive

safetyarxiv-cs-cv
1 Jul 2026
Safety

Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation

DGX agent

arXiv:2606.31184v1 Announce Type: cross Abstract: Adaptive experiments for average treatment effects (ATE) require randomized allocations balancing valid inference with statistical efficiency. The ora

safetyarxiv-cs-ai
1 Jul 2026
Safety

TreeAgent: A Generalizable Multi-Agent Framework for Automated Bias Labeling in Forestry via Compiled Expert Rules and Vision-Language Models

DGX agent

arXiv:2606.31976v1 Announce Type: new Abstract: Human-labeled data are widely used as reference annotations in ML, despite known variability across annotators in many expert-driven domains. In additio

safetyarxiv-cs-ai
1 Jul 2026
Safety

TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning

DGX agent

arXiv:2606.32017v1 Announce Type: cross Abstract: Agentic reinforcement learning requires assigning credit to environment-facing actions such as searches, clicks, edits, navigation commands, and objec

safetyarxiv-cs-ai
1 Jul 2026
Safety

UniTac: A Unified Multimodal Model for Cross-Sensor Tactile Understanding and Generation

DGX agent

arXiv:2606.31451v1 Announce Type: cross Abstract: Unified multimodal models (UMMs) have shown great promise in integrating understanding and generation across diverse modalities. However, existing res

safetyarxiv-cs-ai
1 Jul 2026
Safety

Unsupervised Data-Efficient Cross-Modal Retrieval with Global-Neighborhood Alignment Hashing

DGX agent

arXiv:2606.31517v1 Announce Type: cross Abstract: Compared to supervised cross-modal hashing (CMH), unsupervised CMH reduces the reliance on manual labeling by learning binary codes from unlabeled ima

safetyarxiv-cs-cv
1 Jul 2026
Safety

VIGOR: VIdeo Geometry-Oriented Reward for Temporal Generative Alignment

DGX agent

arXiv:2603.16271v3 Announce Type: replace Abstract: Video diffusion models lack explicit geometric supervision during training, leading to inconsistency artifacts such as object deformation, spatial d

safetyarxiv-cs-cv
1 Jul 2026
Safety

Vision-Language Procedural Reasoning for Context-Aware Reward Modeling of Robotic Endovascular Guidewire Navigation

DGX agent

arXiv:2606.30698v1 Announce Type: new Abstract: Robotic-assisted endovascular interventions demand accurate, stable, and context-aware guidewire navigation in complex and patient-specific vascular ana

safetyarxiv-cs-ro
1 Jul 2026
Safety

Wait, am I Being Fair? Characterizing Deductive Stereotyping and Mitigating It with Fair-GCG

DGX agent

arXiv:2606.30989v1 Announce Type: cross Abstract: Warning: This paper contains several toxic and offensive statements. While reasoning generally improves fairness in recent large language models (LLMs

safetyarxiv-cs-ai
1 Jul 2026
Safety

Warp RL: Reshaping Base Policy Distributions for Dynamics Adaptation

DGX agent

arXiv:2606.31043v1 Announce Type: new Abstract: Residual reinforcement learning adapts a pretrained robot policy by learning an additive correction to its actions. While effective when adaptation amou

safetyarxiv-cs-lg
1 Jul 2026
Safety

What Probing Reveals about Autonomous Driving: Linking Internal Prediction Errors to Ego Planning

DGX agent

arXiv:2606.31106v1 Announce Type: cross Abstract: Large-scale datasets and fast simulators have enabled improvements in driving policies that appear safe and robust, yet strong performance in nominal

safetyarxiv-cs-ai
1 Jul 2026
Safety

A causal modeling perspective on decision theory

DGX agent

arXiv:2606.29911v1 Announce Type: new Abstract: Decision theory provides a formal framework for how agents should make choices under uncertainty, drawing on ideas from philosophy, probability, and cau

safetyarxiv-cs-ai
30 Jun 2026
Safety

A Hybrid Framework for Song Lyric Annotation Based on Human-LLM Alignment

DGX agent

arXiv:2606.29273v1 Announce Type: cross Abstract: Emotion recognition of song lyrics is a challenging task since lyrics may not necessarily align with the overall emotion of a song. As a result, lyric

safetyarxiv-cs-ai
30 Jun 2026
Safety

A3M: Adaptive, Adversarial and Multi-Objective Learning for Strategic Bidding in Repeated Auctions

DGX agent

arXiv:2606.28943v1 Announce Type: new Abstract: Learning to bid in repeated multi-unit auctions with bandit feedback poses a fundamental challenge. Existing methods often rely on rigid explore-then-ex

safetyarxiv-cs-cl
30 Jun 2026
Safety

AB-RAG: Adaptive Budgeted Retrieval-Augmented Generation for Reliable Question Answering

DGX agent

arXiv:2606.29090v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) has become the standard way to ground large language models in external knowledge, yet most systems retrieve a fi

safetyarxiv-cs-ai
30 Jun 2026
Safety

AccelAes: Accelerating Diffusion Transformers for Training-Free Aesthetic-Enhanced Image Generation

DGX agent

arXiv:2603.12575v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) are a dominant backbone for high-fidelity text-to-image generation due to strong scalability and alignment at high res

safetyarxiv-cs-cv
30 Jun 2026
Safety

Accurate Recognition of Pneumonia and COVID-19 by Geometric Shape Normalization of Lung Region using Automatic Landmark Detection and Piecewise Affine Warping

DGX agent

arXiv:2606.29715v1 Announce Type: new Abstract: This paper presents an automatic system for recognizing pulmonary diseases in chest X-rays using geometric normalization of the lung region. The method

safetyarxiv-cs-cv
30 Jun 2026
Safety

ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2606.30072v1 Announce Type: new Abstract: Cooperative tasks in Multi-Agent Reinforcement Learning (MARL) require agents to collectively maximize a shared return. Under the Centralized Training w

safetyarxiv-cs-ai
30 Jun 2026
Safety

Adaptive Block Diffusion: Resolving Training-Inference Mismatch in Diffusion Language Models

DGX agent

arXiv:2606.29275v1 Announce Type: new Abstract: Diffusion Language Models (DLMs) are typically trained under fixed context structures, restricting denoising to predetermined token subsets. This create

safetyarxiv-cs-lg
30 Jun 2026
Safety

AeroPlace-Flow: Language-Grounded Object Placement for Aerial Manipulators via Visual Foresight and Object Flow

DGX agent

arXiv:2603.07744v2 Announce Type: replace Abstract: Precise object placement remains underexplored in aerial manipulation, where most systems rely on predefined target coordinates and focus primarily

safetyarxiv-cs-ro
30 Jun 2026
Safety

Agentic Tool Use in Large Language Models

DGX agent

arXiv:2604.00835v2 Announce Type: replace Abstract: Large language models are increasingly being deployed as autonomous agents yet their real world effectiveness depends on reliable tools for informat

safetyarxiv-cs-cl
30 Jun 2026
Safety

'AI Watermarking': Bridging Policy Discourse and Technical Capabilities

DGX agent

arXiv:2606.28331v1 Announce Type: cross Abstract: The widespread deployment of generative artificial intelligence (AI) models has raised serious concerns about the proliferation of AI-generated conten

safetyarxiv-cs-ai
30 Jun 2026
Safety

An Integrated Machine Learning and Hierarchical Variance Decomposition Pipeline for Student Performance Prediction and Metacognitive Calibration on Multi-Signal Telemetry

DGX agent

arXiv:2606.28881v1 Announce Type: cross Abstract: Predicting student performance and characterizing metacognitive calibration are essential for personalization in intelligent tutoring systems. Prior r

safetyarxiv-cs-ai
30 Jun 2026
Safety

Analytic Concept-Centric Memory for Agentic Embodied Manipulation

DGX agent

arXiv:2606.29774v1 Announce Type: new Abstract: Long-horizon embodied manipulation requires agents to remember persistent objects, track changing scene states, and reuse prior interaction knowledge. H

safetyarxiv-cs-ro
30 Jun 2026
Safety

Analyzing Defensive Misdirection Against Model-Guided Automated Attacks on Agentic AI Systems

DGX agent

arXiv:2606.20470v2 Announce Type: replace-cross Abstract: Agentic AI systems increasingly rely on language-model components to interpret instructions, process external data, invoke tools, and coordina

safetyarxiv-cs-ai
30 Jun 2026
← Previous
1…105106107108109…260
Next →