AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Adaptive 3D-RoPE: Physics-Aligned Rotary Positional Encoding for Wireless Foundation Models

DGX agent

arXiv:2605.00968v1 Announce Type: cross Abstract: Positional encoding plays a pivotal role in determin?ing the extrapolation and generalization performance of wireless foundation models for channel st

safetyarxiv-cs-ai
6 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

ADAPTS: Agentic Decomposition for Automated Protocol-agnostic Tracking of Symptoms

DGX agent

arXiv:2605.03212v1 Announce Type: cross Abstract: Modeling latent clinical constructs from unconstrained clinical interactions is a unique challenge in affective computing. We present ADAPTS (Agentic

safetyarxiv-cs-cl
6 May 2026
Safety

Agentic AI-Based Joint Computing and Networking via Mixture of Experts and Large Language Models

DGX agent

arXiv:2605.02911v1 Announce Type: new Abstract: Future sixth-generation (6G) mobile networks are envisioned to be equipped with a diverse set of powerful, yet highly specialized, optimization experts.

safetyarxiv-cs-lg
6 May 2026
Safety

Aligning Inductive Bias for Data-Efficient Generalization in State Space Models

DGX agent

arXiv:2509.20789v4 Announce Type: replace Abstract: The remarkable success of modern AI has been closely tied to scaling laws, yet the finite supply of high-quality data makes data efficiency--learnin

safetyarxiv-cs-lg
6 May 2026
Safety

Beyond Activation Alignment: The Geometry of Neural Sensitivity

DGX agent

arXiv:2605.03222v1 Announce Type: new Abstract: Activation-alignment measures such as Representational Similarity Analysis (RSA), Canonical Correlation Analysis (CCA), and Centered Kernel Alignment (C

safetyarxiv-cs-lg
6 May 2026
Safety

BifrostUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation

DGX agent

arXiv:2605.03452v1 Announce Type: new Abstract: High-quality data collection is a fundamental cornerstone for training humanoid whole-body visuomotor policies. Current data acquisition paradigms predo

safetyarxiv-cs-ro
6 May 2026
Safety

Bootstrapped Mixed Rewards for RL Post-Training: Injecting Canonical Action Order

DGX agent

arXiv:2512.04277v3 Announce Type: replace Abstract: Post-training with reinforcement learning (RL) typically optimizes a single scalar objective and ignores structure in how solutions are produced. We

safetyarxiv-cs-lg
6 May 2026
Safety

Can Semantic Methods Enhance Team Sports Tactics? A Methodology for Football with Broader Applications

DGX agent

arXiv:2601.00421v2 Announce Type: replace Abstract: This paper explores how semantic-space reasoning, traditionally used in computational linguistics, can be extended to tactical decision-making in te

safetyarxiv-cs-ai
6 May 2026
Safety

Catching the Infection Before It Spreads: Foresight-Guided Defense in Multi-Agent Systems

DGX agent

arXiv:2605.01758v1 Announce Type: new Abstract: Large multimodal model-based Multi-Agent Systems (MASs) enable collaborative complex problem solving through specialized agents. However, MASs are vulne

safetyarxiv-cs-ai
6 May 2026
Safety

Coherent Hierarchical Multi-Label Learning to Defer for Medical Imaging

DGX agent

arXiv:2605.02734v1 Announce Type: new Abstract: Learning to Defer (L2D) enables a model to predict autonomously or defer to an expert, but prior work largely assumes flat label spaces. We study the fi

safetyarxiv-cs-ai
6 May 2026
Safety

Deciphering Shortcut Learning from an Evolutionary Game Theory Perspective

DGX agent

arXiv:2605.02658v2 Announce Type: new Abstract: Shortcut learning causes deep learning models to rely on non-essential features within the data. However, its formation in deep neural network training

safetyarxiv-cs-ai
6 May 2026
Safety

Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning

DGX agent

arXiv:2602.20078v3 Announce Type: replace-cross Abstract: Scaling cooperative multi-agent reinforcement learning (MARL) is fundamentally limited by cross-agent noise. When agents share a common reward

safetyarxiv-cs-lg
6 May 2026
Safety

DGPO: Distribution Guided Policy Optimization for Fine Grained Credit Assignment

DGX agent

arXiv:2605.03327v1 Announce Type: new Abstract: Reinforcement learning is crucial for aligning large language models to perform complex reasoning tasks. However, current algorithms such as Group Relat

safetyarxiv-cs-lg
6 May 2026
Safety

Discovering Reinforcement Learning Interfaces with Large Language Models

DGX agent

arXiv:2605.03408v1 Announce Type: new Abstract: Reinforcement learning systems rely on environment interfaces that specify observations and reward functions, yet constructing these interfaces for new

safetyarxiv-cs-lg
6 May 2026
Safety

DMGD: Train-Free Dataset Distillation with Semantic-Distribution Matching in Diffusion Models

DGX agent

arXiv:2605.03877v1 Announce Type: new Abstract: Dataset distillation enables efficient training by distilling the information of large-scale datasets into significantly smaller synthetic datasets. Dif

safetyarxiv-cs-cv
6 May 2026
Safety

FIBER: A Differentially Private Optimizer with Filter-Aware Innovation Bias Correction

DGX agent

arXiv:2605.03425v1 Announce Type: new Abstract: Differentially private (DP) training protects individual examples by adding noise to gradients, but the injected noise interacts nontrivially with adapt

safetyarxiv-cs-lg
6 May 2026
Safety

FINER-SQL: Boosting Small Language Models for Text-to-SQL

DGX agent

arXiv:2605.03465v1 Announce Type: cross Abstract: Large language models have driven major advances in Text-to-SQL generation. However, they suffer from high computational cost, long latency, and data

safetyarxiv-cs-cl
6 May 2026
Safety

From SFT to RL: Demystifying the Post-Training Pipeline for LLM-based Vulnerability Detection

DGX agent

arXiv:2602.14012v2 Announce Type: replace-cross Abstract: The integration of LLMs into vulnerability detection (VD) has shifted the field toward more interpretable and context-aware analysis. While po

safetyarxiv-cs-ai
6 May 2026
Safety

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling

DGX agent

arXiv:2507.07982v2 Announce Type: replace Abstract: Videos inherently represent 2D projections of a dynamic 3D world. However, our analysis suggests that video diffusion models trained solely on raw v

safetyarxiv-cs-cv
6 May 2026
Safety

Global and Local Topology-Aware Attention with Persistent Homology and Euler Biases for Time-Series Forecasting

DGX agent

arXiv:2605.03163v1 Announce Type: new Abstract: Scientific time series often encode predictive geometric structure, including connectivity, cycles, shell-like geometry, directional changes, and nonlin

safetyarxiv-cs-lg
6 May 2026
Safety

GRAFT: Auditing Graph Neural Networks via Global Feature Attribution

DGX agent

arXiv:2605.03377v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve strong performance on node classification tasks but remain difficult to interpret, particularly with respect to whi

safetyarxiv-cs-lg
6 May 2026
Safety

Grounding Multi-Hop Reasoning in Structural Causal Models via Group Relative Policy Optimization

DGX agent

arXiv:2605.01482v1 Announce Type: new Abstract: Multi-Hop Fact Verification (MHFV) necessitates complex reasoning across disparate evidence, posing significant challenges for Large Language Models (LL

safetyarxiv-cs-ai
6 May 2026
Safety

GRPO-TTA: Test-Time Visual Tuning for Vision-Language Models via GRPO-Driven Reinforcement Learning

DGX agent

arXiv:2605.03403v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has recently shown strong performance in post-training large language models and vision-language models. It ra

safetyarxiv-cs-cv
6 May 2026
Safety

Healthcare AI GYM for Medical Agents

DGX agent

arXiv:2605.02943v1 Announce Type: new Abstract: Clinical reasoning demands multi-step interactions -- gathering patient history, ordering tests, interpreting results, and making safe treatment decisio

safetyarxiv-cs-lg
6 May 2026
Safety

Heterogeneous Graph Importance Scoring and Clustering with Automated LLM-based Interpretation

DGX agent

arXiv:2605.02919v1 Announce Type: new Abstract: Urban bridge networks are critical infrastructure whose disruption can cascade into severe impacts on transportation, emergency services, and economic a

safetyarxiv-cs-lg
6 May 2026
Safety

HiMAC: Hierarchical Macro-Micro Learning for Long-Horizon LLM Agents

DGX agent

arXiv:2603.00977v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents have recently demonstrated strong capabilities in interactive decision-making, yet they remain fundamentally

safetyarxiv-cs-lg
6 May 2026
Safety

Intervention Complexity as a Canonical Reward and a Measure of Intelligence

DGX agent

arXiv:2605.02175v1 Announce Type: new Abstract: The Legg--Hutter universal intelligence measure provides a rigorous scalar assessment of general intelligence as expected reward across all computable e

safetyarxiv-cs-ai
6 May 2026
Model Releases

Jiao: Bridging Isolation and Customization in Mixed Criticality Robotics

DGX agent

arXiv:2605.03641v1 Announce Type: new Abstract: Consumer robotics demands consolidation of safety-critical control, perception pipelines, and user applications on shared multicore platforms. While sta

model-releasesarxiv-cs-ro
6 May 2026
Safety

Khala: Scaling Acoustic Token Language Models Toward High-Fidelity Music Generation

DGX agent

arXiv:2605.01790v1 Announce Type: cross Abstract: A common design pattern in high-quality music generation is to handle structure and fidelity in different representation spaces: a generator first mod

safetyarxiv-cs-ai
6 May 2026
Safety

Large Language Models are Universal Reasoners for Visual Generation

DGX agent

arXiv:2605.04040v1 Announce Type: new Abstract: Text-to-image generation has advanced rapidly with diffusion models, progressing from CLIP and T5 conditioning to unified systems where a single LLM bac

safetyarxiv-cs-cv
6 May 2026
Safety

LLM-XTM: Enhancing Cross-Lingual Topic Models with Large Language Models

DGX agent

arXiv:2605.03299v1 Announce Type: new Abstract: Cross-lingual topic modeling aims to discover shared semantic structures across languages, yet existing models depend on sparse bilingual resources and

safetyarxiv-cs-cl
6 May 2026
Safety

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation

DGX agent

arXiv:2602.05048v2 Announce Type: replace Abstract: Joint planning through language-based interactions is a key area of human-AI teaming. Planning problems in the open world often involve various aspe

safetyarxiv-cs-ai
6 May 2026
Safety

Mix3R: Mixing Feed-forward Reconstruction and Generative 3D Priors for Joint Multi-view Aligned 3D Reconstruction and Pose Estimation

DGX agent

arXiv:2605.03359v1 Announce Type: new Abstract: Recent trends in sparse-view 3D reconstruction have taken two different paths: feed-forward reconstruction that predicts pixel-aligned point maps withou

safetyarxiv-cs-cv
6 May 2026
Safety

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation

DGX agent

arXiv:2605.03058v1 Announce Type: new Abstract: A key goal of explainable AI (XAI) is to express the decision logic of large language models (LLMs) in symbolic form and link it to internal mechanisms.

safetyarxiv-cs-lg
6 May 2026
Safety

Nora: Normalized Orthogonal Row Alignment for Scalable Matrix Optimizer

DGX agent

arXiv:2605.03769v1 Announce Type: new Abstract: Matrix-based optimizers have demonstrated immense potential in training Large Language Models (LLMs), however, designing an ideal optimizer remains a fo

safetyarxiv-cs-lg
6 May 2026
Safety

Normalized Matching Transformer

DGX agent

arXiv:2503.17715v3 Announce Type: replace Abstract: We introduce the Normalized Matching Transformer (NMT), a deep learning approach for efficient and accurate sparse semantic keypoint matching betwee

safetyarxiv-cs-cv
6 May 2026
Safety

OGPO: Sample Efficient Full-Finetuning of Generative Control Policies

DGX agent

arXiv:2605.03065v1 Announce Type: new Abstract: Generative control policies (GCPs), such as diffusion- and flow-based control policies, have emerged as effective parameterizations for robot learning.

safetyarxiv-cs-lg
6 May 2026
Safety

On Surprising Effects of Risk-Aware Domain Randomization for Contact-Rich Sampling-based Predictive Control

DGX agent

arXiv:2605.03290v1 Announce Type: new Abstract: Domain randomization (DR) is widely used in policy learning to improve robustness to modeling error, but remains underexplored in contact-rich sampling-

safetyarxiv-cs-ro
6 May 2026
Safety

Optimal Posterior Sampling for Policy Identification in Tabular Markov Decision Processes

DGX agent

arXiv:2605.03921v1 Announce Type: new Abstract: We study the (arepsilon, elta)-PAC policy identification problem in finite-horizon episodic Markov Decision Processes. Existing approaches provide finit

safetyarxiv-cs-lg
6 May 2026
Safety

Orientation-Aware Unsupervised Domain Adaptation for Brain Tumor Classification Across Multi-Modal MRI

DGX agent

arXiv:2605.03490v1 Announce Type: new Abstract: The clinical integration of deep learning models for brain tumor diagnosis in neuro-oncology is severely constrained by limited expert-annotated MRI dat

safetyarxiv-cs-cv
6 May 2026
Safety

Poly-EPO: Training Exploratory Reasoning Models

DGX agent

arXiv:2604.17654v3 Announce Type: replace Abstract: Exploration is a cornerstone of learning from experience: it enables agents to find solutions to complex problems, generalize to novel ones, and sca

safetyarxiv-cs-ai
6 May 2026
Safety

Population-Aware Imitation Learning in Mean-field Games with Common Noise

DGX agent

arXiv:2605.03357v1 Announce Type: new Abstract: Mean Field Games (MFGs) provide a powerful framework for modeling the collective behavior of large populations of interacting agents. In this paper, we

safetyarxiv-cs-lg
6 May 2026
Safety

Power-Softmax: Towards Secure LLM Inference over Encrypted Data

DGX agent

arXiv:2410.09457v2 Announce Type: replace Abstract: Modern cryptographic methods for implementing privacy-preserving LLMs such as gls{HE} require the LLMs to have a polynomial form. Forming such a rep

safetyarxiv-cs-lg
6 May 2026
Safety

Predicting missing values: A good idea?

DGX agent

arXiv:2605.03733v1 Announce Type: cross Abstract: Minimizing the Mean Squared Error (MSE) is a key objective in machine learning and is commonly used for imputing missing values. While this approach p

safetyarxiv-cs-lg
6 May 2026
Safety

Privacy Preserving Machine Learning Workflow: from Anonymization to Personalized Differential Privacy Budgets in Federated Learning

DGX agent

arXiv:2605.02372v1 Announce Type: cross Abstract: The growing development of artificial intelligence based solutions, together with privacy legislation, has driven the rise of the so-called privacy pr

safetyarxiv-cs-ai
6 May 2026
Safety

Pseudo-differential-enhanced physics-informed neural networks

DGX agent

arXiv:2602.14663v2 Announce Type: replace Abstract: We present pseudo-differential enhanced physics-informed neural networks (PINNs), an extension of gradient enhancement but in Fourier space. Gradien

safetyarxiv-cs-lg
6 May 2026
Safety

Reinforcement Learning Trained Observer Control for Bearings-Only Tracking

DGX agent

arXiv:2605.02120v1 Announce Type: new Abstract: This paper develops a deep reinforcement learning based observer control policy for autonomous bearings-only tracking of a moving target. The observer m

safetyarxiv-cs-ai
6 May 2026
Safety

Resource-Efficient Reinforcement for Reasoning Large Language Models via Dynamic One-Shot Policy Refinement

DGX agent

arXiv:2602.00815v2 Announce Type: replace Abstract: Large language models (LLMs) have exhibited remarkable performance on complex reasoning tasks, with reinforcement learning under verifiable rewards

safetyarxiv-cs-ai
6 May 2026
← Previous
1…198199200201202…257
Next →