AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

When Context Returns: Toward Robust Internalization in On-Policy Distillation

DGX agent

arXiv:2606.11627v1 Announce Type: cross Abstract: Recent work has shown that on-policy distillation can internalize privileged context, such as system prompts or task hints, into a student model so th

safetyarxiv-cs-ai
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

When Researchers Say Mental Model/Theory of Mind of AI, What Are They Really Talking About?

DGX agent

arXiv:2510.02660v2 Announce Type: replace-cross Abstract: When researchers claim AI systems possess ToM or mental models, they are fundamentally discussing behavioral predictions and bias corrections

safetyarxiv-cs-ai
11 Jun 2026
Safety

Which Speech Representation Better Matches Text-Native Reasoning? A Study of Speech-Text Alignment on Frame Rate and Representation

DGX agent

arXiv:2606.12199v1 Announce Type: cross Abstract: Spoken dialogue models typically start from text LLM backbones, yet reasoning often degrades when conditioning on speech instead of text. We attribute

safetyarxiv-cs-cl
11 Jun 2026
Safety

3SPO: State-Score-Supervised Policy Optimization for LLM Agents

DGX agent

arXiv:2606.09961v1 Announce Type: cross Abstract: Training large language models (LLMs) as autonomous agents via reinforcement learning (RL) has enabled frontier models to achieve superhuman performan

safetyarxiv-cs-ai
10 Jun 2026
Safety

A Comprehensive Survey of Direct Preference Optimization: Datasets, Theories, Variants, and Applications

DGX agent

arXiv:2410.15595v4 Announce Type: replace Abstract: With the rapid advancement of large language models (LLMs), aligning policy models with human preferences has become increasingly critical. Direct P

safetyarxiv-cs-ai
10 Jun 2026
Safety

A fine-grained attention and geometric correspondence model for musculoskeletal risk classification in athletes using multimodal visual and skeletal features

DGX agent

arXiv:2509.05913v3 Announce Type: replace Abstract: Musculoskeletal disorders pose significant risks to athletes, and early risk assessment is essential for prevention. However, most existing methods

safetyarxiv-cs-cv
10 Jun 2026
Safety

A Practical Recipe Towards Improving Sim-and-Real Correlation for VLA Evaluation

DGX agent

arXiv:2606.10366v1 Announce Type: cross Abstract: Simulation has become an essential tool for evaluating and improving vision-language-action (VLA) policies, offering scalable, reproducible, and contr

safetyarxiv-cs-ai
10 Jun 2026
Safety

A Source Domain is All You Need: Source-Only Cross-OS Transfer Learning for APT Anomaly Detection via Semantic Alignment and Optimal Transport

DGX agent

arXiv:2606.10216v1 Announce Type: cross Abstract: Advanced Persistent Threats (APTs) are stealthy, multi-stage cyberattacks whose detection is difficult due to scarce labeled traces, severe class imba

safetyarxiv-cs-ai
10 Jun 2026
Safety

A Unified Multi-Modal Framework for Intelligent Financial Systems: Integrating Reinforcement Learning, High-Frequency Trading, and Game-Theoretic Approaches with Cross-Modal Sentiment Analysis

DGX agent

arXiv:2606.10412v1 Announce Type: new Abstract: The rapid evolution of financial technology demands sophisticated artificial intelligence systems capable of handling diverse challenges across multiple

safetyarxiv-cs-ai
10 Jun 2026
Safety

Adoption of Generative Artificial Intelligence in the German Software Engineering Industry: An Empirical Study

DGX agent

arXiv:2601.16700v2 Announce Type: replace-cross Abstract: Generative artificial intelligence (GenAI) tools have seen rapid adoption among software developers. While adoption rates in the industry are

safetyarxiv-cs-ai
10 Jun 2026
Safety

Alignment Defends LLMs from Property Inference Attacks

DGX agent

arXiv:2606.10217v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly fine-tuned on domain-specific datasets that may contain sensitive, dataset-level properties. Recent work h

safetyarxiv-cs-lg
10 Jun 2026
Safety

An LLM-Native Psychometric Instrument Does Not Predict LLM Behavior: Evidence Across 25 Models

DGX agent

arXiv:2606.09843v1 Announce Type: cross Abstract: Large language models (LLMs) produce stable self-reports on personality inventories, but these self-reports do not predict observed behavior. Whether

safetyarxiv-cs-ai
10 Jun 2026
Safety

AnimaSpark: A Feed-Forward Method for Animating Arbitrary 3D Objects

DGX agent

arXiv:2606.10988v1 Announce Type: new Abstract: While recent advancements in generative AI have substantially accelerated static 3D model creation workflows, the synthesis of category-agnostic 3D anim

safetyarxiv-cs-cv
10 Jun 2026
Safety

Architect-Ant: Editable Automatic Furnishing of Architectural Floor Plans

DGX agent

arXiv:2606.10953v1 Announce Type: new Abstract: Furnished floor plans are fundamental to real estate visualization, interior design, and architectural workflows. However, progress in automatic furnitu

safetyarxiv-cs-ai
10 Jun 2026
Safety

ARM: An AutoRegressive Large Multimodal Model with Unified Discrete Representations

DGX agent

arXiv:2606.11188v1 Announce Type: new Abstract: This paper introduces ARM, a discrete representation-based AutoRegressive Model that unifies image understanding, generation, and editing within a next-

safetyarxiv-cs-cv
10 Jun 2026
Safety

AsyncWebRL: Efficient Multi-Step RL for Visual Web Agents

DGX agent

arXiv:2606.05597v2 Announce Type: replace Abstract: Training vision-language web agents with multi-step RL is compute-intensive, with two dominant forms of inefficiency: idle GPUs in synchronous RL, a

safetyarxiv-cs-lg
10 Jun 2026
Safety

Automated Alignment between Elicitation Interviews and Requirements

DGX agent

arXiv:2510.08622v2 Announce Type: replace Abstract: Software requirements are derived from a variety of elicitation techniques, many of which have a conversational nature, like interviews. However, ev

safetyarxiv-cs-cl
10 Jun 2026
Safety

Automated Scoring of Arabic Text Using Large Language Models: A Literature Review

DGX agent

arXiv:2606.09830v1 Announce Type: new Abstract: In modern educational systems, Automatic Text Scoring (ATS) plays a central role by enabling scalable and consistent evaluation of learner responses wit

safetyarxiv-cs-cl
10 Jun 2026
Safety

Baseline-Free Policy Optimization for Neural Combinatorial Optimization

DGX agent

arXiv:2606.10321v1 Announce Type: cross Abstract: Neural combinatorial optimization (NCO) trains autoregressive policies to solve routing problems. The standard training algorithm, REINFORCE with a ro

safetyarxiv-cs-ai
10 Jun 2026
Safety

Bellman-Taylor Score Decoding for Markov Decision Processes with State-Dependent Feasible Action Sets

DGX agent

arXiv:2606.10979v1 Announce Type: new Abstract: Many Markov decision processes (MDPs) in operations research have feasible actions that are state dependent and defined implicitly by various operationa

safetyarxiv-cs-ai
10 Jun 2026
Safety

Beyond Uniform Token-Level Trust Region in LLM Reinforcement Learning

DGX agent

arXiv:2606.10968v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become standard for improving LLM reasoning. However, existing PPO-style trust-region mechan

safetyarxiv-cs-ai
10 Jun 2026
Safety

Bypassing Copyright Protection in Diffusion-based Customization via Two-Stage Latent Feature Optimization

DGX agent

arXiv:2606.09909v1 Announce Type: cross Abstract: With the growing concerns over copyright infringement in diffusion-based customization, adversarial attacks have emerged as a prominent defense strate

safetyarxiv-cs-ai
10 Jun 2026
Safety

Causal Ensemble Agent: Hierarchical Causal Discovery with LLM-guided Expert Reweighting

DGX agent

arXiv:2606.10607v1 Announce Type: cross Abstract: Causal discovery aims to uncover causal structures from observational data, which is crucial for real-world decision-making. However, different causal

safetyarxiv-cs-ai
10 Jun 2026
Safety

Closing the Modality Gap in Zero-Shot HAR: Contrastive Training and Separability-Optimized Prototypes on IMU Data

DGX agent

arXiv:2606.10789v1 Announce Type: new Abstract: Zero-shot learning (ZSL) for inertial measurement unit (IMU)-based human activity recognition (HAR) faces a central challenge: bridging the gap between

safetyarxiv-cs-lg
10 Jun 2026
Safety

Conditional Vendi Score: Prompt-Aware Diversity Evaluation for Generative AI Models and LLMs

DGX agent

arXiv:2411.02817v2 Announce Type: replace-cross Abstract: Generative models guided by text prompts are widely evaluated for fidelity and prompt alignment, yet their ability to produce outputs remains

safetyarxiv-cs-ai
10 Jun 2026
Safety

Contrastive Spectral Rectification: Test-Time Defense towards Zero-shot Adversarial Robustness of CLIP

DGX agent

arXiv:2601.19210v2 Announce Type: replace Abstract: Vision-language models (VLMs) such as CLIP have demonstrated remarkable zero-shot generalization, yet remain highly vulnerable to adversarial exampl

safetyarxiv-cs-cv
10 Jun 2026
Safety

Convergence of Monte Carlo Optimistic Policy Iteration: Beyond Uniform State-Action Updates

DGX agent

arXiv:2606.10580v1 Announce Type: cross Abstract: The asymptotic behaviour of Monte Carlo optimistic policy iteration (MC-O-PI) is a long-standing open question. When the model of the environment is u

safetyarxiv-cs-ai
10 Jun 2026
Safety

Cross-Modal Knowledge Distillation without Paired Data: Theoretical Foundation and Algorithm

DGX agent

arXiv:2606.10504v1 Announce Type: new Abstract: Cross-modal knowledge distillation (CMKD) studies how a (large) teacher model trained on one type of data (e.g., images) can guide a (smaller) student m

safetyarxiv-cs-ai
10 Jun 2026
Safety

Decision-Calibrated Conformal Uncertainty for Pacing Decisions in Streaming Advertising

DGX agent

arXiv:2606.10187v1 Announce Type: cross Abstract: We develop a decision-calibrated conformal framework for pacing decisions in streaming advertising. Pacing depends on uncertain future inventory, dema

safetyarxiv-cs-lg
10 Jun 2026
Safety

Decoupling Thought from Speech: Knowledge-Grounded Counterfactual Reasoning for Resilient Multi-Agent Argumentation

DGX agent

arXiv:2606.10475v1 Announce Type: cross Abstract: Multi-agent debate frameworks have been shown to improve large language model performance in convergent tasks, but they are currently optimized in a w

safetyarxiv-cs-ai
10 Jun 2026
Safety

DeRA-MOS: Optimizing Text-to-Music Evaluation via Decoupled Listwise Ranking and Modality Alignment

DGX agent

arXiv:2606.10010v1 Announce Type: cross Abstract: Evaluating text-to-music (TTM) systems remains expensive because music impression (MI) and text alignment (TA) scores rely on human mean opinion score

safetyarxiv-cs-ai
10 Jun 2026
Safety

Dexterous Point Policy: Learning Point-based Dexterous Hand Policies from Human Demonstrations

DGX agent

arXiv:2606.10614v1 Announce Type: cross Abstract: Robotic foundation models pre-trained on human demonstration videos have shown promise, but a significant embodiment gap remains when the resulting po

safetyarxiv-cs-cv
10 Jun 2026
Safety

Dissect and Prune: Enhancing Robustness in AI-Generated Image Detection

DGX agent

arXiv:2606.10309v1 Announce Type: new Abstract: While existing AI-generated image detectors report high performance, we identify that this is largely driven by a critical prediction asymmetry: a bias

safetyarxiv-cs-cv
10 Jun 2026
Safety

Dropout-GRPO: Variational Stochasticity for Continuous Latent Reasoning

DGX agent

arXiv:2606.10184v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) relies on the diversity of K rollouts within each group; otherwise, the group-mean advantage A^{(k)} = r^{(k

safetyarxiv-cs-ai
10 Jun 2026
Safety

Early-Token Confidence Predicts Reasoning Quality in Multi-Agent LLM Debate

DGX agent

arXiv:2606.10307v1 Announce Type: new Abstract: Evaluating reasoning quality in multi-agent LLM systems is challenging, especially for open-ended tasks without reference answers. We investigate whethe

safetyarxiv-cs-cl
10 Jun 2026
Safety

Embodiment-conditioned Generalist Control for Multirotor Aerial Robots

DGX agent

arXiv:2606.10857v1 Announce Type: cross Abstract: We present a generalist position control policy capable of controlling arbitrary multirotor configurations of a certain rotor count (e.g., hexarotors

safetyarxiv-cs-lg
10 Jun 2026
Safety

Encoding the Euler Characteristic Transform

DGX agent

arXiv:2606.10824v1 Announce Type: new Abstract: The Euler Characteristic Curve (ECC) records the Euler characteristic of a linearly embedded cell complex as a function of filtration height in a given

safetyarxiv-cs-lg
10 Jun 2026
Safety

Enhancing Multilingual LLM-based ASR with Mixture of Experts and Dynamic Downsampling

DGX agent

arXiv:2606.10439v1 Announce Type: cross Abstract: The rapid progress of large language models (LLMs) has opened up a new frontier for automatic speech recognition (ASR), making their effective integra

safetyarxiv-cs-cl
10 Jun 2026
Safety

ERAlign: Energy-based Representation Alignment of GNNs and LLMs on Text-attributed Graphs

DGX agent

arXiv:2606.10461v1 Announce Type: cross Abstract: Text-attributed Graphs (TAGs) incorporate textual node attributes with graph structures to describe rich relational semantics. Recent efforts to integ

safetyarxiv-cs-ai
10 Jun 2026
Safety

Ethical and Technical Limits of Deepfake Speech Datasets

DGX agent

arXiv:2606.10911v1 Announce Type: cross Abstract: Claims about the robustness and fairness of deepfake speech detectors are only as credible as the datasets used to train and evaluate those systems. W

safetyarxiv-cs-ai
10 Jun 2026
Safety

Event-Driven Reinforcement Learning Enables Long-Horizon Control in Semiconductor Fabrication

DGX agent

arXiv:2606.10705v1 Announce Type: cross Abstract: Reinforcement learning promises to optimize sequential decisions in large-scale systems. Semiconductor manufacturing systems are stochastic and highly

safetyarxiv-cs-ai
10 Jun 2026
Safety

Fast and Highly Expressive Policy Learning for Offline Reinforcement Learning via Bootstrapped Flow Q-Learning

DGX agent

arXiv:2606.10613v1 Announce Type: cross Abstract: Diffusion-based Q-learning has emerged as a powerful paradigm for offline reinforcement learning, but its reliance on multi-step denoising makes both

safetyarxiv-cs-ai
10 Jun 2026
Safety

Flow Control: Steering Vision-Language-Action Models with Simple Real-Time Inputs

DGX agent

arXiv:2606.10180v1 Announce Type: cross Abstract: We introduce flow control of vision-language-action (VLA) models, a simple and effective way to steer VLA actions in real-time through generic inputs,

safetyarxiv-cs-ai
10 Jun 2026
Safety

Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models

DGX agent

arXiv:2606.11025v1 Announce Type: new Abstract: Recent work has demonstrated that online reinforcement learning (RL) can substantially improve the quality and alignment of flow matching models for ima

safetyarxiv-cs-lg
10 Jun 2026
Safety

Generalized-CVO: Fast and Correspondence-Free Local Point Cloud Registration with Second Order Riemannian Optimization

DGX agent

arXiv:2606.10019v1 Announce Type: cross Abstract: We propose a fast and correspondence-free local point cloud registration method that leverages geometric surface structure and reproducing kernel Hilb

safetyarxiv-cs-ai
10 Jun 2026
Safety

GHOST: Hierarchical Sub-Goal Policies for Generalizing Robot Manipulation

DGX agent

arXiv:2606.10025v1 Announce Type: cross Abstract: We present GHOST, a framework for learning visuomotor manipulation policies that generalize beyond the training distribution. GHOST factorizes control

safetyarxiv-cs-cv
10 Jun 2026
Safety

Globally Localizing Lunar Rover in Pixels via Graph Alignment

DGX agent

arXiv:2606.10602v1 Announce Type: new Abstract: Precise rover localization is a prerequisite for autonomous lunar exploration, yet the absence of Global Navigation Satellite System (GNSS) signals and

safetyarxiv-cs-cv
10 Jun 2026
Safety

Going with the Flow: Koopman Behavioral Models as Pseudo Planners for Visuo-Motor Dexterity

DGX agent

arXiv:2602.07413v3 Announce Type: replace Abstract: Contemporary visuo-motor dexterity models often rely on expressive policy classes with diffusion and transformer backbones to achieve strong perform

safetyarxiv-cs-ro
10 Jun 2026
← Previous
1…124125126127128…260
Next →