AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

RAG-KT: Cross-platform Explainable Knowledge Tracing with Multi-view Fusion Retrieval Generation

DGX agent

arXiv:2604.10960v1 Announce Type: new Abstract: Knowledge Tracing (KT) infers a student's knowledge state from past interactions to predict future performance. Conventional Deep Learning (DL)-based KT

safetyarxiv-cs-ai
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought

DGX agent

arXiv:2506.16796v4 Announce Type: replace Abstract: Real-World Image Super-Resolution is one of the most challenging task in image restoration. However, existing methods struggle with an accurate unde

safetyarxiv-cs-cv
14 Apr 2026
Safety

Regularized Entropy Information Adaptation with Temporal-Awareness Networks for Simultaneous Speech Translation

DGX agent

arXiv:2604.09916v1 Announce Type: new Abstract: Simultaneous Speech Translation (SimulST) requires balancing high translation quality with low latency. Recent work introduced REINA, a method that trai

safetyarxiv-cs-lg
14 Apr 2026
Safety

Reinforcement Learning for Intensity Control: An Application to Choice-Based Network Revenue Management

DGX agent

arXiv:2406.05358v3 Announce Type: replace Abstract: Intensity control is a class of continuous-time dynamic optimization problems with many important applications in Operations Research including queu

safetyarxiv-cs-lg
14 Apr 2026
Safety

Relative Entropy Pathwise Policy Optimization

DGX agent

arXiv:2507.11019v4 Announce Type: replace Abstract: Score-function based methods for policy learning, such as REINFORCE and PPO, have delivered strong results in game-playing and robotics, yet their h

safetyarxiv-cs-lg
14 Apr 2026
Safety

Reliable Evaluation Protocol for Low-Precision Retrieval

DGX agent

arXiv:2508.03306v4 Announce Type: replace-cross Abstract: Lowering the numerical precision of model parameters and computations is widely adopted to improve the efficiency of retrieval systems. Howeve

safetyarxiv-cs-ai
14 Apr 2026
Safety

Rethinking LLM Watermark Detection in Black-Box Settings: A Non-Intrusive Third-Party Framework

DGX agent

arXiv:2603.14968v2 Announce Type: replace-cross Abstract: While watermarking serves as a critical mechanism for LLM provenance, existing secret-key schemes tightly couple detection with injection, req

safetyarxiv-cs-cl
14 Apr 2026
Safety

Rethinking Token-Level Credit Assignment in RLVR: A Polarity-Entropy Analysis

DGX agent

arXiv:2604.11056v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has substantially improved the reasoning ability of Large Language Models (LLMs). However, its s

safetyarxiv-cs-ai
14 Apr 2026
Safety

Revisiting Compositionality in Dual-Encoder Vision-Language Models: The Role of Inference

DGX agent

arXiv:2604.11496v1 Announce Type: cross Abstract: Dual-encoder Vision-Language Models (VLMs) such as CLIP are often characterized as bag-of-words systems due to their poor performance on compositional

safetyarxiv-cs-cl
14 Apr 2026
Safety

Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?

DGX agent

arXiv:2505.24778v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly used in high-stakes domains, accurately assessing their confidence is crucial. Humans typically exp

safetyarxiv-cs-cl
14 Apr 2026
Safety

SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting

DGX agent

arXiv:2604.10688v1 Announce Type: cross Abstract: On-policy reinforcement learning has become the dominant paradigm for reasoning alignment in large language models, yet its sparse, outcome-level rewa

safetyarxiv-cs-ai
14 Apr 2026
Safety

SEARL: Joint Optimization of Policy and Tool Graph Memory for Self-Evolving Agents

DGX agent

arXiv:2604.07791v2 Announce Type: replace Abstract: Recent advances in Reinforcement Learning with Verifiable Rewards (RLVR) have demonstrated significant potential in single-turn reasoning tasks. Wit

safetyarxiv-cs-ai
14 Apr 2026
Safety

See Fair, Speak Truth: Equitable Attention Improves Grounding and Reduces Hallucination in Vision-Language Alignment

DGX agent

arXiv:2604.09749v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) frequently hallucinate objects that are absent from the visual input, often because attention during decoding i

safetyarxiv-cs-cv
14 Apr 2026
Safety

SHE: Stepwise Hybrid Examination Reinforcement Learning Framework for E-commerce Search Relevance

DGX agent

arXiv:2510.07972v3 Announce Type: replace Abstract: Query-product relevance prediction is vital for AI-driven e-commerce, yet current LLM-based approaches face a dilemma: SFT and DPO struggle with lon

safetyarxiv-cs-ai
14 Apr 2026
Safety

Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD

DGX agent

arXiv:2604.10373v1 Announce Type: cross Abstract: From adversarial robustness to multi-agent learning, many machine learning tasks can be cast as finite-sum min-max optimization or, more generally, as

safetyarxiv-cs-lg
14 Apr 2026
Safety

SIMPLER: H&E-Informed Representation Learning for Structured Illumination Microscopy

DGX agent

arXiv:2604.10334v1 Announce Type: new Abstract: Structured Illumination Microscopy (SIM) enables rapid, high-contrast optical sectioning of fresh tissue without staining or physical sectioning, making

safetyarxiv-cs-cv
14 Apr 2026
Safety

Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents

DGX agent

arXiv:2604.10674v1 Announce Type: cross Abstract: Reinforcement learning (RL) has been widely used to train LLM agents for multi-turn interactive tasks, but its sample efficiency is severely limited b

safetyarxiv-cs-ai
14 Apr 2026
Safety

SLALOM: Simulation Lifecycle Analysis via Longitudinal Observation Metrics for Social Simulation

DGX agent

arXiv:2604.11466v1 Announce Type: cross Abstract: Large Language Model (LLM) agents offer a potentially-transformative path forward for generative social science but face a critical crisis of validity

safetyarxiv-cs-ai
14 Apr 2026
Safety

StaMo: Unsupervised Learning of Generalizable Robot Motion from Compact State Representation

DGX agent

arXiv:2510.05057v2 Announce Type: replace-cross Abstract: A fundamental challenge in embodied intelligence is developing expressive and compact state representations for efficient world modeling and d

safetyarxiv-cs-cv
14 Apr 2026
Safety

Stop Fixating on Prompts: Reasoning Hijacking and Constraint Tightening for Red-Teaming LLM Agents

DGX agent

arXiv:2604.05549v2 Announce Type: replace Abstract: With the widespread application of LLM-based agents across various domains, their complexity has introduced new security threats. Existing red-team

safetyarxiv-cs-cl
14 Apr 2026
Safety

Structural Consequences of Policy-Based Interventions on the Global Supply Chain Network

DGX agent

arXiv:2604.11479v1 Announce Type: new Abstract: As global political tensions rise and the anticipation of additional tariffs from the United States on international trade increases, the issues of econ

safetyarxiv-cs-lg
14 Apr 2026
Safety

Structural Gating and Effect-aligned Lag-resolved Temporal Causal Discovery Framework with Application to Heat-Pollution Extremes

DGX agent

arXiv:2604.10371v1 Announce Type: new Abstract: This study proposes Structural Gating and Effect-aligned Discovery for Temporal Causal Discovery (SGED-TCD), a novel and general framework for lag-resol

safetyarxiv-cs-lg
14 Apr 2026
Safety

Structured Causal Video Reasoning via Multi-Objective Alignment

DGX agent

arXiv:2604.04415v2 Announce Type: replace Abstract: Human understanding of video dynamics is typically grounded in a structured mental representation of entities, actions, and temporal relations, rath

safetyarxiv-cs-cl
14 Apr 2026
Safety

SVD-Prune: Training-Free Token Pruning For Efficient Vision-Language Models

DGX agent

arXiv:2604.11530v1 Announce Type: cross Abstract: Vision-Language Models (VLM) have revolutionized multimodal learning by jointly processing visual and textual information. Yet, they face significant

safetyarxiv-cs-ai
14 Apr 2026
Safety

Taking a Pulse on How Generative AI is Reshaping the Software Engineering Research Landscape

DGX agent

arXiv:2604.11184v1 Announce Type: cross Abstract: Context: Software engineering (SE) researchers increasingly study Generative AI (GenAI) while also incorporating it into their own research practices.

safetyarxiv-cs-ai
14 Apr 2026
Safety

Task2vec Readiness: Diagnostics for Federated Learning from Pre-Training Embeddings

DGX agent

arXiv:2604.10849v1 Announce Type: cross Abstract: Federated learning (FL) performance is highly sensitive to heterogeneity across clients, yet practitioners lack reliable methods to anticipate how a f

safetyarxiv-cs-ai
14 Apr 2026
Safety

TCSA-UDA: Text-Driven Cross-Semantic Alignment for Unsupervised Domain Adaptation in Medical Image Segmentation

DGX agent

arXiv:2511.05782v2 Announce Type: replace Abstract: Unsupervised domain adaptation for medical image segmentation remains a significant challenge due to substantial domain shifts across imaging modali

safetyarxiv-cs-cv
14 Apr 2026
Safety

Teaching the Teacher: The Role of Teacher-Student Smoothness Alignment in Genetic Programming-based Symbolic Distillation

DGX agent

arXiv:2507.22767v3 Announce Type: replace-cross Abstract: Obtaining human-readable symbolic formulas via genetic programming-based symbolic distillation of a deep neural network trained on a target da

safetyarxiv-cs-ai
14 Apr 2026
Safety

The Augmentation Trap: AI Productivity and the Cost of Cognitive Offloading

DGX agent

arXiv:2604.03501v2 Announce Type: replace-cross Abstract: Experimental evidence confirms that AI tools raise worker productivity, but also that sustained use can erode the expertise on which those gai

safetyarxiv-cs-ai
14 Apr 2026
Safety

The Paradox of Professional Input: How Expert Collaboration with AI Systems Shapes Their Future Value

DGX agent

arXiv:2504.12654v1 Announce Type: cross Abstract: This perspective paper examines a fundamental paradox in the relationship between professional expertise and artificial intelligence: as domain expert

safetyarxiv-cs-ai
14 Apr 2026
Safety

The Past Is Not Past: Memory-Enhanced Dynamic Reward Shaping

DGX agent

arXiv:2604.11297v1 Announce Type: cross Abstract: Despite the success of reinforcement learning for large language models, a common failure mode is reduced sampling diversity, where the policy repeate

safetyarxiv-cs-ai
14 Apr 2026
Safety

The Poisoned Apple Effect: Strategic Manipulation of Mediated Markets via Technology Expansion of AI Agents

DGX agent

arXiv:2601.11496v2 Announce Type: replace-cross Abstract: The integration of AI agents into economic markets fundamentally alters the landscape of strategic interaction. We investigate the economic im

safetyarxiv-cs-ai
14 Apr 2026
Safety

The Price of Ignorance: Information-Free Quotation for Data Retention in Machine Unlearning

DGX agent

arXiv:2604.11511v1 Announce Type: cross Abstract: When users exercise data deletion rights under the General Data Protection Regulation (GDPR) and similar regulations, mobile network operators face a

safetyarxiv-cs-lg
14 Apr 2026
Safety

THOM: Generating Physically Plausible Hand-Object Meshes From Text

DGX agent

arXiv:2604.02736v3 Announce Type: replace Abstract: Generating photorealistic 3D hand-object interactions (HOIs) from text is important for applications like robotic grasping and AR/VR content creatio

safetyarxiv-cs-cv
14 Apr 2026
Safety

Thought Branches: Interpreting LLM Reasoning Requires Resampling

DGX agent

arXiv:2510.27484v2 Announce Type: replace-cross Abstract: Most work interpreting reasoning models studies only a single chain-of-thought (CoT), yet these models define distributions over many possible

safetyarxiv-cs-ai
14 Apr 2026
Safety

TInR: Exploring Tool-Internalized Reasoning in Large Language Models

DGX agent

arXiv:2604.10788v1 Announce Type: cross Abstract: Tool-Integrated Reasoning (TIR) has emerged as a promising direction by extending Large Language Models' (LLMs) capabilities with external tools durin

safetyarxiv-cs-ai
14 Apr 2026
Safety

Tracing the Thought of a Grandmaster-level Chess-Playing Transformer

DGX agent

arXiv:2604.10158v1 Announce Type: new Abstract: While modern transformer neural networks achieve grandmaster-level performance in chess and other reasoning tasks, their internal computation process re

safetyarxiv-cs-lg
14 Apr 2026
Safety

Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations

DGX agent

arXiv:2604.10123v1 Announce Type: new Abstract: Dysarthric speech severity assessment typically requires trained clinicians or supervised models built from labelled pathological speech, limiting scala

safetyarxiv-cs-cl
14 Apr 2026
Safety

Triviality Corrected Endogenous Reward

DGX agent

arXiv:2604.11522v1 Announce Type: new Abstract: Reinforcement learning for open-ended text generation is constrained by the lack of verifiable rewards, necessitating reliance on judge models that requ

safetyarxiv-cs-cl
14 Apr 2026
Safety

UNIGEOCLIP: Unified Geospatial Contrastive Learning

DGX agent

arXiv:2604.11668v1 Announce Type: new Abstract: The growing availability of co-located geospatial data spanning aerial imagery, street-level views, elevation models, text, and geographic coordinates o

safetyarxiv-cs-cv
14 Apr 2026
Safety

Utilizing and Calibrating Hindsight Process Rewards via Reinforcement with Mutual Information Self-Evaluation

DGX agent

arXiv:2604.11611v1 Announce Type: new Abstract: To overcome the sparse reward challenge in reinforcement learning (RL) for agents based on large language models (LLMs), we propose Mutual Information S

safetyarxiv-cs-cl
14 Apr 2026
Safety

Variance-Aware Prior-Based Tree Policies for Monte Carlo Tree Search

DGX agent

arXiv:2512.21648v2 Announce Type: replace-cross Abstract: Monte Carlo Tree Search (MCTS) has profoundly influenced reinforcement learning (RL) by integrating planning and learning in tasks requiring l

safetyarxiv-cs-ai
14 Apr 2026
Safety

Variational Latent Entropy Estimation Disentanglement: Controlled Attribute Leakage for Face Recognition

DGX agent

arXiv:2604.11250v1 Announce Type: new Abstract: Face recognition embeddings encode identity, but they also encode other factors such as gender and ethnicity. Depending on how these factors are used by

safetyarxiv-cs-cv
14 Apr 2026
Safety

VeriTrans: Fine-Tuned LLM-Assisted NL-to-PL Translation via a Deterministic Neuro-Symbolic Pipeline

DGX agent

arXiv:2604.10341v1 Announce Type: new Abstract: extbf{VeriTrans} is a reliability-first ML system that compiles natural-language requirements into solver-ready logic with validator-gated reliability.

safetyarxiv-cs-ai
14 Apr 2026
Safety

VideoStir: Understanding Long Videos via Spatio-Temporally Structured and Intent-Aware RAG

DGX agent

arXiv:2604.05418v2 Announce Type: replace-cross Abstract: Scaling multimodal large language models (MLLMs) to long videos is constrained by limited context windows. While retrieval-augmented generatio

safetyarxiv-cs-ai
14 Apr 2026
Safety

ViserDex: Visual Sim-to-Real for Robust Dexterous In-hand Reorientation

DGX agent

arXiv:2604.11138v1 Announce Type: cross Abstract: In-hand object reorientation requires precise estimation of the object pose to handle complex task dynamics. While RGB sensing offers rich semantic cu

safetyarxiv-cs-cv
14 Apr 2026
Safety

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning

DGX agent

arXiv:2604.10500v1 Announce Type: new Abstract: Multimodal latent reasoning has emerged as a promising paradigm that replaces explicit Chain-of-Thought (CoT) decoding with implicit feature propagation

safetyarxiv-cs-cv
14 Apr 2026
Model Releases

VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models

DGX agent

arXiv:2603.22003v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models typically map visual observations and linguistic instructions directly to robotic control signals. This 'black-b

model-releasesarxiv-cs-ro
14 Apr 2026
← Previous
1…230231232233234…257
Next →