AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
14 Apr 2026

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought

SafetyDGX agent

arXiv:2506.16796v4 Announce Type: replace Abstract: Real-World Image Super-Resolution is one of the most challenging task in image restoration. However, existing methods struggle with an accurate unde

Regularized Entropy Information Adaptation with Temporal-Awareness Networks for Simultaneous Speech Translation

SafetyDGX agent

arXiv:2604.09916v1 Announce Type: new Abstract: Simultaneous Speech Translation (SimulST) requires balancing high translation quality with low latency. Recent work introduced REINA, a method that trai

Reinforcement Learning for Intensity Control: An Application to Choice-Based Network Revenue Management

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2406.05358v3 Announce Type: replace Abstract: Intensity control is a class of continuous-time dynamic optimization problems with many important applications in Operations Research including queu

Relative Entropy Pathwise Policy Optimization

SafetyDGX agent

arXiv:2507.11019v4 Announce Type: replace Abstract: Score-function based methods for policy learning, such as REINFORCE and PPO, have delivered strong results in game-playing and robotics, yet their h

Reliable Evaluation Protocol for Low-Precision Retrieval

SafetyDGX agent

arXiv:2508.03306v4 Announce Type: replace-cross Abstract: Lowering the numerical precision of model parameters and computations is widely adopted to improve the efficiency of retrieval systems. Howeve

Rethinking LLM Watermark Detection in Black-Box Settings: A Non-Intrusive Third-Party Framework

SafetyDGX agent

arXiv:2603.14968v2 Announce Type: replace-cross Abstract: While watermarking serves as a critical mechanism for LLM provenance, existing secret-key schemes tightly couple detection with injection, req

Rethinking Token-Level Credit Assignment in RLVR: A Polarity-Entropy Analysis

SafetyDGX agent

arXiv:2604.11056v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has substantially improved the reasoning ability of Large Language Models (LLMs). However, its s

Revisiting Compositionality in Dual-Encoder Vision-Language Models: The Role of Inference

SafetyDGX agent

arXiv:2604.11496v1 Announce Type: cross Abstract: Dual-encoder Vision-Language Models (VLMs) such as CLIP are often characterized as bag-of-words systems due to their poor performance on compositional

Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?

SafetyDGX agent

arXiv:2505.24778v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly used in high-stakes domains, accurately assessing their confidence is crucial. Humans typically exp

SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting

SafetyDGX agent

arXiv:2604.10688v1 Announce Type: cross Abstract: On-policy reinforcement learning has become the dominant paradigm for reasoning alignment in large language models, yet its sparse, outcome-level rewa

SEARL: Joint Optimization of Policy and Tool Graph Memory for Self-Evolving Agents

SafetyDGX agent

arXiv:2604.07791v2 Announce Type: replace Abstract: Recent advances in Reinforcement Learning with Verifiable Rewards (RLVR) have demonstrated significant potential in single-turn reasoning tasks. Wit

See Fair, Speak Truth: Equitable Attention Improves Grounding and Reduces Hallucination in Vision-Language Alignment

SafetyDGX agent

arXiv:2604.09749v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) frequently hallucinate objects that are absent from the visual input, often because attention during decoding i

SHE: Stepwise Hybrid Examination Reinforcement Learning Framework for E-commerce Search Relevance

SafetyDGX agent

arXiv:2510.07972v3 Announce Type: replace Abstract: Query-product relevance prediction is vital for AI-driven e-commerce, yet current LLM-based approaches face a dilemma: SFT and DPO struggle with lon

Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD

SafetyDGX agent

arXiv:2604.10373v1 Announce Type: cross Abstract: From adversarial robustness to multi-agent learning, many machine learning tasks can be cast as finite-sum min-max optimization or, more generally, as

SIMPLER: H&E-Informed Representation Learning for Structured Illumination Microscopy

SafetyDGX agent

arXiv:2604.10334v1 Announce Type: new Abstract: Structured Illumination Microscopy (SIM) enables rapid, high-contrast optical sectioning of fresh tissue without staining or physical sectioning, making

Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents

SafetyDGX agent

arXiv:2604.10674v1 Announce Type: cross Abstract: Reinforcement learning (RL) has been widely used to train LLM agents for multi-turn interactive tasks, but its sample efficiency is severely limited b

SLALOM: Simulation Lifecycle Analysis via Longitudinal Observation Metrics for Social Simulation

SafetyDGX agent

arXiv:2604.11466v1 Announce Type: cross Abstract: Large Language Model (LLM) agents offer a potentially-transformative path forward for generative social science but face a critical crisis of validity

Some opportunities you just can't turn down, like when @GaryMarcus wants to come talk about AI at BugBash. Fewer than 20 tickets left, come …

SafetyDGX agent

Gary Marcus, a prominent AI researcher and critic known for his skepticism of current deep learning approaches, was invited to speak at BugBash, an event hosted by Antithesis. The post, shared by Anti

StaMo: Unsupervised Learning of Generalizable Robot Motion from Compact State Representation

SafetyDGX agent

arXiv:2510.05057v2 Announce Type: replace-cross Abstract: A fundamental challenge in embodied intelligence is developing expressive and compact state representations for efficient world modeling and d

Stop Fixating on Prompts: Reasoning Hijacking and Constraint Tightening for Red-Teaming LLM Agents

SafetyDGX agent

arXiv:2604.05549v2 Announce Type: replace Abstract: With the widespread application of LLM-based agents across various domains, their complexity has introduced new security threats. Existing red-team

Structural Consequences of Policy-Based Interventions on the Global Supply Chain Network

SafetyDGX agent

arXiv:2604.11479v1 Announce Type: new Abstract: As global political tensions rise and the anticipation of additional tariffs from the United States on international trade increases, the issues of econ

Structural Gating and Effect-aligned Lag-resolved Temporal Causal Discovery Framework with Application to Heat-Pollution Extremes

SafetyDGX agent

arXiv:2604.10371v1 Announce Type: new Abstract: This study proposes Structural Gating and Effect-aligned Discovery for Temporal Causal Discovery (SGED-TCD), a novel and general framework for lag-resol

Structured Causal Video Reasoning via Multi-Objective Alignment

SafetyDGX agent

arXiv:2604.04415v2 Announce Type: replace Abstract: Human understanding of video dynamics is typically grounded in a structured mental representation of entities, actions, and temporal relations, rath

SVD-Prune: Training-Free Token Pruning For Efficient Vision-Language Models

SafetyDGX agent

arXiv:2604.11530v1 Announce Type: cross Abstract: Vision-Language Models (VLM) have revolutionized multimodal learning by jointly processing visual and textual information. Yet, they face significant

Taking a Pulse on How Generative AI is Reshaping the Software Engineering Research Landscape

SafetyDGX agent

arXiv:2604.11184v1 Announce Type: cross Abstract: Context: Software engineering (SE) researchers increasingly study Generative AI (GenAI) while also incorporating it into their own research practices.

Task2vec Readiness: Diagnostics for Federated Learning from Pre-Training Embeddings

SafetyDGX agent

arXiv:2604.10849v1 Announce Type: cross Abstract: Federated learning (FL) performance is highly sensitive to heterogeneity across clients, yet practitioners lack reliable methods to anticipate how a f

TCSA-UDA: Text-Driven Cross-Semantic Alignment for Unsupervised Domain Adaptation in Medical Image Segmentation

SafetyDGX agent

arXiv:2511.05782v2 Announce Type: replace Abstract: Unsupervised domain adaptation for medical image segmentation remains a significant challenge due to substantial domain shifts across imaging modali

Teaching the Teacher: The Role of Teacher-Student Smoothness Alignment in Genetic Programming-based Symbolic Distillation

SafetyDGX agent

arXiv:2507.22767v3 Announce Type: replace-cross Abstract: Obtaining human-readable symbolic formulas via genetic programming-based symbolic distillation of a deep neural network trained on a target da

The Augmentation Trap: AI Productivity and the Cost of Cognitive Offloading

SafetyDGX agent

arXiv:2604.03501v2 Announce Type: replace-cross Abstract: Experimental evidence confirms that AI tools raise worker productivity, but also that sustained use can erode the expertise on which those gai

The Paradox of Professional Input: How Expert Collaboration with AI Systems Shapes Their Future Value

SafetyDGX agent

arXiv:2504.12654v1 Announce Type: cross Abstract: This perspective paper examines a fundamental paradox in the relationship between professional expertise and artificial intelligence: as domain expert

The Past Is Not Past: Memory-Enhanced Dynamic Reward Shaping

SafetyDGX agent

arXiv:2604.11297v1 Announce Type: cross Abstract: Despite the success of reinforcement learning for large language models, a common failure mode is reduced sampling diversity, where the policy repeate

The Poisoned Apple Effect: Strategic Manipulation of Mediated Markets via Technology Expansion of AI Agents

SafetyDGX agent

arXiv:2601.11496v2 Announce Type: replace-cross Abstract: The integration of AI agents into economic markets fundamentally alters the landscape of strategic interaction. We investigate the economic im

The Price of Ignorance: Information-Free Quotation for Data Retention in Machine Unlearning

SafetyDGX agent

arXiv:2604.11511v1 Announce Type: cross Abstract: When users exercise data deletion rights under the General Data Protection Regulation (GDPR) and similar regulations, mobile network operators face a

THOM: Generating Physically Plausible Hand-Object Meshes From Text

SafetyDGX agent

arXiv:2604.02736v3 Announce Type: replace Abstract: Generating photorealistic 3D hand-object interactions (HOIs) from text is important for applications like robotic grasping and AR/VR content creatio

Thought Branches: Interpreting LLM Reasoning Requires Resampling

SafetyDGX agent

arXiv:2510.27484v2 Announce Type: replace-cross Abstract: Most work interpreting reasoning models studies only a single chain-of-thought (CoT), yet these models define distributions over many possible

TInR: Exploring Tool-Internalized Reasoning in Large Language Models

SafetyDGX agent

arXiv:2604.10788v1 Announce Type: cross Abstract: Tool-Integrated Reasoning (TIR) has emerged as a promising direction by extending Large Language Models' (LLMs) capabilities with external tools durin

Tracing the Thought of a Grandmaster-level Chess-Playing Transformer

SafetyDGX agent

arXiv:2604.10158v1 Announce Type: new Abstract: While modern transformer neural networks achieve grandmaster-level performance in chess and other reasoning tasks, their internal computation process re

Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations

SafetyDGX agent

arXiv:2604.10123v1 Announce Type: new Abstract: Dysarthric speech severity assessment typically requires trained clinicians or supervised models built from labelled pathological speech, limiting scala

Triviality Corrected Endogenous Reward

SafetyDGX agent

arXiv:2604.11522v1 Announce Type: new Abstract: Reinforcement learning for open-ended text generation is constrained by the lack of verifiable rewards, necessitating reliance on judge models that requ

UNIGEOCLIP: Unified Geospatial Contrastive Learning

SafetyDGX agent

arXiv:2604.11668v1 Announce Type: new Abstract: The growing availability of co-located geospatial data spanning aerial imagery, street-level views, elevation models, text, and geographic coordinates o

Utilizing and Calibrating Hindsight Process Rewards via Reinforcement with Mutual Information Self-Evaluation

SafetyDGX agent

arXiv:2604.11611v1 Announce Type: new Abstract: To overcome the sparse reward challenge in reinforcement learning (RL) for agents based on large language models (LLMs), we propose Mutual Information S

Variance-Aware Prior-Based Tree Policies for Monte Carlo Tree Search

SafetyDGX agent

arXiv:2512.21648v2 Announce Type: replace-cross Abstract: Monte Carlo Tree Search (MCTS) has profoundly influenced reinforcement learning (RL) by integrating planning and learning in tasks requiring l

Variational Latent Entropy Estimation Disentanglement: Controlled Attribute Leakage for Face Recognition

SafetyDGX agent

arXiv:2604.11250v1 Announce Type: new Abstract: Face recognition embeddings encode identity, but they also encode other factors such as gender and ethnicity. Depending on how these factors are used by

VeriTrans: Fine-Tuned LLM-Assisted NL-to-PL Translation via a Deterministic Neuro-Symbolic Pipeline

SafetyDGX agent

arXiv:2604.10341v1 Announce Type: new Abstract: extbf{VeriTrans} is a reliability-first ML system that compiles natural-language requirements into solver-ready logic with validator-gated reliability.

VideoStir: Understanding Long Videos via Spatio-Temporally Structured and Intent-Aware RAG

SafetyDGX agent

arXiv:2604.05418v2 Announce Type: replace-cross Abstract: Scaling multimodal large language models (MLLMs) to long videos is constrained by limited context windows. While retrieval-augmented generatio

ViserDex: Visual Sim-to-Real for Robust Dexterous In-hand Reorientation

SafetyDGX agent

arXiv:2604.11138v1 Announce Type: cross Abstract: In-hand object reorientation requires precise estimation of the object pose to handle complex task dynamics. While RGB sensing offers rich semantic cu

Visual Enhanced Depth Scaling for Multimodal Latent Reasoning

SafetyDGX agent

arXiv:2604.10500v1 Announce Type: new Abstract: Multimodal latent reasoning has emerged as a promising paradigm that replaces explicit Chain-of-Thought (CoT) decoding with implicit feature propagation

VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2603.22003v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models typically map visual observations and linguistic instructions directly to robotic control signals. This 'black-b

Warm-Started Reinforcement Learning for Iterative 3D/2D Liver Registration

SafetyDGX agent

arXiv:2604.10245v1 Announce Type: new Abstract: Registration between preoperative CT and intraoperative laparoscopic video plays a crucial role in augmented reality (AR) guidance for minimally invasiv

WARPED: Wrist-Aligned Rendering for Robot Policy Learning from Egocentric Human Demonstrations

SafetyDGX agent

arXiv:2604.10809v1 Announce Type: new Abstract: Recent advancements in learning from human demonstration have shown promising results in addressing the scalability and high cost of data collection req

What Factors Affect LLMs and RLLMs in Financial Question Answering?

SafetyDGX agent

arXiv:2507.08339v4 Announce Type: replace Abstract: Recently, large language models (LLMs) and reasoning large language models (RLLMs) have gained considerable attention from many researchers. RLLMs e

When Can You Poison Rewards? A Tight Characterization of Reward Poisoning in Linear MDPs

SafetyDGX agent

arXiv:2604.10062v1 Announce Type: new Abstract: We study reward poisoning attacks in reinforcement learning (RL), where an adversary manipulates rewards within constrained budgets to force the target

When simulations look right but causal effects go wrong: Large language models as behavioral simulators

SafetyDGX agent

arXiv:2604.02458v2 Announce Type: replace-cross Abstract: Behavioral simulation is increasingly used to anticipate responses to interventions. Large language models (LLMs) enable researchers to specif

When Valid Signals Fail: Regime Boundaries Between LLM Features and RL Trading Policies

SafetyDGX agent

arXiv:2604.10996v1 Announce Type: cross Abstract: Can large language models (LLMs) generate continuous numerical features that improve reinforcement learning (RL) trading agents? We build a modular pi

Why why why is it so hard to understand that stupid systems can still be dangerous?

SafetyDGX agent

Why why why is it so hard to understand that stupid systems can still be dangerous? @Graffitinights1 For the zillionth time, dumb systems that are empowered can dangerous. like this monstrosity, as an

WM-DAgger: Enabling Efficient Data Aggregation for Imitation Learning with World Models

SafetyDGX agent

arXiv:2604.11351v1 Announce Type: new Abstract: Imitation learning is a powerful paradigm for training robotic policies, yet its performance is limited by compounding errors: minor policy inaccuracies

YIELD: A Large-Scale Dataset and Evaluation Framework for Information Elicitation Agents

SafetyDGX agent

arXiv:2604.10968v1 Announce Type: new Abstract: Most conversational agents (CAs) are designed to satisfy user needs through user-driven interactions. However, many real-world settings, such as academi

ZoomR: Memory Efficient Reasoning through Multi-Granularity Key Value Retrieval

SafetyDGX agent

arXiv:2604.10898v1 Announce Type: new Abstract: Large language models (LLMs) have shown great performance on complex reasoning tasks but often require generating long intermediate thoughts before reac

13 Apr 2026

A Mathematical Framework for Temporal Modeling and Counterfactual Policy Simulation of Student Dropout

SafetyDGX agent

arXiv:2604.08874v1 Announce Type: cross Abstract: This study proposes a temporal modeling framework with a counterfactual policy-simulation layer for student dropout in higher education, using LMS eng

A Representation-Level Assessment of Bias Mitigation in Foundation Models

SafetyDGX agent

arXiv:2604.08561v1 Announce Type: new Abstract: We investigate how successful bias mitigation reshapes the embedding space of encoder-only and decoder-only foundation models, offering an internal audi

← Previous
1…213214215216217…240
Next →