AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
15 May 2026

From Ranking to Reasoning: Explainable Web API Recommendation via Semantic Reasoning

SafetyDGX agent

arXiv:2511.05820v2 Announce Type: replace-cross Abstract: The rapid growth of Web APIs has made automated Web API recommendation essential for efficient mashup development. However, existing approache

@GaryMarcus If token prediction is 'generating thought itself,' then my phone's autocomplete is a philosopher.

SafetyDGX agent

Gary Marcus critiques the claim that token prediction in large language models constitutes genuine thought or reasoning, using the analogy of smartphone autocomplete to illustrate that predictive text

Generative Deep Learning for Computational Destaining and Restaining of Unregistered Digital Pathology Images

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.14251v1 Announce Type: new Abstract: Conditional generative adversarial networks (cGANs) have enabled high-fidelity computational staining and destaining of hematoxylin and eosin (H&E) in d

Google confirms it's testing a new storage policy after some users reported that new Gmail accounts get only 5GB of free storage unless they add a phone number (Akshay Gangwar/Android Authority)

SafetyDGX agent

Akshay Gangwar / Android Authority: Google confirms it's testing a new storage policy after some users reported that new Gmail accounts get only 5GB of free storage unless they add a phone number — Up

Google updates its spam rules to include attempts to ‘manipulate’ AI

SafetyDGX agent

Google updated its spam policy to mark attempts to 'manipulate' its AI model in search results as spam, including results in AI Overview or AI Mode in Search, as Search Engine Land reports: 'In the co

GradShield: Alignment Preserving Finetuning

SafetyDGX agent

arXiv:2605.14194v1 Announce Type: new Abstract: Large Language Models (LLMs) pose a significant risk of safety misalignment after finetuning, as models can be compromised by both explicitly and implic

Grokking Finite-Dimensional Algebra

SafetyDGX agent

arXiv:2602.19533v2 Announce Type: replace-cross Abstract: This paper investigates the grokking phenomenon, which refers to the sudden transition from a long memorization to generalization observed dur

Hand-in-the-Loop: Improving Dexterous VLA via Seamless Interventional Correction

SafetyDGX agent

arXiv:2605.15157v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are prone to compounding errors in dexterous manipulation, where high-dimensional action spaces and contact-rich d

HeatKV: Head-tuned KV-cache Compression for Visual Autoregressive Modeling

SafetyDGX agent

arXiv:2605.14877v1 Announce Type: new Abstract: Visual Autoregressive (VAR) models have recently demonstrated impressive image generation quality while maintaining low latency. However, they suffer fr

Hierarchical Image Tokenization for Multi-Scale Image Super Resolution

SafetyDGX agent

arXiv:2605.14891v1 Announce Type: new Abstract: We introduce a multi-scale Image Super Resolution (ISR) method building on recent advances in Visual Auto-Regressive (VAR) modeling. VAR models break im

Hyperbolic Graph Neural Networks Under the Microscope: The Role of Geometry-Task Alignment

SafetyDGX agent

arXiv:2602.01828v2 Announce Type: replace Abstract: Many complex networks exhibit hierarchical, tree-like structures, making hyperbolic space a natural candidate wherein to learn representations of th

ICED: Concept-level Machine Unlearning via Interpretable Concept Decomposition

SafetyDGX agent

arXiv:2605.14309v1 Announce Type: cross Abstract: Machine unlearning in Vision-Language Models (VLMs) is typically performed at the image or instance level, making it difficult to precisely remove tar

Identifying Culprits Through Deep Deterministic Policy Gradient Deep Learning Investigation

SafetyDGX agent

arXiv:2605.14774v1 Announce Type: new Abstract: In the world of AI and advanced technologies investigation aspects identification of a crime or criminal plays a major problem. In this research we focu

Ideology Prediction of German Political Texts

SafetyDGX agent

arXiv:2605.14352v1 Announce Type: new Abstract: Elections represent a crucial milestone in a nation's ongoing development. To better understand the political rhetoric from various movements, ranging f

'In the past nine months, the United States has produced more AI legislation than in the prior decade,' write @JeffSonnenfeld, @GaryMarcus, …

SafetyDGX agent

'In the past nine months, the United States has produced more AI legislation than in the prior decade,' write @JeffSonnenfeld, @GaryMarcus, and Stephen Henriques in a commentary piece for Fortune. 'No

InfoSFT: Learn More and Forget Less with Information-Aware Token Weighting

SafetyDGX agent

arXiv:2605.14967v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) provides the standard approach for teaching LLMs new behaviors from offline expert demonstrations. However, standard SFT un

Know When To Fold 'Em: Token-Efficient LLM Synthetic Data Generation via Multi-Stage In-Flight Rejection

SafetyDGX agent

arXiv:2605.14062v1 Announce Type: new Abstract: While synthetic data generation with large language models (LLMs) is widely used in post-training pipelines, existing approaches typically generate full

KVPO: ODE-Native GRPO for Autoregressive Video Alignment via KV Semantic Exploration

SafetyDGX agent

arXiv:2605.14278v1 Announce Type: new Abstract: Aligning streaming autoregressive (AR) video generators with human preferences is challenging. Existing reinforcement learning methods predominantly rel

Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space

SafetyDGX agent

arXiv:2605.14531v1 Announce Type: new Abstract: This work reformulates language generation as a stochastic optimal control problem, providing a unified theoretical perspective to analyze autoregressiv

LATERN: Test-Time Context-Aware Explainable Video Anomaly Detection

SafetyDGX agent

arXiv:2605.15054v1 Announce Type: new Abstract: Vision-language models (VLMs) have recently emerged as a promising paradigm for video anomaly detection (VAD) due to their strong visual reasoning abili

Learning from Failures: Correction-Oriented Policy Optimization with Verifiable Rewards

SafetyDGX agent

arXiv:2605.14539v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as an effective paradigm for improving the reasoning capabilities of large language mo

Learning from Language Feedback via Variational Policy Distillation

SafetyDGX agent

arXiv:2605.15113v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) suffers from sparse outcome signals, creating severe exploration bottlenecks on complex reasoning

Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis

SafetyDGX agent

arXiv:2605.14392v1 Announce Type: new Abstract: We pursue a vision for self-improving language models in which the model does not merely generate problems or traces to imitate, but constructs the envi

Let Robots Feel Your Touch: Visuo-Tactile Cortical Alignment for Embodied Mirror Resonance

SafetyDGX agent

arXiv:2605.14571v1 Announce Type: cross Abstract: Observing touch on another's body can elicit corresponding tactile sensations in the observer, a phenomenon termed mirror touch that supports empathy

Logging Policy Design for Off-Policy Evaluation

SafetyDGX agent

arXiv:2605.15108v1 Announce Type: cross Abstract: Off-policy evaluation (OPE) estimates the value of a target treatment policy (e.g., a recommender system) using data collected by a different logging

LPH-VTON: Resolving the Structure-Texture Dilemma of Virtual Try-On via Latent Process Handover

SafetyDGX agent

arXiv:2605.14874v1 Announce Type: new Abstract: Virtual Try-On (VTON) aims to synthesize photorealistic images of garments precisely aligned with a person's body and pose. Current diffusion-based meth

MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2605.14201v1 Announce Type: cross Abstract: Vision-language-action (VLA) models are effective as end-to-end motion planners, but can be brittle when evaluated in closed-loop settings due to bein

Measuring and Mitigating Toxicity in Large Language Models: A Comprehensive Replication Study

SafetyDGX agent

arXiv:2605.14087v1 Announce Type: new Abstract: Large Language Models (LLMs), when trained on web-scale corpora, inherently absorb toxic patterns from their training data. This leads to ``toxic degene

Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems

SafetyDGX agent

arXiv:2605.14744v1 Announce Type: cross Abstract: Large language models in regulated financial workflows are governed by natural-language policies that the same model interprets, creating a principal-

Medical Report Generation: A Hierarchical Task Structure-Based Cross-Modal Causal Intervention Framework

SafetyDGX agent

arXiv:2511.02271v2 Announce Type: replace Abstract: Medical Report Generation (MRG) is a key part of modern medical diagnostics, as it automatically generates reports from radiological images to reduc

MetaBackdoor: Exploiting Positional Encoding as a Backdoor Attack Surface in LLMs

SafetyDGX agent

arXiv:2605.15172v1 Announce Type: cross Abstract: Backdoor attacks pose a serious security threat to large language models (LLMs), which are increasingly deployed as general-purpose assistants in safe

Mitigating Mask Prior Drift and Positional Attention Collapse in Large Diffusion Vision-Language Models

SafetyDGX agent

arXiv:2605.14530v1 Announce Type: new Abstract: Large diffusion vision-language models (LDVLMs) have recently emerged as a promising alternative to autoregressive models, enabling parallel decoding fo

MoRe: Modular Representations for Principled Continual Representation Learning on Squantial Data

SafetyDGX agent

arXiv:2605.14364v1 Announce Type: new Abstract: Continual learning requires models to adapt to new data while preserving previously acquired knowledge. At its core, this challenge can be viewed as pri

Multi-Dimensional Model Integrity and Responsibility Assessment Index and Scoring Framework

SafetyDGX agent

arXiv:2605.14550v1 Announce Type: new Abstract: Artificial intelligence in high-stakes tabular domains cannot be evaluated by predictive performance alone, yet current practice still assesses explaina

Multi-scale Coarse-to-fine Modeling for Test-time Human Motion Control

SafetyDGX agent

arXiv:2605.14935v1 Announce Type: new Abstract: We present MSCoT, a multi-scale, coarse-to-fine model for test-time human motion synthesis and control. Unlike recent approaches that rely on multiple i

Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning

SafetyDGX agent

arXiv:2512.07461v3 Announce Type: replace Abstract: We introduce Native Parallel Reasoner (NPR), a teacher-free framework that enables Large Language Models (LLMs) to self-evolve genuine parallel reas

NEST: Nested Event Stream Transformer for Sequences of Multisets

SafetyDGX agent

arXiv:2602.00520v3 Announce Type: replace Abstract: Event stream data often exhibit hierarchical structure in which multiple events co-occur, resulting in a sequence of multisets (i.e., bags of events

Not All Timesteps Matter Equally: Selective Alignment Knowledge Distillation for Spiking Neural Networks

SafetyDGX agent

arXiv:2605.14252v1 Announce Type: cross Abstract: Spiking neural networks (SNNs), which are brain-inspired and spike-driven, achieve high energy efficiency. However, a performance gap between SNNs and

Novel Dynamic Batch-Sensitive Adam Optimiser for Vehicular Accident Injury Severity Prediction

SafetyDGX agent

arXiv:2605.15083v1 Announce Type: cross Abstract: The choice of optimiser is important in deep learning, as it strongly influences model efficiency and speed of convergence. However, many commonly use

On Strong Equivalence Notions in Logic Programming and Abstract Argumentation

SafetyDGX agent

arXiv:2605.14721v1 Announce Type: new Abstract: Strong equivalence between knowledge bases ensures the possibility of replacing one with the other without affecting reasoning outcomes, in any given co

On the Burden of Achieving Fairness in Conformal Prediction

SafetyDGX agent

arXiv:2605.14260v1 Announce Type: cross Abstract: Conformal prediction is often calibrated with a single pooled threshold, but this can hide cross-group heterogeneity in score distributions and distor

On the Unreasonable Effectiveness of Last-layer Retraining

SafetyDGX agent

arXiv:2512.01766v2 Announce Type: replace Abstract: Last-layer retraining (LLR) methods -- wherein the last layer of a neural network is reinitialized and retrained on a held-out set following ERM tra

One Step to the Side: Why Defenses Against Malicious Finetuning Fail Under Adaptive Adversaries

SafetyDGX agent

arXiv:2605.14605v1 Announce Type: cross Abstract: Model providers increasingly release open weights or allow users to fine-tune foundation models through APIs. Although these models are safety-aligned

OneSearch-V2: The Latent Reasoning Enhanced Self-distillation Generative Search Framework

SafetyDGX agent

arXiv:2603.24422v2 Announce Type: replace-cross Abstract: Generative Retrieval (GR) has emerged as a promising paradigm for modern search systems. Compared to multi-stage cascaded architecture, it off

OpenAI's disavowal of a liability shield in Illinois SB 3444 bill and endorsement of a stronger SB 315 suggest it is open to meaningful AI safety legislation (Transformer)

SafetyDGX agent

Transformer: OpenAI's disavowal of a liability shield in Illinois SB 3444 bill and endorsement of a stronger SB 315 suggest it is open to meaningful AI safety legislation — Transformer Weekly: US-Chin

Performance-Driven Policy Optimization for Speculative Decoding with Adaptive Windowing

SafetyDGX agent

arXiv:2605.14978v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by having a lightweight draft model propose speculative windows of candidate tokens for parallel verifica

Persian MusicGen: A Large-Scale Dataset and Culturally-Aware Generative Model for Persian Music

SafetyDGX agent

arXiv:2605.14765v1 Announce Type: cross Abstract: Persian music, with its unique tonalities, modal systems (Dastgah), and rhythmic structures, presents significant challenges for music generation mode

Phylogenetic Tree Inference with Tropical Axial Attention

SafetyDGX agent

arXiv:2605.13894v1 Announce Type: cross Abstract: In this work, we introduce a Tropical Axial Attention neural reasoning architecture that replaces vanilla softmax dot-product attention with max-plus

Policy Optimization in Hybrid Discrete-Continuous Action Spaces via Mixed Gradients

SafetyDGX agent

arXiv:2605.14297v1 Announce Type: cross Abstract: We study reinforcement learning in hybrid discrete-continuous action spaces, such as settings where the discrete component selects a regime (or index)

Position: Behavioural Assurance Cannot Verify the Safety Claims Governance Now Demands

SafetyDGX agent

arXiv:2605.15164v1 Announce Type: cross Abstract: This position paper argues that behavioural assurance, even when carefully designed, is being asked to carry safety claims it cannot verify. AI govern

Precise Verification of Transformers through ReLU-Catalyzed Abstraction Refinement

SafetyDGX agent

arXiv:2605.14294v1 Announce Type: new Abstract: Formal verification of transformers has become increasingly important due to their widespread deployment in safety-critical applications. Compared to cl

Pro-DG: Procedural Diffusion Guidance for Architectural Facade Generation

SafetyDGX agent

arXiv:2504.01571v2 Announce Type: replace-cross Abstract: We use hierarchical procedural rules for the generation of control maps within the stable diffusion framework to produce photo-realistic archi

Probabilistic Verification of Recurrent Neural Networks for Single and Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.14758v1 Announce Type: new Abstract: History-dependent policies induced by recurrent neural networks (RNNs) rely on latent hidden state dynamics, making verification in partially observable

Progent: Securing AI Agents with Privilege Control

SafetyDGX agent

arXiv:2504.11703v3 Announce Type: replace-cross Abstract: AI agents interact with external environments through tool calls, exposing them to attacks like indirect prompt injection that can trigger una

Prompting Policies for Multi-step Reasoning and Tool-Use in Black-box LLMs with Iterative Distillation of Experience

SafetyDGX agent

arXiv:2605.14443v1 Announce Type: new Abstract: The shift toward interacting with frozen, 'black-box' Large Language Models (LLMs) has transformed prompt engineering from a heuristic exercise into a c

Proximal Action Replacement for Behavior Cloning Actor-Critic in Offline Reinforcement Learning

SafetyDGX agent

arXiv:2602.07441v2 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL), which optimizes policies using a previously collected static dataset, is an important branch of RL. A pop

Proxy Compression for Language Modeling

SafetyDGX agent

arXiv:2602.04289v2 Announce Type: replace Abstract: Modern language models are trained almost exclusively on token sequences produced by a fixed tokenizer, an external lossless compressor often over U

Quantifying and Mitigating Premature Closure in Frontier LLMs

SafetyDGX agent

arXiv:2605.15000v1 Announce Type: cross Abstract: Premature closure, or committing to a conclusion before sufficient information is available, is a recognized contributor to diagnostic error but remai

Quantitative Video World Model Evaluation for Geometric-Consistency

SafetyDGX agent

arXiv:2605.15185v1 Announce Type: cross Abstract: Generative video models are increasingly studied as implicit world models, yet evaluating whether they produce physically plausible 3D structure and m

R2PS: Worst-Case Robust Real-Time Pursuit Strategies under Partial Observability

SafetyDGX agent

arXiv:2511.17367v2 Announce Type: replace Abstract: Computing worst-case robust strategies in pursuit-evasion games (PEGs) is time-consuming, especially when real-world factors like partial observabil

← Previous
1…140141142143144…214
Next →