AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
Human
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
19 May 2026

Spatial Blindness in Whole-Slide Multiple Instance Learning

ResearchDGX agent

arXiv:2605.17449v1 Announce Type: cross Abstract: Whole-slide MIL models are often called context-aware once graphs, Transform ers, or state-space modules are placed above patch embeddings. We show th

Spatially Aware Linear Transformer (SAL-T) for Particle Jet Tagging

Local AiDGX agent

arXiv:2510.23641v2 Announce Type: replace-cross Abstract: Transformers are very effective in capturing both global and local correlations within high-energy particle collisions, but they present deplo

SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning

Model ReleasesDGX agent

arXiv:2605.18209v1 Announce Type: cross Abstract: Spatial question answering over egocentric video is a challenging task that requires Vision-Language Models (VLMs) to reason about 3D object positions

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Spatiotemporal Robustness of Temporal Logic Tasks using Multi-Objective Reasoning

AgentsDGX agent

arXiv:2603.29868v2 Announce Type: replace Abstract: The reliability of autonomous systems depends on their robustness, i.e., their ability to meet their objectives under uncertainty. In this paper, we

Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias

SafetyDGX agent

arXiv:2509.22061v2 Announce Type: replace-cross Abstract: Speech Continuation (SC) is the task of generating a coherent extension of a spoken prompt while preserving both semantic context and speaker

SpecSem-Net: Integrating Spectral and Semantic Features for Robust AI-generated Video Detection

Model ReleasesDGX agent

arXiv:2605.17311v1 Announce Type: new Abstract: The remarkable visual fidelity of recent commercial video generative models, such as Sora and Veo, renders robust AI-generated video detection increasin

Spectral Progressive Diffusion for Efficient Image and Video Generation

ResearchDGX agent

arXiv:2605.18736v1 Announce Type: new Abstract: Diffusion models have been shown to implicitly generate visual content autoregressively in the frequency domain, where low-frequency components are gene

Speech-Guided Multimodal Learning for Vocal Tract Segmentation in Real-Time MRI

Local AiDGX agent

arXiv:2605.18466v1 Announce Type: new Abstract: Segmenting vocal tract articulators in real-time MRI (rtMRI) is a challenging dynamic image segmentation problem characterized by low contrast, rapid mo

Speech-Hands: A Self-Reflection Voice Agentic Approach to Speech Recognition and Audio Reasoning with Omni Perception

AgentsDGX agent

arXiv:2601.09413v2 Announce Type: replace-cross Abstract: We introduce a voice-agentic framework that learns one critical omni-understanding skill: knowing when to trust itself versus when to consult

Spherical Harmonic Optimal Transport: Application to Climate Models Comparisons

HardwareDGX agent

arXiv:2605.18389v1 Announce Type: new Abstract: Optimal transport provides a powerful framework for comparing measures while respecting the geometry of their support, but comes with an expensive compu

Spherical Steering: Geometry-Aware Activation Rotation for Language Models

ResearchDGX agent

arXiv:2602.08169v2 Announce Type: replace-cross Abstract: Inference-time steering offers a promising way to control language models (LMs) without retraining. However, standard approaches typically rel

Spherical VAE with Cluster-Aware Feasible Regions: Guaranteed Prevention of Posterior Collapse

Model ReleasesDGX agent

arXiv:2603.10935v4 Announce Type: replace-cross Abstract: Variational autoencoders (VAEs) frequently suffer from posterior collapse, where the latent variables become uninformative as the approximate

SPIKE: An Adaptive Dual Controller Framework for Cost-Efficient Long-Horizon Game Agents

ResearchDGX agent

arXiv:2605.18636v1 Announce Type: new Abstract: Long-horizon multimodal agents in open-world games must stay goal-directed across many low-level interactions under tight token and latency budgets. Exi

Spiker-LL: An Energy-Efficient FPGA Accelerator Enabling Adaptive Local Learning in Spiking Neural Networks

Local AiDGX agent

arXiv:2605.18003v1 Announce Type: cross Abstract: Deploying adaptive intelligence at the edge remains challenging due to the high computational and energy cost of training neural models. Spiking Neura

SRC-Flow: Compact Semantic Representations Enable Normalizing Flows for Image Generation

TutorialsDGX agent

arXiv:2605.18267v1 Announce Type: new Abstract: Normalizing flows (NFs) provide exact likelihoods and deterministic invertible sampling, but have historically lagged behind diffusion models for large-

SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language Reasoning

SafetyDGX agent

arXiv:2510.16416v4 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have shown remarkable abilities by integrating large language models with visual inputs. However, they often fai

SSTL: Self-Sensing Tendon Loop for Hysteresis Modeling and Compensation in Tendon-Sheath Mechanisms

ResearchDGX agent

arXiv:2605.16870v1 Announce Type: new Abstract: Flexible endoscopic robots enable minimally invasive access through natural orifices, but their control accuracy is limited by configuration-dependent h

ST-BCP: Tightening Coverage Bound for Backward Conformal Prediction via Non-Conformity Score Transformation

ResearchDGX agent

arXiv:2602.01733v2 Announce Type: replace-cross Abstract: Conformal Prediction (CP) provides a statistical framework for uncertainty quantification that constructs prediction sets with coverage guaran

Stabilizing, Scaling & Enhancing MeanFlow for Large-scale Diffusion Distillation

Model ReleasesDGX agent

arXiv:2605.17834v1 Announce Type: new Abstract: Diffusion models exhibit remarkable generative capability, but their high latency limits practical deployment. Many studies have attempted to reduce sam

Stabilizing Temporal Inference Dynamics for Online Surgical Phase Recognition

ResearchDGX agent

arXiv:2605.16387v1 Announce Type: cross Abstract: Online Surgical Phase Recognition (SPR) models can reach high frame-wise accuracy, yet their predictions often lack temporal stability, fragmenting wo

Stable and Near-Reversible Diffusion ODE Solvers for Image Editing

SafetyDGX agent

arXiv:2605.16399v1 Announce Type: new Abstract: The inversion of diffusion models plays a central role in image editing. Algebraically reversible ODE solvers provide an appealing approach to diffusion

Stable Audio 3

HardwareDGX agent

arXiv:2605.17991v1 Announce Type: cross Abstract: Stable Audio 3 is a family of fast latent diffusion models (small, medium, large) for variable-length audio generation and editing. Since our models c

Stable Routing for Mixture-of-Experts in Class-Incremental Learning

SafetyDGX agent

arXiv:2605.17571v1 Announce Type: new Abstract: Class-incremental learning (CIL) requires models to learn new classes sequentially while preserving prior knowledge. Recently, approaches that combine p

StableHand: Quality-Aware Flow Matching for World-Space Dual-Hand Motion Estimation from Egocentric Video

SafetyDGX agent

arXiv:2605.18553v1 Announce Type: cross Abstract: Recovering world space 4D motion of two interacting hands from egocentric video is a fundamental capability for supervising robot policy learning, whe

StableVLA: Towards Robust Vision-Language-Action Models without Extra Data

Model ReleasesDGX agent

arXiv:2605.18287v1 Announce Type: new Abstract: It is infeasible to encompass all possible disturbances within the training dataset. This raises a critical question regarding the robustness of Vision-

StAD: Stein Amortized Divergence for Fast Likelihoods with Diffusion and Flow

TutorialsDGX agent

arXiv:2605.16486v1 Announce Type: cross Abstract: Diffusion and flow-based models are ubiquitously used for generative modelling and density estimation. They admit a deterministic probability flow ord

STAG-CN: Spatio-Temporal Apiary Graph Convolutional Network for Disease Onset Prediction in Beehive Sensor Networks

ResearchDGX agent

arXiv:2603.14462v2 Announce Type: replace-cross Abstract: Honey bee colony losses threaten global pollination services, yet current monitoring systems treat each hive as an isolated unit, ignoring the

Starve to Perceive: Taming Lazy Perception in VLMs with Constrained Visual Bandwidth

TutorialsDGX agent

arXiv:2605.18603v1 Announce Type: new Abstract: Vision-Language Models (VLMs) deployed as situated agents in high-resolution visual environments require active perception -- the ability to dynamically

State-Conditional Adversarial Learning: An Off-Policy Visual Domain Transfer Method for End-to-End Imitation Learning

SafetyDGX agent

arXiv:2512.05335v3 Announce Type: replace Abstract: We study visual domain transfer for end-to-end imitation learning in a realistic and challenging setting where target-domain data are strictly off-p

State Contamination in Memory-Augmented LLM Agents

SafetyDGX agent

arXiv:2605.16746v1 Announce Type: new Abstract: LLM agents increasingly rely on persistent state, including transcripts, summaries, retrieved context, and memory buffers, to support long-horizon inter

State-of-the-Art Claims Require State-of-the-Art Evidence

Model ReleasesDGX agent

arXiv:2605.17273v1 Announce Type: cross Abstract: State-of-the-Art (SOTA) claims pervade Artificial Intelligence (AI) and Machine Learning (ML) research. These claims rest on benchmark evaluations, wh

Statistical Hand Shape Modeling from Clinical CT Scans Using Deep Learning and Implicit Skinning

ResearchDGX agent

arXiv:2605.16980v1 Announce Type: new Abstract: Accurate segmentation and statistical shape modeling of hand anatomy have significant implications for medical diagnostics, ergonomics, and biomechanics

Statistical Limits and Efficient Algorithms for Differentially Private Federated Learning

Model ReleasesDGX agent

arXiv:2605.18656v1 Announce Type: cross Abstract: Federated Learning is a leading framework for training ML and AI models collaboratively across numerous user devices or databases. We study the trade-

Statistical Unlearning of Distributions: A Hypothesis Testing Approach

ResearchDGX agent

arXiv:2605.16645v1 Announce Type: cross Abstract: Machine learning systems increasingly face requirements to forget not only individual data points, but entire domains of information, such as toxic la

StatQAT: Statistical Quantizer Optimization for Deep Networks

ResearchDGX agent

arXiv:2605.17745v1 Announce Type: cross Abstract: Quantization is essential for reducing the computational cost and memory usage of deep neural networks, enabling efficient inference on low-precision

SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation

Model ReleasesDGX agent

arXiv:2511.19320v2 Announce Type: replace Abstract: Preserving first-frame identity while ensuring precise motion control is a fundamental challenge in human image animation. The Image-to-Motion Bindi

Stein Diffusion Guidance: Training-Free Posterior Correction for Sampling Beyond High-Density Regions

ResearchDGX agent

arXiv:2507.05482v3 Announce Type: replace Abstract: Training-free diffusion guidance offers a flexible framework for leveraging off-the-shelf classifiers without additional training. Yet, current appr

Step-wise Rubric Rewards for LLM Reasoning

ResearchDGX agent

arXiv:2605.17291v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is widely used to improve reasoning in large language models, but rewards only final-answer correc

Stochastic Penalty-Barrier Methods for Constrained Machine Learning

SafetyDGX agent

arXiv:2605.18618v1 Announce Type: cross Abstract: Constrained machine learning enables fairness-aware training, physics-informed neural networks, and integration of symbolic domain knowledge into stat

Stochastic Regret Guarantees for Online Zeroth- and First-Order Bilevel Optimization

ResearchDGX agent

arXiv:2511.01126v2 Announce Type: replace Abstract: Online bilevel optimization (OBO) is a powerful framework for machine learning problems where both outer and inner objectives evolve over time, requ

Stop When Reasoning Converges: Semantic-Preserving Early Exit for Reasoning Models

SafetyDGX agent

arXiv:2605.17672v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance by generating long chains of thought (CoT), but often overthink, continuing to reason after a s

Strategic Over-Parameterization for Generalizable Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2605.16470v1 Announce Type: cross Abstract: Adapting large language models (LLMs) to downstream tasks via full fine-tuning is increasingly impractical due to its computational and memory demands

StreamingEffect: Real-Time Human-Centric Video Effect Generation

HardwareDGX agent

arXiv:2605.17019v1 Announce Type: new Abstract: Streaming video effect generation is highly desirable for live human-centric applications such as e-commerce streaming, entertainment, and vlogging, yet

StreamingTalker: Audio-driven 3D Facial Animation with Autoregressive Diffusion Model

ResearchDGX agent

arXiv:2511.14223v3 Announce Type: replace Abstract: This paper focuses on the task of speech-driven 3D facial animation, which aims to generate realistic and synchronized facial motions driven by spee

StreamPro: From Reactive Perception to Proactive Decision-Making in Streaming Video

Model ReleasesDGX agent

arXiv:2605.16381v1 Announce Type: cross Abstract: Proactive streaming video understanding requires models to continuously process video streams and decide when to respond, rather than merely what to r

Stress-Testing Neural Network Verifiers with Provably Robust Instances

ResearchDGX agent

arXiv:2605.17153v1 Announce Type: new Abstract: Neural network verifiers aim to provide formal guarantees on model behavior, but existing verification benchmarks are fundamentally limited by their lac

Stretch-ICP: A Continuous-Trajectory Registration and Deskewing Algorithm in Scenarios of Aggressive Motions

ResearchDGX agent

arXiv:2605.17264v1 Announce Type: new Abstract: Robust robotic autonomy remains challenging in complex environments, where loss of stability on uneven or slippery terrain can induce extreme accelerati

STRIDE: A Self-Reflective Agent Framework for Reliable Automatic Equation Discovery

AgentsDGX agent

arXiv:2605.17790v1 Announce Type: new Abstract: LLM-based equation discovery offers a promising route to recovering symbolic laws from data, but many systems still rely on generation-centered loops th

STRIDE-AI: A Threat Modeling Framework for Generative AI Security Assessment

ApplicationsDGX agent

arXiv:2605.17163v1 Announce Type: cross Abstract: Traditional cybersecurity methodologies target deterministic systems and fail to address the probabilistic nature of AI, leaving systems vulnerable to

StrLoRA: Towards Streaming Continual Visual Instruction Tuning for MLLMs

Model ReleasesDGX agent

arXiv:2605.16353v1 Announce Type: cross Abstract: Continual Visual Instruction Tuning (CVIT) enables Multimodal Large Language Models to incrementally acquire new abilities. However, existing CVIT met

Stroke of Surprise: Progressive Semantic Illusions in Vector Sketching

ResearchDGX agent

arXiv:2602.12280v2 Announce Type: replace Abstract: Visual illusions traditionally rely on spatial manipulations such as multi-view consistency. In this work, we introduce Progressive Semantic Illusio

StructLens: A Structural Lens for Language Models via Maximum Spanning Trees

ResearchDGX agent

arXiv:2603.03328v2 Announce Type: replace-cross Abstract: Language exhibits inherent structures, a property that explains both language acquisition and language change. Given this characteristic, we e

Structure-Aware Masking for Protein Representation Learning

SafetyDGX agent

arXiv:2605.16581v1 Announce Type: new Abstract: Masked language modeling (MLM) is the standard objective for training protein language models, typically implemented by randomly masking individual resi

Structured Labeling Enables Faster Vision-Language Models for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2506.05442v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) offer a promising approach to end-to-end autonomous driving due to their human-like reasoning capabilities. Howe

Structured Neural Marked Point Processes for Interpretable Event Interaction Modeling

Model ReleasesDGX agent

arXiv:2605.17568v1 Announce Type: new Abstract: Multi-class event streams arise in numerous real-world applications, where uncovering structured, interpretable inter-event relationships, together with

STT-Arena: A More Realistic Environment for Tool-Using with Spatio-Temporal Dynamics

Model ReleasesDGX agent

arXiv:2605.18548v1 Announce Type: cross Abstract: Large language models (LLMs) deployed in real-world agentic applications must be capable of replanning and adapting when mid-task disruptions invalida

StyleText: A Large-Scale Dataset and Benchmark for Stylized Scene Text Inpainting

Model ReleasesDGX agent

arXiv:2605.17309v1 Announce Type: cross Abstract: We present StyleText, a large-scale dataset and benchmark for localized scene-text inpainting with style preservation. StyleText contains 28,518 image

Subject-Specific Analysis of Self-Initiated Attention Shifts from EEG with Controlled Internal and External Attention Conditions

ResearchDGX agent

arXiv:2605.18251v1 Announce Type: cross Abstract: Self-initiated attention shifts play a critical role in voluntary behavior but are difficult to study due to the absence of explicit temporal markers.

Supervise Less, See More: Training-free Nuclear Instance Segmentation with Prototype-Guided Prompting

Model ReleasesDGX agent

arXiv:2511.19953v2 Announce Type: replace Abstract: Accurate nuclear instance segmentation is a pivotal task in computational pathology, supporting data-driven clinical insights and facilitating downs

Supervised contrastive learning for cell stage classification of animal embryos

ResearchDGX agent

arXiv:2502.07360v3 Announce Type: replace-cross Abstract: Videomicroscopy, when combined with machine learning, offers a promising approach for studying the early development of in vitro produced (IVP

← Previous
1…665666667668669…1025
Next →