AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
Human
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
14 Apr 2026

Spatio-Temporal Difference Guided Motion Deblurring with the Complementary Vision Sensor

ApplicationsDGX agent

arXiv:2604.10554v1 Announce Type: new Abstract: Motion blur arises when rapid scene changes occur during the exposure period, collapsing rich intra-exposure motion into a single RGB frame. Without exp

Spatiotemporal-Aware Bit-Flip Injection on DNN-based Advanced Driver Assistance Systems (extended version)

SafetyDGX agent

arXiv:2604.03753v2 Announce Type: replace-cross Abstract: Modern advanced driver assistance systems (ADAS) rely on deep neural networks (DNNs) for perception and planning. Since DNNs' parameters resid

Speaking to No One: Ontological Dissonance and the Double Bind of Conversational AI

SafetyDGX agent

arXiv:2604.10833v1 Announce Type: cross Abstract: Recent reports indicate that sustained interaction with conversational artificial intelligence (AI) systems can, in a small subset of users, contribut

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Specificity-aware reinforcement learning for fine-grained open-world classification

TutorialsDGX agent

arXiv:2603.03197v3 Announce Type: replace Abstract: Classifying fine-grained visual concepts under open-world settings, i.e., without a predefined label set, demands models to be both accurate and spe

SpecMoE: A Fast and Efficient Mixture-of-Experts Inference via Self-Assisted Speculative Decoding

Model ReleasesDGX agent

arXiv:2604.10152v1 Announce Type: new Abstract: The Mixture-of-Experts (MoE) architecture has emerged as a promising approach to mitigate the rising computational costs of large language models (LLMs)

Spectral Kernel Dynamics via Maximum Caliber: Fixed Points, Geodesics, and Phase Transitions

ResearchDGX agent

arXiv:2604.09745v1 Announce Type: cross Abstract: We derive a closed-form geometric functional for kernel dynamics on finite graphs by applying the Maximum Caliber (MaxCal) variational principle to th

SpectralLoRA: Is Low-Frequency Structure Sufficient for LoRA Adaptation? A Spectral Analysis of Weight Updates

ResearchDGX agent

arXiv:2604.10649v1 Announce Type: cross Abstract: We present a systematic empirical study of the spectral structure of LoRA weight updates. Through 2D Discrete Cosine Transform (DCT) analysis of train

SpeechLess: Micro-utterance with Personalized Spatial Memory-aware Assistant in Everyday Augmented Reality

ResearchDGX agent

arXiv:2602.00793v2 Announce Type: replace-cross Abstract: Speaking aloud to a wearable AR assistant in public can be socially awkward, and re-articulating the same requests every day creates unnecessa

SPEED-Bench: A Unified and Diverse Benchmark for Speculative Decoding

Model ReleasesDGX agent

arXiv:2604.09557v1 Announce Type: cross Abstract: Speculative Decoding (SD) has emerged as a critical technique for accelerating Large Language Model (LLM) inference. Unlike deterministic system optim

Spoiler Alert: Narrative Forecasting as a Metric for Tension in LLM Storytelling

Model ReleasesDGX agent

arXiv:2604.09854v1 Announce Type: new Abstract: LLMs have so far failed both to generate consistently compelling stories and to recognize this failure--on the leading creative-writing benchmark (EQ-Be

SpotFormer: Multi-Scale Spatio-Temporal Transformer for Facial Expression Spotting

Local AiDGX agent

arXiv:2407.20799v3 Announce Type: replace Abstract: Facial expression spotting, identifying periods where facial expressions occur in a video, is a significant yet challenging task in facial expressio

Spotlight and Shadow: Attention-Guided Dual-Anchor Introspective Decoding for MLLM Hallucination Mitigation

TutorialsDGX agent

arXiv:2604.10071v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable reasoning capabilities yet continue to suffer from hallucination, where generated

SRBench: A Comprehensive Benchmark for Sequential Recommendation with Large Language Models

Model ReleasesDGX agent

arXiv:2604.09553v1 Announce Type: cross Abstract: LLM development has aroused great interest in Sequential Recommendation (SR) applications. However, comprehensive evaluation of SR models remains lack

Stability of a Generalized Debiased Lasso with Applications to Resampling-Based Variable Selection

ResearchDGX agent

arXiv:2405.03063v2 Announce Type: replace-cross Abstract: We propose a generalized debiased Lasso estimator based on a stability principle. When a single column of the design matrix is perturbed, the

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs

ResearchDGX agent

arXiv:2509.22220v2 Announce Type: replace Abstract: Prevalent semantic speech tokenizers, designed to capture linguistic content, are surprisingly fragile. We find they are not robust to meaning-irrel

StableTTA: Training-Free Test-Time Adaptation that Improves Model Accuracy on ImageNet1K to 96%

Model ReleasesDGX agent

arXiv:2604.04552v2 Announce Type: replace-cross Abstract: Ensemble methods are widely used to improve predictive performance, but their effectiveness often comes at the cost of increased memory usage

StaMo: Unsupervised Learning of Generalizable Robot Motion from Compact State Representation

SafetyDGX agent

arXiv:2510.05057v2 Announce Type: replace-cross Abstract: A fundamental challenge in embodied intelligence is developing expressive and compact state representations for efficient world modeling and d

STaR-DRO: Stateful Tsallis Reweighting for Group-Robust Structured Prediction

Model ReleasesDGX agent

arXiv:2604.09737v1 Announce Type: cross Abstract: Structured prediction requires models to generate ontology-constrained labels, grounded evidence, and valid structure under ambiguity, label skew, and

STARS: Skill-Triggered Audit for Request-Conditioned Invocation Safety in Agent Systems

Model ReleasesDGX agent

arXiv:2604.10286v1 Announce Type: new Abstract: Autonomous language-model agents increasingly rely on installable skills and tools to complete user tasks. Static skill auditing can expose capability s

StarVLA-alpha: Reducing Complexity in Vision-Language-Action Systems

Model ReleasesDGX agent

arXiv:2604.11757v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for building general-purpose robotic agents. However, the VLA landsc

Steered LLM Activations are Non-Surjective

SafetyDGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

STGV: Spatio-Temporal Hash Encoding for Gaussian-based Video Representation

ResearchDGX agent

arXiv:2604.10910v1 Announce Type: new Abstract: 2D Gaussian Splatting (2DGS) has recently become a promising paradigm for high-quality video representation. However, existing methods employ content-ag

Stop Fixating on Prompts: Reasoning Hijacking and Constraint Tightening for Red-Teaming LLM Agents

SafetyDGX agent

arXiv:2604.05549v2 Announce Type: replace Abstract: With the widespread application of LLM-based agents across various domains, their complexity has introduced new security threats. Existing red-team

STORM: End-to-End Referring Multi-Object Tracking in Videos

Model ReleasesDGX agent

arXiv:2604.10527v1 Announce Type: cross Abstract: Referring multi-object tracking (RMOT) is a task of associating all the objects in a video that semantically match with given textual queries or refer

StreamServe: Adaptive Speculative Flows for Low-Latency Disaggregated LLM Serving

HardwareDGX agent

arXiv:2604.09562v1 Announce Type: cross Abstract: Efficient LLM serving must balance throughput and latency across diverse, bursty workloads. We introduce StreamServe, a disaggregated prefill decode s

StreetDesignAI: A Multi-Persona Evaluation System for Inclusive Infrastructure Design

ResearchDGX agent

arXiv:2601.15671v2 Announce Type: replace-cross Abstract: Designing cycling infrastructure requires balancing the competing needs of diverse user groups, yet designers often struggle to anticipate how

Structural Consequences of Policy-Based Interventions on the Global Supply Chain Network

SafetyDGX agent

arXiv:2604.11479v1 Announce Type: new Abstract: As global political tensions rise and the anticipation of additional tariffs from the United States on international trade increases, the issues of econ

Structural Gating and Effect-aligned Lag-resolved Temporal Causal Discovery Framework with Application to Heat-Pollution Extremes

SafetyDGX agent

arXiv:2604.10371v1 Announce Type: new Abstract: This study proposes Structural Gating and Effect-aligned Discovery for Temporal Causal Discovery (SGED-TCD), a novel and general framework for lag-resol

Structure-Grounded Knowledge Retrieval via Code Dependencies for Multi-Step Data Reasoning

AgentsDGX agent

arXiv:2604.10516v1 Announce Type: new Abstract: Selecting the right knowledge is critical when using large language models (LLMs) to solve domain-specific data analysis tasks. However, most retrieval-

Structured Causal Video Reasoning via Multi-Objective Alignment

SafetyDGX agent

arXiv:2604.04415v2 Announce Type: replace Abstract: Human understanding of video dynamics is typically grounded in a structured mental representation of entities, actions, and temporal relations, rath

Structured State-Space Regularization for Compact and Generation-Friendly Image Tokenization

TutorialsDGX agent

arXiv:2604.11089v1 Announce Type: new Abstract: Image tokenizers are central to modern vision models as they often operate in latent spaces. An ideal latent space must be simultaneously compact and ge

STS-Mixer: Spatio-Temporal-Spectral Mixer for 4D Point Cloud Video Understanding

ResearchDGX agent

arXiv:2604.11637v1 Announce Type: new Abstract: 4D point cloud videos capture rich spatial and temporal dynamics of scenes which possess unique values in various 4D understanding tasks. However, most

STU-PID: Steering Token Usage via PID Controller for Efficient Large Language Model Reasoning

ResearchDGX agent

arXiv:2506.18831v2 Announce Type: replace Abstract: Large Language Models employing extended chain-of-thought (CoT) reasoning often suffer from the overthinking phenomenon, generating excessive and re

StyleBench: Evaluating thinking styles in Large Language Models

Model ReleasesDGX agent

arXiv:2509.20868v2 Announce Type: replace-cross Abstract: Structured reasoning can improve the inference performance of large language models (LLMs), but it also introduces computational cost and cont

Subargument Argumentation Frameworks: Separating Direct Conflict from Structural Dependency

ResearchDGX agent

arXiv:2601.12038v3 Announce Type: replace Abstract: Dung's abstract argumentation frameworks model acceptability solely in terms of an attack relation, thereby conflating two conceptually distinct asp

Suiren-1.0 Technical Report: A Family of Molecular Foundation Models

ResearchDGX agent

arXiv:2603.21942v2 Announce Type: replace-cross Abstract: We introduce Suiren-1.0, a family of molecular foundation models for the accurate modeling of diverse organic systems. Suiren-1.0 comprising t

Sustainable Transformer Neural Network Acceleration with Stochastic Photonic Computing

HardwareDGX agent

arXiv:2604.09759v1 Announce Type: cross Abstract: Transformers achieve state-of-the-art performance in natural language processing, vision, and scientific computing, but demand high computation and me

SVD-Prune: Training-Free Token Pruning For Efficient Vision-Language Models

SafetyDGX agent

arXiv:2604.11530v1 Announce Type: cross Abstract: Vision-Language Models (VLM) have revolutionized multimodal learning by jointly processing visual and textual information. Yet, they face significant

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning

TutorialsDGX agent

arXiv:2604.10228v1 Announce Type: new Abstract: Current multimodal models often suffer from shallow reasoning, leading to errors caused by incomplete or inconsistent thought processes. To address this

SWE-AGILE: A Software Agent Framework for Efficiently Managing Dynamic Reasoning Context

AgentsDGX agent

arXiv:2604.11716v1 Announce Type: new Abstract: Prior representative ReAct-style approaches in autonomous Software Engineering (SWE) typically lack the explicit System-2 reasoning required for deep an

SwinTextUNet: Integrating CLIP-Based Text Guidance into Swin Transformer U-Nets for Medical Image Segmentation

ResearchDGX agent

arXiv:2604.10000v1 Announce Type: new Abstract: Precise medical image segmentation is fundamental for enabling computer aided diagnosis and effective treatment planning. Traditional models that rely s

Switch-JustDance: Benchmarking Whole Body Motion Tracking Controllers Using a Commercial Console Game

Model ReleasesDGX agent

arXiv:2511.17925v3 Announce Type: replace-cross Abstract: Recent advances in whole-body robot control have enabled humanoid and legged robots to perform increasingly agile and coordinated motions. How

Symmetry-Aware Generative Modeling through Learned Canonicalization

TutorialsDGX agent

arXiv:2501.07773v3 Announce Type: replace Abstract: Generative modeling of symmetric densities has a range of applications in AI for science, from drug discovery to physics simulations. The existing g

SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation

Model ReleasesDGX agent

arXiv:2604.03723v2 Announce Type: replace Abstract: Controlling both camera motion and object dynamics is essential for coherent and expressive video generation, yet current methods typically handle o

SyncFix: Fixing 3D Reconstructions via Multi-View Synchronization

ResearchDGX agent

arXiv:2604.11797v1 Announce Type: new Abstract: We present SyncFix, a framework that enforces cross-view consistency during the diffusion-based refinement of reconstructed scenes. SyncFix formulates r

SynthAgent: Adapting Web Agents with Synthetic Supervision

ResearchDGX agent

arXiv:2511.06101v3 Announce Type: replace-cross Abstract: Web agents struggle to adapt to new websites due to the scarcity of environment specific tasks and demonstrations. Recent works have explored

Synthius-Mem: Brain-Inspired Hallucination-Resistant Persona Memory Achieving 94.4% Memory Accuracy and 99.6% Adversarial Robustness on LoCoMo

Model ReleasesDGX agent

arXiv:2604.11563v1 Announce Type: cross Abstract: Providing AI agents with reliable long-term memory that does not hallucinate remains an open problem. Current approaches to memory for LLM agents -- s

Tackling multiphysics problems via finite element-guided physics-informed operator learning

ResearchDGX agent

arXiv:2603.01420v2 Announce Type: replace Abstract: This work presents a finite element-guied physics-informed learning framework for multiphysics problems with coupled partial differential equations

TacMan-Turbo: Proactive Tactile Control for Robust and Efficient Articulated Object Manipulation

TutorialsDGX agent

arXiv:2508.02204v4 Announce Type: replace Abstract: Adept manipulation of articulated objects is essential for robots to operate successfully in human environments. Such manipulation requires both eff

TaFall: Balance-Informed Fall Detection via Passive Thermal Sensing

ApplicationsDGX agent

arXiv:2604.09693v1 Announce Type: cross Abstract: Falls are a major cause of injury and mortality among older adults, yet most incidents occur in private indoor environments where monitoring must bala

TAG-Head: Time-Aligned Graph Head for Plug-and-Play Fine-grained Action Recognition

Model ReleasesDGX agent

arXiv:2604.11498v1 Announce Type: new Abstract: Fine-grained human action recognition (FHAR) is challenging because visually similar actions differ by subtle spatio-temporal cues. Many recent systems

Tail-Aware Information-Theoretic Generalization for RLHF and SGLD

Model ReleasesDGX agent

arXiv:2604.10727v1 Announce Type: cross Abstract: Classical information-theoretic generalization bounds typically control the generalization gap through KL-based mutual information and therefore rely

Taking a Pulse on How Generative AI is Reshaping the Software Engineering Research Landscape

SafetyDGX agent

arXiv:2604.11184v1 Announce Type: cross Abstract: Context: Software engineering (SE) researchers increasingly study Generative AI (GenAI) while also incorporating it into their own research practices.

TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering

ResearchDGX agent

arXiv:2509.04123v2 Announce Type: replace Abstract: Text-to-story visualization is challenging due to the need for consistent interaction among multiple characters across frames. Existing methods stru

TAMISeg: Text-Aligned Multi-scale Medical Image Segmentation with Semantic Encoder Distillation

ResearchDGX agent

arXiv:2604.10912v1 Announce Type: new Abstract: Medical image segmentation remains challenging due to limited fine-grained annotations, complex anatomical structures, and image degradation from noise,

TAPNext++: What's Next for Tracking Any Point (TAP)?

ResearchDGX agent

arXiv:2604.10582v1 Announce Type: new Abstract: Tracking-Any-Point (TAP) models aim to track any point through a video which is a crucial task in AR/XR and robotics applications. The recently introduc

TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection

ResearchDGX agent

arXiv:2504.04099v2 Announce Type: replace-cross Abstract: Large Vision-Language Models have demonstrated remarkable capabilities, yet they suffer from hallucinations that limit practical deployment. W

Task2vec Readiness: Diagnostics for Federated Learning from Pre-Training Embeddings

SafetyDGX agent

arXiv:2604.10849v1 Announce Type: cross Abstract: Federated learning (FL) performance is highly sensitive to heterogeneity across clients, yet practitioners lack reliable methods to anticipate how a f

TCSA-UDA: Text-Driven Cross-Semantic Alignment for Unsupervised Domain Adaptation in Medical Image Segmentation

SafetyDGX agent

arXiv:2511.05782v2 Announce Type: replace Abstract: Unsupervised domain adaptation for medical image segmentation remains a significant challenge due to substantial domain shifts across imaging modali

Teaching Language Models How to Code Like Learners: Conversational Serialization for Student Simulation

Model ReleasesDGX agent

arXiv:2604.10720v1 Announce Type: new Abstract: Artificial models that simulate how learners act and respond within educational systems are a promising tool for evaluating tutoring strategies and feed

← Previous
1…956957958959960…989
Next →