AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
Human
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
21 Apr 2026

SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe

ApplicationsDGX agent

arXiv:2410.05248v4 Announce Type: replace Abstract: To acquire instruction-following capabilities, large language models (LLMs) undergo instruction tuning, where they are trained on instruction-respon

Sharpening Lightweight Models for Generalized Polyp Segmentation: A Boundary Guided Distillation from Foundation Models

SafetyDGX agent

arXiv:2604.17865v1 Announce Type: new Abstract: Automated polyp segmentation is critical for early colorectal cancer detection and its prevention, yet remains challenging due to weak boundaries, large

Shepherding UAV Swarm with Action Prediction Based on Movement Constraints

SafetyDGX agent

arXiv:2604.17189v1 Announce Type: new Abstract: In this study, we propose a new sheepdog-inspired control method for a swarm of small unmanned aerial vehicles (UAVs), which predicts the swarm behavior

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Shifting the Gradient: Understanding How Defensive Training Methods Protect Language Model Integrity

ApplicationsDGX agent

arXiv:2604.16423v1 Announce Type: new Abstract: Defensive training methods such as positive preventative steering (PPS) and inoculation prompting (IP) offer surprising results through seemingly simila

SHRUG-FM: Reliability-Aware Foundation Models for Earth Observation

Model ReleasesDGX agent

arXiv:2511.10370v2 Announce Type: replace Abstract: Geospatial foundation models (GFMs) for Earth observation often fail to perform reliably in environments underrepresented during pretraining. We int

SIF: Semantically In-Distribution Fingerprints for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.17041v1 Announce Type: new Abstract: The public accessibility of large vision-language models (LVLMs) raises serious concerns about unauthorized model reuse and intellectual property infrin

SigGate-GT: Taming Over-Smoothing in Graph Transformers via Sigmoid-Gated Attention

Model ReleasesDGX agent

arXiv:2604.17324v1 Announce Type: new Abstract: Graph transformers achieve strong results on molecular and long-range reasoning tasks, yet remain hampered by over-smoothing (the progressive collapse o

SIGMA: A Semantic-Grounded Instruction-Driven Generative Multi-Task Recommender at AliExpress

ApplicationsDGX agent

arXiv:2602.22913v2 Announce Type: replace-cross Abstract: With the rapid evolution of Large Language Models (LLMs), generative recommendation is gradually reshaping the paradigm of recommender systems

SignDPO: Multi-level Direct Preference Optimisation for Skeleton-based Gloss-free Sign Language Translation

Local AiDGX agent

arXiv:2604.18034v1 Announce Type: new Abstract: We present SignDPO, a novel multi-level Direct Preference Optimisation (DPO) framework designed to enhance the alignment of skeleton-based Sign Language

Singularity Formation: Synergy in Theoretical, Numerical and Machine Learning Approaches

ResearchDGX agent

arXiv:2604.16842v1 Announce Type: cross Abstract: This thesis develops numerical and theoretical approaches for understanding and analyzing singularity formation in Partial Differential Equations (PDE

SinkRouter: Sink-Aware Routing for Efficient Long-Context Decoding in Large Language and Multimodal Models

Model ReleasesDGX agent

arXiv:2604.16883v1 Announce Type: new Abstract: In long-context decoding for LLMs and LMMs, attention becomes increasingly memory-bound because each decoding step must load a large amount of KV-cache

SkillX: Automatically Constructing Skill Knowledge Bases for Agents

AgentsDGX agent

arXiv:2604.04804v2 Announce Type: replace Abstract: Learning from experience is critical for building capable large language model (LLM) agents, yet prevailing self-evolving paradigms remain inefficie

Sky2Ground: A Benchmark for Site Modeling under Varying Altitude

Model ReleasesDGX agent

arXiv:2603.13740v3 Announce Type: replace Abstract: We introduce Sky2Ground, a three-view dataset designed for varying altitude camera localization, correspondence learning, and reconstruction. The da

SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving

HardwareDGX agent

arXiv:2604.17627v1 Announce Type: new Abstract: Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone sea

SMILE-UHURA Challenge -- Small Vessel Segmentation at Mesoscopic Scale from Ultra-High Resolution 7T Magnetic Resonance Angiograms

ResearchDGX agent

arXiv:2411.09593v2 Announce Type: replace-cross Abstract: The human brain receives nutrients and oxygen through an intricate network of blood vessels. Pathology affecting small vessels, at the mesosco

SmoGVLM: A Small, Graph-enhanced Vision-Language Model

ResearchDGX agent

arXiv:2604.16517v1 Announce Type: cross Abstract: Large vision-language models (VLMs) achieve strong performance on multimodal tasks but often suffer from hallucination and poor grounding in knowledge

Sobolev Gradient Ascent for Optimal Transport: Barycenter Optimization and Convergence Analysis

ResearchDGX agent

arXiv:2505.13660v2 Announce Type: replace-cross Abstract: This paper introduces a new constraint-free concave dual formulation for the Wasserstein barycenter. Tailoring the vanilla dual gradient ascen

Soft Label Pruning and Quantization for Large-Scale Dataset Distillation

SafetyDGX agent

arXiv:2604.18135v1 Announce Type: new Abstract: Large-scale dataset distillation requires storing auxiliary soft labels that can be 30-40x larger on ImageNet-1K and 200x larger on ImageNet-21K than th

Sonata: A Hybrid World Model for Inertial Kinematics under Clinical Data Scarcity

Model ReleasesDGX agent

arXiv:2604.18058v1 Announce Type: new Abstract: We introduce Sonata, a compact latent world model for six-axis trunk IMU representation learning under clinical data scarcity. Clinical cohorts typicall

Source-Free Domain Adaptation with Vision-Language Prior

SafetyDGX agent

arXiv:2604.17748v1 Announce Type: new Abstract: Source-Free Domain Adaptation (SFDA) seeks to adapt a source model, which is pre-trained on a supervised source domain, for a target domain, with only a

SpaceDex: Generalizable Dexterous Grasping in Tiered Workspaces

ApplicationsDGX agent

arXiv:2604.17888v1 Announce Type: new Abstract: Generalizable grasping with high-degree-of-freedom (DoF) dexterous hands remains challenging in tiered workspaces, where occlusion, narrow clearances, a

SparrowSNN: A Hardware/software Co-design for Energy Efficient ECG Classification

ResearchDGX agent

arXiv:2406.06543v2 Announce Type: replace-cross Abstract: Deep learning has driven significant technological advancements, but its high energy consumption limits its use on battery-operated edge devic

(Sparse) Attention to the Details: Preserving Spectral Fidelity in ML-based Weather Forecasting Models

SafetyDGX agent

arXiv:2604.16429v1 Announce Type: cross Abstract: We introduce Mosaic, a probabilistic weather forecasting model that addresses two principal sources of spectral degradation in ML-based weather predic

Sparse Feature Coactivation Reveals Causal Semantic Modules in Large Language Models

ResearchDGX agent

arXiv:2506.18141v3 Announce Type: replace Abstract: We identify semantically coherent, context-consistent network components in large language models (LLMs) using coactivation of sparse autoencoder (S

SPaRSe-TIME: Saliency-Projected Low-Rank Temporal Modeling for Efficient and Interpretable Time Series Prediction

ApplicationsDGX agent

arXiv:2604.17350v1 Announce Type: cross Abstract: Time series forecasting is traditionally dominated by sequence-based architectures such as recurrent neural networks and attention mechanisms, which p

Spatial-Regularization-Aware Dual-Branch Collaborative Inference for Training-Free OVSS in Remote Sensing Imagery

Local AiDGX agent

arXiv:2601.21159v2 Announce Type: replace Abstract: High-resolution remote sensing images contain densely distributed objects with pronounced scale variations and complex boundaries, which impose high

SpatialImaginer: Towards Adaptive Visual Imagination for Spatial Reasoning

ResearchDGX agent

arXiv:2604.17385v1 Announce Type: new Abstract: Spatial intelligence, which refers to the ability to reason about geometric and physical structure from visual observations, remains a core challenge fo

SpatialStack: Layered Geometry-Language Fusion for 3D VLM Spatial Reasoning

Local AiDGX agent

arXiv:2603.27437v2 Announce Type: replace Abstract: Large vision-language models (VLMs) still struggle with reliable 3D spatial reasoning, a core capability for embodied and physical AI systems. This

Spatiotemporal Sycophancy: Negation-Based Gaslighting in Video Large Language Models

Model ReleasesDGX agent

arXiv:2604.17873v1 Announce Type: new Abstract: Video Large Language Models (Vid-LLMs) have demonstrated remarkable performance in video understanding tasks, yet their robustness under conversational

SpeakerSleuth: Can Large Audio-Language Models Judge Speaker Consistency across Multi-turn Dialogues?

Model ReleasesDGX agent

arXiv:2601.04029v2 Announce Type: replace Abstract: Large Audio-Language Models (LALMs) as judges have emerged as a prominent approach for evaluating speech generation quality, yet their ability to as

Spec-o3: A Tool-Augmented Vision-Language Agent for Rare Celestial Object Candidate Vetting via Automated Spectral Inspection

Model ReleasesDGX agent

arXiv:2601.06498v2 Announce Type: replace Abstract: Due to the limited generalization and interpretability of deep learning classifiers, The final vetting of rare celestial object candidates still rel

Spectral bandits for smooth graph functions

SafetyDGX agent

arXiv:2604.18420v1 Announce Type: cross Abstract: Smooth functions on graphs have wide applications in manifold and semi-supervised learning. In this paper, we study a bandit problem where the payoffs

Spectral Forensics of Diffusion Attention Graphs for Copy-Move Forgery Detection

ResearchDGX agent

arXiv:2604.17287v1 Announce Type: new Abstract: Copy-move forgery, where a region within an image is duplicated to hide or fabricate content, remains a persistent threat to visual media integrity. We

Speculative Decoding for Autoregressive Video Generation

ResearchDGX agent

arXiv:2604.17397v1 Announce Type: new Abstract: Autoregressive video diffusion is emerging as a promising paradigm for streaming video synthesis, with step distillation serving as the primary means of

Speculative Verification: Exploiting Information Gain to Refine Speculative Decoding

SafetyDGX agent

arXiv:2509.24328v2 Announce Type: replace Abstract: LLMs have low GPU efficiency and high latency due to autoregressive decoding. Speculative decoding (SD) mitigates this using a small draft model to

SpeechMedAssist: Efficiently and Effectively Adapting Speech Language Models for Medical Consultation

Model ReleasesDGX agent

arXiv:2601.04638v2 Announce Type: replace Abstract: Medical consultations are intrinsically speech-centric. However, most prior works focus on long-text-based interactions, which are cumbersome and pa

SPENCE: A Syntactic Probe for Detecting Contamination in NL2SQL Benchmarks

Model ReleasesDGX agent

arXiv:2604.17771v1 Announce Type: new Abstract: Large language models (LLMs) have achieved strong performance on natural language to SQL (NL2SQL) benchmarks, yet their reported accuracy may be inflate

SpidR-Adapt: A Universal Speech Representation Model for Few-Shot Adaptation

ResearchDGX agent

arXiv:2512.21204v2 Announce Type: replace Abstract: Human infants, with only a few hundred hours of speech exposure, acquire basic units of new languages, highlighting a striking efficiency gap compar

Spike-NVPT: Learning Robust Visual Prompts via Bio-Inspired Temporal Filtering and Discretization

Model ReleasesDGX agent

arXiv:2604.18284v1 Announce Type: new Abstract: Pre-trained vision models have found widespread application across diverse domains. Prompt tuning-based methods have emerged as a parameter-efficient pa

SpiralFormer: Looped Transformers Can Learn Hierarchical Dependencies via Multi-Resolution Recursion

Model ReleasesDGX agent

arXiv:2602.11698v2 Announce Type: replace Abstract: Recursive (looped) Transformers decouple computational depth from parameter depth by repeatedly applying shared layers, providing an explicit archit

SpiralThinker: Latent Reasoning through an Iterative Process with Text-Latent Interleaving

SafetyDGX agent

arXiv:2511.08983v2 Announce Type: replace Abstract: Recent advances in large reasoning models have been driven by reinforcement learning and test-time scaling, accompanied by growing interest in laten

Splatography: Sparse multi-view dynamic Gaussian Splatting for filmmaking challenges

ResearchDGX agent

arXiv:2511.05152v2 Announce Type: replace Abstract: Deformable Gaussian Splatting (GS) accomplishes photorealistic dynamic 3-D reconstruction from dense multi-view video (MVV) by learning to deform a

SPOT: Single-Shot Positioning via Trainable Near-Field Rainbow Beamforming

ResearchDGX agent

arXiv:2511.11391v3 Announce Type: replace Abstract: Phase-time arrays, which integrate phase shifters (PSs) and true-time delays (TTDs), have emerged as a cost-effective architecture for generating fr

Spotlights and Blindspots: Evaluation Machine-Generated Text Detection

ResearchDGX agent

arXiv:2604.16607v1 Announce Type: new Abstract: With the rise of generative language models, machine-generated text detection has become a critical challenge. A wide variety of models is available, bu

SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models

SafetyDGX agent

arXiv:2604.16995v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a promising paradigm for training reasoning-oriented models by leveraging rule-based reward signals. However,

SQL Query Engine: A Self-Healing LLM Pipeline for Natural Language to PostgreSQL Translation

Model ReleasesDGX agent

arXiv:2604.16511v1 Announce Type: cross Abstract: We present SQL Query Engine, an open-source, self-hosted service that translates natural language questions into validated PostgreSQL queries through

ST-pi: Structured SpatioTemporal VLA for Robotic Manipulation

Local AiDGX agent

arXiv:2604.17880v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have achieved great success on general robotic tasks, but still face challenges in fine-grained spatiotemporal man

Stability-Weighted Decoding for Diffusion Language Models

ResearchDGX agent

arXiv:2604.17068v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) enable parallel text generation by iteratively denoising a fully masked sequence, unmasking a subset of masked t

Stable Language Guidance for Vision-Language-Action Models

ResearchDGX agent

arXiv:2601.04052v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have demonstrated impressive capabilities in generalized robotic control; however, they remain notoriously

Stable On-Policy Distillation through Adaptive Target Reformulation

Model ReleasesDGX agent

arXiv:2601.07155v2 Announce Type: replace Abstract: Knowledge distillation (KD) is a widely adopted technique for transferring knowledge from large language models to smaller student models; however,

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement

Model ReleasesDGX agent

arXiv:2604.17887v1 Announce Type: new Abstract: Inverse Dynamics Models (IDMs) map visual observations to low-level action commands, serving as central components for data labeling and policy executio

StableMTL: Repurposing Latent Diffusion Models for Multi-Task Learning from Partially Annotated Synthetic Datasets

ResearchDGX agent

arXiv:2506.08013v2 Announce Type: replace Abstract: Multi-task learning for dense prediction is limited by the need for extensive annotation for every task, though recent works have explored training

STaD: Scaffolded Task Design for Identifying Compositional Skill Gaps in LLMs

Model ReleasesDGX agent

arXiv:2604.18177v1 Announce Type: new Abstract: Benchmarks are often used as a standard to understand LLM capabilities in different domains. However, aggregate benchmark scores provide limited insight

StageMem: Lifecycle-Managed Memory for Language Models

ResearchDGX agent

arXiv:2604.16774v1 Announce Type: new Abstract: Long-horizon language model systems increasingly rely on persistent memory, yet many current designs still treat memory primarily as a static store: wri

StealthGraph: Exposing Domain-Specific Risks in LLMs through Knowledge-Graph-Guided Harmful Prompt Generation

SafetyDGX agent

arXiv:2601.04740v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied in specialized domains such as finance and healthcare, where they introduce unique safety risk

STEP-PD: Stage-Aware and Explainable Parkinson's Disease Severity Classification Using Multimodal Clinical Assessments

ResearchDGX agent

arXiv:2604.17611v1 Announce Type: new Abstract: Parkinson's disease (PD) is a progressive disorder in which symptom burden and functional impairment evolve over time, making severity staging essential

StepPO: Step-Aligned Policy Optimization for Agentic Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.18401v1 Announce Type: new Abstract: General agents have given rise to phenomenal applications such as OpenClaw and Claude Code. As these agent systems (a.k.a. Harnesses) strive for bolder

Still Between Us? Evaluating and Improving Voice Assistant Robustness to Third-Party Interruptions

ApplicationsDGX agent

arXiv:2604.17358v1 Announce Type: new Abstract: While recent Spoken Language Models (SLMs) have been actively deployed in real-world scenarios, they lack the capability to discern Third-Party Interrup

STL-Based Motion Planning and Uncertainty-Aware Risk Analysis for Human-Robot Collaboration with a Multi-Rotor Aerial Vehicle

SafetyDGX agent

arXiv:2509.10692v3 Announce Type: replace Abstract: This paper presents a motion planning and risk analysis framework for enhancing human-robot collaboration with a Multi-Rotor Aerial Vehicle. The pro

Stochastic Control Methods for Optimization

Model ReleasesDGX agent

arXiv:2601.01248v4 Announce Type: replace-cross Abstract: In this work, we investigate a stochastic control framework for global optimization over both Euclidean spaces and the Wasserstein space of pr

← Previous
1…899900901902903…998
Next →