AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
19,193 results
12 May 2026

Regret Minimization in Bilateral Trade With Perturbed Markets

ResearchDGX agent

arXiv:2605.10475v1 Announce Type: cross Abstract: We address the problem of maximizing Gain from Trade (GFT) in repeated buyer-seller exchanges subject to global budget balance constraints. While this

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models

ResearchDGX agent

arXiv:2605.10759v1 Announce Type: cross Abstract: Diffusion and flow-matching models scale because pretraining is supervised regression: a clean sample is noised analytically, and a model regresses ag

RelFlexformer: Efficient Attention 3D-Transformers for Integrable Relative Positional Encodings

ResearchDGX agent

arXiv:2605.10706v1 Announce Type: new Abstract: We present a new class of efficient attention mechanisms applying universal 3D Relative Positional Encoding (RPE) methods given by arbitrary integrable


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ReLibra: Routing-Replay-Guided Load Balancing for MoE Training in Reinforcement Learning

ResearchDGX agent

arXiv:2605.08639v1 Announce Type: new Abstract: Load imbalance is a long-standing challenge in Mixture-of-Experts (MoE) training and is exacerbated in reinforcement learning (RL) for LLMs, where hot e

Remix the Timbre: Diffusion-Based Style Transfer Across Polyphonic Stems

ResearchDGX agent

arXiv:2605.09259v1 Announce Type: cross Abstract: Timbre transfer aims to modify the timbral identity of a musical recording while preserving the original melody and rhythm. While single-instrument ti

Representative Action Selection for Large Action Space Bandit Families

ResearchDGX agent

arXiv:2505.18269v5 Announce Type: replace Abstract: We study the problem of selecting a subset from a large action space shared by a family of bandits. In many natural situations, while the nominal se

Resource-Aware Evolutionary Neural Architecture Search for Cardiac MRI Segmentation

ResearchDGX agent

arXiv:2605.08238v1 Announce Type: cross Abstract: Cardiac magnetic resonance (CMR) segmentation underpins quantitative assessment of ventricular structure and function, yet reliable delineation remain

ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing

ResearchDGX agent

arXiv:2605.08840v1 Announce Type: new Abstract: Large language models (LLMs) face growing challenges in efficient generative inference due to the increasing memory demands of Key-Value (KV) caches, es

Restoration-Aligned Generative Flow Models for Blind Motion Deblurring

ResearchDGX agent

arXiv:2605.08854v1 Announce Type: new Abstract: Generative flow models offer powerful priors learned from large-scale natural images, but directly adapting them to restoration tasks such as motion deb

Restoring Exploration after Post-Training: Latent Exploration Decoding for Large Reasoning Models

ResearchDGX agent

arXiv:2602.01698v3 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have recently achieved strong mathematical and code reasoning performance through Reinforcement Learning (RL) post-tra

Rethinking Constraint Awareness for Efficient State Embedding of Neural Routing Solver

ResearchDGX agent

arXiv:2605.10122v1 Announce Type: new Abstract: Heavy-Encoder-Light-Decoder (HELD) neural routing solvers have emerged as a promising paradigm due to their broad applicability across multiple vehicle

Rethinking Event-Based Object Dtection through Representation-Level Temporal Aggregation and Model-Level Hypergraph Reasoning

ResearchDGX agent

arXiv:2605.08825v1 Announce Type: new Abstract: Event cameras provide microsecond-level temporal resolution, low latency, and high dynamic range, offering potential for perception under fast motion an

Rethinking Expert Trajectory Utilization in LLM Post-training for Mathematical Reasoning

ResearchDGX agent

arXiv:2512.11470v2 Announce Type: replace-cross Abstract: Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) dominate the post-training landscape for mathematical reasoning, yet differ funda

Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models

ResearchDGX agent

arXiv:2602.11824v2 Announce Type: replace Abstract: Despite the advanced capabilities of Large Vision-Language Models (LVLMs), they frequently suffer from object hallucination. One reason is that visu

Revisiting Mixture Policies in Entropy-Regularized Actor-Critic

ResearchDGX agent

arXiv:2605.09157v1 Announce Type: cross Abstract: Mixture policies theoretically offer greater flexibility than unimodal policies in continuous action reinforcement learning, but the practical benefit

Revisiting the syntax of imperatives in Yemeni Arabic: An Agree across phases approach

ResearchDGX agent

arXiv:2605.08447v1 Announce Type: new Abstract: This article revisits the syntax of imperatives in Yemeni Arabic proposing an Agree acros phases (AAP) approach. I argue that the AAP approach successfu

RL Fine-Tuning Heals OOD Forgetting in SFT

ResearchDGX agent

arXiv:2509.12235v3 Announce Type: replace-cross Abstract: Supervised Fine-Tuning (SFT) followed by Reinforcement Learning (RL) is a standard post-training recipe for improving Large Language Models (L

Robust Building Damage Detection in Cross-Disaster Settings Using Domain Adaptation

ResearchDGX agent

arXiv:2603.14694v2 Announce Type: replace-cross Abstract: Rapid structural damage assessment from remote sensing imagery is essential for timely disaster response. Within human-machine systems (HMS) f

S2P-Net: A Spectral-Spatial Polar Network for Rotation-Invariant Object Recognition in Low-Data Regimes

ResearchDGX agent

arXiv:2605.09667v1 Announce Type: cross Abstract: We present S2P-Net (Spectral-Spatial Polar Network), a compact deep learning architecture that achieves mathematically guaranteed rotation invariance

SAFformer:Improving Spiking Transformer via Active Predictive Filtering

ResearchDGX agent

arXiv:2605.08270v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) offer notable advantages in biological plausibility and energy efficiency, making them promising candidates for buildin

SAMOFT: Robust Multi-Object Tracking via Region and Flow

ResearchDGX agent

arXiv:2605.09417v1 Announce Type: new Abstract: Multi-object tracking (MOT) is a fundamental task in computer vision that requires continuously tracking multiple targets while maintaining consistent i

Sanity Checks for Long-Form Hallucination Detection

ResearchDGX agent

arXiv:2605.08346v1 Announce Type: cross Abstract: Hallucination detection methods for large language models increasingly operate on chain-of-thought reasoning traces, yet it remains unclear whether th

Scalable Mamba-Based Message-Passing Neural Decoder for Error-Correcting Codes

ResearchDGX agent

arXiv:2605.10681v1 Announce Type: cross Abstract: Forward error correction is essential for reliable communication over noisy channels. Attention-based model-free neural decoders have shown strong per

SciLT: Long-tailed Image Classification under Scientific Image Domains

ResearchDGX agent

arXiv:2604.03687v2 Announce Type: replace Abstract: Long-tailed recognition has benefited from foundation models and fine-tuning paradigms, yet existing studies and benchmarks are mainly confined to n

Scratchpad Patching: Decoupling Compute from Patch Size in Byte-Level Language Models

ResearchDGX agent

arXiv:2605.09630v1 Announce Type: new Abstract: Tokenizer-free language models eliminate the tokenizer step of the language modeling pipeline by operating directly on bytes; patch-based variants furth

SDG-MoE: Signed Debate Graph Mixture-of-Experts

ResearchDGX agent

arXiv:2605.08322v1 Announce Type: cross Abstract: Sparse MoE models achieve a good balance between capacity and compute by routing each token to a small subset of experts. However, in most MoE archite

SDTalk: Structured Facial Priors and Dual-Branch Motion Fields for Generalizable Gaussian Talking Head Synthesis

ResearchDGX agent

arXiv:2605.09956v1 Announce Type: cross Abstract: High-quality, real-time talking head synthesis remains a fundamental challenge in computer vision. Existing reconstruction- and rendering-based method

SearchSkill: Teaching LLMs to Use Search Tools with Evolving Skill Banks

ResearchDGX agent

arXiv:2605.09038v1 Announce Type: new Abstract: Teaching language models to use search tools is not only a question of whether they search, but also of whether they issue good queries. This is especia

SeasonScapes: Learning Large-scale Re-lightable 3D Landscapes with Seasonal Variation from Sparse Webcams

ResearchDGX agent

arXiv:2605.09039v1 Announce Type: new Abstract: We introduce SeasonScapes framework and a the SeasonScapes dataset: Swiss Sparse-view Mountain Scenes with Seasonal Changes that covers over 50 km x 60

SegSTRONG-C: Segmenting Surgical Tools Robustly On Non-adversarial Generated Corruptions -- An EndoVis'24 Challenge

ResearchDGX agent

arXiv:2407.11906v3 Announce Type: replace Abstract: Surgical data science has seen rapid advancement with the excellent performance of end-to-end deep neural networks (DNNs). Despite their successes,

SEIS: Subspace-based Equivariance and Invariance Scores for Neural Representations

ResearchDGX agent

arXiv:2602.04054v2 Announce Type: replace-cross Abstract: Understanding how neural representations respond to geometric transformations is essential for evaluating whether learned features preserve me

Self-Attention as a Covariance Readout: A Unified View of In-Context Learning and Repetition

ResearchDGX agent

arXiv:2605.10466v1 Announce Type: new Abstract: Large language models (LLMs) exhibit two striking and ostensibly unrelated behaviours: in-context learning (ICL) and repetitive generation. In both, the

Self-Captioning Multimodal Interaction Tuning: Amplifying Exploitable Redundancies for Robust Vision Language Models

ResearchDGX agent

arXiv:2605.08145v1 Announce Type: cross Abstract: Current vision language models face hallucination and robustness issues against ambiguous or corrupted modalities. We hypothesize that these issues ca

Semantic Voting: Execution-Grounded Consensus for LLM Code Generation

ResearchDGX agent

arXiv:2605.08680v1 Announce Type: cross Abstract: LLM code-generation pipelines often sample multiple candidates and select one final answer without access to a complete oracle. Existing pipelines mix

Sequential Membership Inference Attacks

ResearchDGX agent

arXiv:2602.16596v2 Announce Type: replace Abstract: Modern AI models are not static. They go through multiple updates in their lifecycles. We propose to design Sequential Membership Inference (SeMI) a

Set-Based Groupwise Registration for Variable-Length, Variable-Contrast Cardiac MRI

ResearchDGX agent

arXiv:2605.10571v1 Announce Type: cross Abstract: Quantitative cardiac magnetic resonance imaging (MRI) enables non-invasive myocardial tissue characterization but relies on robust motion correction w

Sharp feature-learning transitions and Bayes-optimal neural scaling laws in extensive-width networks

ResearchDGX agent

arXiv:2605.10395v1 Announce Type: cross Abstract: We study the information-theoretic limits of learning a one-hidden-layer teacher network with hierarchical features from noisy queries, in the context

ShifaMind: A Multiplicative Concept Bottleneck for Interpretable ICD-10 Coding

ResearchDGX agent

arXiv:2605.08482v1 Announce Type: cross Abstract: Automated ICD-10 coding from clinical discharge summaries requires models that are both accurate on long-tailed multi-label classification tasks and i

Signature Approach for Contextual Bandits with Nonlinear and Path-dependent Rewards

ResearchDGX agent

arXiv:2605.10313v1 Announce Type: new Abstract: We study contextual bandits with nonlinear and path-dependent rewards through a novel signature-transform-based approach. Leveraging the universal nonli

SimReg: Achieving Higher Performance in the Pretraining via Embedding Similarity Regularization

ResearchDGX agent

arXiv:2605.08809v1 Announce Type: cross Abstract: Pretraining large language models (LLMs) with next-token prediction has led to remarkable advances, yet the context-dependent nature of token embeddin

SLayerGen: a Crystal Generative Model for all Space and Layer Groups

ResearchDGX agent

arXiv:2605.08262v1 Announce Type: cross Abstract: Crystal generative models have shown rapid progress for accelerating the discovery of bulk, periodic materials. However, many material systems such as

Sliced Inner Product Gromov-Wasserstein Distances

ResearchDGX agent

arXiv:2605.08546v1 Announce Type: cross Abstract: The Gromov-Wasserstein (GW) problem provides a framework for aligning heterogeneous datasets by matching their intrinsic geometry, but its statistical

Sliding Window Informative Canonical Correlation Analysis

ResearchDGX agent

arXiv:2507.17921v2 Announce Type: replace-cross Abstract: Canonical correlation analysis (CCA) is a technique for finding correlated sets of features between two datasets. In this paper, we propose a

SlimQwen: Exploring the Pruning and Distillation in Large MoE Model Pre-training

ResearchDGX agent

arXiv:2605.08738v1 Announce Type: cross Abstract: Structured pruning and knowledge distillation (KD) are typical techniques for compressing large language models, but it remains unclear how they shoul

SlimSpec: Low-Rank Draft LM-Head for Accelerated Speculative Decoding

ResearchDGX agent

arXiv:2605.10453v1 Announce Type: cross Abstract: Speculative decoding speeds up autoregressive generation in Large Language Models (LLMs) through a two-step procedure, where a lightweight draft model

Slum Detection and Density Mapping with AlphaEarth Foundations: A Representation Learning Evaluation Across 12 Global Cities

ResearchDGX agent

arXiv:2605.10029v1 Announce Type: new Abstract: Pixel-level slum mapping has long been constrained by limited cross-city generalisation, the absence of continuous density estimation, and weak global c

SMOG: Scalable Meta-Learning for Multi-Objective Bayesian Optimization

ResearchDGX agent

arXiv:2601.22131v2 Announce Type: replace Abstract: Multi-objective optimization aims to solve problems with competing objectives. Evaluating such problems is often slow or expensive, limiting the bud

Smoothing Out the Edges: Continuous-Time Estimation with Gaussian Process Motion Priors on Factor Graphs

ResearchDGX agent

arXiv:2605.09073v1 Announce Type: new Abstract: Continuous-time state estimation is gaining in popularity due to its abilities to provide smooth solutions, handle asynchronous sensors, and interpolate

Social Determinants of Health and Fentanyl Overdose Mortality Across US Counties: An XGBoost and SHAP Analysis Identifying Silent Risk Counties and Treatment Deserts

ResearchDGX agent

arXiv:2605.08230v1 Announce Type: new Abstract: Background: Fentanyl overdose deaths are still increasing across the U.S. We do not fully understand which county-level social and structural conditions

SpaceMind++: Toward Allocentric Cognitive Maps for Spatially Grounded Video MLLMs

ResearchDGX agent

arXiv:2605.09449v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) have made remarkable progress in visual understanding and language-based reasoning, yet they lack a pers

Sparse Layers are Critical to Scaling Looped Language Models

ResearchDGX agent

arXiv:2605.09165v1 Announce Type: cross Abstract: Looped language models repeat a set of transformer layers through depth, reducing memory costs and providing natural early-exit points at loop boundar

Spatial-Frequency Gated Swin Transformer for Remote Sensing Single-Image Super-Resolution

ResearchDGX agent

arXiv:2605.09687v1 Announce Type: new Abstract: Remote Sensing (RS) single-image super-resolution aims to reconstruct high-resolution imagery from low-resolution observations while preserving fine spa

Spatial Priming Outperforms Semantic Prompting: A Grid-Based Approach to Improving LLM Accuracy on Chart Data Extraction

ResearchDGX agent

arXiv:2605.08220v1 Announce Type: new Abstract: The automated extraction of data from scientific charts is a critical task for large-scale literature analysis. While multimodal Large Language Models (

Spectral Condition for muP under Width-Depth Scaling

ResearchDGX agent

arXiv:2603.00541v2 Announce Type: replace Abstract: Generative foundation models are increasingly scaled in both width and depth, posing significant challenges for stable feature learning and reliable

Spectrally-Guided Diffusion Noise Schedules

ResearchDGX agent

arXiv:2603.19222v2 Announce Type: replace Abstract: Denoising diffusion models are widely used for high-quality image and video generation. Their performance depends on noise schedules, which define t

Speech-based Psychological Crisis Assessment using LLMs

ResearchDGX agent

arXiv:2605.10027v1 Announce Type: cross Abstract: Psychological support hotlines provide critical support for individuals experiencing mental health emergencies, yet current assessments largely rely o

Spherical Flows for Sampling Categorical Data

ResearchDGX agent

arXiv:2605.05629v2 Announce Type: replace-cross Abstract: We study the problem of learning generative models for discrete sequences in a continuous embedding space. Whereas prior approaches typically

Split CNN Inference on Networked Microcontrollers

ResearchDGX agent

arXiv:2605.09357v1 Announce Type: cross Abstract: Running deep neural networks on microcontroller units (MCUs) is severely constrained by limited memory resources. While TinyML techniques reduce model

SSA: Improving Performance With a Better Scoring Function

ResearchDGX agent

arXiv:2508.14685v4 Announce Type: replace Abstract: While transformer models exhibit strong in-context learning (ICL) abilities, they often fail to generalize under simple distribution shifts. We anal

Stability of the Monge Map in Semi-Dual Optimal Transport

ResearchDGX agent

arXiv:2605.05569v2 Announce Type: replace-cross Abstract: This paper shows that the semi-dual formulation of the optimal transport problem has a degenerate saddle-point structure, and that its numeric

← Previous
1…226227228229230…320
Next →