AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
19,193 results
13 May 2026

RNA-FM: Flow-Matching Generative Model for Genome-wide RNA-Seq Prediction

ResearchDGX agent

arXiv:2605.11622v1 Announce Type: new Abstract: Histopathology whole-slide images (WSIs) are routinely acquired in clinical practice and contain rich tissue morphology but lack direct molecular archit

Robust Biomedical Publication Type and Study Design Classification with Knowledge-Guided Perturbations

ResearchDGX agent

arXiv:2605.11502v1 Announce Type: new Abstract: Accurately and consistently indexing biomedical literature by publication type and study design is essential for supporting evidence synthesis and knowl

Rollbot: a Spherical Robot Driven by a Single Actuator

ResearchDGX agent

arXiv:2404.05120v2 Announce Type: replace Abstract: Spherical robots typically require at least two actuators to achieve controlled 2D planar motion. Here we present Rollbot, the first spherical robot


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Rotary Masked Autoencoders are Versatile Learners

ResearchDGX agent

arXiv:2505.20535v3 Announce Type: replace Abstract: Applying Transformers to irregular time-series typically requires specializations to their baseline architecture, which can result in additional com

Rotation-Preserving Supervised Fine-Tuning

ResearchDGX agent

arXiv:2605.10973v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) improves in-domain performance but can degrade out-of-domain (OOD) generalization. Prior work suggests that this degradatio

RT-Transformer: The Transformer Block as a Spherical State Estimator

ResearchDGX agent

arXiv:2605.11007v1 Announce Type: new Abstract: We show that the core components of the Transformer block -- attention, residual connections, and normalization -- arise naturally from a single geometr

@SakanaAILabs @NVIDIAAI Sparser, Faster, Lighter Transformer Language Models https://arxiv.org/abs/2603.23198

ResearchDGX agent

This research paper from Sakana AI and NVIDIA explores techniques for creating more efficient transformer language models by reducing sparsity, computational requirements, and model size while maintai

Sampling-Based Follow-the-Leader Motion Planning for Manipulator-Mounted Continuum Robots

ResearchDGX agent

arXiv:2605.11618v1 Announce Type: new Abstract: Follow-the-leader (FTL) motion exploits the unique morphology of continuum robots (CRs) to navigate confined spaces by having the body retrace the path

Scalable Token-Level Hallucination Detection in Large Language Models

ResearchDGX agent

arXiv:2605.12384v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated remarkable capabilities, but they still frequently produce hallucinations. These hallucinations are diffi

ScaleMoGen: Autoregressive Next-Scale Prediction for Human Motion Generation

ResearchDGX agent

arXiv:2605.11704v1 Announce Type: new Abstract: We present ScaleMoGen, a scale-wise autoregressive framework for text-driven human motion generation. Unlike conventional autoregressive approaches that

ScribbleDose: Scribble-Guided Dose Prediction in Radiotherapy

ResearchDGX agent

arXiv:2605.11555v1 Announce Type: new Abstract: Anatomical structure masks are widely adopted in radiotherapy dose prediction, as they provide explicit geometric constraints that facilitate structure-

See the past: Time-Reversed Scene Reconstruction from Thermal Traces Using Visual Language Models

ResearchDGX agent

arXiv:2510.05408v2 Announce Type: replace Abstract: Recovering the past from present observations is an intriguing challenge with potential applications in forensics and scene analysis. Thermal imagin

Self-Distilled Trajectory-Aware Boltzmann Modeling: Bridging the Training-Inference Discrepancy in Diffusion Language Models

ResearchDGX agent

arXiv:2605.11854v1 Announce Type: new Abstract: Diffusion Language Models (DLMs) have recently emerged as a promising alternative to autoregressive language models, offering stronger global awareness

Semi-Supervised Bayesian GANs with Log-Signatures for Uncertainty-Aware Credit Card Fraud Detection

ResearchDGX agent

arXiv:2509.00931v3 Announce Type: replace-cross Abstract: We present a novel deep generative semi-supervised framework for credit card fraud detection, formulated as time series classification task. A

Shapley Value Approximation Based on k-Additive Games

ResearchDGX agent

arXiv:2502.04763v2 Announce Type: replace-cross Abstract: The Shapley value is the prevalent solution for fair division problems in which a payout is to be divided among multiple agents. By adopting a

ShardTensor: Domain Parallelism for Scientific Machine Learning

ResearchDGX agent

arXiv:2605.11111v1 Announce Type: cross Abstract: Scientific Machine Learning (SciML) faces unique challenges for extreme-resolution data, with mitigations that often fail to scale or degrade the accu

Single-Shot HDR Recovery via a Video Diffusion Prior

ResearchDGX agent

arXiv:2605.11628v1 Announce Type: new Abstract: Recent generative methods for single-shot high dynamic range (HDR) image reconstruction show promising results, but often struggle with preserving fidel

Smooth-Rigid-Body Contact as a ReLCP: A Recursively Generated Linear Complementarity Problem

ResearchDGX agent

arXiv:2506.14097v2 Announce Type: replace Abstract: This paper reformulates complementarity-based time-stepping for frictionless nonsmooth contact between smooth rigid bodies as a recursively generate

SOAR: Scale Optimization for Accurate Reconstruction in NVFP4 Quantization

ResearchDGX agent

arXiv:2605.12245v1 Announce Type: new Abstract: NVFP4 has recently emerged as an efficient 4-bit microscaling format for large language models (LLMs), offering superior numerical fidelity with native

Sobolev Regularized MMD Gradient Flow

ResearchDGX agent

arXiv:2605.11884v1 Announce Type: new Abstract: We propose Sobolev-regularized Maximum Mean Discrepancy (SrMMD) gradient flow, a regularized variant of maximum mean discrepancy (MMD) gradient flow bas

SoK: Unlearnability and Unlearning for Model Dememorization

ResearchDGX agent

arXiv:2605.11592v1 Announce Type: new Abstract: Advanced model dememorization methods, including availability poisoning (unlearnability) and machine unlearning, are emerging as key safeguards against

Sparse Attention Remapping with Clustering for Efficient LLM Decoding on PIM

ResearchDGX agent

arXiv:2505.05772v2 Announce Type: replace Abstract: Transformer-based models are the foundation of modern machine learning, but their execution, particularly during autoregressive decoding in large la

SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images

ResearchDGX agent

arXiv:2605.11462v1 Announce Type: new Abstract: Recent advancements in Large Vision-Language Models (VLMs) have demonstrated exceptional semantic understanding, yet these models consistently struggle

Spectral Vision Transformer for Efficient Tokenization with Limited Data

ResearchDGX agent

arXiv:2605.12026v1 Announce Type: new Abstract: We propose a novel spectral vision transformer architecture for efficient tokenization in limited data, with an emphasis on medical imaging. We outline

Split the Differences, Pool the Rest: Provably Efficient Multi-Objective Imitation

ResearchDGX agent

arXiv:2605.12000v1 Announce Type: new Abstract: This work investigates multi-objective imitation learning: the problem of recovering policies that lie on the Pareto front given demonstrations from mul

SRG: Score-based Relaxation-guided Generation for Mixed Integer Linear Programming

ResearchDGX agent

arXiv:2603.24033v2 Announce Type: replace Abstract: We propose Score-based Relaxation-guided Generation (SRG), a generative framework based on an approximate formulation of relaxation-guided stochasti

Stationary MMD Points

ResearchDGX agent

arXiv:2505.20754v3 Announce Type: replace-cross Abstract: Approximation of a target probability distribution using a finite set of points is a problem of fundamental importance in numerical integratio

Steerable Neural ODEs on Homogeneous Spaces

ResearchDGX agent

arXiv:2605.11133v1 Announce Type: new Abstract: We introduce steerable neural ordinary differential equations on homogeneous spaces M=G/H. These models constitute a novel geometric extension of manifo

Steering Without Breaking: Mechanistically Informed Interventions for Discrete Diffusion Language Models

ResearchDGX agent

arXiv:2605.10971v1 Announce Type: cross Abstract: Discrete diffusion language models (DLMs) generate text by iteratively denoising all positions in parallel, offering an alternative to autoregressive

Stop Marginalizing My Dreams: Model Inversion via Laplace Kernel for Continual Learning

ResearchDGX agent

arXiv:2605.11804v1 Announce Type: cross Abstract: Data-free continual learning (DFCIL) relies on model inversion to synthesize pseudo-samples and mitigate catastrophic forgetting. However, existing in

Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding

ResearchDGX agent

arXiv:2602.06412v3 Announce Type: replace Abstract: Masked Diffusion Language Models generate sequences via iterative sampling that progressively unmasks tokens. However, they still recompute the atte

Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space

ResearchDGX agent

arXiv:2605.12412v1 Announce Type: new Abstract: Large Language Models (LLMs) update their behavior in context, which can be viewed as a form of Bayesian inference. However, the structure of the latent

Streaming of rendered content with adaptive frame rate and resolution

ResearchDGX agent

arXiv:2605.10995v1 Announce Type: cross Abstract: Streaming rendered content is an attractive way to bring high-quality graphics to billions of mobile devices that do not have sufficient rendering pow

SurvBench: A Standardised Preprocessing Pipeline for Multi-Modal Electronic Health Record Survival Analysis

ResearchDGX agent

arXiv:2511.11935v2 Announce Type: replace Abstract: Deep-learning survival models for electronic health record (EHR) data are hard to compare across papers because the upstream preprocessing step, whi

Tackling Fake Forgetting through Uncertainty Quantification

ResearchDGX agent

arXiv:2501.19403v3 Announce Type: replace Abstract: Machine unlearning seeks to remove the influence of specified data from a trained model. While the unlearning accuracy provides a widely used metric

Taking the Road Less Scheduled with Adaptive Polyak Steps

ResearchDGX agent

arXiv:2511.07767v2 Announce Type: replace Abstract: Schedule-Free SGD, proposed in [Defazio et al., 2024], achieves optimal convergence rates without requiring the training horizon in advance, by repl

Taming Score-Based Denoisers in ADMM: A Convergent Plug-and-Play Framework

ResearchDGX agent

arXiv:2603.10281v3 Announce Type: replace-cross Abstract: While score-based generative models have emerged as powerful priors for solving inverse problems, directly integrating them into optimization

Targeted Tests for LLM Reasoning: An Audit-Constrained Protocol

ResearchDGX agent

arXiv:2605.11599v1 Announce Type: new Abstract: Fixed reasoning benchmarks evaluate canonical prompts, but semantically valid changes in presentation can still change model behavior. Studies of prompt

Task-Adaptive Embedding Refinement via Test-time LLM Guidance

ResearchDGX agent

arXiv:2605.12487v1 Announce Type: new Abstract: We explore the effectiveness of an LLM-guided query refinement paradigm for extending the usability of embedding models to challenging zero-shot search

TCP-SSM: Efficient Vision State Space Models with Token-Conditioned Poles

ResearchDGX agent

arXiv:2605.11563v1 Announce Type: new Abstract: State Space Models (SSMs) have emerged as a compelling alternative to attention models for long-range vision tasks, offering input-dependent recurrence

Testing General Relativity Through Gravitational Wave Classification: A Convolutional Neural Network Framework

ResearchDGX agent

arXiv:2605.02453v1 Announce Type: cross Abstract: We present a machine learning framework for testing general relativity (GR) with gravitational wave signals from binary black hole mergers. Using the

The Algorithmic Caricature: Auditing LLM-Generated Political Discourse Across Crisis Events

ResearchDGX agent

arXiv:2605.12452v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate fluent political text at scale, raising concerns about synthetic discourse during crises and social conflict.

The Bicameral Model: Bidirectional Hidden-State Coupling Between Parallel Language Models

ResearchDGX agent

arXiv:2605.11167v1 Announce Type: new Abstract: Existing multi-model and tool-augmented systems communicate by generating text, serializing every exchange through the output vocabulary. Can two pretra

The Challenge and Reward of Fair Play in Narrative: A Computational Approach

ResearchDGX agent

arXiv:2507.13841v2 Announce Type: replace Abstract: Good storytelling involves surprise -- unpredictability in how the story unfolds -- and sense-making, the requirement that the story forms a coheren

The Download: making drugs in orbit and NASA’s nuclear-powered spacecraft

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. A plan to make drugs in orbit is going commercial A startup ca

The first global agricultural field boundary map at 10m resolution

ResearchDGX agent

arXiv:2605.11055v1 Announce Type: new Abstract: The agricultural field is the natural unit at which crops are planted, managed, regulated, and reported, yet most global remote-sensing products for agr

The Midas Touch for Metric Depth

ResearchDGX agent

arXiv:2605.11578v1 Announce Type: new Abstract: Recent advances have markedly improved the cross-scene generalization of relative depth estimation, yet its practical applicability remains limited by t

The Offline-Frontier Shift: Diagnosing Distributional Limits in Generative Multi-Objective Optimization

ResearchDGX agent

arXiv:2602.11126v2 Announce Type: replace Abstract: Offline multi-objective optimization (MOO) aims to recover Pareto-optimal designs given a finite, static dataset. Recent generative approaches, incl

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling

ResearchDGX agent

arXiv:2605.05922v2 Announce Type: replace Abstract: Recent advances in generative video models are increasingly driven by post-training and test-time scaling, both of which critically depend on the qu

Today we release Token Superposition Training (TST), a modification to the standard LLM pretraining loop that produces a 2-3× wall-clock spe…

ResearchDGX agent

Today we release Token Superposition Training (TST), a modification to the standard LLM pretraining loop that produces a 2-3× wall-clock speedup at matched FLOPs without changing the model architectur

TRACE: Temporal Routing with Autoregressive Cross-channel Experts for EEG Representation Learning

ResearchDGX agent

arXiv:2605.11380v1 Announce Type: new Abstract: Learning transferable representations for electroencephalography (EEG) remains challenging because EEG signals are inherently multi-channel and non-stat

Training-Inference Consistent Segmented Execution for Long-Context LLMs

ResearchDGX agent

arXiv:2605.11744v1 Announce Type: new Abstract: Transformer-based large language models face severe scalability challenges in long-context generation due to the computational and memory costs of full-

TST works in two phases. In phase 1, which covers the first 20-40% of training, the model reads contiguous bags of k tokens, with input embe…

ResearchDGX agent

TST works in two phases. In phase 1, which covers the first 20-40% of training, the model reads contiguous bags of k tokens, with input embeddings averaged within each bag, and predicts the next bag o

Uniform Scaling Limits in AdamW-Trained Transformers

ResearchDGX agent

arXiv:2605.11059v1 Announce Type: cross Abstract: We study the large-depth limit of transformers trained with AdamW, by modelling the hidden-state dynamics as an interacting particle system (IPS) coup

Unlearning with Asymmetric Sources: Improved Unlearning-Utility Trade-off with Public Data

ResearchDGX agent

arXiv:2605.11170v1 Announce Type: new Abstract: Noise-based certified machine unlearning currently faces a hard ceiling: the noise magnitude required to certify unlearning typically destroys model uti

Unlocking Compositional Generalization in Continual Few-Shot Learning

ResearchDGX agent

arXiv:2605.11710v1 Announce Type: cross Abstract: Object-centric representations promise a key property for few-shot learning: Rather than treating a scene as a single unit, a model can decompose it i

Unpacking the Eye of the Beholder: Social Location, Identity, and the Moving Target of Political Perspectives

ResearchDGX agent

arXiv:2605.11166v1 Announce Type: new Abstract: Political and social identities structure how people evaluate political information, a finding decades deep in political science and routinely discarded

USEMA: a Scalable Efficient Mamba Like Attention for Medical Image Segmentation

ResearchDGX agent

arXiv:2605.11131v1 Announce Type: new Abstract: Accurate medical image segmentation is an integral part of the medical image analysis pipeline that requires the ability to merge local and global infor

Variational Linear Attention: Stable Associative Memory for Long-Context Transformers

ResearchDGX agent

arXiv:2605.11196v1 Announce Type: new Abstract: Linear attention reduces the quadratic cost of softmax attention to O(T), but its memory state grows as O(T) in Frobenius norm, causing progressive inte

Vector Scaffolding: Inter-Scale Orchestration for Differentiable Image Vectorization

ResearchDGX agent

arXiv:2605.11913v1 Announce Type: new Abstract: Differentiable vector graphics have enabled powerful gradient-based optimization of vector primitives directly from raster images. However, existing fra

← Previous
1…217218219220221…320
Next →