AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,898 results
19 May 2026

Learning Unbiased Permutations via Flow Matching

ResearchDGX agent

arXiv:2605.16755v1 Announce Type: cross Abstract: Learning permutations is fundamental to sorting, ranking, and matching, but existing differentiable methods based on entropy-regularized Sinkhorn prod

Learning Variable-Length Tokenization for Generative Recommendation

ResearchDGX agent

arXiv:2605.17779v1 Announce Type: new Abstract: Generative recommendation reformulates recommendation as next-token prediction over discrete semantic identifiers (IDs). A fundamental yet unexplored de

Lever: Speculative LLM Inference on Smartphones

ResearchDGX agent

arXiv:2605.16786v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly needed for interactive mobile applications, but high-quality models exceed the limited DRAM available on s

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Leveraging Error Diversity in Group Rollouts for Reinforcement Learning

ResearchDGX agent

arXiv:2605.17333v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) typically samples multiple responses per prompt and assigns binary rewards based on individual cor

Leveraging Graph Structure in Seq2Seq Models for Knowledge Graph Link Prediction

ResearchDGX agent

arXiv:2605.18211v1 Announce Type: cross Abstract: We introduce Graph-Augmented Sequence-to-Sequence (GA-S2S), a novel framework that integrates a T5-small encoder-decoder with a Relational Graph Atten

Leveraging Multimodal Self-Consistency Reasoning in Coding Motivational Interviewing for Alcohol Use Reduction

ResearchDGX agent

arXiv:2605.12987v2 Announce Type: replace Abstract: BACKGROUND: Coding Motivational Interviewing (MI) sessions is essential for understanding client behaviors and predicting outcomes, but it requires

Lightweight Physics-Aware Zero-Shot Ultrasound Plane-Wave Denoising

ResearchDGX agent

arXiv:2506.21499v2 Announce Type: replace-cross Abstract: Ultrasound Coherent Plane-Wave Compounding (CPWC) enhances image contrast by combining echoes from multiple steered transmissions. While incre

LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs

ResearchDGX agent

arXiv:2603.16947v2 Announce Type: replace-cross Abstract: Although vision-language navigation (VLN) has progressed rapidly, zero-shot VLN in continuous environments (VLN-CE) remains highly challenging

Lipschitz-Guided Design of Interpolation Schedules in Generative Models

ResearchDGX agent

arXiv:2509.01629v3 Announce Type: replace-cross Abstract: We study the design of interpolation schedules in flow and diffusion-based generative models from both statistical and numerical perspectives.

LISA: Language-guided Interference-aware Spatial-Frequency Attention for Driver Gaze Estimation

ResearchDGX agent

arXiv:2605.17287v1 Announce Type: new Abstract: Driver gaze estimation serves as a fundamental metric for evaluating driver attentiveness in modern monitoring systems. Beyond being vulnerable to sudde

LiteFrame: Efficient Vision Encoders Unlock Frame Scaling in Video LLMs

ResearchDGX agent

arXiv:2605.17260v1 Announce Type: new Abstract: The fundamental challenge in scaling Video Large Language Models (Video LLMs) to long-form video lies in managing the explosion of visual-token context

LLM-Safety Evaluations Lack Robustness

SafetyDGX agent

arXiv:2503.02574v2 Announce Type: replace-cross Abstract: In this paper, we argue that current safety alignment research efforts for large language models are hindered by many intertwined sources of n

Long Context Modeling with Ranked Memory-Augmented Retrieval

ResearchDGX agent

arXiv:2503.14800v3 Announce Type: replace-cross Abstract: Effective long-term memory management is crucial for language models handling extended contexts. We introduce the Enhanced Ranked Memory Augme

Long-horizon prediction of three-dimensional wall-bounded turbulence with CTA-Swin-UNet and resolvent analysis

ResearchDGX agent

arXiv:2605.17888v1 Announce Type: cross Abstract: Long-horizon prediction of three-dimensional (3D) wall-bounded turbulence with machine-learning methods remains a challenging task, due to the rapid a

Longwang: Zero-Shot Global Spatiotemporal Precipitation Downscaling with a Latent Generative Prior

ResearchDGX agent

arXiv:2605.17603v1 Announce Type: cross Abstract: High-resolution precipitation information is essential for climate impact assessment, yet global climate models remain too coarse to resolve key small

Lost or Hidden? A Concept-Level Forgetting in Supervised Continual Learning

ResearchDGX agent

arXiv:2605.16374v1 Announce Type: cross Abstract: Continual learning studies how models can adapt to new tasks while retaining previously acquired knowledge. Although a broad spectrum of methods has b

Lotus-2: Advancing Geometric Dense Prediction with Powerful Image Generative Model

ResearchDGX agent

arXiv:2512.01030v3 Announce Type: replace Abstract: Recovering pixel-wise geometric properties from a single image is fundamentally ill-posed due to appearance ambiguity and non-injective mappings bet

LURE: Latent Space Unblocking for Multi-Concept Reawakening in Diffusion Models

ResearchDGX agent

arXiv:2601.14330v2 Announce Type: replace Abstract: Concept erasure aims to suppress sensitive content in diffusion models, but recent studies show that erased concepts can still be reawakened, reveal

Machine Learning-Based Pre-Test Risk Stratification for PCR-Confirmed Chlamydia Using Patient-Reported Data and Urine Biomarkers

ResearchDGX agent

arXiv:2605.16365v1 Announce Type: new Abstract: Early identification of individuals at elevated risk of Chlamydia trachomatis infection may enable optimal use of molecular testing in resource-aware sc

Markerless Motion Capture for Biomechanical Whole-Body Kinematic Estimation in Infants

ResearchDGX agent

arXiv:2605.17120v1 Announce Type: new Abstract: arly identification of motor impairment in infancy relies on expert visual assessment of spontaneous movement, motivating the development of automated,

MARQUIS: A Three-Stage Pipeline for Video Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2605.17640v1 Announce Type: cross Abstract: Retrieval-augmented generation from videos requires systems to retrieve relevant audiovisual evidence from large corpora and synthesize it into cohere

Matrix-Decoupled Concentration for Autoregressive Sequences: Dimension-Free Guarantees for Sparse Long-Context Rewards

ResearchDGX agent

arXiv:2605.06017v2 Announce Type: replace Abstract: Sequence-level evaluations in autoregressive Large Language Models (LLMs) rely on highly dependent token generation. Establishing tight concentratio

Mechanism Learning: Prototype-Anchored Mechanism Inference for Scientific Forecasting

ResearchDGX agent

arXiv:2605.17091v1 Announce Type: new Abstract: Scientific forecasting typically relies on direct state prediction, an approach that grows brittle under data scarcity, extended horizons, non-stationar

Mechanistically Interpretable Neural Encoding Reveals Fine-Grained Functional Selectivity in Human Visual Cortex

ResearchDGX agent

arXiv:2605.16468v1 Announce Type: cross Abstract: A central goal in understanding human vision is to uncover the visual features that drive neuronal activity. A growing body of work has used artificia

MedMIX: Modality-Internal Expert Fusion for Multimodal Medical Diagnosis

ResearchDGX agent

arXiv:2605.16639v1 Announce Type: new Abstract: Multimodal clinical prediction faces three challenges: multiple foundation models (FMs) with complementary strengths per modality, pervasive missing mod

Memory-Augmented Query Intent Understanding for Efficient Chat-based Image Retrieval

ResearchDGX agent

arXiv:2605.17365v1 Announce Type: new Abstract: Different from traditional text-to-image retrieval tasks, chat-based image retrieval allows the human-interactive system to iteratively clarify and refi

Memory-Efficient Differentially Private Training with Gradient Random Projection

ResearchDGX agent

arXiv:2506.15588v2 Announce Type: replace Abstract: Differential privacy (DP) protects sensitive data during neural network training, but standard methods like DP-Adam suffer from high memory overhead

Memory-Guided Tree Search with Cross-Branch Knowledge Transfer for LLM Solver Synthesis

ResearchDGX agent

arXiv:2605.17539v1 Announce Type: new Abstract: Combinatorial optimization (CO) underlies decision-making from logistics to chip design, where infeasible solutions are operationally unusable and small

MetaLab: Few-Shot Game Changer for Image Recognition

ResearchDGX agent

arXiv:2507.22057v2 Announce Type: replace Abstract: Difficult few-shot image recognition has significant application prospects, yet remaining the substantial technical gaps with the conventional large

Metric-Guided Feature Fusion of Visual Foundation Models for Segmentation Tasks

ResearchDGX agent

arXiv:2605.16864v1 Announce Type: cross Abstract: Although large-scale visual foundation models (VFMs) achieve remarkable performance in semantic understanding, they still underperform in instance-awa

MHMamba: Multi-Head Mamba for 3D Brain Tumor Segmentation

ResearchDGX agent

arXiv:2605.16464v1 Announce Type: cross Abstract: Brain tumors exhibit high heterogeneity in morphology and multimodal contrast, making manual slice-by-slice de lineation time-consuming and experience

Mirror Mean-Field Langevin Dynamics

ResearchDGX agent

arXiv:2505.02621v2 Announce Type: replace Abstract: The mean-field Langevin dynamics (MFLD) minimizes an entropy-regularized nonlinear convex functional on the Wasserstein space over R^d, and has gain

Mitigating 3D Prostate Biparametric MRI Data Scarcity through Domain Adaptation using Locally-Trained Latent Diffusion Models for Prostate Cancer Detection

ResearchDGX agent

arXiv:2507.06384v2 Announce Type: replace-cross Abstract: Objective: Latent diffusion models (LDMs) could mitigate data scarcity challenges affecting machine learning development for medical image int

Mixup Barcodes: Quantifying Geometric-Topological Interactions between Point Clouds

ResearchDGX agent

arXiv:2402.15058v3 Announce Type: replace-cross Abstract: We combine standard persistent homology with image persistent homology to define a novel way of characterizing shapes and interactions between

ML-based Fast Simulation of FARICH Responses

ResearchDGX agent

arXiv:2605.17635v1 Announce Type: cross Abstract: A fast simulation of the detector response is a vital task in high-energy physics (HEP). Traditional Monte-Carlo methods form the backbone of modern p

MoCA3D: Monocular 3D Bounding Box Prediction in the Image Plane

ResearchDGX agent

arXiv:2603.19538v2 Announce Type: replace Abstract: Monocular 3D object understanding has largely been cast as a 2D RoI-to-3D box lifting problem. However, emerging downstream applications require ima

Modality vs. Morphology: A Framework for Time Series Classification for Biological Signals

ResearchDGX agent

arXiv:2605.18483v1 Announce Type: cross Abstract: Time series classification (TSC) of biological signals has progressed from handcrafted, modality-specific approaches to deep architectures capable of

Monocular Depth Perception Enhancement Based on Joint Shading/Contrast Model and Motion Parallax (JSM)

ResearchDGX agent

arXiv:2605.17252v1 Announce Type: new Abstract: Stereoscopic 3D displays adopt a binocular depth cue to provide depth perception. However, users should be equipped with expensive special devices to ap

MORN: Metacognitive Object-Goal Regulation for Resource-Rational Long-Horizon Navigation

ResearchDGX agent

arXiv:2605.16932v1 Announce Type: new Abstract: Robots deployed in unstructured human environments must frequently execute long-horizon missions, such as find the mug, then the chair, then the printer

Motion Planning of Cooperative Nonholonomic Mobile Manipulators

ResearchDGX agent

arXiv:2502.05462v2 Announce Type: replace Abstract: We propose a real-time implementable motion planning framework for cooperative object transportation by nonholonomic mobile manipulator robots (MMRs

Multi-hop Relational Contrastive Learning: Extending Spatial Contrastive Pre-training Beyond Pairwise Relations

ResearchDGX agent

arXiv:2605.16456v1 Announce Type: new Abstract: Understanding how objects relate to each other in space is fundamental to scene understanding, yet most contrastive pre-training approaches only model p

Multi-Mode Quantum Annealing for Generative Representation Learning with Boltzmann Priors

ResearchDGX agent

arXiv:2604.00919v2 Announce Type: replace-cross Abstract: Energy-based models provide a natural bridge between statistical physics and machine learning by representing data through structured energy l

Multi-task learning on partially labeled datasets via invariant/equivariant semi-supervised learning

ResearchDGX agent

arXiv:2605.17624v1 Announce Type: cross Abstract: We investigate the potential of invariant and equivariant semi-supervised learning for addressing the challenges of training multi-task models on part

MV-Gate: Insider Threat Detection via Multi-View Behavioral Statistics and Semantic Modeling

ResearchDGX agent

arXiv:2605.17761v1 Announce Type: cross Abstract: Insider threats often reveal early anomalies through disruptions in behavioral statistics-such as altered recurrence patterns or short-versus long-ter

Nash: Neural Adaptive Shrinkage for Structured High-Dimensional Regression

ResearchDGX agent

arXiv:2505.11143v2 Announce Type: replace-cross Abstract: Sparse linear regression is a fundamental tool in data analysis. However, traditional approaches often fall short when covariates exhibit stru

Needles in the Landscape: Semi-Supervised Pseudolabeling for Archaeological Site Discovery under Label Scarcity

ResearchDGX agent

arXiv:2510.16814v2 Announce Type: replace-cross Abstract: Archaeological predictive modelling estimates where undiscovered sites are likely to occur by combining known locations with environmental, cu

NeuroLiDAR: Adaptive Frame Rate Depth Sensing via Neuromorphic Event-LiDAR Fusion

ResearchDGX agent

arXiv:2605.16805v1 Announce Type: new Abstract: LiDARs are widely used for 3D depth reconstruction, but their performance is often limited by inherent hardware constraints that impose trade-offs betwe

NeuroRVQ: Multi-Scale Biosignal Tokenization for Generative Foundation Models

ResearchDGX agent

arXiv:2510.13068v4 Announce Type: replace-cross Abstract: Biosignals such as electroencephalography (EEG), electrocardiography (ECG), and electromyography (EMG) encode physiological activity across mu

New Insight of Variance reduce in Zero-Order Hard-Thresholding: Mitigating Gradient Error and Expansivity Contradictions

ResearchDGX agent

arXiv:2605.18035v1 Announce Type: new Abstract: Hard-thresholding is an important type of algorithm in machine learning that is used to solve ell_0 constrained optimization problems. However, the true

NGM: A Plug-and-Play Training-Free Memory Module for LLMs

ResearchDGX agent

arXiv:2605.16893v1 Announce Type: new Abstract: Recent studies introduce conditional memory modules that decouple knowledge storage from neural computation, enabling more direct knowledge access. Comp

No Plan, Yet Human: A Reactive Robotics Model Predicts Human Planning Failures on a Clinical Task

ResearchDGX agent

arXiv:2605.16514v1 Announce Type: cross Abstract: Understanding why some sequential planning problems are harder than others requires models that go beyond average performance. They should capture the

NOETHER: A Constructive Framework for Metamorphic Pattern Discovery from Operator Algebras

ResearchDGX agent

arXiv:2605.17390v1 Announce Type: cross Abstract: Context. Metamorphic Testing is recognised in IEEE/ISO software-testing standards and increasingly recommended for AI systems, but its progress is bot

Non-Colliding Biometric Identities for Digital Entities: Geometry, Capacity, and Million-Scale Virtual Identity Provisioning

ResearchDGX agent

arXiv:2605.18238v1 Announce Type: new Abstract: Digital entities such as AI agents and humanoid robots increasingly operate alongside real humans, yet their identity infrastructure is based on credent

Nonlinear Bipolar Compensation: Handling Outliers in Post-Training Quantization

ResearchDGX agent

arXiv:2605.16423v1 Announce Type: new Abstract: Network quantization has emerged as one of the most practical model compression techniques, which significantly reduces a model's memory and compute con

Olivia: Harmonizing Time Series Foundation Models with Power Spectral Density

ResearchDGX agent

arXiv:2605.17340v1 Announce Type: new Abstract: Time series foundation models rely on large-scale pretraining over diverse datasets across domains, yet their heterogeneity in temporal patterns could h

Omni-Customizer: End-to-End MultiModal Customization for Joint Audio-Video Generation

ResearchDGX agent

arXiv:2605.17488v1 Announce Type: new Abstract: The landscape of joint audio and video generation has been fundamentally transformed by the advent of powerful foundation models. Despite these strides,

OmniSelect: Dynamic Modality-Aware Token Compression for Efficient Omni-modal Large Language Models

ResearchDGX agent

arXiv:2605.18041v1 Announce Type: new Abstract: Omnimodal large language models (OmniLLMs) have recently gained increasing attention for unified audio-video understanding. However, processing long mul

O(n) alternative to Quantum Fourier Transform with efficient neural net classical post-processing

ResearchDGX agent

arXiv:2605.16998v1 Announce Type: cross Abstract: The Quantum Fourier Transform (QFT) is required by hidden subgroup problem (HSP) algorithms, including Shor's algorithm for factoring. The circuit dep

On Applicability of Synthetic Datasets for Facial Expression Recognition

ResearchDGX agent

arXiv:2605.17483v1 Announce Type: new Abstract: Facial Expression Recognition faces two core challenges. The first is class imbalance in public datasets, which skews the learning process and weakens g

On Gaussian approximation for entropy-regularized Q-learning with function approximation

ResearchDGX agent

arXiv:2605.17678v1 Announce Type: cross Abstract: In this paper, we derive rates of convergence in the high-dimensional central limit theorem for Polyak--Ruppert averaged iterates generated by entropy

← Previous
1…237238239240241…432
Next →