AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,895 results
4 May 2026

End-to-End Autoregressive Image Generation with 1D Semantic Tokenizer

ResearchDGX agent

arXiv:2605.00503v1 Announce Type: new Abstract: Autoregressive image modeling relies on visual tokenizers to compress images into compact latent representations. We design an end-to-end training pipel

Escaping Mode Collapse in LLM Generation via Geometric Regulation

ResearchDGX agent

arXiv:2605.00435v1 Announce Type: new Abstract: Mode collapse is a persistent challenge in generative modeling and appears in autoregressive text generation as behaviors ranging from explicit looping

Exploring the Limits of End-to-End Feature-Affinity Propagation for Single-Point Supervised Infrared Small Target Detection

ResearchDGX agent

arXiv:2605.00722v1 Announce Type: new Abstract: Single-point supervised infrared small target detection (IRSTD) drastically reduces dense annotation costs. Current state-of-the-art (SOTA) methods achi

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Federated Learning with Hypergradient-based Online Update of Aggregation Weights

ResearchDGX agent

arXiv:2605.00458v1 Announce Type: new Abstract: Federated learning using mobile and Internet of Things devices requires not only the ability to handle heterogeneity of clients' data distributions but

Federated Weather Modeling on Sensor Data

ResearchDGX agent

arXiv:2605.00322v1 Announce Type: new Abstract: Federated weather modeling on sensor data is a distributed system underpinned by federated learning, enabling multiple sensor data sources, including gr

Feedback Lunch: Learned Feedback Codes for Secure Communications

ResearchDGX agent

arXiv:2510.16620v3 Announce Type: replace-cross Abstract: We consider reversely-degraded secure-communication channels, for which the secrecy capacity is zero if there is no channel feedback. Specific

Flow matching for Sentinel-2 super-resolution: implementation, application, and implications

ResearchDGX agent

arXiv:2605.00367v1 Announce Type: new Abstract: Developing robust techniques for super-resolution of satellite imagery involves navigating commonly observed trade-offs between spectral fidelity and pe

Free Energy Surface Sampling via Reduced Flow Matching

ResearchDGX agent

arXiv:2605.00337v1 Announce Type: new Abstract: Sampling the free energy surface, namely, the distribution of collective variables (CVs), is a crucial problem in statistical physics, as it underpins a

From Local to Global to Mechanistic: An iERF-Centered Unified Framework for Interpreting Vision Models

ResearchDGX agent

arXiv:2605.00474v1 Announce Type: new Abstract: Modern vision models achieve remarkable accuracy, but explaining where evidence arises, what the model encodes, and how internal computations assemble t

GAFSV-Net: A Vision Framework for Online Signature Verification

ResearchDGX agent

arXiv:2605.00120v1 Announce Type: new Abstract: Online signature verification (OSV) requires distinguishing skilled forgeries from genuine samples under high intra-class variability and with very few

GD4: Graph-based Discrete Denoising Diffusion for MIMO Detection

ResearchDGX agent

arXiv:2605.00423v1 Announce Type: new Abstract: In wireless communications, recovering the optimal solution to the multiple-input multiple-output (MIMO) detection problem is NP-hard. Obtaining high-qu

Generating Statistical Charts with Validation-Driven LLM Workflows

ResearchDGX agent

arXiv:2605.00800v1 Announce Type: new Abstract: Generating diverse, readable statistical charts from tabular data remains challenging for LLMs, as many failures become apparent after rendering and are

GMGaze: MoE-Based Context-Aware Gaze Estimation with CLIP and Multiscale Transformer

ResearchDGX agent

arXiv:2605.00799v1 Announce Type: new Abstract: Gaze estimation methods commonly use facial appearances to predict the direction of a person gaze. However, previous studies show three major challenges

Gradient Regularized Newton Boosting Trees with Global Convergence

ResearchDGX agent

arXiv:2605.00581v1 Announce Type: cross Abstract: Gradient Boosting Decision Trees (GBDTs) dominate tabular machine learning, with modern implementations like XGBoost, LightGBM, and CatBoost being bas

Graph Rewiring in GNNs to Mitigate Over-Squashing and Over-Smoothing: A Survey

ResearchDGX agent

arXiv:2411.17429v2 Announce Type: replace Abstract: Graph Neural Networks are powerful models for learning from graph-structured data, yet their effectiveness is often limited by two critical challeng

H-RAG at SemEval-2026 Task 8: Hierarchical Parent-Child Retrieval for Multi-Turn RAG Conversations

ResearchDGX agent

arXiv:2605.00631v1 Announce Type: new Abstract: We present H-RAG, our submission to SemEval-2026 Task 8 (MTRAGEval), addressing both Task A (Retrieval) and Task C (Generation with Retrieved Passages).

High-Speed Vision Improves Zero-Shot Semantic Understanding of Human Actions

ResearchDGX agent

arXiv:2605.00496v1 Announce Type: new Abstract: Understanding human actions from visual observations is essential for human--robot interaction, particularly when semantic interpretation of unfamiliar

Human-in-the-Loop Meta Bayesian Optimization for Fusion Energy and Scientific Applications

ResearchDGX agent

arXiv:2605.00068v1 Announce Type: new Abstract: Inertial Confinement Fusion (ICF) holds transformative promise for sustainable, near-limitless clean energy, yet remains constrained by prohibitively hi

Information-geometric adaptive sampling for graph diffusion

ResearchDGX agent

arXiv:2605.00250v1 Announce Type: cross Abstract: Standard diffusion models for graph generation typically rely on uniform time-stepping, an approach that overlooks the non-homogeneous dynamics of dis

Information-Theoretic Generalization Bounds for Stochastic Gradient Descent with Predictable Virtual Noise

ResearchDGX agent

arXiv:2605.00064v1 Announce Type: new Abstract: Information-theoretic generalization bounds analyze stochastic optimization by relating expected generalization error to the mutual information between

Intrinsic Gradient Suppression for Label-Noise Prompt Tuning in Vision-Language Models

ResearchDGX agent

arXiv:2605.00591v1 Announce Type: new Abstract: Contrastive vision-language models like CLIP exhibit remarkable zero-shot generalization. However, prompt tuning remains highly sensitive to label noise

Is Textual Similarity Invariant under Machine Translation? Evidence Based on the Political Manifesto Corpus

ResearchDGX agent

arXiv:2605.00618v1 Announce Type: new Abstract: We investigate the extent to which cosine similarity between paragraph embeddings is invariant under machine translation, using the Manifesto Corpus of

It's Never Too Late: Noise Optimization for Collapse Recovery in Trained Diffusion Models

ResearchDGX agent

arXiv:2601.00090v2 Announce Type: replace Abstract: Contemporary text-to-image models exhibit a surprising degree of mode collapse, as can be seen when sampling several images given the same text prom

Knowing when to trust machine-learned interatomic potentials

ResearchDGX agent

arXiv:2605.00640v1 Announce Type: new Abstract: Prevailing machine-learned interatomic potential (MLIP) uncertainty-quantification methods rely on ensembles of independently trained backbones. These m

Koopman-Assisted Reinforcement Learning

ResearchDGX agent

arXiv:2403.02290v2 Announce Type: replace-cross Abstract: The Bellman equation and its continuous form, the Hamilton-Jacobi-Bellman equation, are ubiquitous in reinforcement learning and control theor

LandSegmenter: Towards a Flexible Foundation Model for Land Use and Land Cover Mapping

ResearchDGX agent

arXiv:2511.08156v2 Announce Type: replace Abstract: Land Use and Land Cover (LULC) mapping is a fundamental task in Earth Observation (EO). However, current LULC models are typically developed for a s

LASE: Language-Adversarial Speaker Encoding for Indic Cross-Script Identity Preservation

ResearchDGX agent

arXiv:2605.00777v1 Announce Type: cross Abstract: A speaker encoder used in multilingual voice cloning should treat the same speaker identically regardless of which script the audio was uttered in. Of

Last-Iterate Analyses of FTRL with the 1/2-Tsallis Entropy in Stochastic Bandits

ResearchDGX agent

arXiv:2510.22819v2 Announce Type: replace Abstract: The convergence analysis of online learning algorithms is central to machine learning theory, where the last-iterate convergence is particularly imp

Learning Multimodal Energy-Based Model with Multimodal Variational Auto-Encoder via MCMC Revision

ResearchDGX agent

arXiv:2605.00644v1 Announce Type: new Abstract: Energy-based models (EBMs) are a flexible class of deep generative models and are well-suited to capture complex dependencies in multimodal data. Howeve

Let ViT Speak: Generative Language-Image Pre-training

ResearchDGX agent

arXiv:2605.00809v1 Announce Type: new Abstract: In this paper, we present extbf{Gen}erative extbf{L}anguage-extbf{I}mage extbf{P}re-training (GenLIP), a minimalist generative pretraining framework for

Leveraging Vision-Language Models as Weak Annotators in Active Learning

ResearchDGX agent

arXiv:2605.00480v1 Announce Type: new Abstract: Active learning aims to reduce annotation cost by selectively querying informative samples for supervision under a limited labeling budget. In this work

LLM DNA: Tracing Model Evolution via Functional Representations

ResearchDGX agent

arXiv:2509.24496v3 Announce Type: replace Abstract: The explosive growth of large language models (LLMs) has created a vast but opaque landscape: millions of models exist, yet their evolutionary relat

Lost in State Space: Probing Frozen Mamba Representations

ResearchDGX agent

arXiv:2605.00253v1 Announce Type: new Abstract: Mamba's recurrent state h_t is, by construction, a compressed summary of every token seen so far. This raises a tempting hypothesis: if we extract token

Making Every Verified Token Count: Adaptive Verification for MoE Speculative Decoding

ResearchDGX agent

arXiv:2605.00342v1 Announce Type: new Abstract: Tree-based speculative decoding accelerates autoregressive generation by verifying multiple draft candidates in parallel, but this advantage weakens for

Matroid Algorithms Under Size-Sensitive Independence Oracles

ResearchDGX agent

arXiv:2605.00201v1 Announce Type: cross Abstract: The standard oracle model for matroid algorithms assumes that each independence query can be answered in constant time, regardless of the size of the

Mean-field limit from general mixtures of experts to quantum neural networks

ResearchDGX agent

arXiv:2501.14660v2 Announce Type: replace-cross Abstract: In this work, we study the asymptotic behavior of Mixture of Experts (MoE) trained via gradient flow on supervised learning problems. Our main

MMAudioReverbs: Video-Guided Acoustic Modeling for Dereverberation and Room Impulse Response Estimation

ResearchDGX agent

arXiv:2605.00431v1 Announce Type: cross Abstract: Although recent video-to-audio (V2A) models excelled at synthesizing semantically plausible sounds from visual inputs, they do not explicitly model ro

Modeling Subjective Urban Perception with Human Gaze

ResearchDGX agent

arXiv:2605.00764v1 Announce Type: new Abstract: Urban perception describes how people subjectively evaluate urban environments, shaping how cities are experienced and understood. Existing computationa

Near-optimal and Efficient First-Order Algorithm for Multi-Task Learning with Shared Linear Representation

ResearchDGX agent

arXiv:2605.00473v1 Announce Type: new Abstract: Multi-task learning (MTL) has emerged as a pivotal paradigm in machine learning by leveraging shared structures across multiple related tasks. Despite i

NRGPT: An Energy-based Alternative for GPT

ResearchDGX agent

arXiv:2512.16762v3 Announce Type: replace Abstract: Generative Pre-trained Transformer (GPT) architectures are the most popular design for language modeling. Energy-based modeling is a different parad

Observable Performance Does Not Fully Reflect System Organization: A Multi-Level Analysis of Gait Dynamics Under Occlusal Constraint

ResearchDGX agent

arXiv:2605.00778v1 Announce Type: new Abstract: In biomechanical systems, observable performance is often used as a proxy for underlying system organization. However, this assumption implicitly presum

Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning

ResearchDGX agent

arXiv:2605.00347v1 Announce Type: cross Abstract: Given the rapidly growing capabilities of vision-language models (VLMs), extending them to interactive decision-making tasks such as video games has e

On the Expressive Power of Contextual Relations in Transformers

ResearchDGX agent

arXiv:2603.25860v2 Announce Type: replace-cross Abstract: Transformer architectures have achieved remarkable empirical success in modeling contextual relations, yet a clear understanding of their expr

Optimal hypersurface decision trees

ResearchDGX agent

arXiv:2509.12057v3 Announce Type: replace Abstract: The study of optimal decision trees has gained increasing attention in recent years; however, despite substantial progress, it still suffers from tw

[P] QLoRA Fine-Tuning of Qwen2.5-1.5B for CEFR English Proficiency Classification (A1–C2) [P]

ResearchDGX agent

This post likely describes a machine learning project implementing QLoRA (Quantized Low-Rank Adaptation) fine-tuning on the Qwen2.5-1.5B model to classify English language proficiency levels according

PhysEdit: Physically-Consistent Region-Aware Image Editing via Adaptive Spatio-Temporal Reasoning

ResearchDGX agent

arXiv:2605.00707v1 Announce Type: new Abstract: Image editing instructions are heterogeneous: a color swap, an object insertion, and a physical-action edit all demand different spatial coverage and di

PhysiGen: Integrating Collision-Aware Physical Constraints for High-Fidelity Human-Human Interaction Generation

ResearchDGX agent

arXiv:2605.00517v1 Announce Type: new Abstract: Despite substantial progress in text-driven 3D human motion synthesis, generating realistic multi-person interaction sequences remains challenging. Nota

Possibilistic Predictive Uncertainty for Deep Learning

ResearchDGX agent

arXiv:2605.00600v1 Announce Type: cross Abstract: Deep neural networks achieve impressive results across diverse applications, yet their overconfidence on unseen inputs necessitates reliable epistemic

Posterior Augmented Flow Matching

ResearchDGX agent

arXiv:2605.00825v1 Announce Type: new Abstract: Flow matching (FM) trains a time-dependent vector field that transports samples from a simple prior to a complex data distribution. However, for high-di

Privacy Amplification in Differentially Private Zeroth-Order Optimization with Hidden States

ResearchDGX agent

arXiv:2506.00158v2 Announce Type: replace Abstract: Zeroth-order optimization has emerged as a promising approach for fine-tuning large language models under differential privacy (DP) and memory const

Quantum Gradient-Based Approach for Edge and Corner Detection Using Sobel Kernels

ResearchDGX agent

arXiv:2605.00744v1 Announce Type: new Abstract: Edge detection refers to identifying points in a digital image where intensity changes sharply, indicating object boundaries or structural features. Cor

Quantum Interval Bound Propagation for Certified Training of Quantum Neural Networks

ResearchDGX agent

arXiv:2605.00747v1 Announce Type: cross Abstract: Quantum machine learning is a promising field for efficiently learning features of a dataset to perform a specified task, such as classification. Inte

Randomized Subspace Nesterov Accelerated Gradient

ResearchDGX agent

arXiv:2605.00740v1 Announce Type: cross Abstract: Randomized-subspace methods reduce the cost of first-order optimization by using only low-dimensional projected-gradient information, a feature that i

RAT+: Train Dense, Infer Sparse -- Recurrence Augmented Attention for Dilated Inference

ResearchDGX agent

arXiv:2602.18196v3 Announce Type: replace Abstract: Structured dilated attention has an appealing inference-time efficiency knob: it reduces the FLOPs of attention and the KV cache size by a factor of

REALM: An RGB and Event Aligned Latent Manifold for Cross-Modal Perception

ResearchDGX agent

arXiv:2605.00271v1 Announce Type: new Abstract: Event cameras provide several unique advantages over standard frame-based sensors, including high temporal resolution, low latency, and robustness to ex

Representation in large language models

ResearchDGX agent

arXiv:2501.00885v2 Announce Type: replace Abstract: The extraordinary success of recent Large Language Models (LLMs) on a diverse array of tasks has led to an explosion of scientific and philosophical

Rethinking LLM Ensembling from the Perspective of Mixture Models

ResearchDGX agent

arXiv:2605.00419v1 Announce Type: cross Abstract: Model ensembling is a well-established technique for improving the performance of machine learning models. Conventionally, this involves averaging the

Revealing graph bandits for maximizing local influence

ResearchDGX agent

arXiv:2605.00489v1 Announce Type: new Abstract: We study a graph bandit setting where the objective of the learner is to detect the most influential node of a graph by requesting as little information

Reward Modeling from Natural Language Human Feedback

ResearchDGX agent

arXiv:2601.07349v3 Announce Type: replace Abstract: Reinforcement Learning with Verifiable reward (RLVR) on preference data has become the mainstream approach for training Generative Reward Models (GR

Riemannian MeanFlow

ResearchDGX agent

arXiv:2602.07744v3 Announce Type: replace Abstract: Diffusion and flow models have become the dominant paradigm for generative modeling on Riemannian manifolds, with successful applications in protein

← Previous
1…285286287288289…432
Next →