AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
19,194 results
21 May 2026

Gated DeltaNet has been one of my favorite 'hybrid attention' newcomers in the good old transformer stack. Excited to see Gated DeltaNet-2. …

ResearchDGX agent

Gated DeltaNet has been one of my favorite 'hybrid attention' newcomers in the good old transformer stack. Excited to see Gated DeltaNet-2. Adding it to my reading stack. In the meantime, I have a pri

Gaze into the Details: Locality-Sensitive Enhancement for OCTA Retinal Vessel Segmentation

ResearchDGX agent

arXiv:2605.20651v1 Announce Type: new Abstract: Existing deep learning frameworks for Optical Coherence Tomography Angiography (OCTA) vessel segmentation are largely derived from the U-Net architectur

Generation of Heterogeneous PET Images from Uniform Organ Activity Maps Using a Pretrained Domain-Adapted Diffusion Model

Research

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.20267v1 Announce Type: new Abstract: Synthetic PET images are valuable for quantitative imaging workflow development, scalable virtual imaging trials, and deep learning model training, but

Generative AI Practices, Literacy, and Divides: An Empirical Analysis in the Italian Context

ResearchDGX agent

arXiv:2512.03671v2 Announce Type: replace Abstract: The rise of generative AI (GenAI) chatbots accessible via conversational interfaces is transforming digital interactions and holds economic promise.

Genetic Programming with Transformer-Based Mutation for Approximate Circuit Design

ResearchDGX agent

arXiv:2605.21055v1 Announce Type: cross Abstract: A recent trend is to leverage machine learning models to improve the evolutionary design and optimization process. We propose a novel transformer-base

GeoPT: Scaling Physics Simulation via Lifted Geometric Pre-Training

ResearchDGX agent

arXiv:2602.20399v2 Announce Type: replace Abstract: Neural simulators promise efficient surrogates for physics simulation, but scaling them is bottlenecked by the prohibitive cost of generating high-f

Goodbye Drift: Anchored Tree Sampling for Long-Horizon Video-to-Video Generation

ResearchDGX agent

arXiv:2605.20476v1 Announce Type: new Abstract: Long-horizon video generation suffers from two intertwined issues. First, there is drift, where video quality degrades over time. Second, there are cont

Gradient Scalability and Taylor Surrogation of Quantum Cost Landscapes

ResearchDGX agent

arXiv:2507.06344v3 Announce Type: replace-cross Abstract: Variational Quantum Algorithms are promising candidates for near-term quantum computing, yet they face scalability challenges due to barren pl

GraphCSVAE: Graph Categorical Structured Variational Autoencoder for Spatiotemporal Auditing of Physical Vulnerability Towards Sustainable Post-Disaster Risk Reduction

ResearchDGX agent

arXiv:2509.10308v2 Announce Type: replace Abstract: In the aftermath of disasters, many institutions worldwide face challenges in monitoring changes in disaster risk, limiting assessment of progress t

Grounding Driving VLA via Inverse Kinematics

ResearchDGX agent

arXiv:2605.21061v1 Announce Type: new Abstract: Existing Driving VLAs predict trajectories while largely ignoring their visual tokens -- a phenomenon we trace not to insufficient training but to a str

Group-Aware Matrix Estimation and Latent Subspace Recovery

ResearchDGX agent

arXiv:2605.20559v1 Announce Type: cross Abstract: Modern matrix completion problems often involve heterogeneous data whose rows simultaneously belong to many meta-categories, such as demographic and a

GSA-YOLO: A High-Efficiency Framework via Structured Sparsity and Adaptive Knowledge Distillation for Real-Time X-ray Security Inspection

ResearchDGX agent

arXiv:2605.20669v1 Announce Type: new Abstract: X-ray security inspection requires accurate real-time detection of prohibited items, but existing models often struggle to balance the challenges of sev

HADS-Net:A Hybrid Attention-Augmented Dual-Stream Network with Physics-Informed Augmentation for Breast Ultrasound Image Classification

ResearchDGX agent

arXiv:2605.20536v1 Announce Type: new Abstract: Accurate classification of breast ultrasound images into benign, malignant, and normal categories is a critical clinical task complicated by speckle noi

HAPS: Rethinking Image Similarity for Virtual Staining

ResearchDGX agent

arXiv:2605.20362v1 Announce Type: new Abstract: Virtual staining of histopathology images (e.g., H&E-IHC) is an emerging tool in digital pathology, enabling faster and cheaper workflows by synthesizin

Hiding in Plain Sight: Finding MAHA on Reddit

ResearchDGX agent

arXiv:2605.20435v1 Announce Type: cross Abstract: Make America Healthy Again (MAHA) is a national health movement that encompasses a striking mix of beliefs, from broadly accepted concerns about good

HiRes: Inspectable Precedent Memory for Reaction Condition Recommendation

ResearchDGX agent

arXiv:2605.21420v1 Announce Type: new Abstract: Reaction condition recommendation sits immediately after retrosynthetic disconnection selection, and in practice, chemists require both accurate predict

How Many Human Survey Respondents is a Large Language Model Worth? An Uncertainty Quantification Perspective

ResearchDGX agent

arXiv:2502.17773v5 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate survey responses, but synthetic data can be misaligned with the human populatio

How Open Must Language Models be to Enable Reliable Scientific Inference?

ResearchDGX agent

arXiv:2603.26539v2 Announce Type: replace Abstract: How does the extent to which a model is open or closed impact the scientific inferences that can be drawn from research that involves it? In this pa

How You Move Tells What You'll Do: Trajectory-Conditioned Egocentric Prediction

ResearchDGX agent

arXiv:2605.20388v1 Announce Type: new Abstract: Predicting how a person's first-person view will evolve (what action will follow, what plan completes a task, whether an in-progress shot will score) is

Hybrid Machine Learning Model for Forest Height Estimation from TanDEM-X and Landsat Data

ResearchDGX agent

arXiv:2605.20997v1 Announce Type: new Abstract: Integrating machine learning (ML) with physical models (PM) has emerged as a promising way of retrieving geophysical parameters from remote sensing data

I'm very wrong most of the time

ResearchDGX agent

Francois Chollet reflects on the frequency of his mistakes and incorrect predictions, likely discussing humility in scientific work and the importance of intellectual openness in AI research and devel

Improved Guarantees for Constrained Online Convex Optimization via Self-Contraction

ResearchDGX agent

arXiv:2605.21107v1 Announce Type: new Abstract: We consider Constrained Online Convex Optimization (COCO) with adversarially chosen constraints. At each round, the learner chooses an action before obs

Improving 3D Gaussian Splatting Compression by Scene-Adaptive Lattice Vector Quantization

ResearchDGX agent

arXiv:2509.13482v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) is rapidly gaining popularity for its photorealistic rendering quality and real-time performance, but it generates mass

Instance Discrimination for Link Prediction

ResearchDGX agent

arXiv:2605.20257v1 Announce Type: new Abstract: Recently, instance discrimination models have emerged as a major solution for self-supervised learning. Having already demonstrated its effectiveness in

Interpretable Discriminative Text Representations via Agreement and Label Disentanglement

ResearchDGX agent

arXiv:2605.20693v1 Announce Type: new Abstract: Interpretable text representations should expose coordinates that are not only predictive, but also meaningful enough for independent auditors to apply.

iReasoner: Trajectory-Aware Intrinsic Reasoning Supervision for Self-Evolving Large Multimodal Models

ResearchDGX agent

arXiv:2601.05877v3 Announce Type: replace Abstract: Recent work shows that large multimodal models (LMMs) can self-improve from unlabeled data via self-play and intrinsic feedback. Yet existing self-e

Iterative LLM-based improvement for French Clinical Interview Transcription and Speaker Diarization

ResearchDGX agent

arXiv:2603.00086v2 Announce Type: replace Abstract: Automatic speech recognition for French medical conversations remains challenging, with word error rates often exceeding 30% in spontaneous clinical

It’s pronounced “Hermes”, not “Hermes”

ResearchDGX agent

This post likely clarifies the correct pronunciation of 'Hermes' (the Greek mythological messenger god and luxury brand), humorously highlighting a common mispronunciation among English speakers. The

Large Language Models Unpack Complex Political Opinions through Target-Stance Extraction

ResearchDGX agent

arXiv:2603.23531v2 Announce Type: replace Abstract: Political polarization emerges from a complex interplay of beliefs about policies, figures, and issues. However, most computational analyses reduce

Latent Dynamics for Full Body Avatar Animation

ResearchDGX agent

arXiv:2605.21478v1 Announce Type: new Abstract: Pose-driven full-body avatars built on neural rendering produce high-quality novel views of a captured subject. Yet loose clothing and other dynamic ele

Learning First Integrals via Backward-Generated Data and Guided Reinforcement Learning

ResearchDGX agent

arXiv:2605.21160v1 Announce Type: new Abstract: The discovery of first integrals is of fundamental scientific importance for understanding conservation laws in dynamical systems. However, existing sym

Learning Structural Latent Points for Efficient Visual Representations in Robotic Manipulation

ResearchDGX agent

arXiv:2605.21258v1 Announce Type: new Abstract: Current 3D-aware pretraining methods for embodied perception and manipulation are largely built on differentiable rendering frameworks, producing either

Less Data, Faster Training: repeating smaller datasets speeds up learning via sampling biases

ResearchDGX agent

arXiv:2605.20314v1 Announce Type: new Abstract: This work investigates the ``small-vs-large gap'', where repeating on fewer samples can lead to compute saving during training compared to using a large

Leveraging Large Language Models for Sentiment Analysis: Multi-Modal Analysis of Decentraland's MANA Token

ResearchDGX agent

arXiv:2605.20192v1 Announce Type: new Abstract: Decentraland, a decentralized virtual reality platform operating within the expanding Metaverse ecosystem, utilizes its native MANA token to facilitate

Lisbon Machine Learning School (LxMLS 2026) [D]

ResearchDGX agent

LxMLS 2026 is a 6-day in-person event scheduled for July 20-25 at Instituto Superior Técnico that covers machine learning topics from theory to practice for solving natural language processing problem

Local-sensitive connectivity filter (ls-cf): A post-processing unsupervised improvement of the frangi, hessian and vesselness filters for multimodal vessel segmentation

ResearchDGX agent

arXiv:2605.21251v1 Announce Type: cross Abstract: A retinal vessel analysis is a procedure that can be used as an assessment of risks to the eye. This work proposes an unsupervised multimodal approach

Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning

ResearchDGX agent

arXiv:2605.20201v1 Announce Type: new Abstract: Recent large language models support inputs of up to 10 million tokens, yet they perform poorly on long-context tasks that require complex reasoning. Su

Lowering the Barrier to IREX Participation: Open-Source Algorithms, Toolkit, and Benchmarking for Iris Recognition

ResearchDGX agent

arXiv:2605.20735v1 Announce Type: new Abstract: This paper proposes two new open-source iris recognition algorithms, providing both Python and IREX-compliant C++ implementations to be submitted to the

Machine-Learned Force Fields for Lattice Dynamics at Coupled-Cluster Level Accuracy

ResearchDGX agent

arXiv:2507.06929v2 Announce Type: replace-cross Abstract: We investigate Machine-Learned Force Fields (MLFFs) trained on approximate Density Functional Theory (DFT) and Coupled Cluster (CC) level pote

Machine-Learning-Enhanced Non-Invasive Testing for MASLD Fibrosis: Shallow-Deep Neural Networks Versus FIB-4, Tabular Foundation Models, and Large Language Models

ResearchDGX agent

arXiv:2605.20523v1 Announce Type: new Abstract: Advanced fibrosis is a major determinant of liver-related morbidity in metabolic dysfunction-associated steatotic liver disease (MASLD). FIB-4 is widely

Manga109-v2026: Revisiting Manga109 Annotations for Modern Manga Understanding

ResearchDGX agent

arXiv:2605.21182v1 Announce Type: new Abstract: Manga is a culturally distinctive multimodal medium and one of the most influential forms of Japanese popular culture. As AI systems increasingly target

Map-Mono-Ego: Map-Grounded Global Human Pose Estimation from Monocular Egocentric Video

ResearchDGX agent

arXiv:2605.20889v1 Announce Type: new Abstract: Monocular egocentric human pose estimation is essential for ubiquitous activity monitoring. However, understanding the user's absolute location within t

Matryoshka Concept Bottleneck Models

ResearchDGX agent

arXiv:2605.20612v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) have emerged as a prominent paradigm for interpretable deep learning, learning by grounding predictions in human-unders

Maxitive Donsker-Varadhan Formulation for Possibilistic Variational Inference

ResearchDGX agent

arXiv:2511.21223v2 Announce Type: replace-cross Abstract: Variational inference (VI) is a cornerstone of modern Bayesian learning, enabling approximate inference in complex models. However, its formul

Mercer Large-Scale Kernel Machines from Ridge Function Perspective

ResearchDGX agent

arXiv:2307.11925v3 Announce Type: replace Abstract: To present Mercer large-scale kernel machines from a ridge function perspective, we recall the results by Lin and Pinkus from {it Fundamentality of

MeshTailor: Cutting Seams via Generative Mesh Traversal

ResearchDGX agent

arXiv:2603.27309v2 Announce Type: replace-cross Abstract: We present MeshTailor, the first mesh-native generative framework for synthesizing edge-aligned seams on 3D surfaces. Unlike prior optimizatio

Metaphors in Literary Post-Editing: Opening Pandora's Box?

ResearchDGX agent

arXiv:2605.21178v1 Announce Type: new Abstract: This paper investigates how post-editors of literary texts react and respond to the way metaphors have been translated by Neu ral Machine Translation (N

Mind Your Margin and Boundary: Are Your Distilled Datasets Truly Robust?

ResearchDGX agent

arXiv:2605.20606v1 Announce Type: new Abstract: Dataset distillation (DD) compresses a large training set into a small synthetic set for efficient training, but most DD methods optimize only clean acc

Modeling and Control of a Pneumatic Morphing Soft Quadrotor based on the SOFA Framework for Dynamic Soft Robotic Simulation

ResearchDGX agent

arXiv:2605.21031v1 Announce Type: new Abstract: This article presents a novel SOFA based finite element method for the soft body modeling and the corresponding dynamic simulation and control of a pneu

Modeling Temporal scRNA-seq Data with Latent Gaussian Process and Optimal Transport

ResearchDGX agent

arXiv:2605.20989v1 Announce Type: new Abstract: Single-cell RNA sequencing provides insights into gene expression at single-cell resolution, yet inferring temporal processes from these static snapshot

Modular Multimodal Classification Without Fine-Tuning: A Simple Compositional Approach

ResearchDGX agent

arXiv:2605.20674v1 Announce Type: new Abstract: We introduce CoMET, extit{extbf{C}omposing extbf{M}odality extbf{E}ncoders with extbf{T}abular foundation models}, a simple yet highly competitive metho

Most Transformer Modifications Still Do Not Transfer at 1-3B: A 2020-2026 Update to Narang et al. (2021) with Downstream Evaluation and a Noise Floor

ResearchDGX agent

arXiv:2605.20798v1 Announce Type: cross Abstract: Narang et al. (2021) evaluated 40+ Transformer modifications at T5-base scale and concluded that most did not transfer. Five years later, the typical

Motion-Robust Deep Reconstruction for Free-Breathing Cardiac Cine MRI

ResearchDGX agent

arXiv:2605.20687v1 Announce Type: cross Abstract: Conventional cardiac cine MRI relies on breath-hold Cartesian acquisitions, which are vulnerable to motion artifacts and can be uncomfortable or infea

Multi-Channel Replay Speech Detection using Acoustic Maps

ResearchDGX agent

arXiv:2602.16399v2 Announce Type: replace-cross Abstract: Replay attacks remain a critical vulnerability for automatic speaker verification systems, particularly in real-time voice assistant applicati

Musical Attention Transformer: Music Generation Using a Music-Specific Attention Model

ResearchDGX agent

arXiv:2605.21081v1 Announce Type: cross Abstract: This study aims to enhance the quality of music generation using Transformers by incorporating meta-information. While Transformer-based approaches ar

NeighborDiv: Training-free Zero-shot Generalist Graph Anomaly Detection via Neighbor Diversity

ResearchDGX agent

arXiv:2605.20879v1 Announce Type: new Abstract: Graph Anomaly Detection (GAD) is increasingly shifting to Generalist GAD (GGAD) for cross-domain 'one-for-all' detection, but existing GGAD methods pred

Neural Estimation of Pairwise Mutual Information in Masked Discrete Sequence Models

ResearchDGX agent

arXiv:2605.20187v1 Announce Type: new Abstract: Understanding dependencies between variables is critical for interpretability and efficient generation in masked diffusion models (MDMs), yet these mode

New VIDEO: From LLM Wikis to LLM Artifacts Shared all my thoughts on why LLM wikis and HTML artifacts are a big deal. Plus, new tools to hel…

ResearchDGX agent

New VIDEO: From LLM Wikis to LLM Artifacts Shared all my thoughts on why LLM wikis and HTML artifacts are a big deal. Plus, new tools to help you build wikis and artifacts with agents. Just getting st

Nonlocal operator learning for fMRI encoding and decoding tasks

ResearchDGX agent

arXiv:2605.20389v1 Announce Type: new Abstract: Functional MRI data exhibit high-dimensional spatiotemporal structure, making both prediction and decoding challenging. In this work, we investigate neu

OCTOPUS: Optimized KV Cache for Transformers via Octahedral Parametrization Under optimal Squared error quantization

ResearchDGX agent

arXiv:2605.21226v1 Announce Type: new Abstract: The key-value (KV) cache dominates memory bandwidth and footprint in long-context autoregressive inference. Recent rotation-preconditioned codecs (Turbo

← Previous
1…186187188189190…320
Next →