AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,898 results
21 May 2026

Improving 3D Gaussian Splatting Compression by Scene-Adaptive Lattice Vector Quantization

ResearchDGX agent

arXiv:2509.13482v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) is rapidly gaining popularity for its photorealistic rendering quality and real-time performance, but it generates mass

Instance Discrimination for Link Prediction

ResearchDGX agent

arXiv:2605.20257v1 Announce Type: new Abstract: Recently, instance discrimination models have emerged as a major solution for self-supervised learning. Having already demonstrated its effectiveness in

Interpretable Discriminative Text Representations via Agreement and Label Disentanglement

ResearchDGX agent

arXiv:2605.20693v1 Announce Type: new Abstract: Interpretable text representations should expose coordinates that are not only predictive, but also meaningful enough for independent auditors to apply.

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

iReasoner: Trajectory-Aware Intrinsic Reasoning Supervision for Self-Evolving Large Multimodal Models

ResearchDGX agent

arXiv:2601.05877v3 Announce Type: replace Abstract: Recent work shows that large multimodal models (LMMs) can self-improve from unlabeled data via self-play and intrinsic feedback. Yet existing self-e

Iterative LLM-based improvement for French Clinical Interview Transcription and Speaker Diarization

ResearchDGX agent

arXiv:2603.00086v2 Announce Type: replace Abstract: Automatic speech recognition for French medical conversations remains challenging, with word error rates often exceeding 30% in spontaneous clinical

Large Language Models Unpack Complex Political Opinions through Target-Stance Extraction

ResearchDGX agent

arXiv:2603.23531v2 Announce Type: replace Abstract: Political polarization emerges from a complex interplay of beliefs about policies, figures, and issues. However, most computational analyses reduce

Latent Dynamics for Full Body Avatar Animation

ResearchDGX agent

arXiv:2605.21478v1 Announce Type: new Abstract: Pose-driven full-body avatars built on neural rendering produce high-quality novel views of a captured subject. Yet loose clothing and other dynamic ele

Learning First Integrals via Backward-Generated Data and Guided Reinforcement Learning

ResearchDGX agent

arXiv:2605.21160v1 Announce Type: new Abstract: The discovery of first integrals is of fundamental scientific importance for understanding conservation laws in dynamical systems. However, existing sym

Learning Structural Latent Points for Efficient Visual Representations in Robotic Manipulation

ResearchDGX agent

arXiv:2605.21258v1 Announce Type: new Abstract: Current 3D-aware pretraining methods for embodied perception and manipulation are largely built on differentiable rendering frameworks, producing either

Less Data, Faster Training: repeating smaller datasets speeds up learning via sampling biases

ResearchDGX agent

arXiv:2605.20314v1 Announce Type: new Abstract: This work investigates the ``small-vs-large gap'', where repeating on fewer samples can lead to compute saving during training compared to using a large

Local-sensitive connectivity filter (ls-cf): A post-processing unsupervised improvement of the frangi, hessian and vesselness filters for multimodal vessel segmentation

ResearchDGX agent

arXiv:2605.21251v1 Announce Type: cross Abstract: A retinal vessel analysis is a procedure that can be used as an assessment of risks to the eye. This work proposes an unsupervised multimodal approach

Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning

ResearchDGX agent

arXiv:2605.20201v1 Announce Type: new Abstract: Recent large language models support inputs of up to 10 million tokens, yet they perform poorly on long-context tasks that require complex reasoning. Su

Machine-Learned Force Fields for Lattice Dynamics at Coupled-Cluster Level Accuracy

ResearchDGX agent

arXiv:2507.06929v2 Announce Type: replace-cross Abstract: We investigate Machine-Learned Force Fields (MLFFs) trained on approximate Density Functional Theory (DFT) and Coupled Cluster (CC) level pote

Machine-Learning-Enhanced Non-Invasive Testing for MASLD Fibrosis: Shallow-Deep Neural Networks Versus FIB-4, Tabular Foundation Models, and Large Language Models

ResearchDGX agent

arXiv:2605.20523v1 Announce Type: new Abstract: Advanced fibrosis is a major determinant of liver-related morbidity in metabolic dysfunction-associated steatotic liver disease (MASLD). FIB-4 is widely

Map-Mono-Ego: Map-Grounded Global Human Pose Estimation from Monocular Egocentric Video

ResearchDGX agent

arXiv:2605.20889v1 Announce Type: new Abstract: Monocular egocentric human pose estimation is essential for ubiquitous activity monitoring. However, understanding the user's absolute location within t

Matryoshka Concept Bottleneck Models

ResearchDGX agent

arXiv:2605.20612v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) have emerged as a prominent paradigm for interpretable deep learning, learning by grounding predictions in human-unders

Maxitive Donsker-Varadhan Formulation for Possibilistic Variational Inference

ResearchDGX agent

arXiv:2511.21223v2 Announce Type: replace-cross Abstract: Variational inference (VI) is a cornerstone of modern Bayesian learning, enabling approximate inference in complex models. However, its formul

Mercer Large-Scale Kernel Machines from Ridge Function Perspective

ResearchDGX agent

arXiv:2307.11925v3 Announce Type: replace Abstract: To present Mercer large-scale kernel machines from a ridge function perspective, we recall the results by Lin and Pinkus from {it Fundamentality of

MeshTailor: Cutting Seams via Generative Mesh Traversal

ResearchDGX agent

arXiv:2603.27309v2 Announce Type: replace-cross Abstract: We present MeshTailor, the first mesh-native generative framework for synthesizing edge-aligned seams on 3D surfaces. Unlike prior optimizatio

Metaphors in Literary Post-Editing: Opening Pandora's Box?

ResearchDGX agent

arXiv:2605.21178v1 Announce Type: new Abstract: This paper investigates how post-editors of literary texts react and respond to the way metaphors have been translated by Neu ral Machine Translation (N

Mind Your Margin and Boundary: Are Your Distilled Datasets Truly Robust?

ResearchDGX agent

arXiv:2605.20606v1 Announce Type: new Abstract: Dataset distillation (DD) compresses a large training set into a small synthetic set for efficient training, but most DD methods optimize only clean acc

Modeling and Control of a Pneumatic Morphing Soft Quadrotor based on the SOFA Framework for Dynamic Soft Robotic Simulation

ResearchDGX agent

arXiv:2605.21031v1 Announce Type: new Abstract: This article presents a novel SOFA based finite element method for the soft body modeling and the corresponding dynamic simulation and control of a pneu

Modeling Temporal scRNA-seq Data with Latent Gaussian Process and Optimal Transport

ResearchDGX agent

arXiv:2605.20989v1 Announce Type: new Abstract: Single-cell RNA sequencing provides insights into gene expression at single-cell resolution, yet inferring temporal processes from these static snapshot

Modular Multimodal Classification Without Fine-Tuning: A Simple Compositional Approach

ResearchDGX agent

arXiv:2605.20674v1 Announce Type: new Abstract: We introduce CoMET, extit{extbf{C}omposing extbf{M}odality extbf{E}ncoders with extbf{T}abular foundation models}, a simple yet highly competitive metho

Most Transformer Modifications Still Do Not Transfer at 1-3B: A 2020-2026 Update to Narang et al. (2021) with Downstream Evaluation and a Noise Floor

ResearchDGX agent

arXiv:2605.20798v1 Announce Type: cross Abstract: Narang et al. (2021) evaluated 40+ Transformer modifications at T5-base scale and concluded that most did not transfer. Five years later, the typical

Motion-Robust Deep Reconstruction for Free-Breathing Cardiac Cine MRI

ResearchDGX agent

arXiv:2605.20687v1 Announce Type: cross Abstract: Conventional cardiac cine MRI relies on breath-hold Cartesian acquisitions, which are vulnerable to motion artifacts and can be uncomfortable or infea

Multi-Channel Replay Speech Detection using Acoustic Maps

ResearchDGX agent

arXiv:2602.16399v2 Announce Type: replace-cross Abstract: Replay attacks remain a critical vulnerability for automatic speaker verification systems, particularly in real-time voice assistant applicati

Musical Attention Transformer: Music Generation Using a Music-Specific Attention Model

ResearchDGX agent

arXiv:2605.21081v1 Announce Type: cross Abstract: This study aims to enhance the quality of music generation using Transformers by incorporating meta-information. While Transformer-based approaches ar

NeighborDiv: Training-free Zero-shot Generalist Graph Anomaly Detection via Neighbor Diversity

ResearchDGX agent

arXiv:2605.20879v1 Announce Type: new Abstract: Graph Anomaly Detection (GAD) is increasingly shifting to Generalist GAD (GGAD) for cross-domain 'one-for-all' detection, but existing GGAD methods pred

Neural Estimation of Pairwise Mutual Information in Masked Discrete Sequence Models

ResearchDGX agent

arXiv:2605.20187v1 Announce Type: new Abstract: Understanding dependencies between variables is critical for interpretability and efficient generation in masked diffusion models (MDMs), yet these mode

New VIDEO: From LLM Wikis to LLM Artifacts Shared all my thoughts on why LLM wikis and HTML artifacts are a big deal. Plus, new tools to hel…

ResearchDGX agent

New VIDEO: From LLM Wikis to LLM Artifacts Shared all my thoughts on why LLM wikis and HTML artifacts are a big deal. Plus, new tools to help you build wikis and artifacts with agents. Just getting st

Nonlocal operator learning for fMRI encoding and decoding tasks

ResearchDGX agent

arXiv:2605.20389v1 Announce Type: new Abstract: Functional MRI data exhibit high-dimensional spatiotemporal structure, making both prediction and decoding challenging. In this work, we investigate neu

OCTOPUS: Optimized KV Cache for Transformers via Octahedral Parametrization Under optimal Squared error quantization

ResearchDGX agent

arXiv:2605.21226v1 Announce Type: new Abstract: The key-value (KV) cache dominates memory bandwidth and footprint in long-context autoregressive inference. Recent rotation-preconditioned codecs (Turbo

On the Cost and Benefit of Chain of Thought: A Learning-Theoretic Perspective

ResearchDGX agent

arXiv:2605.21260v1 Announce Type: new Abstract: We develop a learning-theoretic framework for understanding Chain of Thought (CoT). We model CoT as the interaction between an answer map and a chain ru

One Operator to Rule Them All? On Boundary-Indexed Operator Families in Neural PDE Solvers

ResearchDGX agent

arXiv:2603.01406v2 Announce Type: replace Abstract: Neural PDE solvers are often described as learning solution operators that map problem data to PDE solutions. In this work, we argue that this inter

OpenSeisML: Open Large-Scale Real Seismic and well-log Dataset for Generative AI

ResearchDGX agent

arXiv:2605.20539v1 Announce Type: new Abstract: The advent of machine learning (ML) and computer vision has significantly accelerated seismic inversion workflows by reducing the computational cost of

Optimal Query Allocation in Extractive QA with LLMs: A Learning-to-Defer Framework with Theoretical Guarantees

ResearchDGX agent

arXiv:2410.15761v4 Announce Type: replace Abstract: Large Language Models excel in generative tasks but exhibit inefficiencies in structured text selection, particularly in extractive question answeri

Oracle Supervision Transfers for Hyperparameter Prediction in Model-Based Image Denoising

ResearchDGX agent

arXiv:2605.20479v1 Announce Type: new Abstract: Hyperparameter prediction is a critical practical bottleneck for model-based image denoisers, ranging from classical TV/TGV variational solvers to moder

Ordering Matters: Rank-Aware Selective Fusion for Blended Emotion Recognition

ResearchDGX agent

arXiv:2605.21417v1 Announce Type: new Abstract: Blended emotion recognition is challenging because emotions are often expressed as mixtures of subtle and overlapping multimodal cues rather than a sing

PerpetualWonder: Long-Horizon Action-Conditioned 4D Scene Generation

ResearchDGX agent

arXiv:2602.04876v2 Announce Type: replace Abstract: We introduce PerpetualWonder, a hybrid generative simulator that enables long-horizon, action-conditioned 4D scene generation from a single image. C

Physics-informed convolutional neural networks for fluid flow through porous media

ResearchDGX agent

arXiv:2605.20250v1 Announce Type: new Abstract: Accurate simulation of fluid flow in porous media is challenging due to complex pore-space geometries and the computational cost of solving the Navier-S

Platonic Representations in the Human Brain: Unsupervised Recovery of Universal Geometry

ResearchDGX agent

arXiv:2605.20496v1 Announce Type: cross Abstract: The Strong Platonic Representation Hypothesis suggests that representational convergence in artificial neural networks can be harnessed constructively

Playing Devil's Advocate: Off-the-Shelf Persona Vectors Rival Targeted Steering for Sycophancy

ResearchDGX agent

arXiv:2605.21006v1 Announce Type: cross Abstract: We study the effect of different persona on extbf{sycophancy}: model's agreement with users even when the user is incorrect. The standard mitigation,

Plug-and-Play Spiking Operators: Breaking the Nonlinearity Bottleneck in Spiking Transformers

ResearchDGX agent

arXiv:2605.20289v1 Announce Type: new Abstract: ANN-to-SNN conversion offers a practical, training-free route to spiking large language models. However, current pipelines primarily focus on spike-driv

Polynomial-Time Robust Multiclass Linear Classification under Gaussian Marginals

ResearchDGX agent

arXiv:2605.21428v1 Announce Type: new Abstract: We study the task of agnostic learning of multiclass linear classifiers under the Gaussian distribution. Given labeled examples (x, y) from a distributi

Praxium: Diagnosing Cloud Anomalies with AI-based Telemetry and Dependency Analysis

ResearchDGX agent

arXiv:2603.23890v2 Announce Type: replace-cross Abstract: As the modern microservice architecture for cloud applications grows in popularity, cloud services are becoming increasingly complex and more

PREF: Phasorial Embedding Fields for Compact Neural Representations

ResearchDGX agent

arXiv:2205.13524v4 Announce Type: replace Abstract: We present an efficient frequency-based neural representation termed PREF: a shallow MLP augmented with a phasor volume that covers significant bord

PrefixWall: Mitigating Prefix Caching Side Channels in Shared LLM Systems

ResearchDGX agent

arXiv:2603.10726v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) rely on optimizations like Automatic Prefix Caching (APC) to accelerate inference. APC works by reusing previousl

Prism: Structural Symmetry Scanning via Duality-Constrained Laplacian Projection

ResearchDGX agent

arXiv:2605.20245v1 Announce Type: cross Abstract: We introduce extbf{Prism}, a framework for structural symmetry diagnosis in complex networks. Given a graph Laplacian L and a duality operator P (a sy

Prompt Reinjection: Alleviating Prompt Forgetting in Multimodal Diffusion Transformers

ResearchDGX agent

arXiv:2602.06886v3 Announce Type: replace Abstract: Multimodal Diffusion Transformers (MMDiTs) for text-to-image generation maintain separate text and image branches, with bidirectional information fl

ProtoPathway: Biologically Structured Prototype-Pathway Fusion for Multimodal Cancer Survival Prediction

ResearchDGX agent

arXiv:2605.21454v1 Announce Type: new Abstract: We introduce ProtoPathway, an interpretable-by-design multimodal framework for cancer survival prediction that unifies whole slide imaging and transcrip

Provably Learning Diffusion Models under the Manifold Hypothesis: Collapse and Refine

ResearchDGX agent

arXiv:2605.20235v1 Announce Type: new Abstract: Diffusion models generate high-dimensional data with remarkable quality, yet how their training efficiently learns the score function, bypassing the cur

Q-ARVD: Quantizing Autoregressive Video Diffusion Models

ResearchDGX agent

arXiv:2605.21072v1 Announce Type: new Abstract: Autoregressive video diffusion models (ARVDs) have emerged as a promising architecture for streaming video generation, paving the way for real-time inte

Q-SYNTH: Hybrid Quantum-Classical Adversarial Augmentation for Imbalanced Fraud Detection

ResearchDGX agent

arXiv:2605.21164v1 Announce Type: new Abstract: Credit card fraud detection is fundamentally challenged by extreme class imbalance, where fraudulent transactions are rare yet operationally critical. T

Quantifying the cross-linguistic effects of syncretism on agreement attraction

ResearchDGX agent

arXiv:2605.21403v1 Announce Type: new Abstract: Agreement attraction errors, in which a verb erroneously agrees with an intervening noun rather than its grammatical head, are amplified by morphologica

Reasoning-Trace Collapse: Evaluating the Loss of Explicit Reasoning During Fine-Tuning

ResearchDGX agent

arXiv:2605.21127v1 Announce Type: new Abstract: Explicit reasoning models are trained to produce intermediate reasoning traces before final answers, but downstream fine-tuning is often performed on or

Reinforcement Learning-based Control via Y-wise Affine Neural Networks: Comparative Case Studies for Chemical Processes

ResearchDGX agent

arXiv:2605.21211v1 Announce Type: cross Abstract: In this work we present an efficient and practically implementable approach for the application of reinforcement learning (RL)-based control in chemic

Reliable Automated Triage in Spanish Clinical Notes: A Hybrid Framework for Risk-Aware HIV Suspicion Identification

ResearchDGX agent

arXiv:2605.21256v1 Announce Type: new Abstract: Standard clinical Natural Language Processing (NLP) benchmarks often yield inflated metrics by forcing deterministic classification on ambiguous instanc

RelWitness: Open-Vocabulary 3D Scene Graph Generation with Visual-Geometric Relation Witnesses

ResearchDGX agent

arXiv:2605.20823v1 Announce Type: new Abstract: Open-vocabulary 3D scene graph generation seeks to describe object instances and their relations with flexible natural-language predicates. The central

ReMATF: Recurrent Motion-Adaptive Multi-scale Turbulence Mitigation for Dynamic Scenes

ResearchDGX agent

arXiv:2605.21440v1 Announce Type: new Abstract: Atmospheric turbulence severely degrades video quality by introducing distortions such as geometric warping, blur, and temporal flickering, posing signi

← Previous
1…227228229230231…432
Next →