AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,880 results
10 Apr 2026

Cross-Tokenizer LLM Distillation through a Byte-Level Interface

ResearchDGX agent

arXiv:2604.07466v1 Announce Type: new Abstract: Cross-tokenizer distillation (CTD), the transfer of knowledge from a teacher to a student language model when the two use different tokenizers, remains

CRPS-Optimal Binning for Univariate Conformal Regression

ResearchDGX agent

arXiv:2603.22000v2 Announce Type: replace Abstract: We propose a method for non-parametric conditional distribution estimation based on partitioning covariate-sorted observations into contiguous bins

Current LLMs still cannot 'talk much' about grammar modules: Evidence from syntax

ResearchDGX agent

arXiv:2603.20114v4 Announce Type: replace Abstract: We aim to examine the extent to which Large Language Models (LLMs) can 'talk much' about grammar modules, providing evidence from syntax core proper

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DailyArt: Discovering Articulation from Single Static Images via Latent Dynamics

ResearchDGX agent

arXiv:2604.07758v1 Announce Type: new Abstract: Articulated objects are essential for embodied AI and world models, yet inferring their kinematics from a single closed-state image remains challenging

Database Querying under Missing Values Governed by Missingness Mechanisms

ResearchDGX agent

arXiv:2604.06520v1 Announce Type: cross Abstract: We address the problems of giving a semantics to- and doing query answering (QA) on a relational database (RDB) that has missing values (MVs). The cau

DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation

ResearchDGX agent

arXiv:2511.19365v2 Announce Type: replace-cross Abstract: Pixel diffusion aims to generate images directly in pixel space in an end-to-end fashion. This approach avoids the limitations of VAE in the t

Depression Detection at the Point of Care: Automated Analysis of Linguistic Signals from Routine Primary Care Encounters

ResearchDGX agent

arXiv:2604.06193v1 Announce Type: cross Abstract: Depression is underdiagnosed in primary care, yet timely identification remains critical. Recorded clinical encounters, increasingly common with digit

Development of ML model for triboelectric nanogenerator based sign language detection system

ResearchDGX agent

arXiv:2604.06220v1 Announce Type: cross Abstract: Sign language recognition (SLR) is vital for bridging communication gaps between deaf and hearing communities. Vision-based approaches suffer from occ

DHFP-PE: Dual-Precision Hybrid Floating Point Processing Element for AI Acceleration

ResearchDGX agent

arXiv:2604.04507v2 Announce Type: replace-cross Abstract: The rapid adoption of low-precision arithmetic in artificial intelligence and edge computing has created a strong demand for energy-efficient

DietDelta: A Vision-Language Approach for Dietary Assessment via Before-and-After Images

ResearchDGX agent

arXiv:2604.06352v1 Announce Type: cross Abstract: Accurate dietary assessment is critical for precision nutrition, yet most image-based methods rely on a single pre-consumption image and provide only

Differentially Private Language Generation and Identification in the Limit

ResearchDGX agent

arXiv:2604.08504v1 Announce Type: cross Abstract: We initiate the study of language generation in the limit, a model recently introduced by Kleinberg and Mullainathan [KM24], under the constraint of d

Diffusion Language Models Know the Answer Before Decoding

ResearchDGX agent

arXiv:2508.19982v5 Announce Type: replace Abstract: Diffusion language models (DLMs) have recently emerged as an alternative to autoregressive approaches, offering parallel sequence generation and fle

DiffVC: A Non-autoregressive Framework Based on Diffusion Model for Video Captioning

ResearchDGX agent

arXiv:2604.08084v1 Announce Type: new Abstract: Current video captioning methods usually use an encoder-decoder structure to generate text autoregressively. However, autoregressive methods have inhere

DINO-QPM: Adapting Visual Foundation Models for Globally Interpretable Image Classification

ResearchDGX agent

arXiv:2604.07166v1 Announce Type: cross Abstract: Although visual foundation models like DINOv2 provide state-of-the-art performance as feature extractors, their complex, high-dimensional representati

DisCEdge: Distributed Context Management for Large Language Models at the Edge

ResearchDGX agent

arXiv:2511.22599v2 Announce Type: replace-cross Abstract: Deploying Large Language Model (LLM) services at the edge benefits latency-sensitive and privacy-aware applications. However, the stateless na

Distilling Specialized Orders for Visual Generation

ResearchDGX agent

arXiv:2504.17069v2 Announce Type: replace Abstract: Autoregressive (AR) image generators are becoming increasingly popular due to their ability to produce high-quality images and their scalability. Ty

DIVERSED: Relaxed Speculative Decoding via Dynamic Ensemble Verification

ResearchDGX agent

arXiv:2604.07622v1 Announce Type: new Abstract: Speculative decoding is an effective technique for accelerating large language model inference by drafting multiple tokens in parallel. In practice, its

DMin: Scalable Training Data Influence Estimation for Diffusion Models

ResearchDGX agent

arXiv:2412.08637v4 Announce Type: replace Abstract: Identifying the training data samples that most influence a generated image is a critical task in understanding diffusion models (DMs), yet existing

Do We Need Distinct Representations for Every Speech Token? Unveiling and Exploiting Redundancy in Large Speech Language Models

ResearchDGX agent

arXiv:2604.06871v1 Announce Type: cross Abstract: Large Speech Language Models (LSLMs) typically operate at high token rates (tokens/s) to ensure acoustic fidelity, yet this results in sequence length

'Don't Do That!': Guiding Embodied Systems through Large Language Model-based Constraint Generation

ResearchDGX agent

arXiv:2506.04500v3 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have spurred interest in robotic navigation that incorporates complex spatial, mathematica

Double: Breaking the Acceleration Limit via Double Retrieval Speculative Parallelism

ResearchDGX agent

arXiv:2601.05524v2 Announce Type: replace Abstract: Parallel Speculative Decoding (PSD) accelerates traditional Speculative Decoding (SD) by overlapping draft generation with verification. However, it

DP-DeGauss: Dynamic Probabilistic Gaussian Decomposition for Egocentric 4D Scene Reconstruction

ResearchDGX agent

arXiv:2604.07986v1 Announce Type: new Abstract: Egocentric video is crucial for next-generation 4D scene reconstruction, with applications in AR/VR and embodied AI. However, reconstructing dynamic fir

Drifting Fields are not Conservative

ResearchDGX agent

arXiv:2604.06333v1 Announce Type: new Abstract: Drifting models generate high-quality samples in a single forward pass by transporting generated samples toward the data distribution using a vector val

DYCP: Dynamic Context Pruning for Long-Form Dialogue with LLMs

ResearchDGX agent

arXiv:2601.07994v4 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly operate over long-form dialogues with frequent topic shifts. While recent LLMs support extended context wi

E-3DPSM: A State Machine for Event-Based Egocentric 3D Human Pose Estimation

ResearchDGX agent

arXiv:2604.08543v1 Announce Type: new Abstract: Event cameras offer multiple advantages in monocular egocentric 3D human pose estimation from head-mounted devices, such as millisecond temporal resolut

ECLipsE-Gen-Local: Efficient Compositional Local Lipschitz Estimates for Deep Neural Networks

ResearchDGX agent

arXiv:2510.05261v2 Announce Type: replace Abstract: The Lipschitz constant is a key measure for certifying the robustness of neural networks to input perturbations. However, computing the exact consta

Ecological Legacies of Pre-Columbian Settlements Evident in Palm Clusters of Neotropical Mountain Forests

ResearchDGX agent

arXiv:2507.06949v3 Announce Type: replace Abstract: Ancient populations inhabited and transformed neotropical forests, yet the spatial extent of their ecological influence remains underexplored at hig

EEG2Vision: A Multimodal EEG-Based Framework for 2D Visual Reconstruction in Cognitive Neuroscience

ResearchDGX agent

arXiv:2604.08063v1 Announce Type: new Abstract: Reconstructing visual stimuli from non-invasive electroencephalography (EEG) remains challenging due to its low spatial resolution and high noise, parti

Efficient Learned Data Compression via Dual-Stream Feature Decoupling

ResearchDGX agent

arXiv:2604.07239v1 Announce Type: cross Abstract: While Learned Data Compression (LDC) has achieved superior compression ratios, balancing precise probability modeling with system efficiency remains c

Efficient PRM Training Data Synthesis via Formal Verification

ResearchDGX agent

arXiv:2505.15960v3 Announce Type: replace Abstract: Process Reward Models (PRMs) have emerged as a promising approach for improving LLM reasoning capabilities by providing process supervision over rea

Efficient Provably Secure Linguistic Steganography via Range Coding

ResearchDGX agent

arXiv:2604.08052v1 Announce Type: new Abstract: Linguistic steganography involves embedding secret messages within seemingly innocuous texts to enable covert communication. Provable security, which is

Efficient Quantization of Mixture-of-Experts with Theoretical Generalization Guarantees

ResearchDGX agent

arXiv:2604.06515v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) allows scaling of language and vision models efficiently by activating only a small subset of experts per input. While

Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models

ResearchDGX agent

arXiv:2604.06893v1 Announce Type: cross Abstract: Deep convolutional neural networks achieve remarkable performance by exhaustively processing dense spatial feature maps, yet this brute-force strategy

Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models

ResearchDGX agent

arXiv:2604.08456v1 Announce Type: cross Abstract: Despite rapid progress, pretrained vision-language models still struggle when answers depend on tiny visual details or on combining clues spread acros

European leaders are learning there is no point trying to appease Trump. He has an insatiable appetite for flattery no matter how transparen…

ResearchDGX agent

European leaders are learning there is no point trying to appease Trump. He has an insatiable appetite for flattery no matter how transparently false. Like any bully, he sees weakness as an invitation

Evaluating In-Context Translation with Synchronous Context-Free Grammar Transduction

ResearchDGX agent

arXiv:2604.07320v1 Announce Type: cross Abstract: Low-resource languages pose a challenge for machine translation with large language models (LLMs), which require large amounts of training data. One p

Evaluating PQC KEMs, Combiners, and Cascade Encryption via Adaptive IND-CPA Testing Using Deep Learning

ResearchDGX agent

arXiv:2604.06942v1 Announce Type: cross Abstract: Ensuring ciphertext indistinguishability is fundamental to cryptographic security, but empirically validating this property in real implementations an

EviSnap: Faithful Evidence-Cited Explanations for Cold-Start Cross-Domain Recommendation

ResearchDGX agent

arXiv:2604.06172v1 Announce Type: cross Abstract: Cold-start cross-domain recommender (CDR) systems predict a user's preferences in a target domain using only their source-domain behavior, yet existin

EvoFlows: Evolutionary Edit-Based Flow-Matching for Protein Engineering

ResearchDGX agent

arXiv:2603.11703v2 Announce Type: replace Abstract: We introduce EvoFlows, a variable-length protein sequence-to-sequence modeling approach designed for protein engineering. Existing protein language

Faithful-First Reasoning, Planning, and Acting for Multimodal LLMs

ResearchDGX agent

arXiv:2511.08409v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) frequently suffer from unfaithfulness, generating reasoning chains that drift from visual evidence or contr

Fast Spatial Memory with Elastic Test-Time Training

ResearchDGX agent

arXiv:2604.07350v1 Announce Type: cross Abstract: Large Chunk Test-Time Training (LaCT) has shown strong performance on long-context 3D reconstruction, but its fully plastic inference-time updates rem

FBS: Modeling Native Parallel Reading inside a Transformer

ResearchDGX agent

arXiv:2601.21708v2 Announce Type: replace Abstract: Large language models (LLMs) excel across many tasks, yet inference is still dominated by strictly token-by-token autoregression. Existing accelerat

Flemme: A Flexible and Modular Learning Platform for Medical Images

ResearchDGX agent

arXiv:2408.09369v3 Announce Type: replace-cross Abstract: As the rapid development of computer vision and the emergence of powerful network backbones and architectures, the application of deep learnin

Formalizing building-up constructions of self-dual codes through isotropic lines in Lean

ResearchDGX agent

arXiv:2604.08485v1 Announce Type: cross Abstract: The purpose of this paper is two-fold. First we show that Kim's building-up construction of binary self-dual codes is equivalent to Chinburg-Zhang's H

From Classical Machine Learning to Tabular Foundation Models: An Empirical Investigation of Robustness and Scalability Under Class Imbalance in Emergency and Critical Care

ResearchDGX agent

arXiv:2512.21602v2 Announce Type: replace-cross Abstract: Millions of patients pass through emergency departments and intensive care units each year, where clinicians must make high-stakes decisions u

From Exposure to Internalization: Dual-Stream Calibration for In-context Clinical Reasoning

ResearchDGX agent

arXiv:2604.06262v1 Announce Type: cross Abstract: Contextual clinical reasoning demands robust inference grounded in complex, heterogeneous clinical records. While state-of-the-art fine-tuning, in-con

From Load Tests to Live Streams: Graph Embedding-Based Anomaly Detection in Microservice Architectures

ResearchDGX agent

arXiv:2604.06448v1 Announce Type: cross Abstract: Prime Video regularly conducts load tests to simulate the viewer traffic spikes seen during live events such as Thursday Night Football as well as vid

Gaussian Approximation for Asynchronous Q-learning

ResearchDGX agent

arXiv:2604.07323v1 Announce Type: cross Abstract: In this paper, we derive rates of convergence in the high-dimensional central limit theorem for Polyak-Ruppert averaged iterates generated by the asyn

GEAR: GEometry-motion Alternating Refinement for Articulated Object Modeling with Gaussian Splatting

ResearchDGX agent

arXiv:2604.07728v1 Announce Type: new Abstract: High-fidelity interactive digital assets are essential for embodied intelligence and robotic interaction, yet articulated objects remain challenging to

Generative 3D Gaussian Splatting for Arbitrary-ResolutionAtmospheric Downscaling and Forecasting

ResearchDGX agent

arXiv:2604.07928v1 Announce Type: new Abstract: While AI-based numerical weather prediction (NWP) enables rapid forecasting, generating high-resolution outputs remains computationally demanding due to

Generative Phomosaic with Structure-Aligned and Personalized Diffusion

ResearchDGX agent

arXiv:2604.06989v1 Announce Type: cross Abstract: We present the first generative approach to photomosaic creation. Traditional photomosaic methods rely on a large number of tile images and color-base

Great to see this on X. This is something I created during my PhD time 20 years ago.

ResearchDGX agent

Great to see this on X. This is something I created during my PhD time 20 years ago. This is a healing grid by Japanese artist Ryota Kanai. If you stare at the center, the irregularities start to heal

GroundingAnomaly: Spatially-Grounded Diffusion for Few-Shot Anomaly Synthesis

ResearchDGX agent

arXiv:2604.08301v1 Announce Type: new Abstract: The performance of visual anomaly inspection in industrial quality control is often constrained by the scarcity of real anomalous samples. Consequently,

Hallucination as output-boundary misclassification: a composite abstention architecture for language models

ResearchDGX agent

arXiv:2604.06195v1 Announce Type: cross Abstract: Large language models often produce unsupported claims. We frame this as a misclassification error at the output boundary, where internally generated

HCRE: LLM-based Hierarchical Classification for Cross-Document Relation Extraction with a Prediction-then-Verification Strategy

ResearchDGX agent

arXiv:2604.07937v1 Announce Type: new Abstract: Cross-document relation extraction (RE) aims to identify relations between the head and tail entities located in different documents. Existing approache

HOTFLoc++: End-to-End Hierarchical LiDAR Place Recognition, Re-Ranking, and 6-DoF Metric Localisation in Forests

ResearchDGX agent

arXiv:2511.09170v2 Announce Type: replace Abstract: This article presents HOTFLoc++, an end-to-end hierarchical framework for LiDAR place recognition, re-ranking, and 6-DoF metric localisation in fore

How Does Machine Learning Manage Complexity?

ResearchDGX agent

arXiv:2604.07233v1 Announce Type: new Abstract: We provide a computational complexity lens to understand the power of machine learning models, particularly their ability to model complex systems. Mach

HST-HGN: Heterogeneous Spatial-Temporal Hypergraph Networks with Bidirectional State Space Models for Global Fatigue Assessment

ResearchDGX agent

arXiv:2604.08435v1 Announce Type: new Abstract: It remains challenging to assess driver fatigue from untrimmed videos under constrained computational budgets, due to the difficulty of modeling long-ra

Hybrid ResNet-1D-BiGRU with Multi-Head Attention for Cyberattack Detection in Industrial IoT Environments

ResearchDGX agent

arXiv:2604.06481v1 Announce Type: cross Abstract: This study introduces a hybrid deep learning model for intrusion detection in Industrial IoT (IIoT) systems, combining ResNet-1D, BiGRU, and Multi-Hea

Image-Guided Geometric Stylization of 3D Meshes

ResearchDGX agent

arXiv:2604.07795v1 Announce Type: new Abstract: Recent generative models can create visually plausible 3D representations of objects. However, the generation process often allows for implicit control

← Previous
1…343344345346347…432
Next →