AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,881 results
17 Apr 2026

XMark: Reliable Multi-Bit Watermarking for LLM-Generated Texts

ResearchDGX agent

arXiv:2604.05242v2 Announce Type: replace Abstract: Multi-bit watermarking has emerged as a promising solution for embedding imperceptible binary messages into Large Language Model (LLM)-generated tex

Zero-Ablation Overstates Register Content Dependence in DINO Vision Transformers

ResearchDGX agent

arXiv:2604.14433v1 Announce Type: new Abstract: Zero-ablation -- replacing token activations with zero vectors -- is widely used to probe token function in vision transformers. Register zeroing in DIN

Zeroth-Order Optimization at the Edge of Stability

ResearchDGX agent

arXiv:2604.14669v1 Announce Type: new Abstract: Zeroth-order (ZO) methods are widely used when gradients are unavailable or prohibitively expensive, including black-box learning and memory-efficient f

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
16 Apr 2026

3DRealHead: Few-Shot Detailed Head Avatar

ResearchDGX agent

arXiv:2604.13171v1 Announce Type: new Abstract: The human face is central to communication. For immersive applications, the digital presence of a person should mirror the physical reality, capturing t

A 3D SAM-Based Progressive Prompting Framework for Multi-Task Segmentation of Radiotherapy-induced Normal Tissue Injuries in Limited-Data Settings

ResearchDGX agent

arXiv:2604.13367v1 Announce Type: new Abstract: Radiotherapy-induced normal tissue injury is a clinically important complication, and accurate segmentation of injury regions from medical images could

A closer look at how large language models trust humans: patterns and biases

ResearchDGX agent

arXiv:2504.15801v2 Announce Type: replace Abstract: As large language models (LLMs) and LLM-based agents increasingly interact with humans in decision-making contexts, understanding the trust dynamics

A Function-Centric Perspective on Flat and Sharp Minima

ResearchDGX agent

arXiv:2510.12451v2 Announce Type: replace-cross Abstract: Flat minima are strongly associated with improved generalisation in deep neural networks. However, this connection has proven nuanced in recen

A Multi-Stage Optimization Pipeline for Bethesda Cell Detection in Pap Smear Cytology

ResearchDGX agent

arXiv:2604.13939v1 Announce Type: new Abstract: Computer vision techniques have advanced significantly in recent years, finding diverse and impactful applications within the medical field. In this pap

A Multimodal Clinically Informed Coarse-to-Fine Framework for Longitudinal CT Registration in Proton Therapy

ResearchDGX agent

arXiv:2604.13397v1 Announce Type: new Abstract: Proton therapy offers superior organ-at-risk sparing but is highly sensitive to anatomical changes, making accurate deformable image registration (DIR)

A Resource-Efficient Hybrid CNN-LSTM network for image-based bean leaf disease classification

ResearchDGX agent

arXiv:2604.13835v1 Announce Type: new Abstract: Accurate and resource-efficient automated diagnosis is a cornerstone of modern agricultural expert systems. While Convolutional Neural Networks (CNNs) h

A short proof of near-linear convergence of adaptive gradient descent under fourth-order growth and convexity

ResearchDGX agent

arXiv:2604.13393v1 Announce Type: cross Abstract: Davis, Drusvyatskiy, and Jiang showed that gradient descent with an adaptive stepsize converges locally at a nearly-linear rate for smooth functions t

A startup called Sabi just came out of stealth with a beanie that reads your thoughts. 70,000 to 100,000 miniature EEG sensors woven into th…

IndustryDGX agent

A startup called Sabi just came out of stealth with a beanie that reads your thoughts. 70,000 to 100,000 miniature EEG sensors woven into the fabric. You put it on like a winter hat and type by imagin

A transformable slender microrobot inspired by nematode parasites for interventional endovascular surgery

ResearchDGX agent

arXiv:2604.13513v1 Announce Type: new Abstract: Cardiovascular diseases account for around 17.9 million deaths per year globally, the treatment of which is challenging considering the confined space a

A Unified Conditional Flow for Motion Generation, Editing, and Intra-Structural Retargeting

ResearchDGX agent

arXiv:2604.13427v1 Announce Type: cross Abstract: Text-driven motion editing and intra-structural retargeting, where source and target share topology but may differ in bone lengths, are traditionally

A1: A Fully Transparent Open-Source, Adaptive and Efficient Truncated Vision-Language-Action Model

ResearchDGX agent

arXiv:2604.05672v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have emerged as a powerful paradigm for open-world robot manipulation, but their practical deployment is often c

Adaptive Conformal Prediction for Improving Factuality of Generations by Large Language Models

ResearchDGX agent

arXiv:2604.13991v1 Announce Type: new Abstract: Large language models (LLMs) are prone to generating factually incorrect outputs. Recent work has applied conformal prediction to provide uncertainty es

Adaptive Learning via Off-Model Training and Importance Sampling for Fully Non-Markovian Optimal Stochastic Control. Complete version

ResearchDGX agent

arXiv:2604.13147v1 Announce Type: cross Abstract: This paper studies continuous-time stochastic control problems whose controlled states are fully non-Markovian and depend on unknown model parameters.

AIM-CoT: Active Information-driven Multimodal Chain-of-Thought for Vision-Language Reasoning

ResearchDGX agent

arXiv:2509.25699v2 Announce Type: replace Abstract: Interleaved-Modal Chain-of-Thought (I-MCoT) advances vision-language reasoning, such as Visual Question Answering (VQA). This paradigm integrates sp

Any3DAvatar: Fast and High-Quality Full-Head 3D Avatar Reconstruction from Single Portrait Image

ResearchDGX agent

arXiv:2604.13856v1 Announce Type: new Abstract: Reconstructing a complete 3D head from a single portrait remains challenging because existing methods still face a sharp quality-speed trade-off: high-f

Artificial intelligence application in lymphoma diagnosis with Vision Transformer using weakly supervised training

ResearchDGX agent

arXiv:2604.13795v1 Announce Type: new Abstract: Vision transformers (ViT) have been shown to allow for more flexible feature detection and can outperform convolutional neural network (CNN) when pre-tr

Automated co-design of high-performance thermodynamic cycles via graph-based hierarchical reinforcement learning

ResearchDGX agent

arXiv:2604.13133v1 Announce Type: new Abstract: Thermodynamic cycles are pivotal in determining the efficacy of energy conversion systems. Traditional design methodologies, which rely on expert knowle

Better and Worse with Scale: How Contextual Entrainment Diverges with Model Size

ResearchDGX agent

arXiv:2604.13275v1 Announce Type: new Abstract: Larger language models become simultaneously better and worse at handling contextual information -- better at ignoring false claims, worse at ignoring i

Binomial Gradient-Based Meta-Learning for Enhanced Meta-Gradient Estimation

ResearchDGX agent

arXiv:2604.13263v1 Announce Type: new Abstract: Meta-learning offers a principled framework leveraging task-invariant priors from related tasks, with which task-specific models can be fine-tuned on do

BOAT: Navigating the Sea of In Silico Predictors for Antibody Design via Multi-Objective Bayesian Optimization

ResearchDGX agent

arXiv:2604.13980v1 Announce Type: new Abstract: Antibody lead optimization is inherently a multi-objective challenge in drug discovery. Achieving a balance between different drug-like properties is cr

Bridging Compositional and Distributional Semantics: A Survey on Latent Semantic Geometry via AutoEncoder

ResearchDGX agent

arXiv:2506.20083v4 Announce Type: replace Abstract: Integrating compositional and symbolic properties into current distributional semantic spaces can enhance the interpretability, controllability, com

C-voting: Confidence-Based Test-Time Voting without Explicit Energy Functions

ResearchDGX agent

arXiv:2604.13521v1 Announce Type: new Abstract: Neural network models with latent recurrent processing, where identical layers are recursively applied to the latent state, have gained attention as pro

C2: Scalable Rubric-Augmented Reward Modeling from Binary Preferences

ResearchDGX agent

arXiv:2604.13618v1 Announce Type: new Abstract: Rubric-augmented verification guides reward models with explicit evaluation criteria, yielding more reliable judgments than single-model verification. H

Calibrated Speculative Decoding: Frequency-Guided Candidate Selection for Efficient Inference

ResearchDGX agent

arXiv:2604.13634v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by letting draft tokens bypass full verification, but conventional frameworks suffer from fre

Can Cross-Layer Transcoders Replace Vision Transformer Activations? An Interpretable Perspective on Vision

ResearchDGX agent

arXiv:2604.13304v1 Announce Type: new Abstract: Understanding the internal activations of Vision Transformers (ViTs) is critical for building interpretable and trustworthy models. While Sparse Autoenc

Caption First, VQA Second: Knowledge Density, Not Task Format, Drives Multimodal Scaling

ResearchDGX agent

arXiv:2604.13054v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved rapid progress, yet their scaling behavior remains less clearly characterized and often less pred

Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs

ResearchDGX agent

arXiv:2604.13950v1 Announce Type: new Abstract: We show how causal interventions in Transformer models provide insights into English syntax by focusing on a long-standing challenge for syntactic theor

ClipGStream: Clip-Stream Gaussian Splatting for Any Length and Any Motion Multi-View Dynamic Scene Reconstruction

ResearchDGX agent

arXiv:2604.13746v1 Announce Type: new Abstract: Dynamic 3D scene reconstruction is essential for immersive media such as VR, MR, and XR, yet remains challenging for long multi-view sequences with larg

Complex Interpolation of Matrices with an application to Multi-Manifold Learning

ResearchDGX agent

arXiv:2604.14118v1 Announce Type: new Abstract: Given two symmetric positive-definite matrices A, B in R^{n imes n}, we study the spectral properties of the interpolation A^{1-x} B^x for 0 leq x leq 1

Computational framework for multistep metabolic pathway design

ResearchDGX agent

arXiv:2604.13471v1 Announce Type: new Abstract: In silico tools are important for generating novel hypotheses and exploring alternatives in de novo metabolic pathway design. However, while many comput

Concrete Jungle: Towards Concreteness Paved Contrastive Negative Mining for Compositional Understanding

ResearchDGX agent

arXiv:2604.13313v1 Announce Type: new Abstract: Vision-Language Models demonstrate remarkable capabilities but often struggle with compositional reasoning, exhibiting vulnerabilities regarding word or

Convex Hulls of Reachable Sets

ResearchDGX agent

arXiv:2303.17674v5 Announce Type: replace-cross Abstract: We study the convex hulls of reachable sets of nonlinear systems with bounded disturbances and uncertain initial conditions. Reachable sets pl

Creo: From One-Shot Image Generation to Progressive, Co-Creative Ideation

ResearchDGX agent

arXiv:2604.13956v1 Announce Type: cross Abstract: Text-to-image (T2I) systems enable rapid generation of high-fidelity imagery but are misaligned with how visual ideas develop. T2I systems generate ou

Cyclic 2.5D Perceptual Loss for Cross-Modal 3D Medical Image Synthesis: T1w MRI to Tau PET

ResearchDGX agent

arXiv:2406.12632v3 Announce Type: replace-cross Abstract: Positron emission tomography (PET) provides molecular biomarkers for Alzheimer's disease and related dementias (ADRD) and is increasingly used

Dataset-Level Metrics Attenuate Non-Determinism: A Fine-Grained Non-Determinism Evaluation in Diffusion Language Models

ResearchDGX agent

arXiv:2604.13413v1 Announce Type: new Abstract: Diffusion language models (DLMs) have emerged as a promising paradigm for large language models (LLMs), yet the non-deterministic behavior of DLMs remai

Deep Spatially-Regularized and Superpixel-Based Diffusion Learning for Unsupervised Hyperspectral Image Clustering

ResearchDGX agent

arXiv:2604.13307v1 Announce Type: new Abstract: An unsupervised framework for hyperspectral image (HSI) clustering is proposed that incorporates masked deep representation learning with diffusion-base

Dehaze-then-Splat: Generative Dehazing with Physics-Informed 3D Gaussian Splatting for Smoke-Free Novel View Synthesis

ResearchDGX agent

arXiv:2604.13589v1 Announce Type: new Abstract: We present Dehaze-then-Splat, a two-stage pipeline for multi-view smoke removal and novel view synthesis developed for Track~2 of the NTIRE 2026 3D Rest

Delineate Anything Flow: Fast, Country-Level Field Boundary Detection from Any Source

ResearchDGX agent

arXiv:2511.13417v2 Announce Type: replace Abstract: Accurate delineation of agricultural field boundaries from satellite imagery is essential for land management and crop monitoring, yet existing meth

Democratising Pathology Co-Pilots: An Open Pipeline and Dataset for Whole-Slide Vision-Language Modelling

ResearchDGX agent

arXiv:2512.17326v2 Announce Type: replace Abstract: Vision-language models (VLMs) have the potential to become co-pilots for pathologists. However, most VLMs either focus on small regions of interest

Depth-Resolved Coral Reef Thermal Fields from Satellite SST and Sparse In-Situ Loggers Using Physics-Informed Neural Networks

ResearchDGX agent

arXiv:2604.13131v1 Announce Type: cross Abstract: Satellite sea surface temperature (SST) products underpin global coral bleaching monitoring, yet they measure only the ocean skin. Corals inhabit dept

Design and Behavior of Sparse Mixture-of-Experts Layers in CNN-based Semantic Segmentation

ResearchDGX agent

arXiv:2604.13761v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) layers have been shown to substantially increase model capacity without a proportional increase in computational cost an

Design Conditions for Intra-Group Learning of Sequence-Level Rewards: Token Gradient Cancellation

ResearchDGX agent

arXiv:2604.13088v1 Announce Type: new Abstract: In sparse termination rewards, intra-group comparisons have become the dominant paradigm for fine-tuning reasoning models via reinforcement learning. Ho

Destroying a library brings the dark ages.

ResearchDGX agent

Destroying a library brings the dark ages. Destroying the @InternetArchive's @WayBackMachine would be the equivalent of the burning of the Library of Alexandria - one of the worst losses of knowledge

DiffMagicFace: Identity Consistent Facial Editing of Real Videos

ResearchDGX agent

arXiv:2604.13841v1 Announce Type: new Abstract: Text-conditioned image editing has greatly benefitted from the advancements in Image Diffusion Models. However, extending these techniques to facial vid

Diffusion Sequence Models for Generative In-Context Meta-Learning of Robot Dynamics

ResearchDGX agent

arXiv:2604.13366v1 Announce Type: new Abstract: Accurate modeling of robot dynamics is essential for model-based control, yet remains challenging under distributional shifts and real-time constraints.

Does Dimensionality Reduction via Random Projections Preserve Landscape Features?

ResearchDGX agent

arXiv:2604.13230v1 Announce Type: new Abstract: Exploratory Landscape Analysis (ELA) provides numerical features for characterizing black-box optimization problems. In high-dimensional settings, howev

Don't Let the Video Speak: Audio-Contrastive Preference Optimization for Audio-Visual Language Models

ResearchDGX agent

arXiv:2604.14129v1 Announce Type: new Abstract: While Audio-Visual Language Models (AVLMs) have achieved remarkable progress over recent years, their reliability is bottlenecked by cross-modal halluci

DRG-Font: Dynamic Reference-Guided Few-shot Font Generation via Contrastive Style-Content Disentanglement

ResearchDGX agent

arXiv:2604.13797v1 Announce Type: new Abstract: Few-shot Font Generation aims to generate stylistically consistent glyphs from a few reference glyphs. However, capturing complex font styles from a few

DroneScan-YOLO: Redundancy-Aware Lightweight Detection for Tiny Objects in UAV Imagery

ResearchDGX agent

arXiv:2604.13278v1 Announce Type: new Abstract: Aerial object detection in UAV imagery presents unique challenges due to the high prevalence of tiny objects, adverse environmental conditions, and stri

Dual-Enhancement Product Bundling: Bridging Interactive Graph and Large Language Model

ResearchDGX agent

arXiv:2604.14030v1 Announce Type: new Abstract: Product bundling boosts e-commerce revenue by recommending complementary item combinations. However, existing methods face two critical challenges: (1)

Empirical Evidence of Complexity-Induced Limits in Large Language Models on Finite Discrete State-Space Problems with Explicit Validity Constraints

ResearchDGX agent

arXiv:2604.13371v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly described as possessing strong reasoning capabilities, supported by high performance on mathematical, logi

English is Not All You Need: Systematically Exploring the Role of Multilinguality in LLM Post-Training

ResearchDGX agent

arXiv:2604.13286v1 Announce Type: new Abstract: Despite the widespread multilingual deployment of large language models, post-training pipelines remain predominantly English-centric, contributing to p

Enhancing Mixture-of-Experts Specialization via Cluster-Aware Upcycling

ResearchDGX agent

arXiv:2604.13508v1 Announce Type: new Abstract: Sparse Upcycling provides an efficient way to initialize a Mixture-of-Experts (MoE) model from pretrained dense weights instead of training from scratch

Event-Adaptive State Transition and Gated Fusion for RGB-Event Object Tracking

ResearchDGX agent

arXiv:2604.13426v1 Announce Type: new Abstract: Existing Vision Mamba-based RGB-Event(RGBE) tracking methods suffer from using static state transition matrices, which fail to adapt to variations in ev

Exploring Urban Land Use Patterns by Pattern Mining and Unsupervised Learning

ResearchDGX agent

arXiv:2604.13050v1 Announce Type: cross Abstract: Urban areas are intricate systems shaped by socioeconomic, environmental, and infrastructural factors, with land use patterns serving as aspects of ur

FAST: A Synergistic Framework of Attention and State-space Models for Spatiotemporal Traffic Prediction

ResearchDGX agent

arXiv:2604.13453v1 Announce Type: new Abstract: Traffic forecasting requires modeling complex temporal dynamics and long-range spatial dependencies over large sensor networks. Existing methods typical

← Previous
1…326327328329330…432
Next →