AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
19,014 results
17 Apr 2026

DySCO: Dynamic Attention-Scaling Decoding for Long-Context Language Models

ResearchDGX agent

arXiv:2602.22175v2 Announce Type: replace Abstract: Understanding and reasoning over long contexts is a crucial capability for language models (LMs). Although recent models support increasingly long c

Edge-preserving noise for diffusion models

ResearchDGX agent

arXiv:2410.01540v4 Announce Type: replace Abstract: Classical diffusion models typically rely on isotropic Gaussian noise, treating all regions uniformly and overlooking structural information importa

Efficient closed-form approaches for pose estimation using Sylvester forms

ResearchDGX agent

arXiv:2604.14747v1 Announce Type: new Abstract: Solving non-linear least-squares problem for pose estimation (rotation and translation) is often a time consuming yet fundamental problem in several rea


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Efficient Search of Implantable Adaptive Cells for Medical Image Segmentation

ResearchDGX agent

arXiv:2604.14849v1 Announce Type: new Abstract: Purpose: Adaptive skip modules can improve medical image segmentation, but searching for them is computationally costly. Implantable Adaptive Cells (IAC

ELMoE-3D: Leveraging Intrinsic Elasticity of MoE for Hybrid-Bonding-Enabled Self-Speculative Decoding in On-Premises Serving

ResearchDGX agent

arXiv:2604.14626v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have become the dominant architecture for large-scale language models, yet on-premises serving remains fundamentally mem

Enhancing Linguistic Competence of Language Models through Pre-training with Language Learning Tasks

ResearchDGX agent

arXiv:2601.03448v2 Announce Type: replace Abstract: Language models (LMs) are pre-trained on raw text datasets to generate text sequences token-by-token. While this approach facilitates the learning o

Enhancing LLM-Based Neural Network Generation: Few-Shot Prompting and Efficient Validation for Automated Architecture Design

ResearchDGX agent

arXiv:2512.24120v2 Announce Type: replace Abstract: Automated neural network architecture design remains a significant challenge in computer vision. Task diversity and computational constraints requir

Estimating the Diameter at Breast Height of Trees in a Forest from RGB

ResearchDGX agent

arXiv:2505.03093v3 Announce Type: replace Abstract: Forest inventories rely on accurate measurements of the diameter at breast height (DBH) for ecological monitoring, resource management, and carbon a

Evolving Beyond Snapshots: Harmonizing Structure and Sequence via Entity State Tuning for Temporal Knowledge Graph Forecasting

ResearchDGX agent

arXiv:2602.12389v3 Announce Type: replace-cross Abstract: Temporal knowledge graph (TKG) forecasting requires predicting future facts by jointly modeling structural dependencies within each snapshot a

Expert-Guided Class-Conditional Goodness-of-Fit Scores for Interpretable Classification with Informative Missingness: An Application to Seismic Monitoring

ResearchDGX agent

arXiv:2604.14809v1 Announce Type: cross Abstract: We study a classification problem with three key challenges: pervasive informative missingness, the integration of partial prior expert knowledge into

Explain the Flag: Contextualizing Hate Speech Beyond Censorship

ResearchDGX agent

arXiv:2604.14970v1 Announce Type: new Abstract: Hate, derogatory, and offensive speech remains a persistent challenge in online platforms and public discourse. While automated detection systems are wi

Exploring the flavor structure of leptons via diffusion models

ResearchDGX agent

arXiv:2503.21432v2 Announce Type: replace-cross Abstract: We propose a method to explore the flavor structure of leptons using diffusion models, which are known as one of generative artificial intelli

Expressivity of Transformers: A Tropical Geometry Perspective

ResearchDGX agent

arXiv:2604.14727v1 Announce Type: new Abstract: To quantify the geometric expressivity of transformers, we introduce a tropical geometry framework to characterize their exact spatial partitioning capa

Fabricator or dynamic translator?

ResearchDGX agent

arXiv:2604.15165v1 Announce Type: new Abstract: LLMs are proving to be adept at machine translation although due to their generative nature they may at times overgenerate in various ways. These overge

FADPNet: Frequency-Aware Dual-Path Network for Face Super-Resolution

ResearchDGX agent

arXiv:2506.14121v2 Announce Type: replace Abstract: Face super-resolution (FSR) under limited computational budgets remains challenging. Existing methods often treat all facial pixels equally, leading

Faithfulness Serum: Mitigating the Faithfulness Gap in Textual Explanations of LLM Decisions via Attribution Guidance

ResearchDGX agent

arXiv:2604.14325v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong performance and have revolutionized NLP, but their lack of explainability keeps them treated as black boxes,

Federated Breast Cancer Detection Enhanced by Synthetic Ultrasound Image Augmentation

ResearchDGX agent

arXiv:2506.23334v3 Announce Type: replace-cross Abstract: Federated learning enables collaborative training of deep learning models across institutions without sharing sensitive patient data. However,

Feedback Adaptation for Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2604.06647v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems are typically evaluated under static assumptions, despite being frequently corrected through user or ex

Find the Differences: Differential Morphing Attack Detection vs Face Recognition

ResearchDGX agent

arXiv:2604.14734v1 Announce Type: new Abstract: Morphing is a challenge to face recognition (FR) for which several morphing attack detection solutions have been proposed. We argue that face recognitio

Finetuning-Free Diffusion Model with Adaptive Constraint Guidance for Inorganic Crystal Structure Generation

ResearchDGX agent

arXiv:2604.13354v1 Announce Type: cross Abstract: The discovery of inorganic crystal structures with targeted properties is a significant challenge in materials science. Generative models, especially

Flow of Truth: Proactive Temporal Forensics for Image-to-Video Generation

ResearchDGX agent

arXiv:2604.15003v1 Announce Type: new Abstract: The rapid rise of image-to-video (I2V) generation enables realistic videos to be created from a single image but also brings new forensic demands. Unlik

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench

ResearchDGX agent

arXiv:2604.15037v1 Announce Type: cross Abstract: Recent advancements in LLM agents are gradually shifting from reactive, text-based paradigms toward proactive, multimodal interaction. However, existi

From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning

ResearchDGX agent

arXiv:2604.15244v1 Announce Type: new Abstract: Speculative decoding (SD) accelerates large language model inference by allowing a lightweight draft model to propose outputs that a stronger target mod

FSDETR: Frequency-Spatial Feature Enhancement for Small Object Detection

ResearchDGX agent

arXiv:2604.14884v1 Announce Type: new Abstract: Small object detection remains a significant challenge due to feature degradation from downsampling, mutual occlusion in dense clusters, and complex bac

Fundamental Limitations of Favorable Privacy-Utility Guarantees for DP-SGD

ResearchDGX agent

arXiv:2601.10237v2 Announce Type: replace Abstract: Differentially Private Stochastic Gradient Descent (DP-SGD) is the dominant paradigm for private training, but its fundamental limitations under wor

G-MIXER: Geodesic Mixup-based Implicit Semantic Expansion and Explicit Semantic Re-ranking for Zero-Shot Composed Image Retrieval

ResearchDGX agent

arXiv:2604.14710v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) aims to retrieve target images by integrating a reference image with a corresponding modification text. CIR requires join

Gating Enables Curvature: A Geometric Expressivity Gap in Attention

ResearchDGX agent

arXiv:2604.14702v1 Announce Type: new Abstract: Multiplicative gating is widely used in neural architectures and has recently been applied to attention layers to improve performance and training stabi

Gaussian Process Regression of Steering Vectors With Physics-Aware Deep Composite Kernels for Augmented Listening

ResearchDGX agent

arXiv:2509.02571v2 Announce Type: replace-cross Abstract: This paper investigates continuous representations of steering vectors over frequency and microphone/source positions for augmented listening

Generalization in LLM Problem Solving: The Case of the Shortest Path

ResearchDGX agent

arXiv:2604.15306v1 Announce Type: cross Abstract: Whether language models can systematically generalize remains actively debated. Yet empirical performance is jointly shaped by multiple factors such a

Generating Concept Lexicalizations via Dictionary-Based Cross-Lingual Sense Projection

ResearchDGX agent

arXiv:2604.14397v1 Announce Type: new Abstract: We study the task of automatically expanding WordNet-style lexical resources to new languages through sense generation. We generate senses by associatin

Generative Augmented Inference

ResearchDGX agent

arXiv:2604.14575v1 Announce Type: new Abstract: Data-driven operations management often relies on parameters estimated from costly human-generated labels. Recent advances in large language models (LLM

Generative Data Augmentation for Skeleton Action Recognition

ResearchDGX agent

arXiv:2604.14933v1 Announce Type: new Abstract: Skeleton-based human action recognition is a powerful approach for understanding human behaviour from pose data, but collecting large-scale, diverse, an

Generative Modeling of Complex-Valued Brain MRI Data

ResearchDGX agent

arXiv:2604.14800v1 Announce Type: cross Abstract: Objective. Standard Magnetic Resonance Imaging (MRI) reconstruction pipelines discard phase information captured during acquisition, despite evidence

Geometrically Consistent Multi-View Scene Generation from Freehand Sketches

ResearchDGX agent

arXiv:2604.14302v1 Announce Type: new Abstract: We tackle a new problem: generating geometrically consistent multi-view scenes from a single freehand sketch. Freehand sketches are the most geometrical

Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head Avatars

ResearchDGX agent

arXiv:2604.14541v1 Announce Type: new Abstract: We present a framework for explicit emotion control in feed-forward, single-image 3D head avatar reconstruction. Unlike existing pipelines where emotion

Grading the Unspoken: Evaluating Tacit Reasoning in Quantum Field Theory and String Theory with LLMs

ResearchDGX agent

arXiv:2604.14188v1 Announce Type: cross Abstract: Large language models have demonstrated impressive performance across many domains of mathematics and physics. One natural question is whether such mo

Graph-Based Alternatives to LLMs for Human Simulation

ResearchDGX agent

arXiv:2511.02135v2 Announce Type: replace Abstract: Large language models (LLMs) have become a popular approach for simulating human behaviors, yet it remains unclear if LLMs are necessary for all sim

Graph Theoretical Outlier Rejection for 4D Radar Registration in Feature-Poor Environments

ResearchDGX agent

arXiv:2604.14857v1 Announce Type: new Abstract: Automotive 4D imaging radar is well suited for operation in dusty and low-visibility environments, but scan registration remains challenging due to scan

GUI-Perturbed: Domain Randomization Reveals Systematic Brittleness in GUI Grounding Models

ResearchDGX agent

arXiv:2604.14262v1 Announce Type: new Abstract: GUI grounding models report over 85% accuracy on standard benchmarks, yet drop 27-56 percentage points when instructions require spatial reasoning rathe

HAMSA: Scanning-Free Vision State Space Models via SpectralPulseNet

ResearchDGX agent

arXiv:2604.14724v1 Announce Type: new Abstract: Vision State Space Models (SSMs) like Vim, VMamba, and SiMBA rely on complex scanning strategies to adapt sequential SSMs to process 2D images, introduc

Hierarchical Semantic Retrieval with Cobweb

ResearchDGX agent

arXiv:2510.02539v2 Announce Type: replace Abstract: Neural document retrieval often treats a corpus as a flat cloud of vectors scored at a single granularity, leaving corpus structure underused and ex

High Probability Guarantees for Random Reshuffling

ResearchDGX agent

arXiv:2311.11841v4 Announce Type: replace-cross Abstract: We consider the stochastic gradient method with random reshuffling (mathsf{RR}) for tackling smooth nonconvex optimization problems. mathsf{RR

High-Speed Full-Color HDR Imaging via Unwrapping Modulo-Encoded Spike Streams

ResearchDGX agent

arXiv:2604.14632v1 Announce Type: new Abstract: Conventional RGB-based high dynamic range (HDR) imaging faces a fundamental trade-off between motion artifacts in multi-exposure captures and irreversib

Hoi! -- A Multimodal Dataset for Force-Grounded, Cross-View Articulated Manipulation

ResearchDGX agent

arXiv:2512.04884v3 Announce Type: replace Abstract: We present a dataset for force-grounded, cross-view articulated manipulation that couples what is seen with what is done and what is felt during rea

How Retrieved Context Shapes Internal Representations in RAG

ResearchDGX agent

arXiv:2602.20091v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) enhances large language models (LLMs) by conditioning generation on retrieved external documents, but the effec

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds

ResearchDGX agent

arXiv:2604.14268v1 Announce Type: new Abstract: We introduce HY-World 2.0, a multi-modal world model framework that advances our prior project HY-World 1.0. HY-World 2.0 accommodates diverse input mod

'I don't think people are sufficiently prepared notwithstanding what happened in 2021, for the possibility that he will try to fuck with thi…

ResearchDGX agent

'I don't think people are sufficiently prepared notwithstanding what happened in 2021, for the possibility that he will try to fuck with this election. And he will. He's already basically telling us t

I don't understand why we automatically give people credit for 'sincere views.' Like who gives a fuck? If I have a sincere view that a trans…

ResearchDGX agent

I don't understand why we automatically give people credit for 'sincere views.' Like who gives a fuck? If I have a sincere view that a transdimensional vampire attack is imminent it doesn't make it sa

Identifying Information from Observations with Uncertainty and Novelty

ResearchDGX agent

arXiv:2501.09331v3 Announce Type: replace Abstract: A machine that learns a task from observations must encounter and process uncertainty and novelty, especially when it is to maintain performance whe

IMPACTX: improving model performance by appropriately constraining the training with teacher explanations

ResearchDGX agent

arXiv:2502.12222v2 Announce Type: replace Abstract: The eXplainable Artificial Intelligence (XAI) research predominantly concentrates to provide explainations about AI model decisions, especially Deep

Improved Multiscale Structural Mapping with Supervertex Vision Transformer for the Detection of Alzheimer's Disease Neurodegeneration

ResearchDGX agent

arXiv:2604.14837v1 Announce Type: new Abstract: Alzheimer's disease (AD) confirmation often relies on positron emission tomography (PET) or cerebrospinal fluid (CSF) analysis, which are costly and inv

Improving Prostate Gland Segmentation Using Transformer based Architectures

ResearchDGX agent

arXiv:2506.14844v2 Announce Type: replace-cross Abstract: Inter reader variability and cross site domain shift challenge the automatic segmentation of prostate anatomy using T2 weighted MRI images. Th

Improving Sparse Autoencoder with Dynamic Attention

ResearchDGX agent

arXiv:2604.14925v1 Announce Type: new Abstract: Recently, sparse autoencoders (SAEs) have emerged as a promising technique for interpreting activations in foundation models by disentangling features i

International Conference on Learning Representations (ICLR) 2026

ResearchDGX agent

ICLR 2026 is a major international conference on machine learning and deep learning where Apple's research team will present their latest work and innovations in representation learning. The conferenc

JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation

ResearchDGX agent

arXiv:2411.09209v5 Announce Type: replace Abstract: Audio-driven portrait animation has made significant advances with diffusion-based models, improving video quality and lipsync accuracy. However, th

Just Pass Twice: Efficient Token Classification with LLMs for Zero-Shot NER

ResearchDGX agent

arXiv:2604.05158v2 Announce Type: replace Abstract: Large language models encode extensive world knowledge valuable for zero-shot named entity recognition. However, their causal attention mechanism, w

Kernel Neural Operators (KNOs) for Scalable, Memory-efficient, Geometrically-flexible Operator Learning

ResearchDGX agent

arXiv:2407.00809v3 Announce Type: replace Abstract: This paper introduces the Kernel Neural Operator (KNO), a provably convergent operator-learning architecture that utilizes compositions of deep kern

KVNN: Learnable Multi-Kernel Volterra Neural Networks

ResearchDGX agent

arXiv:2604.15141v1 Announce Type: new Abstract: Higher-order learning is fundamentally rooted in exploiting compositional features. It clearly hinges on enriching the representation by more elaborate

Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis

ResearchDGX agent

arXiv:2506.00433v4 Announce Type: replace Abstract: High-resolution image synthesis remains a core challenge in generative modeling, particularly in balancing computational efficiency with the preserv

Lazy or Efficient? Towards Accessible Eye-Tracking Event Detection Using LLMs

ResearchDGX agent

arXiv:2604.13243v1 Announce Type: cross Abstract: Gaze event detection is fundamental to vision science, human-computer interaction, and applied analytics. However, current workflows often require spe

← Previous
1…289290291292293…317
Next →