AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
19,194 results
5 Jun 2026

Narrative Knowledge Weaver: Narrative-Centric Retrieval-Augmented Reasoning for Long-Form Text Understanding

ResearchDGX agent

arXiv:2606.05724v1 Announce Type: new Abstract: Long-form narrative QA requires reasoning over evolving story worlds rather than isolated passages: answers may depend on earlier goals, changing charac

Next-Generation Parallel Decoder for LPDR: Architectural Optimization and Class-Balanced GAN-Augmentation

ResearchDGX agent

arXiv:2606.05785v1 Announce Type: new Abstract: Real-Time License Plate Detection and Recognition (LPDR) forms the backbone of modern smart cities. Although the YOLOV5-PDLPR model substantially improv

NIV: Neural Axis Variations for Variable Font Generation

ResearchDGX agent

arXiv:2606.05261v1 Announce Type: new Abstract: Variable fonts enable continuous variation of glyph geometry along semantic design axes such as weight, width, slant, and optical size. However, constru


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Noise-Adaptive Regularization for Robust Multi-Label Remote Sensing Image Classification

ResearchDGX agent

arXiv:2601.08446v2 Announce Type: replace Abstract: The development of reliable methods for multi-label classification (MLC) has become a prominent research direction in remote sensing (RS). As the sc

ORACLE-CT: Anatomy-Aware Support Pooling for CT Classification

ResearchDGX agent

arXiv:2606.05460v1 Announce Type: new Abstract: Abdominal CT disease classification is challenging because each scan is a large 3D volume with many possible findings, while diagnostic evidence is ofte

PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scene Understanding

ResearchDGX agent

arXiv:2606.06485v1 Announce Type: new Abstract: Recent advances in 3D multimodal large language models (3D-MLLMs) have enabled unified solutions for 3D scene understanding tasks, including visual ques

Parallel Jacobi Decoding for Fast Autoregressive Image Generation

ResearchDGX agent

arXiv:2606.05703v1 Announce Type: new Abstract: Autoregressive (AR) models have demonstrated remarkable performance in generating high-fidelity images. However, their inherently sequential next-token

PHUMA: Physically Reliable Humanoid Locomotion Dataset

ResearchDGX agent

arXiv:2510.26236v2 Announce Type: replace Abstract: Motion imitation is a promising approach for humanoid locomotion, enabling agents to acquire humanlike behaviors. Existing methods typically rely on

Physics in 2-Steps: Locking Motion Priors Before Visual Refinement Erases Them

ResearchDGX agent

arXiv:2606.06361v1 Announce Type: new Abstract: Image-to-Video diffusion models leverage input images to generate visually stunning content, yet frequently produce motion that violates physical laws.

Predictable Scaling Laws of Optimal Hyperparameters for LLM Continued Pre-training

ResearchDGX agent

arXiv:2606.05610v1 Announce Type: new Abstract: The efficacy of continued pre-training for Large Language Models (LLMs) hinges upon hyperparameter configurations, such as learning rate and batch size.

Preserving Full 6-DOF Actuation Under Abrupt Total Rotor Failures: Passive Fault-Tolerant Flight Control Using a Biaxial-Tilt Hexacopter

ResearchDGX agent

arXiv:2606.05663v1 Announce Type: new Abstract: Conventional multirotors suffer from a rapid collapse of attainable wrench space (AWS) under abrupt total rotor failures, rendering full 6-DOF recovery

RealDexUMI: A Wearable Universal Manipulation Interface for Dexterous Robot Learning

ResearchDGX agent

arXiv:2606.06033v1 Announce Type: new Abstract: Learning dexterous manipulation requires demonstrations that preserve fine hand-object interactions while remaining executable at deployment. Existing p

Really excited to be part of this journey and team 🚀 Ever since @SakanaAILabs inception, we have been discovering stepping stones for a fun…

ResearchDGX agent

Really excited to be part of this journey and team 🚀 Ever since @SakanaAILabs inception, we have been discovering stepping stones for a fundamental 'AI² paradigm shift': Leveraging AI systems to impro

Reinforcement Learning Elicits Contextual Learning of Unseen Language Translation

ResearchDGX agent

arXiv:2606.06428v1 Announce Type: new Abstract: Prior work has shown that large language models (LLMs) can translate unseen or low-resource languages by undergoing continued training or even by encodi

ReTreVal: Reasoning Tree with Validation and Cross-Problem Memory for Large Language Models

ResearchDGX agent

arXiv:2601.02880v2 Announce Type: replace-cross Abstract: Every existing inference-time reasoning framework discards all failure context at problem boundaries, leaving a model solving problem 500 no w

ReverseEOL: Improving Training-free Text Embeddings via Text Reversal in Decoder-only LLMs

ResearchDGX agent

arXiv:2606.05858v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have opened new avenues for generating training-free text embeddings. However, the causal attention in d

Revising Context, Shifting Simulated Stance: Auditing LLM-Based Stance Simulation in Online Discussions

ResearchDGX agent

arXiv:2606.06443v1 Announce Type: new Abstract: Large language models are increasingly used to simulate social media users and infer how individuals may respond to online discussions. However, it rema

RQUL-UIE: Revitalizing Quality-Unstable Labels for Underwater Image Enhancement via In-Dataset Self-Supervision

ResearchDGX agent

arXiv:2606.06176v1 Announce Type: new Abstract: Underwater Image Enhancement (UIE) is essential for mitigating degradations caused by water medium. Although learning-based methods have advanced signif

SAM-Flow: Source-Anchored Masked Flow for Training-Free Image Editing

ResearchDGX agent

arXiv:2606.06228v1 Announce Type: new Abstract: Training-free image editing has recently attracted increasing attention due to its ability to modify real images using powerful pre-trained diffusion an

SC-MFJ: A Simple Haptic Quality Metric for Medical Image Segmentation

ResearchDGX agent

arXiv:2606.06199v1 Announce Type: new Abstract: Standard segmentation metrics such as Dice and Hausdorff distance measure geometric overlap but say nothing about whether a segmented surface is suitabl

Self-Learning Expression Deformations for Data-Efficient Gaussian Avatars

ResearchDGX agent

arXiv:2606.05912v1 Announce Type: new Abstract: Modeling dynamic facial expressions using 3D Gaussian representations remains challenging due to their unstructured nature. Conventional Gaussian avatar

Self-supervised Feature Disentanglement and Augmentation Network for One-class Face Anti-spoofing

ResearchDGX agent

arXiv:2503.22929v3 Announce Type: replace Abstract: Face anti-spoofing (FAS) techniques aim to enhance the security of facial identity authentication by distinguishing authentic live faces from decept

Semi-Offline Reinforcement Learning for Optimized Text Generation

ResearchDGX agent

arXiv:2306.09712v2 Announce Type: replace-cross Abstract: In reinforcement learning (RL), there are two major settings for interacting with the environment: online and offline. Online methods explore

Some billionaires (or their foundations) do fund certain areas of basic research. Examples: Simons Foundation, Moore Foundation, Sloan Found…

ResearchDGX agent

Some billionaires (or their foundations) do fund certain areas of basic research. Examples: Simons Foundation, Moore Foundation, Sloan Foundation, Keck Foundation, Schmidt Sciences, and several others

SpanNorm: Reconciling Training Stability and Performance in Deep Transformers

ResearchDGX agent

arXiv:2601.22580v2 Announce Type: replace Abstract: The success of Large Language Models (LLMs) hinges on the stable training of deep Transformer architectures. A critical design choice is the placeme

Tamaththul3D: High-Fidelity 3D Saudi Sign Language Avatars from Monocular Video

ResearchDGX agent

arXiv:2605.05367v2 Announce Type: replace Abstract: Existing 3D sign language avatar reconstruction methods are developed and evaluated exclusively on Western sign languages, and no 3D parametric anno

The Invisible Hand of Physics: When Video Diffusion Models Know More Than They Show

ResearchDGX agent

arXiv:2606.05328v1 Announce Type: cross Abstract: Modern video diffusion models generate increasingly realistic and temporally coherent videos, motivating their use as candidate world simulators. Yet

The Tell-Tale Norm: ell_2 Magnitude as a Signal for Reasoning Dynamics in Large Language Models

ResearchDGX agent

arXiv:2606.06188v1 Announce Type: new Abstract: Recent work has sought to understand Large Language Models (LLMs) reasoning, yet a principled, model-intrinsic signal that captures its layer-wise reaso

Three-Dimensional Retinal Microvasculature Restoration in OCT Angiography

ResearchDGX agent

arXiv:2606.05375v1 Announce Type: new Abstract: Optical coherence tomographic angiography (OCTA) is a powerful technique for imaging retinal microvasculature. However, acquiring reliable quantificatio

To be clear: there should not be a president of AI (let alone me 😅). Which is pretty much one point I make in the video. This reminder is a…

ResearchDGX agent

To be clear: there should not be a president of AI (let alone me 😅). Which is pretty much one point I make in the video. This reminder is about the content of the video, which is 3 years old, but ever

Towards Truly Multilingual ASR: Generalizing Code-Switching ASR to Unseen Language Pairs

ResearchDGX agent

arXiv:2606.05846v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) has become a key technology for human--AI interaction. However, code-switching ASR (CS-ASR) remains particularly chal

UnHype: CLIP-Guided Hypernetworks for Dynamic LoRA Unlearning

ResearchDGX agent

arXiv:2602.03410v2 Announce Type: replace Abstract: Recent advances in large-scale diffusion models have intensified concerns about their potential misuse, particularly in generating realistic yet har

Unsupervised Monocular 3D Keypoint Discovery from Multi-View Diffusion Priors

ResearchDGX agent

arXiv:2507.12336v2 Announce Type: replace Abstract: Most existing 3D keypoint estimation methods rely on manual annotations or calibrated multi-view images, both of which are expensive to collect. Thi

USAD 2.0: Scaling Representation Distillation for Universal Audio Understanding

ResearchDGX agent

arXiv:2606.06444v1 Announce Type: cross Abstract: Audio encoders are critical to modern audio applications as large language models (LLMs) increasingly rely on a single encoder for diverse inputs. Whi

Vavanagi: a Community-run Platform for Documentation of the Hula Language in Papua New Guinea

ResearchDGX agent

arXiv:2603.14210v2 Announce Type: replace Abstract: We present Vavanagi, a community-run platform for Hula (Vula'a), an Austronesian language of Papua New Guinea with approximately 10,000 speakers. Va

Visual Commonsense Driven Knowledge Refinements for Scene Graph Generation

ResearchDGX agent

arXiv:2606.06369v1 Announce Type: new Abstract: Learning-driven Scene Graph Generation (SGG) models excel on frequent relation types but degrade sharply under annotation sparsity, failing to capture r

Wave Focusing in Metamaterials: Tactile Displays Beyond the Diffraction Limit

ResearchDGX agent

arXiv:2606.05572v1 Announce Type: cross Abstract: We address the challenge of engineering distributed haptic displays capable of reproducing multiple localized, independently addressable vibrations --

We had to make some deep level changes to Hermes Update command this morning. It may require a number of you to run hermes update twice in a…

ResearchDGX agent

We had to make some deep level changes to Hermes Update command this morning. It may require a number of you to run hermes update twice in a row (where you'll see an error the first time) Please run i

What is the fast Fourier transform?

ResearchDGX agent

The Fast Fourier Transform (FFT) is a computationally efficient algorithm that converts time-domain signals into their frequency-domain representation, reducing computational complexity from O(n²) to

What Makes Two Language Models Think Alike?

ResearchDGX agent

arXiv:2406.12620v3 Announce Type: replace Abstract: Do architectural and training differences influence the way models represent and process language? Traditional similarity metrics tell us whether tw

What Objects Enable, Not What They Are: Functional Latent Spaces for Affordance Reasoning

ResearchDGX agent

arXiv:2606.05533v1 Announce Type: cross Abstract: Existing robot planning systems rely on appearance-based reasoning, where visual observations are encoded into latent spaces organized around object a

When New Generators Arrive: Lifelong Machine-Generated Text Attribution via Ridge Feature Transfer

ResearchDGX agent

arXiv:2606.05626v1 Announce Type: new Abstract: Machine-generated text (MGT) attribution aims to identify the specific generator responsible for a given text, thereby providing fine-grained evidence f

Where does Absolute Position come from in decoder-only Transformers?

ResearchDGX agent

arXiv:2606.06160v1 Announce Type: cross Abstract: RoPE-trained transformers distinguish absolute position in their attention patterns, even though RoPE encodes only relative offsets in the inner produ

You Only Index Once: Cross-Layer Sparse Attention with Shared Routing

ResearchDGX agent

arXiv:2606.06467v1 Announce Type: new Abstract: Long-context inference in modern LLMs is increasingly constrained by decoding efficiency, especially in reasoning-heavy settings where models generate l

4 Jun 2026

3D Temporal Analysis for Autism Spectrum Disorder Screening During Attention Tasks

ResearchDGX agent

arXiv:2606.04836v1 Announce Type: new Abstract: Accurate Autism Spectrum Disorder (ASD) screening for school-age children is crucial to identify cases that may have been missed earlier and to enable t

A French Corpus Annotated for Multiword Expressions with Adverbial Function

ResearchDGX agent

arXiv:2606.04828v1 Announce Type: new Abstract: This paper presents a French corpus annotated for multiword expressions (MWEs) with adverbial function. This corpus is designed for investigation on inf

A General Framework for Dynamic Consistent Submodular Maximization

ResearchDGX agent

arXiv:2606.04946v1 Announce Type: cross Abstract: Consistency is an important property in dynamic submodular maximization and entails maintaining a near-optimal solution at all times, making only a sm

A Latent Variable Framework for Scaling Laws in Large Language Models

ResearchDGX agent

arXiv:2512.06553v2 Announce Type: replace-cross Abstract: We propose a statistical framework built on latent variable modeling for scaling laws of large language models (LLMs). Our work is motivated b

A Normative Intermediate Representation for ASP-Based Compliance Reasoning

ResearchDGX agent

arXiv:2606.04619v1 Announce Type: new Abstract: We propose MONIR, a Modalized-Output Normative Intermediate Representation for ASP-based compliance reasoning. Its core fragment has a staged operationa

A Systematic Analysis of Linguistic Features in AI-Generated Text Detection Across Domains and Models

ResearchDGX agent

arXiv:2606.04177v1 Announce Type: cross Abstract: Interpretable linguistic features offer a promising approach for explaining why a given text appears machine-generated, particularly for non-expert us

Abduction Prover in Isabelle/HOL

ResearchDGX agent

arXiv:2606.04877v1 Announce Type: cross Abstract: Proof assistants based on expressive logics suffer limited automation for proof search, raising the cost of formal verification based on proof assista

ACAT: A Collaborative Platform for Efficient Aspect-Based Sentiment Dataset Annotation

ResearchDGX agent

arXiv:2606.04189v1 Announce Type: new Abstract: Aspect-Based Sentiment Analysis (ABSA) requires high-quality datasets to train reliable models. However, existing annotation tools treat output as flat

Adalina: Adaptive Linear Approximation for the Shapley Value and Beyond

ResearchDGX agent

arXiv:2604.08438v2 Announce Type: replace Abstract: The Shapley value, and its broader family of semi-values, has received much attention in various attribution problems. A fundamental and long-standi

AI from concrete to abstract: demystifying artificial intelligence to the general public

ResearchDGX agent

arXiv:2006.04013v6 Announce Type: cross Abstract: Artificial Intelligence (AI) has been adopted in a wide range of domains. This shows the imperative need to develop means to endow common people with

AlphaQ: Calibration-Free Bit Allocation for Mixture-of-Experts Quantization

ResearchDGX agent

arXiv:2606.04980v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures scale model capacity through sparse expert activation, but their deployment remains memory-bound because all expe

Answer Self-Consistency with Margin-Triggered Question Re-Arbitration for the CVPR 2026 VidLLMs Challenge

ResearchDGX agent

arXiv:2606.04323v1 Announce Type: new Abstract: In this report, we present our solution for Track 2 of the CVPR 2026 VidLLMs Challenge. This track evaluates visual relational reasoning in videos, wher

AttnRegDeepLab: A Two-Stage Decoupled Framework for Interpretable Embryo Fragmentation Grading

ResearchDGX agent

arXiv:2511.18454v3 Announce Type: replace-cross Abstract: Embryo fragmentation is a morphological indicator critical for evaluating developmental potential in In Vitro Fertilization (IVF). However, ma

Audio Interaction Model

ResearchDGX agent

arXiv:2606.05121v1 Announce Type: cross Abstract: Audio is an inherently interactive modality, yet today's Large Audio Language Models (LALMs) are offline, and streaming audio models each handle only

Automated Lexical Coverage for Language Learning: From General to Specialized Word Lists

ResearchDGX agent

arXiv:2512.15552v2 Announce Type: replace Abstract: A General Service List (GSL) is a commonly used resource for language learners to identify important English words. Traditional GSL creation is reso

AutoNumerics-Zero: Automated Discovery of State-of-the-Art Mathematical Functions

ResearchDGX agent

arXiv:2312.08472v2 Announce Type: replace-cross Abstract: Transcendental functions, such as the exponential, are central to scientific computing, yet they cannot be natively calculated by digital hard

← Previous
1…133134135136137…320
Next →