AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
18,857 results
11 Aug 2026

Unveiling the Secret of AdaLN-Zero in Diffusion Transformer

ResearchDGX agent

arXiv:2608.09438v1 Announce Type: new Abstract: Diffusion transformer (DiT), a rapidly emerging architecture for image generation, has gained much attention. However, despite ongoing efforts to improv

UPolarSQ: Polar Representation Learning for Optic Disc and Peripapillary Atrophy Segmentation and Quantification in Fundus Photographs

ResearchDGX agent

arXiv:2608.08771v1 Announce Type: new Abstract: Myopia-induced posterior-pole remodeling is frequently accompanied by Optic Disc (OD) deformation and Peripapillary Atrophy (PPA), both of which provide

Variance reduction in lattice QCD observables via normalizing flows

ResearchDGX agent

arXiv:2603.02984v2 Announce Type: replace-cross Abstract: Normalizing flows can be used to construct unbiased, reduced-variance estimators for lattice field theory observables that are defined by a de


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

VeriForge: Mitigating Latent Knowledge Gaps in Narrative Drafting via Mixed-Initiative Scaffolding

ResearchDGX agent

arXiv:2608.09698v1 Announce Type: cross Abstract: Great fiction earns its verisimilitude through precise details, from how a longsword is gripped to pierce armor gaps to why a bleeding corpse cannot y

Vision Meets WiFi: Physics-Grounded Estimation of Volumetric Mechanical Properties

ResearchDGX agent

arXiv:2608.07726v1 Announce Type: new Abstract: Estimating volumetric mechanical properties, including Young's modulus, Poisson's ratio, and density at each voxel, is intrinsically ambiguous from visi

VisionSelector: End-to-End Learnable Visual Token Compression for Efficient Multimodal LLMs

ResearchDGX agent

arXiv:2510.16598v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) encounter significant computational and memory bottlenecks from the massive number of visual tokens generat

VOICE: A Vision-Omics Foundation Model Integrating Direct and Retrieval-Based Prediction of In-situ Single-Cell Gene Expression

ResearchDGX agent

arXiv:2608.08366v1 Announce Type: new Abstract: Spatial transcriptomics can resolve gene expression at single-cell resolution, but it is costly, limited to targeted panels of a few hundred to a few th

WA-SpecDec: World-Aware Speculative Decoding for Vision-Language-Action Models

ResearchDGX agent

arXiv:2608.08725v1 Announce Type: new Abstract: Vision-language-action (VLA) policies generate robot controls autoregressively, making closed-loop latency dominated by repeated target-model forward pa

Walk-on-Spheres Monte Carlo and deep neural network approximations of elliptic PDEs with drift and killing

ResearchDGX agent

arXiv:2608.09494v1 Announce Type: cross Abstract: In this paper we provide Monte Carlo and deep neural network approximations for stochastic representations of solutions to linear elliptic partial dif

Walking through Discussions: A Mobile Visual Analytics System for In-Situ Group Discussion Analysis

ResearchDGX agent

arXiv:2608.08617v1 Announce Type: new Abstract: Group discussion-based teaching is widely used to foster collaborative learning, yet teachers in physical classrooms often struggle to simultaneously mo

When Can Fraud Operations Authorize Automation? A Decision-Support Framework for Fresh Audit Evidence and Review Workload

ResearchDGX agent

arXiv:2608.08577v1 Announce Type: new Abstract: Fraud operations must allocate events among automatic approval, analyst review, and automatic blocking even though the labels needed to evaluate these a

When Confidence Fails: Overconfidence in LLMs under Uncertainty and Missing Clinical Information

ResearchDGX agent

arXiv:2608.09080v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved strong performance in medical question answering and clinical reasoning tasks. However, their reliability u

Where Is the Bee? Detecting Tiny Pollinators with a Single Collaborative-Head Transformer

ResearchDGX agent

arXiv:2608.08580v1 Announce Type: new Abstract: The CVPPA@ECCV 2026 BuzzSpot Challenge asks us to detect bees, bumblebees, hoverflies, and moths in 1920x1080 field keyframes. Its annotations carry 2 d

Wisdom in Unity: The Role of Multilingual Training in Figurative Language Identification in Proverbs

ResearchDGX agent

arXiv:2608.08090v1 Announce Type: new Abstract: Although multilingual approaches to figurative language identification are not new, the shift beyond language homogeneous training data requires a clear

World Simulator: Queer Erotica and the Absurdity of AI Video Models That Promise the World

ResearchDGX agent

arXiv:2608.07510v1 Announce Type: cross Abstract: Increasingly, AI video models are marketed as 'world simulators,' suggesting their ability to model infinite realities. Despite such claims, these mod

XEns-CKD: An Explainable Ensemble-Based Approach for Chronic Kidney Disease Stage Detection

ResearchDGX agent

arXiv:2608.07561v1 Announce Type: new Abstract: Chronic kidney disease (CKD) is a silent disease. Its progression may not significantly hamper a person's daily routine. Human kidney function can be cl

You Only Flow Once: Calibrated and Real-Time Radar Pose Estimation with Multi-Hypothesis Normalizing Flows

ResearchDGX agent

arXiv:2608.09579v1 Announce Type: new Abstract: Sparse and noisy millimeter-wave radar point cloud observations often correspond to multiple plausible human poses, making deterministic pose estimation

ZetaGPT: A Reference Implementation of Positional--Encoding--Free State--Space--Attention Language Models

ResearchDGX agent

arXiv:2608.09432v1 Announce Type: cross Abstract: Transformer-based language models rely on self-attention, whose computation is permutation-equivariant and therefore lacks an intrinsic mechanism for

ZOMP: Zeroth-Order Multi-Modal Prompt Tuning for Vision-Language Models

ResearchDGX agent

arXiv:2608.08060v1 Announce Type: new Abstract: Fine-tuning vision-language models such as CLIP typically requires backpropagation (BP) through the full model, which is infeasible when only forward-pa

10 Aug 2026

A Disturbance in the Force: Force Actuation on the RAVEN II Surgical Robot with Parallel Motor-Cable Units

ResearchDGX agent

arXiv:2608.06488v1 Announce Type: new Abstract: Difficulty in haptic feedback for surgical robots has been a long-term problem for decades. In recent years, learning-based force estimation from robot

A Finite E-Group of Nilpotency Class Three

ResearchDGX agent

arXiv:2608.07275v1 Announce Type: cross Abstract: A group is an E-group if every element commutes with each of its endomorphic images. Caranti asked whether a finite E-group can have nilpotency class

A foundation-model approach to pediatric headache classification from rs-fMRI

ResearchDGX agent

arXiv:2608.07287v1 Announce Type: new Abstract: Headache is the most common neurological disorder in children and substantially affects quality of life. We investigated whether resting-state functiona

A Haptic Robot Finger Designed for Guqin Instrument Playing

ResearchDGX agent

arXiv:2608.07002v1 Announce Type: new Abstract: With the rapid advancement of humanoid robotics and embodied intelligence technologies, numerous musical instrument-playing robots have emerged in recen

A Physics-Inspired Classical Digital Twin of Cortical Dynamics: A Band-Stratified Metriplectic Port-Hamiltonian Neural Network Learned from Brain-Computer-Interface EEG

ResearchDGX agent

arXiv:2607.10439v3 Announce Type: replace-cross Abstract: We present a physics-inspired classical digital twin of brain-computer- interface (BCI) data: a graph neural network constrained to a band-str

A primer on optimal transport for causal inference with observational data

ResearchDGX agent

arXiv:2503.07811v3 Announce Type: replace-cross Abstract: The theory of optimal transportation has developed into a powerful and elegant framework for comparing probability distributions, with wide-ra

A proximal subgradient method for nonconvex stochastic optimization under the Kurdyka-{L}ojasiewicz condition

ResearchDGX agent

arXiv:2608.05460v1 Announce Type: cross Abstract: This work introduces a proximal stochastic subgradient method for minimizing the sum of an expected cost, whose integrand is potentially nonsmooth and

A Rate Separation for Agnostic Direct Sums

ResearchDGX agent

arXiv:2608.06951v1 Announce Type: new Abstract: Hanneke, Moran, and Waknine ite{HannekeMoranWaknine2024} asked how the agnostic PAC learning curve of the direct sum C^r depends on the single-instance

A Transferable Autologistic Model for Predicting Rare Failures in Heterogeneous Equipment

ResearchDGX agent

arXiv:2608.06695v1 Announce Type: new Abstract: Predicting failures before they occur remains a major challenge in predictive maintenance, particularly when failures are rare, when equipment of the sa

Adversarial Causal Intervention Falsification

ResearchDGX agent

arXiv:2608.06427v1 Announce Type: new Abstract: Generative models can reproduce an observational distribution while encoding an incorrect causal structure. We study a sequential game in which a struct

AfriNLLB: Efficient Translation Models for African Languages

ResearchDGX agent

arXiv:2602.09373v2 Announce Type: replace Abstract: In this work, we present AfriNLLB, a series of lightweight models for efficient translation from and into African languages. AfriNLLB supports 15 la

AI for science needs reasoning, not just data

ResearchDGX agent

Every few decades, someone announces that science has reached its end. In 1903, the revered physicist Albert Michelson wrote that the “facts of physical science have all been discovered.” In the 1980s

AnyTrack: Unifying Visual Object Tracking with Any Modalities

ResearchDGX agent

arXiv:2608.06773v1 Announce Type: new Abstract: Visual object tracking aims to continuously locate specific targets within sequential frames, evolving from single-modal methods to multi-modal ones. Ho

Are Visual Place Recognition Models Recognizing Places or Conditions? Distractor-Augmented Evaluation and Condition Suppression

ResearchDGX agent

arXiv:2608.06847v1 Announce Type: cross Abstract: Long-term Visual Place Recognition (VPR) is typically evaluated by matching queries from one condition against a database from another. Crowdsourced m

Authoring and Management of Transparent Research Integrity Assessments of Randomised Clinical Trial Publications Using LLM-assisted Tools and Provenance Knowledge Graphs

ResearchDGX agent

arXiv:2608.07202v1 Announce Type: new Abstract: Systematic reviews of Randomised Controlled Trials (RCTs) are routinely used as evidence for clinical care guidelines. Such evidence has to meet high re

BDD2Seq: Enabling Scalable Reversible-Circuit Synthesis via Graph-to-Sequence Learning

ResearchDGX agent

arXiv:2511.08315v2 Announce Type: replace-cross Abstract: Binary Decision Diagrams (BDDs) are instrumental in many electronic design automation (EDA) tasks thanks to their compact representation of Bo

Bend the Basics: Degradation-Aware Deformable Tokenization for All-in-One Image Restoration

ResearchDGX agent

arXiv:2608.06832v1 Announce Type: new Abstract: All-in-one image restoration seeks a single model that can recover images degraded by diverse and spatially non-uniform corruptions. However, many unifi

Better Together: Quantifying the Benefits of AI-Assisted Recruitment

ResearchDGX agent

arXiv:2507.08029v2 Announce Type: replace Abstract: Hiring algorithms have mostly scored the materials recruiters already see. Large language models (LLMs) can instead generate new information about c

Beyond 'AI Language': The case for the idiolectal nature of LLM output

ResearchDGX agent

arXiv:2608.06589v1 Announce Type: cross Abstract: While large language model outputs are frequently analysed as a collective super variety termed 'AI language,' this chapter argues that this perspecti

Beyond Attention: Signed Integrated Gradients Attribution in a BiomeGPT-Style Microbiome Transformer

ResearchDGX agent

arXiv:2608.06486v1 Announce Type: new Abstract: In a feature-tokenized transformer (arXiv:2106.11959) such as BiomeGPT (doi:10.64898/2026.01.05.697599), each input token is built by fusing a fixed ide

Beyond Co-Movement: Locality by Exposures Enables a Joint Factor-Graph Framework for Portfolio Diversification

ResearchDGX agent

arXiv:2608.06618v1 Announce Type: cross Abstract: Current portfolio construction methods are either agnostic to the effects of idiosyncratic shocks (standard factor models) or to the latent data struc

Beyond Isolation: Unlocking Reinforcement Learning Component Synergy for Sample-Efficient Continuous Control

ResearchDGX agent

arXiv:2608.07086v1 Announce Type: cross Abstract: Reinforcement learning systems are significantly more complex than other machine learning paradigms due to inherent properties, causing RL system desi

Beyond Routing Weights: Faithful Response-Level Interpretation of Mixture-of-Experts Reward Models via Contribution Contrast

ResearchDGX agent

arXiv:2608.06400v1 Announce Type: new Abstract: Reward models are central to learning from human preferences, yet identifying what drives their predictions remains challenging. Recent sparse Mixture-o

Bridging the Gap Between Hyperdimensional Computing and Kernel Methods via the Nystrom Method

ResearchDGX agent

arXiv:2608.06860v1 Announce Type: cross Abstract: Hyperdimensional computing (HDC) is an approach from the cognitive science literature for solving information processing tasks using data represented

CAS2UML: A Handwritten Sketch-to-PlantUML Dataset for Class and Activity Diagrams

ResearchDGX agent

arXiv:2608.07036v1 Announce Type: cross Abstract: Automated UML generation from sketches and images is gaining renewed attention with the rise of large language models and multimodal AI. However, repr

CellWorld: From Gene-Level Reconstruction to Latent Cell Prediction in Spatial Transcriptomics Foundation Models

ResearchDGX agent

arXiv:2608.06659v1 Announce Type: new Abstract: This paper shows that latent-space predictive pretraining can provide a scalable route to foundation models for spatial transcriptomics. Existing spatia

Certified Feedforward Tracking for Unknown Nonlinear Systems via Invertible Neural Networks

ResearchDGX agent

arXiv:2608.06419v1 Announce Type: cross Abstract: In this paper, we address the certification of datadriven feedforward control for periodic tracking of unknown nonlinear systems under partial state m

Coding isn't yet another application domain -- it's the meta-skill required for AI to automatically develop its own training material, via s…

ResearchDGX agent

Coding isn't yet another application domain -- it's the meta-skill required for AI to automatically develop its own training material, via symbolic world models. That's how the RSI loop actually kicks

CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG

ResearchDGX agent

arXiv:2608.07458v1 Announce Type: cross Abstract: Recent optimization studies on Retrieval-Augmented Generation (RAG) have exploited chunk-level KV cache reuse to avoid processing long retrieved conte

Conditioning Protein Generation via Hopfield Pattern Multiplicity

ResearchDGX agent

arXiv:2603.20115v2 Announce Type: replace Abstract: Small protein-family alignments often contain a subset of interest but not enough labeled data to train a conditional generator. We condition a trai

Confirming Our Biases? Evaluating the Capabilities, Risks, and Societal Impact of Large Language Models

ResearchDGX agent

arXiv:2608.06977v1 Announce Type: new Abstract: It is well established that large language models (LLMs) are sensitive to prompt framing, reflecting patterns in their training data or prior prompts. I

Conformal Fusion Under Missing Modalities

ResearchDGX agent

arXiv:2608.07183v1 Announce Type: new Abstract: Multimodal fusion architectures typically assume all modalities are available at inference, yet sensor failures, acquisition variability, and cost const

ConstructCIE: A Dataset for Extracting Causal Information from Construction Accident Narratives

ResearchDGX agent

arXiv:2608.06495v1 Announce Type: new Abstract: Construction accident narratives contain rich causal information, but the evidence is often implicit, long-span, and distributed. We introduce Construct

Convergence of Diffusion Models Under the Manifold Hypothesis in High-Dimensions

ResearchDGX agent

arXiv:2409.18804v3 Announce Type: replace-cross Abstract: Denoising Diffusion Probabilistic Models (DDPM) are powerful state-of-the-art methods used to generate synthetic data from high-dimensional da

Counterfactual Simulation Training for Chain-of-Thought Faithfulness

ResearchDGX agent

arXiv:2602.20710v2 Announce Type: replace Abstract: Inspecting Chain-of-Thought reasoning is among the most common means of understanding why an LLM produced its output. But well-known problems with C

Cross-View Action Consistency for Camera-Robust Vision-Language-Action Policies

ResearchDGX agent

arXiv:2608.06965v1 Announce Type: new Abstract: Vision-language-action (VLA) policies fine-tuned from a fixed scene camera can fail when the camera is moved, even when the task, objects, language, and

Debias in Text, Believe Your Eyes: Text-Anchored Cross-Modal Transfer for Visual Counter-Commonsense Reasoning

ResearchDGX agent

arXiv:2608.06938v1 Announce Type: cross Abstract: The visual reasoning ability of multimodal large language models (MLLMs) is crucial for downstream applications, particularly counter-commonsense reas

Depth-Wise Probing and Pruning of the Planning Token in a Driving Vision-Language-Action Model

ResearchDGX agent

arXiv:2608.07361v1 Announce Type: cross Abstract: Vision-language-action (VLA) models route driving decisions through a deep language model, but it is unclear how much of that depth the action itself

Deterministic Preprocessing and Interpretable Fuzzy Banding for Cost-per-Student Reporting from Extracted Records

ResearchDGX agent

arXiv:2603.04905v2 Announce Type: replace-cross Abstract: Administrative extracts are often exchanged as spreadsheets and may be read as reports in their own right during budgeting, workload review, a

Dirichlet Follow-the-Leader Closes the Gap in Simultaneous Multiclass U-Calibration

ResearchDGX agent

arXiv:2608.06656v1 Announce Type: new Abstract: Can one forecaster attain the optimal regret rate for every bounded proper loss and also adapt to every smooth proper loss? Recent work answered this up

DocMemo: Dynamic Evidence Discovery via Probabilistic Memory-Guided Retrieval for Multi-Modal Document Understanding

ResearchDGX agent

arXiv:2608.07067v1 Announce Type: new Abstract: Long-document understanding requires locating sparse and heterogeneous evidence across hundreds of pages, yet existing systems remain limited by static

← Previous
1…89101112…315
Next →