AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,881 results
15 Jul 2026

ABot-3DWorld 0: A Universal World Model to Explore Any 3D Space

ResearchDGX agent

arXiv:2607.11673v2 Announce Type: replace Abstract: We present ABot-3DWorld 0, a universal multimodal 3D world model that turns text, image, and video inputs into high-fidelity, explorable 3D worlds.

Accelerated Mixing Time of Randomized Hamiltonian Monte Carlo

ResearchDGX agent

arXiv:2607.12902v1 Announce Type: cross Abstract: We show the Randomized Hamiltonian Monte Carlo (RHMC) algorithm has accelerated mixing time guarantees for sampling from log-concave probability distr

ACID: Adaptive Caching for vIDeo generation

ResearchDGX agent

arXiv:2607.12358v1 Announce Type: new Abstract: Video diffusion models produce high-quality generations but remain slow at inference due to their sequential denoising procedure. Caching-based accelera

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Adaptive Compute in Latent World Models: When Depth Helps, Hurts, or Doesn't Matter

ResearchDGX agent

arXiv:2607.10203v2 Announce Type: replace-cross Abstract: Adaptive-compute world models -- early-exit or mixture-of-depths predictors that spend variable depth per step -- assume depth buys better pre

Adversarial Attacks on Online Handwriting using Salience-based Temporal Editing

ResearchDGX agent

arXiv:2607.12500v1 Announce Type: cross Abstract: Deep learning models for online handwriting recognition have been shown effective and are increasingly deployed in practical applications. However, th

Affordance-Guided Diffusion Prior for 3D Hand Reconstruction

ResearchDGX agent

arXiv:2510.00506v2 Announce Type: replace Abstract: How can we reconstruct 3D hand poses when large portions of the hand are heavily occluded by itself or by objects? Humans often resolve such ambigui

An Omnilingual-ASR-Based Speech-LLM System for the 2nd MLC-SLM Challenge

ResearchDGX agent

arXiv:2607.12468v1 Announce Type: cross Abstract: We describe our submission to Task 1 of the 2nd MLCSLM Challenge: a cascaded diarization-then-recognition system that combines DiariZen-Large-s80 (Wav

Analysis of Mutual and Referential Human and Robot Gazes in a Collaborative Word Association Game

ResearchDGX agent

arXiv:2607.12181v1 Announce Type: new Abstract: Robot gaze is a major component of human-robot dialogue coordination. Most studies of gaze in human-robot dialogue focus on face-to-face social conversa

Analyzing Image Encoder Choices and Graph Homophily in GCN Frameworks for Breast Ultrasound Classification

ResearchDGX agent

arXiv:2607.12054v1 Announce Type: cross Abstract: Breast ultrasound is widely used for screening, yet automated analysis remains challenging due to speckle noise, acquisition variability, and weak sep

ARDepth: Auto-regressive Monocular Depth Estimation with Progressive Visual Conditioning

ResearchDGX agent

arXiv:2607.12433v1 Announce Type: cross Abstract: Diffusion models have recently become the dominant paradigm for monocular depth estimation (MDE). However, they implicitly assume that depth can be re

Audio-Native Speech Recognition with a Frozen Discrete-Diffusion Language Model

ResearchDGX agent

arXiv:2607.13013v1 Announce Type: new Abstract: Automatic speech recognition is dominated by autoregressive decoders that emit one token at a time. We ask whether a discrete diffusion language model c

AVQ-Attention: Adaptive Vector-Quantized Attention

ResearchDGX agent

arXiv:2607.12789v1 Announce Type: cross Abstract: The O(N^2) complexity of attention over N tokens remains a computational bottleneck in transformer models. Vector-Quantized (VQ) attention reduces thi

BattVAE-GP: Generative Modeling of Long-Horizon Battery Degradation with Uncertainty Quantification

ResearchDGX agent

arXiv:2607.11943v1 Announce Type: cross Abstract: Long-horizon physics-based simulations of battery degradation provide mechanistic insight but remain computationally expensive, limiting their use for

Beyond Logit Adjustment: A Residual Decomposition Framework for Long-Tailed Reranking

ResearchDGX agent

arXiv:2604.01506v2 Announce Type: replace Abstract: Long-tailed classification, where a small number of frequent classes dominate many rare ones, remains challenging because models systematically favo

ChartGenEval: Corruption-Tested Multi-Dimensional Feedback for Rhythm-Game Chart Generation

ResearchDGX agent

arXiv:2607.12857v1 Announce Type: cross Abstract: A generated rhythm-game chart need not reproduce one official note sequence: many note choices can fit the same song and difficulty. Reference-note ag

Cluster-Weighted EDMD

ResearchDGX agent

arXiv:2607.12243v1 Announce Type: new Abstract: Extended Dynamic Mode Decomposition (EDMD) approximates Koopman operators from data, but a single global operator is inefficient when different state-sp

Connected by Construction: Learning Tractable Near-Tour Marginals for Traveling Salesman Problems

ResearchDGX agent

arXiv:2607.12127v1 Announce Type: new Abstract: Learning-based methods for the traveling salesman problem (TSP) are often evaluated through the tours produced after decoding or search, but the learned

Constructed Reality, Contested Priors: Decoupling and the Architecture of Cognitive Relapse Under the Free Energy Principle

ResearchDGX agent

arXiv:2607.11958v1 Announce Type: new Abstract: Under the free energy principle, a predictive system does not observe reality directly; it maintains a generative model of the world and experiences tha

Decentralized Gradient Descent: Bottleneck Regimes and Budget Complexity

ResearchDGX agent

arXiv:2607.12172v1 Announce Type: cross Abstract: Decentralized gradient descent (DGD) is widely used for solving distributed optimization problems over networks of agents. While its convergence prope

Deep Learning-based Surrogate Modelling of the LOD Method for Multiscale Problems

ResearchDGX agent

arXiv:2607.12570v1 Announce Type: cross Abstract: Multiscale problems are notoriously difficult to tackle using traditional numerical methods, as accurately resolving fine-scale features often require

Demonstration of the common dual-channel feature decoupling characteristic of front-door mediation causal inference methods in whole-slice image classification

ResearchDGX agent

arXiv:2607.12376v1 Announce Type: cross Abstract: Causal inference using front door intervention and multi-instance learning (MIL) has advanced the analysis of Whole Slide Images (WSI) in digital path

Do We Really Need Transformers for Global Spatial Information Extraction in Traffic Forecasting?

ResearchDGX agent

arXiv:2607.12462v1 Announce Type: new Abstract: Existing traffic forecasting models commonly focus on extracting spatial dependencies, particularly global spatial information, which characterizes the

Do You Remember? Toward Memory-Centric Multimodal AI

ResearchDGX agent

arXiv:2607.11919v1 Announce Type: cross Abstract: Human memory is reconstructive, not a faithful recording. Current multimodal LLMs (MLLMs) lack this capability: they process images through a frozen v

Domain-Incremental Remote Sensing Change Detection via Difference-Guided Adaptation and Frequency-Decoupled Distillation

ResearchDGX agent

arXiv:2607.12934v1 Announce Type: new Abstract: Remote sensing change detection (RSCD) models are prone to catastrophic forgetting when incrementally adapted to new domains. Existing domain-incrementa

Dynamic Online Processor-Native Inference for State Estimation

ResearchDGX agent

arXiv:2607.12095v1 Announce Type: cross Abstract: Sensor-rich data-driven applications increasingly use Bayesian approaches to infer latent states of dynamic systems from noisy sensor measurements and

DynTrace: Tracking Dynamic Object Evidence for 4D Spatio-Temporal Reasoning in MLLMs

ResearchDGX agent

arXiv:2607.12503v1 Announce Type: new Abstract: 4D spatio-temporal reasoning, jointly modeling 3D spatial structure and temporal evolution, is essential for understanding dynamic worlds and enabling e

Efficient Conformal Prediction for Regression Models under Label Noise

ResearchDGX agent

arXiv:2509.15120v2 Announce Type: replace Abstract: In high-stakes scenarios, such as medical imaging applications, it is critical to equip the predictions of a regression model with reliable confiden

Energy-Based Physics-Informed Form Finding for Clustered Tensegrity Structures

ResearchDGX agent

arXiv:2607.12888v1 Announce Type: new Abstract: Tensegrity form-finding and physical property prediction are fundamental inverse problems in structural mechanics, which aim to determine equilibrium co

Ensemble Controlled-Flow Filtering for Implicit Data Assimilation

ResearchDGX agent

arXiv:2607.12975v1 Announce Type: cross Abstract: Data assimilation estimates the state of a dynamical system from model forecasts and incoming observations. Many observation mechanisms, however, are

Entropy in Semantic Memory Navigation in Blind and Sighted Individuals: The Effect of Visual Experience

ResearchDGX agent

arXiv:2607.12185v1 Announce Type: new Abstract: Embodied accounts of semantic memory highlight the role of sensorimotor systems in acquiring and storing knowledge. Congenitally blind populations offer

Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models

ResearchDGX agent

arXiv:2602.02244v3 Announce Type: replace-cross Abstract: The standard post-training recipe for large reasoning models, supervised fine-tuning followed by reinforcement learning (SFT-then-RL), may lim

Evaluating Nonuniform Dependability Across Response Conditions: A Conditional Generalizability Framework Illustrated in Automated Essay Scoring

ResearchDGX agent

arXiv:2607.11981v1 Announce Type: cross Abstract: Aggregate reliability estimates can obscure heterogeneity in measurement-design burden across response conditions, so a single G- or D-study may misch

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

ResearchDGX agent

arXiv:2509.22415v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved strong vision-language performance, yet their token-level visual evidence remains diffi

Exact and Calibrated Diffusion Reconstruction for Digital Breast Tomosynthesis

ResearchDGX agent

arXiv:2607.12937v1 Announce Type: cross Abstract: Limited-angle digital breast tomosynthesis (DBT) reconstructs a volume from a few low-dose projections over a narrow arc. At a representative nine-vie

Exact and Certified Data Shapley for Weighted k-Nearest-Neighbor Regression and Soft-Label Prediction

ResearchDGX agent

arXiv:2607.11956v1 Announce Type: cross Abstract: Data Shapley is the standard principled answer to which training points are worth what, and its k-nearest-neighbor (KNN) specialization is the version

Excited for our first general model Inkling -- open weights, 975B, natively multimodal (text, image, audio). Available on Tinker, HuggingFac…

ResearchDGX agent

Excited for our first general model Inkling -- open weights, 975B, natively multimodal (text, image, audio). Available on Tinker, HuggingFace and partners. It is yours to personalize and use openly. I

Fast and Accurate Image Restoration and Generation with Rank Enhanced Linear Attention

ResearchDGX agent

arXiv:2505.16157v2 Announce Type: replace Abstract: Transformer-based models have made remarkable progress in image restoration (IR) tasks. However, the quadratic complexity of self-attention in Trans

Flatness-Preserving Residual Learning for Real-Time Tight Quadrotor Formation Flight

ResearchDGX agent

arXiv:2607.12275v1 Announce Type: new Abstract: Quadrotors flying in tight formations are severely affected by turbulent aerodynamic interactions, such as downwash, that can cause catastrophic collisi

From Hindsight to Foresight: Self-Encouraged Hindsight Distillation for Knowledge-based Visual Question Answering

ResearchDGX agent

arXiv:2511.11132v4 Announce Type: replace Abstract: Knowledge-based Visual Question Answering (KBVQA) necessitates external knowledge incorporation beyond cross-modal understanding. Existing KBVQA met

From Observation to Insight: Mechanistic World Models and the Quest for Autonomous Discovery

AgentsDGX agent

arXiv:2607.12474v1 Announce Type: new Abstract: Recent advances in foundation models have transformed AI for Science, enabling remarkably accurate predictive performance across domains ranging from pr

From Reconstruction to Interpretation: Zero-Setup Multi-Phase Segmentation of X-ray Tomography Data

ResearchDGX agent

arXiv:2607.12175v1 Announce Type: cross Abstract: X-ray tomography enables nondestructive characterization of material microstructures, while advances in micro-CT imaging have accelerated volumetric d

From Words to Widgets for Controllable LLM Generation

ResearchDGX agent

arXiv:2604.10925v2 Announce Type: cross Abstract: Natural language remains the predominant way people interact with large language models (LLMs). However, users often struggle to precisely express and

Gaussian Mixture Modeling for Event-Aware Visual Allocation in Long Video Understanding

ResearchDGX agent

arXiv:2607.12557v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) face significant challenges in long video understanding due to the excessive computational cost and information los

Gene Expression-Informed Jointly Controlled Generative Modeling for Precision Molecular Design

ResearchDGX agent

arXiv:2607.11978v1 Announce Type: cross Abstract: Precision molecular design aims to discover personalized drug candidates through joint control of multiple conditions, such as biological relevance an

Generalized Distribution-Free Semi-Supervised Learning with Risk Rewrite

ResearchDGX agent

arXiv:2607.11947v1 Announce Type: cross Abstract: Typical semi-supervised learning (SSL) methods rely on distributional assumptions, and their performance degrades when these are violated. While PNU l

Generating Physically Plausible Parachute Dynamics with Deep Generative Modeling

ResearchDGX agent

arXiv:2607.12143v1 Announce Type: cross Abstract: Accurately modeling the dynamics of planetary parachute and entry vehicle systems is critical for Entry, Descent, and Landing events such as vehicle s

GeoSEAN: Explainable Country-Level Image Geolocation for ASEAN Regions

ResearchDGX agent

arXiv:2607.12284v1 Announce Type: new Abstract: Image geolocation aims to infer the geographic origin of an image from visual content alone. However, this task remains challenging in regions where cou

Good Benchmarks

ResearchDGX agent

arXiv:2607.12217v1 Announce Type: new Abstract: Good tasks are correct, solvable, verifiable, well-specified, and hard for interesting reasons. The best tasks describe a real problem an experienced pr

Graph-Based Detection of Disinformation Narrative Diffusion between Russian and Ukrainian Telegram Channels

ResearchDGX agent

arXiv:2607.11894v1 Announce Type: cross Abstract: Detecting disinformation narratives on social media is challenging due to the scale of amplification, rapid evolution, and linguistic variability of o

Graph Regularized PCA

ResearchDGX agent

arXiv:2601.10199v2 Announce Type: replace Abstract: Multivariate data often exhibit complex dependencies that violate the assumption of isotropic residual noise. For such cases, we introduce Graph Reg

Hierarchical Latent Structures in Data Generation Process Unify Mechanistic Phenomena across Scale

ResearchDGX agent

arXiv:2603.06592v2 Announce Type: replace Abstract: Contemporary studies in mechanistic interpretability have uncovered many puzzling phenomena in the neural information processing of Transformer-base

High-Dimensional Gaussian Mean Estimation under Realizable Contamination

ResearchDGX agent

arXiv:2603.16798v2 Announce Type: replace Abstract: We study mean estimation for a Gaussian distribution with identity covariance in R^d under a missing data scheme termed realizable epsilon-contamina

Higher Embedding Dimension Creates a Stronger World Model for a Simple Sorting Task

ResearchDGX agent

arXiv:2510.18315v2 Announce Type: replace-cross Abstract: We investigate how embedding dimension affects the emergence of an internal 'world model' in a transformer trained with reinforcement learning

How Query Visibility Changes KV-Cache Compression Rankings: A Matched-Budget Audit

ResearchDGX agent

arXiv:2607.11942v1 Announce Type: cross Abstract: KV-cache compression methods are predominantly evaluated with the query appended to the context before compression -- a query-aware protocol. Yet the

Hybrid Continual Learning for Low-Resource Australian Aboriginal Language Identification

ResearchDGX agent

arXiv:2607.11946v1 Announce Type: new Abstract: Language identification is an important step toward integrating endangered Australian Aboriginal languages (AALs) into speech technologies supporting la

I built a new attention mechanism (wave field) — runs 128K context where standard attention OOMs, 80+ tok/s on laptop CPU

Model ReleasesDGX agent

Hey r/LocalLLaMA — solo researcher here. I built a new attention architecture and want independent testers. Wave Field LLM replaces O(N²) dot-product attention with FFT wave convolution on a field. Tr

I’m very excited by this test time training work for robotic learning! It’s an awesome collaboration between @StanfordSVL and @NVIDIARobotic…

ResearchDGX agent

I’m very excited by this test time training work for robotic learning! It’s an awesome collaboration between @StanfordSVL and @NVIDIARobotics ! We scaled a robot model natively to 8,000 timesteps of c

Implicit 4D Gaussian Splatting for Fast Motion with Large Inter-Frame Displacements

ResearchDGX agent

arXiv:2607.12362v1 Announce Type: new Abstract: Recent 4D Gaussian Splatting (4DGS) methods often fail under fast motion with large inter-frame displacements, where Gaussian attributes are poorly lear

Improved Robustness from Biologically Inspired Sparse Contrast Representations

ResearchDGX agent

arXiv:2509.24863v2 Announce Type: replace Abstract: Deep neural networks surpass humans on many vision benchmarks, yet remain far less robust to distribution shifts such as illumination and weather ch

Inhibited Self-Attention: Sharpening Focus in Vision Transformers

ResearchDGX agent

arXiv:2607.12881v1 Announce Type: new Abstract: Vision Transformers (ViTs) have demonstrated remarkable performance in computer vision tasks. However, their self-attention mechanism often diffuses foc

← Previous
1…107108109110111…432
Next →