AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,888 results
8 Jun 2026

MHA-RAG: Improving Efficiency, Accuracy, and Consistency by Encoding Exemplars as Soft Prompts

ResearchDGX agent

arXiv:2510.05363v2 Announce Type: replace Abstract: Adapting Foundation Models to new domains with limited training data is challenging and computationally expensive. While prior work has demonstrated

Mind the Gap: Disentangling Performance Bottlenecks in Video Instance Segmentation

ResearchDGX agent

arXiv:2606.07394v1 Announce Type: new Abstract: In Video Instance Segmentation (VIS), classification, segmentation, and tracking objectives are jointly evaluated, but their individual contributions to

MoDA: Modulation Adapter for Fine-Grained Visual Grounding in Instructional MLLMs

ResearchDGX agent

arXiv:2506.01850v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success in instruction-following tasks by integrating pretrained visual enco

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Modeling AdaGrad, RMSProp, and Adam with Integro-Differential Equations

ResearchDGX agent

arXiv:2411.09734v3 Announce Type: replace Abstract: In this paper, we propose a continuous-time formulation for the AdaGrad, RMSProp, and Adam optimization algorithms by modeling them as first-order i

Modular Monolingual Adaptation using Pretrained Language Models

ResearchDGX agent

arXiv:2606.06738v1 Announce Type: new Abstract: Building monolingual language models (LMs) for low-resource languages typically relies on adapting pretrained language models (PLMs) by finetuning the w

Multi-objective optimization and quantum hybridization of equivariant deep learning interatomic potentials

ResearchDGX agent

arXiv:2602.16908v2 Announce Type: replace-cross Abstract: Allegro is a machine learning interatomic potential model designed to predict atomic properties in molecules using E(3) equivariant neural net

Multi-Robot Planning and Control from CCTV Camera Networks in a Real Warehouse

ResearchDGX agent

arXiv:2606.06762v1 Announce Type: new Abstract: Off-board control of mobile robots from cameras embedded in the environment offers a practical path to scalable autonomy, moving sensing and compute off

Multilingual Multi-Speaker Unit Vocoders: A Systematic Analysis of Discrete Speech Representations

ResearchDGX agent

arXiv:2606.06740v1 Announce Type: cross Abstract: Discrete speech units obtained via k-means clustering of self supervised embeddings entangle phonetic, speaker, and language information, causing spea

Multiscale POD of Transformer Attention Fields: Scale-Selective Analysis via Morlet Scalogram

ResearchDGX agent

arXiv:2606.06573v1 Announce Type: cross Abstract: We introduce scale-selective Proper Orthogonal Decomposition (POD) for transformer attention fields, inspired by the use of POD for extracting energet

OffQ: Taming Structured Outliers in LLM Quantization by Offsetting

ResearchDGX agent

arXiv:2606.07116v1 Announce Type: cross Abstract: Low-bit quantization has been widely adopted to accelerate the inference of large language models (LLMs) by significantly reducing computational cost

On orbital stabilization of a circular motion primitive for a dynamic extension of the Dubins car model

ResearchDGX agent

arXiv:2606.07449v1 Announce Type: cross Abstract: This paper addresses orbital stabilization of a circular motion primitive for a dynamic extension of the Dubins car model within a transverse-lineariz

On the conditional equivalence of phase retrieval algorithms

ResearchDGX agent

arXiv:2606.07257v1 Announce Type: cross Abstract: Phase retrieval - recovering a complex-valued field from intensity measurements - is typically solved using variants of the Gerchberg-Saxton (GS) algo

One Loss to Rule Them All: Marked Time-to-Event for Structured EHR Foundation Models

ResearchDGX agent

arXiv:2602.00541v2 Announce Type: replace Abstract: Clinical events captured in Electronic Health Records (EHR) are irregularly sampled and may consist of a mixture of discrete events and numerical me

OpenACMv2: An Accuracy-Constrained Co-Optimization Framework for Approximate DCiM

ResearchDGX agent

arXiv:2603.13042v2 Announce Type: replace Abstract: Digital Compute-in-Memory (DCiM) accelerates neural networks by reducing data movement. Approximate DCiM can further improve power-performance-area

Optimal Rates for Generalization of Gradient Descent Methods with Deep Neural Networks

ResearchDGX agent

arXiv:2606.06764v1 Announce Type: cross Abstract: Recent progress has been made in understanding the statistical generalization performance of gradient descent methods for overparameterized neural net

P-Cast Precision in FP8 Attention: Sink-Induced Collapse and the Optimality of S=2^8

ResearchDGX agent

arXiv:2606.06521v1 Announce Type: cross Abstract: FP8 (E4M3) acceleration for attention computation offers significant throughput gains, but the 3-bit mantissa introduces precision challenges when the

PARSE: Part-Aware Relational Spatial Modeling

ResearchDGX agent

arXiv:2603.07704v2 Announce Type: replace Abstract: Inter-object relations underpin spatial intelligence, yet existing representations -- linguistic prepositions or object-level scene graphs -- are to

Phonetic Error Analysis of Raw Waveform Acoustic Models

ResearchDGX agent

arXiv:2606.07030v1 Announce Type: cross Abstract: We analyse error patterns of raw waveform acoustic models on TIMIT phone recognition beyond the overall phone error rate (PER). PER is decomposed acro

Phun-Bench: Evaluating LLMs on Phonological Understanding in Chinese

Model ReleasesDGX agent

arXiv:2606.07300v1 Announce Type: new Abstract: Language is a vehicle for thought, intricately tied to sounds, symbols, and meaning. However, most large language model (LLM) research focuses on meanin

Physiologically Constrained Musculoskeletal Neural Network for Multi-DoF Joint Kinematics Estimation from Partially Observed sEMG

ResearchDGX agent

arXiv:2606.07476v1 Announce Type: cross Abstract: This paper investigates multi-degrees of freedom (DoF) joint kinematics estimation under partially observed surface electromyography (sEMG), where onl

PolarQuant: Leveraging Polar Transformation for Efficient Key Cache Quantization and Decoding Acceleration

ResearchDGX agent

arXiv:2502.00527v2 Announce Type: replace-cross Abstract: The KV cache in large language models is a dominant factor in memory usage, limiting their broader applicability. Quantizing the cache to lowe

polyDAG: Polynomial Acyclicity Constraints for Efficient Continuous Causal Discovery in Visual Semantic Graphs

ResearchDGX agent

arXiv:2606.06908v1 Announce Type: new Abstract: Modern image-analysis pipelines often convert images into structured semantic variables, such as facial attributes, object concepts, and scene descripto

Position: A Dynamical Systems Perspective is Needed to Advance Time Series Modeling

ResearchDGX agent

arXiv:2602.16864v2 Announce Type: replace-cross Abstract: Time series (TS) modeling has come a long way from early statistical, mainly linear, approaches to the current trend in TS foundation models.

Predictable Compression Failures: Order Sensitivity and Information Budgeting for Evidence-Grounded Binary Adjudication

ResearchDGX agent

arXiv:2509.11208v3 Announce Type: replace-cross Abstract: Transformers used for evidence-grounded binary adjudication (e.g., support/refute, yes/no, or verifier-backed pass/fail decisions) can be sens

Predictive Style Matching: Natural and Robust Humanoid Locomotion

ResearchDGX agent

arXiv:2606.07083v1 Announce Type: new Abstract: Reinforcement learning has become the prevailing approach to humanoid locomotion control: policies transfer reliably from simulation to hardware and rec

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory

ResearchDGX agent

arXiv:2606.06624v1 Announce Type: new Abstract: In the current era of deep learning and especially generative models, there is significant investment in training very large generative models. Thus far

Privacy Implies Stability: Information-Theoretic Generalization Bounds for Quantum Learning

ResearchDGX agent

arXiv:2602.01177v3 Announce Type: replace-cross Abstract: We develop an information-theoretic framework connecting stability, privacy, and generalization for quantum learning algorithms. Learning proc

Probabilistic Gaussian Homotopy: A Probability-Space Continuation Framework for Nonconvex Optimization

ResearchDGX agent

arXiv:2603.13546v2 Announce Type: replace Abstract: We introduce Probabilistic Gaussian Homotopy (PGH), a probability-space continuation framework for nonconvex optimization. Unlike classical Gaussian

Probabilistic learning to perform pre-onset individualised prediction of disease severity: application to Veno Occlusive Disease

ResearchDGX agent

arXiv:2606.06516v1 Announce Type: cross Abstract: We advance a new probabilistic supervised learning approach that permits reliable, automated, and early individualised prediction of the severity with

pTNAS: Progressive Neural Architecture Search for Tabular Data

ResearchDGX agent

arXiv:2403.10318v3 Announce Type: replace Abstract: Recent advances have shifted the paradigm of tabular learning toward tabular foundation models, yet their accuracy relies on a heavy inference cost

Quantifying Media Representation Dynamics Across 25 Years of News Reporting on Policing-related Deaths

ResearchDGX agent

arXiv:2606.06812v1 Announce Type: new Abstract: We perform the largest known computational analysis of Canadian news narratives about police-involved deaths, spanning 4,000 articles from the last quar

Re-Centering Humans in LLM Personalization

ResearchDGX agent

arXiv:2606.06614v1 Announce Type: cross Abstract: Despite growing interest, most evaluations of large language models' (LLMs') personalization abilities have relied on synthetic data. It remains uncle

Real-Time AttentionBender: Granular Interactive Network Bending of Video Diffusion Transformers

ResearchDGX agent

arXiv:2606.06497v1 Announce Type: cross Abstract: Generative video models have achieved remarkable visual fidelity, yet their prompt-only interface offers thin creative agency and obscures the model's

Reconstructing Multi-Decadal Forest Disturbances: A Spatio-Temporal Transformer Approach

ResearchDGX agent

arXiv:2606.07249v1 Announce Type: new Abstract: Accurate monitoring of forest disturbances is essential for understanding carbon dynamics and land management, yet traditional approaches typically rely

Reference-Free Evaluation of Taxonomies

ResearchDGX agent

arXiv:2505.11470v3 Announce Type: replace Abstract: We introduce two reference-free metrics for quality evaluation of taxonomies in the absence of labels. The first metric evaluates robustness by calc

RePo: Language Models with Context Re-Positioning

ResearchDGX agent

arXiv:2512.14391v3 Announce Type: replace-cross Abstract: In-context learning is fundamental to modern Large Language Models (LLMs); however, prevailing architectures impose a rigid and fixed contextu

Rethinking Genomic Modeling Through Optical Character Recognition

ResearchDGX agent

arXiv:2602.02014v2 Announce Type: replace-cross Abstract: Recent genomic foundation models largely adopt large language model architectures that treat DNA as a one-dimensional token sequence. However,

RigPAPR: Rig-Based Animation of Static Neural Point Clouds from a Fixed-Viewpoint Video

ResearchDGX agent

arXiv:2606.06685v1 Announce Type: new Abstract: Static neural point reconstructions capture a subject at high fidelity from posed images. Given such a reconstruction, we aim to animate it to follow a

S23DR 2026 Winning Solution

ResearchDGX agent

arXiv:2606.06695v1 Announce Type: new Abstract: This text presents the winning solution to the S23DR 2026 challenge for structured 3D wireframe reconstruction from sparse SfM, fitted depth, and semant

Second-Order Path Kernel Interpolation Formulas in Machine Learning

ResearchDGX agent

arXiv:2606.07495v1 Announce Type: new Abstract: Understanding how training data shape neural network predictions is a central problem in modern learning theory. In 2020, Pedro Domingos proposed an int

SEEK: Steering LLM Reasoning for RAG via Internal Reasoning Sketches

ResearchDGX agent

arXiv:2601.09402v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by incorporating external knowledge into the generation process. Benefiti

SegMoTE: Token-Level Mixture of Experts for Medical Image Segmentation

ResearchDGX agent

arXiv:2602.19213v2 Announce Type: replace Abstract: Medical image segmentation is vital for clinical diagnosis and quantitative analysis, yet remains challenging due to the heterogeneity of imaging mo

SHAP-Guided Kernel Actor-Critic for Explainable Reinforcement Learning

ResearchDGX agent

arXiv:2512.05291v3 Announce Type: replace Abstract: Actor-critic (AC) methods are a cornerstone of reinforcement learning (RL) but offer limited interpretability. Current explainable RL methods seldom

Skip a Layer or Loop It? Learning Program-of-Layers in LLMs

ResearchDGX agent

arXiv:2606.06574v1 Announce Type: new Abstract: Large language models (LLMs) perform inference by following a fixed depth and order, non-recurrent execution of all layers. We reveal the wide existence

SleepExplain: Explainable Non-Rapid Eye Movement and Rapid Eye Movement Sleep Stage Classification from EEG Signal

ResearchDGX agent

arXiv:2606.07351v1 Announce Type: cross Abstract: Classification of sleep stages is one of the most important diagnostic approaches for a variety of sleep-related disorders. Electroencephalography (EE

Small Language Model Agents Enable Efficient and High-Quality Knowledge Mining

AgentsDGX agent

arXiv:2510.01427v3 Announce Type: replace Abstract: At the core of Deep Research is knowledge mining, the task of extracting structured information from massive unstructured text in response to user i

Stability beyond Bounded Differences: Sharp Generalization Bounds under Finite L_p Moments

ResearchDGX agent

arXiv:2606.06855v1 Announce Type: cross Abstract: While algorithmic stability is a central tool for understanding generalization of learning algorithms, existing high-probability guarantees typically

Striking paper from Wharton. The big conclusion: AI must increase productivity 2.7x -- and quickly -- or tech companies risk bankruptcy with…

ResearchDGX agent

Striking paper from Wharton. The big conclusion: AI must increase productivity 2.7x -- and quickly -- or tech companies risk bankruptcy with all that entails for the economy. For context: this is how

STRIPS-WM: Learning Grounded Propositional STRIPS-style World Models from Images

ResearchDGX agent

arXiv:2606.06832v1 Announce Type: new Abstract: Robots performing long-horizon visual manipulation observe high-dimensional images, but successful plans depend on action-relevant facts: what can be do

SWE-IF: Aligning Code Evaluation with Human Preference

ResearchDGX agent

arXiv:2510.07315v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have catalyzed vibe coding, where users leverage LLMs to generate and iteratively refine code through natural lan

Synthetic Benchmarks Overstate Forward-Forward Scaling: Real-Data Limits of Layer-Local Training

ResearchDGX agent

arXiv:2606.06539v1 Announce Type: cross Abstract: Forward-Forward (FF) learning [Hinton, 2022] replaces backpropagation with strictly layer-local goodness updates. Recent FF-CNN work has narrowed the

T2LM: Long-Term 3D Human Motion Generation from Multiple Sentences

ResearchDGX agent

arXiv:2406.00636v2 Announce Type: replace Abstract: In this paper, we address the challenging problem of long-term 3D human motion generation. Specifically, we aim to generate a long sequence of smoot

TA-RAG: Tone-Aware Retrieval-Augmented Generation for Peer-Support Health Communication

ResearchDGX agent

arXiv:2606.06794v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) successfully grounds large language model (LLM) outputs in trusted documents, but factual grounding alone is insuff

TabSwift: An Efficient Tabular Foundation Model with Row-Wise Attention

ResearchDGX agent

arXiv:2606.07345v1 Announce Type: new Abstract: Tabular foundation models, exemplified by TabPFN, perform prediction via in-context learning, inferring test labels directly from labeled training examp

Telling stories, making Hanzi: AI-assisted co-creation with elderly migrants in urban China

ResearchDGX agent

arXiv:2507.01548v3 Announce Type: replace-cross Abstract: This paper explores how older migrants in urban China can record stories that everyday language and design often miss. We ran two co-creation

Terastal: Layer-Variant-based Scheduling for Real-Time Multi-DNN Workloads on Heterogeneous Accelerators

ResearchDGX agent

arXiv:2606.06818v1 Announce Type: cross Abstract: Heterogeneous DNN accelerators improve soft real-time multi-DNN execution by mapping each layer to its preferred accelerator to reduce latency. Howeve

The Download: how the World Cup ball will fly and OpenAI’s “super app”

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Why this year’s World Cup ball may not fly as far Much is new

The Dual Mechanisms of Spatial Variable Binding in Vision-Language Models

ResearchDGX agent

arXiv:2603.22278v2 Announce Type: replace Abstract: Many multimodal tasks, such as image captioning and visual question answering, require vision-language models (VLMs) to bind objects with their prop

The Effect of Training Task Diversity on In-Context Learning through the Lens of Low-Dimensional Subspaces

ResearchDGX agent

arXiv:2606.06814v1 Announce Type: cross Abstract: The transformer's emergent ability to perform in-context learning (ICL) has sparked a wide range of studies designed to understand its underlying mech

The Identity Trap in EEG Foundation Models: A Diagnostic Audit

ResearchDGX agent

arXiv:2606.06647v1 Announce Type: new Abstract: Objective. EEG foundation models (FMs) report strong accuracy on clinical resting-state EEG. However, high accuracy under subject-disjoint cross-validat

← Previous
1…174175176177178…432
Next →