AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,885 results
Research

FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation

DGX agent

arXiv:2603.09721v2 Announce Type: replace Abstract: High-fidelity video generation remains challenging for diffusion models due to the difficulty of modeling complex spatio-temporal dynamics efficient

researcharxiv-cs-cv
21 Apr 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

FrameVGGT: Geometry-Aligned Frame-Level Memory for Bounded Streaming VGGT

DGX agent

arXiv:2603.07690v2 Announce Type: replace Abstract: Streaming Visual Geometry Transformers such as StreamVGGT enable strong online 3D perception, but their KV-cache grows unbounded over long streams,

researcharxiv-cs-cv
21 Apr 2026
Research

Frequency-guided Multi-level Reasoning for Scene Graph Generation in Video

DGX agent

arXiv:2604.17298v1 Announce Type: new Abstract: Video Scene Graph Generation aims to obtain structured semantic representations of objects and their relationships in videos for high-level understandin

researcharxiv-cs-cv
21 Apr 2026
Research

From 2:4 to 8:16 sparsity patterns in LLMs for Outliers and Weights with Variance Correction

DGX agent

arXiv:2507.03052v2 Announce Type: replace Abstract: As large language models (LLMs) grow in size, efficient compression techniques like quantization and sparsification are critical. While quantization

researcharxiv-cs-lg
21 Apr 2026
Research

From Attribution to Abstention: Training-Free Attention-Based Auditing for Clinical Summarization

DGX agent

arXiv:2601.16397v2 Announce Type: replace Abstract: Deploying multimodal large language models (MLLMs) for clinical summarization demands not only fluent generation but also transparency about where e

researcharxiv-cs-cl
21 Apr 2026
Research

From Classical to Quantum: Extending Prometheus for Unsupervised Discovery of Phase Transitions in Three Dimensions and Quantum Systems

DGX agent

arXiv:2602.14928v4 Announce Type: replace-cross Abstract: We extend the Prometheus framework for unsupervised phase transition discovery from two-dimensional classical systems to three-dimensional cla

researcharxiv-cs-lg
21 Apr 2026
Research

From Domains to Instances: Dual-Granularity Data Synthesis for LLM Unlearning

DGX agent

arXiv:2601.04278v2 Announce Type: replace Abstract: Although machine unlearning is essential for removing private, harmful, or copyrighted content from LLMs, current benchmarks often fail to faithfull

researcharxiv-cs-cl
21 Apr 2026
Research

From Fallback to Frontline: When Can LLMs be Superior Annotators of Human Perspectives?

DGX agent

arXiv:2604.17968v1 Announce Type: cross Abstract: Although large language models (LLMs) are increasingly used as annotators at scale, they are typically treated as a pragmatic fallback rather than a f

researcharxiv-cs-cl
21 Apr 2026
Research

From Implicit to Explicit: Token-Efficient Logical Supervision for Mathematical Reasoning in LLMs

DGX agent

arXiv:2601.03682v2 Announce Type: replace Abstract: Recent studies reveal that large language models (LLMs) exhibit limited logical reasoning abilities in mathematical problem-solving, instead often r

researcharxiv-cs-cl
21 Apr 2026
Research

From User Recognition to Activity Counting: An Identity-Agnostic Approach to Multi-User WiFi Sensing

DGX agent

arXiv:2604.16572v1 Announce Type: new Abstract: Wi-Fi Channel State Information (CSI) enables device-free human activity recognition, but existing multi-user approaches assume a fixed set of known use

researcharxiv-cs-lg
21 Apr 2026
Research

From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents

DGX agent

arXiv:2603.01455v2 Announce Type: replace-cross Abstract: While multimodal large language models have demonstrated impressive short-term reasoning, they struggle with long-horizon video understanding

researcharxiv-cs-cl
21 Apr 2026
Research

G-PARC: Graph-Physics Aware Recurrent Convolutional Neural Networks for Spatiotemporal Dynamics on Unstructured Meshes

DGX agent

arXiv:2604.16533v1 Announce Type: new Abstract: Physics-aware recurrent convolutional networks (PARC) have demonstrated strong performance in predicting nonlinear spatiotemporal dynamics by embedding

researcharxiv-cs-lg
21 Apr 2026
Research

GaLa: Hypergraph-Guided Visual Language Models for Procedural Planning

DGX agent

arXiv:2604.17241v1 Announce Type: new Abstract: Implicit spatial relations and deep semantic structures encoded in object attributes are crucial for procedural planning in embodied AI systems. However

researcharxiv-cs-ro
21 Apr 2026
Research

GeGS-PCR: Effective and Robust 3D Point Cloud Registration with Two-Stage Color-Enhanced Geometric-3DGS Fusion

DGX agent

arXiv:2604.17721v1 Announce Type: new Abstract: We address the challenge of point cloud registration using color information, where traditional methods relying solely on geometric features often strug

researcharxiv-cs-cv
21 Apr 2026
Research

Geometry-Guided 3D Visual Token Pruning for Video-Language Models

DGX agent

arXiv:2604.18260v1 Announce Type: new Abstract: Multimodal large language models have demonstrated remarkable capabilities in 2D vision, motivating their extension to 3D scene understanding. Recent st

researcharxiv-cs-cv
21 Apr 2026
Research

Get ready for hotter, muggier, stormier summers

DGX agent

A long stretch of humid heat followed by a powerful thunderstorm is a familiar weather pattern in the tropics, but it’s also becoming more common in midlatitude regions such as the US Midwest. A recen

researchmit-tech-review
21 Apr 2026
Research

GoCoMA: Hyperbolic Multimodal Representation Fusion for Large Language Model-Generated Code Attribution

DGX agent

arXiv:2604.16377v1 Announce Type: new Abstract: Large Language Models (LLMs) trained on massive code corpora are now increasingly capable of generating code that is hard to distinguish from human-writ

researcharxiv-cs-cl
21 Apr 2026
Applications

Graph Neural Networks for Graphs with Heterophily: A Survey

DGX agent

arXiv:2202.07082v4 Announce Type: replace Abstract: Recent years have witnessed fast developments of graph neural networks (GNNs) that have benefited myriad graph analytic tasks and applications. Most

applicationsarxiv-cs-lg
21 Apr 2026
Research

HABIT: Chrono-Synergia Robust Progressive Learning Framework for Composed Image Retrieval

DGX agent

arXiv:2604.18037v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) is a flexible image retrieval paradigm that enables users to accurately locate the target image through a multimodal quer

researcharxiv-cs-cv
21 Apr 2026
Research

Hard to Be Heard: Phoneme-Level ASR Analysis of Phonologically Complex, Low-Resource Endangered Languages

DGX agent

arXiv:2604.18204v1 Announce Type: new Abstract: We present a phoneme-level analysis of automatic speech recognition (ASR) for two low-resourced and phonologically complex East Caucasian languages, Arc

researcharxiv-cs-cl
21 Apr 2026
Research

HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference

DGX agent

arXiv:2601.13684v2 Announce Type: replace Abstract: The linear memory growth of the KV cache poses a significant bottleneck for LLM inference in long-context tasks. Existing static compression methods

researcharxiv-cs-cl
21 Apr 2026
Research

HiPrune: Hierarchical Attention for Efficient Token Pruning in Vision-Language Models

DGX agent

arXiv:2508.00553v3 Announce Type: replace Abstract: Vision-Language Models (VLMs) encode images and videos into abundant tokens, which contain substantial redundancy and computation cost. While visual

researcharxiv-cs-cv
21 Apr 2026
Model Releases

HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution

DGX agent

arXiv:2604.17745v1 Announce Type: new Abstract: Recent advances in large language models have highlighted their potential to automate computational research, particularly reproducing experimental resu

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Horospherical Depth and Busemann Median on Hadamard Manifolds

DGX agent

arXiv:2604.18242v1 Announce Type: cross Abstract: We introduce the horospherical depth, an intrinsic notion of statistical depth on Hadamard manifolds, and define the Busemann median as the set of its

researcharxiv-cs-lg
21 Apr 2026
Research

How Much Cache Does Reasoning Need? Depth-Cache Tradeoffs in KV-Compressed Transformers

DGX agent

arXiv:2604.17935v1 Announce Type: new Abstract: The key-value (KV) cache is the dominant memory bottleneck during Transformer inference, yet little is known theoretically about how aggressively it can

researcharxiv-cs-lg
21 Apr 2026
Research

How Much Data is Enough? The Zeta Law of Discoverability in Biomedical Data, featuring the enigmatic Riemann zeta function

DGX agent

arXiv:2604.17581v1 Announce Type: new Abstract: How much data is enough to make a scientific discovery? As biomedical datasets scale to millions of samples and AI models grow in capacity, progress inc

researcharxiv-cs-lg
21 Apr 2026
Research

HPLT 3.0: Very Large-Scale Multilingual Resources for LLMs and MT. Mono- and Bi-lingual Data, Multilingual Evaluation, and Pre-Trained Models

DGX agent

arXiv:2511.01066v3 Announce Type: replace Abstract: We present an ongoing initiative to provide open, very large, high-quality, and richly annotated textual datasets for almost 200 languages. At 30 tr

researcharxiv-cs-cl
21 Apr 2026
Research

HyKey: Hyperspectral Keypoint Detection and Matching in Minimally Invasive Surgery

DGX agent

arXiv:2604.17446v1 Announce Type: new Abstract: Purpose: 3D reconstruction in minimally invasive surgery (MIS) enables enhanced surgical guidance through improved visualisation, tool tracking, and aug

researcharxiv-cs-cv
21 Apr 2026
Research

Hyperspectral Unmixing Hierarchies

DGX agent

arXiv:2604.16969v1 Announce Type: new Abstract: Unmixing reveals the spatial distribution and spectral details of different constituents, called endmembers, in a hyperspectral image. Because unmixing

researcharxiv-cs-cv
21 Apr 2026
Hardware

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, an…

DGX agent

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, and provide a comprehensive tutorial in a Jupyter Notebook fil

hardwareclem-delangue--x
21 Apr 2026
Research

ICLAD: In-Context Learning with Comparison-Guidance for Audio Deepfake Detection

DGX agent

arXiv:2604.16749v1 Announce Type: cross Abstract: Audio deepfakes pose a significant security threat, yet current state-of-the-art (SOTA) detection systems do not generalize well to realistic in-the-w

researcharxiv-cs-cl
21 Apr 2026
Research

I’m hearing there’s renewed lobbying in DC and in state legislatures to ban or severely restrict open-source. Like a few years ago, we’ll ne…

DGX agent

I’m hearing there’s renewed lobbying in DC and in state legislatures to ban or severely restrict open-source. Like a few years ago, we’ll need everyone to help show policymakers why open-source matter

researchyann-lecun--x
21 Apr 2026
Research

IMA-MoE: An Interpretable Modality-Aware Mixture-of-Experts Framework for Characterizing the Neurobiological Signatures of Binge Eating Disorder

DGX agent

arXiv:2604.17028v1 Announce Type: new Abstract: Binge eating disorder (BED) is the most prevalent eating disorder. However, current diagnostic frameworks remain largely grounded in symptom-based crite

researcharxiv-cs-cv
21 Apr 2026
Research

ImpRIF: Stronger Implicit Reasoning Leads to Better Complex Instruction Following

DGX agent

arXiv:2602.21228v2 Announce Type: replace Abstract: As applications of large language models (LLMs) become increasingly complex, the demand for robust complex instruction following capabilities is gro

researcharxiv-cs-cl
21 Apr 2026
Research

Improving Radio Interferometry Imaging by Explicitly Modeling Cross-Domain Consistency in Reconstruction

DGX agent

arXiv:2604.16794v1 Announce Type: new Abstract: Radio astronomy plays a crucial role in understanding the universe, particularly within the realm of non-thermal astrophysics. Images of celestial objec

researcharxiv-cs-cv
21 Apr 2026
Research

Improving reproducibility by controlling random seed stability in machine learning based estimation via bagging

DGX agent

arXiv:2604.17694v1 Announce Type: cross Abstract: Predictions from machine learning algorithms can vary across random seeds, inducing instability in downstream debiased machine learning estimators. We

researcharxiv-cs-lg
21 Apr 2026
Research

Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context

DGX agent

arXiv:2506.10779v2 Announce Type: replace Abstract: Classroom speech and lectures often contain named entities (NEs) such as names of people and special terminology. While automatic speech recognition

researcharxiv-cs-cl
21 Apr 2026
Research

In Search of Lost DNA Sequence Pretraining

DGX agent

arXiv:2604.16570v1 Announce Type: new Abstract: DNA sequence encoding is fundamental to gene function prediction, protein synthesis, and diverse downstream biological tasks. Despite the substantial pr

researcharxiv-cs-lg
21 Apr 2026
Research

In Situ Training of Implicit Neural Compressors for Scientific Simulations via Sketch-Based Regularization

DGX agent

arXiv:2511.02659v3 Announce Type: replace Abstract: Focusing on implicit neural representations, we present a novel in situ training protocol that employs limited memory buffers of full and sketched d

researcharxiv-cs-lg
21 Apr 2026
Research

Incentivizing Parametric Knowledge via Reinforcement Learning with Verifiable Rewards for Cross-Cultural Entity Translation

DGX agent

arXiv:2604.16881v1 Announce Type: new Abstract: Cross-cultural entity translation remains challenging for large language models (LLMs) as literal or phonetic renderings are usually yielded instead of

researcharxiv-cs-cl
21 Apr 2026
Research

IncepDeHazeGAN: Novel Satellite Image Dehazing

DGX agent

arXiv:2604.16609v1 Announce Type: new Abstract: Dehazing is a technique in computer vision for enhancing the visual quality of images captured in cloudy or foggy conditions. Dehazing helps to recover

researcharxiv-cs-cv
21 Apr 2026
Research

Inductive Convolution Nuclear Norm Minimization for Tensor Completion with Arbitrary Sampling

DGX agent

arXiv:2604.17001v1 Announce Type: new Abstract: The recently established Convolution Nuclear Norm Minimization (CNNM) addresses the problem of extit{tensor completion with arbitrary sampling} (TCAS),

researcharxiv-cs-cv
21 Apr 2026
Research

Inference-Time Temporal Probability Smoothing for Stable Video Segmentation with SAM2 under Weak Prompts

DGX agent

arXiv:2604.17115v1 Announce Type: new Abstract: Interactive video segmentation models such as SAM2 have demonstrated strong generalization across diverse visual domains. However, under weak user super

researcharxiv-cs-cv
21 Apr 2026
Research

Instant Colorization of Gaussian Splats

DGX agent

arXiv:2604.17155v1 Announce Type: new Abstract: Gaussian Splatting has recently become one of the most popular frameworks for photorealistic 3D scene reconstruction and rendering. While current raster

researcharxiv-cs-cv
21 Apr 2026
Research

💫 Introducing NeuralSet: a simple, fast, scalable Python package for Neuro-AI 📦 pip install neuralset 📄 https://kingjr.github.io/files/ne…

DGX agent

💫 Introducing NeuralSet: a simple, fast, scalable Python package for Neuro-AI 📦 pip install neuralset 📄 https://kingjr.github.io/files/neuralset.pdf 🔍 https://facebookresearch.github.io/neuroai/neural

researchyann-lecun--x
21 Apr 2026
Research

Inventor recalls eye imaging breakthrough

DGX agent

If you’ve been to an eye doctor and had an image taken of the inside of your eye, chances are good it was done with optical coherence tomography (OCT)—a technology invented by clinician-scientist Davi

researchmit-tech-review
21 Apr 2026
Research

iPhoneme: Brain-to-Text Communication for ALS Using ConformerXL Decoding

DGX agent

arXiv:2604.16441v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) for speech restoration hold transformative potential for the approximately 173,000--232,500 individuals worldwide wit

researcharxiv-cs-cl
21 Apr 2026
Research

Is SAM3 ready for pathology segmentation?

DGX agent

arXiv:2604.18225v1 Announce Type: new Abstract: Is Segment Anything Model 3 (SAM3) capable in segmenting Any Pathology Images? Digital pathology segmentation spans tissue-level and nuclei-level scales

researcharxiv-cs-cv
21 Apr 2026
← Previous
1…394395396397398…540
Next →