AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,881 results
21 Apr 2026

Frequency-guided Multi-level Reasoning for Scene Graph Generation in Video

ResearchDGX agent

arXiv:2604.17298v1 Announce Type: new Abstract: Video Scene Graph Generation aims to obtain structured semantic representations of objects and their relationships in videos for high-level understandin

From 2:4 to 8:16 sparsity patterns in LLMs for Outliers and Weights with Variance Correction

ResearchDGX agent

arXiv:2507.03052v2 Announce Type: replace Abstract: As large language models (LLMs) grow in size, efficient compression techniques like quantization and sparsification are critical. While quantization

From Attribution to Abstention: Training-Free Attention-Based Auditing for Clinical Summarization

ResearchDGX agent

arXiv:2601.16397v2 Announce Type: replace Abstract: Deploying multimodal large language models (MLLMs) for clinical summarization demands not only fluent generation but also transparency about where e

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

From Classical to Quantum: Extending Prometheus for Unsupervised Discovery of Phase Transitions in Three Dimensions and Quantum Systems

ResearchDGX agent

arXiv:2602.14928v4 Announce Type: replace-cross Abstract: We extend the Prometheus framework for unsupervised phase transition discovery from two-dimensional classical systems to three-dimensional cla

From Domains to Instances: Dual-Granularity Data Synthesis for LLM Unlearning

ResearchDGX agent

arXiv:2601.04278v2 Announce Type: replace Abstract: Although machine unlearning is essential for removing private, harmful, or copyrighted content from LLMs, current benchmarks often fail to faithfull

From Fallback to Frontline: When Can LLMs be Superior Annotators of Human Perspectives?

ResearchDGX agent

arXiv:2604.17968v1 Announce Type: cross Abstract: Although large language models (LLMs) are increasingly used as annotators at scale, they are typically treated as a pragmatic fallback rather than a f

From Implicit to Explicit: Token-Efficient Logical Supervision for Mathematical Reasoning in LLMs

ResearchDGX agent

arXiv:2601.03682v2 Announce Type: replace Abstract: Recent studies reveal that large language models (LLMs) exhibit limited logical reasoning abilities in mathematical problem-solving, instead often r

From User Recognition to Activity Counting: An Identity-Agnostic Approach to Multi-User WiFi Sensing

ResearchDGX agent

arXiv:2604.16572v1 Announce Type: new Abstract: Wi-Fi Channel State Information (CSI) enables device-free human activity recognition, but existing multi-user approaches assume a fixed set of known use

From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents

ResearchDGX agent

arXiv:2603.01455v2 Announce Type: replace-cross Abstract: While multimodal large language models have demonstrated impressive short-term reasoning, they struggle with long-horizon video understanding

G-PARC: Graph-Physics Aware Recurrent Convolutional Neural Networks for Spatiotemporal Dynamics on Unstructured Meshes

ResearchDGX agent

arXiv:2604.16533v1 Announce Type: new Abstract: Physics-aware recurrent convolutional networks (PARC) have demonstrated strong performance in predicting nonlinear spatiotemporal dynamics by embedding

GaLa: Hypergraph-Guided Visual Language Models for Procedural Planning

ResearchDGX agent

arXiv:2604.17241v1 Announce Type: new Abstract: Implicit spatial relations and deep semantic structures encoded in object attributes are crucial for procedural planning in embodied AI systems. However

GeGS-PCR: Effective and Robust 3D Point Cloud Registration with Two-Stage Color-Enhanced Geometric-3DGS Fusion

ResearchDGX agent

arXiv:2604.17721v1 Announce Type: new Abstract: We address the challenge of point cloud registration using color information, where traditional methods relying solely on geometric features often strug

Geometry-Guided 3D Visual Token Pruning for Video-Language Models

ResearchDGX agent

arXiv:2604.18260v1 Announce Type: new Abstract: Multimodal large language models have demonstrated remarkable capabilities in 2D vision, motivating their extension to 3D scene understanding. Recent st

Get ready for hotter, muggier, stormier summers

ResearchDGX agent

A long stretch of humid heat followed by a powerful thunderstorm is a familiar weather pattern in the tropics, but it’s also becoming more common in midlatitude regions such as the US Midwest. A recen

GoCoMA: Hyperbolic Multimodal Representation Fusion for Large Language Model-Generated Code Attribution

ResearchDGX agent

arXiv:2604.16377v1 Announce Type: new Abstract: Large Language Models (LLMs) trained on massive code corpora are now increasingly capable of generating code that is hard to distinguish from human-writ

Graph Neural Networks for Graphs with Heterophily: A Survey

ApplicationsDGX agent

arXiv:2202.07082v4 Announce Type: replace Abstract: Recent years have witnessed fast developments of graph neural networks (GNNs) that have benefited myriad graph analytic tasks and applications. Most

HABIT: Chrono-Synergia Robust Progressive Learning Framework for Composed Image Retrieval

ResearchDGX agent

arXiv:2604.18037v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) is a flexible image retrieval paradigm that enables users to accurately locate the target image through a multimodal quer

Hard to Be Heard: Phoneme-Level ASR Analysis of Phonologically Complex, Low-Resource Endangered Languages

ResearchDGX agent

arXiv:2604.18204v1 Announce Type: new Abstract: We present a phoneme-level analysis of automatic speech recognition (ASR) for two low-resourced and phonologically complex East Caucasian languages, Arc

HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference

ResearchDGX agent

arXiv:2601.13684v2 Announce Type: replace Abstract: The linear memory growth of the KV cache poses a significant bottleneck for LLM inference in long-context tasks. Existing static compression methods

HiPrune: Hierarchical Attention for Efficient Token Pruning in Vision-Language Models

ResearchDGX agent

arXiv:2508.00553v3 Announce Type: replace Abstract: Vision-Language Models (VLMs) encode images and videos into abundant tokens, which contain substantial redundancy and computation cost. While visual

HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution

Model ReleasesDGX agent

arXiv:2604.17745v1 Announce Type: new Abstract: Recent advances in large language models have highlighted their potential to automate computational research, particularly reproducing experimental resu

Horospherical Depth and Busemann Median on Hadamard Manifolds

ResearchDGX agent

arXiv:2604.18242v1 Announce Type: cross Abstract: We introduce the horospherical depth, an intrinsic notion of statistical depth on Hadamard manifolds, and define the Busemann median as the set of its

How Much Cache Does Reasoning Need? Depth-Cache Tradeoffs in KV-Compressed Transformers

ResearchDGX agent

arXiv:2604.17935v1 Announce Type: new Abstract: The key-value (KV) cache is the dominant memory bottleneck during Transformer inference, yet little is known theoretically about how aggressively it can

How Much Data is Enough? The Zeta Law of Discoverability in Biomedical Data, featuring the enigmatic Riemann zeta function

ResearchDGX agent

arXiv:2604.17581v1 Announce Type: new Abstract: How much data is enough to make a scientific discovery? As biomedical datasets scale to millions of samples and AI models grow in capacity, progress inc

HPLT 3.0: Very Large-Scale Multilingual Resources for LLMs and MT. Mono- and Bi-lingual Data, Multilingual Evaluation, and Pre-Trained Models

ResearchDGX agent

arXiv:2511.01066v3 Announce Type: replace Abstract: We present an ongoing initiative to provide open, very large, high-quality, and richly annotated textual datasets for almost 200 languages. At 30 tr

HyKey: Hyperspectral Keypoint Detection and Matching in Minimally Invasive Surgery

ResearchDGX agent

arXiv:2604.17446v1 Announce Type: new Abstract: Purpose: 3D reconstruction in minimally invasive surgery (MIS) enables enhanced surgical guidance through improved visualisation, tool tracking, and aug

Hyperspectral Unmixing Hierarchies

ResearchDGX agent

arXiv:2604.16969v1 Announce Type: new Abstract: Unmixing reveals the spatial distribution and spectral details of different constituents, called endmembers, in a hyperspectral image. Because unmixing

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, an…

HardwareDGX agent

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, and provide a comprehensive tutorial in a Jupyter Notebook fil

ICLAD: In-Context Learning with Comparison-Guidance for Audio Deepfake Detection

ResearchDGX agent

arXiv:2604.16749v1 Announce Type: cross Abstract: Audio deepfakes pose a significant security threat, yet current state-of-the-art (SOTA) detection systems do not generalize well to realistic in-the-w

I’m hearing there’s renewed lobbying in DC and in state legislatures to ban or severely restrict open-source. Like a few years ago, we’ll ne…

ResearchDGX agent

I’m hearing there’s renewed lobbying in DC and in state legislatures to ban or severely restrict open-source. Like a few years ago, we’ll need everyone to help show policymakers why open-source matter

IMA-MoE: An Interpretable Modality-Aware Mixture-of-Experts Framework for Characterizing the Neurobiological Signatures of Binge Eating Disorder

ResearchDGX agent

arXiv:2604.17028v1 Announce Type: new Abstract: Binge eating disorder (BED) is the most prevalent eating disorder. However, current diagnostic frameworks remain largely grounded in symptom-based crite

ImpRIF: Stronger Implicit Reasoning Leads to Better Complex Instruction Following

ResearchDGX agent

arXiv:2602.21228v2 Announce Type: replace Abstract: As applications of large language models (LLMs) become increasingly complex, the demand for robust complex instruction following capabilities is gro

Improving Radio Interferometry Imaging by Explicitly Modeling Cross-Domain Consistency in Reconstruction

ResearchDGX agent

arXiv:2604.16794v1 Announce Type: new Abstract: Radio astronomy plays a crucial role in understanding the universe, particularly within the realm of non-thermal astrophysics. Images of celestial objec

Improving reproducibility by controlling random seed stability in machine learning based estimation via bagging

ResearchDGX agent

arXiv:2604.17694v1 Announce Type: cross Abstract: Predictions from machine learning algorithms can vary across random seeds, inducing instability in downstream debiased machine learning estimators. We

Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context

ResearchDGX agent

arXiv:2506.10779v2 Announce Type: replace Abstract: Classroom speech and lectures often contain named entities (NEs) such as names of people and special terminology. While automatic speech recognition

In Search of Lost DNA Sequence Pretraining

ResearchDGX agent

arXiv:2604.16570v1 Announce Type: new Abstract: DNA sequence encoding is fundamental to gene function prediction, protein synthesis, and diverse downstream biological tasks. Despite the substantial pr

In Situ Training of Implicit Neural Compressors for Scientific Simulations via Sketch-Based Regularization

ResearchDGX agent

arXiv:2511.02659v3 Announce Type: replace Abstract: Focusing on implicit neural representations, we present a novel in situ training protocol that employs limited memory buffers of full and sketched d

Incentivizing Parametric Knowledge via Reinforcement Learning with Verifiable Rewards for Cross-Cultural Entity Translation

ResearchDGX agent

arXiv:2604.16881v1 Announce Type: new Abstract: Cross-cultural entity translation remains challenging for large language models (LLMs) as literal or phonetic renderings are usually yielded instead of

IncepDeHazeGAN: Novel Satellite Image Dehazing

ResearchDGX agent

arXiv:2604.16609v1 Announce Type: new Abstract: Dehazing is a technique in computer vision for enhancing the visual quality of images captured in cloudy or foggy conditions. Dehazing helps to recover

Inductive Convolution Nuclear Norm Minimization for Tensor Completion with Arbitrary Sampling

ResearchDGX agent

arXiv:2604.17001v1 Announce Type: new Abstract: The recently established Convolution Nuclear Norm Minimization (CNNM) addresses the problem of extit{tensor completion with arbitrary sampling} (TCAS),

Inference-Time Temporal Probability Smoothing for Stable Video Segmentation with SAM2 under Weak Prompts

ResearchDGX agent

arXiv:2604.17115v1 Announce Type: new Abstract: Interactive video segmentation models such as SAM2 have demonstrated strong generalization across diverse visual domains. However, under weak user super

Instant Colorization of Gaussian Splats

ResearchDGX agent

arXiv:2604.17155v1 Announce Type: new Abstract: Gaussian Splatting has recently become one of the most popular frameworks for photorealistic 3D scene reconstruction and rendering. While current raster

💫 Introducing NeuralSet: a simple, fast, scalable Python package for Neuro-AI 📦 pip install neuralset 📄 https://kingjr.github.io/files/ne…

ResearchDGX agent

💫 Introducing NeuralSet: a simple, fast, scalable Python package for Neuro-AI 📦 pip install neuralset 📄 https://kingjr.github.io/files/neuralset.pdf 🔍 https://facebookresearch.github.io/neuroai/neural

Inventor recalls eye imaging breakthrough

ResearchDGX agent

If you’ve been to an eye doctor and had an image taken of the inside of your eye, chances are good it was done with optical coherence tomography (OCT)—a technology invented by clinician-scientist Davi

iPhoneme: Brain-to-Text Communication for ALS Using ConformerXL Decoding

ResearchDGX agent

arXiv:2604.16441v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) for speech restoration hold transformative potential for the approximately 173,000--232,500 individuals worldwide wit

Is SAM3 ready for pathology segmentation?

ResearchDGX agent

arXiv:2604.18225v1 Announce Type: new Abstract: Is Segment Anything Model 3 (SAM3) capable in segmenting Any Pathology Images? Digital pathology segmentation spans tissue-level and nuclei-level scales

Joint Distillation for Fast Likelihood Evaluation and Sampling in Flow-based Models

ResearchDGX agent

arXiv:2512.02636v3 Announce Type: replace-cross Abstract: Log-likelihood evaluation enables important capabilities in generative models, including model comparison, certain fine-tuning objectives, and

KaLDeX: Kalman Filter based Linear Deformable Cross Attention for Retina Vessel Segmentation

ResearchDGX agent

arXiv:2410.21160v2 Announce Type: replace-cross Abstract: Background and Objective: In the realm of ophthalmic imaging, accurate vascular segmentation is paramount for diagnosing and managing various

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is…

HardwareDGX agent

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is just scratching the surface. Ultimately, it boils down to g

Kevin Warsh just lost me. He argues he's going to be an independent Fed Chair, but refuses to acknowledge that Trump lost the 2020 election.…

ResearchDGX agent

Kevin Warsh just lost me. He argues he's going to be an independent Fed Chair, but refuses to acknowledge that Trump lost the 2020 election. If you can't state simple facts when you're in the politica

L1 Regularization Paths in Linear Models by Parametric Gaussian Message Passing

ResearchDGX agent

arXiv:2604.16949v1 Announce Type: new Abstract: The paper considers the computation of L1 regularization paths in a state space setting, which includes L1 regularized Kalman smoothing, linear SVM, LAS

LASER: Low-Rank Activation SVD for Efficient Recursion

ResearchDGX agent

arXiv:2604.17224v1 Announce Type: new Abstract: Recursive architectures such as Tiny Recursive Models (TRMs) perform implicit reasoning through iterative latent computation, yet the geometric structur

Latent Abstraction for Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2604.17866v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become a standard approach for enhancing large language models (LLMs) with external knowledge, mitigating hallu

Latent-Compressed Variational Autoencoder for Video Diffusion Models

ResearchDGX agent

arXiv:2604.16479v1 Announce Type: new Abstract: Video variational autoencoders (VAEs) used in latent diffusion models typically require a sufficiently large number of latent channels to ensure high-qu

Latent Phase-Shift Rollback: Inference-Time Error Correction via Residual Stream Monitoring and KV-Cache Steering

ResearchDGX agent

arXiv:2604.18567v1 Announce Type: cross Abstract: Large language models frequently commit unrecoverable reasoning errors mid-generation: once a wrong step is taken, subsequent tokens compound the mist

LBFTI: Layer-Based Facial Template Inversion for Identity-Preserving Fine-Grained Face Reconstruction

ResearchDGX agent

arXiv:2604.18358v1 Announce Type: new Abstract: In face recognition systems, facial templates are widely adopted for identity authentication due to their compliance with the data minimization principl

Learning residue level protein dynamics with multiscale Gaussians

ResearchDGX agent

arXiv:2509.01038v2 Announce Type: replace-cross Abstract: Many methods have been developed to predict static protein structures, however understanding the dynamics of protein structure is essential fo

Learning the Riccati solution operator for time-varying LQR via Deep Operator Networks

ResearchDGX agent

arXiv:2604.18507v1 Announce Type: cross Abstract: We propose a computational framework for replacing the repeated numerical solution of differential Riccati equations in finite-horizon Linear Quadrati

Learning to Correct: Calibrated Reinforcement Learning for Multi-Attempt Chain-of-Thought

ResearchDGX agent

arXiv:2604.17912v1 Announce Type: new Abstract: State-of-the-art reasoning models utilize long chain-of-thought (CoT) to solve increasingly complex problems using more test-time computation. In this w

Learning Unanimously Acceptable Lotteries via Queries

ResearchDGX agent

arXiv:2604.17505v1 Announce Type: cross Abstract: Many high-stakes AI deployments proceed only if every stakeholder deems the system acceptable relative to their own minimum standard. With randomizati

← Previous
1…315316317318319…432
Next →