AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
19,193 results
11 May 2026

Goldilocks RL: Tuning Task Difficulty to Escape Sparse Rewards for Reasoning

ResearchDGX agent

arXiv:2602.14868v2 Announce Type: replace-cross Abstract: Reinforcement learning has emerged as a powerful paradigm for unlocking reasoning capabilities in language models. However, relying on sparse

GraphFusion3D: Dynamic Graph Attention Convolution with Adaptive Cross-Modal Transformer for 3D Object Detection

ResearchDGX agent

arXiv:2512.02991v2 Announce Type: replace Abstract: Despite significant progress in 3D object detection, point clouds remain challenging due to sparse data, incomplete structures, and limited semantic

GRASP -- Graph-Based Anomaly Detection Through Self-Supervised Classification

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.07812v1 Announce Type: cross Abstract: Advanced persistent threat (APT) attacks remain difficult to detect due to their stealth, adaptability, and use of legitimate system components. Prove

GRaSp: Automatic Example Optimization for In-Context Learning in Low-Data Tasks

ResearchDGX agent

arXiv:2605.07454v1 Announce Type: new Abstract: In-context learning enables large language models to adapt to new tasks, but their performance is highly sensitive to the selected examples. Finding eff

Great essay by Tobi. Building an AI-native company? Go read it now. I couldn't resist visualizing it with my artifact generator. Biggest tak…

ResearchDGX agent

Great essay by Tobi. Building an AI-native company? Go read it now. I couldn't resist visualizing it with my artifact generator. Biggest takeaway for me: 'The risk isn't that AI does the work. It's th

Hammer and Anvil: Toward a Theory of Backdoors in Federated Learning

ResearchDGX agent

arXiv:2509.08089v2 Announce Type: replace Abstract: Federated Learning (FL) enables distributed model training but is vulnerable to backdoor attacks, where malicious clients embed attacker-controlled

Hierarchical Perfusion Graphs for Tumor Heterogeneity Modeling in Glioma Molecular Subtyping

ResearchDGX agent

arXiv:2605.07156v1 Announce Type: new Abstract: Precise molecular subtyping of gliomas, including isocitrate dehydrogenase (IDH) mutation and 1p/19q codeletion, directly guides surgical and therapeuti

High-Fidelity Surface Splatting-Based 3D Reconstruction from Multi-View Images

ResearchDGX agent

arXiv:2605.07254v1 Announce Type: new Abstract: Multi-view mesh reconstruction remains a core challenge in computer graphics and vision, especially for recovering high-frequency geometry from sparse o

How Do Language Models Compose Functions?

ResearchDGX agent

arXiv:2510.01685v2 Announce Type: replace-cross Abstract: While large language models (LLMs) appear to be increasingly capable of solving compositional tasks, it is an open question whether they do so

How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem

ResearchDGX agent

arXiv:2605.06882v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved great improvements in recent years. Nevertheless, it still remains unclear how good LLMs are for reasoning ta

https://x.com/_alex_kirillov_/status/2053939325818294310

ResearchDGX agent

People talk, listen, watch, think, and collaborate at the same time, in real time. We've designed an AI that works with people the same way. We share our approach, early results, and a quick look at o

ICDAR 2026 Competition on Writer Identification and Pen Classification from Hand-Drawn Circles

ResearchDGX agent

arXiv:2605.07816v1 Announce Type: new Abstract: This paper presents CircleID, a large-scale ICDAR 2026 competition on writer identification and pen classification from scanned hand-drawn circles. The

Identifiability Challenges in Sparse Linear Ordinary Differential Equations

ResearchDGX agent

arXiv:2506.09816v3 Announce Type: replace Abstract: Dynamical systems modeling is a core pillar of scientific inquiry across natural and life sciences. Increasingly, dynamical system models are learne

ImplantMamba: Long-range Sequential Modeling Mamba For Dental Implant Position Prediction

ResearchDGX agent

arXiv:2605.07082v1 Announce Type: new Abstract: In the design of surgical guides for implant placement, determining the precise implant position is a critical step. However, the implant region itself

Improved Model-based Reinforcement Learning with Smooth Kernels

ResearchDGX agent

arXiv:2605.07218v1 Announce Type: new Abstract: For continuous state-action space scenarios, classical reinforcement learning (RL) theory predominantly focuses on low-rank Markov decision processes (M

In modern ML accelerators, FLOPS have absolutely exploded. Often though, the bottleneck is not FLOPS but memory bandwidth. Similarly, model …

ResearchDGX agent

In modern ML accelerators, FLOPS have absolutely exploded. Often though, the bottleneck is not FLOPS but memory bandwidth. Similarly, model intelligence has exploded, causing the bottleneck to be huma

Inference of Qualitative Models from Steady-State Data via Weighted MaxSMT

ResearchDGX agent

arXiv:2605.07433v1 Announce Type: cross Abstract: Qualitative models provide crucial instruments for modelling complex biological systems. While advances in automated reasoning and symbolic encodings

Information-theoretic Limits of Learning and Estimation

ResearchDGX agent

arXiv:2605.06710v1 Announce Type: cross Abstract: Information theory plays a central role in establishing fundamental limits on what any learning or estimation algorithm can -- and cannot -- achieve,

Innovation abounds in device charging

ResearchDGX agent

The changes may be less perceptible than in smartphones, tablets, or wearables, but chargers have also been quietly reinvented over the last decade. At one time a bulky mix of tangled cables and conne

INO-SGD: Addressing Utility Imbalance under Individualized Differential Privacy

ResearchDGX agent

arXiv:2605.07930v1 Announce Type: cross Abstract: Differential privacy (DP) is widely employed in machine learning to protect confidential or sensitive training data from being revealed. As data owner

InsHuman: Towards Natural and Identity-Preserving Human Insertion

ResearchDGX agent

arXiv:2605.07402v1 Announce Type: new Abstract: Human insertion aims to naturally place specific individuals into a target background. Although existing image editing models may have such ability, the

Interactive Jensen–Shannon Divergence Visualisation [P]

ResearchDGX agent

Jensen-Shannon divergence is a method of measuring the similarity between two probability distributions. It is based on the Kullback-Leibler divergence, with the notable difference that it is symmetri

Interpreting Speaker Characteristics in the Dimensions of Self-Supervised Speech Features

ResearchDGX agent

arXiv:2603.03096v2 Announce Type: replace-cross Abstract: How do speech models trained through self-supervised learning structure their representations? Previous studies have looked at how information

Into the Rabbit Hull: From Task-Relevant Concepts in DINO to Minkowski Geometry

ResearchDGX agent

arXiv:2510.08638v3 Announce Type: replace-cross Abstract: DINOv2 is routinely deployed to recognize objects, scenes, and actions; yet the nature of what it perceives remains unknown. As a working base

Is Chain-of-Thought Really Not Explainability? Chain-of-Thought Can Be Faithful without Hint Verbalization

ResearchDGX agent

arXiv:2512.23032v2 Announce Type: replace-cross Abstract: Recent work, using the Biasing Features metric, labels a CoT as unfaithful if it omits a prompt-injected hint that affected the prediction. We

It Just Takes Two: Scaling Amortized Inference to Large Sets

ResearchDGX agent

arXiv:2605.07972v1 Announce Type: cross Abstract: Neural posterior estimation has emerged as a powerful tool for amortized inference, with growing adoption across scientific and applied domains. In ma

Kernel Selection is Model Selection: A Unified Complexity-Penalized Approach for MMD Two-Sample Tests

ResearchDGX agent

arXiv:2605.06883v1 Announce Type: cross Abstract: The Maximum Mean Discrepancy (MMD) is a cornerstone statistic for nonparametric two-sample testing, but its test power is dictated entirely by the cho

Knowledge Transfer Scaling Laws for 3D Medical Imaging

ResearchDGX agent

arXiv:2605.06859v1 Announce Type: cross Abstract: Vision foundation models are increasingly moving beyond 2D to volumetric domains such as 3D medical imaging, where unified pretraining across differen

Koopman Autoencoders with Continuous-Time Latent Dynamics for Fluid Dynamics Forecasting

ResearchDGX agent

arXiv:2602.02832v3 Announce Type: replace Abstract: Forecasting physical systems over long horizons from irregularly sampled observations demands models that are stable, computationally efficient, and

Latent Order Bandits

ResearchDGX agent

arXiv:2605.07304v1 Announce Type: new Abstract: Bandit algorithms solve diverse sequential decision-making problems, but are often too sample-inefficient for from-scratch personalization. To substanti

Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models

ResearchDGX agent

arXiv:2602.01166v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models benefit from chain-of-thought (CoT) reasoning, but existing approaches incur high inference overhead and rely on

Latent-Space Causal Discovery from Indirect Neuroimaging Observations

ResearchDGX agent

arXiv:2602.09034v2 Announce Type: replace-cross Abstract: Neuroimaging does not observe causal variables directly: hemodynamics and volume conduction distort signals so that statistical dependence nee

LaTER: Efficient Test-Time Reasoning via Latent Exploration and Explicit Verification

ResearchDGX agent

arXiv:2605.07315v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning improves large language models (LLMs) on difficult tasks, but it also makes inference expensive because every intermedi

Learning Image-Adaptive Scale Fields for Metric Depth Recovery

ResearchDGX agent

arXiv:2605.07418v1 Announce Type: new Abstract: Monocular depth estimation (MDE) typically produces depth estimations that are defined up to an unknown scale or shift. When only sparse metric anchors

Learning Minimal-Deviation Corrections for Multi-Dimensional Mismodelling in HEP Simulations

ResearchDGX agent

arXiv:2605.07460v1 Announce Type: new Abstract: Accurate Monte Carlo (MC) modelling in high-energy physics is challenging, particularly in complex scenarios where simulations fail to reproduce observe

Learning to Pose Problems: Reasoning-Driven and Solver-Adaptive Data Synthesis

ResearchDGX agent

arXiv:2511.09907v5 Announce Type: replace Abstract: Data synthesis for training large reasoning models offers a scalable alternative to limited, human-curated datasets, enabling the creation of high-q

LENS: Low-Frequency Eigen Noise Shaping for Efficient Diffusion Sampling

ResearchDGX agent

arXiv:2605.07253v1 Announce Type: new Abstract: Distilled diffusion models accelerate image generation by reducing the number of denoising steps, but often suffer from degraded image quality. To mitig

LensVLM: Selective Context Expansion for Compressed Visual Representation of Text

ResearchDGX agent

arXiv:2605.07019v1 Announce Type: cross Abstract: Vision Language Models (VLMs) offer the exciting possibility of processing text as rendered images, bypassing the need for tokenizing the text into lo

Less Random, More Private: What is the Optimal Subsampling Scheme for DP-SGD?

ResearchDGX agent

arXiv:2605.07072v1 Announce Type: new Abstract: Poisson subsampling is the default sampling scheme in differentially private machine learning, largely because its unstructured randomness yields tracta

Limitations on Accurate, Trusted, Human-level Reasoning

ResearchDGX agent

arXiv:2509.21654v2 Announce Type: replace-cross Abstract: We identify a fundamental incompatibility between the goals of accuracy, trust, and human-level reasoning in artificial intelligence (AI) syst

Linear Response Estimators for Singular Statistical Models

ResearchDGX agent

arXiv:2605.07970v1 Announce Type: cross Abstract: We define susceptibilities as a measure of the response of an observable quantity of a parameterized statistical model to a perturbation of the data f

LLMs are not (consistently) Bayesian: Quantifying internal (in)consistencies of LLMs' probabilistic beliefs

ResearchDGX agent

arXiv:2605.06915v1 Announce Type: new Abstract: Modern AI systems are being deployed in complex domains such as medicine, science, and law, where it is important that they not only produce correct ans

Lossy Common Information in a Learnable Gray-Wyner Network

ResearchDGX agent

arXiv:2601.21424v3 Announce Type: replace-cross Abstract: Many computer vision tasks share substantial overlapping information, yet conventional codecs tend to ignore this, leading to redundant and in

LR-SGS: Robust LiDAR-Reflectance-Guided Salient Gaussian Splatting for Self-Driving Scene Reconstruction

ResearchDGX agent

arXiv:2603.12647v2 Announce Type: replace-cross Abstract: Recent 3D Gaussian Splatting (3DGS) methods have demonstrated the feasibility of self-driving scene reconstruction and novel view synthesis. H

MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference

ResearchDGX agent

arXiv:2605.05225v2 Announce Type: replace-cross Abstract: Mixture-of-Experts Multimodal Large Language Models (MoE MLLMs) suffer from a significant efficiency bottleneck during Expert Parallelism (EP)

Making AI Evaluation Deployment Relevant Through Context Specification

ResearchDGX agent

arXiv:2603.06811v3 Announce Type: replace Abstract: With many organizations struggling to gain value from AI deployments, pressure to evaluate AI in an informed manner has intensified. Status quo AI e

MAST: A Multi-fidelity Augmented Surrogate model via Spatial Trust-weighting

ResearchDGX agent

arXiv:2602.20974v2 Announce Type: replace Abstract: In engineering design and scientific computing, computational cost and predictive accuracy are intrinsically coupled. High-fidelity simulations prov

Mean Mode Screaming: Mean--Variance Split Residuals for 1000-Layer Diffusion Transformers

ResearchDGX agent

arXiv:2605.06169v1 Announce Type: cross Abstract: Scaling Diffusion Transformers (DiTs) to hundreds of layers introduces a structural vulnerability: networks can enter a silent, mean-dominated collaps

Measuring and Mitigating the Distributional Gap Between Real and Simulated User Behaviors

ResearchDGX agent

arXiv:2605.07847v1 Announce Type: new Abstract: As user simulators are increasingly used for interactive training and evaluation of AI assistants, it is essential that they represent the diverse behav

Mechanistic Interpretability with Sparse Autoencoder Neural Operators

ResearchDGX agent

arXiv:2509.03738v4 Announce Type: replace-cross Abstract: We introduce sparse autoencoder neural operators (SAE-NOs), a new class of sparse autoencoders that operate in function spaces rather than fix

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models

ResearchDGX agent

arXiv:2605.07721v1 Announce Type: cross Abstract: Recurrent LLM architectures have emerged as a promising approach for improving reasoning, as they enable multi-step computation in the embedding space

Minerva: Reinforcement Learning with Verifiable Rewards for Cyber Threat Intelligence LLMs

ResearchDGX agent

arXiv:2602.00513v3 Announce Type: replace Abstract: Cyber threat intelligence (CTI) analysts routinely convert noisy, unstructured security artifacts into standardized, automation-ready representation

Minimizing Modality Gap from the Input Side: Your Speech LLM Can Be a Prosody-Aware Text LLM

ResearchDGX agent

arXiv:2605.05927v2 Announce Type: replace Abstract: Speech large language models (SLMs) are typically built from text large language model (TLM) checkpoints, yet they still suffer from a substantial m

MIT and IBM renew their research collaboration

ResearchDGX agent

MIT and IBM have renewed their long-standing research collaboration, focusing on advancing computing research and technology development. The partnership brings together expertise from both institutio

Mixture of Masters: Sparse Chess Language Models with Player Routing

ResearchDGX agent

arXiv:2602.04447v2 Announce Type: replace-cross Abstract: Modern chess language models are dense transformers trained on millions of games played by thousands of high-rated individuals. However, these

Motion-o: Trajectory-Grounded Video Reasoning

ResearchDGX agent

arXiv:2603.18856v2 Announce Type: replace-cross Abstract: Recent video reasoning models increasingly produce spatio-temporal evidence chains that localize objects at specific timestamps. While these t

Multi-Dimensional Evaluation of LLMs for Grammatical Error Correction

ResearchDGX agent

arXiv:2605.07635v1 Announce Type: new Abstract: Automated assistants for Grammatical Error Correction are now embedded in educational platforms serving millions of learners, yet three critical gaps re

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation

ResearchDGX agent

arXiv:2411.16748v5 Announce Type: replace Abstract: Long-duration talking video synthesis faces enduring challenges in achieving high video quality, portrait consistency, temporal coherence, and compu

Multimodal Latent Reasoning via Hierarchical Visual Cues Injection

ResearchDGX agent

arXiv:2602.05359v2 Announce Type: replace Abstract: The advancement of multimodal large language models (MLLMs) has enabled impressive perception capabilities. However, their reasoning process often r

Multimodal Stepwise Clinically-Guided Attention Learning for Pathological Complete Response Prediction in Breast Cancer

ResearchDGX agent

arXiv:2605.07561v1 Announce Type: new Abstract: Pathological complete response (pCR) is a key prognostic factor in breast cancer patients undergoing neoadjuvant therapy, strongly associated with long-

← Previous
1…231232233234235…320
Next →