AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,885 results
4 May 2026

Scaling Reasoning Hop Exposes Weaknesses: Demystifying and Improving Hop Generalization in Large Language Models

ResearchDGX agent

arXiv:2601.21214v2 Announce Type: replace Abstract: Chain-of-thought (CoT) reasoning has become the standard paradigm for enabling Large Language Models (LLMs) to solve complex problems. However, rece

Score-based Greedy Search for Structure Identification of Partially Observed Linear Causal Models

ResearchDGX agent

arXiv:2510.04378v2 Announce Type: replace Abstract: Identifying the structure of a partially observed causal system is essential to various scientific fields. Recent advances have focused on constrain

Short Chains, Deep Thoughts: Balancing Reasoning Efficiency and Intra-Segment Capability via Split-Merge Optimization

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.03141v3 Announce Type: replace Abstract: While Large Reasoning Models (LRMs) have demonstrated impressive capabilities in solving complex tasks through the generation of long reasoning chai

SketchGuard: Scaling Byzantine-Robust Decentralized Federated Learning via Sketch-Based Screening

ResearchDGX agent

arXiv:2510.07922v4 Announce Type: replace Abstract: Decentralized Federated Learning (DFL) enables privacy-preserving collaborative training without centralized servers but remains vulnerable to Byzan

Smart Ensemble Learning Framework for Predicting Groundwater Heavy Metal Pollution

ResearchDGX agent

arXiv:2605.00056v1 Announce Type: new Abstract: Groundwater in the Densu Basin is increasingly threatened by heavy metal contamination, but conventional methods fail to capture the statistical complex

Something from Nothing: Data Augmentation for Robust Severity Level Estimation of Dysarthric Speech

ResearchDGX agent

arXiv:2603.15988v2 Announce Type: replace-cross Abstract: Dysarthric speech quality assessment (DSQA) is critical for clinical diagnostics and inclusive speech technologies. However, subjective evalua

Sparse VideoGen2: Accelerate Video Generation with Sparse Attention via Semantic-Aware Permutation

ResearchDGX agent

arXiv:2505.18875v4 Announce Type: replace Abstract: Diffusion Transformers (DiTs) are essential for video generation but suffer from significant latency due to the quadratic complexity of attention. B

Spiking Sequence Machines and Transformers

ResearchDGX agent

arXiv:2605.00662v1 Announce Type: cross Abstract: Sequence learning reduces to similarity-based retrieval over a temporally indexed representation space, a constraint on any sequence model, not a prop

SPLICE: Latent Diffusion over JEPA Embeddings for Conformal Time-Series Inpainting

ResearchDGX agent

arXiv:2605.00126v1 Announce Type: new Abstract: Generative models for time-series imputation achieve strong reconstruction accuracy, yet provide no finite-sample reliability guarantees, a critical lim

Statistical Testing Framework for Clustering Pipelines by Selective Inference

ResearchDGX agent

arXiv:2603.18413v3 Announce Type: replace-cross Abstract: A data analysis pipeline is a structured sequence of steps that transforms raw data into meaningful insights by integrating multiple analysis

Stepper: Stepwise Immersive Scene Generation with Multiview Panoramas

ResearchDGX agent

arXiv:2603.28980v2 Announce Type: replace Abstract: The synthesis of immersive 3D scenes from text is rapidly maturing, driven by novel video generative models and feed-forward 3D reconstruction, with

Structural Prognostic Event Modeling for Multimodal Cancer Survival Analysis

ResearchDGX agent

arXiv:2512.01116v3 Announce Type: replace Abstract: The integration of histology images and gene profiles has shown great promise for improving survival prediction in cancer. However, current approach

Surprisingly High Redundancy in Electronic Structure Data Across Materials Explained by Low Intrinsic Dimensionality

ResearchDGX agent

arXiv:2507.09001v3 Announce Type: replace-cross Abstract: Machine learning (ML) models for electronic structure typically rely on large datasets generated by computationally expensive Kohn-Sham densit

Tailoring AI solutions for health care needs

ResearchDGX agent

The AI market is full of big promises of grand transformation. Health care is a prime target for those promises, beset as it is by financial pressures, labor shortages, and the growing burden of carin

Technical Report: Activation Residual Hessian Quantization (ARHQ) for Low-Bit LLM Quantization

ResearchDGX agent

arXiv:2605.00140v1 Announce Type: cross Abstract: We present Activation Residual Hessian Quantization (ARHQ), a post-training weight splitting method designed to mitigate error propagation in low-bit

Temporal Data Requirement for Predicting Unplanned Hospital Readmissions

ResearchDGX agent

arXiv:2605.00738v1 Announce Type: new Abstract: With the proliferation of Electronic Health Records (EHRs), a critical challenge in building predictive models is determining the optimal historical dat

The Algorithmic Gaze of Image Quality Assessment: An Audit and Trace Ethnography of the LAION-Aesthetics Predictor

ResearchDGX agent

arXiv:2601.09896v4 Announce Type: replace-cross Abstract: Visual generative AI models are trained using a one-size-fits-all measure of aesthetic appeal. However, what is deemed 'aesthetic' is inextric

The distillation panic

ResearchDGX agent

The article discusses concerns about the widespread use of knowledge distillation in AI development, where larger models' outputs are used to train smaller models, potentially creating a cycle of degr

The Seismic Wavefield Common Task Framework

ResearchDGX agent

arXiv:2512.19927v2 Announce Type: replace Abstract: Seismology faces fundamental challenges in state forecasting and reconstruction (e.g., earthquake early warning and ground motion prediction) and ma

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning

ResearchDGX agent

arXiv:2603.17837v3 Announce Type: replace-cross Abstract: During conversational interactions, humans subconsciously engage in concurrent thinking while listening to a speaker. Although this internal c

Thought Graph Traversal for Test-time Scaling in Chest X-ray VLLMs

ResearchDGX agent

arXiv:2506.11989v3 Announce Type: replace Abstract: Test-time scaling offers a promising way to improve the reasoning performance of vision-language large models (VLLMs) without additional training. I

Timing is Everything: Temporal Scaffolding of Semantic Surprise in Humor

ResearchDGX agent

arXiv:2605.00143v1 Announce Type: new Abstract: Humor is a fundamental cognitive phenomenon in which humans derive pleasure from the expectation violations and their resolution, exemplifying the brain

Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token Selection

ResearchDGX agent

arXiv:2602.03216v2 Announce Type: replace Abstract: The quadratic complexity of attention remains the central bottleneck in long-context inference for large language models. Prior acceleration methods

Trading off rewards and errors in multi-armed bandits

ResearchDGX agent

arXiv:2605.00488v1 Announce Type: new Abstract: In multi-armed bandits, the most-explored arms are the most informative, while reward maximization typically pulls only the best arm. We study the trade

Trees to Flows and Back: Unifying Decision Trees and Diffusion Models

ResearchDGX agent

arXiv:2605.00414v1 Announce Type: new Abstract: Decision trees and diffusion models are ostensibly disparate model classes, one discrete and hierarchical, the other continuous and dynamic. This work u

Trident: Improving Malware Detection with LLMs and Behavioral Features

ResearchDGX agent

arXiv:2605.00297v1 Announce Type: cross Abstract: Traditionally, machine learning methods for PE malware detection have relied on static features like byte histograms, string information, and PE heade

Two-View Accumulation as the Primary Training Lever for Hybrid-Capture Gaussian Splatting: A Variance-Decomposition View of When Gradient Surgery Helps

ResearchDGX agent

arXiv:2605.00052v1 Announce Type: new Abstract: Hybrid-capture novel view synthesis combines images at substantially different camera distances (e.g., aerial drone and ground-level views). Standard 3D

Uncertainty Modeling for Multi-Objective RTA Interception with Distillation Acceleration

ResearchDGX agent

arXiv:2511.05582v2 Announce Type: replace Abstract: Real-Time Auction (RTA) Interception aims to filter out invalid or irrelevant traffic to enhance the integrity and reliability of downstream data. H

Unlearning Offline Stochastic Multi-Armed Bandits

ResearchDGX agent

arXiv:2605.00638v1 Announce Type: new Abstract: Machine unlearning aims to unlearn data points from a learned model, offering a principled way to process data-deletion requests and mitigate privacy ri

VANISHING CULTURE is out now! From Internet Archive, this book looks at what is disappearing online. 🌐 Websites vanish 🗞️ News archives go…

ResearchDGX agent

VANISHING CULTURE is out now! From Internet Archive, this book looks at what is disappearing online. 🌐 Websites vanish 🗞️ News archives go offline 🎮 Games become unplayable 📼 Personal media breaks & b

VideoDetective: Clue Hunting via both Extrinsic Query and Intrinsic Relevance for Long Video Understanding

ResearchDGX agent

arXiv:2603.22285v2 Announce Type: replace Abstract: Long video understanding remains challenging for multimodal large language models (MLLMs) due to limited context windows, which necessitate identify

VQ-SAD: Vector Quantized Structure Aware Diffusion For Molecule Generation

ResearchDGX agent

arXiv:2605.00354v1 Announce Type: new Abstract: Many diffusion based molecule generation methods ignore the symbolic information of molecules and represent the atom and bond type as one hot representa

Week one of the Musk v. Altman trial: What it was like in the room

ResearchDGX agent

This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. Two of the most powerful people in AI—Sam Altman and Elon Musk

'What Are You Really Trying to Do?': Co-Creating Life Goals from Everyday Computer Use

ResearchDGX agent

arXiv:2605.00497v1 Announce Type: cross Abstract: Recent advances in user modeling make it feasible to conduct open-ended inference over a person's everyday computer use. Despite longstanding visions

When Structure Doesn't Help: LLMs Do Not Read Text-Attributed Graphs as Effectively as We Expected

ResearchDGX agent

arXiv:2511.16767v2 Announce Type: replace Abstract: Graphs provide a unified representation of semantic content and relational structure, making them a natural fit for domains such as molecular modeli

WildfireVLM: AI-powered Analysis for Early Wildfire Detection and Risk Assessment Using Satellite Imagery

ResearchDGX agent

arXiv:2602.13305v2 Announce Type: replace Abstract: Wildfires are a growing threat to ecosystems, human lives, and infrastructure, with their frequency and intensity rising due to climate change and h

3 May 2026

As always, more details and higher-res versions at https://sebastianraschka.com/llm-architecture-gallery/

ResearchDGX agent

Sebastian Raschka maintains a comprehensive gallery of LLM architecture diagrams and visualizations on his website, with higher-resolution versions available at sebastianraschka.com/llm-architecture-g

Here is a 2nd batch of April architecture drops. What a month! - Ant Ling 2.6 1T - Minimax M2.7 - Xiaomi MiMo V2.5 - Poolside Laguna XS.2 - …

ResearchDGX agent

This post documents a second wave of large language model and AI architecture releases from April, featuring updates from multiple organizations including Ant Ling 2.6 1T, Minimax M2.7, Xiaomi MiMo V2

If GitHub were built in: Japan 🇯🇵 China 🇨🇳 North Korea 🇰🇵 The EU 🇪🇺

ResearchDGX agent

This post is a humorous thought experiment comparing how GitHub's features, policies, and design would differ if the platform were developed in various countries and regions, each reflecting their dis

Releasing a skill to help build LLM Wikis.

ResearchDGX agent

DAIR.AI released a skill or tool designed to assist in building wikis powered by large language models (LLMs), likely enabling users to create structured knowledge bases with AI capabilities. The reso

Struggling with Chebyshev Filter Integration in CNN — Any Advice? [R]

ResearchDGX agent

This Reddit discussion covers technical challenges encountered when attempting to integrate Chebyshev filters—mathematical filters with equiripple characteristics used in signal processing—into convol

2 May 2026

2022: “Stop overreacting, they won’t overturn Roe.” They did. 2023: “Stop overreacting, they won’t let women die rather than get an abortion…

ResearchDGX agent

2022: “Stop overreacting, they won’t overturn Roe.” They did. 2023: “Stop overreacting, they won’t let women die rather than get an abortion.” They did. 2024: “Stop overreacting, they won’t arrest wom

GWの富山。 新緑の先に、まだ雪の立山連峰。

ResearchDGX agent

This post showcases Golden Week (GW) travel in Toyama Prefecture, featuring fresh spring greenery in the foreground contrasted with snow-capped peaks of the Tateyama Mountain Range visible in the dist

'If we are honest — and scientists have to be — we must admit that religion is a jumble of false assertions, with no basis in reality. The v…

ResearchDGX agent

'If we are honest — and scientists have to be — we must admit that religion is a jumble of false assertions, with no basis in reality. The very idea of God is a product of the human imagination. It is

Stanford's latest seminar is a deep dive into the evolution of world modeling in AI. Focuses on the shift in the world model from traditiona…

ResearchDGX agent

Stanford's latest seminar is a deep dive into the evolution of world modeling in AI. Focuses on the shift in the world model from traditional reconstruction methods toward latent space prediction. Cov

This is insane Three Trump judicial nominees refused, over and over, to say Joe Biden won the 2020 election. They either believe Trump’s lie…

ResearchDGX agent

This is insane Three Trump judicial nominees refused, over and over, to say Joe Biden won the 2020 election. They either believe Trump’s lie that the election was stolen, or they’re too afraid of him

1 May 2026

3D-ReGen: A Unified 3D Geometry Regeneration Framework

ResearchDGX agent

arXiv:2604.28134v1 Announce Type: new Abstract: We consider the problem of regenerating 3D objects from 2D images and initial 3D shapes. Most 3D generators operate in a one-shot fashion, converting te

A decision-theoretic approach to dealing with uncertainty in quantum mechanics

ResearchDGX agent

arXiv:2503.20607v3 Announce Type: replace-cross Abstract: We provide a decision-theoretic framework for dealing with uncertainty in quantum mechanics. This uncertainty is two-fold: on the one hand the

A Framework for Variational Inference of Lightweight Bayesian Neural Networks with Heteroscedastic Uncertainties

ResearchDGX agent

arXiv:2402.14532v2 Announce Type: replace Abstract: Obtaining heteroscedastic predictive uncertainties from a Bayesian Neural Network (BNN) is vital to many applications. Often, heteroscedastic aleato

A Gated Hybrid Contrastive Collaborative Filtering Recommendation

ResearchDGX agent

arXiv:2604.27117v1 Announce Type: cross Abstract: Recommender systems increasingly incorporate textual reviews to enrich user and item representations. However, most review-aware models remain optimiz

A new US phone network for Christians aims to block porn and gender-related content

ResearchDGX agent

A new US-wide cell phone network marketed to Christians is set to launch next week. It blocks porn, which experts in network security say marks the first time a US cell plan has used network-level blo

A Novel Computational Framework for Causal Inference: Tree-Based Discretization with ILP-Based Matching

ResearchDGX agent

arXiv:2604.27307v1 Announce Type: cross Abstract: Causal inference is essential for data-driven decision-making, as it aims to uncover causal relationships from observational data. However, identifyin

A Real-time Scale-robust Network for Glottis Segmentation in Nasal Transnasal Intubation

ResearchDGX agent

arXiv:2604.27383v1 Announce Type: cross Abstract: Nasotracheal intubation (NTI) is a critical clinical procedure for establishing and maintaining patient airway patency. Machine-assisted NTI has emerg

A Short Note on Batch-efficient Divide-and-Conquer Algorithm for EigenDecomposition

ResearchDGX agent

arXiv:2604.27325v1 Announce Type: new Abstract: EigenDecomposition (ED) is at the heart of many computer vision algorithms and applications. One crucial bottleneck limiting its usage is the expensive

ActiNet: An Open-Source Tool for Activity Intensity Classification of Wrist-Worn Accelerometry Using Self-Supervised Deep Learning

ResearchDGX agent

arXiv:2510.01712v2 Announce Type: replace Abstract: The use of accurate and reliable open-source human activity recognition (HAR) models on passively collected wrist-accelerometer data is essential in

Activation Function Design Sustains Plasticity in Continual Learning

ResearchDGX agent

arXiv:2509.22562v4 Announce Type: replace-cross Abstract: In independent, identically distributed (i.i.d.) training regimes, activation functions have been benchmarked extensively, and their differenc

AdaBFL: Multi-Layer Defensive Adaptive Aggregation for Bzantine-Robust Federated Learning

ResearchDGX agent

arXiv:2604.27434v1 Announce Type: cross Abstract: Federated learning (FL) is a popular distributed learning paradigm in machine learning, which enables multiple clients to collaboratively train models

Adaptive Nonlinear MPC for Trajectory Tracking of An Overactuated Tiltrotor Hexacopter

ResearchDGX agent

arXiv:2211.06762v2 Announce Type: replace Abstract: Omnidirectional micro aerial vehicles (OMAVs) are more capable of doing environmentally interactive tasks due to their ability to exert full wrenche

Adjoint Inversion Reveals Holographic Superposition and Destructive Interference in CNN Classifiers

ResearchDGX agent

arXiv:2604.27529v1 Announce Type: new Abstract: A foundational assumption in CNN interpretability -- that deep encoders suppress background pixels while classifiers merely select from a cleaned featur

AG-TAL: Anatomically-Guided Topology-Aware Loss for Multiclass Segmentation of the Circle of Willis Using Large-Scale Multi-Center Datasets

ResearchDGX agent

arXiv:2604.27357v1 Announce Type: cross Abstract: Accurate multiclass segmentation of the Circle of Willis (CoW) is essential for neurovascular disease management but remains challenging due to comple

← Previous
1…286287288289290…432
Next →