AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
Tutorials

Exploring the Effectiveness of Using LLMs for Automated Assessment of Student Self Explanations in Programming Education

DGX agent

arXiv:2605.21614v1 Announce Type: cross Abstract: Worked examples are step-by-step solutions to problems in a specific domain, offered to students to acquire domain-specific problem-solving skills. Th

tutorialsarxiv-cs-lg
23 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

extit{BlockFormer} : Transformer-based inference from interaction maps

DGX agent

arXiv:2605.21617v1 Announce Type: new Abstract: Inference from interaction maps, such as centromere identification from genome-wide chromosome conformation capture techniques -- notably Hi-C -- can be

safetyarxiv-cs-lg
23 May 2026
Local Ai

F-TIS: Harnessing Diverse Models in Collaborative GRPO

DGX agent

arXiv:2605.22537v1 Announce Type: new Abstract: Reinforcement learning methods such as GRPO have seen great popularity in LLM post-training. In GRPO, models produce completions to a set of prompts, wh

local-aiarxiv-cs-lg
23 May 2026
Safety

Factored Diffusion Policies:Compositionally Generalized Robot Control with a Single Score Network

DGX agent

arXiv:2605.22596v1 Announce Type: new Abstract: Robotic tasks are typically specified by a tuple of factors, such as the object to be grasped, the obstacles to be avoided, the color of the target, and

safetyarxiv-cs-lg
23 May 2026
Applications

FAME: Failure-Aware Mixture-of-Experts for Message-Level Log Anomaly Detection

DGX agent

arXiv:2605.22779v1 Announce Type: cross Abstract: Production systems generate millions of log lines daily, yet most anomaly detectors operate at the session or window-level, flagging groups of lines r

applicationsarxiv-cs-lg
23 May 2026
Model Releases

FD-Bench: A Modular and Fair Benchmark for Data-driven Fluid Simulation

DGX agent

arXiv:2505.20349v2 Announce Type: replace-cross Abstract: Data-driven modeling of fluid dynamics has advanced rapidly with neural PDE solvers, yet a fair and strong benchmark remains fragmented due to

model-releasesarxiv-cs-lg
23 May 2026
Research

Finite-Particle Convergence Rates for Conservative and Non-Conservative Drifting Models

DGX agent

arXiv:2605.22795v1 Announce Type: cross Abstract: We propose and analyze a conservative drifting method for one-step generative modeling. The method replaces the original displacement-based drifting v

researcharxiv-cs-lg
23 May 2026
Research

Flashlight: PyTorch Compiler Extensions to Accelerate Attention Variants

DGX agent

arXiv:2511.02043v4 Announce Type: replace Abstract: Attention is a fundamental building block of large language models (LLMs), so there have been many efforts to implement it efficiently. For example,

researcharxiv-cs-lg
23 May 2026
Hardware

FlashSinkhorn: IO-Aware Entropic Optimal Transport on GPU

DGX agent

arXiv:2602.03067v3 Announce Type: replace Abstract: Entropic optimal transport (EOT) via Sinkhorn iterations is widely used in modern machine learning, yet GPU solvers remain inefficient at scale. Ten

hardwarearxiv-cs-lg
23 May 2026
Model Releases

Frequency-Domain Regularized Adversarial Alignment for Transferable Attacks against Closed-Source MLLMs

DGX agent

arXiv:2605.21541v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) remain vulnerable to transfer-based targeted attacks, where perturbations optimized on open-source surrogate

model-releasesarxiv-cs-lg
23 May 2026
Research

From Betting to Empirical Bernstein LIL

DGX agent

arXiv:2605.22124v1 Announce Type: cross Abstract: This is a verbatim copy of a technical report I wrote in 2017-2018 to obtain the law of the iterated logarithm using the guarantee on the wealth of an

researcharxiv-cs-lg
23 May 2026
Hardware

From Sequential Nodes to GPU Batches: Parallel Branch and Bound for Optimal k-Sparse GLMs

DGX agent

arXiv:2605.22188v1 Announce Type: new Abstract: GPUs have significantly accelerated first-order methods for large-scale optimization, especially in continuous optimization. However, this success has n

hardwarearxiv-cs-lg
23 May 2026
Safety

From Snapshots to Trajectories: Learning Single-Cell Gene Expression Dynamics via Conditional Flow Matching

DGX agent

arXiv:2605.22340v1 Announce Type: new Abstract: Single-cell RNA sequencing (scRNA-seq) provides high-dimensional profiles of cellular states, enabling data-driven modeling of cellular dynamics over ti

safetyarxiv-cs-lg
23 May 2026
Safety

Generative Modeling by Value-Driven Transport

DGX agent

arXiv:2605.22507v1 Announce Type: new Abstract: We propose a new framework for generative modeling based on a discrete-time stochastic control formulation of measure transport. Adapting classic result

safetyarxiv-cs-lg
23 May 2026
Applications

Geometry-Induced Diffusion on Graphs: A Learnable Weighted Laplacian for Spectral GNNs

DGX agent

arXiv:2602.18141v2 Announce Type: replace Abstract: Long-range graph tasks are challenging for Graph Neural Networks (GNNs): global mechanisms such as attention or rewiring schemes can be computationa

applicationsarxiv-cs-lg
23 May 2026
Research

Gradient Descent as a Perceptron Algorithm: Understanding Dynamics and Implicit Acceleration

DGX agent

arXiv:2512.11587v2 Announce Type: replace Abstract: Even for the gradient descent (GD) method applied to neural network training, understanding its optimization dynamics, including convergence rate, i

researcharxiv-cs-lg
23 May 2026
Research

Graph neural network explanations reveal a topological signature of disease-associated hubs in biological networks

DGX agent

arXiv:2605.21502v1 Announce Type: cross Abstract: Graph neural networks (GNNs) are increasingly used to model biological systems, yet the reliability of post-hoc explanation methods for recovering mea

researcharxiv-cs-lg
23 May 2026
Model Releases

GraphFlow: A Graph-Based Workflow Management for Efficient LLM-Agent Serving

DGX agent

arXiv:2605.22566v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents demonstrate strong reasoning and execution capabilities on complex tasks when guided by structured instructions,

model-releasesarxiv-cs-lg
23 May 2026
Safety

Harnesses for Inference-Time Alignment over Execution Trajectories

DGX agent

arXiv:2605.21516v1 Announce Type: new Abstract: Harness engineering has emerged as an important inference-time technique for large language model (LLM) agents, aiming to improve long-term performance

safetyarxiv-cs-lg
23 May 2026
Applications

Healthcare LLM Benchmarks Are Only as Good as Their Explicit Assumptions

DGX agent

arXiv:2605.22612v1 Announce Type: cross Abstract: Benchmarks are necessary for healthcare evaluation, but are not sufficient for predicting deployment performance. Our position is that the evaluation-

applicationsarxiv-cs-lg
23 May 2026
Safety

HealthMamba: An Uncertainty-aware Spatiotemporal Graph State Space Model for Effective and Reliable Healthcare Facility Visit Prediction

DGX agent

arXiv:2602.05286v3 Announce Type: replace Abstract: Healthcare facility visit prediction is essential for optimizing healthcare resource allocation and informing public health policy. Despite advanced

safetyarxiv-cs-lg
23 May 2026
Safety

Heterogeneous Agent Collaborative Reinforcement Learning

DGX agent

arXiv:2603.02604v2 Announce Type: replace Abstract: We introduce Heterogeneous Agent Collaborative Reinforcement Learning (HACRL), a new Reinforcement Learning from Verifiable Reward (RLVR) problem th

safetyarxiv-cs-lg
23 May 2026
Model Releases

HIDBench: Benchmarking Large Language Models for Host-Based Intrusion Detection

DGX agent

arXiv:2605.21773v1 Announce Type: cross Abstract: Recent benchmark efforts have advanced the evaluation of large language models (LLMs) in cybersecurity, including tasks such as penetration testing an

model-releasesarxiv-cs-lg
23 May 2026
Research

Holographic functions and neural networks

DGX agent

arXiv:2605.22666v1 Announce Type: cross Abstract: A fuzzy Boolean function is a map f:ube^no [0,1], where ninmathbb N. We introduce and compare three ways of saying that such a function has bounded co

researcharxiv-cs-lg
23 May 2026
Model Releases

Holomorphic Neural ODEs with Kolmogorov-Arnold Networks for Interpretable Discovery of Complex Dynamics

DGX agent

arXiv:2605.22235v1 Announce Type: new Abstract: Complex dynamical systems governed by holomorphic maps such as z^2 + c exhibit fractal boundaries with extreme sensitivity to initial conditions. Accura

model-releasesarxiv-cs-lg
23 May 2026
Research

How Many Different Outputs Can a Transformer Generate?

DGX agent

arXiv:2605.22223v1 Announce Type: new Abstract: We study how we can leverage only a handful of characteristics of a transformer's architecture to closely predict the number of different sequences it c

researcharxiv-cs-lg
23 May 2026
Research

How Sparsity Allocation Shapes Label-Free Post-Pruning Recoverability

DGX agent

arXiv:2605.21972v1 Announce Type: new Abstract: Unstructured magnitude pruning at high sparsity can reduce neural network accuracy to near-random performance, while labeled retraining may be unavailab

researcharxiv-cs-lg
23 May 2026
Model Releases

Hybrid Kolmogorov-Arnold Network and XGBoost Framework for Week-Ahead Price Forecasting in Australia's National Electricity Market

DGX agent

arXiv:2605.22387v1 Announce Type: new Abstract: Accurate electricity price forecasting (EPF) is essential for market participants to support operational planning and risk management, yet remains chall

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Hyperparameter Transfer with Mixture-of-Expert Layers

DGX agent

arXiv:2601.20205v3 Announce Type: replace Abstract: Mixture-of-Experts (MoE) layers have emerged as an important tool in scaling up modern neural networks by decoupling total trainable parameters from

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

I-SAFE: Wasserstein Coherence Metrics for Structural Auditing of Scientific AI Models

DGX agent

arXiv:2605.21731v1 Announce Type: new Abstract: Deep learning models are increasingly used in scientific prediction tasks where strong benchmark performance is often interpreted as evidence of scienti

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

IKNO: Infinite-order Kernel Neural Operators

DGX agent

arXiv:2605.22182v1 Announce Type: new Abstract: Neural operators have achieved significant success in modern scientific computing due to their flexibility and strong generalization capabilities. Exist

model-releasesarxiv-cs-lg
23 May 2026
Research

Implicit Regularization of Mini-Batch Training in Graph Neural Networks

DGX agent

arXiv:2605.22480v1 Announce Type: new Abstract: Mini-batch training of Graph Neural Networks (GNNs) is fundamentally different from training on i.i.d. data: sampling a subgraph alters the topology and

researcharxiv-cs-lg
23 May 2026
Hardware

ImplicitTerrainV2: Wavelet-Guided Spatially Adaptive Neural Terrain Representation

DGX agent

arXiv:2605.22556v1 Announce Type: new Abstract: Digital elevation models (DEMs) underpin terrain analysis in Geographic Information Systems (GIS), but in their common raster form, they rely on interpo

hardwarearxiv-cs-lg
23 May 2026
Research

Innovations in Cardless Artificial Intelligence Banking: A Comprehensive Framework for Cyber Secure and Fraud Mitigation using Machine Learning Algorithms

DGX agent

arXiv:2605.22604v1 Announce Type: cross Abstract: The advent of cardless artificial intelligence (AI) banking heralds a paradigm shift in the financial landscape, offering users unprecedented security

researcharxiv-cs-lg
23 May 2026
Model Releases

Integrable Elasticity via Neural Demand Potentials

DGX agent

arXiv:2605.22820v1 Announce Type: new Abstract: We propose the Integrable Context-Dependent Demand Network (ICDN), a demand-first neural model for multiproduct retail demand. The model learns log-dema

model-releasesarxiv-cs-lg
23 May 2026
Research

Interpreting and Steering State-Space Models via Activation Subspace Bottlenecks

DGX agent

arXiv:2602.22719v2 Announce Type: replace Abstract: State-space models (SSMs) have emerged as an efficient strategy for building powerful language models, avoiding the quadratic complexity of computin

researcharxiv-cs-lg
23 May 2026
Safety

Kernel-Based Safe Exploration in Deep Reinforcement Learning

DGX agent

arXiv:2605.22207v1 Announce Type: cross Abstract: Safety has been a major concern when deploying deep reinforcement learning algorithms in the real world. A promising direction that ensures that the l

safetyarxiv-cs-lg
23 May 2026
Applications

LABO: LLM-Accelerated Bayesian Optimization through Broad Exploration and Selective Experimentation

DGX agent

arXiv:2605.22054v1 Announce Type: new Abstract: The high cost and data scarcity in scientific exploration have motivated the use of large language models (LLMs) as knowledge-driven components in Bayes

applicationsarxiv-cs-lg
23 May 2026
Research

Large-scale Score-based Variational Posterior Inference for Bayesian Deep Neural Networks

DGX agent

arXiv:2602.05873v2 Announce Type: replace Abstract: Bayesian (deep) neural networks (BNN) are often more attractive than the vanilla point-estimate deep learning in various aspects including uncertain

researcharxiv-cs-lg
23 May 2026
Agents

LCGuard: Latent Communication Guard for Safe KV Sharing in Multi-Agent Systems

DGX agent

arXiv:2605.22786v1 Announce Type: cross Abstract: Large language model (LLM)-based multi-agent systems increasingly rely on intermediate communication to coordinate complex tasks. While most existing

agentsarxiv-cs-lg
23 May 2026
Tutorials

Learning Causal Orderings for In-Context Tabular Prediction

DGX agent

arXiv:2605.22335v1 Announce Type: new Abstract: In-context learning for tabular data sets strong predictive standards in observational settings; it however primarily relies on correlational structure,

tutorialsarxiv-cs-lg
23 May 2026
Research

Learning Mixture Models via Efficient High-dimensional Sparse Fourier Transforms

DGX agent

arXiv:2601.05157v2 Announce Type: replace-cross Abstract: In this work, we give a {rm poly}(d,k) time and sample algorithm for efficiently learning the parameters of a mixture of k spherical distribut

researcharxiv-cs-lg
23 May 2026
Research

LEMUR: Learned Multi-Vector Retrieval

DGX agent

arXiv:2601.21853v2 Announce Type: replace-cross Abstract: Multi-vector representations generated by late interaction models, such as ColBERT, enable superior retrieval quality compared to single-vecto

researcharxiv-cs-lg
23 May 2026
Research

Leveraging Self-Paced Curriculum Learning for Enhanced Modality Balance in Multimodal Conversational Emotion Recognition

DGX agent

arXiv:2605.21565v1 Announce Type: new Abstract: Multimodal Emotion Recognition in Conversations (MERC) is a crucial task for understanding human interactions, where multimodal approaches integrating l

researcharxiv-cs-lg
23 May 2026
Hardware

LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations

DGX agent

arXiv:2602.01935v2 Announce Type: replace Abstract: LLM-guided compiler optimization has recently shown promise, but existing approaches rely on a single large LLM throughout search, making them expen

hardwarearxiv-cs-lg
23 May 2026
Research

Little by Little: Continual Learning via Incremental Mixture of Rank-1 Associative Memory Experts

DGX agent

arXiv:2506.21035v5 Announce Type: replace Abstract: Continual learning (CL) with large pre-trained models aims to incrementally acquire knowledge without catastrophic forgetting. Existing LoRA-based M

researcharxiv-cs-lg
23 May 2026
Safety

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators

DGX agent

arXiv:2605.22717v1 Announce Type: cross Abstract: Interactive streaming music generation promises the use of generative models for live performance and co-creation that is impossible with offline mode

safetyarxiv-cs-lg
23 May 2026
Applications

Local Covariate Selection for Average Causal Effect Estimation without Pretreatment and Causal Sufficiency Assumptions

DGX agent

arXiv:2605.21548v1 Announce Type: cross Abstract: We study the problem of selecting covariates for unbiased estimation of the total causal effect.Existing approaches typically rely on global causal st

applicationsarxiv-cs-lg
23 May 2026
← Previous
1…786787788789790…1311
Next →