AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
13 Apr 2026

BLEG: LLM Functions as Powerful fMRI Graph-Enhancer for Brain Network Analysis

SafetyDGX agent

arXiv:2604.07361v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have been widely used in diverse brain network analysis tasks based on preprocessed functional magnetic resonance imagi

Boosting Brain-inspired Path Integration Efficiency via Learning-based Replication of Continuous Attractor Neurodynamics

Model ReleasesDGX agent

arXiv:2511.17687v2 Announce Type: replace Abstract: The brain's Path Integration (PI) mechanism offers substantial guidance and inspiration for Brain-Inspired Navigation (BIN). However, the PI capabil

Bridging SFT and RL: Dynamic Policy Optimization for Robust Reasoning

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.08926v1 Announce Type: new Abstract: Post-training paradigms for Large Language Models (LLMs), primarily Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL), face a fundamental dil

Bringing Clustering to MLL: Weakly-Supervised Clustering for Partial Multi-Label Learning

ResearchDGX agent

arXiv:2604.09359v1 Announce Type: new Abstract: Label noise in multi-label learning (MLL) poses significant challenges for model training, particularly in partial multi-label learning (PML) where cand

CERBERUS: A Three-Headed Decoder for Vertical Cloud Profiles

ResearchDGX agent

arXiv:2604.08772v1 Announce Type: cross Abstract: Atmospheric clouds exhibit complex three-dimensional structure and microphysical details that are poorly constrained by the predominantly two-dimensio

Conformal Prediction in Hierarchical Classification with Constrained Representation Complexity

Model ReleasesDGX agent

arXiv:2501.19038v3 Announce Type: replace-cross Abstract: Conformal prediction has emerged as a widely used framework for constructing valid prediction sets in classification and regression tasks. In

Continuous Orthogonal Mode Decomposition: Haptic Signal Prediction in Tactile Internet

ResearchDGX agent

arXiv:2604.09446v1 Announce Type: cross Abstract: The Tactile Internet demands sub-millisecond latency and ultra-high reliability, as high latency or packet loss could lead to haptic control instabili

Contribution of task-irrelevant stimuli to drift of neural representations

AgentsDGX agent

arXiv:2510.21588v2 Announce Type: replace-cross Abstract: Biological and artificial learners are inherently exposed to a stream of data and experience throughout their lifetimes and must constantly ad

Creator Incentives in Recommender Systems: A Cooperative Game-Theoretic Approach for Stable and Fair Collaboration in Multi-Agent Bandits

SafetyDGX agent

arXiv:2604.08643v1 Announce Type: new Abstract: User interactions in online recommendation platforms create interdependencies among content creators: feedback on one creator's content influences the s

Delve into the Applicability of Advanced Optimizers for Multi-Task Learning

Model ReleasesDGX agent

arXiv:2604.08939v1 Announce Type: new Abstract: Multi-Task Learning (MTL) is a foundational machine learning problem that has seen extensive development over the past decade. Recently, various optimiz

Demystifying Mergeability: Interpretable Properties to Predict Model Merging Success

SafetyDGX agent

arXiv:2601.22285v4 Announce Type: replace Abstract: Model merging combines knowledge from separately fine-tuned models, yet success factors remain poorly understood. While recent work treats mergeabil

Differentially Private and Federated Structure Learning in Bayesian Networks

ResearchDGX agent

arXiv:2512.01708v2 Announce Type: replace-cross Abstract: Learning the structure of a Bayesian network from decentralized data poses two major challenges: (i) ensuring rigorous privacy guarantees for

DiffHLS: Differential Learning for High-Level Synthesis QoR Prediction with GNNs and LLM Code Embeddings

ResearchDGX agent

arXiv:2604.09240v1 Announce Type: new Abstract: High-Level Synthesis (HLS) compiles C/C++ into RTL, but exploring pragma-driven optimization choices remains expensive because each design point require

Discrete Meanflow Training Curriculum

ResearchDGX agent

arXiv:2604.08837v1 Announce Type: new Abstract: Flow-based image generative models exhibit stable training and produce high quality samples when using multi-step sampling procedures. One-step generati

Distributed Online Convex Optimization with Compressed Communication: Optimal Regret and Applications

ResearchDGX agent

arXiv:2604.09276v1 Announce Type: new Abstract: Distributed online convex optimization (D-OCO) is a powerful paradigm for modeling distributed scenarios with streaming data. However, the communication

Distribution-free two-sample testing with blurred total variation distance

ResearchDGX agent

arXiv:2602.05862v2 Announce Type: replace-cross Abstract: Two-sample testing, where we aim to determine whether two distributions are equal or not equal based on samples from each one, is challenging

dnaHNet: A Scalable and Hierarchical Foundation Model for Genomic Sequence Learning

ResearchDGX agent

arXiv:2602.10603v3 Announce Type: replace Abstract: Genomic foundation models have the potential to decode DNA syntax, yet face a fundamental tradeoff in their input representation. Standard fixed-voc

Drift-Aware Online Dynamic Learning for Nonstationary Multivariate Time Series: Application to Sintering Quality Prediction

Model ReleasesDGX agent

arXiv:2604.09358v1 Announce Type: new Abstract: Accurate prediction of nonstationary multivariate time series remains a critical challenge in complex industrial systems such as iron ore sintering. In

Dual Mamba for Node-Specific Representation Learning: Tackling Over-Smoothing with Selective State Space Modeling

ResearchDGX agent

arXiv:2511.06756v3 Announce Type: replace Abstract: Over-smoothing remains a fundamental challenge in deep Graph Neural Networks (GNNs), where repeated message passing causes node representations to b

Efficient Hierarchical Implicit Flow Q-learning for Offline Goal-conditioned Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.08960v1 Announce Type: new Abstract: Offline goal-conditioned reinforcement learning (GCRL) is a practical reinforcement learning paradigm that aims to learn goal-conditioned policies from

Efficient RL Training for LLMs with Experience Replay

SafetyDGX agent

arXiv:2604.08706v1 Announce Type: new Abstract: While Experience Replay - the practice of storing rollouts and reusing them multiple times during training - is a foundational technique in general RL,

EngageTriBoost: Predictive Modeling of User Engagement in Digital Mental Health Intervention Using Explainable Machine Learning

ResearchDGX agent

arXiv:2604.08589v1 Announce Type: new Abstract: Mental health challenges among young adults, are on the rise, necessitating effective solutions such as digital mental health interventions (DMHIs). Des

Event-Driven Temporal Graph Networks for Asynchronous Multi-Agent Cyber Defense in NetForge_RL

AgentsDGX agent

arXiv:2604.09523v1 Announce Type: new Abstract: The transition of Multi-Agent Reinforcement Learning (MARL) policies from simulated cyber wargames to operational Security Operations Centers (SOCs) is

EvoLen: Evolution-Guided Tokenization for DNA Language Model

SafetyDGX agent

arXiv:2604.08698v1 Announce Type: new Abstract: Tokens serve as the basic units of representation in DNA language models (DNALMs), yet their design remains underexplored. Unlike natural language, DNA

Feature-Label Modal Alignment for Robust Partial Multi-Label Learning

SafetyDGX agent

arXiv:2604.09064v1 Announce Type: new Abstract: In partial multi-label learning (PML), each instance is associated with a set of candidate labels containing both ground-truth and noisy labels. The pre

Finite-Sample Analysis of Nonlinear Independent Component Analysis:Sample Complexity and Identifiability Bounds

Model ReleasesDGX agent

arXiv:2604.08850v1 Announce Type: new Abstract: Independent Component Analysis (ICA) is a fundamental unsupervised learning technique foruncovering latent structure in data by separating mixed signals

Fisher-Geometric Diffusion in Stochastic Gradient Descent: Optimal Rates, Oracle Complexity, and Information-Theoretic Limits

ResearchDGX agent

arXiv:2603.02417v3 Announce Type: replace-cross Abstract: Classical stochastic-approximation analyses treat the covariance of stochastic gradients as an exogenous modeling input. We show that under ex

FIT-GNN: Faster Inference Time for GNNs that 'FIT' in Memory Using Coarsening

Model ReleasesDGX agent

arXiv:2410.15001v5 Announce Type: replace Abstract: Scalability of Graph Neural Networks (GNNs) remains a significant challenge. To tackle this, methods like coarsening, condensation, and computation

Fully Autonomous Z-Score-Based TinyML Anomaly Detection on Resource-Constrained MCUs Using Power Side-Channel Data

Local AiDGX agent

arXiv:2604.08581v1 Announce Type: new Abstract: This paper presents a fully autonomous Tiny Machine Learning (TinyML) Z-Score-based anomaly detection system deployed on a low-power microcontroller for

Gated-SwinRMT: Unifying Swin Windowed Attention with Retentive Manhattan Decay via Input-Dependent Gating

SafetyDGX agent

arXiv:2604.06014v2 Announce Type: replace Abstract: We introduce Gated-SwinRMT, a family of hybrid vision transformers that combine the shifted-window attention of the Swin Transformer with the Manhat

Geometry-Induced Long-Range Correlations in Recurrent Neural Network Quantum States

SafetyDGX agent

arXiv:2604.08661v1 Announce Type: cross Abstract: Neural Quantum States based on autoregressive recurrent neural network (RNN) wave functions enable efficient sampling without Markov-chain autocorrela

GeoPAS: Geometric Probing for Algorithm Selection in Continuous Black-Box Optimisation

Model ReleasesDGX agent

arXiv:2604.09095v1 Announce Type: new Abstract: Automated algorithm selection in continuous black-box optimisation typically relies on fixed landscape descriptors computed under a limited probing budg

GL-LowPopArt: A Nearly Instance-Wise Minimax-Optimal Estimator for Generalized Low-Rank Trace Regression

Local AiDGX agent

arXiv:2506.03074v5 Announce Type: replace-cross Abstract: We present `GL-LowPopArt`, a novel Catoni-style estimator for generalized low-rank trace regression. Building on `LowPopArt` (Jang et al., 202

Graph Defense Diffusion Model

ApplicationsDGX agent

arXiv:2501.11568v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) are highly vulnerable to adversarial attacks, which can greatly degrade their performance. Existing graph purification

Hierarchical Flow Decomposition for Turning Movement Prediction at Signalized Intersections

ResearchDGX agent

arXiv:2604.09336v1 Announce Type: new Abstract: Accurate prediction of intersection turning movements is essential for adaptive signal control but remains difficult due to the high volatility of direc

Hierarchical Kernel Transformer: Multi-Scale Attention with an Information-Theoretic Approximation Analysis

ResearchDGX agent

arXiv:2604.08829v1 Announce Type: new Abstract: The Hierarchical Kernel Transformer (HKT) is a multi-scale attention mechanism that processes sequences at L resolution levels via trainable causal down

Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling

Model ReleasesDGX agent

arXiv:2604.05072v2 Announce Type: replace Abstract: Recent large language models have shifted SVG generation from differentiable rendering optimization to autoregressive program synthesis. However, ex

High-dimensional inference for the gamma-ray sky with differentiable programming

HardwareDGX agent

arXiv:2604.08648v1 Announce Type: cross Abstract: We motivate the use of differentiable probabilistic programming techniques in order to account for the large model-space inherent to astrophysical gam

How does Chain of Thought decompose complex tasks?

ResearchDGX agent

arXiv:2604.08872v1 Announce Type: new Abstract: Many language tasks can be modeled as classification problems where a large language model (LLM) is given a prompt and selects one among many possible a

Identifying Causal Effects Using a Single Proxy Variable

ResearchDGX agent

arXiv:2604.09135v1 Announce Type: cross Abstract: Unobserved confounding is a key challenge when estimating causal effects from a treatment on an outcome in scientific applications. In this work, we a

IKKA: Inversion Classification via Critical Anomalies for Robust Visual Servoing

ResearchDGX agent

arXiv:2604.08754v1 Announce Type: new Abstract: We introduce IKKA (Inversion Classification via Critical Anomalies), a topologically motivated weighting framework for robust visual servoing under dist

Imitation Learning for Combinatorial Optimisation under Uncertainty

ResearchDGX agent

arXiv:2601.05383v4 Announce Type: replace Abstract: Imitation learning (IL) provides a data-driven framework for approximating policies for large-scale combinatorial optimisation problems formulated a

Implicit Bias in Deep Linear Discriminant Analysis

SafetyDGX agent

arXiv:2603.02622v2 Announce Type: replace Abstract: While the Implicit Bias(or Implicit Regularization) of standard loss functions has been studied, the optimization geometry induced by discriminative

Improving Model Performance by Adapting the KGE Metric to Account for System Non-Stationarity

Model ReleasesDGX agent

arXiv:2604.03906v2 Announce Type: replace Abstract: Geoscientific systems tend to be characterized by pronounced temporal non-stationarity, arising from seasonal and climatic variability in hydrometeo

Inferring Latent Temporal Sparse Coordination Graph for Multi-Agent Reinforcement Learning

Model ReleasesDGX agent

arXiv:2403.19253v3 Announce Type: replace Abstract: Effective agent coordination is crucial in cooperative Multi-Agent Reinforcement Learning (MARL). While agent cooperation can be represented by grap

Integrated electro-optic attention nonlinearities for transformers

ResearchDGX agent

arXiv:2604.09512v1 Announce Type: new Abstract: Transformers have emerged as the dominant neural-network architecture, achieving state-of-the-art performance in language processing and computer vision

Iterative Identification Closure: Amplifying Causal Identifiability in Linear SEMs

ResearchDGX agent

arXiv:2604.09309v1 Announce Type: cross Abstract: The Half-Trek Criterion (HTC) is the primary graphical tool for determining generic identifiability of causal effect coefficients in linear structural

Large Reasoning Models Learn Better Alignment from Flawed Thinking

SafetyDGX agent

arXiv:2510.00938v2 Announce Type: replace Abstract: Large reasoning models (LRMs) 'think' by generating structured chain-of-thought (CoT) before producing a final answer, yet they still lack the abili

Learning Encodings by Maximizing State Distinguishability: Variational Quantum Error Correction

ResearchDGX agent

arXiv:2506.11552v2 Announce Type: replace-cross Abstract: Quantum error correction is crucial for protecting quantum information against decoherence. Traditional codes like the surface code require su

Loom: A Scalable Analytical Neural Computer Architecture

ResearchDGX agent

arXiv:2604.08816v1 Announce Type: new Abstract: We present Loom, a computer architecture that executes programs compiled from C inside a looped transformer whose weights are derived analytically. The

Low Rank Based Subspace Inference for the Laplace Approximation of Bayesian Neural Networks

Model ReleasesDGX agent

arXiv:2502.02345v2 Announce Type: replace Abstract: Subspace inference for neural networks assumes that a subspace of their parameter space suffices to produce a reliable uncertainty quantification. I

Mamba-Based Graph Convolutional Networks: Tackling Over-smoothing with Selective State Space

Model ReleasesDGX agent

arXiv:2501.15461v4 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have shown great success in various graph-based learning tasks. However, it often faces the issue of over-smoothing as

MARBLE: Multi-Armed Restless Bandits in Latent Markovian Environment

SafetyDGX agent

arXiv:2511.09324v2 Announce Type: replace Abstract: Restless Multi-Armed Bandits (RMABs) are powerful models for decision-making under uncertainty, yet classical formulations typically assume fixed dy

MATCHA: Efficient Deployment of Deep Neural Networks on Multi-Accelerator Heterogeneous Edge SoCs

Model ReleasesDGX agent

arXiv:2604.09124v1 Announce Type: cross Abstract: Deploying DNNs on System-on-Chips (SoC) with multiple heterogeneous acceleration engines is challenging, and the majority of deployment frameworks can

Mechanisms of Introspective Awareness

SafetyDGX agent

arXiv:2603.21396v2 Announce Type: replace Abstract: Recent work has shown that LLMs can sometimes detect when steering vectors are injected into their residual stream and identify the injected concept

Memory-Guided Trust-Region Bayesian Optimization (MG-TuRBO) for High Dimensions

ApplicationsDGX agent

arXiv:2604.08569v1 Announce Type: new Abstract: Traffic simulation and digital-twin calibration is a challenging optimization problem with a limited simulation budget. Each trial requires an expensive

Meta-Learned Basis Adaptation for Parametric Linear PDEs

ResearchDGX agent

arXiv:2604.09289v1 Announce Type: new Abstract: We propose a hybrid physics-informed framework for solving families of parametric linear partial differential equations (PDEs) by combining a meta-learn

Modality-Aware Zero-Shot Pruning and Sparse Attention for Efficient Multimodal Edge Inference

ResearchDGX agent

arXiv:2604.08971v1 Announce Type: new Abstract: Edge devices increasingly run multimodal sensing pipelines that must remain accurate despite fluctuating power budgets and unpredictable sensor dropout.

Multi-Agent Decision-Focused Learning via Value-Aware Sequential Communication

AgentsDGX agent

arXiv:2604.08944v1 Announce Type: new Abstract: Multi-agent coordination under partial observability requires agents to share complementary private information. While recent methods optimize messages

Natural Riemannian gradient for learning functional tensor networks

Model ReleasesDGX agent

arXiv:2604.09263v1 Announce Type: cross Abstract: We consider machine learning tasks with low-rank functional tree tensor networks (TTN) as the learning model. While in the case of least-squares regre

← Previous
1…233234235236237…239
Next →