AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
23 May 2026

GraphFlow: A Graph-Based Workflow Management for Efficient LLM-Agent Serving

Model ReleasesDGX agent

arXiv:2605.22566v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents demonstrate strong reasoning and execution capabilities on complex tasks when guided by structured instructions,

Harnesses for Inference-Time Alignment over Execution Trajectories

SafetyDGX agent

arXiv:2605.21516v1 Announce Type: new Abstract: Harness engineering has emerged as an important inference-time technique for large language model (LLM) agents, aiming to improve long-term performance

Healthcare LLM Benchmarks Are Only as Good as Their Explicit Assumptions

ApplicationsDGX agent

arXiv:2605.22612v1 Announce Type: cross Abstract: Benchmarks are necessary for healthcare evaluation, but are not sufficient for predicting deployment performance. Our position is that the evaluation-


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HealthMamba: An Uncertainty-aware Spatiotemporal Graph State Space Model for Effective and Reliable Healthcare Facility Visit Prediction

SafetyDGX agent

arXiv:2602.05286v3 Announce Type: replace Abstract: Healthcare facility visit prediction is essential for optimizing healthcare resource allocation and informing public health policy. Despite advanced

Heterogeneous Agent Collaborative Reinforcement Learning

SafetyDGX agent

arXiv:2603.02604v2 Announce Type: replace Abstract: We introduce Heterogeneous Agent Collaborative Reinforcement Learning (HACRL), a new Reinforcement Learning from Verifiable Reward (RLVR) problem th

HIDBench: Benchmarking Large Language Models for Host-Based Intrusion Detection

Model ReleasesDGX agent

arXiv:2605.21773v1 Announce Type: cross Abstract: Recent benchmark efforts have advanced the evaluation of large language models (LLMs) in cybersecurity, including tasks such as penetration testing an

Holographic functions and neural networks

ResearchDGX agent

arXiv:2605.22666v1 Announce Type: cross Abstract: A fuzzy Boolean function is a map f:ube^no [0,1], where ninmathbb N. We introduce and compare three ways of saying that such a function has bounded co

Holomorphic Neural ODEs with Kolmogorov-Arnold Networks for Interpretable Discovery of Complex Dynamics

Model ReleasesDGX agent

arXiv:2605.22235v1 Announce Type: new Abstract: Complex dynamical systems governed by holomorphic maps such as z^2 + c exhibit fractal boundaries with extreme sensitivity to initial conditions. Accura

How Many Different Outputs Can a Transformer Generate?

ResearchDGX agent

arXiv:2605.22223v1 Announce Type: new Abstract: We study how we can leverage only a handful of characteristics of a transformer's architecture to closely predict the number of different sequences it c

How Sparsity Allocation Shapes Label-Free Post-Pruning Recoverability

ResearchDGX agent

arXiv:2605.21972v1 Announce Type: new Abstract: Unstructured magnitude pruning at high sparsity can reduce neural network accuracy to near-random performance, while labeled retraining may be unavailab

Hybrid Kolmogorov-Arnold Network and XGBoost Framework for Week-Ahead Price Forecasting in Australia's National Electricity Market

Model ReleasesDGX agent

arXiv:2605.22387v1 Announce Type: new Abstract: Accurate electricity price forecasting (EPF) is essential for market participants to support operational planning and risk management, yet remains chall

Hyperparameter Transfer with Mixture-of-Expert Layers

Model ReleasesDGX agent

arXiv:2601.20205v3 Announce Type: replace Abstract: Mixture-of-Experts (MoE) layers have emerged as an important tool in scaling up modern neural networks by decoupling total trainable parameters from

I-SAFE: Wasserstein Coherence Metrics for Structural Auditing of Scientific AI Models

Model ReleasesDGX agent

arXiv:2605.21731v1 Announce Type: new Abstract: Deep learning models are increasingly used in scientific prediction tasks where strong benchmark performance is often interpreted as evidence of scienti

IKNO: Infinite-order Kernel Neural Operators

Model ReleasesDGX agent

arXiv:2605.22182v1 Announce Type: new Abstract: Neural operators have achieved significant success in modern scientific computing due to their flexibility and strong generalization capabilities. Exist

Implicit Regularization of Mini-Batch Training in Graph Neural Networks

ResearchDGX agent

arXiv:2605.22480v1 Announce Type: new Abstract: Mini-batch training of Graph Neural Networks (GNNs) is fundamentally different from training on i.i.d. data: sampling a subgraph alters the topology and

ImplicitTerrainV2: Wavelet-Guided Spatially Adaptive Neural Terrain Representation

HardwareDGX agent

arXiv:2605.22556v1 Announce Type: new Abstract: Digital elevation models (DEMs) underpin terrain analysis in Geographic Information Systems (GIS), but in their common raster form, they rely on interpo

Innovations in Cardless Artificial Intelligence Banking: A Comprehensive Framework for Cyber Secure and Fraud Mitigation using Machine Learning Algorithms

ResearchDGX agent

arXiv:2605.22604v1 Announce Type: cross Abstract: The advent of cardless artificial intelligence (AI) banking heralds a paradigm shift in the financial landscape, offering users unprecedented security

Integrable Elasticity via Neural Demand Potentials

Model ReleasesDGX agent

arXiv:2605.22820v1 Announce Type: new Abstract: We propose the Integrable Context-Dependent Demand Network (ICDN), a demand-first neural model for multiproduct retail demand. The model learns log-dema

Interpreting and Steering State-Space Models via Activation Subspace Bottlenecks

ResearchDGX agent

arXiv:2602.22719v2 Announce Type: replace Abstract: State-space models (SSMs) have emerged as an efficient strategy for building powerful language models, avoiding the quadratic complexity of computin

Kernel-Based Safe Exploration in Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.22207v1 Announce Type: cross Abstract: Safety has been a major concern when deploying deep reinforcement learning algorithms in the real world. A promising direction that ensures that the l

LABO: LLM-Accelerated Bayesian Optimization through Broad Exploration and Selective Experimentation

ApplicationsDGX agent

arXiv:2605.22054v1 Announce Type: new Abstract: The high cost and data scarcity in scientific exploration have motivated the use of large language models (LLMs) as knowledge-driven components in Bayes

Large-scale Score-based Variational Posterior Inference for Bayesian Deep Neural Networks

ResearchDGX agent

arXiv:2602.05873v2 Announce Type: replace Abstract: Bayesian (deep) neural networks (BNN) are often more attractive than the vanilla point-estimate deep learning in various aspects including uncertain

LCGuard: Latent Communication Guard for Safe KV Sharing in Multi-Agent Systems

AgentsDGX agent

arXiv:2605.22786v1 Announce Type: cross Abstract: Large language model (LLM)-based multi-agent systems increasingly rely on intermediate communication to coordinate complex tasks. While most existing

Learning Causal Orderings for In-Context Tabular Prediction

TutorialsDGX agent

arXiv:2605.22335v1 Announce Type: new Abstract: In-context learning for tabular data sets strong predictive standards in observational settings; it however primarily relies on correlational structure,

Learning Mixture Models via Efficient High-dimensional Sparse Fourier Transforms

ResearchDGX agent

arXiv:2601.05157v2 Announce Type: replace-cross Abstract: In this work, we give a {rm poly}(d,k) time and sample algorithm for efficiently learning the parameters of a mixture of k spherical distribut

LEMUR: Learned Multi-Vector Retrieval

ResearchDGX agent

arXiv:2601.21853v2 Announce Type: replace-cross Abstract: Multi-vector representations generated by late interaction models, such as ColBERT, enable superior retrieval quality compared to single-vecto

Leveraging Self-Paced Curriculum Learning for Enhanced Modality Balance in Multimodal Conversational Emotion Recognition

ResearchDGX agent

arXiv:2605.21565v1 Announce Type: new Abstract: Multimodal Emotion Recognition in Conversations (MERC) is a crucial task for understanding human interactions, where multimodal approaches integrating l

LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations

HardwareDGX agent

arXiv:2602.01935v2 Announce Type: replace Abstract: LLM-guided compiler optimization has recently shown promise, but existing approaches rely on a single large LLM throughout search, making them expen

Little by Little: Continual Learning via Incremental Mixture of Rank-1 Associative Memory Experts

ResearchDGX agent

arXiv:2506.21035v5 Announce Type: replace Abstract: Continual learning (CL) with large pre-trained models aims to incrementally acquire knowledge without catastrophic forgetting. Existing LoRA-based M

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators

SafetyDGX agent

arXiv:2605.22717v1 Announce Type: cross Abstract: Interactive streaming music generation promises the use of generative models for live performance and co-creation that is impossible with offline mode

Local Covariate Selection for Average Causal Effect Estimation without Pretreatment and Causal Sufficiency Assumptions

ApplicationsDGX agent

arXiv:2605.21548v1 Announce Type: cross Abstract: We study the problem of selecting covariates for unbiased estimation of the total causal effect.Existing approaches typically rely on global causal st

Long-term Fairness with Selective Labels

SafetyDGX agent

arXiv:2605.22291v1 Announce Type: new Abstract: Long-term fairness algorithms aim to satisfy fairness beyond static and short-term notions by accounting for the dynamics between decision-making polici

Lost in Tokenization: Fundamental Trade-offs in Graph Tokenization for Transformers

ApplicationsDGX agent

arXiv:2605.22471v1 Announce Type: new Abstract: Transformers have become a central architecture for graph learning, but their application to graphs requires first choosing a tokenization: a graph-to-t

Lumberjack: Better Differentially Private Random Forests through Heavy Hitter Detection in Trees

Model ReleasesDGX agent

arXiv:2605.22756v1 Announce Type: new Abstract: Random forests are widely used in fields involving sensitive tabular data, but existing approaches to enforcing differential privacy (DP) typically degr

Machine learning prediction of obstructive coronary artery disease using opportunistic coronary calcium and epicardial fat assessments from CT calcium scoring scans

ResearchDGX agent

arXiv:2605.21762v1 Announce Type: new Abstract: Non-contrast computed tomography calcium scoring (CTCS) is a cost-effective imaging modality widely used to detect coronary artery calcifications. This

MambaGaze: Bidirectional Mamba with Explicit Missing Data Modeling for Cognitive Load Assessment from Eye-Gaze Tracking Data

SafetyDGX agent

arXiv:2605.22775v1 Announce Type: new Abstract: Real-time cognitive load assessment from eye-tracking signals could potentially enable adaptive human-centered-AI such as safety-critical applications s

Manifold-Guided Attention Steering

TutorialsDGX agent

arXiv:2605.21770v1 Announce Type: new Abstract: Large language models frequently produce errors in reasoning tasks despite possessing the underlying knowledge required for correct reasoning. One possi

MapTab: Are MLLMs Ready for Multi-Criteria Route Planning in Heterogeneous Graphs?

Model ReleasesDGX agent

arXiv:2602.18600v3 Announce Type: replace Abstract: Systematic evaluation of Multimodal Large Language Models (MLLMs) is crucial for advancing Artificial General Intelligence (AGI). However, existing

MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models

TutorialsDGX agent

arXiv:2603.16077v3 Announce Type: replace Abstract: Masked diffusion models (MDM) exhibit superior generalization when learned using a Partial masking scheme (Prime). This approach converts tokens int

Measuring Cross-Modal Synergy: A Benchmark for VLM Explainability

Model ReleasesDGX agent

arXiv:2605.22168v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) map complex visual inputs to semantic spaces, but interpreting the cross-modal reasoning of VLMs currently relies on pos

Memory-Efficient LLM Pretraining via Minimalist Optimizer Design

Model ReleasesDGX agent

arXiv:2506.16659v3 Announce Type: replace Abstract: Training large language models (LLMs) relies on adaptive optimizers such as Adam, which introduce extra operations and require significantly more me

Memory-R2: Fair Credit Assignment for Long-Horizon Memory-Augmented LLM Agents

Model ReleasesDGX agent

arXiv:2605.21768v1 Announce Type: new Abstract: Memory-augmented LLM agents enable interactions that extend beyond finite context windows by storing, updating, and reusing information across sessions.

MemReward: Graph-Based Experience Memory for LLM Reward Prediction with Limited Labels

SafetyDGX agent

arXiv:2603.19310v3 Announce Type: replace Abstract: Reinforcement learning has emerged as a powerful paradigm for improving large language model (LLM) reasoning, where rollouts are sampled from the po

MetaDNS: Enhancing Exploration in Discrete Neural Samplers via Well-Tempered Metadynamics

SafetyDGX agent

arXiv:2605.21722v1 Announce Type: cross Abstract: Sampling from discrete distributions with multiple modes and energy barriers is fundamental to machine learning and computational physics. Recent disc

Minimum Description Length based Granular-Ball Tree Regularization for Spectral Clustering

Local AiDGX agent

arXiv:2605.22410v1 Announce Type: new Abstract: Spectral clustering largely depends on the affinity graph, yet constructing a graph that preserves reliable local connectivity while adapting to heterog

MMD-Balls as Credal Sets: A PAC-Bayesian Framework for Epistemic Uncertainty in Test-Time Adaptation

ResearchDGX agent

arXiv:2605.21783v1 Announce Type: new Abstract: Test-time adaptation (TTA) methods improve model performance under distribution shift but lack formal guarantees connecting shift magnitude to predictio

Models Can Model, But Can't Bind: Structured Grounding in Text-to-Optimization

Model ReleasesDGX agent

arXiv:2605.21751v1 Announce Type: new Abstract: Text-to-optimization requires two separable capabilities: modeling -- choosing the right optimization structure -- and binding -- grounding every coeffi

MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents

Model ReleasesDGX agent

arXiv:2602.13372v2 Announce Type: replace-cross Abstract: Evaluating moral alignment in agents navigating conflicting, hierarchically structured human norms is a critical challenge at the intersection

MOSS: Self-Evolution through Source-Level Rewriting in Autonomous Agent Systems

AgentsDGX agent

arXiv:2605.22794v1 Announce Type: cross Abstract: Autonomous agentic systems are largely static after deployment: they do not learn from user interactions, and recurring failures persist until the nex

Multi-Modal Machine Learning for Population- and Subject-Specific lncRNA-Type 2 Diabetes Association Analysis

ResearchDGX agent

arXiv:2605.20747v1 Announce Type: cross Abstract: Long non-coding RNAs (lncRNAs) are emerging regulatory molecules implicated in chronic disease pathogenesis, including Type 2 Diabetes Mellitus (T2D).

Multiple Neural Operators Achieve Near-Optimal Rates for Multi-Task Learning

ResearchDGX agent

arXiv:2605.22724v1 Announce Type: new Abstract: We study the approximation and statistical complexity of learning collections of operators in a shared multi-task setting, with a focus on the Multiple

Near-Optimal Convergence of Accelerated Gradient Methods under Generalized and (L_0, L_1)-Smoothness

ResearchDGX agent

arXiv:2508.06884v2 Announce Type: replace-cross Abstract: We study first-order methods for convex optimization problems with functions f satisfying the recently proposed ell-smoothness condition ||nab

Neural Acceleration for Graph Partitioning

ResearchDGX agent

arXiv:2605.21519v1 Announce Type: cross Abstract: Graph Partitioning is a critical problem in numerous scientific and engineering domains including social network analysis, VLSI design, and many more.

Neural Flow Operators can Approximate any Operator: Abstract Frameworks and Universal Approcimations

ResearchDGX agent

arXiv:2605.22557v1 Announce Type: new Abstract: We introduce an abstract neural flow framework for neural networks and neural operators. The framework contains two continuous-depth models, namely neur

Neuro-Symbolic AI for Analytical Solutions of Differential Equations

ResearchDGX agent

arXiv:2502.01476v4 Announce Type: replace Abstract: Analytical solutions to differential equations offer exact, interpretable insight but are rarely available because discovering them requires expert

No Epoch Like the Present: Robust Climate Emulation Requires Out-of-Distribution Generalisation

ApplicationsDGX agent

arXiv:2605.22248v1 Announce Type: new Abstract: Climate emulation is an out-of-distribution (OOD) projection task. This is precisely the challenge where modern Machine Learning (ML) methods are most p

Noise Schedule Design for Diffusion Models: An Optimal Control Perspective

ResearchDGX agent

arXiv:2605.21911v1 Announce Type: new Abstract: We develop a principled framework for analyzing and designing noise schedules in diffusion models. We show that one can recast this design problem as an

Objective-Induced Bias and Search Dynamics in Multiobjective Unsupervised Feature Selection

SafetyDGX agent

arXiv:2605.21561v1 Announce Type: new Abstract: Unsupervised feature selection is commonly formulated as a multiobjective optimisation problem that jointly optimises subset quality and subset size. Ye

On-Policy Consistency Training Improves LLM Safety with Minimal Capability Degradation

SafetyDGX agent

arXiv:2605.21834v1 Announce Type: new Abstract: Aligned models can misbehave in several ways: they are often sycophantic, fall victim to jailbreaks, or fail to include appropriate safety warnings. Con

On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

Model ReleasesDGX agent

arXiv:2602.12506v3 Announce Type: replace Abstract: Reinforcement learning (RL) finetuning has become a key technique for enhancing large language models (LLMs) on reasoning-intensive tasks, motivatin

← Previous
1…134135136137138…243
Next →