AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
13 May 2026

Exploring Token-Space Manipulation in Latent Audio Tokenizers

ResearchDGX agent

arXiv:2605.11192v1 Announce Type: cross Abstract: Neural audio codecs provide compact discrete representations for speech generation and manipulation. However, most codecs organize tokens as frame-lev

Extending Kernel Trick to Influence Functions

Model ReleasesDGX agent

arXiv:2605.11239v1 Announce Type: new Abstract: In this paper, we present a dual representation of the influence functions, whose computational complexity scales with dataset size rather than model si

Fair Conformal Classification via Learning Representation-Based Groups

SafetyDGX agent

arXiv:2605.12195v1 Announce Type: new Abstract: Conformal prediction methods provide statistically rigorous marginal coverage guarantees for machine learning models, but such guarantees fail to accoun


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Fast MoE Inference via Predictive Prefetching and Expert Replication

HardwareDGX agent

arXiv:2605.11537v1 Announce Type: new Abstract: The Mixture of Experts (MoE) architecture has become a fundamental building block in state-of-the-art large language models (LLMs), improving domain-spe

FastUMAP: Scalable Dimensionality Reduction via Bipartite Landmark Sampling

Model ReleasesDGX agent

arXiv:2605.11428v1 Announce Type: new Abstract: Exploratory analysis of high-dimensional data rarely stops at a single embedding. In practice, analysts rerun dimensionality reduction after changing pr

Fed-BAC: Federated Bandit-Guided Additive Clustering in Hierarchical Federated Learning

SafetyDGX agent

arXiv:2605.11815v1 Announce Type: new Abstract: Hierarchical federated learning (HFL) leverages edge servers for partial aggregation in edge computing. Yet existing FL methods lack mechanisms for join

Federated Client Selection under Partial Visibility: A POMDP Approach with Spatio-Temporal Attention

ResearchDGX agent

arXiv:2605.11752v1 Announce Type: new Abstract: Federated learning relies on effective client selection to alleviate the performance degradation caused by data heterogeneity. Most existing methods ass

FedOUI: OUI-Guided Client Weighting for Federated Aggregation

SafetyDGX agent

arXiv:2605.11571v1 Announce Type: new Abstract: Federated learning usually aggregates client updates using dataset size or gradient-level criteria, while overlooking internal signals about how each cl

FedRot-LoRA: Mitigating Rotational Misalignment in Federated LoRA

Local AiDGX agent

arXiv:2602.23638v2 Announce Type: replace Abstract: Federated LoRA provides a communication-efficient mechanism for fine-tuning large language models on decentralized data. In practice, however, a dis

FedSurrogate: Backdoor Defense in Federated Learning via Layer Criticality and Surrogate Replacement

SafetyDGX agent

arXiv:2605.11122v1 Announce Type: cross Abstract: Federated Learning remains highly susceptible to backdoor attacks--malicious clients inject targeted behaviours into the global model. Existing defens

FERMI: Exploiting Relations for Membership Inference Against Tabular Diffusion Models

TutorialsDGX agent

arXiv:2605.11527v1 Announce Type: new Abstract: Diffusion models are the leading approach for tabular data synthesis and are increasingly used to share sensitive records. Whether they actually protect

Finite and Corruption-Robust Regret Bounds in Online Inverse Linear Optimization under M-Convex Action Sets

AgentsDGX agent

arXiv:2602.01682v2 Announce Type: replace Abstract: We study online inverse linear optimization, also known as contextual recommendation, where a learner sequentially infers an agent's hidden objectiv

Finite Sentence-Interface Control for Learning Bounded-Fan-Out Linear MCFGs under Fixed Monoid Typing

ResearchDGX agent

arXiv:2605.11644v1 Announce Type: cross Abstract: We study positive-data learning of bounded-fan-out linear multiple context-free grammars under a fixed explicit finite monoid homomorphism (h). The ma

Finite Volume-Informed Neural Network Framework for 2D Shallow Water Equations: Rugged Loss Landscapes and the Importance of Data Guidance

Model ReleasesDGX agent

arXiv:2605.11001v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) are a simple surrogate-modelling paradigm for partial differential equations, but their standard strong-form re

FLARE: Adaptive Multi-Dimensional Reputation for Robust Client Reliability in Federated Learning

Model ReleasesDGX agent

arXiv:2511.14715v3 Announce Type: replace Abstract: Federated learning (FL) enables collaborative model training while preserving data privacy. However, it remains vulnerable to malicious clients who

Focusing Influence Mechanism for Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2506.19417v2 Announce Type: replace Abstract: Cooperative multi-agent reinforcement learning (MARL) under sparse rewards remains fundamentally challenging because agents often fail to concentrat

Fractal Graph Contrastive Learning

ApplicationsDGX agent

arXiv:2505.11356v4 Announce Type: replace Abstract: Graph Contrastive Learning (GCL) relies on semantically consistent graph augmentations, but common local perturbations provide limited control over

From Generic Correlation to Input-Specific Credit in On-Policy Self Distillation

SafetyDGX agent

arXiv:2605.11613v1 Announce Type: new Abstract: On-policy self-distillation has emerged as a promising paradigm for post-training language models, in which the model conditions on environment feedback

From Message-Passing to Linearized Graph Sequence Models

SafetyDGX agent

arXiv:2605.12358v1 Announce Type: new Abstract: Message-passing based approaches form the default backbone of most learning architectures on graph-structured data. However, the rapid progress of moder

From Observations to States: Latent Time Series Forecasting

TutorialsDGX agent

arXiv:2602.00297v2 Announce Type: replace Abstract: Deep learning has achieved strong performance in Time Series Forecasting (TSF). However, we identify a critical representation paradox, termed Laten

From raw data to neutrino candidates: a neural-network pipeline for Baikal-GVD

ResearchDGX agent

arXiv:2605.11176v1 Announce Type: cross Abstract: We present a neural-network-based data processing pipeline for Baikal-GVD, designed to improve event reconstruction quality and accelerate neutrino ca

Fused Gromov-Wasserstein Distance with Feature Selection

SafetyDGX agent

arXiv:2605.12161v1 Announce Type: new Abstract: Fused Gromov-Wasserstein (FGW) distances provide a principled framework for comparing objects by jointly aligning structure and node features. However,

Generative climate downscaling enables high-resolution compound risk assessment by preserving multivariate dependencies

SafetyDGX agent

arXiv:2605.11531v1 Announce Type: cross Abstract: Physics-based climate projections using general circulation models are essential for assessing future risks, but their coarse resolution limits region

Generative Diffusion Prior Distillation for Long-Context Knowledge Transfer

Model ReleasesDGX agent

arXiv:2605.11414v1 Announce Type: new Abstract: While traditional time-series classifiers assume full sequences at inference, practical constraints (latency and cost) often limit inputs to partial pre

GeneZip: Region-Aware Compression for Long Context DNA Modeling

Model ReleasesDGX agent

arXiv:2602.17739v3 Announce Type: replace-cross Abstract: Long-context DNA models are limited by token-mixing cost and by how compression allocates representational budget across the genome. Existing

GeomHerd: A Forward-looking Herding Quantification via Ricci Flow Geometry on Agent Interactive Simulations

Model ReleasesDGX agent

arXiv:2605.11645v1 Announce Type: cross Abstract: Herding -- where agents align their behaviors and act collectively -- is a central driver of market fragility and systemic risk. Existing approaches t

Gradient Clipping Beyond Vector Norms: A Spectral Approach for Matrix-Valued Parameters

ResearchDGX agent

arXiv:2605.11838v1 Announce Type: new Abstract: Gradient clipping is a standard safeguard for training neural networks under noisy, heavy-tailed stochastic gradients; yet, most clipping rules treat al

GRAFT-ATHENA: Self-Improving Agentic Teams for Autonomous Discovery and Evolutionary Numerical Algorithms

Model ReleasesDGX agent

arXiv:2605.11117v1 Announce Type: new Abstract: Scientific discovery can be modeled as a sequence of probabilistic decisions that map physical problems to numerical solutions. Recent agentic AI system

GRAFT: Graph-Tokenized LLMs for Tool Planning

SafetyDGX agent

arXiv:2605.11706v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to complete complex tasks by selecting and coordinating external tools across multiple steps. This re

Grid Games: The Power of Multiple Grids for Quantizing Large Language Models

Model ReleasesDGX agent

arXiv:2605.12327v1 Announce Type: new Abstract: A major recent advance in quantization is given by microscaled 4-bit formats such as NVFP4 and MXFP4, quantizing values into small groups sharing a scal

gym-invmgmt: An Open Benchmarking Framework for Inventory Management Methods

Model ReleasesDGX agent

arXiv:2605.11355v1 Announce Type: new Abstract: Inventory-policy comparisons are often difficult to interpret because performance depends on the evaluation contract as much as on the policy itself. Di

HEPA: A Self-Supervised Horizon-Conditioned Event Predictive Architecture for Time Series

ResearchDGX agent

arXiv:2605.11130v1 Announce Type: new Abstract: Critical events in multivariate time series, from turbine failures to cardiac arrhythmias, demand accurate prediction, yet labeled data is scarce becaus

Hierarchical LLM-Driven Control for HAPS-Assisted UAV Networks: Joint Optimization of Flight and Connectivity

Local AiDGX agent

arXiv:2605.11509v1 Announce Type: cross Abstract: Uncrewed aerial vehicles (UAVs) are increasingly deployed in complex networked environments, yet the joint optimization of multi-UAV motion control an

Hierarchical Multi-Scale Graph Neural Networks: Scalable Heterophilous Learning with Oversmoothing and Oversquashing Mitigation

ApplicationsDGX agent

arXiv:2605.10975v1 Announce Type: new Abstract: Graphs with heterophily, where adjacent nodes carry different labels, are prevalent in real-world applications, from social networks to molecular intera

High-arity Sample Compression

ResearchDGX agent

arXiv:2605.12465v1 Announce Type: new Abstract: Recently, a series of works have started studying variations of concepts from learning theory for product spaces, which can be collected under the name

Hindsight Hint Distillation: Scaffolded Reasoning for SWE Agents from CoT-free Answers

SafetyDGX agent

arXiv:2605.11556v1 Announce Type: cross Abstract: Solving complex long-horizon tasks requires strong planning and reasoning capabilities. Although datasets with explicit chain-of-thought (CoT) rationa

Holder Policy Optimisation

Model ReleasesDGX agent

arXiv:2605.12058v1 Announce Type: new Abstract: Group Relative Policy Optimisation (GRPO) enhances large language models by estimating advantages across a group of sampled trajectories. However, mappi

Hypernetworks for Dynamic Feature Selection

ResearchDGX agent

arXiv:2605.12278v1 Announce Type: new Abstract: Dynamic feature selection (DFS) is a machine learning framework in which features are acquired sequentially for individual samples under budget constrai

Improving the Accuracy of Amortized Model Comparison with Self-Consistency

Model ReleasesDGX agent

arXiv:2508.20614v3 Announce Type: replace-cross Abstract: Amortized Bayesian model comparison (BMC) enables fast probabilistic ranking of models via simulation-based training of neural surrogates. How

Improving the Performance and Learning Stability of Parallelizable RNNs Designed for Ultra-Low Power Applications

ResearchDGX agent

arXiv:2605.11855v1 Announce Type: new Abstract: Sequence learning is dominated by Transformers and parallelizable recurrent neural networks (RNNs) such as state-space models, yet learning long-term de

In-context learning to predict critical transitions in dynamical systems

ApplicationsDGX agent

arXiv:2605.12308v1 Announce Type: new Abstract: Critical transitions - abrupt, often irreversible changes in system dynamics - arise across human and natural systems, often with catastrophic consequen

In-Context Multi-Objective Optimization

SafetyDGX agent

arXiv:2512.11114v2 Announce Type: replace Abstract: Balancing competing objectives is omnipresent across disciplines, from drug design to autonomous systems. Multi-objective Bayesian optimization is a

Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning

SafetyDGX agent

arXiv:2605.11889v1 Announce Type: new Abstract: Collaborative machine learning involves training high-quality models using datasets from a number of sources. To incentivize sources to share data, exis

Information-Theoretic Generalization Bounds for Sequential Decision Making

ResearchDGX agent

arXiv:2605.12190v1 Announce Type: cross Abstract: Information-theoretic generalization bounds based on the supersample construction are a central tool for algorithm-dependent generalization analysis i

Information theoretic underpinning of self-supervised learning by clustering

ResearchDGX agent

arXiv:2605.11870v1 Announce Type: new Abstract: Self-supervised learning (SSL) is recognized as an essential tool for building foundation models for Artificial Intelligence applications. The advances

Instruction Lens Score: Your Instruction Contributes a Powerful Object Hallucination Detector for Multimodal Large Language Models

Local AiDGX agent

arXiv:2605.12258v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved remarkable progress, yet the object hallucination remains a critical challenge for reliable deplo

Integral Imprecise Probability Metrics

ResearchDGX agent

arXiv:2505.16156v3 Announce Type: replace-cross Abstract: Quantifying differences between probability distributions is fundamental to statistics and machine learning, primarily for comparing statistic

Intention-Conditioned Flow Occupancy Models

Model ReleasesDGX agent

arXiv:2506.08902v4 Announce Type: replace Abstract: Large-scale pre-training has fundamentally changed how machine learning research is done today: large foundation models are trained once, and then c

Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning

SafetyDGX agent

arXiv:2605.11235v1 Announce Type: new Abstract: In LLM Reinforcement Fine-Tuning (RFT), curriculum learning drives both efficiency and performance. Yet, current methods externalize curriculum judgment

Interpretability Can Be Actionable

ApplicationsDGX agent

arXiv:2605.11161v1 Announce Type: new Abstract: Interpretability aims to explain the behavior of deep neural networks. Despite rapid growth, there is mounting concern that much of this work has not tr

Interpretable EEG Microstate Discovery via Variational Deep Embedding: A Systematic Architecture Search with Multi-Quadrant Evaluation

ResearchDGX agent

arXiv:2605.10947v1 Announce Type: new Abstract: EEG microstate analysis segments continuous brain electrical activity into brief, quasi-stable topographic configurations that reflect discrete function

Interpretable Machine Learning for Spatial Science: A Lie-Algebraic Kernel for Rotationally Anisotropic Gaussian Processes

ResearchDGX agent

arXiv:2605.11179v1 Announce Type: cross Abstract: Many three-dimensional spatial fields are anisotropic, with directions of rapid and slow variation that need not align with the coordinate axes. Stand

Interpretable rainfall modelling reveals rapid reorganisation of Amazonian rainfall under vegetation loss

ResearchDGX agent

arXiv:2605.10948v1 Announce Type: cross Abstract: Understanding how vegetation loss alters rainfall remains a major challenge in climate and hydrological science, as deforestation modifies precipitati

Intrinsic Vicarious Conditioning for Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.12224v1 Announce Type: new Abstract: Advancements in reinforcement learning have produced a variety of complex and useful intrinsic driving forces; crucially, these drivers operate under a

Investigating simple target-covariate relationships for Chronos-2 and TabPFN-TS

Model ReleasesDGX agent

arXiv:2605.12200v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have recently achieved state-of-the-art performance, often outperforming supervised models in zero-shot settings.

Is Monotonic Sampling Necessary in Diffusion Models?

ResearchDGX agent

arXiv:2605.11773v1 Announce Type: new Abstract: Diffusion models generate samples by iteratively denoising a Gaussian prior, traversing a sequence of noise levels that, in every published sampler, dec

Joint Learning of Hierarchical Neural Options and Abstract World Model

AgentsDGX agent

arXiv:2602.02799v2 Announce Type: replace Abstract: Building agents that can perform new skills by composing existing skills is a long-standing goal of AI agent research. Towards this end, we investig

Keeping Score: Efficiency Improvements in Neural Likelihood Surrogate Training via Score-Augmented Loss Functions

Model ReleasesDGX agent

arXiv:2605.12118v1 Announce Type: cross Abstract: For stochastic process models, parameter inference is often severely bottlenecked by computationally expensive likelihood functions. Simulation-based

Language Modeling with Hyperspherical Flows

ResearchDGX agent

arXiv:2605.11125v1 Announce Type: new Abstract: Discrete Diffusion Language Models progressed rapidly as an alternative to autoregressive (AR) models, motivated by their parallel generation abilities.

Latent Chain-of-Thought Improves Structured-Data Transformers

ResearchDGX agent

arXiv:2605.11262v1 Announce Type: new Abstract: Chain-of-thought and more broadly test-time compute are known to augment the expressive capabilities of language models and have led to major innovation

← Previous
1…163164165166167…243
Next →