AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
15 May 2026

bde: A Python Package for Bayesian Deep Ensembles via MILE

TutorialsDGX agent

arXiv:2605.14146v1 Announce Type: new Abstract: bde is a user-friendly Python package for Bayesian Deep Ensembles with a particular focus on tabular data. Built on an efficient JAX implementation of t

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2506.05762v5 Announce Type: replace Abstract: Recent advances in offline Reinforcement Learning (RL) have proven that effective policy learning can benefit from imposing conservative constraints

BOOST: A Data-Driven Framework for the Automated Joint Selection of Kernel and Acquisition Functions in Bayesian Optimization

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2508.02332v3 Announce Type: replace Abstract: The performance of Bayesian optimization (BO), a highly sample-efficient method for expensive black-box problems, is critically governed by the sele

Breaking the Reasoning Horizon in Entity Alignment Foundation Models

Local AiDGX agent

arXiv:2601.21174v2 Announce Type: replace Abstract: Entity alignment (EA) is critical for knowledge graph (KG) fusion. Existing EA models lack transferability and are incapable of aligning unseen KGs

CA2: Code-Aware Agent for Automated Game Testing

AgentsDGX agent

arXiv:2605.13918v1 Announce Type: cross Abstract: Automated game testing is important for verifying game functionality, but it remains a costly and time-consuming process. Manual testing often misses

CAKE: Confidence in Assignments via K-partition Ensembles

ApplicationsDGX agent

arXiv:2602.18435v2 Announce Type: replace Abstract: Clustering is widely used for unsupervised structure discovery, yet it offers limited insight into how reliable each individual assignment is. Diagn

Can Stationary Distributions of Scale-Invariant Neural Networks Be Described by the Thermodynamics of an Ideal Gas?

TutorialsDGX agent

arXiv:2511.07308v2 Announce Type: replace Abstract: Understanding the training dynamics of deep neural networks remains a major open problem, with physics-inspired approaches offering promising insigh

Causal Foundation Models with Continuous Treatments

ResearchDGX agent

arXiv:2605.15133v1 Announce Type: new Abstract: Causal inference, estimating causal effects from observational data, is a fundamental tool in many disciplines. Of particular importance across a variet

Causal Multi-Task Demand Learning

ResearchDGX agent

arXiv:2602.09969v2 Announce Type: replace Abstract: We study a canonical multi-task demand-learning problem motivated by retail pricing, where a firm seeks to estimate heterogeneous linear price-respo

Causal Time Series Generation via Diffusion Models

TutorialsDGX agent

arXiv:2509.20846v3 Announce Type: replace Abstract: Time series generation (TSG) synthesizes realistic sequences and has achieved remarkable success. Among TSG, conditional models generate sequences g

Change of measure through the Legendre transform

ResearchDGX agent

arXiv:2202.05568v2 Announce Type: replace-cross Abstract: PAC-Bayes generalisation bounds are derived via change-of-measure inequalities that transfer concentration properties from a reference measure

Closing the Gap on the Sample Complexity of 1-Identification

AgentsDGX agent

arXiv:2601.15620v2 Announce Type: replace Abstract: The 1-identification problem is a fundamental pure-exploration problem in multi-armed bandits. An agent aims to determine whether there exists an ar

CoCo-InEKF: State Estimation with Learned Contact Covariances in Dynamic, Contact-Rich Scenarios

ApplicationsDGX agent

arXiv:2605.15122v1 Announce Type: cross Abstract: Robust state estimation for highly dynamic motion of legged robots remains challenging, especially in dynamic, contact-rich scenarios. Traditional app

Communication-Efficient Federated Fine-Tuning

Model ReleasesDGX agent

arXiv:2505.04535v3 Announce Type: replace Abstract: Federated Learning (FL) enables the utilization of vast, previously inaccessible data sources. At the same time, pre-trained Language Models (LMs) h

Comparative Evaluation of Machine Learning Approaches for Minority-Class Financial Distress Prediction Under Class Imbalance Constraints

ApplicationsDGX agent

arXiv:2605.14067v1 Announce Type: new Abstract: Financial distress prediction remains a significant challenge in enterprise risk analysis due to the highly imbalanced nature of real-world financial da

Composable Crystals: Controllable Materials Discovery via Concept Learning

Local AiDGX agent

arXiv:2605.14769v1 Announce Type: new Abstract: De novo crystal generation, a central task in materials discovery, aims to generate crystals that are simultaneously valid, stable, unique, and novel. E

Conformal Prediction for Multimodal Regression

ResearchDGX agent

arXiv:2410.19653v3 Announce Type: replace Abstract: This paper introduces multimodal conformal regression. Traditionally confined to scenarios with solely numerical input features, conformal predictio

ContextFlow: Context-Aware Flow Matching For Trajectory Inference From Spatial Omics Data

Local AiDGX agent

arXiv:2510.02952v3 Announce Type: replace Abstract: Inferring trajectories from longitudinal spatially-resolved omics data is fundamental to understanding the dynamics of structural and functional tis

Croissant Baker: Metadata Generation for Discoverable, Governable, and Reusable ML Datasets

ResearchDGX agent

arXiv:2605.15079v1 Announce Type: new Abstract: Croissant has emerged as the metadata standard for machine learning datasets, providing a structured, JSON-LD-based format that makes dataset discovery,

Crys-JEPA: Accelerating Crystal Discovery via Embedding Screening and Generative Refinement

ResearchDGX agent

arXiv:2605.14759v1 Announce Type: new Abstract: De novo crystal generation seeks to discover materials that are not merely realistic, but also stable and novel. However, most existing generative model

CSI-JEPA: Towards Foundation Representations for Ubiquitous Sensing with Minimal Supervision

ApplicationsDGX agent

arXiv:2605.14171v1 Announce Type: new Abstract: Channel state information (CSI) provides a widely available sensing modality for human and environment perception, but existing CSI sensing models usual

DeepTokenEEG Enhancing Mild Cognitive Impairment and Alzheimers Classification via Tokenized EEG Features

ResearchDGX agent

arXiv:2605.15009v1 Announce Type: new Abstract: The detection of Alzheimers disease (AD) is considered crucial, as timely intervention can improve patient outcomes. Electroencephalogram (EEG)-based di

Discovering Physical Directions in Weight Space: Composing Neural PDE Experts

Model ReleasesDGX agent

arXiv:2605.14546v1 Announce Type: new Abstract: Recent advances in neural operators have made partial differential equation (PDE) surrogate modeling increasingly scalable and transferable through larg

Distance-Matrix Wasserstein Statistics for Scalable Gromov--Wasserstein Learning

SafetyDGX agent

arXiv:2605.14981v1 Announce Type: new Abstract: Gromov--Wasserstein (GW) distances compare graphs, shapes, and point clouds through internal distances, without requiring a common coordinate system. Th

Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling

AgentsDGX agent

arXiv:2605.14350v1 Announce Type: new Abstract: Multi-task reinforcement learning (MTRL) aims to train a single agent to efficiently optimize performance across multiple tasks simultaneously. However,

DRL-STAF: A Deep Reinforcement Learning Framework for State-Aware Forecasting of Complex Multivariate Hidden Markov Processes

ResearchDGX agent

arXiv:2605.14632v1 Announce Type: new Abstract: Forecasting multivariate hidden Markov processes is challenging due to nonlinear and nonstationary observations, latent state transitions, and cross-seq

Efficient Online Conformal Selection with Limited Feedback

Model ReleasesDGX agent

arXiv:2605.14953v1 Announce Type: new Abstract: We address the problem of conformal selection, where an agent must select a minimal subset of options to ensure that at least one ``success'' is identif

EMA: Efficient Model Adaptation for Learning-based Systems

HardwareDGX agent

arXiv:2605.13942v1 Announce Type: new Abstract: Machine learning (ML) is increasingly applied to optimize system performance in tasks such as resource management and network simulation. Unlike traditi

Embedding Perturbation may Better Reflect Intermediate-Step Uncertainty in LLM Reasoning

ResearchDGX agent

arXiv:2602.02427v2 Announce Type: replace Abstract: Large language Models (LLMs) have achieved significant breakthroughs across diverse domains; however, they can still produce unreliable or misleadin

EnergyLens: Predictive Energy-Aware Exploration for Multi-GPU LLM Inference Optimization

HardwareDGX agent

arXiv:2605.14249v1 Announce Type: new Abstract: We present EnergyLens, an end-to-end framework for energy-aware large language model (LLM) inference optimization. As LLMs scale, predicting and reducin

Enjoy Your Layer Normalization with the Computational Efficiency of RMSNorm

ResearchDGX agent

arXiv:2605.14521v1 Announce Type: new Abstract: Layer normalization (LN) is a fundamental component in modern deep learning, but its per-sample centering and scaling introduce non-negligible inference

Exemplar Partitioning for Mechanistic Interpretability

Model ReleasesDGX agent

arXiv:2605.14347v1 Announce Type: new Abstract: We introduce Exemplar Partitioning (EP), an unsupervised method for constructing interpretable feature dictionaries from large language model activation

Exploring Geographic Relative Space in Large Language Models through Activation Patching

SafetyDGX agent

arXiv:2605.14535v1 Announce Type: new Abstract: The increased use of Large Language Models (LLMs) in geography raises substantial questions about the safety of integrating these tools across a wide ra

Fair and Calibrated Toxicity Detection with Robust Training and Abstention

SafetyDGX agent

arXiv:2605.14074v1 Announce Type: new Abstract: Fairness in toxicity classification involves three integrated axes: ranking, calibration, and abstention. Training-time interventions and post-hoc safet

Fast Adversarial Attacks with Gradient Prediction

ResearchDGX agent

arXiv:2605.14868v1 Announce Type: new Abstract: Generating adversarial examples at scale is a core primitive for robustness evaluation, adversarial training, and red-teaming, yet even 'fast' attacks s

Feature Visualization Recovers Known Cortical Selectivity from TRIBE v2

ResearchDGX agent

arXiv:2605.13904v1 Announce Type: cross Abstract: Brain encoder models predict cortical fMRI responses from the internal activations of pretrained vision and language networks, and are typically evalu

Finite Sample Bounds for Learning with Score Matching

ResearchDGX agent

arXiv:2605.14168v1 Announce Type: new Abstract: Learning of continuous exponential family distributions with unbounded support remains an important area of research for both theory and applications in

Focused PU learning from imbalanced data

ApplicationsDGX agent

arXiv:2605.14467v1 Announce Type: new Abstract: We propose a new method of learning from positive and unlabeled (PU) examples in highly imbalanced datasets. Many real-world problems, such as disease g

ForcingDAS: Unified and Robust Data Assimilation via Diffusion Forcing

ApplicationsDGX agent

arXiv:2605.14285v1 Announce Type: cross Abstract: Data assimilation (DA) estimates the state of an evolving dynamical system from noisy, partial observations, and is widely used in scientific simulati

Frequency-adaptive tensor neural networks for high-dimensional multi-scale problems

ResearchDGX agent

arXiv:2508.15198v2 Announce Type: replace Abstract: Tensor neural networks (TNNs) have demonstrated their superiority in solving high-dimensional problems. However, similar to conventional neural netw

From Data to Action: Accelerating Refinery Optimization with AI

ResearchDGX agent

arXiv:2605.15085v1 Announce Type: cross Abstract: Nowadays refinery optimization utilizes sheer amounts of data, which can be handled with modern Linear Programming (LP) software, but the interpreting

FrontierSmith: Synthesizing Open-Ended Coding Problems at Scale

ApplicationsDGX agent

arXiv:2605.14445v1 Announce Type: new Abstract: Many real-world coding challenges are open-ended and admit no known optimal solution. Yet, recent progress in LLM coding has focused on well-defined tas

Functional-level Uncertainty Quantification for Calibrated Fine-tuning on LLMs

ResearchDGX agent

arXiv:2410.06431v5 Announce Type: replace Abstract: Accurate uncertainty quantification in large language models (LLMs) is essential for reliable confidence estimation, yet fine-tuned LLMs often becom

GenAI for Energy-Efficient and Interference-Aware Compressed Sensing of GNSS Signals on a Google Edge TPU

HardwareDGX agent

arXiv:2605.14839v1 Announce Type: new Abstract: Traditional methods for classifying global navigation satellite system (GNSS) jamming signals typically involve post-processing raw or spectral data str

Generalizing Score-based generative models for Heavy-tailed Distributions

ResearchDGX agent

arXiv:2603.00772v2 Announce Type: replace-cross Abstract: Score-based generative models (SGMs) have achieved remarkable empirical success, motivating their application to a broad range of data distrib

Generative Bayesian Optimization: Generative Models as Acquisition Functions

ResearchDGX agent

arXiv:2510.25240v3 Announce Type: replace-cross Abstract: We present a general strategy for turning generative models into candidate solution samplers for batch Bayesian optimization (BO). The use of

GFMate: Empowering Graph Foundation Models with Test-time Prompt Tuning

Model ReleasesDGX agent

arXiv:2605.14809v1 Announce Type: new Abstract: Graph prompt tuning has shown great potential in graph learning by introducing trainable prompts to enhance the model performance in conventional single

Guided Diffusion Sampling for Precipitation Forecast Interventions

ResearchDGX agent

arXiv:2605.14317v1 Announce Type: new Abstract: Extreme precipitation causes severe societal and economic damage, and weather control has long been discussed as a potential mitigation strategy. Howeve

Hand-in-the-Loop: Improving Dexterous VLA via Seamless Interventional Correction

SafetyDGX agent

arXiv:2605.15157v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are prone to compounding errors in dexterous manipulation, where high-dimensional action spaces and contact-rich d

How to Scale Mixture-of-Experts: From muP to the Maximally Scale-Stable Parameterization

TutorialsDGX agent

arXiv:2605.14200v1 Announce Type: new Abstract: Recent frontier large language models predominantly rely on Mixture-of-Experts (MoE) architectures. Despite empirical progress, there is still no princi

How well behaved is finite dimensional Diffusion Maps?

ResearchDGX agent

arXiv:2412.03992v3 Announce Type: replace-cross Abstract: Under a set of assumptions on a family of submanifolds subset {mathbb R}^D, we derive a series of geometric properties that remain valid after

Hyperbolic Graph Neural Networks Under the Microscope: The Role of Geometry-Task Alignment

SafetyDGX agent

arXiv:2602.01828v2 Announce Type: replace Abstract: Many complex networks exhibit hierarchical, tree-like structures, making hyperbolic space a natural candidate wherein to learn representations of th

In-Context Learning for Data-Driven Censored Inventory Control

Model ReleasesDGX agent

arXiv:2605.14840v1 Announce Type: new Abstract: We study inventory control with decision-dependent censoring, focusing on the censored or repeated newsvendor (R-NV), where each order quantity determin

Indian Wedding System Optimization (IWSO): A Novel Socially Inspired Metaheuristic with Operational Design and Analysis

Model ReleasesDGX agent

arXiv:2605.13871v1 Announce Type: cross Abstract: This paper presents a novel population-based metaheuristic, Indian Wedding System Optimization (IWSO), inspired by the socio-cultural dynamics of trad

InfoSFT: Learn More and Forget Less with Information-Aware Token Weighting

SafetyDGX agent

arXiv:2605.14967v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) provides the standard approach for teaching LLMs new behaviors from offline expert demonstrations. However, standard SFT un

IsoNet: Spatially-aware audio-visual target speech extraction in complex acoustic environments

ResearchDGX agent

arXiv:2605.14736v1 Announce Type: cross Abstract: Target speech extraction remains difficult for compact devices because monaural neural models lack spatial evidence and classical beamformers lose res

K-Models: a Flexible and Interpretable Method for Ordinal Clustering with Application to Antigen-Antibody Interaction Profiles

Model ReleasesDGX agent

arXiv:2605.14828v1 Announce Type: cross Abstract: Existing clustering methods for functional data often prioritize partitioning accuracy over interpretability, making it challenging to extract meaning

Kairos: Toward Adaptive and Parameter-Efficient Time Series Foundation Models

Model ReleasesDGX agent

arXiv:2509.25826v3 Announce Type: replace Abstract: Inherent temporal heterogeneity, such as varying sampling densities and periodic structures, has posed substantial challenges in zero-shot generaliz

Kolmogorov-Arnold Chemical Reaction Neural Networks for learning pressure-dependent kinetic rate laws

Model ReleasesDGX agent

arXiv:2511.07686v2 Announce Type: replace-cross Abstract: Chemical Reaction Neural Networks (CRNNs) have emerged as an interpretable machine learning framework for discovering reaction kinetics direct

Lang2MLIP: End-to-End Language-to-Machine Learning Interatomic Potential Development with Autonomous Agentic Workflows

AgentsDGX agent

arXiv:2605.14527v1 Announce Type: new Abstract: Developing machine learning interatomic potentials (MLIPs) for complex materials systems remains challenging because it requires expertise in atomistic

← Previous
1…154155156157158…243
Next →