AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
13 May 2026

LEAP: Unlocking dLLM Parallelism via Lookahead Early-Convergence Token Detection

SafetyDGX agent

arXiv:2605.10980v1 Announce Type: new Abstract: Diffusion Language Models (dLLMs) have garnered significant attention for their potential in highly parallel processing. The parallel capabilities of ex

Learning Compact Boolean Networks

Model ReleasesDGX agent

arXiv:2602.05830v2 Announce Type: replace-cross Abstract: Floating-point neural networks dominate modern machine learning but incur substantial inference costs, motivating emerging interest in Boolean

Learning density ratios in causal inference using Bregman-Riesz regression

TutorialsDGX agent

arXiv:2510.16127v2 Announce Type: replace-cross Abstract: The ratio of two probability density functions is a fundamental quantity that appears in many areas of statistics and machine learning, includ


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Learning, Fast and Slow: Towards LLMs That Adapt Continually

Model ReleasesDGX agent

arXiv:2605.12484v1 Announce Type: new Abstract: Large language models (LLMs) are trained for downstream tasks by updating their parameters (e.g., via RL). However, updating parameters forces them to a

Learning Feature Encoder with Synthetic Anomalies for Weakly Supervised Graph Anomaly Detection

TutorialsDGX agent

arXiv:2605.11749v1 Announce Type: new Abstract: Weakly supervised graph anomaly detection aims to unveil unusual graph instances, e.g., nodes, whose behaviors significantly differ from normal ones, gi

Learning Minimally Rigid Graphs with High Realization Counts

SafetyDGX agent

arXiv:2605.12427v1 Announce Type: new Abstract: For minimally rigid graphs, the same edge-length data can admit multiple realizations (up to translations and rotations). Finding graphs with exceptiona

Learning plug-in surrogate endpoints for randomized experiments

ApplicationsDGX agent

arXiv:2605.12051v1 Announce Type: new Abstract: Surrogate endpoints are used in place of long-term outcomes in randomized experiments when observing the real outcome for a large enough cohort is prohi

Learning U-Statistics with Active Inference

ResearchDGX agent

arXiv:2605.11638v1 Announce Type: cross Abstract: U-statistics play a central role in statistical inference. In many modern applications, however, acquiring the labels required for U-statistics is cos

Learning Weakly Communicating Average-Reward CMDPs: Strong Duality and Improved Regret

ResearchDGX agent

arXiv:2605.11586v1 Announce Type: new Abstract: We study infinite-horizon average-reward constrained Markov decision processes (CMDPs) under the weakly communicating assumption. Our contributions are

Learning What Matters: Adaptive Information-Theoretic Objectives for Robot Exploration

Model ReleasesDGX agent

arXiv:2605.12084v1 Announce Type: cross Abstract: Designing learnable information-theoretic objectives for robot exploration remains challenging. Such objectives aim to guide exploration toward data t

Leveraging RAG for Training-Free Alignment of LLMs

SafetyDGX agent

arXiv:2605.11217v1 Announce Type: new Abstract: Large language model (LLM) alignment algorithms typically consist of post-training over preference pairs. While such algorithms are widely used to enabl

LiBaGS: Lightweight Boundary Gap Synthesis for Targeted Synthetic Data Selection

ResearchDGX agent

arXiv:2605.11231v1 Announce Type: new Abstract: Synthetic data is useful only when the added samples fill missing parts of the training distribution that matter for the downstream task. We introduce L

Limits of Learning Linear Dynamics from Experiments

ResearchDGX agent

arXiv:2605.12010v1 Announce Type: new Abstract: Learning governing dynamics from data is a common goal across the sciences, yet it is only well-posed when the underlying mechanisms are identifiable. I

Local and Mixing-Based Algorithms for Gaussian Graphical Model Selection from Glauber Dynamics

Local AiDGX agent

arXiv:2412.18594v3 Announce Type: replace Abstract: Gaussian graphical model selection is usually studied under independent sampling, but in many applications observations arise from dependent dynamic

Localising Dropout Variance in Twin Networks

ApplicationsDGX agent

arXiv:2507.03622v2 Announce Type: replace Abstract: Accurate individual treatment-effect estimation demands not only reliable point predictions but also uncertainty measures that help practitioners lo

Localization Boosting for Growth Markets: Mitigating Cross-Locale Behavioral Bias in Learning-to-Rank

Local AiDGX agent

arXiv:2605.11272v1 Announce Type: new Abstract: Adobe Express is expanding internationally, but the US has a disproportionately large content supply and interaction volume. Learning-to-rank (LTR) mode

LOFT: Low-Rank Orthogonal Fine-Tuning via Task-Aware Support Selection

Model ReleasesDGX agent

arXiv:2605.11872v1 Announce Type: new Abstract: Orthogonal parameter-efficient fine-tuning (PEFT) adapts pretrained weights through structure-preserving multiplicative transformations, but existing me

LoopUS: Recasting Pretrained LLMs into Looped Latent Refinement Models

ResearchDGX agent

arXiv:2605.11011v1 Announce Type: new Abstract: Looped computation shows promise in improving the reasoning-oriented performance of LLMs by scaling test-time compute. However, existing approaches typi

Lower bounds for one-layer transformers that compute parity

ResearchDGX agent

arXiv:2605.12171v1 Announce Type: new Abstract: This note shows that no self-attention layer post-processed by a rational function can sign-represent the parity function unless the product of the numb

LPDP: Inference-Time Reward Control for Variable-Length DNA Generation with Edit Flows

Local AiDGX agent

arXiv:2605.11368v1 Announce Type: new Abstract: We study the application of recent Edit Flows for inference-time reward control for DNA sequence generation. Unlike most reward-guided DNA generation fr

MAC: Masked Agent Collaboration Boosts Large Language Model Medical Decision-Making

AgentsDGX agent

arXiv:2507.21159v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have proven effective in artificial intelligence, where the multi-agent system (MAS) holds considerable promise f

Machine Learning for neutron source distributions

ResearchDGX agent

arXiv:2605.12165v1 Announce Type: cross Abstract: In light of the recent advancements in machine learning, we propose a novel approach to neutron source distribution estimation through the utilisation

Make It Long, Keep It Fast: End-to-End 10k-Sequence Modeling at Billion Scale on Douyin Recommendation

ApplicationsDGX agent

arXiv:2511.06077v2 Announce Type: replace Abstract: Short-video recommenders such as Douyin must exploit extremely long user histories without breaking latency or cost budgets. We present an end-to-en

MambaNetBurst: Direct Byte-level Network Traffic Classification without Tokenization or Pretraining

ResearchDGX agent

arXiv:2605.11034v1 Announce Type: cross Abstract: We present MambaNetBurst, a compact tokenizer-free byte-level sequence classifier for network burst classification based on a Mamba-2 backbone. In con

Manifold Sampling via Entropy Maximization

ResearchDGX agent

arXiv:2605.12338v1 Announce Type: new Abstract: Sampling from constrained distributions has a wide range of applications, including in Bayesian optimization and robotics. Prior work establishes conver

Martingale-Consistent Self-Supervised Learning

ResearchDGX agent

arXiv:2605.11846v1 Announce Type: new Abstract: Self-supervised learning (SSL) is often deployed under changing information, such as shorter histories, missing features, or partially observed images.

Maximin Robust Bayesian Experimental Design

SafetyDGX agent

arXiv:2603.14094v2 Announce Type: replace-cross Abstract: We address the brittleness of Bayesian experimental design under model misspecification by formulating the problem as a max--min game between

MCPShield: Content-Aware Attack Detection for LLM Agent Tool-Call Traffic

AgentsDGX agent

arXiv:2605.11053v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) has become a widely adopted interface for LLM agents to invoke external tools, yet learned monitoring of MCP tool-cal

Measuring Accuracy and Energy-to-Solution of Quantum Fine-Tuning of Foundational AI Models

ApplicationsDGX agent

arXiv:2605.02798v1 Announce Type: cross Abstract: We present an experimental study of energy-to-solution (ETS) of hybrid quantum-classical applications, enabled by direct instrumentation of power cons

Measuring Five-Nines Reliability: Sample-Efficient LLM Evaluation in Saturated Benchmarks

Model ReleasesDGX agent

arXiv:2605.11209v1 Announce Type: new Abstract: While existing benchmarks demonstrate the near-perfect performance of large language models (LLMs) on various tasks, this apparent saturation often obsc

MetaColloc: Optimization-Free PDE Solving via Meta-Learned Basis Functions

ResearchDGX agent

arXiv:2605.12368v1 Announce Type: new Abstract: Solving partial differential equations (PDEs) with machine learning typically requires training a new neural network for every new equation. This optimi

Minimax Rates and Spectral Distillation for Tree Ensembles

ResearchDGX agent

arXiv:2605.11841v1 Announce Type: cross Abstract: Tree ensembles such as random forests (RFs) and gradient boosting machines (GBMs) are among the most widely used supervised learners, yet their theore

Missing Old Logits in Asynchronous Agentic RL: Semantic Mismatch and Repair Methods for Off-Policy Correction

SafetyDGX agent

arXiv:2605.12070v1 Announce Type: new Abstract: Asynchronous reinforcement learning improves rollout throughput for large language model agents by decoupling sample generation from policy optimization

Missingness-MDPs: Bridging the Theory of Missing Data and POMDPs

ResearchDGX agent

arXiv:2605.12262v1 Announce Type: cross Abstract: We introduce missingness-MDPs (miss-MDPs), a novel subclass of partially observable Markov decision processes (POMDPs) that incorporates the theory of

MIST: Reliable Streaming Decision Trees for Online Class-Incremental Learning via McDiarmid Bound

ResearchDGX agent

arXiv:2605.11617v1 Announce Type: new Abstract: Streaming decision trees are natural candidates for open-world continual learning, as they perform local updates, enjoy bounded memory, and static decis

MLCommons Chakra: Advancing Performance Benchmarking and Co-design using Standardized Execution Traces

HardwareDGX agent

arXiv:2605.11333v1 Announce Type: cross Abstract: The fast pace of artificial intelligence~(AI) innovation demands an agile methodology for observation, reproduction and optimization of distributed ma

Model-based Bootstrap of Controlled Markov Chains

SafetyDGX agent

arXiv:2605.12410v1 Announce Type: cross Abstract: We propose and analyze a model-based bootstrap for transition kernels in finite controlled Markov chains (CMCs) with possibly nonstationary or history

Model-Level GNN Explanations via Rule-to-Graph Readout for Logit Reconstruction

ApplicationsDGX agent

arXiv:2503.09051v2 Announce Type: replace Abstract: We propose a novel model-level GNN explanation framework that shifts the explanation target from class-wise rule extraction to rule-based logit reco

Molecular Design beyond Training Data with Novel Extended Objective Functionals of Generative AI Models Driven by Quantum Annealing Computer

ResearchDGX agent

arXiv:2602.15451v3 Announce Type: replace-cross Abstract: Deep generative modeling to stochastically design small molecules is an emerging technology for accelerating drug discovery and development. H

More Than Meets the Eye: A Semantics-Aware Traffic Augmentation Framework for Generalizable Website Fingerprinting

SafetyDGX agent

arXiv:2605.11402v1 Announce Type: new Abstract: Deep learning-based website fingerprinting has emerged as an effective technique for inferring the websites users visit. Although existing methods achie

Multi-Agent System Identification with Nonlinear Sheaf Diffusion

AgentsDGX agent

arXiv:2605.11204v1 Announce Type: cross Abstract: Local interaction laws governing multi-agent systems can be difficult to recover from trajectory data, even when the dynamics are observed faithfully.

Multi-modal Bayesian Neural Network Surrogates with Conjugate Last-Layer Estimation

TutorialsDGX agent

arXiv:2509.21711v2 Announce Type: replace-cross Abstract: As data collection and simulation capabilities advance, multi-modal learning, the task of learning from multiple modalities and sources of dat

Multi-Narrow Transformation as a Single-Model Ensemble: Boundary Conditions, Mechanisms, and Failure Modes

Model ReleasesDGX agent

arXiv:2605.11530v1 Announce Type: new Abstract: Single-model ensembles (SMEs) have attracted attention as a way to approximate some of the benefits of deep ensembles within a single network. However,

Multi-Task Representation Learning for Conservative Linear Bandits

Model ReleasesDGX agent

arXiv:2605.12176v1 Announce Type: new Abstract: This paper presents the Constrained Multi-Task Representation Learning (CMTRL) framework for linear bandits. We consider T linear bandit tasks in a d di

Multi-Timescale Conductance Spiking Networks: A Sparse, Gradient-Trainable Framework with Rich Firing Dynamics for Enhanced Temporal Processing

ResearchDGX agent

arXiv:2605.11835v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) promise low-power event-driven computation for temporally rich tasks, but commonly used neuron models often trade off g

Multi-Variable Conformal Prediction: Optimizing Prediction Sets without Data Splitting

AgentsDGX agent

arXiv:2605.12341v1 Announce Type: cross Abstract: Conformal prediction constructs prediction sets with finite-sample coverage guarantees, but its calibration stage is structurally constrained to a sca

Muon is Not That Special: Random or Inverted Spectra Work Just as Well

Local AiDGX agent

arXiv:2605.11181v1 Announce Type: new Abstract: The recent empirical success of the Muon optimizer has renewed interest in non-Euclidean optimization, typically justified by similarities with second-o

MuonQ: Enhancing Low-Bit Muon Quantization via Directional Fidelity Optimization

Model ReleasesDGX agent

arXiv:2605.11396v1 Announce Type: new Abstract: The Muon optimizer has emerged as a compelling alternative to Adam for training large language models, achieving remarkable computational savings throug

Neural ARFIMA model for forecasting BRIC exchange rates with long memory

Model ReleasesDGX agent

arXiv:2509.06697v2 Announce Type: replace-cross Abstract: Accurate forecasting of exchange rates remains a persistent challenge, particularly for emerging economies such as Brazil, Russia, India, and

Neural-Schwarz Tiling for Geometry-Universal PDE Solving at Scale

Local AiDGX agent

arXiv:2605.12343v1 Announce Type: new Abstract: Most learned PDE solvers follow a global-surrogate paradigm: a neural operator is trained to map full problem descriptions to full solution fields for a

Neural Statistical Functions

ApplicationsDGX agent

arXiv:2605.11327v1 Announce Type: new Abstract: Classical deep learning typically operates on individual cases. Despite its success, real-world usage often requires repeated inference to estimate stat

Newton's Lantern: A Reinforcement Learning Framework for Finetuning AC Power Flow Warm Start Models

SafetyDGX agent

arXiv:2605.11102v1 Announce Type: new Abstract: Neural warm starts can sharply reduce the number of Newton-Raphson iterations required to solve the AC power flow problem, but existing supervised appro

No More, No Less: Task Alignment in Terminal Agents

Model ReleasesDGX agent

arXiv:2605.12233v1 Announce Type: new Abstract: Terminal agents are increasingly capable of executing complex, long-horizon tasks autonomously from a single user prompt. To do so, they must interpret

NOFE -- Neural Operator Function Embedding

ApplicationsDGX agent

arXiv:2605.11970v1 Announce Type: new Abstract: Most dimensionality reduction methods treat data as discrete point clouds, ignoring the continuous domain structure inherent to many real-world processe

Off-Policy Learning with Limited Supply

SafetyDGX agent

arXiv:2603.18702v3 Announce Type: replace Abstract: We study off-policy learning (OPL) in contextual bandits, which plays a key role in a wide range of real-world applications such as recommendation s

Offline Constrained Reinforcement Learning under Partial Data Coverage

SafetyDGX agent

arXiv:2505.17506v2 Announce Type: replace-cross Abstract: We study offline constrained reinforcement learning with general function approximation in discounted constrained Markov decision processes. P

OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning

SafetyDGX agent

arXiv:2605.12400v1 Announce Type: new Abstract: We study {on-policy self-distillation} (OPSD), where a language model improves its reasoning ability by distilling privileged teacher distributions alon

On the Approximation Complexity of Matrix Product Operator Born Machines

ResearchDGX agent

arXiv:2605.11471v1 Announce Type: new Abstract: Matrix product operator Born machines (MPO-BMs) are tractable tensor-network models for probabilistic modeling, but their efficient approximation capabi

On the Importance of Multistability for Horizon Generalization in Reinforcement Learning

SafetyDGX agent

arXiv:2605.12206v1 Announce Type: new Abstract: In reinforcement learning (RL), agents acting in partially observable Markov decision processes (POMDPs) must rely on memory, typically encoded in a rec

On What We Can Learn from Low-Resolution Data

TutorialsDGX agent

arXiv:2605.12168v1 Announce Type: new Abstract: Artificial intelligence systems typically rely on large, centrally collected datasets, a premise that does not hold in many real-world domains such as h

← Previous
1…164165166167168…243
Next →