AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
23 Jun 2026

Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis

SafetyDGX agent

arXiv:2505.12462v3 Announce Type: replace Abstract: Robust reinforcement learning (RL) under the average-reward criterion is essential for long-term decision-making, particularly when the environment

Model Inversion meets Cryptographic Fuzzy Extractors

ResearchDGX agent

arXiv:2510.25687v3 Announce Type: replace-cross Abstract: Model inversion attacks pose an open challenge to privacy-sensitive applications that use machine learning (ML) models. For example, face auth

Model Merging in the Essential Subspace

Model ReleasesDGX agent

arXiv:2602.20208v2 Announce Type: replace Abstract: Model merging aims to integrate multiple task-specific fine-tuned models derived from a shared pre-trained checkpoint into a single multi-task model


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning

SafetyDGX agent

arXiv:2606.21943v1 Announce Type: new Abstract: Reinforcement learning (RL) has become central to LLM post-training, yet the methods that dominate current pipelines, PPO and GRPO, represent only a nar

MORL-A2C: Multi-Objective Reinforcement Learning Reranker for Optimizing Healthiness in MOPI-HFRS

Model ReleasesDGX agent

arXiv:2606.23603v1 Announce Type: new Abstract: Unhealthy dietary behavior continues to be a persistent public health issue in the United States, exacerbated by recommendation systems that prioritize

Multi-Year-to-Decadal Temperature Prediction using a Machine Learning Model-Analog Framework

SafetyDGX agent

arXiv:2502.17583v2 Announce Type: replace-cross Abstract: Multi-year-to-decadal climate predictions are a key tool in understanding the range of potential regional climate futures. Here, we present a

Multigrid Training for Molecular Generation using Graph Neural Networks

Model ReleasesDGX agent

arXiv:2606.22377v1 Announce Type: new Abstract: Deep learning has demonstrated significant success for modeling biochemical molecular systems, where inputs are commonly represented as graphs or 3D gri

Muown Implicitly Performs Angular Step-size Decay

Model ReleasesDGX agent

arXiv:2606.23637v1 Announce Type: new Abstract: Matrix-aware optimizers such as Muon and Muown have recently shown strong empirical performance for pre-training Transformers. In particular, Muown sepa

Music Playlist Captioning at Scale with Large Language Models

ResearchDGX agent

arXiv:2606.22460v1 Announce Type: cross Abstract: Music streaming services such as Deezer often recommend personalized playlists to users. Playlist captioning, which involves describing these playlist

NAC: Neural Action Codec for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2606.21372v1 Announce Type: cross Abstract: Vision-language-action (VLA) models rely on discrete action tokenizers to bridge continuous robot control and autoregressive sequence modeling, yet ex

NASDAQ: Normalized Observation Space Dynamics-Augmented Q-Learning

AgentsDGX agent

arXiv:2606.21297v1 Announce Type: new Abstract: Augmenting model-free reinforcement learning (RL) with representations learned through observation dynamics prediction (observation-predictive RL) can i

Negative Knowledge as Failure-aware Shared Memory for AutoResearch

AgentsDGX agent

arXiv:2606.21024v1 Announce Type: cross Abstract: AI-assisted research systems generate many failed attempts, but those failures rarely become a durable, shared knowledge asset. We propose a negative

Neural Architecture Search of Sample Reweighting Networks for Complex Distribution Shift

ResearchDGX agent

arXiv:2606.22991v1 Announce Type: new Abstract: Sample reweighting is a major approach to addressing distribution shifts, such as label noise and class imbalance. Meta-Weight-Net (MW-Net) is a promisi

Neural Concept Verifier: Scaling Prover-Verifier Games via Concept Encodings

ApplicationsDGX agent

arXiv:2507.07532v4 Announce Type: replace Abstract: While Prover-Verifier Games (PVGs) offer a promising path toward verifiability in nonlinear classification models, they have not yet been applied to

Neural Conjugate Aggregation: Identifiable Unsupervised Multi-Sensor Regression under Heterogeneous Sensor Bias

SafetyDGX agent

arXiv:2606.22200v1 Announce Type: new Abstract: We study regression-based data fusion under uncertainty, where multiple noisy and biased measurement sources are available but ground-truth labels are a

Neural Networks as Linear Regression: An Introduction for Statisticians

ResearchDGX agent

arXiv:2606.23601v1 Announce Type: cross Abstract: Neural networks are a commonly used prediction tool in computer science and statistics. However, the barrier to entry of this interesting field remain

Neural Operator Processes for Probabilistic Operator Learning under Partial Observations

TutorialsDGX agent

arXiv:2606.22946v1 Announce Type: new Abstract: Neural operators learn mappings between function spaces, but are typically developed with dense input-output training fields and fully observed inputs a

Neural Parameter Calibration for Finite-State Mean Field Games

Model ReleasesDGX agent

arXiv:2606.23155v1 Announce Type: cross Abstract: Mean field games efficiently approximate a very large population of strategic agents. While these games can aid the understanding of complex systems,

New Smooth Loss functions for Robust Regression that Closely Approximate Absolute Error and Provide Improved Performance on Datasets With Significant Outliers

ResearchDGX agent

arXiv:2606.22068v1 Announce Type: new Abstract: The performance of supervised machine learning models is directly related to the quality of the training dataset. In particular, the presence of signifi

NewPINNs: Physics-Informing Neural Networks Using Conventional Solvers for Partial Differential Equations

TutorialsDGX agent

arXiv:2601.17207v2 Announce Type: replace Abstract: We introduce NewPINNs, a physics-informing learning framework that couples neural networks with conventional numerical solvers for solving different

Next-Gen CAPTCHAs: Leveraging the Cognitive Gap for Scalable and Diverse GUI-Agent Defense

Model ReleasesDGX agent

arXiv:2602.09012v2 Announce Type: replace Abstract: The rapid evolution of GUI-enabled agents has rendered traditional CAPTCHAs obsolete. While previous benchmarks like OpenCaptchaWorld established a

NNiT: Width-Agnostic Neural Network Generation with Structurally Aligned Weight Spaces

Model ReleasesDGX agent

arXiv:2603.00180v2 Announce Type: replace Abstract: Generative modeling of neural network parameters is often tied to architectures because standard parameter representations rely on known weight-matr

No Reference-Free Generalization in Quantum Machine Learning

ResearchDGX agent

arXiv:2606.22331v1 Announce Type: cross Abstract: Quantum machine learning is often motivated by the exponentially large state space of quantum systems, but this promise leaves a basic generalization

Noise is Signal: Density-Based Outliers as Leading Indicators of Occupational Emergence in Labor Market Text

SafetyDGX agent

arXiv:2606.22769v1 Announce Type: new Abstract: Standard NLP pipelines for occupational clustering discard the 10-15% of job postings that density-based methods assign to noise. We argue this is an er

Non-Adversarial Imitation Learning Provably Free of Compounding Errors: The Value Flow Mechanism

TutorialsDGX agent

arXiv:2603.22713v2 Announce Type: replace Abstract: Adversarial imitation learning (AIL) achieves high-quality imitation by mitigating compounding errors inherent to behavioral cloning (BC), yet its a

Non-asymptotic estimates of the minimal risk in statistical learning

Model ReleasesDGX agent

arXiv:2606.23295v1 Announce Type: new Abstract: In this paper we prove some concentration inequalities for two types of error probabilities in the Empirical Risk Principle (ERP) in statistical learnin

Non-Euclidean SGD for Structured Optimization: Unified Analysis and Improved Rates

ResearchDGX agent

arXiv:2511.11466v2 Announce Type: replace-cross Abstract: Recently, several instances of non-Euclidean SGD, including SignSGD, Lion, and Muon, have attracted significant interest from the optimization

Nonconvex-Nonconcave Min-Max Optimization with a Small Maximization Domain

ResearchDGX agent

arXiv:2110.03950v2 Announce Type: replace-cross Abstract: We study the problem of finding approximate first-order stationary points in optimization problems of the form min_{x in X} max_{y in Y} f(x,y

Nous: A Predictive World Model for Long-Term Agent Memory

Model ReleasesDGX agent

arXiv:2606.22030v1 Announce Type: cross Abstract: We present Nous, a novel agent memory architecture grounded in the principle that knowledge is prediction, not storage. Rather than persisting facts a

Null-Calibrated Conformal Selection via Target-Membership Scores

ResearchDGX agent

arXiv:2606.22336v1 Announce Type: cross Abstract: Conformal selection aims to identify test candidates whose unknown responses fall in a target region while controlling the false discovery rate. Exist

Numerical stability analysis of large language models

Local AiDGX agent

arXiv:2503.10251v2 Announce Type: replace-cross Abstract: Transformers are the state-of-the-art architecture for large language models, and a key to their scalability is the strategic usage of low-pre

O-RAN Xapps Conflict Prediction Using Graph Convolutional Networks

ApplicationsDGX agent

arXiv:2503.03523v3 Announce Type: replace-cross Abstract: O-RAN hosts many intelligent applications known as eXtended Applications (xApps). xApps are applications that leverage advanced algorithms to

Objective-Behavior Alignment: Diagnostics for MORL Policy Selection

SafetyDGX agent

arXiv:2606.21321v1 Announce Type: new Abstract: Real-world decision-making often requires optimizing multiple competing objectives simultaneously. In reinforcement learning (RL), this is typically add

OFMU: Optimization-Driven Framework for Machine Unlearning

SafetyDGX agent

arXiv:2509.22483v2 Announce Type: replace Abstract: Large language models deployed in sensitive applications increasingly require the ability to unlearn specific knowledge, such as user requests, copy

OGD4All: A Framework for Accessible Interaction with Geospatial Open Government Data Based on Large Language Models

Model ReleasesDGX agent

arXiv:2602.00012v3 Announce Type: replace Abstract: We present OGD4All, a transparent, auditable, and reproducible framework based on Large Language Models (LLMs) to enhance citizens' interaction with

Omega: Operator-based Mixture Ensemble for Generative Assimilation

ResearchDGX agent

arXiv:2606.20920v1 Announce Type: new Abstract: Characterizing non-Gaussian posterior distributions in partially observed high-dimensional nonlinear systems remains a fundamental challenge in data ass

On Quantum Perceptron Learning via Quantum Search

ResearchDGX agent

arXiv:2503.17308v2 Announce Type: replace-cross Abstract: With the growing interest in quantum machine learning, the perceptron, a fundamental building block in traditional machine learning, has emerg

On the Curse of Dimensionality in Private Sparse Covariance Estimation and PCA

ResearchDGX agent

arXiv:2606.21951v1 Announce Type: new Abstract: We study high-dimensional differentially private (DP) covariance estimation in the operator norm, and principal component analysis (PCA), under k-row-co

On the Expressive Power of Weight Quantization in Large Language Models

ResearchDGX agent

arXiv:2606.22249v1 Announce Type: new Abstract: In recent years, weight quantization that encodes the learnable parameters of large language models in an n-bit format has garnered significant attentio

On the Limits of Prompt-Conditioned Language Models as General-Purpose Learners

SafetyDGX agent

arXiv:2606.23668v1 Announce Type: new Abstract: Large Language Models (LLMs) are frequently portrayed as general-purpose solvers capable of solving arbitrary tasks. We argue that this view overlooks a

On the Memorization Behavior of LLMs in Generative Recommendation: Observations, Implications, and Training Strategies

TutorialsDGX agent

arXiv:2606.17276v3 Announce Type: replace-cross Abstract: Generative recommendation (GR) has emerged as a promising direction for recommender systems. Recently, large language models (LLMs) have been

On the Position Bias of On-Policy Distillation

SafetyDGX agent

arXiv:2606.22600v1 Announce Type: new Abstract: On-Policy Distillation (OPD) improves the learning efficiency of standard reinforcement learning through dense, token-level supervision from teachers. I

On the Sparsity-Storage-Accuracy Tradeoff in Parsimoniously Activated Dictionary Learning

ResearchDGX agent

arXiv:2606.22352v1 Announce Type: new Abstract: Dictionary learning has long been studied from both optimization and probabilistic perspectives. While formulations with element-wise sparsity regulariz

One Size does not Fit All: Heterogeneous Latent Space Alignment for Unsupervised Domain Adaptation

SafetyDGX agent

arXiv:2606.21415v1 Announce Type: new Abstract: Domain shift remains a major obstacle to the reliable deployment of machine learning models in high-stakes environments such as healthcare. While Domain

One-Step Flow Matching for Generative Modeling of Path-Dependent Physical Fields

ResearchDGX agent

arXiv:2606.22752v1 Announce Type: new Abstract: Physical simulations for intricate geometries with path-dependent constitutive models face difficulties due to the enormous computational cost they requ

Open Problem: Is AdamW Effective Under Heavy-Tailed Noise?

Model ReleasesDGX agent

arXiv:2606.23676v1 Announce Type: new Abstract: AdamW is the de facto optimizer for training large language models (LLMs), yet the theory behind it still lives mostly in finite-variance regimes. This

Orthogonal Discrepancy Kernels for Learning with Partial Physics

Model ReleasesDGX agent

arXiv:2606.21199v1 Announce Type: cross Abstract: We introduce a semi-parametric framework for nonlinear system identification, which decouples discrepancy functions from physics-based components. Ort

Over-the-Air Federated Learning: Rethinking Edge AI Through Signal Processing

Local AiDGX agent

arXiv:2512.03719v2 Announce Type: replace-cross Abstract: Over-the-Air Federated Learning (AirFL) is an emerging paradigm that tightly integrates wireless signal processing and distributed machine lea

OVIG: Optimistic Verification of AI Training Integrity via Gradient Signals

Model ReleasesDGX agent

arXiv:2606.21045v1 Announce Type: cross Abstract: The rapid growth of AI has increased the demand for domain-specific post-training, while the cost and specialization of accelerator infrastructure pus

PACT: Preserving Anchored Cores in Task-vectors for Model Merging

ResearchDGX agent

arXiv:2606.18627v2 Announce Type: replace Abstract: Model merging has emerged as a training-free alternative to multi-task learning, aiming to combine multiple task-specific fine-tuned models into a s

Parameterized Representations via Implicit Stochastic Modulation for High-Dimensional and High-Order Neural PDE Solvers

Model ReleasesDGX agent

arXiv:2606.22150v1 Announce Type: new Abstract: Solving high-dimensional and high-order PDEs is challenged by the coupled growth of spatial dimensionality and derivative order. Recent stochastic deriv

Parametrized Power-Iteration Clustering for Directed Graphs

ApplicationsDGX agent

arXiv:2210.00310v3 Announce Type: replace Abstract: Vertex-level clustering for directed graphs (digraphs) remains challenging as edge directionality breaks the key assumptions underlying popular spec

Patched Flow Matching: Generative Wall-Pressure Reconstruction Beyond Training-Domain Scales from Sparse Sensors

Local AiDGX agent

arXiv:2606.22084v1 Announce Type: cross Abstract: Characterizing the complete wall-pressure spectrum in turbulent wall-bounded flows requires simultaneous access to the viscous-scale high-wavenumber c

Path-dependent program induction under resource constraints explains human sequence learning

ResearchDGX agent

arXiv:2606.20623v1 Announce Type: cross Abstract: How do people build abstract, reusable knowledge from sequential experience under bounded cognitive resources? To answer this question, we integrate r

Patient-Aware Contrastive Learning Preserves Per-Patient Structure in RR-Interval Representations

ResearchDGX agent

arXiv:2606.23570v1 Announce Type: new Abstract: Contrastive representation learning struggles on physiological signals when each subject contributes a distinct baseline pattern. If class differences o

PEAR: Permutation-Equivariant Adaptive Routing Multi-Agent Debate

AgentsDGX agent

arXiv:2606.20621v1 Announce Type: cross Abstract: Multi-agent debate improves the reliability of large language models (LLMs) through iterative peer critiques. However, fixed topologies often introduc

PeLAP-A: Adaptive Latent Pruning for Lightweight Latent Diffusion Models

ResearchDGX agent

arXiv:2606.23086v1 Announce Type: new Abstract: Latent diffusion models achieve strong generative performance by operating in a compressed latent space produced by a variational autoencoder (VAE). How

Physics-Guided Dual-Stream Heterogeneous Graph Neural Network for Predicting Full-Field Structural Response of Stiffened Panels

Model ReleasesDGX agent

arXiv:2606.20916v1 Announce Type: new Abstract: Iterative design and optimization of large, complex structures require fast and accurate prediction of stress, displacement, and other fields. Finite el

Physics-Guided Fully Convolutional Spatiotemporal Learning Toward Digital-Twin-Enabled Microstructure Evolution Prediction

ResearchDGX agent

arXiv:2606.20983v1 Announce Type: new Abstract: Understanding and predicting microstructure evolution is central to materials design, yet purely data-driven spatiotemporal learning models often suffer

Physics-Informed Eikonal Caging for Whole-Arm Manipulation Planning

ApplicationsDGX agent

arXiv:2606.22143v1 Announce Type: cross Abstract: Planning contact-rich whole-arm manipulation is challenging because interactions that involve extended robot geometry give rise to complex contact dyn

← Previous
1…8081828384…243
Next →