AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
8 Jul 2026

Leveraging Extragradient for Effective Sharpness-Aware Minimization in Deep Learning

Model ReleasesDGX agent

arXiv:2607.06151v1 Announce Type: new Abstract: Generalization remains a pivotal challenge in deep learning, where traditional optimizers like Stochastic Gradient Descent (SGD) often converge to sharp

Leveraging Neural Graph Compilers in Machine Learning Research for Edge-Cloud Systems

ResearchDGX agent

arXiv:2504.20198v2 Announce Type: replace-cross Abstract: This work presents a comprehensive evaluation of neural network graph compilers across heterogeneous hardware platforms, addressing the critic

Life Cycle Assessment of Pre-training the Lucie 7B Open-Source Large Language Model on the Jean Zay Supercomputer

HardwareDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.05408v1 Announce Type: cross Abstract: The environmental impact of training large language models (LLMs) is increasingly scrutinised, yet most published estimates focus on operational energ

Low-Overhead Error-Corrected QCNNs Using Bivariate Bicycle Codes

ResearchDGX agent

arXiv:2607.05724v1 Announce Type: new Abstract: Quantum convolutional neural networks (QCNNs) combine the power of quantum computing and classical CNN for computational speedup in classification tasks

MAME: Multidimensional Adaptive Metamer Exploration with Human Perceptual Feedback

SafetyDGX agent

arXiv:2503.13212v3 Announce Type: replace Abstract: Alignment between human brain networks and artificial models has become an active research area in vision science and machine learning. A widely ado

Mitigating Errors in LLM-Generated Web API Invocations via Retrieval-Augmented Generation and Constrained Decoding

ApplicationsDGX agent

arXiv:2607.05936v1 Announce Type: cross Abstract: Integration of web APIs is a cornerstone of modern software systems, yet writing correct web API invocation code remains challenging due to complex an

Modeling Normal Is All You Need: Joint Latent Clustering for Anomaly Detection in Multimodal Cyber-Physical Systems

ResearchDGX agent

arXiv:2607.06094v1 Announce Type: new Abstract: Faults on a cyber-physical system (CPS) are too rare and unrepresentative to characterise, or even to select a model on, so detection must instead model

More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges

Model ReleasesDGX agent

arXiv:2607.05904v1 Announce Type: new Abstract: Training a language model against its own reference-free judgments (the premise of self-rewarding, self-play, and LLM-as-a-judge pipelines) assumes a mo

Multi-Channel Spread-Spectrum Code Watermarking

Model ReleasesDGX agent

arXiv:2607.06009v1 Announce Type: cross Abstract: Attributing code to the large language model that produced it is essential for provenance, licensing, and misuse accountability, yet no deployed water

Multimodal Analytics of Cybersecurity Crisis Preparation Exercises: What Predicts Success?

SafetyDGX agent

arXiv:2603.28553v2 Announce Type: replace-cross Abstract: Instructional alignment, the match between intended cognition and enacted activity, is central to effective instruction but hard to operationa

Multimodal Molecular Representation Learning with Graph Neural Networks, Deep & Cross Networks, and SMILES Embeddings

Model ReleasesDGX agent

arXiv:2607.05736v1 Announce Type: new Abstract: Molecular property prediction often relies on isolated data modalities, where continuous 3D graph neural networks (GNNs) struggle to efficiently capture

No-Regret Gaussian Process Optimization of Time-Varying Functions

ResearchDGX agent

arXiv:2512.00517v3 Announce Type: replace-cross Abstract: Sequential optimization of black-box functions from noisy evaluations has been widely studied, with Gaussian Process bandit algorithms such as

No Subspace to Track: Non-Identifiability and Optimizer State in Low-Rank Training

ResearchDGX agent

arXiv:2607.05872v1 Announce Type: new Abstract: Memory-efficient optimizers such as GaLore train large language models by projecting gradients onto a rank-r subspace recomputed every T steps, assuming

On the Condition Number Upper Bound of the L-BFGS Inverse Hessian Approximation Matrix with a Two-Sided Geometric Envelope Safeguarding Mechanism

ResearchDGX agent

arXiv:2607.05836v1 Announce Type: cross Abstract: The limited-memory BFGS (L-BFGS) algorithm is a cornerstone of large-scale optimization due to its linear memory and computational costs. However, in

On the convergence of graph Laplacians with a symmetric divergence

ResearchDGX agent

arXiv:2607.05892v1 Announce Type: cross Abstract: When analyzing a manifold learning algorithm for data lying on a smooth, compact, connected Riemannian submanifold (M, g) of R^d, a key estimate for t

Orthogonal Dendritic Intrinsic Networks: An Architecture for Significance-Ordered, Orthogonal Latent Spaces

ApplicationsDGX agent

arXiv:2607.05653v1 Announce Type: new Abstract: Principal Component Analysis or PCA-like properties (orthogonality, variance ranking) are seldom realized in deep autoencoder architectures. In this wor

Parameter-Free Encoders Remain Viable for RDB Foundation Models

Model ReleasesDGX agent

arXiv:2607.05476v1 Announce Type: new Abstract: Given a relational database (RDB) storing heterogeneous tabular information, how can we predict missing (or future) values in some target column of inte

Performance Optimization and Comparative Analysis of Generative AI Models on Advanced Accelerators

ResearchDGX agent

arXiv:2607.05400v1 Announce Type: cross Abstract: Generative AI models, such as Large Language Models (LLMs) and diffusion models, have demonstrated impressive performance across a wide range of tasks

Physics-Informed Neural Embeddings of PDE Solution Families

ResearchDGX agent

arXiv:2607.06348v1 Announce Type: new Abstract: We introduce a physics-informed framework for learning finite-dimensional embeddings of solution families of partial differential equations. The method

Quantitative Gaussian-Process limits of Tensor Programs

ResearchDGX agent

arXiv:2607.06290v1 Announce Type: new Abstract: We study the infinite-width Gaussian-process limit of random neural networks through the lens of tensor programs, and we provide a quantitative converge

REAN: Reconstruction-aware ECG Anonymization Based on Privacy--Utility Orthogonality

ResearchDGX agent

arXiv:2607.06037v1 Announce Type: cross Abstract: A shared electrocardiogram (ECG) is itself a biometric fingerprint that can re-identify a patient and reveal personal information. Recent ECG anonymiz

Regularity and Stability Properties of Selective SSMs with Discontinuous Gating

ResearchDGX agent

arXiv:2505.11602v3 Announce Type: replace Abstract: Selective State-Space Models (SSMs) such as Mamba have become central to long-sequence modeling. Still, their stability is poorly understood: their

SafeImpute: Reliable Clinical Data Imputation via Conformal Selection

Model ReleasesDGX agent

arXiv:2607.05613v1 Announce Type: new Abstract: Clinical care often relies on key laboratory indicators, yet real-world patient visits are sparse and tests are ordered irregularly, leading to pervasiv

Scalable Perturbation Learning for Online Self-Supervised Echo State Networks

AgentsDGX agent

arXiv:2607.06079v1 Announce Type: new Abstract: Intelligent systems should not only solve tasks but also adapt under real-world constraints. Autonomous adaptation via self-supervised learning, sequent

Separation Capacity of Scattering Networks on Low-Dimensional Datasets

ResearchDGX agent

arXiv:2607.06048v1 Announce Type: cross Abstract: We aim to identify scattering network architectures that maximize the separation capacity on data with low intrinsic dimension. The networks we consid

SGD-Based Knowledge Distillation with Bayesian Teachers: Theory and Guidelines

ResearchDGX agent

arXiv:2601.01484v2 Announce Type: replace Abstract: Knowledge Distillation (KD) is a central paradigm for transferring knowledge from a large teacher network to a typically smaller student model, ofte

SHARC: SHAP-Based Interpretability in Machine Learning Risk Models for Regulatory Capital under ICAAP and CCAR

Local AiDGX agent

arXiv:2607.05484v1 Announce Type: cross Abstract: The adoption of non-parametric machine learning models for regulatory capital estimation introduces a fundamental governance challenge: the inability

SplineNet: An Isogeometric Deep Learning Method for Complex Shells

ApplicationsDGX agent

arXiv:2607.06026v1 Announce Type: new Abstract: We present a novel isogeometric deep learning method, termed SplineNet, for the seamless design and analysis of shell structures with complex geometries

Stability Annealing Selects the Implicit Bias of Smoothed Sign Descent: A Rate-Indexed Barrier Path on Separable Data

SafetyDGX agent

arXiv:2607.06013v1 Announce Type: new Abstract: Adaptive gradient methods can favor max-margin separators that differ from gradient descent, yet a fixed positive numerical stability constant eventuall

Statistically Meaningful Geometry and Gauge Symmetry Breaking: A Geometric Foundation for Scientific Discovery and Intelligence Emergence

Model ReleasesDGX agent

arXiv:2607.05436v1 Announce Type: new Abstract: The rapid scaling of over-parameterized machine learning architectures, particularly LLMs, raises a profound crisis: do these systems exhibit genuine in

Strategic Bargaining in Multi-Buyer Markets: Reinforcement Learning from Verifiable Rewards for LLM Negotiations

AgentsDGX agent

arXiv:2607.05863v1 Announce Type: new Abstract: Negotiation is a fundamental strategic interaction in management science, characterized by agents attempting to reach agreements while protecting privat

Supervised Reward Inference

Model ReleasesDGX agent

arXiv:2502.18447v2 Announce Type: replace Abstract: Existing approaches to reward inference typically assume that humans provide demonstrations according to specific behavior models. However, humans o

Temporal Variational Implicit Neural Representations

TutorialsDGX agent

arXiv:2506.01544v2 Announce Type: replace Abstract: We introduce Temporal Variational Implicit Neural Representations (TV-INRs), a probabilistic framework for modeling irregular multivariate time seri

TRACE: Trajectory Recovery for Continuous Mechanism Evolution in Causal Representation Learning

ApplicationsDGX agent

arXiv:2601.21135v2 Announce Type: replace Abstract: Temporal causal representation learning methods assume that causal mechanisms switch instantaneously between discrete domains, yet real-world system

Two Sides of the Same Coin: Learning the Backdoor to Remove the Backdoor

TutorialsDGX agent

arXiv:2607.05748v1 Announce Type: new Abstract: The community has recently developed various training-time defenses to counter neural backdoors introduced through data poisoning. In light of the obser

Uncertainty-Aware Velocity Correction for Proprioceptive Vehicle Localization using Evidential Mamba

Local AiDGX agent

arXiv:2607.05669v1 Announce Type: cross Abstract: Reliable localization in GNSS-denied environments remains a fundamental challenge for intelligent vehicles, as inertial navigation systems accumulate

Unequal Uncertainty: Rethinking Algorithmic Interventions for Mitigating Discrimination from AI

ApplicationsDGX agent

arXiv:2508.07872v2 Announce Type: replace-cross Abstract: Uncertainty in artificial intelligence (AI) predictions raises pressing legal and ethical questions for AI-assisted decision-making. This arti

Universality of Benign Overfitting in Binary Linear Classification

ResearchDGX agent

arXiv:2501.10538v3 Announce Type: replace Abstract: The practical success of deep learning has led to the discovery of several surprising phenomena. One of these phenomena, that has spurred intense th

Variational Learning of Disentangled Representations

ResearchDGX agent

arXiv:2506.17182v3 Announce Type: replace Abstract: Disentangled representations separate factors that are shared across conditions from those that are condition-specific. Such separation is needed fo

Width-Robust Learnability in Mean-Field Bayesian Neural Networks

SafetyDGX agent

arXiv:2607.05735v1 Announce Type: cross Abstract: Infinite-width limits are a standard way to reason about neural networks, but it is not automatic that the limiting learner has the same complexity-th

7 Jul 2026

A Bayesian Approach for the Network Reconstruction of Interdependent Critical Infrastructure Systems from Cascading Failures

ApplicationsDGX agent

arXiv:2211.15590v2 Announce Type: replace Abstract: Analyzing the behavior of complex interdependent networks requires complete information about the network topology and the interdependent links acro

A Gradient Flow Perspective on Minimum MMD Estimation

Model ReleasesDGX agent

arXiv:2607.03871v1 Announce Type: new Abstract: Minimum maximum mean discrepancy (MMD) estimation has emerged as a robust and likelihood-free alternative to maximum likelihood estimation for parameter

A Granularity-Aware EEG Feature Framework for Psychopathology Dimension Prediction

ResearchDGX agent

arXiv:2607.02670v1 Announce Type: new Abstract: Electroencephalography (EEG) offers a noninvasive approach for examining neurophysiological correlates of dimensional psychopathology, yet systematic ev

A Hierarchy of Policy Learning Problems

SafetyDGX agent

arXiv:2607.03385v1 Announce Type: cross Abstract: Policy learning has received substantial attention with the goal of learning policies from observational data for decision-making. A majority of work

A Near-Linear-Time Solver for Graph p-Laplacian Semi-Supervised Learning via Continuation in p

Model ReleasesDGX agent

arXiv:2607.03503v1 Announce Type: new Abstract: Graph-based semi-supervised learning (SSL) propagates a few labels over a similarity graph by minimizing a Dirichlet-type energy. The standard quadratic

A Physics-Regulated Neural Framework for Learning 3D Grain Growth Dynamics

ResearchDGX agent

arXiv:2607.04680v1 Announce Type: new Abstract: Grain growth is governed by the reduction in grain boundary energy and exhibits well-established statistical scaling laws. Developing data-driven surrog

A Policy Decomposition Framework for Dynamic Order Fulfillment Operations

SafetyDGX agent

arXiv:2607.04056v1 Announce Type: cross Abstract: Modern supply chains span diverse operational environments, ranging from e-commerce distribution networks to customized production-to-order manufactur

A Provably-Correct and Robust Convex Model for Smooth Separable NMF

ApplicationsDGX agent

arXiv:2511.07109v2 Announce Type: replace-cross Abstract: Nonnegative matrix factorization (NMF) is a linear dimensionality reduction technique for nonnegative data, with applications such as hyperspe

A short tour of operator learning theory: Convergence rates, statistical limits, and open questions

ResearchDGX agent

arXiv:2603.00819v2 Announce Type: replace-cross Abstract: This paper surveys recent developments at the intersection of operator learning, statistical learning theory, and approximation theory. First,

A simplex-based measure of symmetry

ResearchDGX agent

arXiv:2607.03815v1 Announce Type: cross Abstract: For compact convex sets L,K subset R^n, denote by lambda_K(L) the smallest size of a homothet of K that contains L. We define a measure of symmetry ba

A Structural Interpretation of GELU and Threshold-Transmission Activations via the First-Order Loss Function

Model ReleasesDGX agent

arXiv:2607.03664v1 Announce Type: new Abstract: The Gaussian Error Linear Unit is usually motivated as the expected output of an input-dependent stochastic Bernoulli gate. This work gives a complement

A Survey of Reinforcement Learning-Based Motion Planning for Autonomous Driving: Lessons Learned from a Driving Task Perspective

AgentsDGX agent

arXiv:2503.23650v2 Announce Type: replace Abstract: Reinforcement learning (RL), with its ability to explore and optimize policies in complex, dynamic decision-making tasks, has emerged as a promising

A Unified Framework for In-Context Learning with Causal and Masked Language Models

ResearchDGX agent

arXiv:2607.04081v1 Announce Type: new Abstract: In-context learning (ICL) has emerged as a central capability of pretrained language models, yet its theoretical analysis has focused primarily on causa

A Unified Framework for Quantized and Continuous Strong Lottery Tickets

ResearchDGX agent

arXiv:2607.03860v1 Announce Type: new Abstract: The Strong Lottery Ticket Hypothesis (SLTH) asserts that sufficiently overparameterized, randomly initialized neural networks contain sparse subnetworks

ACE: Agentic Control for Embodied Manipulation via Zero-shot Workflow Reasoning

SafetyDGX agent

arXiv:2607.04162v1 Announce Type: cross Abstract: Open-ended tabletop manipulation requires agents to not only understand natural language but also adapt to dynamic environments and execution failures

Active Learning on Adversarially Corrupted Graphs

ApplicationsDGX agent

arXiv:2607.04869v1 Announce Type: new Abstract: Motivated by real-world scenarios where malicious entities tamper with existing networks, we define a model where an adversary seeks to hide a set of co

Adaptive Entropy-Driven Sensor Selection in a Camera-LiDAR Particle Filter for Single-Vessel Tracking

SafetyDGX agent

arXiv:2603.08457v2 Announce Type: replace-cross Abstract: Robust single-vessel tracking from fixed coastal platforms is hindered by modality-specific degradations: cameras suffer from illumination and

Adaptive Loss Balancing for Multi-Task Bioacoustic Classification of Bird Species and Call Types

ResearchDGX agent

arXiv:2607.03304v1 Announce Type: cross Abstract: Reliable analysis of bird vocalisations in passive acoustic monitoring requires models handling multiple, imbalanced annotation targets. We extend Bir

Adaptive Partitioning and Learning for Stochastic Control of Diffusion Processes

SafetyDGX agent

arXiv:2512.14991v2 Announce Type: replace Abstract: We study reinforcement learning for controlled diffusion processes with unbounded continuous state spaces, bounded continuous actions, and polynomia

AdaptiveSD A Stability-Aware, Runtime-Adaptive Speculative Decoding Framework with Multi-Policy Orchestration for CPU-Constrained LLM Inference

Local AiDGX agent

arXiv:2607.03876v1 Announce Type: new Abstract: With the rise of small quantized GGUF-based language models and their increasing use for on-device inference tasks, we have seen the growing need for an

← Previous
1…4748495051…241
Next →