AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
22 Apr 2026

Time-Scale Coupling Between States and Parameters in Recurrent Neural Networks

Model ReleasesDGX agent

arXiv:2508.12121v5 Announce Type: replace Abstract: We show that gating mechanisms in recurrent neural networks (RNNs) induce lag-dependent and direction-dependent effective learning rates, even when

Trainability Beyond Linearity in Variational Quantum Objectives

ResearchDGX agent

arXiv:2604.18846v1 Announce Type: cross Abstract: Barren-plateau results have established exponential gradient suppression as a widely cited obstacle to the scalability of variational quantum algorith

TrEEStealer: Stealing Decision Trees via Enclave Side Channels

ResearchDGX agent

arXiv:2604.18716v1 Announce Type: cross Abstract: Today, machine learning is widely applied in sensitive, security-related, and financially lucrative applications. Model extraction attacks undermine c


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Ultrametric OGP - parametric RDT symmetric binary perceptron connection

ResearchDGX agent

arXiv:2604.19712v1 Announce Type: new Abstract: In [97,99,100], an fl-RDT framework is introduced to characterize statistical computational gaps (SCGs). Studying symmetric binary perceptrons (SBPs), [

Unsupervised Confidence Calibration for Reasoning LLMs from a Single Generation

ResearchDGX agent

arXiv:2604.19444v1 Announce Type: new Abstract: Reasoning language models can solve increasingly complex tasks, but struggle to produce the calibrated confidence estimates necessary for reliable deplo

Virtual boundary integral neural network for three-dimensional exterior acoustic problems

ResearchDGX agent

arXiv:2604.18636v1 Announce Type: cross Abstract: This paper presents a virtual boundary integral neural network (VBINN) for exterior acoustic problems in three dimensions. The method introduces a vir

VoteGCL: Enhancing Graph-based Recommendations with Majority-Voting LLM-Rerank Augmentation

SafetyDGX agent

arXiv:2507.21563v4 Announce Type: replace-cross Abstract: Recommendation systems often suffer from data sparsity caused by limited user-item interactions, which degrade their performance and amplify p

When Active Learning Falls Short: An Empirical Study on Chemical Reaction Extraction

ResearchDGX agent

arXiv:2604.19335v1 Announce Type: new Abstract: The rapid growth of chemical literature has generated vast amounts of unstructured data, where reaction information is particularly valuable for applica

When Langevin Monte Carlo Meets Randomization: Non-asymptotic Error Bounds beyond Log-Concavity and Gradient Lipschitzness

ResearchDGX agent

arXiv:2509.25630v2 Announce Type: replace-cross Abstract: Efficient sampling from complex and high dimensional target distributions turns out to be a fundamental task in diverse disciplines such as sc

Whispers in the Machine: Confidentiality in Agentic Systems

AgentsDGX agent

arXiv:2402.06922v5 Announce Type: replace-cross Abstract: Large language model (LLM)-based agents combine LLMs with external tools to automate tasks such as scheduling meetings, managing documents, or

ZC-Swish: Stabilizing Deep BN-Free Networks for Edge and Micro-Batch Applications

Model ReleasesDGX agent

arXiv:2604.19453v1 Announce Type: new Abstract: Batch Normalization (BN) is a cornerstone of deep learning, yet it fundamentally breaks down in micro-batch regimes (e.g., 3D medical imaging) and non-I

21 Apr 2026

A Discordance-Aware Multimodal Framework with Multi-Agent Clinical Reasoning

AgentsDGX agent

arXiv:2604.16333v1 Announce Type: new Abstract: Knee osteoarthritis frequently exhibits discordance between structural damage observed in imaging and patient-reported symptoms such as pain. This misma

A Machine Learning Approach to Two-Stage Adaptive Robust Optimization

ResearchDGX agent

arXiv:2307.12409v3 Announce Type: replace Abstract: We propose an approach based on machine learning to solve two-stage linear adaptive robust optimization (ARO) problems with binary here-and-now vari

A Mechanism Study of Delayed Loss Spikes in Batch-Normalized Linear Models

ResearchDGX agent

arXiv:2604.16809v1 Announce Type: cross Abstract: Delayed loss spikes have been reported in neural-network training, but existing theory mainly explains earlier non-monotone behavior caused by overly

A Model and Estimation of the Bitcoin Transaction Fee

ResearchDGX agent

arXiv:2604.17183v1 Announce Type: cross Abstract: Bitcoin transaction fees will become more important as the block subsidy declines, but fee formation is hard to study with blockchain data alone becau

A Note on TurboQuant and the Earlier DRIVE/EDEN Line of Work

Model ReleasesDGX agent

arXiv:2604.18555v1 Announce Type: new Abstract: This note clarifies the relationship between the recent TurboQuant work and the earlier DRIVE (NeurIPS 2021) and EDEN (ICML 2022) schemes. DRIVE is a 1-

A Probabilistic Consensus-Driven Approach for Robust Counterfactual Explanations

Model ReleasesDGX agent

arXiv:2604.17494v1 Announce Type: new Abstract: Counterfactual explanations (CFEs) are essential for interpreting black-box models, yet they often become invalid when models are slightly changed. Exis

A proposal for PU classification under Non-SCAR using clustering and logistic model

ResearchDGX agent

arXiv:2604.17130v1 Announce Type: cross Abstract: The present study aims to investigate a cluster cleaning algorithm that is both computationally simple and capable of solving the PU classification wh

A Quasi-Experimental Developer Study of Security Training in LLM-Assisted Web Application Development

SafetyDGX agent

arXiv:2604.17763v1 Announce Type: cross Abstract: This paper presents a controlled quasi-experimental developer study examining whether a layer-based security training package is associated with impro

A Ridge Too Far: Correcting Over-Shrinkage via Negative Regularization

ResearchDGX agent

arXiv:2508.17412v4 Announce Type: replace Abstract: Conventional regularization is designed to control variance, but in small-data regression it can also aggravate underfitting when predictive signal

A Scalable Nystrom-Based Kernel Two-Sample Test with Permutations

ResearchDGX agent

arXiv:2502.13570v4 Announce Type: replace-cross Abstract: Two-sample hypothesis testing-determining whether two sets of data are drawn from the same distribution-is a fundamental problem in statistics

A Sensitivity Approach to Causal Inference Under Limited Overlap

SafetyDGX agent

arXiv:2511.22003v2 Announce Type: replace-cross Abstract: Limited overlap between treated and control groups is a key challenge in observational analysis. Standard approaches like trimming importance

A Sugeno Integral View of Binarized Neural Network Inference

ResearchDGX agent

arXiv:2604.17967v1 Announce Type: cross Abstract: In this article, we establish a precise connection between binarized neural networks (BNNs) and Sugeno integrals. The advantage of the Sugeno integral

A Survey of Reinforcement Learning for Large Language Models under Data Scarcity: Challenges and Solutions

TutorialsDGX agent

arXiv:2604.17312v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a powerful post-training paradigm for enhancing the reasoning capabilities of large language models (LLMs). H

A Systematic Survey and Benchmark of Deep Learning for Molecular Property Prediction in the Foundation Model Era

Model ReleasesDGX agent

arXiv:2604.16586v1 Announce Type: new Abstract: Molecular property prediction integrates quantum chemistry, cheminformatics, and deep learning to connect molecular structure with physicochemical and b

A Two-Phase Deep Learning Framework for Adaptive Time-Stepping in High-Speed Flow Modeling

ResearchDGX agent

arXiv:2506.07969v2 Announce Type: replace Abstract: We consider the problem of modeling high-speed flows using machine learning methods. While most prior studies focus on low-speed fluid flows in whic

A Unification of Discrete, Gaussian, and Simplicial Diffusion

ResearchDGX agent

arXiv:2512.15923v2 Announce Type: replace Abstract: To model discrete sequences such as DNA, proteins, and language using diffusion, practitioners must choose between three major methods: diffusion in

A Unified Compliance Aggregator Framework for Automated Multi-Tool Security Assessment of Linux Systems

ResearchDGX agent

arXiv:2604.17256v1 Announce Type: cross Abstract: Assessing the security posture of modern computing systems typically requires the use of multiple specialized tools. These tools focus on different as

A unified convergence theory for adaptive first-order methods in the nonconvex case, including AdaNorm, full and diagonal AdaGrad, Shampoo and Muo

ResearchDGX agent

arXiv:2604.17423v1 Announce Type: new Abstract: A unified framework for first-order optimization algorithms fornonconvex unconstrained optimization is proposed that uses adaptivelypreconditioned gradi

Adjustment of Cluster-Then-Predict Framework for Multiport Scatterer Load Prediction

ApplicationsDGX agent

arXiv:2602.08129v2 Announce Type: replace-cross Abstract: Predicting interdependent load values in multiport scatterers is challenging due to high dimensionality and complex dependence between impedan

Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition

SafetyDGX agent

arXiv:2604.17803v1 Announce Type: cross Abstract: Post-training Large Language Models requires diverse, high-quality data which is rare and costly to obtain, especially in low resource domains and for

Agentic Risk-Aware Set-Based Engineering Design

Model ReleasesDGX agent

arXiv:2604.16687v1 Announce Type: cross Abstract: This paper introduces a multi-agent framework guided by Large Language Models (LLMs) to assist in the early stages of engineering design, a phase ofte

An Integrated Deep-Learning Framework for Peptide-Protein Interaction Prediction and Target-Conditioned Peptide Generation with ConGA-PePPI and TC-PepGen

Model ReleasesDGX agent

arXiv:2604.18467v1 Announce Type: new Abstract: Motivation: Peptide-protein interactions (PepPIs) are central to cellular regulation and peptide therapeutics, but experimental characterization remains

An Interpretable Framework Applying Protein Words to Predict Protein-Small Molecule Complementary Pairing Rules

Model ReleasesDGX agent

arXiv:2604.16550v1 Announce Type: new Abstract: Despite the high accuracy of 'black box' deep learning models, drug discovery still relies on protein-ligand interaction principles and heuristics. To i

An `Inverse' Experimental Framework to Estimate Market Efficiency

SafetyDGX agent

arXiv:2604.18130v1 Announce Type: new Abstract: Digital marketplaces processing billions of dollars annually represent critical infrastructure in sociotechnical ecosystems, yet their performance optim

An LLM-Guided Query-Aware Inference System for GNN Models on Large Knowledge Graphs

ApplicationsDGX agent

arXiv:2603.04545v2 Announce Type: replace Abstract: Efficient inference for graph neural networks (GNNs) on large knowledge graphs (KGs) is essential for many real-world applications. GNN inference qu

Annotation-Assisted Learning of Treatment Policies From Multimodal Electronic Health Records

SafetyDGX agent

arXiv:2507.20993v3 Announce Type: replace Abstract: We study how to learn treatment policies from multimodal electronic health records (EHRs) that consist of tabular data and clinical text. These poli

AntiPaSTO: Self-Supervised Honesty Steering via Anti-Parallel Representations

Model ReleasesDGX agent

arXiv:2601.07473v4 Announce Type: replace Abstract: As models grow more capable, humans cannot reliably verify what they say. Scalable steering requires methods that are internal, self-supervised, and

AQPIM: Breaking the PIM Capacity Wall for LLMs with In-Memory Activation Quantization

HardwareDGX agent

arXiv:2604.18137v1 Announce Type: cross Abstract: Processing-in-Memory (PIM) architectures offer a promising solution to the memory bottlenecks in data-intensive machine learning, yet often overlook t

ARCS: Autoregressive Circuit Synthesis with Topology-Aware Graph Attention and Spec Conditioning

SafetyDGX agent

arXiv:2603.29068v3 Announce Type: replace Abstract: This paper presents ARCS (Autoregressive Circuit Synthesis), a system for amortized analog circuit generation. ARCS produces complete, SPICE-simulat

ARMove: Learning to Predict Human Mobility through Agentic Reasoning

AgentsDGX agent

arXiv:2604.17419v1 Announce Type: cross Abstract: Human mobility prediction is a critical task but remains challenging due to its complexity and variability across populations and regions. Recently, l

ASTRA: An Automated Framework for Strategy Discovery, Retrieval, and Evolution for Jailbreaking LLMs

SafetyDGX agent

arXiv:2511.02356v2 Announce Type: replace-cross Abstract: Despite extensive safety alignment, Large Language Models (LLMs) remain vulnerable to jailbreak attacks. However, existing methods generally l

Asymptotic behavior of eigenvalues of large rank perturbations of large random matrices

ResearchDGX agent

arXiv:2507.12182v4 Announce Type: replace-cross Abstract: The paper is concerned with deformed Wigner random matrices. These matrices are closely related to Deep Neural Networks (DNNs): weight matrice

Auto-encoder model for faster generation of effective one-body gravitational waveform approximations

Model ReleasesDGX agent

arXiv:2511.12642v2 Announce Type: replace-cross Abstract: Upgrades to current gravitational wave detectors for the next observation run and upcoming third-generation observatories, like the Einstein t

Automated Classification of Plasma Regions at Mars Using Machine Learning

ResearchDGX agent

arXiv:2604.17131v1 Announce Type: cross Abstract: The plasma environment around Mars is highly variable because it is strongly influenced by the solar wind. Accurate identification of plasma regions a

Automatic Dataset Construction (ADC): Sample Collection, Data Curation, and Beyond

Model ReleasesDGX agent

arXiv:2408.11338v2 Announce Type: replace-cross Abstract: Large-scale data collection is essential for developing personalized training data, mitigating the shortage of training data, and fine-tuning

AutoOR: Scalably Post-training LLMs to Autoformalize Operations Research Problems

ApplicationsDGX agent

arXiv:2604.16804v1 Announce Type: new Abstract: Optimization problems are central to decision-making in manufacturing, logistics, scheduling, and other industrial settings. Translating complicated des

AutoPPA: Automated Circuit PPA Optimization via Contrastive Code-based Rule Library Learning

ResearchDGX agent

arXiv:2604.18445v1 Announce Type: new Abstract: Performance, power, and area (PPA) optimization is a fundamental task in RTL design, requiring a precise understanding of circuit functionality and the

Back to Repair: A Minimal Denoising Network for Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.17388v1 Announce Type: new Abstract: We introduce JuRe (Just Repair), a minimal denoising network for time series anomaly detection that exposes a central finding: architectural complexity

Balance-Guided Sparse Identification of Multiscale Nonlinear PDEs with Small-coefficient Terms

ResearchDGX agent

arXiv:2604.18414v1 Announce Type: new Abstract: Data-driven discovery of governing equations has advanced significantly in recent years; however, existing methods often struggle in multiscale systems

Balanced Co-Clustering of Users and Items for Embedding Table Compression in Recommender Systems

Model ReleasesDGX agent

arXiv:2604.18351v1 Announce Type: cross Abstract: Recommender systems have advanced markedly over the past decade by transforming each user/item into a dense embedding vector with deep learning models

Barrier-enforced multi-objective optimization for direct point and sharp interval forecasting

ResearchDGX agent

arXiv:2604.18492v1 Announce Type: new Abstract: This paper proposes a multi-step probabilistic forecasting framework using a single neural-network based model to generate simultaneous point and interv

BASIS: Balanced Activation Sketching with Invariant Scalars for 'Ghost Backpropagation'

SafetyDGX agent

arXiv:2604.16324v1 Announce Type: new Abstract: The activation memory required for exact backpropagation scales linearly with network depth, context length, and feature dimensionality, forming an O(L

Batch-Adaptive Causal Annotations

SafetyDGX agent

arXiv:2502.10605v3 Announce Type: replace-cross Abstract: Estimating the causal effects of interventions is crucial to policy and decision-making, yet outcome data are often missing or subject to non-

Bayesian Neural Networks: An Introduction and Survey

ResearchDGX agent

arXiv:2006.12024v3 Announce Type: replace-cross Abstract: Neural Networks (NNs) have provided state-of-the-art results for many challenging machine learning tasks such as detection, regression and cla

Beam-Plasma Collective Oscillations in Intense Charged-Particle Beams: Dielectric Response Theory, Langmuir Wave Dispersion, and Unsupervised Detection via Prometheus

ResearchDGX agent

arXiv:2603.10457v3 Announce Type: replace-cross Abstract: We develop a theoretical and computational framework for beam-plasma collective oscillations in intense charged-particle beams at intermediate

Benchmarking System Dynamics AI Assistants: Cloud Versus Local LLMs on CLD Extraction and Discussion

Model ReleasesDGX agent

arXiv:2604.18566v1 Announce Type: cross Abstract: We present a systematic evaluation of large language model families -- spanning both proprietary cloud APIs and locally-hosted open-source models -- o

Beyond Feature Fusion: Contextual Bayesian PEFT for Multimodal Uncertainty Estimation

Model ReleasesDGX agent

arXiv:2604.16615v1 Announce Type: new Abstract: We introduce CoCo-LoRA, a multimodal, uncertainty-aware parameter-efficient fine-tuning method for text prediction tasks accompanied by audio context. E

Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling

Local AiDGX agent

arXiv:2508.16745v2 Announce Type: replace Abstract: Reasoning is a core capability of large language models, yet how multi-step reasoning is learned and executed remains unclear. We study this questio

Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents

ResearchDGX agent

arXiv:2604.16335v1 Announce Type: new Abstract: Despite recent progress in Large Language Model (LLM) Agents for Software Engineering (SWE) tasks, end-to-end fine-tuning typically relies on verifiable

← Previous
1…213214215216217…241
Next →