AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
11 May 2026

TraXion: Rethinking Pre-training Frameworks for Mobility and Beyond

ApplicationsDGX agent

arXiv:2605.06906v1 Announce Type: new Abstract: Human mobility differs from text and from generic time series in three structural ways: visits are tuple-valued events whose meaning depends on the join

Tree SAE: Learning Hierarchical Feature Structures in Sparse Autoencoders

TutorialsDGX agent

arXiv:2605.07922v1 Announce Type: new Abstract: Learning hierarchical features in Sparse Autoencoders (SAEs) is essential for capturing the structured nature of real-world data and mitigating issues l

TUANDROMD-X: Advanced Entropy and Visual Analytics Dataset for Enhanced Malware Detection and Classification

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.06718v1 Announce Type: cross Abstract: Malware and malware-based attacks are becoming more prevalent and complex. Attackers regularly come up with new techniques that have the ability to ev

Tyche: One Step Flow for Efficient Probabilistic Weather Forecasting

ResearchDGX agent

arXiv:2605.06916v1 Announce Type: new Abstract: Probabilistic weather forecasting requires not only accurate trajectories, but calibrated distributions over plausible atmospheric futures. Recent data-

Uncovering Hidden Systematics in Neural Network Models for High Energy Physics

TutorialsDGX agent

arXiv:2605.07470v1 Announce Type: new Abstract: Neural networks (NNs) are inherently multidimensional classifiers that learn complex, non-linear relationships among input observables. While their flex

Understanding Robustness of Model Editing in Code LLMs

Model ReleasesDGX agent

arXiv:2511.03182v2 Announce Type: replace-cross Abstract: Large language models (LLMs) for code are increasingly used in software development, but they remain static after pretraining while APIs and s

Upper Generalization Bounds for Neural Oscillators

ResearchDGX agent

arXiv:2603.09742v2 Announce Type: replace Abstract: Neural oscillators that originate from second-order ordinary differential equations (ODEs) have shown competitive performance in learning mappings b

Versatile yet Efficient Network Traffic Analysis: Offloading Network Foundation Model to SmartNIC

HardwareDGX agent

arXiv:2508.02001v2 Announce Type: replace-cross Abstract: Pervasive encryption makes large-scale labeling infeasible for traffic analysis, while security operations demand edge analysis to avert servi

VNN-LIB 2.0: Rigorous Foundations for Neural Network Verification

ResearchDGX agent

arXiv:2605.07451v1 Announce Type: new Abstract: Neural network verification is an active and rapidly maturing research area, with a growing ecosystem of solvers and tools. The VNN-LIB standard was int

When Descent Is Too Stable: Event-Triggered Hamiltonian Learning to Optimize

SafetyDGX agent

arXiv:2605.06868v1 Announce Type: new Abstract: Fixed-budget nonconvex optimization can fail not because local descent is unstable, but because it is too stable: after reaching a nearby stationary poi

When Diffusion Model Can Ignore Dimension: An Entropy-Based Theory

ResearchDGX agent

arXiv:2605.07969v1 Announce Type: new Abstract: Diffusion models perform remarkably well on high-dimensional data such as images, often using only a modest number of reverse-time steps. Despite this p

When Does Embedding Magnitude Matter? A Cross-Task Functional-Symmetry Framework

ResearchDGX agent

arXiv:2602.09229v3 Announce Type: replace Abstract: Cosine similarity normalizes both sides; dot product normalizes neither. We propose a 2x2 framework that independently controls query-side and docum

When Symbol Names Should Not Matter: A Logistic Theory of Fresh-Symbol Classification

TutorialsDGX agent

arXiv:2605.07120v1 Announce Type: new Abstract: Template tasks have emerged as a clean testbed for asking whether transformers reason with abstract symbols rather than concrete token names. We study t

Where to Spend Rollouts: Hit-Utility Optimal Rollout Allocation for Group-Based RLVR

Model ReleasesDGX agent

arXiv:2605.07114v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a central paradigm for improving the reasoning capabilities of large language model

Why Does Agentic Safety Fail to Generalize Across Tasks?

SafetyDGX agent

arXiv:2605.06992v1 Announce Type: new Abstract: AI agents are increasingly deployed in multi-task settings, where the task to perform is specified at test time, and the agent must generalize to unseen

XDecomposer: Learning Prior-Free Set Decomposition for Multiphase X-ray Diffraction

ApplicationsDGX agent

arXiv:2605.05866v1 Announce Type: cross Abstract: Multiphase powder X-ray diffraction (PXRD) analysis remains a fundamental bottleneck in structure identification, as real-world synthesis often produc

You Only Stack Once (YOSO): A Motion-Filtered, Deep-Learning Framework for Detecting Faint Moving Sources

ResearchDGX agent

arXiv:2605.06913v1 Announce Type: cross Abstract: We present You Only Stack Once (YOSO), an automated pipeline designed to detect faint, slow-moving Solar System objects in wide-field astronomical sur

Zero-Shot Imagined Speech Decoding via Imagined-to-Listened MEG Mapping

SafetyDGX agent

arXiv:2605.08075v1 Announce Type: new Abstract: Decoding imagined speech from non-invasive brain recordings is challenging because imagined datasets are scarce and difficult to align temporally across

Zero-Shot Neural Network Evaluation with Sample-Wise Activation Patterns

HardwareDGX agent

arXiv:2605.07378v1 Announce Type: new Abstract: Zero-shot proxies, also known as training-free metrics, are widely adopted to reduce the computational overhead in neural network evaluation for scenari

7 May 2026

A Biased Nonnegative Block Term Tensor Decomposition Model for Dynamic QoS Prediction

Model ReleasesDGX agent

arXiv:2605.04813v1 Announce Type: new Abstract: With the rapid development of cloud computing and Web services, Quality of Service (QoS) has become a key criterion for service selection and recommenda

A Consistency-Centric Approach to Set-Based Optimization with Multiple Models of Unranked Fidelity

ApplicationsDGX agent

arXiv:2605.04051v1 Announce Type: cross Abstract: In complex real-world settings, optimization is challenged by the presence of diverse models of differing fidelity. In many optimization problems, a s

A Foundation Model for Zero-Shot Logical Rule Induction

ApplicationsDGX agent

arXiv:2605.04916v1 Announce Type: cross Abstract: Inductive Logic Programming (ILP) learns interpretable logical rules from data. Existing methods are transductive: their learned parameters are bound

A foundation model of vision, audition, and language for in-silico neuroscience

ResearchDGX agent

arXiv:2605.04326v1 Announce Type: cross Abstract: Cognitive neuroscience is fragmented into specialized models, each tailored to specific experimental paradigms, hence preventing a unified model of co

A geometric relation of the error introduced by sampling a language model's output distribution to its internal state

ResearchDGX agent

arXiv:2605.04899v1 Announce Type: new Abstract: GPT-style language models are sensitive to single-token changes at generation points where the predicted probability distribution is spread across multi

A Harmonic Mean Formulation of Average Reward Reinforcement Learning in SMDPs

ResearchDGX agent

arXiv:2605.04880v1 Announce Type: new Abstract: Recent research has revived and amplified interest in algorithms for undiscounted average reward reinforcement learning in infinite-horizon, non-episodi

A Hybrid Quantum-Classical Framework for Financial Volatility Forecasting Based on Quantum Circuit Born Machines

TutorialsDGX agent

arXiv:2603.09789v2 Announce Type: replace Abstract: Accurate financial volatility forecasting is crucial but challenged by the non-linear, highly correlated nature of market data. Recently, quantum co

A Mean Curvature Approach to Boundary Detection: Geometric Insights for Unsupervised Learning

ApplicationsDGX agent

arXiv:2605.04274v1 Announce Type: new Abstract: Accurate boundary detection in high-dimensional data remains a central challenge in unsupervised learning, particularly in the presence of non-linear st

A Physics-Aware Framework for Short-Term GPU Power Forecasting of AI Data Centers

HardwareDGX agent

arXiv:2605.04074v1 Announce Type: new Abstract: AI data centers experience rapid fluctuations in power demand due to the heterogeneity of computational tasks that they have to support. For example, th

A Provably Convergent and Practical Algorithm for Gromov--Wasserstein Optimal Transport

ResearchDGX agent

arXiv:2605.04175v1 Announce Type: new Abstract: Gromov--Wasserstein optimal transport (GWOT) aligns metric measure spaces by matching their within-domain relational structures, but large-scale GWOT re

A Queueing-Theoretic Framework for Stability Analysis of LLM Inference with KV Cache Memory Constraints

HardwareDGX agent

arXiv:2605.04595v1 Announce Type: new Abstract: The rapid adoption of large language models (LLMs) has created significant challenges for efficient inference at scale. Unlike traditional workloads, LL

A Regulatory Governance Framework for AI-Driven Financial Fraud Detection in U.S. Banking: Integrating OCC, SR 11-7, CFPB, and FinCEN Compliance Requirements for Model Development, Validation, and Monitoring Lifecycles

Model ReleasesDGX agent

arXiv:2605.04076v1 Announce Type: new Abstract: U.S. financial institutions deploying AI-based fraud detection face a fragmented compliance landscape spanning four regulatory frameworks -- OCC Bulleti

A Scalable Multi-Task Model for Virtual Sensors

Model ReleasesDGX agent

arXiv:2601.20634v2 Announce Type: replace Abstract: Virtual sensors replace expensive physical sensors in critical applications through machine learning by predicting target signals from available mea

A Self-Attentive Meta-Optimizer with Group-Adaptive Learning Rates and Weight Decay

Model ReleasesDGX agent

arXiv:2605.04055v1 Announce Type: new Abstract: Adaptive optimizers like AdamW apply uniform hyperparameters across all parameter groups, ignoring heterogeneous optimization dynamics across layers and

A semicontinuous relaxation of Saito's criterion and freeness as angular minimization

ResearchDGX agent

arXiv:2604.02995v2 Announce Type: replace-cross Abstract: We introduce a nonnegative functional mathfrak{S} on the space of line arrangements in P^2 that vanishes precisely on free arrangements, obtai

Actionable Real-Time Modeling of Surgical Team Dynamics via Time-Expanded Interaction Graphs

ResearchDGX agent

arXiv:2605.04169v1 Announce Type: cross Abstract: Surgical team performance arises from complex interactions between technical execution and non-technical skills, including communication and coordinat

Adapt or Forget: Provable Tradeoffs Between Adam and SGD in Nonstationary Optimization

ResearchDGX agent

arXiv:2605.04269v1 Announce Type: cross Abstract: We provide a theoretical analysis of Adam under non-stationary stochastic objectives, separating two regimes: Euclidean tracking under adaptive strong

Adaptive Consensus in LLM Ensembles via Sequential Evidence Accumulation: Automatic Budget Identification and Calibrated Commit Signals

ResearchDGX agent

arXiv:2605.04236v1 Announce Type: new Abstract: Large Language Model ensembles improve reasoning accuracy up to a performance boundary; beyond it, additional deliberation degrades accuracy. Static-bud

Adaptive Ensemble Aggregation for Actor-Critics

Model ReleasesDGX agent

arXiv:2507.23501v2 Announce Type: replace Abstract: Ensembles are ubiquitous in off-policy actor-critic learning, yet their efficacy depends critically on how they are aggregated. Current methods typi

Adaptive Inverted-Index Routing for Granular Mixtures-of-Experts

ResearchDGX agent

arXiv:2605.04952v1 Announce Type: new Abstract: Mixture-of-experts (MoE) models enable scalable transformer architectures by activating only a subset of experts per token. Recent evidence suggests tha

Adaptive Learning Strategies for AoA-Based Outdoor Localization: A Comprehensive Framework

Local AiDGX agent

arXiv:2605.05055v1 Announce Type: new Abstract: Localization in 5G and 6G networks is essential for important use cases such as intelligent transportation, smart factories, and smart cities. Although

Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2605.05123v1 Announce Type: new Abstract: In offline-to-online reinforcement learning (O2O-RL), policies are first safely trained offline using previously collected datasets and then further fin

Adaptivity Under Realizability Constraints: Comparing In-Context and Agentic Learning

AgentsDGX agent

arXiv:2605.04995v1 Announce Type: new Abstract: We compare in-context learning with fixed queries and agentic learning with adaptive queries for uniform approximation of task families. We consider two

Advancing Analytic Class-Incremental Learning through Vision-Language Calibration

SafetyDGX agent

arXiv:2602.13670v2 Announce Type: replace Abstract: Class-incremental learning (CIL) with pre-trained models (PTMs) faces a critical trade-off between efficient adaptation and long-term stability. Whi

Agentic Vulnerability Reasoning on Windows COM Binaries

Model ReleasesDGX agent

arXiv:2605.05000v1 Announce Type: cross Abstract: Windows Component Object Model (COM) services run with elevated privileges and are widely accessible to authenticated users, making race conditions in

Aggressive or Imperceptible, or Both: Network Pruning Assisted Hybrid Byzantines in Federated Learning

Model ReleasesDGX agent

arXiv:2404.06230v3 Announce Type: replace Abstract: In federated learning (FL), profiling and verifying each client is inherently difficult, which introduces a significant security vulnerability: mali

Analogy between Boltzmann machines and Feynman path integrals

ResearchDGX agent

arXiv:2301.06217v1 Announce Type: cross Abstract: We provide a detailed exposition of the connections between Boltzmann machines commonly utilized in machine learning problems and the ideas already we

ANDRE: An Attention-based Neuro-symbolic Differentiable Rule Extractor

TutorialsDGX agent

arXiv:2605.04193v1 Announce Type: cross Abstract: Inductive Logic Programming (ILP) aims to learn interpretable first-order rules from data, but existing symbolic and neuro-symbolic approaches struggl

ARTA: Adversarial-Robust Multivariate Time--Series Anomaly Detection via Sparsity-Constrained Perturbations

Model ReleasesDGX agent

arXiv:2603.25956v2 Announce Type: replace Abstract: Time-series anomaly detection (TSAD) is a critical component in monitoring complex systems, yet modern deep learning-based detectors are often highl

AsymmetryZero: A Framework for Operationalizing Human Expert Preferences as Semantic Evals

Model ReleasesDGX agent

arXiv:2605.04083v1 Announce Type: new Abstract: Much of the focus in RL today is on evaluation design: building meaningful evals that serve simultaneously as benchmarks and as well-defined reward sign

Attention Sinks Induce Gradient Sinks: Massive Activations as Gradient Regulators in Transformers

ResearchDGX agent

arXiv:2603.17771v2 Announce Type: replace Abstract: Attention sinks and massive activations are recurring and closely related phenomena in Transformer models. Existing explanations have largely focuse

Automated Formal Proofs of Combinatorial Identities via Wilf-Zeilberger Guidance and LLMs

Model ReleasesDGX agent

arXiv:2605.04472v1 Announce Type: new Abstract: Automating formal proofs of combinatorial identities is challenging for LLM-based provers, as long-horizon proof planning is required and unconstrained

Average Attention Transformers and Arithmetic Circuits

ResearchDGX agent

arXiv:2605.04683v1 Announce Type: cross Abstract: We analyse the computational power of transformer encoders as sequence-to-sequence functions on vectors. We show that average hard attention can be us

AxMoE: Characterizing the Impact of Approximate Multipliers on Mixture-of-Experts DNN Architectures

ResearchDGX agent

arXiv:2605.04754v1 Announce Type: new Abstract: Deep neural network (DNN) inference at the edge demands simultaneous improvements in accuracy, computational efficiency, and energy consumption. Approxi

Back to Blackwell: Closing the Loop on Intransitivity in Multi-Objective Preference Fine-Tuning

Model ReleasesDGX agent

arXiv:2602.19041v2 Announce Type: replace Abstract: A recurring challenge in preference fine-tuning (PFT) is handling extit{intransitive} (i.e., cyclic) preferences. Intransitive preferences often ste

Bayesian Parameter Shift Rule in Variational Quantum Eigensolvers

Model ReleasesDGX agent

arXiv:2502.02625v2 Announce Type: replace Abstract: Parameter shift rules (PSRs) are key techniques for efficient gradient estimation in variational quantum eigensolvers (VQEs). In this paper, we prop

Benchmarking LLMs on the Massive Sound Embedding Benchmark (MSEB)

Model ReleasesDGX agent

arXiv:2605.04556v1 Announce Type: cross Abstract: The Massive Sound Embedding Benchmark (MSEB) has emerged as a standard for evaluating the functional breadth of audio models. While initial baselines

Beyond Rigid Geometries: The Spline-Pullback Metric for Universal Diffeomorphic SPD Representation Learning

ResearchDGX agent

arXiv:2605.04406v1 Announce Type: new Abstract: The integration of Symmetric Positive Definite (SPD) matrices into deep learning has historically relied on fixed algebraic Riemannian metrics. Analogou

Beyond Variance: Prompt-Efficient RLVR via Rare-Event Amplification and Bidirectional Pairing

ResearchDGX agent

arXiv:2602.03452v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) is effective for training large language models on deterministic outcome reasoning tasks. Prio

Bilinear Mamba-Koopman Neural MPC for Varying Dynamics

ResearchDGX agent

arXiv:2605.04793v1 Announce Type: new Abstract: Koopman-based neural MPC models generate time-varying dynamics from historical data, but preserve convexity by enforcing that the system operator is ind

BOOOM: Loss-Function-Agnostic Black-Box Optimization over Orthonormal Manifolds for Machine Learning and Statistical Inference

ResearchDGX agent

arXiv:2605.04087v1 Announce Type: cross Abstract: Optimization over the Stiefel manifold St(p,d), the set of p imes d column-orthonormal matrices, is fundamental in statistics, machine learning, and s

← Previous
1…178179180181182…241
Next →