AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
Model Releases

Inverse Problem for Partial Differential Equations with Jump Discontinuities in Coefficients by Two-stage Physics-Informed Deep Learning and Statistical Mixture Models

DGX agent

arXiv:2510.14656v2 Announce Type: replace-cross Abstract: This work proposes a two-stage physics-informed deep learning framework that combines neural-network-based sampling with statistical inference

model-releasesarxiv-cs-lg
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Inverting the Bellman Equation: From Q-Values to World Models

DGX agent

arXiv:2606.21173v1 Announce Type: new Abstract: Model-based and model-free reinforcement learning are traditionally viewed as separate paradigms: instead of learning a model of the transition kernel P

safetyarxiv-cs-lg
23 Jun 2026
Safety

IRumAI: Reinforcement Learning for Indian Rummy

DGX agent

arXiv:2606.21975v1 Announce Type: cross Abstract: Despite its massive player base and complex hidden-information dynamics, Indian Rummy has received no reinforcement learning attention. Existing agent

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Is Our Benchmark Enough? An Analysis of Continual Learning for MLLMs

DGX agent

arXiv:2606.20961v1 Announce Type: new Abstract: Continual adaptation is essential for multimodal large language models (MLLMs) deployed across evolving domains, but the state-of-the-art MR-LoRA method

model-releasesarxiv-cs-lg
23 Jun 2026
Local Ai

ITNet: A Learnable Integral Transform That Subsumes Convolution, Attention, and Recurrence

DGX agent

arXiv:2606.19538v2 Announce Type: replace-cross Abstract: Convolutional networks, recurrent networks, and transformers each encode different inductive biases -- locality, sequential memory, and conten

local-aiarxiv-cs-lg
23 Jun 2026
Model Releases

It's Much Easier for Neural Networks to learn Game of Life Dynamics with the Right Activation Function: Polynomial Kolmogorov-Arnold Networks

DGX agent

arXiv:2606.23587v1 Announce Type: new Abstract: Previous work has found a gap between the scale of neural networks that reliably learn Conway's Game of Life, and minimal networks capable of representi

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Jacobian-Adaptive Weighting for Stability: Enhancing Long-term Rollout of Neural Partial Differential Equation Solvers via Spatially-Adaptive Regularization

DGX agent

arXiv:2603.05538v3 Announce Type: replace Abstract: Data-driven surrogate models can significantly accelerate the simulation of continuous dynamical systems, yet the step-wise accumulation of errors d

researcharxiv-cs-lg
23 Jun 2026
Research

Kernel of Partition Paths: A Unified Representation for Tree Ensembles

DGX agent

arXiv:2606.18853v2 Announce Type: replace-cross Abstract: A recent line of work has reframed individual decision trees as linear models on engineered features associated with their splits, opening rou

researcharxiv-cs-lg
23 Jun 2026
Research

Kiwano: A Cutting-Edge Open-Source Toolkit for Speaker Verification

DGX agent

arXiv:2606.22369v1 Announce Type: cross Abstract: In this paper, we present Kiwano, an open-source toolkit designed to advance research and evaluation for speaker verification. Kiwano provides a light

researcharxiv-cs-lg
23 Jun 2026
Research

Kolmogorov-Arnold Reservoir Computing

DGX agent

arXiv:2606.19984v2 Announce Type: replace Abstract: Reservoir computing offers a lightweight framework for forecasting dynamical systems but may struggle to capture long-range dependencies due to limi

researcharxiv-cs-lg
23 Jun 2026
Research

Ky Fan Norms and Beyond: Dual Norms and Combinations for Matrix Optimization

DGX agent

arXiv:2512.09678v2 Announce Type: replace-cross Abstract: In this article, we explore the use of various matrix norms for optimizing functions of weight matrices, a crucial problem in deep learning. M

researcharxiv-cs-lg
23 Jun 2026
Model Releases

L20-Edu-135M: An Auditable Single-GPU Study of Data-Efficient Small Language Modeling

DGX agent

arXiv:2606.22189v1 Announce Type: new Abstract: Small language models are cheap to serve and feasible on local hardware, but strong public 135M-class systems are commonly trained with hundreds of bill

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Lane Change Intention Prediction of two distinct Populations using a Transformer

DGX agent

arXiv:2509.06529v2 Announce Type: replace Abstract: In complex traffic scenarios, intention prediction of surrounding vehicles can improve the strategy of automated driving functions. Existing work on

researcharxiv-cs-lg
23 Jun 2026
Local Ai

Latent Goal Prediction from Language for Model-Based Planning

DGX agent

arXiv:2606.20627v1 Announce Type: cross Abstract: Planning with world models is bottlenecked by compounding prediction errors and the difficulty of defining optimizable goals. Visual targets provide p

local-aiarxiv-cs-lg
23 Jun 2026
Model Releases

LAYUP: Asynchronous decentralized gradient descent with LAYer-wise UPdates

DGX agent

arXiv:2410.05985v4 Announce Type: replace Abstract: The increasing size of deep learning models has made distributed training across multiple devices essential. Synchronous, centralized methods incur

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Learned Controllers for Agile Quadrotors in Pursuit-Evasion Games

DGX agent

arXiv:2506.02849v3 Announce Type: replace-cross Abstract: In this letter we study 1v1 quadrotor pursuit-evasion, where a pursuer and an evader are trained via reinforcement learning (RL) by competing

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Learning a Normal World Model for Few-Shot Boundary-Calibrated Abnormality Detection

DGX agent

arXiv:2606.22261v1 Announce Type: new Abstract: Abnormality detection in complex systems faces two practical barriers: abnormal labels are scarce, and binary labels do not quantify how far an event ha

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Learning-Augmented Algorithms for Online Vertex Cover

DGX agent

arXiv:2606.22831v1 Announce Type: cross Abstract: This paper studies learning-augmented online weighted vertex cover with advice and a parameter lambda in (0,1). We consider two graph cases: bipartite

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Learning Bug Context for PyTorch-to-JAX Translation with LLMs

DGX agent

arXiv:2510.09898v2 Announce Type: replace Abstract: Large language models (LLMs) have shown strong performance on code translation between widely used programming languages. However, translation becom

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Learning by Shifting: Temporal View Construction for Time Series Contrastive Learning

DGX agent

arXiv:2606.21957v1 Announce Type: new Abstract: Supervised learning demands large quantities of labeled data, a bottleneck that is expensive and reliant on domain-specific expertise. Self-supervised l

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Learning Chern Numbers of Topological Insulators with Gauge Equivariant Neural Networks

DGX agent

arXiv:2502.15376v2 Announce Type: replace Abstract: Equivariant network architectures are a well-established tool for predicting invariant or equivariant quantities. However, almost all learning probl

researcharxiv-cs-lg
23 Jun 2026
Research

Learning Expressive Random Feature Models via Parametrized Activations

DGX agent

arXiv:2411.19468v4 Announce Type: replace Abstract: The random feature (RF) method is a powerful kernel approximation technique, but it typically uses fixed activation functions, limiting its adaptabi

researcharxiv-cs-lg
23 Jun 2026
Research

Learning Graphs through Continuous Information Entropy Fields

DGX agent

arXiv:2606.22895v1 Announce Type: new Abstract: Graph theory is inherently descriptive, capturing what relationships exist but not why they arise, because it treats edges as primitive constructs. This

researcharxiv-cs-lg
23 Jun 2026
Safety

Learning Process Rewards via Success Visitation Matching for Efficient RL

DGX agent

arXiv:2606.23640v1 Announce Type: new Abstract: In many modern applications of reinforcement learning (RL), the natural reward for a task of interest is inherently sparse: a reward of 0 is given every

safetyarxiv-cs-lg
23 Jun 2026
Research

Learning thermodynamic master equations for open quantum systems

DGX agent

arXiv:2506.01882v3 Announce Type: replace-cross Abstract: The characterization of Hamiltonians and other components of open quantum dynamical systems plays a crucial role in quantum computing and othe

researcharxiv-cs-lg
23 Jun 2026
Tutorials

Learning through Internalization

DGX agent

arXiv:2606.20937v1 Announce Type: new Abstract: We study internalization processes, by which neural-network-based systems absorb an explicit computational procedure into their own weights, and how the

tutorialsarxiv-cs-lg
23 Jun 2026
Safety

Learning to Place Guards by Reinforcement: A Geo-Free Neural Policy for the Vertex-Guard Art Gallery Problem

DGX agent

arXiv:2606.21604v1 Announce Type: new Abstract: Neural combinatorial optimization (NCO) has shown that policies trained by reinforcement can construct strong solutions to NP-hard problems directly fro

safetyarxiv-cs-lg
23 Jun 2026
Research

Learning with Multiple Correct Answers -- Regret Bounds under Different Feedback Models

DGX agent

arXiv:2602.09402v2 Announce Type: replace Abstract: We study the problem of learning with multiple correct answers, where each instance admits a set of valid labels. We primarily focus on the online s

researcharxiv-cs-lg
23 Jun 2026
Tutorials

Leveraging AutoML for Sustainable Deep Learning: A Multi-Objective HPO Approach on Deep Shift Neural Networks

DGX agent

arXiv:2606.23208v1 Announce Type: new Abstract: Deep Learning (DL) has advanced various fields by extracting complex patterns from large datasets. However, the computational demands of DL models pose

tutorialsarxiv-cs-lg
23 Jun 2026
Model Releases

Leveraging LaBSE with Progressive Curriculum Learning for Multicultural Polarization

DGX agent

arXiv:2606.21718v1 Announce Type: cross Abstract: Detecting online polarization remains a critical challenge, particularly in multilingual and multicultural contexts where intergroup hostility is prev

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Leveraging Similarities in Multi-Armed Bandits

DGX agent

arXiv:2606.23414v1 Announce Type: new Abstract: In many online learning and bandit problems, the actions we consider possess inherent similarities--for instance because they share latent traits, tags,

researcharxiv-cs-lg
23 Jun 2026
Research

LIG: Layer-wise Integrated Gradients for Within-Layer Flow Analysis in Transformers

DGX agent

arXiv:2606.21564v1 Announce Type: new Abstract: Transformers achieve strong performance, but their internal computations remain opaque. We view each Transformer layer as a dynamic graph whose nodes ar

researcharxiv-cs-lg
23 Jun 2026
Research

LK Jam: System Architecture and Implementation of a Real-Time Human-AI Interactive Music Generation System using Role-Aware GRU

DGX agent

arXiv:2606.21018v1 Announce Type: cross Abstract: As artificial intelligence advances into the era of Embodied AI, live musical interaction urgently needs to break free from the limitations of offline

researcharxiv-cs-lg
23 Jun 2026
Tutorials

LLM-Aided A* Search in Non-Geometric Network Graphs

DGX agent

arXiv:2606.23136v1 Announce Type: cross Abstract: Finding the shortest path in non-geometric network graphs, where edge weights encode arbitrary metrics such as latency or monetary cost rather than sp

tutorialsarxiv-cs-lg
23 Jun 2026
Agents

LLM-Guided Test-Time Discovery of Quantum-Chemical Approximation Algorithms

DGX agent

arXiv:2606.20729v1 Announce Type: cross Abstract: Quantum chemistry simulations underpin modern materials discovery, yet their impact is limited by steep computational cost and dependence on fixed app

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

Load Testing for Machine Learning Model Serving Systems at Scale

DGX agent

arXiv:2606.22013v1 Announce Type: new Abstract: Machine learning (ML) model serving has become a dominant consumer of GPU infrastructure, yet capacity planning in these systems remains largely ad hoc.

model-releasesarxiv-cs-lg
23 Jun 2026
Local Ai

Local Causal Attribution of Chain-of-Thought Reasoning

DGX agent

arXiv:2606.21821v1 Announce Type: new Abstract: Understanding the causal structure of a language model's thought process is a problem of significant importance for both transparency and safety. In thi

local-aiarxiv-cs-lg
23 Jun 2026
Local Ai

Local Flow Matching Generative Models

DGX agent

arXiv:2410.02548v4 Announce Type: replace-cross Abstract: Flow Matching (FM) is a simulation-free method for learning a continuous, invertible flow that interpolates between two distributions, and in

local-aiarxiv-cs-lg
23 Jun 2026
Model Releases

Localizing and Editing Knowledge in Large Audio-Language Models

DGX agent

arXiv:2603.14343v2 Announce Type: replace Abstract: Large Audio-Language Models (LALMs) have shown strong performance in speech understanding, making speech a natural interface for accessing factual i

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

LOLLA: Deep Reinforcement Learning for Closed-Loop Link Adaptation Towards a GPU-Accelerated AI-RAN

DGX agent

arXiv:2606.23110v1 Announce Type: cross Abstract: Outer-loop link adaptation (OLLA) is widely deployed in 5G NR to track channel variations, yet its reliance on first-order, single-bit feedback degrad

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Low-variance estimators overcome the phase-gradient bottleneck in complex-valued neural quantum states

DGX agent

arXiv:2606.13912v2 Announce Type: replace-cross Abstract: Complex neural quantum states are difficult to optimize when their wavefunction phase carries gauge, chiral, fermionic, or topological structu

model-releasesarxiv-cs-lg
23 Jun 2026
Research

LSTM Variants for Chaotic Dynamical Systems: An Empirical Study on the Lorenz Attractor

DGX agent

arXiv:2606.22662v1 Announce Type: new Abstract: Forecasting chaotic dynamical systems such as the Lorenz attractor is notoriously difficult: small numerical errors are amplified exponentially over lon

researcharxiv-cs-lg
23 Jun 2026
Research

Machine Learning Classification of Cryopathy Syndromes: A Comprehensive Comparative Study

DGX agent

arXiv:2606.20874v1 Announce Type: new Abstract: Cryopathy syndromes are difficult to classify because laboratory patterns often overlap across diagnostic categories, while some diagnoses are rare. Thi

researcharxiv-cs-lg
23 Jun 2026
Safety

MAGNIFIED: RL Fine-tuning of Multimodal Large Language Models for Motion Planning

DGX agent

arXiv:2606.20641v1 Announce Type: cross Abstract: Multi-modal Large Language Models (MLLMs) have demonstrated remarkable capabilities in semantic understanding and common sense reasoning, making them

safetyarxiv-cs-lg
23 Jun 2026
Agents

MAS-PromptBench: When Does Prompt Optimization Improve Multi-Agent LLM Systems?

DGX agent

arXiv:2606.23664v1 Announce Type: new Abstract: Multi-agent systems (MAS) offer a scalable path forward for agentic AI, comprising multiple LLM-based agents, each assigned a system prompt and a positi

agentsarxiv-cs-lg
23 Jun 2026
Research

Massive Activations Are Architecturally Robust: A Controlled Scratch/Commitment Residual Stream Test

DGX agent

arXiv:2606.20743v1 Announce Type: new Abstract: Trained transformers reliably develop massive activations, a small number of hidden dimensions whose magnitude is far above the median and which concent

researcharxiv-cs-lg
23 Jun 2026
Model Releases

Mat-Pref: Verifiable-Reward Training Improves Compositional Reasoning in Inorganic Materials

DGX agent

arXiv:2606.21830v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) has driven rapid progress in mathematical and code reasoning, but when extended to science, existi

model-releasesarxiv-cs-lg
23 Jun 2026
Tutorials

MAVRL: Learning Reward Functions from Multiple Feedback Types with Amortized Variational Inference

DGX agent

arXiv:2602.15206v2 Announce Type: replace Abstract: Reward learning typically relies on a single feedback type or combines multiple feedback types using manually weighted loss terms. Currently, it rem

tutorialsarxiv-cs-lg
23 Jun 2026
← Previous
1…99100101102103…304
Next →