AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
23 Jun 2026

Inverse Problem for Partial Differential Equations with Jump Discontinuities in Coefficients by Two-stage Physics-Informed Deep Learning and Statistical Mixture Models

Model ReleasesDGX agent

arXiv:2510.14656v2 Announce Type: replace-cross Abstract: This work proposes a two-stage physics-informed deep learning framework that combines neural-network-based sampling with statistical inference

Inverting the Bellman Equation: From Q-Values to World Models

SafetyDGX agent

arXiv:2606.21173v1 Announce Type: new Abstract: Model-based and model-free reinforcement learning are traditionally viewed as separate paradigms: instead of learning a model of the transition kernel P

IRumAI: Reinforcement Learning for Indian Rummy

SafetyDGX agent

arXiv:2606.21975v1 Announce Type: cross Abstract: Despite its massive player base and complex hidden-information dynamics, Indian Rummy has received no reinforcement learning attention. Existing agent


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Is Our Benchmark Enough? An Analysis of Continual Learning for MLLMs

Model ReleasesDGX agent

arXiv:2606.20961v1 Announce Type: new Abstract: Continual adaptation is essential for multimodal large language models (MLLMs) deployed across evolving domains, but the state-of-the-art MR-LoRA method

ITNet: A Learnable Integral Transform That Subsumes Convolution, Attention, and Recurrence

Local AiDGX agent

arXiv:2606.19538v2 Announce Type: replace-cross Abstract: Convolutional networks, recurrent networks, and transformers each encode different inductive biases -- locality, sequential memory, and conten

It's Much Easier for Neural Networks to learn Game of Life Dynamics with the Right Activation Function: Polynomial Kolmogorov-Arnold Networks

Model ReleasesDGX agent

arXiv:2606.23587v1 Announce Type: new Abstract: Previous work has found a gap between the scale of neural networks that reliably learn Conway's Game of Life, and minimal networks capable of representi

Jacobian-Adaptive Weighting for Stability: Enhancing Long-term Rollout of Neural Partial Differential Equation Solvers via Spatially-Adaptive Regularization

ResearchDGX agent

arXiv:2603.05538v3 Announce Type: replace Abstract: Data-driven surrogate models can significantly accelerate the simulation of continuous dynamical systems, yet the step-wise accumulation of errors d

Kernel of Partition Paths: A Unified Representation for Tree Ensembles

ResearchDGX agent

arXiv:2606.18853v2 Announce Type: replace-cross Abstract: A recent line of work has reframed individual decision trees as linear models on engineered features associated with their splits, opening rou

Kiwano: A Cutting-Edge Open-Source Toolkit for Speaker Verification

ResearchDGX agent

arXiv:2606.22369v1 Announce Type: cross Abstract: In this paper, we present Kiwano, an open-source toolkit designed to advance research and evaluation for speaker verification. Kiwano provides a light

Kolmogorov-Arnold Reservoir Computing

ResearchDGX agent

arXiv:2606.19984v2 Announce Type: replace Abstract: Reservoir computing offers a lightweight framework for forecasting dynamical systems but may struggle to capture long-range dependencies due to limi

Ky Fan Norms and Beyond: Dual Norms and Combinations for Matrix Optimization

ResearchDGX agent

arXiv:2512.09678v2 Announce Type: replace-cross Abstract: In this article, we explore the use of various matrix norms for optimizing functions of weight matrices, a crucial problem in deep learning. M

L20-Edu-135M: An Auditable Single-GPU Study of Data-Efficient Small Language Modeling

Model ReleasesDGX agent

arXiv:2606.22189v1 Announce Type: new Abstract: Small language models are cheap to serve and feasible on local hardware, but strong public 135M-class systems are commonly trained with hundreds of bill

Lane Change Intention Prediction of two distinct Populations using a Transformer

ResearchDGX agent

arXiv:2509.06529v2 Announce Type: replace Abstract: In complex traffic scenarios, intention prediction of surrounding vehicles can improve the strategy of automated driving functions. Existing work on

Latent Goal Prediction from Language for Model-Based Planning

Local AiDGX agent

arXiv:2606.20627v1 Announce Type: cross Abstract: Planning with world models is bottlenecked by compounding prediction errors and the difficulty of defining optimizable goals. Visual targets provide p

LAYUP: Asynchronous decentralized gradient descent with LAYer-wise UPdates

Model ReleasesDGX agent

arXiv:2410.05985v4 Announce Type: replace Abstract: The increasing size of deep learning models has made distributed training across multiple devices essential. Synchronous, centralized methods incur

Learned Controllers for Agile Quadrotors in Pursuit-Evasion Games

SafetyDGX agent

arXiv:2506.02849v3 Announce Type: replace-cross Abstract: In this letter we study 1v1 quadrotor pursuit-evasion, where a pursuer and an evader are trained via reinforcement learning (RL) by competing

Learning a Normal World Model for Few-Shot Boundary-Calibrated Abnormality Detection

Model ReleasesDGX agent

arXiv:2606.22261v1 Announce Type: new Abstract: Abnormality detection in complex systems faces two practical barriers: abnormal labels are scarce, and binary labels do not quantify how far an event ha

Learning-Augmented Algorithms for Online Vertex Cover

Model ReleasesDGX agent

arXiv:2606.22831v1 Announce Type: cross Abstract: This paper studies learning-augmented online weighted vertex cover with advice and a parameter lambda in (0,1). We consider two graph cases: bipartite

Learning Bug Context for PyTorch-to-JAX Translation with LLMs

Model ReleasesDGX agent

arXiv:2510.09898v2 Announce Type: replace Abstract: Large language models (LLMs) have shown strong performance on code translation between widely used programming languages. However, translation becom

Learning by Shifting: Temporal View Construction for Time Series Contrastive Learning

Model ReleasesDGX agent

arXiv:2606.21957v1 Announce Type: new Abstract: Supervised learning demands large quantities of labeled data, a bottleneck that is expensive and reliant on domain-specific expertise. Self-supervised l

Learning Chern Numbers of Topological Insulators with Gauge Equivariant Neural Networks

ResearchDGX agent

arXiv:2502.15376v2 Announce Type: replace Abstract: Equivariant network architectures are a well-established tool for predicting invariant or equivariant quantities. However, almost all learning probl

Learning Expressive Random Feature Models via Parametrized Activations

ResearchDGX agent

arXiv:2411.19468v4 Announce Type: replace Abstract: The random feature (RF) method is a powerful kernel approximation technique, but it typically uses fixed activation functions, limiting its adaptabi

Learning Graphs through Continuous Information Entropy Fields

ResearchDGX agent

arXiv:2606.22895v1 Announce Type: new Abstract: Graph theory is inherently descriptive, capturing what relationships exist but not why they arise, because it treats edges as primitive constructs. This

Learning Process Rewards via Success Visitation Matching for Efficient RL

SafetyDGX agent

arXiv:2606.23640v1 Announce Type: new Abstract: In many modern applications of reinforcement learning (RL), the natural reward for a task of interest is inherently sparse: a reward of 0 is given every

Learning thermodynamic master equations for open quantum systems

ResearchDGX agent

arXiv:2506.01882v3 Announce Type: replace-cross Abstract: The characterization of Hamiltonians and other components of open quantum dynamical systems plays a crucial role in quantum computing and othe

Learning through Internalization

TutorialsDGX agent

arXiv:2606.20937v1 Announce Type: new Abstract: We study internalization processes, by which neural-network-based systems absorb an explicit computational procedure into their own weights, and how the

Learning to Place Guards by Reinforcement: A Geo-Free Neural Policy for the Vertex-Guard Art Gallery Problem

SafetyDGX agent

arXiv:2606.21604v1 Announce Type: new Abstract: Neural combinatorial optimization (NCO) has shown that policies trained by reinforcement can construct strong solutions to NP-hard problems directly fro

Learning with Multiple Correct Answers -- Regret Bounds under Different Feedback Models

ResearchDGX agent

arXiv:2602.09402v2 Announce Type: replace Abstract: We study the problem of learning with multiple correct answers, where each instance admits a set of valid labels. We primarily focus on the online s

Leveraging AutoML for Sustainable Deep Learning: A Multi-Objective HPO Approach on Deep Shift Neural Networks

TutorialsDGX agent

arXiv:2606.23208v1 Announce Type: new Abstract: Deep Learning (DL) has advanced various fields by extracting complex patterns from large datasets. However, the computational demands of DL models pose

Leveraging LaBSE with Progressive Curriculum Learning for Multicultural Polarization

Model ReleasesDGX agent

arXiv:2606.21718v1 Announce Type: cross Abstract: Detecting online polarization remains a critical challenge, particularly in multilingual and multicultural contexts where intergroup hostility is prev

Leveraging Similarities in Multi-Armed Bandits

ResearchDGX agent

arXiv:2606.23414v1 Announce Type: new Abstract: In many online learning and bandit problems, the actions we consider possess inherent similarities--for instance because they share latent traits, tags,

LIG: Layer-wise Integrated Gradients for Within-Layer Flow Analysis in Transformers

ResearchDGX agent

arXiv:2606.21564v1 Announce Type: new Abstract: Transformers achieve strong performance, but their internal computations remain opaque. We view each Transformer layer as a dynamic graph whose nodes ar

LK Jam: System Architecture and Implementation of a Real-Time Human-AI Interactive Music Generation System using Role-Aware GRU

ResearchDGX agent

arXiv:2606.21018v1 Announce Type: cross Abstract: As artificial intelligence advances into the era of Embodied AI, live musical interaction urgently needs to break free from the limitations of offline

LLM-Aided A* Search in Non-Geometric Network Graphs

TutorialsDGX agent

arXiv:2606.23136v1 Announce Type: cross Abstract: Finding the shortest path in non-geometric network graphs, where edge weights encode arbitrary metrics such as latency or monetary cost rather than sp

LLM-Guided Test-Time Discovery of Quantum-Chemical Approximation Algorithms

AgentsDGX agent

arXiv:2606.20729v1 Announce Type: cross Abstract: Quantum chemistry simulations underpin modern materials discovery, yet their impact is limited by steep computational cost and dependence on fixed app

Load Testing for Machine Learning Model Serving Systems at Scale

Model ReleasesDGX agent

arXiv:2606.22013v1 Announce Type: new Abstract: Machine learning (ML) model serving has become a dominant consumer of GPU infrastructure, yet capacity planning in these systems remains largely ad hoc.

Local Causal Attribution of Chain-of-Thought Reasoning

Local AiDGX agent

arXiv:2606.21821v1 Announce Type: new Abstract: Understanding the causal structure of a language model's thought process is a problem of significant importance for both transparency and safety. In thi

Local Flow Matching Generative Models

Local AiDGX agent

arXiv:2410.02548v4 Announce Type: replace-cross Abstract: Flow Matching (FM) is a simulation-free method for learning a continuous, invertible flow that interpolates between two distributions, and in

Localizing and Editing Knowledge in Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2603.14343v2 Announce Type: replace Abstract: Large Audio-Language Models (LALMs) have shown strong performance in speech understanding, making speech a natural interface for accessing factual i

LOLLA: Deep Reinforcement Learning for Closed-Loop Link Adaptation Towards a GPU-Accelerated AI-RAN

SafetyDGX agent

arXiv:2606.23110v1 Announce Type: cross Abstract: Outer-loop link adaptation (OLLA) is widely deployed in 5G NR to track channel variations, yet its reliance on first-order, single-bit feedback degrad

Low-variance estimators overcome the phase-gradient bottleneck in complex-valued neural quantum states

Model ReleasesDGX agent

arXiv:2606.13912v2 Announce Type: replace-cross Abstract: Complex neural quantum states are difficult to optimize when their wavefunction phase carries gauge, chiral, fermionic, or topological structu

LSTM Variants for Chaotic Dynamical Systems: An Empirical Study on the Lorenz Attractor

ResearchDGX agent

arXiv:2606.22662v1 Announce Type: new Abstract: Forecasting chaotic dynamical systems such as the Lorenz attractor is notoriously difficult: small numerical errors are amplified exponentially over lon

Machine Learning Classification of Cryopathy Syndromes: A Comprehensive Comparative Study

ResearchDGX agent

arXiv:2606.20874v1 Announce Type: new Abstract: Cryopathy syndromes are difficult to classify because laboratory patterns often overlap across diagnostic categories, while some diagnoses are rare. Thi

MAGNIFIED: RL Fine-tuning of Multimodal Large Language Models for Motion Planning

SafetyDGX agent

arXiv:2606.20641v1 Announce Type: cross Abstract: Multi-modal Large Language Models (MLLMs) have demonstrated remarkable capabilities in semantic understanding and common sense reasoning, making them

MAS-PromptBench: When Does Prompt Optimization Improve Multi-Agent LLM Systems?

AgentsDGX agent

arXiv:2606.23664v1 Announce Type: new Abstract: Multi-agent systems (MAS) offer a scalable path forward for agentic AI, comprising multiple LLM-based agents, each assigned a system prompt and a positi

Massive Activations Are Architecturally Robust: A Controlled Scratch/Commitment Residual Stream Test

ResearchDGX agent

arXiv:2606.20743v1 Announce Type: new Abstract: Trained transformers reliably develop massive activations, a small number of hidden dimensions whose magnitude is far above the median and which concent

Mat-Pref: Verifiable-Reward Training Improves Compositional Reasoning in Inorganic Materials

Model ReleasesDGX agent

arXiv:2606.21830v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) has driven rapid progress in mathematical and code reasoning, but when extended to science, existi

MAVRL: Learning Reward Functions from Multiple Feedback Types with Amortized Variational Inference

TutorialsDGX agent

arXiv:2602.15206v2 Announce Type: replace Abstract: Reward learning typically relies on a single feedback type or combines multiple feedback types using manually weighted loss terms. Currently, it rem

Measuring Intent Comprehension in LLMs

Model ReleasesDGX agent

arXiv:2506.16584v3 Announce Type: replace-cross Abstract: People judge interactions with large language models (LLMs) as successful when outputs match what they want, not what they type. Yet LLMs are

Measuring Model-Induced Discrimination via Efficient Fairness Approximation

SafetyDGX agent

arXiv:2405.09251v2 Announce Type: replace Abstract: Providing various machine learning (ML) applications in the real world, concerns about discrimination hidden in ML models are growing, particularly

MedFedPure: A Medical Federated Framework with MAE-based Detection and Diffusion Purification for Inference-Time Attacks

Local AiDGX agent

arXiv:2511.11625v2 Announce Type: replace Abstract: Artificial intelligence (AI) has shown great potential in medical imaging, particularly for brain tumor detection using Magnetic Resonance Imaging (

MedTS-TTT: Test-Time Training for Medical Time Series Classification

Model ReleasesDGX agent

arXiv:2606.21329v1 Announce Type: new Abstract: Medical time series (MedTS) signals such as electroencephalography (EEG) and electrocardiography (ECG) support many clinical applications. However, subs

Memory Contagion: Cross-Temporal Propagation of Evaluator Bias via Agent Memory

SafetyDGX agent

arXiv:2606.23195v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly rely on memory systems to maintain long-term coherence. Recent work shows that agent memories degrade dur

Memory Is No Longer a Bottleneck: Memory-Efficient Graph Filtering for Scalable Collaborative Filtering

ResearchDGX agent

arXiv:2606.21540v1 Announce Type: new Abstract: Graph convolutional networks (GCNs) have demonstrated significant success in capturing complex user-item relationships for collaborative filtering (CF).

Meta-Reinforcement Learning via Evolution for Multi-Objective Combinatorial Supply Chain Optimisation

SafetyDGX agent

arXiv:2606.22146v1 Announce Type: new Abstract: Meta-reinforcement learning is a promising approach to multi-objective optimisation because it enables rapid policy adaptation across changing environme

Mind the Noise: Sensitivity of Transformer-based Interaction-Aware Trajectory Prediction Models to Noisy Data

Local AiDGX agent

arXiv:2606.21344v1 Announce Type: cross Abstract: Trajectory prediction allows autonomous vehicles to anticipate the future behavior of surrounding objects (or agents) and, accordingly, maximize the s

Minimax Quantile Lower Bounds for Interactive Statistical Decision Making with Privacy

ResearchDGX agent

arXiv:2606.23096v1 Announce Type: new Abstract: Minimax risk and regret are expectation-based criteria and do not capture rare but consequential failures. To address this concern, we develop a elta-ex

Mixture-of-Experts Graph Transformers for Interpretable Particle Collision Detection

ResearchDGX agent

arXiv:2501.03432v3 Announce Type: replace Abstract: The Large Hadron Collider at CERN produces immense volumes of complex data from high-energy particle collisions, demanding sophisticated analytical

mlx-vis: GPU-Native Dimensionality Reduction on Apple Silicon

Local AiDGX agent

arXiv:2603.04035v4 Announce Type: replace Abstract: Dimensionality reduction is a foundational tool for visualizing high-dimensional data, yet its reference implementations span a fragmented stack of

MMGNN: Multi-level, multi-color graph neural networks for molecular property prediction

ResearchDGX agent

arXiv:2606.20906v1 Announce Type: new Abstract: Molecular message-passing neural networks commonly propagate chemically diverse interactions through a single graph, which may mix interaction-specific

← Previous
1…7980818283…243
Next →