AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
23 Jun 2026

Structure-Aware Compound-Protein Affinity Prediction via Graph Neural Networks with Group Lasso Regularization

ResearchDGX agent

arXiv:2507.03318v3 Announce Type: replace Abstract: Explainable artificial intelligence approaches accelerate drug discovery by improving molecular representation learning, identifying key molecular s

Structure-Aware Graph Multi-Task Learning for Dynamic Sparse OD Demand Prediction

ApplicationsDGX agent

arXiv:2606.21022v1 Announce Type: new Abstract: Origin-Destination (OD) demand prediction is fundamental to intelligent transportation systems, yet real-world OD flows are often dynamically sparse, lo

Sub-Billion, Super-Frontier: Small Language Models Rival Zero-Shot Frontier LLMs on General and Literary Relation Extraction

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.22606v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong relation extraction (RE), but their computational demands and reliance on proprietary APIs limit deploymen

Sublinearly Structured Deep Neural Networks Achieve Feature Learning Consistency for Compositional Functions

ResearchDGX agent

arXiv:2606.23477v1 Announce Type: cross Abstract: Over the past decade, deep neural networks (DNNs) have achieved remarkable success on complex machine-learning tasks, yet the theoretical foundations

Subsampling for supervised learning in reproducing kernel Hilbert spaces

ApplicationsDGX agent

arXiv:2606.21260v1 Announce Type: cross Abstract: In the era of big data, subsampling became a common practice in statistical learning. By selecting a subgroup of individuals based on which the learne

Subspace-Constrained Federated Learning with Low-Rank Adaptation

Local AiDGX agent

arXiv:2606.22724v1 Announce Type: new Abstract: Federated low-rank adaptation methods are attractive for fine-tuning large models under communication and privacy constraints, but heterogeneous client

Substitution-Based Analysis of Structural Novelty for Generative Models of Materials

SafetyDGX agent

arXiv:2606.23166v1 Announce Type: new Abstract: There has been rapid progress in generative artificial intelligence (AI) models for inorganic crystal design, which can efficiently generate large numbe

SuperCond-GNN: Scalable Graph Neural Network Surrogate for Superconducting Circuit Simulations

Local AiDGX agent

arXiv:2606.23548v1 Announce Type: cross Abstract: This paper presents SuperCond-GNN, a graph neural network-based surrogate model for predicting the voltage distribution in high-temperature supercondu

Superhuman AI for Generals.io Using Self-Play Reinforcement Learning

SafetyDGX agent

arXiv:2606.23348v1 Announce Type: new Abstract: We present a superhuman AI agent for Generals.io, a real-time strategy game that requires both long-horizon planning and short-term tactics under strong

Surprise-Guided MergeSort: Budget-Efficient Human-in-the-Loop Ranking via Adaptive Comparison Scheduling

SafetyDGX agent

arXiv:2606.15623v2 Announce Type: replace Abstract: Pairwise comparison is the gold standard for subjective ranking tasks; however, exhaustive annotation requires a massive number of human comparisons

SVD-Surgeon: Optimal Singular-Value Surgery for Large Language Model Compression

Model ReleasesDGX agent

arXiv:2606.23568v1 Announce Type: new Abstract: Large language models (LLMs) achieve remarkable performance across a wide range of tasks, but their deployment is constrained by substantial memory and

Synthetic Network Packet Generation through Statistical Learning and Genetic Algorithms

ResearchDGX agent

arXiv:2606.20864v1 Announce Type: cross Abstract: Developing robust intrusion detection systems (IDS) for IoT environments requires large, labeled datasets capturing realistic traffic distributions ac

TaigiSpeech: A Low-Resource Real-World Speech Intent Dataset and Preliminary Results with Scalable Data Mining In-the-Wild

Model ReleasesDGX agent

arXiv:2603.21478v2 Announce Type: replace-cross Abstract: Speech technologies have advanced rapidly and serve diverse populations worldwide. However, many languages remain underrepresented due to limi

TaLK: Text-attributed Graph Dataset Distillation via Coupling Language Model with Graph-Aware Kernel

ApplicationsDGX agent

arXiv:2606.22975v1 Announce Type: new Abstract: Text-attributed graphs (TAGs) are widely used in many real-world domains, and learning on TAGs requires jointly modeling text semantics and graph struct

Tapered Language Models

Model ReleasesDGX agent

arXiv:2606.23670v1 Announce Type: new Abstract: Modern language models, including transformer, recurrent, and memory-based variants, share a common chassis: a stack of identical layers in which parame

Target-Aware Linear Regression Under Distribution Shift

Model ReleasesDGX agent

arXiv:2606.22775v1 Announce Type: cross Abstract: Distribution shift between training and deployment is a pervasive challenge for modern AI systems. In many cases, the target marginals of covariates a

Task-Agnostic Federated Continual Learning via Replay-Free Gradient Projection

ResearchDGX agent

arXiv:2509.21606v3 Announce Type: replace Abstract: Federated continual learning (FCL) enables collaborative model training across distributed clients on sequentially arriving tasks without revisiting

Task-Differentiated Atomic Skill Expansion and Routing for Continual Learning Across Highly Heterogeneous Tasks

Model ReleasesDGX agent

arXiv:2606.21307v1 Announce Type: new Abstract: Continual learning (CL) is commonly studied under the assumption that sequential tasks are semantically related or structurally similar. However, in hig

Tell Me: An LLM-powered Mental Well-being Assistant with RAG, Synthetic Dialogue Generation, and Agentic Planning

AgentsDGX agent

arXiv:2511.14445v2 Announce Type: replace-cross Abstract: We present Tell Me, a mental well-being system that leverages advances in large language models to provide accessible, context-aware support f

Temper-Then-Tilt: Principled Unlearning for Generative Models through Tempering and Classifier Guidance

Model ReleasesDGX agent

arXiv:2602.10217v2 Announce Type: replace Abstract: We study machine unlearning in large generative models by framing the task as density ratio estimation to a target distribution rather than supervis

Temporal Causal Prior-Data Fitted Networks for Panel Data with Learned Reliability Signals

Model ReleasesDGX agent

arXiv:2606.20889v1 Announce Type: new Abstract: Estimating causal effects in industrial time series requires handling temporal dynamics, time-varying treatments, and unobserved confounders. Existing c

Temporal Graph Pattern Machine

Model ReleasesDGX agent

arXiv:2601.22454v3 Announce Type: replace Abstract: Temporal graph learning is pivotal for deciphering dynamic systems, where the core challenge lies in explicitly modeling the underlying evolving pat

Temporal-Spectral Alignment with Frequency Adaptation for Source-Free Time-Series Adaptation

Model ReleasesDGX agent

arXiv:2606.23120v1 Announce Type: new Abstract: The goal of source-free domain adaptation (SFDA) for time-series data is to transfer knowledge from a pre-trained source model to an unlabeled target do

Tensor Train Decomposition-based 3D Implicit Full Waveform Inversion with Multi-scale Structural Similarity

ResearchDGX agent

arXiv:2606.22867v1 Announce Type: cross Abstract: Three-dimensional full waveform inversion (3DFWI) is a powerful technique for reconstructing high-resolution subsurface velocity models. However, its

TF-SNO: Time-Frequency Gated Spectral Neural Operators for Learning Non-Stationary Partial Differential Equations

ResearchDGX agent

arXiv:2606.21189v1 Announce Type: new Abstract: Non-stationary partial differential equations (PDEs) arise throughout scientific computing, where the dominant frequency content and energy distribution

The Alignment Problem in Constrained Code Generation

SafetyDGX agent

arXiv:2606.21619v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in code generation, but their outputs frequently contain syntax or type errors that

The Anatomy of the CTC Oracle Gap: Acoustic Exhaustion and Linguistic Recovery

ResearchDGX agent

arXiv:2606.23306v1 Announce Type: cross Abstract: We study the limits of CTC-internal scoring for N-best hypothesis selection and locate the information bottleneck separating acoustic confidence from

The Cost Geometry of Belief: finite-resource inference under noisy observation

ResearchDGX agent

arXiv:2606.21585v1 Announce Type: new Abstract: We equip the space of beliefs with a cost geometry (what it costs to pass from one belief to another): optimal transport in Wasserstein space, reweighte

The Efficiency Frontier: Classical Shadows versus Direct Quantum Measurement

ApplicationsDGX agent

arXiv:2509.06218v3 Announce Type: replace-cross Abstract: Interfacing quantum and classical processors is an important subroutine in full-stack quantum algorithms. The so-called ``classical shadow'' m

The Energy Consumption of Transformer Fine-Tuning: A Roofline-Inspired Scaling Model

ResearchDGX agent

arXiv:2606.23546v1 Announce Type: new Abstract: Transformer-based models underpin modern natural language processing but incur rapidly growing computational and energy costs. As training scales in bot

The Fractal Neural Operator: Overcoming Spectral Bias in Chaotic Attractors via Prime-Harmonic Weierstrass Encodings

SafetyDGX agent

arXiv:2606.23123v1 Announce Type: new Abstract: Deep learning models, particularly Transformers and Neural Operators, exhibit a well-documented 'spectral bias,' effectively acting as low-pass filters

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs

Model ReleasesDGX agent

arXiv:2606.22686v1 Announce Type: cross Abstract: Modern Large Language Models (LLMs) rely on extensive safety alignment, yet the mechanistic basis of refusal remains opaque. In this work, we investig

The Metanym Game: A Self-Contained, Self-Consistent LLM Peer-Community Benchmark for Structural Intelligence

Model ReleasesDGX agent

arXiv:2606.21008v1 Announce Type: cross Abstract: The metanym game is a competitive word game for LLMs that measures structural intelligence against established cognitive-science constructs. No conten

The New Associationism: Lessons from Deep Learning

TutorialsDGX agent

arXiv:2606.20600v1 Announce Type: cross Abstract: What can the success of modern AI tell us about how humans learn? This paper argues that taking AI seriously as a model of human learning supports a m

The Optimal Token Baseline: Variance Reduction for Long-Horizon LLM-RL

ResearchDGX agent

arXiv:2602.07078v2 Announce Type: replace Abstract: Reinforcement Learning (RL) for Large Language Models (LLMs) often suffers from training collapse in long-horizon tasks due to exploding gradient va

The Optimization Landscape of Caratheodory Decomposition of Toeplitz Covariances

ResearchDGX agent

arXiv:2511.01605v2 Announce Type: replace Abstract: Toeplitz covariance estimation is a classical problem in statistical signal processing, yet the geometry of the Gaussian maximum-likelihood objectiv

The Pitfall of Scaling Up: Uncovering and Mitigating Popularity Bias Amplification in Scaling Transformer-based Recommenders

SafetyDGX agent

arXiv:2606.21911v1 Announce Type: cross Abstract: We identify a critical pitfall in scaling transformer-based sequential recommenders: while increasing model size improves recommendation accuracy, it

The Reservoir Attention Network: Cross-Pass State in Pretrained Transformers via Content-Addressable Reservoir Injection

HardwareDGX agent

arXiv:2606.15678v2 Announce Type: replace Abstract: A feasibility and dynamics study of the Reservoir Attention Network (RAN), an architecture that injects a fixed, randomly-initialized reservoir into

The Trilemma of Truth in Large Language Models

ResearchDGX agent

arXiv:2506.23921v5 Announce Type: replace-cross Abstract: The public often attributes human-like qualities to large language models (LLMs), assuming that they 'know' certain things. In reality, LLMs e

The Two-Hump Problem: Bridging the Difficulty Gap in Mathematical Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.21611v1 Announce Type: new Abstract: Mathematical search problems present a unique challenge for Reinforcement Learning (RL) due to vast search spaces and sparse rewards. In previous works,

The Unseen Hand: Manipulating Model Fairness and SHAP with Targeted Identity Re-Association Attacks

SafetyDGX agent

arXiv:2606.22858v1 Announce Type: new Abstract: As machine learning models grow more influential and opaque, algorithmic fairness and explainability are critical for ensuring accountability. However,

ThermoLLM: Thermodynamics-Aware HVAC Control with Spatial-Semantic Knowledge Graph

ResearchDGX agent

arXiv:2606.22911v1 Announce Type: cross Abstract: Multi-zone HVAC control is a spatial decision problem in which indoor thermal evolution and control decisions depend not only on outdoor conditions an

Time Series Classification through Diffeomorphic Time Warping (DiffTW)

SafetyDGX agent

arXiv:2606.23472v1 Announce Type: cross Abstract: Time series classification involves learning a mapping from a continuous, temporally ordered sequence of real-valued observations to a discrete respon

TIP-Search: Time-Predictable Inference Scheduling for Market Prediction under Uncertain Load

SafetyDGX agent

arXiv:2506.08026v3 Announce Type: replace-cross Abstract: Real-time market prediction services need correct predictions before a decision deadline; a correct prediction delivered late is not a usable

Token Factory: Efficiently Integrating Diverse Signals into Large Recommendation Models

ApplicationsDGX agent

arXiv:2606.19635v2 Announce Type: replace-cross Abstract: Large Recommendation Models (LRMs) have demonstrated promising capabilities in industry-scale recommendation tasks. However, holistically inte

Topological Neural Dynamics: A Neuron-wise Framework for Sequence Modeling

SafetyDGX agent

arXiv:2606.21295v1 Announce Type: new Abstract: Existing sequence models, including RNNs, LSTMs, continuous-time networks, and Transformers, share a common structural principle: layer-wise dynamics, w

Topological Out-of-Domain Generalization in Dynamical Systems Reconstruction

Model ReleasesDGX agent

arXiv:2606.22969v1 Announce Type: new Abstract: Predicting the behavior of dynamical systems (DS) beyond the dynamical and parameter regimes observed in training is a pivotal and essentially unresolve

Toward Non-Expert Customized Congestion Control: Large Language Model-Assisted CCA Code Generation with eBPF Deployment

ApplicationsDGX agent

arXiv:2601.22461v2 Announce Type: replace-cross Abstract: General-purpose congestion control algorithms (CCAs) are designed to achieve general congestion control goals, but they may not meet the speci

Towards Adaptive Categories: Dimensional Governance for Agentic AI

AgentsDGX agent

arXiv:2505.11579v3 Announce Type: replace-cross Abstract: As AI systems evolve from static tools to dynamic agents, traditional categorical governance frameworks -- based on fixed risk tiers, levels o

Towards CSI-Native Foundation Models: A Channel-Adaptive Roadmap for 6G

ResearchDGX agent

arXiv:2606.20670v1 Announce Type: new Abstract: Wireless foundation models offer a path toward reusable channel state information (CSI) intelligence for sixth-generation (6G) systems. However, existin

Towards Robust Personalized Federated Learning: Vulnerability Assessment and Defense Co-Design

Model ReleasesDGX agent

arXiv:2606.22782v1 Announce Type: new Abstract: The proliferation of IoT devices has fueled distributed edge systems to collect vast amounts of sensitive data, creating fertile ground for on-device ma

Towards Robust Training in NNGPT AutoML Pipeline: A Loss-Optimizer Pairing Selection Study

ResearchDGX agent

arXiv:2606.20933v1 Announce Type: new Abstract: The choice of loss function and optimizer is an important decision, that shapes further model training. Yet automated architecture search pipelines (Aut

Towards Understanding the Power and Limits of the Muon Optimizer: A River-Valley Perspective

ResearchDGX agent

arXiv:2606.21514v1 Announce Type: new Abstract: Recently, Muon has gained substantial attention as an appealing alternative to Adam-like optimizers, with many works highlighting its advantages through

Towards Whole Hand and Wrist Kinematic Tracking with a Wearable A-Mode Ultrasound Probe

Local AiDGX agent

arXiv:2606.22333v1 Announce Type: cross Abstract: A-mode ultrasound (US) has emerged as a promising modality for hand and wrist motion tracking. Prior works have mainly addressed static gesture classi

Training Diffusion Policies via Prior-Mapping Co-Evolution

SafetyDGX agent

arXiv:2512.02581v3 Announce Type: replace Abstract: Reinforcement learning (RL) faces a persistent tension: policies that are stable to optimize (e.g., Gaussians) are often too simple to represent the

Training-free Task Classification for Multi-Task Model Merging

Model ReleasesDGX agent

arXiv:2606.22589v1 Announce Type: new Abstract: Ever since the advent of foundation models and the pre-training-finetuning paradigm, there have been numerous efforts to merge multiple task-specific ex

Transcribing Bengali Text with Regional Dialects to IPA using District Guided Tokens

Local AiDGX agent

arXiv:2403.17407v4 Announce Type: replace-cross Abstract: Accurate transcription of Bengali text to the International Phonetic Alphabet (IPA) is a challenging task due to the complex phonology of the

TROPT: An Open Framework for Unifying and Advancing Discrete Text Optimization

SafetyDGX agent

arXiv:2606.23496v1 Announce Type: new Abstract: Discrete text-trigger optimization -- searching for text sequences that, when ingested by a model, steer it toward a specified objective -- underpins mo

Turning Tabular Foundation Models into Graph Foundation Models

ResearchDGX agent

arXiv:2508.20906v3 Announce Type: replace Abstract: While foundation models have revolutionized fields such as natural language processing and computer vision, their potential in graph machine learnin

Two-Bridge: Exclusive Objectives and Extended Horizon StarCraft II Benchmark

Model ReleasesDGX agent

arXiv:2603.06608v2 Announce Type: replace-cross Abstract: The research community lacks a middle ground between StarCraft II full game and its mini-games. The full-game's sprawling state-action space r

← Previous
1…8384858687…243
Next →