AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
21 Apr 2026

Harness as an Asset: Enforcing Determinism via the Convergent AI Agent Framework (CAAF)

Model ReleasesDGX agent

arXiv:2604.17025v1 Announce Type: cross Abstract: Large Language Models (LLMs) produce a controllability gap in safety-critical engineering: even low rates of undetected constraint violations render a

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment

SafetyDGX agent

arXiv:2604.17928v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Reward (RLVR) has proven effective for training reasoning-oriented large language models, but existing methods la

Heterogeneous Self-Play for Realistic Highway Traffic Simulation

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.16406v1 Announce Type: cross Abstract: Realistic highway simulation is critical for scalable safety evaluation of autonomous vehicles, particularly for interactions that are too rare to stu

Horizon-Aware Forecasting of Passenger Assistance Demand for Rail Station Workforce Planning

ApplicationsDGX agent

arXiv:2604.16464v1 Announce Type: cross Abstract: Passenger assistance services are essential for accessible rail travel, yet demand varies substantially across stations and over time, creating challe

Horospherical Depth and Busemann Median on Hadamard Manifolds

ResearchDGX agent

arXiv:2604.18242v1 Announce Type: cross Abstract: We introduce the horospherical depth, an intrinsic notion of statistical depth on Hadamard manifolds, and define the Busemann median as the set of its

How Much Cache Does Reasoning Need? Depth-Cache Tradeoffs in KV-Compressed Transformers

ResearchDGX agent

arXiv:2604.17935v1 Announce Type: new Abstract: The key-value (KV) cache is the dominant memory bottleneck during Transformer inference, yet little is known theoretically about how aggressively it can

How Much Data is Enough? The Zeta Law of Discoverability in Biomedical Data, featuring the enigmatic Riemann zeta function

ResearchDGX agent

arXiv:2604.17581v1 Announce Type: new Abstract: How much data is enough to make a scientific discovery? As biomedical datasets scale to millions of samples and AI models grow in capacity, progress inc

How Robustly do LLMs Understand Execution Semantics?

Model ReleasesDGX agent

arXiv:2604.16320v1 Announce Type: cross Abstract: LLMs demonstrate remarkable reasoning capabilities, yet whether they utilize internal world models or rely on sophisticated pattern matching remains o

How to Approximate Inference with Subtractive Mixture Models

TutorialsDGX agent

arXiv:2604.16714v1 Announce Type: new Abstract: Classical mixture models (MMs) are widely used tractable proposals for approximate inference settings such as variational inference (VI) and importance

Hybrid Spectro-Temporal Fusion Framework for Structural Health Monitoring

SafetyDGX agent

arXiv:2604.16589v1 Announce Type: new Abstract: Structural health monitoring plays a critical role in ensuring structural safety by analyzing vibration responses from engineering systems. This paper p

IDOBE: Infectious Disease Outbreak forecasting Benchmark Ecosystem

Model ReleasesDGX agent

arXiv:2604.18521v1 Announce Type: new Abstract: Epidemic forecasting has become an integral part of real-time infectious disease outbreak response. While collaborative ensembles composed of statistica

Implicit neural representations as a coordinate-based framework for continuous environmental field reconstruction from sparse ecological observations

SafetyDGX agent

arXiv:2604.18083v1 Announce Type: new Abstract: Reconstructing continuous environmental fields from sparse and irregular observations remains a central challenge in environmental modelling and biodive

Improving reproducibility by controlling random seed stability in machine learning based estimation via bagging

ResearchDGX agent

arXiv:2604.17694v1 Announce Type: cross Abstract: Predictions from machine learning algorithms can vary across random seeds, inducing instability in downstream debiased machine learning estimators. We

In-Context Learning Under Regime Change

ApplicationsDGX agent

arXiv:2604.16988v1 Announce Type: new Abstract: Non-stationary sequences arise naturally in control, forecasting, and decision-making. The data-generating process shifts at unknown times, and models m

In-Context Symbolic Regression for Robustness-Improved Kolmogorov-Arnold Networks

Model ReleasesDGX agent

arXiv:2603.15250v2 Announce Type: replace Abstract: Symbolic regression aims to replace black-box predictors with concise analytical expressions that can be inspected and validated in scientific machi

In Search of Lost DNA Sequence Pretraining

ResearchDGX agent

arXiv:2604.16570v1 Announce Type: new Abstract: DNA sequence encoding is fundamental to gene function prediction, protein synthesis, and diverse downstream biological tasks. Despite the substantial pr

In Situ Training of Implicit Neural Compressors for Scientific Simulations via Sketch-Based Regularization

ResearchDGX agent

arXiv:2511.02659v3 Announce Type: replace Abstract: Focusing on implicit neural representations, we present a novel in situ training protocol that employs limited memory buffers of full and sketched d

Incremental learning for audio classification with Hebbian Deep Neural Networks

TutorialsDGX agent

arXiv:2604.18270v1 Announce Type: cross Abstract: The ability of humans for lifelong learning is an inspiration for deep learning methods and in particular for continual learning. In this work, we app

Inter-Agent Relative Representations for Multi-Agent Option Discovery

SafetyDGX agent

arXiv:2512.24827v3 Announce Type: replace Abstract: Temporally extended actions improve the ability to explore and plan in single-agent settings. In multi-agent settings, the exponential growth of the

Interpolating Discrete Diffusion Models with Controllable Resampling

Model ReleasesDGX agent

arXiv:2604.17310v1 Announce Type: new Abstract: Discrete diffusion models form a powerful class of generative models across diverse domains, including text and graphs. However, existing approaches fac

Introducing the O-Value: A Universal Standardization for Confusion-Matrix-Based Classification Performance Metrics

ApplicationsDGX agent

arXiv:2505.07033v2 Announce Type: replace-cross Abstract: Many classification performance metrics exist, each suited to a specific application. However, these metrics often differ in scale and can exh

Knowledge without Wisdom: Measuring Misalignment between LLMs and Intended Impact

Model ReleasesDGX agent

arXiv:2603.00883v2 Announce Type: replace Abstract: LLMs increasingly excel on AI benchmarks, but doing so does not guarantee validity for downstream tasks. This study contrasts LLM alignment on bench

L1 Regularization Paths in Linear Models by Parametric Gaussian Message Passing

ResearchDGX agent

arXiv:2604.16949v1 Announce Type: new Abstract: The paper considers the computation of L1 regularization paths in a state space setting, which includes L1 regularized Kalman smoothing, linear SVM, LAS

LASER: Low-Rank Activation SVD for Efficient Recursion

ResearchDGX agent

arXiv:2604.17224v1 Announce Type: new Abstract: Recursive architectures such as Tiny Recursive Models (TRMs) perform implicit reasoning through iterative latent computation, yet the geometric structur

Late Fusion Neural Operators for Extrapolation Across Parameter Space in Partial Differential Equations

Model ReleasesDGX agent

arXiv:2604.16721v1 Announce Type: new Abstract: Developing neural operators that accurately predict the behavior of systems governed by partial differential equations (PDEs) across unseen parameter re

Learning-Based Sparsification of Dynamic Graphs in Robotic Exploration Algorithms

SafetyDGX agent

arXiv:2604.16509v1 Announce Type: cross Abstract: Many robotic exploration algorithms rely on graph structures for frontier-based exploration and dynamic path planning. However, these graphs grow rapi

Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes

ApplicationsDGX agent

arXiv:2604.18381v1 Announce Type: cross Abstract: Fine-tuning Large Language Models (LLMs) typically relies on large quantities of high-quality annotated data, or questions with well-defined ground tr

Learning Invariant Modality Representation for Robust Multimodal Learning from a Causal Inference Perspective

TutorialsDGX agent

arXiv:2604.18460v1 Announce Type: new Abstract: Multimodal affective computing aims to predict humans' sentiment, emotion, intention, and opinion using language, acoustic, and visual modalities. Howev

Learning residue level protein dynamics with multiscale Gaussians

ResearchDGX agent

arXiv:2509.01038v2 Announce Type: replace-cross Abstract: Many methods have been developed to predict static protein structures, however understanding the dynamics of protein structure is essential fo

Learning Stable Predictors from Weak Supervision under Distribution Shift

Model ReleasesDGX agent

arXiv:2604.05002v2 Announce Type: replace Abstract: Learning from weak, proxy, or relative supervision is common when ground-truth labels are unavailable, but robustness under distribution shift remai

Learning the Riccati solution operator for time-varying LQR via Deep Operator Networks

ResearchDGX agent

arXiv:2604.18507v1 Announce Type: cross Abstract: We propose a computational framework for replacing the repeated numerical solution of differential Riccati equations in finite-horizon Linear Quadrati

Learning to Correct: Calibrated Reinforcement Learning for Multi-Attempt Chain-of-Thought

ResearchDGX agent

arXiv:2604.17912v1 Announce Type: new Abstract: State-of-the-art reasoning models utilize long chain-of-thought (CoT) to solve increasingly complex problems using more test-time computation. In this w

Learning to Trade Like an Expert: Cognitive Fine-Tuning for Stable Financial Reasoning in Language Models

AgentsDGX agent

arXiv:2604.16862v1 Announce Type: new Abstract: Recent deployments of large language models (LLMs) as autonomous trading agents raise questions about whether financial decision-making competence gener

Learning Unanimously Acceptable Lotteries via Queries

ResearchDGX agent

arXiv:2604.17505v1 Announce Type: cross Abstract: Many high-stakes AI deployments proceed only if every stakeholder deems the system acceptable relative to their own minimum standard. With randomizati

LEPO: nderline{L}atent Rnderline{e}asoning nderline{P}olicy nderline{O}ptimization for Large Language~Models

ResearchDGX agent

arXiv:2604.17892v1 Announce Type: new Abstract: Recently, latent reasoning has been introduced into large language models (LLMs) to leverage rich information within a continuous space. However, withou

Leveraging Kernel Symmetry for Joint Compression and Error Mitigation in Edge Model Transfer

ResearchDGX agent

arXiv:2604.17371v1 Announce Type: cross Abstract: This paper investigates communication-efficient neural network transmission by exploiting structured symmetry constraints in convolutional kernels. In

Light-Adapted Electroretinogram and Oscillatory Potentials (LEOPs) Dataset for Autism Spectrum Disorder and Typically Developing Individuals

ResearchDGX agent

arXiv:2604.16981v1 Announce Type: cross Abstract: The LEOPs (Light-ERG-Oscillatory Potentials) dataset provides light-adapted (LA) electroretinogram (ERG) and Oscillatory Potentials (OPs) waveforms fo

Lightweight Cybersickness Detection based on User-Specific Eye and Head Tracking Data in Virtual Reality

ApplicationsDGX agent

arXiv:2604.17158v1 Announce Type: cross Abstract: The occurrence of cybersickness in virtual reality (VR) significantly impairs users' perception and sense of immersion. Therefore, timely detection of

Live LTL Progress Tracking: Towards Task-Based Exploration

AgentsDGX agent

arXiv:2604.17106v1 Announce Type: new Abstract: Motivated by the challenge presented by non-Markovian objectives in reinforcement learning (RL), we present a novel framework to track and represent the

LiveGraph: Active-Structure Neural Re-ranking for Exercise Recommendation

ApplicationsDGX agent

arXiv:2602.17036v2 Announce Type: replace-cross Abstract: The continuous expansion of digital learning environments has catalyzed the demand for intelligent systems capable of providing personalized e

LLM-AUG: Robust Wireless Data Augmentation with In-Context Learning in Large Language Models

ResearchDGX agent

arXiv:2604.17770v1 Announce Type: new Abstract: Data scarcity remains a fundamental bottleneck in applying deep learning to wireless communication problems, particularly in scenarios where collecting

LLM-Extracted Covariates for Clinical Causal Inference: Rethinking Integration Strategies

SafetyDGX agent

arXiv:2604.16763v1 Announce Type: new Abstract: Causal inference from electronic health records (EHR) is fundamentally limited by unmeasured confounding: critical clinical states such as frailty, goal

LLMs can persuade only psychologically susceptible humans on societal issues, via trust in AI and emotional appeals, amid logical fallacies

ResearchDGX agent

arXiv:2604.16935v1 Announce Type: cross Abstract: Scarce longitudinal evidence examines LLMs' persuasiveness and humanness along time-evolving psychological frameworks. We introduce Talk2AI, a longitu

Local Inconsistency Resolution: The Interplay between Attention and Control in Probabilistic Models

ResearchDGX agent

arXiv:2604.17140v1 Announce Type: cross Abstract: We present a generic algorithm for learning and approximate inference with an intuitive epistemic interpretation: iteratively focus on a subset of the

Local learning for stable backpropagation-free neural network training towards physical learning

ApplicationsDGX agent

arXiv:2603.24790v2 Announce Type: replace Abstract: While backpropagation and automatic differentiation have driven deep learning's success, the physical limits of chip manufacturing and rising enviro

LoRaQ: Optimized Low Rank Approximation for 4-bit Quantization

ResearchDGX agent

arXiv:2604.18117v1 Announce Type: new Abstract: Post-training quantization (PTQ) is essential for deploying large diffusion transformers on resource-constrained hardware, but aggressive 4-bit quantiza

LoReC: Rethinking Large Language Models for Graph Data Analysis

ResearchDGX agent

arXiv:2604.17897v1 Announce Type: new Abstract: The advent of Large Language Models (LLMs) has fundamentally reshaped the way we interact with graphs, giving rise to a new paradigm called GraphLLM. As

Low-rank Orthogonalization for Large-scale Matrix Optimization with Applications to Foundation Model Training

Model ReleasesDGX agent

arXiv:2509.11983v2 Announce Type: replace Abstract: Neural network (NN) training is inherently a large-scale matrix optimization problem, yet the matrix structure of NN parameters has long been overlo

Lower Bounds and Proximally Anchored SGD for Non-Convex Minimization Under Unbounded Variance

ResearchDGX agent

arXiv:2604.16620v1 Announce Type: new Abstract: Analysis of Stochastic Gradient Descent (SGD) and its variants typically relies on the assumption of uniformly bounded variance, a condition that freque

M100: An Orchestrated Dataflow Architecture Powering General AI Computing

Model ReleasesDGX agent

arXiv:2604.17862v1 Announce Type: new Abstract: As deep learning-based AI technologies gain momentum, the demand for general-purpose AI computing architectures continues to grow. While GPGPU-based arc

Machine Learning Based Prediction of Proton Conductivity in Metal-Organic Frameworks

ResearchDGX agent

arXiv:2407.09514v3 Announce Type: replace-cross Abstract: Recently, metal-organic frameworks (MOFs) have demonstrated their potential as solid-state electrolytes in proton exchange membrane fuel cells

Machine Learning Hamiltonian Dynamical Systems with Sparse and Noisy Data

Model ReleasesDGX agent

arXiv:2604.17470v1 Announce Type: new Abstract: Machine learning has become a powerful tool for discovering governing laws of dynamical systems from data. However, most existing approaches degrade sev

MASPO: Unifying Gradient Utilization, Probability Mass, and Signal Reliability for Robust and Sample-Efficient LLM Reasoning

SafetyDGX agent

arXiv:2602.17550v3 Announce Type: replace Abstract: Existing Reinforcement Learning with Verifiable Rewards (RLVR) algorithms, such as GRPO, rely on rigid, uniform, and symmetric trust region mechanis

Matched-Learning-Rate Analysis of Attention Drift and Transfer Retention in Fine-Tuned CLIP

ResearchDGX agent

arXiv:2604.16410v1 Announce Type: new Abstract: CLIP adaptation can improve in-domain accuracy while degrading out-of-domain transfer, but comparisons between Full Fine-Tuning (Full FT) and LoRA are o

MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval

Model ReleasesDGX agent

arXiv:2604.18584v1 Announce Type: cross Abstract: Mathematical problem solving remains a challenging test of reasoning for large language and multimodal models, yet existing benchmarks are limited in

Matlas: A Semantic Search Engine for Mathematics

ResearchDGX agent

arXiv:2604.17484v1 Announce Type: cross Abstract: Retrieving mathematical knowledge is a central task in both human-driven research, such as determining whether a result already exists, finding relate

MerLin: A Discovery Engine for Photonic and Hybrid Quantum Machine Learning

Model ReleasesDGX agent

arXiv:2602.11092v2 Announce Type: replace Abstract: Identifying where quantum models may offer practical benefits in near term quantum machine learning (QML) requires moving beyond isolated algorithmi

MeSH: Memory-as-State-Highways for Recursive Transformers

Model ReleasesDGX agent

arXiv:2510.07739v2 Announce Type: replace Abstract: Recursive transformers reuse parameters and iterate over hidden states multiple times, decoupling compute depth from parameter depth. However, under

Method for Aggregating Unstructured Data Using Large Language Models

Model ReleasesDGX agent

arXiv:2604.16425v1 Announce Type: cross Abstract: This paper presents a method for the automated collection and aggregation of unstructured data from diverse web sources, utilizing Large Language Mode

mlr3torch: A Deep Learning Framework in R based on mlr3 and torch

TutorialsDGX agent

arXiv:2604.18152v1 Announce Type: cross Abstract: Deep learning (DL) has become a cornerstone of modern machine learning (ML) praxis. We introduce the R package mlr3torch, which is an extensible DL fr

← Previous
1…216217218219220…241
Next →