AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
9 Jun 2026

Automating the Expert Eye: A System-Agnostic Deep Learning Framework for Rare Event Discovery in Imbalanced Force Spectroscopy

ResearchDGX agent

arXiv:2606.09541v1 Announce Type: cross Abstract: Single-Molecule Force Spectroscopy (SMFS) provides unprecedented insights into biomolecular mechanics, yet the high-throughput generation of force-ext

AutoMegaKernel: A Statically-Checked Agent Harness for Self-Retargeting Megakernel Synthesis

Model ReleasesDGX agent

arXiv:2606.09682v1 Announce Type: new Abstract: AutoMegaKernel (AMK) compiles a HuggingFace Llama-family model into a single persistent cooperative CUDA kernel that runs the whole forward pass in one

Autonomous Aerial Manipulation via Contextual Contrastive Meta Reinforcement Learning

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.08533v1 Announce Type: new Abstract: Unmanned aerial vehicles (UAVs) are increasingly being deployed in logistics, service robotics, and other real-world applications, creating a growing de

Backward Coherence and Hidden-State Stability in Recurrent Neural Networks: A Quasi-Reverse-Martingale Theory

ResearchDGX agent

arXiv:2606.08934v1 Announce Type: new Abstract: Recurrent neural networks maintain a hidden state h_t, but its probabilistic meaning is often unclear. We study hidden-state stability through backward

Barycentric Projections of Optimal Transport Plans on Riemannian Manifolds

ResearchDGX agent

arXiv:2606.07926v1 Announce Type: cross Abstract: Optimal transport couplings are probabilistic objects, while many learning pipelines require deterministic maps. In Euclidean space, barycentric proje

Bayesian Optimization of a Multi-Product Chemical Reactor Using Composite Models and Partial Physics Knowledge

Model ReleasesDGX agent

arXiv:2606.08611v1 Announce Type: cross Abstract: We study data-driven real-time economic optimization of a multi-product chemical reactor when no reliable first-principles model is available beyond a

Benchmark Datasets for Lead-Lag Forecasting on Social Platforms

Model ReleasesDGX agent

arXiv:2511.03877v2 Announce Type: replace Abstract: Social and collaborative platforms emit multivariate time-series traces in which early interactions -- such as views, likes, or downloads -- are fol

Benchmarking Empirical Privacy Protection for Adaptations of Large Language Models

Model ReleasesDGX agent

arXiv:2606.09401v1 Announce Type: new Abstract: Recent work has applied differential privacy (DP) to adapt large language models (LLMs) for sensitive applications, offering theoretical guarantees. How

Beyond Convolution: Advancing Hypergraph Neural Networks with Hypergraph U-Nets

Local AiDGX agent

arXiv:2606.09051v1 Announce Type: new Abstract: Convolutions have successfully transitioned from image processing to the complex realm of non-Euclidean higher-order domains, particularly in hypergraph

Beyond Fixed Rounds: Data-Free Early Stopping for Practical Federated Learning

ResearchDGX agent

arXiv:2601.22669v3 Announce Type: replace Abstract: Federated Learning (FL) facilitates decentralized collaborative learning without transmitting raw data. However, reliance on fixed global rounds or

Beyond FLOPs: Benchmarking Real Inference Acceleration of LLM Pruning under a GEMM-Centric Taxonomy

ResearchDGX agent

arXiv:2606.09080v1 Announce Type: new Abstract: Pruning has emerged as a dominant paradigm for accelerating large language model (LLM) inference, spanning a broad spectrum of methods that remove compu

Beyond Homophily: Towards Generalized Graph Reconstruction Attack and Defense

SafetyDGX agent

arXiv:2606.08067v1 Announce Type: new Abstract: Graph neural networks (GNNs) are widely deployed on relational data, yet they can leak sensitive or proprietary information about the training graph adj

Beyond Linear Activation Steering: Invertible Latent Transformations for Controlling LLM Behavior

SafetyDGX agent

arXiv:2606.08454v1 Announce Type: new Abstract: Activation steering provides a lightweight inference-time mechanism for controlling large language models (LLMs) by modifying their internal activation

Beyond Neural Collapse: Task-Intrinsic Geometry Governs Neural Representations in Modular Arithmetic

SafetyDGX agent

arXiv:2606.08985v1 Announce Type: new Abstract: While neural collapse (NC) predicts that a K-class-balanced classifier should organize terminal representations as a (K-1)-dimensional simplex equiangul

Biological Reasoning-Informed Regression for Interpretable Regulatory DNA Activity Prediction

ResearchDGX agent

arXiv:2606.08147v1 Announce Type: cross Abstract: DNA cis-regulatory elements (CREs) such as enhancers control gene expression levels. Accurately predicting regulatory activity from DNA sequences is v

BlendServe: Optimizing Offline Inference for Auto-regressive Large Models with Resource-aware Batching

ResearchDGX agent

arXiv:2411.16102v2 Announce Type: replace Abstract: Offline batch inference, which leverages the flexibility of request batching to achieve higher throughput and lower costs, is becoming more popular

Boundary Variance Inflation Causes Acquisition Bias in Gaussian Processes

SafetyDGX agent

arXiv:2606.07561v1 Announce Type: new Abstract: Gaussian processes with stationary kernels on bounded domains exhibit inflated posterior variance near the boundary. Despite being a long-recognized art

BrainSurgery: Reproducible and Reliable Declarative Weight Manipulations for Model Editing and Upcycling

ResearchDGX agent

arXiv:2606.09707v1 Announce Type: new Abstract: As deep learning models scale, managing, inspecting, and modifying large checkpoints has become increasingly challenging. Researchers often need to alte

Breaking the Bubble: Asynchronous Pipeline Parallel Training with Bounded Weight Inconsistency

Model ReleasesDGX agent

arXiv:2606.07881v1 Announce Type: new Abstract: Pipeline parallelism is essential for training large neural networks, but existing schedules trade off throughput, memory, and optimization consistency.

Breaking the Tokenizer Barrier: On-Policy Distillation across Model Families

SafetyDGX agent

arXiv:2606.09456v1 Announce Type: new Abstract: On-Policy Distillation (OPD) has become a core technique in the post-training of Large Language Models (LLMs) for transferring knowledge from domain exp

BUDDY: BUdget-Driven DYnamic Depth Routing for Adaptive Large Language Model Inference

Model ReleasesDGX agent

arXiv:2606.09514v1 Announce Type: new Abstract: Large language models (LLMs) incur high inference cost due to their depth and parameter scale. Depth pruning can reduce latency by skipping redundant Tr

Bulk-boundary decomposition of neural networks

ResearchDGX agent

arXiv:2511.02003v2 Announce Type: replace Abstract: We present the bulk--boundary decomposition as a new framework for understanding the training dynamics of deep neural networks. Starting from the st

Byzantine Cheap Talk: Adversarial Resilience and Topology Effects in LLM Coordination Games

AgentsDGX agent

arXiv:2606.07790v1 Announce Type: new Abstract: Multi-agent LLM systems increasingly rely on communication protocols for coordination, yet their robustness under adversarial and structural constraints

CAAL: Contextual Bandits based Online Hand-Craft Active Learning Strategy Selection

ResearchDGX agent

arXiv:2606.07910v1 Announce Type: new Abstract: The challenge with active learning algorithms is the uncertainty of the statistical distribution of unlabeled data, making it difficult to choose the be

Can LLMs extract scientific consensus? A case study in high-temperature superconductivity

ApplicationsDGX agent

arXiv:2606.07570v1 Announce Type: cross Abstract: Scientific knowledge is increasingly dispersed across vast and heterogeneous scientific literature, where important claims are often implicit, evolvin

CATPO: Critique-Augmented Tree Policy Optimization

Model ReleasesDGX agent

arXiv:2606.08346v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a dominant paradigm for improving the reasoning capabilities of large language models

Causal Longitudinal Prior-Fitted Networks for Counterfactual Outcome Prediction

Model ReleasesDGX agent

arXiv:2606.05797v2 Announce Type: replace Abstract: Longitudinal treatment decisions from multivariate time-series data require predicting potential outcomes under future treatment sequences in the pr

Causal Representation Learning from Network Data

ResearchDGX agent

arXiv:2509.01916v2 Announce Type: replace Abstract: Causal disentanglement from soft interventions is identifiable under the assumptions of linear interventional faithfulness and availability of both

Causal Semantic Alignment for LLM-based Time Series Forecasting

SafetyDGX agent

arXiv:2606.08262v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have opened new possibilities for time series forecasting by enabling alignment between temporal pattern

Characterizing the Discrete Geometry of ReLU Networks

ApplicationsDGX agent

arXiv:2606.07728v1 Announce Type: new Abstract: It is well established that ReLU networks define continuous piecewise-linear functions, and that their linear regions are polyhedra in the input space.

Claw-R1: A Step-Level Data Middleware System for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.09138v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has become an important post-training paradigm for turning LLMs from static chatbots into interactive agents, giving

Code Is More Than Text: Uncertainty Estimation for Code Generation

SafetyDGX agent

arXiv:2606.09577v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as code generators, where silently wrong programs pose real safety and reliability risks. Relia

Communication-Efficient Federated Learning under Dynamic Device Arrival and Departure: Convergence Analysis and Algorithm Design

Local AiDGX agent

arXiv:2410.05662v4 Announce Type: replace Abstract: Most federated learning (FL) approaches assume a fixed device set. However, real-world scenarios often involve devices dynamically joining or leavin

Community-Specific Slang and Entity Detection via Semantic Shift in Fine-Tuned Language Models

ResearchDGX agent

arXiv:2606.07522v1 Announce Type: cross Abstract: We propose an unsupervised method of resolving slang, unique entities, and folklore from online communities by isolating words in the lexicon that hav

Compositional Approximation Can Strictly Outperform Superpositional Approximation

ResearchDGX agent

arXiv:2606.08727v1 Announce Type: cross Abstract: Many classically studied function classes are known to be approximated optimally by superpositional methods, i.e. with approximants constructed as the

Conditional Normalizing Flows for Forward and Backward Joint State and Parameter Estimation

Model ReleasesDGX agent

arXiv:2601.07013v2 Announce Type: replace-cross Abstract: Traditional filtering algorithms for state estimation -- such as classical Kalman filtering, unscented Kalman filtering, and particle filters

Conditional Random Ordered Transport Spaces

ResearchDGX agent

arXiv:2606.08113v1 Announce Type: new Abstract: A small Wasserstein distance does not certify that a transformation is admissible. In evidence-constrained, semantic, causal, physical, monotone, or ris

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning

SafetyDGX agent

arXiv:2606.08088v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has recently become a key paradigm for improving the reasoning abilities of Large Language Models

Constrained user-item allocation for e-commerce marketing campaigns

SafetyDGX agent

arXiv:2606.09623v1 Announce Type: new Abstract: When running marketing campaigns, retailers must decide which products to promote and which users to target. These decisions are inherently coupled: eff

Constraint-Aware Optimization for Robust Protein Stability Prediction

SafetyDGX agent

arXiv:2606.08100v1 Announce Type: new Abstract: Multimodal DeltaDelta G predictors integrating protein language models with inverse-folding representations achieve strong in-distribution accuracy on t

Continuous Language Diffusion as a Decoder-Interface Problem

Local AiDGX agent

arXiv:2606.08810v1 Announce Type: cross Abstract: Gaussian-corrupted sentence embeddings have no direct linguistic interpretation, yet continuous diffusion language models can generate fluent text fro

Contrast encodes inductive bias: separating slow noise from dynamics in predictive representation learning

SafetyDGX agent

arXiv:2606.07770v1 Announce Type: new Abstract: Self-supervised methods that learn representations and predict dynamics fully in the latent space, such as JEPA, have been shown to confuse slowly varyi

Convergence Bound and Critical Batch Size of Muon Optimizer

Model ReleasesDGX agent

arXiv:2507.01598v5 Announce Type: replace Abstract: Muon, a recently proposed optimizer that leverages the inherent matrix structure of neural network parameters, has demonstrated strong empirical per

Convolutional Sparse Coding via the Locally Competitive Algorithm on Loihi 2

Model ReleasesDGX agent

arXiv:2606.08584v1 Announce Type: new Abstract: Sparse coding provides a principled framework for signal representation by expressing an input as a linear combination of only a small number of basis f

Counterfactual Transport Flows for Offline Conservative Trajectory Refinement

Model ReleasesDGX agent

arXiv:2606.09115v1 Announce Type: new Abstract: Offline reinforcement learning (RL) offers a path to policy improvement from logged data alone, using historical returns or other measurable outcomes as

Cryptographic Backdoor for Neural Networks: Boon and Bane

ResearchDGX agent

arXiv:2509.20714v2 Announce Type: replace-cross Abstract: In this paper we show that cryptographic backdoors in a neural network (NN) can be highly effective in two directions, namely mounting the att

CTS-Bench: Benchmarking Graph Coarsening Trade-offs for GNNs in Clock Tree Synthesis

Model ReleasesDGX agent

arXiv:2602.19330v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) are increasingly explored for physical design analysis in Electronic Design Automation, particularly for modeling Clock

Curvature-Guided LoRA: Matching Full Fine-Tuning in Function Space

Model ReleasesDGX agent

arXiv:2603.29824v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning methods such as LoRA enable efficient adaptation of large pretrained models, but often lag behind full fine-tuning i

Cutting LLM Evaluation Costs with SySRs: A Bandit Algorithm that Provably Exploits Model Similarity

ApplicationsDGX agent

arXiv:2606.07726v1 Announce Type: new Abstract: Large Language Models are typically benchmarked by evaluating every model on every test query. For practitioners seeking the best model to deploy, this

Data augmented bootstrap: Unifying confidence interval construction by approximate invariance

ResearchDGX agent

arXiv:2606.09049v1 Announce Type: cross Abstract: We propose the data augmented bootstrap (DAB), a framework for constructing confidence intervals from approximately invariant transformations of the d

Data-driven discovery of governing differential equations across physical systems

ResearchDGX agent

arXiv:2606.09638v1 Announce Type: new Abstract: Differential equations play a critical role in scientific discovery because they provide a mathematical framework to describe the behaviour of physical

De novo molecular generation with optical property preconditioning at the token level

Model ReleasesDGX agent

arXiv:2606.08221v1 Announce Type: new Abstract: Designing OLED molecules with targeted optical properties remains challenging due to the scarcity of high-quality data and the limited reliability of co

Decentralized Online Riemannian Optimization Beyond Hadamard Manifolds

ResearchDGX agent

arXiv:2509.07779v2 Announce Type: replace-cross Abstract: We study decentralized online Riemannian optimization over manifolds with possibly positive curvature, going beyond the Hadamard manifold sett

Decision-Focused Continual Learning for Seaport Power-Logistics Scheduling: Generalization across Varying Tasks

ResearchDGX agent

arXiv:2511.07938v3 Announce Type: replace Abstract: Power-logistics scheduling in modern seaports typically follows a predict-then-optimize pipeline. To enhance the decision quality of predictions, de

Declarative Outcome-Conformant Synthesis: Exact, Closed-Form Specification Satisfaction and a Conformance Benchmark

Model ReleasesDGX agent

arXiv:2606.08736v1 Announce Type: new Abstract: We study a capability the dominant paradigm in synthetic tabular data does not provide: exact satisfaction of a declared analytical outcome with no sour

Decoding Naturalistic Emotion Dynamics from the Brain: An LLM-Enhanced Regression Framework

ResearchDGX agent

arXiv:2606.07707v1 Announce Type: new Abstract: Decoding emotional states from neural signals has been typically framed as a discrete, single-label classification task based on emotionally stable stim

Decomposable Neuro Symbolic Regression

Model ReleasesDGX agent

arXiv:2511.04124v3 Announce Type: replace Abstract: Symbolic regression (SR) models complex systems by discovering mathematical expressions that capture underlying relationships in observed data. Howe

Decoy-Calibrated Failure Audits for Language Models

ResearchDGX agent

arXiv:2606.09046v1 Announce Type: new Abstract: Useful audits reveal not only how often a model fails, but also where its failures concentrate. An auditor may test many candidate explanations: long in

Deep reinforcement learning for process design: Review and perspective

AgentsDGX agent

arXiv:2308.07822v2 Announce Type: replace Abstract: The transformation towards renewable energy and feedstock supply in the chemical industry requires new conceptual process design approaches. Recentl

Dendrograms of Mixing Measures for Softmax-Gated Gaussian Mixture of Experts: Consistency Without Model Sweeps

Model ReleasesDGX agent

arXiv:2510.12744v2 Announce Type: replace-cross Abstract: We develop a unified statistical framework for softmax-gated Gaussian mixture of experts (SGMoE) that addresses three long-standing obstacles

← Previous
1…9091929394…243
Next →