AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
9 Jun 2026

SAD-Flower: Flow Matching for Safe, Admissible, and Dynamically Consistent Planning

SafetyDGX agent

arXiv:2511.05355v3 Announce Type: replace Abstract: Flow matching (FM) has shown promising results in data-driven planning. However, it inherently lacks formal guarantees for ensuring state and action

SAEExplainer: Interpreting SAE Features with Activation-Guided Preference Optimization

ResearchDGX agent

arXiv:2606.08496v1 Announce Type: cross Abstract: Although Sparse Autoencoders (SAEs) have mitigated the opacity of large language models (LLMs) by decomposing dense representations into sparse featur

SC3: The Multi-Solvent Solubility Challenge and Benchmark

Model ReleasesDGX agent

arXiv:2606.07656v1 Announce Type: cross Abstract: Solubility prediction is a standard benchmark in computational chemistry, yet multi-solvent models which reportedly approach the experimental-noise ce


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Scaling Laws for Masked-Reconstruction Transformers on Single-Cell Transcriptomics

Model ReleasesDGX agent

arXiv:2602.15253v2 Announce Type: replace Abstract: Neural scaling laws -- power-law relationships between loss, model size, and data -- have been extensively documented for language and vision transf

scCBGM: Interpretable Single-Cell Counterfactual Editing

Model ReleasesDGX agent

arXiv:2606.07760v1 Announce Type: new Abstract: Understanding cellular phenotypes and how they respond to perturbations is critical for disease biology and therapeutic design. Single-cell RNA sequenci

Self-Consistent Generative Paths via Admissible Random Variational Transport

Local AiDGX agent

arXiv:2606.08953v1 Announce Type: new Abstract: Modern generative models often define an entire probability path from a simple prior to the data law, rather than only an endpoint map. Diffusion models

Self-Supervised Dynamical System Representations for Physiological Time-Series

ApplicationsDGX agent

arXiv:2512.00239v2 Announce Type: replace Abstract: The effectiveness of self-supervised learning (SSL) for physiological time series depends on the ability of a pretraining objective to preserve info

Sequential statistical inference for Large Language Models: Representation, validity, and monitoring

SafetyDGX agent

arXiv:2606.07624v1 Announce Type: new Abstract: This discussion argues that sequential statistical inference can naturally contribute to LLM trustworthiness. In deployment, LLM systems are queried rep

SFILES 2.0: An extended text-based flowsheet representation

ResearchDGX agent

arXiv:2208.00778v2 Announce Type: replace-cross Abstract: SFILES are a text-based notation for chemical process flowsheets. They were originally proposed by d'Anterroches (Process flow sheet generatio

SG-OPD: Sign-Gated On-Policy Distillation via Sign-Consistency Gating and Phased Teacher Sampling

SafetyDGX agent

arXiv:2606.09304v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories with dense per-token supervision from a stronger teacher, and often outperforms

Shared Semantics, Divergent Mechanisms: Unsupervised Feature Discovery by Aligning Semantics and Mechanisms

ResearchDGX agent

arXiv:2606.08236v1 Announce Type: cross Abstract: As large language models are increasingly deployed in high-stakes settings, there is a growing need for tools that audit not only model outputs but al

Shortcuts in the Tail: Debiasing via Post-Hoc Spectral Compression of Fine-Tuning Updates

TutorialsDGX agent

arXiv:2606.07596v1 Announce Type: new Abstract: Fine-tuning often introduces spurious correlations alongside task knowledge, causing systematic failures on underrepresented groups. Existing mitigation

Should Demand Models Incorporate Competitor Prices? Oblivious Learning and Algorithmic Collusion

ResearchDGX agent

arXiv:2606.05363v2 Announce Type: replace-cross Abstract: On a platform with many sellers, should a pricing algorithm explicitly model competitors' prices when learning demand? Classical learning argu

SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History

AgentsDGX agent

arXiv:2606.08671v1 Announce Type: new Abstract: Agent skills extend language-model agents with task-specific procedures, scripts, and references, but the tasks and environments they target continually

SNN-MLIR: An MLIR Dialect for Compiling Neuromorphic SNNs from NIR to Bare-Metal C

Model ReleasesDGX agent

arXiv:2606.09213v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) are increasingly trained in a wide range of frameworks (SnnTorch, Lava, Norse, and others) each with its own model form

SoK: Reconstruction Attacks on Synthetic Tabular Data (Insights from Winning the NIST CRC)

Model ReleasesDGX agent

arXiv:2606.08372v1 Announce Type: cross Abstract: Synthetic data is increasingly promoted as a privacy-preserving substitute for releasing sensitive tabular records, yet its central adversarial threat

Solving Inverse Problems with Flow-based Models via Model Predictive Control

Model ReleasesDGX agent

arXiv:2601.23231v2 Announce Type: replace-cross Abstract: Flow-based generative models provide strong unconditional priors for inverse problems, but guiding their dynamics for conditional generation r

Speaker-Invariant Representation Learning for Spoofing Detection via Gradient Reversal and A Variational Information Bottleneck

SafetyDGX agent

arXiv:2606.08678v1 Announce Type: cross Abstract: Sophisticated generative speech technology can undermined the reliability of voice biometrics. While spoofing detection systems excel when assessed un

Spectral Truncation Kernels: Noncommutativity in C^*-algebraic Kernel Machines

Local AiDGX agent

arXiv:2405.17823v5 Announce Type: replace-cross Abstract: A central question in vector- and function-valued learning is how to design kernels that capture both local and non-local interactions while r

SpectraLDS: Provable Distillation for Linear Dynamical Systems

TutorialsDGX agent

arXiv:2505.17868v2 Announce Type: replace Abstract: We present the first provable method for identifying symmetric linear dynamical systems (LDS) with accuracy guarantees that are independent of the s

SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving

Model ReleasesDGX agent

arXiv:2606.08635v1 Announce Type: new Abstract: Prefill-decode (PD) disaggregation decouples prompt processing from token generation, but it also turns the key-value (KV) cache into a network payload.

SPIN: Decentralized Swarm Control via Tensorized Policy Coordination

SafetyDGX agent

arXiv:2606.07557v1 Announce Type: new Abstract: Decentralized multi-agent swarm coordination on resource-constrained edge platforms remains fundamentally bottlenecked by the exponential scaling of joi

Stable and Scalable Probabilistic Numerical Solvers for Stiff and High-Dimensional ODEs

ResearchDGX agent

arXiv:2606.08203v1 Announce Type: cross Abstract: Filtering-based probabilistic numerical solvers for ordinary differential equations (ODEs) have been established as a flexible and efficient simulatio

STARIXNet: Multivariate and Multi-attribute Deep Learning Approach to Real-Time Resource Allocation in Cloud Platforms

SafetyDGX agent

arXiv:2606.07565v1 Announce Type: new Abstract: Intelligent scaling of microservices in cloud platforms is crucial for mitigating escalating compute costs while avoiding service disruptions. Current s

State Backdoor: Towards Stealthy Real-world Poisoning Attack on Vision-Language-Action Model in State Space

SafetyDGX agent

arXiv:2601.04266v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models are widely deployed in safety-critical embodied AI applications such as robotics. However, their complex m

Statistical Decision Theory with Counterfactual Loss

TutorialsDGX agent

arXiv:2505.08908v3 Announce Type: replace-cross Abstract: Many researchers apply classical statistical decision theory to evaluate treatment choices and learn optimal policies. However, because this f

Still: Amortized KV Cache Compaction in a Single Forward Pass

Model ReleasesDGX agent

arXiv:2606.07878v1 Announce Type: new Abstract: The KV cache is the memory bottleneck of long-horizon language model deployment. Practically, a deployable compactor must be lightweight enough to call

Stochastic Dimension Implicit Functional Projections for Global Integral Conservation in High-Dimensional PINNs

ResearchDGX agent

arXiv:2603.29237v2 Announce Type: replace Abstract: Enforcing prescribed global integral constraints in mesh-free neural PDE solvers is challenging in high-dimensional domains. Existing projection met

Structural Decoupling: A Scaffold-Flow Theory of Generalization and Alignment

Local AiDGX agent

arXiv:2506.20699v2 Announce Type: replace Abstract: Learning in non-stationary and multi-context environments requires more than ordinary within-task generalization. A system must also discover which

Structural Grid Descriptors Predict Within-Task Solver Success on ARC-AGI

ResearchDGX agent

arXiv:2606.09026v1 Announce Type: new Abstract: We ask whether structural properties of intermediate grid states predict whether a symbolic ARC-AGI solver will succeed, framed as a test of conditional

Structure-Aware Modeling of Multiple-Choice Questions Improves Automatic Difficulty Estimation

ResearchDGX agent

arXiv:2606.08988v1 Announce Type: cross Abstract: Automatic Question Difficulty Estimation (AQDE) holds growing promise for educational assessment because it has the potential to yield difficulty esti

Synthetic but Not Realistic: The Evaluation Challenge in Generative Modelling for Structured Electronic Medical Records

ApplicationsDGX agent

arXiv:2606.08903v1 Announce Type: new Abstract: Synthetic healthcare data are widely proposed as privacy-preserving substitutes for real patient data, yet their evaluation remains dominated by statist

TAMUNA: Doubly Accelerated Distributed Optimization under Partial Participation

Local AiDGX agent

arXiv:2302.09832v4 Announce Type: replace Abstract: In distributed optimization and federated learning, slow and costly communication between parallel devices and the central server constitutes the pr

Teacher-Free Self-Training Amplifies but Does Not Compound: A Pass@K Crossover on a Free-Verifier Domain

Local AiDGX agent

arXiv:2606.07856v1 Announce Type: new Abstract: When a language model trains on its own verified outputs, does it acquire capability beyond its base, or merely get better at expressing capability the

Temporal Coverage over Density: Parsimonious Training-Set Design for ML Climate Downscaling

ResearchDGX agent

arXiv:2606.07898v1 Announce Type: new Abstract: High-resolution regional climate simulations provide critical information for climate impacts assessments but remain computationally expensive, motivati

Tensorizing Engram: Sharing Latents Across N-Gram Embeddings is Beneficial in LLMs

ResearchDGX agent

arXiv:2606.08347v1 Announce Type: cross Abstract: Modern language models represent text using discrete token-level embeddings, which forces recurring multi-token patterns to be learned implicitly acro

The Easy, the Hard, and the Learnable: Confidence and Difficulty-Adaptive Policy Optimization for LLM Reasoning

SafetyDGX agent

arXiv:2606.07950v1 Announce Type: new Abstract: RL with verifiable rewards can substantially improve LLM reasoning, yet standard GRPO-style training often treats easy, hard, and learnable questions al

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning

SafetyDGX agent

arXiv:2606.09078v1 Announce Type: new Abstract: Process Reward Models (PRMs) improve credit assignment for reasoning by providing step-level feedback. However, we identify a hidden bias in PRMs caused

The Injection Paradox: Brand-Level Suppression in Safety-Trained LLM Recommendations via RAG Context Injection

Model ReleasesDGX agent

arXiv:2606.09204v1 Announce Type: new Abstract: We present a reproducible failure mode of safety training in RAG-based LLM recommendation -- the Injection Paradox -- in which prompt injections embedde

The Label Horizon Paradox: Rethinking Supervision Targets in Financial Forecasting

ResearchDGX agent

arXiv:2602.03395v4 Announce Type: replace Abstract: While deep learning has revolutionized financial forecasting through sophisticated architectures, the design of the supervision signal itself is rar

The Mirrored Influence Hypothesis: Efficient Data Influence Estimation by Harnessing Forward Passes

ResearchDGX agent

arXiv:2402.08922v3 Announce Type: replace Abstract: Large-scale black-box models have become ubiquitous across numerous applications. Understanding the influence of individual training data sources on

The Routing Plateau: Understanding and Breaking the Accuracy Limits of LLM Routers

TutorialsDGX agent

arXiv:2606.07587v1 Announce Type: new Abstract: LLM routing has become a popular approach to improve the cost-quality trade-off of LLM services by dynamically selecting a model for each query. Recent

The Sample Complexity of Parameter-Free Stochastic Convex Optimization

Model ReleasesDGX agent

arXiv:2506.11336v2 Announce Type: replace Abstract: We study the sample complexity of stochastic convex optimization when problem parameters such as the distance to optimality and the Lipschitz consta

The Spectral Dynamics and Noise Geometry of Muon

SafetyDGX agent

arXiv:2606.08388v1 Announce Type: new Abstract: Muon replaces a matrix gradient G=USigma V^op by its polar factor UV^op. This keeps the singular directions selected by the gradient, but makes the upda

The Value of Personalized Recommendations: Evidence from Netflix

ResearchDGX agent

arXiv:2511.07280v5 Announce Type: replace-cross Abstract: Personalized recommendation systems shape much of user choice online, yet their targeted nature makes separating out the value of recommendati

Theoretical Foundations of Continual Learning via Drift-Plus-Penalty

Model ReleasesDGX agent

arXiv:2606.08452v1 Announce Type: new Abstract: In many real-world settings, data streams are nonstationary and arrive sequentially, requiring learning systems to adapt continuously without retraining

Thresholded Local Hyper-Flow Diffusion

ResearchDGX agent

arXiv:2606.09340v1 Announce Type: new Abstract: Local Hyper-Flow Diffusion (HFD) gives an edge-size-independent Cheeger-type guarantee for seeded clustering in general submodular hypergraphs, but exis

Tight Sample Complexity of Transformers

ResearchDGX agent

arXiv:2606.09731v1 Announce Type: new Abstract: We tightly characterize the VC dimension of depth-L Transformers with a total of W parameters, mapping an input sequence of length T to a single output,

TinyJudge: Unverifiable Constraint Alignment via Lightweight Specialist Ensembles

SafetyDGX agent

arXiv:2606.07520v1 Announce Type: cross Abstract: Instruction Following (IF) is a core capability of LLMs, requiring strict adherence to diverse constraints, ranging from verifiable ones (e.g., output

Titans-as-a-Layer: Test-Time Memory for Conversational Speech Emotion Recognition

ResearchDGX agent

arXiv:2606.08573v1 Announce Type: new Abstract: Speech emotion recognition (SER) is commonly formulated as utterance-level classification, although conversational emotion depends on a speaker's usual

Token Sample Complexity of Attention

Model ReleasesDGX agent

arXiv:2512.10656v3 Announce Type: replace Abstract: As context windows in large language models continue to expand, it is essential to characterize how attention behaves at extreme sequence lengths. W

Toward Compiler World Models: Learning Latent Dynamics for Efficient Tensor Program Search

HardwareDGX agent

arXiv:2606.09312v1 Announce Type: new Abstract: Tensor program optimization is essential for modern machine learning systems, but its search space is enormous. Existing auto-schedulers reduce measurem

Towards Automated Kernel Generation in the Era of LLMs

HardwareDGX agent

arXiv:2601.15727v3 Announce Type: replace Abstract: The performance of modern AI systems is fundamentally constrained by the quality of their underlying GPU kernels, which translate high-level algorit

Towards End to End Motion Planning and Execution for Autonomous Underwater Vehicles Using Reinforcement Learning

SafetyDGX agent

arXiv:2606.08513v1 Announce Type: cross Abstract: Autonomous Underwater Vehicles (AUVs) traditionally rely on complex, heavily engineered pipelines for perception, path planning, and motion control. T

Towards Graph Foundation Models for Dynamics in Complex Networked Systems: Lessons from Super-Spreader Identification in Multilayer Networks

ApplicationsDGX agent

arXiv:2606.08306v1 Announce Type: new Abstract: Network dynamics - including spreading, influence maximisation, and epidemic modelling - remain largely confined to the transductive paradigm, where mod

Towards Personalized Bangla Book Recommendation: A Large-Scale Heterogeneous Book Graph Dataset

Model ReleasesDGX agent

arXiv:2602.12129v2 Announce Type: replace-cross Abstract: Personalized book recommendation in Bangla literature has been constrained by the lack of structured, large-scale, and publicly available data

Trajectory Geometry of Transformer Representations Across Layers

ResearchDGX agent

arXiv:2606.09287v1 Announce Type: new Abstract: Understanding how transformer representations evolve across layers, not merely what they encode, remains an open problem in mechanistic interpretability

Transfer learning for causal forest

ApplicationsDGX agent

arXiv:2606.07693v1 Announce Type: cross Abstract: Transfer learning addresses the challenge of transfering knowledge from one domain to another. Traditional transfer learning focuses on adapting model

TriHead-GAN: A Generative Adversarial Network with Triple-Head Discriminator for Carbon Emission Time Series Generation

Model ReleasesDGX agent

arXiv:2606.07569v1 Announce Type: new Abstract: Accurate carbon emission monitoring is critical for climate policy and emerging regulatory mechanisms such as the EU Carbon Border Adjustment Mechanism,

TRUST-SCF: Transformer-based Risk Understanding and Scoring for Transactional Supply Chain Finance

SafetyDGX agent

arXiv:2606.08140v1 Announce Type: new Abstract: Supply Chain Finance (SCF) and LendTech platforms need credit scoring systems that respond to evolving transaction behavior, repayment delays, and activ

← Previous
1…9495969798…243
Next →