AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
5 May 2026

Minimum Specification Perturbation: Robustness as Distance-to-Falsification in Causal Inference

Model ReleasesDGX agent

arXiv:2605.01579v1 Announce Type: cross Abstract: Empirical causal claims depend on many analyst decisions, from selecting covariates to choosing estimators. Existing robustness tools summarize how re

MIRA: A Score for Conditional Distribution Accuracy and Model Comparison

SafetyDGX agent

arXiv:2605.02014v1 Announce Type: cross Abstract: We introduce Mira, a sample-based score for assessing the accuracy of a candidate conditional distribution using only joint samples from the true data

Misclassification Rate and Privacy-Utility Trade-offs in Graph Convolutional Networks via Subsampling Stability

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.01987v1 Announce Type: new Abstract: We study differential privacy (DP) in Graph Convolutional Networks (GCNs) through the framework of extit{subsampling stability}. We derive upper bounds

Missingness-aware Data Imputation via AI-powered Bayesian Generative Modeling

ResearchDGX agent

arXiv:2605.01676v1 Announce Type: cross Abstract: Missing data imputation remains a fundamental challenge in modern data science, especially when uncertainty quantification is essential. In this work,

Model-Based Proactive Cost Generation for Learning Safe Policies Offline with Limited Violation Data

SafetyDGX agent

arXiv:2605.01356v1 Announce Type: new Abstract: Learning constraint-satisfying policies from offline data without risky online interaction is crucial for safety-critical decision making. Conventional

Model Merging: Foundations and Algorithms

Model ReleasesDGX agent

arXiv:2605.01580v1 Announce Type: new Abstract: Modern deep learning usually treats models as separate artifacts: trained independently, specialized for particular purposes, and replaced when improved

Molecular Representations for Large Language Models

Model ReleasesDGX agent

arXiv:2605.01822v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly being used to support scientific discovery. In chemistry, tasks such as reaction prediction and structure

MPCS: Neuroplastic Continual Learning via Multi-Component Plasticity and Topology-Aware EWC

Model ReleasesDGX agent

arXiv:2605.02509v1 Announce Type: new Abstract: Continual learning systems face a fundamental tension between plasticity -- acquiring new knowledge -- and stability -- retaining prior knowledge. We in

MSMixer: Learned Multi-Scale Temporal Mixing with Complementary Linear Shortcut for Long-Term Time Series Forecasting

ResearchDGX agent

arXiv:2605.02689v1 Announce Type: new Abstract: Long-term time series forecasting requires models that simultaneously capture rapid oscillations, medium-range periodicities, and slowly evolving macro-

MU-SHOT-Fi: Self-Supervised Multi-User Wi-Fi Sensing with Source-free Unsupervised Domain Adaptation

TutorialsDGX agent

arXiv:2605.01369v1 Announce Type: cross Abstract: Deep learning has been widely adopted for WiFi CSI-based human activity recognition (HAR) due to its ability to learn spatio-temporal features in a pr

Multi-fidelity surrogates for mechanics of composites: from co-kriging to multi-fidelity neural networks

Model ReleasesDGX agent

arXiv:2605.02871v1 Announce Type: cross Abstract: Composite materials exhibit strongly hierarchical and anisotropic properties governed by coupled mechanisms spanning constituents, plies, laminates, s

Multi-Objective Instruction-Aware Representation Learning in Procedural Content Generation RL

ResearchDGX agent

arXiv:2508.09193v2 Announce Type: replace Abstract: Recent advancements in generative modeling emphasize the importance of natural language as a highly expressive and accessible modality for controlli

Multi-Perspective Transformers in ARC-AGI-2 Challenge

Model ReleasesDGX agent

arXiv:2605.01154v1 Announce Type: new Abstract: ARC-AGI-2 is a benchmark of human-intuitive visual puzzles that measures a machine's ability to generalize from limited examples, interpret symbolic mea

Multi-User Dueling Bandits: A Fair Approach using Nash Social Welfare

SafetyDGX agent

arXiv:2605.01961v1 Announce Type: new Abstract: Learning from human preference data is becoming a useful tool, from fine-tuning large language models to training reinforcement learning agents. However

Multimodal Data Curation Through Ranked Retrieval

SafetyDGX agent

arXiv:2605.01163v1 Announce Type: cross Abstract: Shared embedding spaces are widely used for multimodal search and data curation. In practice, two problems often limit how well this works. First, emb

NAPS: Attention-Based Fusion of Heterogeneous Physiological Signals

ApplicationsDGX agent

arXiv:2511.03488v2 Announce Type: replace Abstract: Physiological signals are inherently heterogeneous: they are collected under diverse acquisition setups, differ in the number and type of modalities

NaviMaster: Learning a Unified Policy for GUI and Embodied Navigation Tasks

SafetyDGX agent

arXiv:2508.02046v4 Announce Type: replace-cross Abstract: Recent advances in Graphical User Interface (GUI) and embodied navigation have driven progress, yet these domains have largely evolved in isol

Near-Optimal Privacy-Preserving Learning for Max-Min Fair Multi-Agent Bandits

AgentsDGX agent

arXiv:2306.04498v3 Announce Type: replace Abstract: We study fair multi-agent multi-armed bandit learning under collision-only coordination. Agents cannot communicate explicitly during learning and ob

Nearly-Optimal Bandit Learning in Stackelberg Games with Side Information

ResearchDGX agent

arXiv:2502.00204v3 Announce Type: replace Abstract: We study the problem of online learning in Stackelberg games with side information between a leader and a sequence of followers. In every round the

Networked Information Aggregation for Binary Classification

AgentsDGX agent

arXiv:2605.01082v1 Announce Type: new Abstract: We study networked binary classification on a directed acyclic graph (DAG) where each agent observes only a subset of the feature columns of a shared da

NeuroViz: Real-time Interactive Visualization of Forward and Backward Passes in Neural Network Training

ResearchDGX agent

arXiv:2605.02044v1 Announce Type: new Abstract: Training neural networks is difficult to interpret, particularly for newcomers. We introduce NeuroViz, an interactive visualization tool that supports r

New Bounds for Kernel Sums via Fast Spherical Embeddings

ResearchDGX agent

arXiv:2605.01263v1 Announce Type: cross Abstract: We study query time bounds for the fundamental problem of estimating the kernel mean frac1{|X|}sum_{xin X}mathbf{k}(x,y) of a query y in a finite data

On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization

Model ReleasesDGX agent

arXiv:2605.02141v1 Announce Type: new Abstract: Kullback-Leibler (KL) regularization is widely used in offline decision-making and offers several benefits, motivating recent work on the sample complex

On Training Large Language Models for Long-Horizon Tasks: An Empirical Study of Horizon Length

ResearchDGX agent

arXiv:2605.02572v1 Announce Type: cross Abstract: Large language models (LLMs) have shown promise as interactive agents that solve tasks through extended sequences of environment interactions. While p

Online Generalised Predictive Coding

TutorialsDGX agent

arXiv:2605.02675v1 Announce Type: cross Abstract: This paper introduces an extension of generalised filtering for online applications. Generalised filtering refers to data assimilation schemes that jo

Optimistic {epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2502.03506v2 Announce Type: replace-cross Abstract: The Centralized Training with Decentralized Execution (CTDE) paradigm is widely used in cooperative multi-agent reinforcement learning. Howeve

P1-KAN: an effective Kolmogorov-Arnold network with application to hydraulic valley optimization

ResearchDGX agent

arXiv:2410.03801v5 Announce Type: replace Abstract: A new Kolmogorov-Arnold network (KAN) is proposed to approximate potentially irregular functions in high dimensions. We provide error bounds for thi

P3-LLM: An Integrated NPU-PIM Accelerator for Edge LLM Inference Using Hybrid Numerical Formats

ResearchDGX agent

arXiv:2511.06838v4 Announce Type: replace-cross Abstract: The substantial memory bandwidth and computational demands of large language models (LLMs) present critical challenges for efficient inference

PACE: Parameter Change for Unsupervised Environment Design

Model ReleasesDGX agent

arXiv:2605.01358v1 Announce Type: new Abstract: Unsupervised Environment Design (UED) offers a promising paradigm for improving reinforcement learning generalization by adaptively shaping training env

Pandora's Regret: A Proper Scoring Rule for Evaluating Sequential Search

Model ReleasesDGX agent

arXiv:2605.01936v1 Announce Type: new Abstract: In sequential search, alternatives are tested until the true class is found. Standard proper scoring rules like log loss are local, ignoring the ranking

Parameter Space Analysis through Guided Visual Interpolations

Model ReleasesDGX agent

arXiv:2509.19202v2 Announce Type: replace-cross Abstract: We propose Parameter Space Analysis through Guided Visual Interpolations (ParamInter), a novel tool for high-dimensional input parameter space

ParaRNN: An Interpretable and Parallelizable Recurrent Neural Network for Time-Dependent Data

ResearchDGX agent

arXiv:2605.02692v1 Announce Type: cross Abstract: The proliferation of large-scale and structurally complex data has spurred the integration of machine learning methods into statistical modeling. Recu

PepSpecBench: A Unified Evaluation Benchmark for Peptide Tandem Mass Spectrometry Prediction

Model ReleasesDGX agent

arXiv:2605.01945v1 Announce Type: new Abstract: Tandem mass spectrometry provides a high-throughput framework for identifying and quantifying proteins in complex biological samples. In computational p

Personalized Federated Learning for Gradient Alignment

Local AiDGX agent

arXiv:2605.02143v1 Announce Type: new Abstract: Personalized federated learning (pFL) aims to adapt models to client specific data distributions, yet it often fails to reliably preserve personalized i

Perturb and Correct: Post-Hoc Ensembles using Affine Redundancy

ResearchDGX agent

arXiv:2605.01632v1 Announce Type: new Abstract: Models that are indistinguishable on in-distribution data can behave very differently under distribution shift. We introduce Perturb-and-Correct (P&C),

PhaseNet++: Phase-Aware Frequency-Domain Anomaly Detection for Industrial Control Systems via Phase Coherence Graphs

Model ReleasesDGX agent

arXiv:2605.00929v1 Announce Type: new Abstract: Multivariate time series anomaly detection in ICS has attracted growing attention due to the increasing threat of cyber-physical attacks on critical inf

phi-Table: A Statistical Explanation for Global SHAP

ResearchDGX agent

arXiv:2512.07578v3 Announce Type: replace-cross Abstract: Global SHAP explanations are typically presented as feature-importance rankings, which identify variables that matter to a black-box model but

Physics-Informed Neural Learning for State Reconstruction and Parameter Identification in Coupled Greenhouse Climate Dynamics

Model ReleasesDGX agent

arXiv:2605.02524v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have recently emerged as a promising framework for integrating data-driven learning with physical knowledge. In

Physiology-Aware Masked Cross-Modal Reconstruction for Biosignal Representation Learning

ResearchDGX agent

arXiv:2605.00973v1 Announce Type: new Abstract: Biosignals acquired from different locations on the body often provide temporally ordered views of the same underlying physiological process. However, m

Pi-Change: A Prior-Informed Multiple Change Point Detection Algorithm

ResearchDGX agent

arXiv:2605.01003v1 Announce Type: cross Abstract: Statistical change point (CP) detection methods typically rely on likelihood-based inference and ignore contextual information about plausible CP loca

PiCSRL: Physics-Informed Contextual Spectral Reinforcement Learning

TutorialsDGX agent

arXiv:2603.26816v2 Announce Type: replace Abstract: High-dimensional low-sample-size (HDLSS) datasets constrain reliable environmental model development, where labeled data remain sparse. Reinforcemen

Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning

Model ReleasesDGX agent

arXiv:2605.02168v1 Announce Type: cross Abstract: Language model (LM)-based agents have demonstrated promising capabilities in automating complex tasks from natural language instructions, yet they con

Polynomial-Time Optimal Group Selection via the Double-Commutator Eigenvalue Problem

ResearchDGX agent

arXiv:2605.00834v1 Announce Type: new Abstract: The algebraic diversity framework replaces temporal averaging over multiple observations with algebraic group action on a single observation for second-

Poodle: Seamlessly Scaling Down Large Language Models with Just-in-Time Model Replacement

ResearchDGX agent

arXiv:2512.05525v2 Announce Type: replace-cross Abstract: Businesses increasingly rely on large language models (LLMs) to automate simple repetitive tasks instead of developing custom machine learning

PPO guided Agentic Pipeline for Adaptive Prompt Selection and Test Case Generation

Model ReleasesDGX agent

arXiv:2605.00942v1 Announce Type: cross Abstract: Developing effective test cases capable of thoroughly exercising large-scale software systems is inherently difficult, especially if such systems have

PRCD-MAP: Learning How Much to Trust Imperfect Priors in Causal Discovery

SafetyDGX agent

arXiv:2605.01669v1 Announce Type: cross Abstract: External priors of unknown reliability create a brittle trade-off in causal discovery: blind trust amplifies errors, blind rejection wastes signal. Re

Predicting Post Virality with Temporal Cross-Attention over Trend Signals

ApplicationsDGX agent

arXiv:2605.02358v1 Announce Type: new Abstract: Current models for predicting social media virality rely heavily on static textual and structural features, effectively ignoring the highly dynamic natu

Pretraining on Sleep Data Improves non-Sleep Biosignal Tasks

ResearchDGX agent

arXiv:2605.02500v1 Announce Type: new Abstract: Sleep foundation models have recently demonstrated strong performance on in-domain polysomnography tasks, including sleep staging, apnea detection, and

PRIME: Protein Representation via Physics-Informed Multiscale Equivariant Hierarchies

Model ReleasesDGX agent

arXiv:2605.01625v1 Announce Type: new Abstract: Proteins are inherently multiscale physical systems whose functional properties emerge from coordinated structural organization across multiple spatial

Principles and Guidelines for Randomized Controlled Trials in AI Evaluation

ResearchDGX agent

arXiv:2605.02050v1 Announce Type: cross Abstract: This work establishes a foundational framework for standardizing AI evaluation RCTs (sometimes called human uplift studies). Drawing on established ex

Probe-Geometry Alignment: Erasing the Cross-Sequence Memorization Signature Below Chance

Model ReleasesDGX agent

arXiv:2605.01699v1 Announce Type: new Abstract: Recent attacks show that behavioural unlearning of large language models leaves internal traces recoverable by adversarial probes. We characterise where

Projection-Free Transformers via Gaussian Kernel Attention

Model ReleasesDGX agent

arXiv:2605.02144v1 Announce Type: new Abstract: Self-attention in Transformers is typically implemented as softmax(QK^op/sqrt{d})V, where Q=XW_Q, K=XW_K, and V=XW_V are learned linear projections of t

ProPACT: A Proactive AI-Driven Adaptive Collaborative Tutor for Pair Programming

SafetyDGX agent

arXiv:2605.02703v1 Announce Type: cross Abstract: Effective pair programming depends on coordination of attention, cognitive effort, and joint regulation over time, yet most adaptive learning systems

Protein-Conditioned Multi-Objective Reinforcement Learning for Full-Length mRNA Design

SafetyDGX agent

arXiv:2605.01513v1 Announce Type: new Abstract: Designing therapeutic messenger RNA (mRNA) requires creating full-length transcripts that carefully balance stability, translation efficiency, and immun

Provable Benefit of Curriculum in Transformer Tree-Reasoning Post-Training

ResearchDGX agent

arXiv:2511.07372v3 Announce Type: replace Abstract: Recent curriculum techniques in the post-training stage of LLMs have been empirically observed to outperform non-curriculum approaches in improving

Provably Learning Attention with Queries

TutorialsDGX agent

arXiv:2601.16873v2 Announce Type: replace Abstract: We study the problem of learning Transformer-based sequence models with black-box access to their outputs. In this setting, a learner may adaptively

Pruning Federated Models through Loss Landscape Analysis and Client Agreement Scoring

ResearchDGX agent

arXiv:2405.10271v4 Announce Type: replace Abstract: The practical deployment of Federated Learning (FL) on resource-constrained devices is fundamentally limited by the high cost of training large mode

Q-RAG: Long Context Multi-step Retrieval via Value-based Embedder Training

ResearchDGX agent

arXiv:2511.07328v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) methods enhance LLM performance by efficiently filtering relevant context for LLMs, reducing hallucinations and

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL

ApplicationsDGX agent

arXiv:2605.01862v1 Announce Type: new Abstract: Offline goal-conditioned RL (GCRL) learns goal-reaching policies from static datasets, but real-world datasets are often partially observable and histor

Quant VideoGen: Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization

HardwareDGX agent

arXiv:2602.02958v4 Announce Type: replace Abstract: Despite rapid progress in autoregressive video diffusion, an emerging system algorithm bottleneck limits both deployability and generation capabilit

← Previous
1…190191192193194…241
Next →