AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
Model Releases

Multi-Perspective Transformers in ARC-AGI-2 Challenge

DGX agent

arXiv:2605.01154v1 Announce Type: new Abstract: ARC-AGI-2 is a benchmark of human-intuitive visual puzzles that measures a machine's ability to generalize from limited examples, interpret symbolic mea

model-releasesarxiv-cs-lg
5 May 2026
Safety

Multi-User Dueling Bandits: A Fair Approach using Nash Social Welfare

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.01961v1 Announce Type: new Abstract: Learning from human preference data is becoming a useful tool, from fine-tuning large language models to training reinforcement learning agents. However

safetyarxiv-cs-lg
5 May 2026
Safety

Multimodal Data Curation Through Ranked Retrieval

DGX agent

arXiv:2605.01163v1 Announce Type: cross Abstract: Shared embedding spaces are widely used for multimodal search and data curation. In practice, two problems often limit how well this works. First, emb

safetyarxiv-cs-lg
5 May 2026
Applications

NAPS: Attention-Based Fusion of Heterogeneous Physiological Signals

DGX agent

arXiv:2511.03488v2 Announce Type: replace Abstract: Physiological signals are inherently heterogeneous: they are collected under diverse acquisition setups, differ in the number and type of modalities

applicationsarxiv-cs-lg
5 May 2026
Safety

NaviMaster: Learning a Unified Policy for GUI and Embodied Navigation Tasks

DGX agent

arXiv:2508.02046v4 Announce Type: replace-cross Abstract: Recent advances in Graphical User Interface (GUI) and embodied navigation have driven progress, yet these domains have largely evolved in isol

safetyarxiv-cs-lg
5 May 2026
Agents

Near-Optimal Privacy-Preserving Learning for Max-Min Fair Multi-Agent Bandits

DGX agent

arXiv:2306.04498v3 Announce Type: replace Abstract: We study fair multi-agent multi-armed bandit learning under collision-only coordination. Agents cannot communicate explicitly during learning and ob

agentsarxiv-cs-lg
5 May 2026
Research

Nearly-Optimal Bandit Learning in Stackelberg Games with Side Information

DGX agent

arXiv:2502.00204v3 Announce Type: replace Abstract: We study the problem of online learning in Stackelberg games with side information between a leader and a sequence of followers. In every round the

researcharxiv-cs-lg
5 May 2026
Agents

Networked Information Aggregation for Binary Classification

DGX agent

arXiv:2605.01082v1 Announce Type: new Abstract: We study networked binary classification on a directed acyclic graph (DAG) where each agent observes only a subset of the feature columns of a shared da

agentsarxiv-cs-lg
5 May 2026
Research

NeuroViz: Real-time Interactive Visualization of Forward and Backward Passes in Neural Network Training

DGX agent

arXiv:2605.02044v1 Announce Type: new Abstract: Training neural networks is difficult to interpret, particularly for newcomers. We introduce NeuroViz, an interactive visualization tool that supports r

researcharxiv-cs-lg
5 May 2026
Research

New Bounds for Kernel Sums via Fast Spherical Embeddings

DGX agent

arXiv:2605.01263v1 Announce Type: cross Abstract: We study query time bounds for the fundamental problem of estimating the kernel mean frac1{|X|}sum_{xin X}mathbf{k}(x,y) of a query y in a finite data

researcharxiv-cs-lg
5 May 2026
Model Releases

On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization

DGX agent

arXiv:2605.02141v1 Announce Type: new Abstract: Kullback-Leibler (KL) regularization is widely used in offline decision-making and offers several benefits, motivating recent work on the sample complex

model-releasesarxiv-cs-lg
5 May 2026
Research

On Training Large Language Models for Long-Horizon Tasks: An Empirical Study of Horizon Length

DGX agent

arXiv:2605.02572v1 Announce Type: cross Abstract: Large language models (LLMs) have shown promise as interactive agents that solve tasks through extended sequences of environment interactions. While p

researcharxiv-cs-lg
5 May 2026
Tutorials

Online Generalised Predictive Coding

DGX agent

arXiv:2605.02675v1 Announce Type: cross Abstract: This paper introduces an extension of generalised filtering for online applications. Generalised filtering refers to data assimilation schemes that jo

tutorialsarxiv-cs-lg
5 May 2026
Agents

Optimistic {epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2502.03506v2 Announce Type: replace-cross Abstract: The Centralized Training with Decentralized Execution (CTDE) paradigm is widely used in cooperative multi-agent reinforcement learning. Howeve

agentsarxiv-cs-lg
5 May 2026
Research

P1-KAN: an effective Kolmogorov-Arnold network with application to hydraulic valley optimization

DGX agent

arXiv:2410.03801v5 Announce Type: replace Abstract: A new Kolmogorov-Arnold network (KAN) is proposed to approximate potentially irregular functions in high dimensions. We provide error bounds for thi

researcharxiv-cs-lg
5 May 2026
Research

P3-LLM: An Integrated NPU-PIM Accelerator for Edge LLM Inference Using Hybrid Numerical Formats

DGX agent

arXiv:2511.06838v4 Announce Type: replace-cross Abstract: The substantial memory bandwidth and computational demands of large language models (LLMs) present critical challenges for efficient inference

researcharxiv-cs-lg
5 May 2026
Model Releases

PACE: Parameter Change for Unsupervised Environment Design

DGX agent

arXiv:2605.01358v1 Announce Type: new Abstract: Unsupervised Environment Design (UED) offers a promising paradigm for improving reinforcement learning generalization by adaptively shaping training env

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Pandora's Regret: A Proper Scoring Rule for Evaluating Sequential Search

DGX agent

arXiv:2605.01936v1 Announce Type: new Abstract: In sequential search, alternatives are tested until the true class is found. Standard proper scoring rules like log loss are local, ignoring the ranking

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Parameter Space Analysis through Guided Visual Interpolations

DGX agent

arXiv:2509.19202v2 Announce Type: replace-cross Abstract: We propose Parameter Space Analysis through Guided Visual Interpolations (ParamInter), a novel tool for high-dimensional input parameter space

model-releasesarxiv-cs-lg
5 May 2026
Research

ParaRNN: An Interpretable and Parallelizable Recurrent Neural Network for Time-Dependent Data

DGX agent

arXiv:2605.02692v1 Announce Type: cross Abstract: The proliferation of large-scale and structurally complex data has spurred the integration of machine learning methods into statistical modeling. Recu

researcharxiv-cs-lg
5 May 2026
Model Releases

PepSpecBench: A Unified Evaluation Benchmark for Peptide Tandem Mass Spectrometry Prediction

DGX agent

arXiv:2605.01945v1 Announce Type: new Abstract: Tandem mass spectrometry provides a high-throughput framework for identifying and quantifying proteins in complex biological samples. In computational p

model-releasesarxiv-cs-lg
5 May 2026
Local Ai

Personalized Federated Learning for Gradient Alignment

DGX agent

arXiv:2605.02143v1 Announce Type: new Abstract: Personalized federated learning (pFL) aims to adapt models to client specific data distributions, yet it often fails to reliably preserve personalized i

local-aiarxiv-cs-lg
5 May 2026
Research

Perturb and Correct: Post-Hoc Ensembles using Affine Redundancy

DGX agent

arXiv:2605.01632v1 Announce Type: new Abstract: Models that are indistinguishable on in-distribution data can behave very differently under distribution shift. We introduce Perturb-and-Correct (P&C),

researcharxiv-cs-lg
5 May 2026
Model Releases

PhaseNet++: Phase-Aware Frequency-Domain Anomaly Detection for Industrial Control Systems via Phase Coherence Graphs

DGX agent

arXiv:2605.00929v1 Announce Type: new Abstract: Multivariate time series anomaly detection in ICS has attracted growing attention due to the increasing threat of cyber-physical attacks on critical inf

model-releasesarxiv-cs-lg
5 May 2026
Research

phi-Table: A Statistical Explanation for Global SHAP

DGX agent

arXiv:2512.07578v3 Announce Type: replace-cross Abstract: Global SHAP explanations are typically presented as feature-importance rankings, which identify variables that matter to a black-box model but

researcharxiv-cs-lg
5 May 2026
Model Releases

Physics-Informed Neural Learning for State Reconstruction and Parameter Identification in Coupled Greenhouse Climate Dynamics

DGX agent

arXiv:2605.02524v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have recently emerged as a promising framework for integrating data-driven learning with physical knowledge. In

model-releasesarxiv-cs-lg
5 May 2026
Research

Physiology-Aware Masked Cross-Modal Reconstruction for Biosignal Representation Learning

DGX agent

arXiv:2605.00973v1 Announce Type: new Abstract: Biosignals acquired from different locations on the body often provide temporally ordered views of the same underlying physiological process. However, m

researcharxiv-cs-lg
5 May 2026
Research

Pi-Change: A Prior-Informed Multiple Change Point Detection Algorithm

DGX agent

arXiv:2605.01003v1 Announce Type: cross Abstract: Statistical change point (CP) detection methods typically rely on likelihood-based inference and ignore contextual information about plausible CP loca

researcharxiv-cs-lg
5 May 2026
Tutorials

PiCSRL: Physics-Informed Contextual Spectral Reinforcement Learning

DGX agent

arXiv:2603.26816v2 Announce Type: replace Abstract: High-dimensional low-sample-size (HDLSS) datasets constrain reliable environmental model development, where labeled data remain sparse. Reinforcemen

tutorialsarxiv-cs-lg
5 May 2026
Model Releases

Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning

DGX agent

arXiv:2605.02168v1 Announce Type: cross Abstract: Language model (LM)-based agents have demonstrated promising capabilities in automating complex tasks from natural language instructions, yet they con

model-releasesarxiv-cs-lg
5 May 2026
Research

Polynomial-Time Optimal Group Selection via the Double-Commutator Eigenvalue Problem

DGX agent

arXiv:2605.00834v1 Announce Type: new Abstract: The algebraic diversity framework replaces temporal averaging over multiple observations with algebraic group action on a single observation for second-

researcharxiv-cs-lg
5 May 2026
Research

Poodle: Seamlessly Scaling Down Large Language Models with Just-in-Time Model Replacement

DGX agent

arXiv:2512.05525v2 Announce Type: replace-cross Abstract: Businesses increasingly rely on large language models (LLMs) to automate simple repetitive tasks instead of developing custom machine learning

researcharxiv-cs-lg
5 May 2026
Model Releases

PPO guided Agentic Pipeline for Adaptive Prompt Selection and Test Case Generation

DGX agent

arXiv:2605.00942v1 Announce Type: cross Abstract: Developing effective test cases capable of thoroughly exercising large-scale software systems is inherently difficult, especially if such systems have

model-releasesarxiv-cs-lg
5 May 2026
Safety

PRCD-MAP: Learning How Much to Trust Imperfect Priors in Causal Discovery

DGX agent

arXiv:2605.01669v1 Announce Type: cross Abstract: External priors of unknown reliability create a brittle trade-off in causal discovery: blind trust amplifies errors, blind rejection wastes signal. Re

safetyarxiv-cs-lg
5 May 2026
Applications

Predicting Post Virality with Temporal Cross-Attention over Trend Signals

DGX agent

arXiv:2605.02358v1 Announce Type: new Abstract: Current models for predicting social media virality rely heavily on static textual and structural features, effectively ignoring the highly dynamic natu

applicationsarxiv-cs-lg
5 May 2026
Research

Pretraining on Sleep Data Improves non-Sleep Biosignal Tasks

DGX agent

arXiv:2605.02500v1 Announce Type: new Abstract: Sleep foundation models have recently demonstrated strong performance on in-domain polysomnography tasks, including sleep staging, apnea detection, and

researcharxiv-cs-lg
5 May 2026
Model Releases

PRIME: Protein Representation via Physics-Informed Multiscale Equivariant Hierarchies

DGX agent

arXiv:2605.01625v1 Announce Type: new Abstract: Proteins are inherently multiscale physical systems whose functional properties emerge from coordinated structural organization across multiple spatial

model-releasesarxiv-cs-lg
5 May 2026
Research

Principles and Guidelines for Randomized Controlled Trials in AI Evaluation

DGX agent

arXiv:2605.02050v1 Announce Type: cross Abstract: This work establishes a foundational framework for standardizing AI evaluation RCTs (sometimes called human uplift studies). Drawing on established ex

researcharxiv-cs-lg
5 May 2026
Model Releases

Probe-Geometry Alignment: Erasing the Cross-Sequence Memorization Signature Below Chance

DGX agent

arXiv:2605.01699v1 Announce Type: new Abstract: Recent attacks show that behavioural unlearning of large language models leaves internal traces recoverable by adversarial probes. We characterise where

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Projection-Free Transformers via Gaussian Kernel Attention

DGX agent

arXiv:2605.02144v1 Announce Type: new Abstract: Self-attention in Transformers is typically implemented as softmax(QK^op/sqrt{d})V, where Q=XW_Q, K=XW_K, and V=XW_V are learned linear projections of t

model-releasesarxiv-cs-lg
5 May 2026
Safety

ProPACT: A Proactive AI-Driven Adaptive Collaborative Tutor for Pair Programming

DGX agent

arXiv:2605.02703v1 Announce Type: cross Abstract: Effective pair programming depends on coordination of attention, cognitive effort, and joint regulation over time, yet most adaptive learning systems

safetyarxiv-cs-lg
5 May 2026
Safety

Protein-Conditioned Multi-Objective Reinforcement Learning for Full-Length mRNA Design

DGX agent

arXiv:2605.01513v1 Announce Type: new Abstract: Designing therapeutic messenger RNA (mRNA) requires creating full-length transcripts that carefully balance stability, translation efficiency, and immun

safetyarxiv-cs-lg
5 May 2026
Research

Provable Benefit of Curriculum in Transformer Tree-Reasoning Post-Training

DGX agent

arXiv:2511.07372v3 Announce Type: replace Abstract: Recent curriculum techniques in the post-training stage of LLMs have been empirically observed to outperform non-curriculum approaches in improving

researcharxiv-cs-lg
5 May 2026
Tutorials

Provably Learning Attention with Queries

DGX agent

arXiv:2601.16873v2 Announce Type: replace Abstract: We study the problem of learning Transformer-based sequence models with black-box access to their outputs. In this setting, a learner may adaptively

tutorialsarxiv-cs-lg
5 May 2026
Research

Pruning Federated Models through Loss Landscape Analysis and Client Agreement Scoring

DGX agent

arXiv:2405.10271v4 Announce Type: replace Abstract: The practical deployment of Federated Learning (FL) on resource-constrained devices is fundamentally limited by the high cost of training large mode

researcharxiv-cs-lg
5 May 2026
Research

Q-RAG: Long Context Multi-step Retrieval via Value-based Embedder Training

DGX agent

arXiv:2511.07328v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) methods enhance LLM performance by efficiently filtering relevant context for LLMs, reducing hallucinations and

researcharxiv-cs-lg
5 May 2026
Applications

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL

DGX agent

arXiv:2605.01862v1 Announce Type: new Abstract: Offline goal-conditioned RL (GCRL) learns goal-reaching policies from static datasets, but real-world datasets are often partially observable and histor

applicationsarxiv-cs-lg
5 May 2026
Hardware

Quant VideoGen: Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization

DGX agent

arXiv:2602.02958v4 Announce Type: replace Abstract: Despite rapid progress in autoregressive video diffusion, an emerging system algorithm bottleneck limits both deployability and generation capabilit

hardwarearxiv-cs-lg
5 May 2026
← Previous
1…238239240241242…301
Next →