AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
Research

Learning with Monotone Adversarial Corruptions

DGX agent

arXiv:2601.02193v2 Announce Type: replace Abstract: We study the extent to which standard machine learning algorithms rely on exchangeability and independence of data by introducing a monotone adversa

researcharxiv-cs-lg
25 Jun 2026
Tutorials

Lifelong In-Context Learning with Transformers Requires Parametric Forms of Attention

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.25342v1 Announce Type: new Abstract: Lifelong continual learning remains an obstacle on the path to human-like intelligence. Modern transformers show sparks of intelligence with in-context

tutorialsarxiv-cs-lg
25 Jun 2026
Research

Limitations of SGD for Multi-Index Models Beyond Statistical Queries

DGX agent

arXiv:2602.05704v2 Announce Type: replace Abstract: Understanding the limitations of gradient methods, and stochastic gradient descent (SGD) in particular, is a central challenge in learning theory. T

researcharxiv-cs-lg
25 Jun 2026
Applications

LLM Evolution as an Industry-Scale Ecosystem: A Lifecycle Perspective on Continual Learning

DGX agent

arXiv:2606.24901v1 Announce Type: new Abstract: Continual learning capability is critical for Industrial LLMs, as deployed models must be continuously updated to meet evolving requirements and environ

applicationsarxiv-cs-lg
25 Jun 2026
Tutorials

LLM Program Optimization via Retrieval Augmented Search

DGX agent

arXiv:2501.18916v2 Announce Type: replace Abstract: Recent work has demonstrated the potential of large language models (LLMs) for program optimization, a key challenge in programming languages. We pr

tutorialsarxiv-cs-lg
25 Jun 2026
Research

Logit Distance Bounds Representational Similarity

DGX agent

arXiv:2602.15438v3 Announce Type: replace Abstract: For a broad family of discriminative models that includes autoregressive language models, identifiability results imply that if two models induce th

researcharxiv-cs-lg
25 Jun 2026
Safety

Low-Complexity Policy Tessellations in Structured Markov Decision Processes

DGX agent

arXiv:2606.25593v1 Announce Type: new Abstract: We study optimal-policy geometry in structured Markov decision processes. While approximate dynamic programming and reinforcement learning typically app

safetyarxiv-cs-lg
25 Jun 2026
Research

Low-Cost High-Order Singular Value Decomposition for Tensor-Based Reconstruction from Sparse Sensor Measurements: Urban Flow and Air-Quality Applications

DGX agent

arXiv:2606.24989v1 Announce Type: new Abstract: Urban flow and air-quality simulations generate high-dimensional datasets describing velocity and pollutant transport across multiple spatial, temporal,

researcharxiv-cs-lg
25 Jun 2026
Safety

Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning

DGX agent

arXiv:2606.25526v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning assumes each agent shares the same reward function and can be trained effectively using the Trust Region

safetyarxiv-cs-lg
25 Jun 2026
Model Releases

MacroLens: A Multi-Task Benchmark for Contextual Financial Reasoning under Macroeconomic Scenarios

DGX agent

arXiv:2606.24950v1 Announce Type: new Abstract: Financial decision-making is contextual: forecasting prices, valuing companies, and assessing event exposure weigh price history, accounting fundamental

model-releasesarxiv-cs-lg
25 Jun 2026
Applications

Margin in Abstract Spaces

DGX agent

arXiv:2603.07221v2 Announce Type: replace Abstract: Margin-based learning, exemplified by linear and kernel methods, is one of the few classical settings where generalization guarantees are independen

applicationsarxiv-cs-lg
25 Jun 2026
Model Releases

Memory-Efficient Policy Libraries with Low-Rank Adaptation in Reinforcement Learning

DGX agent

arXiv:2606.25700v1 Announce Type: new Abstract: When fine-tuning Large Language Models (LLMs), there has been success in minimizing both memory usage and computation with Parameter-Efficient Fine-Tuni

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

MINIF2F-DAFNY: LLM-Guided Mathematical Theorem Proving via Auto-Active Verification

DGX agent

arXiv:2512.10187v3 Announce Type: replace Abstract: LLMs excel at reasoning, but validating their steps remains challenging. Formal verification offers a solution through mechanically checkable proofs

model-releasesarxiv-cs-lg
25 Jun 2026
Safety

Minimax PAC Bounds for Learning in Exogenous Contextual MDPs

DGX agent

arXiv:2606.25170v1 Announce Type: cross Abstract: We study PAC learning in tabular discounted Markov decision processes with exogenous i.i.d. contexts, with discount factor gamma, finite state space m

safetyarxiv-cs-lg
25 Jun 2026
Safety

MiniOpt: Reasoning to Model and Solve General Optimization Problems with Limited Resources

DGX agent

arXiv:2606.25832v1 Announce Type: new Abstract: Achieving strong optimization generalization across diverse optimization problems while requiring limited training resources remains a challenging probl

safetyarxiv-cs-lg
25 Jun 2026
Model Releases

Model-agnostic Mitigation Strategies of Data Imbalance for Regression

DGX agent

arXiv:2506.01486v2 Announce Type: replace Abstract: Data imbalance persists as a pervasive challenge in regression tasks, introducing bias in model performance and undermining predictive reliability.

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Model Forensics: Investigating Whether Concerning Behavior Reflects Misalignment

DGX agent

arXiv:2606.26071v1 Announce Type: new Abstract: A central goal of safety research is determining whether a model is misaligned. Prior work has largely focused on detecting concerning behavior. But beh

model-releasesarxiv-cs-lg
25 Jun 2026
Research

MorphStrata: Layer-Specific Perturbations for Generating Morphence Students in Time-Series Moving Target Defense

DGX agent

arXiv:2606.17435v2 Announce Type: replace Abstract: Time-series forecasting models remain vulnerable to gradient-based adversarial attacks while existing defense mechanisms typically incur a trade-off

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Multi-Agent Goal Recognition with Team- and Goal-Conditioned Reinforcement Learning and Factorized Branch-and-Bound

DGX agent

arXiv:2606.25978v1 Announce Type: cross Abstract: Multi-agent goal recognition asks an observer to jointly infer which agents act together and what each team is trying to achieve, so the hypothesis sp

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Multi-Stream Temporal Fusion for Financial Fraud Detection

DGX agent

arXiv:2606.25007v1 Announce Type: new Abstract: Financial fraud detection in digital banking requires reasoning over multiple heterogeneous event streams -- transactions, login sessions, risk signals

model-releasesarxiv-cs-lg
25 Jun 2026
Tutorials

Multifidelity-Augmented Gaussian Process Inputs for Surrogate Modeling from Scarce Data

DGX agent

arXiv:2603.22050v2 Announce Type: replace-cross Abstract: Supervised machine learning describes the practice of fitting a parameterized model to labeled input-output data. Supervised machine learning

tutorialsarxiv-cs-lg
25 Jun 2026
Safety

Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents

DGX agent

arXiv:2606.26080v1 Announce Type: new Abstract: Process reward models enable fine-grained, step-level evaluation of LLMs, yet building them for agentic settings remains prohibitively difficult: long-h

safetyarxiv-cs-lg
25 Jun 2026
Research

Neural operator-based digital twins for modeling amyloid-eta and tau propagation and treatment optimization in Alzheimer's disease

DGX agent

arXiv:2606.25185v1 Announce Type: new Abstract: Accurately predicting the spatiotemporal evolution of amyloid-eta and tau proteins at the individual level is critical for improving the diagnosis and t

researcharxiv-cs-lg
25 Jun 2026
Research

OmegAMP: Targeted AMP Discovery via Biologically Informed Generation

DGX agent

arXiv:2504.17247v3 Announce Type: replace Abstract: Deep learning-based antimicrobial peptide (AMP) discovery faces critical challenges such as limited controllability, lack of representations that ef

researcharxiv-cs-lg
25 Jun 2026
Model Releases

On-Device Neural Architecture Search

DGX agent

arXiv:2606.24900v1 Announce Type: new Abstract: This paper proposes a new approach to near-sensor computing, in which a lightweight Neural Architecture Search (NAS) is performed directly on the deploy

model-releasesarxiv-cs-lg
25 Jun 2026
Safety

On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity

DGX agent

arXiv:2606.26091v1 Announce Type: new Abstract: On-policy self-distillation achieves strong pass@1 accuracy by using a single model as both teacher and student, with the teacher conditioned on a corre

safetyarxiv-cs-lg
25 Jun 2026
Applications

OncoSynth: Synthetic data generation for treatment effect estimation in oncology

DGX agent

arXiv:2606.25762v1 Announce Type: new Abstract: In oncology, access to patient-level data is often restricted. Synthetic data provides an alternative for analyzing treatment effectiveness, but existin

applicationsarxiv-cs-lg
25 Jun 2026
Model Releases

Operator Boosting Produces Pareto-Efficient PDE Surrogates

DGX agent

arXiv:2606.17460v2 Announce Type: replace Abstract: Neural operators are widely used as surrogate solution maps for partial differential equations (PDEs), but full-size models can be costly to store,

model-releasesarxiv-cs-lg
25 Jun 2026
Safety

PERRY: Policy Evaluation with Confidence Intervals using Auxiliary Data

DGX agent

arXiv:2507.20068v3 Announce Type: replace Abstract: Off-policy evaluation (OPE) methods estimate the value of a new reinforcement learning (RL) policy prior to deployment. Recent advances have shown t

safetyarxiv-cs-lg
25 Jun 2026
Research

PERTINENCE: Input-based Opportunistic Neural Network Dynamic Execution

DGX agent

arXiv:2507.01695v3 Announce Type: replace Abstract: Deep neural networks (DNNs) are widely used for their ability to model complex patterns across domains such as computer vision, speech recognition,

researcharxiv-cs-lg
25 Jun 2026
Research

Physics-conforming Latent Twins

DGX agent

arXiv:2606.15053v2 Announce Type: replace Abstract: Surrogate models are central to scientific machine learning, where they enable fast prediction, simulation, inference, and control for complex physi

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Project Auto-World: Towards Automated Benchmarking of Neural Relational Reasoners

DGX agent

arXiv:2606.24965v1 Announce Type: cross Abstract: Reasoning about relational structures remains a significant challenge for neural models, particularly when they must systematically apply learned know

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

PVF:Understanding AI Vulnerability Against SDCs

DGX agent

arXiv:2405.01741v4 Announce Type: replace-cross Abstract: Reliability of AI systems is a fundamental concern for the successful deployment and widespread adoption of AI technologies. Unfortunately, th

model-releasesarxiv-cs-lg
25 Jun 2026
Applications

Quantifying Explainable AI-introduced signal noise on ECG data with Spectral Entropy

DGX agent

arXiv:2606.24974v1 Announce Type: new Abstract: Explainability techniques are used to assess the output of various deep learning models. This is especially true in healthcare, where models need to be

applicationsarxiv-cs-lg
25 Jun 2026
Agents

Quantization Inflates Reasoning: Token Inflation as a Hidden Cost of Low-Bit Reasoning Models

DGX agent

arXiv:2606.25519v1 Announce Type: cross Abstract: Quantization is widely used to reduce the inference cost of large language models, but its effect on reasoning models is not fully captured by final-a

agentsarxiv-cs-lg
25 Jun 2026
Research

Randomized Kriging Believer for Parallel Bayesian Optimization with Regret Bounds

DGX agent

arXiv:2603.01470v3 Announce Type: replace Abstract: We consider the optimization problem of an expensive-to-evaluate black-box function, in which we can obtain noisy function values in parallel. For t

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Rational Neural Networks have Expressivity Advantages

DGX agent

arXiv:2602.12390v2 Announce Type: replace Abstract: We study neural networks with trainable low-degree rational activation functions and show that they are more expressive and parameter-efficient than

model-releasesarxiv-cs-lg
25 Jun 2026
Research

Recursive QLSTM with Dynamic Variational Quantum Circuit Adaptation

DGX agent

arXiv:2606.24932v1 Announce Type: cross Abstract: Recent advances in quantum computing and machine learning have motivated the development of quantum models for sequential data processing. In this pap

researcharxiv-cs-lg
25 Jun 2026
Applications

Reliable Conformal Prediction for Ordinal Classification Using the Ranked Probability Score

DGX agent

arXiv:2606.24959v1 Announce Type: new Abstract: Ordinal classification (OC) arises in high-stakes domains such as medicine and finance, where uncertainty quantification must account for the severity o

applicationsarxiv-cs-lg
25 Jun 2026
Research

Retrieval-Augmented Personalization with Foundation Models for Wearable Stress Detection

DGX agent

arXiv:2606.24985v1 Announce Type: new Abstract: Personalization in wearable-based stress detection remains challenging due to substantial inter-individual variability in physiological and behavioral r

researcharxiv-cs-lg
25 Jun 2026
Model Releases

RevengeBench: Reverse Engineering Code-Space Policies from Behavioral Experiments

DGX agent

arXiv:2606.26094v1 Announce Type: new Abstract: For most of scientific history, researchers studying behavior could only infer hidden mechanisms from outward actions: an inverse problem that becomes m

model-releasesarxiv-cs-lg
25 Jun 2026
Safety

Reward-Conditioned Attention: How Reward Design Shapes What Autonomous Driving Agents See

DGX agent

arXiv:2606.25127v1 Announce Type: new Abstract: We investigate how reward design shapes the internal attention patterns of reinforcement learning agents trained for autonomous driving. Using three Per

safetyarxiv-cs-lg
25 Jun 2026
Safety

RN-D: Discretized Categorical Actors for On-Policy Reinforcement Learning

DGX agent

arXiv:2601.23075v2 Announce Type: replace Abstract: On-policy Reinforcement Learning (RL) remains a dominant paradigm for continuous control, yet standard implementations rely on Gaussian actors and r

safetyarxiv-cs-lg
25 Jun 2026
Safety

ROAD-VLA: Robust Online Adaptation via Self-Distillation for Vision-Language-Action Models

DGX agent

arXiv:2606.25800v1 Announce Type: new Abstract: Effective online adaptation of vision-language-action (VLA) models remains challenging, as sparse rewards provide weak supervision for high-dimensional

safetyarxiv-cs-lg
25 Jun 2026
Research

Robust Linear Predictions: Analyses of Uniform Concentration, Fast Rates and Model Misspecification

DGX agent

arXiv:2201.01973v3 Announce Type: replace-cross Abstract: The problem of linear predictions has been extensively studied for the past century under pretty generalized frameworks. Recent advances in th

researcharxiv-cs-lg
25 Jun 2026
Research

RotRNN: Modelling Long Sequences with Rotations

DGX agent

arXiv:2407.07239v3 Announce Type: replace Abstract: Linear recurrent neural networks, such as State Space Models (SSMs) and Linear Recurrent Units (LRUs), have recently shown state-of-the-art performa

researcharxiv-cs-lg
25 Jun 2026
Safety

Safe Learning Control with Optimality and Stability Guarantees

DGX agent

arXiv:2501.15373v2 Announce Type: replace-cross Abstract: Merely pursuing performance may adversely affect safety, while a conservative policy for safe exploration will degrade the performance. How to

safetyarxiv-cs-lg
25 Jun 2026
Research

Sample complexity of unbalanced entropic OT

DGX agent

arXiv:2606.24987v1 Announce Type: cross Abstract: Optimal transport (OT) has become a central language for comparing probability measures, but exact balanced OT is often both too rigid for data with m

researcharxiv-cs-lg
25 Jun 2026
← Previous
1…8889909192…304
Next →