AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
Safety

Beyond Linear Activation Steering: Invertible Latent Transformations for Controlling LLM Behavior

DGX agent

arXiv:2606.08454v1 Announce Type: new Abstract: Activation steering provides a lightweight inference-time mechanism for controlling large language models (LLMs) by modifying their internal activation

safetyarxiv-cs-lg
9 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Beyond Neural Collapse: Task-Intrinsic Geometry Governs Neural Representations in Modular Arithmetic

DGX agent

arXiv:2606.08985v1 Announce Type: new Abstract: While neural collapse (NC) predicts that a K-class-balanced classifier should organize terminal representations as a (K-1)-dimensional simplex equiangul

safetyarxiv-cs-lg
9 Jun 2026
Research

Biological Reasoning-Informed Regression for Interpretable Regulatory DNA Activity Prediction

DGX agent

arXiv:2606.08147v1 Announce Type: cross Abstract: DNA cis-regulatory elements (CREs) such as enhancers control gene expression levels. Accurately predicting regulatory activity from DNA sequences is v

researcharxiv-cs-lg
9 Jun 2026
Research

BlendServe: Optimizing Offline Inference for Auto-regressive Large Models with Resource-aware Batching

DGX agent

arXiv:2411.16102v2 Announce Type: replace Abstract: Offline batch inference, which leverages the flexibility of request batching to achieve higher throughput and lower costs, is becoming more popular

researcharxiv-cs-lg
9 Jun 2026
Safety

Boundary Variance Inflation Causes Acquisition Bias in Gaussian Processes

DGX agent

arXiv:2606.07561v1 Announce Type: new Abstract: Gaussian processes with stationary kernels on bounded domains exhibit inflated posterior variance near the boundary. Despite being a long-recognized art

safetyarxiv-cs-lg
9 Jun 2026
Research

BrainSurgery: Reproducible and Reliable Declarative Weight Manipulations for Model Editing and Upcycling

DGX agent

arXiv:2606.09707v1 Announce Type: new Abstract: As deep learning models scale, managing, inspecting, and modifying large checkpoints has become increasingly challenging. Researchers often need to alte

researcharxiv-cs-lg
9 Jun 2026
Model Releases

Breaking the Bubble: Asynchronous Pipeline Parallel Training with Bounded Weight Inconsistency

DGX agent

arXiv:2606.07881v1 Announce Type: new Abstract: Pipeline parallelism is essential for training large neural networks, but existing schedules trade off throughput, memory, and optimization consistency.

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

Breaking the Tokenizer Barrier: On-Policy Distillation across Model Families

DGX agent

arXiv:2606.09456v1 Announce Type: new Abstract: On-Policy Distillation (OPD) has become a core technique in the post-training of Large Language Models (LLMs) for transferring knowledge from domain exp

safetyarxiv-cs-lg
9 Jun 2026
Model Releases

BUDDY: BUdget-Driven DYnamic Depth Routing for Adaptive Large Language Model Inference

DGX agent

arXiv:2606.09514v1 Announce Type: new Abstract: Large language models (LLMs) incur high inference cost due to their depth and parameter scale. Depth pruning can reduce latency by skipping redundant Tr

model-releasesarxiv-cs-lg
9 Jun 2026
Research

Bulk-boundary decomposition of neural networks

DGX agent

arXiv:2511.02003v2 Announce Type: replace Abstract: We present the bulk--boundary decomposition as a new framework for understanding the training dynamics of deep neural networks. Starting from the st

researcharxiv-cs-lg
9 Jun 2026
Agents

Byzantine Cheap Talk: Adversarial Resilience and Topology Effects in LLM Coordination Games

DGX agent

arXiv:2606.07790v1 Announce Type: new Abstract: Multi-agent LLM systems increasingly rely on communication protocols for coordination, yet their robustness under adversarial and structural constraints

agentsarxiv-cs-lg
9 Jun 2026
Research

CAAL: Contextual Bandits based Online Hand-Craft Active Learning Strategy Selection

DGX agent

arXiv:2606.07910v1 Announce Type: new Abstract: The challenge with active learning algorithms is the uncertainty of the statistical distribution of unlabeled data, making it difficult to choose the be

researcharxiv-cs-lg
9 Jun 2026
Applications

Can LLMs extract scientific consensus? A case study in high-temperature superconductivity

DGX agent

arXiv:2606.07570v1 Announce Type: cross Abstract: Scientific knowledge is increasingly dispersed across vast and heterogeneous scientific literature, where important claims are often implicit, evolvin

applicationsarxiv-cs-lg
9 Jun 2026
Model Releases

CATPO: Critique-Augmented Tree Policy Optimization

DGX agent

arXiv:2606.08346v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a dominant paradigm for improving the reasoning capabilities of large language models

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Causal Longitudinal Prior-Fitted Networks for Counterfactual Outcome Prediction

DGX agent

arXiv:2606.05797v2 Announce Type: replace Abstract: Longitudinal treatment decisions from multivariate time-series data require predicting potential outcomes under future treatment sequences in the pr

model-releasesarxiv-cs-lg
9 Jun 2026
Research

Causal Representation Learning from Network Data

DGX agent

arXiv:2509.01916v2 Announce Type: replace Abstract: Causal disentanglement from soft interventions is identifiable under the assumptions of linear interventional faithfulness and availability of both

researcharxiv-cs-lg
9 Jun 2026
Safety

Causal Semantic Alignment for LLM-based Time Series Forecasting

DGX agent

arXiv:2606.08262v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have opened new possibilities for time series forecasting by enabling alignment between temporal pattern

safetyarxiv-cs-lg
9 Jun 2026
Applications

Characterizing the Discrete Geometry of ReLU Networks

DGX agent

arXiv:2606.07728v1 Announce Type: new Abstract: It is well established that ReLU networks define continuous piecewise-linear functions, and that their linear regions are polyhedra in the input space.

applicationsarxiv-cs-lg
9 Jun 2026
Safety

Claw-R1: A Step-Level Data Middleware System for Agentic Reinforcement Learning

DGX agent

arXiv:2606.09138v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has become an important post-training paradigm for turning LLMs from static chatbots into interactive agents, giving

safetyarxiv-cs-lg
9 Jun 2026
Safety

Code Is More Than Text: Uncertainty Estimation for Code Generation

DGX agent

arXiv:2606.09577v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as code generators, where silently wrong programs pose real safety and reliability risks. Relia

safetyarxiv-cs-lg
9 Jun 2026
Local Ai

Communication-Efficient Federated Learning under Dynamic Device Arrival and Departure: Convergence Analysis and Algorithm Design

DGX agent

arXiv:2410.05662v4 Announce Type: replace Abstract: Most federated learning (FL) approaches assume a fixed device set. However, real-world scenarios often involve devices dynamically joining or leavin

local-aiarxiv-cs-lg
9 Jun 2026
Research

Community-Specific Slang and Entity Detection via Semantic Shift in Fine-Tuned Language Models

DGX agent

arXiv:2606.07522v1 Announce Type: cross Abstract: We propose an unsupervised method of resolving slang, unique entities, and folklore from online communities by isolating words in the lexicon that hav

researcharxiv-cs-lg
9 Jun 2026
Research

Compositional Approximation Can Strictly Outperform Superpositional Approximation

DGX agent

arXiv:2606.08727v1 Announce Type: cross Abstract: Many classically studied function classes are known to be approximated optimally by superpositional methods, i.e. with approximants constructed as the

researcharxiv-cs-lg
9 Jun 2026
Model Releases

Conditional Normalizing Flows for Forward and Backward Joint State and Parameter Estimation

DGX agent

arXiv:2601.07013v2 Announce Type: replace-cross Abstract: Traditional filtering algorithms for state estimation -- such as classical Kalman filtering, unscented Kalman filtering, and particle filters

model-releasesarxiv-cs-lg
9 Jun 2026
Research

Conditional Random Ordered Transport Spaces

DGX agent

arXiv:2606.08113v1 Announce Type: new Abstract: A small Wasserstein distance does not certify that a transformation is admissible. In evidence-constrained, semantic, causal, physical, monotone, or ris

researcharxiv-cs-lg
9 Jun 2026
Safety

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning

DGX agent

arXiv:2606.08088v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has recently become a key paradigm for improving the reasoning abilities of Large Language Models

safetyarxiv-cs-lg
9 Jun 2026
Safety

Constrained user-item allocation for e-commerce marketing campaigns

DGX agent

arXiv:2606.09623v1 Announce Type: new Abstract: When running marketing campaigns, retailers must decide which products to promote and which users to target. These decisions are inherently coupled: eff

safetyarxiv-cs-lg
9 Jun 2026
Safety

Constraint-Aware Optimization for Robust Protein Stability Prediction

DGX agent

arXiv:2606.08100v1 Announce Type: new Abstract: Multimodal DeltaDelta G predictors integrating protein language models with inverse-folding representations achieve strong in-distribution accuracy on t

safetyarxiv-cs-lg
9 Jun 2026
Local Ai

Continuous Language Diffusion as a Decoder-Interface Problem

DGX agent

arXiv:2606.08810v1 Announce Type: cross Abstract: Gaussian-corrupted sentence embeddings have no direct linguistic interpretation, yet continuous diffusion language models can generate fluent text fro

local-aiarxiv-cs-lg
9 Jun 2026
Safety

Contrast encodes inductive bias: separating slow noise from dynamics in predictive representation learning

DGX agent

arXiv:2606.07770v1 Announce Type: new Abstract: Self-supervised methods that learn representations and predict dynamics fully in the latent space, such as JEPA, have been shown to confuse slowly varyi

safetyarxiv-cs-lg
9 Jun 2026
Model Releases

Convergence Bound and Critical Batch Size of Muon Optimizer

DGX agent

arXiv:2507.01598v5 Announce Type: replace Abstract: Muon, a recently proposed optimizer that leverages the inherent matrix structure of neural network parameters, has demonstrated strong empirical per

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Convolutional Sparse Coding via the Locally Competitive Algorithm on Loihi 2

DGX agent

arXiv:2606.08584v1 Announce Type: new Abstract: Sparse coding provides a principled framework for signal representation by expressing an input as a linear combination of only a small number of basis f

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Counterfactual Transport Flows for Offline Conservative Trajectory Refinement

DGX agent

arXiv:2606.09115v1 Announce Type: new Abstract: Offline reinforcement learning (RL) offers a path to policy improvement from logged data alone, using historical returns or other measurable outcomes as

model-releasesarxiv-cs-lg
9 Jun 2026
Research

Cryptographic Backdoor for Neural Networks: Boon and Bane

DGX agent

arXiv:2509.20714v2 Announce Type: replace-cross Abstract: In this paper we show that cryptographic backdoors in a neural network (NN) can be highly effective in two directions, namely mounting the att

researcharxiv-cs-lg
9 Jun 2026
Model Releases

CTS-Bench: Benchmarking Graph Coarsening Trade-offs for GNNs in Clock Tree Synthesis

DGX agent

arXiv:2602.19330v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) are increasingly explored for physical design analysis in Electronic Design Automation, particularly for modeling Clock

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Curvature-Guided LoRA: Matching Full Fine-Tuning in Function Space

DGX agent

arXiv:2603.29824v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning methods such as LoRA enable efficient adaptation of large pretrained models, but often lag behind full fine-tuning i

model-releasesarxiv-cs-lg
9 Jun 2026
Applications

Cutting LLM Evaluation Costs with SySRs: A Bandit Algorithm that Provably Exploits Model Similarity

DGX agent

arXiv:2606.07726v1 Announce Type: new Abstract: Large Language Models are typically benchmarked by evaluating every model on every test query. For practitioners seeking the best model to deploy, this

applicationsarxiv-cs-lg
9 Jun 2026
Research

Data augmented bootstrap: Unifying confidence interval construction by approximate invariance

DGX agent

arXiv:2606.09049v1 Announce Type: cross Abstract: We propose the data augmented bootstrap (DAB), a framework for constructing confidence intervals from approximately invariant transformations of the d

researcharxiv-cs-lg
9 Jun 2026
Research

Data-driven discovery of governing differential equations across physical systems

DGX agent

arXiv:2606.09638v1 Announce Type: new Abstract: Differential equations play a critical role in scientific discovery because they provide a mathematical framework to describe the behaviour of physical

researcharxiv-cs-lg
9 Jun 2026
Model Releases

De novo molecular generation with optical property preconditioning at the token level

DGX agent

arXiv:2606.08221v1 Announce Type: new Abstract: Designing OLED molecules with targeted optical properties remains challenging due to the scarcity of high-quality data and the limited reliability of co

model-releasesarxiv-cs-lg
9 Jun 2026
Research

Decentralized Online Riemannian Optimization Beyond Hadamard Manifolds

DGX agent

arXiv:2509.07779v2 Announce Type: replace-cross Abstract: We study decentralized online Riemannian optimization over manifolds with possibly positive curvature, going beyond the Hadamard manifold sett

researcharxiv-cs-lg
9 Jun 2026
Research

Decision-Focused Continual Learning for Seaport Power-Logistics Scheduling: Generalization across Varying Tasks

DGX agent

arXiv:2511.07938v3 Announce Type: replace Abstract: Power-logistics scheduling in modern seaports typically follows a predict-then-optimize pipeline. To enhance the decision quality of predictions, de

researcharxiv-cs-lg
9 Jun 2026
Model Releases

Declarative Outcome-Conformant Synthesis: Exact, Closed-Form Specification Satisfaction and a Conformance Benchmark

DGX agent

arXiv:2606.08736v1 Announce Type: new Abstract: We study a capability the dominant paradigm in synthetic tabular data does not provide: exact satisfaction of a declared analytical outcome with no sour

model-releasesarxiv-cs-lg
9 Jun 2026
Research

Decoding Naturalistic Emotion Dynamics from the Brain: An LLM-Enhanced Regression Framework

DGX agent

arXiv:2606.07707v1 Announce Type: new Abstract: Decoding emotional states from neural signals has been typically framed as a discrete, single-label classification task based on emotionally stable stim

researcharxiv-cs-lg
9 Jun 2026
Model Releases

Decomposable Neuro Symbolic Regression

DGX agent

arXiv:2511.04124v3 Announce Type: replace Abstract: Symbolic regression (SR) models complex systems by discovering mathematical expressions that capture underlying relationships in observed data. Howe

model-releasesarxiv-cs-lg
9 Jun 2026
Research

Decoy-Calibrated Failure Audits for Language Models

DGX agent

arXiv:2606.09046v1 Announce Type: new Abstract: Useful audits reveal not only how often a model fails, but also where its failures concentrate. An auditor may test many candidate explanations: long in

researcharxiv-cs-lg
9 Jun 2026
Agents

Deep reinforcement learning for process design: Review and perspective

DGX agent

arXiv:2308.07822v2 Announce Type: replace Abstract: The transformation towards renewable energy and feedstock supply in the chemical industry requires new conceptual process design approaches. Recentl

agentsarxiv-cs-lg
9 Jun 2026
Model Releases

Dendrograms of Mixing Measures for Softmax-Gated Gaussian Mixture of Experts: Consistency Without Model Sweeps

DGX agent

arXiv:2510.12744v2 Announce Type: replace-cross Abstract: We develop a unified statistical framework for softmax-gated Gaussian mixture of experts (SGMoE) that addresses three long-standing obstacles

model-releasesarxiv-cs-lg
9 Jun 2026
← Previous
1…113114115116117…304
Next →