AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
30 Jul 2026

From Conceptual Hydrologic Models to Conceptually Interpretable Neural Networks: A Snow-Water Mass-Conserving-Perceptron Framework for Discovering Catchment-Scale Precipitation-Storage-Runoff Representations

AgentsDGX agent

arXiv:2607.26492v1 Announce Type: new Abstract: The Mass-Conserving Perceptron (MCP) establishes a modeling paradigm in which conceptual hydrologic models can be reformulated as physically constrained

From Interface to Inference: Eliciting Any-Order Inference from Any-Order Models

Local AiDGX agent

arXiv:2607.26504v1 Announce Type: new Abstract: Many discrete reasoning tasks, such as code generation, are inherently non-causal: programmers move between high-level structure and local details, a pr

From Tokens to Watt-hours: Analytical Energy Estimation for LLM Inference on Modern GPUs


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2607.26571v1 Announce Type: new Abstract: The operational energy consumption of large language model (LLM) inference is becoming an increasingly important component of the environmental footprin

From Unsupervised Subgroups to Hypothetical State-Intervention Policies: An Evaluation of Selected Subgrouping Methods in Observational Health Data

SafetyDGX agent

arXiv:2607.26521v1 Announce Type: new Abstract: Conventional subgroup analyses can yield unstable and difficult-to-interpret conclusions, especially in observational biomedical data where each individ

Gated Adaptation for Continual Learning in Human Activity Recognition

Model ReleasesDGX agent

arXiv:2603.10046v2 Announce Type: replace Abstract: Wearable sensors in Internet of Things (IoT) ecosystems increasingly support applications such as remote health monitoring, elderly care, and smart

GEqTrain: A Configuration-Driven Framework for Retargeting Equivariant Graph Neural Networks Across 3D Scientific Tasks

Model ReleasesDGX agent

arXiv:2607.19083v2 Announce Type: replace Abstract: Equivariant graph neural networks provide a powerful modeling language for three-dimensional scientific data, but their reuse is often limited by im

GPTQ-2D: Cubic-Time Two-Sided Adaptive Rounding

ResearchDGX agent

arXiv:2607.27042v1 Announce Type: cross Abstract: Adaptive rounding methods such as GPTQ, or equivalently Babai's nearest plane algorithm, round a real matrix to integers under a quadratic metric. The

Graph Signal Diffusion Models for Wireless Resource Allocation

SafetyDGX agent

arXiv:2604.05175v2 Announce Type: replace-cross Abstract: We consider constrained ergodic resource optimization in wireless networks with graph-structured interference. We train a diffusion model poli

Harnessing Large Language Models for Intelligent Resource Allocation in the Internet of Everything

ResearchDGX agent

arXiv:2607.26602v1 Announce Type: cross Abstract: The rapid development of the Internet of Everything (IoE) is accelerating the adoption of intelligent applications. However, the massive number of con

HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring

ApplicationsDGX agent

arXiv:2509.07260v5 Announce Type: replace-cross Abstract: Mobile and wearable healthcare monitoring play a vital role in facilitating timely interventions, managing chronic health conditions, and ulti

Hierarchical Spatio-Temporal Transformer for Coherent Emergency Department Forecasting

Local AiDGX agent

arXiv:2607.27106v1 Announce Type: new Abstract: Emergency Departments (EDs) are critical access points in healthcare systems, yet they face persistent pressure from unpredictable patient demand, seaso

HiFloat4 Format for End-To-End Reinforcement Learning Post-Training of Large Language Models

SafetyDGX agent

arXiv:2607.26515v1 Announce Type: new Abstract: We present, to our knowledge, the first end-to-end FP4 RL post-training, in which both the rollout and training policies, including their forward and ba

High-Order Markov Blanket Discovery via a k-Order Relaxation of the Faithfulness Assumption

ResearchDGX agent

arXiv:2607.26357v1 Announce Type: new Abstract: The problem of learning the graphical Markov blanket (MB) of a variable from data has applications in many areas such as structure learning for Bayesian

HoF-Bench: Rediscovering Real AI-Discovered CVEs Without Frontier Models

Model ReleasesDGX agent

arXiv:2607.27030v1 Announce Type: cross Abstract: LLM-based analyzers have begun finding real vulnerabilities in mature open-source projects: AISLE's analyzer is credited with more than 280 CVEs acros

Incast-Free MoE Rate-Based Scheduling

ResearchDGX agent

arXiv:2607.26340v1 Announce Type: cross Abstract: Mixture of Experts (MoE) architectures have become key to large language models; however, their typical round-robin (RR) scheduling introduces signifi

InferScale: GPU-Native KV Injection for Personalized LLM Serving

Model ReleasesDGX agent

arXiv:2607.27090v1 Announce Type: cross Abstract: Large language models are increasingly deployed with persistent personalized context, such as accumulated memory profiles or long conversation histori

Inverse Learning of Latent Risk-Neutral Densities from Irregular Option Quotes

Model ReleasesDGX agent

arXiv:2607.27188v1 Announce Type: new Abstract: Accurate option prices do not imply accurate recovery of the latent risk-neutral density. We study this distinction with two complementary benchmarks. A

Investigating reservoir computing for branch predictionin pipelined processors using emerging CMOS memristor devices

Model ReleasesDGX agent

arXiv:2607.27140v1 Announce Type: cross Abstract: This project aimed to develop a novel reservoir compute (RC) implementation framework targeting high-speed operation and integration with CMOS digital

Journey Operators for Structured Multi-Axis Composition

SafetyDGX agent

arXiv:2607.26775v1 Announce Type: new Abstract: Many kinds of data have structure along one or more axes: words in a sentence, pixels in an image, nodes in a tree, frames in audio, or cells in a 3D vo

Kairos: Numerically Robust News Recommendation under Item Cold-Start via Cholesky-based LinUCB

ApplicationsDGX agent

arXiv:2607.26832v1 Announce Type: new Abstract: Algorithmic news personalization in regional markets often fails because modern deep learning models require massive interaction data while real-world n

Learning Controlled Stochastic Differential Equations

ResearchDGX agent

arXiv:2411.01982v2 Announce Type: replace-cross Abstract: We study the problem of learning controlled stochastic differential equations (SDEs) [ dX_t = b(t,X_t,u_t),dt + sigma(t,X_t,u_t),dW_t, ] whose

Learning Implicit Causal World Models from Multi-Agent Demonstrations

SafetyDGX agent

arXiv:2607.26336v1 Announce Type: new Abstract: In model-based reinforcement learning, world models exist as internal simulators, but their training often conflates statistical correlations with causa

Learning the Word Problem: Geodesic Lengths and Cryptographic Applications

ResearchDGX agent

arXiv:2607.26241v1 Announce Type: cross Abstract: The Word Problem has been a subject of intensive mathematical study for over a century, initially driving advances in combinatorial group theory and m

Lilith: Backdoor Generalization under Training-Inference Trigger Shift

SafetyDGX agent

arXiv:2607.26099v1 Announce Type: cross Abstract: Machine-learning services increasingly rely on public data, third-party providers, and outsourced training, creating opportunities for data-poisoning

LLMET: Enabling Cross-Layer Evaluation of Emerging M3D Memories for Energy-Efficient LLM Serving

Model ReleasesDGX agent

arXiv:2607.26491v1 Announce Type: cross Abstract: The energy consumption of Large Language Model (LLM) serving is becoming a major system challenge as deployment scales, driven by hardware power and t

Low-cost Embedded Breathing Rate Determination Using 802.15.4z IR-UWB Hardware for Remote Healthcare

Model ReleasesDGX agent

arXiv:2504.03772v3 Announce Type: replace-cross Abstract: Respiratory diseases account for a significant portion of global mortality. Affordable and early detection is an effective way of addressing t

Low-Precision Training of Large Language Models: Methods, Challenges, and Opportunities

ResearchDGX agent

arXiv:2505.01043v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved impressive performance across various domains. However, the substantial hardware resources required for t

MetaKoopman: Bayesian Meta-Learning of Koopman Operators for Modeling Structured Dynamics under Distribution Shifts

AgentsDGX agent

arXiv:2607.26345v1 Announce Type: new Abstract: Modeling and forecasting nonlinear dynamics under distribution shifts is essential for robust decision-making in real-world systems. In this work, we pr

Minimal Markovization via Stable Quotients in Holonomy-Cover Decision Processes

AgentsDGX agent

arXiv:2607.27132v1 Announce Type: new Abstract: An agent acting under partial observability must retain a recursively updateable statistic of history that restores the Markov property, but the smalles

Mixture-of-experts for handwriting trajectory reconstruction from IMU sensors

Model ReleasesDGX agent

arXiv:2607.26708v1 Announce Type: new Abstract: The use of digital pens for online handwriting trajectory reconstruction is a prevalent method for human-computer interaction. In this study, we focus o

Neural Architecture Search for Traffic Prediction: A Survey of Methods, Challenges, and Future Directions

ResearchDGX agent

arXiv:2607.26467v1 Announce Type: new Abstract: Traffic prediction is a core task in intelligent transportation systems, supporting applications such as adaptive signal control, route guidance, and ri

No Data Is Not No Risk: Visibility Aware Graph-Based Inference of Business Conduct Risk

ResearchDGX agent

arXiv:2607.26859v1 Announce Type: cross Abstract: The monitoring of business conduct risk is hindered by sparse, uneven, and visibility-biased data. Prior studies show that business conduct risk infor

On the Rademacher Complexity of Graph Neural Networks: Unifying Expressivity and Geometry

ResearchDGX agent

arXiv:2510.10101v4 Announce Type: replace Abstract: Understanding the interplay between generalization, expressivity, and the geometry of the input space is a central challenge in graph learning. The

On the robustness of noisy solutions in non-convex neural networks

ResearchDGX agent

arXiv:2607.27000v1 Announce Type: cross Abstract: Optimization in non-convex neural network models is strongly influenced by the geometry of the solution space: sparse, isolated, point-like clusters a

Ontology-driven personalized information retrieval for XML documents

ResearchDGX agent

arXiv:2603.21139v2 Announce Type: replace-cross Abstract: This paper addresses the challenge of improving information retrieval from semi-structured eXtensible Markup Language (XML) documents. Traditi

Origins and mitigation of double descent in reduced order modeling

Local AiDGX agent

arXiv:2607.26414v1 Announce Type: cross Abstract: Latent low-dimensional structure in datasets of natural and engineered systems enables their sparse sensing, or full-state reconstruction from histori

Parameter-Free Dynamic Regret for Online Convex Optimization under Heavy-Tailed Noise

Model ReleasesDGX agent

arXiv:2607.27073v1 Announce Type: new Abstract: We study online convex optimization (OCO) in non-stationary environments under heavy-tailed noise, where the stochastic gradient oracle admits only a fi

Parameterized Fair Resource Allocation under Diversity Constraints

SafetyDGX agent

arXiv:2607.26485v1 Announce Type: cross Abstract: Resource allocation across multiple agent groups arises in many applications including e-commerce recommendation systems, housing assignment, and cour

Persistence Spheres: a Bi-continuous Linear Representation of Measures for Partial Optimal Transport

Model ReleasesDGX agent

arXiv:2603.15384v2 Announce Type: replace-cross Abstract: We improve and extend persistence spheres, introduced in~ite{pegoraro2025persistence}. Persistence spheres map an integrable measure mu on the

Physics-Informed Graph Neural Networks for Robust AC-Optimal Power Flow

ResearchDGX agent

arXiv:2410.04818v2 Announce Type: replace-cross Abstract: We present PINCO, an unsupervised learning framework that integrates Graph Neural Networks with physics-informed neural networks for AC optima

Physics-Informed Singular-Value Learning for Cross-Covariances Forecasting in Financial Markets

ResearchDGX agent

arXiv:2601.07687v3 Announce Type: replace-cross Abstract: Recent advances in nonlinear shrinkage yield asymptotically optimal cleaners for large covariance matrices and have been extended to empirical

PIKS: Universal Physics-Informed Kernel Methods

ResearchDGX agent

arXiv:2607.27062v1 Announce Type: cross Abstract: Physics-informed machine learning incorporates physical principles --often expressed via differential operators-- into data-driven models. While physi

Post-Training at the Edge of Detectability: A Game-Theoretic Approach to Fine-Tuning

Model ReleasesDGX agent

arXiv:2607.26358v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning is widely used in language model training to improve model performance on a target task while limiting drift fro

PowerAtlas: Towards Electricity-Computing Co-Scheduling for Power Systems

Model ReleasesDGX agent

arXiv:2607.26710v1 Announce Type: new Abstract: The rapid growth of AI workloads is turning data centers into large-scale, volatile, yet spatiotemporally flexible grid loads, creating an urgent need f

Projective Graph Residualization: Variation-Allocation Frontiers for Control-Function IV

Model ReleasesDGX agent

arXiv:2606.14636v2 Announce Type: replace Abstract: Control-function instrumental-variable estimators pass an estimated first-stage residual to an outcome model. The residual must retain the latent co

Q-Steer: Action-Value Guidance for Molecular Policy Optimization

Local AiDGX agent

arXiv:2607.26391v1 Announce Type: new Abstract: Oracle-limited molecular optimization gives reward only after a complete molecule is generated, while each rollout requires many local next-token decisi

RAG-HAR+: Towards Cost-Efficient LLM-Based Human Activity Recognition for Edge Deployment

AgentsDGX agent

arXiv:2607.26631v1 Announce Type: new Abstract: Human Activity Recognition (HAR) from wearable sensors supports applications in healthcare, rehabilitation, fitness tracking, and smart environments. Ye

RAGuard: A Layered Defense Framework for Retrieval-Augmented Generation Systems Against Data Poisoning

Model ReleasesDGX agent

arXiv:2607.26339v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems ground large language models (LLMs) in external corpora, but this reliance exposes them to corpus poisoning

Randomizing the Number of Centers in k-means++

ResearchDGX agent

arXiv:2607.26202v1 Announce Type: cross Abstract: The k-means++ algorithm is a standard and widely used seeding method for k-means clustering, but for a fixed number k of centers its worst-case expect

ReCo: Reweighting GRPO Against Distributional Concentration

Model ReleasesDGX agent

arXiv:2607.26862v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has become a standard reinforcement learning method for post-training language models. Recent work shows that

Recover, Decode, Reguard: Guard-Agnostic Defense Amplification againstEncoded VLM Jailbreaks

SafetyDGX agent

arXiv:2607.26574v1 Announce Type: cross Abstract: Safety classifiers ('guards') are the dominant black-box defense for vision-language models, yet they judge an input's surface form, not its meaning:

ReDiSC: A Reparameterized Masked Diffusion Model for Scalable Node Classification with Structured Predictions

ResearchDGX agent

arXiv:2507.14484v2 Announce Type: replace Abstract: In recent years, graph neural networks (GNN) have achieved unprecedented successes in node classification tasks. Although GNNs inherently encode spe

ResearchArena: Evaluating Sabotage and Monitoring in Automated AI R&D

SafetyDGX agent

arXiv:2607.19321v2 Announce Type: replace-cross Abstract: As AI agents begin to automate AI R&D, we need ways to assess whether their outputs are safe to deploy, even when the agents themselves may be

Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting

Model ReleasesDGX agent

arXiv:2607.26643v1 Announce Type: cross Abstract: Enabling large language model (LLM) agents to accumulate and reuse experience from past interactions remains a central challenge in real-world applica

Retrospective Orthogonal Design: Response-Surface Reconstruction from Observational Data

SafetyDGX agent

arXiv:2607.26219v1 Announce Type: cross Abstract: Regression estimates from observational data can depend on specification under multicollinearity, while sequential sums of squares (SS) depend on term

Scores Are Not Decisions: Cost-Aware Stopping for Tool Acquisition in LLM Agents

AgentsDGX agent

arXiv:2607.27083v1 Announce Type: new Abstract: As LLM agents increasingly depend on diverse external services such as search engines, databases, and connectors, agent harnesses face a fundamental too

SCOUT: Per-Context Reset Curricula for Sparse-Reward Reinforcement Learning

TutorialsDGX agent

arXiv:2607.26417v1 Announce Type: new Abstract: Sparse-reward reinforcement learning often fails because rollouts from the unassisted evaluation start rarely reach later task stages. Reset curricula a

Self-Adaptive Learning and Model Predictive Control for Tracking Unknown Dynamics with No Regret

SafetyDGX agent

arXiv:2607.26370v1 Announce Type: cross Abstract: We propose a self-adaptive online learning for control method for tracking unknown target dynamics. The target dynamics can exhibit switching behavior

SENSE: Efficient EEG-to-Text via Privacy-Preserving Semantic Retrieval

Local AiDGX agent

arXiv:2603.17109v2 Announce Type: replace Abstract: Decoding brain activity into natural language is a major challenge in AI with important applications in assistive communication, neurotechnology, an

Shared SFT Lessons Across Alignment, Model Organisms, and Toy Models

Model ReleasesDGX agent

arXiv:2607.26173v1 Announce Type: new Abstract: Alignment training, model organisms, and toy models are usually treated as separate research areas. But projects in all three frequently use supervised

← Previous
1…2526272829…241
Next →