AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
16 Apr 2026

Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

Model ReleasesDGX agent

arXiv:2604.13472v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning (MARL) is widely used to address large joint observation and action spaces by decomposing a centralized c

C-voting: Confidence-Based Test-Time Voting without Explicit Energy Functions

ResearchDGX agent

arXiv:2604.13521v1 Announce Type: new Abstract: Neural network models with latent recurrent processing, where identical layers are recursively applied to the latent state, have gained attention as pro

Can Coding Agents Be General Agents?

AgentsDGX agent

arXiv:2604.13107v1 Announce Type: cross Abstract: As coding agents have seen rapid capability and adoption gains, users are applying them to general tasks beyond software engineering. In this post, we


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Character Beyond Speech: Leveraging Role-Playing Evaluation in Audio Large Language Models via Reinforcement Learning

SafetyDGX agent

arXiv:2604.13804v1 Announce Type: new Abstract: The rapid evolution of multimodal large models has revolutionized the simulation of diverse characters in speech dialogue systems, enabling a novel inte

Complex Interpolation of Matrices with an application to Multi-Manifold Learning

ResearchDGX agent

arXiv:2604.14118v1 Announce Type: new Abstract: Given two symmetric positive-definite matrices A, B in R^{n imes n}, we study the spectral properties of the interpolation A^{1-x} B^x for 0 leq x leq 1

Composite Silhouette: A Subsampling-based Aggregation Strategy

SafetyDGX agent

arXiv:2604.13816v1 Announce Type: new Abstract: Determining the number of clusters is a central challenge in unsupervised learning, where ground-truth labels are unavailable. The Silhouette coefficien

Computational framework for multistep metabolic pathway design

ResearchDGX agent

arXiv:2604.13471v1 Announce Type: new Abstract: In silico tools are important for generating novel hypotheses and exploring alternatives in de novo metabolic pathway design. However, while many comput

Concrete Jungle: Towards Concreteness Paved Contrastive Negative Mining for Compositional Understanding

ResearchDGX agent

arXiv:2604.13313v1 Announce Type: new Abstract: Vision-Language Models demonstrate remarkable capabilities but often struggle with compositional reasoning, exhibiting vulnerabilities regarding word or

Convex Hulls of Reachable Sets

ResearchDGX agent

arXiv:2303.17674v5 Announce Type: replace-cross Abstract: We study the convex hulls of reachable sets of nonlinear systems with bounded disturbances and uncertain initial conditions. Reachable sets pl

Counterfactual Peptide Editing for Causal TCR--pMHC Binding Inference

Model ReleasesDGX agent

arXiv:2604.13256v1 Announce Type: new Abstract: Neural models for TCR-pMHC binding prediction are susceptible to shortcut learning: they exploit spurious correlations in training data -- such as pepti

Covariance-adapting algorithm for semi-bandits with application to sparse rewards

Model ReleasesDGX agent

arXiv:2604.13738v1 Announce Type: cross Abstract: We investigate stochastic combinatorial semi-bandits, where the entire joint distribution of outcomes impacts the complexity of the problem instance (

Cross-Layer Co-Optimized LSTM Accelerator for Real-Time Gait Analysis

ApplicationsDGX agent

arXiv:2604.13543v1 Announce Type: cross Abstract: Long Short-Term Memory (LSTM) neural networks have penetrated healthcare applications where real-time requirements and edge computing capabilities are

Data-driven Learning of Probabilistic Model of Binary Droplet Collision for Spray Simulation

Model ReleasesDGX agent

arXiv:2604.13594v1 Announce Type: cross Abstract: Binary droplet collisions are ubiquitous in dense sprays. Traditional deterministic models cannot adequately represent transitional and stochastic beh

Data-Efficient RLVR via Off-Policy Influence Guidance

SafetyDGX agent

arXiv:2510.26491v2 Announce Type: replace Abstract: Data selection is a critical aspect of Reinforcement Learning with Verifiable Rewards (RLVR) for enhancing the reasoning capabilities of large langu

Dataset-Level Metrics Attenuate Non-Determinism: A Fine-Grained Non-Determinism Evaluation in Diffusion Language Models

ResearchDGX agent

arXiv:2604.13413v1 Announce Type: new Abstract: Diffusion language models (DLMs) have emerged as a promising paradigm for large language models (LLMs), yet the non-deterministic behavior of DLMs remai

Decentralized Rank Scheduling for Energy-Constrained Multi-Task Federated Fine-Tuning in Edge-Assisted IoV Networks

ApplicationsDGX agent

arXiv:2508.09532v2 Announce Type: replace Abstract: Federated fine-tuning has emerged as a promising approach for adapting foundation models (FMs) to diverse downstream tasks in edge environments. In

Design Conditions for Intra-Group Learning of Sequence-Level Rewards: Token Gradient Cancellation

ResearchDGX agent

arXiv:2604.13088v1 Announce Type: new Abstract: In sparse termination rewards, intra-group comparisons have become the dominant paradigm for fine-tuning reasoning models via reinforcement learning. Ho

Design Space Exploration of Hybrid Quantum Neural Networks for Chronic Kidney Disease

Model ReleasesDGX agent

arXiv:2604.13608v1 Announce Type: new Abstract: Hybrid Quantum Neural Networks (HQNNs) have recently emerged as a promising paradigm for near-term quantum machine learning. However, their practical pe

Diagnostics for Individual-Level Prediction Instability in Machine Learning for Healthcare

ApplicationsDGX agent

arXiv:2603.00192v2 Announce Type: replace Abstract: In healthcare, predictive models increasingly inform patient-level decisions, yet little attention is paid to the variability in individual risk est

Diffusion Sequence Models for Generative In-Context Meta-Learning of Robot Dynamics

ResearchDGX agent

arXiv:2604.13366v1 Announce Type: new Abstract: Accurate modeling of robot dynamics is essential for model-based control, yet remains challenging under distributional shifts and real-time constraints.

DiPO: Disentangled Perplexity Policy Optimization for Fine-grained Exploration-Exploitation Trade-Off

SafetyDGX agent

arXiv:2604.13902v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has catalyzed significant advances in the reasoning capabilities of Large Language Models (LLMs).

Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching

SafetyDGX agent

arXiv:2509.21912v2 Announce Type: replace Abstract: Guidance provides a simple and effective framework for posterior sampling by steering the generation process towards the desired distribution. When

Does Dimensionality Reduction via Random Projections Preserve Landscape Features?

ResearchDGX agent

arXiv:2604.13230v1 Announce Type: new Abstract: Exploratory Landscape Analysis (ELA) provides numerical features for characterizing black-box optimization problems. In high-dimensional settings, howev

Driving Engagement in Daily Fantasy Sports with a Scalable and Urgency-Aware Ranking Engine

Local AiDGX agent

arXiv:2604.13796v1 Announce Type: cross Abstract: In daily fantasy sports (DFS), match participation is highly time-sensitive. Users must act within a narrow window before a game begins, making match

Drowsiness-Aware Adaptive Autonomous Braking System based on Deep Reinforcement Learning for Enhanced Road Safety

Model ReleasesDGX agent

arXiv:2604.13878v1 Announce Type: new Abstract: Driver drowsiness significantly impairs the ability to accurately judge safe braking distances and is estimated to contribute to 10%-20% of road acciden

EMGFlow: Robust and Efficient Surface Electromyography Synthesis via Flow Matching

Model ReleasesDGX agent

arXiv:2604.13685v1 Announce Type: cross Abstract: Deep learning-based surface electromyography (sEMG) gesture recognition is frequently bottlenecked by data scarcity and limited subject diversity. Whi

Empowering Targeted Neighborhood Search via Hyper Tour for Large-Scale TSP

TutorialsDGX agent

arXiv:2510.20169v3 Announce Type: replace Abstract: Traveling Salesman Problem (TSP) is a classic NP-hard problem that has garnered significant attention from both academia and industry. While neural-

Enhancing Confidence Estimation in Telco LLMs via Twin-Pass CoT-Ensembling

Model ReleasesDGX agent

arXiv:2604.13271v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly applied to complex telecommunications tasks, including 3GPP specification analysis and O-RAN network troub

Enhancing Reinforcement Learning for Radiology Report Generation with Evidence-aware Rewards and Self-correcting Preference Learning

SafetyDGX agent

arXiv:2604.13598v1 Announce Type: new Abstract: Recent reinforcement learning (RL) approaches have advanced radiology report generation (RRG), yet two core limitations persist: (1) report-level reward

Estimating Continuous Treatment Effects with Two-Stage Kernel Ridge Regression

SafetyDGX agent

arXiv:2604.13410v1 Announce Type: cross Abstract: We study the problem of estimating the effect function for a continuous treatment, which maps each treatment value to a population-averaged outcome. A

Evaluating Supervised Machine Learning Models: Principles, Pitfalls, and Metric Selection

Model ReleasesDGX agent

arXiv:2604.13882v1 Announce Type: new Abstract: The evaluation of supervised machine learning models is a critical stage in the development of reliable predictive systems. Despite the widespread avail

Event Tensor: A Unified Abstraction for Compiling Dynamic Megakernel

HardwareDGX agent

arXiv:2604.13327v1 Announce Type: cross Abstract: Modern GPU workloads, especially large language model (LLM) inference, suffer from kernel launch overheads and coarse synchronization that limit inter

Exploring Urban Land Use Patterns by Pattern Mining and Unsupervised Learning

ResearchDGX agent

arXiv:2604.13050v1 Announce Type: cross Abstract: Urban areas are intricate systems shaped by socioeconomic, environmental, and infrastructural factors, with land use patterns serving as aspects of ur

FAST: A Synergistic Framework of Attention and State-space Models for Spatiotemporal Traffic Prediction

ResearchDGX agent

arXiv:2604.13453v1 Announce Type: new Abstract: Traffic forecasting requires modeling complex temporal dynamics and long-range spatial dependencies over large sensor networks. Existing methods typical

Fast and Simple Densest Subgraph with Predictions

ApplicationsDGX agent

arXiv:2505.12600v3 Announce Type: replace-cross Abstract: We study the densest subgraph problem and its NP-hard densest at-most-k subgraph variant through the lens of learning-augmented algorithms. We

Fast training of accurate physics-informed neural networks without gradient descent

Model ReleasesDGX agent

arXiv:2405.20836v3 Announce Type: replace-cross Abstract: Solving time-dependent Partial Differential Equations (PDEs) is one of the most critical problems in computational science. While Physics-Info

Fast Voxelization and Level of Detail for Microgeometry Rendering

Local AiDGX agent

arXiv:2604.13191v1 Announce Type: cross Abstract: Many materials show anisotropic light scattering patterns due to the shape and local alignment of their underlying micro structures: surfaces with sma

First-See-Then-Design: A Multi-Stakeholder View for Optimal Performance-Fairness Trade-Offs

SafetyDGX agent

arXiv:2604.14035v1 Announce Type: new Abstract: Fairness in algorithmic decision-making is often defined in the predictive space, where predictive performance - used as a proxy for decision-maker (DM)

FlexGuard: Continuous Risk Scoring for Strictness-Adaptive LLM Content Moderation

Model ReleasesDGX agent

arXiv:2602.23636v3 Announce Type: replace Abstract: Ensuring the safety of LLM-generated content is essential for real-world deployment. Most existing guardrail models formulate moderation as a fixed

Flow-based Generative Modeling of Potential Outcomes and Counterfactuals

Model ReleasesDGX agent

arXiv:2505.16051v4 Announce Type: replace-cross Abstract: Predicting potential and counterfactual outcomes from observational data is central to individualized decision-making, particularly in clinica

Fluids You Can Trust: Property-Preserving Operator Learning for Incompressible Flows

HardwareDGX agent

arXiv:2602.15472v4 Announce Type: replace-cross Abstract: We present a novel property-preserving kernel-based operator learning method for incompressible flows governed by the incompressible Navier--S

From Alignment to Prediction: A Study of Self-Supervised Learning and Predictive Representation Learning

SafetyDGX agent

arXiv:2604.13518v1 Announce Type: new Abstract: Self-supervised learning has emerged as a major technique for the task of learning from unlabeled data, where the current methods mostly revolve around

From Order to Distribution: A Spectral Characterization of Forgetting in Continual Learning

ResearchDGX agent

arXiv:2604.13460v1 Announce Type: new Abstract: A central challenge in continual learning is forgetting, the loss of performance on previously learned tasks induced by sequential adaptation to new one

Geminet: Learning the Duality-based Iterative Process for Lightweight Traffic Engineering in Changing Topologies

ResearchDGX agent

arXiv:2506.23640v2 Announce Type: replace-cross Abstract: Recently, researchers have explored ML-based Traffic Engineering (TE), leveraging neural networks to solve TE problems traditionally addressed

Generalization Guarantees on Data-Driven Tuning of Gradient Descent with Langevin Updates

TutorialsDGX agent

arXiv:2604.13130v1 Announce Type: new Abstract: We study learning to learn for regression problems through the lens of hyperparameter tuning. We propose the Langevin Gradient Descent Algorithm (LGD),

Golden Handcuffs make safer AI agents

SafetyDGX agent

arXiv:2604.13609v1 Announce Type: new Abstract: Reinforcement learners can attain high reward through novel unintended strategies. We study a Bayesian mitigation for general environments: we expand th

Gradient Descent's Last Iterate is Often (slightly) Suboptimal

ResearchDGX agent

arXiv:2604.13870v1 Announce Type: cross Abstract: We consider the well-studied setting of minimizing a convex Lipschitz function using either gradient descent (GD) or its stochastic variant (SGD), and

Graph In-Context Operator Networks for Generalizable Spatiotemporal Prediction

ApplicationsDGX agent

arXiv:2603.12725v3 Announce Type: replace Abstract: In-context operator learning enables neural networks to infer solution operators from contextual examples without weight updates. While prior work h

Guided Transfer Learning for Discrete Diffusion Models

ApplicationsDGX agent

arXiv:2512.10877v4 Announce Type: replace Abstract: Discrete diffusion models (DMs) have achieved strong performance in language and other discrete domains, offering a compelling alternative to autore

Hardware-Efficient Neuro-Symbolic Networks with the Exp-Minus-Log Operator

SafetyDGX agent

arXiv:2604.13871v1 Announce Type: new Abstract: Deep neural networks (DNNs) deliver state-of-the-art accuracy on regression and classification tasks, yet two structural deficits persistently obstruct

Hierarchical DLO Routing with Reinforcement Learning and In-Context Vision-language Models

AgentsDGX agent

arXiv:2510.19268v2 Announce Type: replace-cross Abstract: Long-horizon routing tasks of deformable linear objects (DLOs), such as cables and ropes, are common in industrial assembly lines and everyday

Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM Agents

ResearchDGX agent

arXiv:2604.05808v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents have demonstrated strong capabilities in complex interactive decision-making tasks. However, existing LLM ag

Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation

Model ReleasesDGX agent

arXiv:2604.14032v1 Announce Type: cross Abstract: Reinforcement learning has shown promise for automating power-grid operation tasks such as topology control and congestion management. However, its de

HINTBench: Horizon-agent Intrinsic Non-attack Trajectory Benchmark

Model ReleasesDGX agent

arXiv:2604.13954v1 Announce Type: new Abstract: Existing agent-safety evaluation has focused mainly on externally induced risks. Yet agents may still enter unsafe trajectories under benign conditions.

HUANet: Hard-Constrained Unrolled ADMM for Constrained Convex Optimization

ResearchDGX agent

arXiv:2604.13179v1 Announce Type: cross Abstract: This paper presents HUANet, a constrained deep neural network architecture that unrolls the iterations of the Alternating Direction Method of Multipli

Hybrid Attention Model Using Feature Decomposition and Knowledge Distillation for Glucose Forecasting

ApplicationsDGX agent

arXiv:2411.10703v3 Announce Type: replace Abstract: The availability of continuous glucose monitors as over-the-counter commodities have created a unique opportunity to monitor a person's blood glucos

ID and Graph View Contrastive Learning with Multi-View Attention Fusion for Sequential Recommendation

Model ReleasesDGX agent

arXiv:2604.14114v1 Announce Type: cross Abstract: Sequential recommendation has become increasingly prominent in both academia and industry, particularly in e-commerce. The primary goal is to extract

Identifiability of Potentially Degenerate Gaussian Mixture Models With Piecewise Affine Mixing

ResearchDGX agent

arXiv:2604.13218v1 Announce Type: cross Abstract: Causal representation learning (CRL) aims to identify the underlying latent variables from high-dimensional observations, even when variables are depe

Irregularly Sampled Time Series Interpolation for Binary Evolution Simulations Using Dynamic Time Warping

SafetyDGX agent

arXiv:2604.13604v1 Announce Type: cross Abstract: Binary stellar evolution simulations are computationally expensive. Stellar population synthesis relies on these detailed evolution models at a fundam

Joint Representation Learning and Clustering via Gradient-Based Manifold Optimization

Model ReleasesDGX agent

arXiv:2604.13484v1 Announce Type: cross Abstract: Clustering and dimensionality reduction have been crucial topics in machine learning and computer vision. Clustering high-dimensional data has been ch

← Previous
1…224225226227228…239
Next →