AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
17 Apr 2026

Curvature-Aligned Probing for Local Loss-Landscape Stabilization

Model ReleasesDGX agent

arXiv:2604.14870v1 Announce Type: new Abstract: Local loss-landscape stabilization under sample growth is typically measured either pointwise or through isotropic averaging in the full parameter space

DEEP-GAP: Deep-learning Evaluation of Execution Parallelism in GPU Architectural Performance

HardwareDGX agent

arXiv:2604.14552v1 Announce Type: cross Abstract: Modern datacenters increasingly rely on low-power, single-slot inference accelerators to balance performance, energy efficiency, and rack density cons

Deep Learning for Subspace Regression

ResearchDGX agent

arXiv:2509.23249v4 Announce Type: replace Abstract: It is often possible to perform reduced order modelling by specifying linear subspace which accurately captures the dynamics of the system. This app


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Dense Neural Networks are not Universal Approximators

ResearchDGX agent

arXiv:2602.07618v3 Announce Type: replace Abstract: We investigate the approximation capabilities of dense neural networks. While universal approximation theorems establish that sufficiently large arc

Deployment of AI-Assisted Interventions: Capacity Constraints and Noisy Compliance

TutorialsDGX agent

arXiv:2604.14370v1 Announce Type: cross Abstract: AI tools increasingly guide targeted interventions in healthcare, education, and recruiting. Algorithms score individuals, trigger outreach to those a

Differentially Private Conformal Prediction

TutorialsDGX agent

arXiv:2604.14621v1 Announce Type: cross Abstract: Conformal prediction (CP) has attracted broad attention as a simple and flexible framework for uncertainty quantification through prediction sets. In

Direct Preference Optimization for Primitive-Enabled Hierarchical RL: A Bilevel Approach

SafetyDGX agent

arXiv:2411.00361v4 Announce Type: replace Abstract: Hierarchical reinforcement learning (HRL) enables agents to solve complex, long-horizon tasks by decomposing them into manageable sub-tasks. However

DLink: Distilling Layer-wise and Dominant Knowledge from EEG Foundation Models

ResearchDGX agent

arXiv:2604.15016v1 Announce Type: new Abstract: EEG foundation models (FMs) achieve strong cross-subject and cross-task generalization but impose substantial computational and memory costs that hinder

Do Not Step Into the Same River Twice: Learning to Reason from Trial and Error

SafetyDGX agent

arXiv:2510.26109v4 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly boosted the reasoning capability of language models (LMs). However, existing

Does RL Expand the Capability Boundary of LLM Agents? A PASS@(k,T) Analysis

AgentsDGX agent

arXiv:2604.14877v1 Announce Type: new Abstract: Does reinforcement learning genuinely expand what LLM agents can do, or merely make them more reliable? For static reasoning, recent work answers the se

Doubly Outlier-Robust Online Infinite Hidden Markov Model

ResearchDGX agent

arXiv:2604.14322v1 Announce Type: cross Abstract: We derive a robust update rule for the online infinite hidden Markov model (iHMM) for when the streaming data contains outliers and the model is missp

DPQuant: Efficient and Differentially-Private Model Training via Dynamic Quantization Scheduling

ResearchDGX agent

arXiv:2509.03472v2 Announce Type: replace Abstract: Differentially-Private SGD (DP-SGD) and its adaptive variant DP-Adam are powerful techniques to protect user privacy when using sensitive data to tr

DPSQL+: A Differentially Private SQL Library with a Minimum Frequency Rule

Model ReleasesDGX agent

arXiv:2602.22699v2 Announce Type: replace-cross Abstract: SQL is the de facto interface for exploratory data analysis; however, releasing exact query results can expose sensitive information through m

EEGDM: Learning EEG Representation with Latent Diffusion Model

Local AiDGX agent

arXiv:2508.20705v3 Announce Type: replace Abstract: Recent advances in self-supervised learning for EEG representation have largely relied on masked reconstruction, where models are trained to recover

ELMoE-3D: Leveraging Intrinsic Elasticity of MoE for Hybrid-Bonding-Enabled Self-Speculative Decoding in On-Premises Serving

ResearchDGX agent

arXiv:2604.14626v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have become the dominant architecture for large-scale language models, yet on-premises serving remains fundamentally mem

Enabling Agents to Communicate Entirely in Latent Space

AgentsDGX agent

arXiv:2511.09149v4 Announce Type: replace Abstract: While natural language is the de facto communication medium for LLM-based agents, it presents a fundamental constraint. The process of downsampling

Enhancing LLM-based Search Agents via Contribution Weighted Group Relative Policy Optimization

SafetyDGX agent

arXiv:2604.14267v1 Announce Type: new Abstract: Search agents extend Large Language Models (LLMs) beyond static parametric knowledge by enabling access to up-to-date and long-tail information unavaila

Expert-Guided Class-Conditional Goodness-of-Fit Scores for Interpretable Classification with Informative Missingness: An Application to Seismic Monitoring

ResearchDGX agent

arXiv:2604.14809v1 Announce Type: cross Abstract: We study a classification problem with three key challenges: pervasive informative missingness, the integration of partial prior expert knowledge into

Explainable Graph Neural Networks for Interbank Contagion Surveillance: A Regulatory-Aligned Framework for the U.S. Banking Sector

Model ReleasesDGX agent

arXiv:2604.14232v1 Announce Type: new Abstract: The Spatial-Temporal Graph Attention Network (ST-GAT) framework was created to serve as an explainable GNN-based solution for detecting bank distress ea

Exploiting Correlations in Federated Learning: Opportunities and Practical Limitations

Model ReleasesDGX agent

arXiv:2604.14751v1 Announce Type: cross Abstract: The communication bottleneck in federated learning (FL) has spurred extensive research into techniques to reduce the volume of data exchanged between

Exploring the flavor structure of leptons via diffusion models

ResearchDGX agent

arXiv:2503.21432v2 Announce Type: replace-cross Abstract: We propose a method to explore the flavor structure of leptons using diffusion models, which are known as one of generative artificial intelli

Expressivity of Transformers: A Tropical Geometry Perspective

ResearchDGX agent

arXiv:2604.14727v1 Announce Type: new Abstract: To quantify the geometric expressivity of transformers, we introduce a tropical geometry framework to characterize their exact spatial partitioning capa

Fall into a Pit, Gain in a Wit: Cognitive-Guided Harmful Meme Detection via Misjudgment Risk Pattern Retrieval

Model ReleasesDGX agent

arXiv:2510.15946v3 Announce Type: replace Abstract: Internet memes have emerged as a popular multimodal medium, yet they are increasingly weaponized to convey harmful opinions through subtle rhetorica

Federated Multi-Task Clustering

TutorialsDGX agent

arXiv:2512.22897v3 Announce Type: replace Abstract: Spectral clustering has emerged as one of the most effective clustering algorithms due to its superior performance. However, most existing models ar

FedIDM: Achieving Fast and Stable Convergence in Byzantine Federated Learning through Iterative Distribution Matching

Model ReleasesDGX agent

arXiv:2604.15115v1 Announce Type: new Abstract: Most existing Byzantine-robust federated learning (FL) methods suffer from slow and unstable convergence. Moreover, when handling a substantial proporti

Flow with the Force Field: Learning 3D Compliant Flow Matching Policies from Force and Demonstration-Guided Simulation Data

SafetyDGX agent

arXiv:2510.02738v3 Announce Type: replace-cross Abstract: While visuomotor policy has made advancements in recent years, contact-rich tasks still remain a challenge. Robotic manipulation tasks that re

From Risk to Rescue: An Agentic Survival Analysis Framework for Liquidation Prevention

SafetyDGX agent

arXiv:2604.14583v1 Announce Type: new Abstract: Decentralized Finance (DeFi) lending protocols like Aave v3 rely on over-collateralization to secure loans, yet users frequently face liquidation due to

From Tokens to Layers: Redefining Stall-Free Scheduling for MoE Serving with Layered Prefill

ApplicationsDGX agent

arXiv:2510.08055v2 Announce Type: replace Abstract: Large Language Model (LLM) inference in production must meet stringent service-level objectives for both time-to-first-token (TTFT) and time-between

Fundamental Limitations of Favorable Privacy-Utility Guarantees for DP-SGD

ResearchDGX agent

arXiv:2601.10237v2 Announce Type: replace Abstract: Differentially Private Stochastic Gradient Descent (DP-SGD) is the dominant paradigm for private training, but its fundamental limitations under wor

Gating Enables Curvature: A Geometric Expressivity Gap in Attention

ResearchDGX agent

arXiv:2604.14702v1 Announce Type: new Abstract: Multiplicative gating is widely used in neural architectures and has recently been applied to attention layers to improve performance and training stabi

Gaussian Process Regression of Steering Vectors With Physics-Aware Deep Composite Kernels for Augmented Listening

ResearchDGX agent

arXiv:2509.02571v2 Announce Type: replace-cross Abstract: This paper investigates continuous representations of steering vectors over frequency and microphone/source positions for augmented listening

Generalization in LLM Problem Solving: The Case of the Shortest Path

ResearchDGX agent

arXiv:2604.15306v1 Announce Type: cross Abstract: Whether language models can systematically generalize remains actively debated. Yet empirical performance is jointly shaped by multiple factors such a

Generative Augmented Inference

ResearchDGX agent

arXiv:2604.14575v1 Announce Type: new Abstract: Data-driven operations management often relies on parameters estimated from costly human-generated labels. Recent advances in large language models (LLM

Generative Models and Connected and Automated Vehicles: A Survey in Exploring the Intersection of Transportation and AI

SafetyDGX agent

arXiv:2403.10559v3 Announce Type: replace Abstract: This report investigates the history and impact of Generative Models and Connected and Automated Vehicles (CAVs), two groundbreaking forces pushing

GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification

SafetyDGX agent

arXiv:2604.14258v1 Announce Type: cross Abstract: Large language models are typically post-trained using supervised fine-tuning (SFT) and reinforcement learning (RL), yet effectively unifying efficien

Graph-Based Fraud Detection with Dual-Path Graph Filtering

ApplicationsDGX agent

arXiv:2604.14235v1 Announce Type: new Abstract: Fraud detection on graph data can be viewed as a demanding task that requires distinguishing between different types of nodes. Because graph neural netw

GUI-Perturbed: Domain Randomization Reveals Systematic Brittleness in GUI Grounding Models

ResearchDGX agent

arXiv:2604.14262v1 Announce Type: new Abstract: GUI grounding models report over 85% accuracy on standard benchmarks, yet drop 27-56 percentage points when instructions require spatial reasoning rathe

Heat and Matern Kernels on Matchings

SafetyDGX agent

arXiv:2604.14331v1 Announce Type: new Abstract: Applying kernel methods to matchings is challenging due to their discrete, non-Euclidean nature. In this paper, we develop a principled framework for co

HELENA: High-Efficiency Learning-based channel Estimation using dual Neural Attention

Local AiDGX agent

arXiv:2506.13408v2 Announce Type: replace-cross Abstract: Accurate channel estimation is critical for high-performance Orthogonal Frequency-Division Multiplexing systems such as 5G New Radio, particul

High Probability Guarantees for Random Reshuffling

ResearchDGX agent

arXiv:2311.11841v4 Announce Type: replace-cross Abstract: We consider the stochastic gradient method with random reshuffling (mathsf{RR}) for tackling smooth nonconvex optimization problems. mathsf{RR

How Embeddings Shape Graph Neural Networks: Classical vs Quantum-Oriented Node Representations

Model ReleasesDGX agent

arXiv:2604.15273v1 Announce Type: new Abstract: Node embeddings act as the information interface for graph neural networks, yet their empirical impact is often reported under mismatched backbones, spl

Identifying Information from Observations with Uncertainty and Novelty

ResearchDGX agent

arXiv:2501.09331v3 Announce Type: replace Abstract: A machine that learns a task from observations must encounter and process uncertainty and novelty, especially when it is to maintain performance whe

IMPACTX: improving model performance by appropriately constraining the training with teacher explanations

ResearchDGX agent

arXiv:2502.12222v2 Announce Type: replace Abstract: The eXplainable Artificial Intelligence (XAI) research predominantly concentrates to provide explainations about AI model decisions, especially Deep

Improving Clean Accuracy via a Tangent-Space Perspective on Adversarial Training

Model ReleasesDGX agent

arXiv:2408.14728v2 Announce Type: replace Abstract: Adversarial training has proven effective in improving the robustness of deep neural networks against adversarial attacks. However, this enhanced ro

Improving Machine Learning Performance with Synthetic Augmentation

SafetyDGX agent

arXiv:2604.14498v1 Announce Type: cross Abstract: Synthetic augmentation is increasingly used to mitigate data scarcity in financial machine learning, yet its statistical role remains poorly understoo

Improving Sparse Autoencoder with Dynamic Attention

ResearchDGX agent

arXiv:2604.14925v1 Announce Type: new Abstract: Recently, sparse autoencoders (SAEs) have emerged as a promising technique for interpreting activations in foundation models by disentangling features i

Interpretable and Explainable Surrogate Modeling for Simulations: A State-of-the-Art Survey and Perspectives on Explainable AI for Decision-Making

AgentsDGX agent

arXiv:2604.14240v1 Announce Type: cross Abstract: The simulation of complex systems increasingly relies on sophisticated but fundamentally opaque computational black-box simulators. Surrogate models p

Kernel Neural Operators (KNOs) for Scalable, Memory-efficient, Geometrically-flexible Operator Learning

ResearchDGX agent

arXiv:2407.00809v3 Announce Type: replace Abstract: This paper introduces the Kernel Neural Operator (KNO), a provably convergent operator-learning architecture that utilizes compositions of deep kern

Layered Mutability: Continuity and Governance in Persistent Self-Modifying Agents

SafetyDGX agent

arXiv:2604.14717v1 Announce Type: cross Abstract: Persistent language-model agents increasingly combine tool use, tiered memory, reflective prompting, and runtime adaptation. In such systems, behavior

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers

HardwareDGX agent

arXiv:2509.23638v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models face memory and PCIe latency bottlenecks when deployed on commodity hardware. Offloading expert weights to CPU memor

Learning Ad Hoc Network Dynamics via Graph-Structured World Models

SafetyDGX agent

arXiv:2604.14811v1 Announce Type: new Abstract: Ad hoc wireless networks exhibit complex, innate and coupled dynamics: node mobility, energy depletion and topology change that are difficult to model a

Learning and Generating Mixed States Prepared by Shallow Channel Circuits

ResearchDGX agent

arXiv:2604.01197v2 Announce Type: replace-cross Abstract: Learning quantum states from measurement data is a central problem in quantum information and computational complexity. In this work, we study

Learning temporal embeddings from electronic health records of chronic kidney disease patients

TutorialsDGX agent

arXiv:2601.18675v2 Announce Type: replace Abstract: We investigate whether temporal embedding models trained on longitudinal electronic health records can learn clinically meaningful representations w

Learning to Concatenate Quantum Codes

ResearchDGX agent

arXiv:2604.14931v1 Announce Type: cross Abstract: Concatenating quantum error correction codes scales error correction capability by driving logical error rates down double-exponentially across levels

Learning to Plan, Planning to Learn: Adaptive Hierarchical RL-MPC for Sample-Efficient Decision Making

TutorialsDGX agent

arXiv:2512.17091v2 Announce Type: replace Abstract: We propose a new approach for solving planning problems with a hierarchical structure, fusing reinforcement learning and MPC planning. Our formulati

Leveraging graph neural networks and mobility data for COVID-19 forecasting

ResearchDGX agent

arXiv:2501.11711v2 Announce Type: replace Abstract: The COVID-19 pandemic has claimed millions of lives, spurring the development of diverse forecasting models. In this context, the true utility of co

LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking

Model ReleasesDGX agent

arXiv:2604.15149v1 Announce Type: new Abstract: As reinforcement Learning with Verifiable Rewards (RLVR) has become the dominant paradigm for scaling reasoning capabilities in LLMs, a new failure mode

Logo-LLM: Local and Global Modeling with Large Language Models for Time Series Forecasting

Local AiDGX agent

arXiv:2505.11017v2 Announce Type: replace Abstract: Time series forecasting is critical across multiple domains, where time series data exhibit both local patterns and global dependencies. While Trans

Low-Cost System for Automatic Recognition of Driving Pattern in Assessing Interurban Mobility using Geo-Information

ResearchDGX agent

arXiv:2604.15216v1 Announce Type: cross Abstract: Mobility in urban and interurban areas, mainly by cars, is a day-to-day activity of many people. However, some of its main drawbacks are traffic jams

Magnitude Is All You Need? Rethinking Phase in Quantum Encoding of Complex SAR Data

Model ReleasesDGX agent

arXiv:2604.14229v1 Announce Type: cross Abstract: Synthetic Aperture Radar (SAR) data is inherently complex-valued, while quantum machine learning (QML) models naturally operate in complex Hilbert spa

← Previous
1…221222223224225…239
Next →