AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
29 Apr 2026

Application of a Mixture of Experts-based Foundation Model to the GlueX DIRC Detector

Model ReleasesDGX agent

arXiv:2604.24775v1 Announce Type: cross Abstract: We present a Mixture-of-Experts-based foundation model applied to the GlueX DIRC detector at Jefferson Lab, demonstrating its utility as a unified fra

Architecture Determines Observability in Transformers

Model ReleasesDGX agent

arXiv:2604.24801v1 Announce Type: new Abstract: Autoregressive transformers make confident errors, but activation monitoring can catch them only if the model preserves an internal signal that output c

Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories

ApplicationsDGX agent

arXiv:2411.05174v2 Announce Type: replace Abstract: We consider the problem of estimating the transition dynamics T^* from near-optimal expert trajectories in the context of offline model-based reinfo


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Biased Dreams: Limitations to Epistemic Uncertainty Quantification in Latent Space Models

ResearchDGX agent

arXiv:2604.25416v1 Announce Type: new Abstract: Model-Based Reinforcement Learning distinguishes between physical dynamics models operating on proprioceptive inputs and latent dynamics models operatin

Bug-Report-Driven Fault Localization: Industrial Benchmarking and Lesson Learned at ABB Robotics

Local AiDGX agent

arXiv:2604.25700v1 Announce Type: cross Abstract: Software quality assurance remains a major challenge in industrial environments, where large-scale and long-lived systems inevitably accumulate defect

Calibrated Fusion for Heterogeneous Graph-Vector Retrieval in Multi-Hop QA

ResearchDGX agent

arXiv:2603.28886v2 Announce Type: replace-cross Abstract: Graph-augmented retrieval combines dense similarity with graph-based relevance signals such as Personalized PageRank (PPR), but these scores h

CAN-QA: A Question-Answering Benchmark for Reasoning over In-Vehicle CAN Traffic

Model ReleasesDGX agent

arXiv:2604.24935v1 Announce Type: cross Abstract: The Controller Area Network (CAN) is a safety-critical in-vehicle communication protocol that lacks built-in security mechanisms, making intrusion det

Carbon-Taxed Transformers: A Green Compression Pipeline for Overgrown Language Models

SafetyDGX agent

arXiv:2604.25903v1 Announce Type: cross Abstract: The accelerating adoption of Large Language Models (LLMs) in software engineering (SE) has brought with it a silent crisis: unsustainable computationa

Categorical Optimization with Bayesian Anchored Latent Trust Regions for Structural Design under High-Dimensional Uncertainty

ResearchDGX agent

arXiv:2604.25241v1 Announce Type: new Abstract: Categorical structural optimization under aleatoric uncertainty is challenging because each design variable must be selected from a finite catalog of ad

CHUCKLE -- When Humans Teach AI To Learn Emotions The Easy Way

SafetyDGX agent

arXiv:2510.09382v2 Announce Type: replace Abstract: Curriculum learning (CL) structures training from simple to complex samples, facilitating progressive learning. However, existing CL approaches for

CiteRadar: A Citation Intelligence Platform for Researcher Profiling and Geographic Visualization

ResearchDGX agent

arXiv:2604.25057v1 Announce Type: new Abstract: Understanding the geographic reach and community structure of one's scholarly citations is increasingly valuable for career development, grant applicati

Comparative Study of Bending Analysis using Physics-Informed Neural Networks and Numerical Dynamic Deflection in Perforated nanobeam

ResearchDGX agent

arXiv:2604.24768v1 Announce Type: new Abstract: In this chapter, we investigate the bending behavior of a perforated nanobeam subjected to sinusoidal loading using an efficient and computationally rob

Comparing Data Assimilation and Likelihood-Based Inference on Latent State Estimation in Agent-Based Models

Model ReleasesDGX agent

arXiv:2509.17625v2 Announce Type: replace Abstract: In this paper, we present the first systematic comparison of Data Assimilation (DA) and Likelihood-Based Inference (LBI) in the context of an Agent-

Compute Aligned Training: Optimizing for Test Time Inference

SafetyDGX agent

arXiv:2604.24957v1 Announce Type: new Abstract: Scaling test-time compute has emerged as a powerful mechanism for enhancing Large Language Model (LLM) performance. However, standard post-training para

Conditional Flow Matching for Probabilistic Downscaling of Maximum 3-day Snowfall in Alaska

ResearchDGX agent

arXiv:2604.25172v1 Announce Type: cross Abstract: Precipitation in complex terrain is governed by orographic processes operating at scales of a few kilometers, yet climate models typically run at reso

Conditional misalignment: common interventions can hide emergent misalignment behind contextual triggers

SafetyDGX agent

arXiv:2604.25891v1 Announce Type: new Abstract: Finetuning a language model can lead to emergent misalignment (EM) [Betley et al., 2025b]. Models trained on a narrow distribution of misaligned behavio

Contrast-Enhanced Gating in GRUs for Robust Low-Data Sequence Learning

Model ReleasesDGX agent

arXiv:2402.09034v3 Announce Type: replace Abstract: Activation functions govern how recurrent networks regulate and transmit information across temporal dependencies. Despite advances in sequence mode

Contrastive Image-Metadata Pre-Training for Materials Transmission Electron Microscopy

TutorialsDGX agent

arXiv:2604.24909v1 Announce Type: new Abstract: The vast majority of transmission electron microscopy (TEM) data never gets published and ends up on a backup drive until deleted to free up space. Thes

CoreFlow: Low-Rank Matrix Generative Models

ResearchDGX agent

arXiv:2604.24959v1 Announce Type: new Abstract: Learning matrix-valued distributions from high-dimensional and possibly incomplete training data is challenging: ambient-space generative modeling is co

Cornserve: A Distributed Serving System for Any-to-Any Multimodal Models

ResearchDGX agent

arXiv:2603.12118v2 Announce Type: replace Abstract: Any-to-Any models are an emerging class of multimodal models that accept combinations of multimodal data (e.g., text, image, video, audio) as input

Curl Descent: Non-Gradient Learning Dynamics with Sign-Diverse Plasticity

ResearchDGX agent

arXiv:2510.02765v4 Announce Type: replace Abstract: Gradient-based algorithms are a cornerstone of artificial neural network training, yet it remains unclear whether biological neural networks use sim

Curriculum-guided multimodal representation learning enables generalizable prediction of nanomaterial-protein interactions

ResearchDGX agent

arXiv:2507.14245v2 Announce Type: replace Abstract: Nanomaterial-protein interactions (NPI) are pivotal to realizing the therapeutic and diagnostic potential of nanomaterials. Although AI promises to

Data-Driven Hamiltonian Reduction for Superconducting Qubits via Meta-Learning

Model ReleasesDGX agent

arXiv:2604.24912v1 Announce Type: cross Abstract: We introduce HAML (Hamiltonian Adaptation via Meta-Learning), a framework for fast online adaptation of effective Hamiltonian models of superconductin

DCD: Decomposition-based Causal Discovery from Autocorrelated and Non-Stationary Temporal Data

ApplicationsDGX agent

arXiv:2602.01433v2 Announce Type: replace Abstract: Multivariate time series in domains such as finance, climate science, and healthcare often exhibit long-term trends, seasonal patterns, and short-te

Deflation-Free Optimal Scoring

ApplicationsDGX agent

arXiv:2604.25664v1 Announce Type: cross Abstract: Sparse Optimal Scoring (SOS) reformulates linear discriminant analysis to enable feature selection through elastic net regularization, making it well-

DGLight: DQN-Guided GRPO Fine-Tuning of Large Language Models for Traffic Signal Control

SafetyDGX agent

arXiv:2604.25259v1 Announce Type: new Abstract: Traffic signal control (TSC) plays a central role in reducing congestion and maintaining urban mobility. This dissertation introduces DGLight, a critic-

Dictionary learning for Kernel EDMD

Model ReleasesDGX agent

arXiv:2604.25572v1 Announce Type: cross Abstract: Studying nonlinear dynamical systems through their state space behavior can be challenging, and one possible alternative is to analyze them via their

Diffusion Model for Manifold Data: Score Decomposition, Curvature, and Statistical Complexity

TutorialsDGX agent

arXiv:2603.20645v2 Announce Type: replace Abstract: Diffusion models have become a leading framework in generative modeling, yet their theoretical understanding -- especially for high-dimensional data

Digitizing Nepal's Written Heritage: A Comprehensive HTR Pipeline for Old Nepali Manuscripts

ResearchDGX agent

arXiv:2512.17111v2 Announce Type: replace Abstract: This paper presents the first end-to-end pipeline for Handwritten Text Recognition (HTR) for Old Nepali, a historically significant but low-resource

DiRe-RAPIDS: Topology-faithful dimensionality reduction at scale

Model ReleasesDGX agent

arXiv:2604.25209v1 Announce Type: new Abstract: Dimensionality reduction methods such as UMAP and t-SNE are central tools for visualising high-dimensional data, but their local-neighborhood objectives

Drivetrain simulation using variational autoencoders

ApplicationsDGX agent

arXiv:2501.17653v3 Announce Type: replace Abstract: This work proposes variational autoencoders (VAEs) to predict a vehicle's jerk signals from torque demand in the context of limited real-world drive

Dyna-Style Safety Augmented Reinforcement Learning: Staying Safe in the Face of Uncertainty

SafetyDGX agent

arXiv:2604.25508v1 Announce Type: new Abstract: Safety remains an open problem in reinforcement learning (RL), especially during training. While safety filters are promising to address safe exploratio

Dynamic Regret for Online Regression in RKHS via Discounted VAW and Subspace Approximation

ResearchDGX agent

arXiv:2604.25021v1 Announce Type: new Abstract: We study online regression with the square loss in a reproducing kernel Hilbert space under a dynamic regret criterion. The learner is compared with a t

Egocentric Tactile and Proximity Sensors as Observation Priors for Humanoid Collision Avoidance

Model ReleasesDGX agent

arXiv:2604.25554v1 Announce Type: cross Abstract: Collision-free motion is often aided by tactile and proximity sensors distributed on the body of the robot due to their resistance to occlusion as opp

Elite-Driven Support Vector Machines for Classification

Model ReleasesDGX agent

arXiv:2604.25158v1 Announce Type: cross Abstract: Support vector machines (SVMs) are a standard tool for binary classification, but their classical formulations are purely data-driven and offer no dir

Emergent Self-Attention from Astrocyte-Gated Associative Memory Dynamics

ResearchDGX agent

arXiv:2604.25481v1 Announce Type: cross Abstract: We introduce a Hopfield-type associative memory in which effective connectivity is multiplicatively modulated by astrocytic gains evolving under an en

Enhancing SignSGD: Small-Batch Convergence Analysis and a Hybrid Switching Strategy

ResearchDGX agent

arXiv:2604.25550v1 Announce Type: new Abstract: SignSGD compresses each stochastic gradient coordinate to a single bit, offering substantial memory and communication savings, but its 1-bit quantizatio

Evaluating LLM Safety Under Repeated Inference via Accelerated Prompt Stress Testing

Model ReleasesDGX agent

arXiv:2602.11786v2 Announce Type: replace Abstract: Traditional benchmarks for large language models (LLMs), such as HELM and AIR-BENCH, primarily assess safety through breadth-oriented evaluation acr

Evaluation without Generation: Non-Generative Assessment of Harmful Model Specialization with Applications to CSAM

ResearchDGX agent

arXiv:2604.25119v1 Announce Type: new Abstract: Auditing the fine-tunes of open-weight generative models for harmful specialization has become a new governance challenge for model hosting platforms. T

Evolving Multi-Channel Confidence-Aware Activation Functions for Missing Data with Channel Propagation

ResearchDGX agent

arXiv:2602.13864v2 Announce Type: replace-cross Abstract: Learning in the presence of missing data can result in biased predictions and poor generalizability, among other difficulties, which data impu

EvoTSC: Evolving Feature Learning Models for Time Series Classification via Genetic Programming

Model ReleasesDGX agent

arXiv:2604.25499v1 Announce Type: new Abstract: Time series classification is an important analytical task across diverse domains. However, its practical application is often hindered by the scarcity

Explainable AI for Jet Tagging: A Comparative Study of GNNExplainer, GNNShap, and GradCAM for Jet Tagging in the Lund Jet Plane

ResearchDGX agent

arXiv:2604.25885v1 Announce Type: cross Abstract: Graph neural networks such as ParticleNet and transformer based networks on point clouds such as ParticleTransformer achieve state-of-the-art performa

FARM: Enhancing Molecular Representations with Functional Group Awareness

Model ReleasesDGX agent

arXiv:2410.02082v4 Announce Type: replace Abstract: We introduce Functional Group-Aware Representations for Small Molecules (FARM), a novel foundation model designed to bridge the gap between SMILES,

Fast Geometric Embedding for Node Influence Maximization

ResearchDGX agent

arXiv:2506.07435v3 Announce Type: replace-cross Abstract: Computing classical centrality measures such as betweenness and closeness is computationally expensive on large-scale graphs. In this work, we

Feasible-First Exploration for Constrained ML Deployment Optimization in Crash-Prone Hierarchical Search Spaces

Model ReleasesDGX agent

arXiv:2604.25073v1 Announce Type: new Abstract: Deploying machine learning models under production constraints requires joint optimization over model family, quantization scheme, runtime backend, and

FED-FSTQ: Fisher-Guided Token Quantization for Communication-Efficient Federated Fine-Tuning of LLMs on Edge Devices

Model ReleasesDGX agent

arXiv:2604.25421v1 Announce Type: new Abstract: Federated fine-tuning provides a practical route to adapt large language models (LLMs) on edge devices without centralizing private data, yet in mobile

FGDM: Reasoning Aware Multi-Agentic Framework for Software Bug Detection using Chain of Thought and Tree of Thought Prompting

AgentsDGX agent

arXiv:2604.24831v1 Announce Type: cross Abstract: Deep Learning methods are becoming prominent in automated software bug detection; however, they lack the global understanding of the given code. Conse

Fractionally Supervised Classification with Maxima Nominated Samples

ResearchDGX agent

arXiv:2604.25145v1 Announce Type: cross Abstract: Fractionally supervised classification (FSC) offers a flexible framework for combining labeled and unlabeled data in model-based classification, but e

From Cursed to Competitive: Closing the ZO-FO Gap via Input-to-State Stability

ResearchDGX agent

arXiv:2604.25372v1 Announce Type: cross Abstract: While it is generally understood that zeroth-order (ZO) algorithms have an extra dependency on their number of iterations for any choice of parameters

From Soliloquy to Agora: Memory-Enhanced LLM Agents with Decentralized Debate for Optimization Modeling

AgentsDGX agent

arXiv:2604.25847v1 Announce Type: cross Abstract: Optimization modeling underpins real-world decision-making in logistics, manufacturing, energy, and public services, but reliably solving such problem

Frontier Coding Agents Can Now Implement an AlphaZero Self-Play Machine Learning Pipeline For Connect Four That Performs Comparably to an External Solver

Model ReleasesDGX agent

arXiv:2604.25067v1 Announce Type: cross Abstract: Forecasting when AI systems will become capable of meaningfully accelerating AI research is a central challenge for AI safety. Existing benchmarks mea

GCA-BULF: A Bottom-Up Framework for Short-Term Load Forecasting Using Grouped Critical Appliances

ResearchDGX agent

arXiv:2604.24766v1 Announce Type: new Abstract: With the rise of time-of-use and tiered electricity pricing, energy consumers are encouraged to adopt peak-shifting strategies by automatically controll

Generative diffusion models for spatiotemporal influenza forecasting

TutorialsDGX agent

arXiv:2604.24913v1 Announce Type: new Abstract: Forecasting infectious disease incidence can provide important information to guide public health planning, yet is difficult because epidemic dynamics a

Gradient-Direction Sensitivity Reveals Linear-Centroid Coupling Hidden by Optimizer Trajectories

Model ReleasesDGX agent

arXiv:2604.25143v1 Announce Type: new Abstract: We show that replacing the rolling SVD of AdamW updates with a rolling SVD of loss gradients changes the diagnostic by 1-2 orders of magnitude. Performi

GraphPL: Leveraging GNN for Efficient and Robust Modalities Imputation in Patchwork Learning

Model ReleasesDGX agent

arXiv:2604.25352v1 Announce Type: new Abstract: Current research on distributed multi-modal learning typically assumes that clients can access complete information across all modalities, which may not

Grothendieck Graph Neural Networks Framework: An Algebraic Platform for Crafting Topology-Aware GNNs

ResearchDGX agent

arXiv:2412.08835v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) are almost universally built on a single primitive: the neighborhood. Regardless of architectural variations, message p

Heterogeneous Variational Inference for Markov Degradation Hazard Models: Discretized Mixture with Interpretable Clusters

ApplicationsDGX agent

arXiv:2604.24818v1 Announce Type: new Abstract: Bayesian finite mixture models can identify discrete risk clusters (low-risk vs. high-risk equipment), but face three critical bottlenecks: (1) insuffic

Hidden States as Early Signals: Step-level Trace Evaluation and Pruning for Efficient Test-Time Scaling

Model ReleasesDGX agent

arXiv:2601.09093v2 Announce Type: replace Abstract: Large Language Models (LLMs) can enhance reasoning capabilities through test-time scaling by generating multiple traces. However, the combination of

Hierarchical Reinforcement Learning for the Dynamic VNE with Alternatives Problem

Model ReleasesDGX agent

arXiv:2512.05207v2 Announce Type: replace-cross Abstract: Virtual Network Embedding (VNE) is a key enabler of network slicing, yet most formulations assume that each Virtual Network Request (VNR) has

How Can Reinforcement Learning Achieve Expert-level Placement?

TutorialsDGX agent

arXiv:2604.25191v1 Announce Type: cross Abstract: Chip placement is a critical step in physical design. While reinforcement learning (RL)-based methods have recently emerged, their training primarily

← Previous
1…200201202203204…241
Next →