AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
16 Jul 2026

Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations

SafetyDGX agent

arXiv:2607.13399v1 Announce Type: cross Abstract: On-policy distillation (OPD) has become a key paradigm in LLM post-training, yet its training dynamics remain poorly understood. We present a systemat

Design-Specification Tiling for ICL-based CAD Code Generation

ResearchDGX agent

arXiv:2603.12712v2 Announce Type: replace-cross Abstract: Large language models~(LLMs) have demonstrated remarkable capabilities in code generation, yet their performance remains limited on domain-spe

Distributionally Robust and Safe Imitation Learning

SafetyDGX agent

arXiv:2607.13436v1 Announce Type: new Abstract: Imitation learning (IL) has achieved remarkable success in complex decision-making tasks. However, its performance is highly sensitive to distribution s


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Efficiency, Feasibility, and Incentive-Awareness in Constrained Online Resource Allocation

ResearchDGX agent

arXiv:2507.09473v2 Announce Type: replace-cross Abstract: We study the dynamic allocation of indivisible resources to strategic agents under long-term constraints, where the planner aims to maximize s

EM-GANSim: Real-time and Accurate EM Simulation Using Conditional GANs for 3D Indoor Scenes

ResearchDGX agent

arXiv:2405.17366v3 Announce Type: replace Abstract: We present a novel machine-learning (ML) approach (EM-GANSim) for real-time electromagnetic (EM) propagation that is used for wireless communication

Evaluating Frontier AI Agents as Autonomous Clinical Security Auditors

Model ReleasesDGX agent

arXiv:2607.13411v1 Announce Type: cross Abstract: Clinical AI models can expose patients to harm when adversarial vulnerabilities go undetected, yet formal security auditing requires statistical exper

EXPLORE: Exploration with Guided Search for Analog Topology Generation using Language Models

Model ReleasesDGX agent

arXiv:2607.13416v1 Announce Type: new Abstract: Automating analog circuit topology design is essential to reduce the extensive manual effort required to meet increasingly diverse and customized applic

Factorized Spectral Representations for Reinforcement Learning

ResearchDGX agent

arXiv:2607.13498v1 Announce Type: new Abstract: Learning a compact model of the world from interaction data is central to sample-efficient deep reinforcement learning. Spectral representation methods

Flow-aware Optimal Navigation in Unsteady Flows through Reinforcement Learning

AgentsDGX agent

arXiv:2607.13553v1 Announce Type: cross Abstract: Autonomous robotic navigation in nonstationary time-varying fluid flows remains a fundamental challenge due to partial observability and the unpredict

From Novice to Expert: Cost-Aware Bandits for Evolving Worker Performance in Crowdsensing

TutorialsDGX agent

arXiv:2607.13546v1 Announce Type: new Abstract: Mobile crowdsensing (MC) recruits mobile users to perform sensing tasks using their smartphones, enabling large-scale applications such as traffic monit

Gauge-Invariant, Parameter-Insensitive Regularization for Potential Recovery from Flow on Directed Graphs

Model ReleasesDGX agent

arXiv:2607.13609v1 Announce Type: new Abstract: Recovering a latent potential from observed flow on a directed graph (a discrete Poisson problem with Dirichlet boundaries) is ill-posed, and the standa

GFlowRL: Scaling Distribution-Matching RL to Large Language Models

ResearchDGX agent

arXiv:2607.13394v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) offer a promising alternative to reward-maximizing reinforcement learning (RL) for large reasoning models, encour

Graded Entity-Familiarity Readouts in Language Models: Polish Adaptation, Cross-Language Robustness, and Refusal Steering

Model ReleasesDGX agent

arXiv:2607.13568v1 Announce Type: cross Abstract: Can a language model estimate its familiarity with an entity before generating an answer? We study activations at the final prompt token in twelve ins

Graph Partitioning with Demands: Generalized Conductance and its Applications

ResearchDGX agent

arXiv:2607.13218v1 Announce Type: cross Abstract: In this work, we study various graph partitioning problems under a general demand model. In each such task, we are given a graph G=(V,E,c,w) with a ca

Heavy-Tailed Flow Matching via Random Clocks

Model ReleasesDGX agent

arXiv:2607.13841v1 Announce Type: new Abstract: Heavy-tailed data arise in many domains where rare events carry disproportionate importance, such as imbalanced image datasets, financial returns, and w

HEDGEHOG: Hierarchical Evaluation of Drug Generators Through Rigorous Filtration

Model ReleasesDGX agent

arXiv:2607.13155v1 Announce Type: new Abstract: Generative molecular models can support early drug discovery by proposing new candidate compounds de novo. In practice, useful candidates must balance t

Hierarchical F-Clustering: Approximation and Hardness of Clustering into Trees and Bounded Diameter Graphs

ResearchDGX agent

arXiv:2607.13217v1 Announce Type: cross Abstract: Consider the following variation on the Hierarchical Clustering problem: Usually, while building a hierarchical clustering, one recursively partitions

How the Hessian-Spectrum of Neural Networks Depends on Data

ResearchDGX agent

arXiv:2607.13631v1 Announce Type: new Abstract: The Hessian matrix is an important quantity of interest when it comes to studying the loss landscape and optimization dynamics in deep learning, as well

HRIBench: Benchmarking Interaction-Centric Human-Robot Collaboration

Model ReleasesDGX agent

arXiv:2607.13056v1 Announce Type: cross Abstract: Current vision-language-action (VLA) benchmarks primarily evaluate isolated manipulation skills while leaving human-robot interaction structure largel

Implementations of Quantum and Classical Topology-Aligned Architectures for Molecular Property Prediction

Model ReleasesDGX agent

arXiv:2607.13737v1 Announce Type: new Abstract: For low-data and resource-constrained regimes typical of quantum chemistry, parameter-efficient learning is a key objective. Here, we propose a topology

Lag Operator SSMs: A Geometric Framework for Structured State Space Modeling

ResearchDGX agent

arXiv:2512.18965v2 Announce Type: replace Abstract: Structured State Space Models (SSMs), which are at the heart of the recently popular Mamba architecture, are powerful tools for sequence modeling. H

Learned Pairwise Deep Dual-Optimal Inequalities for Stabilizing Column Generation

ResearchDGX agent

arXiv:2607.13373v1 Announce Type: cross Abstract: Column generation (CG) is central to many large-scale optimization algorithms, including branch-price-and-cut methods for vehicle routing problems, bu

Leveraging Differentiable PDE Solvers for Semi-Neural Spatial Reconstruction From Sparse Measurements

ResearchDGX agent

arXiv:2601.20496v2 Announce Type: replace-cross Abstract: Generating dense physical fields from sparse measurements is a fundamental question in sampling, signal processing, and many other application

Leveraging unlabelled data for generalizable neural population decoding

ResearchDGX agent

arXiv:2607.14086v1 Announce Type: new Abstract: Robust and accurate neural decoders are integral to neurotechnologies such as brain-computer interfaces and closed-loop experiments. Recent work has sho

Lighthouse RL: Sample-Efficient Circuit Optimization via Strategic Reset Points

Model ReleasesDGX agent

arXiv:2607.14008v1 Announce Type: new Abstract: In this paper, we introduce Lighthouse RL, a sample-efficient reinforcement learning (RL) approach for analog circuit sizing. Traditional methods lack g

Linear Independent Component Analysis via Optimal Transport

ResearchDGX agent

arXiv:2607.14081v1 Announce Type: new Abstract: Linear Independent Component Analysis (ICA) recovers jointly independent source signals from their linear mixtures. To achieve this, classical ICA algor

Local Redundancy: An Information-Theoretic Measure of Plasticity from Synthetic Memorization

Local AiDGX agent

arXiv:2607.13432v1 Announce Type: new Abstract: Plasticity -- a neural network's ability to adapt to new tasks -- is critical for continual and transfer learning. Existing measures, such as effective

Lyapunov Exponent as Physics-Informed Dense Reward: RL Discovery of Stabilization Beyond the Kapitza Pendulum

AgentsDGX agent

arXiv:2607.14001v1 Announce Type: new Abstract: We suggest using the Lyapunov characteristic exponent (LCE) as a dense reward signal for the reinforcement learning problem of stabilizing the inverted

M+Adam: Low-Precision Training via Additive-Multiplicative Optimization

Model ReleasesDGX agent

arXiv:2607.10611v2 Announce Type: replace Abstract: Training with quantized weights can reduce costs but often results in degraded accuracy, especially when optimization is carried out in low precisio

Maximally Robust Satisficing Bayesian Optimization

ResearchDGX agent

arXiv:2607.13652v1 Announce Type: new Abstract: Many design tasks can be cast as black-box function optimization, enabling use of Bayesian optimization to find an ideal design with minimal number of t

MetaPerch: Learning from metadata for bioacoustics foundation models

ApplicationsDGX agent

arXiv:2607.14072v1 Announce Type: new Abstract: Bioacoustic foundation models rely on large-scale citizen science platforms like Xeno-Canto for geographically and ecologically diverse data. Recent wor

Microstructure-Conditioned Surrogate Models for Graded Multiscale Optimization of Mycelium Composites

ApplicationsDGX agent

arXiv:2607.13688v1 Announce Type: new Abstract: Emerging sustainable materials increasingly rely on engineered hierarchy and microstructure to achieve control of their properties and mechanical behavi

Mixed-Timescale Differential Coding for Downlink Model Broadcast in Wireless Federated Learning

Model ReleasesDGX agent

arXiv:2607.13119v1 Announce Type: cross Abstract: In standard federated learning systems, the parameter server broadcasts the global model to the participating devices in every iteration. Motivated by

Mono-Z Dark Matter Search with Neural Spline Flows Using CMS Run 2015D Open Data

Model ReleasesDGX agent

arXiv:2607.13771v1 Announce Type: new Abstract: We report a search for dark matter (DM) produced in association with a leptonically decaying (Z) boson at (sqrt{s}=13) TeV using CMS Run 2015D open data

Multi-Dictionary Learning for Low Rank Sparse Coding

TutorialsDGX agent

arXiv:2509.10033v2 Announce Type: replace Abstract: Sparse dictionary coding represents signals as linear combinations of a few dictionary atoms. It has been applied to images, time series, graph sign

Multimodal Empirical Bayes Variational Autoencoders for Joint Longitudinal and Time-to-Event Modeling

ResearchDGX agent

arXiv:2607.13984v1 Announce Type: cross Abstract: Longitudinal tumor measurements, dropout information, and genetic covariates provide complementary information about treatment response, but integrati

New universal operator approximation theorem for encoder-decoder architectures

ResearchDGX agent

arXiv:2503.24092v2 Announce Type: replace-cross Abstract: Motivated by the rapidly growing field of mathematics for operator approximation with neural networks, we present a novel universal operator a

Non-Expansive Two-Time-Scale Stochastic Approximation: A Fixed-Schedule One-Quarter Barrier and Bias-Corrected Acceleration

SafetyDGX agent

arXiv:2607.13414v1 Announce Type: cross Abstract: Non-expansive two-time-scale stochastic approximation is governed by a slow stochastic Krasnoselskii--Mann fixed-point iteration rather than by contra

Not All Retrievals are Useful: Cross-Attention for Input-Aware RAG in Time Series Forecasting

ResearchDGX agent

arXiv:2603.14709v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) enhances zero-shot time series (TS) forecasting by leveraging external knowledge bases, yet existing approaches

On the Sublinear Regret of Continuous K-Max Bandits

SafetyDGX agent

arXiv:2502.13467v2 Announce Type: replace Abstract: The K-Max combinatorial multi-armed bandit problem arises in applications such as recommendation and distributed decision making, where the reward i

Optimal and Efficient Contextual Combinatorial Semi-bandits with General Function Approximation

SafetyDGX agent

arXiv:2607.13686v1 Announce Type: new Abstract: We study the contextual combinatorial semi-bandit (CCSB) problem with general reward function approximation. At each round, the learner observes a conte

Optimizing Binary and Ternary Neural Network Inference on RRAM Crossbars using CIM-Explorer

ResearchDGX agent

arXiv:2505.14303v3 Announce Type: replace-cross Abstract: Using Resistive Random Access Memory (RRAM) crossbars in Computing-in-Memory (CIM) architectures offers a promising solution to overcome the v

OrDA: Orthogonal Disentanglement of Access Habits Framework for Homepage Marketing Block Recommendations

SafetyDGX agent

arXiv:2607.13420v1 Announce Type: new Abstract: Clicks on homepage marketing blocks are driven by a dual-mechanism of content interest and access habits. However, habitual clicks often create Pseudo-P

Parallel gradient boosting for flexible estimation of conditional distributions

ResearchDGX agent

arXiv:2607.13550v1 Announce Type: cross Abstract: Boosting is one of the most successful learning techniques for standard classification and regression tasks. Its extension to multi-output prediction

Plausible Deniability Guarantees for Whistleblowers

ResearchDGX agent

arXiv:2607.13928v1 Announce Type: cross Abstract: Whistleblowers are a key safeguard against organizational wrongdoing, but the threat of retaliation deters reporting. Existing whistleblower-protectio

Power Homotopy for Zeroth-Order Non-Convex Optimizations

ResearchDGX agent

arXiv:2511.13592v2 Announce Type: replace-cross Abstract: The existing method of GS-PowerOpt solves the non-convex optimization problem of the form max_{oldsymbol{x} in R^d} f(oldsymbol{x}) through ma

PQFA: Parallel Quantum Feature Augmentation of Fused Representations for Multimodal Classification

Model ReleasesDGX agent

arXiv:2607.13466v1 Announce Type: new Abstract: Most multimodal learning methods improve how heterogeneous representations are aligned and fused, while post-fusion enhancement remains less explored. W

Precomputing the Future-Offset Average in TriAttention

ResearchDGX agent

arXiv:2607.13051v1 Announce Type: cross Abstract: TriAttention is a recent method for shrinking the KV cache of long-reasoning LLMs: it scores each cached key by how much attention it is likely to rec

Pretraining in Actor-Critic Reinforcement Learning for Locomotion

SafetyDGX agent

arXiv:2510.12363v4 Announce Type: replace-cross Abstract: The pretraining-finetuning paradigm has facilitated numerous transformative advancements in artificial intelligence research in recent years.

PUe: Biased Positive-Unlabeled Learning Enhancement by Causal Inference

SafetyDGX agent

arXiv:2607.13428v1 Announce Type: new Abstract: Positive-Unlabeled (PU) learning aims to achieve high-accuracy binary classification with limited labeled positive examples and numerous unlabeled ones.

Quantum Topological Data Encoding

ResearchDGX agent

arXiv:2607.13847v1 Announce Type: cross Abstract: Many datasets encountered across a wide range of domains possess rich geometric and topological structure that is difficult to capture using conventio

Relevance-Aware Rule: Structural Deletion of Irrelevant Conditions in Decision Trees

ResearchDGX agent

arXiv:2607.13874v1 Announce Type: new Abstract: Decision trees generate interpretable if--then rules, yet they contain irrelevant conditions (IRCs). These IRCs arise from the structural mechanism of t

RF-Informed Graph Neural Networks for Accurate and Data-Efficient Circuit Performance Prediction

ResearchDGX agent

arXiv:2508.16403v3 Announce Type: replace Abstract: Accurately predicting the performance of active radio frequency (RF) circuits is essential for modern wireless systems but remains challenging due t

RF Spectrogram Anomaly Detection with Quantum Kitchen Sinks: Architecture, Representation, and Hardware Validation

Model ReleasesDGX agent

arXiv:2607.13897v1 Announce Type: new Abstract: The broadcast nature of wireless channels exposes radio-frequency (RF) networks to anomalous and malicious transmissions, making anomaly detection a fun

SCOPE-RL: Optimizing Reasoning Paths Before and After Success

SafetyDGX agent

arXiv:2607.11506v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) optimizes LLMs using sparse verifiable final-answer rewards. This sparse anchor reliably verif

Screening of Biosecurity Features in Metagenomic Data with Evo 2 Probes

TutorialsDGX agent

arXiv:2607.14070v1 Announce Type: cross Abstract: Genomic foundation models such as Evo 2 learn rich sequence representations, but their value for biosecurity screening is largely unexplored. We ask h

Securing LLMs in the Wild: Privacy and Security Challenges at the Edge

Model ReleasesDGX agent

arXiv:2607.13088v1 Announce Type: cross Abstract: Large Language Models (LLMs) are rapidly moving from research settings into the wild, deployed on enterprise infrastructure, personal devices, and edg

Self-Improving is Often Sudden: Enlightenment-style Finetuning for Large-Scale Models

ResearchDGX agent

arXiv:2607.13395v1 Announce Type: new Abstract: The pursuit of autonomously self-improving models has attracted growing interest in the era of large-scale foundation models. Drawing inspiration from t

Smooth Quasar-Convex Optimization with Constraints

ResearchDGX agent

arXiv:2510.01943v3 Announce Type: replace-cross Abstract: Quasar-convex functions form a broad nonconvex class with applications to linear dynamical systems, generalized linear models, and Riemannian

Structured Reinforcement Learning for Bayesian Persuasion : Application to Intelligent Interactive Driving

SafetyDGX agent

arXiv:2607.13576v1 Announce Type: new Abstract: Interactive driving, wherein an intelligent lead vehicle equipped with real-time traffic data coordinates route choices of connected vehicles, offers a

← Previous
1…3940414243…241
Next →