AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
23 Jun 2026

From Knowing to Acting: Benchmarking Self-Awareness Capability of LLM Agents

Model ReleasesDGX agent

arXiv:2606.20661v1 Announce Type: cross Abstract: The integration of external tools has transitioned LLM agents from passive responders to autonomous systems. However, current benchmarks prioritize ex

From Markov to Laplace: How Mamba In-Context Learns Markov Chains

ResearchDGX agent

arXiv:2502.10178v2 Announce Type: replace Abstract: While transformer-based language models have driven the AI revolution thus far, their computational complexity has spurred growing interest in viabl

From Speech to Text Corpora: Evaluating ASR-Based Data Acquisition for Low-Resource Fongbe and Hausa

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.22274v1 Announce Type: cross Abstract: Low-resource African languages lack text corpora needed for language model training. We investigate whether ASR pipelines can extend text resources fo

From Zipf's Law to Neural Scaling through Heaps' Law and Hilberg's Hypothesis

ResearchDGX agent

arXiv:2512.13491v3 Announce Type: replace-cross Abstract: We inspect the deductive connection between the neural scaling law and Zipf's law -- two statements discussed in machine learning and quantita

Fusing Backdoors, Machine Learning, and Optimization for Large-Scale Parametric Mixed-Integer Programs

ApplicationsDGX agent

arXiv:2606.21440v1 Announce Type: new Abstract: Large-scale optimization problems are often solved repeatedly under similar structural conditions, leading to substantial computational overhead. This o

GAC: Stabilizing Asynchronous RL Training for LLMs via Gradient Alignment Control

SafetyDGX agent

arXiv:2603.01501v2 Announce Type: replace Abstract: Asynchronous execution is essential for scaling reinforcement learning (RL) to modern large model workloads, including large language models and AI

GARIP: A Running-Average Moving Reference for Last-Iterate Self-Play in Two-Player Zero-Sum Games

SafetyDGX agent

arXiv:2606.22688v1 Announce Type: cross Abstract: Self-play with naive gradient ascent cycles in two-player zero-sum games: the last iterate orbits the equilibrium. Modern methods restore last-iterate

Gated MLPs as Symmetry-Broken Rank-1 Bilinear Attention

ResearchDGX agent

arXiv:2606.22172v1 Announce Type: new Abstract: We show that the conventional gated MLP can be viewed as a rank-1 approximation to a bilinear attention mechanism with two distinct factors correspondin

Gaussian process policy iteration with additive Schwarz acceleration for forward and inverse HJB and mean field game problems

SafetyDGX agent

arXiv:2505.00909v3 Announce Type: replace Abstract: In this paper, we propose a Gaussian Process (GP)-based policy iteration framework for addressing both forward and inverse problems in Hamilton--Jac

Generalized nonparametric regression in reproducing kernel Hilbert spaces: Consistency and rates of convergence

Model ReleasesDGX agent

arXiv:2606.22993v1 Announce Type: cross Abstract: We develop a comprehensive theory for regularized M-estimation in reproducing kernel Hilbert spaces. Under mild conditions on the loss we establish ex

Generative Modeling via Kernelized Stochastic Interpolants

ResearchDGX agent

arXiv:2602.20070v3 Announce Type: replace Abstract: We develop a kernel method for generative modeling within the stochastic interpolant framework, replacing neural network training with linear system

Generative Robust Optimisation

ApplicationsDGX agent

arXiv:2606.22536v1 Announce Type: new Abstract: Classical uncertainty sets for robust optimisation impose fixed geometric shapes that cannot represent the complex dependencies present in real-world da

Geometric and Information Compression of Representations in Deep Learning

ResearchDGX agent

arXiv:2606.21593v1 Announce Type: new Abstract: Deep neural networks transform input data into latent representations that support a wide range of downstream tasks. These representations can be charac

GeoRouteNet: Geometry-Enhanced Non-Autoregressive Neural Solver for the Traveling Salesman Problem

Model ReleasesDGX agent

arXiv:2606.22776v1 Announce Type: new Abstract: The traveling salesman problem (TSP) is a canonical NP-hard combinatorial optimization benchmark that tests the representational capacity and generaliza

GeoTransolver: Learning Physics on Irregular Domains Using Multi-scale Geometry Aware Physics Attention Transformer

Model ReleasesDGX agent

arXiv:2512.20399v3 Announce Type: replace Abstract: We present GeoTransolver, a multiscale geometry-aware physics attention transformer for Computer Aided Engineering (CAE). GeoTransolver extends the

GitOfThoughts: Version-Controlled Reasoning and Agent Memory You Can Replay, Diff, and Merge

Model ReleasesDGX agent

arXiv:2606.14470v2 Announce Type: replace-cross Abstract: Large language model reasoning leaves no trace once it is done. The steps of a chain of thought disappear when the context window closes, a pr

Good-Enough LLM Obfuscation (GELO)

Model ReleasesDGX agent

arXiv:2603.05035v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly served on shared accelerators where an adversary with read access to device memory can observe K

GRADE: Graph Representation of LLM Agent Dependency and Execution

AgentsDGX agent

arXiv:2606.22741v1 Announce Type: new Abstract: Can one graph represent every kind of LLM agent's run? A trace records what each step did, never what it relied on, the state it read, and the results i

Gradient-Descent Steps to Success over Mean Accuracy: A Paradigm Shift for ML

ResearchDGX agent

arXiv:2606.22053v1 Announce Type: new Abstract: Traditional evaluation of machine learning (ML) models typically focuses on achieving the maximum possible accuracy irrespective of the computational co

Gradient Flow Through Diagram Expansions: Learning Regimes and Explicit Solutions

Model ReleasesDGX agent

arXiv:2602.04548v2 Announce Type: replace Abstract: We develop a general mathematical framework to analyze scaling regimes and derive explicit analytic solutions for gradient flow (GF) in large learni

Gradient-Free Warm-Start Library Recovery: an Amortized-Regret Separation

AgentsDGX agent

arXiv:2606.21253v1 Announce Type: new Abstract: Continual learning that is gradient-free, local, online, and append-only is attractive for edge and streaming deployment, but its value is usually argue

Gradual Capacity Growth for Sparse Network Discovery

ResearchDGX agent

arXiv:2509.25665v2 Announce Type: replace Abstract: Sparse neural network methods typically assume that the target sparsity (or density) is fixed in advance, even though the relationship between netwo

GRAG: Generic Response-Augmented Generation Framework for Personalized Conversational Systems

Model ReleasesDGX agent

arXiv:2606.21097v1 Announce Type: cross Abstract: Deploying highly capable personalized conversational agents in resource-constrained or privacy-sensitive environments remains a significant challenge.

GRAIN: Group Aggregation via Min-Norm Objective

ResearchDGX agent

arXiv:2606.22917v1 Announce Type: new Abstract: Learning instability is a long-standing problem across machine learning, but it is especially acute in the overparameterized regime that defines modern

GraphPFN: A Prior-Data Fitted Graph Foundation Model

ApplicationsDGX agent

arXiv:2509.21489v3 Announce Type: replace Abstract: Graph foundation models face several fundamental challenges including transferability across diverse domains and data scarcity, which calls into que

GRIMIP: A General Framework for Instance-Specific Configuration of MIP Solvers Using LLMs

ResearchDGX agent

arXiv:2606.23299v1 Announce Type: new Abstract: Configuring the hyperparameters of Mixed-integer programming (MIP) solvers is a high-dimensional, instance-dependent optimization problem where suboptim

GRINQH: Graded Input-based Quantization Hierarchy for Efficient LLM Generation

HardwareDGX agent

arXiv:2606.23419v1 Announce Type: new Abstract: Autoregressive decoding with LLMs is primarily bottlenecked by GPU memory bandwidth, especially in edge-computing settings. While quantization is essent

Group-Graph Policy Optimization for Long-Horizon Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.22995v1 Announce Type: new Abstract: Group-based Reinforcement Learning (RL) has significantly enhanced Large Language Models (LLMs) in agentic scenarios. To achieve finer-grained policy up

Grouped Query Experts: Mixture-of-Experts on GQA Self-Attention

Model ReleasesDGX agent

arXiv:2606.20945v1 Announce Type: new Abstract: Self-attention is central to Transformer performance and is often the most expensive part of the Transformer at long context lengths because its pairwis

Hard or Just Unreached? Diagnosing the Sampling Blind Spot in Math-Reasoning Difficulty Estimation

ResearchDGX agent

arXiv:2606.19636v2 Announce Type: replace Abstract: Math and science reasoning benchmarks rely on pass@k, the fraction of sampled chains that reach gold, as the canonical per-example difficulty signal

Harnessing Agent Skills: Architectural Patterns and a Reference Architecture for Skill-Mediated LLM Agents

AgentsDGX agent

arXiv:2606.20631v1 Announce Type: cross Abstract: Agent skills externalise reusable agent-facing behavioural knowledge and guidance as persistent artefacts that can be discovered, activated, and inter

HEAS: Hierarchical Evolutionary Agent-Based Simulation Framework for Multi-Objective Policy Search

SafetyDGX agent

arXiv:2508.15555v4 Announce Type: replace-cross Abstract: HEAS is a Python framework that connects agent-based simulation, evolutionary search, and scenario-based evaluation in a single reproducible p

HERALD: High-Throughput Block Diffusion LLM Serving via CPU-GPU Cooperative KV Cache Retrieval

HardwareDGX agent

arXiv:2606.21633v1 Announce Type: new Abstract: Diffusion LLMs (dLLMs) improve GPU utilization over autoregressive decoding by generating multiple tokens per forward pass, but their KV cache still gro

Hierarchical Adversarial Bandits for Online Configuration Optimization

Model ReleasesDGX agent

arXiv:2505.19061v2 Announce Type: replace Abstract: Motivated by Online Configuration Optimization in large, dynamic parameter spaces, this work studies the nonstochastic multi-armed bandit (MAB) prob

Hierarchical Pooling for Sheaf Neural Networks

ResearchDGX agent

arXiv:2606.20932v1 Announce Type: new Abstract: Sheaf Neural Networks (SNNs) generalize Graph Neural Networks (GNNs) by replacing scalar node signals with stalk-valued signals and by using restriction

Hierarchical Reinforcement Learning for Sparse-Reward Search in Commutative Algebra

SafetyDGX agent

arXiv:2606.22922v1 Announce Type: new Abstract: Applying machine learning techniques to solving long-standing mathematical conjectures can be particularly challenging due to their extreme reward spars

Hierarchical Sparse Circuit Extraction from Billion-Parameter Language Models through Scalable Attribution Graph Decomposition

Model ReleasesDGX agent

arXiv:2601.12879v2 Announce Type: replace Abstract: Extracting sparse circuits from billion-parameter transformers is constrained by O(2^n) search cost and pervasive feature reuse across co-active pat

High-Dimensional Differentially Private Quantile Regression: Distributed Estimation and Statistical Inference

ResearchDGX agent

arXiv:2508.05212v2 Announce Type: replace-cross Abstract: With the development of big data and machine learning, privacy concerns have become increasingly critical, especially when handling heterogene

Horizon Adaptive Offline Policy Learning via Value Stitching

SafetyDGX agent

arXiv:2606.21136v1 Announce Type: new Abstract: Learning accurate value functions plays a decisive role for reinforcement learning (RL) agents to solve long-horizon, complex tasks. Conventional tempor

How Should a Simulation-to-Reality Transfer Budget Be Spent?

Model ReleasesDGX agent

arXiv:2606.22062v1 Announce Type: cross Abstract: Simulation-to-reality transfer, often called sim-to-real transfer, is a central challenge in robot learning. Yet, the tradeoff between measuring a sys

How Well Do Self-Supervised Speech Models Encode Age and Gender in Children's Speech? A Layer-Wise Analysis Across Multiple Architectures

Model ReleasesDGX agent

arXiv:2606.22177v1 Announce Type: cross Abstract: Self-supervised learning (SSL) models have become a central component of modern speech processing systems, as they enable the learning of rich acousti

HyperQuant: A Rate-Distortion-Optimal Quantization Pipeline for Large Language and Diffusion Models

Model ReleasesDGX agent

arXiv:2606.23406v1 Announce Type: new Abstract: We present HyperQuant (Hadamard, optimallY Packing, Entropy Rice-coding), a unified post-training quantization pipeline for the weights and the KV cache

Hypothesis-Disciplined Multi-Agent Automated Formalization of Asymptotic Statistical Theory

AgentsDGX agent

arXiv:2606.20642v1 Announce Type: cross Abstract: Asymptotic statistical theory is a challenging domain for AI-assisted formalization: its central results mix convergence statements, asymptotic expans

Imitation from Heterogeneous Demonstrations using Grounded Latent-Action World Models

SafetyDGX agent

arXiv:2606.21672v1 Announce Type: cross Abstract: Imitation learning has emerged as a powerful paradigm for learning visuomotor policies, but its generalisation and stability are limited by the scale

Improved Algorithms for Nash Welfare in Linear Bandits

SafetyDGX agent

arXiv:2601.22969v2 Announce Type: replace Abstract: Nash regret has recently emerged as a principled fairness-aware performance metric for stochastic multi-armed bandits, motivated by the Nash Social

Improving Text-to-Music Generation with Human Preference Rewards

Model ReleasesDGX agent

arXiv:2606.21670v1 Announce Type: cross Abstract: We describe our entry to the efficiency track of the Academic Text-to-Music (ATTM) Grand Challenge at ICME 2026. Beyond the challenge protocol's FAD-C

In-Context Molecular Property Prediction with LLMs: A Blinding Study on Memorization and Knowledge Conflicts

Model ReleasesDGX agent

arXiv:2603.25857v2 Announce Type: replace Abstract: The capabilities of large language models (LLMs) have expanded beyond natural language processing to scientific prediction tasks, including molecula

In LLM Reasoning, there is Irrationality on top of Value Misalignment

Model ReleasesDGX agent

arXiv:2606.20624v1 Announce Type: cross Abstract: Significant progress has been made in aligning LLMs with target value functions. We argue that, even when an LLM has been well aligned in (post-)train

Incremental Learning in Mirror Flows

ResearchDGX agent

arXiv:2606.23198v1 Announce Type: cross Abstract: We study mirror flows generated by a convex quadratic loss and a general convex lower semicontinuous mirror potential. We show that, when initialized

IndicGuard: A Multilingual Safety Guard Model and Dataset for Indic Languages

Model ReleasesDGX agent

arXiv:2606.22841v1 Announce Type: cross Abstract: As Large Language Models (LLMs) achieve widespread integration across diverse linguistic landscapes, ensuring their safety and alignment with regional

Inductive Generalization for Robotic Manipulation

SafetyDGX agent

arXiv:2606.20999v1 Announce Type: cross Abstract: Understanding the generalization capabilities of visuomotor policies is essential in the development of capable robotic agents. Generalizable models l

Influencer Cartels

SafetyDGX agent

arXiv:2405.10231v3 Announce Type: replace-cross Abstract: Social media influencers account for a growing share of marketing worldwide. We demonstrate the existence of a novel form of market failure in

Input-schema identifiability limits in physics-informed surrogates for mechanics-governed flow

ResearchDGX agent

arXiv:2606.20655v1 Announce Type: cross Abstract: Physics-informed and data-driven surrogates are increasingly used to approximate mechanics-governed flow fields, but the target quantities assigned to

Integrated Marketing Attribution: A Bayesian Framework for Privacy-Safe Granular Measurement Anchored in MMM

ResearchDGX agent

arXiv:2606.16878v3 Announce Type: replace Abstract: Retail marketing measurement increasingly requires granular campaign-level insights without relying on user-level tracking. However, the two dominan

Intent-Handover: Grounding Language in Human-Usage Regions for Trustworthy Robot-to-Human Handovers

SafetyDGX agent

arXiv:2503.03579v2 Announce Type: replace-cross Abstract: Spoken instructions in robot-to-human handovers may specify either an object ('the cup') or an intended use ('pour water'); in both cases, suc

Interleaved Speech Language Models Latently Work In Text

ResearchDGX agent

arXiv:2606.22473v1 Announce Type: cross Abstract: Speech language models (SLMs) have been extensively studied, with the common paradigm incorporating text data and pre-trained text LMs. A leading appr

Interpretable Kolmogorov-Arnold Network with Feature-Isolated Temporal Attention Mechanism for Electricity Load Forecasting

ResearchDGX agent

arXiv:2606.23425v1 Announce Type: new Abstract: Accurate electricity load forecasting is a crucial prerequisite for stable power system operations. While prevalent deep learning models present competi

Interpretable Machine Learning for Predicting Startup Funding, Patenting, and Exits

ApplicationsDGX agent

arXiv:2510.09465v2 Announce Type: replace Abstract: This study develops an interpretable machine learning framework to forecast startup outcomes, including funding, patenting, and exit. A firm-quarter

Interpretable machine learning of halo gas density profiles: a sensitivity analysis of cosmological hydrodynamical simulations

ResearchDGX agent

arXiv:2512.09021v3 Announce Type: replace-cross Abstract: Stellar and AGN-driven feedback processes affect the distribution of gas on a wide range of scales, from within galaxies well into the interga

Intrinsic Flow Matching on Quantum Pure-State Manifolds with Phase-Aligned Transport

ResearchDGX agent

arXiv:2606.21256v1 Announce Type: new Abstract: Quantum pure-state ensembles live on complex projective space, making flat Euclidean generative modeling geometrically mismatched. We introduce Intrinsi

← Previous
1…7879808182…243
Next →