AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
26,078 results
26 May 2026

TGFormer: Towards Temporal Graph Transformer with Auto-Correlation Mechanism

ResearchDGX agent

arXiv:2605.24971v1 Announce Type: cross Abstract: The growing interest in Temporal Graph Neural Networks (TGNNs) stems from their ability to model complex dynamics and deliver superior performance. Ho

Thaka at KSAA-2026 Task 2: Regularized Fine-Tuning for Arabic Speech Diacritization

ResearchDGX agent

arXiv:2605.25928v1 Announce Type: new Abstract: We describe the winning system for Task 2 of the KSAA-2026 Shared Task on Arabic Speech Dictation with Automatic Diacritization. The task requires produ

The Download: puncturing the AI jobs panic

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. A reality check on the AI jobs hysteria Despite the growing hy

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Impact of Large Language Models on Open-source Innovation: Evidence from GitHub Copilot

ResearchDGX agent

arXiv:2409.08379v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are reshaping knowledge work, yet their impact on voluntary, self-guided open innovation forums (contributors cho

The meaning of prompts and the prompts of meaning: Semiotic reflections and modelling

ResearchDGX agent

arXiv:2509.14250v2 Announce Type: replace Abstract: This paper explores prompts and prompting in large language models (LLMs) as dynamic semiotic phenomena, drawing on Peirce's triadic model of signs,

The Normalized Maximum Likelihood for Regular Non-Smooth Models: Measure-Theoretic Foundations and Geometric Sampling

ResearchDGX agent

arXiv:2605.24477v1 Announce Type: new Abstract: The Normalized Maximum Likelihood (NML) codelength, or stochastic complexity, represents a principled criterion for universal coding. While recent coare

The Quantization Benefits of Residual-Free Transformers

ResearchDGX agent

arXiv:2605.25880v1 Announce Type: new Abstract: Large-scale transformer training and deployment are increasingly constrained by the transfer of activations, gradients, and optimizer states across acce

The Telegraph: Macron tore up 65 years of doctrine to defend Europe with French nukes, with or without the US. Poland, Germany, the Netherla…

ResearchDGX agent

The Telegraph: Macron tore up 65 years of doctrine to defend Europe with French nukes, with or without the US. Poland, Germany, the Netherlands, Belgium, Greece, Sweden, Denmark, and now the Czech Rep

The Timing Dependencies of Trust: Speed, Accuracy, and cBCI Neuro-Decoupling in Human-AI Teams

ResearchDGX agent

arXiv:2605.25868v1 Announce Type: cross Abstract: The speed and accuracy of an artificial teammate fundamentally alter the failure states of Human-AI integration. While high-speed AI interventions ris

The Tokenizer Tax Across 25 European Languages: Domain Invariance, Cross-Lingual Few-Shot Effects, and the Ukrainian Penalty

ResearchDGX agent

arXiv:2605.24718v1 Announce Type: new Abstract: Tokenizer fertility the number of tokens per word imposes a hidden cost on non-English NLP. We measure fertility for ten foundation models across 25 Eur

Theoretical Analysis of Sparse Optimization with Reparameterization, Weight Decay, and Adaptive Learning Rate

ResearchDGX agent

arXiv:2605.25134v1 Announce Type: cross Abstract: Sparse optimization is a fundamental challenge in various practical applications. A popular approach to sparse optimization is ell_p regularization. H

They Are Not the Same: Direct Causes Are Not Grounded Emotion Explanations

ResearchDGX agent

arXiv:2605.25208v1 Announce Type: new Abstract: Emotion-Cause Pair Extraction (ECPE) was introduced to explain why an emotion occurs, but this goal is now often reduced to binary pair/non-pair predict

TIGER: Text-Informed Generalized Enzyme-Reaction Retrieval

ResearchDGX agent

arXiv:2605.24489v1 Announce Type: new Abstract: Enzyme-reaction retrieval is a fundamental problem in computational biology, underpinning enzyme characterization, reaction mechanism elucidation, and t

.@tinkerapi by TML is a highly underrated product. It should be part of the “starter post training” infra for any newco

ResearchDGX agent

TinkerAPI, developed by TML, is a post-training infrastructure tool that Soumith Chintala recommends as an essential component for new AI companies during their startup phase. The tool appears to addr

TinyFormer: Preserving Tiny Objects in YOLO-DETRHybridReal-time Detectors

ResearchDGX agent

arXiv:2605.25046v1 Announce Type: cross Abstract: YOLO-series and DETR-based detectors struggle with tiny-object detection. YOLO-style models benefit from efficient dense prediction, but their large-s

Towards end-to-end LLM-based censoring-aware survival analysis

ResearchDGX agent

arXiv:2605.25399v1 Announce Type: new Abstract: Objective: Survival analysis is central to medical prediction, yet large language models (LLMs) are rarely used as end-to-end survival models because ce

Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild

ResearchDGX agent

arXiv:2605.24213v1 Announce Type: cross Abstract: Evaluation harnesses are software systems that orchestrate model evaluation by managing model invocation, data loading, metric computation, and result

Towards Long-Horizon Interpretability: Efficient and Faithful Multi-Token Attribution for Reasoning LLMs

ResearchDGX agent

arXiv:2602.01914v2 Announce Type: replace Abstract: Token attribution methods provide intuitive explanations for language model outputs by identifying causally important input tokens. However, as mode

TRAFA: Anticipating User Actions to Reduce Errors in Procedural Tasks with Predictive Feedback

ResearchDGX agent

arXiv:2605.24526v1 Announce Type: cross Abstract: Interactive assistance systems typically provide feedback after an action has been completed, supporting error recovery but not preventing the error i

Train-Free Segmentation in MRI with Cubical Persistent Homology

ResearchDGX agent

arXiv:2401.01160v3 Announce Type: replace-cross Abstract: We investigate a framework for train-free MRI segmentation based on Topological Data Analysis. The pipeline proceeds in three steps, first ide

Trained quantum neural networks are Gaussian processes

ResearchDGX agent

arXiv:2402.08726v2 Announce Type: replace-cross Abstract: We study quantum neural networks made by parametric one-qubit gates and fixed two-qubit gates in the limit of infinite width, where the genera

Trajectory-Based Difficulty Scoring for Reliable Learning on Tabular Data

ResearchDGX agent

arXiv:2605.24680v1 Announce Type: new Abstract: Gradient-boosted trees achieve strong performance on tabular data, yet often leave a long tail of poorly predicted instances. We introduce a Trajectory-

Transformer-based few-shot learning for modeling Electricity Consumption Profiles with minimal data across thousands of domains

ResearchDGX agent

arXiv:2408.08399v3 Announce Type: replace Abstract: Electricity Consumption Profiles (ECPs) are crucial for operating and planning power distribution systems, especially with the increasing number of

Triplet-Block Diffusion RWKV

ResearchDGX agent

arXiv:2605.25969v1 Announce Type: new Abstract: Causal Transformer language models suffer from strictly sequential decoding and a quadratic per-step attention cost. While linear-time causal models and

TSFLora: Token-Compressed Split Fine-Tuning for Wireless Edge Networks

ResearchDGX agent

arXiv:2605.23988v1 Announce Type: cross Abstract: Adapting large AI models (LAMs) to personalized edge data is challenging because wireless devices have limited memory, computation, and uplink capacit

TUBE: Tangent Upper Bound on Evidence for Discrete Diffusion Language Models

ResearchDGX agent

arXiv:2605.24292v1 Announce Type: new Abstract: Log-likelihood is a standard metric for evaluating generative models. Unfortunately, in contrast to autoregressive models (ARMs), discrete diffusion mod

Turn-Based Structural Triggers: Prompt-Free Backdoors in Multi-Turn LLMs

ResearchDGX agent

arXiv:2601.14340v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are widely integrated into interactive systems such as dialogue agents and task-oriented assistants. This growing

Uncovering Autoregressive LLM Knowledge of Thematic Fit in Event Representation

ResearchDGX agent

arXiv:2410.15173v4 Announce Type: replace-cross Abstract: The thematic fit estimation task measures semantic arguments' compatibility with a given semantic role for a given predicate. We investigate i

Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence

ResearchDGX agent

arXiv:2509.23573v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to help security analysts manage the surge of cyber threats, automating tasks from vulnerab

Understanding, Accelerating, and Improving MeanFlow Training

ResearchDGX agent

arXiv:2511.19065v2 Announce Type: replace-cross Abstract: MeanFlow promises high-quality generative modeling in few steps, by jointly learning instantaneous and average velocity fields. Yet, the under

Understanding the Impact of Geometric Foundation Models on Vision-Language-Action Models

ResearchDGX agent

arXiv:2605.24642v1 Announce Type: cross Abstract: Recent work explores new opportunities at the intersection of vision-language-action models (VLAs) and geometric foundation models (GFMs) for 3D recon

Universal Activation Verbalizer: A Unified Framework for Cross-Model Activation Explanation

ResearchDGX agent

arXiv:2605.25903v1 Announce Type: new Abstract: Activation verbalization explains hidden representations in natural language, but existing methods are mostly limited to self-explanation, where each mo

Variable Clustering via Distributionally Robust Nodewise Regression

ResearchDGX agent

arXiv:2212.07944v3 Announce Type: replace Abstract: We study a multi-factor block model for variable clustering and connect it to regularized subspace clustering through a distributionally robust vers

Verified SHAP: Provable Bounds for Exact Shapley Values of Neural Networks

ResearchDGX agent

arXiv:2605.24084v1 Announce Type: cross Abstract: Shapley additive explanations (SHAP) are widely recognised as computationally intractable for neural networks, since they induce an exponential search

Volatility Surface Reconstruction using Deep Learning under No-Arbitrage Constraints

ResearchDGX agent

arXiv:2605.24031v1 Announce Type: cross Abstract: We study the reconstruction of implied volatility surfaces from sparse and noisy option quotes using deep learning models under no-arbitrage constrain

Voting with the Graph: Stable RLAIF via Topological Consistency Maximization

ResearchDGX agent

arXiv:2510.15514v3 Announce Type: replace Abstract: Reinforcement Learning from AI Feedback (RLAIF) relies on LLM judges as preference measurement instruments, yet these instruments are fundamentally

What Questions Should Robots Be Able to Answer? A Dataset of User Questions for Explainable Robotics

ResearchDGX agent

arXiv:2510.16435v2 Announce Type: replace-cross Abstract: With the growing use of large language models and conversational interfaces in human-robot interaction, robots' ability to answer user questio

When Does Synthetic Patent Data Help? Volume-Fidelity Trade-offs in Low-Resource Multi-Label Classification

ResearchDGX agent

arXiv:2605.24296v1 Announce Type: new Abstract: We study when LLM-generated synthetic data helps low-resource multi-label patent classification, separating true synthetic value from the confound that

When Gradients Collide: Failure Modes of Multi-Objective Prompt Optimization for LLM Judges

ResearchDGX agent

arXiv:2605.26046v1 Announce Type: cross Abstract: Customizing an LLM judge to a specific task or domain often involves optimizing its prompt across multiple evaluation criteria simultaneously. Textual

When In-Distribution Gains Fail: Evaluating Weak-to-Strong Reward Models under Preference Shift

ResearchDGX agent

arXiv:2605.25629v1 Announce Type: new Abstract: Weak-to-strong (W2S) generalization is a promising framework for scalable oversight, yet existing evaluations often test students under matched train--t

Which Is Better For Reducing Outdated and Vulnerable Dependencies: Pinning or Floating?

ResearchDGX agent

arXiv:2510.08609v3 Announce Type: replace-cross Abstract: Developers consistently use version constraints to specify acceptable versions of the dependencies for their project. Pinning dependencies can

WINO: A Weak-Form Physics Informed Neural Operator for Hyperelasticity on Variable Domains

ResearchDGX agent

arXiv:2605.24651v1 Announce Type: cross Abstract: We propose a Weak-form Physics-Informed Neural Operator (WINO), a data-free framework that combines the efficiency of neural operators with the geomet

Word Class Representations Spontaneously Emerge from Successor Representations Trained on Natural Language

ResearchDGX agent

arXiv:2605.24585v1 Announce Type: new Abstract: Language models are typically trained to predict the next token in a sequence. Here, we explore an alternative predictive principle from reinforcement l

WTKO-CNN: Deep Learning Reveals Sequence Motifs Distinguishing Wild-Type and Knockout ATAC-seq Peaks

ResearchDGX agent

arXiv:2605.24034v1 Announce Type: cross Abstract: Chromatin regulators can alter transcriptional programs by modifying the accessibility of regulatory DNA elements. Understanding how regulatory sequen

You Can Ground Earlier than See: An Effective and Efficient Pipeline for Temporal Sentence Grounding in Compressed Videos

ResearchDGX agent

arXiv:2303.07863v3 Announce Type: replace-cross Abstract: Given an untrimmed video, temporal sentence grounding (TSG) aims to locate a target moment semantically according to a sentence query. Althoug

Zero-Shot Parkinson's Disease Detection from Speech: Comparing Large Audio and Language Models

ResearchDGX agent

arXiv:2605.24806v1 Announce Type: cross Abstract: Large audio and language models have recently demonstrated zero-shot reasoning capabilities across various domains. However, it remains unclear how th

Zeroth-Order Nonconvex Nonsmooth Optimization with Heavy-Tailed Noise

ResearchDGX agent

arXiv:2605.24513v1 Announce Type: new Abstract: This paper considers the nonconvex nonsmooth problem in which the objective function is Lipschitz continuous. We focus on the stochastic setting where t

25 May 2026

5月25日、内閣総理大臣官邸で開催された車座対談に、Sakana AIの伊藤錬が参加しました。 金融など各産業に特化した高度なAI実装の実例をご紹介しました。また、海外の先端モデルも活用しつつ、日本独自の技術で、我が国の防衛の自律性やデータ主権を確保する具体的方法についても高市総…

ResearchDGX agent

5月25日、内閣総理大臣官邸で開催された車座対談に、Sakana AIの伊藤錬が参加しました。 金融など各産業に特化した高度なAI実装の実例をご紹介しました。また、海外の先端モデルも活用しつつ、日本独自の技術で、我が国の防衛の自律性やデータ主権を確保する具体的方法についても高市総理と貴重な意見 交換の機会をいただきました。 本日、スタートアップ4社(Atomis、Sakana AI、Oceanic

A comprehensive evaluation of pretraining strategies for channel-agnostic contrastive self-supervision of biosignals

ResearchDGX agent

arXiv:2410.19842v2 Announce Type: replace-cross Abstract: Contrastive learning yields impressive results for self-supervision in computer vision. The approach relies on the creation of positive pairs,

A drone-based framework for coral habitat mapping via weakly supervised segmentation

ResearchDGX agent

arXiv:2508.18958v2 Announce Type: replace-cross Abstract: Obtaining pixel-level annotations over large spatial extents remains a major bottleneck for deploying machine learning in ecological applicati

A Fine-Tuned BERT Classifier for Personal-Letter Titles in Late-Ming and Early-Qing Collected Works

ResearchDGX agent

arXiv:2605.23103v1 Announce Type: cross Abstract: I present Lepton (Letter Prediction), a fine-tuned BERT classifier that predicts whether a title in a Classical Chinese wenji table of contents is a p

A graph-based analysis of semantic types and coercion in contextualized word embeddings

ResearchDGX agent

arXiv:2605.23710v1 Announce Type: new Abstract: Semantic type mismatch between a noun and its context is central to coercion phenomena. This paper introduces a graph-based method to examine how lexica

A Simple Plug-in for Improving Eviction-Based KV Cache Compression

ResearchDGX agent

arXiv:2605.23258v1 Announce Type: new Abstract: KV cache growth is a major bottleneck for long-context inference in large language models. Existing methods are often dominated by binary eviction or re

A solution to generalized learning from small training sets found in infant repeated visual experiences of individual objects

ResearchDGX agent

arXiv:2510.15060v3 Announce Type: replace Abstract: One-year-old infants rapidly form and generalize categories of the everyday objects they encounter. Here we provide evidence on infants daily-life v

A Survey of Text and Speech Resources for Hausa and Fongbe: Availability, Quality, and Gaps for NLP Development

ResearchDGX agent

arXiv:2605.22828v1 Announce Type: new Abstract: This survey provides a comprehensive catalog of publicly available text and speech resources for two West African languages: Hausa, an Afroasiatic langu

Accelerating ground state search of spatial photonic Ising machines with genetic-simulated annealing hybrid algorithm

ResearchDGX agent

arXiv:2605.23295v1 Announce Type: cross Abstract: Spatial photonic Ising machines (SPIMs) based on spatial light modulators (SLMs) have emerged as highly effective solvers for many tasks, including co

Active Sensing Subserves Task-Level Control

ResearchDGX agent

arXiv:2605.22988v1 Announce Type: cross Abstract: Active sensing is traditionally defined as the expenditure of energy, typically in the form of movement, for obtaining information. Here, we propose t

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning

ResearchDGX agent

arXiv:2605.23200v1 Announce Type: cross Abstract: The linear growth of the Key-Value (KV) cache is a critical bottleneck in long-form LLM inference. Existing KV compression methods mitigate this by ev

Adversarial Vulnerability Under Temporal Concept Drift: A Longitudinal Study of Android Malware Detection

ResearchDGX agent

arXiv:2605.23623v1 Announce Type: cross Abstract: We present a longitudinal, drift-aware evaluation of adversarial robustness across more than a decade of Android applications using static and dynamic

AGZO: Activation-Guided Zeroth-Order Optimization for LLM Fine-Tuning

ResearchDGX agent

arXiv:2601.17261v4 Announce Type: replace Abstract: Zeroth-Order (ZO) optimization has emerged as a promising solution for fine-tuning LLMs under strict memory constraints, as it avoids the prohibitiv

← Previous
1…220221222223224…435
Next →