AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
28 May 2026

SafeMed-R1: Clinician-Audited Safety and Ethics Alignment for Medical Large Language Models

SafetyDGX agent

arXiv:2605.28338v1 Announce Type: new Abstract: Large language models(LLMs) increasingly match expert performance on licensing examinations, yet routine clinical use remains limited because governance

SANTS: A State-Adaptive Scheduler for World Action Models

ResearchDGX agent

arXiv:2605.27947v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot manipulation by using video-based future representations to condition action generation. In pixel-space WAMs, h

Semantic-level Backdoor Attack against Text-to-Image Diffusion Models

ResearchDGX agent

arXiv:2602.04898v3 Announce Type: replace-cross Abstract: Text-to-image (T2I) diffusion models are widely adopted for their strong generative capabilities, yet remain vulnerable to backdoor attacks. E

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models

ResearchDGX agent

arXiv:2510.20665v3 Announce Type: replace Abstract: Evaluating the quality of reasoning traces from large language models remains understudied, labor-intensive, and unreliable: current practice relies

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable …

HardwareDGX agent

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable consumer GPU can do LLM fine-tuning will make a video on thi

27 May 2026

A Unified Framework for Diffusion Model Unlearning with f-Divergence

ResearchDGX agent

arXiv:2509.21167v2 Announce Type: replace-cross Abstract: Most existing methods for concept unlearning in text-to-image diffusion models minimize a mean squared error (MSE) loss between the denoiser o

Agile Online Model Selection: Resolving Adaptation Lag via Safeguarded Large Learning Rates

ApplicationsDGX agent

arXiv:2605.26919v1 Announce Type: new Abstract: Maintaining predictive accuracy in non-stationary environments requires online model selection to adapt autonomously to unknown distribution shifts. How

Agreement Between Large Language Models and Human Raters in Essay Scoring: A Research Synthesis

ResearchDGX agent

arXiv:2512.14561v2 Announce Type: replace Abstract: Despite the growing promise of large language models (LLMs) in automated essay scoring (AES), empirical findings regarding their reliability compare

An In-Vitro Study on Cross-Lingual Generalization in Language Models

ResearchDGX agent

arXiv:2605.26683v1 Announce Type: cross Abstract: Cross-lingual transfer in language models is difficult to study in natural corpora because lexical overlap, morphology, data imbalance, and tokenizati

Clozing the Gap: Exploring Why Language Model Surprisal Outperforms Cloze Surprisal

ResearchDGX agent

arXiv:2601.09886v2 Announce Type: replace Abstract: How predictable a word is can be quantified in two ways: using human responses to the cloze task or using probabilities from language models (LMs).W

Conv-to-Bench: Evaluating Language Models Via User-Assistant Dialogues In Code Tasks

SafetyDGX agent

arXiv:2605.26440v1 Announce Type: new Abstract: The rapid advancement of Large Language Models (LLMs) has outpaced the scalability of traditional evaluation benchmarks, which remain heavily dependent

DinoComplete: 3D Shape Completion with Distilled Semantic Priors and State Space Models

ApplicationsDGX agent

arXiv:2605.26949v1 Announce Type: new Abstract: 3D shape completion from partial scans remains challenging for unseen categories and noisy real-world observations, where geometry alone is often insuff

Explainable Comparison of Feature-Based and Deep Learning Models for TROPOMI Methane Plume Screening

ResearchDGX agent

arXiv:2605.27236v1 Announce Type: new Abstract: Continuous and global detection of large methane emissions is a crucial step for global warming mitigation. Satellite observations, such as from S5P/TRO

FTibSuite: A Comprehensive Resource Suite for Tibetan Vision-Language Modeling

SafetyDGX agent

arXiv:2605.26601v1 Announce Type: new Abstract: Vision-language models have progressed rapidly, but Tibetan remains a severely underserved low-resource language due to the lack of reproducible trainin

GICDM: Mitigating Hubness for Reliable Distance-Based Generative Model Evaluation

SafetyDGX agent

arXiv:2602.16449v2 Announce Type: replace-cross Abstract: Generative model evaluation commonly relies on high-dimensional embedding spaces to compute distances between samples. We show that dataset re

KREA 2 Image is now a Partner Node in ComfyUI KREA's first foundation image model — trained from scratch — with tunable creativity, style re…

Local AiDGX agent

KREA 2 Image, KREA's foundational image generation model trained from scratch with adjustable creativity and style parameters, has been integrated as a partner node within ComfyUI. This integration al

KZ-SafetyPrompts: A Kazakh Safety Evaluation Prompt Dataset for Large Language Models

SafetyDGX agent

arXiv:2605.26947v1 Announce Type: new Abstract: Kazakh is underrepresented in resources for evaluating the safety behavior of large language models. We present KZ-SafetyPrompts, a Kazakh prompt datase

On the Robustness of Machine Unlearning for Vision-Language Models

ResearchDGX agent

arXiv:2605.26992v1 Announce Type: new Abstract: Vision-language models (VLMs) may memorize undesirable information from training data, motivating growing interest in machine unlearning. In this work,

Personalizing Embodied Multimodal Large Language Model Agents over Long-term User Interactions

AgentsDGX agent

arXiv:2605.26256v1 Announce Type: new Abstract: Multimodal large language model (MLLM)-based embodied agents have shown strong potential for solving complex tasks in physical environments. However, pe

RT-Lynx: Putting the GEMM Sparsity In a Right Way for Diffusion Models

HardwareDGX agent

arXiv:2605.26632v1 Announce Type: new Abstract: Diffusion Transformers (DiT) achieve strong performance in image generation but incur substantial inference costs. While prior work has reduced this cos

Stabilizing Recurrent Dynamics for Test-Time Scalable Latent Reasoning in Looped Language Models

ResearchDGX agent

arXiv:2605.26733v1 Announce Type: cross Abstract: Looped Language Models (LoopLMs) enable efficient latent reasoning through depth recurrence, yet exhibit unreliable test-time scaling behavior: perfor

The AI Cognitive Trojan Horse: How Large Language Models May Bypass Human Epistemic Vigilance

SafetyDGX agent

arXiv:2601.07085v2 Announce Type: replace-cross Abstract: Large language model (LLM)-based conversational AI systems present a challenge to human cognition that current frameworks for understanding mi

The best research labs are building what comes after static models. Congrats to @trajectorylabs on the launch! Excited to have them training…

ToolsDGX agent

The best research labs are building what comes after static models. Congrats to @trajectorylabs on the launch! Excited to have them training on the AI Native Cloud as they push the frontier on Continu

The Labyrinth and the Thread: Rethinking Regularizations in Sequential Knowledge Editing for Large Language Models

SafetyDGX agent

arXiv:2605.26670v1 Announce Type: cross Abstract: Sequential editing of structured knowledge in large language models allows targeted factual updates without retraining, yet existing methods often rel

The open source community has been delivering on LTX 2.3 LoRAs. Fine tuning LTX 2.3 unlocks control that you can't get from closed models. H…

Local AiDGX agent

The open source community has been delivering on LTX 2.3 LoRAs. Fine tuning LTX 2.3 unlocks control that you can't get from closed models. Here are 7 LoRAs that can save your footage. There are too ma

TSFMAudit: Data Contamination Auditing in Forecasting Time Series Foundation Models

ResearchDGX agent

arXiv:2605.26161v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) are increasingly pretrained on large corpora, raising concerns that evaluation datasets may have been exposed du

26 May 2026

A Signal-Language Foundation Model for Broad-Spectrum Cardiovascular Assessment from Routine Electrocardiography

ResearchDGX agent

arXiv:2605.25446v1 Announce Type: new Abstract: Electrocardiography (ECG) is central to cardiovascular care, but conventional AI models are often restricted to common arrhythmias and may generalize po

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models

SafetyDGX agent

arXiv:2605.26013v1 Announce Type: cross Abstract: We introduce AdvantageFlow, a forward-process reinforcement learning algorithm for rectified flow models. Unlike Flow-GRPO, which optimizes the revers

Boosting Inference with Guided Reasoning: Stochastic Exploration for Recursive Models

Local AiDGX agent

arXiv:2605.25230v1 Announce Type: new Abstract: Recent work on recursive architectures has shown that tiny neural networks can be surprisingly powerful on structured reasoning tasks. The trick is to m

Confidence and Calibration of Activation Oracles for Reliable Interpretation of Language Model Internals

ResearchDGX agent

arXiv:2605.26045v1 Announce Type: cross Abstract: Activation oracles aim to make the activations of other models legible to humans and yield promising results compared to white-box interpretability te

Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation

ResearchDGX agent

arXiv:2601.21406v3 Announce Type: replace-cross Abstract: Unified Multimodal Models (UMMs) integrate both visual understanding and generation within a single framework. Their ultimate aspiration is to

Generative modeling of granular flow on inclined planes using conditional flow matching

ResearchDGX agent

arXiv:2604.04453v2 Announce Type: replace-cross Abstract: Granular flows govern many natural and industrial processes, yet their interior kinematics and mechanics remain largely unobservable, as exper

Harmony in Diversity: Multi-domain Contrastive Policy Optimization for Large Reasoning Models

SafetyDGX agent

arXiv:2605.25443v1 Announce Type: new Abstract: Post-training has significantly enhanced the reasoning capability of Large Reasoning Models (LRMs), especially with Reinforcement Learning (RL) like Gro

Human in the loop 🤖 Made in @ComfyUI with a laundry list of tools: @ltx_model 2.3 + @thesystms FLW LoRA, @Alibaba_Wan 2.2 I2V + T2V, Floren…

Local AiDGX agent

Human in the loop 🤖 Made in @ComfyUI with a laundry list of tools: @ltx_model 2.3 + @thesystms FLW LoRA, @Alibaba_Wan 2.2 I2V + T2V, Florence2, @Meta Sapiens2, @OpenAI GPT Image 2.0, @suno , @AdobeAE

HyperGuide: Hyperbolic Guidance for Efficient Multi-Step Reasoning in Large Language Models

TutorialsDGX agent

arXiv:2605.24140v1 Announce Type: new Abstract: Multi-step reasoning remains a central challenge for large language models: single-pass generation is efficient but lacks accuracy; tree-search methods

Infinite context windows seem to present a very large problem to using AI. Today's models already leak too much old information into current…

ApplicationsDGX agent

Infinite context windows seem to present a very large problem to using AI. Today's models already leak too much old information into current responses, a distraction that is part of why they are cogni

Language Movement Primitives: Grounding Language Models in Robot Motion

ApplicationsDGX agent

arXiv:2602.02839v3 Announce Type: replace Abstract: Enabling robots to perform novel manipulation tasks from natural language instructions remains a fundamental challenge in robotics, despite signific

Launching our new paper on arXiv: we trained the largest multilingual food model ever built. 4.1M recipes. 7 languages. 1,790 ingredients. 3…

AgentsDGX agent

Launching our new paper on arXiv: we trained the largest multilingual food model ever built. 4.1M recipes. 7 languages. 1,790 ingredients. 300 dimensions. All of human cooking compressed into 2 megaby

Locality Matters for Training-Free Audio Token Compression in Audio-Language Models

SafetyDGX agent

arXiv:2605.25179v1 Announce Type: new Abstract: Audio-language models (ALMs) are increasingly used for audio captioning, question answering, and open-ended audio understanding, but their inference cos

Model test: SenseNova U1 vs GPT Image2 vs Nano Banana in Infographic generation

Local AiDGX agent

A comparative test of three AI image generation models—SenseNova U1, GPT Image 2, and Nano Banana 2—evaluating their performance on infographic generation tasks. GPT Image 2 proved more reliable for e

OmniSapiens: A Foundation Model for Social Behavior Processing via Heterogeneity-Aware Relative Policy Optimization

SafetyDGX agent

arXiv:2602.10635v2 Announce Type: replace Abstract: Socially intelligent AI systems must entail reasoning across diverse human behavioral tasks, and generalization to new contexts. However, AI has yet

PageLLM: A Multi-Grained Reward Framework for Whole-Page Optimization with Large Language Models

SafetyDGX agent

arXiv:2506.09084v2 Announce Type: replace-cross Abstract: Whole-page optimization (WPO) decides how search and recommendation results are surfaced to users, and large language models (LLMs) open a new

SPA-Cache: Singular Proxies for Adaptive Caching in Diffusion Language Models

ResearchDGX agent

arXiv:2602.02544v2 Announce Type: replace-cross Abstract: While Diffusion Language Models (DLMs) offer a flexible, arbitrary-order alternative to the autoregressive paradigm, their non-causal nature p

The goal of training a machine learning model is loss minimization. That was just being applied to algorithms, and now it's being applied to…

IndustryDGX agent

The goal of training a machine learning model is loss minimization. That was just being applied to algorithms, and now it's being applied to whole enterprises. Loss minimization is one thing that the

The Normalized Maximum Likelihood for Regular Non-Smooth Models: Measure-Theoretic Foundations and Geometric Sampling

ResearchDGX agent

arXiv:2605.24477v1 Announce Type: new Abstract: The Normalized Maximum Likelihood (NML) codelength, or stochastic complexity, represents a principled criterion for universal coding. While recent coare

Uncertainty Reasoning with Large Language Models for Explainable Disease Diagnosis

TutorialsDGX agent

arXiv:2605.25566v1 Announce Type: new Abstract: Clinical decision-making requires reasoning over incomplete, imprecise, and linguistically expressed patient narratives. While large language models (LL

Your Embedding Model is SMARTer Than You Think

Local AiDGX agent

arXiv:2605.24938v1 Announce Type: cross Abstract: Multimodal retrieval relies heavily on single-vector retrievers, which compress rich, sequential token sequences into one single global representation

25 May 2026

Causal Additive Models with Unobserved Causal Paths and Backdoor Paths

ResearchDGX agent

arXiv:2502.07646v3 Announce Type: replace Abstract: Causal additive models provide a tractable yet expressive framework for causal discovery in the presence of hidden variables. When unobserved backdo

Disentangling Interaction and Bias Effects in Opinion Dynamics of Large Language Models

SafetyDGX agent

arXiv:2509.06858v2 Announce Type: replace-cross Abstract: Large Language Models are increasingly used to simulate human opinion dynamics, yet the effect of genuine interaction is often obscured by sys

Do Language Models Know What Not to Say? Causal Evidence for Statistical Preemption in LLMs

ResearchDGX agent

arXiv:2605.23039v1 Announce Type: cross Abstract: How do learners acquire knowledge of what is unacceptable without negative evidence? Construction Grammar proposes statistical preemption: exposure to

Empirical Bayes Conformal Prediction for Vision and Language Models

ResearchDGX agent

arXiv:2605.23189v1 Announce Type: new Abstract: Conformal prediction (CP) gives distribution-free coverage for modern vision and language models, but it is often forced to make a ranking decision from

Entropy-Aware On-Policy Distillation of Language Models

SafetyDGX agent

arXiv:2603.07079v2 Announce Type: replace-cross Abstract: On-policy distillation is a promising approach for transferring knowledge between language models, where a student learns from dense token-lev

Evaluating Counterfactual Strategic Reasoning in Large Language Models

ResearchDGX agent

arXiv:2603.19167v2 Announce Type: replace Abstract: We evaluate Large Language Models (LLMs) in repeated game-theoretic settings to assess whether strategic performance reflects genuine reasoning or r

GILT: An LLM-Free, Tuning-Free Graph Foundational Model for In-Context Learning

ResearchDGX agent

arXiv:2510.04567v2 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) are powerful tools for processing relational data but often struggle to generalize to unseen graphs, giving rise

Grok foundation model V9-Medium (1.5T) has finished training. Evals look good. A lot of Cursor data was added in supplementary training and …

ApplicationsDGX agent

Grok foundation model V9-Medium (1.5T) has finished training. Evals look good. A lot of Cursor data was added in supplementary training and there is more to come. Fine-tuning is underway and reinforce

Improving Sampling for Masked Diffusion Models via Information Gain

ResearchDGX agent

arXiv:2602.18176v3 Announce Type: replace Abstract: Masked Diffusion Models (MDMs) enable flexible decoding orders, yet existing samplers remain largely greedy, selecting locally certain tokens withou

Operationalizing Individual Fairness via Gradient Descent and Bradley-Terry Models

SafetyDGX agent

arXiv:2605.23145v1 Announce Type: cross Abstract: Individual fairness, the notion that 'similar individuals should be treated similarly,' provides a strong and flexible fairness guarantee for algorith

Preisach Attention: A Hysteretic Model of Sequential Memory

Local AiDGX agent

arXiv:2605.23603v1 Announce Type: cross Abstract: We introduce the Preisach Attention Layer (PAL), a novel sequence modelling architecture grounded in the classical Preisach hysteresis operator from m

Real-Time Earthquake Magnitude Classification from Initial P-Waves: Models, Dataset, and Comparative Analysis for South Asia

ResearchDGX agent

arXiv:2605.22836v1 Announce Type: cross Abstract: Rapid earthquake magnitude estimation is crucial for effective early warning systems that can save lives and reduce economic damage. In this paper, we

Scalable Heterogeneous Graph Foundation Models for Data-Driven Optimal Power Flow in Smart Grids

ResearchDGX agent

arXiv:2605.23194v1 Announce Type: cross Abstract: Fast and reliable optimal power flow (OPF) approximation is essential for reliable smart-grid operation, yet many learning-based surrogates either fla

← Previous
1…214215216217218…1017
Next →