AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Model Releases

Do LLMs Build World Models From Text? A Multilingual Diagnostic of Spatial Reasoning

DGX agent

arXiv:2605.28277v1 Announce Type: new Abstract: Whether large language models (LLMs) construct internal spatial world models from pure-text descriptions remains contested, and whether such capabilitie

model-releasesarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Towards Faithful Agentic XAI: A Verification Method and an Open-World Benchmark for Better Model Faithfulness

DGX agent

arXiv:2605.27879v1 Announce Type: new Abstract: Explainable AI (XAI) helps users interpret model behavior and identify potential faults. Agentic XAI systems use Large Language Models (LLMs) to make ex

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Benchmarking Convolutional, Transformer, Hybrid, and Vision Language Models for Multi Disease Retinal Screening

DGX agent

arXiv:2605.26283v1 Announce Type: new Abstract: Modern deep learning offers powerful tools for automated retinal screening, but it remains unclear how different visual model families compare in realis

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Black-box Membership Inference Attacks on the Pre-training Data of Image-generation Models

DGX agent

arXiv:2605.27020v1 Announce Type: cross Abstract: The rapid advancement of diffusion-based image generation models has raised serious concerns regarding potential copyright and privacy infringements i

model-releasesarxiv-cs-ai
27 May 2026
Safety

Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization

DGX agent

arXiv:2605.26282v1 Announce Type: new Abstract: Model-based reinforcement learning (RL) can be effectively supported at scale through the use of world models. However, in practice, scaling such approa

safetyarxiv-cs-lg
27 May 2026
Model Releases

Benchmarking Patent Embeddings: A Multi-Task Evaluation of 22 Models Across Retrieval, Classification, and Clustering

DGX agent

arXiv:2605.24297v1 Announce Type: cross Abstract: Which fine-tuning signals improve patent embedding models, and do gains transfer across patent landscapes? We benchmark 22 embedding models, from 22M-

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

DGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

model-releasesarxiv-cs-ai
26 May 2026
Research

Mimir: Large-scale Multilingual Concept Modeling

DGX agent

arXiv:2605.25263v1 Announce Type: cross Abstract: Current language modeling approaches are built around tokens. Text corpora are split into tokens, and models are trained by performing computations on

researcharxiv-cs-ai
26 May 2026
Model Releases

SomaliBench Eval: Measuring English-to-Somali Refusal Gaps in Open-Weight Language Models

DGX agent

arXiv:2605.25420v1 Announce Type: cross Abstract: Large language model safety evaluation remains heavily English-centered, leaving low-resource languages under-measured even when models are deployed g

model-releasesarxiv-cs-ai
26 May 2026
Agents

Why We Need World Models for AGI: Where LLMs Fail and How World Models May Outperform

DGX agent

arXiv:2605.23972v1 Announce Type: new Abstract: Large language models achieve strong performance in language generation and knowledge-intensive tasks, yet remain limited in settings requiring causal r

agentsarxiv-cs-ai
26 May 2026
Model Releases

A Comparative Evaluation of Structural Topic Models and BERTopic for Short, Open-Ended Survey Responses

DGX agent

arXiv:2605.23093v1 Announce Type: new Abstract: Topic modeling in applied psychology increasingly spans two methodological traditions: probabilistic bag-of-words models and newer embedding-based appro

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models

DGX agent

arXiv:2605.23238v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as economic agents in marketplaces, auctions, and bidding settings. Anticipating their behavior i

model-releasesarxiv-cs-ai
25 May 2026
Agents

Latent Cache Flow: Model-to-Model Communication Without Text

DGX agent

arXiv:2605.22863v1 Announce Type: new Abstract: LLM agents today communicate via text, which incurs considerable latency and information loss due to the need to autoregressively decode the sharer mode

agentsarxiv-cs-lg
25 May 2026
Local Ai

The Attribution Contract: Feature Attribution for Generative Language Models

DGX agent

arXiv:2605.23080v1 Announce Type: new Abstract: Feature attribution methods promise to identify which input features matter for a model output. In generative language models, however, it is often uncl

local-aiarxiv-cs-lg
25 May 2026
Model Releases

Provable Joint Decontamination for Benchmarking Multiple Large Language Models

DGX agent

arXiv:2605.21543v1 Announce Type: new Abstract: Benchmark data contamination has become a central challenge in LLM evaluation: when evaluation examples appear in the training data of one or more audit

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Hy-MT2: A Family of Fast, Efficient and Powerful Multilingual Translation Models in the Wild

DGX agent

arXiv:2605.22064v1 Announce Type: new Abstract: Hy-MT2 is a family of fast-thinking multilingual translation models designed for complex real-world scenarios. It includes three model sizes: 1.8B, 7B,

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Chronicle: A Multimodal Foundation Model for Joint Language and Time Series Understanding

DGX agent

arXiv:2605.20268v1 Announce Type: cross Abstract: Real-world time series come with text: metadata, descriptions, news, reports. Yet time series foundation models process numerical sequences in isolati

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Memory Grafting: Scaling Language Model Pre-training via Offline Conditional Memory

DGX agent

arXiv:2605.20948v1 Announce Type: new Abstract: Scaling conditional memory offers a promising way to increase language-model capacity, but existing methods such as Engram learn large memory tables fro

model-releasesarxiv-cs-cl
21 May 2026
Hardware

Towards Multi-Model LLM Schedulers: Empirical Insights into Offloading and Preemption

DGX agent

arXiv:2605.19593v1 Announce Type: new Abstract: Modern deployments of Large Language Models (LLMs) increasingly require serving multiple models with diverse architectures, sizes, and specialization on

hardwarearxiv-cs-ai
20 May 2026
Model Releases

Unlocking the Potential of Continual Model Merging: An ODE Perspective

DGX agent

arXiv:2605.19409v1 Announce Type: cross Abstract: Continual Model Merging (CMM) enables rapid customization of foundation models across sequentially arriving tasks, offering a scalable alternative to

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Extending Pretrained 10-Second ECG Foundation Models to Longer Horizons

DGX agent

arXiv:2605.16975v1 Announce Type: cross Abstract: Electrocardiogram (ECG) foundation models pretrained on typical diagnostic 10-second ECG segments, have demonstrated strong transferability across a r

model-releasesarxiv-cs-ai
19 May 2026
Research

Foundation Models for Credit Risk Prediction: A Game Changer?

DGX agent

arXiv:2605.18147v1 Announce Type: new Abstract: Predictive models play a pivotal role in credit risk management, guiding critical decisions through accurate estimation of default probabilities and los

researcharxiv-cs-lg
19 May 2026
Model Releases

GenTS: A Comprehensive Benchmark Library for Generative Time Series Models

DGX agent

arXiv:2605.17804v1 Announce Type: new Abstract: Generative models have demonstrated remarkable potential in time series analysis tasks, like synthesis, forecasting, imputation, etc. However, offering

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency

DGX agent

arXiv:2605.18732v1 Announce Type: cross Abstract: While scaling laws govern aggregate large language model performance, no scaling law has linked factual recall to both model size and training-data co

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Beyond Mode-Seeking RL: Trajectory-Balance Post-Training for Diffusion Language Models

DGX agent

arXiv:2605.13935v1 Announce Type: cross Abstract: Diffusion language models are a promising alternative to autoregressive models, yet post-training methods for them largely adapt reward-maximizing obj

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Hidden State Poisoning Attacks against Mamba-based Language Models

DGX agent

arXiv:2601.01972v4 Announce Type: cross Abstract: State space models (SSMs) like Mamba offer efficient alternatives to Transformer-based language models, with linear time complexity. Yet, their advers

model-releasesarxiv-cs-ai
15 May 2026
Safety

Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations

DGX agent

arXiv:2605.14937v1 Announce Type: cross Abstract: Predictive world models enable agents to model scene dynamics and reason about the consequences of their actions. Inspired by human perception, object

safetyarxiv-cs-ai
15 May 2026
Model Releases

Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling

DGX agent

arXiv:2605.13062v1 Announce Type: new Abstract: Recent image editing models have achieved remarkable progress in instruction following, multimodal understanding, and complex visual editing. However, e

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Understanding and Accelerating the Training of Masked Diffusion Language Models

DGX agent

arXiv:2605.13026v1 Announce Type: cross Abstract: Masked diffusion models (MDMs) have emerged as a promising alternative to autoregressive models (ARMs) for language modeling. However, MDMs are known

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

When is Warmstarting Effective for Scaling Language Models?

DGX agent

arXiv:2605.13405v1 Announce Type: new Abstract: Model growth from a given checkpoint aims to accelerate training of a larger model, offering potential resource savings. Despite recent interest, warmst

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models

DGX agent

arXiv:2605.11887v1 Announce Type: new Abstract: Large language models have achieved remarkable capabilities across diverse tasks, yet their internal decision-making processes remain largely opaque, li

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

READ: Recurrent Adapter with Partial Video-Language Alignment for Parameter-Efficient Transfer Learning in Low-Resource Video-Language Modeling

DGX agent

arXiv:2312.06950v3 Announce Type: replace-cross Abstract: Fully fine-tuning pretrained large-scale transformer models has become a popular paradigm for video-language modeling tasks, such as temporal

model-releasesarxiv-cs-cl
13 May 2026
Research

A Single-Layer Model Can Do Language Modeling

DGX agent

arXiv:2605.10643v1 Announce Type: new Abstract: Modern language models scale depth by stacking layers, each holding its own state - a per-layer KV cache in transformers, a per-layer matrix in Mamba, G

researcharxiv-cs-cl
12 May 2026
Model Releases

Aligning Agents via Planning: A Benchmark for Trajectory-Level Reward Modeling

DGX agent

arXiv:2604.08178v2 Announce Type: replace Abstract: In classical Reinforcement Learning from Human Feedback (RLHF), Reward Models (RMs) serve as the fundamental signal provider for model alignment. As

model-releasesarxiv-cs-ai
12 May 2026
Safety

dFlowGRPO: Rate-Aware Policy Optimization for Discrete Flow Models

DGX agent

arXiv:2605.09291v1 Announce Type: new Abstract: Discrete flow models (DFMs) are a class of flexible generative models for generating discrete data, and diffusion large language models (dLLMs) can be v

safetyarxiv-cs-lg
12 May 2026
Model Releases

GIFT: Guided Importance-Aware Fine-Tuning for Diffusion Language Models

DGX agent

arXiv:2509.20863v3 Announce Type: replace Abstract: Diffusion models have recently shown strong potential in language modeling, offering faster generation compared to traditional autoregressive approa

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Hunyuan3D 2.0: Scaling Diffusion Models for High Resolution Textured 3D Assets Generation

DGX agent

arXiv:2501.12202v4 Announce Type: replace Abstract: We present Hunyuan3D 2.0, an advanced large-scale 3D synthesis system for generating high-resolution textured 3D assets. This system includes two fo

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Is Your Driving World Model an All-Around Player?

DGX agent

arXiv:2605.10858v1 Announce Type: new Abstract: Today's driving world models can generate remarkably realistic dash-cam videos, yet no single model excels universally. Some generate photorealistic tex

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Model-Free Neural Filtering: A Comparison with Classical Filters in Nonlinear Systems

DGX agent

arXiv:2601.21266v3 Announce Type: replace Abstract: Neural network models are increasingly used for state estimation in control and decision-making, yet it remains unclear to what extent they behave a

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Benchmarking World-Model Learning with Environment-Level Queries

DGX agent

arXiv:2510.19788v4 Announce Type: replace Abstract: World models are central to building AI agents capable of flexible reasoning and planning. Yet current evaluations (i) test only properties measurab

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Fine-tuning a vision-language model for fracture-surface morphology recognition

DGX agent

arXiv:2605.07145v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown strong potential for scientific image understanding, but general-purpose models often lack the domain-specifi

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

A Domain Incremental Continual Learning Benchmark for ICU Time Series Model Transportability

DGX agent

arXiv:2605.03832v1 Announce Type: new Abstract: In recent years, machine learning has made significant progress in clinical outcome prediction, demonstrating increasingly accurate results. However, th

model-releasesarxiv-cs-lg
6 May 2026
Research

Mechanism-Faithful Queueing Simulation Model Translation with Large Language Model Support

DGX agent

arXiv:2601.06543v2 Announce Type: replace Abstract: Queueing simulation studies often require substantial manual effort to translate conceptual system descriptions into executable programs and to veri

researcharxiv-cs-cl
6 May 2026
Model Releases

StateVLM: A State-Aware Vision-Language Model for Robotic Affordance Reasoning

DGX agent

arXiv:2605.03927v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown remarkable performance in various robotic tasks, as they can perceive visual information and understand natural

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

CNN-based Multi-In-Multi-Out Model for Efficient Spatiotemporal Prediction

DGX agent

arXiv:2605.01277v1 Announce Type: new Abstract: Recently, Convolutional Neural Network (CNN) or Transformer architecture based models have been proposed to overcome the limitations of Recurrent Neural

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Dispersion Loss Counteracts Embedding Condensation and Improves Generalization in Small Language Models

DGX agent

arXiv:2602.00217v2 Announce Type: replace Abstract: Large language models (LLMs) achieve remarkable performance through ever-increasing parameter counts, but scaling incurs steep computational costs.

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Fine-Tuning Impairs the Balancedness of Foundation Models in Long-tailed Personalized Federated Learning

DGX agent

arXiv:2605.02247v1 Announce Type: new Abstract: Personalized federated learning (PFL) with foundation models has emerged as a promising paradigm enabling clients to adapt to heterogeneous data distrib

model-releasesarxiv-cs-cv
5 May 2026
Tutorials

GIFT: Guided Fine-Tuning and Transfer for Enhancing Instruction-Tuned Language Models

DGX agent

arXiv:2605.01256v1 Announce Type: new Abstract: A promising paradigm for adapting instruction-tuned language models is to learn task-specific updates on a pretrained base model and subsequently merge

tutorialsarxiv-cs-cl
5 May 2026
← Previous
1…1314151617…1012
Next →