AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,399 results
11 Jun 2026

Generalization Hacking: Models Can Game Reinforcement Learning by Preventing Behavioral Generalization

ResearchDGX agent

arXiv:2606.12016v1 Announce Type: cross Abstract: Model post-training, and in particular reinforcement learning (RL), is one of the primary mechanisms by which developers can shape models' values and

Making Models Unmergeable via Scaling-Sensitive Loss Landscape

Model ReleasesDGX agent

arXiv:2601.21898v2 Announce Type: replace Abstract: The rise of model hubs has made it easier to access reusable model components, making model merging a practical tool for combining capabilities. Yet

Mapping Scientific Literature with Large Language Models and Topic Modeling

ResearchDGX agent

arXiv:2510.16152v2 Announce Type: replace-cross Abstract: Scientific literature is increasingly fragmented by disciplinary boundaries, specialized terminology, and potentially sparse keyword systems,

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Physics-Distilled Neural Network enabled by Large Language Models for Manufacturing Process-Property Predictive Modeling

ApplicationsDGX agent

arXiv:2606.11605v1 Announce Type: cross Abstract: Predicting process-property relationships in manufacturing is often challenged by high experimental costs and the limited interpretability of complex

9 Jun 2026

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling

SafetyDGX agent

arXiv:2606.07988v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on reward models to align their outputs with diverse user preferences. While personalized reward models a

Zero and Few Shot Load Forecasting with Large Language Models

ApplicationsDGX agent

arXiv:2411.11350v2 Announce Type: replace Abstract: Deep learning models have shown strong performance in load forecasting, but they generally require large amounts of data for model training before b

8 Jun 2026

A Conformation-Centric Generative Foundation Model for Linear Polymer Modeling and Design

ResearchDGX agent

arXiv:2510.16023v2 Announce Type: replace Abstract: Linear polymers, macromolecules formed from monomers covalently bonded into continuous chains, underpin countless technologies and are indispensable

Local models to turn images into 3d models?

Local AiDGX agent

AI 2D to 3D image converter tools use artificial intelligence to transform two-dimensional images into three-dimensional models by analyzing depth, textures, and shapes, simplifying the traditionally

Modeling semantic association in self-paced reading with language model embeddings

ResearchDGX agent

arXiv:2606.07066v1 Announce Type: new Abstract: Semantic association between a word and its context has been identified as an important component of reading comprehension, even when word predictabilit

6 Jun 2026

Integrating Mechanistic and Data-Driven Models for Neurological Disorders through Differentiable Programming

ResearchDGX agent

arXiv:2606.06094v1 Announce Type: new Abstract: Advances in computational modeling, neuroimaging, and artificial intelligence are revolutionizing the modeling of neurological disorders for improved di

Your margin is my opportunity: AI version… The biggest surprise of 2026 is that the capability gap between the best open-weight/source model…

Model ReleasesDGX agent

Your margin is my opportunity: AI version… The biggest surprise of 2026 is that the capability gap between the best open-weight/source models and the best closed models has narrowed much faster than t

4 Jun 2026

An Empirical Study of Data Scale, Model Complexity, and Input Modalities in Visual Generalization

Model ReleasesDGX agent

arXiv:2606.04409v1 Announce Type: cross Abstract: Modern deep neural networks usually have large parameter scales and nonlinear hierarchical structures, and they have achieved strong performance in co

Geometry-Preserving Unsupervised Alignment for Heterogeneous Foundation Models

Model ReleasesDGX agent

arXiv:2606.04385v1 Announce Type: new Abstract: Foundation models have driven rapid progress in computer vision, yet the two dominant paradigms, vision-language foundation models (VLMs) and vision-onl

NVIDIA’s Nemotron 3 Ultra is available on Ollama’s cloud! Try it 👇 Claude Code: ollama launch claude --model nemotron-3-ultra:cloud Hermes …

Model ReleasesDGX agent

NVIDIA’s Nemotron 3 Ultra is available on Ollama’s cloud! Try it 👇 Claude Code: ollama launch claude --model nemotron-3-ultra:cloud Hermes Agent: ollama launch hermes --model nemotron-3-ultra:cloud Op

3 Jun 2026

Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path Forward

Model ReleasesDGX agent

arXiv:2602.12430v4 Announce Type: replace-cross Abstract: The transition from monolithic language models to modular, skill-equipped agents marks a defining shift in how large language models (LLMs) ar

Cosmos 3: Omnimodal World Models for Physical AI

Model ReleasesDGX agent

arXiv:2606.02800v1 Announce Type: cross Abstract: We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences

How Quantization Changes Interpretable Features: A Sparse Autoencoder Analysis of Language Models

Model ReleasesDGX agent

arXiv:2606.03002v1 Announce Type: cross Abstract: Quantization is a standard path to deploying large language models, and a quantized model is typically judged acceptable when its perplexity or downst

2 Jun 2026

Balancing Accuracy and Efficiency: Adaptive Dynamics Orchestration for Model Predictive Control

SafetyDGX agent

arXiv:2606.00085v1 Announce Type: new Abstract: Model Predictive Control (MPC) for autonomous navigation faces a fundamental trade-off between model accuracy and real-time efficiency. High-fidelity dy

Expected Value Alignment for Generative Reward Modeling in Formal Mathematics Verification

SafetyDGX agent

arXiv:2606.01160v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used with formal interactive theorem provers such as Lean 4. Scaling these systems with reinforcement lear

Finding the Minimal Parameter Budget for Implicit Reasoning: A Data Complexity Driven Scaling Law for Language Models

Model ReleasesDGX agent

arXiv:2504.03635v4 Announce Type: replace Abstract: Reasoning is a core capability of language models (LMs), yet it remains unclear how much model capacity is necessary to support reasoning during pre

Interpretable Modeling of Driver Attention Shifts with a Vision--Language Model

SafetyDGX agent

arXiv:2508.05852v2 Announce Type: replace Abstract: Driver gaze is commonly modeled as a spatial heatmap, but heatmaps alone are difficult for humans to interpret because they do not explain which roa

RynnVLA-002: A Unified Vision-Language-Action and World Model

Model ReleasesDGX agent

arXiv:2511.17502v3 Announce Type: replace Abstract: We introduce RynnVLA-002, a unified Vision-Language-Action (VLA) and world model. The world model leverages action and visual inputs to predict futu

Uncovering Competency Gaps in Large Language Models and Their Benchmarks

Model ReleasesDGX agent

arXiv:2512.20638v2 Announce Type: replace-cross Abstract: The evaluation of large language models relies heavily on standardized benchmarks. These benchmarks provide useful aggregated metrics, but can

1 Jun 2026

Dreaming Of Others: Latent Teammate Modeling In World Models For Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2605.31361v1 Announce Type: cross Abstract: In cooperative multi-agent reinforcement learning (MARL), agents must coordinate with partners whose internal policies and intentions are not directly

EvoDefense: Co-Evolving Black-Box Defense with Large Language Models

Model ReleasesDGX agent

arXiv:2605.31140v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain highly vulnerable to diverse attacks, particularly in black-box settings where the internals of target models are

How Trustpilot built a real-time architecture for data enrichment using Gemma

Model ReleasesDGX agent

Processing millions of user reviews in real-time, under strict latency and cost constraints, is no easy task. Trustpilot has been doing exactly that with custom machine learning since long before larg

29 May 2026

Chess-World-Model: A 10M-Game Benchmark for Exact State Tracking from Chess Move Sequences

Model ReleasesDGX agent

arXiv:2605.30100v1 Announce Type: new Abstract: World models require state tracking, which is the ability to maintain a correct latent state across action sequences. Existing benchmarks are often synt

Do Physics Foundation Models Learn Generalizable Physics? A Bias-Aware Benchmark Across Physical Regimes and Distribution Shifts

Model ReleasesDGX agent

arXiv:2605.29283v1 Announce Type: cross Abstract: Recent physics foundation models claim general spatiotemporal forecasting ability, yet their evaluations often collapse performance into a single aver

Draft-OPD: On-Policy Distillation for Speculative Draft Models

SafetyDGX agent

arXiv:2605.29343v1 Announce Type: new Abstract: Speculative decoding accelerates large language model inference by pairing a target model with a lightweight draft model whose proposed tokens are verif

RightNow-Arabic-0.5B-Turbo: An Open Sub-1B Arabic Language Model via Vocabulary Injection and Edge-First Deployment

Model ReleasesDGX agent

arXiv:2605.28827v1 Announce Type: new Abstract: Open Arabic large language models split into two classes: sub-1B multilingual models that treat Arabic as an afterthought (Qwen2.5-0.5B, Falcon-H1-0.5B)

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

Model ReleasesDGX agent

arXiv:2605.29744v1 Announce Type: new Abstract: The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-speci

28 May 2026

Code as a Weapon: A Consensus-Labeled Prompt Bank for Measuring Coding-Model Compliance with Malicious-Code Requests

Model ReleasesDGX agent

arXiv:2605.28734v1 Announce Type: cross Abstract: A general-purpose language model that answers a harmful question returns text; a coding model that complies with a malicious request can return a work

Continual Learning in Modern Hopfield Networks with an Application to Diffusion Models

ResearchDGX agent

arXiv:2605.27975v1 Announce Type: new Abstract: Generative models, including diffusion models, are increasingly used as foundation models and adapted through sequential fine-tuning, making continual l

Do LLMs Build World Models From Text? A Multilingual Diagnostic of Spatial Reasoning

Model ReleasesDGX agent

arXiv:2605.28277v1 Announce Type: new Abstract: Whether large language models (LLMs) construct internal spatial world models from pure-text descriptions remains contested, and whether such capabilitie

Towards Faithful Agentic XAI: A Verification Method and an Open-World Benchmark for Better Model Faithfulness

Model ReleasesDGX agent

arXiv:2605.27879v1 Announce Type: new Abstract: Explainable AI (XAI) helps users interpret model behavior and identify potential faults. Agentic XAI systems use Large Language Models (LLMs) to make ex

27 May 2026

Benchmarking Convolutional, Transformer, Hybrid, and Vision Language Models for Multi Disease Retinal Screening

Model ReleasesDGX agent

arXiv:2605.26283v1 Announce Type: new Abstract: Modern deep learning offers powerful tools for automated retinal screening, but it remains unclear how different visual model families compare in realis

Black-box Membership Inference Attacks on the Pre-training Data of Image-generation Models

Model ReleasesDGX agent

arXiv:2605.27020v1 Announce Type: cross Abstract: The rapid advancement of diffusion-based image generation models has raised serious concerns regarding potential copyright and privacy infringements i

Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization

SafetyDGX agent

arXiv:2605.26282v1 Announce Type: new Abstract: Model-based reinforcement learning (RL) can be effectively supported at scale through the use of world models. However, in practice, scaling such approa

26 May 2026

Benchmarking Patent Embeddings: A Multi-Task Evaluation of 22 Models Across Retrieval, Classification, and Clustering

Model ReleasesDGX agent

arXiv:2605.24297v1 Announce Type: cross Abstract: Which fine-tuning signals improve patent embedding models, and do gains transfer across patent landscapes? We benchmark 22 embedding models, from 22M-

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

Model ReleasesDGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

Mimir: Large-scale Multilingual Concept Modeling

ResearchDGX agent

arXiv:2605.25263v1 Announce Type: cross Abstract: Current language modeling approaches are built around tokens. Text corpora are split into tokens, and models are trained by performing computations on

SomaliBench Eval: Measuring English-to-Somali Refusal Gaps in Open-Weight Language Models

Model ReleasesDGX agent

arXiv:2605.25420v1 Announce Type: cross Abstract: Large language model safety evaluation remains heavily English-centered, leaving low-resource languages under-measured even when models are deployed g

Why We Need World Models for AGI: Where LLMs Fail and How World Models May Outperform

AgentsDGX agent

arXiv:2605.23972v1 Announce Type: new Abstract: Large language models achieve strong performance in language generation and knowledge-intensive tasks, yet remain limited in settings requiring causal r

25 May 2026

A Comparative Evaluation of Structural Topic Models and BERTopic for Short, Open-Ended Survey Responses

Model ReleasesDGX agent

arXiv:2605.23093v1 Announce Type: new Abstract: Topic modeling in applied psychology increasingly spans two methodological traditions: probabilistic bag-of-words models and newer embedding-based appro

GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2605.23238v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as economic agents in marketplaces, auctions, and bidding settings. Anticipating their behavior i

Latent Cache Flow: Model-to-Model Communication Without Text

AgentsDGX agent

arXiv:2605.22863v1 Announce Type: new Abstract: LLM agents today communicate via text, which incurs considerable latency and information loss due to the need to autoregressively decode the sharer mode

The Attribution Contract: Feature Attribution for Generative Language Models

Local AiDGX agent

arXiv:2605.23080v1 Announce Type: new Abstract: Feature attribution methods promise to identify which input features matter for a model output. In generative language models, however, it is often uncl

23 May 2026

Provable Joint Decontamination for Benchmarking Multiple Large Language Models

Model ReleasesDGX agent

arXiv:2605.21543v1 Announce Type: new Abstract: Benchmark data contamination has become a central challenge in LLM evaluation: when evaluation examples appear in the training data of one or more audit

22 May 2026

Hy-MT2: A Family of Fast, Efficient and Powerful Multilingual Translation Models in the Wild

Model ReleasesDGX agent

arXiv:2605.22064v1 Announce Type: new Abstract: Hy-MT2 is a family of fast-thinking multilingual translation models designed for complex real-world scenarios. It includes three model sizes: 1.8B, 7B,

21 May 2026

Chronicle: A Multimodal Foundation Model for Joint Language and Time Series Understanding

Model ReleasesDGX agent

arXiv:2605.20268v1 Announce Type: cross Abstract: Real-world time series come with text: metadata, descriptions, news, reports. Yet time series foundation models process numerical sequences in isolati

Memory Grafting: Scaling Language Model Pre-training via Offline Conditional Memory

Model ReleasesDGX agent

arXiv:2605.20948v1 Announce Type: new Abstract: Scaling conditional memory offers a promising way to increase language-model capacity, but existing methods such as Engram learn large memory tables fro

20 May 2026

Towards Multi-Model LLM Schedulers: Empirical Insights into Offloading and Preemption

HardwareDGX agent

arXiv:2605.19593v1 Announce Type: new Abstract: Modern deployments of Large Language Models (LLMs) increasingly require serving multiple models with diverse architectures, sizes, and specialization on

Unlocking the Potential of Continual Model Merging: An ODE Perspective

Model ReleasesDGX agent

arXiv:2605.19409v1 Announce Type: cross Abstract: Continual Model Merging (CMM) enables rapid customization of foundation models across sequentially arriving tasks, offering a scalable alternative to

19 May 2026

Extending Pretrained 10-Second ECG Foundation Models to Longer Horizons

Model ReleasesDGX agent

arXiv:2605.16975v1 Announce Type: cross Abstract: Electrocardiogram (ECG) foundation models pretrained on typical diagnostic 10-second ECG segments, have demonstrated strong transferability across a r

Foundation Models for Credit Risk Prediction: A Game Changer?

ResearchDGX agent

arXiv:2605.18147v1 Announce Type: new Abstract: Predictive models play a pivotal role in credit risk management, guiding critical decisions through accurate estimation of default probabilities and los

GenTS: A Comprehensive Benchmark Library for Generative Time Series Models

Model ReleasesDGX agent

arXiv:2605.17804v1 Announce Type: new Abstract: Generative models have demonstrated remarkable potential in time series analysis tasks, like synthesis, forecasting, imputation, etc. However, offering

Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency

Model ReleasesDGX agent

arXiv:2605.18732v1 Announce Type: cross Abstract: While scaling laws govern aggregate large language model performance, no scaling law has linked factual recall to both model size and training-data co

Seems a lot of autoregressive models will be converted to diffusion models

IndustryDGX agent

Emad Mostaque, CEO of Stability AI, suggests that many autoregressive models may transition to or be replaced by diffusion-based approaches. This reflects speculation or prediction about a potential s

15 May 2026

Beyond Mode-Seeking RL: Trajectory-Balance Post-Training for Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.13935v1 Announce Type: cross Abstract: Diffusion language models are a promising alternative to autoregressive models, yet post-training methods for them largely adapt reward-maximizing obj

Hidden State Poisoning Attacks against Mamba-based Language Models

Model ReleasesDGX agent

arXiv:2601.01972v4 Announce Type: cross Abstract: State space models (SSMs) like Mamba offer efficient alternatives to Transformer-based language models, with linear time complexity. Yet, their advers

← Previous
1…1718192021…990
Next →