AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,588 results
26 May 2026

The Impact of Large Language Models on Open-source Innovation: Evidence from GitHub Copilot

ResearchDGX agent

arXiv:2409.08379v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are reshaping knowledge work, yet their impact on voluntary, self-guided open innovation forums (contributors cho

Understanding and Mitigating Premature Confidence for Better LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.24396v1 Announce Type: new Abstract: Long chains of thought (CoT) from current language models frequently contain logical gaps and unjustified leaps, limiting the gains from additional test

Universal Activation Verbalizer: A Unified Framework for Cross-Model Activation Explanation

ResearchDGX agent

arXiv:2605.25903v1 Announce Type: new Abstract: Activation verbalization explains hidden representations in natural language, but existing methods are mostly limited to self-explanation, where each mo

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

When In-Distribution Gains Fail: Evaluating Weak-to-Strong Reward Models under Preference Shift

ResearchDGX agent

arXiv:2605.25629v1 Announce Type: new Abstract: Weak-to-strong (W2S) generalization is a promising framework for scalable oversight, yet existing evaluations often test students under matched train--t

Why Agentic Theorem Prover Works: A Statistical Provability Theory of Mathematical Reasoning Models

AgentsDGX agent

arXiv:2602.10538v3 Announce Type: replace-cross Abstract: Agentic theorem provers combine a reasoning model, retrieval, search, and a proof assistant verifier, yet it remains unclear which components

World-VLA-Loop: Closed-Loop Learning of Video World Model and VLA Policy

SafetyDGX agent

arXiv:2602.06508v2 Announce Type: replace Abstract: Reinforcement learning (RL) can refine Vision-Language-Action (VLA) policies beyond behavior cloning, but real-world RL remains expensive due to ext

Zero-Shot Parkinson's Disease Detection from Speech: Comparing Large Audio and Language Models

ResearchDGX agent

arXiv:2605.24806v1 Announce Type: cross Abstract: Large audio and language models have recently demonstrated zero-shot reasoning capabilities across various domains. However, it remains unclear how th

25 May 2026

A look at the UK's AI Safety Institute, whose researchers probe AI models for safety gaps, as its work becomes a blueprint for other governments' AI policies (New York Times)

SafetyDGX agent

New York Times: A look at the UK's AI Safety Institute, whose researchers probe AI models for safety gaps, as its work becomes a blueprint for other governments' AI policies — The government's A.I. Se

ChainFlow-VLA: Causal Flow Planning with Vision-Language Models

SafetyDGX agent

arXiv:2605.23270v1 Announce Type: cross Abstract: Current end-to-end autonomous driving systems are fundamentally limited by a mismatch between temporal causal reasoning and global trajectory consiste

Convex Compositional Reasoning Models

ResearchDGX agent

arXiv:2605.23395v1 Announce Type: new Abstract: Compositional energy-based models can generalize to larger combinatorial reasoning problems by reusing a learned factor energy across many local constra

Decomposing and Measuring Evaluation Awareness

Model ReleasesDGX agent

arXiv:2605.23055v1 Announce Type: cross Abstract: Frontier language models sometimes recognize that they are being evaluated and adjust their behavior, undermining validity of benchmark results. Yet t

From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills

AgentsDGX agent

arXiv:2605.23899v1 Announce Type: new Abstract: Language agents increasingly improve by reusing skills -- structured procedural artifacts distilled from past experience. In particular, domain-level an

Hierarchical Concept Geometry in Language Models Emerges from Word Co-occurrence

Model ReleasesDGX agent

arXiv:2605.23821v1 Announce Type: new Abstract: We propose a distributional theory of how hypernymy -- the ``is-a'' relation between general and specific concepts -- is encoded geometrically in langua

Is Dimensionality a Barrier for Retrieval Models?

ResearchDGX agent

arXiv:2605.23556v1 Announce Type: new Abstract: Why does the low dimensionality of representations, typically dapprox 1000, not prevent modern embedding-based retrieval models from scaling to billions

It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amplified by the language of the prompt

Model ReleasesDGX agent

arXiv:2605.23825v1 Announce Type: cross Abstract: It has generally been assumed that geopolitical bias in language models originates from the training data used during the pre-training phase. We teste

MadEvolve: Evolutionary Optimization of Trading Systems with Large Language Models

Model ReleasesDGX agent

arXiv:2605.23007v1 Announce Type: cross Abstract: We explore the application of LLM-driven algorithm optimization to several common tasks in quantitative finance. MadEvolve, a general-purpose algorith

Next-Latent Prediction Transformers Learn Compact World Models

SafetyDGX agent

arXiv:2511.05963v2 Announce Type: replace Abstract: Transformers replace recurrence with a memory that grows with sequence length and self-attention that enables ad-hoc lookups over past tokens. Conse

Reading Calibrated Uncertainty from Language Model Trajectories

ResearchDGX agent

arXiv:2605.22864v1 Announce Type: new Abstract: The maximum softmax probability (MSP) represents a default approach when evaluating uncertainty quantification for language model generation with struct

WMAttack: Automated Attack Search for Adversarial Evaluation of World-Model Agents

ResearchDGX agent

arXiv:2605.23220v1 Announce Type: new Abstract: Despite the growing use of world models as decision-making agents, their adversarial robustness remains underexplored due to the lack of dedicated autom

World Machine: Towards Generative World Modeling for Time-Series

ResearchDGX agent

arXiv:2605.23025v1 Announce Type: new Abstract: World models represent a paradigm shift in generative AI, pursuing predictive understanding and controllable simulation of environments in a structured

24 May 2026

My daily average local model token burn is 17M They have become tremendously useful.

Local AiDGX agent

Clem Delangue reports consuming approximately 17 million tokens daily when running local language models, indicating substantial usage of on-device AI inference. He expresses satisfaction with the uti

Thinking Machines is impressive. In a couple hours I just fine tuned my own Qwen3.5-397B model this afternoon. Fast usable multimodal is als…

ResearchDGX agent

Thinking Machines is impressive. In a couple hours I just fine tuned my own Qwen3.5-397B model this afternoon. Fast usable multimodal is also going to enable very mind-blowing personal AI. People talk

23 May 2026

A 6-person team is building task-specific AI models that are 4-8x faster than anything from OpenAI or Anthropic. 500K downloads on HuggingFa…

IndustryDGX agent

A 6-person team is building task-specific AI models that are 4-8x faster than anything from OpenAI or Anthropic. 500K downloads on HuggingFace. No hype. Just better engineering winning on the merits.

[AINews] All Model Labs are now Agent Labs

AgentsDGX agent

Latent Space reports on a rebranding or reorganization where Model Labs have been renamed or converted into Agent Labs, reflecting a shift in focus toward AI agent development and capabilities. This c

Disentanglement Beyond Generative Models with Riemannian ICA

Local AiDGX agent

arXiv:2605.22531v1 Announce Type: new Abstract: There is a gap between the theoretical foundations of disentanglement and the practice of modern representation learning. Existing theoretical framework

EnCAgg: Enhanced Clustering Aggregation for Robust Federated Learning against Dynamic Model Poisoning

Local AiDGX agent

arXiv:2605.22506v1 Announce Type: cross Abstract: Federated learning faces increasing threats from model poisoning attacks, which harms its application to improve privacy. Existing defense methods typ

Finite-Particle Convergence Rates for Conservative and Non-Conservative Drifting Models

ResearchDGX agent

arXiv:2605.22795v1 Announce Type: cross Abstract: We propose and analyze a conservative drifting method for one-step generative modeling. The method replaces the original displacement-based drifting v

Generative Modeling by Value-Driven Transport

SafetyDGX agent

arXiv:2605.22507v1 Announce Type: new Abstract: We propose a new framework for generative modeling based on a discrete-time stochastic control formulation of measure transport. Adapting classic result

Google’s new anything-to-anything AI model is wild

Model ReleasesDGX agent

Last year I deepfaked my kid's stuffed animal to make it look like his plush deer was on vacation. It was an experiment to see if I could re-create the events depicted in a Gemini ad Google was runnin

it’s truly astonishing to see the entire field move in the neurosymbolic and world model directions I advocated in 2019 and 2020 and then si…

SafetyDGX agent

it’s truly astonishing to see the entire field move in the neurosymbolic and world model directions I advocated in 2019 and 2020 and then simultaneously see people claim (invariably without specific e

SepsisAI Orchestrator: A Containerized and Scalable Platform for Deploying AI Models and Real-Time Monitoring in Early Sepsis Detection

ResearchDGX agent

arXiv:2605.22331v1 Announce Type: new Abstract: Despite strong predictive results in the clinical machine learning literature, the translation of these models into bedside use remains limited by syste

22 May 2026

A Tutorial on Diffusion Theory: From Differential Equations to Diffusion Models

TutorialsDGX agent

arXiv:2605.22586v1 Announce Type: cross Abstract: This tutorial develops diffusion models from the viewpoint of differential equations. We begin with the conditional Gaussian forward process and show

An Open Multi-Center Whole-Body FDG PET/CT Foundation Model for Tumor Segmentation

ResearchDGX agent

arXiv:2605.21835v1 Announce Type: cross Abstract: The synergistic interpretation of anatomical information from computed tomography (CT) and metabolic information from positron emission tomography (PE

Cohere Command A+ is now available in Microsoft Foundry as a Managed Compute offer. Cohere’s latest open-source model is built for enterpris…

AgentsDGX agent

Cohere Command A+ is now available in Microsoft Foundry as a Managed Compute offer. Cohere’s latest open-source model is built for enterprise-grade agentic AI workloads, bringing together reasoning, m

Conceptualizing Embeddings: Sparse Disentanglement for Vision-Language Models

TutorialsDGX agent

arXiv:2605.22679v1 Announce Type: new Abstract: Vision-language models learn powerful multimodal embeddings, yet their internal semantics remain opaque. While sparse autoencoders (SAEs) can extract in

From Correlation to Cause: A Five-Stage Methodology for Feature Analysis in Transformer Language Models

ResearchDGX agent

arXiv:2605.22462v1 Announce Type: new Abstract: We propose a five-stage methodology for causal feature analysis in transformer language models (probe design, feature extraction, causal validation, rob

i need someone at @OpenAI and @AnthropicAI to teach the models that while prototyping, backwards compatibility is just a bad idea

TutorialsDGX agent

Jeremy Howard argues that AI model developers at OpenAI and Anthropic should prioritize breaking backwards compatibility during the prototyping phase rather than maintaining it, suggesting that backwa

In Silico Modeling of the RAMPHO Buffer: Dissociating Informational and Energetic Masking via Phonetic Entropy in Deep Neural Networks

ResearchDGX agent

arXiv:2605.22465v1 Announce Type: new Abstract: The fundamental challenge of listening in multi-talker environments is a cognitive bottleneck, defined by the Ease of Language Understanding (ELU) model

NEW paper worth reading. A full agentic workflow can be distilled into model weights and run at roughly 100x lower inference cost while pres…

AgentsDGX agent

NEW paper worth reading. A full agentic workflow can be distilled into model weights and run at roughly 100x lower inference cost while preserving near-frontier task quality. The workflow includes mul

21 May 2026

3D Reconstruction and Knowledge Distillation to Improve Multi-View Image Models to Explore Spike Volume Estimation in Wheat

SafetyDGX agent

arXiv:2605.20940v1 Announce Type: new Abstract: Accurate estimation of wheat spike volume is important for yield component analysis and stress resilience assessment, yet field-based measurement remain

Accelerating Video Inverse Problem Solvers with Autoregressive Diffusion Models

HardwareDGX agent

arXiv:2605.20624v1 Announce Type: new Abstract: Diffusion models provide powerful priors for zero-shot video inverse problems, but their real-time deployment is hindered by two inefficiencies: high in

AirfoilGen: A valid-by-construction and performance-aware latent diffusion model for airfoil generation

ResearchDGX agent

arXiv:2605.20303v1 Announce Type: new Abstract: Airfoil shape design is a fundamental task in aerospace engineering, with a direct impact on flight stability and fuel consumption. Deep learning has re

CoarseSoundNet: Building a reliable model for ecological soundscape analysis

ApplicationsDGX agent

arXiv:2605.21143v1 Announce Type: cross Abstract: A soundscape is composed of three types of sound: biophony (sounds made by animals), geophony (natural abiotic sounds) and anthropophony (sounds made

Efficient Learning of Deep State Space Models via Importance Smoothing

ResearchDGX agent

arXiv:2605.21108v1 Announce Type: new Abstract: Latent state space systems are ubiquitous in statistical modelling, arising naturally when a time series is observed through a noisy measurement functio

Focus-then-Context: Subject-Centric Progressive Visual Token Reduction for Vision-Language Models

ResearchDGX agent

arXiv:2605.20950v1 Announce Type: new Abstract: Vision-Language Models (VLMs) face a bottleneck of prohibitive computational costs arising from massive visual token sequences during inference. Existin

Generation of Heterogeneous PET Images from Uniform Organ Activity Maps Using a Pretrained Domain-Adapted Diffusion Model

ResearchDGX agent

arXiv:2605.20267v1 Announce Type: new Abstract: Synthetic PET images are valuable for quantitative imaging workflow development, scalable virtual imaging trials, and deep learning model training, but

How Well Do Vision-Language Models Understand Sequential Driving Scenes? A Sensitivity Study

AgentsDGX agent

arXiv:2604.06750v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) are increasingly proposed for autonomous driving tasks, yet their performance on sequential driving scenes remains poo

Inference Time Policy Optimization for Offline RL with Differentiable World Models

SafetyDGX agent

arXiv:2603.22430v2 Announce Type: replace Abstract: Offline Reinforcement Learning (RL) learns optimal policies from fixed datasets, training a policy once and deploying it at inference time without f

Large-Step Training Dynamics of a Two-Factor Linear Transformer Model

Model ReleasesDGX agent

arXiv:2605.21292v1 Announce Type: cross Abstract: Gradient-flow analyses show that simplified linear transformers can learn the in-context linear-regression algorithm, but they do not explain the fini

Linear-DPO: Linear Direct Preference Optimization for Diffusion and Flow-Matching Generative Models

SafetyDGX agent

arXiv:2605.21123v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is successful for alignment in LLMs but still faces challenges in text-to-image generation. Existing studies are co

Managed Deep Agents is now in Private Beta ICYMI: It’s managed, model-agnostic infra for deep agents you can deploy with a single line of co…

ApplicationsDGX agent

Managed Deep Agents is a new LangChain feature in private beta that provides managed infrastructure for deep agents, designed to be model-agnostic and deployable with minimal code. The service aims to

Memorisation, convergence and generalisation in generative models

TutorialsDGX agent

arXiv:2605.21402v1 Announce Type: cross Abstract: Generative neural networks learn how to produce highly realistic images from a large, but finite number of examples - or do they simply memorise their

Oracle Supervision Transfers for Hyperparameter Prediction in Model-Based Image Denoising

ResearchDGX agent

arXiv:2605.20479v1 Announce Type: new Abstract: Hyperparameter prediction is a critical practical bottleneck for model-based image denoisers, ranging from classical TV/TGV variational solvers to moder

Provably Learning Diffusion Models under the Manifold Hypothesis: Collapse and Refine

ResearchDGX agent

arXiv:2605.20235v1 Announce Type: new Abstract: Diffusion models generate high-dimensional data with remarkable quality, yet how their training efficiently learns the score function, bypassing the cur

The top announcements for startups from Google I/O ‘26

Model ReleasesDGX agent

Many of the world’s fastest-growing AI startups are choosing to build their future — and the world’s — on Google Cloud because of our complete and open AI stack. We embed AI into every layer of our ar

Uni-Edit: Intelligent Editing Is A General Task For Unified Model Tuning

ResearchDGX agent

arXiv:2605.21487v1 Announce Type: new Abstract: Currently, enhancing Unified Multimodal Models (UMMs) with image understanding, generation, and editing capabilities mainly relies on mixed multi-task t

20 May 2026

A Geometric Analysis of Small-sized Language Model Hallucinations

AgentsDGX agent

arXiv:2602.14778v3 Announce Type: replace-cross Abstract: Hallucinations -- plausible but factually incorrect responses -- pose a major challenge to the reliability of Large Language Models (LLMs), es

An Objective Performance Evaluation of the LSTM Networks in Time Series Classification

ResearchDGX agent

arXiv:2605.19311v1 Announce Type: new Abstract: The rapid adoption of deep learning has increasingly led to data-driven models replacing classical model-based algorithms, even in domains governed by w

Attention-Guided Reward for Reinforcement Learning-based Jailbreak against Large Reasoning Models

SafetyDGX agent

arXiv:2605.19485v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated remarkable capabilities in solving complex problems by generating structured, step-by-step reasoning con

Bayesian Latent Space Models for Graphs Are Misspecified: Toward Robust Inference via Generalized Posteriors

ApplicationsDGX agent

arXiv:2605.18927v1 Announce Type: cross Abstract: Bayesian latent space models offer a principled approach to network representation, but rely on correct specification of both geometry and link functi

← Previous
1…183184185186187…1010
Next →