AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,553 results
Model Releases

Mistral debuts OCR 4, a model featuring structured document extraction with bounding boxes, block classification, and inline confidence scores, in 170 languages (Mistral AI Blog)

DGX agent

Mistral AI Blog: Mistral debuts OCR 4, a model featuring structured document extraction with bounding boxes, block classification, and inline confidence scores, in 170 languages — Today, we're releasi

model-releasestechmeme
23 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

NeuroShield: A Device-Agnostic Foundation Model for EEG Authentication

DGX agent

arXiv:2606.20673v1 Announce Type: cross Abstract: A central challenge in EEG authentication is that models are typically tied to the acquisition settings in which they are trained. In particular, vari

researcharxiv-cs-cv
23 Jun 2026
Model Releases

OGD4All: A Framework for Accessible Interaction with Geospatial Open Government Data Based on Large Language Models

DGX agent

arXiv:2602.00012v3 Announce Type: replace Abstract: We present OGD4All, a transparent, auditable, and reproducible framework based on Large Language Models (LLMs) to enhance citizens' interaction with

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Prompting Diffusion Models for Zero-Shot Instance Segmentation

DGX agent

arXiv:2606.22660v1 Announce Type: new Abstract: Several disruptive research directions have recently emerged in computer vision, including foundation models achieving previously unseen zero-shot perfo

researcharxiv-cs-cv
23 Jun 2026
Research

Quantum Convolutional Neural Networks for Groundwater Heat Plume Prediction: A Surrogate Modeling Approach

DGX agent

arXiv:2606.23411v1 Announce Type: cross Abstract: Quantum machine learning methods are increasingly explored for modeling complex environmental systems, including groundwater heat plume dynamics. In t

researcharxiv-cs-lg
23 Jun 2026
Model Releases

Reinforcement learning to improve large language model-based automated code compliance systems

DGX agent

arXiv:2606.22402v1 Announce Type: cross Abstract: Large language model (LLM)-based approaches for automated code compliance (ACC) of building regulations are prone to generating incorrect and hallucin

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Test-Time Alignment of Text-to-Image Diffusion Models via Null-Text Embedding Optimisation

DGX agent

arXiv:2511.20889v2 Announce Type: replace Abstract: Test-time alignment (TTA) aims to adapt models to specific rewards during inference. However, existing methods tend to either under-optimise or over

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

Zero-Shot Vision-Language Models for Classroom Engagement Recognition: A Benchmark Study of Prompt Sensitivity and Cross-Dataset Generalization

DGX agent

arXiv:2606.21861v1 Announce Type: new Abstract: Automated classroom engagement recognition holds substantial promise for scalable learning analytics, yet the suitability of modern Vision-Language Mode

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

My parallel agent side-project today was having Claude Code port the new Moebius image pinpointing model to ONNX in order to run it entirely…

DGX agent

My parallel agent side-project today was having Claude Code port the new Moebius image pinpointing model to ONNX in order to run it entirely in the browser https://simonwillison.net/2026/Jun/22/portin

model-releasessimon-willison--x
22 Jun 2026
Model Releases

Sakana AI launches Fugu, a multi-agent orchestration system accessible through a single model API, claiming Fugu Ultra matches Fable and Mythos on benchmarks (Carl Franzen/VentureBeat)

DGX agent

Carl Franzen / VentureBeat: Sakana AI launches Fugu, a multi-agent orchestration system accessible through a single model API, claiming Fugu Ultra matches Fable and Mythos on benchmarks — Last night,

model-releasestechmeme
22 Jun 2026
Local Ai

Let’s go open models! ❤️

DGX agent

Ollama, an open-source platform for running large language models locally, announced support or enthusiasm for open models on X (formerly Twitter). The post likely promotes the benefits of open-source

local-aiollama--x
21 Jun 2026
Model Releases

Adaptive Multi-Resolution Procedural Knowledge Compression for Large Language Models

DGX agent

arXiv:2606.12203v1 Announce Type: new Abstract: Large language models (LLMs) are widely used to tackle complex tasks with autonomous workflows. Recently, reusable natural language skills have emerged

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Bridging the Morphology Gap: Adapting VLA Models to Dexterous Manipulation via Intent-Conditioned Fine-Tuning

DGX agent

arXiv:2606.12109v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable zero-shot generalization in robotic manipulation, yet the vast majority of pre-traine

model-releasesarxiv-cs-ai
11 Jun 2026
Research

EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards

DGX agent

arXiv:2511.16672v4 Announce Type: replace Abstract: Recent advances in large multimodal models (LMMs) have enabled impressive reasoning and perception abilities, yet most existing training pipelines s

researcharxiv-cs-cv
11 Jun 2026
Model Releases

PermDoRA -- Understanding Adapter Interference in Language Models: Limits of Parameter-Space Geometry

DGX agent

arXiv:2606.11262v1 Announce Type: cross Abstract: Access control in large language models (LLMs) requires modular mechanisms to enable domain-specific behavior without retraining or cross-domain inter

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models

DGX agent

arXiv:2606.12412v1 Announce Type: cross Abstract: Vision-language models (VLMs) project images into hundreds to thousands of visual tokens, making decoder inference expensive in both attention computa

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Sparse probes and murky physics: a case study of interpretability challenges in a foundation model for continuum dynamics

DGX agent

arXiv:2606.11657v1 Announce Type: cross Abstract: Generative AI emulators are increasingly used in scientific domains where we already have strong theory, benchmarks, and physical intuition. This rais

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Task-Aligned Stability Analysis of Vision-Language Models for Autonomous Driving Hazard Detection

DGX agent

arXiv:2606.11889v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used for scene understanding in autonomous driving, but robustness analysis often relies on task-agnost

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of…

DGX agent

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of code that’s not well understood that you will have to pay l

model-releasesjerry-liu--x
11 Jun 2026
Model Releases

Towards Data-free and Training-free Compression for Speech Foundation Models Using Parameter Clustering

DGX agent

arXiv:2606.11836v1 Announce Type: cross Abstract: This paper presents a novel data-free and training-free compression approach for speech foundation models using channelwise clustering via k-means. Mo

model-releasesarxiv-cs-ai
11 Jun 2026
Research

A Continuous-Time Markov Chain Framework for Insertion Language Models

DGX agent

arXiv:2606.10199v1 Announce Type: cross Abstract: Insertion Language Models (ILMs) offer several advantages over left-to-right generation and mask-based generation. However, existing formulations of i

researcharxiv-cs-cl
10 Jun 2026
Model Releases

Benchmarking stereo reconstruction for 3D printable Martian terrain models

DGX agent

arXiv:2606.10364v1 Announce Type: new Abstract: Reconstructing printable 3D models from Mars rover imagery is challenging because Martian terrain is low-texture, irregular, and partially observed. We

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

CITRAS-FM: Tiny Time Series Foundation Model for Covariate-Informed Zero-Shot Forecasting

DGX agent

arXiv:2606.10798v1 Announce Type: new Abstract: Pretrained time series foundation models (TSFMs) have enabled zero-shot forecasting on unseen target series. However, existing TSFMs often incur high co

model-releasesarxiv-cs-lg
10 Jun 2026
Safety

Conditional Vendi Score: Prompt-Aware Diversity Evaluation for Generative AI Models and LLMs

DGX agent

arXiv:2411.02817v2 Announce Type: replace-cross Abstract: Generative models guided by text prompts are widely evaluated for fidelity and prompt alignment, yet their ability to produce outputs remains

safetyarxiv-cs-ai
10 Jun 2026
Local Ai

Density Field State Space Models: 1-Bit Distillation, Efficient Inference, and Knowledge Organization in Mamba-2

DGX agent

arXiv:2606.10932v1 Announce Type: new Abstract: We present Density Field State Space Models (DF-SSM), a framework for compressing SSMs to a 1-bit scaffold with int8 low-rank correction. Applied to Mam

local-aiarxiv-cs-cl
10 Jun 2026
Safety

Does Reasoning Preserve Alignment? On the Trustworthiness of Large Reasoning Models

DGX agent

arXiv:2606.11046v1 Announce Type: new Abstract: Instruction-tuned LLMs are increasingly converted into reasoning models through post-training to improve multi-step task performance. This conversion is

safetyarxiv-cs-cl
10 Jun 2026
Research

Few-step Generative Models as Lossy Compression

DGX agent

arXiv:2606.10450v1 Announce Type: new Abstract: DiffC provides a principled way to reuse pre-trained diffusion models for lossy compression, but its encoding and decoding procedures remain slow becaus

researcharxiv-cs-cv
10 Jun 2026
Safety

Going with the Flow: Koopman Behavioral Models as Pseudo Planners for Visuo-Motor Dexterity

DGX agent

arXiv:2602.07413v3 Announce Type: replace Abstract: Contemporary visuo-motor dexterity models often rely on expressive policy classes with diffusion and transformer backbones to achieve strong perform

safetyarxiv-cs-ro
10 Jun 2026
Local Ai

I took Andrej Karpathy's LLM Council concept to the next level (Docker, MCP, Skill, Search, local (Ollama)/cloud model support and much more)

DGX agent

A developer expanded on Andrej Karpathy's LLM Council concept by implementing an enhanced system with Docker containerization, Model Context Protocol (MCP) integration, skill modules, web search capab

local-air-ollama
10 Jun 2026
Safety

MedFeat: Model-Aware and Explainability-Driven Feature Engineering with LLMs for Clinical Tabular Prediction

DGX agent

arXiv:2603.02221v2 Announce Type: replace-cross Abstract: In clinical tabular prediction, classical machine learning models with feature engineering often outperform neural methods. LLMs are increasin

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

Meet DiffusionGemma! An experimental open model that explores a fast approach to text generation, released under an Apache 2.0 license. Movi…

DGX agent

Meet DiffusionGemma! An experimental open model that explores a fast approach to text generation, released under an Apache 2.0 license. Moving beyond sequential, token-by-token processes to generate e

model-releasesclem-delangue--x
10 Jun 2026
Safety

Multi-Faceted Interactivity Alignment in Full-Duplex Speech Models

DGX agent

arXiv:2606.11167v1 Announce Type: new Abstract: Full-duplex spoken dialogue models can listen and speak simultaneously, making them a promising architecture for natural conversation. However, current

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

OpenRTLSet: A Fully Open-Source Dataset for Large Language Model-based Verilog Module Design

DGX agent

arXiv:2606.10285v1 Announce Type: new Abstract: OpenRTLSet introduces the largest fully open-source dataset for hardware design, offering over 131,000 diverse Verilog code samples to the research comm

model-releasesarxiv-cs-cl
10 Jun 2026
Research

PRISM: Parallel Residual Iterative Sequence Model

DGX agent

arXiv:2602.10796v3 Announce Type: replace Abstract: Generative sequence modeling faces a fundamental tension between the expressivity of Transformers and the efficiency of linear sequence models. Exis

researcharxiv-cs-lg
10 Jun 2026
Model Releases

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models

DGX agent

arXiv:2606.11164v1 Announce Type: new Abstract: Long chain-of-thought (CoT) trajectories in large language model (LLM) reasoning cause severe inference bottlenecks due to rapid key-value (KV) cache gr

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sample Where You Struggle: Sharpening Base Model Reasoning via Entropy-Guided Power Sampling

DGX agent

arXiv:2606.09926v1 Announce Type: cross Abstract: Sampling from the sequence-level power distribution p^alpha elicits RL-level reasoning from base language models without any parameter updates, but th

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SSR-Merge: Subspace Signal Routing for Training-Free LoRA Merging in Diffusion Models

DGX agent

arXiv:2606.10617v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) merging can efficiently combine diverse generative capabilities from multiple trained LoRAs for a diffusion model. However, e

model-releasesarxiv-cs-cv
10 Jun 2026
Agents

the cost of fable is going to make smart model routing impossible to ignore

DGX agent

This post discusses how the pricing model of Fable (likely an AI/LLM service) creates economic incentives that make intelligent routing between different AI models a necessary optimization strategy ra

agentsjerry-liu--x
10 Jun 2026
Agents

The Model Lab vs Agent Lab distinction is one of the clearest frameworks I've seen for understanding where AI value actually lives right now…

DGX agent

The Model Lab vs Agent Lab distinction is one of the clearest frameworks I've seen for understanding where AI value actually lives right now. TL;DR from @latentspacepod: • Model Labs compete on capabi

agentsswyx--x
10 Jun 2026
Safety

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models

DGX agent

arXiv:2606.10740v1 Announce Type: new Abstract: Failures in multi-turn reasoning models are largely invisible to terminal-score evaluation. A model can lock onto an unsafe stance early in a long dialo

safetyarxiv-cs-ai
10 Jun 2026
Research

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling

DGX agent

arXiv:2512.14614v2 Announce Type: replace Abstract: This paper presents WorldPlay, a streaming video diffusion model that enables real-time, interactive world modeling with long-term geometric consist

researcharxiv-cs-cv
10 Jun 2026
Model Releases

Wrote up my initial impressions of Claude Fable 5 - it has a big model smell: slow, expensive and capable of crunching through pretty much e…

DGX agent

Wrote up my initial impressions of Claude Fable 5 - it has a big model smell: slow, expensive and capable of crunching through pretty much everything I threw at it https://simonwillison.net/2026/Jun/9

model-releasessimon-willison--x
10 Jun 2026
Research

3D Oral Modelling with Improved Vertex Distribution Using Matching-Based Learning

DGX agent

arXiv:2606.07907v1 Announce Type: cross Abstract: In our previous work, a deep learning-based framework for 3D intraoral reconstruction was proposed. The model directly predicts explicit 3D point clou

researcharxiv-cs-ai
9 Jun 2026
Local Ai

A systematic investigation of molecular encoding methods for drug property predictions across neural network and Transformer encoder-based model

DGX agent

arXiv:2606.08973v1 Announce Type: cross Abstract: Fundamental investigations into how different molecular encoding methods affect molecular property prediction remain relatively limited. In this study

local-aiarxiv-cs-lg
9 Jun 2026
Model Releases

A TIL on using http://agentsview.io to calculate token spending with Claude Fable 5 despite that model not yet being included in the AgentsV…

DGX agent

A TIL on using http://agentsview.io to calculate token spending with Claude Fable 5 despite that model not yet being included in the AgentsView pricing database https://til.simonwillison.net/llms/agen

model-releasessimon-willison--x
9 Jun 2026
Applications

Adversarial Robustness of Activation Steering in Large Language Models

DGX agent

arXiv:2606.07696v1 Announce Type: cross Abstract: Activation steering has become a popular training-free method to control LLM behavior by injecting precomputed direction vectors into the model's resi

applicationsarxiv-cs-ai
9 Jun 2026
Agents

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models

DGX agent

arXiv:2606.08952v1 Announce Type: new Abstract: Multimodal Foundation Models (MFMs) have made substantial progress, yet remain fragile in spatial reasoning over the physical world. A key bottleneck li

agentsarxiv-cs-ai
9 Jun 2026
Research

An Effective Router for Vision-Language Model Selection

DGX agent

arXiv:2606.08970v1 Announce Type: new Abstract: Vision-language models (VLMs) with varying performance and resource requirements are widely deployed, making it difficult for users to select the most a

researcharxiv-cs-ai
9 Jun 2026
← Previous
1…132133134135136…1262
Next →