AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,929 results
Model Releases

Introducing Muse Glimmer

DGX agent

Introducing Muse Glimmer Meta are back in the open weights game! Muse Glimmer is a brand new 30B model under a clean Apache 2.0 license (a step up from the janky Llama licenses of old). They claim to

model-releasessimon-willison
10 Aug 2026
Model Releases

Introducing Muse Glimmer: an open-weight model optimized for always-on local agent workflows

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

Hi r/LocalLLaMA 👋 Today we’re excited to release Muse Glimmer, a 30B open-weight model built specifically for local agent workflows. We’re releasing the weights to the community under a permissive Apa

model-releasesr-localllama
10 Aug 2026
Model Releases

Retrofitting Linear Attention into Diffusion Language Models

DGX agent

arXiv:2608.06628v1 Announce Type: new Abstract: Diffusion language models (dLLMs) offer a promising alternative to autoregressive models by accelerating inference through parallel decoding. Recent dLL

model-releasesarxiv-cs-lg
10 Aug 2026
Agents

SimWAM: A Simple World Action Model for End-to-End Autonomous Driving

DGX agent

arXiv:2608.07468v1 Announce Type: new Abstract: World-Action Models (WAMs) improve end-to-end autonomous driving by transferring video dynamics priors to action prediction, but existing methods requir

agentsarxiv-cs-cv
10 Aug 2026
Tutorials

TaskSense: Focusing on What Matters in World Models

DGX agent

arXiv:2608.06544v1 Announce Type: new Abstract: World models for visual control typically learn compact latent states by reconstructing observations, implicitly encouraging representations to preserve

tutorialsarxiv-cs-ai
10 Aug 2026
Model Releases

TEMPO: Semantic-Action Decoupled RL Post-Training for Vision-Language-Action Models

DGX agent

arXiv:2608.07314v1 Announce Type: cross Abstract: Vision-language-action (VLA) models are commonly adapted to downstream manipulation tasks via supervised fine-tuning (SFT) or online reinforcement lea

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

WorldMark: A Plug-and-Play World Knowledge Interface for Cross-Host Language Model Watermarking

DGX agent

arXiv:2608.06416v1 Announce Type: cross Abstract: Watermarking traces the provenance of text produced by large language models by embedding statistically detectable signals during decoding. Existing s

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Beyond Next-Token Prediction: A Performance Characterization of Diffusion versus Autoregressive Language Models

DGX agent

Large Language Models (LLMs) have achieved state-of-the-art performance on a broad range of Natural Language Processing (NLP) tasks, including document processing and code generation. Autoregressive L

researchapple-ml-research
7 Aug 2026
Model Releases

C^3PO: Evaluating Cross-Modal Composition and Counterfactual Performance in Omnimodal Models

DGX agent

arXiv:2608.05381v1 Announce Type: new Abstract: Current Multimodal Large Language Models (MLLMs) can process diverse sensory inputs, yet their reasoning remains heavily biased toward a dominant modali

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Deep Generalised Mixed Models: a Novel Neural Network Structure for Analysing Hierarchical Data

DGX agent

arXiv:2608.05930v1 Announce Type: cross Abstract: The experience sampling method (ESM) is a longitudinal research design where participants report their thoughts, emotional states and behaviours multi

model-releasesarxiv-cs-lg
7 Aug 2026
Safety

From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models

DGX agent

arXiv:2608.06020v1 Announce Type: new Abstract: Economic World Models (EWMs) are generative economic models that simulate how economies evolve from within by modeling heterogeneous agents, their belie

safetyarxiv-cs-ai
7 Aug 2026
Research

HERA: Historical Evidence Routing Adapter for Physical Prediction in Latent World Models

DGX agent

arXiv:2608.05523v1 Announce Type: new Abstract: Predictive video models have emerged as promising world models by learning latent visual dynamics from large-scale video. Yet these models remain challe

researcharxiv-cs-cv
7 Aug 2026
Model Releases

MameLoshnLM: Yiddish Language Model and Evaluation Benchmark

DGX agent

arXiv:2608.05850v1 Announce Type: cross Abstract: We present MameLoshnLM, the first open-source 8B-parameter language model built specifically for Yiddish. Despite Yiddish's rich textual tradition, it

model-releasesarxiv-cs-ai
7 Aug 2026
Research

Small Foundation Models of Human Cognition and Behaviour

DGX agent

arXiv:2608.05224v1 Announce Type: new Abstract: Large language models fine-tuned on human behavioural data have emerged as general-purpose cognitive proxies, but the scale this requires, and whether t

researcharxiv-cs-ai
7 Aug 2026
Model Releases

Adversarially Robust Abductive Fusion of Pre-trained Transformer-based Perception Models

DGX agent

arXiv:2608.04190v1 Announce Type: new Abstract: Deploying pre-trained perception models in novel environments degrades their accuracy under distributional shift, and assembling them alone does not rec

model-releasesarxiv-cs-ai
6 Aug 2026
Research

Emergence of Hierarchical Emotion Organization in Large Language Models

DGX agent

arXiv:2507.10599v3 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly power conversational agents, understanding how they model users' emotional states is critical for

researcharxiv-cs-ai
6 Aug 2026
Model Releases

Galaxy Phase-Space and Field-Level Cosmology: The Strength of Semi-Analytic Models

DGX agent

arXiv:2512.10222v2 Announce Type: replace-cross Abstract: Semi-analytic models are a widely used approach to simulate galaxy properties within a cosmological framework, relying on simplified yet physi

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

HelloWorld: Enabling Socially Interactive Characters in Video World Models

DGX agent

arXiv:2608.05070v1 Announce Type: new Abstract: Despite the remarkable recent progress of video world models, social interaction between users and the characters within these worlds remains unsupporte

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Radar4D-VLM: Proposal-Grounded Temporal 4D Radar Reasoning Across Frozen Language Models

DGX agent

arXiv:2608.04130v1 Announce Type: new Abstract: Vision-language models for autonomous driving primarily rely on cameras and LiDAR, leaving 4D radar largely unexplored as a standalone perceptual modali

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

SIGNPOST-Bench: Benchmarking Text-Vision Conflict Resolution in Multimodal Large Language Models

DGX agent

arXiv:2608.04244v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) make grounded predictions in real-world scenes by combining visual and textual cues, yet existing benchmarks

model-releasesarxiv-cs-cl
6 Aug 2026
Research

Simile Understanding in Text-to-Image Models: An Evaluation Framework

DGX agent

arXiv:2608.04750v1 Announce Type: cross Abstract: Similes provide a compact and expressive way to describe visual characteristics in text prompts. Recent text-to-image models (t2i models) can produce

researcharxiv-cs-cl
6 Aug 2026
Model Releases

Toward Federated Large Language Models in Medicine: A Parameter-Efficient Framework for Privacy-Preserving, Multi-Institutional Adaptation

DGX agent

arXiv:2601.22124v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly adapted for medical applications, but most are trained using data from a single institution because pr

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Confident but Unreliable: A Behavioral Safety Audit of Vision-Language Models on Brain MRI

DGX agent

arXiv:2608.02790v1 Announce Type: new Abstract: Vision-language models (VLMs), including medical specialists, are increasingly proposed for medical imaging, yet their stated confidence is rarely evalu

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

From Generator to Embedder: Harnessing Innate Abilities of Multimodal LLMs via Building Zero-Shot Discriminative Embedding Model

DGX agent

arXiv:2508.00955v3 Announce Type: replace-cross Abstract: Adapting generative Multimodal Large Language Models (MLLMs) into universal embedding models typically demands resource-intensive contrastive

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Learning a Vector-Symbolic Model for Socio-Cultural Tasks

DGX agent

arXiv:2608.02807v1 Announce Type: cross Abstract: How can we better represent the impact of sociocultural structures on decision making in computational cognitive models? Modeling this impact requires

researcharxiv-cs-ai
5 Aug 2026
Model Releases

TimeRLM: Recursive Language Models Enable Precise Anomaly Localization in Long-Context Time-Series

DGX agent

arXiv:2608.03391v1 Announce Type: new Abstract: Precise anomaly localization over long-context time series is a crucial task in monitoring applications across clinical care, industrial operations, fin

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

A General-Purpose VLM Can Teach an Astronomy Foundation Model to Better Recognize Galaxy Morphology

DGX agent

arXiv:2608.02300v1 Announce Type: new Abstract: Existing astronomy foundation models provide strong galaxy representations, but adapting them to new survey conditions and survey-specific morphology re

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

DE-NER : Zero-shot Named Entity Recognition via Dialogue Elicitation of Large Language Models

DGX agent

arXiv:2608.00538v1 Announce Type: new Abstract: Recent advancements of zero-shot Named Entity Recognition (NER) establish strong baselines by formulating sequence labeling into question answering wher

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Divergent large language model predictions from convergent representations in ambiguous word pairs

DGX agent

arXiv:2608.01816v1 Announce Type: new Abstract: In this work we investigate how decoder-only transformers resolve lexical ambiguity through layer-by-layer analysis of three models spanning three param

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Do Maps Still Matter for Machines: Revisiting the Role of Choropleth Maps in Foundation Model Spatial Understanding

DGX agent

arXiv:2607.17999v2 Announce Type: replace-cross Abstract: Spatial understanding is crucial for foundation models (FMs), and maps have long helped humans organize and reason about geographic informatio

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Exploring More to Solve More: Boosting Diversity in Text Diffusion Models via Entropy-Based Guidance

DGX agent

arXiv:2608.00024v1 Announce Type: new Abstract: Although diffusion models have revolutionized continuous domains like image synthesis through high quality generations and controllable guidance mechani

researcharxiv-cs-cl
4 Aug 2026
Model Releases

From field-scale to large-scale spectral libraries: Tabular foundation models in soil spectroscopy

DGX agent

arXiv:2608.00608v1 Announce Type: new Abstract: Visible and near-infrared (vis-NIR) and mid-infrared (MIR) spectroscopy enable rapid, cost-effective prediction of soil properties. Yet, translating hig

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Harnessing Adversarial Distillation to Customise Debiased, Disease-Specific Pathology Foundation Models for Breast Cancer

DGX agent

arXiv:2608.01356v1 Announce Type: new Abstract: Pathology foundation models (PFMs) provide strong tissue representations and have become central to digital pathology. However, deployment in disease-sp

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, Aug…

DGX agent

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, August 6th. Sign up below! Next session, August 6 we'll be co l

model-releasesfireworks-ai--x
4 Aug 2026
Model Releases

Models as Tools: An Agentic Coordination Framework for Unified Multimodal Visual Tracking

DGX agent

arXiv:2608.00847v1 Announce Type: new Abstract: Most current visual trackers adopt a matching-based architecture trained exclusively on tracking datasets, whose performance gains depend heavily on the

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Picking the right agent harness is now a crucial skill for any AI engineer. Imagine using the same model, same task, and same prompt. Now mo…

DGX agent

Picking the right agent harness is now a crucial skill for any AI engineer. Imagine using the same model, same task, and same prompt. Now move it between two agent harnesses and the cost per success c

model-releasesdair-ai--x
4 Aug 2026
Model Releases

Quality-Diversity Red-Teaming: Automated Generation of High-Quality and Diverse Attackers for Large Language Models

DGX agent

arXiv:2506.07121v2 Announce Type: replace Abstract: Ensuring the safety and robustness of large language models (LLMs) is a fundamental challenge and a critical prerequisite for the responsible deploy

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Qwen 3.8-Max, the latest model from @Alibaba_Qwen, is now available in Hermes Agent at 20% off. > hermes update

DGX agent

Qwen 3.8-Max, the latest model from @Alibaba_Qwen, is now available in Hermes Agent at 20% off. > hermes update 📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3

model-releasesnous-research--x
4 Aug 2026
Model Releases

SAFE-Merge: Data-Free Continual Model Merging with General Knowledge Preservation

DGX agent

arXiv:2608.01184v1 Announce Type: new Abstract: Data-free continual model merging must incorporate a stream of specialized models while retaining both pretrained general knowledge and previously acqui

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

Self-Correction Bench: Uncovering and Addressing the Self-Correction Blind Spot in Large Language Models

DGX agent

arXiv:2507.02778v3 Announce Type: replace Abstract: Although large language models (LLMs) have transformed AI, they still make errors and follow unproductive reasoning paths. Self-correction is vital

safetyarxiv-cs-cl
4 Aug 2026
Applications

TIDES: A Longitudinal Bilingual Dataset for Modeling Multi-Party Social Dynamics

DGX agent

arXiv:2608.01724v1 Announce Type: new Abstract: Group conversations are fundamental to human collaboration, yet standard large language models (LLMs) still struggle with the complexities of multi-part

applicationsarxiv-cs-cl
4 Aug 2026
Model Releases

Video Models as Native 4D Renderers: World-Grounded Conditioning from Animated Mesh

DGX agent

arXiv:2608.00094v1 Announce Type: new Abstract: Pretrained video diffusion models can act as renderers when the desired scene state is already specified by an animated mesh, a camera trajectory, and a

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs

DGX agent

arXiv:2608.00013v1 Announce Type: new Abstract: Choosing the right large language model (LLM) backbone is the most consequential decision when building a vision-language model (VLM), yet it remains fu

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Why Large Language Models Fail at Tabular Prediction

DGX agent

arXiv:2608.02412v1 Announce Type: new Abstract: Large language models (LLMs) have become the default tool for a remarkable range of tasks, yet they have had conspicuously little success at one of the

model-releasesarxiv-cs-lg
4 Aug 2026
Local Ai

an espresso Q/A model running fully offline on an ESP32S3

DGX agent

i already had an esp32 generating stories, but generating text is not the same as receiving a question and giving a useful answer. barista v0.1, a small model trained for espresso troubleshooting and

local-air-localllama
3 Aug 2026
Model Releases

Can Large Language Models Derive New Knowledge? A Dynamic Benchmark for Biological Knowledge Discovery

DGX agent

arXiv:2603.03322v2 Announce Type: replace-cross Abstract: Recent advancements in Large Language Model (LLM) agents have demonstrated remarkable potential in automatic knowledge discovery. However, rig

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

even within math, frontier models are far from being general intelligences. and we still need people to assess which of their output are tru…

DGX agent

even within math, frontier models are far from being general intelligences. and we still need people to assess which of their output are trustworthy and which are not. All frontier AI models (includin

model-releasesgary-marcus--x
3 Aug 2026
Model Releases

Harnessing the Wisdom of LLM Crowds through Complementarity-Driven Iterative Collaboration

DGX agent

arXiv:2607.29087v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in enterprise settings, yet individual models remain bounded by model-specific capability limitat

model-releasesarxiv-cs-ai
3 Aug 2026
← Previous
1…5657585960…1249
Next →