AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,929 results
13 Aug 2026

— Google AI Pro and Ultra subscribers can experience 3.7 Flash today via Spark in the @GeminiApp — Access the model in the Gemini Enterprise…

Model ReleasesDGX agent

— Google AI Pro and Ultra subscribers can experience 3.7 Flash today via Spark in the @GeminiApp — Access the model in the Gemini Enterprise Agent Platform and Gemini Enterprise app — Build in the Gem

How to Spend Your Oracle Budget: Practical Guidance for Protein Structure Prediction Models

Model ReleasesDGX agent

arXiv:2608.12192v1 Announce Type: new Abstract: Foundation models for protein structure prediction remain unreliable on certain targets. External oracles can flag and correct these failures, but biolo

StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2608.11671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models can follow instructions and manipulate objects, but their performance often collapses out of distribution (OOD), whe

12 Aug 2026

Attention-Path Fragility as an Uncertainty Signal in Large Language Models

Model ReleasesDGX agent

arXiv:2608.11138v1 Announce Type: cross Abstract: We propose that a model's uncertainty about a token is reflected not only in the breadth of its output distribution but also in whether a confident pr

Azure Content Understanding GPT-5 Series Guide: Model Selection, Grounding Improvements, and Confidence Enhancements

Model ReleasesDGX agent

Enterprise content is no longer just something people consume. As organizations increasingly rely on AI to extract and act on information from documents, images, audio, and video, Azure Content Unders

CurveFP: Rational-Radix Logarithmic Datatypes with Closed Products for Language Models

Model ReleasesDGX agent

arXiv:2608.10010v1 Announce Type: new Abstract: Low-precision datatypes reduce language-model cost, but most formats optimize scalar fidelity while leaving the arithmetic induced by their products unc

Generator-Guided Inverse Sampling for Levy-Driven Generative Models

ResearchDGX agent

arXiv:2608.10384v1 Announce Type: new Abstract: This paper studies inverse sampling for Levy-driven generative models from the perspective of Markov generators. Unlike conventional diffusion models, L

Grok 4.6 is an excellent model. I’ve been using it heavily for the past couple of weeks and it handles everything from simple coding & code …

Model ReleasesDGX agent

Grok 4.6 is an excellent model. I’ve been using it heavily for the past couple of weeks and it handles everything from simple coding & code review all the way to designing and debugging complex system

imo GDPVal is probably the most important benchmark, it measures the performance of models on real world tasks Big leap in performance here …

Model ReleasesDGX agent

imo GDPVal is probably the most important benchmark, it measures the performance of models on real world tasks Big leap in performance here to top it at a great price, congrats to @SpaceXAI team & loo

Mixture-of-Experts-based Entropy Model for Learned Image Compression

ResearchDGX agent

arXiv:2608.10947v1 Announce Type: new Abstract: Learned image compression has seen significant progress in recent years with the development of end-to-end learned models that achieve better compressio

MRIComp4Flow: Compression of 3D Brain MRI for Training Multi-Modal Generative Models

TutorialsDGX agent

arXiv:2608.10291v1 Announce Type: cross Abstract: Large-scale multi-modal MRI datasets impose substantial storage and I/O costs, limiting the training of 3D generative models on commodity infrastructu

Never Stop Speaking: a Denial-of-Service Attack on End-to-End Speech Language Models

SafetyDGX agent

arXiv:2608.10405v1 Announce Type: cross Abstract: Many studies have shown that specially crafted inputs can induce large language models (LLMs) to generate excessively long outputs, resulting in signi

PRMU: A Corpus-Free Benchmark for Person-Centric Knowledge Unlearning in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2608.11149v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable capabilities in storing and recalling rich person-related knowledge, raising incre

Serve Qwen3.8-2.4T-A95B, a 2.4T-Parameter Model, with Configurable Reasoning on NVIDIA GB300 NVL72

Model ReleasesDGX agent

Alibaba released the open‑weights Qwen3.8‑2.4T‑A95B (Qwen3.8‑Max), a fine‑grained mixture‑of‑experts model with 2.4 trillion parameters, hybrid full‑ and linear‑attention, a one‑million‑token context

SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information

Model ReleasesDGX agent

arXiv:2608.10692v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as mobile assistants, where a key challenge is leveraging personal information scattered across

v0.32.10

Model ReleasesDGX agent

What's Changed Models that don't set a repeat_penalty now default to 1.0 (off) instead of 1.1, matching other engines and speeding up speculative decoding; set a per-model parameter if an older model

When should you start post-training your own models? @FireworksAI_HQ CEO @lqiao’s answer: after product-market fit. Not because it's hard...…

SafetyDGX agent

When should you start post-training your own models? @FireworksAI_HQ CEO @lqiao’s answer: after product-market fit. Not because it's hard... but because only after PMF is the data coming off your prod

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models

Model ReleasesDGX agent

arXiv:2608.11024v1 Announce Type: new Abstract: Attribute hallucination---where vision-language models (VLMs) correctly identify an object but mischaracterize its properties---is prevalent yet mechani

11 Aug 2026

A continually expandable foundation model for brain MRI

ResearchDGX agent

arXiv:2608.08319v1 Announce Type: new Abstract: Brain magnetic resonance imaging (MRI) is central to neuroscience and clinical assessment, but models are commonly developed for individual diseases, po

Anyone Using (Koreas) 'Solar Open 2' (250B, 15B) Model?

Model ReleasesDGX agent

I just heard of this model. Seems to be a competitor to DeepSeek V4 Flash. About the same size and active parameters. Anyone tested it compared to V4 Flash? Link: https://huggingface.co/upstage/Solar-

CMU-Drive and V2V-VLA: Cooperative Multi-agent Unified Driving with Reasoning Benchmark and Vehicle-to-Vehicle Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2608.07621v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently achieved impressive performance for end-to-end autonomous driving, yet existing approaches are primari

Continued development of the model based on the SSN [D]

ResearchDGX agent

Back after ~6 months — rebuilding my spiking language model around CPU-first inference Hey everyone. It’s been around six months since I last posted anything about this project here. Some of you might

Distilling CT Foundation Models into Editable Concept Bottlenecks for Lung Nodule Malignancy Prediction

ResearchDGX agent

arXiv:2608.07857v1 Announce Type: cross Abstract: Foundation models provide transferable CT representations, but predictions based directly on these embeddings are difficult to interpret. We developed

From Evaluated Models to Evaluation Aids: A Multi-Evidence Study of LLM-Based Difficulty Calibration for Programming Examinations

Model ReleasesDGX agent

arXiv:2608.07523v1 Announce Type: cross Abstract: Difficulty differences across parallel-class programming examinations affect the fairness of course assessment. This study repositions large language

I’ve been thinking about how agents can learn inside world models for years. We decided to scale up our RSI Lab to bridge recursive self-imp…

TutorialsDGX agent

I’ve been thinking about how agents can learn inside world models for years. We decided to scale up our RSI Lab to bridge recursive self-improvement with physical AI and robotics. We are looking for f

LegoLM: Structured Weight Sharing for Large Language Models

Model ReleasesDGX agent

arXiv:2608.08652v1 Announce Type: cross Abstract: We present LegoLM{}, a structured weight-sharing compression framework for large language models grounded in a systematic study of why global weight s

Memorization Dynamics in Knowledge Distillation for Language Models

Model ReleasesDGX agent

arXiv:2601.15394v2 Announce Type: replace Abstract: Knowledge Distillation (KD) is increasingly adopted to transfer capabilities from large language models to smaller ones, offering significant improv

Nemotron 3.5 Lightning is available on Together AI through Dedicated Model Inference, giving teams reserved capacity and predictable perform…

Model ReleasesDGX agent

Nemotron 3.5 Lightning is available on Together AI through Dedicated Model Inference, giving teams reserved capacity and predictable performance for high-volume agent workloads. Start building: https:

NVIDIA Nemotron 3.5 Lightning is now live on Together AI. The fastest open model in its class is built for always-on agents that need to com…

Model ReleasesDGX agent

NVIDIA’s Nemotron 3.5 Lightning—a fast open AI model designed for always‑on agents that perform high‑volume, specialized work—has been launched on the Together AI platform as of 11 August 2026. NVIDIA

OpenMHC: Accelerating the Science of Wearable Foundation Models

Model ReleasesDGX agent

arXiv:2607.16235v3 Announce Type: replace-cross Abstract: Mobile and wearable devices offer an unprecedented opportunity for continuous, passive health monitoring and active health coaching. However,

Prompt engineering does not universally improve Large Language Model performance across clinical decision-making tasks

Model ReleasesDGX agent

arXiv:2512.22966v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated promise in medical knowledge assessments, yet their practical utility in real-world clinical decision

Reducing Pretraining-Generation Mismatch in Diffusion Language Models

Model ReleasesDGX agent

arXiv:2608.09424v1 Announce Type: new Abstract: Autoregressive language models align training and use: generation conditions on a clean prompt, and training predicts future tokens from clean left cont

Scaling Inherently Interpretable Language Models

ResearchDGX agent

arXiv:2608.07594v1 Announce Type: cross Abstract: Interpretability is often treated as a tax on capability: language models are trained as opaque systems, then explained after the fact, with methods w

SpikeWorld: Fast-State Adaptation for Frozen Spiking World Models

Model ReleasesDGX agent

arXiv:2608.07712v1 Announce Type: cross Abstract: A predictive model receives a self-supervised signal whenever the consequence of an action is observed. Using that signal after deployment is difficul

Stealing Reasoning Traces from Proprietary LLM APIs

Model ReleasesDGX agent

Stealing Reasoning Traces from Proprietary LLM APIs A vanity domain name (stolen-thoughts.com) for a neat paper: Anthropic, OpenAI, and Google return encrypted chain-of-thought blocks to clients that

Thinking Hard, Not Smart: Reasoning Models Fail to Ration Test-Time Compute Across Questions

TutorialsDGX agent

arXiv:2608.07968v1 Announce Type: cross Abstract: Reasoning language models increasingly use test-time compute to improve performance, but existing evaluations typically study this compute one questio

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2608.08168v1 Announce Type: new Abstract: While Large Language Models (LLMs) employing Chain-of-Thought (CoT) exhibit superior reasoning capabilities, the neural mechanisms distinguishing this e

TrustRoboReward: Preference-Ordered Isotonic Score Editing for Multi-Paradigm Robot Reward Models

Model ReleasesDGX agent

arXiv:2608.08491v1 Announce Type: new Abstract: Reward models are a bottleneck for reinforcement learning in embodied AI. Long-horizon robotic manipulation requires scalable vision feedback beyond han

Unified Hallucination Fuzzing for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2608.07525v1 Announce Type: cross Abstract: Hallucination remains a persistent challenge for Multimodal Large Language Models (MLLMs), severely limiting their reliability in high-stakes applicat

Unsure but Certain: Uncovering the Representation-Confidence Gap in Diffusion Language Models

ResearchDGX agent

arXiv:2608.08791v1 Announce Type: new Abstract: Diffusion language models use broad context to create text, suggesting they might handle input noise better than standard models. Testing reveals this i

VectraYX-Vision-1B: A Sub-2B Spanish/LATAM Cybersecurity Vision-Language Model with Structured Visual Reasoning and Native Tool Use

Model ReleasesDGX agent

arXiv:2608.08477v1 Announce Type: new Abstract: We present VectraYX-Vision-1B, a sub-2B vision-language model (VLM) for Spanish/LATAM cybersecurity imagery, coupling a frozen SigLIP-so400m encoder to

Vid2WAM: Distilling Video Diffusion Priors into World Action Models

SafetyDGX agent

arXiv:2608.08558v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling future visual dynamics and actions. However, their scalability and generali

10 Aug 2026

CoDAT: Collaborative Dual-Attention Transformer with Low-Cost Temporal Modeling for Efficient Edge Action Recognition

Model ReleasesDGX agent

arXiv:2608.06691v1 Announce Type: new Abstract: Real-time human action recognition on Internet-of-Things (IoT) edge devices requires models that capture rich spatio-temporal cues within strict latency

CrossTracer: Cross-Embodiment Navigation via VLA Model Reasoning and Trace Residuals Adapting

Model ReleasesDGX agent

arXiv:2608.06688v1 Announce Type: new Abstract: Vision-language-action (VLA) models provide strong semantic priors for robot navigation, but they often ignore embodiment-specific mobility constraints.

EchoVLA: Robotic Vision-Language-Action Model with Synergistic Declarative Memory for Mobile Manipulation

Model ReleasesDGX agent

arXiv:2511.18112v3 Announce Type: replace Abstract: Recent progress in Vision-Language-Action (VLA) models has enabled embodied agents to interpret multimodal instructions and perform complex tasks. H

Game-Theoretic Inverse Reinforcement Learning for Modeling Competitive Human Driving: A Cut-in Prediction Study

Model ReleasesDGX agent

arXiv:2608.06445v1 Announce Type: cross Abstract: Capturing the strategic decision-making inherent in competitive human driving is critical for autonomous vehicle safety and traffic simulation. This s

Geo-Spatial Concept Probing of Large Language Models: Abstraction, Compositionality, and Grounding

Model ReleasesDGX agent

arXiv:2608.07353v1 Announce Type: cross Abstract: Understanding concepts is fundamental to generalization. Despite their impressive performance on a wide range of tasks, Large Language Models (LLMs) s

GraphVerse: A Comprehensive Visual Graph Reasoning Benchmark for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2608.06769v1 Announce Type: new Abstract: Recent Multimodal Large Language Models (MLLMs) have achieved remarkable progress across diverse vision-language tasks, creating an urgent need for more

LoRAScan: Detecting Backdoor Prompts in Low-Rank Adapters for Large Language Models via Down-Projection Activation Spikes

TutorialsDGX agent

arXiv:2608.06795v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) enables efficient specialization and distribution of large language models through compact adapters. However, untrusted ada

Meta returns to open weights: Muse Glimmer, its first open-weights release since Llama 4, scores 35 on the Artificial Analysis Intelligence …

Model ReleasesDGX agent

Meta returns to open weights: Muse Glimmer, its first open-weights release since Llama 4, scores 35 on the Artificial Analysis Intelligence Index. It is a 30B-parameter model, and the first from Meta

Pathryoshka: Compressing Pathology Foundation Models via Multi-Teacher Knowledge Distillation with Nested Embeddings

ResearchDGX agent

arXiv:2511.23204v2 Announce Type: replace Abstract: Pathology foundation models (FMs) have driven significant progress in computational pathology. However, these high-performing models can easily exce

Symbolic Graphics Programming with Large Language Models

Model ReleasesDGX agent

arXiv:2509.05208v2 Announce Type: replace Abstract: Large language models (LLMs) excel at program synthesis, yet their ability to produce symbolic graphics programs (SGPs) that render into precise vis

Towards Multi-Label Graph Foundation Models: from Single-Vector Representation Learning to Multi-Semantic Basis Learning

ResearchDGX agent

arXiv:2608.06394v1 Announce Type: new Abstract: Multi-label node classification is an important yet challenging task in graph learning, where nodes exhibit multiple semantics simultaneously. Existing

9 Aug 2026

Anyone already used a model imported directly in the ollama cloud

Local AiDGX agent

Ollama allons you to import model but have you ever tried doing so ? Like running model imported from hugging face or you own model ? Any use case you wanna share ? Very curious about that submitted b

I Turned My Underused Gaming Laptop Into a Local AI Workstation

Local AiDGX agent

TL;DR: I am building a Windows-first local AI setup for people who want to try local LLMs without spending days choosing models, setting up Ollama, Docker, WSL, Open WebUI, agents, and tool permission

7 Aug 2026

Big, Bright, or Invisible: A Frozen-Feature Benchmark of 3D CT Foundation Models

Model ReleasesDGX agent

arXiv:2608.05960v1 Announce Type: cross Abstract: Routine CT interpretation is inherently comprehensive, capturing incidental findings across the entire scan volume. 3D CT foundation models could assi

Human-Like Anaphor Resolution in Large Language Models

Model ReleasesDGX agent

arXiv:2608.05630v1 Announce Type: new Abstract: Anaphors are expressions that refer to other expressions, called antecedents. The process of connecting the two is called resolution. Cognitive science

KV-Skill: Forging Expertise in the Model's Native Language

Model ReleasesDGX agent

arXiv:2608.05475v1 Announce Type: new Abstract: Task knowledge is commonly stored either as text in the prompt or as an update to model weights. Text is modular but must be interpreted on every use, w

MASS: Multiplayer World Models with Authoritative Shared State

Model ReleasesDGX agent

arXiv:2608.06257v1 Announce Type: new Abstract: Current video world models struggle in multiplayer environments because they entangle world state with view-dependent visual latents, leading to redunda

MetaboLLM: a metabolomics-specialized large language model for biochemical knowledge integration and predictive metabolite graph construction

Model ReleasesDGX agent

arXiv:2608.06253v1 Announce Type: new Abstract: Metabolomics knowledge is distributed across heterogeneous resources and remains difficult to translate into predictive representations. We developed Me

← Previous
1…5859606162…999
Next →