AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
All
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,507 results
Model Releases

Selective Depthwise Separable Convolution for Lightweight Joint Source-Channel Coding in Wireless Image Transmission

DGX agent

arXiv:2604.22338v1 Announce Type: cross Abstract: Depthwise separable convolutional (DSConv) layers have been successfully applied to deep learning (DL)-based joint source-channel coding (JSCC) scheme

model-releasesarxiv-cs-cv
27 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Shaken or Stirred? An Analysis of MetaFormer's Token Mixing for Medical Imaging

DGX agent

arXiv:2510.05971v3 Announce Type: replace Abstract: The generalization of the Transformer architecture via MetaFormer has reshaped our understanding of its success in computer vision. By replacing sel

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

SHAPE: Unifying Safety, Helpfulness and Pedagogy for Educational LLMs

DGX agent

arXiv:2604.22134v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely explored in educational scenarios. We identify a critical vulnerability in current educational LLMs, pedag

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Sovereign Agentic Loops: Decoupling AI Reasoning from Execution in Real-World Systems

DGX agent

arXiv:2604.22136v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly issue API calls that mutate real systems, yet many current architectures pass stochastic model outputs

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

SpaMEM: Benchmarking Dynamic Spatial Reasoning via Perception-Memory Integration in Embodied Environments

DGX agent

arXiv:2604.22409v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced static visual--spatial reasoning, yet they often fail to preserve long-horizon spatial coherence

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Spend Less, Fit Better: Budget-Efficient Scaling Law Fitting via Active Experiment Selection

DGX agent

arXiv:2604.22753v1 Announce Type: new Abstract: Scaling laws are used to plan multi-million-dollar training runs, but fitting those laws can itself cost millions. In modern large-scale workflows, asse

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Sum-of-Checks: Structured Reasoning for Surgical Safety with Large Vision-Language Models

DGX agent

arXiv:2604.22156v1 Announce Type: cross Abstract: Purpose: Accurate assessment of the Critical View of Safety (CVS) during laparoscopic cholecystectomy is essential to prevent bile duct injury, a comp

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Test-Time Matching: Unlocking Compositional Reasoning in Multimodal Models

DGX agent

arXiv:2510.07632v2 Announce Type: replace Abstract: Frontier AI models have achieved remarkable progress, yet recent studies suggest they struggle with compositional reasoning, often performing at or

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

The DeepSeek V4 garbled output bug in open source inference engine is fixed in SGLang. To everyone affected over the weekend, sorry for the …

DGX agent

The DeepSeek V4 garbled output bug in open source inference engine is fixed in SGLang. To everyone affected over the weekend, sorry for the trouble. Huge thanks to @Ant_Group for landing the fix PR. I

model-releasesollama--x
27 Apr 2026
Model Releases

The Download: DeepSeek’s latest AI breakthrough, and the race to build world models

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Three reasons why DeepSeek’s new model matters On Friday, Chin

model-releasesmit-tech-review
27 Apr 2026
Model Releases

The next phase of the Microsoft OpenAI partnership

DGX agent

Microsoft and OpenAI announced an expanded partnership extending their collaboration on AI development and deployment. The partnership likely involves increased investment from Microsoft, expanded int

model-releasesopenai
27 Apr 2026
Model Releases

The 'triple usage' period for GLM-5.1 and GLM-5-Turbo is now extended to June 30. Availability: Anytime except 2-6 AM ET.

DGX agent

The 'triple usage' period for GLM-5.1 and GLM-5-Turbo is now extended to June 30. Availability: Anytime except 2-6 AM ET. Usage limits tripled for GLM-5-Turbo in GLM Coding Plan! Enjoy the same high-v

model-releaseszhipu-ai--x
27 Apr 2026
Model Releases

This is kinda interesting. Anthro probably needs to scale up Account Support via Claude or via humans (traditional account managers). Would …

DGX agent

This is kinda interesting. Anthro probably needs to scale up Account Support via Claude or via humans (traditional account managers). Would be funny if they chose humans. Alternatively, more and more

model-releasessoumith-chintala--x
27 Apr 2026
Model Releases

Time-Localized Parametric Decomposition of Respiratory Airflow for Sub-Breath Analysis

DGX agent

arXiv:2604.22695v1 Announce Type: cross Abstract: Respiratory airflow signals provide critical insight into breathing mechanics, yet conventional analysis methods remain limited in their ability to ch

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Toward Automated Robustness Evaluation of Mathematical Reasoning

DGX agent

arXiv:2506.05038v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in various reasoning-intensive tasks. However, these models exhibit unexpecte

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Towards Adaptive Continual Model Merging via Manifold-Aware Expert Evolution

DGX agent

arXiv:2604.22464v1 Announce Type: new Abstract: Continual Model Merging (CMM) sequentially integrates task-specific models into a unified architecture without intensive retraining. However, existing C

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Towards Temporal Compositional Reasoning in Long-Form Sports Videos

DGX agent

arXiv:2604.22226v1 Announce Type: new Abstract: Sports videos are a challenging domain for multimodal understanding because they involve complex and dynamic human activities. Despite rapid progress in

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

TRACE: Topology-aware Reconstruction of Accidents in CARLA for AV Evaluation

DGX agent

arXiv:2604.22068v1 Announce Type: cross Abstract: Validating Autonomous Vehicles (AVs) requires exposure to rare, safety-critical scenarios, infrequent in routine driving data. Existing benchmarks add

model-releasesarxiv-cs-ro
27 Apr 2026
Model Releases

TreeCoder: Systematic Exploration and Optimisation of Decoding and Constraints for LLM Code Generation

DGX agent

arXiv:2511.22277v2 Announce Type: replace Abstract: Large language models (LLMs) have shown remarkable ability to generate code, yet their outputs often violate syntactic or semantic constraints when

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

TS-Arena -- A Live Forecast Pre-Registration Platform

DGX agent

arXiv:2512.20761v3 Announce Type: replace-cross Abstract: Time Series Foundation Models (TSFMs) are transforming the field of forecasting. However, evaluating them on historical data is increasingly d

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

TuneForge: an MCP server that lets your coding agent (Claude, Cursor, etc.) handle dataset generation, LoRA fine-tuning, RL, and evaluation directly in chat

DGX agent

TuneForge is an MCP (Model Context Protocol) server that enables coding agents like Claude and Cursor to perform machine learning operations directly within chat interfaces, including dataset generati

model-releasesr-ollama
27 Apr 2026
Model Releases

UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents

DGX agent

arXiv:2602.07038v2 Announce Type: replace-cross Abstract: Key Information Extraction (KIE) from real-world documents remains challenging due to substantial variations in layout structures, visual qual

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Universal Transformers Need Memory: Depth-State Trade-offs in Adaptive Recursive Reasoning

DGX agent

arXiv:2604.21999v1 Announce Type: cross Abstract: We study learned memory tokens as computational scratchpad for a single-block Universal Transformer (UT) with Adaptive Computation Time (ACT) on Sudok

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

UR^2: Unify RAG and Reasoning through Reinforcement Learning

DGX agent

arXiv:2508.06165v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown strong capabilities through two complementary paradigms: Retrieval-Augmented Generation (RAG) for know

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

We are now enabling a queue for DeepSeek v4 Pro, expect longer time-to-first-token instead of degrading service. please bear with us 🙏🙏🙏…

DGX agent

Ollama has implemented a queue system for DeepSeek v4 Pro to manage high demand, which will result in longer initial response times rather than service degradation. Users are requested to be patient d

model-releasesollama--x
27 Apr 2026
Model Releases

When AI Speaks, Whose Values Does It Express? A Cross-Cultural Audit of Individualism-Collectivism Bias in Large Language Models

DGX agent

arXiv:2604.22153v1 Announce Type: cross Abstract: When you ask an AI assistant for advice about your career, your marriage, or a conflict with your family, does it give you the same answer regardless

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

When Cow Urine Cures Constipation on YouTube: Limits of LLMs in Detecting Culture-specific Health Misinformation

DGX agent

arXiv:2604.22002v1 Announce Type: new Abstract: Social media platforms have become primary channels for health information in the Global South. Using gomutra (cow urine) discourse on YouTube in India

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

When Does LLM Self-Correction Help? A Control-Theoretic Markov Diagnostic and Verify-First Intervention

DGX agent

arXiv:2604.22273v1 Announce Type: new Abstract: Iterative self-correction is widely used in agentic LLM systems, but when repeated refinement helps versus hurts remains unclear. We frame self-correcti

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Wiggle and Go! System Identification for Zero-Shot Dynamic Rope Manipulation

DGX agent

arXiv:2604.22102v1 Announce Type: cross Abstract: Many robotic tasks are unforgiving; a single mistake in a dynamic throw can lead to unacceptable delays or unrecoverable failure. To mitigate this, we

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Xiaomi MiMo-V2.5 is now officially open-sourced! MIT License, supporting commercial deployment, continued training, and fine-tuning - no add…

DGX agent

Xiaomi MiMo-V2.5 is now officially open-sourced! MIT License, supporting commercial deployment, continued training, and fine-tuning - no additional authorization required. Two models, both supporting

model-releasesjeremy-howard--x
27 Apr 2026
Model Releases

❤️ @ying11231, @BanghuaZ and @lmsysorg @sgl_project @radixark Let's go DeepSeek v4 Pro!

DGX agent

❤️ @ying11231, @BanghuaZ and @lmsysorg @sgl_project @radixark Let's go DeepSeek v4 Pro! The DeepSeek V4 garbled output bug in open source inference engine is fixed in SGLang. To everyone affected over

model-releasesollama--x
27 Apr 2026
Model Releases

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threa…

DGX agent

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threads. I don’t think we’ve cracked the right UI for managing ag

model-releasesallie-k--miller--x
26 Apr 2026
Model Releases

April was a pretty strong month for LLM releases: - Gemma 4 - GLM-5.1 - Qwen3.6 - Kimi K2.6 - DeepSeek V4 All are now added to the LLM Archi…

DGX agent

April saw significant activity in large language model releases, with five major models introduced including Gemma 4, GLM-5.1, Qwen 3.6, Kimi K2.6, and DeepSeek V4. These releases have been added to a

model-releasessebastian-raschka--x
26 Apr 2026
Model Releases

@badlogicgames Yup: https://x.com/antirez/status/2048344770234249588 The GGUF tool calling template is wrong but I'm uploading a new GGUF fi…

DGX agent

@badlogicgames Yup: https://x.com/antirez/status/2048344770234249588 The GGUF tool calling template is wrong but I'm uploading a new GGUF file. Otherwise there is the right template file in one of the

model-releasesclem-delangue--x
26 Apr 2026
Model Releases

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate'…

DGX agent

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate' and have that thing be infinitely scalable. See below image

model-releasesallie-k--miller--x
26 Apr 2026
Model Releases

🔥DeepSeek Input Cache Price Drop! Effective immediately, the price for input cache hits across the ENTIRE DeepSeek API series is reduced to…

DGX agent

🔥DeepSeek Input Cache Price Drop! Effective immediately, the price for input cache hits across the ENTIRE DeepSeek API series is reduced to just 1/10th of the original price! Build more efficiently fo

model-releasesjeremy-howard--x
26 Apr 2026
Model Releases

DeepSeek v4 Flash with *local inference* after 24h of playing with that: even with the 2 bit selective quantization GGUF, iti is the FIRST t…

DGX agent

DeepSeek v4 Flash with *local inference* after 24h of playing with that: even with the 2 bit selective quantization GGUF, iti is the FIRST time I feel I have a frontier model running on my computer. T

model-releasesclem-delangue--x
26 Apr 2026
Model Releases

extremely happy that we are in q2 2026, and engineers i look up to are plowing a path for the local model future. yes, we are still beholden…

DGX agent

extremely happy that we are in q2 2026, and engineers i look up to are plowing a path for the local model future. yes, we are still beholden to some lab publishing weights. but i take that over compan

model-releasesclem-delangue--x
26 Apr 2026
Model Releases

FYI Claude Code is mostly a vibe-coded product (as they say, 100% written by Claude) It's the worst harness for Opus 4.6 among ANY harness o…

DGX agent

FYI Claude Code is mostly a vibe-coded product (as they say, 100% written by Claude) It's the worst harness for Opus 4.6 among ANY harness on Terminal-Bench 2 I feel sorry for Claude Code I know they'

model-releasesjeremy-howard--x
26 Apr 2026
Model Releases

🚨GUYS QWEN3.6 27B/35B @unslothai MLX QUANTS @Brooooook_lyn cooked 🔥 Apple Silicon users, the exact models y’all have been asking for LANDE…

DGX agent

🚨GUYS QWEN3.6 27B/35B @unslothai MLX QUANTS @Brooooook_lyn cooked 🔥 Apple Silicon users, the exact models y’all have been asking for LANDED. Unsloth mixed-precision quants in native MLX format. >Fast.

model-releasesclem-delangue--x
26 Apr 2026
Model Releases

Higher res figures (and summaries) in the LLM architecture gallery: https://sebastianraschka.com/llm-architecture-gallery/#card-deepseek-v4-…

DGX agent

Sebastian Raschka has updated his LLM architecture gallery with higher resolution figures and improved summaries, including coverage of the DeepSeek V4 model architecture. This resource provides visua

model-releasessebastian-raschka--x
26 Apr 2026
Model Releases

I actually switched my personal Claude subscription to this (currently using Mimo v2)

DGX agent

I actually switched my personal Claude subscription to this (currently using Mimo v2) Nous Portal offers everything you need to build with Hermes Agent in one easy subscription: → 300+ models from eve

model-releasesnous-research--x
26 Apr 2026
Model Releases

I am increasingly bullish on open source harnesses (like OpenCode) not because they will be better than SOTA closed harnesses, but because t…

DGX agent

I am increasingly bullish on open source harnesses (like OpenCode) not because they will be better than SOTA closed harnesses, but because they will never pull shady stuff like what Claude Code and ot

model-releasesjeremy-howard--x
26 Apr 2026
Model Releases

Inside Andon Market, billed as the first retail boutique run by an AI agent; the Andon Labs experiment uses a Claude Sonnet 4.6-based agent to run the boutique (Heather Knight/New York Times)

DGX agent

Heather Knight / New York Times: Inside Andon Market, billed as the first retail boutique run by an AI agent; the Andon Labs experiment uses a Claude Sonnet 4.6-based agent to run the boutique — Andon

model-releasestechmeme
26 Apr 2026
Model Releases

Model is available here @simonw @ivanfioravanti https://huggingface.co/mlx-community/DeepSeek-V4-Flash-2bit-DQ

DGX agent

A quantized version of DeepSeek-V4-Flash has been released on Hugging Face by the MLX community in 2-bit format with dynamic quantization, making the model more efficient for inference on resource-con

model-releasesclem-delangue--x
26 Apr 2026
Model Releases

NEW paper from Alibaba. A 30B MoE with only 3B active params matches Qwen3-235B on real tool-use workloads. AgenticQwen-30B-A3B: 50.2 averag…

DGX agent

NEW paper from Alibaba. A 30B MoE with only 3B active params matches Qwen3-235B on real tool-use workloads. AgenticQwen-30B-A3B: 50.2 average on TAU-2 + BFCL-V4 Multi-Turn. AgenticQwen-8B: 47.4. Both

model-releasesdair-ai--x
26 Apr 2026
Model Releases

No more updating needed for new model releases via Nous Portal and OpenRouter!

DGX agent

No more updating needed for new model releases via Nous Portal and OpenRouter! Hermes will no longer have to be updated to receive model list curation updates for several providers, including Nous Por

model-releasesnous-research--x
26 Apr 2026
Model Releases

OpenAI released a new dataset on the hub 👀 Benchmark for 'Making ChatGPT better for clinicians' https://huggingface.co/datasets/openai/heal…

DGX agent

OpenAI released a new dataset on Hugging Face Hub designed to improve ChatGPT's performance in clinical settings, addressing the specific needs and workflows of healthcare professionals. The dataset s

model-releasesclem-delangue--x
26 Apr 2026
← Previous
1…390391392393394…469
Next →