AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
Model Releases

Recent Advances in Multi-Agent Human Trajectory Prediction: A Comprehensive Review

DGX agent

arXiv:2506.14831v3 Announce Type: replace Abstract: With the emergence of powerful data-driven methods in human trajectory prediction (HTP), gaining a finer understanding of multi-agent interactions l

model-releasesarxiv-cs-cv
27 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition

DGX agent

arXiv:2604.22390v1 Announce Type: new Abstract: Visual Place Recognition (VPR) determines a query image's geographic location by matching it against geotagged databases. However, existing methods stru

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Regularized Meta-Learning for Improved Generalization

DGX agent

arXiv:2602.12469v2 Announce Type: replace Abstract: Deep ensemble methods often improve predictive performance, yet they suffer from three practical limitations: redundancy among base models that infl

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Relaxation-Informed Training of Neural Network Surrogate Models

DGX agent

arXiv:2604.22746v1 Announce Type: cross Abstract: ReLU neural networks trained as surrogate models can be embedded exactly in mixed-integer linear programs (MILPs), enabling global optimization over t

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Reliability Auditing for Downstream LLM tasks in Psychiatry: LLM-Generated Hospitalization Risk Scores

DGX agent

arXiv:2604.22063v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly utilized in clinical reasoning and risk assessment. However, their interpretive reliability in critical

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

ResRank: Unifying Retrieval and Listwise Reranking via End-to-End Joint Training with Residual Passage Compression

DGX agent

arXiv:2604.22180v1 Announce Type: cross Abstract: Large language model (LLM) based listwise reranking has emerged as the dominant paradigm for achieving state-of-the-art ranking effectiveness in infor

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Rethinking Publication: A Certification Framework for AI-Enabled Research

DGX agent

arXiv:2604.22026v1 Announce Type: new Abstract: AI research pipelines now produce a growing share of publishable academic output, including work that meets existing peer-review standards for quality a

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Rethinking Retrieval-Augmented Generation as a Cooperative Decision-Making Problem

DGX agent

arXiv:2602.18734v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has demonstrated strong effectiveness in knowledge-intensive tasks by grounding language generation in ex

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

RTX 5090 users: TensorRT-LLM vs llama.cpp (GGUF) for Coding Agents (Cline/RooCode) – Is the speed worth the VRAM limit?

DGX agent

This post compares TensorRT-LLM and llama.cpp (GGUF) as inference frameworks for running coding agents like Cline and RooCode on RTX 5090 GPUs, examining the tradeoff between inference speed and VRAM

model-releasesr-ollama
27 Apr 2026
Model Releases

Running Qwen3.5-397B-A17B (4bit quants, 177 GB) on two DGX Sparks using llama.cpp with RPC and RDMA:

DGX agent

This post documents a technical demonstration of running the large Qwen3.5-397B-A17B model across distributed hardware using llama.cpp with advanced networking protocols. The approach leverages 4-bit

model-releasesgeorgi-gerganov--x
27 Apr 2026
Model Releases

Scam Altman

DGX agent

Scam Altman Interesting how it works Elon puts up his own money, rounds up the absolute best AI talent on the planet, leverages every connection he has to secure serious resources, and launches OpenAI

model-releaseselon-musk--x
27 Apr 2026
Model Releases

Scam Altman and Greg Stockman stole a charity. Full stop. Greg got tens of billions of stock for himself and Scam got dozens of OpenAI side …

DGX agent

Scam Altman and Greg Stockman stole a charity. Full stop. Greg got tens of billions of stock for himself and Scam got dozens of OpenAI side deals with a piece of the action for himself, Y Combinator s

model-releaseselon-musk--x
27 Apr 2026
Model Releases

Score-based Membership Inference on Diffusion Models

DGX agent

arXiv:2509.25003v2 Announce Type: replace-cross Abstract: Membership inference attacks (MIAs) against Diffusion Models (DMs) raise pressing privacy concerns by revealing whether a sample was part of t

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Selective Depthwise Separable Convolution for Lightweight Joint Source-Channel Coding in Wireless Image Transmission

DGX agent

arXiv:2604.22338v1 Announce Type: cross Abstract: Depthwise separable convolutional (DSConv) layers have been successfully applied to deep learning (DL)-based joint source-channel coding (JSCC) scheme

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Shaken or Stirred? An Analysis of MetaFormer's Token Mixing for Medical Imaging

DGX agent

arXiv:2510.05971v3 Announce Type: replace Abstract: The generalization of the Transformer architecture via MetaFormer has reshaped our understanding of its success in computer vision. By replacing sel

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

SHAPE: Unifying Safety, Helpfulness and Pedagogy for Educational LLMs

DGX agent

arXiv:2604.22134v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely explored in educational scenarios. We identify a critical vulnerability in current educational LLMs, pedag

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Sovereign Agentic Loops: Decoupling AI Reasoning from Execution in Real-World Systems

DGX agent

arXiv:2604.22136v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly issue API calls that mutate real systems, yet many current architectures pass stochastic model outputs

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

SpaMEM: Benchmarking Dynamic Spatial Reasoning via Perception-Memory Integration in Embodied Environments

DGX agent

arXiv:2604.22409v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced static visual--spatial reasoning, yet they often fail to preserve long-horizon spatial coherence

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Spend Less, Fit Better: Budget-Efficient Scaling Law Fitting via Active Experiment Selection

DGX agent

arXiv:2604.22753v1 Announce Type: new Abstract: Scaling laws are used to plan multi-million-dollar training runs, but fitting those laws can itself cost millions. In modern large-scale workflows, asse

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Sum-of-Checks: Structured Reasoning for Surgical Safety with Large Vision-Language Models

DGX agent

arXiv:2604.22156v1 Announce Type: cross Abstract: Purpose: Accurate assessment of the Critical View of Safety (CVS) during laparoscopic cholecystectomy is essential to prevent bile duct injury, a comp

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Test-Time Matching: Unlocking Compositional Reasoning in Multimodal Models

DGX agent

arXiv:2510.07632v2 Announce Type: replace Abstract: Frontier AI models have achieved remarkable progress, yet recent studies suggest they struggle with compositional reasoning, often performing at or

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

The DeepSeek V4 garbled output bug in open source inference engine is fixed in SGLang. To everyone affected over the weekend, sorry for the …

DGX agent

The DeepSeek V4 garbled output bug in open source inference engine is fixed in SGLang. To everyone affected over the weekend, sorry for the trouble. Huge thanks to @Ant_Group for landing the fix PR. I

model-releasesollama--x
27 Apr 2026
Model Releases

The Download: DeepSeek’s latest AI breakthrough, and the race to build world models

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Three reasons why DeepSeek’s new model matters On Friday, Chin

model-releasesmit-tech-review
27 Apr 2026
Model Releases

The next phase of the Microsoft OpenAI partnership

DGX agent

Microsoft and OpenAI announced an expanded partnership extending their collaboration on AI development and deployment. The partnership likely involves increased investment from Microsoft, expanded int

model-releasesopenai
27 Apr 2026
Model Releases

The 'triple usage' period for GLM-5.1 and GLM-5-Turbo is now extended to June 30. Availability: Anytime except 2-6 AM ET.

DGX agent

The 'triple usage' period for GLM-5.1 and GLM-5-Turbo is now extended to June 30. Availability: Anytime except 2-6 AM ET. Usage limits tripled for GLM-5-Turbo in GLM Coding Plan! Enjoy the same high-v

model-releaseszhipu-ai--x
27 Apr 2026
Model Releases

This is kinda interesting. Anthro probably needs to scale up Account Support via Claude or via humans (traditional account managers). Would …

DGX agent

This is kinda interesting. Anthro probably needs to scale up Account Support via Claude or via humans (traditional account managers). Would be funny if they chose humans. Alternatively, more and more

model-releasessoumith-chintala--x
27 Apr 2026
Model Releases

Time-Localized Parametric Decomposition of Respiratory Airflow for Sub-Breath Analysis

DGX agent

arXiv:2604.22695v1 Announce Type: cross Abstract: Respiratory airflow signals provide critical insight into breathing mechanics, yet conventional analysis methods remain limited in their ability to ch

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Toward Automated Robustness Evaluation of Mathematical Reasoning

DGX agent

arXiv:2506.05038v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in various reasoning-intensive tasks. However, these models exhibit unexpecte

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Towards Adaptive Continual Model Merging via Manifold-Aware Expert Evolution

DGX agent

arXiv:2604.22464v1 Announce Type: new Abstract: Continual Model Merging (CMM) sequentially integrates task-specific models into a unified architecture without intensive retraining. However, existing C

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Towards Temporal Compositional Reasoning in Long-Form Sports Videos

DGX agent

arXiv:2604.22226v1 Announce Type: new Abstract: Sports videos are a challenging domain for multimodal understanding because they involve complex and dynamic human activities. Despite rapid progress in

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

TRACE: Topology-aware Reconstruction of Accidents in CARLA for AV Evaluation

DGX agent

arXiv:2604.22068v1 Announce Type: cross Abstract: Validating Autonomous Vehicles (AVs) requires exposure to rare, safety-critical scenarios, infrequent in routine driving data. Existing benchmarks add

model-releasesarxiv-cs-ro
27 Apr 2026
Model Releases

TreeCoder: Systematic Exploration and Optimisation of Decoding and Constraints for LLM Code Generation

DGX agent

arXiv:2511.22277v2 Announce Type: replace Abstract: Large language models (LLMs) have shown remarkable ability to generate code, yet their outputs often violate syntactic or semantic constraints when

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

TS-Arena -- A Live Forecast Pre-Registration Platform

DGX agent

arXiv:2512.20761v3 Announce Type: replace-cross Abstract: Time Series Foundation Models (TSFMs) are transforming the field of forecasting. However, evaluating them on historical data is increasingly d

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

TuneForge: an MCP server that lets your coding agent (Claude, Cursor, etc.) handle dataset generation, LoRA fine-tuning, RL, and evaluation directly in chat

DGX agent

TuneForge is an MCP (Model Context Protocol) server that enables coding agents like Claude and Cursor to perform machine learning operations directly within chat interfaces, including dataset generati

model-releasesr-ollama
27 Apr 2026
Model Releases

UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents

DGX agent

arXiv:2602.07038v2 Announce Type: replace-cross Abstract: Key Information Extraction (KIE) from real-world documents remains challenging due to substantial variations in layout structures, visual qual

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Universal Transformers Need Memory: Depth-State Trade-offs in Adaptive Recursive Reasoning

DGX agent

arXiv:2604.21999v1 Announce Type: cross Abstract: We study learned memory tokens as computational scratchpad for a single-block Universal Transformer (UT) with Adaptive Computation Time (ACT) on Sudok

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

UR^2: Unify RAG and Reasoning through Reinforcement Learning

DGX agent

arXiv:2508.06165v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown strong capabilities through two complementary paradigms: Retrieval-Augmented Generation (RAG) for know

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

We are now enabling a queue for DeepSeek v4 Pro, expect longer time-to-first-token instead of degrading service. please bear with us 🙏🙏🙏…

DGX agent

Ollama has implemented a queue system for DeepSeek v4 Pro to manage high demand, which will result in longer initial response times rather than service degradation. Users are requested to be patient d

model-releasesollama--x
27 Apr 2026
Model Releases

When AI Speaks, Whose Values Does It Express? A Cross-Cultural Audit of Individualism-Collectivism Bias in Large Language Models

DGX agent

arXiv:2604.22153v1 Announce Type: cross Abstract: When you ask an AI assistant for advice about your career, your marriage, or a conflict with your family, does it give you the same answer regardless

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

When Cow Urine Cures Constipation on YouTube: Limits of LLMs in Detecting Culture-specific Health Misinformation

DGX agent

arXiv:2604.22002v1 Announce Type: new Abstract: Social media platforms have become primary channels for health information in the Global South. Using gomutra (cow urine) discourse on YouTube in India

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

When Does LLM Self-Correction Help? A Control-Theoretic Markov Diagnostic and Verify-First Intervention

DGX agent

arXiv:2604.22273v1 Announce Type: new Abstract: Iterative self-correction is widely used in agentic LLM systems, but when repeated refinement helps versus hurts remains unclear. We frame self-correcti

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Wiggle and Go! System Identification for Zero-Shot Dynamic Rope Manipulation

DGX agent

arXiv:2604.22102v1 Announce Type: cross Abstract: Many robotic tasks are unforgiving; a single mistake in a dynamic throw can lead to unacceptable delays or unrecoverable failure. To mitigate this, we

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Xiaomi MiMo-V2.5 is now officially open-sourced! MIT License, supporting commercial deployment, continued training, and fine-tuning - no add…

DGX agent

Xiaomi MiMo-V2.5 is now officially open-sourced! MIT License, supporting commercial deployment, continued training, and fine-tuning - no additional authorization required. Two models, both supporting

model-releasesjeremy-howard--x
27 Apr 2026
Model Releases

❤️ @ying11231, @BanghuaZ and @lmsysorg @sgl_project @radixark Let's go DeepSeek v4 Pro!

DGX agent

❤️ @ying11231, @BanghuaZ and @lmsysorg @sgl_project @radixark Let's go DeepSeek v4 Pro! The DeepSeek V4 garbled output bug in open source inference engine is fixed in SGLang. To everyone affected over

model-releasesollama--x
27 Apr 2026
Model Releases

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threa…

DGX agent

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threads. I don’t think we’ve cracked the right UI for managing ag

model-releasesallie-k--miller--x
26 Apr 2026
Model Releases

April was a pretty strong month for LLM releases: - Gemma 4 - GLM-5.1 - Qwen3.6 - Kimi K2.6 - DeepSeek V4 All are now added to the LLM Archi…

DGX agent

April saw significant activity in large language model releases, with five major models introduced including Gemma 4, GLM-5.1, Qwen 3.6, Kimi K2.6, and DeepSeek V4. These releases have been added to a

model-releasessebastian-raschka--x
26 Apr 2026
Model Releases

@badlogicgames Yup: https://x.com/antirez/status/2048344770234249588 The GGUF tool calling template is wrong but I'm uploading a new GGUF fi…

DGX agent

@badlogicgames Yup: https://x.com/antirez/status/2048344770234249588 The GGUF tool calling template is wrong but I'm uploading a new GGUF file. Otherwise there is the right template file in one of the

model-releasesclem-delangue--x
26 Apr 2026
Model Releases

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate'…

DGX agent

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate' and have that thing be infinitely scalable. See below image

model-releasesallie-k--miller--x
26 Apr 2026
← Previous
1…390391392393394…470
Next →