AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
Model Releases

Cloud-hosted models will become more niche/seldom-used as much cheaper local models fulfill that which 90% of consumers need from AI. Lots o…

DGX agent

Cloud-hosted models will become more niche/seldom-used as much cheaper local models fulfill that which 90% of consumers need from AI. Lots of models to buy/use, but very few hardware options upon whic

model-releasesclem-delangue--x
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CLT-Optimal Parameter Error Bounds for Linear System Identification

DGX agent

arXiv:2604.21270v1 Announce Type: cross Abstract: There has been remarkable progress over the past decade in establishing finite-sample, non-asymptotic bounds on recovering unknown system parameters f

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Cognitive Amplification vs Cognitive Delegation in Human-AI Systems: A Metric Framework

DGX agent

arXiv:2603.18677v2 Announce Type: replace-cross Abstract: Artificial intelligence is increasingly embedded in human decision making. In some cases, it enhances human reasoning. In others, it fosters e

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

come to the most ai pilled country per capita @anthropicai we want u

DGX agent

This post from Swyx appears to be a humorous or provocative invitation directed at Anthropic AI, suggesting a particular country has the highest per-capita adoption or focus on AI technology and encou

model-releasesswyx--x
24 Apr 2026
Model Releases

Compliance Moral Hazard and the Backfiring Mandate

DGX agent

arXiv:2604.21789v1 Announce Type: cross Abstract: Competing firms that serve shared customer populations face a fundamental information aggregation problem: each firm holds fragmented signals about ri

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Concurrence: A dependence criterion for time series, applied to biological data

DGX agent

arXiv:2512.16001v2 Announce Type: replace-cross Abstract: Measuring the statistical dependence between observed signals is a primary tool for scientific discovery. However, biological systems often ex

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Congrats to @deepseek_ai team! Doing the numbers I would estimate: Pro < 14m for the final training run Flash < 4m Ratio of active params …

DGX agent

Congrats to @deepseek_ai team! Doing the numbers I would estimate: Pro < 14m for the final training run Flash < 4m Ratio of active params x total training tokens vs v3 Total compute costs (data prep,

model-releasesemad-mostaque--x
24 Apr 2026
Model Releases

Context Is What You Need: The Maximum Effective Context Window for Real World Limits of LLMs

DGX agent

arXiv:2509.21361v2 Announce Type: replace-cross Abstract: Large language model (LLM) providers boast big numbers for maximum context window sizes. To test the real world use of context windows, we 1)

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors

DGX agent

arXiv:2604.21241v1 Announce Type: cross Abstract: Vision--Language--Action (VLA) models often use intermediate representations to connect multimodal inputs with continuous control, yet spatial guidanc

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Counterfactual Segmentation Reasoning: Diagnosing and Mitigating Pixel-Grounding Hallucination

DGX agent

arXiv:2506.21546v4 Announce Type: replace-cross Abstract: Segmentation Vision-Language Models (VLMs) have significantly advanced grounded visual understanding, yet they remain prone to pixel-grounding

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Cross-Entropy Is Load-Bearing: A Pre-Registered Scope Test of the K-Way Energy Probe on Bidirectional Predictive Coding

DGX agent

arXiv:2604.21286v1 Announce Type: cross Abstract: Cacioli (2026) showed that the K-way energy probe on standard discriminative predictive coding networks reduces approximately to a monotone function o

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Cross-Session Threats in AI Agents: Benchmark, Evaluation, and Algorithms

DGX agent

arXiv:2604.21131v1 Announce Type: cross Abstract: AI-agent guardrails are memoryless: each message is judged in isolation, so an adversary who spreads a single attack across dozens of sessions slips p

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

CSC: Turning the Adversary's Poison against Itself

DGX agent

arXiv:2604.21416v1 Announce Type: cross Abstract: Poisoning-based backdoor attacks pose significant threats to deep neural networks by embedding triggers in training data, causing models to misclassif

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Data-Driven Open-Loop Simulation for Digital-Twin Operator Decision Support in Wastewater Treatment

DGX agent

arXiv:2604.20935v1 Announce Type: cross Abstract: Wastewater treatment plants (WWTPs) need digital-twin-style decision support tools that can simulate plant response under prescribed control plans, to

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

DAVIS: OOD Detection via Dominant Activations and Variance for Increased Separation

DGX agent

arXiv:2601.22703v2 Announce Type: replace Abstract: Detecting out-of-distribution (OOD) inputs is a critical safeguard for deploying machine learning models in the real world. However, most post-hoc d

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

🎉 Day-0 support for @deepseek_ai V4 Pro and Flash on vLLM — a new generation of DeepSeek model, purpose-built for tasks up to 1M tokens. Al…

DGX agent

🎉 Day-0 support for @deepseek_ai V4 Pro and Flash on vLLM — a new generation of DeepSeek model, purpose-built for tasks up to 1M tokens. Alongside the release, we're publishing a first-principles walk

model-releasesdylan-patel--x
24 Apr 2026
Model Releases

Decoupled DiLoCo for Resilient Distributed Pre-training

DGX agent

arXiv:2604.21428v1 Announce Type: new Abstract: Modern large-scale language model pre-training relies heavily on the single program multiple data (SPMD) paradigm, which requires tight coupling across

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Deep FinResearch Bench: Evaluating AI's Ability to Conduct Professional Financial Investment Research

DGX agent

arXiv:2604.21006v1 Announce Type: new Abstract: We introduce Deep FinResearch Bench, a practical and comprehensive evaluation framework for deep research (DR) agents in financial investment research.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

DeepSeek open-sources V4 large language model series

DGX agent

Chinese artificial intelligence developer DeepSeek today released a new series of open-source large language models. V4, as the algorithm family is called, comprises two LLMs on launch. There’s the fl

model-releasessiliconangle
24 Apr 2026
Model Releases

DeepSeek releases its new flagship models V4 Pro and V4 Flash in preview, saying V4 Pro trails the performance of state-of-the-art models by about 3 to 6 months (Bloomberg)

DGX agent

Bloomberg: DeepSeek releases its new flagship models V4 Pro and V4 Flash in preview, saying V4 Pro trails the performance of state-of-the-art models by about 3 to 6 months — DeepSeek rolled out previe

model-releasestechmeme
24 Apr 2026
Model Releases

DeepSeek V4 - almost on the frontier, a fraction of the price

DGX agent

Chinese AI lab DeepSeek's last model release was V3.2 (and V3.2 Speciale) last December. They just dropped the first of their hotly anticipated V4 series in the shape of two preview models, DeepSeek-V

model-releasessimon-willison
24 Apr 2026
Model Releases

DeepSeek-V4: a million-token context that agents can actually use

DGX agent

DeepSeek-V4 is an advanced language model featuring a million-token context window that enables practical agentic applications beyond simple retrieval. The model demonstrates improved efficiency and u

model-releaseshugging-face
24 Apr 2026
Model Releases

DeepSeek V4 by @deepseek_ai just dropped! SGLang is ready on Day 0 with a full stack of optimizations from architectures to low-level kernel…

DGX agent

DeepSeek V4 by @deepseek_ai just dropped! SGLang is ready on Day 0 with a full stack of optimizations from architectures to low-level kernels. We also deliver a verified RL training pipeline in Miles

model-releasesdylan-patel--x
24 Apr 2026
Model Releases

deepseek-v4-flash is now available on Ollama's cloud! Hosted in the US. Try it with Claude Code: ollama launch claude --model deepseek-v4-fl…

DGX agent

deepseek-v4-flash is now available on Ollama's cloud! Hosted in the US. Try it with Claude Code: ollama launch claude --model deepseek-v4-flash:cloud Try it with OpenClaw: ollama launch openclaw --mod

model-releasesollama--x
24 Apr 2026
Model Releases

DEEPSEEK-V4 IS RELEASED

DGX agent

DeepSeek-V4 is a newly released AI model announced by Clem Delangue on X (formerly Twitter). The release likely represents an updated version of the DeepSeek model series with improvements in capabili

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

DeepSeek v4 just dropped

DGX agent

DeepSeek has released v4, its latest model iteration. The announcement was made by Clem Delangue on X (formerly Twitter). This likely represents a significant update to DeepSeek's AI capabilities, tho

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

DeepSeek-V4 just dropped on Hugging Face https://huggingface.co/collections/deepseek-ai/deepseek-v4

DGX agent

DeepSeek-V4, a new model release from DeepSeek AI, has been made available on Hugging Face's model hub. The release was announced by Clem Delangue and includes model weights and resources accessible t

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T t…

DGX agent

🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T total / 49B active params. Performance rivaling the world's top

model-releasesjeremy-howard--x
24 Apr 2026
Model Releases

Deepseek v4 Pro

DGX agent

DeepSeek-V4-Pro is a Mixture-of-Experts language model with 1.6 trillion total parameters and 49 billion activated per token, supporting a 1 million token context length. Released under the MIT Licens

model-releasesr-ollama
24 Apr 2026
Model Releases

DeepSeek V4 Pro costs 1.74/1M input tokens and 3.48/1M output tokens, while V4 Flash costs 0.14/1M and 0.28/1M; both models are the cheapest in their class (Simon Willison/Simon Willison's Weblog)

DGX agent

Simon Willison / Simon Willison's Weblog: DeepSeek V4 Pro costs 1.74/1M input tokens and 3.48/1M output tokens, while V4 Flash costs 0.14/1M and 0.28/1M; both models are the cheapest in their class —

model-releasestechmeme
24 Apr 2026
Model Releases

DeepSeek V4 Pro has 1.6T total parameters, its largest model by the metric, and V4 Flash has 284B parameters; both models have a context window of 1M tokens (Vincent Chow/South China Morning Post)

DGX agent

Vincent Chow / South China Morning Post: DeepSeek V4 Pro has 1.6T total parameters, its largest model by the metric, and V4 Flash has 284B parameters; both models have a context window of 1M tokens —

model-releasestechmeme
24 Apr 2026
Model Releases

DeepSeek V4 Pro is now available on Together AI. DeepSeek V4 Flash coming soon. Try it now: http://www.together.ai/models/deepseek-v4-pro#

DGX agent

DeepSeek V4 Pro is now available through Together AI's model platform, with the faster DeepSeek V4 Flash variant expected to launch soon. Together AI is offering users the ability to access and test D

model-releasestogether-ai--x
24 Apr 2026
Model Releases

DenoiseRank: Learning to Rank by Diffusion Models

DGX agent

arXiv:2604.20852v1 Announce Type: cross Abstract: Learning to rank (LTR) is one of the core tasks in Machine Learning. Traditional LTR models have made great progress, but nearly all of them are imple

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles

DGX agent

arXiv:2604.21152v1 Announce Type: cross Abstract: As state-of-the-art Large Language Models (LLMs) have become ubiquitous, ensuring equitable performance across diverse demographics is critical. Howev

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Differentially Private Model Merging

DGX agent

arXiv:2604.20985v1 Announce Type: cross Abstract: In machine learning applications, privacy requirements during inference or deployment time could change constantly due to varying policies, regulation

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Diplomatic cable: US State Department has ordered a global push to bring attention to what it says are efforts by Chinese companies to steal IP from US AI labs (Raphael Satter/Reuters)

DGX agent

Raphael Satter / Reuters: Diplomatic cable: US State Department has ordered a global push to bring attention to what it says are efforts by Chinese companies to steal IP from US AI labs — The U.S. Sta

model-releasestechmeme
24 Apr 2026
Model Releases

Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models

DGX agent

arXiv:2507.04023v3 Announce Type: replace Abstract: Large language models (LLMs) achieve impressive performance on complex mathematical benchmarks yet sometimes fail on basic math reasoning while gene

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision

DGX agent

arXiv:2604.21461v1 Announce Type: new Abstract: Egocentric AI agents, such as smart glasses, rely on pointing gestures to resolve referential ambiguities in natural language commands. However, despite

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Domain-Aware Hierarchical Contrastive Learning for Semi-Supervised Generalization Fault Diagnosis

DGX agent

arXiv:2604.20928v1 Announce Type: cross Abstract: Fault diagnosis under unseen operating conditions remains highly challenging when labeled data are scarce. Semi-supervised domain generalization fault

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Dr. Assistant: Enhancing Clinical Diagnostic Inquiry via Structured Diagnostic Reasoning Data and Reinforcement Learning

DGX agent

arXiv:2601.13690v2 Announce Type: replace Abstract: Clinical Decision Support Systems (CDSSs) provide reasoning and inquiry guidance for physicians, yet they face notable challenges, including high ma

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Droplet-LNO: Physics-Informed Laplace Neural Operators for Accurate Prediction of Droplet Spreading Dynamics on Complex Surfaces

DGX agent

arXiv:2604.20993v1 Announce Type: new Abstract: Spreading of liquid droplets on solid substrates constitutes a classic multiphysics problem with widespread applications ranging from inkjet printing, s

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Drug Synergy Prediction via Residual Graph Isomorphism Networks and Attention Mechanisms

DGX agent

arXiv:2604.21473v1 Announce Type: cross Abstract: In the treatment of complex diseases, treatment regimens using a single drug often yield limited efficacy and can lead to drug resistance. In contrast

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

EARL-BO: Reinforcement Learning for Multi-Step Lookahead, High-Dimensional Bayesian Optimization

DGX agent

arXiv:2411.00171v2 Announce Type: replace Abstract: To avoid myopic behavior, multi-step lookahead Bayesian optimization (BO) algorithms consider the sequential nature of BO and have demonstrated prom

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Efficient Multi-Source Knowledge Transfer by Model Merging

DGX agent

arXiv:2508.19353v2 Announce Type: replace-cross Abstract: While transfer learning is an effective strategy, it often overlooks the opportunity to leverage knowledge from numerous available models onli

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Empirical Comparison of Agent Communication Protocols for Task Orchestration

DGX agent

arXiv:2603.22823v3 Announce Type: replace Abstract: Context. The problem of comparative evaluation of communication protocols for task orchestration by large language model (LLM) agents is considered.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

EngramaBench: Evaluating Long-Term Conversational Memory with Structured Graph Retrieval

DGX agent

arXiv:2604.21229v1 Announce Type: cross Abstract: Large language model assistants are increasingly expected to retain and reason over information accumulated across many sessions. We introduce Engrama

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Enhancing Science Classroom Discourse Analysis through Joint Multi-Task Learning for Reasoning-Component Classification

DGX agent

arXiv:2604.21137v1 Announce Type: cross Abstract: Analyzing the reasoning patterns of students in science classrooms is critical for understanding knowledge construction mechanism and improving instru

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Evaluating AI Meeting Summaries with a Reusable Cross-Domain Pipeline

DGX agent

arXiv:2604.21345v1 Announce Type: new Abstract: We present a reusable evaluation pipeline for generative AI applications, instantiated for AI meeting summaries and released with a public artifact pack

model-releasesarxiv-cs-ai
24 Apr 2026
← Previous
1…393394395396397…470
Next →