AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
All
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,499 results
Model Releases

DeepSeek-V4: a million-token context that agents can actually use

DGX agent

DeepSeek-V4 is an advanced language model featuring a million-token context window that enables practical agentic applications beyond simple retrieval. The model demonstrates improved efficiency and u

model-releaseshugging-face
24 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

DeepSeek V4 by @deepseek_ai just dropped! SGLang is ready on Day 0 with a full stack of optimizations from architectures to low-level kernel…

DGX agent

DeepSeek V4 by @deepseek_ai just dropped! SGLang is ready on Day 0 with a full stack of optimizations from architectures to low-level kernels. We also deliver a verified RL training pipeline in Miles

model-releasesdylan-patel--x
24 Apr 2026
Model Releases

deepseek-v4-flash is now available on Ollama's cloud! Hosted in the US. Try it with Claude Code: ollama launch claude --model deepseek-v4-fl…

DGX agent

deepseek-v4-flash is now available on Ollama's cloud! Hosted in the US. Try it with Claude Code: ollama launch claude --model deepseek-v4-flash:cloud Try it with OpenClaw: ollama launch openclaw --mod

model-releasesollama--x
24 Apr 2026
Model Releases

DEEPSEEK-V4 IS RELEASED

DGX agent

DeepSeek-V4 is a newly released AI model announced by Clem Delangue on X (formerly Twitter). The release likely represents an updated version of the DeepSeek model series with improvements in capabili

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

DeepSeek v4 just dropped

DGX agent

DeepSeek has released v4, its latest model iteration. The announcement was made by Clem Delangue on X (formerly Twitter). This likely represents a significant update to DeepSeek's AI capabilities, tho

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

DeepSeek-V4 just dropped on Hugging Face https://huggingface.co/collections/deepseek-ai/deepseek-v4

DGX agent

DeepSeek-V4, a new model release from DeepSeek AI, has been made available on Hugging Face's model hub. The release was announced by Clem Delangue and includes model weights and resources accessible t

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T t…

DGX agent

🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T total / 49B active params. Performance rivaling the world's top

model-releasesjeremy-howard--x
24 Apr 2026
Model Releases

Deepseek v4 Pro

DGX agent

DeepSeek-V4-Pro is a Mixture-of-Experts language model with 1.6 trillion total parameters and 49 billion activated per token, supporting a 1 million token context length. Released under the MIT Licens

model-releasesr-ollama
24 Apr 2026
Model Releases

DeepSeek V4 Pro costs 1.74/1M input tokens and 3.48/1M output tokens, while V4 Flash costs 0.14/1M and 0.28/1M; both models are the cheapest in their class (Simon Willison/Simon Willison's Weblog)

DGX agent

Simon Willison / Simon Willison's Weblog: DeepSeek V4 Pro costs 1.74/1M input tokens and 3.48/1M output tokens, while V4 Flash costs 0.14/1M and 0.28/1M; both models are the cheapest in their class —

model-releasestechmeme
24 Apr 2026
Model Releases

DeepSeek V4 Pro has 1.6T total parameters, its largest model by the metric, and V4 Flash has 284B parameters; both models have a context window of 1M tokens (Vincent Chow/South China Morning Post)

DGX agent

Vincent Chow / South China Morning Post: DeepSeek V4 Pro has 1.6T total parameters, its largest model by the metric, and V4 Flash has 284B parameters; both models have a context window of 1M tokens —

model-releasestechmeme
24 Apr 2026
Model Releases

DeepSeek V4 Pro is now available on Together AI. DeepSeek V4 Flash coming soon. Try it now: http://www.together.ai/models/deepseek-v4-pro#

DGX agent

DeepSeek V4 Pro is now available through Together AI's model platform, with the faster DeepSeek V4 Flash variant expected to launch soon. Together AI is offering users the ability to access and test D

model-releasestogether-ai--x
24 Apr 2026
Model Releases

DenoiseRank: Learning to Rank by Diffusion Models

DGX agent

arXiv:2604.20852v1 Announce Type: cross Abstract: Learning to rank (LTR) is one of the core tasks in Machine Learning. Traditional LTR models have made great progress, but nearly all of them are imple

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles

DGX agent

arXiv:2604.21152v1 Announce Type: cross Abstract: As state-of-the-art Large Language Models (LLMs) have become ubiquitous, ensuring equitable performance across diverse demographics is critical. Howev

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Differentially Private Model Merging

DGX agent

arXiv:2604.20985v1 Announce Type: cross Abstract: In machine learning applications, privacy requirements during inference or deployment time could change constantly due to varying policies, regulation

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Diplomatic cable: US State Department has ordered a global push to bring attention to what it says are efforts by Chinese companies to steal IP from US AI labs (Raphael Satter/Reuters)

DGX agent

Raphael Satter / Reuters: Diplomatic cable: US State Department has ordered a global push to bring attention to what it says are efforts by Chinese companies to steal IP from US AI labs — The U.S. Sta

model-releasestechmeme
24 Apr 2026
Model Releases

Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models

DGX agent

arXiv:2507.04023v3 Announce Type: replace Abstract: Large language models (LLMs) achieve impressive performance on complex mathematical benchmarks yet sometimes fail on basic math reasoning while gene

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision

DGX agent

arXiv:2604.21461v1 Announce Type: new Abstract: Egocentric AI agents, such as smart glasses, rely on pointing gestures to resolve referential ambiguities in natural language commands. However, despite

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Domain-Aware Hierarchical Contrastive Learning for Semi-Supervised Generalization Fault Diagnosis

DGX agent

arXiv:2604.20928v1 Announce Type: cross Abstract: Fault diagnosis under unseen operating conditions remains highly challenging when labeled data are scarce. Semi-supervised domain generalization fault

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Dr. Assistant: Enhancing Clinical Diagnostic Inquiry via Structured Diagnostic Reasoning Data and Reinforcement Learning

DGX agent

arXiv:2601.13690v2 Announce Type: replace Abstract: Clinical Decision Support Systems (CDSSs) provide reasoning and inquiry guidance for physicians, yet they face notable challenges, including high ma

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Droplet-LNO: Physics-Informed Laplace Neural Operators for Accurate Prediction of Droplet Spreading Dynamics on Complex Surfaces

DGX agent

arXiv:2604.20993v1 Announce Type: new Abstract: Spreading of liquid droplets on solid substrates constitutes a classic multiphysics problem with widespread applications ranging from inkjet printing, s

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Drug Synergy Prediction via Residual Graph Isomorphism Networks and Attention Mechanisms

DGX agent

arXiv:2604.21473v1 Announce Type: cross Abstract: In the treatment of complex diseases, treatment regimens using a single drug often yield limited efficacy and can lead to drug resistance. In contrast

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

EARL-BO: Reinforcement Learning for Multi-Step Lookahead, High-Dimensional Bayesian Optimization

DGX agent

arXiv:2411.00171v2 Announce Type: replace Abstract: To avoid myopic behavior, multi-step lookahead Bayesian optimization (BO) algorithms consider the sequential nature of BO and have demonstrated prom

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Efficient Multi-Source Knowledge Transfer by Model Merging

DGX agent

arXiv:2508.19353v2 Announce Type: replace-cross Abstract: While transfer learning is an effective strategy, it often overlooks the opportunity to leverage knowledge from numerous available models onli

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Empirical Comparison of Agent Communication Protocols for Task Orchestration

DGX agent

arXiv:2603.22823v3 Announce Type: replace Abstract: Context. The problem of comparative evaluation of communication protocols for task orchestration by large language model (LLM) agents is considered.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

EngramaBench: Evaluating Long-Term Conversational Memory with Structured Graph Retrieval

DGX agent

arXiv:2604.21229v1 Announce Type: cross Abstract: Large language model assistants are increasingly expected to retain and reason over information accumulated across many sessions. We introduce Engrama

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Enhancing Science Classroom Discourse Analysis through Joint Multi-Task Learning for Reasoning-Component Classification

DGX agent

arXiv:2604.21137v1 Announce Type: cross Abstract: Analyzing the reasoning patterns of students in science classrooms is critical for understanding knowledge construction mechanism and improving instru

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Evaluating AI Meeting Summaries with a Reusable Cross-Domain Pipeline

DGX agent

arXiv:2604.21345v1 Announce Type: new Abstract: We present a reusable evaluation pipeline for generative AI applications, instantiated for AI meeting summaries and released with a public artifact pack

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents

DGX agent

arXiv:2604.21890v1 Announce Type: new Abstract: Event extraction identifies the central aspects of events from text. It supports event understanding and analysis, which is crucial for tasks such as in

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

FairyFuse: Multiplication-Free LLM Inference on CPUs via Fused Ternary Kernels

DGX agent

arXiv:2604.20913v1 Announce Type: new Abstract: Large language models are increasingly deployed on CPU-only platforms where memory bandwidth is the primary bottleneck for autoregressive generation. We

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Fake or Real, Can Robots Tell? Evaluating VLM Robustness to Domain Shift in Single-View Robotic Scene Understanding

DGX agent

arXiv:2506.19579v3 Announce Type: replace-cross Abstract: Robotic scene understanding increasingly relies on Vision-Language Models (VLMs) to generate natural language descriptions of the environment.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Feature request for @huggingface - add a repository size option to the sort menu, I want to see the DeepSeek quantized models that take up t…

DGX agent

Simon Willison requested that Hugging Face add a repository size sorting option to help users find and filter models by storage requirements, specifically mentioning interest in locating DeepSeek quan

model-releasessimon-willison--x
24 Apr 2026
Model Releases

Federated Co-tuning Framework for Large and Small Language Models

DGX agent

arXiv:2411.11707v3 Announce Type: replace-cross Abstract: By adapting Large Language Models (LLMs) to domain-specific tasks or enriching them with domain-specific knowledge, we can fully harness the c

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Federated Learning for Surgical Vision in Appendicitis Classification: Results of the FedSurg EndoVis 2024 Challenge

DGX agent

arXiv:2510.04772v2 Announce Type: replace-cross Abstract: Developing generalizable surgical AI requires multi-institutional data, yet patient privacy constraints preclude direct data sharing, making F

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Fine-Tuning Regimes Define Distinct Continual Learning Problems

DGX agent

arXiv:2604.21927v1 Announce Type: new Abstract: Continual learning (CL) studies how models acquire tasks sequentially while retaining previously learned knowledge. Despite substantial progress in benc

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Flow Matching for Conditional MRI-CT and CBCT-CT Image Synthesis

DGX agent

arXiv:2510.04823v2 Announce Type: replace Abstract: Generating synthetic CT (sCT) from MRI or CBCT plays a crucial role in enabling MRI-only and CBCT-based adaptive radiotherapy, improving treatment p

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some…

DGX agent

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some insane ones I saw: 1. A new AI paradigm from scratch — tryi

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

for the record, i put 'undecided' as the last choice, not the second choice, @nikitabier sorry i know this is minor but it should be a 2 sec…

DGX agent

for the record, i put 'undecided' as the last choice, not the second choice, @nikitabier sorry i know this is minor but it should be a 2 second fix in cursor https://x.com/kr0der/status/20477027526678

model-releasesswyx--x
24 Apr 2026
Model Releases

From Codebooks to VLMs: Evaluating Automated Visual Discourse Analysis for Climate Change on Social Media

DGX agent

arXiv:2604.21786v1 Announce Type: new Abstract: Social media platforms have become primary arenas for climate communication, generating millions of images and posts that - if systematically analysed -

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

From Research Question to Scientific Workflow: Leveraging Agentic AI for Science Automation

DGX agent

arXiv:2604.21910v1 Announce Type: new Abstract: Scientific workflow systems automate execution -- scheduling, fault tolerance, resource management -- but not the semantic translation that precedes it.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure

DGX agent

arXiv:2512.22274v2 Announce Type: replace Abstract: We introduce GeCo, a geometry-grounded metric for jointly detecting geometric deformation and occlusion-inconsistency artifacts in static scenes. By

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning

DGX agent

arXiv:2604.21495v1 Announce Type: cross Abstract: Numerical reasoning over expert-domain tables often exhibits high in-domain accuracy but limited robustness to domain shift. Models trained with super

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Geo-R1: Improving Few-Shot Geospatial Referring Expression Understanding with Reinforcement Fine-Tuning

DGX agent

arXiv:2509.21976v3 Announce Type: replace-cross Abstract: Referring expression understanding in remote sensing poses unique challenges, as it requires reasoning over complex object-context relationshi

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Geometric Characterisation and Structured Trajectory Surrogates for Clinical Dataset Condensation

DGX agent

arXiv:2604.21638v1 Announce Type: new Abstract: Dataset condensation constructs compact synthetic datasets that retain the training utility of large real-world datasets, enabling efficient model devel

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Geometric Monomial (GEM): a family of rational 2N-differentiable activation functions

DGX agent

arXiv:2604.21677v1 Announce Type: cross Abstract: The choice of activation function plays a crucial role in the optimization and performance of deep neural networks. While the Rectified Linear Unit (R

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GeoMind: An Agentic Workflow for Lithology Classification with Reasoned Tool Invocation

DGX agent

arXiv:2604.21501v1 Announce Type: new Abstract: Lithology classification in well logs is a fundamental geoscience data mining task that aims to infer rock types from multi dimensional geophysical sequ

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR

DGX agent

arXiv:2601.09361v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a key paradigm for improving large-scale reasoning models. Unlike supervised fine-tun

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GerAV: Towards New Heights in German Authorship Verification using Fine-Tuned LLMs on a New Benchmark

DGX agent

arXiv:2601.13711v2 Announce Type: replace Abstract: Authorship verification (AV) is the task of determining whether two texts were written by the same author and has been studied extensively, predomin

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

GiVA: Gradient-Informed Bases for Vector-Based Adaptation

DGX agent

arXiv:2604.21901v1 Announce Type: cross Abstract: As model sizes continue to grow, parameter-efficient fine-tuning has emerged as a powerful alternative to full fine-tuning. While LoRA is widely adopt

model-releasesarxiv-cs-ai
24 Apr 2026
← Previous
1…392393394395396…469
Next →