AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,292 results
5 Aug 2026

Anthropic confirms it is building an in-house silicon team to design custom chips for Claude, co-designing hardware and models and using a 'multi-chip approach' (Tom Carter/Business Insider)

Model ReleasesDGX agent

Tom Carter / Business Insider: Anthropic confirms it is building an in-house silicon team to design custom chips for Claude, co-designing hardware and models and using a “multi-chip approach” — - Anth

Anyone interested in building a harness-only benchmark?

Model ReleasesDGX agent

There are a lot of LLM benchmarks but few, if any, harness benchmarks. I am thinking this would be a really good community project to build one. End goal: a leaderboard of harness performance (multipl

Approximate Speculative Decoding

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.03447v1 Announce Type: cross Abstract: Speculative decoding accelerates autoregressive generation by verifying a draft block with a target model in parallel. Under standard greedy verificat

ArtECulture: Benchmarking Culture-Conditioned Visual Emotion Understanding in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2608.03358v1 Announce Type: new Abstract: Existing visual emotion understanding methods typically ignore cultural variations in emotional perception. We introduce culture-conditioned visual emot

AS-FedBridge: Pseudo-Spike Bridge Distillation for Heterogeneous ANN-SNN Federated Learning

Model ReleasesDGX agent

arXiv:2608.03324v1 Announce Type: new Abstract: Federated learning enables collaborative model training across distributed edge devices while strictly preserving data privacy. To facilitate practical

Assessment of Conditional Diffusion Model for Synthetic Histopathology Image Generation

Model ReleasesDGX agent

arXiv:2608.03990v1 Announce Type: new Abstract: Synthetic histopathology image generation has emerged as an approach that may address data scarcity in computational pathology, yet current evaluation m

ATFlash: Per-RoPE-Wavelength Attention Windows for Compute/Memory-Efficient LLM Inference

Model ReleasesDGX agent

arXiv:2608.02947v1 Announce Type: cross Abstract: The attention score with rotary position embeddings (RoPE) decomposes exactly into a sum over its 2D-rotation frequency pairs, and each pair's wavelen

Automated Visualization Code Synthesis via Multi-Path Reasoning and Feedback-Driven Optimization

Model ReleasesDGX agent

arXiv:2502.11140v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become a cornerstone for automated visualization code generation, enabling users to create charts through na

AWS partners with Anthropic and OpenAI to bring Continuum into coding tools

Model ReleasesDGX agent

Amazon Web Services Inc. today said it has partnered with Anthropic PBC and OpenAI Group PBC to wire AWS Continuum for code vulnerabilities directly into the tools developers write code in. The integr

b10276

Model ReleasesDGX agent

Prefer npm ci over install for security (#26601) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramew

b10278

Model ReleasesDGX agent

build : remove GGML_METAL_USE_BF16 from all build scripts (#26604) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel

b10280

Model ReleasesDGX agent

vendor : apply patches for subprocess.h (#26606) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramew

b10282

Model ReleasesDGX agent

server: Adding spec-decode counters to /metrics endpoint (#26389) server: add spec-decode counters to /metrics endpoint server: fixed review comments and now aligned param names exactly with vLLM. Web

b10284

Model ReleasesDGX agent

fit: Fix memory allocation for MTP layers (#26605) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFram

b10285

Model ReleasesDGX agent

mtmd: support multi-row batching for deepseek-ocr (#26154) mtmd: support multi-row batching for deepseek-ocr mtmd: weave deepseek-ocr rows in one shot instead of per row (#26615) Co-authored-by: Saba

b10286

Model ReleasesDGX agent

grammar : degrade max repetition >= 2000 to unbounded (#26613) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64

b10287

Model ReleasesDGX agent

mtmd: Unlimited-OCR fix max_tiles, setting in converter (#25614) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x

b10288

Model ReleasesDGX agent

tests: re-enable MiniMax M3 in test-llama-archs (#26633) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS

b10289

Model ReleasesDGX agent

server: harden the file_glob_search directory walk (#26626) server: don't walk Windows junctions in file_glob_search std::filesystem reports a junction as a plain directory, so the symlink guard misse

Balancing Efficiency and Efficacy: Training-Free Attention-Guided Switching Between Explicit and Latent Thoughts for MLLMs

Model ReleasesDGX agent

arXiv:2608.03450v1 Announce Type: cross Abstract: Reasoning in Multimodal Large Language Models (MLLMs) requires both fine-grained visual perception and rigorous logical deduction. Explicit text-based

BanglaWild: An In-the-Wild Bengali Scene Text Recognition Benchmark for OCR and Vision-Language Models

Model ReleasesDGX agent

arXiv:2608.03884v1 Announce Type: cross Abstract: In-the-wild Bengali scene text recognition is largely unmeasured: existing resources target handwritten documents or constrained sign-board parsing, r

BBOWP-Bench: Evaluating LLMs on Black-Box Optimization Word Problems

Model ReleasesDGX agent

arXiv:2608.02612v1 Announce Type: new Abstract: Formulating an optimization problem strongly affects the quality of the final solution, yet good formulations usually require substantial expertise. Rec

Beyond Either-Or Reasoning: Transduction and Induction as Cooperative Problem-Solving Paradigms

Model ReleasesDGX agent

arXiv:2505.14744v3 Announce Type: replace-cross Abstract: Traditionally, in Programming-by-example (PBE) the goal is to synthesize a program from a small set of input-output examples. Lately, PBE has

Beyond Initialization Loss: A Systematic Study of Token Embedding Initialization Strategies for LLM Vocabulary Extension

Model ReleasesDGX agent

arXiv:2608.03494v1 Announce Type: new Abstract: Vocabulary extension is an efficient way to adapt pretrained large language models (LLMs) to new languages, but the initialization of newly added token

Beyond Representational Similarity: Source-Conditioned Description-Length Gain for Generative Plagiarism Detection and Candidate Source Reranking

Model ReleasesDGX agent

arXiv:2608.03859v1 Announce Type: cross Abstract: Large language models (LLMs) pose challenges to academic integrity and peer review. Yet generative plagiarism detection remains an underexplored and l

Beyond Simulations: What 20,000 Real Conversations Reveal About Mental Health AI Safety

Model ReleasesDGX agent

arXiv:2601.17003v2 Announce Type: replace-cross Abstract: Mental-health AI safety is typically evaluated with small, simulation-based benchmarks that may not reflect the linguistic and contextual dive

Beyond the Gegenbauer Paradigm: q-Orthogonal Kernels for Machine Learning

Model ReleasesDGX agent

arXiv:2608.03482v1 Announce Type: new Abstract: The performance of Support Vector Machines (SVMs) critically depends on the kernel function choice, which enables implicit mapping of data into high-dim

Beyond the Single Camera: Agentic Multi-View Reasoning in Sports Video Understanding

Model ReleasesDGX agent

arXiv:2607.11844v2 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) achieve strong performance on single-view video understanding benchmarks. However, sports videos inv

Bridging Online and Offline Handwriting via Differentiable Physical Rendering

Model ReleasesDGX agent

arXiv:2608.03198v1 Announce Type: new Abstract: Realistic handwritten text generation plays an important role in numerous applications, such as font design, biometric authentication, and robotic calli

Bro, I need you to FOCUS.

Model ReleasesDGX agent

Come on guys, we could get deepseek cheaper with more usage directly with them if we wanted. WHERE IS KIMI K3 WITHOUT EXTRA USAGE NEEDED !!? This is pissing me off like crazy submitted by /u/Other_Che

Building a Fully Local PDF Read-Aloud & PDF-to-Audiobook Desktop App with Kokoro 82M, Qwen, and llama.cpp

Model ReleasesDGX agent

Hey everyone, I’ve been building Speechfony - a desktop app for reading PDFs (and EPUBs) with offline text-to-speech. Open a document, listen sentence-by-sentence with highlighting, or export selected

BulkPR-Bench: Benchmarking Queue-Level Governance of Interacting Pull Requests

Model ReleasesDGX agent

arXiv:2608.02685v1 Announce Type: cross Abstract: Coding-agent benchmarks increasingly cover long-horizon, end-to-end, and interactive development, but typically retain one requested outcome or a fixe

CADET: Physics-Grounded Causal Auditing and Training-Free Deconfounding of End-to-End Driving Planners

Model ReleasesDGX agent

arXiv:2606.14438v3 Announce Type: replace-cross Abstract: End-to-end (E2E) autonomous-driving planners trained by imitation are prone to statistical shortcuts: they associate scene elements that merel

Can Large Language Models Recover Semantic Optimization Opportunities That Compilers Miss?

Model ReleasesDGX agent

arXiv:2608.03983v1 Announce Type: cross Abstract: Optimizing compilers miss profitable transformations when their enabling semantics are absent from the analyzed program representation. We ask whether

Can LLM design high-quality experiments? A Comprehensive and Systematic Benchmark on Autonomous Experimental Design

Model ReleasesDGX agent

arXiv:2608.03501v1 Announce Type: new Abstract: AI for Research (AI4Research) leverages AI to automate and improve scientific workflows. While experimental design is a critical stage of the research p

Can LLMs Test Terminal User Interfaces?

Model ReleasesDGX agent

arXiv:2608.03743v1 Announce Type: cross Abstract: Terminal User Interfaces (TUIs) combine the stateful, screen-oriented behaviour of GUIs with terminal deployment and are now common in developer tools

Can Text-to-Image Models Draw from the Right Frame of Reference?

Model ReleasesDGX agent

arXiv:2608.03357v1 Announce Type: new Abstract: Spatial instruction following has become a crucial requirement for text-to-image (T2I) generation. A common challenge arises when directional expression

CARE-Bench: Benchmarking Patient-Facing LLM Triage

Model ReleasesDGX agent

arXiv:2608.03731v1 Announce Type: new Abstract: Patient-facing medical LLMs and agents increasingly answer symptom questions before clinician contact, where the key safety question is what action the

ChartAnno: Evaluating MLLMs for Chart Annotation Generation

Model ReleasesDGX agent

arXiv:2608.03464v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have made significant progress in chart understanding, generation, and editing, but their ability to annotate e

ChiEngMixBench: Evaluating Large Language Models on Expert-Style Chinese-English Terminology Mixing

Model ReleasesDGX agent

arXiv:2601.16217v2 Announce Type: replace-cross Abstract: Large language models increasingly mediate multilingual professional communication, where useful generation requires adapting to community con

Cloudflare launches Cloudflare OS: an open-source AI agentic workspace for the enterprise

Model ReleasesDGX agent

Cloudflare Inc. today announced the launch of Cloudflare OS, an open-source artificial intelligence agentic workspace available through the browser, filled with custom shared micro-applications for en

Cloudflare launches Identity-Aware AI Gateway to track who is using AI

Model ReleasesDGX agent

Cloudflare Inc. today launched Identity-Aware AI Gateway, a service that attaches a verified identity to every artificial intelligence request leaving a company network. Information technology and sec

Cloudflare open sources a new version of Cloudflare OS, a browser-accessible AI agentic workspace for enterprises that lets employees build custom micro-apps (Kyt Dotson/SiliconANGLE)

Model ReleasesDGX agent

Kyt Dotson / SiliconANGLE: Cloudflare open sources a new version of Cloudflare OS, a browser-accessible AI agentic workspace for enterprises that lets employees build custom micro-apps — Cloudflare In

Confident but Unreliable: A Behavioral Safety Audit of Vision-Language Models on Brain MRI

Model ReleasesDGX agent

arXiv:2608.02790v1 Announce Type: new Abstract: Vision-language models (VLMs), including medical specialists, are increasingly proposed for medical imaging, yet their stated confidence is rarely evalu

ConlangBench: Exploring Language Knowledge and Learning in LLMs through Diverse Constructed Languages

Model ReleasesDGX agent

arXiv:2608.03505v1 Announce Type: new Abstract: Constructed languages (conlangs) are intentionally created human languages with a rich tradition of linguistic creativity. Despite their potential for s

CorePath: A Breast-Specialized Pathology Foundation Model for Core Needle Biopsy Diagnosis and Risk-Controlled Report Generation

Model ReleasesDGX agent

arXiv:2608.03079v1 Announce Type: cross Abstract: Breast core needle biopsy (CNB) is central to breast cancer diagnosis yet remains challenging because limited tissue sampling, lesion heterogeneity, a

Cross-Lingual Bias in Large Language Models: A Comparative Analysis of English and Swahili

Model ReleasesDGX agent

arXiv:2608.03532v1 Announce Type: new Abstract: Large language models are increasingly deployed in multilingual contexts, yet safety alignment and bias evaluation remain overwhelmingly English-centric

CrossScope: A Role-Asymmetric World Model for Joint Dual-Scope Surgical Video Prediction

Model ReleasesDGX agent

arXiv:2608.03211v1 Announce Type: new Abstract: Visual world models typically learn future dynamics from a single observation stream, limiting their ability to model cooperative systems with multiple

CUADebug: Diagnosing and Repairing Computer-Use Agent Failures

Model ReleasesDGX agent

arXiv:2608.02643v1 Announce Type: cross Abstract: Computer-use agents (CUAs) operate real desktop and web interfaces through screenshots, mouse and keyboard actions, and stateful UI feedback, yet thei

CUDA MPC: A GPU-Native Solver for Model Predictive Control

Model ReleasesDGX agent

arXiv:2608.03051v1 Announce Type: new Abstract: Model Predictive Control (MPC) delivers constraint-aware control, but its reliance on online optimization limits its use on systems with fast dynamics,

Cura 1T: Specialized Model for Agentic Healthcare

Model ReleasesDGX agent

arXiv:2607.15314v2 Announce Type: replace Abstract: Healthcare AI agents handle patient consultation, clinical reasoning over text and images, interactive diagnosis, and electronic health record (EHR)

DAIF: A Data-Driven Intermediate Fusion Framework for Multimodal Supervised Learning via Approximate Message Passing

Model ReleasesDGX agent

arXiv:2608.02769v1 Announce Type: cross Abstract: Multimodal supervised learning seeks to leverage multiple heterogeneous data sources to improve predictive performance. A central challenge is determi

DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces

Model ReleasesDGX agent

arXiv:2608.03451v1 Announce Type: new Abstract: Data agents enable natural-language analytics over organizational workspaces, where relevant evidence may be scattered across databases, structured file

DeepSeek V4 Flash 0731 at 10–17 t/s (nothink) on MacBook M5 Pro **64GB***, partly via SSD streaming

Model ReleasesDGX agent

Inspired by a post from u/giveen I motivated claude (no patinence on my side to work through everything myself) to help me get DS running on my MacBook M5 Pro 64GB and it exceeded my expectations.. be

DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO fro…

Model ReleasesDGX agent

DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO from the managed UI. Built for coding agents and high-volume pr

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available …

Model ReleasesDGX agent

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available in Pi across a range of providers. If you’ve never tried an

Deepseek V4 Flash just hit Colibri, does anyone have numbers?

Model ReleasesDGX agent

I'm mosty interested in 128-192GB VRAM with 128-256GB RAM to spare, so SSD streaming is basically not even necessary. Seems only FP4 is supported, so older hardware will likely be slow - no Unsloth GG

DeepSeek-V4-Flash on SM89 4x48gb 4090s with DSpark

Model ReleasesDGX agent

https://github.com/yhfgyyf/vllm-deepseek-v4-sm89 I couldn't believe that someone actually got vLLM working with this particular set of GPUs, but here it is. The video is from right after I got it work

DenialRAG: Single-Document RAG Poisoning via Embedded Parametric Denial

Model ReleasesDGX agent

arXiv:2608.02678v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are vulnerable to corpus poisoning: an attacker who inserts a crafted document into the retrieval corpus

Detecting Hallucinations and Recovering Verified Answers in Arabic Islamic Question Answering

Model ReleasesDGX agent

arXiv:2608.03720v1 Announce Type: new Abstract: Large language models can generate fluent responses to Islamic questions while introducing factual errors that are difficult to identify. This paper pre

← Previous
1…2829303132…372
Next →