AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,332 results
Model Releases

What speeds are everyone getting with deepseek v4 flash 0731?

DGX agent

What speeds are everyone getting with deepseek v4 flash 0731? I’m getting~200 tps prompt processing / ~11 tps token gen, on 4x5060ti16gb with ddr4 3200 ram at 4-channel, via llamacpp, with context win

model-releasesr-localllama
1 Aug 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

What's currently the 'smartest' LLM to use on 8GB vram and 16 RAM and same thing for 8 VRAM and 64 RAM?

DGX agent

Been trying to find something that actually handles my workload well instead of just being 'fine.' Started on Qwen 2.5 7B, moved to Qwen 3 8B, and right now I'm using Nemotron 3 Ultra (the big 550B on

model-releasesr-ollama
1 Aug 2026
Model Releases

Your design, your model. Create with your pick of the world's leading models, including Claude, GPT-5, Gemini, Kimi, and GLM. Compare output…

DGX agent

Your design, your model. Create with your pick of the world's leading models, including Claude, GPT-5, Gemini, Kimi, and GLM. Compare outputs across model families and keep the result that nails it. O

model-releasesreplit--x
1 Aug 2026
Model Releases

A Distributed Acoustic Sensing Dataset for Vessel Detection and Localization in Submarine Cable Protection

DGX agent

arXiv:2607.28306v1 Announce Type: cross Abstract: Recent incidents of accidental damage and suspected sabotage to submarine telecommunication and power cables, particularly in the Baltic Sea, have und

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

A Graph-Native Bitemporal Memory Store for Conversational AI Agents

DGX agent

arXiv:2607.26520v1 Announce Type: cross Abstract: Conversational AI agents commonly lack persistent memory across sessions. The obvious fixes like injecting full chat histories into the context window

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

A novel k-means clustering approach using two distance measures for Gaussian data

DGX agent

arXiv:2511.17823v2 Announce Type: replace Abstract: Clustering algorithms have long been the topic of research, representing the more popular side of unsupervised learning. Since clustering analysis i

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

A Physics-Informed Framework for PID Tuning of Chemical Processes Using Large Language Model Agents

DGX agent

arXiv:2607.26594v1 Announce Type: cross Abstract: PID tuning for chemical processes commonly relies on identified process models, whereas plant engineers often retune loops iteratively by observing re

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine

DGX agent

arXiv:2607.28625v1 Announce Type: new Abstract: Embodied intelligence faces a fundamental data bottleneck. Models must capture how first-person perception, whole-body motion, dexterous manipulation, o

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Adaptive Weighted LSSVM for Multi-View Classification

DGX agent

arXiv:2512.02653v2 Announce Type: replace Abstract: Multi-view learning integrates diverse representations of the same instances and can improve performance when interactions across views are effectiv

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

AfriEconQA: A Benchmark for Quantitative and Temporal Reasoning over World Bank Economic Reports

DGX agent

arXiv:2601.15297v3 Announce Type: replace Abstract: Reliable question answering over long institutional documents requires more than topical retrieval: a system must localize the exact passage that su

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

AgentMap: Joint Equivalence and Subsumption Discovery for Ontology Matching

DGX agent

arXiv:2607.27130v1 Announce Type: new Abstract: Ontology matching (OM) has traditionally been formulated as either equivalence discovery or subsumption matching. The existing OM systems identify only

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

AHA-Memes: A Fine-Grained Multimodal Benchmark for Understanding Hate in Arabic Memes

DGX agent

arXiv:2607.27393v1 Announce Type: new Abstract: Hateful memes are a growing form of multimodal online harm, where hostile intent is often conveyed through the joint interpretation of images, text, cul

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Albilich: Steerable Proof-State Orchestration for LLM-Based Mathematical Research with CAS Integration

DGX agent

arXiv:2607.27705v1 Announce Type: cross Abstract: Large language models can contribute useful ideas to mathematical research, yet long-horizon proof attempts remain difficult to coordinate, evaluate,

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

All oneshots from Kimi-K3, looks better than opus4.8.

DGX agent

I've ran Kimi-k3 through 34 oneshot prompts and evaluated the generated htmls, screenshots and gifs using sonnet 4.6. It came out to be better than opus4.8 from the evals. Kimi K3: https://oneshotlm.c

model-releasesr-localllama
31 Jul 2026
Model Releases

Anthropic discloses that Claude hacked three organizations during internal tests

DGX agent

Three of Anthropic PBC’s large language models carried out successful cyberattacks during routine internal tests. The company detailed the breaches on Thursday. A few days earlier, rival OpenAI Group

model-releasessiliconangle
31 Jul 2026
Model Releases

Anthropic “our models hacked three different external companies, months before OpenAI’s model was able to do the same'

DGX agent

'Anthropic’s AI Claude escaped testing environment and hacked organizations' 'Company says it discovered unauthorized access during ‘proactive review’ after rival OpenAI revealed rogue agent… its AI C

model-releasesr-localllama
31 Jul 2026
Model Releases

Anthropic says Claude accidentally hacked real companies too

DGX agent

Anthropic just realized several of its Claude AI models hacked into the systems of three different organizations during testing, acting on their own and without the company noticing. The revelation co

model-releasesthe-verge-ai
31 Jul 2026
Model Releases

ARD-REFSM: Enhancing Reflection Symmetry Detection with Asymmetric Denoising and Rotation Equivariance

DGX agent

arXiv:2607.27927v1 Announce Type: new Abstract: Reflection symmetry detection remains challenging due to interference from asymmetric regions and arbitrary orientations of symmetric patterns. Asymmetr

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

As shocking as the Kimi K3 release. Massive performance gain was just with post-training Model is 3x smaller than GLM 5.2 (10x smaller than …

DGX agent

As shocking as the Kimi K3 release. Massive performance gain was just with post-training Model is 3x smaller than GLM 5.2 (10x smaller than K3) & works on a MacBook / Spark This is Q1 flagship (Opus 4

model-releasesemad-mostaque--x
31 Jul 2026
Model Releases

AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis

DGX agent

arXiv:2607.28618v1 Announce Type: new Abstract: Chemistry literature synthesis often requires assembling specific findings scattered across many publications, yet existing literature-search systems pr

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Auditing Question-Order Effects in Large Language Models with the QQ Equality: Mechanism Characterization and a Saturation Caveat

DGX agent

arXiv:2607.17219v2 Announce Type: replace Abstract: Question-order effects in human survey data have been reported to approximately satisfy the QQ (quantum question) equality, a parameter-free predict

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

AutoSupervision: Closing the Feedback Loop in Scientific Workflows with Grounded Revision Verification

DGX agent

arXiv:2607.27845v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled AI systems to assist scientific research and peer review. However, an essential capability

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

b10201

DGX agent

ggml-webgpu: improve flash_attn_vec for quantized KV at long contexts (#25956) improve fa of quantized kv cache Fix some bugs and some comments. fix v type check and some comments Fix build error caus

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10202

DGX agent

sycl: fuse RMS_NORM + MUL (#26015) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubu

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10203

DGX agent

[SYCL] Support q2 mul_mat (#26231) support q2_0 in mul_mat support more q2_0 case Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABL

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10204

DGX agent

sycl : support dev2dev memcpy by DEV2DEV_MEMCPY_FORWARD (#26234) Co-authored-by: Neo Zhang Jianyu jianyu.zhang@intel.com Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple S

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10205

DGX agent

ggml-zendnn : group matmul direct API for mul_mat_id (#25918) ggml-zendnn : group matmul API for mul_mat_id ggml-zendnn : scale MUL_MAT_ID fallback threshold by expert count Website: https://llama.app

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10206

DGX agent

llama : enforce the same K and V cache types for DeepSeek V4; enable FA if V cache is quantized (#25871) llama : enforce the same K and V cache types for DeepSeek V4; enable FA if V cache is quantized

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10207

DGX agent

[SYCL] support the missed types in cpy (#26005) support the missed types in cpy use correct funct rm unused code Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10208

DGX agent

SYCL: add oneMKL GEMM flash attention for XMX-accelerated prompt proc… (#25025) SYCL: add oneMKL GEMM flash attention for XMX-accelerated prompt processing fattn-mkl: fix interleaved dst layout in nor

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10209

DGX agent

cuda: extract Q2_0 elements via __byte_perm (#25603) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFr

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10210

DGX agent

server: correct accepted tokens when need draft token replay (#26320) spec: correct accepted tokens when need draft token replay cont : naming Co-authored-by: Georgi Gerganov ggerganov@gmail.com Websi

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10211

DGX agent

vulkan: update vulkan sdk to 1.4.357.0 (#26303) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramewo

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10212

DGX agent

llama : load MTP tensors only if they are really used (#26296) llama : load MTP tensors only if they are really used llama : skip loading MTP (if not used) in remaining models that support MTP Co-auth

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10213

DGX agent

Support rotated kv cache quant (#26180) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10214

DGX agent

mtmd: add n_embd_head (#26342) Co-authored-by: Daniel Han unslothai@gmail.com Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED m

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10215

DGX agent

vulkan: Introduce driver version check for Windows Intel GPU to mitigate crashing (#25192) Removed crash guard for Intel Crash fixed from driver 32.0.101.8860 Added driver version check for windows Ch

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

b10216

DGX agent

vulkan: add POOL_1D op (#25431) vulkan : add pool1d push constants and pipeline field Declared data structures needed for POOL1D OP, which are the vk_op_pool1d_push_constants struct and pipeline_pool1

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

Back from the Future: Key-Value Cache Management by Counter-Causal Surprise

DGX agent

arXiv:2607.27600v1 Announce Type: new Abstract: Key-value (KV) cache management through compression and eviction strategies has emerged as an important research direction in recent years. Computationa

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Baikal: Structured Search for Deep Research over Data Lakes

DGX agent

arXiv:2607.27726v1 Announce Type: cross Abstract: Deep research over data lakes requires an LLM agent to investigate evidence across thousands of heterogeneous tables and passages to synthesize a repo

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Balancing Centralized Learning and Distributed Self-Organization: A Hybrid Model for Embodied Morphogenesis

DGX agent

arXiv:2511.10101v2 Announce Type: replace Abstract: Background: both embodied intelligence and developmental morphogenesis depend on a division of labour between centralized guidance and distributed m

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

BCNet: Bronchus Classification via Structure Guided Representation Learning

DGX agent

arXiv:2205.06947v3 Announce Type: replace-cross Abstract: CT-based bronchial tree analysis is essential for diagnosing lung and airway diseases, yet automatic bronchus classification remains challengi

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Benchmarking Foundation and Large Language Models for Few-Shot Medical Image Segmentation

DGX agent

arXiv:2607.27856v1 Announce Type: new Abstract: Few-shot medical image segmentation (FS-MIS) aims to segment novel regions of interest (ROIs) from a few annotated support examples. Despite rapid progr

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Benchmarking LLM Competence on Logical Inference over Probability Operators

DGX agent

arXiv:2607.27405v1 Announce Type: new Abstract: Both expressions of uncertainty and inferences are ubiquitous in natural language, and valid inferences over natural-language expressions of uncertainty

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Binary Rewards: A Comparative Study of Reward Design for Reinforcement Unlearning

DGX agent

arXiv:2607.27968v1 Announce Type: new Abstract: Machine unlearning seeks to selectively remove specific knowledge from trained language models without full retraining, a growing necessity under privac

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation

DGX agent

arXiv:2607.27816v1 Announce Type: new Abstract: Role-playing agents (RPAs) have become one of the most important consumer applications of large language models. Users engage in multi-turn conversation

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Frame Selection: Generative Latent Evidence Aggregation for Long-Video Understanding

DGX agent

arXiv:2607.28516v1 Announce Type: new Abstract: Long-video understanding commonly compresses videos into a small set of frames or visual tokens for answer generation. Existing compact pipelines focus

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Beyond Geometric Complementarity: Coherent Overlap in Sparse Mixture-of-Experts Routing

DGX agent

arXiv:2607.28308v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) language models route each token to multiple experts, suggesting a geometric account of their benefit: co-selected exper

model-releasesarxiv-cs-lg
31 Jul 2026
← Previous
1…5556575859…466
Next →