AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,570 results
1 Aug 2026

DeepSeek-V4-Flash-0731-UD-Q3_K_XL 3x3090 test results

Model ReleasesDGX agent

For anyone interested, here are the llama-bench results on 3 bit K_XL quantization. I think this could be pushed further but no luck so far. CURRENT RESULTS: full moe offloading Prefill suffers 116 --

DeepSeek-V4-Flash-Q4KExperts-F16HC-F16Compressor-F16Indexer-Q8Attn-Q8Shared-Q8Out-chat-v2-imatrix-0731.gguf

Model ReleasesDGX agent

Antirez stealthily uploaded the new weights in the old folder... and there we were tapping our fingers. https://huggingface.co/antirez/deepseek-v4-gguf/tree/main submitted by /u/challis88ocarina [link

DS4 flash 0731 - Acquarium Panel Failure - Q3_K_XL Unsloth

Model ReleasesDGX agent

https://preview.redd.it/1a39x4zivqgh1.png?width=1550&format=png&auto=webp&s=de591c039cc18782a6b5d8e402fdc1594be05132 start C:llmllamam5uildinllama-server.exe --model 'H:UD-Q3_K_XLDeepSeek-V4-Flash-073

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Fascinating: OpenAI’s @deanwball is saying Astra can do anything, and it’s not even clear it can do “anything” in math (let alone anything i…

Model ReleasesDGX agent

Fascinating: OpenAI’s @deanwball is saying Astra can do anything, and it’s not even clear it can do “anything” in math (let alone anything in more or open-ended, less formalizable domains). I dropped

For MoE models the arithmetic splits in two: capacity follows total params, speed follows active

Local AiDGX agent

A few people asked for this after the bandwidth thread, so here it is on its own instead of buried in a comment. The dense rule was simple: every token reads every weight, so tokens/sec ≈ bandwidth ÷

@FredKSchott @cramforce @matei_zaharia i am making clanker blog all decisions going forward https://forge.smol.ai/blog/every-repository-gets…

AgentsDGX agent

On July 24, @swyx announced that he had started work on 'forge agents' and outlined four new features for SmolForge: customizable skins and spritesheet animations. He also referenced an upcoming blog

Github repo to learn the OPD/OPSD and how they perform compared to GRPO, on a consumer grade GPU [P]

SafetyDGX agent

I am trying to learn concepts like On Policy Distillation (OPD), On Policy Self Distillation (OPSD) and how do they compare to RL algorithms like GRPO. There are a lot of papers on this, but because o

> Google was too nervous to release it and DeepMind was blocked from shipping products that could disrupt Google. bookmark for the next vc t…

ToolsDGX agent

> Google was too nervous to release it and DeepMind was blocked from shipping products that could disrupt Google. bookmark for the next vc that asks you 'what if <incumbent> builds this?' @_chenglou I

great write up, evals are hard! here are 2 broad buckets we use to evaluate agents: 1. Measure the State of the World 2. Agent as a Judge on…

AgentsDGX agent

great write up, evals are hard! here are 2 broad buckets we use to evaluate agents: 1. Measure the State of the World 2. Agent as a Judge on the Trajectory 1. Measure the state of the environment befo

Grok Build can do almost anything you can think of http://X.ai/cli

AgentsDGX agent

Grok Build can do almost anything you can think of http://X.ai/cli Most people seriously underestimate what Grok Build can do They assume an AI coding agent is only useful for building apps or writing

Hot take on OpenAI’s Astra: - Obviously impressive - But math is different from most other problems in that it is more amenable to to formal…

ApplicationsDGX agent

Hot take on OpenAI’s Astra: - Obviously impressive - But math is different from most other problems in that it is more amenable to to formal verification and synthetic data. How well it works in open-

Huawei 96gb

Local AiDGX agent

Ciao a tutti ho appena comprato 4 Huawei duo 96 GB a poco più di 4mila euro qualcuno ha già utilizzato CANN consigli da darmi? E la prima volta che utilizzo questo framework e non so proprio da dove i

I built Vao2, an open-source personal feed for News, YouTube, GitHub and more with local AI summaries

Local AiDGX agent

GitHub: https://github.com/Loann110/Vao2 Hello everyone, I’m currently developing Vao2, an open-source application that brings together news, YouTube channels, GitHub repositories and more (to be adde

If Leopold had read this on June 26 and trimmed his bets accordingly, SALP would not have melted down. I laid everything out. https://open.s…

SafetyDGX agent

If Leopold had read this on June 26 and trimmed his bets accordingly, SALP would not have melted down. I laid everything out. https://open.substack.com/pub/garymarcus/p/the-month-generative-ai-lost-it

If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, …

Model ReleasesDGX agent

If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, 17 real tasks from 3 repositories, with context-injection st

In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep u…

ToolsDGX agent

In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep understanding of your users. The winners will be the companie

Inside Larry Ellison's debt-fueled push to turn Oracle into an AI juggernaut by aligning with Trump, backing Project Stargate, and partnering with OpenAI (New York Times)

IndustryDGX agent

New York Times: Inside Larry Ellison's debt-fueled push to turn Oracle into an AI juggernaut by aligning with Trump, backing Project Stargate, and partnering with OpenAI — the first full day of the se

Is there a point where models just cannot get any smaller without losing intelligence?

Model ReleasesDGX agent

DeepSeek V4 Flash got me thinking... We keep seeing smaller models get way better. A model at a certain parameter count today can be much smarter than a model of the same size from a year or two ago.

Kimi K3 has set a new bar for OSS model intelligence! 2.8T params, 1M context, OpenAI-compatible API. Complete guide to running Kimi K3 on T…

TutorialsDGX agent

Kimi K3 has set a new bar for OSS model intelligence! 2.8T params, 1M context, OpenAI-compatible API. Complete guide to running Kimi K3 on Together AI 👇 👏👏 @Kimi_Moonshot 👏👏 https://www.together.ai/bl

Kimi K3: The Complete Developer Guide

TutorialsDGX agent

**Kimi K3 is Moonshot AI’s 2.8‑trillion‑parameter open‑weight language model—the largest ever released—designed for frontier tasks such as long‑horizon coding and deep reasoning.** Its architecture us

Local Ollama models

Local AiDGX agent

Hello all - i just got a new mac mini with 24GB of RAM and wanting to run local AI for Home Assistant and Hermes. I have been struggling to find a snapy model that will work with my machine. Currently

OpenAI really ought give @HuggingFace a $100M grant, much as @ClementDelangue is asking. And anyone who wants to understand the attack from …

AgentsDGX agent

OpenAI really ought give @HuggingFace a $100M grant, much as @ClementDelangue is asking. And anyone who wants to understand the attack from HF’s side should read this. Great walkthrough; A+ for visual

Our Twitch streams with @Cohere_Labs are #1 in the tech category⚡️ Cohere Labs' free ML Summer School series has brought together viewers fr…

TutorialsDGX agent

Our Twitch streams with @Cohere_Labs are #1 in the tech category⚡️ Cohere Labs' free ML Summer School series has brought together viewers from 20+ countries to learn about everything from NLP in LLMs

// Persistent Workspaces for Long-Lived Claude Code Agent Teams // Four issues to be aware of: > Working state vanishes when a terminal clos…

Model ReleasesDGX agent

// Persistent Workspaces for Long-Lived Claude Code Agent Teams // Four issues to be aware of: > Working state vanishes when a terminal closes and the team cannot be resumed. > Compaction condenses th

Possible to create accurate medieval woodcut style art?

Model ReleasesDGX agent

Wondering if it's possible to actually produce ai art works that are indistinguishable from authentic medieval woodcut illustrations like the one attached. All the AI attempts I've seen at re creating

Quoting Greg Brockman

ToolsDGX agent

at openai, many people hook their chatgpt up to slack. people really don't like when a coworker's chatgpt contacts them asking for help with a task, even when they'd be perfectly happy doing that same

Qwen 3.6 27B Q5 on 3x2080ti: 55tps with llama.cpp. Can I squeeze out more?

Model ReleasesDGX agent

CPU: Threadripper 3970X RAM: 128GB DDR4 GPUs: 3x2080ti 11GB The current best parameters to run it: llama-server --model Qwen3.6-27B-Q5_K_S.gguf --n-gpu-layers 999 --split-mode tensor --flash-attn on -

so cool to just see this pop up on my feed. hermes has built such an organic and creative community. there are so many bells and whistles to…

AgentsDGX agent

so cool to just see this pop up on my feed. hermes has built such an organic and creative community. there are so many bells and whistles to explore and being able to find those through natural langua

so much for the “general” part of general intelligence

Model ReleasesDGX agent

so much for the “general” part of general intelligence I have tried many times to get ChatGPT, Claude, Grok or Gemini to write scripts for my YouTube videos. It is still a complete failure. For one th

Some things to note: 1) AI is getting very good at math and science. 2) Two years ago LLMs could not do basic math consistently 3) According…

ApplicationsDGX agent

Some things to note: 1) AI is getting very good at math and science. 2) Two years ago LLMs could not do basic math consistently 3) According to @polynoamial this cost less than $2000 in current API co

Sources detail how OpenAI fell behind Anthropic in revenue growth and valuation after prioritizing consumer chatbots and flashy side projects over coding tools (Berber Jin/Wall Street Journal)

IndustryDGX agent

Berber Jin / Wall Street Journal: Sources detail how OpenAI fell behind Anthropic in revenue growth and valuation after prioritizing consumer chatbots and flashy side projects over coding tools — Bets

“Stochastic parrots” is not my term (it’s @emilymbender’s). But a lot of people today commenting on it are confused. To some extent (though …

Model ReleasesDGX agent

“Stochastic parrots” is not my term (it’s @emilymbender’s). But a lot of people today commenting on it are confused. To some extent (though I don’t think it’s a perfect metaphor, and have said that be

Ten advances in mathematics and theoretical computer science

Model ReleasesDGX agent

Ten advances in mathematics and theoretical computer science A few days ago it was Anthropic discovering cryptographic weaknesses with Claude using Mythos Preview, spending 100,000 on tokens and with

The “AGI-is-near” community keeps committing the same logical fallacy over and over; I have seen it at least half a dozen times today alone.…

SafetyDGX agent

The “AGI-is-near” community keeps committing the same logical fallacy over and over; I have seen it at least half a dozen times today alone. Every time there’s an advance, I see the same error. Here’s

There's no 'one weird trick” for prompting Krea 2 art styles—just many guidelines [WF included]

TutorialsDGX agent

TLDR: There is no one prompting trick that will result in Krea 2 Turbo giving you exactly the style you want and across the whole image. Instead, if you are trying to achieve styles without the use of

ThreatLocker raised a $190M Series F led by Elephant as it looks to extend its zero-trust enterprise security platform to protect against AI-related risks (Kyle Alspach/CRN)

AgentsDGX agent

Kyle Alspach / CRN: ThreatLocker raised a $190M Series F led by Elephant as it looks to extend its zero-trust enterprise security platform to protect against AI-related risks — The cybersecurity vendo

Trying to understand VRAM usage and find the sweet spot for Wan/SCAIL-2 (or other models) on a GPU

Model ReleasesDGX agent

So upfront I'll admit that this is a ChatGPT summary of my chat with it about this idea i had, but this post wouldn't exist any other way, so... I’m trying to get a better understanding of how VRAM is

V4 flash vs V4 Flash (0731). Guys, new DeepSeek V4 Flash(0731) is now free on InferX

Model ReleasesDGX agent

DeepSeek V4 Flash is now available on InferX, and it’s free to use. We’re continuing to add GPU capacity as demand grows. While we’re bringing additional capacity online, you may occasionally see high

Vivid depiction of a dark scenario for Nvidia. “Indeed, CEO Jensen Huang has been so aggressive in funding the AI ecosystem that the company…

HardwareDGX agent

Vivid depiction of a dark scenario for Nvidia. “Indeed, CEO Jensen Huang has been so aggressive in funding the AI ecosystem that the company is functioning almost like a bank; with annual free cash fl

Wake me when Astra solves a significant open-world problem that doesn’t revolve around formal verification. Or at least fixes poor @skdh’s v…

Model ReleasesDGX agent

Wake me when Astra solves a significant open-world problem that doesn’t revolve around formal verification. Or at least fixes poor @skdh’s video problems. I have tried many times to get ChatGPT, Claud

We are going to see a lot of vertically focused AI native companies accelerate. Routers, open-source models and specialized post-training en…

ApplicationsDGX agent

We are going to see a lot of vertically focused AI native companies accelerate. Routers, open-source models and specialized post-training enabled by companies like @FireworksAI_HQ have all made dramat

What speeds are everyone getting with deepseek v4 flash 0731?

Model ReleasesDGX agent

What speeds are everyone getting with deepseek v4 flash 0731? I’m getting~200 tps prompt processing / ~11 tps token gen, on 4x5060ti16gb with ddr4 3200 ram at 4-channel, via llamacpp, with context win

What's currently the 'smartest' LLM to use on 8GB vram and 16 RAM and same thing for 8 VRAM and 64 RAM?

Model ReleasesDGX agent

Been trying to find something that actually handles my workload well instead of just being 'fine.' Started on Qwen 2.5 7B, moved to Qwen 3 8B, and right now I'm using Nemotron 3 Ultra (the big 550B on

Your design, your model. Create with your pick of the world's leading models, including Claude, GPT-5, Gemini, Kimi, and GLM. Compare output…

Model ReleasesDGX agent

Your design, your model. Create with your pick of the world's leading models, including Claude, GPT-5, Gemini, Kimi, and GLM. Compare outputs across model families and keep the result that nails it. O

You've reached your usage limit, please upgrade to continue

Local AiDGX agent

I am a pro subscriber and have 0 usage, yet I cannot even chat with Ollama because of the message in the title that I keep receiving. This is ridiculous. Anyone else experience this issue and have a s

31 Jul 2026

4DHumanDiff: Direct Text-to-4DGS Generation for Consistent 360-Degree Dynamic Humans

ResearchDGX agent

arXiv:2607.27634v1 Announce Type: new Abstract: Generating high-quality 360-degree dynamic human assets from text prompts is challenging. Existing methods usually synthesize monocular or multi-view vi

A Distributed Acoustic Sensing Dataset for Vessel Detection and Localization in Submarine Cable Protection

Model ReleasesDGX agent

arXiv:2607.28306v1 Announce Type: cross Abstract: Recent incidents of accidental damage and suspected sabotage to submarine telecommunication and power cables, particularly in the Baltic Sea, have und

A First Look at Coding Agents' Compliance with AI Contribution Rules in Open-Source Communities

AgentsDGX agent

arXiv:2607.26819v1 Announce Type: cross Abstract: Open source communities have been flooded with AI-generated contributions. In defense, they have written contribution rules to regulate coding agents'

A Graph-Native Bitemporal Memory Store for Conversational AI Agents

Model ReleasesDGX agent

arXiv:2607.26520v1 Announce Type: cross Abstract: Conversational AI agents commonly lack persistent memory across sessions. The obvious fixes like injecting full chat histories into the context window

A Lightweight Foundation Model for Collider Physics with Multi-Domain Adaptation

ResearchDGX agent

arXiv:2607.27501v1 Announce Type: new Abstract: We present a lightweight approach to foundation modeling (extbf{NEXUS}) that leverages pre-trained learning from collider physics data towards out-of-do

A Methodology for Designing Knowledge-Driven Missions for Robots

AgentsDGX agent

arXiv:2601.20797v1 Announce Type: cross Abstract: This paper presents a comprehensive methodology for implementing knowledge graphs in ROS 2 systems, aiming to enhance the efficiency and intelligence

A Montage-Agnostic Encoder for Calibration-Light Cross-User Gesture Recognition from Surface Electromyography

ResearchDGX agent

arXiv:2607.27565v1 Announce Type: new Abstract: Pattern-recognition control promises a myoelectric prosthesis that responds to many intended gestures rather than one or two, but the promise has stayed

A novel k-means clustering approach using two distance measures for Gaussian data

Model ReleasesDGX agent

arXiv:2511.17823v2 Announce Type: replace Abstract: Clustering algorithms have long been the topic of research, representing the more popular side of unsupervised learning. Since clustering analysis i

A Physics-Informed Framework for PID Tuning of Chemical Processes Using Large Language Model Agents

Model ReleasesDGX agent

arXiv:2607.26594v1 Announce Type: cross Abstract: PID tuning for chemical processes commonly relies on identified process models, whereas plant engineers often retune loops iteratively by observing re

A Query-Efficient Stochastic Volume Rendering Framework for Time-Varying Implicit Neural Volumes

HardwareDGX agent

arXiv:2607.28047v1 Announce Type: cross Abstract: Time-varying implicit neural representations (INRs) provide a compact representation of scientific volumes and, for modalities such as dynamic X-ray c

A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models

ResearchDGX agent

arXiv:2607.26102v1 Announce Type: cross Abstract: Mathematical chain of thought (CoT) evaluation is commonly reduced to whether the final answer matches a reference. This conflates producing a correct

A Robust Placeability Metric for Model-Free Unified Pick-and-Place Reasoning

AgentsDGX agent

arXiv:2510.14584v3 Announce Type: replace Abstract: Reliable manipulation of previously unseen objects remains a fundamental challenge for autonomous robotic systems operating in unstructured environm

A Sparse Glimpse of the Whole: Train-Free Self-Speculative Decoding

ResearchDGX agent

arXiv:2607.27735v1 Announce Type: new Abstract: Speculative decoding alleviates the memory-bandwidth bottleneck in large language model inference, but its acceleration is jointly constrained by drafti

A Systems Engineering Framework for Vision-Language-Enabled UAV Triage and Disaster Response

SafetyDGX agent

arXiv:2607.27597v1 Announce Type: new Abstract: Recent advances in Vision Language Models (VLMs) have created new opportunities for disaster response, where responders must interpret large volumes of

Accelerating SGDM via Learning Rate and Batch Size Schedules: A Lyapunov-Based Analysis

ResearchDGX agent

arXiv:2508.03105v3 Announce Type: replace Abstract: We analyze the convergence behavior of stochastic gradient descent with momentum (SGDM) under dynamic learning-rate and batch-size schedules by intr

← Previous
1…139140141142143…1410
Next →