AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,083
  • Agents7,615
  • Applications5,445
  • Concepts5
  • Hardware1,866
  • Industry6,184
  • Local Ai4,979
  • Model Releases24,164
  • Research20,258
  • Safety13,457
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,083
  • Agents7,615
  • Applications5,445
  • Concepts5
  • Hardware1,866
  • Industry6,184
  • Local Ai4,979
  • Model Releases24,164
  • Research20,258
  • Safety13,457
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlog
89,083Total entries
1Added by human
89,082Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,196 results
Research

Prompts in the Wild: A Large Analyzed Collection of Transactional Prompts in Code

DGX agent

arXiv:2608.12905v1 Announce Type: new Abstract: The behavior of contemporary generative Large Language Models (LLMs) is directly shaped by prompts, unstructured texts that describe the desired output

researcharxiv-cs-cl
14 Aug 2026
Model Releases

Qwen 30b MoE - 30tps - 6GB vram - Done!

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

So, I have been dreaming of getting 17 tokens per second using my RTX 3050 6GB version on a decent context window for Hermes needed above 60k. The hope is that has was a 22GB of DDR 4, hoping they can

model-releasesr-localllama
14 Aug 2026
Model Releases

Qwen 3.8-Max is live on Modal, with the full 1M context window and a custom DFlash speculator under the hood. Love seeing our launch partner…

DGX agent

Qwen 3.8-Max is live on Modal, with the full 1M context window and a custom DFlash speculator under the hood. Love seeing our launch partner Modal go all in from Day 0!⚡️ 2.4T parameters. 1M context.

model-releasesqwen--x
14 Aug 2026
Model Releases

Qwen Live now, an appetizer 👀

DGX agent

Qwen Live now, an appetizer 👀 Ready for Qwen Live EP2?We will start at 10:00AM! UTC+8 https://x.com/i/broadcasts/1mGPaZanWBqJN Qwen Cloud:https://www.qwencloud.com/?utm_content=g_20000002312 #QwenClou

model-releasesqwen--x
14 Aug 2026
Model Releases

Qwen3.8-Max is live on Together AI. Together AI is with us as a Day 0 launch partner, and we couldn’t ask for a better name to share Day 0 w…

DGX agent

Qwen3.8-Max is live on Together AI. Together AI is with us as a Day 0 launch partner, and we couldn’t ask for a better name to share Day 0 with. 2.4T parameters, 95B active, 1M context — all together

model-releasesqwen--x
14 Aug 2026
Model Releases

Regular reminder -- the set of public ARC 3 games is called 'demonstration set', not 'eval set' nor 'training set'. It is not meant to be us…

DGX agent

Regular reminder -- the set of public ARC 3 games is called 'demonstration set', not 'eval set' nor 'training set'. It is not meant to be used as training data, and it is not meant to be used as an ev

model-releasesfrancois-chollet--x
14 Aug 2026
Model Releases

Robust data-driven discovery of fractional differential equations via weak formulations and Pareto-based subset selection

DGX agent

arXiv:2608.12879v1 Announce Type: new Abstract: Fractional partial differential equations describe nonlocal dynamics, but discovering them from noisy data is difficult because fractional differentiati

model-releasesarxiv-cs-lg
14 Aug 2026
Model Releases

Robust Dempster-Shafer Evidence Fusion with Chaos-Conflict Measurement and Historical-Experience Weighting

DGX agent

arXiv:2608.13108v1 Announce Type: new Abstract: Multi-source evidence fusion under Dempster-Shafer theory faces two persistent challenges: existing conflict measures assess inter-evidence inconsistenc

model-releasesarxiv-cs-ai
14 Aug 2026
Research

SAEVerbalizer: Generating Explanations for Sparse Autoencoder Features via Representation Verbalization

DGX agent

arXiv:2608.13538v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) are proposed to extract numerous features from large language model (LLM) representations, yet explaining these features stil

researcharxiv-cs-cl
14 Aug 2026
Model Releases

SCULPT: Subtractive Composition for 3D Part Generation

DGX agent

arXiv:2608.13541v1 Announce Type: new Abstract: Part-aware 3D generation aims to create digital assets that are coherent as complete objects while exposing structural parts for editing, material assig

model-releasesarxiv-cs-cv
14 Aug 2026
Model Releases

Sign Language Video Synthesis via Loss-Guided Multi-Expert GANs

DGX agent

arXiv:2608.13368v1 Announce Type: cross Abstract: This preliminary technical report presents a framework for sign language video synthesis using a loss-guided multi-expert Generative Adversarial Netwo

model-releasesarxiv-cs-ai
14 Aug 2026
Agents

@skills: Attention is all you have

DGX agent

arXiv:2608.12610v1 Announce Type: new Abstract: There are 56,804 public agent skills today, and teams write many more privately. The dominant delivery model is installation: once installed, a skill's

agentsarxiv-cs-ai
14 Aug 2026
Tutorials

Sparse Orthogonal Regression Technique: A Spectral Framework for Equation Discovery, Approximation, and Integration

DGX agent

arXiv:2608.13504v1 Announce Type: new Abstract: We develop the Sparse Orthogonal Regression Technique (SORT), a sparse spectral framework for learning orthonormal-basis expansions from noisy and irreg

tutorialsarxiv-cs-lg
14 Aug 2026
Model Releases

SSPO: Structure-Aware Similarity-Weighted Preference Optimization for Neural Combinatorial Optimization

DGX agent

arXiv:2608.12443v1 Announce Type: cross Abstract: Neural combinatorial optimization (NCO) relies on parallel solution sampling for training, yet existing methods fail to fully exploit the rich informa

model-releasesarxiv-cs-ai
14 Aug 2026
Research

The Hidden Evolution of Disguised Visual Context inside the VLM

DGX agent

arXiv:2606.20077v2 Announce Type: replace-cross Abstract: Visual tokens enter Large Language Models (LLMs) as raw, foreign signals. How they are transformed into meaningful representations and interac

researcharxiv-cs-ai
14 Aug 2026
Research

Thought-Aware KV Cache Compaction for Reasoning via Adaptive Attention Matching

DGX agent

arXiv:2608.12331v1 Announce Type: cross Abstract: Reasoning language models generate lengthy chain-of-thought (CoT) sequences whose key-value (KV) cache grows linearly and becomes a memory bottleneck

researcharxiv-cs-ai
14 Aug 2026
Model Releases

TopoIntent: Compiling Security Intent into Executable, Compliance-Checked Network Topologies

DGX agent

arXiv:2608.13389v1 Announce Type: new Abstract: Enterprise security topology design requires translating business intent, regulatory requirements, and risk assumptions into zones, boundary devices, in

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

Training AI Scientists to Replicate Research

DGX agent

arXiv:2608.13331v1 Announce Type: cross Abstract: The replicability of papers is a cornerstone of scientific knowledge, ensuring the reliability of existing results and providing a base for further ex

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

/ultrafast

DGX agent

OpenAI announced the “Ultrafast” mode (GPT‑5.6 Sol), which can deliver responses up to 14× faster than standard API calls. The feature will first be released to a limited group of API users, with expa

model-releasessam-altman--x
14 Aug 2026
Model Releases

Vero: Can AI Agents Build Formally Verified Software Repositories?

DGX agent

arXiv:2608.13522v1 Announce Type: cross Abstract: AI agents are increasingly used for programming, but do not provide any guarantee on the correctness of generated code. Verified code generation, in w

model-releasesarxiv-cs-ai
14 Aug 2026
Safety

vToken: Token-Level Virtualization for Reclaimable KV Caches

DGX agent

arXiv:2608.13263v1 Announce Type: new Abstract: Large language model serving faces a critical memory bottleneck: the KV cache grows with sequence length and batch size. PagedAttention uses fixed-size

safetyarxiv-cs-ai
14 Aug 2026
Model Releases

What can I realistically do?

DGX agent

I am currently building up a local assistant profile on my MacBook Pro M2 with 32gbs. With Claude I am building out this Hermes agent to be my assistant. I am using Qwen3.6 A3B 4bit. We have Frankenst

model-releasesr-ollama
14 Aug 2026
Model Releases

What's new in LangChain? 🚀 🔌 Support for OpenAI's 3.0 SDK (using httpx2) ✨ Support for gemini-3.7-flash Plus a wave of core reliability fi…

DGX agent

What's new in LangChain? 🚀 🔌 Support for OpenAI's 3.0 SDK (using httpx2) ✨ Support for gemini-3.7-flash Plus a wave of core reliability fixes from external contributors: 🛠️ Tool calling & structured o

model-releasesharrison-chase--x
14 Aug 2026
Model Releases

When Can You Trust Offline Evaluation of Equal-Cost Top-k Allocation? A Controlled, Reproducible Benchmark and Practitioner's Guide

DGX agent

arXiv:2608.12489v1 Announce Type: new Abstract: Organizations decide whom to treat under a budget and want to know what a targeting rule would have earned before deploying it. Off-policy evaluation pr

model-releasesarxiv-cs-lg
14 Aug 2026
Model Releases

When Your Agent Opens the Chat App: Agent-Controlled Search over Raw Chat Logs Rivals Structured Memory

DGX agent

arXiv:2608.12888v1 Announce Type: new Abstract: Agent-memory systems increasingly buy retrieval quality with structure, transforming raw conversation histories into summaries, embeddings, trees, or kn

model-releasesarxiv-cs-cl
14 Aug 2026
Local Ai

Yes, we are back👑, with 206 tok/s on a single RTX 5090! Amazing Day-0 work from the SGLang team. Give it a try~@sgl_project

DGX agent

Yes, we are back👑, with 206 tok/s on a single RTX 5090! Amazing Day-0 work from the SGLang team. Give it a try~@sgl_project The king of small models is back! Qwen3.8-27B from @Alibaba_Qwen is open sou

local-aiqwen--x
14 Aug 2026
Model Releases

You can now turn off Google Gemini’s visible watermarks

DGX agent

Google will now allow you to remove visible watermarks from the images, videos, and music made with AI tools. With the update, you can toggle off a new 'Media watermark' setting in Gemini and Google's

model-releasesthe-verge-ai
14 Aug 2026
Agents

3D Scene Generation: A Survey

DGX agent

arXiv:2505.05474v2 Announce Type: replace Abstract: 3D scene generation seeks to synthesize spatially structured, semantically meaningful, and photorealistic environments for applications such as imme

agentsarxiv-cs-cv
13 Aug 2026
Model Releases

A 124B emitted 15,128 tokens in a single response on one DGX Spark, decode went 35.62 → 35.68 tok/s across the whole thing

DGX agent

Throughput observation more than a demo. Box and recording are sudoingX's on X, shared with his okay; I work on Ling at inclusionAI. He handed the web UI on his llama-server a 33-token prompt — build

model-releasesr-localllama
13 Aug 2026
Model Releases

A positive Claude watermark doesn’t prove it’s entirely AI-generated and the lack of a watermark doesn’t prove it’s entirely human-generated…

DGX agent

A positive Claude watermark doesn’t prove it’s entirely AI-generated and the lack of a watermark doesn’t prove it’s entirely human-generated. The two best ways to tell if something is AI-generated: 1)

model-releasesallie-k--miller--x
13 Aug 2026
Model Releases

A weird experiment I've been trying the last few weeks is having Claude take over day-to-day maintenance of our apps. Seeing early signs of …

DGX agent

A weird experiment I've been trying the last few weeks is having Claude take over day-to-day maintenance of our apps. Seeing early signs of life that this might be possible. The setup is straightforwa

model-releasesboris-cherny--x
13 Aug 2026
Agents

AgonAlpha: Autonomous Alpha Discovery via Prompt Economy and Scalable Agentic Search

DGX agent

arXiv:2608.11250v1 Announce Type: new Abstract: Language models can propose many plausible trading factors, but an autonomous research system must also allocate its evaluation budget, verify its own e

agentsarxiv-cs-ai
13 Aug 2026
Model Releases

Anthropic could be worth $2 trillion when it goes public

DGX agent

Anthropic is expected to float at a valuation of about 2 trillion or more in an October IPO, making it the largest initial public offering ever and surpassing SpaceX’s market cap. Investors cite rapid

model-releasesars-technica
13 Aug 2026
Research

Asymptotic Risk Calibration for Selective Question Answering

DGX agent

arXiv:2608.12008v1 Announce Type: new Abstract: Large language models (LLMs) may generate fluent but incorrect answers, making uncertainty quantification important for reliable question answering. How

researcharxiv-cs-cl
13 Aug 2026
Model Releases

AVA-Encoder: Towards Agent-Native Video Representation Learning

DGX agent

arXiv:2608.12313v1 Announce Type: cross Abstract: Creative agents still lack an effective way to learn from high-quality human films, limiting their ability to produce cinematic-grade videos. A key ch

model-releasesarxiv-cs-cl
13 Aug 2026
Model Releases

b10400

DGX agent

ggml : fix arm builds, unused var (#26991) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Li

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

b10405

DGX agent

ggml-hip : remove -funsafe-math-optimizations (#26696) It enables -fassociative-math, which reassociates FP reductions and can flip greedy argmax on RDNA3.5 (e.g. MTP speculative decode diverging from

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

b10408

DGX agent

sycl : Add DMMV ESIMD Q3_K kernel (#26251) Add DMMV Q4_K and Q6_K ESIMD kernels Configure cmake build with -DGGML_SYCL_ESIMD=ON to enable. Signed-off-by: Todd Malsbary todd.malsbary@intel.com Refactor

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

b10410

DGX agent

sycl: remove separate fp32 type promotion in gemm non-oneDNN path (#26372) sycl: use automatic fp16 promotion in gemm sycl: remove redundant comment Website: https://llama.app macOS/iOS: macOS Apple S

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

b10411

DGX agent

ggml-cpu/ops: vectorize flash-attention V-cache F16 to F32 conversion (#26947) Co-authored-by: jinzihao jinzihao.jzh@alibaba-inc.com Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) m

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

b10412

DGX agent

spec: enable backend sampling for both dflash & dspark (#26958) dflash: enable backend sampling for both dflash & dspark enable p_min > 0 in backend sampling and add guard cont : add TODO Co-authored-

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

b10414

DGX agent

metal : add TQ2_0 support (#26980) metal: add TQ2_0 support Add support for the GGML_TYPE_TQ2_0 (ternary, 2 bits per element) type in the Metal backend. Assisted-by: llama.cpp:DeepSeek-v4-Flash-0731 c

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

b10416

DGX agent

server : serve index.html with no-cache (#27006) index.html was served with max-age=31536000, immutable like the hashed assets, but its name is stable while its contents change every build, so a cache

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

b10417

DGX agent

chat : fix LFM2 tool call arg name prefix ambiguity (#26960) Assisted-by: Claude Opus 5 Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled)

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

b10418

DGX agent

[SYCL] Support host pinned mem to improve SYCL Host-to-Device Memory Access (#26789) support host pinned mem, ggml_backend_sycl_host_buffer_type_get_max_size, fix the thread-safe issue Website: https:

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

b10423

DGX agent

common: apply CPU parameters across tools (#27026) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFram

model-releasesllama-cpp-releases
13 Aug 2026
Applications

Calibration Bets on the Past: Post-Training Quantization for Financial Time-Series Forecasting

DGX agent

arXiv:2608.12259v1 Announce Type: new Abstract: Financial forecasting models are typically developed in full precision, yet production deployment often requires low-precision inference to reduce memor

applicationsarxiv-cs-lg
13 Aug 2026
Model Releases

ChatGPT can now remember your activity across the apps and websites on your computer. With Computer History in the desktop app, future inter…

DGX agent

ChatGPT can now remember your activity across the apps and websites on your computer. With Computer History in the desktop app, future interactions feel more personalized and require less explanation.

model-releasesopenai--x
13 Aug 2026
← Previous
1…758759760761762…1338
Next →