AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “applications”

GridTimelineEvolution
137 results
CompaniesToolsTechniques

Each lane shows up to 8 recent matching entries, ordered from earlier to later. Tracks load separately to keep the 75,000+ entry wiki fast.

Companies

CompanyAnthropic8 recent entries
4 May 2026Sentinel: an open-source local-first desktop app for AI coding

Sentinel is a local-first AI coding desktop application built with Rust and Tauri that automatically routes coding tasks to appropriate models based on task complexity. Instead of swapping between mul

→5 May 2026Parllama -- a terminal UI for Ollama model management and multi-provider LLM chat

Parllama is a TUI (Text UI) application designed for easy management and use of Ollama-based LLMs that also works with major cloud-provided LLMs. It provides core model management features including f

3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
→18 May 2026Witchcraft, fast local semantic search on top of SQLite [P]

Witchcraft is a Rust reimplementation of Stanford's XTR-Warp semantic search engine that uses a single-file SQLite database for storage, enabling client-side deployment. The system operates completely

→22 Jul 2026Stuck scaling a Next.js app on M3 Pro (36GB) using local Qwen 3.6 + VS Code Copilot. Should I switch extensions or go paid?

Hey everyone, I’m a Full-Stack Developer with 6+ years of experience. I’m relatively new to AI-assisted development workflows and want to build a production-ready, enterprise-level Next.js web applica

→23 Jul 2026contrib: allow all AI-generated code in general by ngxson · Pull Request #26012 · ggml-org/llama.cpp

Having read some merged PRs in the past, I know that they were fully written by Claude Code (or similar), so this basically fixes the delusion. But at the same time, we might start seeing more AI slop

→26 Jul 202616 bit better than lower quants for Qwen3.6-27B

I am writing a fairly complex C++ windows MFC application. I have a few 3090s and can run F16 Qwen3.6-27B with 256K context and MTP. The quality of code is exceptional with this quant vs its lower qua

→30 Jul 2026P.A.I. — Sleek Native Desktop AI Overlayer for Local Ollama Models 🤖⚡

Greetings Community! 👋 I hope everyone is doing well! I'm Tauhid — Senior EEE student from a Bangladeshi University Today I'd like to share an open-source project I’ve been developing called P.A.I. (P

→5 Aug 2026Qwen Developers' responses from their recent Twitter/X AMA

Questions & Responses(in BOLD) below. Favorite question(s) moved to end of the thread with combined responses(removed duplicates). Be optimistic folks. I'm sure we're getting other models too apart fr

CompanyOpenAI8 recent entries
7 Jun 2026OpenAI plots biggest ChatGPT overhaul since launch

OpenAI is planning its biggest ChatGPT overhaul yet, aiming to turn it into a 'superapp' with coding tools and AI agents to boost revenue ahead of a potential stock market listing. The redesigned Chat

→22 Jul 2026Stuck scaling a Next.js app on M3 Pro (36GB) using local Qwen 3.6 + VS Code Copilot. Should I switch extensions or go paid?

Hey everyone, I’m a Full-Stack Developer with 6+ years of experience. I’m relatively new to AI-assisted development workflows and want to build a production-ready, enterprise-level Next.js web applica

→22 Jul 2026How to configure a custom OpenAI-compatible API in Cursor?

Hi everyone, I have access to a self-hosted (or third-party) LLM that exposes an OpenAI-compatible API. I have both the API URL and an API token, and the provider states that it's fully compatible wit

→24 Jul 2026Who has set up ChatGPT Finance?

It’s been out for Plus members for a little while now. Has anyone here connected it to their accounts? I myself haven’t done it, even though it’s only read access I’m seriously hesitant to hand over t

→25 Jul 2026Benchmarks: TensorSharp vs. llama.cpp

Cuda and Vulkan Benchmark: TensorSharp vs. llama.cpp I would like to share my latest open source local Unsloth (GGUF) LLM inference engine and applications. It supports many models from Unsloth, like

→30 Jul 2026P.A.I. — Sleek Native Desktop AI Overlayer for Local Ollama Models 🤖⚡

Greetings Community! 👋 I hope everyone is doing well! I'm Tauhid — Senior EEE student from a Bangladeshi University Today I'd like to share an open-source project I’ve been developing called P.A.I. (P

→31 Jul 2026Open Source Ternary LLM Engine in Rust/CUDA for Quantization, Serving, and Training of models on consumer GPUs, called Tritium (Apache 2.0)

This post was not written by a clanker. Hey guys, I'm a comp sci major who wanted to introduce a cool project I built for quantizing models to ternary (1.58 bit) with as minimal of loss as possible, a

→5 Aug 2026Qwen Developers' responses from their recent Twitter/X AMA

Questions & Responses(in BOLD) below. Favorite question(s) moved to end of the thread with combined responses(removed duplicates). Be optimistic folks. I'm sure we're getting other models too apart fr

CompanyGoogle6 recent entries
10 Apr 2026Is google deepmind known to ghost applicants? [D]

A thread on r/MachineLearning asks whether Google DeepMind is known for 'ghosting' applicants — i.e., failing to provide timely or any rejection responses after interviews or application submission...

→11 Apr 2026PhD or Masters for Computational Cognitive Science [R]

This Reddit post on r/MachineLearning discusses the decision between pursuing a PhD versus a Master's degree for those interested in computational cognitive science, an interdisciplinary field combini

→15 Apr 2026How much harder is it these days to get into a PhD program without having a high ranking degree for UG? [D]

This Reddit discussion thread from r/MachineLearning explores the growing challenges faced by applicants from non-elite undergraduate institutions when applying to PhD programs in machine learning and

→16 Apr 2026Google’s DeepMind just released new 4B and 27B MedGemma models!

Google DeepMind released MedGemma, a collection of medical vision-language foundation models based on Gemma 3 in 4B and 27B parameter sizes, demonstrating advanced medical understanding and reasoning

→5 May 2026I asked ChatGPT and Gemini to do this weird perspective portrait with my face. I gave it a front face picture and a profile, including the example artwork. That was the result.

This post documents a user's experiment comparing ChatGPT and Gemini's ability to create perspective portrait artwork using two reference images (a front-facing photo and a profile photo) along with a

→22 Jul 2026Cactus Hybrid: We taught Gemma 4 to know when it's wrong

Hey HN, Henry & Roman here from Cactus. A small, on-device model is fast and private, but sometimes wrong, but frontier models are getting expensive pretty fast. So, we post-trained Gemma 4 E2B post-t

CompanyMeta6 recent entries
14 Apr 2026Agents in Ollama and Langflow

This Reddit post from r/ollama likely discusses how to build and run AI agents locally by combining Ollama — which handles local model serving to keep data private — with Langflow's visual, drag-and-d

→1 May 2026Need suggestions for Rag model

A Reddit discussion in the r/ollama community seeking recommendations for RAG (Retrieval-Augmented Generation) models, likely addressing model selection for local LLM-based document retrieval systems.

→25 Jul 2026Benchmarks: TensorSharp vs. llama.cpp

Cuda and Vulkan Benchmark: TensorSharp vs. llama.cpp I would like to share my latest open source local Unsloth (GGUF) LLM inference engine and applications. It supports many models from Unsloth, like

→30 Jul 2026P.A.I. — Sleek Native Desktop AI Overlayer for Local Ollama Models 🤖⚡

Greetings Community! 👋 I hope everyone is doing well! I'm Tauhid — Senior EEE student from a Bangladeshi University Today I'd like to share an open-source project I’ve been developing called P.A.I. (P

→30 Jul 2026Mechanistic interpretability streamlined for everyday users like us😎 🧠

Context: I want to give the community an Open Research (well open under Apache 2.0 clause) - tool that allows everyday users like us to look deeper into the local models we use consistently. Mechanist

→2 Aug 2026Conclusion: r/LocalLLaMA still has brilliant open-weight research, but finding it requires wading through endless benchmark drama, non-local Discussion Points and repetitive hardware flexes.

I let Gemma4-31b run on my laptop for like almost a day using a heavily altered pi to do a deep dive on our beloved Llama tangentially related Subreddit, and this was the conclusion. Feels pretty accu

CompanyMistral2 recent entries
14 Apr 2026Agents in Ollama and Langflow

This Reddit post from r/ollama likely discusses how to build and run AI agents locally by combining Ollama — which handles local model serving to keep data private — with Langflow's visual, drag-and-d

→6 Aug 2026nvidia/NVIDIA-Nemotron-Parse-2.0 · Hugging Face

NVIDIA Nemotron Parse 2.0 transforms document images into structured, machine-readable representations with text, layout classes, bounding boxes, and reading-order information. Given a Red, Green, Blu

CompanyxAI1 recent entries
5 May 2026Parllama -- a terminal UI for Ollama model management and multi-provider LLM chat

Parllama is a TUI (Text UI) application designed for easy management and use of Ollama-based LLMs that also works with major cloud-provided LLMs. It provides core model management features including f

CompanyDeepSeek8 recent entries
4 May 2026Sentinel: an open-source local-first desktop app for AI coding

Sentinel is a local-first AI coding desktop application built with Rust and Tauri that automatically routes coding tasks to appropriate models based on task complexity. Instead of swapping between mul

→5 May 2026Parllama -- a terminal UI for Ollama model management and multi-provider LLM chat

Parllama is a TUI (Text UI) application designed for easy management and use of Ollama-based LLMs that also works with major cloud-provided LLMs. It provides core model management features including f

→22 Jul 2026Cactus Hybrid: We taught Gemma 4 to know when it's wrong

Hey HN, Henry & Roman here from Cactus. A small, on-device model is fast and private, but sometimes wrong, but frontier models are getting expensive pretty fast. So, we post-trained Gemma 4 E2B post-t

→26 Jul 202616 bit better than lower quants for Qwen3.6-27B

I am writing a fairly complex C++ windows MFC application. I have a few 3090s and can run F16 Qwen3.6-27B with 256K context and MTP. The quality of code is exceptional with this quant vs its lower qua

→2 Aug 2026Real-world reality check on Qwen for autonomous coding agents

TLDR below 👇🏼 I’ve seen a lot of hype around Qwen 3.6 35B and 3.5 120B lately, especially regarding coding and tool-use capabilities. On this subreddit it is the defacto recommended model for everyone

→5 Aug 2026Qwen Developers' responses from their recent Twitter/X AMA

Questions & Responses(in BOLD) below. Favorite question(s) moved to end of the thread with combined responses(removed duplicates). Be optimistic folks. I'm sure we're getting other models too apart fr

→10 Aug 2026DeepSeek V4 Flash 0731 is the ‘killer app’ that is going to sell A LOT of DGX Sparks

Having a ‘Killer Application’ that everyone wants to use helps sell hardware, plain and simple. DeepSeek V4 Flash 0731 isn’t an app of course, but I think it’s going to be the major catalyst for getti

→11 Aug 2026I gave DeepSeek V4 Flash basic vision by training a 40M connector on 100K examples

I wanted to find out whether a huge text-only MoE could be given basic vision without retraining the language model itself. The short answer is yes. I froze DeepSeek V4 Flash and a 417M-parameter Moon

CompanyNVIDIA8 recent entries
1 Jun 2026Nvidia releases Cosmos3-Super-Image2Video . 64B parametres

Cosmos3-Super-Image2Video is a 64B model for temporally coherent image-to-video generation . NVIDIA's Cosmos platform is designed to accelerate Physical AI development by enabling machines to understa

→15 Jul 2026Audio perception layer for LLM agents, with a memory that grows through use

LLMs handle speech well once you run speech-to-text. They don't hear the rest: a bird outside, a glass breaking two rooms away, a smoke alarm two floors down. I've been working on an experimental open

→25 Jul 2026Benchmarks: TensorSharp vs. llama.cpp

Cuda and Vulkan Benchmark: TensorSharp vs. llama.cpp I would like to share my latest open source local Unsloth (GGUF) LLM inference engine and applications. It supports many models from Unsloth, like

→28 Jul 2026Manga Coloring Tool 2

Hey everyone! 👋 I'm excited to announce the official release of Manga Coloring Tool 2.0, a completely free, local, open-source web application designed to colorize manga pages and chapters effortlessl

→6 Aug 2026nvidia/NVIDIA-Nemotron-Parse-2.0 · Hugging Face

NVIDIA Nemotron Parse 2.0 transforms document images into structured, machine-readable representations with text, layout classes, bounding boxes, and reading-order information. Given a Red, Green, Blu

→9 Aug 2026I Turned My Underused Gaming Laptop Into a Local AI Workstation

TL;DR: I am building a Windows-first local AI setup for people who want to try local LLMs without spending days choosing models, setting up Ollama, Docker, WSL, Open WebUI, agents, and tool permission

→10 Aug 2026DeepSeek V4 Flash 0731 is the ‘killer app’ that is going to sell A LOT of DGX Sparks

Having a ‘Killer Application’ that everyone wants to use helps sell hardware, plain and simple. DeepSeek V4 Flash 0731 isn’t an app of course, but I think it’s going to be the major catalyst for getti

→12 Aug 2026LiquidAI/LFM2.5-VL-3B · Hugging Face

LFM2.5-VL-3B is a multimodal variant of LFM2.5, a family of hybrid models designed for on-device deployment. It builds on LFM2-VL-3B with further mid- and post-training. LFM2.5-VL-3B can process both