b9891
b9891 is a build release of llama.cpp, the open-source C/C++ inference engine for running large language models locally. This release enables LLM inference with minimal setup and state-of-the-art perf
Knowledge catalogue
b9891 is a build release of llama.cpp, the open-source C/C++ inference engine for running large language models locally. This release enables LLM inference with minimal setup and state-of-the-art perf
Release b9892 is a version identifier for llama.cpp, an open-source C/C++ framework for running large language model inference on consumer hardware. This specific build (b9892) represents a snapshot o
Barracuda has integrated with Databricks' AI assistant Genie to enable natural language querying of security logs, allowing security teams to analyze and extract insights from log data through convers
Big gains in PotatoMonetBench from Opus 4.5 to Fable: 'There are two controls. One is slider, in which one side is labelled Maximum Potato and the other is Formalware, there are four positions. The ot
big things happened over the weekend! OpenWiki is at 1.7k stars in just 3 days! Right now it's just for codebases, but we're working to expand it to everything for memory. What do you want to see in a
Big update to managing your past sessions in Hermes Agent. Now you can prune or archive with a huge array of filters to clean up your session DB without losing anything you don't want to lose. Get rid
Building the Document Context Layer for AI Agents AI Agents are the new knowledge workers, but the vast majority of knowledge work depends on unstructured documents. If you don’t build the proper tool
II-Agent is a platform where users can submit their creations or projects to a showcase feature, earning credits and potential visibility while allowing other users to explore and replay their work se
This post discusses how observing Claude's internal activation patterns (J-space) reveals the model's hidden reasoning processes, such as detecting code bugs and analyzing images, even when these step
Claude Fable running ComfyUI workflows through the Comfy MCP is a cheat code. → Claude pulled the shot from the web, auto-detected the scene cuts + trimmed it → ran my saved depthanything v3 + openpos
Armin Ronacher / Armin Ronacher's Thoughts and Writings: Claude Opus 4.8 and Sonnet 5 seem worse at tool calls than older models, likely due to post-training that assumes Claude Code-like harnesses as
Come join us on Thursday! It'll be a great session you'll want to add to your MEMORY.md 😄 does your agent need a wiki? should humans and agent use the same wiki? should you have one wiki or many wikis
Databricks describes contextual policies in Omnigent, a framework that leverages session state to implement more effective governance and control mechanisms for AI agents. By maintaining and utilizing
Cool open-source release. Physical AI is the next frontier, so things are shifting from thinking to taking action. LingBot-Vision looks like a strong visual foundation model and shows progress in cons
Creative structures are needed to get GPUs in the hands of startups + other companies that aren't Meta, OpenAI, Anthropic, SpaceXAI, Microsoft, Amazon, Google AI Debt Financing will be over $7T of deb
Creator seungho__yeo ( IG ) transformed one of most iconic scenes in About Time into an Arcane-inspired visual inside ComfyUI. The original performance, camera movement, and emotional pacing remain in
In this post, you deploy a two-phase infrastructure for multi-turn RL using Amazon Nova Forge on Amazon SageMaker HyperPod. By the end, you have an event-driven pipeline that starts training when you
devin has a hot take that wikis are NOT the right abstraction should be a fun webinar!! does your agent need a wiki? should humans and agent use the same wiki? should you have one wiki or many wikis?
Eric Katz / NOTUS: Document: a draft US Treasury Department report is set to warn about the AI market's risks, likening some key aspects to the dotcom crash in the early 2000s — Publicly, the Trump ad
does your agent need a wiki? should humans and agent use the same wiki? should you have one wiki or many wikis? if you're wiki-curious, join @BraceSproul @hwchase17 and I on Thursday! you don't want t
DRIVE-BY MEDIA: Most Americans believe that black women are primarily murdered by racist white men. It turns out that 99% are committed by blacks. 91% are committed by black men and 8% are committed b
During a Bloomberg interview, Yann LeCun (@ylecun ) explains why LLMs are limited in terms of real-world intelligence during a Bloomberg interview. 'Language is a very approximate, reduced, quantized,
🚨Emergency webinar: LLM Wikis and how to give your agent memory LLM Wikis are so hot right now - OpenWiki by @BraceSproul up to nearly 7k GitHub stars in less than a week I'll be chatting with Brace a
#ENGMEX likely refers to a project, initiative, or announcement related to England and Mexico, possibly involving technology, business, or partnership collaboration. Without access to the specific twe
Nonuniform Tensor Parallelism (NTP) enables large-scale LLM training jobs to dynamically adapt tensor parallelism degree in response to transient GPU unavailability, ensuring sustained Goodput and min
Race control displayed a 'Safety Car in this lap' message during the 2026 British Grand Prix at Silverstone, raising expectations for a final lap restart that never materialized, as the message was er
Gary Marcus argues that Geoff Hinton, rather than Dario Amodei, was the originator of making exaggerated claims about neural networks' impact on employment. This appears to be part of a broader discus
The FCC has proposed to streamline its broadband labeling rules by eliminating several requirements it views as unduly burdensome, including itemized pass-through fees . Instead of itemizing, ISPs wou
Ryan Vlastelica / Bloomberg: Filing: Brookfield-backed data center company Csquare seeks to raise up to 1.35B in an IPO, selling 50M shares at 23 to 27 each for an up to 4.18B valuation — Csquare Inc.
From a few hundred thousand to 100M AI builders in 2028? AI will create tons of jobs if we make it distributed instead of monopolized by a few companies! 'A few years ago, there were a few hundred tho
Today, we’re excited to announce a deep-link integration between Hugging Face and Amazon SageMaker AI. Developers can now go from model discovery to hands-on experimentation in SageMaker Studio with a
GenAI isn’t good enough to replace millions of employees, so there is basically no way the capex is going to earn out. If AI isn't going to wipe out millions of jobs every year, it has to generate pro
'Ghost memory' is a real problem with agents. You might have seen the issue where a long-running agent still confidently repeats a user fact that stopped being true weeks ago? New research names the f
This post appears to document a personal account from a San Francisco-based individual describing an experience they interpret as a divine sign, potentially related to concepts of artificial superinte
Ethan Mollick discusses Google's potential to develop frontier AI models, suggesting that Google has the capability to compete in advanced AI development if it chooses to re-enter or prioritize this s
ha ha Geoff Hinton in 2016 completely overhyping where deep learning was at then Geoffrey Hinton explains the coyote test for medical AI: radiology is already over the cliff, it just has not looked do
How do you trace one number in a 200 page ESG report back to the exact page it came from? We dug into that with the @llama_index team behind LiteParse. We tested five ways to retrieve evidence across
David Keohane / Financial Times: How Masayoshi Son is remaking SoftBank in his own image, betting on AI; SoftBank trades at a ~50% discount to its net asset value, as some question the strategy — The
Nations have long invested in domestic infrastructure to advance their economies, protect and use their data, and take advantage of technology opportunities in areas such as transportation, communicat
Every year, the International Conference on Machine Learning (ICML) reveals where thousands of AI researchers have decided to put their work. This year’s accepted papers reveal a clear direction: open
This article from Modal explains the pricing models and cost considerations for serverless GPU services, likely covering how providers like Modal calculate charges based on GPU usage, execution time,
Hugging Face has just been sued for alleged copyright infringement for hosting & distributing copyrighted images, used for AI training. It's been almost a *year* since I flagged to their CEO that they
Hy3, the new 295B MoE model from @TencentHunyuan, is now free in Nous Portal for the next two weeks! It is focused on cost-effective agentic use, and particularly strong on coding, tool-calling reliab
I can basically guarantee you are not being ambitious enough with the work you are assigning Fable Start asking for the maximum possible thing to figure out the far edges of what it can do. After that
I cannot believe how good GLM 5.2 is. Several weeks in now and it's mostly all I have been using. It doesn't have the alignment issue of cloud AI, it's much more clear what it can do and can't because
I predict 50% of companies will need new leadership, because the old management style won't work in the era of AI. That's exactly why I'm seeing 95%+ of so-called 'AI Transformation' initiatives fail.
I've long argued that Hollywood has simultaneously set and ruined our expectations for smart glasses. But after binge-watching two seasons of Netflix's A Man on the Inside, this is perhaps the first t
I talk to a lot of companies that still have active efforts to build GPTs. (It remains weird that OpenAI abandoned GPTs after rolling them out. They were the precursor to Skills & could have been a br
I wonder if Fable might do better at tasks it 'likes' than those that it doesn't. If you ask it to do something 'Fable-y' it reacts with enthusiasm and promises to overdeliver. Otherwise, it just gets
If that happens, where does it leave Anthropic and OpenAI? And Oracle, CoreWeave, Nebius, xAI, etc? Fable 5 probably running locally in about two years. That is the projection in this r/LocalLLaMA cha
If there is ANYTHING I have learned in my early Fable 5 testing, it’s that the judgement calls from AI are getting better and better. Not perfect and not (yet?) a substitute for a human’s taste, conte
If you want to start a startup, don't learn 'entrepreneurship.' Learn how to build things. The hard part of startups is not 'entrepreneurship' but product: to know what to build, and to be able to bui
This Import AI newsletter issue covers Fable's development of GPU kernel writing capabilities, likely discussing how AI systems can automatically generate optimized code for graphics processors, along
// In-context Retrieval at Million-token Scale // Great study providing better understanding of retrieval at million-token scale. They run the first systematic study of in-context retrieval at the sca
This post showcases experimental video transition techniques created using ComfyUI and Adobe After Effects, utilizing AI models including GPT Image 2.0, Wan 2.2 (for image-to-video, text-to-video, and
Interesting stuff. And the visualization at the end is worth trying: https://www.neuronpedia.org/qwen3.6-27b/jlens New Anthropic research: A global workspace in language models. Of everything happenin
Is this what founder mode looks like Sources: President Trump, commerce secretary Howard Lutnick, and White House task force head Andrew Giuliani put together a team of elite lawyers — from outside th
Together AI, in partnership with NVIDIA and Lyra Lab, hosted a fireside chat led by Tri Dao discussing current trends and future directions in AI research and infrastructure development. The event bro
Katalyst Space Technologies' LINK spacecraft successfully launched on July 3, 2026 , and will take approximately one month to rendezvous with NASA's Swift Observatory . The mission aims to boost Swift
Hugging Face announced major updates to their Kernels platform, which is their cloud-based computational environment for machine learning and data science projects. The revamp likely includes improvem