b9387
B9387 is a release of llama.cpp, the main playground for developing features for the GGML library, with the goal of enabling LLM inference with minimal setup and state-of-the-art performance on a wide
Knowledge catalogue
B9387 is a release of llama.cpp, the main playground for developing features for the GGML library, with the goal of enabling LLM inference with minimal setup and state-of-the-art performance on a wide
The search didn't return specific details about the b9388 release. Based on the available information about llama.cpp, this is likely a commit/release entry for the llama.cpp project. Here's a summary
b9391 is a build release of llama.cpp, an open-source C/C++ inference framework for running large language models locally. llama.cpp provides lightweight, optimized model inference with support for mu
B9393 is a build release of llama.cpp, the C/C++ inference framework for running large language models locally. llama.cpp enables LLM inference with minimal setup and state-of-the-art performance on a
b9394 is a release of llama.cpp, a project for LLM inference in C/C++ . The release represents a specific build or version update from the ggml-org/llama.cpp repository. Without access to the specific
b9401 is a release build of llama.cpp, an LLM inference framework in C/C++ that provides tools for running large language models locally. As an intermediate build in the llama.cpp project, it includes
B9402 is a release of llama.cpp, an open-source library that performs inference on large language models such as Llama and was developed in pure C/C++ with no dependencies. The release includes comman
B9403 is an intermediate build release of llama.cpp, an open-source project enabling LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cl
B9406 is a release build number from the llama.cpp project, a C/C++ implementation enabling LLM inference . Based on the release numbering pattern observed in the project's history, this represents an
b9410 is a release of llama.cpp, a C/C++ project that enables LLM inference with minimal setup and state-of-the-art performance on various hardware platforms. This build identifier represents a specif
b9411 is a release version of llama.cpp, a C/C++ implementation for LLM inference . This specific release build number represents updates to the open-source project hosted on GitHub, which provides op
b9412 is a release of llama.cpp, an open-source project for LLM inference in C/C++ . Build identifier b9412 represents a specific commit or version in the llama.cpp development lifecycle, following th
Release b9413 includes a CUDA fix that checks PTX version on the host side to guard PDL dispatch, addressing an issue where incorrect dispatching could occur on newer GPU architectures like sm_90/sm_1
b9414 is a release build of llama.cpp that includes improvements to CUDA PTX version checking , which helps prevent incorrect kernel dispatch on different GPU architectures. This build also adds suppo
arXiv:2605.28855v1 Announce Type: new Abstract: Temporal-difference learning with function approximation can be unstable under off-policy sampling. TDC stabilizes off-policy TD through an auxiliary co
arXiv:2602.14307v3 Announce Type: replace Abstract: As frontier Large Language Models (LLMs) increasingly saturate new benchmarks shortly after they are published, benchmarking itself is at a juncture
arXiv:2605.29233v1 Announce Type: cross Abstract: Diffusion language models (dLLMs) generate text by iteratively denoising multiple token positions in parallel, offering an attractive alternative to s
CineStill 800T Night Film LoRA is a machine learning model addon for Stable Diffusion that applies a cinematic night aesthetic to AI-generated images, using the trigger word 'c1n3st1ll' at strength 0.
arXiv:2605.29638v1 Announce Type: new Abstract: E-learning systems should deliver contents that reflect various phenomena of the language as it is used. In addition to formal Korean, e-learning system
ComfyUI just added @OpenRouter support. Instead of being locked into a single LLM, you can now access 20+ models directly inside Comfy. More flexibility, less friction, same workflow. Links to the wor
ComfyUI's Mini Story Generator is a node designed to generate concise and engaging stories based on a given theme. The tool is versatile for creating succinct narratives with minimal user input, stand
arXiv:2605.30334v1 Announce Type: new Abstract: Large Language Models (LLMs) have revolutionized various fields, yet their training efficiency is heavily reliant on effective data curation. While data
arXiv:2605.30152v1 Announce Type: cross Abstract: Proactive agents read user activity as text and call an LLM on every event to decide whether to act. But user activity is not natively text: it is a s
arXiv:2605.29511v1 Announce Type: cross Abstract: Tackling complex reasoning tasks typically relies on massive monolithic LLMs, which suffer from severe computational redundancy. While task decomposit
arXiv:2605.29663v1 Announce Type: new Abstract: Ground robots often carry payloads, implements, or other attachments that turn their effective footprint into complex, non-convex shapes. Navigating saf
arXiv:2605.29460v1 Announce Type: new Abstract: Federated fine-tuning of foundation models with Low-Rank Adaptation (LoRA) provides an efficient solution for reducing communication and computation cos
arXiv:2605.29695v1 Announce Type: new Abstract: Approximately 10% of newborns require assistance to initiate breathing at birth, and around 5% need ventilation support. Fetal heart rate (FHR) monitori
arXiv:2507.16880v3 Announce Type: replace-cross Abstract: Text-to-image diffusion models (DMs) have achieved remarkable success in image generation. However, concerns about data privacy and intellectu
arXiv:2605.29952v1 Announce Type: new Abstract: Accurate long-range prediction of geophysical systems is difficult due to strongly nonlinear dynamics, the high computational cost of full-physics simul
I don't have current information about this specific Reddit discussion or the installation process for WAN 2.2 into Forge Neo. This appears to be a technical support question from the StableDiffusion
arXiv:2605.29734v1 Announce Type: new Abstract: High-performance GPU kernels are essential for efficient LLM deployment, yet optimizing them remains expertise-intensive. Recent LLM-based code generati
A developer created an open-source GUI application with a brutalist design for interacting with Ollama models and the Personal Iris (PI) platform. The application is compatible with both ARM64 and Int
arXiv:2411.00278v4 Announce Type: replace Abstract: Time series anomaly detection (TSAD) underpins real-time monitoring in cloud services and web systems, allowing rapid identification of anomalies to
arXiv:2512.21311v2 Announce Type: replace Abstract: Solving partial differential equations (PDEs) on shapes underpins many shape analysis and engineering tasks; yet, prevailing PDE solvers operate on
This template provides a ComfyUI workflow configuration for integrating OpenRouter's API to enable large language model (LLM) capabilities within ComfyUI's node-based interface. It demonstrates how to
arXiv:2605.30335v1 Announce Type: new Abstract: Multi-component LLM agents assemble probabilistic claims from components that each see only part of a joint problem; the composition can violate basic p
arXiv:2603.27150v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized medical reasoning tasks, yet single-agent systems often falter on complex, interdisciplinary proble
arXiv:2605.30159v1 Announce Type: new Abstract: Memory-augmented LLM agents tackle complex long-horizon tasks by recursively summarizing interaction trajectories into compact memory. However, existing
arXiv:2605.29622v1 Announce Type: new Abstract: Coupled-cluster (CC) theory is often considered the gold standard of quantum chemistry, but its high computational cost limits routine access to accurat
arXiv:2601.05149v2 Announce Type: replace Abstract: Autoregressive (AR) models have achieved remarkable success in image synthesis, yet their sequential nature imposes significant latency constraints.
ComfyUI's MultiGPU feature, which allows users to distribute memory between VRAM and RAM effectively , has been integrated into the core ComfyUI codebase as native support. This enhancement enables CU
Liquid AI released LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and lightweight server-side use-cases. The model is a fast, memory-
OpenJarvis: a local-first personal AI is now available to run with Ollama Built by Stanford’s @HazyResearch and Scaling Intelligence labs, as part of their “Intelligence Per Watt” research into effici
PlagueKind Nodes is a ComfyUI custom node designed to facilitate stacking multiple LTX LoRAs (Low-Rank Adaptations) within a single node, offering streamlined management for adding and removing these
arXiv:2605.29275v1 Announce Type: new Abstract: Open-ended post-training benefits from rewards that make prompt-specific success conditions explicit, rather than relying only on post-hoc scalar scores
arXiv:2605.29158v1 Announce Type: new Abstract: Protein homology search underlies function annotation, structure prediction, and evolutionary analysis, but remains challenging in the 'twilight zone,'
arXiv:2605.30075v1 Announce Type: new Abstract: Quantum Federated Learning (QFL) offers a promising framework to train quantum models across distributed clients while keeping data strictly local. Due
arXiv:2605.29538v1 Announce Type: new Abstract: With the emergence of wireless applications in three-dimensional environments, such as the low-altitude airspace and 3D heterogeneous networks, radio ma
arXiv:2605.30327v1 Announce Type: cross Abstract: Frontier reasoning models are produced by posttraining base language models with reinforcement learning. Recent work has challenged this by showing th
arXiv:2605.28831v1 Announce Type: cross Abstract: Long-horizon interactive agents often accumulate large trajectory histories yet still fail to answer questions about earlier events reliably. We argue
arXiv:2603.13249v2 Announce Type: replace-cross Abstract: Activation steering offers a computationally efficient mechanism for controlling Large Language Models (LLMs) without fine-tuning. While effec
arXiv:2605.29826v1 Announce Type: cross Abstract: Existing methods in Multimodal Knowledge Editing (MKE) have advanced the ability to correct outdated or inaccurate knowledge in Multimodal Large Langu
arXiv:2509.24895v2 Announce Type: replace Abstract: While protein language models (PLMs) are one of the most promising avenues of research for future de novo protein design, the way in which they tran
arXiv:2605.29534v1 Announce Type: new Abstract: Recent advances in mobile GUI agents have shown strong potential for automating mobile tasks, but most effective systems still depend on large vision-la
arXiv:2605.30227v1 Announce Type: cross Abstract: While Multi-Agent Systems (MAS) empower Large Language Models to tackle complex reasoning tasks through collaborative interaction, optimizing their dy
arXiv:2605.29287v1 Announce Type: cross Abstract: Item-to-Item (I2I) retrieval is a fundamental part of modern content platforms, supporting critical industrial workflows from recommendation engines t
arXiv:2605.29691v1 Announce Type: new Abstract: Self-supervised learning (SSL) has produced a diverse landscape of vision transformers (ViTs) whose pretrained representations support a wide range of d
An updated MarkItDown API Server integrates Microsoft MarkItDown for converting PDF files, images, and Word documents to Markdown, with Ollama and LLaVA for generating image descriptions. This project
This discussion likely covers methods and tools for combining or merging multiple video files generated with Stable Diffusion, an AI image and video generation model. Community members likely share re
Web Chat para Ollama is a tool that enables users to run a private, free AI chat interface locally on their own machines using Ollama, an open-source framework for running large language models. The p