b8861
Release b8861 of llama.cpp removed /api endpoints from the server, including the /api/tags endpoint. The release was published on April 20, 2026. This is a maintenance update to the llama.cpp project,
Knowledge catalogue
Release b8861 of llama.cpp removed /api endpoints from the server, including the /api/tags endpoint. The release was published on April 20, 2026. This is a maintenance update to the llama.cpp project,
Release b8862 of llama.cpp includes a fix for 'mtmd: correct get_n_pos / get_decoder_pos' and provides compiled binaries for multiple platforms including macOS, Linux, Windows, Android, and openEuler
arXiv:2604.15414v1 Announce Type: cross Abstract: Continual reinforcement learning must balance retention with adaptation, yet many methods still rely on single-model preservation, committing to one e
Color shift flicker in Stable Diffusion refers to a known bug where images shift toward magenta and darker shadows during img2img processing . ControlNet can help address color shift by improving colo
arXiv:2604.15814v1 Announce Type: new Abstract: Hand-eye calibration through visual localization is a critical capability for robotic manipulation in open-world environments. However, most deep learni
arXiv:2604.15768v1 Announce Type: cross Abstract: AI-driven methods have demonstrated considerable success in tackling the central challenge of accurately solving the Schrodinger equation for complex
arXiv:2604.15594v1 Announce Type: cross Abstract: Modern datacenters schedule heterogeneous workloads across geo-distributed sites with diverse compute capacities, electricity prices, and thermal cond
arXiv:2510.20299v3 Announce Type: replace-cross Abstract: Brain tumors are a challenging problem in neuro-oncology, where early and precise diagnosis is important for successful treatment. Deep learni
arXiv:2403.18026v2 Announce Type: replace-cross Abstract: High-throughput biological imaging is often constrained by a trade-off between acquisition speed and image quality. Fast imaging modalities, s
arXiv:2604.15750v1 Announce Type: cross Abstract: Diffusion language models (DLMs) have emerged as a promising alternative to autoregressive language generation due to their potential for parallel dec
DFlash for Qwen3.6-35B-A3B just dropped ⚡ The community was running the day-1 preview before we even finished training. Now it's done: ✅ Training complete ✅ Validation passed ✅ Weights finalized ↓ Go
arXiv:2604.16083v1 Announce Type: new Abstract: With the rapid advancement of deep generative models, realistic fake images have become increasingly accessible, yet existing localization methods rely
arXiv:2604.16104v1 Announce Type: cross Abstract: Lung cancer remains one of the leading causes of cancer-related mortality worldwide. Conventional computed tomography (CT) imaging, while essential fo
arXiv:2503.13543v2 Announce Type: replace-cross Abstract: Federated Prototype Learning (FedPL) has emerged as an effective strategy for handling data heterogeneity in Federated Learning (FL). In FedPL
arXiv:2604.15560v1 Announce Type: cross Abstract: NASA's Transiting Exoplanet Survey Satellite (TESS) has identified thousands of exoplanet candidates, yet many remain unconfirmed due to the limitatio
arXiv:2604.15795v1 Announce Type: new Abstract: 3D object detection models trained in one server plays an important role in autonomous driving, robotics manipulation, and augmented reality scenarios.
arXiv:2604.15775v1 Announce Type: new Abstract: Learning with large-scale datasets and information-critical applications, such as in High Energy Physics (HEP), demands highly complex, large-scale mode
arXiv:2503.07520v5 Announce Type: replace Abstract: Traditional supervised drone-view geo-localization (DVGL) methods heavily depend on paired training data and encounter difficulties in learning cros
Getting started: - Update ComfyUI to the latest version (or visit Comfy Cloud) - Search “Quiver” in the Template Library - Load a workflow and generate For more details: https://links.comfy.org/4vQQDW
arXiv:2604.15495v1 Announce Type: new Abstract: Navigating complex, densely packed environments like retail stores, warehouses, and hospitals poses a significant spatial grounding challenge for humans
arXiv:2604.15699v1 Announce Type: new Abstract: Graph self-supervised learning can reduce the need for labeled graph data and has been widely used in recommendation, social networks, and other web app
Image to SVG: Drop in a raster image (a photo, a sketch, a screenshot, or the output of another ComfyUI node) and Quiver traces it into a clean SVG. Properly vectorized paths, not an embedded bitmap.
This Reddit post from r/StableDiffusion discusses the challenges of achieving high-quality hair rendering in AI-generated images, with the user comparing their generated results to a reference image a
Kimi K2.6, 1T params runs on my desktop GPU workstation. 192GB VRAM + 475GB RAM Serious contender to be the best open coding LLM on Earth. We'll see I'm checking their default harness. Goes on and on
Kimi K2.6 is an open-source model featuring advanced coding, long-horizon execution, and agent swarm capabilities that is now available via Ollama Cloud . The model excels in coding and agentic tools
arXiv:2602.07303v3 Announce Type: replace-cross Abstract: Log anomaly detection is crucial for uncovering system failures and security risks. Although logs originate from nested component executions w
A benchmarking post testing 63 different samplers combined with the linear_quadratic noise scheduler for LTX-2.3 (Lightricks' audio-video generation model) . The linear_quadratic scheduler was origina
arXiv:2604.15729v1 Announce Type: cross Abstract: Whole Slide Image (WSI) analysis is pivotal in computational pathology, enabling cancer diagnosis by integrating morphological and architectural cues
Ollama has made the Kimi K2.6 model available in its library, allowing users to run this model locally through the Ollama platform. The announcement highlights expanded integration options documented
ComfyUI-KleinRefGrid is a custom node release for ComfyUI that provides a reference grid interface for convenient image referencing during the generative AI workflow. The node extends ComfyUI's node-b
arXiv:2604.15893v1 Announce Type: new Abstract: Intelligent fetal ultrasound (US) interpretation is crucial for prenatal diagnosis, but high annotation costs and operator-induced variance make unsuper
arXiv:2510.22149v3 Announce Type: replace-cross Abstract: Federated learning (FL) has emerged as a promising paradigm for decentralized model training, enabling multiple clients to collaboratively lea
arXiv:2604.15385v1 Announce Type: cross Abstract: Software documentation is essential for program comprehension, developer onboarding, code review, and long-term maintenance. Yet producing quality doc
arXiv:2510.06953v3 Announce Type: replace Abstract: The Uniform Information Density (UID) hypothesis proposes that effective communication is achieved by maintaining a stable flow of information. In t
arXiv:2604.16263v1 Announce Type: new Abstract: Coordinating multi-robot systems (MRS) to search in unknown environments is particularly challenging for tasks that require semantic reasoning beyond ge
Sharing an OS tool I have been building for a few weeks: http://opentraces.ai - git for agent traces. With OT, you have a local client that helps you santise, review and securely share traces from you
arXiv:2603.19339v2 Announce Type: replace-cross Abstract: Dimensionality reduction is critical for deploying dense retrieval systems at scale, yet mainstream post-hoc methods face a fundamental trade-
arXiv:2511.20834v2 Announce Type: replace-cross Abstract: Sparse Convolution (SpC) powers 3D point cloud networks widely used in autonomous driving and augmented/virtual reality. SpC builds a kernel m
arXiv:2604.16070v1 Announce Type: new Abstract: We present TableSeq, an image-only, end-to-end framework for joint table structure recognition, content recognition, and cell localization. The model fo
arXiv:2508.16739v2 Announce Type: replace Abstract: Unmanned Aerial Vehicles (UAVs) have become increasingly important in disaster emergency response by facilitating aerial video analysis. Due to the
ComfyUI's HY 3D 3.0 nodes enable rapid 3D asset generation from single or multi-view images, producing high-fidelity 3D models suitable for applications including product visualization and game develo
arXiv:2601.12193v3 Announce Type: replace Abstract: Modern video retrieval systems are expected to handle diverse tasks ranging from corpus-level retrieval, fine-grained moment localization to flexibl
arXiv:2604.15613v1 Announce Type: cross Abstract: We present VoodooNet, a non-iterative neural architecture that replaces the stochastic gradient descent (SGD) paradigm with a closed-form analytic sol
arXiv:2604.16248v1 Announce Type: new Abstract: Image geolocalization has traditionally been addressed through retrieval-based place recognition or geometry-based visual localization pipelines. Recent
This post compares two high-end graphics cards—the AMD Radeon RX 7900 XTX and NVIDIA GeForce RTX 4070 Ti Super—across multiple demanding workloads including gaming, AI image generation with Comfy UI,
b8841 is a release of llama.cpp dated April 18, 2026 , a C/C++ library for large language model inference. The main goal of llama.cpp is to enable LLM inference with minimal setup and state-of-the-art
b8842 is a release of llama.cpp, a C/C++ implementation for LLM inference. The llama.cpp project publishes multiple releases in a single day as part of its active development cycle. This specific rele
B8846 is a release version of llama.cpp, a C/C++ inference engine for running large language models locally. The release includes pre-built binaries and libraries for multiple platforms including macO
b8848 is a release from llama.cpp, an open-source C/C++ project for LLM inference . This release builds on the rapidly-developing codebase with regular updates that include bug fixes, feature improvem
b8850 is a release of llama.cpp that includes CUDA refactoring for AMD matrix multiplication acceleration, with fixes for CDNA and RDNA3 GPU architectures . The release provides precompiled binaries a
The llama.cpp project is the main playground for developing new features for the ggml library , and the main goal of llama.cpp is to enable LLM inference with minimal setup and state-of-the-art perfor
Congratulations to the @ollama team for shipping the Copilot CLI support! ollama launch copilot Ollama now supports GitHub's Copilot CLI, the terminal agent that works directly with repositories on Gi
An open-source debugging tool designed for AI pipelines that provides trace timeline visualization, differential analysis, node replay capabilities, and operates without telemetry, licensed under MIT.
This post discusses Graph RAG, an approach that uses local LLMs with Ollama to build graph-based knowledge indexes from source documents by deriving entity knowledge graphs and pregenerating community
Good morning, everyone! TLDR: full base weights for healed 18b merge are LIVE! I am so blessed and excited by all of the support for my frankenmerge of Jackrongs models. Positive and negative feedback
Chroma1-HD is recommended for quick fine-tuning or LoRA on high-resolution 1024x1024 images . The model applies a quadratic remapping to timestep sampling to ensure training visits tail regions more o
This GitHub project (AG-X) implements deterministic safety checks for Ollama agents, designed to catch malformed JSON responses, detect prompt injection attempts, and handle model refusals before they
This resource provides documentation on integrating Ollama with Copilot CLI, enabling users to leverage Ollama's local language models within the Copilot command-line interface. The integration allows
Oh, hi! 👋 Copilot in Ollama ollama launch copilot Ollama now supports GitHub's Copilot CLI, the terminal agent that works directly with repositories on GitHub. You can use it to: Explore issues and PR
ollama launch copilot Ollama now supports GitHub's Copilot CLI, the terminal agent that works directly with repositories on GitHub. You can use it to: Explore issues and PRs. Search across repos by la