Ollama Cloud has become unbearably slow
This Reddit thread from r/ollama discusses user-reported degradation in performance on Ollama Cloud, with community members experiencing notably sluggish response times. The complaints likely center o
Knowledge catalogue
This Reddit thread from r/ollama discusses user-reported degradation in performance on Ollama Cloud, with community members experiencing notably sluggish response times. The complaints likely center o
An open-source agent self-reflection harness built on top of Ollama, shared in the r/ollama community, that enables locally run LLMs to evaluate and iteratively refine their own outputs. The project p
OllamaGPT is a self-hosted, ChatGPT-style web interface that runs locally using Ollama as its backend LLM engine, requiring no API keys or external services. It is designed for developers and power us
arXiv:2604.12100v1 Announce Type: new Abstract: Whole-slide image (WSI) classification in computational pathology is commonly formulated as slide-level Multiple Instance Learning (MIL) with a single g
This r/ollama Reddit post reflects a common community frustration around finding clear, architecture-specific documentation for writing Ollama Modelfiles. A Modelfile is the blueprint used to create a
Babel-Brief is a self-hosted Python project shared on r/ollama that acts as a local AI 'secretary' for Telegram, designed to monitor and summarize high-volume or noisy group chats using a locally runn
Qwen3 is a series of large language models spanning both dense and Mixture-of-Experts (MoE) architectures, with parameter scales ranging from 0.6B to 235B, and a key innovation being the integration o
arXiv:2604.12820v1 Announce Type: new Abstract: Large language models (LLMs) inherently absorb harmful knowledge, misinformation, and personal data during pretraining on large-scale web corpora, with
arXiv:2604.12768v1 Announce Type: new Abstract: Federated learning (FL) is a distributed paradigm that coordinates massive local clients to collaboratively train a global model via stage-wise local tr
arXiv:2604.12219v1 Announce Type: cross Abstract: Video Diffusion Transformers have revolutionized high-fidelity video generation but suffer from the massive computational burden of self-attention. Wh
arXiv:2510.00310v3 Announce Type: replace Abstract: Federated inference, in the form of one-shot federated learning, edge ensembles, or federated ensembles, has emerged as an attractive solution to co
A post from the r/StableDiffusion subreddit titled '*rubs hands together*,' likely featuring an AI-generated image created using Stable Diffusion, possibly showcasing a character or figure in an antic
A Reddit post from r/ollama in which a user shares their experience running a 31B parameter model locally using Ollama, reflecting on the surprisingly demanding hardware and infrastructure requirement
arXiv:2601.07177v2 Announce Type: replace-cross Abstract: Federated learning (FL) addresses privacy and data-silo issues in the training of large language models (LLMs). Most prior work focuses on imp
This Reddit post on r/ollama likely discusses techniques for using a locally-run Ollama LLM to sanitize, filter, or process incoming SMS messages — for example, detecting spam, extracting relevant inf
arXiv:2604.12600v1 Announce Type: new Abstract: The core challenge of hyperspectral image denoising is striking the right balance between data fidelity and noise prior modeling. Most existing methods
arXiv:2602.18109v2 Announce Type: replace Abstract: Real-time schedulers must reason about tight deadlines under strict compute budgets. We present TempoNet, a reinforcement learning scheduler that pa
arXiv:2512.03963v3 Announce Type: replace Abstract: Enhancing the temporal understanding of Multimodal Large Language Models (MLLMs) is essential for advancing long-form video analysis, enabling tasks
Tencent HY-World 2.0 is an open-source AI world model that generates real 3D scenes directly usable in game engines like Unreal Engine and Unity; unlike previous versions which produced video-based wo
A Reddit post from the r/ollama community documents a user's experiment running Google's Gemma 4 model locally via Ollama with web search (internet access) enabled, which resulted in the model becomin
arXiv:2502.07415v2 Announce Type: replace-cross Abstract: Inferring the mechanical properties of soft tissues from measured deformations is a fundamental challenge in elastography. A rarely examined a
arXiv:2508.05461v3 Announce Type: replace Abstract: Likelihood-based deep generative models have been widely investigated for Image Anomaly Detection (IAD), particularly Normalizing Flows, yet their s
arXiv:2512.05812v5 Announce Type: replace-cross Abstract: Scalable multi-agent driving simulation requires behavior models that are both realistic and computationally efficient. We address this by opt
arXiv:2603.18846v2 Announce Type: replace Abstract: Foundation models are used to extract transferable representations from large amounts of unlabeled data, typically via self-supervised learning (SSL
arXiv:2604.12149v1 Announce Type: new Abstract: Trajectory optimization depends heavily on initialization. In particular, sampling-based approaches are highly sensitive to initial solutions, and limit
arXiv:2604.12208v1 Announce Type: cross Abstract: Global navigation information and local scene understanding are two crucial components of autonomous driving systems. However, our experimental result
arXiv:2604.12887v1 Announce Type: new Abstract: Visual tokenizers map high-dimensional raw pixels into a compressed representation for downstream modeling. Beyond compression, tokenizers dictate what
arXiv:2604.12148v1 Announce Type: new Abstract: Video Large Language Models (VideoLLMs) excel at video understanding tasks where outputs are textual, such as Video Question Answering and Video Caption
The VNCCS PoseStudio LoRA for QIE2511 (Qwen-Image-Edit-2511) is a LoRA adapter designed to enhance pose control within the VNCCS (Visual Novel Character Creation Suite) ecosystem for ComfyUI. VNCCS Po
arXiv:2512.22799v2 Announce Type: replace Abstract: Vision-Language Tracking aims to continuously localize objects described by a visual template and a language description. Existing methods, however,
Ollama announced they are expanding cloud infrastructure capacity by adding more GPUs to meet demand, asking users for patience during the scaling process. This indicates Ollama's cloud service experi
We believe the future of intelligence is decentralized. Tonight we're opening the new @agi_inc office in SF with a dinner, gathering chipmakers, OEM partners, and investors to discuss what happens whe
The Ollama Cloud Subscription includes a feature described as 'Run cloud models at a time,' which refers to how many AI models a user can have simultaneously loaded and running in the cloud at any giv
This r/StableDiffusion post likely showcases a creative AI-generated video or audio experiment combining the poetry or likeness of Scottish poet Robert Burns with two AI generation tools: LTX (Lightri
This r/ollama thread is a community discussion where users share their preferred cloud-based AI models for coding tasks and compare their reasoning capabilities. The conversation likely highlights pop
A Reddit thread on r/StableDiffusion questioning why JoyAI Image Edit — a unified multimodal foundation model for image understanding, text-to-image generation, and instruction-guided image editing —
A Reddit post in r/ollama where a user reports that installing or running WSL (Windows Subsystem for Linux) caused Wi-Fi connectivity issues on their Windows machine. The discussion likely covers the
This Reddit thread from r/StableDiffusion discusses community questions and updates around **ZiB (Z-Image-Base)**, an AI image model in the Z-Image family (which includes Z-Image-Base, Z-Image-Edit, a
This r/StableDiffusion Reddit thread likely discusses the topic of hand generation quality when using ZIB and ZIT — two AI image generation models from the Z-Image ecosystem used in Stable Diffusion a
This Reddit post from r/StableDiffusion is a community-shared resource offering a collection of over 300 ready-to-use Stable Diffusion prompts organized across four visual disciplines: photography, ci
arXiv:2604.11146v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative model training across distributed clients without sharing raw data, thereby preserving privacy. However, F
arXiv:2604.09685v1 Announce Type: new Abstract: We describe a zero-shot pipeline developed for the ACCIDENT @ CVPR 2026 challenge. The challenge requires predicting when, where, and what type of traff
arXiv:2603.12221v2 Announce Type: replace Abstract: This paper addresses the expression (EXPR) recognition challenge in the 10th Affective Behavior Analysis in-the-Wild (ABAW) workshop and competition
This r/ollama discussion covers 'abliterated' models — LLMs that have had their built-in refusal mechanisms removed through a technique called abliteration, allowing them to respond to prompts without
arXiv:2604.10096v1 Announce Type: new Abstract: Current embodied intelligent systems still face a substantial gap between high-level reasoning and low-level physical execution in open-world environmen
This Reddit post from r/ollama likely discusses how to build and run AI agents locally by combining Ollama — which handles local model serving to keep data private — with Langflow's visual, drag-and-d
This Reddit thread from r/StableDiffusion discusses the community's search for AI tools capable of analyzing an existing video and automatically generating a descriptive text prompt from it — essentia
arXiv:2604.10884v1 Announce Type: cross Abstract: Automated generation of executable Business Process Model and Notation (BPMN) models from natural-language specifications is increasingly enabled by l
A Reddit post from the r/StableDiffusion community likely depicting a user's node-based workflow (such as in ComfyUI) where most nodes are displayed in bright red, typically indicating errors, broken
Face and head swapping capabilities are available for LTX 2.3, primarily through external LoRA models and ComfyUI workflows rather than built-in native features. A dedicated face replacement workflow
arXiv:2604.10766v1 Announce Type: new Abstract: Open-set 3D macromolecule detection in cryogenic electron tomography eliminates the need for target-specific model retraining. However, strict VRAM cons
arXiv:2510.17934v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) has shown some success in augmenting large language models (LLMs) with external knowledge. However, as a
Build b8783 is a sequential incremental release of llama.cpp, the open-source C/C++ framework for running LLM inference locally and in the cloud. As with nearby builds in the b87xx series, it likely i
Build b8784 is a tagged release of llama.cpp, the open-source C/C++ library for efficient LLM inference maintained by ggml-org on GitHub. Like other incremental builds in the project's continuous rele
Build **b8786** is an incremental release of [llama.cpp](https://github.com/ggml-org/llama.cpp), the open-source C/C++ library for local LLM inference. Like other builds in the project's continuous re
Build b8787 is a tagged release of llama.cpp, an open-source C/C++ library for running large language model (LLM) inference locally or in the cloud with minimal setup. As with all llama.cpp builds, it
Build b8788 is an incremental release of llama.cpp, the open-source C/C++ framework for efficient LLM inference developed by ggml-org. Like other builds in the project's continuous release cycle, it l
Build b8789 is an incremental release of llama.cpp, the open-source C/C++ library for running large language model inference locally and in the cloud. Like other builds in the project's continuous rel
Build b8790 is an incremental automated release of llama.cpp, the open-source C/C++ library for efficient LLM inference on local hardware. Like other builds in the project's continuous release cycle,
Build b8791 is an incremental release of llama.cpp, the open-source C/C++ library for local LLM inference maintained by ggml-org on GitHub. Like other numbered builds in the project's rapid release ca