ToolClaude Code6 recent entries10 Apr 2026OpenClaude com Ollama CloudOpenClaude is an open-source coding-agent CLI, forked from the Claude Code source, that adds an OpenAI-compatible provider shim enabling use of GPT-4o, DeepSeek, Gemini, Ollama local models, and 20...→15 Apr 2026Which cloud model do you use for coding? Which one got better reasoning?This r/ollama thread is a community discussion where users share their preferred cloud-based AI models for coding tasks and compare their reasoning capabilities. The conversation likely highlights pop→
ToolCursor2 recent entries22 Jul 2026How to configure a custom OpenAI-compatible API in Cursor?Hi everyone, I have access to a self-hosted (or third-party) LLM that exposes an OpenAI-compatible API. I have both the API URL and an API token, and the provider states that it's fully compatible wit→5 Aug 202640% speedup of MoE training with faster megakernel, by cursor, of all people (for B200s)daily reminder not to trust benchmarks and run it yourself. claimed e2e speedup is ~40%, forwards are ~140% faster I would wager that compared to a naive kernel anyone can write it's more in the range
ToolLangChain2 recent entries14 Apr 2026Love seeing this open-sourced. Had a great chat with @nicoalbanese10 some weeks ago where he hinted to something like this. Great reference …Love seeing this open-sourced. Had a great chat with @nicoalbanese10 some weeks ago where he hinted to something like this. Great reference architecture for cloud coding agents. Open Agents gives you →24 Apr 2026DeepAgents Sandbox — a lightweight native Linux sandbox for deepagents It gives full isolation for safely running LLM/agent-generated code —…DeepAgents Sandbox — a lightweight native Linux sandbox for deepagents It gives full isolation for safely running LLM/agent-generated code — no Docker, no VMs, no cloud dependency. Super fast and secu
ToolOllama8 recent entries9 Aug 2026Anyone already used a model imported directly in the ollama cloudOllama allons you to import model but have you ever tried doing so ? Like running model imported from hugging face or you own model ? Any use case you wanna share ? Very curious about that submitted b→10 Aug 2026You can now try offline computer use with cua-driver + Muse Glimmer, through our friends at @ollama 🦙 How-to: https://cua.ai/docs/how-to-gu…Francesco @francedot reported that the first fully offline computer‑based LLM experience was achieved by running Muse Glimmer 30B (from @AIatMeta) locally on macOS, using Cua Driver to control native →10 Aug 2026v0.32.8v0.32.8 is an Oct 10, 2023 release of the ollama repository on GitHub, following a pre‑release tag v0.32.8‑rc0. The update adds Muse Glimmer support for NVIDIA, AMD and additional platforms, with the →10 Aug 2026Rätt kontakt för rätt person.Söker en riktigt vass programmerare – jag har ett projekt jag tror kan bli stort. Jag letar efter en extremt kunnig utvecklare som vill hoppa på ett projekt från ett tidigt skede. Jag kan inte avslöja→10 Aug 2026RAG-art: Build Your Own Art Expert with ollamaI built myself a personal AI art history assistant https://github.com/lololerigolo60/RAG-art/tree/main I love art history but I have way too many books, PDFs, and notes scattered everywhere. So I buil→10 Aug 2026How to prevent LLM to act like a robot/assistant?I'm playing with a conversational agent I made using either api/generate or api/chats. In both case I do ask him to not ask follow up question, to not act like an assistant, etc. Either from a system →10 Aug 2026Dual-Node NVIDIA DGX Spark over Tailscale: A Remote-Access Testbed for Distributed LLM Training and Cyber-Threat-Intelligence Fine-TuningarXiv:2608.07226v1 Announce Type: cross Abstract: Compact AI systems make local language-model experimentation increasingly accessible, yet practical evidence for multi-node training on desktop-class →12 Aug 2026You can now use Ollama as a provider in GitHub Copilot for JetBrains. https://github.blog/changelog/2026-08-11-copilot-memory-and-ollama-in-…GitHub announced on August 12 2026 that users can now integrate Ollama as a provider in **GitHub Copilot for JetBrains**. This update allows JetBrains developers to switch to or add locally‑hosted (or
ToolVercel AI8 recent entries4 May 2026A Comparative Analysis of Machine Learning Models for Intrusion Detection in Intelligent Transport SystemsarXiv:2605.00279v1 Announce Type: cross Abstract: AI-powered edge computing security is moving Intelligent Transportation Systems (ITS) from passive, rule-based protections to proactive, smart, zero-t→11 May 2026Federated Spatiotemporal Graph Learning for Passive Attack Detection in Smart GridsarXiv:2510.02371v2 Announce Type: replace-cross Abstract: Smart grids are exposed to passive eavesdropping, where attackers listen silently to communication links. Although no data is actively altered→25 May 2026ObjectCache: Layerwise Object-Storage Retrieval for KV Cache ReusearXiv:2605.22850v1 Announce Type: cross Abstract: Prefix KV caching has become a key mechanism in LLM serving: it reduces time to first token (TTFT) by avoiding redundant computation across requests t→3 Jul 2026I'm interested in how you're all running Hermes day to day. drop your setup below, I'm mapping what the community reaches for. I'm mostly cu…I'm interested in how you're all running Hermes day to day. drop your setup below, I'm mapping what the community reaches for. I'm mostly curious about: - model: your daily driver, plus MoA or a local→7 Jul 2026When Words Predict WorkloadarXiv:2607.04951v1 Announce Type: cross Abstract: Standard distributed ac{llm} schedulers rely on static token counts or rolling latency averages, making them susceptible to failures on statutorily co→25 Jul 2026Modelos do ollama cloud perdendo qualidade?A algumas semanas, percebi algo diferente, aparentemente modelos de qualidade, em especial o GLM 5.1 ficou mais burro, e começou a mandar caracteres em mandarim para mim sem eu nem usar eles, e isto n→27 Jul 2026Give any Ollama-compatible client session memory + a shared knowledge wiki by swapping the chat URLHey folks — I built ContextMemory, an open-source agentic context gateway for apps that already talk to LLMs. The idea is simple: keep your existing POST /api/chat client (Ollama wire format), point i→5 Aug 2026How we built an MCP bridge to give our AgentCore-hosted AI agent access to local MCP toolsAI agents on Amazon Bedrock AgentCore run in the cloud, but users' tools and files live on their laptops. Learn how to build a secure MCP bridge that lets a cloud-hosted agent call local MCP servers b
ToolHugging Face8 recent entries5 Aug 2026MiniMax are issuing takedowns on decensor/explicit H3 LoRAsI saw that someone had uploaded an experimental decensor LoRA for H3 on HuggingFace earlier today. Shortly afterwards, in the discussions, a MiniMax employee issued a warning that if it was not remove→6 Aug 2026Actual is a local inference stack optimized to let you utilize your personal compute. With its low CPU impact while inferencing and multi-pl…Actual is a local inference stack optimized to let you utilize your personal compute. With its low CPU impact while inferencing and multi-platform portability, it’s a great pair for your local Hermes →9 Aug 2026Open-weight video gen that actually delivers. Five days with MiniMax H3 on local hardware.H3 weights went live on HuggingFace August 3rd and I started pulling them immediately. An omni-modal video model with native stereo audio in the same forward pass, where audio can actually drive the v→9 Aug 2026Anyone already used a model imported directly in the ollama cloudOllama allons you to import model but have you ever tried doing so ? Like running model imported from hugging face or you own model ? Any use case you wanna share ? Very curious about that submitted b→10 Aug 2026omlab/VLX-Seek-1.5-10B · Hugging FaceVLX-Seek-1.5-10B VLX-Seek-1.5-10B is the open-source 10B model in the VLX-Seek 1.5 family, designed for fine-grained perception and visual grounding in embodied scenarios. It targets practical setting→10 Aug 2026Muse Glimmer, the new 30B model, is available on Hugging Face right now - here's the GGUF version: https://huggingface.co/meta-models/Muse-G…Muse Glimmer, the new 30B model, is available on Hugging Face right now - here's the GGUF version: https://huggingface.co/meta-models/Muse-Glimmer-30B-GGUF 1/ big announcement today: we will be releas→12 Aug 2026Local AI is exploding! Transformers.js, that we've been building @huggingface for the past three years has now become the most popular open-…Local AI is exploding! Transformers.js, that we've been building @huggingface for the past three years has now become the most popular open-source library to run AI models directly in your browser and→12 Aug 2026LiquidAI/LFM2.5-VL-3B · Hugging FaceLFM2.5-VL-3B is a multimodal variant of LFM2.5, a family of hybrid models designed for on-device deployment. It builds on LFM2-VL-3B with further mid- and post-training. LFM2.5-VL-3B can process both