Foundry Local is now Generally Available
Microsoft Foundry Local is now generally available as an end-to-end local AI solution that enables developers to bring AI inference directly into their applications with no cloud dependency, no net...
Knowledge catalogue
Microsoft Foundry Local is now generally available as an end-to-end local AI solution that enables developers to bring AI inference directly into their applications with no cloud dependency, no net...
The open source ecosystem is making it easier for AI enthusiasts and developers to build, customize and run increasingly capable agents locally. Throughout August, NVIDIA is celebrating the partners a
Cade Metz / New York Times: xAI co-founder Igor Babuschkin's River AI raised $1B led by General Catalyst to build home or small business computer servers capable of running AI locally — Igor Babuschki
What’s the best way to recommend products to little-known users? We’ve spent our careers trying to solve this problem for major companies like Spotify and Priceline, and it’s why Sidd founded Malachyt
Shane Burke / The Information: US data center bans top 500, up from 300+ in late June, as New York and Texas join cities and counties pushing back against data center development — Local government re
AI agents on Amazon Bedrock AgentCore run in the cloud, but users' tools and files live on their laptops. Learn how to build a secure MCP bridge that lets a cloud-hosted agent call local MCP servers b
Welcome to the second Cloud CISO Perspectives for July 2026. Today, Chris Betz, CISO, Google Cloud, and Alicja Cade, Senior Director, Office of the CISO, Google Cloud, explain what boards of directors
Physical AI is forcing the technology industry to rethink the entire computing stack. Robots, autonomous systems and intelligent devices need economical inference, secure data access and infrastructur
Siri Expressive Voices synthesize rich, configurable speech in real time and entirely on device, powered by AFM 3 Core Advanced, Apple’s most powerful on-device foundation model. This work presents th
Perplexity has expanded its agentic Personal Computer tool to Windows, allowing computers running the world's most popular OS to be used as a locally run AI system. Like the Mac version that Perplexit
As AI moves beyond chatbots toward autonomous agents, attention is shifting enterprise AI PCs as a new layer of AI infrastructure. That transition is driving demand for hardware and software designed
NVIDIA founder and CEO Jensen Huang today visited the Naval Postgraduate School in Monterey, California, to commission an NVIDIA DGX GB300 system — bringing one of the world’s most powerful AI platfor
Nativ: Run AI models locally on your Mac Prince Canuma is the developer behind the excellent MLX-VLM Python library for running vision-LLMs using MLX on a Mac. I'm really excited about his new project
Erin Davis calls it the “SuperDuperPOD.” That’s two things in one name: pharmaceutical giant Bristol Myers Squibb (BMS) already runs one of the largest AI clusters in life sciences, with serious resul
General-purpose robots and autonomous machines are moving from research labs to real-world mass-market deployment, creating demand for compact, power-efficient AI supercomputers capable of running fou
A local coding agent, an in-app customer assistant, and an AI SRE triaging production logs may all use the same model class—but not the same harness, eval plan, or rollout risk. Mastra CEO Sam Bhagwat
The race to build the fastest AI infrastructure is reshaping the semiconductor industry, with inference speed emerging as the defining competitive dimension of the AI era. As AI model wars intensify a
In this post, we walk through five capabilities now available in SageMaker HyperPod inference: multi-tier data capture for auditing and model improvement, direct deployment from Hugging Face Hub, loca
Gao Yuan / Bloomberg: Survey: Chinese companies plan to allocate 46% of their AI accelerator budget to domestic products in the next 12 months, up from 30% today, a shift from Nvidia — Chinese compani
A chipmaker called Syntiant Corp. that specializes in making low-powered processors that run artificial intelligence locally on devices has filed to go public. The company filed its initial public off
Jason Hiner / The Deep View: Q&A with Doug Brooks, senior product manager of Apple silicon, about Mac minis becoming preferred AI agent machines, future of on-device AI, and more — W — alk into any of
Financial Times: BYD, Nio, and other Chinese carmakers are rushing to design and increase the use of locally developed chips with AI functions in a bid for chip self-sufficiency — EV makers such as gl
Ahmad Osman discusses the technological and practical reasons why locally-run AI models are becoming increasingly competitive with cloud-based alternatives, likely covering improvements in model effic
The financial services industry (FSI) operates under a unique set of non-negotiable requirements: the need for strict regulatory compliance, sub-millisecond transactional speeds, and security that ver
Mark Gurman / Bloomberg: Sources: Apple plans to skip higher-end M6 chips and launch its next Pro and Max chips in 2027 as part of the M7 lineup, to boost on-device AI capabilities — Apple Inc. is mak
Bloomberg: Digital advocacy firms like CiviClick and Influent appear to use AI to generate mass public comments on local energy projects, mostly favoring fossil fuel use — From California to North Car
Eleanor Olcott / Financial Times: Sources: prices for Nvidia's AI chips on China's black market have more than doubled amid a US export crackdown; its flagship DGX B300 server has risen to $1.1M — US
NVIDIA DGX Spark Enterprise Manageability enables IT administrators to plan, configure, and manage DGX Spark deployments at scale with coverage of workflows, provisioning, update procedures, and syste
Our next generation of Apple Intelligence is centered around our users, integrated deeply into our operating systems, and powered by a bold new architecture with privacy at its core. At the heart of t
Why edge AI development is still hard AI is no longer confined to cloud experiments. Developers are increasingly expected to deliver AI inside apps, devices, and edge systems where responsiveness, pri
Lauren Feiner / The Verge: Google unveils five AI data center water commitments, including becoming water positive by 2030, local infrastructure investment, and transparency about usage — The company
Paul Thurrott / Thurrott: Microsoft announces new on-device AI updates for Edge: a dev preview of a new SLM called Aion-1.0-Instruct, Language Detector and Translator APIs, and more — Tied to the earl
Tom Warren / The Verge: Microsoft announces the new Surface RTX Spark Dev Box for local AI development, powered by Nvidia's new Arm-based RTX Spark chips and 128GB of unified memory — The Surface RTX
Microsoft only just announced a new Surface Laptop Ultra at the weekend, and it's now revealing a miniature Surface PC aimed at developers. The new Surface RTX Spark Dev Box is powered by Nvidia's new
The agentic AI moment has arrived, but delivering on its promise requires more than good models. It also takes fast hardware, secure runtimes, a responsive data layer and models tuned for long-running
Anna Tong / Forbes: Thrive Holdings, a spinoff of Thrive Capital, commits $1B to acquire local accounting firms through its subsidiary, Current, and use AI to automate them — In Thrive Holdings' live-
Personal agents are exploding in popularity, with open source projects like OpenClaw and Hermes seeing rapid adoption by AI developer communities on GitHub. Built to adapt to individual preferences an
Wall Street Journal: As Google plans to build a $15B AI data center hub in India's Visakhapatnam, locals and rights groups are raising concerns over “extremely high” water stress — Developing-world na
As enterprises confront rising memory and storage costs alongside mounting pressure to run AI workloads on-premises, the fundamental architecture of the private cloud data center is changing beneath t
Enterprises are adopting an agentic AI PC strategy as Copilot+ PCs shift AI workloads from expensive cloud inference to secure, high-performance on-device processing. As agentic AI moves from experime
Most production AI features don't need a frontier model. Here's how capability evals and prompt engineering can help ship a local SLM that matches frontier-model quality with lower latency and cost. T
Written by: Jamie Collier While Russian-speaking threat actors have historically dominated the phishing-as-a-service (PhaaS) landscape, a rival ecosystem is rapidly growing within the Chinese-language
The AI PC is being fundamentally redefined as agentic workloads push the boundaries of what local compute can deliver — and as runaway cloud token costs force enterprises to rethink where inference ac
The NVIDIA GB200 NVL72 is a rack-scale GPU supercomputer leveraging Blackwell architecture with NVLink switches for high-density computing , and topology-aware block scheduling in Slurm can align larg
This article examines key emerging trends that are transforming AI infrastructure, likely covering topics such as distributed computing, edge AI deployment, GPU optimization, and evolving cloud infras
Dell Technologies Inc. today is kicking off its Dell Technologies World conference by expanding its artificial intelligence portfolio with enhancements aimed at helping enterprises move AI projects fr
Telco cloud modernization has become an urgent operational imperative for operators burdened by decades of siloed infrastructure and the demands of 5G, 6G and edge AI. A unified platform approach is n
Agentic AI is changing the way users get work done. Following the success of OpenClaw, the community is embracing new open source agentic frameworks. The latest is Hermes Agent, which crossed 140,000
This article examines how agentic AI systems are being deployed to edge computing environments, likely challenging common assumptions about what this deployment actually entails. It discusses practica
Google has offered Gemini Nano for Chrome since 2024 as a lightweight, on-device model , but users reasonably expect the visible AI Mode to use the on-device model with queries staying local, when in
Post-training quantization is a technique that reduces model size and improves inference performance by converting weights and activations to lower precision formats after training is complete. NVIDIA
Popular NoSQL-based database company MongoDB Inc. today announced a new set of capabilities during the company’s .Local conference in London, bringing together everything software and artificial intel
Google Chrome may be taking up more of your storage than expected thanks to a large on-device AI model file that, in some cases, is being automatically downloaded to the browser's system folders. User
Manus, an AI company Meta acquired for $2 billion last year is running ads promising quick, easy money with AI: Find local businesses without websites or with bad websites, have AI build them one, the
Mark Gurman / Bloomberg: Sources: Apple plans an AI overhaul for photo editing in iOS 27, including using on-device AI models to extend, enhance, and reframe photos — Apple Inc. is planning a major ov
Kim Mackrael / Wall Street Journal: ASML says it plans to make at least 60 of its standard EUV machines this year, 36% more than it sold in 2025, as it races to meet demand for making AI chips — ASML
John Higgins / The Verge: Anker announces Thus, a compute-in-memory chip it says will bring on-device AI to its products and accessories, starting with its upcoming Soundcore earbuds — The Thus chip
When we launched Microsoft Agent Framework last October, we made a promise: building production-grade AI agents should feel as natural and structured as building any other software. Today, we’re deliv
Clara Hernanz Lizarraga / Bloomberg: Big Tech companies say a $90B data center buildout in Spain's Aragón, one of Europe's fastest-growing hubs, should be an EU model, as local residents push back — B
Automated lint: 51 errors, 15 warnings, 3 info