CompanyAnthropic3 recent entries30 Apr 2026Creating highly efficient agents: 450M tool-calling tokens distilled for post-training from top open-source modelsHarnesses If you've used Claude Code or Codex, you've used a harness. A harness is the infrastructure layer that wraps an AI coding agent and decides how it operates, what it can touch, and how you me→25 Jun 2026What happens when Claude Code gets an experiment trackerAt CVPR 2026, Lambda ran a live demo for two and a half days: Claude Code teaching Google's Gemma 4 to play a Tetris-like game. Claude Code started with a Gemma 4 model that couldn't play at all. It p
CompanyOpenAI1 recent entries9 Jul 2026GLM 5.2: a new rise of open-weight agentic modelsOn June 16th, Z.ai released GLM 5.2, its latest flagship model. At the time of announcement, it advertised scores at or near Anthropic and OpenAI's models, and far ahead of GLM 5.1. In the world of us
CompanyDeepSeek2 recent entries22 May 2026DeepSeek v4: the most expected open-source model ever released, and the quietest landingAfter 15 months of incremental updates, leaks, and rumored leaks, DeepSeek released version 4. It arrived without the fanfare R1 and R1-preview commanded in early 2025. That quiet reception is the mos→9 Jul 2026GLM 5.2: a new rise of open-weight agentic modelsOn June 16th, Z.ai released GLM 5.2, its latest flagship model. At the time of announcement, it advertised scores at or near Anthropic and OpenAI's models, and far ahead of GLM 5.1. In the world of us
CompanyNVIDIA6 recent entries27 Apr 2026FlashAttention-4 gives the NVIDIA Blackwell platform its most optimized attention kernel yetOn March 5, 2026, the much-anticipated paper for FlashAttention-4 (FA4) was published. The code was dropped on GitHub months ago; early benchmarks circulated, and preliminary results were presented at→19 May 2026Lambda’s NVIDIA HGX 8xB200 on STAC-AI™ LANG6What the numbers mean for financial services Executive summary Lambda is the first to publish an audited STAC-AI™ LANG6 result on NVIDIA HGX 8xB200, with independently verified performance data that F→21 May 2026Lambda Bare Metal Instances: full hardware control with API-driven operationsThe unit of AI compute has shifted from single hosts to rack-scale systems that integrate NVIDIA GPUs, CPUs, scale-up networking fabrics, and liquid cooling, such as the NVIDIA GB300 NVL72 and NVIDIA →22 May 2026DeepSeek v4: the most expected open-source model ever released, and the quietest landingAfter 15 months of incremental updates, leaks, and rumored leaks, DeepSeek released version 4. It arrived without the fanfare R1 and R1-preview commanded in early 2025. That quiet reception is the mos→1 Jun 2026Unbox one of NVIDIA's first co-packaged optics samples with us. See why we bet on CPO early.When we design large GPU clusters, the network is no longer a background system. It's part of the compute envelope. At the 800G and NVIDIA GB300 NVL72 scale, the back-end fabric accounts for 86% of ne→12 Aug 2026Lambda prices $926 million senior secured term loan B facility, the first investment-grade-rated term loan B financing by a private neocloudFirst private neocloud to execute an investment-grade-rated financing in the term loan B market, further broadening the investor base for AI infrastructure financing. Facility supports the purchase an