Weekly Claw · episode 25

14 Aug 2026

The model is the product no more

Watch

What this episode covers

  • DeepSeek shipped the full stack in one move: open weights, the harness, and both Open Responses and Anthropic Messages API dialects under a single MIT umbrella.
  • Z.ai pushed GLM-5.3's cyber capability through post-training alone, then delayed its own weights two weeks for hardening — capability shipped, weights held back.
  • Qwen3.8-27B released as Apache-2.0 open weights while Gemini 3.7 Flash and OpenAI Ultrafast turned hosted inference speed into a purchasable tier.
  • Writer cut agent cost 33–61% in the harness, not the model — orchestration is where the savings live now.
  • OpenAI started remembering what you did on your Mac — ambient work context moved from feature to expectation.
  • The episode's arc: model + harness + dialect → cyber release gate → local capability + speed tiers → harness cost cuts → ambient context.
  • Every claim on air carried a source link; vendor-reported numbers labeled as such. Sponsors: Heritage Telecom and Herald Labs.

Published record

Transcript unavailable

No published transcript is recorded for this episode. We will show it here when a verified transcript is published.