Weekly Claw · episode 25
14 Aug 2026
The model is the product no more
Watch
What this episode covers
- DeepSeek shipped the full stack in one move: open weights, the harness, and both Open Responses and Anthropic Messages API dialects under a single MIT umbrella.
- Z.ai pushed GLM-5.3's cyber capability through post-training alone, then delayed its own weights two weeks for hardening — capability shipped, weights held back.
- Qwen3.8-27B released as Apache-2.0 open weights while Gemini 3.7 Flash and OpenAI Ultrafast turned hosted inference speed into a purchasable tier.
- Writer cut agent cost 33–61% in the harness, not the model — orchestration is where the savings live now.
- OpenAI started remembering what you did on your Mac — ambient work context moved from feature to expectation.
- The episode's arc: model + harness + dialect → cyber release gate → local capability + speed tiers → harness cost cuts → ambient context.
- Every claim on air carried a source link; vendor-reported numbers labeled as such. Sponsors: Heritage Telecom and Herald Labs.
Published record
Transcript unavailable
No published transcript is recorded for this episode. We will show it here when a verified transcript is published.