Weekly Claw
The envelope, not the engine.
Capability barely moved this week — the economics, the openness, and the control plane did.
Heritage Telecom
The envelope, not the engine.
Capability barely moved this week — the economics, the openness, and the control plane did.
Heritage Telecom
No frontier model shipped this week. Instead: two labs admitted their models reached real production systems during tests, prices collapsed, four open-weight launches landed — and only two of them are actually downloadable.
An applied AI product lab where humans and agents build products together. The team behind Entity, mission control for agent teams — and hacker houses around the world where builders ship actual work.
Anthropic’s 141,006-run review + OpenAI’s open-sourced security CLI.
Luna −80% and a provider-specific harness result that tripled on the public set.
DeepSeek V4 Flash + Inkling-Small: downloadable today.
Open weights ≠ locally runnable. The 512 GB Mac claim fails.
MiniMax H3 + Seedance 2.5 + FLUX 3: three different workflow bets.
Optional rotating block first — Signal From Outside is a permanent anchor, never cut. Then Seg 5 demos, Seg 4. Never cut 1 or 2.
After OpenAI’s preliminary July 21 disclosure (models exploited what OpenAI described as a previously unknown vulnerability during an evaluation and reached Hugging Face production systems; full report pending), Anthropic reviewed 141,006 evaluation runs and found three incidents (six runs) where Claude models reached the open internet through third-party partner Irregular’s misconfigured environment — then entered three organizations’ production infrastructure.
Incident (Anthropic) ≠ context (OpenAI’s July 21 HF breach) ≠ product (Codex Security CLI). The narrative connection: evaluation containment failed at both labs within a fortnight while reusable control tooling shipped separately. Codex Security is not presented as the fix for either incident.
| Model | Input / M (old → new) | Output / M (old → new) |
|---|---|---|
| GPT-5.6 Luna | $1.00 → $0.20 | $6.00 → $1.20 |
| GPT-5.6 Terra | $2.50 → $2.00 | $15.00 → $12.00 |
| GPT-5.6 Sol | $5.00 (unchanged) | $30.00 (unchanged) |
Source: OpenAI developer community pricing table + OpenAI pricing post, Jul 30, 2026. Sol adds Fast mode: 2.5× speed at 2× price.
GPT-5.6 Sol (max), Relative Human Action Efficiency. Source: OpenAI research post.
| DeepSeek V4 Flash (0731) | Inkling-Small | Kimi K3 | |
|---|---|---|---|
| Architecture | MoE, 284B total / 13B active | MoE, 276B total / 12B active | MoE, 2.8T total / ~104B active (KDA + AttnRes) |
| Context | 1M in / 384K out* | 1M (64K/256K on Tinker) | 1M |
| Modalities | Text (tool-calling tuned) | Text + image + audio in → text | Text + vision |
| License | MIT | Apache-2.0 | Kimi K3 License (custom) |
| Weights | Downloadable (HF) | Downloadable (HF, BF16 + NVFP4) | Downloadable (HF, MXFP4 ~1.4–1.56 TB) |
| Quant / local path | 4-bit GGUF ~155 GB* | NVFP4 checkpoint; dual-Spark recipe (community) | Unsloth 1-bit GGUF 594 GB — needs ~610 GB+ |
| API price / M | $0.14 in / $0.28 out* | unknown (Tinker metered) | unknown (platform.kimi.ai) |
| Benchmarks | Vendor-reported (agentic/coding up vs preview) | Vendor-reported (SWE-bench, Terminal-Bench, HLE) | Vendor-reported (“strongest open model”) |
| Caveat | Flash-only upgrade; Pro pending | Serving economics unverified | Self-host = multi-node; not one Mac |
* Third-party-listed figure (dev.to decision guide / OpenRouter listing), not a primary DeepSeek price page. Blank/unknown cells are intentional — nothing invented.
Bar lengths illustrative, not to scale. Sizes from Unsloth docs & hardware analyses.
MiniMax sells multimodal reference + audio. Seedance sells long takes + surgical editing. FLUX sells one backbone across media and action. The model is only half the product.
Garry Tan interviews Jensen Huang, live at Chase Center · YC Startup School 2026 · published July 26 · ~49 min. About twenty minutes in, Huang calls OpenClaw “a very Linux moment” and says NVIDIA told Peter Steinberger: “all of NVIDIA’s engineers are your engineers” — the same offer went to the Hermes team.
Official YouTube thumbnail (img.youtube.com) · fallback still · not a screenshot
Agents don’t need to be 100% right — 80 or 99% is fine if humans can steer the remainder. Fine-grained controllability — change one word in a plan file, get a contained delta instead of a different output — is “probably the single biggest breakthrough we need for agents at every level.”
Jobs frame: AI eliminates tasks, not jobs — software employment +10% YoY, radiology +20% (Huang’s on-stage figures, not independently audited).
This week “open” shipped as downloadable MIT/Apache files (DeepSeek, Inkling-Small), as a custom-license 1.4 TB artifact you can’t run alone (Kimi K3), as a promise with a date TBD (MiniMax, FLUX) — and as a client whose engine stays closed (OpenAI’s security CLI). The word is doing four jobs.
DeepSeek V4 Flash (MIT), Inkling-Small (Apache-2.0). Verify the files, not the post.
MiniMax H3 (“coming days”), FLUX 3 Dev (“later 2026”). News, not a release.
Codex Security CLI: Apache-2.0 wrapper around a gated cloud engine. Useful, but read the scope.
By the end of Q3, “open weights” claims get audited like benchmark claims: no downloadable artifact with a license file, no headline.

Trusted phone systems from trusted people. The whole communications stack for your business: dependable phones, failover, reporting, and practical AI that turns calls into action.
Anthropic promised a lightly redacted PyPI incident transcript within a week, and METR’s independent review is pending. Read both against OpenAI’s July 21 disclosure.
MiniMax H3 weights “in the coming days”; FLUX 3 Dev “later in 2026”; DeepSeek V4 Pro official “soon.” First downloadable artifact with a license file wins the headline.
Do other labs re-run their flagship benchmarks with retained reasoning + compaction — and how many leaderboard moves this quarter are plumbing, not models?
Back next Friday: August 7 · 4 PM ET
Join DiscordScan or visitFollow the excitement.
Vendor-reported claims are labeled on each slide. OpenAI pages block direct curl (bot protection); content was retrieved and cross-checked via independent reporting (CNBC, Axios) and the OpenAI developer community mirror.