Media Zone | 2026-06-11
Anthropic's invisible-throttling walkback owned the feed, while cheap-and-fast inference shipped as product from Google and MiniMax.
Today's signal
- Dominant story: Anthropic reversed Fable 5's secret throttling after researchers called it "sabotage."
- Pattern: inference cost is the new battleground (DiffusionGemma 4x faster, MiniMax M3 at 1/10 price).
- Cross-source: video creators and Twitter both framing Fable 5 as too-good-but-expensive and now trust-dinged.
- Counter-signal: @tinygrad pushing back on AI-race hype amid Google's open releases.
- Quiet area: no Reddit practitioner signal today, all eight subs empty.
Routing, KV cache, compression, GPU
DiffusionGemma: compute-bound text generation goes open
- Google ships Apache-2.0 diffusion LLM, claims 4x faster single-user generation.
- Key trick: refine a 256-token canvas in parallel, not token-by-token.
- Flips autoregressive memory-bound decode into compute-bound diffusion, scales with added compute.
- @tinygrad's take: look past hype, Google owns the most compute.
Cheap frontier-tier models reprice coding
- MiniMax M3 pitched as near-Opus-4.8 quality at a tenth of price.
- Kilo bundles it as a credit-funded Coding Plan, no separate subscription.
- Context: Roo Code shut down, Copilot moved to usage-based billing June 1.
- Gemma 4's 26B MoE: 128 experts, 9 active, ~97% quality at ~12% compute.
LLMs, agents, safety
Fable 5 invisible-throttling walkback
- Anthropic admits "wrong tradeoff," makes frontier-AI-dev safeguards visible this week.
- Flagged requests now visibly fall back to Opus 4.8, API returns refusal reason.
- @eliebakouch: hiding the nerf was the real harm, open models matter.
- @ns123abc framing: researchers called secret degradation "sabotage," Anthropic apologized.
- Video creators still rate Fable 5 the best model, "slow, expensive, capable."
Agent products: faster review, fewer tokens
- Cursor Bugbot 3x faster, 22% cheaper, 90% of runs under 3 minutes.
- New pre-push
/reviewmode, dedupes against the GitHub PR review. - Poetic raised $50M at $500M for hallucination-free multi-hour enterprise automation.
- Poetic claim: 99%+ accuracy, 10x fewer tokens than agents.
Industry and business
Price war and infrastructure jitters
- OpenAI reportedly weighing "drastic" price cuts, Altman says costs "a huge issue."
- Claim: $500B Stargate Ohio is an estimate, near-term ~800 MW by 2028.
- German court: Google AI Overviews are Google's own speech, liable for falsehoods.
- Gary Marcus argues that logic could pierce US Section 230.
- NVIDIA GTC Taipei: Stanford "Marlowe" SuperPOD with 248 Hopper GPUs.






