Claude Fable 5.1 shipped on September 1 at the same $10 / $50 per MTok as Fable 5. The one price that moved is the cached one: prompt cache reads are $0.25 per MTok, which is 0.025x the input rate where every other Claude model charges 0.1x, and a quarter of what Fable 5 charged. Anthropic puts the net effect at roughly 25% cheaper for typical workloads and up to 45% for agentic work that re-reads a cached prefix every turn. Cache writes did not move ($12.50 at 5m, $20 at 1h), so the saving lands entirely on the re-read. The window is 1M tokens as both default and maximum with flat per-token pricing across the whole range, 128K max output, adaptive thinking always-on at a default effort of high, and a June 2026 knowledge cutoff. Retirement is no sooner than September 1, 2027.

GPT-6 Astra landed two days later at exactly the same headline price: $10 in, $50 out. The two diverge in the fine print, and in the same direction. Astra charges $1 per MTok for cached input where Fable charges $0.25, and adds a 2x input and 1.5x output surcharge above 272K tokens. On a 200K prefix re-read across fifty turns -- 10M cached tokens, all of it under Astra's threshold -- that is $2.50 against $10.00.

The other big one is corporate. Nvidia is acquiring Hugging Face for $12.93 billion -- signed September 2, announced September 3, about $11.9B to stockholders plus up to ~$1.0B in employee retention equity per the Form 8-K. It buys 18M+ developers, 3M+ models, 500K datasets and 200K+ companies, and Nvidia says the platform stays open to the whole ecosystem. Close is expected in the first half of 2027, subject to regulators. Chinese open-weight models currently lead Hugging Face downloads.


Agent Skills became an MCP extension

From us

The model week

Elsewhere