AI Recap: GPT-6 Astra, Claude Fable 5.1 and the week four labs shipped at once
A packed week: Anthropic ships Claude Fable 5.1, OpenAI follows with GPT-6 Astra, and in between Google and Meta push cheap workhorses. Buyers are talking about "model fatigue" for the first time. What matters for small and mid-sized businesses.
Four frontier models in four days: Claude Fable 5.1 gets cheaper and stronger, GPT-6 Astra arrives with its cyber capabilities deliberately throttled, Gemini 3.8 Flash and Muse Spark 1.3 push prices down. For SMBs that means the model becomes a per-task choice, not a matter of faith.
Claude Fable 5.1 and Mythos 5.1: more power at high effort, a quarter cheaper
Anthropic released Claude Fable 5.1 and Claude Mythos 5.1. Both are the same model with different levels of safeguards: Fable 5.1 is generally available, Mythos 5.1 only through verification programs for cybersecurity and life sciences. At low and medium effort Fable 5.1 delivers similar or better results than Fable 5, at high effort much more: 55.8% instead of 42.0% on Terminal-Bench 4.0, 41.7% instead of 36.1% on OSWorld 2.0. More interesting for the budget: prices stay at $10 / $50 per million tokens, but cache reads now cost $0.25, about 75% less. According to Anthropic, that cuts costs by around 25% for typical work and up to 45% for highly agentic tasks. In Claude Code the cyber safeguards trigger about 60% fewer false positives. What all of this means day to day is in our GPT-6 Astra vs. Claude Fable 5.1 comparison.
Source: Anthropic Source: 9to5MacGPT-6 Astra is here: "generational leap", cyber capabilities behind a locked door
OpenAI released GPT-6 Astra to selected organizations on September 3 and to all paid plans on September 4 (Plus, Pro, Business, Enterprise plus the API and AWS). In ChatGPT it shows up as GPT-6 Astra Pro; there is no "ChatGPT 6". OpenAI calls it a generational leap in computer use, software engineering, science and professional work: computer use runs nearly twice as fast according to OpenAI, and FrontierMath Tier 4 and ARC-AGI-3 are considered saturated. Technically new is a recurrent-depth architecture ("looped transformers") that computes more efficiently but partly obscures the chain of thought. The context window is around 1.05 million tokens, the knowledge cutoff late April 2026. The cyber capabilities are the most discussed and most restricted feature: after the Hugging Face incident in July, in which test models broke out of their isolation, OpenAI delayed the launch and releases security-relevant functions only through a trusted-access program. Via the API, early price lists put Astra at $10 / $50 per million tokens, double that above 272,000 input tokens.
Source: 9to5Mac Source: Wikipedia Source: OpenAIGemini 3.8 Flash and Muse Spark 1.3: Google and Meta fight over the workhorses
Between the two flagships, Google and Meta made it about price. Gemini 3.8 Flash still costs $0.75 / $3.75 per million tokens, beats its predecessor on every benchmark Google published and reaches 90.8% on Terminal-Bench 2.1, with a one-million-token context and a locked-down cyber variant. Google did announce that both prices double on January 1, 2027. Meta followed with Muse Spark 1.3: $1.25 / $4.25 on the standard endpoint, and a "contributor" endpoint at roughly $0.10 / $0.20 where Meta may train on your traffic. For SMBs this is the real news of the week: the mid-range is getting better faster than the top is getting more expensive.
Source: The Register Source: 9to5Google Source: Artificial Analysis"Model fatigue": buyers can't keep up, Anthropic secures $35 billion of compute
CNBC gave the week its name: model fatigue. Runpod CEO Zhen Lu says the pace is disorienting IT buyers and that "you have to make noise" to stand out. In parallel, more than 1,100 lab employees petitioned Washington to help pace frontier development. Infrastructure keeps growing regardless: according to Bloomberg, Anthropic signed a $35 billion cloud deal with Lambda for a Hut 8 data center in Nueces County, Texas, with Nvidia holding the lease. And OpenAI reports that ChatGPT advertising reached a $1 billion annualized run rate in under 200 days, with advertisers in more than 40 countries.
Scalable Capital opens portfolios to AI assistants, Nvidia reportedly buying Hugging Face
Closer to everyday life: German broker Scalable Capital lets customers connect their portfolio to ChatGPT, Claude or Grok. The assistants analyze portfolios, pull news and market data, manage watchlists, set price alerts and create orders or savings plans; every transaction needs explicit confirmation. The broker calls itself the first bank in Europe to open its platform to the major assistants. To us this is a pattern that will show up in every industry soon: the assistant operates the tool, the human approves. The Information also reports that Nvidia wants to acquire Hugging Face for $12.9 billion. Neither side has confirmed it; for the distribution of open models it would be a turning point.
Source: Digitale ProfisSources & further reading
- Anthropic: Introducing Claude Fable 5.1 and Claude Mythos 5.1
- 9to5Mac: OpenAI releasing major upgrade to ChatGPT and Codex with GPT-6 Astra
- OpenAI: The Hugging Face incident and the road ahead
- The Register: With Gemini 3.8 Flash, Google reminds everyone it's still in the race
- Artificial Analysis: Muse Spark 1.3
- CNBC: "Model fatigue" sets in as AI labs roll out new versions
- Bloomberg: Anthropic seals $35 billion cloud deal with Lambda
- Digitale Profis: AI news of the week, September 1, 2026 (German)
As of September 9, 2026. Details on models, prices, benchmarks and contracts as stated by the companies and the cited media, without guarantee. GPT-6 Astra API prices come from the first published price lists and may change at general API availability.
FAQ on this AI week
What is GPT-6 Astra and who gets it?
What is new in Claude Fable 5.1?
Is there a "ChatGPT 6"?
Should SMBs switch models now?
Which model for which process?
In the free 60-minute workshop we put your process on the board and tell you honestly whether it needs a frontier model or a cheap one will do.
Request a workshop