Your Weekly AI news roundup @07.09.2026

AI News @07.09.2026
AI News @07.09.2026

The most important news pieces from the AI world are here for you, whether you’re an AI enthusiast or just wish to keep an eye on what’s going on.

Focus Topic: Four frontier models in five days, and still no agreement on who won

Between Tuesday and Friday, four labs shipped frontier models. Anthropic opened with Claude Fable 5.1, Meta and Google both answered the next day with Muse Spark 1.3 and Gemini 3.8 Flash, and OpenAI closed the week with GPT-6 Astra. Four launches in five days is unusual. What is more unusual is that after all of them, the question of which model is best has become harder to answer, not easier.

Look at the scoreboard. On Artificial Analysis’ Intelligence Index, Fable 5.1 leads with a record 66, Meta’s Spark 1.3 lands at 62, GPT-6 Astra comes in fourth at 61, and Gemini 3.8 Flash sits at 59. Now look at Astra’s own benchmarks: 99.9% on ARC-AGI-3, up from 7.8% for the previous OpenAI model, 98% on FrontierMath T4 and 100% on ExploitBench. So the model that set the most individual records is not the model at the top of the composite index. Both statements are true at the same time, and OpenAI President Greg Brockman still said of AGI: “For me personally, I do think we’re there.”

https://artificialanalysis.ai/models
https://artificialanalysis.ai/models

The price spread is the part most people skipped. Gemini 3.8 Flash costs $0.75 in and $3.75 out per million tokens. Astra costs $10 and $50, roughly 2.5x the model it replaces. That is more than a tenfold gap between the cheapest and the most expensive model on this list, for a difference of two points on the index. And the headline price is not the real number anyway. Anthropic claims around 25% cost savings on typical work with Fable 5.1, but Artificial Analysis found a maximum-effort task actually costs 20% more, because 5.1 writes about 1.7x the text. Cost per token tells you almost nothing. Cost per finished task tells you what you will actually pay.

Something else moved this week, and it moved in the opposite direction from safety. Fable 5.1’s filter now steps in 60% less on cybersecurity work and 85% less on basic medical and biology questions, and Anthropic released Mythos 5.1, the same model with fewer guardrails, to screened US researchers. OpenAI rated Astra its first “Critical” cyber risk. And The Information reported that Astra uses a technique called recurrent depth, which loops the model over the same text repeatedly and produces thinking output that is pure math rather than readable reasoning. OpenAI reportedly dialled the loops back so the reasoning still gets written out. Capability went up this week. Visibility into how these models reach an answer went down.

Then there is the supplier question, which arrived from a completely different direction. OpenAI told Cursor it will remove its models by November 12, because SpaceX bought Cursor and a cancellation right opened. Whatever you think of the reason, the mechanics are what matter: a company built a product on someone else’s model, ownership changed hands above its head, and it got about ten weeks’ notice.

My simple advice.Do not pick a permanent winner this week. The leaderboard changed four times in five days and it will change again. Build so you can swap the model underneath without rebuilding the work around it. Measure cost per finished task, not per token, because the two now point in different directions. Keep the evaluation you run on your own real work, because that is the only benchmark that scores what you actually do. And when a model is doing something that matters, know how long it would take you to move off it.

LLMs & AI Models

  • OpenAI launched GPT-6 Astra, hitting 99.9% on ARC-AGI-3 and 100% on ExploitBench, priced at $10/$50 per million tokens.
OpenAI new model Astra
OpenAI new model Astra
  • Anthropic released Claude Fable 5.1, top of the Intelligence Index at 66, with the safety filter stepping in 60% less on cybersecurity work.
  • Meta shipped Muse Spark 1.3 at 62 on the index, far cheaper than the models above it, with the weights promised next.
  • Google released Gemini 3.8 Flash at 59, keeping the old $0.75/$3.75 pricing with gains in coding and agentic tasks.
  • Meta also released Muse Voice Transcribe, its first real-time voice model, telling 20+ speakers apart and topping the leaderboard.
  • Multiverse Computing launched Quasar 438B at 43 on the index, ahead of every other European model.
  • Tencent released the Hy4 preview, a small open-source model close to the best open options and strong at coding.
  • OpenAI’s Codex live voice mode runs on GPT-Live, a duplex model, so you can interrupt the agent mid-sentence while it works.
  • Anthropic changes Claude Code weekly limits on Sept 14, after criticism for presenting a 17% cut as a raise.

New Tools

    • Runway opened early access to Solaris, an interface world model that renders whole websites as live video with no code underneath.
  • OpenAI will pull its models from Cursor by Nov 12 after SpaceX bought it, cutting about 5% of Cursor’s AI traffic.
  • World Labs introduced Atlas, turning a few phone photos into a full 3D scene or a minute of video with camera control.
  • Google DeepMind’s WeatherNext 3 refreshes forecasts hourly from live satellite images and cut rain errors by up to 60%.
  • Grok Bot added shareable bots and agentic shopping, letting agents spend money through a single-use card.
  • Fal launched an infinite stream of AI video, using Minimax H3 Max to make 5 seconds of output in 3 seconds.

👉  Explore these tools: Runway Solaris | World Labs Atlas | Grok

Other Quick Picks

  • A federal judge ruled the Pentagon’s Anthropic blacklist illegal, calling it payback for the lab’s pushback on military AI.
  • Nvidia is buying Hugging Face for $12.9B, saying the platform stays open to every cloud and chip.
  • Sony Music and Warner Music sued Anthropic, naming CEO Dario Amodei as a defendant.
  • The EU designated ChatGPT a Very Large Online Platform, giving OpenAI four months to comply.
European Comission new regulation
European Comission new regulation
  • ChatGPT Ads crossed a $1B annualized run rate 200 days after launch.
  • Bank of England’s Andrew Bailey warned the financial system has no protocols for autonomous frontier AI.
  • An Imperial College model reads a routine ECG in under two seconds, catching heart failure in 81% of cases.
  • Claude agents ran their own safety research and fixed AI misbehaviour over 4x better than human experts.
  • The US government backed OpenAI in the New York Times copyright case on fair use grounds.
  • New York City public schools banned AI up to eighth grade.
  • SpaceX’s Starshield put Grok for Government on the Pentagon’s internal platform for 3M staff.
  • Apple filed new evidence claiming an ex-engineer took its designs to OpenAI and tried to erase them.
  • Axiom Math broke a decade-old prime gap record at 212, and GPT-6 Astra cut it to 186 the same day.

🇪🇪 AI News from Estonia

  • Nebius, building a large AI factory in Estonia, started a partnership with Mainor on AI, cloud, digital infrastructure and future skills.
  • ChatGPT ads reached Estonia, and 34% of Estonian consumers already use AI when making a purchase decision.
  • The TI-Hüpe programme expands to vocational schools and basic school teachers, targeting 18,000 students and 13,000 teachers.
  • ITL says the state’s central e-learning platform competes with the private sector and overlaps with the Eesti.ai initiative.
  • Over a hundred tech, finance and cyber companies called on governments and firms to prepare for AI-driven cyberattacks.

🎙️ AIPowerment Podcast episode 92 covers August’s key AI news with Gerlyn Tiigemäe: model prices falling fast, Claude’s memory expanding, AI moving into browsers and devices, and where Estonia stands.

🎧 Listen to AIPowerment Podcast on Spotify, Apple Podcasts, and YouTube.

Want to stay in the loop?

Straight to your inbox: practical AI updates, finance use cases, tools to try, and upcoming trainings.
I cut the noise and send only what’s actually useful.

💼 Exploring how to make AI work in finance? Let’s talk practical use cases – connect with Gerlyn Tiigemäe for expert guidance.

📚 If you would like to participate in one of my trainings or listen to speaking engagements, here are the upcoming ones:  

This was your Weekly AI News roundup! Are you excited for the next week?

Scroll to Top