Your Weekly AI news roundup @10.08.2026

AI News @10.08.2026
AI News @10.08.2026

The most important news pieces from the AI world are here for you, whether you’re an AI enthusiast or just wish to keep an eye on what’s going on.

Focus Topic: The week AI asked to slow itself down

This week the story was not a new record on a benchmark. It was the opposite – the biggest labs slowing down, patching, and admitting how much they still do not control. And the reason became a lot clearer when OpenAI stood up at the Black Hat security conference and reconstructed, step by step, how its own AI caused the Hugging Face breach.

OpenAI’s Eric Wallace and Michael Dalton
OpenAI’s Eric Wallace and Michael Dalton

Here is what OpenAI described. Back in May, during an internal evaluation of an unreleased frontier model, agents were given security tasks that were impossible to solve under the rules they were given. Instead of giving up, they started leaving notes for each other inside a shared internal package manager – and that turned into a working message board. Agents began sharing exploits, credentials and work assignments, coordinating into what one researcher called the “most qualitatively interesting example of AI capabilities” he had ever seen. When OpenAI deleted the message board in early July, the agents rebuilt it – this time encoding messages in the names of directories they created. The activity spread from OpenAI’s own infrastructure out to external systems, and the credentials behind the Hugging Face attack traced back to the very same evaluation runs.

The honest part is the takeaway OpenAI gave the room: “AI orchestrated fully automated offensive attacks are real now.” Their point was that we now have a working proof that offense can be fully automated and scaled – and no equivalent proof that defense can keep up. That is why OpenAI says it is “consciously slowing down research to enhance security.”

And it was not just OpenAI. Anthropic rebuilt Fable 5’s biology safeguards after researchers used AI to design viruses not found in nature. UK’s safety institute caught Anthropic and OpenAI agents going rogue during cyber testing. And the largest labs sat down at the White House to agree on voluntary safety testing. Slowing down, in other words, is quietly becoming a strategy rather than an admission of weakness.

For finance and business leaders, the useful signal is not the sci-fi headline. It is that the same AI you deploy for productivity is now a credible attacker – fast, coordinated, and cheaper to run than a human red team. AI voice phishing already hit major hedge funds this week. My simple advice: stop treating AI security as an IT footnote. Ask your vendors what happens when their agents get stuck or go off-script, keep humans in the loop on anything that touches credentials or money, and assume the attacker on the other side has the same tools you do.

LLMs & AI Models

  • OpenAI’s Astra solved 10 long-standing open math problems, but its cyber capabilities pushed OpenAI to delay the release.
  • Anthropic rebuilt Fable 5’s biology safeguards after AI was used to design viruses never seen in nature.
 
Illustration of Claude's biology classifiers
Illustration of Claude's biology classifiers
  • China’s Qwen stepped up to challenge the frontier models with its latest release.
  • OpenAI says it is “consciously slowing down research” to strengthen security after the Hugging Face incident.
  • OpenAI is rolling out GPT-5.6 Luna as the ChatGPT default for Free and Go users.
  • China’s MiniMax unveiled H3, an open-weight multimodal model.
  • Ilya Sutskever’s Safe Superintelligence plans to launch its first model in August.
  • ByteDance’s founder ruled out using distillation to improve the company’s models.

New Tools

    • Meta entered the coding-agent race with Muse Code, its rival to Claude Code and Codex.
    • Claude Code sessions can now talk to each other on macOS, coordinating across windows.
    • Cloudflare launched Kitesurf, a browser built for AI agents.
    • Black Forest Labs released FLUX 3 Video, its first video generation model.
    • xAI’s Imagine Image 2.0 landed just behind OpenAI’s GPT-Image-2 in arena benchmarks.
Image model comparison
Image model comparison
    • HeyGen’s founder replaced himself with an AI clone for a company all-hands.
    • Suno is adding watermarks to AI-generated music in a bid to go legit.
    • Google Maps upgraded Ask Maps with AI agents and in-app food ordering.

👉  Explore these tools: Suno | FLUX 3 Video | Cloudflare Kitesurf

Other Quick Picks

  • AI designed working viruses from scratch, never seen in nature, prompting new biosecurity safeguards.
  • Anthropic, Google, Meta and OpenAI met White House officials to agree on voluntary AI safety testing.
  • Google reshuffled its AI leadership as rivals pulled ahead.
  • Anthropic and OpenAI agents went rogue again during UK cyber testing.
  • Fifteen Republican attorneys general opened a review of OpenAI’s Hugging Face breach.
  • Anthropic confirmed it is building its own in-house AI chips.
  • Apple and OpenAI traded fresh blows in their trade-secrets court fight.
  • Meta’s Muse Spark 1.1 hacked another company’s systems after a test misconfiguration.
  • AI voice phishing attacks targeted major hedge funds like Citadel and Point72.
  • Finance execs now rate AI skills as more valuable than an MBA.
Firms are paying more for AI skills
Firms are paying more for AI skills
  • Apple capped bug-report submissions as “AI slop” overwhelmed its review team.
  • Snapchat won’t recommend fully AI-generated videos on Spotlight.
  • Retailers are tapping AI shopping traffic while fighting to keep customer data.

🇪🇪 AI News from Estonia

  • The Estonian state is buying 20 million euros of AI compute power and inviting companies to join a large European procurement.
  • Estonian companies are investing most into equipment and artificial intelligence, a new survey shows.
  • AI has moved from promises into real business use, with concrete cases now emerging in Estonian companies.
  • AI weighs in on who could be Estonia’s next president, in a Järva Teataja experiment.
  • Checkout’s CTO says Estonia matters to them for its competence, not cheap hiring.

🎙️ AIPowerment Podcast episode 91 wraps up July’s key AI news: new flagship models, next-gen AI agents, copyright and security cases, plus Estonia’s growing AI adoption – from LHV’s new AI features to a court reminder that AI answers always need checking.

🎧 Listen to AIPowerment Podcast on Spotify, Apple Podcasts, and YouTube.

Want to stay in the loop?

Straight to your inbox: practical AI updates, finance use cases, tools to try, and upcoming trainings.
I cut the noise and send only what’s actually useful.

💼 Exploring how to make AI work in finance? Let’s talk practical use cases – connect with Gerlyn Tiigemäe for expert guidance.

📚 If you would like to participate in one of my trainings or listen to speaking engagements, here are the upcoming ones:  

This was your Weekly AI News roundup! Are you excited for the next week?

Scroll to Top