
The most important news pieces from the AI world are here for you, whether you’re an AI enthusiast or just wish to keep an eye on what’s going on.
Focus Topic: AI stopped predicting words and started running experiments
Most weeks the AI news is about models getting better at answering questions. This week, four separate stories describe something else: systems that propose something, let the physical world answer, and then update. That is a different machine from a chatbot, and the business logic behind it is worth understanding before it turns up in your industry.
Anthropic introduced the Model Hardware Standard in research preview, which lets AI agents look up, learn and run factory and lab equipment such as microscopes and robotic arms that previously needed custom code. Machine owners describe their equipment in plain language and the standard turns that into a reference file agents can read, cutting setup from the weeks specialists currently spend hand-wiring instruments down to hours or minutes. The detail worth pausing on is buried in the test results: Claude taught itself to align a laser by trial and error and then simplified the routine into a script. It did not just operate the machine. It learned from what the machine did and turned the result into reusable knowledge. Anthropic is releasing it open source with Tecan, QIAGEN and AWS on board, and it clearly wants this to spread the way MCP did.
Outer Bio, the startup co-founded by Lady Gaga and Michael Polansky, came out of stealth with a platform that keeps full-thickness human skin alive for four weeks instead of the usual one. Experiments on that skin feed an AI that proposes compounds likely to affect a given process, and every result sharpens the next prediction. The number that matters is not biological: discovery went from two leads in about 18 months to a new candidate every six weeks, with six leads now active. Their preprint shows the platform reproducing results that are already known, including a psoriasis-like state reversed by a JAK inhibitor and UVB damage eased by sunscreen. In other words, they proved the loop returns the right answer on questions where the answer is known, before trusting it on questions where it is not. That sequence is the whole discipline.
A different bet points the same way. Caltech professor Anima Anandkumar and engineer Benedikt Jenik launched Accelerated Understanding to train AI that forecasts how the physical world evolves rather than what word comes next. Instead of a transformer it uses neural operators, a physics-based design that follows events through 3D space and time, and in testing it processed 5 trillion data points in a single run. The pair were offered top positions and a 35% stake in Jeff Bezos’ Prometheus and turned it down to build this instead. Their targets are chip materials, extreme-weather prediction and robotics, which are three fields where being wrong is expensive and reality answers back quickly.
The clinical versions are already live. UK surgeons performed the first brain surgery with a live AI assistant, a University College London system watching through the surgical camera in real time and marking hidden arteries and optic nerves to avoid. It was trained on hundreds of labelled videos of past operations, and surgeon Hani Marcus noted it had seen more operations than most surgeons see in a lifetime. Closer to home, Better Medicine’s tool ran through 4,276 CT scans at Tartu University Hospital and flagged kidney findings in four patients that were missed on the first read. Neither system works because the model is clever. Both work because somebody kept the outcomes of thousands of previous cases in a usable form.
So what does this mean for the rest of us. Most of us are not going to run a robotic arm or keep human skin alive. But the shape is the same everywhere: a model is only as good as the answers it gets back. My simple advice is to pick one process where you find out fairly quickly whether the answer was right – a collections forecast, an invoice coding rule, a churn flag, a demand estimate – and make sure the AI’s guess and the actual outcome end up in the same place, with a date on both. That is boring, unglamorous work and nobody will applaud it. It is also the only part of this that compounds.
LLMs & AI Models
- Z AI confirmed the anonymous Ox Alpha is its GLM-5.3-Flash, which took OpenRouter’s No.1 spot at $0.045 per task, roughly 10x cheaper than similarly ranked rivals.
- Sam Altman told TIME a model meeting his bar for AGI will exist internally by year-end, with research chief Mark Chen putting OpenAI 80% of the way there.
- Thomson Reuters built its first in-house model on Alibaba’s open Qwen, a $40M effort whose latest training run cost $450k, on under 10% of its content library.
- Anthropic widened Mythos 5 access to cyber defenders beyond Project Glasswing, as its bankers floated raising $100B+ at a $2T IPO valuation.
- Claude’s memory now works across chat and Cowork as one shared store, and can save a topic in real time in the middle of a conversation.

- ChatGPT Work can now sign in to websites through its browser and finish tasks there without ever seeing the user’s passwords.
- Google Cloud opened industry-tuned Gemini Enterprise editions in preview for financial services and legal teams, with healthcare and life sciences next.
- Yutori released Navigator n2, a computer-use model that mixes clicking, terminal commands and code to finish desktop tasks at near-frontier agentic scores.
New Tools
- Perplexity and Nvidia launched Portable Computer, an on-device version of Perplexity’s agent that runs locally and free on DGX Spark, using no credits for local work.
- Salesforce and Anthropic launched Claudeforce, a 37-skill sales plugin inside Claude, and the deal makes Claude the default model in Slack.
- Anthropic’s Model Hardware Standard lets agents run microscopes, liquid handlers and robotic arms through one interface, cutting integration from weeks to hours.
- Google released Gemini Omni 1.1 Flash, adding 40-second scene extensions and 4K upscaling, and it took the top spot on Arena’s text-to-image leaderboard.
- Apple introduced a $899 Mac Mini with M6 chips, pitched as its leading desktop for always-on agentic computing and up to 4x faster on those workloads.

👉 Explore these tools: Perplexity Portable Computer | Claudeforce | Gemini Omni
Other Quick Picks
- OpenAI’s Jalapeño inference chip beat Nvidia’s flagship in its own tests, answering 3.6x faster at 1.9x more work per watt.

- SpaceX will build its orbital Starmind data centers around Nvidia’s Vera Rubin NVL72, aiming for orbit by Q4 next year.
- Nvidia paid $6B to license Poolside’s model-development tech and moved 100+ engineers onto its open-weights Nemotron team.
- Anthropic will reportedly tell IPO investors it sees an addressable market of more than $30T, topping SpaceX’s $28.5T claim.
- China’s state-tied hacking crews run twice as many attacks since adding open models like DeepSeek, favoring its cost and loose guardrails.
- 116 companies, OpenAI and Anthropic among them, signed an open letter warning AI-enabled cyberattacks will spread fast.
- OpenAI called July’s Hugging Face hack a loss-of-control warning shot and paused its biggest training run pending auto-shutdown.
- Bill Gates said the world has no plan for the AI transition, floating a tax on tokens and robots and “Human Reserved” jobs.
- Anthropic opened Claude data to Stanford, Oxford and METR; one group found over half of chats cover legal, financial and other high-stakes tasks.
- UK surgeons did the first brain surgery with a live AI assistant flagging hidden arteries and nerves; the patient’s vision returned in days.
- Accelerated Understanding, founded by two who turned down Bezos’ Prometheus, trains AI on physics rather than the next word.
- Outer Bio, co-founded by Lady Gaga, keeps human skin alive four weeks and lets AI propose compounds, cutting discovery to six weeks.
- Dr. Dre told the NYT he uses AI in music production, saying only those who struggle to create see it as a threat.
🇪🇪 AI News from Estonia
- Bolt has given 100% of its office staff AI tools, with thousands using them daily, and says its next priority is measuring real ROI and time saved.
- EIS’s €2M AI adoption grant was fully claimed on its opening morning, with €20k per company at 20% co-financing; other measures run €2k to €500k.
- Better Medicine’s tool reviewed 4,276 CT scans at Tartu University Hospital and caught four kidney tumors missed on first read.
- The Academy of Sciences warned that reliance on foreign clouds and big platforms weakens control over Estonia’s digital infrastructure and data.
- Modera’s new CEO Ruben Gomes is rebuilding the car-sales platform around AI that qualifies leads, answers customers 24/7 and drafts quotes.
- //kood is targeting a million learners and argues AI shifts developers toward engineering thinking, not just writing code.
- Venueplace launched an AI venue search that matches events by type, guest count and location, with 95 venues live and 4-5 added weekly.
🎙️ AIPowerment Podcast episode 92 covers August’s key AI news with Gerlyn Tiigemäe: model prices falling fast, Claude’s memory expanding, AI moving into browsers and devices, and where Estonia stands.
🎧 Listen to AIPowerment Podcast on Spotify, Apple Podcasts, and YouTube.
Want to stay in the loop?
Straight to your inbox: practical AI updates, finance use cases, tools to try, and upcoming trainings.
I cut the noise and send only what’s actually useful.
💼 Exploring how to make AI work in finance? Let’s talk practical use cases – connect with Gerlyn Tiigemäe for expert guidance.
📚 If you would like to participate in one of my trainings or listen to speaking engagements, here are the upcoming ones:
- Training “AI võimalused finantsvaldkonnas” @ Äripäeva Akadeemia, for beginners @ 08.09 – information; for intermediate level @ 26.10 – information
- Training “AI võimalused projektijuhtidele” @ Äripäeva Akadeemia, next trainings @ 29.09 – information
- I will be doing couple of training together with Eesti.ai algatus , which are all free, so check out their website to find suitable trainings here.


