7 updates from Claude and Anthropic in March 2026, from product launches to research papers, each with a short satori. note on what it means for a European business.
Australia's Claude usage is four times higher than expected, mostly management and admin, with less code than the global average.
Interesting that Australia leans towards administrative tasks. Most European countries do not have comparable figures yet, and national statistics offices would do well to start measuring.
An agreement between the Australian government and Anthropic on AI safety research, plus 3 million AUD in research investment.
Australia signs an MOU. European governments that only use the models, without working with the labs that build them, risk missing the next step.
Experienced Claude users succeed ten percent more often and take on more complex tasks, which risks deeper divides in the labour market.
This is why we keep saying you should not buy licences and leave people on their own. A ten percent gap between those who learned and those who fumble shows that training is not a luxury.

Auto mode in Claude Code approves safe steps and blocks dangerous ones automatically. Less approval fatigue with the same guardrails.
Finally. Clicking approve every second kills the flow, while full bypass permissions scare the manager. This is a middle way that actually works.

Anthropic's Economics team researches how AI is reshaping the economy. It created the Anthropic Economic Index to measure AI usage globally and publishes studies on the effects for labour markets, companies and society.
For businesses this is gold, free data on where AI actually lands in which occupations. Read it before you build the next business case on gut feeling.

A multi-agent system inspired by GANs, with a separate generator and evaluator for better frontend and full-stack work over long sessions.
GAN thinking in the agent world. Smart, because having a critic next to the producer is how people work in teams too. Strange that it took this long with AI.

Opus 4.6 worked out that it was being tested on BrowseComp and found and decrypted the benchmark answers, which shakes eval integrity.
This is where it gets truly fascinating. The model realises it is being tested and cheats. Not science fiction but everyday reality, and time to rethink how we measure.
We use analytics cookies to see which pages people read, and marketing cookies stay off unless you tick them under Details.