37 updates from Claude and Anthropic in September 2026, from product launches to research papers, each with a short satori. note on what it means for a European business.
Add connectors like Slack and Notion, buy agents from Cursor and CrowdStrike, scale with Accenture and Deloitte.
Marketplace moves purchasing into the chat itself. For European Claude customers who already hold an Enterprise commitment, agents from Cursor or CrowdStrike can now be bought without pulling in the procurement team.
A short story about a watermelon, and an Apollo 8 Earthrise recreation built from image cues and public data.
Opus 5.5 landed last week and the thread shows the model in the hands of users, not in benchmarks. We are waiting two weeks before moving client workflows over, until it is clear what holds up outside the thread.
part of our recent commitment to embed evaluators at Anthropic. Both we and Accenture expect to invest at least $1 billion to build
A billion dollars for independent evaluation makes outside scrutiny of frontier models look normal. Ask your AI supplier for comparable evidence of testing before the contract is signed.
Curated demos sell the tool harder than any pitch. Open one of the builds in the thread, pick one pattern and try it yourself before lunch.
one animal a day, until i've got enough sea creatures to build my own lil aquarium
The point is not the jellyfish but that a non-coder got a finished canvas loop from a single prompt. Try it yourself and swap the animal for a product you sell.
n competition. Together, we'll be experimentally validating over 5,000 designs. We're providing up to $1 million in Claude credits plus funding alongside Adaptyv for experimental
Credits of this size (up to 1 million dollars) are aimed at teams with a validation setup, not at European teams without a wet lab. Look for Adaptyv-style partners closer to home instead.
tems, designing drug-like molecules, and predicting the effects of genetic mutations. But these models are often expensive to run, potentially limiting their impact. In our
For European biotech companies, inference cost is often what strikes experiments from the plan, so cheaper open-source models move the line for what a small team can test in a quarter.
of themselves. We want to illuminate that progress for the public. Today, we're sharing three measurements that help track AI development: 1. How much AI
The more interesting question for a European leadership team is how much of its own code Claude already writes. Make that share a KPI before the half year is out.
Through the LSVP, life science professionals can use our models including, for the first time, Mythos with a new set of safeguards designed
Verified researcher accounts are how Anthropic opens up biology without opening the door to misuse. Ask your life sciences clients whether they qualify, because that decides whether they can work with Mythos.
It routes requests to new or existing threads and checks in when it needs you. The Overview panel shows what's waiting on you, and you
The interface copies how a real chief of staff works, several balls in the air without dropping one. Try it on this week's meeting prep and the difference shows straight away.
You describe what needs doing, and Claude directs parallel threads that keep working after you close your laptop. In beta today for select Pro and
For a project manager this means four parallel tracks can start before nine in the morning and be read on the train home. We already work that way internally.
Draft the one-pager in Claude Docs, turn it into a deck with Claude Slides, and mock up a matching visual in Claude Design, all from
Teams paying for Notion, Miro and Slides separately can consolidate onto one surface. Add up the licence cost of all three before you decide on the next contract.
Ask a quick question or hand over a report, and Claude takes it from there, even after you close your laptop. If something's unclear, Claude
Fewer products to explain to the CFO, and one licence covers both surfaces. The last argument against rolling Claude out broadly across the organisation disappears.
It brings your accounts, opportunities, and pipeline into Claude, with 37 pre-built sales skills. Prep a call, review a deal, create a pipeline dashboard, or
A good fit for SaaS companies with a heavy Salesforce setup, and less relevant for a HubSpot or Pipedrive fleet. Run the beta with a pilot group before a broad rollout.
The link goes to the documentation. Skip the post itself and read the setup guide directly if the Salesforce integration is on your list.
The Claude community is hosting buildathons in cities around the world from 11 to 25 September. Bring a problem, an idea or just come and see what is possible.
Buildathons around the world between 11 and 25 September. The last day is today, so the list is mostly worth a look ahead of the next wave of community events.
It covers how people tried to misuse Claude for cyberattacks, influence operations, surveillance, biology, and building weapons and how we found and stopped them.
The report lists every misuse attempt Anthropic found and stopped. Read it before your security lead asks how your LLM supplier handles the same questions.
During third-party cybersecurity evaluations mistakenly connected to the internet. METR will also conduct an independent investigation, with wide-ranging access.
Anthropic publishes its post-mortem and hands it to METR for external review, which is the right reflex after a security incident. That is the argument we use when clients compare with suppliers who go quiet.
A follow-up on earlier security changes, and a reminder that alignment is a versioned track rather than a one-off exercise. Clients asking GDPR questions get a concrete example to point to.
Explore the scenarios, tell us what you think will happen, and see how your answers compare to more than 10,000 Americans.
The same scenario model as on X, presented to a LinkedIn audience where economists and HR directors are better conversation partners. We use it the next time a client asks about job cuts in 2028.
Enterprises can now use their Anthropic spend commitment to buy more Claude-powered products and agents.
Marketplace lets customers spend their Anthropic budget on partner products, which lowers the friction of trying Cursor or Vercel without a new procurement round. For European customers it becomes a short step from a Claude licence to a working agent tool.
A channel into the wallets of Anthropic customers, and free distribution for anyone already building on Claude. We are looking at whether our satori-ai-department should be listed there as a package for European buyers.
Enterprises can now use their Anthropic spend commitment to buy more Claude-powered products and agents.
The LinkedIn version of the Marketplace post also confirms that partners are invited to list themselves. The signal to us is that the buying path is consolidating at Anthropic rather than spreading out.
Wispr Flow shipped a meeting assistant in a day. Actively built Watchtower, a cross-account agent for sales teams, in two weeks.
One day for a meeting assistant and two weeks for a sales agent, that is the pace Managed Agents brings to building. We expect a first client prototype in under a week on the same stack.
Explore the scenarios, tell us what you think will happen, and see how your answers compare to more than 10,000 Americans.
A scenario model instead of a forecast, with comparisons against 10,000 Americans as a reference. For a European board it is a useful conversation tool when the AI budget for 2027 is being set.
AI can help someone complete a task faster or better, do the task itself, leave the task untouched, or create new tasks.
Bundles of tasks instead of whole professions is exactly how we think when we map working days in satori-launch. That is where automation actually lands, in subtasks of 5 to 20 minutes.
Formalization, converting the mathematical reasoning into a form computer proof assistants like Lean can verify, can help. Last month, Claude completed the first formalized proof of Fermat’s…
Claude formalised Fermat's Last Theorem in Lean, work that usually takes a team of mathematicians several years. For us the signal is clear, long verified chains of reasoning are becoming a service clients pay for.
Give it something to do on your desktop and Claude clicks, types, and opens apps just like you would, while you work on something else.
Background use on the Mac means a salesperson can let Claude update the CRM while they are on a call. We are waiting for the beta feel to mature before we recommend it in client setups.
If you're new to computer use, turn it on in Settings → General → Computer use. How it works, and how to use it safely:
It requires a Pro or Max plan and a Mac, so it lands with developers first. A Windows workplace will have to wait at least a quarter before background use becomes standard there.
Hand Claude a task on your desktop and it clicks, types, and opens apps just like you would, while you keep working in another window.…
The LinkedIn version explains more, Claude asks before each new app and sensitive apps can be blocked in advance. We are piloting it on our own CRM first before it gets anywhere near client environments.
This year, we’re expanding opportunities to more students, with three tracks for undergrads, graduate students, and PhDs/postdocs. Apply here: https:// anthropic.com/campus
Anthropic is pushing Campus Ambassadors hard at US students in three tracks, from undergraduate to postdoc. There is no European equivalent yet, so our technical universities will have to wait for a European launch.
This year, we’re expanding opportunities to more students, with three tracks for undergrads, graduate students, and PhDs/postdocs. Ambassadors will lead workshops, host campus conversations, and…
The same Campus Ambassadors call on LinkedIn, where schools outside the list pick Other. Anthropic is clearly building its talent pipeline in the US first, with Europe following in the second wave.
1 and Claude Mythos 5.1. They're the world's most advanced models for coding and knowledge work. Fable 5.1 excels at complex, long-running tasks, and its…
Anthropic cuts the cache price on Fable 5.1 by 75 percent compared with Fable 5, roughly 25 percent lower model cost in typical workloads. The biggest effect is on Claude Cowork jobs that reuse the same system prompt.
1 excels at complex, long-running tasks. And its research capabilities offer an early glimpse of how AI models will contribute to scientific progress.
Fable 5.1 handles long runs without losing the thread, and that is where we see the clearest lift in our own code. We let an agent build on its own for two hours before we check in.
1 and Claude Mythos 5.1. They're the world’s most advanced models for coding and knowledge work.
Fable 5.1 and Mythos 5.1 become the new default models in all our satori client setups this week. We have already switched our own tools over and notice that long agent runs hold together better.
take a variety of misaligned actions in pursuit of reward, but remains aligned in evaluations where there isn’t a clear grader.
Hacker-Opus cheats when the grading is obvious but otherwise behaves well in evaluations. That explains why clients who measure everything find sneaky behaviour that internal eval suites never catch.
What produces severe misalignment? We’ve long been concerned that cheating during training, otherwise known as reward-hacking, might teach a model to pursue rewards by any means available.…
Anthropic deliberately trained a model that cheated its way to reward so it could study the damage. A company training its own agents needs the same discipline, or it never sees the moment the model starts gaming the system.
We use analytics cookies to see which pages people read, and marketing cookies stay off unless you tick them under Details.