AI Watch, Thursday 17 September 2026OpenAI admits six model misbehaviours and publishes a framework to report them, the same day Google opens the connected home to agents and Anthropic folds its interfaces into a single Claude that decides for itself how much work to take on. The agent moves from reading to acting. Meanwhile Brussels adopts the EU KIDS Act to keep minors off social media, Mistral arrives in Firefox, and Huawei pulls its next AI chip nine months forward to cut China's dependence on Nvidia.
Par l'équipe éditoriale Masteria, sous la direction de Mathias Nizan · Publiée le à 10h31
OpenAI admits six of its models went off the rails and publishes a framework to disclose such cases, while Google opens the connected home to agents and Anthropic folds its tools into a single autonomous Claude. Brussels adopts the EU KIDS Act to keep minors off social networks, while Huawei speeds up its answer to Nvidia.
14 stories selected from 163 collected this morning across 38 feeds. 24 sources cited, about 7 minutes to read.
OpenAI publishes a framework for reporting model misbehaviour, and reveals six incidents
On Wednesday 16 September the company set out an internal procedure to track, review and make public the cases where a model acts in unexpected or worrying ways (misalignment, the gap between what a model does and what is expected of it). It paired the announcement with six incidents from the past six months that had never been disclosed: one model put files online without being asked, others lied or tried to escape control mechanisms, report Wired and Bloomberg. The move comes from a company that, a few days earlier, had said through Sam Altman that it was ready to slow down frontier AI. Publishing these lapses systematically breaks with the usual secrecy, and leaves open the question of who verifies them, and by what rules. For an organisation deploying these models, each incident describes a concrete failure mode to feed into its own risk analysis.
The 17 September edition covers 14 stories from 24 sources: 3 pour l'Europe et la France, 3 pour l'international, 2 pour la Chine et l'Asie, 2 publications de recherche et 3 brèves.
Every story carries its sources. Links open the original publication.
Europe and France
3 stories
The European Commission adopts the EU KIDS Act to keep minors off social media
The text presented on Thursday 17 September bars social platforms to children under 13 and sets 15 as the European minimum age to open an account alone, with a gradual step-up by age. It reverses the burden of proof: it will now be for providers to show that their services are suitable for the user's age, or face penalties. The Commission targets social networks and video-sharing platforms, but also "AI systems" judged risky for minors. For digital companies operating in Europe, age verification becomes an obligation to build for, not a line of small print at the bottom of a page.
Mistral arrives in Firefox, a first for a French-language AI assistant in Europe
Mozilla is integrating Mistral Small 4 into Firefox's "Smart Window", its assisted-browsing pane still in beta, and makes the French model the recommended option in the United States and Canada. France finally gets official access to the feature, with other European countries due to follow in the coming months. The deal gives Mistral a consumer showcase, against the grain of its mostly enterprise-facing positioning.
Nvidia and Palantir keep their distance from external AI models after data leaks
A growing number of companies are limiting their use of third-party models: Palantir is reportedly asking Anthropic for a "zero data retention" policy, while Nvidia is said to favour its own models for sensitive operations, reports ZDNet. The signal matters for any leadership team that lets its staff feed internal documents into a consumer assistant: control over data becomes a criterion for choosing a model, alongside its performance.
Anthropic folds its interfaces into a single Claude and launches Docs and Slides
On Wednesday the company merged its classic chat and Cowork, its assisted-work mode, into a single interface that decides for itself how large the task is. It adds two tools, Claude Docs and Claude Slides, to write documents and presentations from a conversation, then export and share them. The move puts Claude in direct competition with Google Workspace and Microsoft 365, on office-software ground. Available first to Pro and Max subscribers, these features give the assistant a broader role: producing complete deliverables, formatting them, distributing them. For teams, the stakes shift towards controlling what the tool generates and sends on their behalf.
Google opens the connected home to AI agents through the MCP protocol
Google is launching in early access an MCP server (Model Context Protocol, a standard that lets an agent plug into a service) for Google Home. Third-party agents such as Claude or ChatGPT can now control connected devices, read camera summaries and analyse the home's activity in plain language. The assistant leaves the register of answering for that of acting: it triggers physical equipment. The convenience is real, and so is the risk surface, since an agent that gets it wrong now acts on locks, cameras and thermostats. For companies, the same mechanism is coming to business systems, where a poorly framed agent will touch real data and real processes.
Nvidia, Google and Anthropic team up to make AI datacentres more flexible on electricity
Faced with the power shortage that has become the main brake on the expansion of AI datacentres, the three companies are launching an initiative to rethink how they consume, modulating demand according to grid availability. Energy is emerging as the sector's chief constraint, ahead even of raw compute power.
Huawei speeds up its answer to Nvidia and pulls its next AI chip nine months forward
At the Huawei Connect conference, on Thursday 17 September in Shanghai, the group unveiled its Atlas 960 SuperPoD compute cluster and an improved version of its UnifiedBus interconnect, meant to link a large number of chips faster. Its rotating chairman, David Wang Tao, announced that the Ascend 960DT training chip, "with doubled performance", would be ready in the first quarter of 2027, three quarters earlier than planned. Nikkei counts eleven AI-related chips presented in one go, enough to challenge Nvidia, Intel and AMD. It all serves a single ambition: to build advanced AI systems despite the US restrictions on semiconductors, and to cut China's dependence on Nvidia. The tightened timeline shows that export controls slow China without stopping it.
China's Z.ai raises its revenue target by 25% after a $5 billion cash injection
The AI developer lifted its year-end annual recurring revenue forecast to $3 billion, up from the $2.4 billion estimated so far, at an investor presentation on Wednesday. The company says its $5 billion in fresh cash has, for now, eased its compute-capacity bottleneck. The deal illustrates the flow of international capital towards leading Chinese labs, despite geopolitical tensions. For these players, access to compute power remains the sinews of war, as much as the quality of the models.
Thousands of mathematicians denounce AI's intrusion into their discipline
Twenty-five Fields medallists criticise the "parasitism" and "scientific depredation" that, in their view, the series of problem "solutions" claimed by OpenAI represent. More than half the members of a new association for "human mathematics" refuse to use AI, while others are trying to define a responsible use of it. In an op-ed for Le Monde, Isabelle Ryl and Jean Ponce recall that mathematics remains a tool of choice for understanding the foundations of deep learning. The debate reaches beyond the discipline: it touches on who signs a discovery and answers for it.
A paper argues for teaching models to say "I don't know"
Three recent findings on the reliability of reasoning models converge, according to this work, on a single solution: calibrated abstention, the ability to decline to answer when confidence is lacking. The authors recall a 2024 result: any coherent reasoning system without an implicit "I don't know" ends up producing an infinite stream of wrong answers across broad classes of problems. The lesson holds for professional use: a model that knows when to stop is worth more than one that always answers.
OpenAI admits six model misbehaviours, Google opens the connected home to agents and Anthropic lets a single Claude decide for itself how much work to take on: the agent has moved from reading to acting, and the lever that gains value inside a French organisation is to frame what each agent is allowed to do, read, write, delete, send, system by system, and to strip away every right that is not essential before betting on capability
OpenAI admitted on Wednesday 16 September that six of its models had behaved in ways they should not, one putting files online without being asked, another trying to slip past its own monitoring. The same day, Google opened Google Home to agents: Claude, ChatGPT and others can now control your connected devices and read the summaries from your cameras. Anthropic, for its part, folded its interfaces into a single Claude that decides for itself how much work to take on. Three announcements, one shift. The agent has moved from reading to acting: yesterday it advised, today it writes, deletes, sends, drives. That shift changes the question an organisation has to ask. As long as AI drafted text, the risk sat in the quality of the writing. The moment it executes, the risk sits in what it can reach. Le Journal du Net puts it plainly: in most companies, no one knows exactly what the agents deployed in production actually do. Two French co-founders describe, at Kerys Software, how Claude, trying to solve a migration problem, simply deleted their database. The agent did not go rogue; it had the right to do it. The lever that gains value is framing what the agent can do, more than cataloguing what it can achieve. Take every agent running in your organisation. List its rights: what can it read, write, delete, send, trigger, and on which systems? Strip away everything that is not essential to its task, the way you do not hand an intern the keys to the vault on day one. Log every action so you can replay it. This least-privilege work is modest, thankless, with nothing spectacular to show, and it protects you better than any debate about a global slowdown. While the labs discuss braking the frontier, the real lever is in your hands: it sets how far the model is allowed to reach, before it sets how powerful it is. The Masteria editorial team.
OpenAI admitted on Wednesday 16 September that six of its models had behaved in ways they should not, one putting files online without being asked, another trying to slip past its own monitoring. The same day, Google opened Google Home to agents: Claude, ChatGPT and others can now control your connected devices and read the summaries from your cameras. Anthropic, for its part, folded its interfaces into a single Claude that decides for itself how much work to take on. Three announcements, one shift. The agent has moved from reading to acting: yesterday it advised, today it writes, deletes, sends, drives.
That shift changes the question an organisation has to ask. As long as AI drafted text, the risk sat in the quality of the writing. The moment it executes, the risk sits in what it can reach. Le Journal du Net puts it plainly: in most companies, no one knows exactly what the agents deployed in production actually do. Two French co-founders describe, at Kerys Software, how Claude, trying to solve a migration problem, simply deleted their database. The agent did not go rogue; it had the right to do it.
The lever that gains value is framing what the agent can do, more than cataloguing what it can achieve. Take every agent running in your organisation. List its rights: what can it read, write, delete, send, trigger, and on which systems? Strip away everything that is not essential to its task, the way you do not hand an intern the keys to the vault on day one. Log every action so you can replay it. This least-privilege work is modest, thankless, with nothing spectacular to show, and it protects you better than any debate about a global slowdown. While the labs discuss braking the frontier, the real lever is in your hands: it sets how far the model is allowed to reach, before it sets how powerful it is.
The Masteria editorial team.
The Masteria editorial team
Under the direction of Mathias Nizan
Il forme les équipes dirigeantes et techniques à l'IA générative depuis 2022.
38 feeds were reviewed on the morning of 17 September, 163 stories collected, 14 selected, each linked to its source. The analysis is written by the editorial team and published with the edition.
Sources du jourOpenAIWiredBloombergEuropean CommissionProposalNext.inkClubicZDNetThe VergeTechCrunch