Masteria, Centre de formation IA certifié Qualiopi
Certifié QualiopiFinançable OPCOFrance · Suisse · Belgique
Français
Édition du Wednesday 26 August 2026

Open-weight models pass proprietary ones on Vercel
AI Watch, Wednesday 26 August 2026

Par l'équipe éditoriale Masteria, sous la direction de Mathias Nizan · Publiée le à 9h02

OpenAI unveils Jalapeño, its first in-house inference chip, rated faster and more power-efficient than the best hardware on the market. The same day, open models overtake proprietary ones on the Vercel developer platform, DeepSeek moves toward a stock listing valuing it at $74 billion, and Apple puts on sale a 512-gigabyte Mac Studio built for local AI.

14 stories selected from 123 collected this morning across 38 feeds. 20 sources cited, about 7 minutes to read.

14 stories20 sources7 min read
Top story
Top story

OpenAI unveils Jalapeño, its first in-house inference chip, and rates it ahead of the best hardware on the market

Jalapeño is a chip designed by OpenAI for inference (running an already-trained model so it produces its answers). On SemiAnalysis's independent InferenceX benchmark, it delivers both more tokens per user and more throughput per kilowatt than the best hardware available, according to the tests reported by TechCrunch and The Verge. Richard Ho, OpenAI's head of hardware, claims "the best of both worlds," lower latency and higher throughput at the same time. The announcement comes with a piece by chief financial officer Sarah Friar on "the full stack behind abundant intelligence": OpenAI wants to own the entire chain, from chips to product, to bring down the cost of each answer. Designing its own inference silicon is a bid to loosen its dependence on Nvidia and take back control of the bill weighing on its margins.

Today's detail

Stories from 26 August

The 26 August edition covers 14 stories from 20 sources: 3 pour l'Europe et la France, 4 pour l'international, 3 pour la Chine et l'Asie, 1 publication de recherche et 2 brèves.

Every story carries its sources. Links open the original publication.

Europe and France

3 stories

Microsoft gets the green light for an AI data centre in Mulhouse

The American company will be able to build new infrastructure dedicated to AI and cloud in the Haut-Rhin, Clubic reports. The project does not have unanimous local support, at a time when the water and electricity consumption of these facilities is already fuelling sharp tensions. Every siting of this kind plants a concrete question in France for local authorities and neighbouring businesses: which resources for what territorial benefit.

SourceClubic

LinkedIn reports more than a million AI contents flagged by its members

The button to report a post judged AI-generated, opened in July, has been activated more than a million times, says Hari Srinivasan, the network's head of product, quoted by Siècle Digital and Clubic. LinkedIn now adds a notice in post analytics to warn the author concerned. The figure measures a fatigue: professionals are rejecting mass-produced text, including on the network that pushed assisted writing hardest.

Nscale prepares a $3 billion listing to fund its AI infrastructure

The specialist in cloud infrastructure dedicated to AI is aiming for this raise in an offering planned for September, ZDNet reports. The move illustrates the rise of the "neoclouds," those specialised operators that stand against the established cloud giants to rent out compute. An IPO of this size signals that investors keep betting heavily on compute demand, at a time when hardware costs are climbing.

SourceZDNet

International

4 stories

Repris dans l'analyse

Open-weight models overtake proprietary models on the Vercel platform

On Vercel's AI Gateway, a widely used web development service, open models (which you download and run yourself) accounted for 54% of the token volume processed on Tuesday, against 46% for proprietary models, the South China Morning Post reports from the platform's public data. The share of open models even reached a record 62% on one given day. DeepSeek's most recent lightweight model is driving this shift. In the same move, corporate demand for Fable 5, Anthropic's most powerful model, is stalling because of its high price. The trade-off between top performance and acceptable cost is thus being replayed in the field, request by request, and it is tilting toward open. This is the subject of our analysis.

Repris dans l'analyse

Apple puts on sale a Mac Studio with 512 GB of memory, built for local AI

Two weeks before its autumn keynote, Apple unveiled by surprise new Mac mini and Mac Studio models and two chips, the M6 and the M5 Ultra, Numerama and ZDNet report. The M5 Ultra Mac Studio version goes up to 512 gigabytes of unified memory and supports eight external displays; Apple claims AI performance four times that of the M3 Ultra. Prices start at $2,499 for the M5 Max and $5,499 for the M5 Ultra, available on 22 September. That much memory targets a precise use: running large models locally, on a desktop, without going through the cloud.

OpenAI loses a key data centre executive as the run of departures continues

Chris Malone, who helped lead the construction of OpenAI's data centres, has left the company, Bloomberg and TechCrunch report. Asked about the departure, the company says it "recently reorganised" its infrastructure organisation "to support the scale and pace" of its work. Malone joins a list of high-profile departures in recent months, at the very moment OpenAI is committing tens of billions to its compute capacity. These moves raise a question of continuity: building data centres at this speed requires a stable team in charge, and the company is losing its people one after another.

Hexaware's boss expects a 25% drop in the value of IT services under the effect of AI

AI could cut the price of routine IT projects by up to 25%, the head of the Indian digital services group Hexaware told Bloomberg. The observation targets a sector built on labour-intensive tasks, where automating part of the coding and testing compresses the bill charged to the client. For providers, value shifts from the volume of hours sold toward the ability to deliver faster at equal quality. For French companies that outsource their development, it is a signal to renegotiate: yesterday's rate no longer holds if AI does part of the work.

SourceBloomberg

China and Asia

3 stories

Repris dans l'analyse

DeepSeek nears a funding round valuing it at $74 billion ahead of a stock listing

The Chinese lab is finalising a raise that values it at about 500 billion yuan, or $74 billion, pre-money, the South China Morning Post reports. DeepSeek was seeking to raise nearly 50 billion yuan in this round, expected to close before the end of August, notably from its existing shareholders. The move sets up a listing targeted for 2027 on Shanghai's Star Market, China's technology exchange. The timing is no accident: DeepSeek's lightweight model is the one pushing open models ahead of proprietary models on Western platforms. A Chinese lab funded at this level, whose models spread free of charge internationally, is redrawing the map of value in AI.

Xiaomi doubles down on in-house chips despite falling profit

Xiaomi is pursuing its bet on proprietary silicon to prepare for complex AI workloads on smartphones and cars, the South China Morning Post reports, as its profit declines and component costs soar. On Monday the group presented its Xring O3 AI processor, built on a 3-nanometre process, due in September aboard the Xiaomi 18 Fold, along with a Xring O100 on 6 nanometres. Designing its own chips is expensive and weighs on margins in the short term, but reduces dependence on American suppliers subject to export controls. The approach matches that of several Chinese giants: bring compute in-house to last over time.

Alibaba raises $10.2 billion for its AI infrastructure

The Chinese giant issued 710 million shares at $14.4 apiece in Hong Kong, the largest deal of its kind ever carried out by a company on the exchange, Siècle Digital reports. The full amount will fund its AI infrastructure, while its net profit fell 75% in the last quarter. In the same move, founder Jack Ma and the group's executives bought back several hundred million dollars of shares, a signal of confidence sent to the market after a week of decline in the stock.

Research

Research and papers

Pour les équipes techniques

An MIT Media Lab study shows that getting your news through a chatbot degrades judgement

Participants who spent four weeks assessing news headlines and images, presented in pairs, were 21% more accurate at the start at telling true from false; the regular use of a conversational assistant as a source of information reduced that accuracy, the MIT Technology Review sums up from the work of Pattie Maes and her colleagues. The mechanism is familiar to researchers: delegating the sorting of information to a machine lowers cognitive effort, and critical attention dulls. For the training and change-management professions, the result invites treating the chatbot as a synthesis tool to be checked, not as a source to take on trust. The stakes go beyond general knowledge: a team that loses the habit of checking makes worse decisions.

The rest of the news

In brief

  • OpenAI restores the five-hour limit on Codex and ChatGPT Work

    Six weeks after removing it, OpenAI is reactivating for Plus subscribers a quota calculated over a five-hour window, on top of the weekly cap, Numerama reports.

    SourceNumerama
  • A study estimates that a third of new web pages are written by AI

    More than a third of the pages created are said to be produced with the help of AI, according to results relayed by Siècle Digital, a sign of the speed at which generated text is filling the web.

The Masteria read

Open-weight models pass proprietary ones on Vercel, DeepSeek heads for a $74 billion listing, and Apple sells a 512 GB Mac Studio

On Vercel, a widely used web development platform, open-weight models (which you download and run yourself) overtook proprietary models on Tuesday: 54% of the token volume processed against 46%, with a peak at 62%. DeepSeek's lightweight model is driving this shift, while corporate demand for Fable 5, Anthropic's most powerful model, stalls on its price. The same day, DeepSeek is preparing a raise that values it at $74 billion ahead of a Shanghai listing, and Apple puts on sale a Mac Studio with 512 gigabytes of unified memory, enough to run a full large open model on a desktop computer. Three signals, one direction: the open layer, cheap and self-hostable, is becoming the default choice for real production work. The skill gaining value in a French team is the ability to run open models in-house: take one production use case this quarter, deploy an open model on your server or a local machine, measure it on your own set of cases against the API you pay for today, compare quality, latency and cost per request, then decide use case by use case between hosting and renting. Many organisations find that half of their calls do not need the most expensive model. The technical frontier keeps advancing; autonomy is built now, one use case at a time.

On Vercel, a widely used web development platform, open-weight models, the ones you download and run yourself, overtook proprietary models on Tuesday: 54% of the token volume processed against 46%, with a peak at 62% on one day. DeepSeek's lightweight model is driving this shift, while corporate demand for Fable 5, Anthropic's most powerful model, stalls on its price. The same day, DeepSeek is preparing a raise that values it at $74 billion ahead of a Shanghai listing, and Apple puts on sale a Mac Studio with 512 gigabytes of unified memory, enough to run a full large open model on a desktop computer. Three signals, one direction: the open layer, cheap and self-hostable, is becoming the default choice for real production work.

The usual story pits closed labs, sole holders of performance, against second-rate open models. In the field, sufficient performance costs less and less and is shifting toward models you install yourself. For a French team, the advantage goes to whoever can pick the right open model for each use and run it on their own premises, on their server or on a machine sitting in the office.

The skill gaining value is the ability to run open models in-house. Take one production use case this quarter, deploy an open model on it, and measure it on your own set of cases against the API you pay for today. Compare quality, latency and cost per request, then decide use case by use case: host or rent. Many organisations then find that half of their calls do not need the most expensive model. The technical frontier keeps advancing; autonomy is built now, one use case at a time.

The Masteria editorial team

Mathias Nizan, fondateur de Masteria
The Masteria editorial team
Under the direction of Mathias Nizan

Il forme les équipes dirigeantes et techniques à l'IA générative depuis 2022.

Son parcours
Method

Comment cette édition a été produite

38 feeds were reviewed on the morning of 26 August, 123 stories collected, 14 selected, each linked to its source. The analysis is written by the editorial team and published with the edition.

Sources du jourOpenAITechCrunchThe VergeClubicSiècle DigitalZDNetSouth China Morning PostNumeramaBloombergMIT Technology Review

What this changes for your teams

Masteria forme dirigeants, chefs de projet, développeurs et juristes sur l'IA générative, et développe les solutions qui vont avec.

Parler de votre projet

Réponse sous 24 h · Organisme certifié Qualiopi · Lyon, France, Suisse, Belgique