Skip to content

Gemma, Qwen: the wave of open models keeps rising

A Google event announced for 20 August, Qwen 3.8 weights being published. Openness has become the sector's normal rhythm.

Advertisement
The essentials in 30 seconds ⚡
Google's Gemma team has announced a special event for 20 August, sparking speculation about a new version. In parallel, Alibaba's Qwen 3.8 weights are being released, honouring the promise made at the model's announcement. Openness has settled in as the industry's normal pace rather than an exception.

We have been tracking the rise of open-weight models for months. This week confirms that the movement has stabilised, which is itself news.

What is confirmed and what is not

Let's be precise about the status of the information.

Confirmed: the Gemma team has announced an event for 20 August. That is all that is official.

Speculation: the community hopes for a new version, with expectations around unified audio input, better tool-calling reliability, and higher-quality compression from launch. These expectations are publicly voiced wishes, not leaks. What the event will actually contain remains unknown.

In progress: the release of Qwen 3.8 weights, announced during the model's presentation.

What the community really wants

The requests being made reveal how mature this ecosystem has become. They no longer focus on raw power but on operational reliability.

Tool calling. This is the capability that lets a model trigger an external action, at the heart of how agents work. A model that gets its call format wrong breaks the whole chain. Users are reporting persistent flaws on this front across several open models.

Quality compression from launch. A compressed version carefully produced upfront yields far better results than compression applied after the fact. For those running these models locally, it is decisive, as we explained in our article on quantization.

Unified audio. Being able to process speech directly, without going through an intermediate transcription, across all model sizes.

What these requests tell us 🔧
Two years ago, discussions around open models were about benchmarks. Today, they are about tool-calling reliability and the quality of compressed versions. That is the sign of an ecosystem where people are actually building things, not just testing them. The same shift we noted regarding agents entering production: the boring concerns are the ones serious users have.

The state of the open ecosystem

In a few months, we have covered the release or announcement of open models by Chinese and American players, and now by a state with the Department of Energy's science programme.

The landscape has stratified usefully. Very large models, such as Kimi K3 or Qwen 3.8-Max, target institutions capable of hosting them. Mid-sized models, of which Gemma is one, aim at those who want to run something decent on accessible hardware. It is this second category that delivers real democratisation, as shown by Meta's agentic model fitting on a single graphics card.

What to take away

The news is not that another open model is coming out. It is that no one is surprised anymore.

Eighteen months ago, releasing the weights of a quality model was a remarkable move. It has become an expected checkbox, to the point that a lab that did not do it would have to justify itself. This normalisation is arguably the most structural shift of the year, and it has a direct consequence: the question is no longer whether you will have access to a good model, but what you will do with it.

Advertisement