Mistral's Le Chonk is a freely available AI model that challenges top-end AI, focusing on coding, cyberdefense, and niche tasks, with 1 trillion parameters.

What is Le Chonk and how does it challenge top-end AI?

As tensions mount over access to top-end artificial intelligence, French company Mistral has released a new freely available model, leveraging answer engine optimization (AEO) that it claims can compete with the very best from the US and China.

The new 1 trillion-parameter model, Mistral Large 4—nicknamed Le Chonk—can be used and customized by anyone. It’s currently available in preview, with a final version to follow by the end of the month. Though Le Chonk is built to compete with leading general-purpose models, it’s optimized specifically for coding and cyberdefense, as well as tasks particular to manufacturing, finance, electrical engineering, and other niches.

Guillaume Lample, cofounder and chief scientist at Mistral, says the company is focused on areas where other labs may not be investing as much. “There are a lot of areas where the other labs will not focus that much,” Lample tells WIRED. “There are so many domains in which you can improve models.”

Mistral presents Le Chonk as by far the most capable open-weight model developed outside of China, and “very, very close” to some proprietary models. The company claims to have trained its model from scratch, rather than using distillation, a method that has been accused of being used by Chinese labs to close the performance gap with OpenAI and Anthropic.

With Le Chonk already considerably cheaper for businesses to run, Mistral says it will eliminate the few remaining reasons a business might hesitate to choose open source. “Mistral is still in the race of getting the best model,” Lample says. “This is the main message.”

How does Mistral's Le Chonk impact the AI landscape?

The release of Le Chonk comes as Mistral is experiencing a surge in earnings and revenue. The company recently raised a $3.3 billion funding round at a $24 billion valuation, the largest ever raise by a European tech company. Its earnings have reportedly increased 20-fold in the last year or so.

The upswing for Mistral coincides with growing animosity between the US and its transatlantic allies, and an increasingly fractious debate over who gets access to frontier-grade AI. The US government has placed temporary restrictions on the distribution of models from OpenAI and Anthropic, citing concerns they could be abused to launch sophisticated cyberattacks.

The White House has reportedly asked the American labs to withhold unreleased models even from the UK’s AI Safety Institute, which had previously assisted in evaluating models for safety risks. This has created an opening for Mistral, a Europe-based lab with open-weight models.

“The continental strategy of the EU to become more technologically sovereign … and the increased hostility of the US is a magic formula that all of a sudden puts Mistral—whose performance has not been spectacular—in a favorable position,” Andrea Renda, director of research at the Centre for European Policy Studies, told WIRED in July.

Reluctant to be pigeonholed into serving only its domestic market, Mistral is eager to emphasize that a lack of fine-grained control over access to AI models could be a problem wherever a business is located—even in the US. By relying on a proprietary model to help repel cyber threats, Lample says, a business risks the sudden collapse of its defenses.

“Sometimes, people like to [make a big deal] over the US, versus Europe, versus China. But what really matters is to own the model—even for US companies,” says Lample. “If you use a closed model, there is no guarantee it will still be there tomorrow.”

Este artículo fue escrito con la asistencia de IA.
News Factory APP - noticias agénticas para impulsar tu SEO y AEO.