Вернуться к новостям
ПродуктAI Understanding брифинг

Mistral выпускает Le Chonk, мультимодальную модель искусственного интеллекта с 1 триллионом параметров

Mistral AI представила Le Chonk, мультимодальную модель с 1 триллионом параметров, как часть своей версии Mistral Large 4, чтобы конкурировать с моделями с открытым весом.

4 min readRead the linked source
Source-page capture accompanying Mistral releases Le Chonk, a 1 trillion-parameter multimodal AI model
Ссылка на источникИсточник записан
Издатель
thehindu.com
Тип источника
Связанный источник — статус первоисточника не установлен.
КонтекстПоймите это за 60 секунд

Ключевые термины

Параметр
Изученный вес внутри модели, который влияет на ее выходные данные.
Мультимодальная модель
Модель, которая может обрабатывать или генерировать несколько типов данных, таких как текст, изображение и аудио.
Вывод
Фаза выполнения, на которой обученная модель генерирует прогнозы или выходные данные.

Что случилось

Mistral AI has officially released its latest , Mistral Large 4, internally referred to as 'Le Chonk.' The company describes the model as its most capable to date, featuring a 1 trillion- architecture with 49 billion active parameters. According to The Hindu, the open weights for this model are scheduled for release at the end of October 2026.

Mistral AI announced the release of Mistral Large 4, also known as 'Le Chonk,' on October 6, 2026. The company characterizes this as a major milestone in its product roadmap, signaling a shift toward a new generation of specialized and optimized models.

The model is defined by its 1 trillion- total size, utilizing a mixture-of-experts approach that activates 49 billion parameters per . This design is intended to provide the reasoning capabilities of a massive model while maintaining the efficiency of a smaller, more active parameter set.

While the model has been announced, the open weights are not yet available for public download. Mistral has confirmed that these weights will be made accessible to the developer community by the end of October 2026.

Подробности об источнике: thehindu.com ↗

Почему это важно

The release of Le Chonk represents a strategic move by Mistral to challenge the dominance of high-performance open-weight models, particularly those emerging from Chinese developers. By utilizing a mixture-of-experts architecture with 49 billion active parameters, Mistral aims to balance massive scale with operational efficiency. This development is significant for the open-weight ecosystem, as it provides developers with a high-capacity alternative to proprietary models, potentially shifting the competitive landscape for sovereign and specialized AI deployments. The model's performance and accessibility will be critical factors for organizations seeking to integrate large-scale multimodal capabilities without relying on closed-source infrastructure.

The introduction of Le Chonk is explicitly positioned to compete with the growing number of high-performance open-weight models originating from China. This competition is reshaping the global AI market by offering powerful, transparent alternatives to the closed-source models typically provided by major US-based tech giants.

For the broader AI industry, the release highlights the ongoing trend of 'sovereign AI,' where companies and nations prioritize the use of models that can be audited and hosted independently. Mistral's ability to scale to 1 trillion parameters while maintaining an open-weight strategy provides a significant tool for researchers and enterprises that require high-end performance without vendor lock-in.

Interactive Mechanism

Интерактивный механизм: как он на самом деле работает

Изучите технологию, лежащую в основе этой разработки, в интерактивном режиме.

Thinking Budget (Test-Time Tokens):1,024 tokens
Complex Accuracy79%Math & Code Logic
Latency3.2sTime to first full output
Inference Cost$0.0092Per query estimated
Reasoning StyleStep VerificationInternal chain depth
Active Thinking Trace:
1Deconstruct user problem into formal constraints
2Propose candidate hypotheses & step-by-step calculation
3Self-correction: Backtrack and refute subtle edge cases
4Exhaustive consistency check & final output synthesis
Core takeaway: Test-time compute fundamentally changes AI economics. Instead of only scaling during pre-training, giving reasoning models more tokens at inference time allows them to systematically solve PhD-level STEM problems.
Интерактивная проверка концепции+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

Что посмотреть дальше

The primary focus will be the official release of the model's open weights at the end of October 2026. Observers should monitor how the model performs in independent benchmarks compared to existing Chinese and Western open-weight counterparts. Additionally, the practical implications of its 49 billion active count on hardware requirements for local or private cloud deployment remain a key area for technical evaluation.

The most immediate milestone is the end-of-October release of the model's weights. Once released, the community will be able to verify Mistral's claims regarding the model's capabilities and efficiency.

Industry analysts will be watching for performance comparisons against other recent large-scale open-weight models, such as those from Reflection AI or other emerging competitors in the Chinese market. The actual utility of the model in real-world multimodal tasks will be the ultimate test of its market impact.

Сопутствующие руководства и викторины

Нашли это полезным?