Retour aux Actualités
ProduitBriefing AI Understanding

Aleph Alpha lance Kolibri, un modèle d'IA à poids ouvert 78B

Aleph Alpha a publié Kolibri, un modèle d'IA ouvert de 78,1 milliards de paramètres avec une fenêtre contextuelle d'un million de jetons, formé en Europe en mettant l'accent sur la langue allemande et les secteurs à forte conformité.

4 min readRead the linked source
Source-provided image accompanying Aleph Alpha releases Kolibri, a 78B open-weight AI model
Référence sourceSource enregistrée
Éditeur
cryptobriefing.com
Lien source
cryptobriefing.comhttps://cryptobriefing.com/aleph-alpha-releases-kolibri-ai-model/
Type de source
Source liée : le statut de source principale n'a pas été établi.
ContexteComprenez cela en 60 secondes

Commencez ici

Termes clés

Poids
Une valeur numérique apprise qui met à l'échelle les signaux transitant par un réseau neuronal.
API (interface de programmation d'applications)
Une manière structurée permettant à un système logiciel d'envoyer des requêtes et de recevoir des réponses d'un autre système.
Mélange d'experts (MoE)
Une architecture avec des sous-réseaux spécialisés où seuls des experts sélectionnés s'exécutent par entrée.
Testez-vousQuiz sur les modèles d'IA expliqués

Que s'est-il passé

Aleph Alpha, a Heidelberg-based AI company, released Kolibri, a 78.1-billion-parameter Mixture-of-Experts (MoE) transformer model. The model features a context window of up to 1,048,576 tokens and is available under the Apache 2.0 license on Hugging Face. It was trained on infrastructure in Germany and Finland using 768 NVIDIA B200 GPUs, with a specific emphasis on German content and compliance for sectors like public administration and defense.

Aleph Alpha released Kolibri, a 78.1-billion-parameter AI model with open weights. The model is a Mixture-of-Experts (MoE) transformer, meaning it has 78.1 billion total parameters but only approximately 3.46 billion are active per token. This architecture is designed to balance computational efficiency with model capacity.

A key feature of Kolibri is its context window, which supports up to 1,048,576 tokens (approximately 1 million). The model was trained on 20 to 24 trillion tokens, with a heavy emphasis on German content. It is tuned for both English and German, targeting sectors where compliance is critical, including public administration, industry, aerospace, and defense.

The training process took place on infrastructure located in Germany and Finland, utilizing 768 NVIDIA B200 GPUs. This European-based training aligns with the company's pitch to organizations concerned with data sovereignty and regional regulatory compliance.

Kolibri is distributed under the Apache 2.0 license and is available on Hugging Face. Aleph Alpha also published a 189-page technical report detailing the model's architecture, training methods, and evaluation results. According to the company's own reporting, Kolibri scored 96.9% on the AIME 2025 math competition benchmark.

Détails de la source: cryptobriefing.com ↗

Pourquoi c'est important

The release of Kolibri represents a significant development in the European AI landscape, offering a high-capacity, open- model that addresses specific regional and regulatory needs. By training on European infrastructure and prioritizing German language capabilities, Aleph Alpha provides an alternative to US-centric models for organizations with strict data sovereignty and compliance requirements. The 1-million-token context window and MoE architecture, which activates only 3.46 billion parameters per token, suggest a design optimized for efficiency and long-context tasks, potentially lowering inference costs for complex applications.

The release of Kolibri provides a substantial open- alternative for European organizations that require high-performance AI models with strong German language capabilities. The emphasis on compliance and training on European infrastructure addresses specific concerns about data sovereignty and regulatory adherence, which are often barriers to adopting US-based AI models in sensitive sectors.

The 1-million-token context window is a significant technical specification that allows the model to process and reason over very large documents or datasets in a single pass. This capability is particularly useful for legal, financial, and administrative tasks that involve extensive documentation.

The MoE architecture, with only 3.46 billion active parameters per token, suggests that Kolibri may offer more efficient inference compared to dense models of similar total parameter size. This could translate to lower operational costs for organizations deploying the model at scale, although independent verification of these efficiency claims is not provided in the source.

As an open- model under the Apache 2.0 license, Kolibri allows for broad adoption and modification, fostering a more competitive and diverse AI ecosystem in Europe. This contrasts with proprietary models and supports the broader trend of open-source AI development.

Interactive Mechanism

Mécanisme interactif : comment cela fonctionne réellement

Explorez de manière interactive la technologie sous-jacente à ce développement.

Thinking Budget (Test-Time Tokens):1,024 tokens
Complex Accuracy79%Math & Code Logic
Latency3.2sTime to first full output
Inference Cost$0.0092Per query estimated
Reasoning StyleStep VerificationInternal chain depth
Active Thinking Trace:
1Deconstruct user problem into formal constraints
2Propose candidate hypotheses & step-by-step calculation
3Self-correction: Backtrack and refute subtle edge cases
4Exhaustive consistency check & final output synthesis
Core takeaway: Test-time compute fundamentally changes AI economics. Instead of only scaling during pre-training, giving reasoning models more tokens at inference time allows them to systematically solve PhD-level STEM problems.
Vérification de concept interactive+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

Que regarder ensuite

Independent benchmarks comparing Kolibri's performance against other open- models, particularly in German language tasks and long-context reasoning. Adoption rates among European public sector and defense entities. Clarification on access conditions and pricing for API services, as the source only confirms open-weight availability on Hugging Face.

Independent evaluations of Kolibri's performance, particularly in German language tasks and long-context reasoning, will be crucial for validating Aleph Alpha's claims. The source only provides the company's self-reported benchmark results, such as the 96.9% score on AIME 2025.

Adoption by European public sector and defense entities will indicate the model's practical impact and acceptance in compliance-heavy industries. Monitoring for official deployments or partnerships in these sectors will provide concrete evidence of its utility.

Clarification on access conditions and pricing for any API services or hosted versions of Kolibri is needed. The source confirms open- availability on Hugging Face but does not specify if there are paid tiers or enterprise support options.

The publication of the 189-page technical report will likely lead to further analysis by the AI community, potentially revealing insights into the model's training data, architectural choices, and limitations that are not covered in the initial release announcement.

Guides et quiz associés

Modèles d'IA expliquésTransformateursÉthique de l'IATestez ce que vous savez : essayez un quiz gratuit sur l'IARecherchez un terme d'IA dans notre glossaireSuivez le suivi des versions du modèle AI
Vous avez trouvé cela utile ?