Skip to main content
Back to News
Aleph Alpha releases Kolibri with open weights and support for 1M-token contexts
Product
2 min read
DE

Aleph Alpha releases Kolibri with open weights and support for 1M-token contexts

The AMW Read

Kolibri meaningfully updates Aleph Alpha's foundation-model offering through a deliberate Apache 2.0 open-weight release and bilingual specialization, although segment-level impact depends on independent performance and deployment validation.
NoveltySignificance
Foundation Models · Player MapScaling Laws
Aleph Alpha
Aleph Alpha

Foundation Models / LLMs

View Company Profile

Aleph Alpha releases Kolibri with open weights and support for 1M-token contexts

Aleph Alpha has released Kolibri, an English-German model targeting government and regulated industries, with downloadable weights under Apache 2.0. The Mixture-of-Experts model has 78.1 billion total parameters and activates 3.46 billion per token. It supports tool calling, four reasoning settings, and contexts up to 1,048,576 tokens. Customers can deploy it on-premises using Aleph Alpha's inference package and a Kolibri-specific vLLM plugin. The company built the model in Germany and trained it in Germany and Finland.

The release positions Aleph Alpha in the foundation-model market around deployment control and bilingual specialization. Open weights let customers operate the model on their own hardware, while German represents 21.3% of its pre-training tokens. This gives buyers evaluating sovereign AI a concrete alternative to sending internal data to a third-party inference service. Its architecture combines sparse expert activation with sliding-window attention in 40 of 50 layers to contain serving costs. Aleph Alpha claims competitive quality relative to serving cost, but the reported math, coding, and tool-use results are vendor-run benchmarks, with the highest reasoning setting used where applicable.

For builders, the immediate task is to validate long-context reliability and deployment economics on their own workloads. Long-context adaptation reached 256,000 tokens; the one-million-token ceiling comes from the supplied serving settings, so maximum supported length should not be treated as evidence of equally reliable retrieval across that entire window. Teams should also test abstention and document grounding before relying on outputs in mission-critical workflows. The 3.46 billion active parameters describe computation per token, while the full model still has 78.1 billion parameters to accommodate in deployment planning.

#AlephAlpha #Kolibri #OpenWeights #FoundationModels #SovereignAI

#Aleph Alpha#Kolibri#open-weight models#sovereign AI

How This Connects

Based on Foundation Models · Player Map

  1. 1d agoMistral AI previews Large 4, with open weights planned for late OctoberMistral AI
  2. 1d agoMistral AI previews Large 4 with open weights planned for OctoberMistral AI
  3. 1d agoMistral Releases Large 4 Through Guardrailed Access, With Open Weights PlannedMistral
  4. 3d agoAnthropic infrastructure financing reportedly reaches $60B with Broadcom supportAnthropic
  5. 4d agoAleph Alpha releases Kolibri with open weights and support for 1M-token contexts · THIS ARTICLE
  6. 6d agoAnthropic reportedly secures up to $42B in Broadcom financing for AI infrastructureAnthropic

More news from Aleph Alpha

Stay updated with the latest news and announcements from Aleph Alpha.

View all Aleph Alpha news

Discover AI Startups

Explore 5,000+ AI companies with VC-grade analysis, funding data, and investment insights.

Explore Dashboard