Kolibri 1 is Aleph Alpha’s open-weight German-English reasoning and tool-use model, trained from scratch with a German-focused tokenizer and data mixture. Its hybrid-attention mixture-of-experts Transformer has 78B total parameters and approximately 3.46B active per token. The native 262,144-token context can extrapolate to 1,048,576 tokens, but the publisher recommends remaining within the native window for complex tasks and serving efficiency. Apache-licensed weights support private and sovereign deployments. Sparse activation reduces computation without eliminating the memory needed for the full weights; the published knowledge cutoff is June 18, 2026 for both English and German.
Typetext
Parameters78B total, approximately 3.46B active