gekro
GitHub LinkedIn
News

AI News

Aleph Alpha releases Kolibri open-weight MoE; Google tightens Gemini access

Aleph Alpha released Kolibri, a 78.1B-parameter open-weight MoE model with 3.46B active parameters; Google restricts free Gemini access to its weakest tier.

1 min read 2 sources

Aleph Alpha has released Kolibri, a 78.1B-parameter English-German Mixture-of-Experts model that activates only 3.46B parameters per token (MarkTechPost). The model supports a 1M-token context window and per-request reasoning effort, with Apache 2.0 licensed FP8 weights capable of running on a single B200 or H200 GPU (MarkTechPost). The release targets practitioners seeking efficiency gains through conditional computation on inference-constrained hardware.

Google has restructured its Gemini access tiers, restricting free users to Flash-Lite, the smallest available model, while Flash and Pro models are reserved for paid subscribers (The Decoder). The change, effective in October 2026, may signal preparation for launching Gemini 4 Argon, a more resource-intensive model (The Decoder).

Compiled automatically from the linked sources and published without manual editing - a neutral summary of third-party reporting, for information only. Every claim links to its origin. Not original reporting.