Cohere Debuts Open-Weight 218B Mixture-of-Experts Machine Translation Model

Loading…

Cohere has released an open-weight 218 billion parameter Mixture-of-Experts model specifically designed for machine translation, making it one of the largest openly available models in this category. The MoE architecture means that despite the large parameter count, only a subset of parameters are active per inference pass, making deployment more compute-efficient than dense models of comparable size. This is directly relevant to developers building multilingual applications, localization pipelines, or cross-language enterprise tools who want a high-quality, self-hostable translation backbone. The open-weight release also allows fine-tuning on domain-specific parallel corpora, which is critical for technical, legal, or medical translation use cases. Cohere continues to differentiate through open releases targeting enterprise NLP workflows.