Aya Expanse

Open-weight multilingual models from Cohere Labs at 8B and 32B parameters covering 23 languages.

Open SourceSelf HostedOffline CapableGPU Required
0.0 (0)

About

Aya Expanse comes out of Cohere Labs' Aya initiative, an open multilingual research effort with thousands of collaborators, and ships open weights at 8B and 32B parameters covering 23 languages including Arabic, Chinese, Hindi, Japanese, Korean, and Vietnamese. Context windows differ by size: 128K tokens on the 32B model and 8K on the 8B. The models combine several research contributions: data arbitrage for sourcing synthetic multilingual training data, multilingual preference training, safety tuning, and model merging, applied on top of Cohere's Command family architecture. Cohere reports the 32B model winning head-to-head multilingual comparisons against larger open models including Llama 3.1 70B, Mixtral 8x22B, and Gemma 2 27B. Weights are on Hugging Face under CC-BY-NC-4.0 plus Cohere Lab's Acceptable Use Policy, so commercial use is off the table, but research, prototyping, and evaluation are permitted. There is no companion GitHub repository; the models load through standard Transformers tooling and quantized community builds. The wider Aya family now also spans Aya Vision and the compact Tiny Aya models.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Intermediate (3/5)
License
CC-BY-NC-4.0
Added
Jul 29, 2026

Related Tools

Featured

Lightweight open-weight LLM by Google available in 1B to 27B sizes.

Open SourceSelf HostedOfflineGPU 4GB+
Easy
0.0 (0)

Open-source code LLM family by IBM for enterprise code generation.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)

Open-weight LLM by Meta in 8B and 70B sizes with strong general capabilities.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Featured

High-performance open-weight MoE LLM with 671B total parameters.

Open SourceSelf HostedOfflineGPU 24GB+
Advanced
0.0 (0)

Hybrid SSM-Transformer model by AI21 Labs combining Mamba with attention layers.

Open SourceSelf HostedOfflineGPU 12GB+
Intermediate
0.0 (0)

Open-weight code LLM trained on 2 trillion tokens of code and natural language.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Browse all Large Language Models (LLMs) tools