Model
1 entry · 22 August 2024
Two sizes — a 94B and a 12B active-parameter mixture-of-experts model — built on AI21's Mamba-Transformer hybrid, both offering a 256K-token context window.
Models & capabilities · Open weights & ecosystem