Model
1 entry · 14 June 2024
The 236B-parameter mixture-of-experts model scored 90.2% on HumanEval, edging out GPT-4-Turbo's 88.2%, while running with only 21B parameters active per token.
Open weights & ecosystem · Models & capabilities