Model
1 entry · 26 November 2024
Trained on up to 5 trillion tokens, AI2 said the 7B and 13B models beat Llama 3.1 8B and Qwen 2.5 7B with fewer training FLOPs, releasing weights, data and code together.
Open weights & ecosystem · Models & capabilities