Introducing Small 3, our most efficient and versatile model yet! Pre-trained and instructed version, Apache 2.0, 24B, 81% MMLU, 150 tok/s. No synthetic data so great base for anything reasoning - happy building! https://t.co/8WfgzuuwRk

1 min read Original article ↗

Introducing Small 3, our most efficient and versatile model yet! Pre-trained and instructed version, Apache 2.0, 24B, 81% MMLU, 150 tok/s. No synthetic data so great base for anything reasoning - happy building!

Mistral Small 3 | Mistral AI

From mistral.ai

2:16 PM · Jan 30, 2025629.4KViews