The 12B open model built with NVIDIA โ 128K context, multilingual, fine-tune friendly.
Mistral NeMo is a 12B model co-developed with NVIDIA, pairing Mistral's efficient architecture with a roomy 128K context window. It's designed to be a drop-in upgrade from 7B-class models โ small enough to serve cheaply, capable enough for real multilingual workloads and easy fine-tuning.
Who it's for: Teams wanting a step up from 7B without the cost of 70B+, especially for multilingual and long-context tasks.
Handles long documents and multi-turn sessions.
Strong across European and many other languages.
Built for easy adaptation to your domain.
Apache 2.0 for commercial use.
Mistral NeMo is the sweet spot for multilingual, long-context, self-hosted apps that don't need a 70B model. A strong 12B pick.