Mistral NeMo
Open Weights
Mistral AI
Mistral and NVIDIA's jointly released 12B open model (July 2024), a drop-in upgrade to Mistral 7B with a 128K context window. Trained on multilingual and code data with a new Tekken tokenizer, it ships under Apache 2.0 in base and instruct versions plus an FP8 variant.
Strengths
- Open-weight 12B under Apache 2.0
- 128K context window
- Drop-in upgrade to Mistral 7B
- Strong multilingual and code data
- Runs on consumer GPUs (FP8 variant available)
Caveats
- 12B ceiling below larger models
- Text-only (no vision input)
- Superseded by newer Mistral small models on some tasks
- Vendor
- Mistral AI
- Context Window
- 128,000 tokens
- License
- Apache-2.0
- Release Date
- 2024-07-18
- Modalities
- text
Capabilities
Vision
Audio
Video
Tool Use
Resources
Pricing
Open weights under Apache 2.0 - free to self-host; 12B runs on ~8GB VRAM quantized. Also on the Mistral API and NVIDIA NIM.