Skip to main content
llm.info

Mistral NeMo

Open Weights

Mistral AI

Mistral and NVIDIA's jointly released 12B open model (July 2024), a drop-in upgrade to Mistral 7B with a 128K context window. Trained on multilingual and code data with a new Tekken tokenizer, it ships under Apache 2.0 in base and instruct versions plus an FP8 variant.

Strengths

  • Open-weight 12B under Apache 2.0
  • 128K context window
  • Drop-in upgrade to Mistral 7B
  • Strong multilingual and code data
  • Runs on consumer GPUs (FP8 variant available)

Caveats

  • 12B ceiling below larger models
  • Text-only (no vision input)
  • Superseded by newer Mistral small models on some tasks
Vendor
Mistral AI
Context Window
128,000 tokens
License
Apache-2.0
Release Date
2024-07-18
Modalities
text

Capabilities

Vision
Audio
Video
Tool Use

Pricing

Open weights under Apache 2.0 - free to self-host; 12B runs on ~8GB VRAM quantized. Also on the Mistral API and NVIDIA NIM.

Reviews

Comments