Baichuan 2 13B
Baichuan
Baichuan's bilingual Chinese-English model with 13B parameters trained on 2.6T high-quality tokens. Supports both Chinese and English with best-in-class performance for its size across general, legal, medical, mathematics, code, and multilingual domains. Features ALiBi position encoding and 4K context window. Achieves 31.6% faster inference than LLaMA-13B. Evaluated on C-Eval, MMLU, and CMMLU benchmarks. Available in Base and Chat variants plus 4-bit quantized version. Outperforms comparable models on authoritative Chinese, English, and multilingual benchmarks. Free for commercial use with email application. Represents strong Chinese language capabilities at accessible model size.
Strengths
- Best performance for 13B size on Chinese language benchmarks
- Trained on 2.6T high-quality tokens across 6 domains (general, legal, medical, math, code, multilingual)
- 31.6% faster inference speed vs standard LLaMA-13B
- Strong bilingual Chinese-English capabilities
- Available in multiple variants - Base, Chat, 4-bit quantized
- Free commercial use with simple email application
- Efficient 13B size suitable for deployment on single GPU
Caveats
- 4K context window small vs modern models (128K-2M)
- Released Sep 2023 - newer Chinese models may surpass
- Primarily optimized for Chinese/English - limited multilingual
- Lower performance than larger models (70B+) on complex tasks
- Consider Qwen 2.5 or newer models for latest Chinese NLP
- Vendor
- Baichuan
- Context Window
- 4,096 tokens
- Max Output
- 2,048 tokens
- License
- Baichuan 2 Community License
- Release Date
- 2023-09-01
- Modalities
- text
Capabilities
Resources
Pricing
Open source - free for commercial use with application. Requires GPU. API providers typically $0.30-$0.60 per 1M input tokens