Vicuna 33B
LMSYS
LMSYS's open-source chatbot fine-tuned from LLaMA with supervised instruction fine-tuning on ~125K user-shared conversations from ShareGPT.com. Auto-regressive language model based on transformer architecture with 33B parameters. Evaluated with standard benchmarks, human preference studies, and LLM-as-a-judge (MT-Bench). Achieved 1091 ELO rating on Chatbot Arena. Scores: 62.1 Arc, 83 HellaSwag, 59.2 MMLU, 56.2 TruthfulQA, 77 WinoGrande, 13.7 GSM8K. Released as part of FastChat platform for training and evaluating large language models. Represents important early work in open-source chat model fine-tuning. Non-commercial license. Part of research leading to Chatbot Arena benchmark.
Strengths
- Early influential open-source chat model (June 2023)
- Fine-tuned on 125K real user conversations from ShareGPT
- Evaluated via comprehensive MT-Bench and Chatbot Arena
- 1091 ELO demonstrates strong conversational ability for era
- Part of FastChat platform enabling training and evaluation
- Contributed to development of LLM-as-a-judge methodology
- Open weights enable research and experimentation
Caveats
- Released June 2023 - heavily superseded by modern chat models
- Non-commercial license restricts commercial deployment
- 2K context window extremely limited vs modern models (128K-2M)
- Lower scores than current models (13.7 GSM8K vs GPT-4o 96.4)
- Consider Llama 3.1 70B Instruct or Qwen 2.5 for current chat models
- Vendor
- LMSYS
- Context Window
- 2,000 tokens
- Max Output
- 2,000 tokens
- License
- Non-commercial (LLaMA-derived)
- Release Date
- 2023-06-22
- Modalities
- text
| Benchmark | Score |
|---|---|
| Chatbot Arena ELO | 1091.0 |
| MMLU | 59.2 |
| HellaSwag | 83.0 |
Capabilities
Resources
Pricing
Open source (non-commercial) - free for research. Requires 65GB VRAM. Not available for commercial use