Skip to main content
llm.info

Vicuna 33B

Open Weights

LMSYS

LMSYS's open-source chatbot fine-tuned from LLaMA with supervised instruction fine-tuning on ~125K user-shared conversations from ShareGPT.com. Auto-regressive language model based on transformer architecture with 33B parameters. Evaluated with standard benchmarks, human preference studies, and LLM-as-a-judge (MT-Bench). Achieved 1091 ELO rating on Chatbot Arena. Scores: 62.1 Arc, 83 HellaSwag, 59.2 MMLU, 56.2 TruthfulQA, 77 WinoGrande, 13.7 GSM8K. Released as part of FastChat platform for training and evaluating large language models. Represents important early work in open-source chat model fine-tuning. Non-commercial license. Part of research leading to Chatbot Arena benchmark.

Strengths

  • Early influential open-source chat model (June 2023)
  • Fine-tuned on 125K real user conversations from ShareGPT
  • Evaluated via comprehensive MT-Bench and Chatbot Arena
  • 1091 ELO demonstrates strong conversational ability for era
  • Part of FastChat platform enabling training and evaluation
  • Contributed to development of LLM-as-a-judge methodology
  • Open weights enable research and experimentation

Caveats

  • Released June 2023 - heavily superseded by modern chat models
  • Non-commercial license restricts commercial deployment
  • 2K context window extremely limited vs modern models (128K-2M)
  • Lower scores than current models (13.7 GSM8K vs GPT-4o 96.4)
  • Consider Llama 3.1 70B Instruct or Qwen 2.5 for current chat models
Vendor
LMSYS
Context Window
2,000 tokens
Max Output
2,000 tokens
License
Non-commercial (LLaMA-derived)
Release Date
2023-06-22
Modalities
text
BenchmarkScore
Chatbot Arena ELO1091.0
MMLU59.2
HellaSwag83.0

Capabilities

Vision
Audio
Video
Tool Use

Pricing

Open source (non-commercial) - free for research. Requires 65GB VRAM. Not available for commercial use

Reviews

Comments