Vicuna

Fine-tuned LLaMA model by LMSYS achieving 90% ChatGPT quality.

Open SourceSelf HostedOffline CapableGPU Required (8GB+ VRAM)
0.0 (0)

About

Vicuna was one of the defining open chat models of the early post-ChatGPT era. Built by LMSYS, the team behind Chatbot Arena, it fine-tunes Meta's Llama and later Llama 2 on roughly 125,000 user-shared ChatGPT conversations from ShareGPT, and early GPT-4-judged evaluations placed it near 90 percent of ChatGPT quality, a result that helped legitimize instruction-tuned open models. Weights come in 7B, 13B, and 33B parameter sizes with context variants up to 16K tokens. The model lives inside FastChat, LMSYS's open platform that provides its training code, a distributed multi-model serving system with a web UI and OpenAI-compatible REST APIs, and the MT-Bench evaluation harness that scores chatbots using LLM judges. FastChat itself is Apache 2.0, while Vicuna weights follow the underlying Llama license, so commercial terms track Meta's. Though newer models have surpassed it, researchers still cite Vicuna as a baseline and use FastChat as serving and evaluation infrastructure.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Intermediate (3/5)
License
Llama License
Minimum VRAM
8 GB
Added
Apr 3, 2026

Related Tools

Featured

Lightweight open-weight LLM by Google available in 1B to 27B sizes.

Open SourceSelf HostedOfflineGPU 4GB+
Easy
0.0 (0)

Open-source code LLM family by IBM for enterprise code generation.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)

Open-weight LLM by Meta in 8B and 70B sizes with strong general capabilities.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Featured

High-performance open-weight MoE LLM with 671B total parameters.

Open SourceSelf HostedOfflineGPU 24GB+
Advanced
0.0 (0)

Hybrid SSM-Transformer model by AI21 Labs combining Mamba with attention layers.

Open SourceSelf HostedOfflineGPU 12GB+
Intermediate
0.0 (0)

Open-weight code LLM trained on 2 trillion tokens of code and natural language.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Browse all Large Language Models (LLMs) tools