Model description

This is a Vicuna-like model with only 68M parameters, which is fine-tuned from LLaMA-68m on ShareGPT data.

The training setup follows the Vicuna suite.

The model is mainly developed as a base Small Speculative Model in the MCSD paper. As a comparison, it can be better aligned to the Vicuna models than LLaMA-68m with little loss of alignment to the LLaMA models.

Draft Model	Target Model	Alignment
LLaMA-68/160M	LLaMA-13/33B	😃
LLaMA-68/160M	Vicuna-13/33B	😟
Vicuna-68/160M	LLaMA-13/33B	😃
Vicuna-68/160M	Vicuna-13/33B	😃

Downloads last month: 6,295

Safetensors

Model size

68M params

Tensor type

F32

Model tree for double7/vicuna-68m

Quantizations

1 model

Dataset used to train double7/vicuna-68m

Collection including double7/vicuna-68m

SSM

Collection

Small Speculative Models • 2 items • Updated May 23, 2025

Paper for double7/vicuna-68m

Multi-Candidate Speculative Decoding

Paper • 2401.06706 • Published Jan 12, 2024 • 1