README.md · double7/vicuna-160m at ac8867c1c565d0a6de41ef4d60e782b4d03158a8

metadata

license: apache-2.0
datasets:
  - anon8231489123/ShareGPT_Vicuna_unfiltered
language:
  - en
pipeline_tag: text-generation

Model description

This is a Vicuna-like model with only 160M parameters, which is fine-tuned from LLaMA-160m on ShareGPT data.

The training setup follows the Vicuna suite.

The model is mainly developed as a base Small Speculative Model. As a comparison, it can be better aligned to the Vicuna models than LLaMA-160m with little loss of alignment to the LLaMA models.