Checkpoints for https://huggingface.co/keeeeenw/Llama-3.2-1B-Instruct-Open-R1-Distill
I am keeping it in a separate repo to improve the download speed for the original repo.
This is useful for folks who want to probe into the process of how an LLM learns to reason or continued training for this model.
- Downloads last month
- 3
Inference Providers
NEW
This model is not currently available via any of the supported third-party Inference Providers, and
HF Inference API was unable to determine this model's library.