Checkpoints for https://huggingface.co/keeeeenw/Llama-3.2-1B-Instruct-Open-R1-Distill

I am keeping it in a separate repo to improve the download speed for the original repo.

This is useful for folks who want to probe into the process of how an LLM learns to reason or continued training for this model.

Downloads last month
3
Safetensors
Model size
1.24B params
Tensor type
BF16
·
Inference Providers NEW
This model is not currently available via any of the supported third-party Inference Providers, and HF Inference API was unable to determine this model's library.