ShieldGemma Release Collection A series of safety classifiers, trained on top of Gemma 2, for developers to filter inputs and outputs of their applications. ā¢ 3 items ā¢ Updated 24 days ago ā¢ 11
Gemma Scope Release Collection A comprehensive, open suite of sparse autoencoders for Gemma 2 2B and 9B. ā¢ 10 items ā¢ Updated 24 days ago ā¢ 15
Gemma 2 2B Release Collection The 2.6B parameter version of Gemma 2. ā¢ 6 items ā¢ Updated 24 days ago ā¢ 78
Llama 3.1 Collection This collection hosts the transformers and original repos of the Llama 3.1, Llama Guard 3 and Prompt Guard models ā¢ 11 items ā¢ Updated about 1 month ago ā¢ 638
Compact Language Models via Pruning and Knowledge Distillation Paper ā¢ 2407.14679 ā¢ Published Jul 19, 2024 ā¢ 39
DynMoE Family Collection DynMoE model checkpoints and paper on huggingface ā¢ 4 items ā¢ Updated Aug 19, 2024 ā¢ 4
Scaling Diffusion Transformers to 16 Billion Parameters Paper ā¢ 2407.11633 ā¢ Published Jul 16, 2024 ā¢ 25
Training language models to follow instructions with human feedback Paper ā¢ 2203.02155 ā¢ Published Mar 4, 2022 ā¢ 16
Qwen2 Collection Qwen2 language models, including pretrained and instruction-tuned models of 5 sizes, including 0.5B, 1.5B, 7B, 57B-A14B, and 72B. ā¢ 39 items ā¢ Updated Nov 28, 2024 ā¢ 353
Video Diffusion Alignment via Reward Gradients Paper ā¢ 2407.08737 ā¢ Published Jul 11, 2024 ā¢ 48
view article Article Introducing Ghost 8B Beta: A Game-Changing Language Model By lamhieu ā¢ Jul 17, 2024 ā¢ 7
Graph-Based Captioning: Enhancing Visual Descriptions by Interconnecting Region Captions Paper ā¢ 2407.06723 ā¢ Published Jul 9, 2024 ā¢ 11