allenai
/

Llama-3.1-Tulu-3.1-8B

Text Generation

text-generation-inference

Inference Endpoints

Model card Files Files and versions Community

natolambert commited on 10 days ago

Commit

46239c2

·

verified ·

1 Parent(s): 11a6310

Update README.md

Files changed (1) hide show

README.md +3 -2

README.md CHANGED Viewed

@@ -14,8 +14,9 @@ library_name: transformers
 # Llama-3.1-Tulu-3.1-8B
-Tülu3 is a leading instruction following model family, offering fully open-source data, code, and recipes designed to serve as a comprehensive guide for modern post-training techniques.
-Tülu3 is designed for state-of-the-art performance on a diversity of tasks in addition to chat, such as MATH, GSM8K, and IFEval.
 **Version 3.1 update**: The new version of our Tülu model is from an improvement only in the final RL stage of training.
 We switched from PPO to GRPO (no reward model) and did further hyperparameter tuning to achieve substantial performance improvements across the board over the original Tülu 3 8B model,

 # Llama-3.1-Tulu-3.1-8B
+Tülu 3 is a leading instruction following model family, offering a post-training package with fully open-source data, code, and recipes designed to serve as a comprehensive guide for modern techniques.
+This is one step of a bigger process to training fully open-source models, like our [OLMo](https://allenai.org/olmo) models.
+Tülu 3 is designed for state-of-the-art performance on a diversity of tasks in addition to chat, such as MATH, GSM8K, and IFEval.
 **Version 3.1 update**: The new version of our Tülu model is from an improvement only in the final RL stage of training.
 We switched from PPO to GRPO (no reward model) and did further hyperparameter tuning to achieve substantial performance improvements across the board over the original Tülu 3 8B model,