allenai
/

OLMo-2-1124-7B-SFT-Preview

@@ -14,13 +14,12 @@ library_name: transformers
 OLMo-2 7B SFT November 2024 is finetuned variant of the [OLMo-2 7B November 2024](https://huggingface.co/allenai/OLMo2-7B-1124) model, which has undergone supervised finetuning on the [Tülu 3 dataset](https://huggingface.co/datasets/allenai/tulu-3-sft-mixture).
 Tülu 3 is designed for state-of-the-art performance on a diversity of tasks in addition to chat, such as MATH, GSM8K, and IFEval.
-We use a
 OLMo is a series of **O**pen **L**anguage **Mo**dels designed to enable the science of language models.
 These models are trained on the Dolma dataset. We are releasing all code, checkpoints, logs (coming soon), and associated training details.
 The core models released in this batch include the following:
-TODO: model family details.
 | **Stage**           | **OLMo-2 7B**                                                                                          | **OLMo-2 7B**                                                                                         |
 |----------------------|----------------------------------------------------------------------------------------------------------|----------------------------------------------------------------------------------------------------------|
@@ -46,7 +45,7 @@ TODO: model family details.
     - Core repo (training, inference, fine-tuning etc.): https://github.com/allenai/OLMo
     - Evaluation code: https://github.com/allenai/olmes
     - Further fine-tuning code: https://github.com/allenai/open-instruct
-- **Paper:** Coming soon!
 - **Demo:** https://playground.allenai.org/
 ## Using the model

 OLMo-2 7B SFT November 2024 is finetuned variant of the [OLMo-2 7B November 2024](https://huggingface.co/allenai/OLMo2-7B-1124) model, which has undergone supervised finetuning on the [Tülu 3 dataset](https://huggingface.co/datasets/allenai/tulu-3-sft-mixture).
 Tülu 3 is designed for state-of-the-art performance on a diversity of tasks in addition to chat, such as MATH, GSM8K, and IFEval.
+Check out [the OLMo-2 paper](https://TODO) or [Tülu 3 paper](https://arxiv.org/abs/2411.15124) for more details!
 OLMo is a series of **O**pen **L**anguage **Mo**dels designed to enable the science of language models.
 These models are trained on the Dolma dataset. We are releasing all code, checkpoints, logs (coming soon), and associated training details.
 The core models released in this batch include the following:
 | **Stage**           | **OLMo-2 7B**                                                                                          | **OLMo-2 7B**                                                                                         |
 |----------------------|----------------------------------------------------------------------------------------------------------|----------------------------------------------------------------------------------------------------------|
     - Core repo (training, inference, fine-tuning etc.): https://github.com/allenai/OLMo
     - Evaluation code: https://github.com/allenai/olmes
     - Further fine-tuning code: https://github.com/allenai/open-instruct
+- **Paper:** Coming soon! TODO
 - **Demo:** https://playground.allenai.org/
 ## Using the model