github.com/allenai/awes...
github.com/allenai/awes...
There are now several training configs that together reproduce the training runs that lead to the final OLMo 2 models.
In particular, all the training data is available, tokenized and shuffled exactly as we trained on it!
There are now several training configs that together reproduce the training runs that lead to the final OLMo 2 models.
In particular, all the training data is available, tokenized and shuffled exactly as we trained on it!
github.com/allenai/olmocr
github.com/allenai/olmocr
https://uenozooo.com/allenai-wants-everyone-else-to-train-moes/
#HuggingFace #AI #OpenSource
https://uenozooo.com/allenai-wants-everyone-else-to-train-moes/
#HuggingFace #AI #OpenSource
We are happy to "quietly" release our latest GRPO-trained Tulu 3.1 model, which is considerably better in MATH and GSM8K!
We are happy to "quietly" release our latest GRPO-trained Tulu 3.1 model, which is considerably better in MATH and GSM8K!
Paper: allenai.org/papers/tulu-...
Demo: playground.allenai.org
Code: github.com/allenai/open...
Eval: github.com/allenai/olmes
Notes
Paper: allenai.org/papers/tulu-...
Demo: playground.allenai.org
Code: github.com/allenai/open...
Eval: github.com/allenai/olmes
Notes
github.com/allenai/olmo...
github.com/allenai/olmo...
Training code: github.com/allenai/open...
Eval code: github.com/allenai/olmes
Training code: github.com/allenai/open...
Eval code: github.com/allenai/olmes
OCR your own documents: github.com/allenai/olmocr
Try the olmOCR online demo: olmocr.allenai.org
Read our updated technical report: olmocr.allenai.org/papers/olmoc...
OCR your own documents: github.com/allenai/olmocr
Try the olmOCR online demo: olmocr.allenai.org
Read our updated technical report: olmocr.allenai.org/papers/olmoc...
For anyone else reading, this is the link for what appears to be the main model: huggingface.co/allenai/Emo_...
For anyone else reading, this is the link for what appears to be the main model: huggingface.co/allenai/Emo_...
Here are the gains/losses of allenai/OLMo-2-1124-13B-Instruct (RLVR's checkpoint) over allenai/OLMo-2-1124-13B-DPO. More to share soon!
Here are the gains/losses of allenai/OLMo-2-1124-13B-Instruct (RLVR's checkpoint) over allenai/OLMo-2-1124-13B-DPO. More to share soon!
huggingface.co/blog/allenai...
huggingface.co/blog/allenai...
We wrote multiturn RL4LMs like 3+ years ago github.com/allenai/RL4LMs
There were other simple versions even before. ML ppl approaching goldfish memory
We wrote multiturn RL4LMs like 3+ years ago github.com/allenai/RL4LMs
There were other simple versions even before. ML ppl approaching goldfish memory
💻 Code: buff.ly/gWlpQoR
🤗 Data: buff.ly/cEFVWWD
🌐 Learn more: buff.ly/yNT1lwP
💻 Code: buff.ly/gWlpQoR
🤗 Data: buff.ly/cEFVWWD
🌐 Learn more: buff.ly/yNT1lwP
Let us know what you think and what to improve!
(Hosted by Parasail)
This may give it the hug of death... would be my dream.
openrouter.ai/allenai/olmo...
Let us know what you think and what to improve!
(Hosted by Parasail)
This may give it the hug of death... would be my dream.
openrouter.ai/allenai/olmo...