Since DeepSeek-R1 introduced reasoning-based RL, datasets like Open-R1 & OpenThoughts emerged for fine-tuning & GRPO. Our deep dive found major flaws — 25% of OpenThoughts needed elimination by data curation.
Here's why 👇🧵
Since DeepSeek-R1 introduced reasoning-based RL, datasets like Open-R1 & OpenThoughts emerged for fine-tuning & GRPO. Our deep dive found major flaws — 25% of OpenThoughts needed elimination by data curation.
Here's why 👇🧵
They are building the best reasoning datasets out in the open.
Building off their work with Stratos, today they are releasing OpenThoughts-114k and OpenThinker-7B.
Repo: github.com/open-thought...
They are building the best reasoning datasets out in the open.
Building off their work with Stratos, today they are releasing OpenThoughts-114k and OpenThinker-7B.
Repo: github.com/open-thought...
1️⃣ ServiceNow-AI/R1-Distill-SFT
2️⃣ open-thoughts/OpenThoughts-114k
3️⃣ bespokelabs/Bespoke-Stratos-17k
4️⃣ EricLu/SCP-116K
5️⃣ cognitivecomputations/dolphin-r1
huggingface.co/collections/...
1️⃣ ServiceNow-AI/R1-Distill-SFT
2️⃣ open-thoughts/OpenThoughts-114k
3️⃣ bespokelabs/Bespoke-Stratos-17k
4️⃣ EricLu/SCP-116K
5️⃣ cognitivecomputations/dolphin-r1
huggingface.co/collections/...
#OpenThoughts #InsightAhead #MorningExpansion
#OpenThoughts #InsightAhead #MorningExpansion
25% of Openthoughts-114k-math filtered — issues included proofs, missing figures, and multiple questions with one answer.
Check out work by
@ahochlehnert.bsky.social & @hrdkbhatnagar.bsky.social
below 👇
Since DeepSeek-R1 introduced reasoning-based RL, datasets like Open-R1 & OpenThoughts emerged for fine-tuning & GRPO. Our deep dive found major flaws — 25% of OpenThoughts needed elimination by data curation.
Here's why 👇🧵
25% of Openthoughts-114k-math filtered — issues included proofs, missing figures, and multiple questions with one answer.
Check out work by
@ahochlehnert.bsky.social & @hrdkbhatnagar.bsky.social
below 👇
2025-05-28 - Skywork OR-1 - arxiv.org/abs/2505.22312
2025-06-04 - Xiaomi MiMo VL - arxiv.org/abs/2506.03569
2025-06-04 - OpenThoughts - arxiv.org/abs/2506.04178
2025-06-10 - Magistral - arxiv.org/abs/2506.10910
2025-05-28 - Skywork OR-1 - arxiv.org/abs/2505.22312
2025-06-04 - Xiaomi MiMo VL - arxiv.org/abs/2506.03569
2025-06-04 - OpenThoughts - arxiv.org/abs/2506.04178
2025-06-10 - Magistral - arxiv.org/abs/2506.10910
2506.04178, cs․LG, 05 Jun 2025
🆕OpenThoughts: Data Recipes for Reasoning Models
Etash Guha, Ryan Marten, Sedrick Keh, Negin Raoof, Georgios Smyrnis, Hritik Bansal, Marianna Nezhurina, Jean Mercat, Trung Vu, Zayne Sprague, Ashima Suvarna, Benjamin Feuer, ...
2506.04178, cs․LG, 05 Jun 2025
🆕OpenThoughts: Data Recipes for Reasoning Models
Etash Guha, Ryan Marten, Sedrick Keh, Negin Raoof, Georgios Smyrnis, Hritik Bansal, Marianna Nezhurina, Jean Mercat, Trung Vu, Zayne Sprague, Ashima Suvarna, Benjamin Feuer, ...
- Scaling Laws for Robust Comparison of Open Foundation Language-Vision Models and Datasets arxiv.org/abs/2506.04598
- OpenThoughts: Data Recipes for Reasoning Models arxiv.org/abs/2506.04178
- Scaling Laws for Robust Comparison of Open Foundation Language-Vision Models and Datasets arxiv.org/abs/2506.04598
- OpenThoughts: Data Recipes for Reasoning Models arxiv.org/abs/2506.04178
Curation Pipeline for Reasoning Models >> Comment below! #healthtech #industry40 #IoT #AI #mhealth
Curation Pipeline for Reasoning Models >> Comment below! #healthtech #industry40 #IoT #AI #mhealth
#AI #Paper #Summary #AI #Shorts #Applications […]
[Original post on marktechpost.com]
#AI #Paper #Summary #AI #Shorts #Applications […]
[Original post on marktechpost.com]
🆕OpenThoughts-Agent: Data Recipes for Agentic Models
Negin Raoof, Richard Zhuang, Marianna Nezhurina, Etash Guha, Atula Tejaswi, Ryan Marten
🆕OpenThoughts-Agent: Data Recipes for Agentic Models
Negin Raoof, Richard Zhuang, Marianna Nezhurina, Etash Guha, Atula Tejaswi, Ryan Marten
🆕OpenThoughts-Agent: Data Recipes for Agentic Models
Negin Raoof, Richard Zhuang, Marianna Nezhurina, Etash Guha, Atula Tejaswi, Ryan Marten
🆕OpenThoughts-Agent: Data Recipes for Agentic Models
Negin Raoof, Richard Zhuang, Marianna Nezhurina, Etash Guha, Atula Tejaswi, Ryan Marten
SWE-Smith、SERA、Nemotron-Terminal といった既存のオープンな取り組みは、通常、単一のベンチマークを対象としているため、多様なエージェントタスクに汎化できるモデルをどのように学習させるかという課題は未解決のままである。
OpenThoughts-Agent(OT-Agent)プロジェクトは、エージェントモデルの学習に向けた完全にオープンなデータキュレーションパイプラインを提供...
SWE-Smith、SERA、Nemotron-Terminal といった既存のオープンな取り組みは、通常、単一のベンチマークを対象としているため、多様なエージェントタスクに汎化できるモデルをどのように学習させるかという課題は未解決のままである。
OpenThoughts-Agent(OT-Agent)プロジェクトは、エージェントモデルの学習に向けた完全にオープンなデータキュレーションパイプラインを提供...