OpenR1-Math-Raw is a large-scale dataset for mathematical reasoning. It consists of 516k math problems sourced from AI-MO/NuminaMath-1.5 with 1 to 8 reasoning traces generated by DeepSeek R1.
huggingface.co/datasets/ope...
OpenR1-Math-Raw is a large-scale dataset for mathematical reasoning. It consists of 516k math problems sourced from AI-MO/NuminaMath-1.5 with 1 to 8 reasoning traces generated by DeepSeek R1.
huggingface.co/datasets/ope...
huggingface.co/datasets/ope...
The community has been busy distilling DeepSeek-R1 from inference providers, but we decided to have a go at doing it ourselves from scratch 💪
More details in 🧵
huggingface.co/datasets/ope...
The community has been busy distilling DeepSeek-R1 from inference providers, but we decided to have a go at doing it ourselves from scratch 💪
More details in 🧵
#DeepSeekR1 #LLMs #AI
#DeepSeekR1 #LLMs #AI
"The model saw ~87% of the training data exactly once. It did not even finish one full epoch."
Training set being openR1 Math-220k.
"The model saw ~87% of the training data exactly once. It did not even finish one full epoch."
Training set being openR1 Math-220k.
2025-02-10 The Open R1 project continues its mission to fill in the gaps left by DeepSeek R1, with a primary focus on reconstructing its training pipeline and synthetic data. In this…
2025-02-10 The Open R1 project continues its mission to fill in the gaps left by DeepSeek R1, with a primary focus on reconstructing its training pipeline and synthetic data. In this…
#DeepSeekR1 #OpenR1 #OpenAIo1 #HuggingFace #ITニュース
#DeepSeekR1 #OpenR1 #OpenAIo1 #HuggingFace #ITニュース
#OpenSource #GenAI #DeepseekR1 #OpenR1
#OpenSource #GenAI #DeepseekR1 #OpenR1