AG2 now has surpassed the average gold-medalist in solving Olympiad geometry problems, w/ a solve rate of 84% compared to 54% previously!
Paper: arxiv.org/abs/2502.03544
See full list of authors on link
AG2 now has surpassed the average gold-medalist in solving Olympiad geometry problems, w/ a solve rate of 84% compared to 54% previously!
Paper: arxiv.org/abs/2502.03544
See full list of authors on link
go.nature.com/4jSPnfJ
go.nature.com/4jSPnfJ
Using the other place as the only place is bad for the field!
Using the other place as the only place is bad for the field!
Main Link | Techmeme Permalink
Main Link | Techmeme Permalink
garymarcus.substack.com/p/alphageome...
garymarcus.substack.com/p/alphageome...
AlphaGeometry2 (AG2) surpasses IMO gold medalists in solving Olympiad geometry problems by expanding its problem domain, enhancing symbolic reasoning, and integrating a stronger language model.
AlphaGeometry2 (AG2) surpasses IMO gold medalists in solving Olympiad geometry problems by expanding its problem domain, enhancing symbolic reasoning, and integrating a stronger language model.
🚀This week, AI is leveling up—from solving Olympiad geometry with AlphaGeometry2 to generating motion with VideoJAM, while also mastering the art of reasoning and adversarial resilience with tools like DeepRAG and LIMO. 🚨
🚀This week, AI is leveling up—from solving Olympiad geometry with AlphaGeometry2 to generating motion with VideoJAM, while also mastering the art of reasoning and adversarial resilience with tools like DeepRAG and LIMO. 🚨
- OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models
- Gold-medalist Performance in Solving Olympiad Geometry with AlphaGeometry2
- DeepRAG: Thinking to Retrieval Step by Step for Large Language Models
- OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models
- Gold-medalist Performance in Solving Olympiad Geometry with AlphaGeometry2
- DeepRAG: Thinking to Retrieval Step by Step for Large Language Models
But the interesting part isn't the model — it's the architecture.
LLM + symbolic reasoning + search.
Here’s a deep dive into how the system works:
luhuidev.medium.com/alphageometr...
But the interesting part isn't the model — it's the architecture.
LLM + symbolic reasoning + search.
Here’s a deep dive into how the system works:
luhuidev.medium.com/alphageometr...
1. 第一次达到了 IMO 金牌选手的平均水平
2. 50道 2000年 到 2024年的 IMO 几何题目,OpenAI o1 一道都解决不了,而 AG2 可以解决 42 道,AG1 的水平是 27 道
3. “... many AlphaGeometry solutions to exhibit superhuman creativity”
arxiv.org/pdf/2502.03544
1. 第一次达到了 IMO 金牌选手的平均水平
2. 50道 2000年 到 2024年的 IMO 几何题目,OpenAI o1 一道都解决不了,而 AG2 可以解决 42 道,AG1 的水平是 27 道
3. “... many AlphaGeometry solutions to exhibit superhuman creativity”
arxiv.org/pdf/2502.03544
Google DeepMind's AlphaGeometry2 (AG2) AI model solved 84% of the geometry problems from the last 25 years of International Math Olympiads (IMO), outperforming the average human gold-medalist per…
#ai #deepmind #news
www.scientificamerican.com/article/goog...
www.scientificamerican.com/article/goog...
#NeuroSymbolic #AI #FutureOfAI
#InspiredbyNature
the-decoder.com/alphageometr...
#NeuroSymbolic #AI #FutureOfAI
#InspiredbyNature
the-decoder.com/alphageometr...