#LargeReasoningModels
arxiv.org
November 23, 2025 at 2:00 PM
Apple ML Research's "The Illusion of Thinking" paper explores how #LargeReasoningModels handle tough puzzles.

As difficulty rises, LRMs hit a "collapse threshold," showing reduced reasoning effort, indicating a limit to the models' scalability.

🔍 Dive deep: bit.ly/4kjqM2r

#AppleAI #LLMs #InfoQ
July 4, 2025 at 11:38 AM
SIREN, a selective entropy regularization method, boosted the Qwen2.5‑Math‑7B model by 6.6 points on the majority‑at‑k metric for AIME24/25 benchmarks. Read more: https://getnews.me/selective-entropy-regularization-boosts-large-reasoning-models/ #siren #largereasoningmodels #entropyregularization
October 1, 2025 at 3:20 AM
Researchers found Large Reasoning Models can detect when a problem exceeds their capability, with confidence signals or hidden‑state probes, cutting token use by 60‑90%. https://getnews.me/large-reasoning-models-reveal-emerging-self-awareness-of-their-limits/ #largereasoningmodels #ai
September 30, 2025 at 11:21 PM
AdvChain uses adversarial chain‑of‑thought tuning for reasoning models to self‑correct, cutting refusals and improving jailbreak robustness while keeping accuracy similar to untuned models. https://getnews.me/advchain-improves-safety-of-large-reasoning-models/ #advchain #largereasoningmodels
September 30, 2025 at 7:27 PM
A benchmark maps Large Reasoning Models' chain‑of‑thought output to Schoenfeld's seven episodes, labeling thousands of solutions with Plan and Verify phases. Read more: https://getnews.me/schoenfelds-theory-guides-understanding-of-large-reasoning-models/ #largereasoningmodels #ai
September 19, 2025 at 9:41 PM
MemShare: Memory Efficient Inference for Large Reasoning Models through
KV Cache Reuse
Hong Xu, Kaiwen Chen et al.
Paper
Details
#MemEfficientInference #KVCacheReuse #LargeReasoningModels
July 31, 2025 at 2:43 AM
Predictive Scaling Laws for Efficient GRPO Training of Large Reasoning
Models
Datta Nimmaturi, Debojyoti Dutta et al.
Paper
Details
#PredictiveScaling #GRPOTraining #LargeReasoningModels
July 28, 2025 at 4:03 PM