
💥 Karpathy’s 2025 LLM summary puts RL with verifiable rewards (RLVR) first — hinting at a new scaling law. 🤗 We just open-sourced CoPaRe (Coordinated Parallel Reasoning): scale test-time compute, generate longer reasoning chains, and unlock big gains. 🔢 An 8B model beats GPT-5 Thinking on math and it’s already gaining strong traction on HuggingFace. 🚀 Dive in and try it now! https://t.co/hDgopViIuS
CoPaRe scales test time compute through coordinated parallel reasoning chains, and StepFun reports an 8B model outscoring GPT-5 Thinking on math with it, released openly so the claim can be tested.
article👏🏻Congratulations!Step3-VL-10B was selected for HuggingFace Daily Papers…
postStep-Audio-R1.1 opens weights for a speech model that reasons in real timeChecking sign-in…
Loading comments…