11X cheaper than ChatGPT: Tiny 150M model just proved AI doesn't need to "think out loud" to be smart
Source ↗
👁 0
💬 0
A 150M model reached 29.5% while costing just $0.0007 per taskChatGPT scored higher, yet its comparable reasoning runs cost substantially moreBDH-CQ performs reasoning internally instead of generating lengthy intermediate textPathway, an AI lab focused on building Post-Transformer architectures, has released new benchmark results for its BDH-CQ reasoning model.According to the researchers, their 150M-parameter model scored 29.5% pass@2 on the public ARC-AGI-1 evaluation set.It achieved this at a
Comments (0)