eicker@lemmy.world to Technology@lemmy.worldEnglish · 12 days agoNvidia just showed that the harness, not the AI model, is now the real hero: researchers got Claude Opus 5 to achieve a 100% score on the interactive reasoning benchmark ARC-AGI-3.techcrunch.comexternal-linkmessage-square4linkfedilinkarrow-up119
arrow-up119external-linkNvidia just showed that the harness, not the AI model, is now the real hero: researchers got Claude Opus 5 to achieve a 100% score on the interactive reasoning benchmark ARC-AGI-3.techcrunch.comeicker@lemmy.world to Technology@lemmy.worldEnglish · 12 days agomessage-square4linkfedilink
minus-squareeicker@lemmy.worldOPlinkfedilinkEnglisharrow-up4·12 days agoPerhaps this is an important point, as I’ve no idea why this post is being downvoted so much?!
minus-squareDiurnambule@jlai.lulinkfedilinkEnglisharrow-up4·12 days agoAi = bad. Buy yeah, this open bar g doors for selfhosting LLMs. That mean frontier model are not that usefull
Perhaps this is an important point, as I’ve no idea why this post is being downvoted so much?!
Ai = bad. Buy yeah, this open bar g doors for selfhosting LLMs. That mean frontier model are not that usefull