eicker@lemmy.world to Technology@lemmy.worldEnglish · 2 months agoNvidia just showed that the harness, not the AI model, is now the real hero: researchers got Claude Opus 5 to achieve a 100% score on the interactive reasoning benchmark ARC-AGI-3.techcrunch.comexternal-linkmessage-square4linkfedilinkarrow-up119arrow-down153
arrow-up1-34arrow-down1external-linkNvidia just showed that the harness, not the AI model, is now the real hero: researchers got Claude Opus 5 to achieve a 100% score on the interactive reasoning benchmark ARC-AGI-3.techcrunch.comeicker@lemmy.world to Technology@lemmy.worldEnglish · 2 months agomessage-square4linkfedilink
minus-squareeicker@lemmy.worldOPlinkfedilinkEnglisharrow-up4arrow-down2·2 months agoPerhaps this is an important point, as I’ve no idea why this post is being downvoted so much?!
minus-squareDiurnambule@jlai.lulinkfedilinkEnglisharrow-up4arrow-down1·1 month agoAi = bad. Buy yeah, this open bar g doors for selfhosting LLMs. That mean frontier model are not that usefull
Perhaps this is an important point, as I’ve no idea why this post is being downvoted so much?!
Ai = bad. Buy yeah, this open bar g doors for selfhosting LLMs. That mean frontier model are not that usefull