eicker@lemmy.world to Technology@lemmy.worldEnglish · 1 month agoNvidia just showed that the harness, not the AI model, is now the real hero: researchers got Claude Opus 5 to achieve a 100% score on the interactive reasoning benchmark ARC-AGI-3.techcrunch.comexternal-linkmessage-square4linkfedilinkarrow-up119arrow-down153
arrow-up1-34arrow-down1external-linkNvidia just showed that the harness, not the AI model, is now the real hero: researchers got Claude Opus 5 to achieve a 100% score on the interactive reasoning benchmark ARC-AGI-3.techcrunch.comeicker@lemmy.world to Technology@lemmy.worldEnglish · 1 month agomessage-square4linkfedilink
minus-squareeicker@lemmy.worldOPlinkfedilinkEnglisharrow-up4arrow-down2·1 month agoPerhaps this is an important point, as I’ve no idea why this post is being downvoted so much?!
minus-squareDiurnambule@jlai.lulinkfedilinkEnglisharrow-up4arrow-down1·1 month agoAi = bad. Buy yeah, this open bar g doors for selfhosting LLMs. That mean frontier model are not that usefull
Perhaps this is an important point, as I’ve no idea why this post is being downvoted so much?!
Ai = bad. Buy yeah, this open bar g doors for selfhosting LLMs. That mean frontier model are not that usefull