sanitation@lemmy.today to Technology@lemmy.worldEnglish · 3 个月前Nvidia’s CEO Jensen Huang says electricians and plumbers will be needed by the hundreds of thousands in the new working worldfortune.comexternal-linkmessage-square106linkfedilinkarrow-up1248arrow-down139
arrow-up1209arrow-down1external-linkNvidia’s CEO Jensen Huang says electricians and plumbers will be needed by the hundreds of thousands in the new working worldfortune.comsanitation@lemmy.today to Technology@lemmy.worldEnglish · 3 个月前message-square106linkfedilink
minus-squareshrugs@lemmy.worldlinkfedilinkEnglisharrow-up5·3 个月前Don’t antropomorphize AI! An AI doesn’t evaluate anything, an AI doesn’t check for consequences. All AI does is predicting the next word. Do I take the car to the carwash or do i walk? If it’s only 300m away, you should walk Sure, now predict the future please *facepalm*
minus-squareboonhet@sopuli.xyzlinkfedilinkEnglisharrow-up3arrow-down2·3 个月前The carwash thing applies to low end models and older models. Here’s Claude from lowest to highest model, ignoring the banned Fable
minus-squarereplicat@lemmy.worldlinkfedilinkEnglisharrow-up5·3 个月前They altered the training data to address this challenge. The underlying issue wasn’t solved in any way. Don’t be naive.
minus-squareboonhet@sopuli.xyzlinkfedilinkEnglisharrow-up1·3 个月前Takes months to train a model, there were already models that got it right when the question was popular, as long as thinking was enabled. Also if they were optimising for this question, why not update their lower end model (Haiku) as well? The interesting question would be what percent of humans get it wrong. Smaller than LLMs for sure, but I somehow doubt it’s 0.
minus-squaremabeledo@lemmy.worlddeleted by creatorlinkfedilinkEnglisharrow-up1·edit-26 小时前deleted by creator
Don’t antropomorphize AI!
An AI doesn’t evaluate anything, an AI doesn’t check for consequences. All AI does is predicting the next word.
Do I take the car to the carwash or do i walk?
Sure, now predict the future please *facepalm*
The carwash thing applies to low end models and older models. Here’s Claude from lowest to highest model, ignoring the banned Fable
They altered the training data to address this challenge. The underlying issue wasn’t solved in any way. Don’t be naive.
Takes months to train a model, there were already models that got it right when the question was popular, as long as thinking was enabled.
Also if they were optimising for this question, why not update their lower end model (Haiku) as well?
The interesting question would be what percent of humans get it wrong. Smaller than LLMs for sure, but I somehow doubt it’s 0.
deleted by creator