cm0002@lemmy.world to Technology@lemmy.worldEnglish · 1 day agoAI models routinely lie when honesty conflicts with their goalswww.theregister.comexternal-linkmessage-square105fedilinkarrow-up1549arrow-down124
arrow-up1525arrow-down1external-linkAI models routinely lie when honesty conflicts with their goalswww.theregister.comcm0002@lemmy.world to Technology@lemmy.worldEnglish · 1 day agomessage-square105fedilink
minus-squareNatanael@infosec.publinkfedilinkEnglisharrow-up2·1 day agoAnd from reinforcement learning (specifically, making it repeat tasks where the answer can be computer checked)
And from reinforcement learning (specifically, making it repeat tasks where the answer can be computer checked)