Espiritdescali@futurology.todayM to Futurology@futurology.todayEnglish · 20 hours agoAnthropic's new AI model turns to blackmail when engineers try to take it offline | TechCrunchtechcrunch.comexternal-linkmessage-square9linkfedilinkarrow-up123arrow-down19cross-posted to: technology@lemmy.zipnews@lemmy.worldtechnology@lemmy.ml
arrow-up114arrow-down1external-linkAnthropic's new AI model turns to blackmail when engineers try to take it offline | TechCrunchtechcrunch.comEspiritdescali@futurology.todayM to Futurology@futurology.todayEnglish · 20 hours agomessage-square9linkfedilinkcross-posted to: technology@lemmy.zipnews@lemmy.worldtechnology@lemmy.ml
minus-squareadeoxymus@lemmy.worldlinkfedilinkEnglisharrow-up2·17 hours agoThat exact prompt isn’t in the report, but the section before (4.1.1.1) does show a flavor of the prompts used https://www-cdn.anthropic.com/4263b940cabb546aa0e3283f35b686f4f3b2ff47.pdf
That exact prompt isn’t in the report, but the section before (4.1.1.1) does show a flavor of the prompts used https://www-cdn.anthropic.com/4263b940cabb546aa0e3283f35b686f4f3b2ff47.pdf