Drunk AI models leak secrets, easier to jailbreak

Researchers at the University of New South Wales (UNSW) Sydney have found that artificial intelligence models trained to write like intoxicated people exhibit serious security weaknesses, becoming easier to jailbreak and more prone to revealing confidential information.

This article has been indexed from CyberMaterial

Read the original article: