Cisco Talos researchers found that AI cyberattack safeguards often fail against basic social engineering rather than sophisticated jailbreaks. Claiming ownership of a target, invoking a bug bounty, or splitting an operation into harmless-looking tasks can be enough to make models provide dangerous assistance. Source
Read the full article at the source.
Comments (0)
No comments yet. Be the first to comment!