A fundamental flaw leaves LLMs strikingly vulnerable to attack
via role-confusion.github.io
Short excerpt below. Read at the original source.
It is impossible to make large language models fully secure against hacks because of a fundamental flaw in how they work, a team of researchers argue in a paper presented at the International Conference on Machine Learning, a top AI conference, this month. The claim has huge implications for the safety of this technology, which […]