A fundamental flaw leaves LLMs strikingly vulnerable to attack
Researchers argue that a fundamental flaw in how large language models operate makes it impossible to fully secure them against hacks. This claim was presented in a paper at the International Conference on Machine Learning, raising significant safety concerns.
Why it matters: If true, this claim could have major implications for the safety and deployment of large language models.
Full story at: MIT Technology Review / AI ↗