Tests raise fears open AI models could be ‘used at scale for harm’
A testing tool built by researchers at Waterloo Engineering shows that leading open-weight artificial intelligence (AI) models are alarmingly vulnerable to malicious tampering.
The researchers collaborated with experts at FAR.AI, a non-profit AI security research group, to rigorously test 21 of the most popular open-weight large language models (LLMs) and found their built-in safeguards could all be defeated.
