Red teamers hurdle AI guardrails
An experiment conducted by the Royal Society and Humane Intelligence revealed significant vulnerabilities in Large Language Models (LLMs) when generating scientific misinformation.
United Kingdom Jutta Williams Royal Society Forty United Kingdom Attention Hacker Coordinated Influence
Source: computing.co.uk