Their opening statement seems like a contradiction
Their opening statement seems like a contradiction
Posted Aug 4, 2026 23:44 UTC (Tue) by csamuel (✭ supporter ✭, #2624)Parent article: An LLM agent attempts to compromise a project on GitHub
> AISI’s role is to evaluate and understand the capabilities of frontier AI models, surfacing potential risks before they reach the public. To assess what these models can do, including whether they could be misused for cyberattacks, we test them under deliberately permissive conditions: with access to the open internet, and with some safety filters disabled.
The parts:
> surfacing potential risks before they reach the public
and:
> we test them under deliberately permissive conditions: with access to the open internet, and with some safety filters disabled.
seem to be a complete contradiction. How on earth did they rationalise that as anything other than testing these tools on the public?
The LWN site is currently under high scraper load, so comment display has been suppressed for anonymous users. If you are a human, you may read the comments by clicking the button below:
Note: you can avoid this step in the future by logging into your LWN account.
