Skip to content
Simon Willison·· 3 d ago

Quoting @joedaroo

Quoting @joedaroo

AI summary

OpenAI agent safety team member @joedaroo said the speed of model capability gains related to cyber, swarming and message boards was surprising. Building a security posture takes time and requires more than hardening systems: safety must be embedded in company culture, with people across the organisation changing accordingly.

Selection record

Not admittedSum of both 104 < twice the threshold 152

Source tier
Media and individuals; this tier's threshold is 76
Pre-filter
passed:OpenAI安全人员谈AI能力突增与应对

A model scores each item twice, independently, against one written standard, out of 100. An item is admitted only when the two scores add up to twice the threshold. The threshold is set per source tier.

Source: Simon Willison · simonwillison.net