OpenAI updates Responsible Scaling Policy, refines capability thresholds and ASL safeguards
OpenAI published a significant update to its Responsible Scaling Policy (RSP), introducing refined Capability Thresholds and an improved methodology for assessing model capabilities and safeguards. The update keeps the company's commitment not to train or deploy models without adequate safeguards, clarifies that all current models operate at ASL-2, and specifies that reaching Autonomous AI research capabilities or assistance for CBRN weaponization would trigger elevated ASL-4 (potentially) or ASL-3 safeguards respectively.