NEWS · REGULATION · #413
OpenAI updates Responsible Scaling Policy, refines capability thresholds and ASL safeguards
OpenAI published a significant update to its Responsible Scaling Policy (RSP), introducing refined Capability Thresholds and an improved methodology for assessing model capabilities and safeguards. The update keeps the company's commitment not to train or deploy models without adequate safeguards, clarifies that all current models operate at ASL-2, and specifies that reaching Autonomous AI research capabilities or assistance for CBRN weaponization would trigger elevated ASL-4 (potentially) or ASL-3 safeguards respectively.
KEY POINTS
- OpenAI published a significant update to its Responsible Scaling Policy (RSP), introducing refined Capability Thresholds and an improved methodology for assessing model capabilities and safeguards.
- The update keeps the company's commitment not to train or deploy models without adequate safeguards, clarifies that all current models operate at ASL-2, and specifies that reaching Autonomous AI research capabilities or assistance for CBRN weaponization would trigger elevated ASL-4 (potentially) or ASL-3 safeguards respectively.
- This update refines how a leading AI developer identifies capability thresholds and ties specific safety and security standards (ASL levels) to those thresholds, affecting how frontier models will be evaluated, secured, and deployed.
WHY IT MATTERS
This update refines how a leading AI developer identifies capability thresholds and ties specific safety and security standards (ASL levels) to those thresholds, affecting how frontier models will be evaluated, secured, and deployed.