Responsible Scaling Policies (RSP)
Pioneered by Anthropic and adopted across frontier labs, RSPs define explicit capability triggers (ASL-1 to ASL-4). If a model demonstrates dangerous autonomous replication or biological synthesis, the lab is contractually and operationally bound to halt training or release until safeguards match the threat.