Anthropic Launches AI Safety Levels to Mitigate Advanced AI Risks
Anthropic has introduced its Responsible Scaling Policy (RSP) to manage the risks of advanced AI systems. The RSP features AI Safety Levels (ASL), inspired by biosafety standards, categorizing AI based on potential risk from ASL-1 to ASL-4. Current models, such as Claude, are classified as ASL-2, indicating early hazardous capabilities. The policy aims to minimize catastrophic risks while promoting beneficial AI applications, halting model training if safety standards can’t be met. Continuous evaluation, approved by the board, ensures rigorous safety measures akin to those in automotive and aviation industries. The RSP is a dynamic framework, adapting to new AI developments.