Why you should worry about Anthropic, OpenAI’s proposed AI risk evaluators

Anthropic and OpenAI are proposing embedded AI evaluators to help manage risk of models causing catastrophic harm to society, but the idea has some issues.

Leave a Reply

Your email address will not be published. Required fields are marked *