RT @ClementDelangue: It's now clear that:
- alignment is critical to making AI safe
- alignment won't be solved behind the closed doors of a handful of frontier labs
So we're taking two steps:
- launching the Open Alignment Initiative, led by @Thom_Wolf @huggingface.
- asking to be part of the "embedded evaluators" program that @DarioAmodei just committed to.
Let's make AI safer by making it more transparent!
引用推文
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so.
Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.
You can read the full post here: https://t.co/OGyPb7yaYt