模型多源确认精选93°

Anthropic 将为第三方提供员工级系统访问权限以评估安全

I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions w...

精选理由

Anthropic要给第三方安全专家开权限了,让他们能直接检查模型训练过程,确保安全,这和之前说的要放慢AI发展节奏有关。

Anthropic宣布将允许第三方评估员获得永久、员工级别的系统访问权限,用于验证其安全措施、报告事件并评估模型对齐情况。这是该公司在AI发展速度上采取的举措之一。

原文 · Sam Altman

I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions w...

I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions we've had at OpenAI in recent weeks. Committing to having independent evaluators with employee-like access is a great idea, and we will do the same. We'll have more to share soon. Dario Amodei @DarioAmodei We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must-p… 🔗 View Quoted Tweet 💬 1732 🔄 2038 ❤️ 20523 👀 2121692 📊 3130 ⚡