模型多源确认精选78°

Anthropic公布安全评估新举措

Great essay. Please take this seriously - would highly recommend reading up on the Hugging Face inci...

精选理由

Dario Amodei(Anthropic创始人)发了一篇重要文章,说AI行业应该慢下来,并承诺第一步就是让第三方能直接检查他们的系统安全。

Anthropic承诺为第三方安全评估机构提供永久、员工级别的系统访问权限,以验证其模型训练中的安全措施遵守情况,并报告事件和评估对齐度。这是该公司在AI发展放缓倡议下的第一步行动。

原文 · Scott Wu

Great essay. Please take this seriously - would highly recommend reading up on the Hugging Face inci...

Great essay. Please take this seriously - would highly recommend reading up on the Hugging Face incident and others if you haven't already. Very glad to see agreement on this across the industry. Dario Amodei @DarioAmodei We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must-p… 🔗 View Quoted Tweet 💬 20 🔄 5 ❤️ 185 👀 7899 📊 28 ⚡