模型

微软发布新AI行为准则禁止模型黑客攻击或欺骗人类

Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans

精选理由

微软发了新AI行为准则,要求模型不能去攻击系统或欺骗人,这个挺有意思的。

微软的新AI行为准则规定其模型不应进行系统攻击或欺骗人类,同时强调模型应支持人类而非取代人类,并加速人类繁荣。

图片来源 · techcrunch
原文 · techcrunch

Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans

The code of conduct lays out general principles that Microsoft AI models should uphold — supporting humans rather than replacing them, for instance, and accelerating human flourishing — as well as specific safety constraints meant to implement those principles.