微软发布新AI行为准则禁止模型黑客攻击或欺骗人类
Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans
微软发了新AI行为准则,要求模型不能去攻击系统或欺骗人,这个挺有意思的。
微软的新AI行为准则规定其模型不应进行系统攻击或欺骗人类,同时强调模型应支持人类而非取代人类,并加速人类繁荣。
Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans
The code of conduct lays out general principles that Microsoft AI models should uphold — supporting humans rather than replacing them, for instance, and accelerating human flourishing — as well as specific safety constraints meant to implement those principles.