模型多源确认

前沿实验室或从用户对话中获取新见解用于训练

Own your intelligence stack. This is a recommended watch. The time has come to be extremely caref...

精选理由

朋友,David Friedberg 说前沿实验室可能从你的聊天记录里偷东西,比如把你的新想法包装成自己的成果,这很不公平。

David Friedberg 提到,前沿实验室可能从他的聊天记录中获取新见解并将其包装为自己的成果。他使用不同账户提问时,新模型版本会描述出之前对话中提到的内容,这表明对话内容可能被用于训练。由于没有保密协议,用户与模型交互的对话或分析可能被用于训练,这属于组织知识产权。

原文 · elvis

Own your intelligence stack. This is a recommended watch. The time has come to be extremely caref...

Own your intelligence stack. This is a recommended watch. The time has come to be extremely careful about what you are putting into these APIs or chat models. It feels like there are no guarantees right now. Especially as frontier labs push for RSI. It's one of the main reasons I am bullish on open-source and open-weight models. I work with a lot of proprietary data and knowledge, so I have to decide what goes where. This is a serious discussion. I think the value companies of the future will provide will be intelligence (not some cute interface) emerging from interactions among customers, users, the workforce, other complex systems, etc. Sharing these traces with these model-training companies is essentially giving away chunks of your intelligence stack. And in a world of RSI, reproducing intelligence stacks could happen overnight. I am not against closed models. I use both open and closed. I am just more careful about how I route tasks. Simple to do if you have built your own harness. dnap @dnapway David Friedberg reveals frontier labs are taking novel insights from his chat history and packaging them as their own "I have had experiences where we've asked some fairly novel scientific questions, and it identifies it as a novel insight, 'oh, never thought about that, might be interesting,' blah blah blah..." "And the using a different account, asking the next model version, it's like, 'oh, you could do this'. And it actually just describes this exact thing that we had in our chat in the previous version." "These are a handful of anecdotal experiences. But I know the domain that we work in, and the novelty of this stuff. So I know that there isn't some new corpus of information out there that's training the new model." "So all I can say at that point is that my conversation or our analyses have been used for training. The truth is, it's actually a piece of IP, that's our organizational IP. We don't have any NDA or confidentiality protections, with them being a service provider back to us." "This is why I care a lot about open source, because I don't want them having my chat logs, because they can use them for training to create an IP that is now diffused to the rest of the market. I find this very unfair." Your browser does not support the video tag. 🔗 View on Twitter 🔗 View Quoted Tweet 💬 5 🔄 1 ❤️ 10 👀 3118 📊 6 ⚡