Transformers 与 GGML 合作优化性能
transformers x @ggml_org collab 🐐
Transformers 现在使用 GGML 内核,性能提升到 llama.cpp 水平,开发者值得关注。
Transformers 库现已通过 GGML 内核实现与 llama.cpp 相同的性能表现。GGUF 文件支持已存在数年,此次合作引入了 `kernels` 库。GGML 组织提供了这些高性能内核,使 Transformers 能够运行更高效。
transformers x @ggml_org collab 🐐
transformers x @ggml_org collab 🐐 Lysandre @LysandreJik Transformers has supported loading GGUF files for a few years now, by unquantizing them. Thanks to @_marcsun , we're now using GGML kernels through the `kernels` library to run at the same performance as llama.cpp Huge kudos to the entire @ggml_org for making these kernels! 🔗 View Quoted Tweet 💬 0 🔄 0 ❤️ 4 👀 525 📊 1 ⚡
- Hugging Face: Blog00:00原文