我们非常高兴地宣布,Llama.cpp 的创建者 GGML 团队正式加入 Hugging Face,共同致力于推动未来 AI 的开放发展。🔥 Georgi Gerganov 及其团队将加入 HF,目标是随着本地 AI 在未来几年持续取得指数级进步,扩大并支持 ggml 和 llama.cpp 背后的社区。
我们与 Georgi 及其团队已经合作了相当长一段时间(团队中甚至已有像 Son 和 Alek 这样出色的 llama.cpp 核心贡献者),因此这一过程非常自然。
llama.cpp 是本地推理的基础构建模块,而 transformers 是模型定义的基础构建模块,这简直是天作之合。❤️
对于 llama.cpp 这个开源项目及其社区来说,会有什么变化?
变化不大——Georgi 和团队仍将 100% 的时间投入到维护 llama.cpp 上,并在技术方向和社区管理方面拥有完全的自主权和领导权。HF 将为该项目提供长期可持续的资源支持,提升项目成长和繁荣的机会。该项目将继续像现在一样,保持 100% 开源并由社区驱动。
技术重点
llama.cpp 是本地推理的基础构建模块,而 transformers 是模型和架构定义的基础构建模块,因此我们将致力于确保未来能够尽可能无缝地(几乎是“一键式”)将 transformers 库中作为模型定义“唯一真实来源”的新模型部署到 llama.cpp 中。
此外,我们还将改进基于 ggml 的软件的打包和用户体验。随着我们进入本地推理成为云端推理有意义且具有竞争力的替代方案这一阶段,改进和简化普通用户部署及访问本地模型的方式至关重要。我们将努力让 llama.cpp 变得无处不在,随时可用。
我们的长期愿景
我们的共同目标是,为社区提供基础构建模块,以便在未来几年内让开源超级智能惠及全世界。
我们将与不断壮大的本地 AI 社区携手实现这一目标,持续打造能在我们设备上尽可能高效运行的终极推理栈。
We are super happy to announce that GGML, creators of Llama.cpp, are joining HF in order to keep future AI open. 🔥 Georgi Gerganov and team are joining HF with the goal of scaling and supporting the community behind ggml and llama.cpp as Local AI continues to make exponential progress in the coming years.
We've been working with Georgi and team for quite some time (we even have awesome core contributors to llama.cpp like Son and Alek in the team already) so this has been a very natural process.
llama.cpp is the fundamental building block for local inference, and transformers is the fundamental building block for model definition, so this is basically a match made in heaven. ❤️
What will change for llama.cpp, the open source project and the community?
Not much – Georgi and team still dedicate 100% of their time maintaining llama.cpp and have full autonomy and leadership on the technical directions and the community. HF is providing the project with long-term sustainable resources, improving the chances of the project to grow and thrive. The project will continue to be 100% open-source and community driven as it is now.
Technical focus
llama.cpp is the fundamental building block for local inference, and transformers is the fundamental building block for definition of models and architectures, so we’ll work on making sure it’s as seamless as possible in the future (almost “single-click”) to ship new models in llama.cpp from the transformers library ‘source of truth’ for model definitions.
Additionally, we will improve packaging and user experience of ggml-based software. As we enter the phase in which local inference becomes a meaningful and competitive alternative to cloud inference, it is crucial to improve and simplify the way in which casual users deploy and access local models. We will work towards making llama.cpp ubiquitous and readily available everywhere.
Our long term vision
Our shared goal is to provide the community with the building blocks to make open-source superintelligence accessible to the world over the coming years.
We will achieve this together with the growing Local AI community, as we continue to build the ultimate inference stack that runs as efficiently as possible on our devices.