Simon Willison 评 EmbeddingGemma 2:Apache 2.0 许可对嵌入模型至关重要
EmbeddingGemma 2
Simon Willison 评论 EmbeddingGemma 2 采用 Apache 2.0 许可,认为嵌入模型不应使用闭源托管方案。他指出嵌入模型通常需计算并存储数千乃至数百万条向量,一旦厂商停供,用户需为重新计算存量向量付费;2024 年 4 月 OpenAI 曾提出承担用户用新模型重新嵌入内容的费用,但他认为不能指望所有供应商都如此。
My comment on EmbeddingGemma 2 — Hacker News.
I really appreciate that EmbeddingGemma 2 is under the Apache 2.0 license.
For embedding models in particular, I don't think it makes sense to use a closed, proprietary, hosted-only model.
Most applications of embedding models involve calculating thousands or even millions of embedding vectors and storing them for later comparison.
If your model is proprietary, the vendor is likely someday going to decide to stop offering that model. They'll have a better model to replace it, but you still need to pay to re-calculate those millions of stored existing vectors.
(In April 2024 OpenAI offered to "cover the financial cost of users re-embedding content with these new models" - https://openai.com/index/gpt-4-api-general-availability/ - but I don't think that's something we can rely on from every provider.)
Notably, I don't want to host the model myself. I'd much rather pay a provider for a hosted model while knowing that if they ever stop hosting it I can run the open weights version myself - or find another vendor who can do that for me.
Tags: google, ai, generative-ai, embeddings, gemma
来源:Simon Willison 博客 · simonwillison.net