模型精选

Perplexity 发布 9B 与 0.6B 搜索嵌入模型,Q2D-Web 基准超过 nemotron-embed-8b

精选理由

Perplexity 放出了自家的搜索嵌入模型评测,0.6B 小模型就追平了大号 nemotron-embed-8b,做 RAG 检索的可以看看。

Perplexity 在 X 上公布基于 Q2D-Web 基准的评测结果,该基准包含约 70,000 条来自生产环境的搜索查询,覆盖 190M 网页文档。其 9B 和 0.6B 模型分别取得 74.8% 和 73.6% 的 Recall @1000。作为对比,nemotron-embed-8b 为 69.3%。0.6B 小模型以更少参数接近 9B 版本的成绩。

原文 · Perplexity

On Q2D-Web, about 70,000 production-derived search queries over 190M web documents, the 9B and 0.6B models reach 74.8% and 73.6% Recall @1000 , against 69.3% for nemotron-embed-8b. 💬 2 🔄 0 ❤️ 5 👀 536 📊 2 ⚡