Aravind Srinivas 宣布开源多模态嵌入模型 pplx-embed-v2-late
中文全文 · AI 翻译
我们正在开源 pplx-embed-v2-late,这是用于文本和图像的多向量嵌入,9B 和 0.6B 参数,在一个共享的嵌入空间中。您可以使用 9B 来索引多模态数据,并使用 0.6B 在设备上进行查询。这还使您能够无需 OCR 即可搜索 PDF 页面。在 MADQA 上得分 92.4%,在 BrowseComp+ 上得分 64%。权重现已在 @huggingface 上可用。
对照原文
We’re open-sourcing pplx-embed-v2-late, multi-vector embeddings for text and images, 9B and 0.6B, in one shared embedding space. You can use these to index multimodal data with 9B, and query on device with 0.6B. This also enables you to search over PDF pages with no OCR. And scores 92.4% on MADQA, 64% on BrowseComp+. Weights available on @huggingface now.