返回 X 名人动态

Jerry Liu 介绍 Extract v2.5 文档提取代理

中文全文 · AI 翻译

今天我们推出 Extract v2.5 - 一系列针对文档提取优化的前沿代理。

这些代理(成本效益型、代理型、代理增强型)针对提取值的准确性与来源对应 进行了优化。我们的提取代理在性能上超越 Opus 5.5 和 GPT-6 Sol,同时成本可降低 30%,最高可降至原来的四分之一。

我们在复杂提取方面取得了巨大改进,包括:

  • 长列表(86.1% -> 95.5%,在我们的代理层级上)
  • 跨页记录(85.5% -> 96.5%)
  • 扫描表单(90.9% -> 95.7%)

我们还推出了以下功能: ✅ 高级引用:我们为推断字段的所有支持值定位边界框,即使没有精确匹配。 ✅ 结构化推理:我们根据类型、布局和信息定制文档提取算法

我们的提取代理在 ExtractBench 上达到了价格-性能的最先进水平,覆盖广泛的成本点。我们是处理任何复杂性文档提取的最佳工具。

博客:https://www.llamaindex.ai/blog/introducing-extract-v2-5

所有这些功能均可在 LlamaParse 上使用:https://cloud.llamaindex.ai/ 。快来试试吧!

对照原文

Today we’re introducing Extract v2.5 - a series of frontier agents tuned for document extraction. The agents (cost-effective, agentic, agentic plus) are tuned for value accuracy and grounding. Our extraction agents outperform Opus 5.5 and GPT-6 Sol while being 30%-4x cheaper. We’ve made massive improvements on complex extraction over * long lists (86.1% -> 95.5% on our agentic tier) * records spanning pages (85.5% -> 96.5%) * scanned forms (90.9% -> 95.7%) We’ve also launched the following features: ✅ Advanced citations: we locate bounding boxes for all supporting values for an inferred field, even if there's not an exact match. ✅ Structural Reasoning: we tailor document extraction algorithms depending on the type, layout, and information Our extraction agents are SOTA in price-performance on ExtractBench, across a wide range of cost points. We are the best tool for document extraction across documents of any complexity. Blog: https://t.co/TTlTdECU03 All of these are available on LlamaParse: https://t.co/XYZmx5TFz8 . Come check it out!

老杨AI实操

微信扫一扫,添加好友

老杨AI实操的微信好友二维码

手机可长按保存图片,再到微信中识别二维码

保存二维码