论文精选

LlamaIndex发布文档提取基准测试

We built one of the most comprehensive benchmarks for document extraction, and evaluated it across a...

精选理由

LlamaIndex发布文档提取基准测试,全面评估多种系统,了解提取难度和成本与准确性的权衡,不容错过!

AI 摘要

LlamaIndex发布文档提取基准测试,评估了多种系统,包括VLMs和LlamaParse,覆盖370个企业文档和67种文档类型。CTO和联合创始人Simon Suo将进行技术讲解,探讨提取难度、现有基准测试的不足以及成本与准确性的权衡。

原文 · Jerry Liu

We built one of the most comprehensive benchmarks for document extraction, and evaluated it across a...

We built one of the most comprehensive benchmarks for document extraction, and evaluated it across a lot of different systems: ✅ one-shot frontier VLMs ✅ frontier VLMs + coding agent harnesses ✅ one-shot open weight VLMs ✅ other document extraction tools (including LlamaParse) Document extraction is an extremely diverse task that covers many different types of docs. From simple schemas over short docs (e.g. resume extraction) to complex extraction docs (credit agreements, data room bundles). @disiok is leading this webinar. Come check it out! watch.getcontrast.io/register/llama… LlamaIndex 🦙 @llama_index Every extraction API demos well on a clean invoice. But what about the scanned form, the nested table, the 40-page financial report with merged headers? We tested 14 frontier systems to find out. ExtractBench evaluates schema-guided extraction across 370 enterprise documents, 67 document types, and 4,800+ pages. Join Simon Suo, CTO and co-founder of LlamaIndex, for a technical walkthrough of the results: what makes extraction hard, where existing benchmarks miss it, and how cost trades off against accuracy across VLMs, coding agents, and specialized APIs. Wednesday, Aug 26 · 9:00 AM PT / 12:00 PM ET Save your seat 👉 lnkd.in/gMbFNHPg 8 🔗 View Quoted Tweet 💬 1 🔄 0 ❤️ 2 👀 1039 📊 1 ⚡

LlamaIndex发布文档提取基准测试 · AI 热点