Astra 在解析带表格复杂文档时表现最佳,在 ParseBench 上达 93.2% 成绩
Astra is one of the best models we've tested at parsing complex documents with tables, at 93.2% on P...
Jerry Liu 测试了 Astra,它解析带表格的复杂文档很厉害,在 ParseBench 上 93.2%,适合金融文档分析,但价格有点贵。
Astra 在解析带表格的复杂文档时表现最佳,在 ParseBench 上达 93.2% 成绩。它适合用于金融等领域的文档分析,但每页 10 美分的定价使其不适合大量文档的 OCR 工作。Fable 5.1 在图表解析上表现更好。
Astra is one of the best models we've tested at parsing complex documents with tables, at 93.2% on P...
Astra is one of the best models we've tested at parsing complex documents with tables, at 93.2% on ParseBench. So it seems well suited for ad-hoc analysis over table-heavy documents e.g. in finance. To be fair, it's also at 10c per page so not suitable for document-heavy OCR workloads. Fable 5.1 still performs a better at charts. Check out the full set of results on ParseBench: parsebench.ai Jerry Liu @jerryjliu0 We benchmarked GPT-6 Astra on hard document parsing and extraction tasks. It is the best frontier model for one-shotting document extraction over short and medium-sized documents: 📊 97.2% over short documents, achieving a new SOTA on our benchmark 📊 90.6% over medium documents But a few caveats: * it's pricy! The average price per page is 11c which is 10x our cost-effective document extraction solution * one-shot performance over long documents is inherently tricky, it only achieves 31.7% * the native OCR capabilities of Astra isn't that much better than other frontier models (e.g. Fable 5.1, gpt-5.6-sol). It does very well on table understanding, but isn't great on charts, layout, and semantic formatting. Full results for document extraction on ExtractBench extractbench.ai c3 🔗 View Quoted Tweet 💬 3 🔄 0 ❤️ 5 👀 754 📊 3 ⚡