百度出了个Unlimited OCR,一次扫描几十页文档,内存不涨还拿了OCR基准第一,处理长文档效率提升明显。
百度发布了Unlimited OCR模型,一次性可处理数十页文档,而以往系统最多处理约10页。该模型通过改进注意力机制,使内存使用量不随页数增加而增长。Unlimited OCR目前在最重要的OCR基准测试中排名第一。
Baidu's "Unlimited OCR" processes dozens of document pages in one pass by treating memory like human forgetting
Baidu's Unlimited OCR reads dozens of document pages in a single pass, where previous systems topped out at about ten. A modified attention mechanism keeps memory use flat no matter how many pages the model processes. It currently holds the top spot on the most important OCR benchmark. The article Baidu's "Unlimited OCR" processes dozens of document pages in one pass by treating memory like human forgetting appeared first on The Decoder .