AI 导读
PaddlePaddle 正式发布 PaddleOCR-VL 1.6,在 OmniDocBench 上达到 96.33% 的 SOTA,超过开源和商业方案。该版本在表格识别、经典文本和稀有字符识别上有明显提升,印章检测与图表理解也更稳,并与 1.5 版本完全架构兼容,零迁移成本。作者认为这降低了将文档智能喂给 LLM 的门槛,RAG 的上限往往取决于输入数据的干净程度。
正文
最近开发了一个OCR的 工具,疯狂给干法律的客户案例!
效果非常好,很合适~
但也遇到有些错乱和不好的结果
金融合同、法律文件、研究报告、历史档案,这些东西里公式、表格、印章、稀有字符混在一起,传统工具经常认错或者直接漏掉,导致后续LLM输出质量直接拉低。
今天PaddlePaddle把PaddleOCR-VL 1.6正式发布了。
它在OmniDocBench上刷到96.33%的SOTA,把开源和商业方案同时甩在身后。
表格识别、经典文本、稀有字符都有明显提升,印章检测、图表理解也更稳。
最实用的是,它和1.5版本完全架构兼容,零迁移成本,拿来就能用。
以前大家总觉得RAG的瓶颈在模型参数或者检索算法,现在看,真正决定上限的往往是输入数据的干净程度。
这份高精度解析能力,直接把文档智能喂给LLM的门槛又往下拉了一大截。
🚀PaddleOCR-VL 1.6 Officially Released! We are thrilled to announce the official release of PaddleOCR-VL 1.6 — this version has set a new SOTA record of 96.33% on OmniDocBench, outperforming both open-source and proprietary solutions in text, formula, and table recognition. 📊 Key Highlights: 🔸 Ranked #1 on OmniDocBench v1.5 and Real5-OmniDocBench as well 🔸 Significant improvements in table, classic text, and rare character recognition 🔸 Enhanced seal, spotting, and chart recognition 🔸 Fully compatible with v1.5 architecture — zero migration, plug-and-play From financial contracts and legal documents to research reports and historical archives — empower your document intelligence workflows by providing high-quality data to large language models (LLMs) and retrieval-augmented generation (RAG) systems. ✨ Industry-leading accuracy. Zero migration. Plug-and-play. 🔗 Read more: github.com/PaddlePaddle/Padd… #PaddlePaddle #PaddleOCR #AI #ComputerVision在 X 查看被引用的帖子
来源:@berryxia · x.com