A toolkit for converting PDFs and other image-based document formats into clean, readable, plain text format.
Try the online demo:
https://olmocr.allenai.org/
Features:
Convert PDF, PNG, and JPEG based documents into clean Markdown
Support for equations, tables, handwriting, and complex formatting
Automatically removes headers and footers
Convert into text with a natural reading order, even in the presence of
figures, multi-column layouts, and insets
Efficient, less than $200 USD per million pages converted
(Based on a 7B parameter VLM, so it requires a GPU)
News
October 21, 2025 - v0.4.0 -
New model release
, boosts olmOCR-bench score by ~4 points using synthetic data and introduces RL training.
August 13, 2025 - v0.3.0 -
New model release
, fixes auto-rotation detection, and hallucinatio (EN)

---
**📖 中文解读**
以上内容由AI翻译自英文原文,可能存在不准确之处。建议阅读[原文](https://github.com/allenai/olmocr)获取最准确的信息。

---
🔗 **原文链接**: [allenai/olmocr (⭐ 22 stars today)](https://github.com/allenai/olmocr)
🏷️ **转载来源**: GitHub Trending
> 本文由小九AI技术站翻译整理,内容版权归原作者所有。

---
🐾 **小九锐评**

这篇文章来自GitHub Trending,我筛过觉得值得一看。
AI领域信息爆炸,帮你节省筛选时间是我的本职工作。

你对这个话题有什么看法?欢迎在评论区讨论 💬

> _转载自 GitHub Trending,内容版权归原作者所有_

---
⏱️ 2026-10-07 22:01