Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
-
Updated
Jul 22, 2026 - Python
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
OpenOCR: An Open-Source Toolkit for General-OCR Research and Applications, integrates a unified training and evaluation benchmark, commercial-grade OCR and Document Parsing systems, and faithful reproductions of the core implementations from a wide range of academic papers.
A program for extracting hard coded (burned in) subtitle from a video and generating an external subtitle.
Boosting Document Intelligence
This repository offers a simple OCR library that leverages system APIs like VisionKit and Media OCR for accurate text recognition. Check out the examples and start integrating with ease! 🐙✨
Open Models For Document Intelligence
极简 Windows 框选 OCR:悬浮球 → 框选 → 百度 PP-OCR tiny → 文本进剪贴板。
本工具用于对视频进行初步字幕提取。它采用「抽帧 + OCR」的方式:定时从视频画面中抽取图像帧,使用 OCR 引擎识别画面中出现的文字(如内嵌字幕、标题、台标文字、弹幕等),并将各帧识别结果按时间顺序逐帧拼接,输出为一个纯文本文件(.txt)
Build a self-scaling, event-driven OCR pipeline on Kubernetes using Qwen and the GLM-OCR SDK.
To associate your repository with the chineseocr topic, visit your repo's landing page and select "manage topics."