Kymata Labs/The Living IndexesBuilt by tekvisions ↗
The Document Index / OCR Engines / #10
PaddlePaddle

PaddlePaddle/PaddleOCR

by PaddlePaddle · OCR Engines · updated 1mo ago

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

78
momentum
89,377
stars
11,322
forks
#10
rank
ai4sciencechineseocrdocument-parsingdocument-translationkieocrpaddleocr-vlpdf-extractor-ragpdf-parserpdf2markdownpp-ocrpp-structure
View on GitHub →