Kymata Labs/The Living IndexesBuilt by tekvisions ↗
The Document Index / OCR Engines / #134
NanoNets

NanoNets/docstrange

by NanoNets · OCR Engines · updated 10mo ago

Extract and convert data from any document, images, pdfs, word doc, ppt or URL into multiple formats (Markdown, JSON, CSV, HTML) with intelligent structured data extraction and advanced OCR.

35
momentum
1,553
stars
138
forks
#134
rank
aidocument-parserdocument-parsingimage-to-markdownllmmarkdownocrpdf-parserpdf-to-jsonpdf-to-markdownstructured-datastructured-data-capture
View on GitHub →