Choose for a MIT-licensed Python library with a simpler API for general documents; it lacks PDF/UA accessibility tagging.
Extract and convert data from any document, images, pdfs, word doc, ppt or URL into multiple formats (Markdown, JSON, CSV, HTML) with intelligent structured data extraction and advanced OCR.
- 1.5k
- Python
- MIT