Extract structured data from documents using custom or auto-generated schemas to process various file formats including PDF, images, and Office documents.
MIT
Congressional Documents — full-text search and retrieval over the official
High-fidelity PDF to structured Markdown conversion and document field extraction.