PDF to JSON Extractor — Tables, Fields & Structure
Input: PDF URLs, many per run. Output: one row per document — tables as rows and columns, text in reading order, Markdown for RAG, metadata, outline, form fields, key fields (invoice no., dates, totals, IBAN, VAT). No OCR: a scanned page is flagged and not charged.
33%
users