PDFCraft
Extraction guides

PDF to JSON, by document type

One page per shape, and every number on them was produced by running the live engine at build time — not written from memory. Each links the sample PDF it was measured against.

DocumentFieldsTablesTime
Extract a commercial invoice61 (9 rows)211 ms
Extract a bank statement41 (38 rows)35 ms
Extract a packing slip51 (12 rows)12 ms
Extract a purchase order51 (8 rows)13 ms
Extract a remittance advice41 (14 rows)15 ms
Extract a utility bill42 (9 rows)11 ms
Extract a payslip52 (6 rows)10 ms
Extract a bill of lading71 (6 rows)9 ms
Extract a insurance claim form61 (7 rows)14 ms
Extract a rent roll41 (24 rows)13 ms
Extract a lab report51 (7 rows)12 ms
Extract a timesheet41 (7 rows)9 ms
Extract a credit note51 (6 rows)8 ms
Extract a expense report51 (11 rows)9 ms
Extract a quotation51 (7 rows)9 ms
Extract a balance sheet41 (11 rows)6 ms
Extract a profit and loss statement31 (9 rows)5 ms
Extract a bill of materials41 (14 rows)9 ms
Extract a certificate of analysis51 (6 rows)8 ms
Extract a work order62 (7 rows)10 ms
Extract a aged receivables report31 (18 rows)7 ms
Extract a dividend statement41 (12 rows)16 ms
Extract a price list41 (16 rows)9 ms
Extract a explanation of benefits51 (5 rows)9 ms
Extract a delivery note51 (9 rows)6 ms
Extract a inspection report51 (8 rows)7 ms
Extract a shipping manifest61 (30 rows)10 ms

Sample documents are synthetic — real ones of these types cannot be published — but the extraction output is the engine’s own, regenerated on every build.