Jetformat

Extract detected PDF tables as data

Use jetformat pdf tables to extract detected PDF tables as data.

By Jetformat

Table detection is heuristic and reads an existing text layer. Merged cells, unusual spacing, and scanned pages need additional review.

Steps

  1. Start with a local PDF. Replace the sample paths and identifiers with your own.
  2. Run the command below. Read the returned data before choosing a follow-up edit.
  3. Check the result as described below before using it in another job.
jetformat pdf tables book.pdf

Options for this task

Option Meaning
--json Write JSON output

Check the result

Compare headers, row counts, and several numeric values against the PDF before using the table in calculations.

Notes and limits

This is a free utility or read operation. It does not need a paid conversion to inspect the result.

Run jetformat pdf tables --help to compare options with your installed version.

See the pdf command reference for the surrounding workflow.

Last updated