What PDF is
PDF describes a page: glyphs at coordinates, not rows and columns. A table in a PDF is a visual arrangement, not a data structure.
Reports, invoices, statements and anything sent to be read rather than processed.
What JSON Lines is
JSON Lines is one complete JSON object per line, with no enclosing array and no commas between records. The file is not valid JSON as a whole, and that is the point.
Log pipelines, machine learning datasets and anything that appends records or reads them in a stream.
What changes when you convert PDF to JSON Lines
Every character in a text-based PDF carries its position on the page. Characters sharing a baseline are one row. Columns are found from the vertical strips no character ever occupies, which is why extraction works cleanly on a document laid out with spacing and badly on one laid out with ruled lines. A header repeated at the top of each page is detected and dropped rather than landing in the middle of the data.
Every row becomes one line. You can append to the file forever without rewriting it, and a reader can process it a line at a time without holding the whole thing in memory.
What carries over from PDF to JSON Lines
PDF records 2 things about a table that JSON Lines has no way to hold.
JSON Lines wants a type for each column, and PDF does not record one, so each column is typed from what its values look like. A column of digits that should stay text - a zip code, a phone number, a leading-zero id - is the usual thing to check afterwards.
A line break inside a cell ends the row in JSON Lines, so line breaks are replaced rather than carried through.
Anything visual in the PDF - weight, alignment, colour, column widths - has no counterpart in JSON Lines and is dropped. The values are what survives.
The result can be read a record at a time, and appended to by adding to the end of it. That is worth having for a table too large to hold in memory, and it means a JSON Lines file that is cut off part way through still gives you every record before the cut.
This is the direction that recovers structure: PDF has no table in it to read, only an arrangement that looks like one, so the rows and columns are inferred rather than read.