Sign in Sign up free Account

PDF to JSON Converter

Convert PDF to JSON free: every page's text line by line, and each table as an array of objects with numbers typed. Or just the tables. In your browser; nothing is uploaded.

Paste cells or a Google Sheets link

Output


        

How it works

  1. Add your PDF

    Drop in the PDF. The text is read from the file itself - nothing is guessed from a picture - and rebuilt into rows and columns.

  2. Check the table

    Every column is shown with what it holds - numbers, dates or text - so you can see it was read right before you save it.

  3. Download the JSON

    Or copy it, or send it straight to a new Google Sheet. Nothing is uploaded at any point.

PDF to JSON Converter: a practical guide

Invoices, statements and reports arrive as PDFs, and code wants JSON. PDF to JSON reads the text inside the file and hands it back as data: every page, every line, and every table as rows keyed by their headers.

Numbers are typed as numbers, and the PDF is read in your browser, so nothing is uploaded on the way to your script.

PDF and JSON: what changes

  • PDF: A page description: geometry, text, images and fonts, fixed so it looks the same everywhere. A page, not artwork. It can equally hold one full-page bitmap and be no more useful than a JPG.
  • JSON: Structured data as nested objects and lists, the language of web APIs. Exact structure and types that every programming language reads.

How the PDF to JSON converter works

A PDF is read in your browser: the characters stored in the file are read, not guessed from a picture, and tables are found from the positions of the text on the page, row by row and column by column. Text that wraps inside a cell is joined to its row, and values like 1,234.50, $1,200 and 7% are recognised as numbers.

It is written as JSON with numbers written as numbers, and a preview of the first rows is shown before you download.

What the JSON is used for

  • APIs and apps
  • Configuration
  • Data exports

Why Veconvert for PDF to JSON

  • Nothing silently changed. Zeros, IDs and quoted text survive.
  • Preview first. See the rows before you save.
  • Private by design. Bank statements and customer lists never leave your computer.

Tips for the best JSON

  • Use a PDF with real text: if you cannot select the words in a PDF reader, it is a scan and needs OCR first.
  • Make sure the first row holds the column names.
  • Remove totals rows and notes above the table for the cleanest result.

On the blog: How to Convert PDF to JSON Free, Step by Step

PDF to JSON converter: text and tables as data

  • pdf to json
  • convert pdf to json
  • pdf to json converter
  • pdf table to json
  • extract data from pdf to json

PDF to JSON here reads the text stored in the PDF itself and writes it as data you can use in code: one object per page with its lines of text from top to bottom, and the tables on that page as arrays of objects keyed by their header row.

To convert pdf to json, drop in the PDF, pick the pages if you only want some, and download the .json file or copy it. JSON Lines gives one line per page, for scripts and pipelines that read a page at a time.

Only need the numbers? Choose Array of objects and this pdf to json converter writes the table rows alone, the same table the PDF to CSV page finds: wrapped cell text joined to its row and a header repeated on every page kept once. That is the quickest way to turn a pdf table to json.

Numbers come out as JSON numbers and yes/no columns as true/false, unless you untick Typed values. Because the characters are read from the file rather than guessed from a picture, figures are exact. A scanned PDF has no text to read: run it through OCR first.

How to convert PDF to JSON

  1. Drop in the PDF (a password-protected one asks for its password).
  2. Choose Whole document for pages, text and tables, or Array of objects for the tables only.
  3. Download the .json file, or copy it.

People also search for

  • pdf to csv
  • pdf to excel
  • json to csv
  • pdf to text
  • csv to json

Why it matters

Private by design

Bank statements, payroll and customer lists never leave your computer: the file is read and written in your browser.

Nothing silently changed

Leading zeros, long ID numbers, commas inside quotes and line breaks inside cells come through exactly as they were.

Numbers stay numbers

Values like 1,234.50, $1,200 and 7% are recognised, so Excel and JSON get real numbers rather than text.

Reviews

No reviews yet. Used this tool? Yours could be the first.

Used this tool? Sign in to write a review.

Questions

How do I convert PDF to JSON?

Drop your PDF into the box at the top of this page and press Convert to JSON. The conversion runs in your browser in a few seconds. Check the result, then download a JSON.

What is the best PDF to JSON converter?

One that gives you your data in the new format, exactly as it was, shows the result before you pay, and does not upload your file. That is what this page is built to do.

What does the JSON look like?

Whole document gives {"file", "pageCount", "pages": [{"page": 1, "text": [lines], "tables": [[{header: value}]]}]}. Array of objects gives the table rows only, one object per row.

Can it read a scanned PDF?

No. A scan is a picture of the page with no text inside it. Run it through OCR (Adobe Acrobat, or open it with Google Docs from Google Drive) and convert the result here.

Is my PDF uploaded?

No. The PDF is read in your browser, so bank statements, invoices and reports never leave your computer.

Does it work on scanned PDFs?

Only when the scan has a text layer (most scanners and Adobe Acrobat add one with OCR). A scan that is only a picture has no text to read; the page says so rather than handing back an empty file. Google Drive can add the text layer: upload the PDF and open it with Google Docs.

What if a table spans several pages?

Every page is read and the rows are joined into one table. A header row repeated at the top of each page is kept once.

What about text that wraps inside a cell?

A line that has nothing in the first column, directly under a row that does, is that row's wrapped text and is joined back to it. Switch that off if your table has genuinely empty first cells.

Are commas, quotes and line breaks inside cells handled?

Yes. Cells are read and written to the CSV standard (RFC 4180): a cell with a comma, a quote or a line break is quoted, and quotes inside are doubled, so nothing splits into the wrong column.

Will leading zeros and long numbers survive?

Yes. Cells are kept as the text they were: 00123 stays 00123 and a 16-digit account number is not turned into 1.23E+15. In Excel output, such cells are stored as text for exactly that reason.

Does it work with Google Sheets?

Paste a Google Sheets link (shared as "Anyone with the link") or copy the cells and paste them. To send a result to Google Sheets, use Open in Google Sheets: the table is copied and a new sheet opens, ready for Ctrl+V.

Is my data uploaded?

No. The whole conversion runs in your browser, so the file never leaves your computer.

Can it open a password-protected PDF or Excel file?

Yes. Drop it in and you are asked for the password; the file is unlocked in your browser and the password is never sent or saved. PDFs locked with RC4 or AES (up to AES-256, PDF 2.0) and Excel workbooks saved with Encrypt with Password both open. A PDF that only restricts copying or printing opens without asking. Without the password nothing is guessed: the file stays shut.

How big a file can it handle?

Hundreds of thousands of rows are fine on a normal laptop; the preview shows the first rows and the download has them all.