ToolCabana

Document to structured JSON

Export PDF text blocks with page numbers and bounding boxes.

Favorites are saved in this browser. Find them in My favorites.

Document to structured JSON

Your input stays in this browser unless stated otherwise
Clear your input, files and result, and restore the default settings.

Drop your files here

or choose files from your device

Com utilitzar Document to structured JSON

Document to structured JSON exports the text of a PDF as JSON, listing every page with its width and height and each text block with its text, x and y position, width and height. Blocks are sorted top to bottom and left to right. You can process up to 100 pages, and the file downloads as document.json.

  1. Choose a source file. If the tool accepts multiple files, arrange them in the order you want.
  2. Set maximum pages.
  3. Run document to structured json, review the output, then use the available copy or download controls.

Què admet aquesta eina

Extracts the existing PDF text layer in coordinate order. Scans require OCR first. Columns and complex layouts may need manual corrections; JSON retains page numbers and bounding boxes.

Maximum pages
Range: 1 to 100

Límits i processament

Image inputs: up to 32 megapixels, batches up to 10 images. Audio/video uploads: 50 MB; transcription: 120 seconds. Edited videos: 20 seconds, 4 fps and 640 pixels wide. Translation: English/Spanish, 30,000 characters; subtitles: 200 cues. Neural speech: 1,000 English characters. Model downloads and memory requirements vary by device.

Font de l'exemple
A useful tool makes everyday work easier. Review the results before sharing.

Preguntes freqüents

What does the JSON output of a PDF look like?

It is an object with a pages array. Each page has page, width, height and blocks, and each block has text, x, y, width and height, so you can locate every text run on the page.

What units are the coordinates in the PDF JSON?

Positions and sizes use PDF points at a scale of 1, matching the page width and height in the same output. The y value is measured from the top of the page, so larger values sit lower down.

Can I get JSON from a scanned PDF?

Only after OCR. The tool reads the existing text layer and reports that no text layer was found when pages are images. Make the PDF searchable with OCR first, then export it.

Where is my input processed?

Processes source in your browser. Copy and download are explicit actions; source is not saved to account history.

What are the input limits?

Image inputs: up to 32 megapixels, batches up to 10 images. Audio/video uploads: 50 MB; transcription: 120 seconds. Edited videos: 20 seconds, 4 fps and 640 pixels wide. Translation: English/Spanish, 30,000 characters; subtitles: 200 cues. Neural speech: 1,000 English characters. Model downloads and memory requirements vary by device.