PDF reading order inspector
Inspect coordinate-sorted PDF text blocks and page bounds.
PDF reading order inspector
Drop your files here
or choose files from your device
Sida loo isticmaalo PDF reading order inspector
PDF reading order inspector shows how the text blocks on each PDF page are ordered once sorted by position, along with each page's width and height. Blocks whose vertical positions are within 3 points count as one line and are read left to right. The JSON report covers 1 to 100 pages and helps diagnose jumbled text extraction.
- Choose a source file. If the tool accepts multiple files, arrange them in the order you want.
- Set maximum pages.
- Run pdf reading order inspector, review the output, then use the available copy or download controls.
Waxa qalabkani taageero
Extracts the existing PDF text layer in coordinate order. Scans require OCR first. Columns and complex layouts may need manual corrections; JSON retains page numbers and bounding boxes.
- Maximum pages
- Range: 1 to 100
Xadka iyo habaynta
Image inputs: up to 32 megapixels, batches up to 10 images. Audio/video uploads: 50 MB; transcription: 120 seconds. Edited videos: 20 seconds, 4 fps and 640 pixels wide. Translation: English/Spanish, 30,000 characters; subtitles: 200 cues. Neural speech: 1,000 English characters. Model downloads and memory requirements vary by device.
Isha tusaalaha
A useful tool makes everyday work easier. Review the results before sharing.
Su’aalaha badanaa la isweydiiyo
Why is the text from my PDF extracted in the wrong order?
PDFs store text as positioned fragments, not as a reading sequence. This inspector lists each block with its coordinates in sorted order, so you can see where columns, sidebars or headers interleave with the main text.
How does the inspector decide the reading order?
Blocks are sorted by their distance from the top of the page. Blocks less than 3 points apart vertically are treated as the same line and ordered left to right. Multi-column layouts may still need manual correction.
Can I check the reading order of a scanned PDF?
Not until it has a text layer. Scanned pages contain only images, so the tool reports that no text layer was found. Run OCR on the PDF first, then inspect the result.
Where is my input processed?
Processes source in your browser. Copy and download are explicit actions; source is not saved to account history.
What are the input limits?
Image inputs: up to 32 megapixels, batches up to 10 images. Audio/video uploads: 50 MB; transcription: 120 seconds. Edited videos: 20 seconds, 4 fps and 640 pixels wide. Translation: English/Spanish, 30,000 characters; subtitles: 200 cues. Neural speech: 1,000 English characters. Model downloads and memory requirements vary by device.