# Translate scanned document

> Translate extracted PDF text into Markdown.

[Open tool](https://www.toolcabana.com/bm/document-translate) · [PDF and documents](https://www.toolcabana.com/bm/category/pdf-and-documents)

Tool ID: document-translate. Requested language: bm. Description language: en. Complete guide translation: no; untranslated sections use English.

## Overview (en)

Translate scanned document extracts the text layer from a PDF and translates it between English and Spanish with Marian models that run in your browser after a first download. It processes 1 to 30 pages, 10 by default, and up to 30,000 characters per run. The translation downloads as Markdown with a Page label for each page.

## Supported tasks (en)

English ↔ Spanish translation uses browser Marian models, downloaded on first run. Document translation exports extracted text in Markdown and does not rebuild page layout.

## Steps (en)

1. Choose a source file. If the tool accepts multiple files, arrange them in the order you want.
2. Set source language code, target language code, maximum pages.
3. Run translate scanned document, review the output, then use the available copy or download controls.

## Settings

- Source language code (en; key: source; type: text)
- Target language code (en; key: target; type: text)
- Maximum pages (en; key: maxPages; type: number); minimum: 1; maximum: 30

## Limitations (en)

Image inputs: up to 32 megapixels, batches up to 10 images. Audio/video uploads: 50 MB; transcription: 120 seconds. Edited videos: 20 seconds, 4 fps and 640 pixels wide. Translation: English/Spanish, 30,000 characters; subtitles: 200 cues. Neural speech: 1,000 English characters. Model downloads and memory requirements vary by device.

## Example input

```text
A useful tool makes everyday work easier. Review the results before sharing.
```

## Privacy and connections (en)

Source files and text stay in this browser. The first run downloads public models from Hugging Face and local runtime assets from the site. Models may remain in the browser cache. No source upload to Vercel, Supabase or a conversion service is required.

## Questions (en)

### Which languages can the document translator handle?

Only English to Spanish and Spanish to English. Set the source and target codes to en and es in either order; any other combination is rejected with an error.

### Will the translated document keep its original layout?

No. The tool extracts the text, translates it and exports Markdown labeled by page. Images, columns, fonts and page design are not rebuilt, so the result is a text draft to review.

### Can I translate a scanned PDF that has no text layer?

Not directly. The tool reads existing PDF text and reports when no text layer is found. Run PDF OCR first to make the scan searchable, then translate the resulting file.

### Where is my input processed?

Source files and text stay in this browser. The first run downloads public models from Hugging Face and local runtime assets from the site. Models may remain in the browser cache. No source upload to Vercel, Supabase or a conversion service is required.

### What are the input limits?

Image inputs: up to 32 megapixels, batches up to 10 images. Audio/video uploads: 50 MB; transcription: 120 seconds. Edited videos: 20 seconds, 4 fps and 640 pixels wide. Translation: English/Spanish, 30,000 characters; subtitles: 200 cues. Neural speech: 1,000 English characters. Model downloads and memory requirements vary by device.

## Related tools

- [PDF to Markdown](https://www.toolcabana.com/bm/pdf-to-markdown)
- [Document to structured JSON](https://www.toolcabana.com/bm/document-to-json)
- [Table image to CSV](https://www.toolcabana.com/bm/table-image-to-csv)
- [Invoice field extractor](https://www.toolcabana.com/bm/invoice-extractor)
