How to Translate a Large PDF (Without It Falling Apart)
File-size caps, silent timeouts, and layouts that fall apart at scale — why large documents defeat most translation tools, and a workflow that holds up.
Translating a two-page letter is a solved problem. Translating a 120-page manual, a 300-page book scan, or a 60-page contract is where tools start failing in ways they don't warn you about: uploads rejected for size, jobs that spin forever and time out, output where the first thirty pages look fine and the rest never arrived. Large PDFs aren't just more pages — they hit different limits entirely.
This guide covers what those limits actually are, which approaches survive contact with a genuinely large document, and how to keep tables, figures, and page structure intact across hundreds of pages — including when the PDF is a scan with no text layer at all.
Why large PDFs break translators
- Hard file-size caps. Google Translate's document upload stops at 10 MB — a limit a scanned document blows through in a dozen pages, since each page is a full image. Many free web tools cap even lower.
- Timeouts. Tools built for short documents process pages one after another; at 100+ pages the job outlives the server's patience and dies partway, often without an error message.
- Memory and layout collapse. Rebuilding formatting takes real work per page. Tools that shortcut it return plain text, or a document where columns and tables drift further out of place the deeper you go.
- Scans multiply everything. An image-only PDF needs OCR before any translation can happen, which multiplies processing cost per page — so limits that were annoying at 10 pages become walls at 100.
First, check what kind of large PDF you have
Open the file and try to select a sentence with your cursor. If text selects, you have a digital PDF — size limits will be your main enemy, and copy-paste workflows are at least possible as a fallback. If nothing selects, you have a scanned (image-only) PDF: every workflow that starts with 'copy the text' is closed to you, and the tool you pick must OCR the pages first. Most large-document horror stories come from feeding a 200 MB scan into a tool designed for 5 MB digital files. If that's your case, start with how to translate a scanned PDF — the constraints there apply to every page of a long document.
The approaches, honestly compared
Google Translate / free web tools
Fine for small digital files, and the price is right. But the 10 MB cap rules out most scans, the output is text-first with formatting loosely reconstructed, and there's no recovery when a big job dies mid-way. If your document is under the cap and layout doesn't matter, start here; otherwise skip.
Chat AI (ChatGPT, Claude, Gemini)
Vision models genuinely read scanned pages now, and for a few pages the translation quality is good. At scale the shape of the tool betrays you: you're pasting or uploading in chunks, the model summarizes or drifts as context fills up, and what comes back is chat text — you'd still have to rebuild a 100-page document by hand. Use chat AI to understand a large document, not to produce a translated copy of one.
Human translation services
The quality ceiling, and required for certified work. But large documents are exactly where per-word pricing hurts: a 300-page book runs to thousands of dollars and weeks of turnaround. Even when you'll ultimately need a human pass, a machine-translated draft that preserves the layout makes the human step faster and cheaper.
A purpose-built document translator
Tools built for documents (rather than text) OCR each page, translate, and re-typeset the translation back into the original layout — so page 200 comes out as clean as page 2. This is the only approach where a large scanned PDF comes back as a usable document rather than a wall of text. Reglyph is built this way: it processes up to 50 pages per upload, keeps tables, stamps, and figures where they were, and gives you a side-by-side bilingual export for checking the result.
A workflow that holds up at 100+ pages
- 1Split the document into parts. For anything over ~50 pages, split the PDF into chunks at natural boundaries — chapters, sections, or exhibits. Free tools (pdftk, ilovepdf, macOS Preview) do this in seconds. Chunks also mean one failed part doesn't sink the whole job.
- 2Run a small test batch first. Translate 3–5 representative pages — one dense table, one figure-heavy page, one plain-text page — before committing the whole document. Five minutes of testing tells you whether the tool's output is worth 300 pages of waiting.
- 3Translate the chunks. Upload each part and let the tool work through it. With Reglyph, each 50-page part comes back with the original layout preserved, so the parts remain visually consistent with each other.
- 4Spot-check with a bilingual view. On a large document you won't proofread every line. Check the riskiest pages — numbers in tables, section headings, anything legally load-bearing — against the original. A side-by-side original/translation export makes this fast.
- 5Merge the parts back. The same PDF tools that split will merge. Because each part kept its layout, the reassembled document reads like the original — same page count, same structure.
Processing cost scales with pages, not megabytes. A sensible way to evaluate any tool on a big job: translate a handful of pages free first (Reglyph gives you 5 free pages every day, no credit card), confirm the layout survives, then run the full document on a paid plan — from $5 — rather than discovering problems at page 250.
Keeping the layout is the hard part — and the point
On a large document, formatting isn't cosmetic. A manual's step numbers, a contract's clause references, a book's figure captions — lose the layout and readers can no longer cross-reference anything against the original. That's why the erase-and-retypeset approach matters at scale: the original text is removed from each page image and the translation is placed where it was, so the 200-page structure the authors built survives translation intact.
If your document is in one specific language, the language pages cover the quirks that show up at volume — a Japanese to English PDF converter has to cope with vertical text and dense figure callouts, a Spanish to English PDF converter with accents and official form grids, and an Arabic to English scanner with the right-to-left flip on every page.
Translate your scanned document now
Upload a scanned PDF or a photo — Reglyph OCRs it, translates it, and rebuilds the page so tables, stamps, and figures stay exactly where they were.
Translate 5 pages free, every dayFrequently asked
How many pages can I translate at once with Reglyph?
Up to 50 pages per upload. For longer documents, split the PDF into parts (chapters or sections), translate each part, and merge — each part keeps the original layout, so the reassembled document stays consistent.
Can I translate a large scanned PDF for free?
You can test free: Reglyph translates 5 pages every day at no cost, with no credit card — enough to verify quality and layout on your document's hardest pages. Full large documents run on paid plans starting at $5.
Why does Google Translate fail on my large PDF?
Its document upload caps at 10 MB and expects a text layer. Scanned pages are full-page images, so even a modest scan exceeds the cap — and image-only pages need OCR that the upload flow doesn't provide.
Will tables and figures stay in place across hundreds of pages?
With a document-native tool, yes — each page is rebuilt independently, so page 200 is laid out as faithfully as page 1. Text-first tools tend to drift: the deeper into the document, the worse the reconstruction.
Reglyph