SReader Blog
How to Convert PDFs to EPUB Without Wrecking the Layout
A file-by-file process for turning PDFs into reflowable EPUBs: triage, OCR for scans, structure rebuilds, quality checks, and when to keep the PDF.
How to Convert PDF to EPUB Without Wrecking the Layout
A small-font PDF on an eInk reader is a specific kind of misery. The page is fixed at print dimensions and does not reflow, so footnotes and dense comparison tables stay at whatever size they were printed, sometimes to the point of being simply illegible. The fix seems to be a bigger screen or a different format. That is the moment most people decide to convert PDF to EPUB.
The promise is real: text that rewraps to your screen and fonts you can enlarge. The failure mode is just as real. A careless conversion merges paragraphs, drops images, scrambles tables, and hands you an EPUB that is harder to read than the PDF you started with.
The difference between those outcomes is mostly the decisions made before any tool opens the file. This guide works through them in order: why PDFs resist reflow, a 60-second triage test, what scanned PDF OCR conversion requires, how to preserve structure during conversion, and how to verify the result before trusting it. Some files should never be converted, and recognizing them is half the skill.
SReader imports readable text from PDFs and EPUBs on your device, with no account required. It is a focused text reader, not a PDF-to-EPUB converter or a viewer that preserves page layout. Keep the original file for diagrams, tables, and formatting.
Fixed layout vs reflowable ebooks: why PDFs refuse to reflow
The root cause of broken conversions is structural, not a matter of picking the right tool. PDF, defined by the ISO 32000 standard, stores every character at a fixed coordinate on a fixed-size page. EPUB 3.3, a W3C Recommendation, is reflowable XHTML and CSS packaged inside a ZIP archive, so the same text rewraps to any screen and any font size. As one conversion guide puts it:
A PDF knows that a particular string of characters appears at coordinates (72, 640) on page 14 in 11-point Garamond. It does not know whether that string is a chapter heading, a body paragraph, a footnote, or a running header.
Conversion is therefore interpretation, not translation. A converter guesses which strings are headings based on font size and weight, merges lines that sat close together on the page, and decides where paragraphs begin and end, sometimes splitting them at page boundaries. The guesswork degrades as layout complexity increases, which is why complex PDFs produce the worst conversions.
Reflow is still worth pursuing when the source cooperates. A reflowable EPUB rewraps the same text to whatever screen and font size you prefer, which is exactly what a fixed PDF page cannot do. If you are weighing the two formats in general rather than for a single file, the trade-offs are covered in Choosing Between EPUB and PDF for Comfortable Reading.
The 60-second triage: is your PDF text-based or scanned?
Before choosing any tool, run one test. Open the PDF and try to select, copy, or search the text. If highlighting works and search finds words, the file is text-based and converts directly, with no OCR needed. If the page behaves like a photograph, with no selectable text and no search hits, you have a scanned image PDF that requires OCR before any conversion, as this scan-versus-text breakdown explains.
Searchability and copy-paste are the practical differentiators between the two types. This single gate determines the entire path a file takes, so run it on every file before doing anything else. It takes under a minute and prevents wasted effort on the wrong workflow.
Scanned PDF OCR conversion: two stages, not one
A scanned PDF is a collection of page images, not a text document. Convert it straight to EPUB and the output comes out garbled or blank, so the workflow needs two stages in strict order: OCR, optical character recognition, first turns the page images into editable text; the format conversion runs only afterward (a scanned-PDF conversion guide covers the sequencing).
Keep expectations calibrated. One guide to converting physical books frames it well: the scan is a rescue vehicle, not the product, and nothing from it survives into the finished file except the corrected words and any images you keep. The same guide's limits hold across OCR tools: simple tables come out acceptably, complex tables need manual adjustment, and complex math formulas are usually preserved as images rather than recognized.
Plan to proofread and correct OCR errors before the text is library-ready. Skip that pass and the errors flow silently into your EPUB.
How to convert PDF to EPUB without losing chapters, footnotes, and images
For text-based PDFs, Calibre is the tool of record: free, open-source, and available on Windows, macOS, and Linux. Output quality depends heavily on the source, though. Clean, logically structured text-based PDFs convert well, while scanned or visually complex PDFs need preparation first and may still yield broken lines, missing images, strange spacing, or poor chapter structure (a Calibre walkthrough documents the failure modes).
The conversions that feel native rather than transplanted are the ones where structure gets rebuilt deliberately:
- Tag chapter headings as real headings, not body text in a large font.
- Replace the page-numbered print table of contents with a linked one.
- Turn endnotes and footnotes into tappable links.
Then accept what cannot survive. KDP's formatting guidance says page numbers, headers, and footers do not apply to reflowable ebooks, and IngramSpark's specifications bar page-number references anywhere in the file, including the table of contents. Kobo's technical guide for authors gives the reason:
Reflowable ePubs can therefore not have any header or footer elements other than those automatically inserted by the eReader. There is no such thing as a “page” in a reflowable book. You are simply seeing one chunk of the never-ending flow of text.
Larger fonts also mean more pages, since page count is fluid in a reflowable book. Let go of print pagination entirely, or the converted file will keep fighting you.
The post-conversion quality checklist: verify before you trust
Before the original PDF goes anywhere, walk the converted EPUB through the failure modes above:
- Chapter boundaries: each chapter starts where it should and is tagged as a heading, not styled body text.
- Paragraph integrity: no lines merged across what were separate paragraphs, no paragraphs split at old page boundaries.
- Tables and equations: simple tables are acceptable, complex tables get adjusted by hand, math appears as images.
- Images and captions: present, near the text they belong to, with captions not orphaned by the reflow.
- Searchability: any text you could find in the PDF should still be findable in the EPUB.
Keep the original PDF until the file passes every item. The checklist is what separates a repeatable process from a one-off gamble.
When layout is the point: fixed-layout EPUBs for image-heavy books
Reflow has a blind spot. When a PDF contains photos, illustrations, tables, charts, screenshots, or equations, a reflowable conversion risks misaligning captions and relocating images, because the layout changes with each reader's screen and font settings (Zamzar's guide to preserving page formatting describes the risk).
A fixed-layout EPUB is the middle path: the page still scales to fit the screen, but the placement of text and images stays identical to the original. Good candidates include autobiographies with photos, recipe books, travel guides, children's picture books, comics and graphic novels, art books, textbooks with charts and equations, and manuals with screenshots, per the fixed-versus-reflowable comparison. Chart-heavy documents such as statistical reports and financial statements should keep charts as images in a fixed layout or simply remain PDFs, a route the Calibre source-quality guide also endorses.
The trade-off is honest: fixed layout preserves the design but gives up adjustable fonts, line spacing, and margins.
When to convert PDF to EPUB and when to keep the original
Convert when the document is text-dominant: the fonts are too small for your reader, or adjustable type would ease eye strain.
Keep the original when layout carries the meaning: charts, complex tables, equations, and image-heavy pages where a relocated caption damages comprehension more than a small font ever did.
SReader can import readable text from either format, so creating an EPUB is not required just to read that text. Use a dedicated PDF viewer when you need the original page layout. The principle underneath every step above is simple: convert only when it earns it.
PDF reflow reading without conversion: focal points and pace control
Not every reading problem is a layout problem. Sometimes the text is legible and the page is still loud, with dense columns and sidebars competing for attention. SReader extracts readable text into a separate reading view. Word, Line, and Book modes let you choose how much text to display and adjust your pace. This does not preserve the original PDF layout or its diagrams; keep the source open separately when those details matter. The same technique works in converted EPUBs, so it complements conversion rather than competing with it. How to Read Dense Academic PDFs Without Visual Overwhelm covers that workflow in more depth.
SReader applies this locally, with no ads, no accounts, and no data harvesting. Your PDFs and converted EPUBs stay in your library rather than in someone else's analytics.
Your repeatable process for a comfortable library
The whole guide compresses into five steps:
- Run the 60-second triage: try to select or search the text.
- If the PDF is scanned, OCR first, then proofread; never skip the correction pass.
- Convert only clean, text-based sources, and rebuild structure: real headings, a linked table of contents, tappable notes.
- Verify against the checklist before deleting anything.
- Keep the original whenever layout carries the meaning.
A comfortable library is built one file at a time, and organizing across both formats is its own project; Build a Calm Digital Library Across EPUB and PDF Files covers it. When you want a reader that keeps both formats on your device, Try SReader.