Long PDFs — textbooks, manuals, lengthy reports — often contain far more than any one reader needs at a given moment. Extracting just the relevant pages, rather than sharing or storing the entire document, makes the result easier to use and easier to send, but only if you can actually identify which pages you need in the first place.

Using a table of contents or bookmark panel to find page numbers

Most long, professionally produced PDFs include either a table of contents page, a bookmark panel in the PDF viewer's sidebar, or both, listing sections alongside their starting page numbers. Checking this first, rather than scrolling manually through dozens or hundreds of pages, is by far the fastest way to identify the specific range you need before opening any extraction tool.

Using search when there's no usable table of contents

For a document without clear navigation aids, most PDF viewers support a text search (commonly Ctrl+F or Cmd+F) that jumps directly to pages containing a specific word or phrase, which is often faster than browsing even when a table of contents does exist, especially if you know a specific term likely to appear only in the section you need. This only works on PDFs with actual underlying text, not on scanned image-only pages, unless OCR has already been run on the file. Building a habit of running OCR on any scanned document as soon as it's captured, rather than waiting until a specific need for searchability comes up, means this option is always available later without having to circle back and do it retroactively.

Accounting for the gap between printed and actual page numbers

A document's printed page numbers (the ones shown at the bottom of each page) often don't match its actual position within the file, since cover pages, tables of contents, or blank separator pages frequently come before page "1" starts. Checking your PDF viewer's own page counter, not just the printed number on the page, avoids extracting the wrong pages due to this common offset.

Extracting multiple separate sections at once

If the pages you need aren't in one continuous block — an introduction section plus a specific appendix much later in the document, for instance — most extraction tools accept a combined range like "1-5,88-92" that pulls both sections into a single output file in the order specified, saving you from running two separate extractions and then merging the results back together.

Double-checking the result before relying on it

Once extracted, opening the new smaller file and confirming it actually starts and ends where intended catches the most common mistake in this whole process: an off-by-one error from the printed-versus-actual page number gap discussed above. This check takes moments but avoids sending or submitting a file that's missing its first or last page.

Try the Split & Extract tool yourself — free, instant, nothing uploaded.

Open the tool

Frequently asked questions

Does text search work on scanned PDFs?

Only if the scanned document has already had OCR (optical character recognition) run on it to add a searchable text layer; a purely image-based scan with no text layer can't be searched, since there's no underlying text for the search function to match against.

Why did I get the wrong pages even though I used the printed page number?

This usually happens because a document's actual file position doesn't match its printed page numbers, due to cover pages, tables of contents, or other unnumbered pages before the numbering starts — check your PDF viewer's own page counter instead of the printed number for accuracy.

Can I extract pages from several different sections into one combined file?

Yes, most extraction tools accept a range covering multiple separate sections at once, like "1-5,88-92,150," pulling them together into a single new output file in the order you list them.

What if the document has no table of contents and no searchable text at all?

In that case, manually scrolling through to identify the pages you need is the only reliable option, though running OCR on the document first, if it's a scan, adds a searchable text layer that makes future extractions from that same original file considerably faster and noticeably more convenient to work with going forward, especially for a document you expect to need again.