A 300-page scanned bundle arrives and you need pages 41 to 58 of it. Or a report has its appendix in the wrong place. Or a scanner has produced one enormous file containing eleven separate documents. These all sound like different problems and they are solved by the same small set of operations.
The trick is picking the right one first, because doing it in the wrong order means counting page numbers twice.
The four operations
Split
Split PDF divides one document into several. Use it when the result should be multiple files: a scanner output containing many documents, a year of statements in one file, a bundle that needs distributing in sections. You either give it the points to cut at, or ask for a file per page.
Extract
Extract Pages pulls a selection out into a new document and leaves the original alone. This is the one to use when you want a few pages from something — a single form from a pack, one chapter, the pages your colleague actually asked for. It is also the quickest way to make a large file small.
Remove
Remove Pages is extract's mirror image: keep the document, delete what you do not want. Reach for it when the pages to discard are fewer than the pages to keep — blank separator sheets from a scanner, a cover page, the duplicated appendix.
Extract and remove can express the same result. Choose whichever means typing fewer numbers, because every page number you type is an opportunity to get one wrong.
Organise
Organize Pages shows every page as a thumbnail so you can drag them into a new order, rotate individuals, and delete as you go. When the job is visual — "this section belongs before that one", "these two are upside down" — working with thumbnails is far more reliable than reasoning about numbers.
Page ranges, and the traps in them
Most of these tools accept a range expression such as 1-3, 7, 12-15. Worth knowing
how they behave:
- Ranges are inclusive.
4-8is five pages, not four. This is the single most common off-by-one. - Page numbers are physical, not printed. Page 1 is the first sheet in the file. If the document has roman-numbered front matter, the printed "page 1" might be the twelfth sheet. Always check against the page counter in your viewer rather than the number printed on the paper.
- Order is usually preserved as given where a tool supports reordering, so
5, 2, 9can mean exactly that. Where it does not, ranges are normalised — so if reordering is the point, use Organise instead.
Getting the order of operations right
The reason to think about sequence is simple: every operation that adds or removes pages renumbers everything after it. Do two operations in the wrong order and the second one is working from numbers that no longer mean what you thought.
A sequence that avoids this:
- Rotate first. Fix sideways pages before anything else — it never changes page numbers, and it makes the rest legible enough to work with. (Rotate PDF.)
- Extract the region you care about, if you only need part of a large document. Everything after this step happens on a smaller file with simpler numbering.
- Remove the rubbish — blank separators, duplicate covers, scanner test pages.
- Reorder, visually, once the set of pages is final.
- Split or merge to produce the files you actually need.
- Compress last, if size is a constraint, so you are not compressing pages you are about to delete.
Rotate early, compress late, and renumber as rarely as possible.
Scanner output, specifically
Bulk-scanned bundles have their own pattern, and it is worth having a routine:
- Blank pages between documents are usually the scanner's separator sheets. Remove them, and note that their positions tell you where the document boundaries are.
- Every other page upside down means a duplex scanner fed the reverse sides rotated. Organise lets you fix them individually.
- Scans are images, so nothing is searchable until you run OCR — and run it after splitting, so you are only processing the pages you are keeping.
Keep the original
All of these operations write a new file rather than modifying the source, which is a good default and easy to undermine by saving over the original in your downloads folder. Keep the untouched scan until the reorganised version is confirmed good — page numbers are easy to get wrong, and the cheapest fix is starting again from the source.
Why none of this needs an upload
Rearranging pages is one of the lightest operations in the format: the page objects are copied into a new document in a different order, without re-rendering or re-compressing anything. Quality is untouched by definition. It is also work a browser can do comfortably, so a 300-page bundle never has to leave your device — which matters, because bundles like that tend to be the scanned records you would least like to hand over.
In short
Extract when you want a few pages out, remove when you want a few pages gone, split when you need several files, and organise when the job is visual. Rotate first, compress last, and check page numbers against your viewer's counter rather than the numbers printed on the page.