Split PDF by Text
Split a PDF by text: a new file starts at every page containing the word or phrase you choose — every “Invoice No.”, every “Employee:” — so a batch of documents scanned or exported together comes apart on its own.
Fit Meter
Settings appear once you add a PDF.
Advanced 1
Still going. Your own device is doing this, so a long scan or a big batch takes as long as it takes — and none of it is being uploaded.
Bytes of your files sent anywhere: 0 B · nothing leaves this tab
Every request this page makes is listed as it happens. You’ll see our own code loading, and for some formats a codec engine, but never your file going out.
- Waiting for network activity…
Check it yourself: open your browser’s developer tools on the Network tab, or turn off Wi-Fi once the page has loaded. The tools keep working.
How to split a PDF by text
-
Add your PDF
Drop the PDF — for example a year of invoices exported as one file.
-
Check the settings
Type the word or phrase that appears on the first page of each document. Matching pages are listed as you type.
-
Save the result
Save. Each file runs from one match to the page before the next, named after the matching line.
Why it doesn’t upload
Batches of invoices, payslips and statements are exactly the kind of document you don’t want on a stranger’s server. The text is read and the file split in your browser.
Check it in your browser’s Network tabWorth knowing
- Only text the PDF contains can be matched. Scanned pages need OCR first.
- Pages before the first match become their own file.
- Spaces and line breaks in the phrase are matched loosely, so “Invoice No.” also matches across a line break.
Under the hood
What happens to your file
pdf.js extracts the text of each page with positions, and a new file starts at every page where your text appears — an invoice number, "Statement", a heading. Scanned pages have no text layer to search, so they never trigger a split.
Reference: pdf.js, Mozilla
Questions
How do I split a PDF of invoices into separate files?
Type something that appears once at the top of every invoice — like Invoice No. — and press Split by text. Each invoice becomes its own PDF, named after its matching line.
Can I split a PDF where a keyword appears?
Yes, that’s exactly this tool: every page containing the keyword starts a new file.
How are the split files named?
They are numbered in order and, with Name each file after the line that matched on, named after that line — for example 01-Invoice-No.-1042.pdf — so you can tell them apart without opening them. Turn it off to use the original file name with a number.
Does it work on scanned PDFs?
Only once they have a text layer. A scan is a picture, so there’s no text to match until OCR reads it.