Skip to content
Merge & Organize

Split PDF by Text

Split a PDF by text: a new file starts at every page containing the word or phrase you choose — every “Invoice No.”, every “Employee:” — so a batch of documents scanned or exported together comes apart on its own.

Settings appear once you add a PDF.

Name each file after the line that matched
Advanced 1
Match capital letters exactly

Done on your device No upload, no queue, no sign-up.

Bytes sent anywhere: 0 B

How to split a PDF by text

Your PDF goes into this browser tab, is split wherever your text appears on your device, and comes back as one PDF per section. 0 bytes uploaded
  1. Add your PDF

    Drop the PDF — for example a year of invoices exported as one file.

  2. Check the settings

    Type the word or phrase that appears on the first page of each document. Matching pages are listed as you type.

  3. Save the result

    Save. Each file runs from one match to the page before the next, named after the matching line.

Why it doesn’t upload

Batches of invoices, payslips and statements are exactly the kind of document you don’t want on a stranger’s server. The text is read and the file split in your browser.

Check it in your browser’s Network tab

Worth knowing

  • Only text the PDF contains can be matched. Scanned pages need OCR first.
  • Pages before the first match become their own file.
  • Spaces and line breaks in the phrase are matched loosely, so “Invoice No.” also matches across a line break.

Under the hood

What happens to your file

pdf.js extracts the text of each page with positions, and a new file starts at every page where your text appears — an invoice number, "Statement", a heading. Scanned pages have no text layer to search, so they never trigger a split.

Runs on pdf.jspdf-lib

Reference: pdf.js, Mozilla

Questions

How do I split a PDF of invoices into separate files?

Type something that appears once at the top of every invoice — like Invoice No. — and press Split by text. Each invoice becomes its own PDF, named after its matching line.

Can I split a PDF where a keyword appears?

Yes, that’s exactly this tool: every page containing the keyword starts a new file.

How are the split files named?

They are numbered in order and, with Name each file after the line that matched on, named after that line — for example 01-Invoice-No.-1042.pdf — so you can tell them apart without opening them. Turn it off to use the original file name with a number.

Does it work on scanned PDFs?

Only once they have a text layer. A scan is a picture, so there’s no text to match until OCR reads it.

Read more about this

More merge & organize tools

All merge & organize tools