Split PDF by Text: Start a New File at Every Match

Got one PDF with many invoices, payslips or certificates in it? Type the text that appears at the start of each one, such as "Invoice No", and the tool starts a new file at every page where it finds it. You see the files it will make before you split. Everything runs in your browser.

Loading PDF tool...

Will this split my PDF?

A PDF made on a computer (exported invoices, payslips, letters)

Its text can be read, so every page with your text starts a new file.

A scanned document or phone photos

There is no readable text to search. Make it searchable with OCR first, then split.

OCR PDF

A book or manual with chapter bookmarks

Splitting at the bookmarks is usually quicker and names files after the chapters.

Split by bookmarks

A password-protected PDF

Remove the password first, then split.

Unlock the PDF

Features

  • Start a new file at every page that contains your text

  • Plain text or a regular expression, with or without matching case

  • Live preview of every file and the line that matched

  • Name files from the matched line, for example invoice-no-1043.pdf

  • Pages before the first match are kept as their own file

  • Flags scanned pages that have no readable text

When people use it

Merged invoices or bills

Split a batch export into one PDF per invoice, named by invoice number.

Payslips and Form 16

Separate a combined payroll PDF at every "Employee Name" or employee code.

Certificates and letters

Turn one long print file of certificates or offer letters into a file per person.

How splitting by text works

Each match starts a new file

The tool reads the text on every page. Each page that contains your text starts a new file, which runs until the page before the next match. If the first pages do not match (a cover sheet, say), they are kept together as their own file.

Matching is done one line at a time. A phrase broken across two lines on the page will not be found, so search for the part that sits on one line.

Naming files from the match

With "Name files from the matched line" on, each file is named from the line where your text was found, starting at the match: "Invoice No: 1043" gives invoice-no-1043.pdf. With a regular expression, the first group in brackets is used instead, so "Invoice No:\s*(\d+)" gives 1043.pdf. Repeated names get -2, -3 added.

What it cannot read

Split PDF by Text: Start a New File at Every Match - FAQ

Automate Your Business with AI

Love our free tools? Discover our AI-powered products that help businesses automate customer communication.