PDF OCR
Pull selectable text out of scanned PDFs, 100+ languages.
Finds which lines changed between two versions, by comparing the text rather than the pixels.
Original
The version you are comparing from.
Revised
The newer version.
The version you are comparing from.
Both are read for their text layer.
Additions and removals listed, with a count of each.
By text. A pixel comparison of a re-flowed document is mostly noise - a single inserted word shifts everything after it. Comparing the extracted text finds the edits people actually care about.
No. Scans have no text layer, so there is nothing to compare. Run OCR on both first, then compare the results.
A line that appears in one document but not the other. A line with one word altered shows as one removal and one addition, since this is a line-level comparison.
Very long documents are rejected rather than left to hang, because the comparison is quadratic in the number of lines. Split them into sections if you hit that.
Compare two PDFs and see exactly which lines were added or removed. The comparison works on the extracted text using a longest-common-subsequence diff - the same approach git takes - rather than comparing pixels, because a single inserted word reflows everything after it and a visual diff would then report the entire remainder of the document as changed. Both files need a real text layer, so scanned documents need OCR run on them first.
The last couple are from other categories - there are 81 tools here in total.
Pull selectable text out of scanned PDFs, 100+ languages.
Pull the text layer out of a PDF instantly.
Shrink PDFs to fit email, chat, or upload limits.
Hit an exact file size for government forms and portals.
Shrink JPG, PNG or WebP without visible quality loss.
Never uploadedStrong random passwords, generated in your browser.
Never uploaded