Paste the source
Copy selectable text from a PDF reader and paste it into the Source pane. The original remains untouched while you compare results.
Preserve explicit paragraph breaks, join likely visual wraps, repair likely split words, and inspect the change report—right in your browser.
Research notes often keep the visual
line wraps used on a printed page.
Research notes often keep the visual line wraps used on a printed page.
Paste copied PDF text to begin.
The report comes from the same deterministic operations that created the result. It does not infer changes after the fact.
Load the example or paste your own copied text, then choose Clean text.
We do not intentionally send your pasted text to our servers, analytics, URLs, or local storage. Clear removes it from this page’s runtime state.
Every rule reruns from what you pasted.
Uncertain boundaries stay for review.
Copy and download match the result.
CopyPrune follows a deterministic set of readable rules. The cleaner protects obvious structure, makes only the repairs allowed by your mode and settings, and builds the proof report from the actual operations.
Copy selectable text from a PDF reader and paste it into the Source pane. The original remains untouched while you compare results.
Careful is the default. Standard joins more probable wraps. Flatten is available when you genuinely want one continuous block.
See recorded wraps, repaired words, normalized glyphs, protected structure, and boundaries the careful mode declined to guess about.
The examples below describe what this browser utility is designed to handle. CopyPrune does not upload PDFs, perform OCR, or reconstruct a multi-column page.
A printed page wraps text at a fixed width even when the paragraph continues.
A printed page wraps text at a fixed width even when the paragraph continues.
Likely visual wraps become spaces; explicit blank-line paragraphs remain.
The report demon- strated a repeatable result.
The report demonstrated a repeatable result.
Likely end-of-line splits can lose the hyphen. Inspect recorded changes for legitimate compounds.
NEXT STEPS 1. Review the text 2. Copy the result
NEXT STEPS 1. Review the text 2. Copy the result
Obvious structure is protected instead of being flattened into prose.
The final file contains hidden spacing.
The final file contains hidden spacing.
Common PDF ligatures and invisible formatting characters become ordinary text.
Careful mode protects uncertainty. The stronger modes exist for messier source text, but they should receive more human review.
Joins likely visual wraps, repairs likely split words, and protects explicit paragraph breaks plus enabled headings, lists, and structured rows.
Best for research notes and ordinary prose that you can review.Joins more single line boundaries while still protecting the clearest document structure. Its changes are labeled medium confidence.
Best when the copied paragraph is visibly fragmented.Repairs likely split words when that rule is enabled, then replaces the remaining line and paragraph breaks with spaces. Layout structure is intentionally removed.
Best only when one continuous text block is required.PDFs describe visual placement, not always the reading structure a writer expects. CopyPrune works after a PDF reader has already provided selectable text.
See every rule and limitation →CopyPrune’s cleaning engine runs on your device. We do not intentionally send document text to our servers, put it in URLs, retain it in local storage, or include it in analytics events.
Google Analytics loads only after a visitor accepts analytics. It may receive ordinary device, network, and usage information, but never the source or cleaned text.
Read the privacy policy →Paste selectable PDF text into the Source pane, keep Careful mode selected, and choose Clean text. CopyPrune joins likely visual line wraps while keeping explicit blank-line paragraph breaks and enabled document structure.
Keep Repair split words enabled. CopyPrune can join a line-ending hyphen with a likely lowercase continuation and records the change for inspection. Because legitimate hyphenated compounds exist, review the cleaned text before using it.
No. CopyPrune intentionally accepts only text you have already copied. This avoids misleading promises about OCR, reading order, tables, and scanned pages.
No tool can recover every PDF’s intended structure from copied text alone. Careful mode protects blank-line paragraphs and obvious structure, then flags uncertain boundaries instead of silently joining them.
No. CopyPrune uses deterministic browser-based rules. The same source, mode, and rule settings produce the same output and proof ledger.
CopyPrune does not claim compliance or suitability for confidential legal, medical, financial, or regulated information. Follow your organization’s data-handling requirements.
After confirmation, the source, cleaned result, and proof report are removed from this page’s runtime state. Document text is not saved in browser storage.
The cleaner is an editorial heuristic, not a language model or laboratory-validated reconstruction system. Its rule definitions and limitations are published so results can be evaluated honestly.
Technical references: Unicode normalization and browser clipboard behavior.
Paste selectable text, choose the level of cleanup, and inspect the result before copying.