Paste the source
Copy selectable text from a PDF reader and paste it into the Source pane. The original remains untouched while you compare results.
CopyPrune is a free PDF text cleaner for text you have already copied. It repairs likely visual line wraps, line-ending word splits, ligatures, and hidden spacing, preserves explicit structure, and shows every recorded change in your browser.
Research notes often keep the visual
line wraps used on a printed page.
Research notes often keep the visual line wraps used on a printed page.
Paste copied PDF text to begin.
The report comes from the same deterministic operations that created the result. It does not infer changes after the fact.
Load the example or paste your own copied text, then choose Clean text.
We do not intentionally send your pasted text to our servers, analytics, URLs, or local storage. Clear removes it from this page’s runtime state.
Use a focused workflow when you already know which PDF copy problem you need to solve.
Every rule reruns from what you pasted.
Uncertain boundaries stay for review.
Copy and download match the result.
CopyPrune follows a deterministic set of readable rules. The cleaner protects obvious structure, makes only the repairs allowed by your mode and settings, and builds the proof report from the actual operations.
Copy selectable text from a PDF reader and paste it into the Source pane. The original remains untouched while you compare results.
Careful is the default. Standard joins more probable wraps. Flatten is available when you genuinely want one continuous block.
See recorded wraps, repaired words, normalized glyphs, protected structure, and boundaries the careful mode declined to guess about.
The examples below describe what this browser utility is designed to handle. CopyPrune does not upload PDFs, perform OCR, or reconstruct a multi-column page.
A printed page wraps text at a fixed width even when the paragraph continues.
A printed page wraps text at a fixed width even when the paragraph continues.
Likely visual wraps become spaces; explicit blank-line paragraphs remain.
The report demon- strated a repeatable result.
The report demonstrated a repeatable result.
Likely end-of-line splits can lose the hyphen. Inspect recorded changes for legitimate compounds.
NEXT STEPS 1. Review the text 2. Copy the result
NEXT STEPS 1. Review the text 2. Copy the result
Obvious structure is protected instead of being flattened into prose.
The final file contains hidden spacing.
The final file contains hidden spacing.
Common PDF ligatures and invisible formatting characters become ordinary text.
Careful mode protects uncertainty. The stronger modes exist for messier source text, but they should receive more human review.
Joins likely visual wraps, repairs likely split words, and protects explicit paragraph breaks plus enabled headings, lists, and structured rows.
Best for research notes and ordinary prose that you can review.Joins more single line boundaries while still protecting the clearest document structure. Its changes are labeled medium confidence.
Best when the copied paragraph is visibly fragmented.Repairs likely split words when that rule is enabled, then replaces the remaining line and paragraph breaks with spaces. Line-break structure is intentionally removed.
Best only when one continuous text block is required.PDFs describe visual placement, not always the reading structure a writer expects. CopyPrune works after a PDF reader has already provided selectable text.
Read the method and limitations →CopyPrune’s cleaning engine runs on your device. We do not intentionally send document text to our servers, put it in URLs, retain it in local storage, or include it in analytics events.
Google Analytics loads only after a visitor accepts analytics. It may receive ordinary device, network, and usage information, but never the source or cleaned text.
Read the privacy policy →CopyPrune can repair likely visual line wraps and safer line-ending word splits, expand common ligatures, remove soft hyphens, and normalize selected hidden spacing. It protects explicit paragraphs, headings, lists, and structured rows when those protections are enabled.
Use Remove Line Breaks Online when line wrapping is the only problem. The full PDF Text Cleaner can handle line wraps together with split words, ligatures, and hidden spacing.
Use Dehyphenate PDF Text to review individual line-ending hyphens and soft hyphens. The full cleaner can also repair likely split words and record each change for inspection.
A PDF can store text as visually positioned page content rather than the logical paragraph structure a writer expects. When a reader exposes that text for copying, printed line endings and page-layout artifacts can come with it.
No. CopyPrune intentionally accepts only text you have already copied. This avoids misleading promises about OCR, reading order, tables, and scanned pages.
No tool can recover every PDF’s intended structure from copied text alone. Careful mode protects blank-line paragraphs and obvious structure, then flags uncertain boundaries instead of silently joining them.
No. CopyPrune uses deterministic browser-based rules. The same source, mode, and rule settings produce the same output and proof ledger.
CopyPrune does not claim compliance or suitability for confidential legal, medical, financial, or regulated information. Follow your organization’s data-handling requirements.
The cleaner is an editorial heuristic, not a language model or laboratory-validated reconstruction system. Its rule definitions and limitations are published in the cleaning methodology so results can be evaluated honestly.
Primary references: Unicode line breaking and soft hyphens and browser clipboard writes.
Paste selectable text, choose the level of cleanup, and inspect the result before copying.