How to Copy Text from a PDF Without Broken Lines

Copying from a PDF often gives broken lines and jumbled columns. Extract clean text in reading order instead — as paragraphs or original lines.

3 min read ·

Copying text out of a PDF viewer often goes wrong: every line ends with a hard break, hyphenated words stay split, and two-column pages come out interleaved. To get clean text, open PDF to Text, add the PDF, choose Paragraphs, and copy the result or download it as a .txt file. Lines are joined back into paragraphs and words broken across lines are rejoined.

Extract text step by step

  1. Open PDF to Text and add your PDF.
  2. Choose a layout:
    • Paragraphs — joins lines into flowing paragraphs. Best for pasting into an email, a document or a translation tool.
    • Original lines — keeps every line exactly as on the page. Best for poems, code, addresses and lists.
  3. To take only some pages, type them in the pages box — for example 1-3, 7 — or leave it empty for the whole document. Turn on page markers if you want to see where each page starts.
  4. Click Copy text, or download the .txt file.

Why copying from a PDF breaks

A PDF doesn't store paragraphs. It stores fragments of text and the exact position of each one on the page. When you select and copy in a viewer, it has to guess how those fragments fit together, and the guess is often poor:

  • Line breaks: each visual line becomes a separate line, so pasted text is chopped up.
  • Hyphens: "infor-" and "mation" stay apart.
  • Columns: the viewer may read straight across both columns, mixing them.
  • Headers and footers: page numbers and running titles appear in the middle of sentences.

PDF to Text reads the fragments in reading order and rebuilds paragraphs, which fixes most of these problems in one go.

If you get no text at all

When the tool finds almost no text, the PDF is probably a scan — a picture of a page. There's no text in it to extract. Use PDF OCR, which reads the letters in the images. It works in 29 languages and runs in your browser too.

If you need formatting

Plain text drops bold, headings and fonts. When you want to keep the structure and edit the document, convert it with PDF to Word instead.

If you just need the gist

For a 60-page report you don't have time to read, extracting every word isn't the point. The AI PDF Summarizer gives you a short summary, key points and deadlines — and you can ask it questions about the document.

Tips for specific documents

  • Academic papers: use Paragraphs, then remove the reference list if you only need the body text.
  • Contracts: Original lines keeps clause numbering on its own lines, which makes quoting easier.
  • Tables: text extraction flattens tables into lines; for rows and columns, use the PDF to Excel option of the Universal Document Converter.
  • Text that comes out as gibberish: some PDFs use fonts that hide which letters they draw. The text looks fine on screen but extracts as symbols. Run PDF OCR on those pages — it reads the shapes on the page rather than the hidden codes.

Frequently asked questions

Is my PDF uploaded?

No. Text extraction happens in your browser, so confidential documents stay on your device.

Does it work on a phone?

Yes. On iPhone or Android, extract the text and tap Copy text, then paste it into Notes, Mail or any app.

Can I extract text from a password-protected PDF?

Remove the password first with Unlock PDF if you know it. Some PDFs aren't password-protected but block copying; the tool can still read their text.

Tools mentioned in this guide

← All guides