Skip to content
Free research tool · No sign-up

Plain Text Cleaner: strip formatting in one click

Paste text copied from Word, Google Docs, a PDF or a web page and instantly remove HTML tags, smart quotes, extra spaces, broken line breaks and hidden characters — leaving clean, publication-ready plain text you can drop straight into a journal portal or manuscript.

Manuscript Formatting & Plain Text Cleaner

Strip messy styling, HTML tags, smart quotes, broken line breaks and hidden characters before submitting to journals or publishers.

100% private — your text is cleaned in your browser and never uploaded.
Removal rules & transformations 5 activeHide options
Text case conversion
Raw input text
Words: 0Chars: 0
Sentences: 0Lines: 0
Sanitized result
Switch to “Diff” after cleaning to highlight what was removed or changed…
Words: 0Chars: 00 chars stripped
Sentences: 0Lines: 0
Done.

Tip: turn on only the rules you need. “Strip citation markers” and “Strip bullets” are off by default so you don’t accidentally remove content you want to keep. Everything runs in your browser — your text is never uploaded.

The basics

What is a plain text cleaner?

A plain text cleaner is a free tool that strips hidden formatting out of text and returns clean, unstyled plain text. When you copy from Word, Google Docs, a PDF or a web page, you also copy invisible baggage — HTML tags, “curly” smart quotes, non-breaking spaces, zero-width characters, and broken line breaks — that can corrupt a journal submission portal, a CMS, or a code editor. This tool removes all of it in one click.

ManuscriptLab’s cleaner is built for academic writing: alongside the usual remove-formatting options, it can unwrap the hard line breaks a PDF adds, standardise punctuation, and optionally strip bullet markers and citation brackets. Everything happens locally in your browser, so your unpublished manuscript is never uploaded or stored.

Quick reference

What the text cleaner removes and fixes

Building the figure by hand is slow and error-prone. This tool makes it fast, accurate, and journal-ready.

RuleWhat it fixesExample
Strip HTML / XML tagsRemoves tags pasted from web pages or rich editors.<p>text</p> → text
Normalize spacesCollapses double spaces and trims line padding.word   word → word word
Unwrap hard line breaksJoins lines broken mid-paragraph by PDF copy-paste.broken
line → broken line
Clean smart quotes & punctuationConverts curly quotes, em-dashes and ellipses to plain forms.“ ” ‘ ’ — … → " ' - ...
Strip bullets & list markersRemoves leading •, -, *, 1., a) prefixes.• Item → Item
Strip citation markersRemoves inline [1], [2-4] and (Author, 2020).result [1] → result
Collapse empty linesLimits runs of blank lines to a single break.3+ blank lines → 1
Text case conversionSwitches to Sentence case, Title Case, lower or UPPER.choose from the dropdown

Also removes invisible control characters: non-breaking spaces (U+00A0), zero-width spaces, and byte-order marks that break submission portals.

Why it matters

Why plain text matters for manuscript submission

Most journal submission systems, plagiarism checkers, reference managers and typesetting pipelines expect clean text. Hidden formatting is a common, avoidable cause of problems at submission.

Smart quotes and em-dashes can render as garbled symbols in a portal’s text box. Non-breaking spaces and zero-width characters can throw off word counts and search-and-replace. Hard line breaks copied from a two-column PDF turn a paragraph into a ladder of short lines. Stray HTML from a web source can be rejected outright. Cleaning your text to plain UTF-8 first removes these failure points, keeps your word count accurate, and makes the file behave predictably wherever you paste it — the submission box, your reference manager, or a co-author’s editor.

Common problem

How to clean text copied from Word or a PDF

Copying from Microsoft Word or a PDF is the number-one source of hidden formatting. Here’s the fastest way to fix it.

Paste the copied text into the cleaner’s left pane. Leave Strip HTML, Normalize spaces, Unwrap hard line breaks and Clean smart quotes switched on — these handle the four problems Word and PDFs cause most often: leftover styling, double spaces, mid-paragraph line breaks, and curly punctuation. Your clean text appears on the right instantly. Use the Diff view to confirm exactly what changed, then click Copy or download it as a .txt file. Because the tool outputs pure plain text, pasting it back into Word (or a submission portal) gives you a clean slate with none of the original’s hidden characters.

Why use it

Why researchers use this text cleaner

One-click cleanup

Live auto-clean rewrites your text as you paste — no buttons to hunt for, results in an instant.

Removes hidden characters

Clears non-breaking spaces, zero-width characters and byte-order marks that silently break portals.

Before/after diff

See exactly what was stripped or changed, word by word, so you never lose content by accident.

Toggle every rule

Turn each transformation on or off, and convert case, so the output is exactly what you want.

Copy or download

Copy clean text to the clipboard or export a ready-to-use .txt file in one click.

Private & free

All processing runs in your browser. No sign-up, no cost, and your text is never uploaded or stored.

What’s inside

Key features at a glance

FeatureWhat it does for you
Live auto-cleanCleans instantly as you paste or type — toggle off for manual control.
8 removal & transform rulesHTML, spaces, line breaks, smart quotes, bullets, citations, blank lines, case.
Dual-pane workspaceRaw input on the left, sanitised output on the right, side by side.
Diff viewWord-level highlighting of everything removed or changed.
Live statisticsWords, characters, sentences, lines, and characters saved.
Copy & downloadClipboard copy or .txt export of the cleaned text.
Case conversionSentence case, Title Case, lowercase or UPPERCASE.
Client-side privacyNothing is uploaded — safe for unpublished manuscripts.
It fixes these

Common formatting problems it solves

Garbled quotes in the submission box.

Curly “smart” quotes and em-dashes turn into odd symbols in many portals — this converts them to plain characters.

Paragraphs broken into short lines.

Two-column PDF copy adds a line break after every line. Unwrapping restores flowing paragraphs.

Wrong word count.

Non-breaking and zero-width spaces distort counts and search-and-replace. Removing them fixes both.

Leftover HTML from the web.

Tags copied from a web page are stripped so only readable text remains.

Double and trailing spaces.

Multiple spaces are collapsed to one and trailing spaces removed for consistent formatting.

Who it’s for

Use cases

Journal submission prep

Clean punctuation, spaces and hidden characters before pasting into a manuscript portal.

Word & PDF cleanup

Strip the formatting baggage that comes with copying from Microsoft Word or a PDF.

Plagiarism & AI checks

Feed checkers clean plain text so hidden characters don’t skew the analysis.

Reference managers & CMS

Paste tidy text into EndNote, a website, or a CMS without dragging in styling.

Coding & data entry

Convert smart quotes to straight quotes so text works in code and spreadsheets.

Reformatting drafts

Reset messy, inherited formatting to a clean slate before you restyle a document.

FAQs

Frequently asked questions

How do I remove formatting from text?

Paste your text into the cleaner above. It automatically strips HTML tags, converts smart quotes to straight quotes, collapses extra spaces, unwraps broken line breaks and removes hidden characters — leaving clean plain text. Copy the result or download it as a .txt file.

How do I convert text to plain text?

Plain text is unstyled text with no HTML, fonts, or special characters. Paste your content into the tool and it outputs plain UTF-8 text you can copy or download. This is the format most journal portals, plagiarism checkers and code editors expect.

Does it remove smart quotes and em-dashes?

Yes. With “Clean smart quotes & punctuation” enabled, curly double and single quotes become straight quotes, em- and en-dashes become hyphens, and the ellipsis character becomes three periods — the plain forms that display correctly everywhere.

Can it fix line breaks copied from a PDF?

Yes. “Unwrap hard line breaks” joins lines that were broken mid-paragraph — the classic problem when copying from a two-column PDF — while preserving the blank lines between real paragraphs.

Is my text private?

Completely. All cleaning happens locally in your browser using JavaScript. Your text is never uploaded, saved, or sent to any server, so it’s safe to use on unpublished or confidential manuscripts.

Will it delete my citations or references?

Only if you choose to. “Strip citation markers” is off by default. Turn it on to remove inline markers like [1] or (Author, 2020); leave it off to keep every citation exactly as written.

Is the plain text cleaner free?

Yes — it’s completely free with no sign-up and no limits. You can clean as much text as you like, as often as you like.

Can you edit my manuscript for me?

The tool handles formatting cleanup. For a human editor to copyedit and proofread your writing to journal standard, see ManuscriptLab’s copyediting and proofreading services.

This Plain Text Cleaner processes your text locally in your browser to remove formatting and hidden characters. Always keep a copy of your original file, and review the cleaned output before submitting. Need a human to polish the writing itself? Explore our copyediting and proofreading services.

Want your manuscript formatted and polished by an editor?

This tool clears the hidden formatting — but if you’d like a professional to copyedit your writing, standardise your style to a journal’s guidelines, and proofread the final draft, our editors do exactly that. Real human editors, no AI, matched to your subject area.

Get a Custom Quote

Tell us a bit about your project and we’ll get back to you within 24 hours.