Text Cleaning and Line Counts

Paste text and choose which line-by-line changes to apply. The source stays visible for review, and output line breaks use LF.

Switching languages opens a new page and does not transfer the current source text or result. Copy or download anything you need first.

Text cleaning could not start. Check that JavaScript is enabled, or use a browser that supports local processing.

Source text

01 / INPUT

Each input must be no more than 200,000 UTF-16 code units and no more than 256 KiB in UTF-8. These limits apply before newline normalization. Split text that exceeds either limit.

Source metrics

Code points
0
Lines
0
UTF-8 bytes
0

Source metrics use LF-normalized text (CRLF and CR become LF).

Cleaning options

Work status

Paste text and choose how to clean it.

Cleaned result

02 / REVIEW

Output line breaks use LF.

Result metrics

Code points
0
Lines
0
UTF-8 bytes
0

How to use this tool

  1. Paste the source

    Paste text and review its metrics. Each input must be no more than 200,000 UTF-16 code units and no more than 256 KiB in UTF-8. Split the text if it exceeds either limit.

  2. Choose cleaning options

    Trim spaces or tabs at line edges, remove empty lines, or remove duplicate lines. Duplicate matching is exact after the selected cleanup, keeps the first line, and distinguishes letter case.

  3. Clean and review

    Select Clean text and check the result and metrics. Output line breaks use LF. You can cancel while processing; editing the source or options invalidates an earlier result.

  4. Copy, download, or clear

    Copy the result when it is ready. Only a non-empty result can be downloaded. Clear all removes the source, result, metrics, and selected options from this page.

Frequently asked questions

Will this rewrite my content?

Only the selected line-by-line cleanup is applied, and line breaks are normalized to LF. Words, letter case, and writing system are not rewritten.

Why does the character count differ from my word processor?

This tool counts Unicode code points, including spaces and line breaks. An emoji or a character with combining marks can contain multiple code points. Code points are not the same as linguistic words.

How are duplicate lines identified?

Lines are compared exactly after the selected cleanup, and the first occurrence is kept. Apple and apple differ; if trimming is off, a difference in surrounding spaces also makes lines distinct.

Will my text still be here after I leave?

The tool does not actively save text or provide history to restore. Do not expect it to recover text after reloading or leaving. Downloaded files and clipboard contents are managed by your device.

Data handling and clearing

Text is cleaned in your browser, and the tool does not actively save source text or cleaning history. Clear all removes the source, result, metrics, options, and errors from this page. It does not delete downloaded files or clear your device-managed clipboard. On a shared device, check downloaded files and clipboard contents before clearing text from the page.