Line Tools
Sort or deduplicate lines with complete comparison, blank-line and retention controls.
Sorting treats only zero-length lines as empty. Deduplication can exclude every whitespace-only candidate before grouping; trimming changes comparison keys only. The report includes exact source, every parsed line, every kept and removed source position, options and full output.
Runs locally in this browser; no text is uploaded.
A QUICK WALKTHROUGH
How to use this tool
- Source text — Input: at most 100,000 UTF-16 units and 50,000 lines. Output: at most 1,000,000 units. Report: at most 4,000,000 units. The whole operation is rejected on overflow.
- Operation — Sort lines / Remove duplicate lines
- Transform → Complete transformation report → Copy transformed text / Download text (UTF-8 TXT)
Comparison method
Original retained spelling is preserved. Sorting is stable for equal keys; deduplication keeps first or last occurrences in their source order. English collation is fixed across page languages; case-insensitive collation uses base sensitivity and can also equate accents. Code-point order compares full Unicode code points.
Newline policy
CRLF and CR become LF. A trailing newline creates an empty final line. Output joins retained lines with LF without adding a newline.
Complete transformation report
Sorting treats only zero-length lines as empty. Deduplication can exclude every whitespace-only candidate before grouping; trimming changes comparison keys only. The report includes exact source, every parsed line, every kept and removed source position, options and full output.
Limits and local processing
Input: at most 100,000 UTF-16 units and 50,000 lines. Output: at most 1,000,000 units. Report: at most 4,000,000 units. The whole operation is rejected on overflow. Lone UTF-16 surrogates cannot be preserved in UTF-8 TXT. Replace them with valid Unicode scalar values. Runs locally in this browser; no text is uploaded.
GOOD TO KNOW
Common questions
Does trimming rewrite retained lines?
Original retained spelling is preserved. Sorting is stable for equal keys; deduplication keeps first or last occurrences in their source order. English collation is fixed across page languages; case-insensitive collation uses base sensitivity and can also equate accents. Code-point order compares full Unicode code points.
How are blank lines and removed records reported?
Sorting treats only zero-length lines as empty. Deduplication can exclude every whitespace-only candidate before grouping; trimming changes comparison keys only. The report includes exact source, every parsed line, every kept and removed source position, options and full output.
Is an oversized result shortened?
Input: at most 100,000 UTF-16 units and 50,000 lines. Output: at most 1,000,000 units. Report: at most 4,000,000 units. The whole operation is rejected on overflow.